Qwen/Qwen2.5-Math-72B
TEXT GENERATIONConcurrency Cost:4Model Size:72.7BQuant:FP8Ctx Length:32kPublished:Sep 16, 2024License:qwenArchitecture:Transformer0.0K Warm

Qwen/Qwen2.5-Math-72B is a 72.7 billion parameter mathematical language model developed by Qwen, specifically designed for solving math problems in both English and Chinese. It supports Chain-of-Thought (CoT) and Tool-integrated Reasoning (TIR) for enhanced computational accuracy and algorithmic manipulation. This model is optimized for mathematical tasks and serves as a strong base for fine-tuning, offering significant performance improvements over its predecessor on mathematical benchmarks.

Loading preview...

Popular Sampler Settings

Top 3 parameter combinations used by Featherless users for this model. Click a tab to see each config.

temperature
top_p
top_k
frequency_penalty
presence_penalty
repetition_penalty
min_p