Qwen2.5 72B Instruct
Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2:
-
Significantly more knowledge and has greatly improved capabilities in coding and mathematics, thanks to our specialized expert models in these domains.
-
Significant improvements in instruction following, generating long texts (over 8K tokens), understanding structured data (e.g, tables), and generating structured outputs especially JSON. More resilient to the diversity of system prompts, enhancing role-play implementation and condition-setting for chatbots.
-
Long-context Support up to 128K tokens and can generate up to 8K tokens.
-
Multilingual support for over 29 languages, including Chinese, English, French, Spanish, Portuguese, German, Italian, Russian, Japanese, Korean, Vietnamese, Thai, Arabic, and more.
Usage of this model is subject to Tongyi Qianwen LICENSE AGREEMENT.
32,768 Token Context
Process and analyze large documents and conversations.
Advanced Coding
Improved capabilities in front-end development and full-stack updates.
Agentic Workflows
Autonomously navigate multi-step processes with improved reliability.
Available On
Provider | Model ID | Context | Max Output | Input Cost | Output Cost | Throughput | Latency |
---|---|---|---|---|---|---|---|
DeepInfra | deepInfra | 33K | 16K | $0.12/M | $0.39/M | 34.8 t/s | 598 ms |
Nebius AI Studio | nebiusAiStudio | 131K | - | $0.13/M | $0.40/M | 19.3 t/s | 1428 ms |
NovitaAI | novitaAi | 32K | 4K | $0.38/M | $0.40/M | 22.0 t/s | 1583 ms |
Hyperbolic | hyperbolic | 131K | - | $0.40/M | $0.40/M | 22.8 t/s | 2022 ms |
Fireworks | fireworks | 33K | - | $0.90/M | $0.90/M | 73.7 t/s | 637 ms |
Together | together | 131K | 2K | $1.20/M | $1.20/M | 104.9 t/s | 667 ms |
per 1K tokens
per 1K tokens