Skip to main content
Qwenqwen-2.5-72b
Context window32,768 tokens
StreamingJSON modeTool callingMultilingual
Maximum output: 8,192 tokens.

Pricing

Typical request (5K input / 1K output tokens): ~$0.0022. Live pricing.

Use this model

Set up your SDK client, then run:

Supported parameters

messagestemperaturemax_completion_tokenstop_pstopfrequency_penaltypresence_penaltyseedstreamuserroutingtoolstool_choiceresponse_format
Send only parameters from this list. routing.require_parameters (default true) skips any provider rail that cannot honor every requested parameter, so requests keep the behavior you asked for.

Routing & fallbacks

Pin qwen-2.5-72b with model when you want this model’s behavior, or list it in an ordered models array. With ninja/auto first, the router’s top pick runs first and qwen-2.5-72b is the explicit fallback; ninja/auto acts as the router only when it is the first entry.
You are billed at the resolved model’s token rates, and only for the successful execution.
TypeScript SDK
Guides: Text generation · Fallbacks · Smart routing · Pricing · All models