Achieves effective integration of thinking and non-thinking modes, allowing mode switching during conversations. Its reasoning ability significantly surpasses QwQ, and its general capability significantly exceeds Qwen2.5-32B-Instruct, reaching state-of-the-art (SOTA) levels among industry models of the same scale.
Pricing
- Input Tokens: $0.32 /M tokens
- Output Tokens: $3.2 /M tokens
- Cache Read: $0 /M tokens
Input Modalities
Output Modalities
- Text
Try this model
Python