Achieves effective integration of thinking and non-thinking modes, allowing mode switching during conversations. Its reasoning ability matches that of QwQ-32B with a smaller parameter size, and its general capability significantly surpasses Qwen2.5-14B, reaching state-of-the-art (SOTA) levels among industry models of the same scale.
Pricing
- Input Tokens: $0.12 /M tokens
- Output Tokens: $1.2 /M tokens
- Cache Read: $0 /M tokens
Input Modalities
Try this model
Python