tencent/Hunyuan-A13B-Instruct
HunyuanHunyuan-A13B-Instruct has 8 billion parameters and can match larger models by activating only 1.3 billion parameters, supporting "fast thinking/slow thinking" hybrid inference. It offers stable long text understanding. Verified by BFCL-v3 and τ-Bench, its Agent capabilities are leading in the field. Combined with GQA and multiple quantization formats, it enables efficient inference.
Pricing
- Input Tokens: $0.14 /M tokens
- Output Tokens: $0.56 /M tokens
Input Modalities
Try this model
Python