tencent/Hunyuan-A13B-Instruct
Hunyuan · → text
Hunyuan-A13B-Instruct has 8 billion parameters and can match larger models by activating only 1.3 billion parameters, supporting "fast thinking/slow thinking" hybrid inference. It offers stable long text understanding. Verified by BFCL-v3 and τ-Bench, its Agent capabilities are leading in the field. Combined with GQA and multiple quantization formats, it enables efficient inference.
Input$0.14 /M
Output$0.56 /M
