The model is built on the Mixture of Experts architecture with a total size of 406B parameters and 32B active ones.
The model supports a context of 256K tokens. HY 2.0 shows significant improvements on key benchmarks.
Main achievements of HY 2.0:
🧠 Reasoning: a score of 73.4 on IMO AnswerBench - almost a 20 percent increase, securing the model among the leaders in mathematical and scientific reasoning.
🛠 Coding and Agents: a leap in SWE Bench Verified from 6.0 to 53.0, and Tau2 Bench grew from 17.1 to 72.4.
⚡ Instruction Following: more stable execution of complex instructions and a natural style of responses.
The model is released in two versions:
• HY 2.0 Think - for deep reasoning, code generation, and complex tasks
• HY 2.0 Instruct - for dialogue, creative writing, and multi-turn contextual conversations
Website, API Access, Documentation
#AI #Tencent #Hunyuan #HY2 #LLM #MoE #DeepLearning #AIModels
••••••••••••••••••••••••••••••••••••••••••••••••••••
🤖 Data Science, ML & Big Data with @DataXplore
