Qwen2.5 Technical Report

Benchmark Model Rank Results
code-generation-on-humanevalQwen2.5-Plus–Pass@1: 87.8
mathematical-reasoning-on-aime24Qwen2.5-72B-Instruct#5Acc: 23.3
question-answering-on-gpqaQwen2.5-72B-Instruct#4Accuracy: 49