Scores updated · 6 tests both models report · How we compare
Every test, side by side
All 6 tests both models report. The winning score is in its model's colour; marks a score checked independently.
Other results6 tests
- BigCodeBench (Full)Yi Coder 9B Chat by 11.937.149+11.9
- MHPPSeed Coder 8B Instruct by 9.536.226.7+9.5
- BigCodeBench (Hard)Seed Coder 8B Instruct by 8.826.417.6+8.8
- LiveCodeBench (v5)Seed Coder 8B Instruct by 7.224.717.5+7.2
- MBPPYi Coder 9B Chat by 2.479.682+2.4
- HumanEvaltie82.982.3tie
Questions people ask
How do you compare the two?
We use the 6 benchmark tests both models have published scores on. None of them is in the eight capability areas we count, so this page lists them without a verdict. Each score is the one shown on the model's own page.