Across 6 shared benchmarks, InternVL3.5-8B-Instruct scores higher on 0 and Qwen3.5 122B A10B on 6. The widest gap is SimpleVQA, where Qwen3.5 122B A10B scores 61.7 against 40.8.
Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.