Issue 2026-09-28 · Industry · 研究 · 模型发布 · 开源

Qwen3-VL 8B on laptop competes with top models on messy docs, fails on date formats

A benchmark shows Qwen3-VL 8B (Q4, on a laptop) got 59% of 137 messy documents fully right, beating GPT-5.6 Terra's 57% but trailing Claude Opus 5.5 (89%) and Sonnet 5 (85%). It excelled on US tax forms (21/32 vs 7/32) but struggled with Indian date formats and long contracts.

r/MachineLearning13 d ago
Read original ↗