AI article
Benchmarking Frontier AI on Curved Tire Sidewall OCR
Community description: This is a submission for the Kaggle Benchmarking Challenge What I Benchmarked In...
Dev.to | Oct 11, 2026 | Malawige Inusha Thathsara Gunasekara
Automated excerpt
Anthropic: Claude Opus 5. 5, Claude Sonnet 5. 5, and Claude Haiku 5. 5. It matched proprietary frontier flagships Claude Sonnet 5. 5 and Claude Haiku 5. 5, while beating OpenAI GPT-6. 1 Sol (76. 47%). Claude Haiku 5. 5 Punched Way Above Its Weight Anthropic's Claude Haiku 5. 5 tied Claude Sonnet 5. 5 exactly at 41 / 51 (80. 39%).
Selected automatically from source text; not independently written or fact-checked. Read the original for full context.