AI article

Benchmarking Frontier AI on Curved Tire Sidewall OCR

Community description: This is a submission for the Kaggle Benchmarking Challenge What I Benchmarked In...

Dev.to | Oct 11, 2026 | Malawige Inusha Thathsara Gunasekara

Automated excerpt

Anthropic: Claude Opus 5. 5, Claude Sonnet 5. 5, and Claude Haiku 5. 5. It matched proprietary frontier flagships Claude Sonnet 5. 5 and Claude Haiku 5. 5, while beating OpenAI GPT-6. 1 Sol (76. 47%). Claude Haiku 5. 5 Punched Way Above Its Weight Anthropic's Claude Haiku 5. 5 tied Claude Sonnet 5. 5 exactly at 41 / 51 (80. 39%).

Selected automatically from source text; not independently written or fact-checked. Read the original for full context.

Read the original article

Read next

AI briefing: recent picks

More stories to explore