Model
Opus 5 Released: Two-Month Performance Leap, ARC-AGI 3 Surpasses 30%
Just two months after its release, Opus 5's ARC-AGI 3 benchmark performance jumped from less than 5% on Opus 4.8 to over 30%. According to Artificial Benchmark, Opus 5's intelligence level is comparable to Fable 5, but with 26% lower cost per task. The model outperforms Fable 5 on most evaluations, marking rapid iteration of AI capabilities.
Read the original (opens in a new tab)
News stream data aggregated by AI HOT