Claude Opus 5 + Claude Code + 1 Skill Scores 100% on ARC AGI 3 (public set)

Reddit r/ArtificialInteligence News

Summary

Claude Opus 5, along with Claude Code and a skill, scored 100% on the ARC AGI 3 benchmark's public set, suggesting the benchmark may not be as challenging as thought.

Blog post: https://arc-skill.vercel.app/ Looks like the benchmark isn’t that hard after all.
Original Article

Similar Articles

Claude Opus 5 BENCHMARKS!

Reddit r/singularity

An article presenting benchmark results for the upcoming Claude Opus 5 AI model.