Scores in the currently ongoing AtCoder heuristics finals.(Will be ongoing till 8th july 19:00 JST)
Summary
Scores update from the ongoing AtCoder heuristics finals, which will continue until July 8, 19:00 JST.
Similar Articles
How can Deepseek v4 top the coding leaderboards and still sit 8 months behind the frontier?
Analysis of DeepSeek V4's top coding scores versus its reported 8-month gap behind the frontier, highlighting differences between narrow benchmark optimization and broader reasoning tests, plus the practical performance hit when running quantized local versions.
humanity's last exam current benchmarks thoughts?
Discussion of recent AI model scores on the 'humanity's last exam' benchmark, noting improvement from GPT-4o's 2.7% in May 2024 to around 45% by June 2026, questioning the exam's difficulty.
@BenjaminDEKR: So it's just over? GPT5.6 Sol Ultra scores 91.9% on TerminalBench Coding is approaching solved, the same way arithmetic…
GPT5.6 Sol Ultra achieves 91.9% on TerminalBench coding benchmark, suggesting coding tasks are approaching solved.
What are you doing this weekend?
A developer shares their weekend project of building a low-level infix language that compiles to WebAssembly, and offers a personal ranking of AI coding tools from contextual autocomplete to frontier models.
Claude Opus 4.8 scores over 1% on ARC-AGI 3 !!
Claude Opus 4.8 achieves a score of over 1% on the ARC-AGI 3 benchmark, demonstrating slight progress on a difficult AI reasoning test.