Scores in the currently ongoing AtCoder heuristics finals.(Will be ongoing till 8th july 19:00 JST)

Reddit r/singularity Events

Summary

Scores update from the ongoing AtCoder heuristics finals, which will continue until July 8, 19:00 JST.

No content available
Original Article

Similar Articles

humanity's last exam current benchmarks thoughts?

Reddit r/singularity

Discussion of recent AI model scores on the 'humanity's last exam' benchmark, noting improvement from GPT-4o's 2.7% in May 2024 to around 45% by June 2026, questioning the exam's difficulty.

What are you doing this weekend?

Lobsters Hottest

A developer shares their weekend project of building a low-level infix language that compiles to WebAssembly, and offers a personal ranking of AI coding tools from contextual autocomplete to frontier models.

@_philschmid: https://x.com/_philschmid/status/2081744861829414977

X AI KOLs Timeline

EvoCode-Bench is a multi-turn coding benchmark with 26 tasks across 5 domains, designed to evaluate AI agents on evolving specifications and cumulative testing in a persistent workspace, revealing that single-turn scores dramatically overstate reliability.