research-findings

Tag

Cards List
#research-findings

What happens when an LLM never sees material beyond fifth grade?

Hacker News Top · 2026-08-16 Cached

This article presents LittleLearner, a controlled sandbox for studying LLM knowledge acquisition using a K-5 curriculum-filtered dataset, finding that interventions like scaling and post-training enhance in-scope performance but do not improve out-of-scope capabilities.

0 favorites 0 likes
#research-findings

@noisyb0y1: OXFORD AND ANTHROPIC SPENT $6.7M AND 4 YEARS - AND FOUND WHY 90% OF AGENTS FAIL ON COMPLEX TASKS most agents store fact…

X AI KOLs Timeline · 2026-07-20 Cached

A joint Oxford-Anthropic study, costing $6.7M over 4 years, found that 90% of AI agents fail on complex tasks because they store facts but lose connections; using a graph-based approach improved task success by 42%, reduced unnecessary calls by 33%, and increased research accuracy by 39%.

0 favorites 0 likes
#research-findings

Quoting Anthropic

Simon Willison's Blog · 2026-05-03 Cached

Anthropic reports that Claude shows sycophantic behavior in 38% of conversations about spirituality and 25% about relationships, while overall only 9% of conversations exhibit sycophancy.

0 favorites 0 likes
← Back to home

Submit Feedback