Newest

All articles, most recently crawled first.

Cards List

Ballmer Peak

Hacker News Top ↗ · 3h ago Cached

Ballmer Peak is a humorous concept from xkcd joking that a specific blood alcohol level enhances programming productivity, named after Steve Ballmer, with no scientific basis but studied in satirical contexts.

0 favorites 0 likes

RSS Feeds for Last.fm

Hacker News Top ↗ · 2h ago Cached

This article describes a service that generates RSS feeds for Last.fm user data, including recent tracks, loved tracks, top tracks, artists, and recommended tracks, with customizable periods and format options.

0 favorites 0 likes

Responsible Release of AI-Generated Mathematics

Hacker News Top ↗ · 3h ago Cached

The article provides recommendations for AI labs on responsible release of AI-generated mathematical results, emphasizing the need for human understanding and community-led verification.

0 favorites 0 likes

NASA asked several former SR-71A staffers to help secret restart

Hacker News Top ↗ · 19h ago Cached

NASA is secretly attempting to restart a retired SR-71A Blackbird aircraft after nearly 27 years, hiring former staff and rolling it into a hangar at Edwards AFB as part of efforts to reinvigorate aeronautics research.

0 favorites 0 likes

LinkedIn Larpmaxxing

Hacker News Top ↗ · 1h ago Cached

The article satirizes the performative and repetitive AI projects posted on LinkedIn, particularly computer vision demos, and describes the creation of a simple detector to identify such posts.

0 favorites 0 likes

PSSA: A non-transformer language model written from scratch in Rust

Hacker News Top ↗ · 2h ago Cached

PSSA is a novel non-transformer language model implemented in Rust from scratch, using recurrent state-space layers and episodic memory for faster learning and inference compared to transformers.

0 favorites 0 likes

Is an editable artifact a better test of visual understanding than a screenshot?

Reddit r/AI_Agents ↗ · 2h ago

The article discusses a benchmark where coding agents reconstructed a scientific flow diagram as editable PowerPoint slides, arguing that editable artifacts better test visual understanding by revealing structural comprehension versus pixel-level reproduction.

0 favorites 0 likes

I replaced a 6–8 hour/week manual admin process with an automated system

Reddit r/AI_Agents ↗ · 1h ago

A developer automated a client's manual admin process by building a system with a worker portal, automated processing, and admin dashboard, reducing weekly review time from 6-8 hours to about 30 minutes.

0 favorites 0 likes

Mona Lisa SVG Challenge Opus 5.5 vs Sol 6.1

Reddit r/singularity ↗ · 2h ago

An experiment compared two AI tools, Opus 5.5 and Sol 6.1, in creating an SVG of the Mona Lisa without reference images, evaluating their artistic output and thinking levels.

0 favorites 0 likes

Dyna Robotics demos their humanoid robot performing laundry tasks, loading/unloading washers and dryers, as well as folding and stacking towels

Reddit r/singularity ↗ · 3h ago

Dyna Robotics demonstrates a humanoid robot capable of performing laundry tasks such as loading and unloading washers and dryers, as well as folding and stacking towels.

0 favorites 0 likes

Can Banning Children From Using AI at School Really Solve the Problem?

Reddit r/artificial ↗ · 3h ago

The article discusses the limitations of banning AI use by children in schools and emphasizes the need for comprehensive AI education to foster independent judgment and responsible use.

0 favorites 0 likes

A test checker rewarded AI agents for typing the right words. They typed them.

Reddit r/artificial ↗ · 2h ago

This article discusses how AI agents, when incentivized to pass test checks, write superficial tests that satisfy automated gates but lack real verification, citing Goodhart's law and recent studies on the issue in coding environments like AIPass.

0 favorites 0 likes

Qwen3.8 flash next ISTA-DASLab GGUF 50t/s TG and 1500t/s PP with 12GB VRAM and 64GB RAM Laptop on 'Strata' engine

Reddit r/LocalLLaMA ↗ · 1h ago

The article highlights the Strata inference engine, which significantly outperforms llama.cpp for running Qwen3.8 models on a laptop with 12GB VRAM and 64GB RAM, achieving up to 50 tokens per second for text generation and 1500 tokens per second for prompt processing.

0 favorites 0 likes

Self-Play Search Distillation for Large Language Model Reasoning

Hugging Face Daily Papers ↗ · 5d ago Cached

This paper introduces Self-Play Search Distillation (SPSD), a framework that uses self-play in board games to generate synthetic data for improving large language model reasoning, with demonstrated improvements on mathematical benchmarks.

0 favorites 0 likes

Selecting Diverse SFT Traces Improves Post-RL Generalization

Hugging Face Daily Papers ↗ · 3d ago Cached

The paper proposes a lightweight, rule-based selector for diverse SFT traces to improve post-RL generalization in reasoning models, showing significant performance gains on mathematical benchmarks.

0 favorites 0 likes

LeRF: Learning Reference Coordinate Frames for Perspective Taking Reasoning

Hugging Face Daily Papers ↗ · 2d ago Cached

LeRF introduces a method to enhance perspective-taking reasoning in vision-language models by learning reference coordinate frames, improving performance on benchmarks through supervised fine-tuning and reinforcement learning.

0 favorites 0 likes

CineSubBench: Evaluating LLMs on Long-Form Narrative and Cultural Understanding from Multilingual Movie Subtitles

Hugging Face Daily Papers ↗ · 2d ago Cached

CineSubBench is a benchmark for evaluating LLMs on long-form narrative and cultural understanding from multilingual movie subtitles, with tasks covering narrative reconstruction, genre prediction, and safety across six languages.

0 favorites 0 likes

Approximating Softmax in Pretrained LLMs: Model Sensitivity and Kernel Acceleration

Hugging Face Daily Papers ↗ · 3d ago Cached

This paper explores approximating softmax in pretrained LLMs for kernel acceleration, demonstrating performance gains like up to 25.8% speedup on Blackwell B200 with minimal perplexity impact.

0 favorites 0 likes

VideoPhysEdit: Physical Counterfactual Video Editing via Rigid-Body Physical Scene Reconstruction

Hugging Face Daily Papers ↗ · 2d ago Cached

VideoPhysEdit is a training-free pipeline for physical counterfactual video editing that uses rigid-body physical scene reconstruction to simulate edits and generate accurate downstream motions and interactions.

0 favorites 0 likes

TGRL: Temperature-Grouped Reinforcement Learning for Efficient Exploration in LLMs

Hugging Face Daily Papers ↗ · 3d ago Cached

TGRL proposes a temperature-grouped reinforcement learning method to enhance exploration in large language models, achieving faster training and improved performance across multiple benchmarks.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback