Reddit

Articles from Reddit

Cards List

Apéry irrationality marked solved on FrontierMath

Reddit r/singularity ↗ · 7h ago Cached

The irrationality of ζ(5) has been proved using an AI-assisted method, and the solution is marked as solved on FrontierMath despite not following the Apéry-style proof as initially sought.

0 favorites 0 likes

Nonobench v1.2: 43 LLMs on nonogram puzzles. Open-weight DeepSeek V4 Pro ties for 4th, and no open model solves the new 20×20 Hard mode

Reddit r/LocalLLaMA ↗ · 8h ago

Nonobench v1.2 benchmarks 43 LLMs on nonogram puzzles, with GPT-6 Astra achieving the first perfect score on 15x15 puzzles and open-weight models failing the new 20×20 hard mode.

0 favorites 0 likes

We added the "Vibe" into vibecoding - Introducing the first Malleable AI workstation

Reddit r/AI_Agents ↗ · 8h ago

Introducing the first Malleable AI workstation that integrates multiple AI models like Gemini, Claude, Codex, and Grok, allowing users to work across sessions and devices in parallel and build custom widgets with prompts.

0 favorites 0 likes

How is GPT-6-Luna so good?!

Reddit r/openclaw ↗ · 8h ago

A user enthusiastically reviews GPT-6-Luna and Sol, highlighting their superior performance, context management, and cost-effectiveness compared to other AI models and services.

0 favorites 0 likes

AI agents that allows you to DO LESS

Reddit r/AI_Agents ↗ · 9h ago

The author discusses how most AI agents increase workload, but highlights the product 'Catch' as it helps reduce tasks by handling reminders, scheduling, and emails without constant supervision.

0 favorites 0 likes

Claude Opus 5.5 controlled a robot arm to copy Michelangelo, noticed it had left a broken line, and went back to fix the mistake on its own

Reddit r/singularity ↗ · 9h ago

Claude Opus 5.5 demonstrated autonomous control of a robot arm to replicate Michelangelo's work and independently noticed and corrected a mistake.

0 favorites 0 likes

I think the missing layer in agent systems is the organisation itself

Reddit r/AI_Agents ↗ · 9h ago

The author shares insights from building OSIO, an operating system for human-AI collaboration in businesses, arguing that organizations should persist above individual AI agents to manage capability, authority, and memory.

0 favorites 0 likes

The Intelligence Age: Judgment

Reddit r/ArtificialInteligence ↗ · 9h ago

AI drastically reduces the cost of generating representations like business plans and code, shifting the organizational bottleneck from creation to validation and judgment in real-world contexts.

0 favorites 0 likes

Scotland's hidden pipeline of 'phantom' AI data centres

Reddit r/artificial ↗ · 9h ago Cached

Scotland may have around 30 AI data centres in the pipeline with a potential electricity demand of 9.7 GW, but many are not publicly registered, raising concerns about transparency and energy planning.

0 favorites 0 likes

[Open Source] Browser-Only Tool To Give Your Agent a stable UI with 1 click, for quick testing, or just visual feeling

Reddit r/AI_Agents ↗ · 9h ago

A browser-only open-source tool that connects to an OpenAI-compatible agent and provides a stable, polished UI for testing, with side-by-side performance metrics like spans, turns, and tokens.

0 favorites 0 likes

'Godfather of AI' has a plan for humanity to survive AI

Reddit r/ArtificialInteligence ↗ · 9h ago Cached

Geoffrey Hinton, the 'godfather of AI,' warns that AI could wipe out humanity and suggests tech companies build 'maternal instincts' into AI models to ensure they care about humans.

0 favorites 0 likes

How do you find out afterwards whether your agent's memory was right?

Reddit r/AI_Agents ↗ · 10h ago

The author discusses the challenge of verifying the correctness of AI agents' long-term memory and asks for community experiences on tracking and correcting memory errors.

0 favorites 0 likes

Power Limits, Local AI, and Questionable Uses of My Free Time

Reddit r/LocalLLaMA ↗ · 10h ago

A user shares detailed benchmarking data and personal insights on running AI models locally with varying GPU power limits, evaluating models like gemma4 and qwen3.5 on a modest hardware setup.

0 favorites 0 likes

Permission to Act Is Not the Same as Evidence to Act

Reddit r/AI_Agents ↗ · 10h ago

The article distinguishes between permission for AI agents to act and the evidence required to justify action, introducing the Organic Intelligence Protocol (OIP) to establish evidence boundaries, illustrated by a recent audit case study.

0 favorites 0 likes

Finnish Study of 2,000+ Workers Found No Clear Link Between Frequent AI Use and Exhaustion — Social Comparison Was More Consistently Associated

Reddit r/ArtificialInteligence ↗ · 10h ago Cached

A Finnish study of over 2,000 workers found no clear link between frequent AI use and workplace exhaustion, but social comparison was consistently associated with higher exhaustion levels.

0 favorites 0 likes

Qwen, where's the small stuff? (1B/2B/4B)

Reddit r/LocalLLaMA ↗ · 10h ago

The article questions why Qwen hasn't released new small-scale models (1B/2B/4B), which affects accessibility and development for hardware-limited users, and mentions a similar issue with Google's Gemma series.

0 favorites 0 likes

Claude Opus 5.5 designed a processor faster and smaller than the human-made one on the HWE benchmark

Reddit r/singularity ↗ · 10h ago

Claude Opus 5.5 has designed a processor that is faster and smaller than the human-made VexRiscv on the HWE benchmark.

0 favorites 0 likes

herdr, read at one pinned commit: strong on agent coordination, no record of cost or unattended work

Reddit r/AI_Agents ↗ · 11h ago

herdr is a Rust binary that serves as a background server for running coding agents, praised for its pane state model and agent-driven coordination, but criticized for missing cost accounting and unattended run logging.

0 favorites 0 likes

Meta launched a personal AI agent with its own computer. I tested it.

Reddit r/artificial ↗ · 12h ago

Meta launched Muse, a personal AI agent with its own computer, persistent memory, and tools like a terminal and browser. The author tested it by generating a 30-minute video from a single prompt, and it's officially available in the US/Canada with a VPN workaround for India.

0 favorites 0 likes

US and Russia cut human review of AI targets from UN draft

Reddit r/artificial ↗ · 12h ago

US and Russia removed the requirement for human review of AI-generated targets from a UN draft, potentially weakening global efforts to regulate autonomous weapons.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback