Newest

All articles, most recently crawled first.

Cards List

Palomar: A registry of Lean verified mathematics

Hacker News Top · 6h ago Cached

Palomar is a new registry for Lean verified mathematics proofs, designed to help validate formal proofs using mechanical checks and AI-assisted methods.

0 favorites 0 likes

Cerebras CS-4

Hacker News Top · 9h ago Cached

Cerebras launches the CS-4, a rack-scale AI system with WSE-3 Turbo technology claiming up to 30x faster inference than GPUs, featuring a modular design for efficient hyperscale deployment.

0 favorites 0 likes

OpenLogi

Hacker News Top · 7h ago Cached

OpenLogi is an open-source, local-first tool written in Rust for configuring Logitech mice over HID++, offering button remapping, DPI control, and per-app profiles without requiring an account or telemetry.

0 favorites 0 likes

Meta's blockbuster trial draws parallels to big tobacco

Hacker News Top · 7h ago

Meta is involved in a major trial that draws parallels to historical big tobacco cases, indicating significant legal scrutiny for the tech industry.

0 favorites 0 likes

Dynamic Multi-Byte Prediction With Hierarchical Language Models

Hugging Face Daily Papers · 3d ago Cached

The paper introduces multi-byte prediction to speed up inference in byte-level language models by generating multiple bytes in parallel with minimal performance impact.

0 favorites 0 likes

Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Requirements

Hugging Face Daily Papers · yesterday Cached

This paper proposes Agentic ESOpt, a method using evolution strategies to enable scalable full-parameter fine-tuning of long-horizon LLM agents with minimal GPU memory requirements.

0 favorites 0 likes

From Sequence to Structure: Relational Uncertainty Propagation for LLM Agents

Hugging Face Daily Papers · 2d ago Cached

The paper proposes RUPA, a framework that models LLM agent execution as a dependency graph to propagate uncertainty, improving failure detection and confidence estimation in long trajectories.

0 favorites 0 likes

aDSL: Agentic 3D Creation via Joint Agent-Program Design

Hugging Face Daily Papers · yesterday Cached

The paper introduces aDSL, a co-designed domain-specific language and multi-agent system that improve LLM-driven 3D program synthesis through relational operators and iterative feedback, enhancing robustness and controllability in 3D content creation.

0 favorites 0 likes

StartupBench: Benchmarking General-Purpose Agents on Market-Validated End-to-End Workflows

Hugging Face Daily Papers · yesterday Cached

StartupBench introduces a benchmark for evaluating general-purpose AI agents on real-world startup workflows, revealing that top models complete only about 30% of tasks due to gaps in complex instruction following and domain-specific expertise.

0 favorites 0 likes

Abra: Scaling Diffusion Image Training

Hugging Face Daily Papers · yesterday Cached

This paper presents a systematic scaling law study for text-to-image diffusion models, showing they scale predictably but require significantly more data per parameter than language models for optimal training.

0 favorites 0 likes

ASI-Bench: At the Dawn of Artificial Superintelligence

Hugging Face Daily Papers · yesterday Cached

ASI-Bench is a new benchmark designed to evaluate AI systems' capabilities in innovative exploration and autonomous scientific execution across 11 scientific domains, revealing current AI's heavy dependence on human guidance.

0 favorites 0 likes

Security Assessment of DeepSeek Harness with A.I.G: Evaluating Resistance to Indirect Prompt Injection

Hugging Face Daily Papers · yesterday Cached

This paper evaluates indirect prompt injection risks in DeepSeek Harness using AI-Infra-Guard for controlled testing, finding notable attack success rates and recommending security controls.

0 favorites 0 likes

GS-Voxel: Fitting-Free Structured Latents for Large-Scale 3DGS Generation

Hugging Face Daily Papers · yesterday Cached

GS-Voxel introduces a fitting-free framework to convert 3D Gaussian Splatting reconstructions into structured latents, enabling scalable generation of large-scale aerial 3D scenes via flow models and tiled inference.

0 favorites 0 likes

From Corpora to Co-Evolving Capabilities: Capability-Centric Data Design for Generalist Image Generation

Hugging Face Daily Papers · yesterday Cached

The paper introduces a capability-driven data infrastructure with curriculum scheduling to train generalist image generation models using heterogeneous supervision for diverse generative tasks.

0 favorites 0 likes

Energy-Guided Flow Matching

Hugging Face Daily Papers · 2026-08-07 Cached

Energy-Guided Flow Matching improves generative image quality by using a moving endpoint and adaptive scheduling, achieving state-of-the-art FID scores with reduced training cost.

0 favorites 0 likes

Embodied-Navigator: Point, Think, Memorize, and Align for Efficient Navigation

Hugging Face Daily Papers · yesterday Cached

TAMP-Nav is a unified framework that enhances embodied navigation by aligning vision-language models with 2D visual prompting, selective reasoning with compressed memory, and policy optimization, achieving state-of-the-art performance with high runtime and sample efficiency.

0 favorites 0 likes

FreeToken: Efficient Edge-Native MoE Serving with Bandwidth-Adaptive Execution

Hugging Face Daily Papers · 2d ago Cached

FreeToken is an edge-native serving system that dynamically maps computation and model state onto heterogeneous local hardware to run large open-weight models on personal machines, enabling efficient execution of models up to 753B on a single GPU.

0 favorites 0 likes

Unifying Graph Neural Networks Through a Common Layer Equation

Hugging Face Daily Papers · 2d ago Cached

The paper introduces a common layer equation that unifies graph neural networks into seven components, enabling architectural comparison, theoretical analysis, and insights into issues like oversmoothing and expressivity.

0 favorites 0 likes

When AI art has no author: Study finds generated images often can’t be traced to training data

MIT News — Artificial Intelligence · 17h ago Cached

MIT CSAIL researchers discovered 'attribution decay' in large generative AI models, where generated images often cannot be traced back to individual training data, using a novel diffusion ensemble architecture.

0 favorites 0 likes

Vim wants you to control, VSCode wants you to consume

Hillel Wayne — Computer Things · 17h ago Cached

The article contrasts Vim's focus on user control through programmatic customization with VSCode's emphasis on consumption via heavier plugin development.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback