All articles, most recently crawled first.
Ant Group has joined the PyTorch Foundation as a Gold Member, working on open AI models and infrastructure to make AI more accessible, and is investing in the AReaL project.
A tweet by user Qinjianbo1984 sharing a link to external content, likely related to technology or AI.
CoreWeave achieved a quarterly revenue of $2.6 billion in about 25 quarters, a feat that AWS took 40 quarters to accomplish since commercialization, demonstrating the rapid growth of Neocloud. AI computing power demand is compressing the growth cycle of cloud computing.
The tweet expresses enthusiasm for AI research opportunities, particularly with AI agents, and highlights the rapid progress in machine learning.
A tiny UK-developed satellite, CosmoCube, will use the far side of the Moon as a shield to detect faint signals from the early universe, aiming to study the cosmic dark ages and the role of dark matter.
The article explains CUDA shared memory swizzling techniques to optimize GPU memory access patterns, with code examples demonstrating performance improvements.
Scientists found that children's lung growth, previously stunted by air pollution, showed rapid recovery in London after the introduction of an ultra-low emission zone, suggesting clean air policies can quickly improve child health.
Palomar is a new registry for Lean verified mathematics proofs, designed to help validate formal proofs using mechanical checks and AI-assisted methods.
Cerebras launches the CS-4, a rack-scale AI system with WSE-3 Turbo technology claiming up to 30x faster inference than GPUs, featuring a modular design for efficient hyperscale deployment.
OpenLogi is an open-source, local-first tool written in Rust for configuring Logitech mice over HID++, offering button remapping, DPI control, and per-app profiles without requiring an account or telemetry.
Meta is involved in a major trial that draws parallels to historical big tobacco cases, indicating significant legal scrutiny for the tech industry.
The paper introduces multi-byte prediction to speed up inference in byte-level language models by generating multiple bytes in parallel with minimal performance impact.
This paper proposes Agentic ESOpt, a method using evolution strategies to enable scalable full-parameter fine-tuning of long-horizon LLM agents with minimal GPU memory requirements.
The paper proposes RUPA, a framework that models LLM agent execution as a dependency graph to propagate uncertainty, improving failure detection and confidence estimation in long trajectories.
The paper introduces aDSL, a co-designed domain-specific language and multi-agent system that improve LLM-driven 3D program synthesis through relational operators and iterative feedback, enhancing robustness and controllability in 3D content creation.
StartupBench introduces a benchmark for evaluating general-purpose AI agents on real-world startup workflows, revealing that top models complete only about 30% of tasks due to gaps in complex instruction following and domain-specific expertise.
This paper presents a systematic scaling law study for text-to-image diffusion models, showing they scale predictably but require significantly more data per parameter than language models for optimal training.
ASI-Bench is a new benchmark designed to evaluate AI systems' capabilities in innovative exploration and autonomous scientific execution across 11 scientific domains, revealing current AI's heavy dependence on human guidance.
This paper evaluates indirect prompt injection risks in DeepSeek Harness using AI-Infra-Guard for controlled testing, finding notable attack success rates and recommending security controls.
GS-Voxel introduces a fitting-free framework to convert 3D Gaussian Splatting reconstructions into structured latents, enabling scalable generation of large-scale aerial 3D scenes via flow models and tiled inference.