Tag
Enabling PCI-E peer-to-peer (P2P) for consumer Nvidia GPUs with patched drivers and vLLM environment variables yields roughly 25% prefill throughput improvement for free, as demonstrated by benchmarks.
Firebird launched the CIS region's largest AI factory in Armenia, powered by NVIDIA accelerated computing and Dell infrastructure, with plans to deploy over 70,000 NVIDIA GPUs and 300 MW of capacity by 2027. NVIDIA also intends to invest in Firebird as part of a broader 2-gigawatt roadmap across Armenia, Kazakhstan, and other markets.
All panels and presentations from the Local AI Summit at AIE's World's Fair 2026 are now available to watch online, showcasing talks on making local AI the default.
NVIDIA released the NeMo Gym conversational tool-use assets on Hugging Face, including golden policy/tool reference pairs and prompt histories for the pipeline.
An educational tweet explaining the three levels of NVIDIA's software stack (CUDA C++, PTX, SASS) and how CUDA's abstraction creates a moat, while mentioning Luminal's automatic compiler search.
NVIDIA's entire speech stack—ASR, TTS, and codec—is now quantized to GGUF and runs locally on-device via NeMo-Speech.cpp, with new model releases for Magpie-TTS, Nemotron Speech Streaming, and Parakeet.
NVIDIA highlights how its Nemotron open models let teams build specialized, trustworthy AI tailored to their business data and workflows.
NVIDIA releases Nemotron Parse 2.0, a document image parsing model that converts scanned PDFs and images into structured text with layout, bounding boxes, and reading order, adding multilingual OCR improvements and chart-aware parsing.
GeForce NOW adds 26 new games in August, including World of Warships: Legends, alongside QuakeCon demos and new titles across Steam, Epic, and Xbox.
NVIDIA highlights the importance of open world models for physical AI, introducing the NVIDIA Cosmos 3 open model family and Omniverse libraries for simulation and model specialization.
Discussion about DeepSeek's price increases and free tier downgrades pushing users toward local hardware, potentially benefiting NVIDIA hardware sales.
Jensen Huang announced that NVIDIA has officially open-sourced its autonomous driving reasoning model Alpamayo 2 Super. The model can understand complex scenes and think before acting, and is suitable for Robotaxi, trucks, delivery vehicles, etc. It is open for commercial use under the OpenMDW-1.1 license.
An analytical critique of NVIDIA's Vera whitepaper, examining the Olympus core's impressive architecture while arguing that the paper's anti-x86 narrative and benchmark claims are overstated, with independent testing suggesting the hardware is genuinely strong.
NVIDIA shares five lessons from over 5,000 Kagglers who fine-tuned reasoning models using LoRA adapters and synthetic chain-of-thought data in the Nemotron Model Reasoning Challenge, focusing on verifiable data, token budget, and infrastructure.
Elon Musk announced that SpaceX will use Nvidia GPUs exclusively, citing they are the best.
Elon Musk says the Starmind V1 satellite design will be adapted for ground deployment in data centers to improve efficiency. SpaceX is partnering with NVIDIA on the Starmind AI1 satellite compute payload using Rubin GPUs and Vera CPUs.
Anthropic signs a reported $10B deal with AI cloud startup Volta to secure 133 MW of compute capacity over six years, with Bitdeer building a Norway data center powered by Nvidia Vera Rubin systems.
The week-old Open Secure AI Alliance (OSAA), led by Nvidia and now including over 120 companies, is already presenting proposals for AI security and collecting open-source contributions from members like Amazon, Red Hat, and Okta, while notable firms such as OpenAI and Google have not yet joined.
NVIDIA joins the NSF State and Regional AI Infrastructure Hubs program to expand access to AI computing, data, and expertise for research and education across US colleges and universities.
Alpamayo 2 Super, an open reasoning model for autonomous vehicles, is released on Hugging Face under the Linux Foundation's permissive OpenMDW-1.1 license, enabling commercial adaptation and redistribution.