Tag
Cargo thieves are using violent tactics like PIT maneuvers and hit-and-runs to steal high-value AI hardware shipments, alarming freight security experts who say such brazen thefts are unprecedented.
Nvidia is reportedly testing lower-memory variants of its upcoming Rubin Ultra accelerator, including configurations with 192 GB or 256 GB and a switch from HBM4E to HBM4, due to HBM supply constraints.
ARK Invest shares key points on Tesla's potential reduced reliance on China, covering merger timing, AI costs, market winners, and AI hardware in a video breakdown.
Stoa Markets (YC S26) launches a marketplace for buying and selling GPUs and AI servers, offering verified counterparties, firm quotes, and structured settlement to address fragmented supply and opaque pricing in the AI hardware market.
Recommends the nearly 5-hour interview with Liao Heng, Chief Scientist of Huawei Semiconductor. The content is detailed, showcasing the confidence of China's AI hardware development, and compares with Liang Wenfeng representing the software side of China's AI.
A hands-on review of OpenAI's $230 Codex Micro keyboard, praising its joystick and hardware but noting unfinished software, bugs, and a freeze during a Codex run.
Tibo hints that OpenAI's first device could arrive in 2–3 months and be fully agentic, suggesting a major evolution in AI interaction beyond laptops.
China's DFSX claims its TY64 SuperNode, built with 14nm DF2000 chips using a 3.5D Infinity Chiplet layout, offers 960TB/s memory bandwidth—2x that of NVIDIA's GB200 NVL72—though with lower raw compute (64 PFLOPS BF16 vs 360 PFLOPS).
Intel is advancing chip packaging technologies like Foveros and EMIB to enable multi-chip packages for AI workloads, scaling beyond traditional reticle limits.
Ahmad announces a Local AI Hardware Arena using ODS to benchmark LLMs on hardware like RTX PRO 6000, DGX Spark, Strix Halo, M5 MacBook Pro, and ChatGPT, inviting community input for future comparisons.
NVIDIA showcases the Jetson platform for edge AI and robotics, highlighting the compact yet powerful Jetson Orin Nano Super developer kit that enables building AI agents and robots anywhere.
AMD has achieved a 46% revenue share in x86 server CPUs, up from nearly zero in 2017, driven by its new Venice processors that offer 2.2x performance per core over Nvidia's comparable chip. The shift highlights a growing preference for modular CPU+GPU setups in AI agentic workloads over Nvidia's vertically integrated racks.
OrangePi releases AI Studio Pro, a single-board computer optimized for running the Qwen3.5-122B-A10B large language model.
Etched announces breakthroughs in low voltage inference and cluster scale memory for their AI inference hardware, with first racks shipping this summer and $1B in customer contracts.
Researchers at Fudan University have developed a two-dimensional flash memory chip that stores data using a single electron, achieving a tenfold voltage improvement and potentially enabling efficient AI on mobile devices.
IneffableLabs received the first Vera Rubin NVL72 clusters from Google Cloud and NVIDIA, marking a leap in AI hardware.
NVIDIA announces Vera Rubin platform, delivering 10x more throughput per megawatt than Blackwell and claiming lowest token cost for AI factories through extreme co-design across seven chips and five rack trays.
Advanced materials are critical for enabling next-generation AI by addressing performance challenges in semiconductor manufacturing and data center infrastructure. The article highlights how materials innovation in polymers, fluids, and thermal management supports the physical demands of increasingly powerful AI systems.
Z.ai completed a 1-gigawatt data center powered entirely by Chinese-made chips, expanding computing infrastructure for training its advanced GLM models.
Chinese factories are repurposing consumer RTX 5090 GPUs into 128GB server cards as a workaround to export controls on dedicated AI chips, creating a grey-market competitor to official enterprise offerings.