Tag
NVIDIA is integrating Groq technology into rack-scale products to enhance agentic AI performance by splitting workloads across specialized processors, improving token generation latency and overall responsiveness.
At Startup School 2026, Google's Chief Scientist Jeff Dean recounts the napkin math that led to Google's search index fitting in RAM and the TPU development, discussing inference hardware specialization and how startups can still compete.
Etched, an AI inference hardware startup, exited stealth after raising $800M and securing over $1B in customer contracts. Their first racks ship this summer, claiming state-of-the-art throughput, latency, and power efficiency.