Tag
A cheaper alternative to Groq for hosting the open-source GPT OSS 120B model in a development environment.
A single viral post from investor Matt Shumer dramatically boosted Groq's business by effectively demonstrating the value of fast inference, a point Groq founder Jonathan Ross had struggled to convey for years.
Groq founder reveals that West Coast VCs rejected funding due to herd behavior, ultimately securing funds from East Coast VCs.
Groq founder Jonathan Ross discusses how West Coast VCs missed investing in Groq due to herd mentality, contrasting with East Coast VCs who do independent analysis. The conversation also covers Groq's $20 billion partnership with NVIDIA.
Announcement of a podcast episode featuring Jonathan Ross, founder of Groq, discussing his $20 billion partnership with NVIDIA and his decade building Groq.
Shawn Presser, a veteran AI researcher with 25 years of programming experience who was employee #2 at Carmack's Keen AI lab and contributed to the LPU chip design at Groq, is publicly asking for a job on X. He says he'll face homelessness if he can't find work and is willing to accept below-market pay. The community has taken notice.
Bash4LLM+ is a lightweight, dependency-free Bash wrapper for LLM APIs, offering secure and auditable interaction with Groq and other providers, with features like dynamic model lists, streaming, and extensible extras.
Groq raises $650M to pivot to its neocloud business after Nvidia's $20B licensing deal and talent poaching, hiring new executives and expanding data centers.
MeetMemory is a tool that gives AI permanent memory for meetings, built on Hindsight and Groq, allowing instant recall across all past conversations without manual search or note-taking.
Groq is raising $650M despite its technology licensing to Nvidia because the corporate entity retained its datacenter operations and inference API, focusing on fast small-model inference.
A progress update on reinventing Groq's LPU, with a redesigned vector execution module to better support overlap operations and self-attention.
Analyzes the overseas free AI inference ecosystem, including services like Groq offering free requests, and points out that Chinese developers have a significant information gap on this.
AI chip startup Groq is reportedly raising $650M from existing investors to grow its inference cloud business, following a $20B technology licensing deal with Nvidia.
Google demonstrated Gemini Flash model achieving 600-1400 tokens per second on TPU 8i, rivaling Groq's inference speeds.