@PyTorch: How do we get more useful work—not just more tokens—from every AI dollar? This Wednesday at 11:50 AM at @AMD #Advancing…
Summary
PyTorch Foundation CTO Matt White will speak at AMD's AdvancingAI event about optimizing AI inference economics using open-source tools like vLLM and SGLang, advocating for right-sized models and intelligent routing to improve dollar per intelligence.
View Cached Full Text
Cached at: 07/21/26, 06:49 PM
How do we get more useful work—not just more tokens—from every AI dollar? This Wednesday at 11:50 AM at @AMD #AdvancingAI, PyTorch Foundation CTO @matthew_d_white will explore how open source, @vllm_project and @radixark’s @sgl_project are reshaping inference economics.
In “Accelerating Open Source AI: Just Enough Intelligence,” Matt will outline a practical alternative to using frontier models for every task: → Decompose workflows → Select right-sized models → Route and cache intelligently → Escalate only when necessary → Measure cost per successfully completed task
The goal is better Dollar per Intelligence: more correct, business-relevant outcomes from every dollar spent, with failures, retries, and review included in the calculation.
Matt opens a series of talks in the AI Training & Inference track ahead of @simon_mo_, co-founder and CEO of @inferact and lead contributor to vLLM, and @ying11231, co-founder and CEO of RadixArk and co-creator of SGLang.
Moscone West, San Francisco Wednesday, July 22 11:50 AM PDT
Explore the track:
Similar Articles
@AMD: From bring-up to tuning, AMD and @OpenAI engineers are sharing insights to push performance further. Go behind the coll…
AMD and OpenAI engineers collaborate to share insights on performance optimization, featuring a behind-the-scenes look with OpenAI's VP of Compute Strategy, Sachin Katti.
@rohanpaul_ai: Sam Altman on how enormous inference demand will finance OpenAI's frontier training without requiring high margins. “We…
Sam Altman explains how massive inference demand will finance OpenAI's frontier model training without requiring high margins, and predicts intelligence becoming fungible with advantage shifting to the largest cheapest compute fleets.
AMD AI ENGAGE
The article discusses the AMD AI Engage Program, a community initiative for AI developers offering prizes, credits, and networking opportunities for building LLM apps and GenAI workflows.
@PyTorch: PyTorch Foundation is a Gold Sponsor of Agentic AI Summit 2026. Matt White, CTO of PyTorch Foundation, will lead “The O…
PyTorch Foundation is a Gold Sponsor of Agentic AI Summit 2026, a two-day event hosted by Berkeley RDI. Matt White, CTO of PyTorch Foundation, will lead a workshop on building AI systems with open source and composability.
@latkins: Little late notice but I’ll be speaking here in 30 minutes. An updated version of my Trinity Large talk with some tease…
Alatkins announces a last-minute talk at AI4 Conference in Vegas, covering an updated version of the Trinity Large talk with teasers about training a 400B MoE model to 17T tokens without loss spikes.