Tag
PyTorch Foundation Ambassador Abdulsalam Bande will present a poster on ERRC (Entropy-Reinvested Residual Correction) at PyTorch Conference North America 2026 in San Jose, compressing inter-GPU communication to boost multi-GPU LLM inference efficiency, alongside the broader conference program covering the open-source AI stack.
At PyTorch Conference North America 2026, AMD's Liz Li and Jiahui Cao will present "Extending TorchInductor with FlyDSL," an MLIR-native GPU kernel backend for high-performance GEMMs that integrates with TorchInductor's compilation and autotuning pipeline, with benchmark comparisons against Triton on AMD Instinct GPUs.
PyTorch blog publishes a Ray-focused guide to PyTorch Conference North America 2026 in San Jose, highlighting keynotes and sessions on co-evolving Ray and Kubernetes, elastic training stacks, and scalable RL from speakers at Anyscale, Google, LinkedIn, Pinterest, and Uber.
This post promotes vLLM, a high-throughput inference engine for LLMs, and details its sessions and activities at PyTorchCon North America.
The article promotes a talk at the PyTorch Conference North America focused on making enterprise agentic inference production-ready using PyTorch and vLLM, covering ecosystem updates and registration details.
The article promotes a presentation on Elastic Expert Parallelism in vLLM at the PyTorch Conference North America 2026, discussing how to dynamically add or remove GPUs in Mixture-of-Experts deployments with minimal downtime.
An announcement for a talk at PyTorch Conference North America where Ziming will discuss using OpGuard for bitwise debugging in LLM training.
At PyTorch Conference North America, Maajid Khan from Fujitsu Research India will present a talk on optimizing Mixture of Experts LLM inference on ARM CPUs using vLLM and OpenVINO.
The article promotes a poster presentation at PyTorch Conference North America on enabling the open-source vime RL post-training framework on AMD Instinct GPUs using ROCm, and provides registration details for the conference.
The PyTorch Conference North America 2026 will feature a session by Red Hat engineers on Cross-Repository CI Relay to streamline CI integration for downstream repositories. Registration details and event highlights are provided.
At PyTorch Conference North America 2026, Dhritiman Das will present a torch-native retrieval engine that uses PyTorch as the primary runtime for retrieval, filtering, and ranking, designed for scalable use cases like feed and search at LinkedIn.
The PyTorch Conference North America in 2026 will feature a talk by Chris Lattner on a unified software stack to address fragmentation in the AI ecosystem, with registration details provided.
A keynote at PyTorch Conference North America will showcase an open software platform for heterogeneous compute powered by Mojo and MAX, addressing ecosystem challenges in deploying AI models across diverse hardware.
This article promotes the PyTorch Conference North America, highlighting a session by Witold Czubala from UBS on using PyTorch-based transformer models for time-to-event prediction in wealth management from longitudinal email data.
Thomas Cottenier from Arm will present at PyTorch Conference North America on how AI agents can synthesize customized runtimes using PyTorch components for optimized inference on edge hardware.
At PyTorch Conference North America, Ricardo Noriega de Soto and Alexander Brooks will demonstrate how extending vLLM's prefix caching mechanism to multistage pipelines boosts inference speeds while reducing GPU memory overhead, providing practical strategies for optimizing complex AI workloads.
Lu Fang and Ilina Mitra from Meta will present a deep dive on building high-performance recommendation inference systems at PyTorch Conference 2026, detailing end-to-end production workflows and advanced optimizations for large-scale platforms.
The PyTorch Conference North America will be held in San Jose, featuring sessions on agentic search, hardware-guided workflows, and AI performance optimization with speakers from major tech companies.
Keynote speaker announced for PyTorch Conference North America 2026, featuring Red Hat's SVP & AI CTO, with full schedule and registration details provided.
Promotes a talk at PyTorchCon North America where AMD's Primus Tuning Agent is demonstrated to efficiently train a 671-billion-parameter model across 1000 GPUs by predicting optimal configurations, saving compute time.