kernel-benchmarks

Tag

Cards List
#kernel-benchmarks

LLM4LLM: Bridging Kernel Benchmarks and Real Deployment via Closed-Loop Agentic Optimization

arXiv cs.AI · 2026-08-25 Cached

LLM4LLM introduces a deployment-aware closed-loop optimization framework to bridge kernel benchmarks and real LLM inference, achieving up to 6.98x speedups on H100 GPUs.

0 favorites 0 likes
← Back to home

Submit Feedback