kog

Tag

Cards List
#kog

@HotAisle: This is awesome. I wonder who's MI300x they used... ;-)

X AI KOLs Following · 2026-05-29 Cached

Kog announces real-time LLM inference achieving 3000+ output tokens per second per request on standard datacenter GPUs, bringing high-speed inference previously limited to custom silicon to production hardware.

0 favorites 0 likes
← Back to home

Submit Feedback