Tag
Emad Mostaque demonstrated Taalas generating at 14,000 tokens per second, significantly faster than ChatGPT's 50-150 tokens per second, as showcased on The Peter McCormack Show.
An analysis of AI inference hardware, comparing Taalas and Groq's approaches to etching model weights into silicon, and noting recent investments by Nvidia and AMD.
AMD has reached a definitive agreement to acquire Toronto-based AI chip startup Taalas, which specializes in hardwiring AI models onto custom silicon for efficient inference. The deal aims to strengthen AMD's AI portfolio with differentiated inference performance and efficiency.
AMD acquires AI chip startup Taalas, which etches model weights directly into silicon, promising order-of-magnitude inference performance boosts. The deal is seen as AMD's move to challenge Nvidia's dominance in AI hardware.