ai-accelerators

Tag

Cards List
#ai-accelerators

Chinese companies are ditching Nvidia’s advanced accelerators for domestic AI suppliers

Reddit r/ArtificialInteligence · 2026-07-08 Cached

Chinese companies are increasingly allocating AI accelerator budgets to domestic suppliers like Huawei and Hygon, reducing reliance on Nvidia amid US-China tensions. A Bloomberg survey shows 46% of spending will go to domestic products in the next year, up from 30%.

0 favorites 0 likes
#ai-accelerators

@yishan: Meta really fumbled this guy.

X AI KOLs Timeline · 2026-07-07 Cached

John Carmack comments on memory cost and capacity issues for AI accelerators, noting that model inference can have deterministic memory access patterns, contrasting with game rendering.

0 favorites 0 likes
#ai-accelerators

@QuixiAI: https://x.com/QuixiAI/status/2073936537213915611

X AI KOLs Following · 2026-07-06 Cached

QuixiAI released QuixiCore, a family of native high-performance AI kernel libraries for modern accelerators, with standalone implementations for CUDA, Metal, ROCm, XPU, and Gaudi backends, all sharing a common contract but no shared code.

0 favorites 0 likes
#ai-accelerators

7 Chinese companies are already shipping H100/H200-class AI chips, most IPO'd in the last 6 months. I mapped all of them.

Reddit r/LocalLLaMA · 2026-06-23

At least seven Chinese companies are shipping H100/H200-class AI accelerators, most having recently IPO'd, with several founded by former NVIDIA/AMD architects. Huawei's Ascend 950 targets H200-class performance, and China's domestic market share is rising as NVIDIA's declines.

0 favorites 0 likes
#ai-accelerators

Buying AI accelerators/GPUs in China...

Reddit r/LocalLLaMA · 2026-06-15

A user asks about buying Chinese AI accelerators/GPUs for inference, specifically looking for Huawei alternatives to Nvidia, with support for vLLM or Llama.cpp.

0 favorites 0 likes
#ai-accelerators

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators

arXiv cs.LG · 2026-06-03 Cached

KForge is a cross-platform framework that uses two collaborating LLM-based agents to automatically generate and optimize high-performance compute kernels for diverse AI accelerators, achieving significant speedups on NVIDIA B200 and Intel Arc B580 hardware.

0 favorites 0 likes
#ai-accelerators

TRAM: Training Approximate Multiplier Structures for Low-Power AI Accelerators

arXiv cs.LG · 2026-05-12 Cached

This paper introduces TRAM, a method that jointly optimizes approximate multiplier structures and AI model parameters to reduce power consumption in AI accelerators while maintaining accuracy.

0 favorites 0 likes
#ai-accelerators

AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization

Hugging Face Daily Papers · 2026-04-15 Cached

AccelOpt is a self-improving LLM agentic system that autonomously optimizes AI accelerator kernels through iterative generation and optimization memory, achieving 49-61% peak throughput improvements on AWS Trainium while being 26x cheaper than Claude Sonnet 4.

0 favorites 0 likes
← Back to home

Submit Feedback