ai-accelerators

Tag

Cards List
#ai-accelerators

Memory is now 63% of what an AI accelerator costs to build, up from 52% in early 2024

Reddit r/ArtificialInteligence ↗ · 2026-09-13

Memory now accounts for 63% of AI accelerator costs, up from 52% in early 2024, shifting industry priorities toward data movement optimization over process shrinks.

0 favorites 0 likes
#ai-accelerators

Samsung Debuts zHBM Prototype, Stacking Memory Directly on AI Accelerators

Hacker News Top ↗ · 2026-09-07 Cached

Samsung unveiled zHBM and zNAND-O memory technologies at the FMS 2026 conference, designed to stack memory directly on AI accelerators for up to 8x higher performance and improved power efficiency in AI systems.

0 favorites 0 likes
#ai-accelerators

@Modular: Portability across CPUs, GPUs, and accelerators sounds simple until you've tried to build it. At #ModCon2026, @clattner…

X AI KOLs Timeline ↗ · 2026-08-24 Cached

Modular CEO Chris Lattner and Qualcomm CEO Cristiano Amon discussed portability across CPUs, GPUs, and accelerators at ModCon2026, emphasizing their collaboration to improve software for heterogeneous compute following Qualcomm's acquisition of Modular.

0 favorites 0 likes
#ai-accelerators

Mojo is now open source

Lobsters Hottest ↗ · 2026-08-18 Cached

Mojo, a programming language designed for AI accelerators and GPUs, is now fully open source under the Apache 2.0 license, with its compiler and toolchain available on GitHub for development and contribution.

0 favorites 0 likes
#ai-accelerators

Show HN: Fixing optical computing jitter via fluid dynamics in GPU registers

Hacker News Top ↗ · 2026-08-09 Cached

This repository presents a proof of concept for a hardware-native optical timing-frozen control plane engine that uses fluid dynamics modeling in GPU registers to reduce jitter in optical data centers for distributed AI architectures.

0 favorites 0 likes
#ai-accelerators

Chinese companies are ditching Nvidia’s advanced accelerators for domestic AI suppliers

Reddit r/ArtificialInteligence ↗ · 2026-07-08 Cached

Chinese companies are increasingly allocating AI accelerator budgets to domestic suppliers like Huawei and Hygon, reducing reliance on Nvidia amid US-China tensions. A Bloomberg survey shows 46% of spending will go to domestic products in the next year, up from 30%.

0 favorites 0 likes
#ai-accelerators

@yishan: Meta really fumbled this guy.

X AI KOLs Timeline ↗ · 2026-07-07 Cached

John Carmack comments on memory cost and capacity issues for AI accelerators, noting that model inference can have deterministic memory access patterns, contrasting with game rendering.

0 favorites 0 likes
#ai-accelerators

@QuixiAI: https://x.com/QuixiAI/status/2073936537213915611

X AI KOLs Following ↗ · 2026-07-06 Cached

QuixiAI released QuixiCore, a family of native high-performance AI kernel libraries for modern accelerators, with standalone implementations for CUDA, Metal, ROCm, XPU, and Gaudi backends, all sharing a common contract but no shared code.

0 favorites 0 likes
#ai-accelerators

7 Chinese companies are already shipping H100/H200-class AI chips, most IPO'd in the last 6 months. I mapped all of them.

Reddit r/LocalLLaMA ↗ · 2026-06-23

At least seven Chinese companies are shipping H100/H200-class AI accelerators, most having recently IPO'd, with several founded by former NVIDIA/AMD architects. Huawei's Ascend 950 targets H200-class performance, and China's domestic market share is rising as NVIDIA's declines.

0 favorites 0 likes
#ai-accelerators

Buying AI accelerators/GPUs in China...

Reddit r/LocalLLaMA ↗ · 2026-06-15

A user asks about buying Chinese AI accelerators/GPUs for inference, specifically looking for Huawei alternatives to Nvidia, with support for vLLM or Llama.cpp.

0 favorites 0 likes
#ai-accelerators

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators

arXiv cs.LG ↗ · 2026-06-03 Cached

KForge is a cross-platform framework that uses two collaborating LLM-based agents to automatically generate and optimize high-performance compute kernels for diverse AI accelerators, achieving significant speedups on NVIDIA B200 and Intel Arc B580 hardware.

0 favorites 0 likes
#ai-accelerators

TRAM: Training Approximate Multiplier Structures for Low-Power AI Accelerators

arXiv cs.LG ↗ · 2026-05-12 Cached

This paper introduces TRAM, a method that jointly optimizes approximate multiplier structures and AI model parameters to reduce power consumption in AI accelerators while maintaining accuracy.

0 favorites 0 likes
#ai-accelerators

AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization

Hugging Face Daily Papers ↗ · 2026-04-15 Cached

AccelOpt is a self-improving LLM agentic system that autonomously optimizes AI accelerator kernels through iterative generation and optimization memory, achieving 49-61% peak throughput improvements on AWS Trainium while being 26x cheaper than Claude Sonnet 4.

0 favorites 0 likes
← Back to home

Submit Feedback