gpu-programming

Tag

Cards List
#gpu-programming

Teaching GPU programming in p5.js: now with compute shaders

Lobsters Hottest ↗ · 2d ago Cached

The article discusses the integration of compute shaders into p5.js to simplify teaching GPU programming through scaffolded learning, addressing educational challenges in computer graphics.

0 favorites 0 likes
#gpu-programming

Introducing CUDA Rust: Two Tracks for Writing GPU Kernels

Lobsters Hottest ↗ · 2026-09-09 Cached

NVIDIA introduces CUDA Rust with two tracks (SIMT and Tile) for writing GPU kernels natively in Rust, enabling performance and developer experience improvements in AI systems.

0 favorites 0 likes
#gpu-programming

@levidiamode: Day 248/365 of GPU Programming The Cerebras CTO has some other great talks online. For example, one from two years ago …

X AI KOLs Timeline ↗ · 2026-09-08 Cached

A social media post highlighting talks by Cerebras CTO on GPU programming and AI hardware architecture, noting their low view counts despite being informative.

0 favorites 0 likes
#gpu-programming

What happens when a GPU writes memory

Hacker News Top ↗ · 2026-09-08 Cached

This article details the process of GPU memory writes, tracing the STG.E instruction through various hardware components like the load/store unit, coalescer, and L1 cache on an RTX 4090.

0 favorites 0 likes
#gpu-programming

Declarative WebGPU with S-expressions

Lobsters Hottest ↗ · 2026-08-23 Cached

Pngine is a declarative format and runtime for WebGPU that uses S-expressions to simplify declaring and validating WebGPU configurations, enabling cross-platform sharing and export to formats like HTML or PNG with embedded runtime.

0 favorites 0 likes
#gpu-programming

Mojo🔥 is now open source

Simon Willison's Blog ↗ · 2026-08-18 Cached

Mojo programming language has been open-sourced under an Apache 2 license, fulfilling a long-standing promise with the release of its compiler and toolchain.

0 favorites 0 likes
#gpu-programming

CUDA Shared Memory Swizzling

Hacker News Top ↗ · 2026-08-13 Cached

The article explains CUDA shared memory swizzling techniques to optimize GPU memory access patterns, with code examples demonstrating performance improvements.

0 favorites 0 likes
#gpu-programming

@vivekgalatage: GPU Programming Fundamentals https://youtu.be/Cl2B_hmg4gA William Brandon, a performance engineer at Anthropic, outline…

X AI KOLs Timeline ↗ · 2026-08-04 Cached

A summary of William Brandon's (performance engineer at Anthropic) GPU programming fundamentals lecture, emphasizing that understanding the streaming multiprocessor (SM) structure of GPU hardware is key to predicting performance, rather than starting solely from the software abstraction of thread blocks/threads.

0 favorites 0 likes
#gpu-programming

@v0xium: If you are looking for an article going in details regarding basics of CUDA, please spend an hour reading this. Link to…

X AI KOLs Timeline ↗ · 2026-07-20 Cached

An updated beginner-friendly tutorial on CUDA programming, covering how to write a simple kernel to add arrays on the GPU.

0 favorites 0 likes
#gpu-programming

@TheAhmadOsman: You wanna learn how do these AI kernels work? Start here

X AI KOLs Following ↗ · 2026-07-15

Tweet by @TheAhmadOsman pointing to a resource for learning how AI kernels work.

0 favorites 0 likes
#gpu-programming

@Modular: Week 2 of Mojo 101 goes live on Thursday. This week, we're deep diving on value ownership and metaprogramming on our li…

X AI KOLs Following ↗ · 2026-07-14 Cached

Modular announces Week 2 of Mojo 101, a free four-week live course teaching Mojo programming from the engineers who built it, covering value ownership and metaprogramming in the upcoming session.

0 favorites 0 likes
#gpu-programming

@kalyan_kpl: What happens when you run a CUDA Kernel A CUDA kernel is a specialized code written to execute parallel computations on…

X AI KOLs Timeline ↗ · 2026-07-11 Cached

This article provides a detailed walkthrough of what happens when a CUDA kernel is compiled and executed on an NVIDIA GPU, covering compilation to PTX and SASS, and the underlying hardware interaction.

0 favorites 0 likes
#gpu-programming

@Modular: Mojo 101: From Syntax to GPU Programming [Session 1]

X AI KOLs Following ↗ · 2026-07-09 Cached

Modular is hosting a session called Mojo 101 covering syntax to GPU programming, aimed at teaching the Mojo language.

0 favorites 0 likes
#gpu-programming

@CUDAHandbook: BREAKING: The CUDA Handbook text is now available on the website, https://cudahandbook.com! Svbstack article in the fir…

X AI KOLs Timeline ↗ · 2026-07-08 Cached

The CUDA Handbook, a comprehensive guide to GPU programming with CUDA, is now available online for free reading with ads or ad-free via membership. The book covers architecture, APIs, and algorithms, with open-source code.

0 favorites 0 likes
#gpu-programming

@PatrickToulme: This exercise makes me believe the future of DSLs and compilers is very much agentic. Programming languages and DSLs bu…

X AI KOLs Timeline ↗ · 2026-07-07 Cached

Claude Fable used the pyptx DSL to write a FlashAttention forward kernel for NVIDIA B200 that achieves near-parity performance with the hand-tuned CUTLASS kernel, demonstrating the potential for AI agents in compiler and DSL design.

0 favorites 0 likes
#gpu-programming

@ManningBooks: Deep learning frameworks make building models easier, but they also make it easy to treat GPUs like a black box. CUDA f…

X AI KOLs Timeline ↗ · 2026-07-07 Cached

Manning Books promotes Elliot Arledge's 'CUDA for Deep Learning,' a book that teaches GPU-level programming with CUDA to optimize deep learning model performance, available at a discount until Sunday.

0 favorites 0 likes
#gpu-programming

@reprompting: https://x.com/reprompting/status/2074133435401064486

X AI KOLs Timeline ↗ · 2026-07-06 Cached

A detailed thread summarizing the book 'Programming Massively Parallel Processors', focusing on CUDA and GPU programming concepts, optimization techniques, and parallel patterns.

0 favorites 0 likes
#gpu-programming

@levidiamode: 183/365 of GPU Programming This 4.5 hour lesson on CUDA + ThunderKittens by @bfspector (TK co-author, Stanford PhD stud…

X AI KOLs Timeline ↗ · 2026-07-05 Cached

A highly recommended 4.5-hour GPU programming lesson on CUDA and ThunderKittens by Ben Spector, offering an in-depth, behind-the-scenes look at kernel optimization.

0 favorites 0 likes
#gpu-programming

@h100envy: CMU PhD who built the kernels NVIDIA now ships in TensorRT-LLM explained fast attention in 68 minutes - better than $12…

X AI KOLs Timeline ↗ · 2026-07-02 Cached

A CMU PhD who developed the kernels now used by NVIDIA in TensorRT-LLM explains fast attention, covering fused CUDA kernels, FlashInfer, Triton, and paged-KV attention, enabling more tokens per second on the same GPU.

0 favorites 0 likes
#gpu-programming

@h100envy: PyTorch core engineer at Meta turned CUDA kernel writing into a sport in 13 minutes - better than $1500 GPU programming…

X AI KOLs Timeline ↗ · 2026-06-30 Cached

A PyTorch core engineer at Meta demonstrated a fast CUDA kernel optimization loop that outperforms expensive bootcamps, with the winning code merged into PyTorch via the KernelBot competition.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback