laptop-inference

Tag

Cards List
#laptop-inference

I tried running a 1.56TB MoE model on a 6GB RTX 4050 Laptop, Here’s the result

Reddit r/LocalLLaMA · 4d ago

Testing a 1.56TB Mixture-of-Experts model on a 6GB RTX 4050 laptop, requiring patched memory streaming with NVMe to achieve 0.106 tokens/s decode speed.

0 favorites 0 likes
#laptop-inference

@Gradio: A hackathon called "Build Small" max 32B params. the model fits on a laptop. somehow that pitch got us OpenAI, NVIDIA, …

X AI KOLs Timeline · 2026-05-28 Cached

A hackathon called 'Build Small' with a maximum of 32B parameters, designed to run on a laptop, has attracted sponsors including OpenAI, NVIDIA, OpenBMB, and Cohere, offering over $40k cash, RTX 5080s, and codex credits.

0 favorites 0 likes
← Back to home

Submit Feedback