ada-gpu

Tag

Cards List
#ada-gpu

Deepseek V4 Flash ~105 t/s on two Nvidia 4090d 48G (ada) in vLLM

Reddit r/LocalLLaMA · 6d ago

Technical post detailing how to run DeepSeek V4 Flash on two Nvidia 4090d GPUs using custom Triton kernels and vLLM, achieving ~105 tokens/second with 262k context.

0 favorites 0 likes
← Back to home

Submit Feedback