machine-learning-inference

Tag

Cards List
#machine-learning-inference

ggml-cuda: hip: add missing AMD GCN MMQ config by thelittlefireman · Pull Request #27841 · ggml-org/llama.cpp - PP improvements for RDNA2(MI50, MI60)

Reddit r/LocalLLaMA · 2d ago Cached

This pull request adds missing AMD GCN MMQ configuration to ggml-cuda for HIP, enhancing prefill performance for RDNA2 GPUs such as MI50 and MI60 in the llama.cpp inference library.

0 favorites 0 likes
← Back to home

Submit Feedback