pull-request

Tag

Cards List
#pull-request

CUDA: add fast walsh-hadamard transform by am17an · Pull Request #23615 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-05-25 Cached

This pull request adds a fast Walsh-Hadamard transform implementation for CUDA in llama.cpp, a popular open-source LLM inference engine. The optimization enhances performance for certain computational operations on NVIDIA GPUs.

0 favorites 0 likes
#pull-request

For everyone that uses OpenCode / Pi - Heres your promptprocessing fix!

Reddit r/LocalLLaMA · 2026-05-21

A pull request for llama.cpp fixes the constant prompt processing issue that occurs when using OpenCode or Pi with the library.

0 favorites 0 likes
#pull-request

llama: avoid copying logits during prompt decode in MTP by am17an · Pull Request #23198 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-05-17 Cached

This pull request optimizes llama.cpp by avoiding unnecessary copying of logits during prompt decode in multi-token prediction, improving inference performance.

0 favorites 0 likes
#pull-request

MTP support merged into llama.cpp

Reddit r/LocalLLaMA · 2026-05-16

The pull request adding MTP (Multi-Token Prediction) support to llama.cpp has been merged into the master branch.

0 favorites 0 likes
#pull-request

MTP PR Merged!!!

Reddit r/LocalLLaMA · 2026-05-16

A pull request for MTP (likely a model training pipeline or similar) related to LLaMA models has been merged, marking a milestone.

0 favorites 0 likes
#pull-request

llama + spec: MTP Support by am17an · Pull Request #22673 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-05-16 Cached

Pull request adding Multi-Token Prediction (MTP) support to llama.cpp, enabling speculative decoding for faster inference.

0 favorites 0 likes
#pull-request

@LangChain: This AI watches its own codebase, flags missing monitors, and opens PRs to fix bugs it finds. @Shevchenkoaalex on @TryR…

X AI KOLs Following · 2026-05-11 Cached

An AI agent built with LangChain continuously monitors its own codebase, flags missing monitors, and automatically opens PRs to fix bugs it finds, as described by Alex Shevchenko from Ramp.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback