pull-request

Tag

Cards List
#pull-request

model: add NVIDIA Nemotron-3-Puzzle-75B-A9B (NemotronHPuzzle) support by YanissAmz · Pull Request #25444 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-09-03 Cached

A pull request adds support for the NVIDIA Nemotron-3-Puzzle-75B-A9B model in the llama.cpp inference tool.

0 favorites 0 likes
#pull-request

qwen4exp fixes in llama.cpp

Reddit r/LocalLLaMA · 2026-09-01

This article reports recent bug fixes and updates in llama.cpp for the Qwen Flash Next model, advising users to update their builds frequently.

0 favorites 0 likes
#pull-request

Support for DFlash2 in llama.cpp has been merged! - spec : add DFlash2 support (local convolution + candidate selector) by SubSir · Pull Request #27342 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-08-27 Cached

Support for DFlash2 has been merged into llama.cpp via pull request #27342, adding local convolution and candidate selector features to the LLM inference tool.

0 favorites 0 likes
#pull-request

@markfenner: Devin can turn “tested” into a video receipt. After it creates a PR, click Test the app. @DevinAI reads the diff, creat…

X AI KOLs Timeline · 2026-08-21 Cached

Devin, an AI tool, now provides video receipts for app testing by generating test plans, executing user flows, and delivering annotated recordings to visually verify work.

0 favorites 0 likes
#pull-request

@shl: Code is mostly solved, but design isn’t. Introducing Tastelint! An agent that runs on every PR and surfaces design feed…

X AI KOLs Following · 2026-08-18 Cached

Introducing Tastelint, an AI agent that runs on every pull request to provide design feedback, aiming to bridge the automation gap in product design compared to code development.

0 favorites 0 likes
#pull-request

llama.cpp adaptive MTP PR#27210

Reddit r/LocalLLaMA · 2026-08-17 Cached

A pull request (PR#27210) for adaptive MTP has been submitted to llama.cpp, a C/C++ implementation for LLM inference with minimal setup and high performance.

0 favorites 0 likes
#pull-request

@ericzakariasson: i turned this into an npm package you drop a widget in your app and when someone files feedback, a cloud agent opens a …

X AI KOLs Timeline · 2026-08-15 Cached

An npm package that automates feedback handling in apps by opening pull requests through a cloud agent.

0 favorites 0 likes
#pull-request

We let agents run tickets to PR unattended. The thing that made it work wasn't a better prompt, it was deleting a tool.

Reddit r/AI_Agents · 2026-08-14

The article describes a system enabling AI agents to autonomously handle software tickets to pull requests by removing the 'ask-the-user' tool and implementing an assumption budget, reducing interruptions and improving efficiency in a production codebase.

0 favorites 0 likes
#pull-request

spec: add DSpark speculative decoding by wjinxu · Pull Request #25173 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-07-28 Cached

Adds DSpark speculative decoding support to llama.cpp via pull request, enhancing inference performance.

0 favorites 0 likes
#pull-request

Minimax M3 support with MSA has been merged into llama.cpp

Reddit r/LocalLLaMA · 2026-07-26 Cached

Minimax M3 support with MSA has been merged into llama.cpp, enabling inference for the Minimax M3 model using the MSA architecture.

0 favorites 0 likes
#pull-request

@GergelyOrosz: OK this I loved: my backend (for my admin portal) had an error popping up, and Sentry was bugging me about it. Sentry n…

X AI KOLs Following · 2026-07-22 Cached

Sentry released an AI agent called 'Seer' that analyzes backend errors, determines root cause, drafts a fix, and automatically opens a pull request for review.

0 favorites 0 likes
#pull-request

Add support for Laguna XS.2 & M.1 by joerowell · Pull Request #25165 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-07-22 Cached

This pull request adds support for Laguna XS.2 & M.1 hardware in llama.cpp, expanding compatibility.

0 favorites 0 likes
#pull-request

There's a new PR for llamacpp claiming to boost prompt processing with rocm by around 15%, also fixes a bug which makes Q2_K 28x faster

Reddit r/LocalLLaMA · 2026-07-21 Cached

A new PR for llama.cpp boosts prompt processing on ROCm by ~15% and fixes a bug making Q2_K quantization 28x faster.

0 favorites 0 likes
#pull-request

model: add Hy3 (hy_v3) support with MTP speculative decoding by satindergrewal · Pull Request #25395 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-07-14 Cached

This pull request adds support for the Hy3 (hy_v3) model with MTP speculative decoding to llama.cpp, enabling efficient inference for this architecture.

0 favorites 0 likes
#pull-request

I handed a live product to an AI system and let it improve itself, here's what actually happened

Reddit r/ArtificialInteligence · 2026-07-13

The author built an autonomous AI system that runs a live product, generating work, quality-gating it, opening pull requests, and self-improving based on analytics. The main challenge was making the system trustworthy rather than making the model smarter.

0 favorites 0 likes
#pull-request

I built an agent that improves its own pipeline, not just one that completes tasks

Reddit r/AI_Agents · 2026-07-13

The author built an autonomous agent that not only completes tasks but also improves its own code and product by observing results, making changes via pull requests, and verifying each change with a ledger. The key insight is that a rigorous verify step—concluding confirmed, rejected, or inconclusive—is essential for the system to truly learn.

0 favorites 0 likes
#pull-request

Initial ET backend by marty1885 · Pull Request #24179 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-07-10 Cached

A pull request adds an initial ET backend to llama.cpp, expanding hardware support for LLM inference.

0 favorites 0 likes
#pull-request

@rajistics: OpenHands Enterprise now works with Azure DevOps. So now if you comment on a work item or a PR, and OpenHands picks it …

X AI KOLs Following · 2026-07-07 Cached

OpenHands Enterprise now integrates with Azure DevOps, enabling users to comment on work items or PRs and have OpenHands automatically perform the work and open a pull request in Azure Repos.

0 favorites 0 likes
#pull-request

Gemini Code Assist will be shut down on July 17

Hacker News Top · 2026-07-03 Cached

Google announces the shutdown of the consumer version of Gemini Code Assist on GitHub on July 17, while the enterprise version remains available.

0 favorites 0 likes
#pull-request

DeepSeek V4 by am17an · Pull Request #24162 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-06-29 Cached

Pull request adding support for DeepSeek V4 model in llama.cpp, enabling inference of this model on various hardware.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback