code-search

Tag

Cards List
#code-search

PatchHolmes: Agentic Patch Retrieval via Listwise Selection

Hugging Face Daily Papers ↗ · 3d ago Cached

PatchHolmes, accepted at AACL-IJCNLP 2026, uses an agentic listwise selection approach to pair software vulnerabilities with their fixing commits, achieving 59.95% Recall@1 versus 34.61% for pointwise baselines by having the agent inspect the full candidate list and reading ~96k tokens per query on a frozen open-weight model.

0 favorites 0 likes
#code-search

@dzhng: jevgrep 0.5 released! Optimized for efficiency - it's jev cost is now 59% lower Ironically, the per task cost stayed th…

X AI KOLs Timeline ↗ · 4d ago Cached

jevgrep 0.5 is released with efficiency optimizations, reducing cost by 59% by returning less data while maintaining the same intelligence for AI-assisted code search in repositories.

0 favorites 0 likes
#code-search

I ACCIDENTALLY MADE A LOCAL RETRIEVAL TOOL WHICH I THINK IS WORKING! NEED HELP FOR IDEAS TO TEST IT

Reddit r/AI_Agents ↗ · 2026-09-04

The author accidentally developed a local retrieval tool for debugging in agentic AI software that efficiently narrows down code functions to find bugs, and is seeking more testing ideas to validate its effectiveness.

0 favorites 0 likes
#code-search

Synthetic Semantic Supervision for Contrastive Code Representation Learning in Small Transformers: An Empirical Study

arXiv cs.AI ↗ · 2026-09-04 Cached

This paper empirically studies contrastive pretraining with synthetic semantic supervision for code embeddings in small transformers, showing significant gains over baselines and competitiveness with larger models.

0 favorites 0 likes
#code-search

@ZvecAI: 🚀 zg (zvec-grep) is now open source! We built a local search tool for people and AI agents— indexing and search run on…

X AI KOLs Timeline ↗ · 2026-09-02 Cached

ZvecAI has open-sourced zg (zvec-grep), a local-first search tool that integrates semantic search, BM25, and ripgrep for efficient indexing and retrieval on-device, designed for both humans and AI agents.

0 favorites 0 likes
#code-search

@Robro612: Check out this low friction way to experience the speed and precision of @mixedbreadai agentic search for yourself!

X AI KOLs Following ↗ · 2026-08-21 Cached

Introducing Repofetch, a developer tool that allows interaction with GitHub repositories via agentic search powered by mixedbreadai's toast-1 model for efficient code queries.

0 favorites 0 likes
#code-search

Inventory

Product Hunt ↗ · 2026-08-02

Inventory is a product that lets you search across every AI Agent and IDE conversation, making it easier to find past interactions and code discussions.

0 favorites 0 likes
#code-search

Don't stop early: Case-folding source code at memory speed

Hacker News Top ↗ · 2026-07-31 Cached

GitHub engineering describes how they optimized case-folding for their code search engine by removing early-exit branches, achieving memory-speed ASCII folding, and open-sourcing the result as a Rust crate called casefold.

0 favorites 0 likes
#code-search

DenseOn with the LateOn: Fully Open Dense and Late-Interaction Models for Multilingual, Long-Context, and Code Search

arXiv cs.CL ↗ · 2026-07-30 Cached

This paper presents fully open DenseOn and LateOn retrieval models, trained on curated English data and extended to multilingual settings via translate-train, achieving state-of-the-art BEIR results for their parameter size.

0 favorites 0 likes
#code-search

How AST-grep Rewrote Tree-sitter in Rust and Made It 30% Faster

Hacker News Top ↗ · 2026-07-26 Cached

ast-grep rewrote Tree-sitter's C core in Rust, achieving up to 30% faster parsing and 22% faster end-to-end performance in ast-grep, at the cost of slightly higher memory usage.

0 favorites 0 likes
#code-search

Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable

arXiv cs.AI ↗ · 2026-07-16 Cached

The Harness Handbook is a behavior-centric representation synthesized from agent harness codebases using static program analysis and LLM assistance, helping developers and coding agents locate code implementing specific behaviors. It introduces Behavior-Guided Progressive Disclosure (BGPD) to guide agents from high-level descriptions to relevant implementation details, improving localization accuracy and edit-plan quality.

0 favorites 0 likes
#code-search

Show HN: Capn-hook for coding agents – don't grep the same mystery twice

Hacker News Top ↗ · 2026-07-12 Cached

Capn-hook is a CLI tool that gives coding agents persistent memory, saving files that answer questions about a codebase so agents don't re-explore the same mysteries across sessions, reducing token usage by 77% on repeat questions.

0 favorites 0 likes
#code-search

@geekbb: A CLI tool written in Go that integrates three search capabilities: Web search (Brave/DDG/SearXNG/Exa), code search (Grep/Sourcegraph/GitHub), and library documentation query (Context7). It also supports web scraping and site crawling. For AI...

X AI KOLs Timeline ↗ · 2026-06-30 Cached

A blazing-fast, stateless CLI tool written in Go that integrates Web search, code search, and library documentation query. It supports web scraping and site crawling, designed for AI agents and terminal use.

0 favorites 0 likes
#code-search

@Ryrenz: The token-saving artifact for coding agents has arrived—cocoindex-code, one command for semantic search on your codebase. Just open-sourced, quickly gaining stars with the selling point of "saving 70% tokens". Enabling agents like Claude Code, Codex…

X AI KOLs Timeline ↗ · 2026-06-29 Cached

cocoindex-code is an AST-based semantic code search tool that can be quickly integrated into coding agents, saving up to 70% tokens and improving search efficiency.

0 favorites 0 likes
#code-search

Recall Before Rerank: Benchmarking Deep Learning Models for Large-Scale Code-to-Code Retrieval

arXiv cs.CL ↗ · 2026-06-29 Cached

This paper benchmarks 17 deep learning models for first-stage recall in large-scale code-to-code retrieval, evaluating their precision, efficiency, and scalability across multiple programming languages and datasets. It introduces LLM-based code normalization and query rewriting schemes that improve precision for lower-performing models.

0 favorites 0 likes
#code-search

@Chenzeze777: Guys, I was totally stunned scrolling through GitHub today. Headroom gained 14k stars in a week, absolutely blowing up in the overseas developer circle. I initially thought it was just another PPT open-source project, but after a close look at the real-world test data—code search compressed from 17k tokens to 1,400, with the answer unchanged word for word. Let me...

X AI KOLs Timeline ↗ · 2026-06-08 Cached

Headroom is an open-source tool that compresses token usage in code search results and AI conversations by up to 92% (e.g., from 17k to 1,400 tokens) while maintaining answer quality. It supports multiple platforms and runs locally for free.

0 favorites 0 likes
#code-search

I built an open-source coding agent that makes context visible and editable — you curate exactly what the LLM sees

Reddit r/AI_Agents ↗ · 2026-05-31

The author built Nice Coding Agent, an open-source coding workbench with a visible and editable context stack, allowing users to curate exactly what the LLM sees. It features local-first retrieval, sandboxed execution, and hybrid code search, aiming to give developers control and visibility over context assembly.

0 favorites 0 likes
#code-search

@Trtd6Trtd: https://github.com/MinishLab/semble… High-speed code search library specialized for AI Compared to grep + reading, it s…

X AI KOLs Timeline ↗ · 2026-05-20 Cached

Semble 是一个面向 AI 代理的高效代码搜索库,使用模型如 Model2Vec 或 BM25 实现快速索引和检索,比 grep+read 节省约 98% 的 token,支持 MCP 服务器和 CLI 集成。

0 favorites 0 likes
#code-search

@aigclink: An Agent-oriented code search tool: Semble. It uses natural language to search codebases and returns precise code snippets, saving 98% token consumption compared to grep+read. The method lets Agents use natural language to directly locate the most relevant lines of code, without guessing keywords or reading entire files. Speed: indexing a typical…

X AI KOLs Timeline ↗ · 2026-05-19 Cached

Semble is an Agent-oriented code search tool that supports natural language queries, accurately returns semantically complete code snippets, saves 98% token consumption compared to traditional grep+read methods, and features intelligent chunking, dual-path retrieval, and code-aware re-ranking.

0 favorites 0 likes
#code-search

Built a local-first context engine for AI coding agents — symbol graph + semantic search, no cloud

Reddit r/artificial ↗ · 2026-05-18

Argyph is an open-source MCP server that provides AI coding agents with structured codebase understanding via a symbol graph and semantic search, running entirely locally with no cloud dependencies.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback