tool-selection

Tag

Cards List
#tool-selection

ToolSearcher: Optimizing Tool Selection at Scale via Reinforcement Learning

arXiv cs.CL ↗ · 3d ago Cached

ToolSearcher is a reinforcement learning framework designed to optimize tool selection for large language models in large-scale tool repositories, using techniques such as category-constrained discrimination and event-level search modeling to improve performance in iterative search and tool composition tasks.

0 favorites 0 likes
#tool-selection

Harness Router

Product Hunt ↗ · 3d ago Cached

Harness Router is a decision layer for AI coding agents that improves tool selection before execution using Jev for ambiguity and MCTS for multi-step consequences.

0 favorites 0 likes
#tool-selection

If Jev picks the tool, how does the LLM ask for another one?

Reddit r/AI_Agents ↗ · 4d ago

The article discusses design patterns for integrating Jev with LLMs in AI agents, specifically how agents handle tool selection and planning when only one tool schema is exposed at a time.

0 favorites 0 likes
#tool-selection

Toollery: Scaling LLM Agents to Thousands of Skills and Tools

arXiv cs.LG ↗ · 2026-09-22 Cached

Toollery is a training-free candidate-compression framework that improves scalability and efficiency for LLM agent tool and skill selection through retrieval-based methods.

0 favorites 0 likes
#tool-selection

Update: parked HITL writes. Cheap Flash for a read-only analytics agent — dates, bilingual routing, and Gemini cache at 0%

Reddit r/AI_Agents ↗ · 2026-09-16

The author updates on a read-only analytics agent project using Gemini Flash Lite, discusses caching inefficiencies and bilingual routing challenges, and seeks advice on date handling.

0 favorites 0 likes
#tool-selection

The Router Within: Eliciting Native Skill Routing from a Frozen LLM

Hugging Face Daily Papers ↗ · 2026-09-14 Cached

The paper introduces Gavel, a method that elicits native skill routing from frozen LLMs via linear projections, enabling efficient tool selection without context overload and outperforming existing pipelines on benchmarks.

0 favorites 0 likes
#tool-selection

Built a read-only analytics agent (route → fetch → narrate → ground). Before I let it write anything, what am I missing?

Reddit r/AI_Agents ↗ · 2026-09-09

A developer describes building a read-only analytics agent for restaurant POS systems using Node.js, TypeScript, and Gemini Flash, seeking advice on tool selection, grounding techniques, and patterns for implementing write actions.

0 favorites 0 likes
#tool-selection

Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out

Hacker News Top ↗ · 2026-09-03 Cached

A study measured 16,893 sessions to analyze how AI coding agents like Claude Code, Codex, and Cursor select tools such as databases, finding consistent recommendations across varied contexts.

0 favorites 0 likes
#tool-selection

How to choose an AI gent tool? Sep 2026 edition

Reddit r/AI_Agents ↗ · 2026-09-01

This article offers a practical framework for evaluating AI agent tools by focusing on key questions about job clarity, access needs, autonomy, maintenance, cost, and exit options, moving beyond feature comparisons.

0 favorites 0 likes
#tool-selection

Gave an agent 30 tools. It got worse at using the 3 that mattered.

Reddit r/AI_Agents ↗ · 2026-08-28

The author observed that adding more tools to an AI agent decreased its accuracy in selecting the correct tool due to increased classification complexity, and found that using multiple smaller agents with narrower toolsets improved reliability.

0 favorites 0 likes
#tool-selection

Representation Affects Retrieval: A Case Study of Skill Discovery and Routing in a Multimodal Agent Harness

arXiv cs.AI ↗ · 2026-08-24 Cached

This paper presents a case study on skill discovery and routing in a multimodal agent harness, showing that partial in-prompt exposure of skills can create lexical competition that hinders correct selection, linking small-scale in-context retrieval to large-scale approaches.

0 favorites 0 likes
#tool-selection

@helloiamleonie: OpenMed Liquid AI

X AI KOLs Timeline ↗ · 2026-08-18 Cached

OpenMed demonstrated using Liquid AI's LFM2.5-VL-3B model to analyze a generated skin image, mapping regions and measuring dimensions locally on a Mac Studio for visual review.

0 favorites 0 likes
#tool-selection

@hanakoxbt: your agent has thirty tools. it calls two of them. the other twenty eight are not sitting idle somewhere. they are in t…

X AI KOLs Following ↗ · 2026-08-08 Cached

Unused tools in an AI agent's toolset still consume tokens and add noise to tool selection, so agents should load only the tools required for the current task.

0 favorites 0 likes
#tool-selection

Diagnosing Tool-Selection Reasoning in LLM Agents with Canary Tools

arXiv cs.AI ↗ · 2026-08-06 Cached

This paper introduces 'canary tools' as diagnostic probes to identify specific tool-selection reasoning failures in LLM agents, proposing a six-type taxonomy and evaluating eight models across 8,640 task runs to show capability-tier disparities and robustness.

0 favorites 0 likes
#tool-selection

@alex_prompter: My agents kept getting dumber every time I gave them more tools. The reason is mechanical. Every MCP server you connect…

X AI KOLs Timeline ↗ · 2026-07-21 Cached

Ratel is an open-source tool that reduces input tokens by 79% and improves tool selection accuracy for AI agents by loading only needed tools using a BM25 index, instead of all available tools.

0 favorites 0 likes
#tool-selection

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth

arXiv cs.LG ↗ · 2026-07-21 Cached

DocOCR-Eval proposes an annotation-free framework that uses a correction and ranking strategy to evaluate and select OCR tools without ground truth labels, showing that aggregating multiple multimodal large language models improves alignment with human rankings.

0 favorites 0 likes
#tool-selection

Where MCP tool-selection actually breaks: retrieval-based fixes cap at ~23% of failures

Reddit r/AI_Agents ↗ · 2026-07-13

Recent analysis reveals that retrieval-based tool selection for LLM agents caps out at recovering ~23% of failures, while readout-side interventions addressing attention biases recover 59-91% of failures, indicating that the real bottleneck is in the model's output processing rather than input filtering.

0 favorites 0 likes
#tool-selection

@liquidai: Storing too many tools in your context window increases latency and can lead to wrong tool selection. In this demo, we …

X AI KOLs Following ↗ · 2026-06-19 Cached

Liquid AI demonstrates using LFM2.5-ColBERT-350M as a filter to select only the five most relevant tools from 151 options, reducing latency and improving tool selection accuracy.

0 favorites 0 likes
#tool-selection

@maximelabonne: LFM2.5-ColBERT-350M is a surprisingly reliable smart tool selector. We gave it 151 tools, and it consistently surfaces …

X AI KOLs Following ↗ · 2026-06-18 Cached

LFM2.5-ColBERT-350M is a model that reliably selects the most relevant tools from a set of 151, saving tokens and improving accuracy, ideal for agentic edge models.

0 favorites 0 likes
#tool-selection

When Lower Privileges Suffice: Investigating Over-Privileged Tool Selection in LLM Agents

Hugging Face Daily Papers ↗ · 2026-06-18 Cached

This paper investigates over-privileged tool selection in LLM agents, introducing ToolPrivBench to evaluate and mitigate unnecessary use of high-privilege tools. It finds that safety alignment does not ensure least-privilege choices, and proposes a post-training defense that reduces excessive privilege use without sacrificing performance.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback