function-calling

Tag

Cards List
#function-calling

Extending FunctionGemma for Practical On-Device Mobile Function Calling

arXiv cs.LG ↗ · 2d ago Cached

This research paper extends FunctionGemma 270M for practical on-device Android workflows by introducing a synthetic dataset and fine-tuning the model, achieving improved accuracy for function calling while balancing performance and coverage.

0 favorites 0 likes
#function-calling

@googleaidevs: Ever get stuck on a bug and wish the exact problem could be pointed out to you on screen? Watch how we used Gemini 3.8 …

X AI KOLs Timeline ↗ · 4d ago Cached

Google AI developers demonstrated how they used Gemini 3.8 Live Extended Thinking to build a coding tutor that assists with debugging by analyzing the screen and referencing the p5.js library.

0 favorites 0 likes
#function-calling

Closed-World Resolution Against Tool Hallucination in LLM Agents

arXiv cs.AI ↗ · 2026-09-18 Cached

This paper introduces a closed-world resolution method to combat tool hallucination in LLM agents, offering a taxonomy and benchmark for measuring and addressing fabricated tool calls.

0 favorites 0 likes
#function-calling

Cactus Needle 3: A Sliceable 8-29MB Automation Foundation Model That Matches DeepSeek v4 Flash

Reddit r/LocalLLaMA ↗ · 2026-09-17

Cactus Needle 3 is a small, sliceable foundation model for automation tasks that runs on-device, achieving performance comparable to larger models on function calling and structured extraction.

0 favorites 0 likes
#function-calling

@miiura: Function calling without waiting for Enter. Answers only when needed. @typesafeai Jev understands intent and arguments …

X AI KOLs Following ↗ · 2026-09-17 Cached

Jev by typesafe AI enables function calling to execute in real-time as users type, understanding intent and arguments without waiting for Enter.

0 favorites 0 likes
#function-calling

Carbon-Aware Routing for Function Calling in Edge-Cloud LLM Systems

arXiv cs.AI ↗ · 2026-09-15 Cached

This paper introduces a carbon-aware routing framework for function-calling LLMs in edge-cloud systems, reducing operational carbon emissions by an average of 4× while maintaining cloud-level accuracy.

0 favorites 0 likes
#function-calling

From Production Traffic to Post-Training: Building a Self-Hosted LLM That Covers the Corporate Request Mix

Hugging Face Daily Papers ↗ · 2026-09-01 Cached

A research paper presents a method for training a smaller self-hosted LLM using separate GRPO experts merged via SLERP, which outperforms a larger baseline on enterprise tasks and serves half of platform traffic at lower cost.

0 favorites 0 likes
#function-calling

@ramin_m_h: yesterday we made them more compressed! today we make them faster than ever with speculative decoding! up to 4x decode …

X AI KOLs Timeline ↗ · 2026-08-20 Cached

Liquid AI releases DSpark draft models for their LFM series, incorporating speculative decoding to achieve up to 4x decode speedup on device while maintaining output quality.

0 favorites 0 likes
#function-calling

Repair, Not Improvement: Decomposing Constrained Decoding in Tool-Call Abstention

arXiv cs.CL ↗ · 2026-08-17 Cached

This paper investigates the impact of constrained decoding on tool-call abstention, decomposing the effects of grammar constraints into stop and emission components, and evaluates performance on small open-weight models across languages.

0 favorites 0 likes
#function-calling

LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge

Hugging Face Blog ↗ · 2026-08-12 Cached

LiquidAI announces LFM2.5-VL-3B, an efficient vision-language model for edge hardware with improved screen understanding, grounding, multi-image input, and function calling, trained with 4x more vision data and post-training via SFT and RL.

0 favorites 0 likes
#function-calling

1.3B activated params out of 7.9B total, aimed at agent work. Where does this curve flatten?

Reddit r/ArtificialInteligence ↗ · 2026-08-07

InclusionAI (Ant Group's lab) launched Ling 3.0 Tiny, a closed API model with 1.3B activated params out of 7.9B total, 256K context, native function calling, and a thinking/instant mode, aimed at multi-turn agent tool loops. The article questions the efficiency curve and notes no independent evals or open weights.

0 favorites 0 likes
#function-calling

Data Turnstile: A Scalable Open Framework for Function-Calling Data Generation

arXiv cs.CL ↗ · 2026-08-03 Cached

Data Turnstile is an open-source framework for generating high-quality synthetic function-calling training data from API specifications. Fine-tuning small language models with this data significantly improves their tool-use performance, closing the gap with much larger models.

0 favorites 0 likes
#function-calling

SAAG: Structured Agent Assessment and Grounding

arXiv cs.AI ↗ · 2026-07-22 Cached

SAAG proposes a cascaded diagnostic framework for evaluating LLM agent function calling by decomposing evaluation into registry conformance, structural completeness, and argument grounding stages, enabling interpretable diagnostics and iterative self-repair. Experiments with sub-4B models show improved argument precision and reduced value hallucination compared to single-pass evaluation.

0 favorites 0 likes
#function-calling

GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF

Hugging Face Models Trending ↗ · 2026-07-13 Cached

GnLOLot releases GGUF quantizations of the MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking model, a 1B parameter thinking model fine-tuned on Fable 5 data with improved tool/function calling compared to V1, designed for local deployment via llama.cpp and compatible runtimes.

0 favorites 0 likes
#function-calling

@ando_w: https://x.com/ando_w/status/2075468963098546520

X AI KOLs Timeline ↗ · 2026-07-10 Cached

This article introduces how to upgrade single-turn RAG to Agentic RAG, by allowing the LLM to autonomously decide on multiple retrievals and tool calls to solve multi-step reasoning for complex problems. It provides code examples and implementation ideas based on Qwen3.7-Max.

0 favorites 0 likes
#function-calling

Expanding Managed Agents in Gemini API: background tasks, remote MCP and more

Google AI Blog ↗ · 2026-07-07 Cached

Google is expanding Managed Agents in the Gemini API with new capabilities including background execution, remote MCP server integration, custom function calling, and credential refresh, enabling more reliable and production-ready agents.

0 favorites 0 likes
#function-calling

@ModelScope2022: Introducing Agents-A1, A 35B MoE agentic model built for long-horizon tasks across search, engineering, scientific rese…

X AI KOLs Timeline ↗ · 2026-06-30 Cached

ModelScope introduces Agents-A1, a 35B MoE agentic model with 256K context and function calling, achieving SOTA on long-horizon tasks and instruction following.

0 favorites 0 likes
#function-calling

@BrianRoemmele: BOOM! Meet the open source Cambrian Explosion of repulsion of Anthropic! Meet Qwythos 9B, a Qwen3.5 based GGUF that's b…

X AI KOLs Timeline ↗ · 2026-06-28 Cached

Qwythos 9B is a new open-source, uncensored reasoning model based on Qwen3.5, offering GGUF quantizations, 1 million token context, vision, and function calling, with significant performance improvements over the base model.

0 favorites 0 likes
#function-calling

@fahdmirza: We trained a 9B model on Claude Mythos traces — here's what happened Meet Qwythos 9B: open-source, local, and punching …

X AI KOLs Timeline ↗ · 2026-06-27 Cached

Trained a 9B model (Qwythos 9B) on Claude Mythos traces, achieving strong results in bug-finding, SQL, and function calling while running locally on an A6000 with llama.cpp.

0 favorites 0 likes
#function-calling

empero-ai/Qwythos-9B-Claude-Mythos-5-1M-GGUF

Hugging Face Models Trending ↗ · 2026-06-19 Cached

Empero AI releases Qwythos-9B-Claude-Mythos-5-1M-GGUF, a 9B parameter reasoning model fine-tuned on 500M+ tokens of Claude Mythos/Fable traces with chain-of-thought, achieving significant gains over Qwen3.5-9B and supporting 1M-token context via YaRN rope-scaling. The GGUF quantizations enable local inference on llama.cpp and compatible runtimes.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback