model-selection

Tag

Cards List
#model-selection

FlexRouter: Learning Complementary Model Sets for Flexible LLM Routing

arXiv cs.LG ↗ · yesterday Cached

FlexRouter 提出一种显式建模模型互补性的 LLM 路由框架,将路由建模为基于覆盖的子集选择问题,并用 Determinantal Point Processes 对模型能力与冗余进行联合建模,同时通过边缘化失败集合的目标函数直接优化答案覆盖率。在 RouterEval 基准上,该方法在域内与域外任务上以更低冗余实现了更高的覆盖率,且推理成本灵活可控。

0 favorites 0 likes
#model-selection

You're Hired: Strategic Model Selection for LLM Collaboration

arXiv cs.CL ↗ · 2d ago Cached

This paper introduces a taxonomy of 9 model selection algorithms for multi-LLM collaboration, showing that capability-aware selection strategies outperform random or heuristic team assembly by up to 36.1% across math, coding, QA, and reasoning tasks.

0 favorites 0 likes
#model-selection

Making AI an asset, not an expense

MIT Technology Review ↗ · 3d ago Cached

The article discusses economic strategies for enterprises transitioning AI from experimentation to production, focusing on evaluating ownership versus consumption models based on workload demands and cost predictability.

0 favorites 0 likes
#model-selection

Pretrain Once, Route Anywhere: Towards a Foundation Model for LLM Routing

Hugging Face Daily Papers ↗ · 4d ago Cached

The paper introduces RouteFM, a foundation model for LLM routing that learns reusable routing capabilities via episodic pretraining across heterogeneous environments, allowing a frozen router to adapt to new domains, modalities, and candidate pools through behavioral context alone. RouteFM outperforms the strongest baseline by 2.23 quality points on MMR-Bench with only eight observations per candidate, supporting a 'pretrain once, route anywhere' paradigm.

0 favorites 0 likes
#model-selection

FlexRouter: Learning Complementary Model Sets for Flexible LLM Routing

Hugging Face Daily Papers ↗ · 4d ago Cached

FlexRouter proposes a coverage-oriented LLM routing framework that uses Determinantal Point Processes to model complementarity among models, maximizing the probability that at least one selected model answers correctly while avoiding redundant selections and fixed budgets.

0 favorites 0 likes
#model-selection

Routing Should Pay for Itself: Sparse Supervision for Economical LLM Routing

Hugging Face Daily Papers ↗ · 4d ago Cached

The paper proposes SaveRouter, a sparse-supervision LLM routing framework that selectively acquires query-model feedback and shares capability information across related queries, cutting supervision costs while maintaining competitive routing quality and reducing break-even deployment volume by 1.9-9.5x.

0 favorites 0 likes
#model-selection

SeLMRoute: Probabilistic Semantic Evidence for Large Language Model Routing

Hugging Face Daily Papers ↗ · 5d ago Cached

SeLMRoute introduces an LLM routing framework that separates candidate-independent semantic evidence extraction from performance learning, achieving 72.08% average accuracy on LLMRouterBench across 15 datasets and 20 candidate models, outperforming the strongest fixed candidate (69.23%) and enabling both performance-oriented and cost-aware routing decisions.

0 favorites 0 likes
#model-selection

A learned LLM router scored 0.84 AUC. Shuffling the labels within each task still scored 0.838 [R]

Reddit r/MachineLearning ↗ · 6d ago

A study revealed that an LLM router trained to select between models learned task recognition instead of difficulty, causing poor generalization on held-out tasks, but deferral based on the cheap model's output yielded better performance.

0 favorites 0 likes
#model-selection

Rank Portability Does Not Imply Feasibility Portability: Target-Specific Evaluation of Joint Hardware Constraints

arXiv cs.LG ↗ · 2026-09-22 Cached

The paper challenges the assumption that rank portability implies feasibility portability in cross-device hardware evaluation, using benchmarks to show high rank correlation does not guarantee safe deployment decisions under joint constraints.

0 favorites 0 likes
#model-selection

@GergelyOrosz: Talked with a company where, a year ago they had unlimited AI budgets + CEO is technical and very bullish in AI “We now…

X AI KOLs Timeline ↗ · 2026-09-20 Cached

A company with a strong AI focus has implemented daily budget limits for state-of-the-art models, shifting to cheaper models for most tasks, indicating a trend in AI cost management.

0 favorites 0 likes
#model-selection

(Genuinely asking) Are smaller quantized models becoming the real sweet spot for local AI?

Reddit r/LocalLLaMA ↗ · 2026-09-19

The article questions whether smaller quantized models are becoming the preferred choice for local AI applications, emphasizing their balance of VRAM usage, performance, and capability like tool calling.

0 favorites 0 likes
#model-selection

FedFIbOS: Fisher Importance based Optimal Submodelling for Heterogeneous Federated Learning

arXiv cs.LG ↗ · 2026-09-18 Cached

FedFIbOS proposes a Fisher importance-based method for optimal submodel selection in heterogeneous federated learning, theoretically grounded and achieving about 10% higher accuracy than state-of-the-art methods under non-IID settings.

0 favorites 0 likes
#model-selection

@eve: Automatic tool approvals powered by @typesafeai’s Jev. https://eve.dev/docs/guides/evaluate#evaluate-tool-approvals…

X AI KOLs Timeline ↗ · 2026-09-17 Cached

Eve introduces automatic tool approvals and model selection using TypeSafe AI's Jev, enabling dynamic choice of AI models for different tasks via the AI SDK evaluation API.

0 favorites 0 likes
#model-selection

Optimal Model Activation Policies for Inference Networks of Large Language Models

arXiv cs.CL ↗ · 2026-09-16 Cached

The paper introduces inference networks, a graph-based framework for optimizing the use of multiple LLMs in inference, with optimal activation policies that minimize cost while meeting performance targets.

0 favorites 0 likes
#model-selection

Save Money Automagically by Auto Routing LLM Choice

Reddit r/AI_Agents ↗ · 2026-09-12

The author built an automated system to route LLM choices based on cost and performance data, aiming to optimize AI spending, with plans to open-source the tool.

0 favorites 0 likes
#model-selection

anypick - Library for filtering and selecting LLMs

Reddit r/AI_Agents ↗ · 2026-09-08

A Python and TypeScript library for downloading LLM catalogs and building pipelines to filter and select models based on criteria like price, latency, and benchmarks.

0 favorites 0 likes
#model-selection

My RULE of Thumb of choosing a models

Reddit r/LocalLLaMA ↗ · 2026-09-03

The author shares personal experience showing how LLMs like Qwen 27B drastically reduce programming task time, offering rules of thumb for model selection.

0 favorites 0 likes
#model-selection

SCX Router: Streaming Zero-Shot Model Selection with a Decoder-KV Classifier and a Real-World Task Ontology

arXiv cs.AI ↗ · 2026-09-03 Cached

SCX Router introduces a lightweight GLiClass-based model selection tool that uses a decoder-KV classifier and a task ontology to route LLM tasks, optimizing for speed, cost, and quality without autoregressive generation.

0 favorites 0 likes
#model-selection

I'm building profile-guided optimization for AI agents

Reddit r/AI_Agents ↗ · 2026-09-02

The author is developing Agent-PGO, a tool that profiles AI agent executions to dynamically substitute cheaper models for less critical tasks while maintaining quality through evaluation benchmarks.

0 favorites 0 likes
#model-selection

EEG-AS: Instance-Level Foundation Model Selection for EEG Foundation Models via Behavior Reconstruction

arXiv cs.LG ↗ · 2026-09-02 Cached

The paper proposes EEG-AS, an algorithm selection framework that enables instance-level selection among multiple EEG foundation models by reconstructing their behaviors, thereby improving neural decoding performance.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback