model-selection

Tag

Cards List
#model-selection

You're Hired: Strategic Model Selection for LLM Collaboration

arXiv cs.CL ↗ · yesterday Cached

This paper introduces a taxonomy of 9 model selection algorithms for multi-LLM collaboration, showing that capability-aware selection strategies outperform random or heuristic team assembly by up to 36.1% across math, coding, QA, and reasoning tasks.

0 favorites 0 likes
#model-selection

Making AI an asset, not an expense

MIT Technology Review ↗ · 2d ago Cached

The article discusses economic strategies for enterprises transitioning AI from experimentation to production, focusing on evaluating ownership versus consumption models based on workload demands and cost predictability.

0 favorites 0 likes
#model-selection

Pretrain Once, Route Anywhere: Towards a Foundation Model for LLM Routing

Hugging Face Daily Papers ↗ · 3d ago Cached

The paper introduces RouteFM, a foundation model for LLM routing that learns reusable routing capabilities via episodic pretraining across heterogeneous environments, allowing a frozen router to adapt to new domains, modalities, and candidate pools through behavioral context alone. RouteFM outperforms the strongest baseline by 2.23 quality points on MMR-Bench with only eight observations per candidate, supporting a 'pretrain once, route anywhere' paradigm.

0 favorites 0 likes
#model-selection

FlexRouter: Learning Complementary Model Sets for Flexible LLM Routing

Hugging Face Daily Papers ↗ · 3d ago Cached

FlexRouter proposes a coverage-oriented LLM routing framework that uses Determinantal Point Processes to model complementarity among models, maximizing the probability that at least one selected model answers correctly while avoiding redundant selections and fixed budgets.

0 favorites 0 likes
#model-selection

Routing Should Pay for Itself: Sparse Supervision for Economical LLM Routing

Hugging Face Daily Papers ↗ · 3d ago Cached

The paper proposes SaveRouter, a sparse-supervision LLM routing framework that selectively acquires query-model feedback and shares capability information across related queries, cutting supervision costs while maintaining competitive routing quality and reducing break-even deployment volume by 1.9-9.5x.

0 favorites 0 likes
#model-selection

SeLMRoute: Probabilistic Semantic Evidence for Large Language Model Routing

Hugging Face Daily Papers ↗ · 4d ago Cached

SeLMRoute introduces an LLM routing framework that separates candidate-independent semantic evidence extraction from performance learning, achieving 72.08% average accuracy on LLMRouterBench across 15 datasets and 20 candidate models, outperforming the strongest fixed candidate (69.23%) and enabling both performance-oriented and cost-aware routing decisions.

0 favorites 0 likes
#model-selection

A learned LLM router scored 0.84 AUC. Shuffling the labels within each task still scored 0.838 [R]

Reddit r/MachineLearning ↗ · 5d ago

A study revealed that an LLM router trained to select between models learned task recognition instead of difficulty, causing poor generalization on held-out tasks, but deferral based on the cheap model's output yielded better performance.

0 favorites 0 likes
#model-selection

Rank Portability Does Not Imply Feasibility Portability: Target-Specific Evaluation of Joint Hardware Constraints

arXiv cs.LG ↗ · 2026-09-22 Cached

The paper challenges the assumption that rank portability implies feasibility portability in cross-device hardware evaluation, using benchmarks to show high rank correlation does not guarantee safe deployment decisions under joint constraints.

0 favorites 0 likes
#model-selection

@GergelyOrosz: Talked with a company where, a year ago they had unlimited AI budgets + CEO is technical and very bullish in AI “We now…

X AI KOLs Timeline ↗ · 2026-09-20 Cached

A company with a strong AI focus has implemented daily budget limits for state-of-the-art models, shifting to cheaper models for most tasks, indicating a trend in AI cost management.

0 favorites 0 likes
#model-selection

(Genuinely asking) Are smaller quantized models becoming the real sweet spot for local AI?

Reddit r/LocalLLaMA ↗ · 2026-09-19

The article questions whether smaller quantized models are becoming the preferred choice for local AI applications, emphasizing their balance of VRAM usage, performance, and capability like tool calling.

0 favorites 0 likes
#model-selection

FedFIbOS: Fisher Importance based Optimal Submodelling for Heterogeneous Federated Learning

arXiv cs.LG ↗ · 2026-09-18 Cached

FedFIbOS proposes a Fisher importance-based method for optimal submodel selection in heterogeneous federated learning, theoretically grounded and achieving about 10% higher accuracy than state-of-the-art methods under non-IID settings.

0 favorites 0 likes
#model-selection

@eve: Automatic tool approvals powered by @typesafeai’s Jev. https://eve.dev/docs/guides/evaluate#evaluate-tool-approvals…

X AI KOLs Timeline ↗ · 2026-09-17 Cached

Eve introduces automatic tool approvals and model selection using TypeSafe AI's Jev, enabling dynamic choice of AI models for different tasks via the AI SDK evaluation API.

0 favorites 0 likes
#model-selection

Optimal Model Activation Policies for Inference Networks of Large Language Models

arXiv cs.CL ↗ · 2026-09-16 Cached

The paper introduces inference networks, a graph-based framework for optimizing the use of multiple LLMs in inference, with optimal activation policies that minimize cost while meeting performance targets.

0 favorites 0 likes
#model-selection

Save Money Automagically by Auto Routing LLM Choice

Reddit r/AI_Agents ↗ · 2026-09-12

The author built an automated system to route LLM choices based on cost and performance data, aiming to optimize AI spending, with plans to open-source the tool.

0 favorites 0 likes
#model-selection

anypick - Library for filtering and selecting LLMs

Reddit r/AI_Agents ↗ · 2026-09-08

A Python and TypeScript library for downloading LLM catalogs and building pipelines to filter and select models based on criteria like price, latency, and benchmarks.

0 favorites 0 likes
#model-selection

My RULE of Thumb of choosing a models

Reddit r/LocalLLaMA ↗ · 2026-09-03

The author shares personal experience showing how LLMs like Qwen 27B drastically reduce programming task time, offering rules of thumb for model selection.

0 favorites 0 likes
#model-selection

SCX Router: Streaming Zero-Shot Model Selection with a Decoder-KV Classifier and a Real-World Task Ontology

arXiv cs.AI ↗ · 2026-09-03 Cached

SCX Router introduces a lightweight GLiClass-based model selection tool that uses a decoder-KV classifier and a task ontology to route LLM tasks, optimizing for speed, cost, and quality without autoregressive generation.

0 favorites 0 likes
#model-selection

I'm building profile-guided optimization for AI agents

Reddit r/AI_Agents ↗ · 2026-09-02

The author is developing Agent-PGO, a tool that profiles AI agent executions to dynamically substitute cheaper models for less critical tasks while maintaining quality through evaluation benchmarks.

0 favorites 0 likes
#model-selection

EEG-AS: Instance-Level Foundation Model Selection for EEG Foundation Models via Behavior Reconstruction

arXiv cs.LG ↗ · 2026-09-02 Cached

The paper proposes EEG-AS, an algorithm selection framework that enables instance-level selection among multiple EEG foundation models by reconstructing their behaviors, thereby improving neural decoding performance.

0 favorites 0 likes
#model-selection

Stochastic complexity of vectors containing cluster structure

arXiv cs.LG ↗ · 2026-09-02 Cached

This paper introduces a recursion formula for efficiently computing the stochastic complexity of vectors with cluster structure using the Normalized Maximum Likelihood model, reducing time complexity from polynomial to linear.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback