multi-llm

Tag

Cards List
#multi-llm

MIDAS: Multi-LLM Iterative Data-Adaptive Summarization

arXiv cs.CL · 2026-08-06 Cached

This paper proposes MIDAS, a multi-LLM framework for data-adaptive summarization that automates prompt optimization for domain-specific enterprise use cases, achieving strong improvements over prior methods on customer ticket summarization benchmarks.

0 favorites 0 likes
#multi-llm

Opti-Q: A Constraint-Based Optimization Framework for Multi-LLM Question Planning

arXiv cs.AI · 2026-07-28 Cached

Opti-Q is a database-inspired optimizer for multi-LLM question answering that plans execution DAGs to optimize answer quality under cost, latency, and energy constraints, achieving significant improvements on benchmarks.

0 favorites 0 likes
#multi-llm

Uncertainty-Aware Trust Estimation for Multi-LLM Systems via Structured Expert Judgement

arXiv cs.LG · 2026-07-24 Cached

This paper introduces an uncertainty-aware trust estimation method for aggregating predictions from multiple LLMs, adapting structured expert judgment with Cooke-style log weighting to penalize overconfident incorrect predictions. Evaluations on MMLU and MMLU-Pro show that this approach achieves superior accuracy-reliability balance under heterogeneous and contaminated expert panels.

0 favorites 0 likes
#multi-llm

SynthAVE: Scalable Synthetic Labeling for E-Commerce with LLM-Arena Validation

arXiv cs.CL · 2026-07-09 Cached

This paper presents SynthAVE, a large-scale human-validated benchmark for attribute value extraction in e-commerce, using a multi-LLM arena framework with 21 judge configurations to validate synthetic labels efficiently and cost-effectively while maintaining quality parity with human review.

0 favorites 0 likes
#multi-llm

@degenrsc: https://x.com/degenrsc/status/2064714047241736302

X AI KOLs Timeline · 2026-06-10 Cached

A detailed guide on building an agentic research framework using a multi-LLM system with persistent memory, allowing researchers to avoid re-explaining context across sessions by leveraging file-based identity, project docs, and memory indices.

0 favorites 0 likes
← Back to home

Submit Feedback