truthfulqa

Tag

Cards List
#truthfulqa

IDEEA: training-free Input-Dependent stEEring via Activation cluster matching

arXiv cs.CL · 2026-09-03 Cached

IDEEA proposes a training-free, input-dependent steering method for large language models that clusters activations and uses optimal matching to improve truthfulness in TruthfulQA by up to 23.5% over baselines.

0 favorites 0 likes
#truthfulqa

Attention-Guided Layer Selection for Contrastive Decoding in Large Language Models

arXiv cs.CL · 2026-07-28 Cached

Proposes three attention-guided strategies for layer selection in contrastive decoding for large language models, improving factuality on TruthfulQA over the DoLa baseline.

0 favorites 0 likes
#truthfulqa

Are Diversity Metrics Measuring Diversity? A Capability-Controlled Audit of Majority-Vote Gain in LLM Ensembles

arXiv cs.CL · 2026-07-24 Cached

This paper audits five diversity measures for LLM ensembles, finding that their associations with majority-vote gain are heavily entangled with model capability and are unstable after controlling for capability. The only robust signal is a modest residual pairwise co-failure association.

0 favorites 0 likes
#truthfulqa

Comprehensive Evaluation of Large Language Model Responses: A Multi-Factor Scoring System

arXiv cs.CL · 2026-07-09 Cached

This paper proposes a multi-factor scoring system for evaluating LLM responses, integrating accuracy, conciseness, factual consistency, readability, and coherence. Applied to the TruthfulQA dataset, it reveals strengths and limitations of mainstream models, offering a transparent evaluation framework.

0 favorites 0 likes
← Back to home

Submit Feedback