uncertainty

Tag

Cards List
#uncertainty

Actionable Hallucination Detection: Translating Latent Uncertainty into Agentic Critique

arXiv cs.LG · 8h ago Cached

This paper introduces the Latent Critic, a lightweight LoRA adapter that detects hallucinated agent actions in real time by restructuring the transformer's residual stream into localized natural-language feedback, achieving 0.966 AUROC and enabling self-correction.

0 favorites 0 likes
#uncertainty

MIDAS: Mutual Information Disentanglement with Uncertainty-Aware Fusion for Incomplete Multimodal Sentiment Analysis

arXiv cs.AI · 8h ago Cached

This paper proposes MIDAS, a unified framework for incomplete multimodal sentiment analysis that uses mutual information disentanglement and uncertainty-aware fusion to robustly represent and integrate modalities under missing-data conditions.

0 favorites 0 likes
#uncertainty

Can Gemma and Qwen models catch hallucinations by looking at their own logprobs?

Reddit r/LocalLLaMA · 14h ago

The author shares experiments using a custom WebUI to let Gemma and Qwen models inspect their own logprobs to detect hallucinations. Initial observations suggest that first-recall token probabilities can indicate uncertainty, though both models struggle to read their own logprobs.

0 favorites 0 likes
#uncertainty

From Uncertainty to Failure Attribution: Self-Diagnosing Models for Failure Attribution under Distribution Shift

arXiv cs.LG · yesterday Cached

The paper introduces self-diagnosing models that attribute model failures under distribution shift, linking uncertainty estimation with failure attribution.

0 favorites 0 likes
#uncertainty

Quantization Damage Is Multiplicative, Not Additive

arXiv cs.LG · 2d ago Cached

This preprint challenges the common assumption that quantization damage is additive noise, showing instead that it multiplies decision margins and shrinks them with bit-width, leading to silent failures in tool-use and safety decisions. The authors propose a fitted multiplicative model that predicts flip rates well.

0 favorites 0 likes
#uncertainty

Risk-Aware Decision Policies for Agents Under Noisy Perception

arXiv cs.LG · 2d ago Cached

This paper presents an artificial life predator-prey model of foraging under noisy perception, showing that uncertainty-aware decision policies significantly improve survival compared to blindly trusting perceptual labels, and that agents transition from exploratory to conservative strategies as uncertainty increases.

0 favorites 0 likes
#uncertainty

A Unified Risk View of Uncertainty: Posterior Risk for Disentanglement and Evaluation Beyond Proxies

arXiv cs.LG · 5d ago Cached

This paper proposes a unified definition of uncertainty as pointwise posterior risk and introduces a theory-backed benchmark using semi-synthetic datasets to directly compute oracle epistemic and aleatoric uncertainty, enabling fine-grained evaluation beyond proxy tasks.

0 favorites 0 likes
#uncertainty

PPDL: LLM-Based Flows as Probabilistic Programs

arXiv cs.LG · 5d ago Cached

This paper introduces PPDL, a probabilistic language for programming LLM-based flows that enables developers to quantify and propagate uncertainty throughout applications, with experimental and case studies on theorem proving.

0 favorites 0 likes
#uncertainty

Uncertainty-Aware World Model for Aerial Image-Goal Navigation

Hugging Face Daily Papers · 6d ago Cached

Presents UA-NWM, an uncertainty-aware latent world model for aerial image-goal navigation that decomposes prediction-goal discrepancy into uncertainty-explainable and unexplainable components, enabling robust trajectory scoring without multiple future samples.

0 favorites 0 likes
#uncertainty

Revisiting TD Target Aggregation under Uncertainty in Q-Learning

arXiv cs.LG · 2026-08-05 Cached

The paper proposes SADQ, a modification to Q-learning that uses one-step rollout predictions from a dynamics model to regularize TD target aggregation, reducing bootstrap-induced overestimation and improving training stability across benchmarks.

0 favorites 0 likes
#uncertainty

Subtype Robustness Is Not Just Accuracy: Calibration Under Unseen Subtype Shift

arXiv cs.LG · 2026-08-04 Cached

This paper presents the first systematic study of calibration under unseen subtype shift, showing that models become overconfident on novel subtypes within known coarse categories, and argues that subtype robustness should be evaluated with calibration metrics rather than accuracy alone.

0 favorites 0 likes
#uncertainty

DeepLook: Deeper Thinking with Lookahead

arXiv cs.AI · 2026-07-28 Cached

DeepLook is a training-free framework that improves LLM reasoning by allocating compute at uncertainty bottlenecks, reducing token generation by 87.3% on average while improving accuracy on competition math benchmarks.

0 favorites 0 likes
#uncertainty

Did the OpenAIs models actually manage to obtain the ExploitGym solutions?

Reddit r/singularity · 2026-07-27

The article questions whether OpenAI's models actually obtained solutions from ExploitGym, noting confusion amid news reports.

0 favorites 0 likes
#uncertainty

ProbSPARQL: Querying Knowledge Graphs with Multi-dimensional, Uncertain Numeric Data

arXiv cs.AI · 2026-07-22 Cached

ProbSPARQL is an upward-compatible SPARQL extension that models uncertain numeric values as random variables with probabilistic RDF literal datatypes, enabling distribution-aware queries, probabilistic filters, and divergence-based joins. Implemented on Apache Jena ARQ, it addresses challenges in querying multi-dimensional uncertain measurement data from circular manufacturing knowledge graphs.

0 favorites 0 likes
#uncertainty

Dynamic Loss Balancing for Joint SOH and RUL Prediction of Lithium-Ion Batteries via a Rotary SOH-Injected Prior Battery Transformer

arXiv cs.LG · 2026-07-22 Cached

Proposes RoSIP-Batt, a Transformer-based model for joint State of Health and Remaining Useful Life prediction of lithium-ion batteries, using dynamic loss balancing and rotary position embeddings, achieving state-of-the-art results on multiple datasets.

0 favorites 0 likes
#uncertainty

When LLMs Agree, Are They Right? Auditing Self-Consistency and Cross-Model Agreement as Confidence Signals

arXiv cs.AI · 2026-07-10 Cached

This paper audits whether self-consistency and cross-model agreement are reliable indicators of correctness in LLMs, finding that agreement is a weak, regime-dependent proxy and that frontier models exhibit overconfidence.

0 favorites 0 likes
#uncertainty

Robust Human-AI Complementarity under Uncertainty

arXiv cs.LG · 2026-07-09 Cached

This paper investigates how uncertainty about AI prediction quality affects human decision makers' ability to benefit from complementary information, finding that negative error correlation between human and AI predictions enables robust improvement strategies.

0 favorites 0 likes
#uncertainty

A toy framework for single and multi-agent human-AI curiosity ecosystems

arXiv cs.AI · 2026-07-08 Cached

This paper introduces a toy framework that models curiosity as an ecosystem in single and multi-agent settings, exploring how agents weigh immediate uncertainty reduction, costs, delayed returns, and the value of keeping questions open. It aims to inform future multi-agent AI systems for discovery.

0 favorites 0 likes
#uncertainty

Robustness Meets Uncertainty: Evidential Adversarial Training for Robust Selective Classification

arXiv cs.LG · 2026-07-07 Cached

This paper introduces Evidential Adversarial Training (EV-AT), a method that improves the robustness-uncertainty trade-off in classifiers by combining an evidence-based loss with robust evidence alignment, achieving state-of-the-art results on selective classification benchmarks.

0 favorites 0 likes
#uncertainty

Consistent but Miscalibrated: Evaluating LLM Limitations for Risk Communication in Natural Language

arXiv cs.CL · 2026-07-07 Cached

This paper evaluates nine LLMs on their ability to accurately communicate probabilistic predictions in natural language, finding that models are consistent but miscalibrated, particularly for uncertainty tasks.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback