arxiv

Tag

Cards List
#arxiv

Learning from Environmental Feedback: Credit Assignment across Multiple Timescales for Agentic Reinforcement Learning

arXiv cs.LG · yesterday Cached

This paper introduces EFCA, a multi-timescale credit assignment method for agentic reinforcement learning that uses short-term feedback and medium-term state-history signals from environment interaction to improve task success and quality on ALFWorld and WebShop.

0 favorites 0 likes
#arxiv

Deployable Per-Instance Multi-Layer Activation Steering for Large Language Models

arXiv cs.CL · yesterday Cached

This paper introduces a deployable per-instance, multi-layer activation steering technique for large language models, showing that optimal layer selection varies per input and can be predicted from the prompt embedding without gold labels at inference.

0 favorites 0 likes
#arxiv

Instability of LLM Pre-Pretraining: It Doesn't Always Help. An Investigation on Multiple Languages

arXiv cs.CL · yesterday Cached

This paper investigates whether pretraining LLMs on artificial languages (pre-pretraining) consistently improves token efficiency across multiple natural languages, finding that gains are highly dependent on experimental setup and random seed, though stable gains appear for small models with the Llama tokenizer.

0 favorites 0 likes
#arxiv

STEMMA: An Adversarial Multi-Agent Framework for Evaluating Self-Identity Consistency in LLMs

arXiv cs.CL · yesterday Cached

The paper introduces STEMMA, a multi-agent framework that adversarially probes self-identity consistency in LLMs, motivated by concerns that knowledge distillation may transfer behavioral traits like identity representation from teacher to student models.

0 favorites 0 likes
#arxiv

NeuPAT: Neuron-aware Plasticity Allocation Tuning for Language-Preserving MLLMs

arXiv cs.CL · yesterday Cached

NeuPAT is a lightweight, architecture-agnostic framework that allocates neuron-wise update constraints during multimodal instruction tuning to preserve language capabilities in MLLMs, recovering 94.5% of language degradation from vanilla tuning.

0 favorites 0 likes
#arxiv

Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States

arXiv cs.CL · yesterday Cached

This paper introduces Prompt Embedding Probes (PEP), a parameter-efficient extension of linear probes that uses learnable prompt embeddings on hidden states to detect hallucinations in frozen LLMs. Evaluations on TriviaQA, GSM8K, and MedQA with Qwen3 models show improvements over standard linear probes, including in pre-generation and cross-model settings.

0 favorites 0 likes
#arxiv

Thinking Hard, Not Smart: Reasoning Models Fail to Ration Test-Time Compute Across Questions

arXiv cs.CL · yesterday Cached

This paper introduces an exam-style evaluation to study how reasoning models allocate a shared test-time compute budget across multiple questions. It finds that models fail to strategically ration compute, instead prioritizing questions by presentation order and ignoring value or difficulty.

0 favorites 0 likes
#arxiv

Finite Constant Frontiers and Auditable Regret Certificates for Average-Reward Reinforcement Learning

arXiv cs.LG · yesterday Cached

This paper introduces a constant-aware comparison protocol for average-reward reinforcement learning regret bounds, deriving an explicit finite lower certificate for communicating MDPs and improving published coefficients.

0 favorites 0 likes
#arxiv

CODS: Iterative Bellman-Residual Data Selection for Reusable Offline Reinforcement Learning

arXiv cs.LG · yesterday Cached

Introduces CODS, an iterative critic-guided data selection method for offline reinforcement learning that retains task performance at low data budgets by selecting high-residual transitions over multiple rounds.

0 favorites 0 likes
#arxiv

SkillConsist: Detecting Inconsistencies in Agent Skills via Bidirectional Graph Alignment

arXiv cs.LG · yesterday Cached

The paper introduces SkillConsist, a method using bidirectional graph alignment to detect inconsistencies between declared and implemented behavior in LLM agent skills, achieving strong F1 improvements over baselines.

0 favorites 0 likes
#arxiv

When Is Benchmark Contamination Detectable? Information Limits and Power-Calibrated Audits

arXiv cs.AI · yesterday Cached

This paper formalizes when benchmark contamination is detectable, deriving information-theoretic limits and proposing power-calibrated audits that distinguish a clean benchmark from a powerless detector. It reports two-sided empirical findings on calibration efficacy and validity gates.

0 favorites 0 likes
#arxiv

GRACE: LLM-Grounded Semantic Metric Spaces for Scalable Mixed-Data Clustering

arXiv cs.AI · yesterday Cached

This paper introduces GRACE, a framework that uses LLM-generated semantic descriptions at the attribute-value level to create unified metric spaces for clustering mixed tabular data, achieving scalability comparable to statistical baselines while improving clustering accuracy.

0 favorites 0 likes
#arxiv

When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-Judge Failure in Reasoning Pipelines

arXiv cs.AI · yesterday Cached

This paper shows that LLM judges embedded in reasoning pipelines often make poor decisions, and proposes Evidence-Locked Derive–Gate–Repair (EL-DGR) to constrain judge overrides with evidence certificates, improving accuracy over majority vote and first-candidate baselines.

0 favorites 0 likes
#arxiv

AndroidReality: How Far Are Mobile Agents from the Real World?

arXiv cs.AI · yesterday Cached

Introduces AndroidReality, a perturbation-based framework for evaluating and improving the robustness of mobile agents, with a taxonomy of real-world interface perturbations and a training-free Test-Time Introspective Recovery (TTIR) mechanism.

0 favorites 0 likes
#arxiv

QuantumMind: Constraint-Grounded Agentic Reasoning for Speedup Analysis in Quantum Computing

arXiv cs.AI · yesterday Cached

QuantumMind presents an auditable agentic workflow that automatically generates and screens quantum speedup hypotheses using typed role-specialized actions and a deterministic validator.

0 favorites 0 likes
#arxiv

Mendel G\"odel Machine: Recursive Self-Improving Coding Agents via Comparative Evolution

arXiv cs.AI · yesterday Cached

This paper introduces the Mendel Gödel Machine, a recursive self-improving framework that applies comparative evolution to iteratively improve coding agents.

0 favorites 0 likes
#arxiv

Agent-MD: Selective LLM Intervention with Event-Driven Escalation for Stateful GCMC--MD Campaigns

arXiv cs.AI · yesterday Cached

Agent-MD is a framework that selectively applies LLM reasoning to long-running molecular simulation campaigns, using a deterministic rule-based agent for routine tasks and event-triggered LLM review for exceptional conditions. Demonstrated in GCMC–MD water-vapor desorption simulations, it shows that auditable, reproducible scientific workflows can avoid placing every operation inside an LLM reasoning loop.

0 favorites 0 likes
#arxiv

When LLM Agents Negotiate: Private Information and Dynamic Bargaining in Supply Chains

arXiv cs.AI · yesterday Cached

This paper studies how LLM agents negotiate in a dynamic supply chain bargaining problem, benchmarking nine models from OpenAI, Google, and Alibaba against a Bayesian equilibrium and finding that capability, provider identity, and prompt design shape surplus creation and division.

0 favorites 0 likes
#arxiv

Determinization in Structure Theories: A Unified Framework via Closure, Comparability, and Joint Admissibility

arXiv cs.AI · yesterday Cached

This paper presents a formal framework for constructing canonical interpretations from plural structure theories, motivated by structural failures in LLM-assisted reasoning. It distinguishes types of non-determinism and provides conditions for licensed canonicalization, without establishing full determinization for all cases.

0 favorites 0 likes
#arxiv

Flow-by-Flow:Content-Judgment Bypass for Governing AI Output in High-Loss Domains

arXiv cs.AI · yesterday Cached

This arXiv paper proposes a 'flow-by-flow' content-judgment bypass approach to govern AI outputs in high-loss domains, aiming to reduce harm from incorrect or unsafe AI-generated content.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback