frozen-llm

Tag

Cards List
#frozen-llm

LODESTAR: Trustworthy Entropy Is Navigated, Not Merely Measured -- Reinforced Polarizer Keeps a Frozen LLM from Being Confidently Misled by the Wrong Evidence

arXiv cs.CL · 2026-08-13 Cached

This paper introduces Lodestar, a method that uses reinforcement learning to train a short polarizer prompt string that helps a frozen LLM avoid being misled by misleading retrieved passages in RAG question answering. It improves F1 and exact match scores across five QA benchmarks compared to existing entropy-based selection rules.

0 favorites 0 likes
#frozen-llm

Hierarchical Self-Improvement: A Framework for Task-Specific Evolvable Agent Harnesses

Hugging Face Daily Papers · 2026-08-09 Cached

The paper introduces Hierarchical Self-Improvement (HSI), a framework that enhances frozen LLM agents by evolving task-specific harnesses through hierarchical self-modification, achieving substantial gains on moderate tasks while being limited by feedback quality and backbone capabilities.

0 favorites 0 likes
#frozen-llm

IRIS: Reusable Identity Representations from Frozen LLMs for Entity Alignment

arXiv cs.CL · 2026-07-29 Cached

IRIS is a training-free framework that uses frozen large language models to construct reusable identity representations for entities in knowledge graphs, enabling efficient entity alignment across different KGs without pair-dependent processing.

0 favorites 0 likes
#frozen-llm

Latent Bridges for Multi-Table Question Answering

arXiv cs.CL · 2026-06-30 Cached

GRAB uses a GNN encoder to convert relational tables into latent tokens for frozen LLMs, achieving significant performance gains in multi-table question answering.

0 favorites 0 likes
#frozen-llm

PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents

arXiv cs.CL · 2026-05-12 Cached

This paper introduces PYTHALAB-MERA, an external controller for frozen local LLMs that uses validation-grounded memory and retrieval to improve coding agent performance. It demonstrates superior success rates in strict validation tasks compared to self-refinement baselines by leveraging execution feedback and temporal difference learning.

0 favorites 0 likes
← Back to home

Submit Feedback