alignment-tax

Tag

Cards List
#alignment-tax

MemSFT: Mitigating Alignment Tax with an External Parametric Memory

Hugging Face Daily Papers · 2026-07-28 Cached

MemSFT is a research paper proposing to mitigate the alignment tax in LLM fine-tuning by using an external parametric memory that decouples domain specialization from backbone parameter updates, enabling reuse across different LLM sizes while preserving general performance.

0 favorites 0 likes
#alignment-tax

Deeper is Not Always Better: Mitigating the Alignment Tax via Confident Layer Decoding

Hugging Face Daily Papers · 2026-06-20 Cached

This paper introduces Confident Decoding, a training-free decoding strategy that dynamically selects the most reliable intermediate layer in LLMs using entropy-guided search, mitigating the alignment tax and improving reasoning performance on benchmarks like GPQA-Diamond and Omni-MATH with negligible overhead.

0 favorites 0 likes
#alignment-tax

SDOF: Taming the Alignment Tax in Multi-Agent Orchestration with State-Constrained Dispatch

arXiv cs.AI · 2026-05-18 Cached

SDOF is a framework that treats multi-agent execution as a constrained state machine, using an online-RLHF specialized intent router and state-aware dispatcher to enforce business process stage constraints, achieving 86.5% task completion on a recruitment system with 6,000+ enterprises.

0 favorites 0 likes
← Back to home

Submit Feedback