adapter-architecture

标签

Cards List
#adapter-architecture

Correcting Suppressed Log-Probabilities in Language Models with Post-Transformer Adapters

arXiv cs.CL · 2026-04-20 缓存

This paper demonstrates that a small post-transformer adapter (786K parameters) can correct suppressed log-probabilities in alignment-tuned language models, particularly on politically sensitive topics. The adapter shows 31-39% generalization to held-out facts across Qwen3 models while maintaining coherent generation when applied at the final prediction position.

0 人收藏 0 人点赞
← 返回首页

提交意见反馈