self-alignment

Tag

Cards List
#self-alignment

From SRA to Self-Flow: Data Augmentation or Self-Supervision?

Hugging Face Daily Papers · 2026-07-02 Cached

This paper investigates the mechanisms behind self-alignment methods in diffusion transformers, revealing that performance improvements from methods like Self-Flow primarily come from data augmentation along the noise dimension rather than token interactions between noise levels. The authors introduce Attention Separation to demonstrate this and propose an effective design combining self-representation alignment with dual-timestep augmentation.

0 favorites 0 likes
#self-alignment

Emergent Alignment

arXiv cs.AI · 2026-06-20 Cached

This paper introduces Emergent Alignment, a self-supervised method that endows LLMs with a conscience step to review their own outputs and uses Direct Preference Optimization to steer away from unethical behavior, enabling online alignment without external judges.

0 favorites 0 likes
#self-alignment

LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward Decomposition

arXiv cs.AI · 2026-05-26 Cached

LC-ERD is a framework that mines latent logic from LLM-generated reasoning chains to decompose global rewards into step-level signals, enabling self-evolving reasoning without human annotation. It addresses label noise, coarse supervision, and distributional collapse via variational logic potential and multi-agent value decomposition.

0 favorites 0 likes
← Back to home

Submit Feedback