self-supervised

Tag

Cards List
#self-supervised

DREAM: Dense Retrieval Embeddings via Autoregressive Modeling

Hugging Face Daily Papers · 2026-06-23 Cached

DREAM trains dense retrieval embeddings by using autoregressive language model attention to supervise query-document similarity, eliminating the need for labeled data. It consistently outperforms baselines on BEIR and RTEB benchmarks across model scales.

0 favorites 0 likes
#self-supervised

BadWorld: Adversarial Attacks on World Models

Hugging Face Daily Papers · 2026-06-15 Cached

BadWorld is a label-free adversarial framework that reveals structural vulnerabilities in visual world models by generating imperceptible perturbations that cause catastrophic failures in future rollouts.

0 favorites 0 likes
#self-supervised

The Art of Interrogation: Consistency Amplifies Factuality in Spatial Reasoning

arXiv cs.AI · 2026-06-11 Cached

This paper proposes a self-supervised reinforcement learning framework that uses consistency verifiers—reward functions checking geometric and semantic consistency under transformations—to improve spatial reasoning in large reasoning models without requiring ground-truth annotations. The method approaches the accuracy of supervised fine-tuning and generalizes across diverse tasks.

0 favorites 0 likes
#self-supervised

Pretrained self-supervised speech models can recognize unseen consonants

arXiv cs.CL · 2026-06-11 Cached

This paper investigates whether pretrained self-supervised speech models like Wav2Vec2 and HuBERT can accurately recognize click consonants, which are rare in training data, by fine-tuning on Khoisan languages. Results show the models recognize clicks more accurately than non-clicks, indicating generalization to uncommon phonemes.

0 favorites 0 likes
#self-supervised

Multilingual Word-Level Forced Alignment with Self-Supervised Representations and Learned Dynamic Programming

arXiv cs.CL · 2026-06-10 Cached

A novel method for multilingual word-level forced alignment combines self-supervised representations from MMS and a phoneme boundary detector with a learned dynamic programming decoder, outperforming existing aligners on English and unseen languages without further training.

0 favorites 0 likes
#self-supervised

MaskAlign: Token-Subset Representation Alignment for Efficient Diffusion Training

Hugging Face Daily Papers · 2026-06-07 Cached

MaskAlign proposes a token-subset representation alignment method that improves diffusion transformer training by reducing reliance on complete token sets and maintaining stable alignment under perturbations.

0 favorites 0 likes
#self-supervised

Self-supervised User Profile Generation for Personalization

arXiv cs.CL · 2026-06-05 Cached

Introduces BUMP, a self-supervised framework for training a profile generator for LLM personalization without task labels, using bidirectional in-batch ranking and GRPO. It matches or outperforms supervised methods on the LaMP benchmark.

0 favorites 0 likes
#self-supervised

Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Trajectory Rollouts

Hugging Face Daily Papers · 2026-06-04 Cached

Retrospective Harness Optimization (RHO) is a self-supervised method that improves LLM agent performance using only past trajectories, achieving a 78% pass rate on SWE-Bench Pro without external grading.

0 favorites 0 likes
#self-supervised

Unsupervised Skill Discovery for Agentic Data Analysis

Hugging Face Daily Papers · 2026-06-04

DataCOPE is an unsupervised verifier-guided skill discovery framework for data-analytic agents that derives verifier signals from exploration trajectories without labeled supervision. It improves performance by 9.71% and 32.30% on report-style and reasoning-style data analysis tasks respectively.

0 favorites 0 likes
#self-supervised

MemTrain: Self-Supervised Context Memory Training

arXiv cs.CL · 2026-06-03 Cached

MemTrain proposes a self-supervised training framework that uses masked reconstruction and intermediate memory recall proxy tasks on Wikipedia corpora to enhance LLM agents' context memory, achieving up to 17.67 point gains on downstream memory-intensive QA benchmarks.

0 favorites 0 likes
#self-supervised

MindZero: Learning Online Mental Reasoning With Zero Annotations

arXiv cs.AI · 2026-06-02 Cached

MindZero introduces a self-supervised reinforcement learning framework that trains multimodal large language models for efficient and robust online mental reasoning without requiring mental state annotations, outperforming model-based methods in accuracy and efficiency.

0 favorites 0 likes
#self-supervised

RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video

Hugging Face Daily Papers · 2026-05-29 Cached

RayDer is a unified feed-forward transformer that consolidates camera estimation, scene reconstruction, and rendering for self-supervised novel view synthesis from real-world video, achieving clean power-law scaling and strong zero-shot performance.

0 favorites 0 likes
#self-supervised

The Flip Side of RLHF: On-Policy Feedback for Reward Model Self-Supervised Improvement

Hugging Face Daily Papers · 2026-05-29 Cached

The SAVE framework improves reward model training by using value functions to grade on-policy responses and update models through contrastive objectives, achieving outperforming results across six benchmarks.

0 favorites 0 likes
#self-supervised

ChildVox: A Speech, Audio, and Large Audio-Language Model Benchmark in Understanding and Characterizing Sound across Childhood

Hugging Face Daily Papers · 2026-05-28 Cached

ChildVox presents a comprehensive benchmark for analyzing children's acoustic communication across developmental stages, integrating over 20 sub-tasks from 17 child-centered audio and speech datasets.

0 favorites 0 likes
#self-supervised

PilotWiMAE: Pilot-Native Representation Learning for Wireless Channels

arXiv cs.AI · 2026-05-25 Cached

PilotWiMAE introduces a self-supervised framework that directly ingests noisy pilot observations for wireless channel representation learning, removing the unrealistic full-CSI assumption and enabling robust cross-frequency beam selection and channel estimation that beats supervised baselines.

0 favorites 0 likes
#self-supervised

Self-Improving In-Context Learning

arXiv cs.CL · 2026-05-25 Cached

This paper proposes a method to improve in-context learning by optimizing the continuous embeddings of a fixed few-shot prompt at test time, using a self-supervised confidence proxy derived from the model's log-probabilities without requiring fine-tuning or token generation.

0 favorites 0 likes
#self-supervised

NITP: Next Implicit Token Prediction for LLM Pre-training

Hugging Face Daily Papers · 2026-05-24 Cached

Next Implicit Token Prediction (NITP) enhances language model pre-training by adding dense continuous supervision in representation space, improving generalization and performance across model sizes with minimal computational overhead.

0 favorites 0 likes
#self-supervised

Temporal Contrastive Transformer for Financial Crime Detection: Self-Supervised Sequence Embeddings via Predictive Contrastive Coding

arXiv cs.LG · 2026-05-22 Cached

Introduces the Temporal Contrastive Transformer (TCT), a self-supervised framework for learning temporal embeddings from financial transactions for fraud detection. Achieves AUC 0.8644 with embeddings alone but does not improve over strong engineered features (AUC 0.9205 vs 0.9245), indicating learned representations overlap with existing features.

0 favorites 0 likes
#self-supervised

@stephenbtl: My talk at @aiDotEngineer is now online. I talked about our research and where @bfl_ml is heading. Thanks @swyx for the…

X AI KOLs Following · 2026-05-11 Cached

Black Forest Labs shared the evolution of the Flux series models at the AI Engineer Conference and released the SelfFlow research paper, proposing a self-supervised multimodal training method that does not require external encoders.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback