representation-learning

Tag

Cards List
#representation-learning

CausalPOI: Spatio-Temporal Graph-Based Causal Modeling for Cold-Start POI Check-in Forecasting

arXiv cs.LG ↗ · 2026-06-05 Cached

Introduces CausalPOI, a spatio-temporal graph-based causal representation learning framework for cold-start POI check-in forecasting, which outperforms state-of-the-art baselines on real-world SafeGraph datasets.

0 favorites 0 likes
#representation-learning

Predict and Reconstruct: Joint Objectives for Self-Supervised Language Representation Learning

arXiv cs.CL ↗ · 2026-06-05 Cached

The paper proposes a hybrid pre-training objective combining JEPA latent-space prediction with MLM reconstruction for language models, showing improved embedding uniformity and semantic-lexical balance.

0 favorites 0 likes
#representation-learning

The Loss Is Not Enough: Sampling Conditions and Inductive Bias in Contrastive Representation Learning

arXiv cs.LG ↗ · 2026-06-04 Cached

This paper develops a measure-theoretic framework analyzing when contrastive learning recovers meaningful latent geometry, introducing a 'diversity condition' on positive-pair sampling and a support-corrected InfoNCE variant, with experiments validating that sampling diversity and architectural inductive bias interact critically in contrastive representation learning.

0 favorites 0 likes
#representation-learning

A Geometric View of Counterfactual Behavior: Interaction of Boundary Proximity and Local Support

arXiv cs.LG ↗ · 2026-06-04 Cached

This paper examines counterfactual behavior in ML models through a geometric lens, showing that models with similar predictive performance can differ substantially in counterfactual outcomes due to the interaction between decision-boundary proximity and local data support. The findings identify counterfactual behavior as a distinct dimension from predictive performance, with implications for model selection and reliability of counterfactual explanation methods.

0 favorites 0 likes
#representation-learning

Dual Advantage Fields

arXiv cs.LG ↗ · 2026-06-04 Cached

Dual Advantage Fields (DAF) is a policy-extraction method for offline goal-conditioned RL that converts a bilinear dual value model into a local advantage signal by learning an action-effect model predicting feature displacement and scoring actions by alignment with the goal direction. Accepted at the ICML 2026 Workshop on Decision Making, DAF shows improved performance on OGBench locomotion, manipulation, and puzzle tasks.

0 favorites 0 likes
#representation-learning

KODA: Contrastive Representation Comparison and Alignment for Vision-Language Foundation Models

arXiv cs.LG ↗ · 2026-06-04 Cached

This paper introduces KODA (Kernel Optimization for Discrepancy Analysis), a kernel-based framework for comparing and aligning vision-language model representations by identifying sample subsets that are clustered differently across models like CLIP, SigLIP, and BLIP. The method uses contrastive embedding clustering and randomized low-dimensional approximations to scale to large datasets while providing interpretable structural differences between representations.

0 favorites 0 likes
#representation-learning

Bayes-Sufficient Representations in Supervised Learning

arXiv cs.LG ↗ · 2026-06-04 Cached

This paper formalizes the concept of Bayes-sufficient representations in supervised learning, defining when a representation retains exactly the information needed for Bayes-optimal prediction under a given loss function. It introduces the Bayes quotient as a canonical loss-dependent object and connects the framework to property elicitation, illustrating distinctions between sufficiency, minimality, and excess retained information through experiments.

0 favorites 0 likes
#representation-learning

OPRD: On-Policy Representation Distillation

Hugging Face Daily Papers ↗ · 2026-06-04

OPRD proposes a new knowledge distillation method that aligns student and teacher hidden states across layers during on-policy rollouts, eliminating sampling variance from token-space KL estimation. Empirically, OPRD outperforms output-space baselines on math reasoning benchmarks (AIME 2024/2025, AIMO) while being 1.44x faster and using 54% less memory.

0 favorites 0 likes
#representation-learning

Neural Networks Provably Learn Spectral Representations for Group Composition

arXiv cs.LG ↗ · 2026-06-03 Cached

This paper theoretically demonstrates that two-layer neural networks trained on group composition tasks learn spectral representations, with neurons converging to irreducible representations and achieving rotational rank-one alignment, providing a representation-theoretic account of feature learning.

0 favorites 0 likes
#representation-learning

Forgetting is Not Erasure: Recovering Latent Knowledge via Transport Keys

arXiv cs.LG ↗ · 2026-06-03 Cached

This paper argues that catastrophic forgetting in neural networks is not erasure but an interface alignment problem. It introduces 'transport keys' to recover latent task-specific features from sequentially trained models, demonstrating significant performance recovery on split CIFAR-100.

0 favorites 0 likes
#representation-learning

QUIVER: Quantum-Informed Views for Enhanced Representations in Large ML Models

arXiv cs.LG ↗ · 2026-06-03 Cached

This paper introduces Quiver, a paradigm that enriches classical machine learning models with quantum-inspired features derived from the quantum Fisher information matrix, demonstrating improvements on molecule property prediction and jet flavor classification benchmarks.

0 favorites 0 likes
#representation-learning

Before Fusion, Ask What to Keep: Contextual Calibration of Multimodal Signals

arXiv cs.LG ↗ · 2026-06-03 Cached

This paper introduces a plug-in calibration module that adjusts multimodal representations before fusion, using cross-modal context to suppress misleading signals and emphasize reliable ones, improving performance on multiple benchmarks.

0 favorites 0 likes
#representation-learning

BRepCLIP: Contrastive Multimodal Pretraining on BRep Primitives for CAD Understanding

Hugging Face Daily Papers ↗ · 2026-06-03 Cached

BRepCLIP introduces contrastive multimodal pretraining on boundary representation (BRep) primitives for CAD understanding, aligning BRep geometry with language and image embeddings to achieve state-of-the-art retrieval and zero-shot classification.

0 favorites 0 likes
#representation-learning

The role of class encoding in neural collapse

arXiv cs.LG ↗ · 2026-06-02 Cached

This paper investigates how class label encoding influences neural collapse in neural network classifiers, showing that with one-hot encoding and balanced data, uncentered mean features transition from a simplex equiangular tight frame to an orthogonal frame as bias regularization increases.

0 favorites 0 likes
#representation-learning

When Softmax Fails at the Top: Extreme Value Corrections for InfoNCE

arXiv cs.LG ↗ · 2026-06-02 Cached

The paper identifies a misalignment between the softmax-based InfoNCE loss and the normalized embedding setting in modern contrastive learning. It proposes WEINCE, a simple modification that blends softmax logits with an endpoint shortfall correction using extreme value theory, yielding consistent improvements across vision benchmarks.

0 favorites 0 likes
#representation-learning

DLLM-JEPA: Joint Embedding Predictive Architectures for Masked Diffusion Language Models

arXiv cs.CL ↗ · 2026-06-02 Cached

Introduces DLLM-JEPA, a JEPA formulation for masked diffusion language models that constructs two views from a single input via the diffusion noise schedule, reducing training FLOPs by 33% relative to LLM-JEPA and improving fine-tuning performance on tasks like GSM8K.

0 favorites 0 likes
#representation-learning

Neural Networks Provably Learn Spectral Representations for Group Composition

Hugging Face Daily Papers ↗ · 2026-06-02

This paper provides a theoretical analysis of how neural networks learn structured representations during group composition tasks, proving that training dynamics drive neurons to converge to irreducible group representations with exponential convergence rates. The work establishes a representation-theoretic account of feature learning and characterizes a low-rank compression phenomenon for matrix-valued group representations.

0 favorites 0 likes
#representation-learning

Bridging the Gap Between Natural Language and Market Dynamics via High-Dimensional Representation Learning

arXiv cs.LG ↗ · 2026-06-01 Cached

This paper explores replacing scalar sentiment scores with high-dimensional FinBERT embeddings in a Transformer-based architecture for short-term stock price prediction, showing improved accuracy with Siamese-optimized embeddings.

0 favorites 0 likes
#representation-learning

Parametric Social Identity Injection and Diversification in Public Opinion Simulation

Hugging Face Daily Papers ↗ · 2026-06-01 Cached

This paper proposes Parametric Social Identity Injection (PSII), a framework that injects parametric representations of demographic attributes into LLM hidden states to improve diversity in public opinion simulation. Experiments on the World Values Survey show it reduces KL divergence and enhances diversity compared to prompt-based methods.

0 favorites 0 likes
#representation-learning

Learning Robust and Task-Invariant Functional Representation from fMRI through Siamese Self-Supervised Learning

arXiv cs.LG ↗ · 2026-05-29 Cached

This paper introduces BrainSimSiam, a lightweight self-supervised framework using siamese networks to learn robust fMRI representations from positive-only pairs, achieving strong performance on downstream tasks even with limited data.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback