representation-learning

Tag

Cards List
#representation-learning

ViQ: Text-Aligned Visual Quantized Representations at Any Resolution

Hugging Face Daily Papers ↗ · 2026-06-25 Cached

ViQ presents a visual quantization framework that balances semantic richness and detail preservation in discrete representations, enabling efficient multimodal training with native-resolution inputs by using text-aligned pre-training and proximal representation learning.

0 favorites 0 likes
#representation-learning

@k_solidified_: https://arxiv.org/abs/2106.10165 All of humanity should read this

X AI KOLs Timeline ↗ · 2026-06-24 Cached

This book develops an effective theory for deep neural networks, showing that their predictions are nearly-Gaussian and governed by the depth-to-width ratio, and introduces representation group flow to analyze signal propagation and learning dynamics.

0 favorites 0 likes
#representation-learning

PORTER: Language-Grounded Event Representations for Portable Structured EHR Foundation Models

arXiv cs.CL ↗ · 2026-06-24 Cached

PORTER is a language-grounded structured EHR foundation model that represents clinical events through text descriptions and numeric values, enabling vocabulary-independent transfer across institutions without retraining. On pediatric prediction tasks, PORTER matches fixed-vocabulary models and recovers 97.1% of AUROC when transferred to unseen event descriptions.

0 favorites 0 likes
#representation-learning

The Galaxy's Guide to the Tokenizer: A Benchmark for Scientific Foundation Models

Hugging Face Daily Papers ↗ · 2026-06-24 Cached

This paper compares four tokenization methods (Affine, AIM, JetFormer, VQ-VAE) for astronomical images within a unified transformer framework, using 640,000 galaxy images to evaluate reconstruction quality, physical property prediction, and morphological preservation. It finds that no single method excels across all tasks, highlighting trade-offs in representation learning.

0 favorites 0 likes
#representation-learning

@ChrisInterno: Signals of physical plausibility are hiding in the geometry of frozen image encoders. No video training. No physics sup…

X AI KOLs Following ↗ · 2026-06-21 Cached

The tweet highlights a research finding that signals of physical plausibility can be extracted from the geometry of frozen image encoders without video training or physics supervision.

0 favorites 0 likes
#representation-learning

DVD-JEPA: an open-source, fully-reproducible JEPA world model [P]

Reddit r/MachineLearning ↗ · 2026-06-20

DVD-JEPA is an open-source, minimal JEPA world model that learns representations from video by predicting future embeddings rather than pixels. It uses a bouncing DVD logo to demonstrate position recovery, dreaming, and anomaly detection, all running in a browser.

0 favorites 0 likes
#representation-learning

EvoEmbedding: Evolvable Representations for Long-Context Retrieval and Agentic Memory

Hugging Face Daily Papers ↗ · 2026-06-19 Cached

EvoEmbedding is a dynamic embedding model that maintains a continuously updated latent memory to generate adaptive representations for long-context retrieval, outperforming larger specialist models and improving agentic workflows.

0 favorites 0 likes
#representation-learning

PoLAR: Factorizing Extent and Mode in Latent Actions for Robot Policy Learning

Hugging Face Daily Papers ↗ · 2026-06-19 Cached

PoLAR introduces a geometrically structured latent action representation in hyperbolic space that separates transition extent from mode, improving robotic policy learning performance.

0 favorites 0 likes
#representation-learning

Beyond Tokenization: Direct Timestep Embedding and Contrastive Alignment for Time-Series Question Answering

arXiv cs.CL ↗ · 2026-06-18 Cached

This paper introduces CADE, a framework for time-series question answering that maps each timestep directly into the LLM embedding space and uses a one-directional supervised contrastive loss to align time-series representations with frozen text anchors, outperforming existing baselines on the Time-MQA benchmark.

0 favorites 0 likes
#representation-learning

Concept Modulation Models: A Unified Framework for Identifiability and Extrapolation

arXiv cs.LG ↗ · 2026-06-18 Cached

This paper introduces concept modulation models (CMMs), a unified framework for identifiability and extrapolation in conditional generative models. It shows that feature agreement on observed attributes induces constraints through attribute potentials, enabling algebraic extrapolation criteria that recover and generalize existing results.

0 favorites 0 likes
#representation-learning

Structured Representation Learning with Locally Linear Embeddings and Adaptive Feature Fusion

arXiv cs.LG ↗ · 2026-06-18 Cached

Proposes a reinforcement learning framework that uses locally linear embeddings to capture environment dynamics and an attention mechanism to adaptively fuse dynamics-specific and reward-specific features, inspired by neural principles, improving learning efficiency.

0 favorites 0 likes
#representation-learning

When, Where, and How: Adaptive Binning for Tabular Self-Supervised Learning

Hugging Face Daily Papers ↗ · 2026-06-18 Cached

This paper proposes Adaptive Binning, a learning-coupled feature-wise coarse-to-fine curriculum for tabular self-supervised learning that adaptively discretizes features, improving representations on medical datasets and establishing a unified benchmark.

0 favorites 0 likes
#representation-learning

Next-Latent Prediction Transformers [R]

Reddit r/MachineLearning ↗ · 2026-06-17

Microsoft Research introduces Next-Latent Prediction (NextLat), a self-supervised method that trains transformers to predict their own next latent state, enabling compact world models for reasoning and planning and achieving up to 3.3x faster inference via self-speculative decoding.

0 favorites 0 likes
#representation-learning

Learning task-specific subspaces via interventional post-training of speech foundation models

arXiv cs.CL ↗ · 2026-06-17 Cached

This paper proposes a post-training refinement approach using interventional contrastive learning to disentangle speech foundation model representations into separate content and speaker subspaces. The method shows improved out-of-domain speaker verification performance and evidence of successful separation.

0 favorites 0 likes
#representation-learning

MoCo-AIS: A Contrastive Learning Framework for Similarity Computation of Vessel Trajectories

arXiv cs.AI ↗ · 2026-06-17 Cached

MoCo-AIS is a unified contrastive learning framework for computing similarity of vessel trajectories, evaluated on large-scale AIS datasets.

0 favorites 0 likes
#representation-learning

Probing, Fusion, and Trustworthiness: A Systematic Evaluation of Foundation Model Representations for Multimodal Cancer Analysis

arXiv cs.LG ↗ · 2026-06-17 Cached

This paper systematically evaluates foundation model representations for multimodal cancer analysis, benchmarking unimodal and multimodal fusion strategies on real-world cohorts, and assessing trustworthiness via conformal prediction.

0 favorites 0 likes
#representation-learning

@AlexiGlad: Progress in AI is driven by approaches that make weaker assumptions, which allows for better scaling But representation…

X AI KOLs Following ↗ · 2026-06-16 Cached

Introduces Temporal Difference in Vision (TDV), a new paradigm for representation learning that relies solely on causality, eliminating the need for augmentations, masking, or cropping, and matches state-of-the-art methods like DINO and iBOT on dense spatial tasks.

0 favorites 0 likes
#representation-learning

@ninaddaithankar: Can a vision model learn to see with no augmentations, no masking, no cropping, no reconstruction? It can! Introducing …

X AI KOLs Timeline ↗ · 2026-06-16 Cached

Introduces Temporal Difference in Vision (TDV), a novel visual representation learning paradigm that learns useful representations without augmentations, masking, cropping, or reconstruction, and matches state-of-the-art methods on dense spatial tasks.

0 favorites 0 likes
#representation-learning

Overcoming the Impedance Mismatch: A Theoretical Roadmap for Fusing Foundation Models and Knowledge Graphs

arXiv cs.AI ↗ · 2026-06-16 Cached

This paper formalizes the 'Impedance Mismatch' between foundation models and knowledge graphs, and proposes a theoretical roadmap for neuro-symbolic fusion using structured residual streams, vector symbolic architectures, and orthogonal subspace editing.

0 favorites 0 likes
#representation-learning

AI Engram: In Search of Memory Traces in Artificial Intelligence

arXiv cs.AI ↗ · 2026-06-16 Cached

Introduces a geometric framework to identify 'AI engrams' – memory traces in deep neural networks – formalizing neuroscientific criteria into a closed-form estimator, enabling surgical memory manipulation in models from MLPs to LLMs.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback