representations

Tag

Cards List
#representations

Transformer Geometry Observatory TGO-IV: Developmental Topology Observatory

arXiv cs.LG · yesterday Cached

This paper presents TGO-IV, a topological framework using persistent homology to analyze how transformer representations evolve across layers, complementing prior spectral and geometric observatories.

0 favorites 0 likes
#representations

Attention-based representations for multi-task computation

arXiv cs.LG · 2026-08-06 Cached

This paper establishes theoretical bounds on the number of attention heads needed to produce vector representations that support multiple tasks, such as computing min/max and XOR, showing trade-offs between head count, embedding dimension, and precision.

0 favorites 0 likes
#representations

From Representations to Behaviors: Exploring the Person-Situation-Behavior Triad in LLMs

arXiv cs.CL · 2026-07-30 Cached

This paper adapts Funder's personality triad framework to LLMs, using sparse autoencoders to discover and validate trait-like internal representations, and demonstrating controllable bidirectional behavioral shifts through feature-level interventions.

0 favorites 0 likes
#representations

The strength of clinical evidence is recoverable from language model representations but not from their stated grades

arXiv cs.CL · 2026-06-30 Cached

This paper demonstrates that large language models internally encode the strength of clinical evidence for claims, yet fail to accurately express this strength when asked, with stated evidence grades performing near chance.

0 favorites 0 likes
#representations

When Probing Accuracy Saturates, Fragility Resolves: A Complementary Metric for LLM Pre-Training Analysis

arXiv cs.CL · 2026-06-11 Cached

This paper introduces 'fragility', a complementary metric to probe accuracy that measures activation-noise level at which probe accuracy collapses, enabling analysis of representation evolution during LLM pre-training even after accuracy saturates.

0 favorites 0 likes
#representations

On the Persistent Effects of Lexicality in Large Language Mod

arXiv cs.CL · 2026-06-03 Cached

This paper investigates how lexical overlap, rather than semantic content, influences LLM representations across layers and architectures, and demonstrates that this lexical effect persists even in models trained for semantic similarity, leading to degraded performance on downstream tasks.

0 favorites 0 likes
#representations

Platonic Representations in the Human Brain: Unsupervised Recovery of Universal Geometry

Hugging Face Daily Papers · 2026-05-19 Cached

This paper investigates whether fMRI representations from different subjects' visual cortices can be aligned using unsupervised geometric methods, finding evidence for approximately isometric structure across individuals, extending the Platonic Representation Hypothesis to human brains.

0 favorites 0 likes
#representations

@Julian_a42f9a: Late-interaction retrieval models are widely used for their strong performance, but their representations can be utiliz…

X AI KOLs Following · 2026-04-17 Cached

A new paper shows that late-interaction retrieval model representations can effectively replace raw document text in RAG tasks, extending their utility beyond retrieval.

0 favorites 0 likes
← Back to home

Submit Feedback