Brain-CLIPLM: Decoding Compressed Semantic Representations in EEG for Language Reconstruction
Summary
Researchers propose Brain-CLIPLM, a two-stage EEG-to-text decoding framework using contrastive learning for semantic anchor extraction and a retrieval-grounded LLM with Chain-of-Thought reasoning, achieving 67.55% top-5 sentence retrieval accuracy and suggesting EEG-to-text decoding should focus on recovering compressed semantic content rather than full sentence reconstruction.
Similar Articles
Sparse Autoencoders Map Brain-LLM Alignment onto Cortical Semantic Topography
This paper uses sparse autoencoders to decompose LLMs into interpretable features and shows that semantic features explain brain alignment with cortical semantic topography, generalizing across English, Chinese, and French.
Beyond Parallel Tracking: Interactive Multi-Feature Fusion Drives Semantic Reconstruction from Non-invasive Brain Recordings
This paper introduces a multi-feature fusion framework for semantic reconstruction from non-invasive brain recordings, combining static lexical (Word2Vec) and dynamic contextual (GPT) representations via cross-attention, achieving state-of-the-art performance in brain-to-text decoding.
Encoding EEG Signals to Examine Human-Like Next-Word Prediction Behaviour in Language Models
This paper investigates whether language models' next-word prediction aligns with human cognitive processing by analyzing EEG signals and event-related potentials, finding that only surprisal correlates with human brain responses, especially for open-class words.
Interpreting Brain Responses to Language with Sparse Features from Language Models
This paper introduces Augmented Sparse Encoding Models to interpret brain responses to language using sparse features from language models, validated on high-field 7T fMRI data. It recovers known neural tuning properties and discovers a new voxel population tuned to people-related content.
The Capacity of Thought: Benchmarking Llama 3.2 in Semantic fMRI Neural Language Decoding and Improving the Huth Encoding-Model Baseline
This paper improves the Huth encoding pipeline for fMRI decoding and introduces fMRIFlamingo, a direct fMRI-to-text framework using Llama 3.2, but finds that decoding success is driven by the language prior rather than neural input.