NeuroCogMap Reveals Cognitive Organization of Large Language Models
Summary
NeuroCogMap is a cognitive neuroscience-inspired framework that maps the internal features of large language models into functional parcels, linking them to interpretable cognitive functions and revealing signatures of model failures like hallucination and bias.
View Cached Full Text
Cached at: 07/14/26, 04:13 AM
Paper page - NeuroCogMap Reveals Cognitive Organization of Large Language Models
Source: https://huggingface.co/papers/2607.00397 Published on Jul 1
·
Submitted byhttps://huggingface.co/Jeryi
SunZXon Jul 14
Authors:
,
,
,
,
,
,
,
,
,
,
,
,
Abstract
Understandinghowcomplexcognitivefunctionsareorganizedwithinartificialsystemsiscentraltointerpretinglargelanguagemodels(LLMs)andrelatingthemtobiologicalcognition.YetalthoughLLMsexhibitbroadcognitive-likebehaviours,itremainsunclearwhethertheirinternalrepresentationsformreproduciblefunctionalsystemsthatexplainbehaviour,failureandlinkstohumancognition.HerewepresentNeuroCogMap,acognitiveneuroscience-inspiredframeworkthatorganizesinternalfeaturesofLLMsintofunctionalparcelsandlinksthemtointerpretablefunctions,cognitivecapabilitiesandacognitivehierarchy.Theseparcelsformastableandsemanticallycoherentorganizationthatispartlyconservedacrossmodelsandfunctionallylinkedtomodeloutputs.Withinthisorganization,majorLLMfailures,includinghallucination,bias,refusalfailureandsycophancy,correspondtodistinctdisruptionsinrepresentationalandbehavioural-controlsystems,yieldinginternalsignaturesformechanism-guideddetectionandtargetedintervention.Beyondmodelbehaviour,NeuroCogMapimprovespredictionofhumancorticalresponsesduringnaturalisticlanguagecomprehension,withthestrongestcorrespondenceinhigher-orderassociationcortex.Atthecognitivelevel,itsinternalsignaturesexposelatentstrategiesthatguiderefinementsofclassicalmodelsofhumandecision-making.Together,thesefindingsestablishNeuroCogMapasasystem-levelframeworkformappingfunctionalorganizationinartificialsystemsandforrelatingthisorganizationtohumancorticalfunctionandcognitivebehaviour.
View arXiv pageView PDFProject pageAdd to collection
Models citing this paper0
No model linking this paper
Cite arxiv.org/abs/2607.00397 in a model README.md to link it from this page.
Datasets citing this paper0
No dataset linking this paper
Cite arxiv.org/abs/2607.00397 in a dataset README.md to link it from this page.
Spaces citing this paper0
No Space linking this paper
Cite arxiv.org/abs/2607.00397 in a Space README.md to link it from this page.
Collections including this paper0
No Collection including this paper
Add this paper to acollectionto link it from this page.
Similar Articles
Modular Cognitive Architecture Emerges in Large Language Models
This paper investigates whether modular cognitive architecture emerges in large language models, finding that LLMs develop specialized neural networks mirroring human brain organization across cognitive domains, suggesting modularity is a fundamental property of intelligent systems.
CogArena: A Multimethod Evaluation of Cognitive Ability Structure in Large Language Models
CogArena introduces a procedurally generated 13-paradigm benchmark to evaluate whether LLMs exhibit separable cognitive abilities or a single general competence, finding only weak support for stable five-dimensional profiles across 55 models.
CogGym: Towards Large-Scale Comparative Evaluation of Human and Machine Cognition
CogGym is a scalable framework for comparing human and AI cognition using cognitive experiments, revealing that larger language models better mimic human reasoning but still lag behind formal benchmarks.
Heterogeneous Neural Predictivity from Language Models During Naturalistic Comprehension
This paper investigates how language model representations predict neural activity during naturalistic language comprehension across MEG, ECoG, and other recordings. The findings demonstrate that language model features serve as useful neural predictors, but caution against overinterpreting predictive success as evidence for shared neural organization.
Decomposing and Steering Functional Metacognition in Large Language Models
This research paper investigates functional metacognition in Large Language Models, demonstrating that internal states like evaluation awareness and self-assessed capability are linearly decodable from residual stream activations. The authors propose a mechanistic framework to steer these states, showing causal control over reasoning behaviors, verbosity, and safety responses.