Human-like object concept representations emerge naturally in multimodal large language models

📅 2024-07-01
🏛️ arXiv.org
📈 Citations: 1
Influential: 0
📄 PDF
🤖 AI Summary
This study investigates whether multimodal large language models (MLLMs) spontaneously develop human-like object concept representations. Method: We conducted behavioral triplet-judgment experiments to construct a 66-dimensional semantic embedding space covering 1,854 natural objects, and performed functional magnetic resonance imaging (fMRI) representational similarity analysis (RSA) to systematically assess alignment between MLLM embeddings and neural activity in human visual cortex regions—including the fusiform face area (FFA), parahippocampal place area (PPA), and extrastriate body area (EBA). Contribution/Results: We report the first evidence that MLLM-implicit conceptual structure closely matches human semantic intuition and functional brain organization: significant neural alignment (p < 0.001, R² ≥ 0.62) is observed across key category-selective regions; moreover, these representations are stable and predictive. This work provides the first cross-modal, empirically testable evidence for deep commonalities between machine cognition and human perception.

Technology Category

Application Category

📝 Abstract
The conceptualization and categorization of natural objects in the human mind have long intrigued cognitive scientists and neuroscientists, offering crucial insights into human perception and cognition. Recently, the rapid development of Large Language Models (LLMs) has raised the attractive question of whether these models can also develop human-like object representations through exposure to vast amounts of linguistic and multimodal data. In this study, we combined behavioral and neuroimaging analysis methods to uncover how the object concept representations in LLMs correlate with those of humans. By collecting large-scale datasets of 4.7 million triplet judgments from LLM and Multimodal LLM (MLLM), we were able to derive low-dimensional embeddings that capture the underlying similarity structure of 1,854 natural objects. The resulting 66-dimensional embeddings were found to be highly stable and predictive, and exhibited semantic clustering akin to human mental representations. Interestingly, the interpretability of the dimensions underlying these embeddings suggests that LLM and MLLM have developed human-like conceptual representations of natural objects. Further analysis demonstrated strong alignment between the identified model embeddings and neural activity patterns in many functionally defined brain ROIs (e.g., EBA, PPA, RSC and FFA). This provides compelling evidence that the object representations in LLMs, while not identical to those in the human, share fundamental commonalities that reflect key schemas of human conceptual knowledge. This study advances our understanding of machine intelligence and informs the development of more human-like artificial cognitive systems.
Problem

Research questions and friction points this paper is trying to address.

Do LLMs develop human-like object representations from data
Compare model embeddings with human neural activity patterns
Explore similarity between AI and human conceptual knowledge
Innovation

Methods, ideas, or system contributions that make the work stand out.

Multimodal LLMs derive object similarity embeddings
Behavioral and neuroimaging analyses validate model representations
Model embeddings align with human neural activity patterns
🔎 Similar Papers
Changde Du
Changde Du
Institute of Automation, Chinese Academy of Sciences
machine learningcomputer visioncomputational neurosciencebrain-computer interface(BCI)artificial intelligence
K
Kaicheng Fu
Laboratory of Brain Atlas and Brain-inspired Intelligence, Institute of Automation, Chinese Academy of Sciences; Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology, Chinese Academy of Sciences; University of Chinese Academy of Sciences
B
Bincheng Wen
Laboratory of Brain Atlas and Brain-inspired Intelligence, Institute of Automation, Chinese Academy of Sciences; Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology, Chinese Academy of Sciences; Center for Excellence in Brain Science and Intelligence Technology, Institute of Neuroscience, Chinese Academy of Sciences; University of Chinese Academy of Sciences
Y
Yi Sun
Laboratory of Brain Atlas and Brain-inspired Intelligence, Institute of Automation, Chinese Academy of Sciences; Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology, Chinese Academy of Sciences; University of Chinese Academy of Sciences
J
Jie Peng
Laboratory of Brain Atlas and Brain-inspired Intelligence, Institute of Automation, Chinese Academy of Sciences; Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology, Chinese Academy of Sciences; University of Chinese Academy of Sciences
W
Wei Wei
Laboratory of Brain Atlas and Brain-inspired Intelligence, Institute of Automation, Chinese Academy of Sciences; Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology, Chinese Academy of Sciences; University of Chinese Academy of Sciences
Ying Gao
Ying Gao
Shell, Imperial College London
Pore scale imagingReservoir engineeringMultiphase flow in porous mediaX-ray imaging
S
Shengpei Wang
Laboratory of Brain Atlas and Brain-inspired Intelligence, Institute of Automation, Chinese Academy of Sciences; Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology, Chinese Academy of Sciences; University of Chinese Academy of Sciences
C
Chuncheng Zhang
Laboratory of Brain Atlas and Brain-inspired Intelligence, Institute of Automation, Chinese Academy of Sciences; Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology, Chinese Academy of Sciences; University of Chinese Academy of Sciences
J
Jinpeng Li
School of Automation Science and Engineering, South China University of Technology
Shuang Qiu
Shuang Qiu
City University of Hong Kong
Reinforcement LearningAgentic AILarge Language ModelsEmbodied AI
L
Le Chang
Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology, Chinese Academy of Sciences; Center for Excellence in Brain Science and Intelligence Technology, Institute of Neuroscience, Chinese Academy of Sciences; University of Chinese Academy of Sciences
Huiguang He
Huiguang He
Institute of Automation, Chinese Academy of Scineces
Artificial Intelligencemedical image processingBrain Computer Interface