返回
A self-supervised domain-general learning framework for human ventral stream representation
DOI:10.1038/s41467-022-28091-4.png)
摘要
En 中文
Anterior regions of the ventral visual stream encode substantial information about object categories. Are top-down category-level forces critical for arriving at this representation, or can this representation be formed purely through domain-general learning of natural image structure? Here we present a fully self-supervised model which learns to represent individual images, rather than categories, such that views of the same image are embedded nearby in a low-dimensional feature space, distinctly from other recently encountered views. We find that category information implicitly emerges in the local similarity structure of this feature space. Further, these models learn hierarchical features which capture the structure of brain responses across the human ventral visual stream, on par with category-supervised models. These results provide computational support for a domain-general framework guiding the formation of visual representation, where the proximate goal is not explicitly about category information, but is instead to learn unique, compressed descriptions of the visual world. It is unknown whether object category learning can be formed purely through domain general learning of natural image structure. Here the authors show that human visual brain responses to objects are well-captured by self-supervised deep neural network models trained without labels, supporting a domain-general account.
Keyword:
DEEP NEURAL-NETWORKS
FUNCTIONAL ARCHITECTURE
TEMPORAL CORTEX
VISUAL FEATURES
HUMAN BRAIN
RESPONSES
ORGANIZATION
VISION
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
15.7
论文数:
9.4W
被引数:
91.2W
机构
引用论文
More tricks with tetramers: a practical guide to staining T cells with peptide–MHC multimers
Immunology
IF0
Conceptual Distinctiveness Supports Detailed Visual Long-Term Memory for Real-World Objects概念独特性支持真实世界对象的详细视觉长期记忆

