CogSci 2025

•

August 02, 2025

•

San Francisco, United States

keywords:

semantics of language

cognitive neuroscience

fmri

natural language processing

Brain-to-Image reconstruction aims to recover visual stimuli perceived by humans from brain activity. However, the reconstructed visual stimuli often missing details and semantic inconsistencies, which may be attributed to insufficient semantic information. To address this issue, we propose an approach named Fine-grained Brain-to-Image reconstruction (FgB2I), which employs fine-grained text as bridge to improve image reconstruction. FgB2I comprises three key stages: detail enhancement, decoding fine-grained text descriptions, and text-bridged brain-to-image reconstruction. In the detail-enhancement stage, we leverage large vision–language models to generate fine-grained captions for visual stimuli and experimentally validate its importance. We propose three reward metrics (object accuracy, text-image semantic similarity, and image-image semantic similarity) to guide the language model in decoding fine-grained text descriptions from fMRI signals. The fine-grained text descriptions can be integrated into existing reconstruction methods to achieve fine-grained Brain-to-Image reconstruction.

Downloads

PaperTranscript English (automatic)

Next from CogSci 2025

Dual-Path Parallel Graph Convolution Combining Brain Region Partitioning and Data-Driven Learning for EEG Emotion Recognition
poster

Dual-Path Parallel Graph Convolution Combining Brain Region Partitioning and Data-Driven Learning for EEG Emotion Recognition

CogSci 2025

+2
Zirui Xiang and 4 other authors

02 August 2025

Similar lecture

MapGuide: A Simple yet Effective Method to Reconstruct Continuous Language from Brain Activities
poster

MapGuide: A Simple yet Effective Method to Reconstruct Continuous Language from Brain Activities

NAACL 2024

+3
Xinpei Zhao and 5 other authors

14 June 2024