Score: 1

Interpretable EEG-to-Image Generation with Semantic Prompts

Published: July 9, 2025 | arXiv ID: 2507.07157v1

By: Arshak Rezvani , Ali Akbari , Kosar Sanjar Arani and more

Potential Business Impact:

Lets computers guess what you see from brain waves.

Business Areas:
Image Recognition Data and Analytics, Software

Decoding visual experience from brain signals offers exciting possibilities for neuroscience and interpretable AI. While EEG is accessible and temporally precise, its limitations in spatial detail hinder image reconstruction. Our model bypasses direct EEG-to-image generation by aligning EEG signals with multilevel semantic captions -- ranging from object-level to abstract themes -- generated by a large language model. A transformer-based EEG encoder maps brain activity to these captions through contrastive learning. During inference, caption embeddings retrieved via projection heads condition a pretrained latent diffusion model for image generation. This text-mediated framework yields state-of-the-art visual decoding on the EEGCVPR dataset, with interpretable alignment to known neurocognitive pathways. Dominant EEG-caption associations reflected the importance of different semantic levels extracted from perceived images. Saliency maps and t-SNE projections reveal semantic topography across the scalp. Our model demonstrates how structured semantic mediation enables cognitively aligned visual decoding from EEG.

Country of Origin
🇨🇦 Canada

Page Count
6 pages

Category
Computer Science:
CV and Pattern Recognition