Score: 0

Quantum Visual Word Sense Disambiguation: Unraveling Ambiguities Through Quantum Inference Model

Published: December 31, 2025 | arXiv ID: 2512.24687v1

By: Wenbo Qiao, Peng Zhang, Qinghua Hu

Visual word sense disambiguation focuses on polysemous words, where candidate images can be easily confused. Traditional methods use classical probability to calculate the likelihood of an image matching each gloss of the target word, summing these to form a posterior probability. However, due to the challenge of semantic uncertainty, glosses from different sources inevitably carry semantic biases, which can lead to biased disambiguation results. Inspired by quantum superposition in modeling uncertainty, this paper proposes a Quantum Inference Model for Unsupervised Visual Word Sense Disambiguation (Q-VWSD). It encodes multiple glosses of the target word into a superposition state to mitigate semantic biases. Then, the quantum circuit is executed, and the results are observed. By formalizing our method, we find that Q-VWSD is a quantum generalization of the method based on classical probability. Building on this, we further designed a heuristic version of Q-VWSD that can run more efficiently on classical computing. The experiments demonstrate that our method outperforms state-of-the-art classical methods, particularly by effectively leveraging non-specialized glosses from large language models, which further enhances performance. Our approach showcases the potential of quantum machine learning in practical applications and provides a case for leveraging quantum modeling advantages on classical computers while quantum hardware remains immature.

Integrating Symbolic Natural Language Understanding and Language Models for Word Sense Disambiguation

Computation and Language

Helps computers understand words with many meanings.

20 Nov 2025 0

87%

SANDWiCH: Semantical Analysis of Neighbours for Disambiguating Words in Context ad Hoc

Computation and Language

Helps computers understand words like people do.

7 Mar 2025 3

87%

Building Reasonable Inference for Vision-Language Models in Blind Image Quality Assessment

CV and Pattern Recognition

Makes AI judge picture quality more like people.

10 Dec 2025 0

View PDF Login to Bookmark

Quantum Visual Word Sense Disambiguation: Unraveling Ambiguities Through Quantum Inference Model

Technical Abstract

Integrating Symbolic Natural Language Understanding and Language Models for Word Sense Disambiguation

SANDWiCH: Semantical Analysis of Neighbours for Disambiguating Words in Context ad Hoc

Building Reasonable Inference for Vision-Language Models in Blind Image Quality Assessment