Score: 2

UniEval: Unified Holistic Evaluation for Unified Multimodal Understanding and Generation

Published: May 15, 2025 | arXiv ID: 2505.10483v1

By: Yi Li , Haonan Wang , Qixiang Zhang and more

Potential Business Impact:

Tests AI that understands and makes pictures and words.

Business Areas:

Usability Testing Data and Analytics, Design

The emergence of unified multimodal understanding and generation models is rapidly attracting attention because of their ability to enhance instruction-following capabilities while minimizing model redundancy. However, there is a lack of a unified evaluation framework for these models, which would enable an elegant, simplified, and overall evaluation. Current models conduct evaluations on multiple task-specific benchmarks, but there are significant limitations, such as the lack of overall results, errors from extra evaluation models, reliance on extensive labeled images, benchmarks that lack diversity, and metrics with limited capacity for instruction-following evaluation. To tackle these challenges, we introduce UniEval, the first evaluation framework designed for unified multimodal models without extra models, images, or annotations. This facilitates a simplified and unified evaluation process. The UniEval framework contains a holistic benchmark, UniBench (supports both unified and visual generation models), along with the corresponding UniScore metric. UniBench includes 81 fine-grained tags contributing to high diversity. Experimental results indicate that UniBench is more challenging than existing benchmarks, and UniScore aligns closely with human evaluations, surpassing current metrics. Moreover, we extensively evaluated SoTA unified and visual generation models, uncovering new insights into Univeral's unique values.

MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models

CV and Pattern Recognition

Tests AI that understands and creates with images and words.

4 Apr 2025 0

90%

UmniBench: Unified Understand and Generation Model Oriented Omni-dimensional Benchmark

Artificial Intelligence

Tests AI that sees and creates together.

19 Dec 2025 0

89%

Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark

CV and Pattern Recognition

Tests how well AI can see and create.

15 Oct 2025 1

View PDF Login to Bookmark

Country of Origin

🇭🇰 Hong Kong

Repos / Data Links

github.com

Page Count

35 pages

UniEval: Unified Holistic Evaluation for Unified Multimodal Understanding and Generation

Tests AI that understands and makes pictures and words.

Technical Abstract

MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models

UmniBench: Unified Understand and Generation Model Oriented Omni-dimensional Benchmark

Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark