Abstract
This paper presents a summary of the VQualA 2025 Challenge on Visual Quality Comparison for Large Multimodal Models (LMMs), hosted as part of the ICCV 2025 Work-shop on Visual Quality Assessment. The challenge aims to evaluate and enhance the ability of state-of-the-art LMMs to perform open-ended and detailed reasoning about visual quality differences across multiple images. To this end, the competition introduces a novel benchmark comprising thousands of coarse-to-fine grained visual quality comparison tasks, spanning single images, pairs, and multi-image groups. Each task requires models to provide accurate quality judgments. The competition emphasizes holistic evaluation protocols, including 2AFC-based binary preference and multi-choice questions (MCQs). Around 100 participants submitted entries, with five models demonstrating the emerging capabilities of instruction-tuned LMMs on quality assessment. This challenge marks a significant step toward open-domain visual quality reasoning and comparison and serves as a catalyst for future research on inter-pretable and human-aligned quality evaluation systems. © 2025 IEEE.
| Original language | English |
|---|---|
| Title of host publication | Proceedings - 2025 IEEE/CVF International Conference on Computer Vision Workshops |
| Subtitle of host publication | ICCV-W 2025 |
| Publisher | IEEE |
| Pages | 3383-3393 |
| ISBN (Electronic) | 9798331589882 |
| ISBN (Print) | 979-8-3315-8989-9 |
| DOIs | |
| Publication status | Published - Oct 2025 |
| Event | 2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCV-W 2025) - Honolulu, United States Duration: 19 Oct 2025 → 23 Oct 2025 https://iccv.thecvf.com/ |
Publication series
| Name | Proceedings - IEEE/CVF International Conference on Computer Vision Workshops, ICCV-W |
|---|---|
| ISSN (Print) | 2473-9936 |
| ISSN (Electronic) | 2473-9944 |
Workshop
| Workshop | 2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCV-W 2025) |
|---|---|
| Abbreviated title | ICCVW 2025 |
| Place | United States |
| City | Honolulu |
| Period | 19/10/25 → 23/10/25 |
| Internet address |
Funding
The research was partially supported by the RGC General Research Fund 11200323 and NSFC/RGC JRS Project N-CityU198/24.
Research Keywords
- Image Quality Assessment
- Large Multimodal Models
Fingerprint
Dive into the research topics of 'VQualA 2025 Challenge on Visual Quality Comparison for Large Multimodal Models: Methods and Results'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver