Skip to main navigation Skip to search Skip to main content

VQualA 2025 Challenge on Visual Quality Comparison for Large Multimodal Models: Methods and Results

  • Hanwei Zhu
  • , Haoning Wu
  • , Zicheng Zhang
  • , Lingyu Zhu
  • , Yixuan Li
  • , Peilin Chen
  • , Shiqi Wang*
  • , Chris Wei Zhou
  • , Linhan Cao
  • , Wei Sun
  • , Xiangyang Zhu
  • , Weixia Zhang
  • , Yucheng Zhu
  • , Jing Liu
  • , Dandan Zhu
  • , Guangtao Zhai
  • , Xiongkuo Min
  • , Zhichao Zhang
  • , Xinyue Li
  • , Shubo Xu
  • Anh Dao, Yifan Li, Hongyuan Yu, Jiaojiao Yi, Yiding Tian, Yupeng Wu, Feiran Sun, Lijuan Liao, Song Jiang
*Corresponding author for this work

Research output: Chapters, Conference Papers, Creative and Literary WorksRGC 32 - Refereed conference paper (with host publication)peer-review

Abstract

This paper presents a summary of the VQualA 2025 Challenge on Visual Quality Comparison for Large Multimodal Models (LMMs), hosted as part of the ICCV 2025 Work-shop on Visual Quality Assessment. The challenge aims to evaluate and enhance the ability of state-of-the-art LMMs to perform open-ended and detailed reasoning about visual quality differences across multiple images. To this end, the competition introduces a novel benchmark comprising thousands of coarse-to-fine grained visual quality comparison tasks, spanning single images, pairs, and multi-image groups. Each task requires models to provide accurate quality judgments. The competition emphasizes holistic evaluation protocols, including 2AFC-based binary preference and multi-choice questions (MCQs). Around 100 participants submitted entries, with five models demonstrating the emerging capabilities of instruction-tuned LMMs on quality assessment. This challenge marks a significant step toward open-domain visual quality reasoning and comparison and serves as a catalyst for future research on inter-pretable and human-aligned quality evaluation systems. © 2025 IEEE.
Original languageEnglish
Title of host publicationProceedings - 2025 IEEE/CVF International Conference on Computer Vision Workshops
Subtitle of host publicationICCV-W 2025
PublisherIEEE
Pages3383-3393
ISBN (Electronic)9798331589882
ISBN (Print)979-8-3315-8989-9
DOIs
Publication statusPublished - Oct 2025
Event2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCV-W 2025) - Honolulu, United States
Duration: 19 Oct 202523 Oct 2025
https://iccv.thecvf.com/

Publication series

NameProceedings - IEEE/CVF International Conference on Computer Vision Workshops, ICCV-W
ISSN (Print)2473-9936
ISSN (Electronic)2473-9944

Workshop

Workshop2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCV-W 2025)
Abbreviated titleICCVW 2025
PlaceUnited States
CityHonolulu
Period19/10/2523/10/25
Internet address

Funding

The research was partially supported by the RGC General Research Fund 11200323 and NSFC/RGC JRS Project N-CityU198/24.

Research Keywords

  • Image Quality Assessment
  • Large Multimodal Models

Fingerprint

Dive into the research topics of 'VQualA 2025 Challenge on Visual Quality Comparison for Large Multimodal Models: Methods and Results'. Together they form a unique fingerprint.

Cite this