Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Depicting Beyond Scores: Advancing Image Quality Assessment through Multi-modal Language Models

About

We introduce a Depicted image Quality Assessment method (DepictQA), overcoming the constraints of traditional score-based methods. DepictQA allows for detailed, language-based, human-like evaluation of image quality by leveraging Multi-modal Large Language Models (MLLMs). Unlike conventional Image Quality Assessment (IQA) methods relying on scores, DepictQA interprets image content and distortions descriptively and comparatively, aligning closely with humans' reasoning process. To build the DepictQA model, we establish a hierarchical task framework, and collect a multi-modal IQA training dataset. To tackle the challenges of limited training data and multi-image processing, we propose to use multi-source training data and specialized image tags. These designs result in a better performance of DepictQA than score-based approaches on multiple benchmarks. Moreover, compared with general MLLMs, DepictQA can generate more accurate reasoning descriptive languages. We also demonstrate that our full-reference dataset can be extended to non-reference applications. These results showcase the research potential of multi-modal IQA methods. Codes and datasets are available in https://depictqa.github.io.

Zhiyuan You, Zheyuan Li, Jinjin Gu, Zhenfei Yin, Tianfan Xue, Chao Dong• 2023

Related benchmarks

TaskDatasetResultRank
Distortion IdentificationPANDABENCH Easy
Accuracy75
14
Distortion type classificationPANDABENCH (Hard set)
Accuracy22
14
Distortion Severity PredictionPANDABENCH Easy
Accuracy55
13
Severity level classificationPANDABENCH (Hard set)
Accuracy30
13
Comparative Relationship PredictionPANDABENCH Easy
Accuracy49
9
Quality Score AssessmentPANDABENCH Easy
SRCC78
9
Quality ScoringPANDABENCH (Hard set)
SRCC0.18
9
Region-wise comparison assessmentPANDABENCH (Hard set)
Accuracy33
9
Image Quality DescriptionKonIQ
Accuracy5.4
8
Image Quality DescriptionSPAQ
Accuracy6.04
8
Showing 10 of 11 rows

Other info

Follow for update