Enhancing Descriptive Image Quality Assessment With a Large-Scale Multi-Modal Dataset
A multi-functional VLM-based IQA model with a 495K large-scale dataset, surpassing GPT-4V
Research archive · 2022—present
Peer-reviewed work across image restoration and enhancement, vision-language image quality understanding, efficient architectures, model interpretation, and controllable generation.
A multi-functional VLM-based IQA model with a 495K large-scale dataset, surpassing GPT-4V
An efficient three-step SDRTV-to-HDRTV conversion method with only 35K parameters for global color mapping
Unidirectional information flow reduces GPU memory by 1/3 and speeds up training by 2.3x for diffusion model control
Introduce causality theory to interpret low-level vision models with a model-/task-agnostic method
Propose a general backbone for image restoration
Propose a language-based image quality assessment method
The most powerful image restoration and enhancement tool that is nearly ready for commercial use
A survey on Single Image Super-Resolution
Winner in NTIRE 2022 Efficient SR sub-track: Model Complexity