Wenjie Liao   |   廖文杰🧑‍🚀

I am a master's student in Computer Science and Technology at Nankai University logo Nankai University, supervised by Prof. Chongyi Li. I also received my bachelor's degree in Data Science and Big Data Technology from Nankai University. My research focuses on multimodal large language models (MLLMs), agentic systems, and explainable image quality assessment (IQA).

Previously, I was an algorithm intern at the ByteDance logo ByteDance Multimedia Evaluation Lab, where I worked on VLMs for image quality understanding and assessment, covering data annotation and curation, chain-of-thought data construction, SFT and RL post-training (GRPO, DAPO), and deployment.

Photo of Wenjie Liao
News

  • Apr. 2026 Our team won 2nd place in the NTIRE 2026 RAIM Challenge (Track 1: Professional Image Quality Assessment), and our agentic pairwise IQA paper was accepted by CVPRW 2026.
  • Feb. 2026 DNF-SR was accepted by CVPR 2026.
  • Oct. 2025 The report of the MIPI 2025 Challenge on Detailed Image Quality Assessment, which I organized, was presented at the ICCV 2025 workshop.
  • Sep. 2025 I started my master's study at Nankai University.
  • Jul. 2025 Our team won 3rd place in the VQualA 2025 GenAI-Bench AIGC video quality assessment challenge (ICCVW 2025).
  • Publications

    * Corresponding author. First-author papers are highlighted. See also my Google Scholar.

    ViDA-UGC data structure ViDA-UGC: Detailed Image Quality Analysis via Visual Distortion Assessment for User-Generated Content
    Wenjie Liao, Jieyu Yuan, Yifang Xu, Chunle Guo, Jihong Li, Zilong Zhang, Jiachen Fu, Haotian Fan, Chongyi Li*
    Homepage / Code / Dataset / Model
    Submitted to IEEE TIP

    A large-scale visual distortion assessment instruction-tuning dataset and benchmark for UGC images (11.5K images, 36K distortion boxes, 534K instruction data) that unifies distortion grounding, low-level perception, and quality description, built with a grounding- and perception-guided chain-of-thought pipeline.

    Planner, Executor and Summarizer modules of the agentic system A VLM-Driven Agentic System for Multi-Dimensional Pairwise Image Quality Comparison
    Wenjie Liao, Shuhao Han, Jiaxin Song, Jieyu Yuan*, Chunle Guo, Chongyi Li
    NTIRE 2026 RAIM Challenge, Track 1: 2nd place
    CVPRW 2026

    A modular agentic framework for explainable pairwise IQA: a Planner, an Executor, and a Summarizer collaborate with dimension-specific analysis tools, trained with SFT and RL, to reach traceable, expert-aligned preference decisions.

    DNF-SR architecture compared with prior one-step SR models DNF-SR: Dual-Input and Negative-Aware Feature Fine-Tuning for Real-World Image Super-Resolution
    Shuhao Han, Wenjie Liao, Hayden Vance, Hang Dong, Rui Zhang, Chun-Le Guo, Chongyi Li*
    Code
    CVPR 2026

    A one-step diffusion Real-ISR framework that feeds both the LR image and its noisy version into an image-editing diffusion model, with Negative-aware Feature Fine-Tuning (NF²T) that uses aggregated IQA rewards to build positive and negative optimization directions.

    ViDA-UGC data structure used in the MIPI 2025 challenge MIPI 2025 Challenge on Detailed Image Quality Assessment: Methods and Results
    Wenjie Liao, Haotian Fan, Yifang Xu, Meijia Song, Qiufang Ma, Shuhao Han, Chunle Guo, Chongyi Li, et al.
    Paper / Challenge / Code
    ICCVW 2025

    Report of the Detailed IQA track of MIPI 2025, which I organized; the track benchmarks distortion grounding, low-level perception, and quality reasoning.

    Overall pipeline of IOVQA Better Supervised Fine-tuning for VQA: Integer-Only Loss
    Baihong Qian, Haotian Fan, Wenjie Liao, Yunqiu Wang, Tao Li, Junhui Cui
    Paper
    ICCVW 2025

    IOVQA fine-tunes Qwen2.5-VL for AIGC video quality scoring with integer-only labels and a target-mask loss on the score tokens; 3rd place in the VQualA 2025 GenAI-Bench challenge.


    Education and Experience
    Nankai University logo Nankai University, Tianjin, China
    2025.09 - 2028.06 (expected)
    M.S. in Computer Science and Technology
    Advisor: Prof. Chongyi Li
    ByteDance logo ByteDance, Multimedia Evaluation Lab, Shanghai, China
    2024.12 - 2026.04
    Algorithm Intern
    VLMs for image quality understanding and assessment: data curation, CoT data construction, SFT and RL post-training, and deployment.
    Nankai University logo Nankai University, Tianjin, China
    2021.09 - 2025.06
    B.S. in Data Science and Big Data Technology

    Projects

    🚀 Build in Open. Grow Together.

    Loading projects...


    Service

  • Organizer: MIPI 2025 Challenge on Detailed Image Quality Assessment (ICCV 2025 Workshop)

  • Awards and Honors

  • 2026: Runner-up, NTIRE 2026 RAIM Challenge, Track 1: Professional Image Quality Assessment (CVPRW 2026)
  • 2025: Third place, VQualA 2025 GenAI-Bench AIGC Video Quality Assessment Challenge (ICCVW 2025)
  • 2025: Nankai University Postgraduate Recommendation Scholarship
  • ✅复制完成