Skip to content

Author

Huiyu Duan

We have 5 of 148 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Aug 2026

FMReward: Aligning and Evaluating Audio-Driven 3D Facial Animation with Human Preferences.

Audio-driven 3D facial animation is essential for advancing immersion and interactivity in virtual experiences. Although recent advances have shown promising capabilities, the training and evaluation of existing methods typically rely on ground-truth-based errors, which fall short of aligning with human preferences. To...

Si-Jing Wu, Yunhao Li, Zhilin Gao et al. · 1 citation
Preprint Sep 2026

Evaluating the Evaluators: Diagnosing Large Multimodal Models for AI-Generated Image Assessment

With the rapid advancement of text-to-image (T2I) generation, robust evaluation becomes critical yet challenging, as traditional metrics fail to capture fine-grained alignment and generative artifacts. While large multimodal models (LMMs) are increasingly adopted as evaluators, existing benchmarks typically study seman...

Yu Zhao, Jia-Rui Wang, Hui-Yu Duan et al. · 0 citations
Jul 2026

Multi-Dimensional Quality Assessment for AI-Generated Human-Centric Videos: Dataset and Model

AI-generated human-centric videos play a crucial role in a wide range of modern applications. However, they often suffer from quality issues and semantic mismatches, underscoring the importance of effective quality assessment for such videos. To this end, we extend our previous dataset HVEval with pairwise preference a...

Sijing Wu, Yunhao Li, Huiyu Duan et al. · 4 citations
Preprint Aug 2026

CamWorldQA: Perceptual Quality Assessment of Camera-Controlled World Video Generation

Recent advances in generative video models have enabled camera-controlled world video generation, allowing models to synthesize videos under user-defined camera trajectories. However, existing video quality assessment (VQA) methods are mainly developed for natural videos and fail to capture the unique perceptual charac...

Yunhe Li, Likun Wu, Sijing Wu et al. · 0 citations
Preprint Aug 2026

MIEScore: Human-Aligned Evaluation for Multi-Source Image Editing

MIE-Bench is introduced, the first large-scale multiple image editing benchmark with fine-grained human preference annotations and MIEScore, a multimodal large language model (MLLM)-based evaluation model enhanced with skill optimization and multi-dimensional supervised fine-tuning, to provide human-aligned feedback fo...

Zi-Tong Xu, Huiyu Duan, Xinyu Zhang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.