About
I am a Ph.D. student in the Department of Computer Science and Engineering at the University at Buffalo, where I am advised by Prof. Chunming Qiao and Prof. Junsong Yuan. Before that, I received my M.S. in Software Engineering from Xi'an Jiaotong University and my B.S. from Sichuan Agricultural University.
My research lies in computer vision and machine learning. Recently I have been working on reward design and post-training for text-to-image generation, perceptual quality metrics for 3D shapes, benchmarking multimodal large language models, and salient object detection. I have interned at TikTok, Microsoft, and Microsoft Research Asia.
News
- RubricRL (text-to-image rewards) accepted to ECCV 2026.
- Learning 3D Shape Fidelity Metric from Real-world Distortions accepted to CVPR 2026.
- Joined TikTok (Trust and Safety) as a Research Scientist Intern for summer 2026.
- Two papers accepted to AAAI 2026: Textured Geometry Evaluation, and SRAM (AIA track).
- GeoRemover accepted to NeurIPS 2025 as a Spotlight.
- Pluralistic Salient Object Detection published in IEEE TIP; Benchmarking Large and Small MLLMs published in Machine Vision and Applications.
Education
-
2023 – Present
University at Buffalo
Ph.D. in Computer Science and Engineering
-
2020 – 2023
Xi'an Jiaotong University
M.S. in Software Engineering, specialty in Intelligent Systems
-
2016 – 2020
Sichuan Agricultural University
B.S. in Electrical Engineering, specialty in Internet of Things Engineering
GPA 4.05 / 5.0, ranked 3 of 156
Experience
-
May 2026 – Aug 2026
TikTok USA
Research Scientist Intern, Trust and Safety
Multilingual ASR post-training with supervised fine-tuning and data-centric optimization for large-scale live content moderation.
-
May 2025 – May 2026
Microsoft USA
Part-time Intern Researcher
Text-to-image generation with large language models; led to RubricRL (ECCV 2026).
-
May 2024 – Nov 2024
Microsoft USA
Part-time Intern Researcher
Benchmarking large and small multimodal LLMs across a wide range of tasks (published in Machine Vision and Applications).
-
Dec 2022 – May 2023
Microsoft Research Asia Beijing
Part-time Intern Researcher
Re-defined salient object detection as a multi-output prediction task with a new evaluation protocol (IEEE TIP). Designed an asymmetric VQGAN that recovers details lost to quantization.
-
Dec 2021 – Aug 2022
Microsoft Research Asia Beijing
Intern Researcher
Optical flow models for dense correspondence tracking of the human body.
-
Sep 2020 – Jun 2023
Xi'an Jiaotong University Xi'an
Research Assistant
Local-to-global feature learning for salient object detection (Pattern Recognition Letters).
Publications
Author name in bold indicates me.
2026
-
Learning 3D Shape Fidelity Metric from Real-world DistortionsCVPR 2026 IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 28391–28401
-
RubricRL: Simple Generalizable Rewards for Text-to-Image GenerationECCV 2026 European Conference on Computer Vision
-
Textured Geometry Evaluation: Perceptual 3D Textured Shape Metric via 3D Latent-Geometry NetworkAAAI 2026 arXiv:2512.01380
-
SRAM: Shape-Realism Alignment Metric for No Reference 3D Shape EvaluationAAAI 2026 AIA arXiv:2512.01373
2025
-
GeoRemover: Removing Objects and Their Causal Visual ArtifactsNeurIPS 2025 Spotlight Advances in Neural Information Processing Systems
-
Benchmarking Large and Small MLLMsMVA Machine Vision and Applications · DOI
-
Pluralistic Salient Object DetectionIEEE TIP IEEE Transactions on Image Processing
2024
-
Exploring Pre-trained Text-to-Video Diffusion Models for Referring Video Object SegmentationECCV 2024 European Conference on Computer Vision
Preprints
-
Designing a Better Asymmetric VQGAN for StableDiffusionPreprint arXiv:2306.04632, 2023
-
HybridPersona: Combining the Visual Characteristics of Humans and Stylized AvatarsPreprint
Earlier
-
Local to Global Feature Learning for Salient Object DetectionPRL Pattern Recognition Letters, vol. 162, pp. 81–88, 2022
-
Seismic Facies Recognition Based on Prestack Data Using Two-dimensional Gabor Transform and Unsupervised ClusteringICICSP 2019 pp. 509–513 · DOI
-
Seismic Facies Recognition Based on Prestack Data Using Two-dimensional Convolutional Auto-encoder and Cluster AnalysisICICSP 2019 pp. 430–434 · DOI
Patents
- Pluralistic Salient Object Detection. Dongdong Chen, Yunsheng Li, Lu Yuan, Xuelu Feng. US Patent Application 18/431,912, published 2025-08-07.
- Runtime Ranking of Object Detection. Dongdong Chen, Yunsheng Li, Lu Yuan, Xuelu Feng. US Patent Application 18/431,913, published 2025-08-07.
Professional Service
Journal Reviewer
- IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI)
- International Journal of Computer Vision (IJCV)
- IEEE Transactions on Image Processing (TIP)
Conference Reviewer
- IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
- European Conference on Computer Vision (ECCV)
- International Conference on Pattern Recognition (ICPR)
Awards
- Outstanding Graduate, Sichuan Agricultural University, 2019
- Chen Yu Scholarship for Outstanding Students (Third Class), 2019
- Guangdong Wen's Outstanding Student Scholarship, 2018
- Outstanding Student, 2016–2017 and 2017–2018
- Third Prize, 4th ACM Programming Contest (school level)
- Excellence Award, "Zhengda Cup" College Student Innovation Practice Competition, Sichuan Final, 2019