Xiao Guo (郭晓)

I am an Applied Scientist in Amazon Trust & Store Integrity (TSI), working on forgery detection for seller verification. I received my Ph.D. in Computer Science from Michigan State University, where I was a member of the Computer Vision Lab, advised by Prof. Xiaoming Liu, and worked closely with Prof. Anil K. Jain. Before that, I spent several years as a Research Programmer at USC/ISI, working with Prof. Iacopo Masi and Prof. Wael AbdAlmageed. I received my M.S. and B.S. degrees from the University of Southern California and Wuhan University of Technology, respectively.

I have contributed to several major U.S. government-sponsored research projects, including MediFor, ODIN, and RED.

My research interests include:

Email  /  CV  /  Scholar  /  Github

profile photo

News:

  • 2026-06: Interviewed as a “Researcher on the Rise” by the IEEE Biometrics Council Newsletter.
  • 2026-02: Two papers are accepted to CVPR 2026.
  • 2026-01: Gave an online lecture for the Vision and Language class at UNC-Chapel Hill.
  • 2025-10: Two papers are accepted to WACV 2026.
  • 2025-06: One paper is accepted to ICCV 2025, and I attended the ICCV 2025 Doctoral Consortium.
  • 2025-03: Our M2F2-Det is selected as an oral presentation at CVPR 2025.
  • 2024-10: Our HiFi-Net++ is accepted to IJCV 2025.
  • 2024-09: Two papers, MM-Det and LGPN, are accepted to NeurIPS 2024.
  • 2024-09: Completed my internship at Amazon One.
  • 2023-09: Completed my internship at Amazon Alexa.
  • 2022-05: Our ECCV 2022 paper is selected as an oral presentation.

Selected Publications:

M2F2-Det teaser
Rethinking Vision-Language Model in Face Forensics: Multi-Modal Interpretable Forged Face Detector
Xiao Guo, Xiufeng Song, Yue Zhang, Xiaohong Liu, Xiaoming Liu
CVPR, 2025
Oral Presentation · 0.7% of total submissions
project page / codeGitHub stars / arXiv

Formulating a deepfake detection task with Large Language Models.

HiFi-Net teaser
Hierarchical FineGrained Image Forgery Detection and Localization
Xiao Guo, Xiaohong Liu, Zhiyuan Ren, Steven Grosz, Iacopo Masi, Xiaoming Liu
CVPR, 2023; extended version in IJCV, 2025 (IF=11.6)
code GitHub stars / arXiv

An image forgery detection and localization method for both digital manipulation and image editing domains.

MDFAS teaser
Multi-domain Learning for Updating Face Anti-spoofing Models
Xiao Guo, Yaojie Liu, Anil Jain, Xiaoming Liu
ECCV, 2022
Oral Presentation · 2.3% of total submissions
codeGitHub stars / arXiv

A new model for multi-domain face anti-spoofing, which addresses the forgetting issue when learning new domain data.

Relation extraction teaser
Discourse-level Relation Extraction via Graph Pooling
I-Hung Hsu, Xiao Guo, Prem Natarajan, Nanyun Peng
AAAI Workshop on Deep Learning on Graphs, 2022
Best Paper Award
arXiv / Workshop Page

Other Publications:

FusionAgent teaser
FusionAgent: A Multimodal Agent with Dynamic Model Selection for Human Recognition
Jie Zhu, Xiao Guo, Yiyang Su, Anil Jain, Xiaoming Liu
CVPR, 2026
codeGitHub stars / arXiv

A MLLM agent to dynamically select the optimal model combination per sample, achieving robust and explainable human recognition via adaptive score fusion.

DDVQA-BLIP teaser
Common Sense Reasoning for Deepfake Detection
Yue Zhang, Ben Colman, Xiao Guo, Ali Shahriyari, Gaurav Bharaj
ECCV, 2024
codeGitHub stars / arXiv

Fine-tuning BLIP for deepfake detection VQA.

MM-Det teaser
On Learning Multi-Modal Forgery Representation for Diffusion Generated Video Detection
Xiufeng Song, Xiao Guo, others , Xiaoming Liu Guangtao Zhai, Xiaohong Liu
NeurIPS, 2024
codeGitHub stars / arXiv

A forgery video detection method based the LLama-2.

HiFi-Net++ teaser
Language-guided Hierarchical Finegrained Image Forgery Detection and Localization
Xiao Guo, Xiaohong Liu, Iacopo Masi, Xiaoming Liu
IJCV, 2025
codeGitHub stars / arXiv

Image Forgery Detection and Localization; An extension work of HiFi-Net (CVPR23).

LGPN teaser
Tracing Hyperparameter Dependencies for Model Parsing via Learnable Graph Pooling Network
Xiao Guo, Vishal Asnani, Sijia Liu, Xiaoming Liu
NeurIPS, 2024
arXiv

Introduce a learnable graph pooling network for the model parsing.

DenseFace teaser
Dense-Face: Personalized Face Generation Model via Dense Annotation Prediction
Xiao Guo, Manh Tran, Jiaxin Cheng, Xiaoming Liu
arXiv, 2024
project page / code GitHub stars / arXiv

A personalized face generation T2I diffusion model via dense landmarks prediction.

SeaCLIP teaser
Sea-CLIP: Mining Semantic-Aware Representations for Few-Shot Anomaly Detection with CLIP
Xiao Guo, Zhimin Chen, Carlos D. Castillo, Hongcheng Wang, Xiaoming Liu
WACV, 2026
arXiv

Semantic-aware representations for few-shot anomaly detection with CLIP.

Human motion prediction teaser
Human motion prediction via learning local structure representations and temporal dependencies
Xiao Guo, Jongmoo Choi
AAAI 2019
codeGitHub stars / arXiv

Academic Services:

I regularly review papers for the following conferences and journals:

  • Conferences: CVPR, ICCV, ECCV, AAAI, NeurIPS, ICLR, etc.
  • Journals: T-PAMI, IJCV, TIFS, etc.