I studied at Sichuan University (SCU) from 2020 to 2024,
where I majored in Computer Science & Technology. My
Major GPA (CS courses):
3.79/4, 89.39/100;
Overall GPA: 3.78/4, 89.25/100
During my time at Sichuan University, I worked as a research assistant at
MachineILab
from 2022 to 2024, advised by
Prof. JiZhe Zhou.
I participated in one National Natural Science Foundation of China project and one National Key
R&D
Program of China.
🤝 Seeking Internship Opportunities
I am currently seeking internship opportunities in large language models (LLMs) and AI agents.
Please feel free to reach out via Email or
WeChat.
Please note: search, advertising, and recommendation opportunities are not a fit.
🤝 实习机会
我目前正在寻找大语言模型(LLM)与 AI Agent 方向的实习机会,欢迎通过邮件或微信联系交流。
注:搜广推方向请勿联系。
Research Topics研究方向
My research centers on large language models (LLMs), AI agents, and retrieval-augmented
generation (RAG), with a particular interest in turning LLM reasoning into reliable,
efficient systems for large-scale decision-making and personalized user experiences.
My previous research was primarily focused on topics within computer vision, such as tampering
detection and object recognition tasks. I have contributed to the design of high-impact
benchmarks,
such as HiBench, and comprehensive surveys like WebAgents.
My work has been published in top-tier conferences, including
NeurIPS 2024 (spotlight) , KDD 2025, KDD 2026 and AAAI 2025, and has accumulated
500+ citations, with an h-index of 6.
🏆 2026-09-08 - Our paper "Do You Truly Love Me?" Benchmarking LLM Capability on Hierarchical Pragmatic Tactic Conversations was accepted to AACL-IJCNLP 2026 Main! 🎉
🏆2026-04 - Our paper SUPERGLASSES was accepted by CVPR 2026 Findings! 🎉
💼 2025-10 - Will join Kuaishou E-commerce Team as a
research intern.
📝 2025-09 - Appointed as Topic Coordinator for
Frontiers in Artificial Intelligence (Impact Factor: 4.7, CiteScore: 7.3,
Logic and Reasoning Section) and Frontiers in Big Data (Impact Factor: 2.3,
CiteScore: 6.1).
🏆 2025-08-20 - Our work QA-Dragon was accepted by KDD
2025 Workshop for Multimodal Retrieval Augmented Generation.
🎤 2025-08-07 - Invited to give a talk "Understanding Hierarchical Data with
Large Language Models: RAG, Structural Reasoning, and Future Directions" at KDD 2025 Reasoning Day in Toronto,
Canada! 🎙️
🏅 2025-06-18 - Achieved 12th place globally in KDD Cup 2025
-
Meta CRAG-MM Multimodal Retrieval Challenge among hundreds of international teams! 🌍
🏆 2025-05-16 - Our benchmark paper HiBench was accepted to KDD Benchmark
Track! 🎉
📝 2024-09-15 - 担任 Frontiers in Artificial
Intelligence(影响因子:4.7,CiteScore:7.3,Logic and Reasoning专栏)与 Frontiers in Big
Data(影响因子:2.3,CiteScore:6.1)期刊的Topic Coordinator。
[CVPR'26 Findings] SUPERGLASSES: Benchmarking Vision Language Models as Intelligent Agents for AI Smart Glasses Zhuohang Jiang, Xu Yuan, Haohao Qu, Shanru Lin, Kanglong Liu, Wenqi Fan, Qing Li
AI Smart Glasses · VQA Benchmark · Vision-Language Models · Multimodal RAG
[KDD'25 Workshop] QA‑Dragon: Query‑Aware Dynamic RAG System for Knowledge‑Intensive Visual
Question Answering Zhuohang Jiang, Pangjing Wu, Xu Yuan, Wenqi Fan, Qing Li
Query-Aware RAG · Knowledge-Intensive VQA · Multi-Hop Reasoning
[KDD'25] HiBench: Benchmarking LLMs Capability on Hierarchical Structure Reasoning
Zhuohang Jiang, Pangjing Wu, Ziran Liang, Peter Q. Chen, Xu Yuan, Ye Jia, Jiancheng Tu,
Chen
Li, Peter H.F. Ng, Qing Li
Hierarchical Reasoning · LLM Benchmark · 30 Tasks · 39,519 Queries
[AACL-IJCNLP'26 Main] "Do You Truly Love Me?" Benchmarking LLM Capability on Hierarchical
Pragmatic Tactic Conversations
Peter Q. Chen, Pangjing Wu, Zhuohang Jiang, Xiaodong Li, Peter H. F. Ng
LLM Evaluation · Hierarchical Pragmatics · Conversational Reasoning
[KDD'25] A Survey of WebAgents: Towards Next-Generation AI Agents for Web Automation with
Large
Foundation Models
Liangbo Ning, Ziran Liang, Zhuohang Jiang, Haohao Qu, Yujuan Ding, Wenqi Fan, Xiao-yong Wei,
Shanru Lin, Hui Liu, Philip S. Yu, Qing Li
Web Agents · Foundation Models · Web Automation · Trustworthiness
[AAAI'25] Mesoscopic Insights: Orchestrating Multi-Scale & Hybrid Architecture for Image
Manipulation Localization
Xuekang Zhu, Xiaochen Ma, Lei Su, Zhuohang Jiang, Bo Du, Xiwen Wang, Zeyu Lei, Wentao Feng,
Chi-Man Pun, Jizhe Zhou
Image Manipulation Localization · Multi-Scale Features · Hybrid Architecture
[NIPS'24] IMDL-BenCo: A Comprehensive Benchmark and Codebase for Image Manipulation
Detection &
Localization
Xiaochen Ma, Xuekang Zhu, Lei Su, Bo Du, Zhuohang Jiang, Bingkui Tong, Zeyu Lei, Xinyu
Yang,
Chi-Man Pun, Jiancheng Lv, Jizhe Zhou
Image Forensics · Modular Benchmark · Robustness Evaluation · Reproducibility
[ICONIP'23] TPTGAN: Two-Path Transformer-Based Generative Adversarial Network Using
Joint Magnitude Masking and Complex Spectral Mapping for Speech Enhancement
Zhaoyi Liu, Zhuohang Jiang, Wendian Luo, Zhuoyao Fan, Haoda Di, Yufan Long, Haizhou Wang
Speech Enhancement · Two-Path Transformer · GAN · Spectral Mapping
TMSC: A Tri-Regional Mask-Guided Semantic Calibration Framework for Zero-Shot Anomaly Detection
Wendian Luo, Tong Yu, Junnan Li, Shengxin Dai, Bing Guo, and Zhuohang Jiang
A/B Agent: A Self-Evolving Agent for Strategy Iteration in Industrial A/B Testing Zhuohang Jiang, Yuxin Chen, Yongsen Pan, Zheng Hu, Wenqi Fan, Qing Li, Hongyang Wang, Jun Wang, Wenwu Ou
Self-Evolving Agents · Tree-RAG · A/B Testing · Strategy Optimization
A Survey on AI Smart Glasses for Wearable Intelligence: From Egocentric Sensing to Agentic Personalization
Xu Yuan, Yi Wang, Zhuohang Jiang, Haohao Qu, Yujuan Ding, Shanru Lin, Guoliang Xing, Hongxia Yang, Jiannong Cao, Qing Li, Wenqi Fan
AI Smart Glasses · Egocentric Sensing · Multimodal Agents · Personalization
Beyond Visual Appearances: Privacy-sensitive Objects Identification via Hybrid Graph
Reasoning Zhuohang Jiang, Bingkui Tong, Xia Du, Ahmed Alhammadi, Jizhe Zhou
Privacy-Sensitive Objects · Scene Graphs · Hybrid Graph Reasoning
IML-ViT: Benchmarking Image Manipulation Localization by Vision Transformer
Xiaochen Ma, Bo Du, Zhuohang Jiang, Ahmed Y. Al Hammadi, Jizhe Zhou
Image Manipulation Localization · Vision Transformer · Multi-Scale Features · Edge Supervision
Perceptual MAE for Image Manipulation Localization: A High-level Vision Learner Focusing
on Low-level Features
Xiaochen Ma, Zhuohang Jiang, Xiong Xu, Chi-Man Pun, Jizhe Zhou
Image Manipulation Localization · Masked Autoencoders · Perceptual Supervision
Selected Projects主要项目
HiBench: Benchmark for Hierarchical ReasoningHiBench:层次化推理基准 First Author & Team Leader • KDD 2025 Benchmark Track • 2025
Understanding Hierarchical Data with Large Language Models: RAG,
Structural Reasoning, and Future Directions用大语言模型理解层次化数据:RAG、结构推理与未来方向 Invited Talks • Reasoning Day @ KDD 2025 • Toronto, ON, Canada • Aug 2025
WebAgents: Towards Next-Generation AI Agents for Web Automation with Large Foundation Models
Yujuan Ding, Liangbo Ning, Ziran Liang, Zhuohang Jiang, Haohao Qu, Wenqi Fan, Qing Li, Hui Liu, Xiaoyong Wei, Philip S. Yu
KDD 2025 Tutorial • The 31st ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD 2025) • 2025
KDD 2025 Tutorial · Web Agents · Foundation Models · Web Automation
KDD 2025 教程 · 网页智能体 · 基础模型 · 网页自动化
Education教育背景
Sichuan University, Chengdu, Sichuan, China
B.E. in Computer Science and Technology • Sep. 2020 to Jun. 2024
Hong Kong Polytechnic University, Hongkong, China
PHD. in Computer Science and Technology • Sep. 2024 to Present
四川大学,成都,四川,中国
计算机科学与技术学士 • 2020年9月至2024年6月
香港理工大学,香港,中国
计算机科学与技术博士 • 2024年9月至今
Experience工作经历
National University of Singapore (NUS) Summer School Participant • Aug. 2023
• Participated in intensive research program at School of Computing
• Completed face recognition project using CNN-based feature extraction and similarity matching
• Gained international research experience and cross-cultural collaboration skills
DICALab, Sichuan University Research Assistant • Sep. 2022 to Jun. 2024
Advisor: Prof. JiZhe
Zhou
• Developed graph-based frameworks for privacy-sensitive object detection
• Participated in National Natural Science Foundation of China project
• Contributed to National Key R&D Program of China
• Co-authored multiple publications in top-tier conferences and journals
Kuaishou E-commerce Team, Kuaishou Inc. Research Intern • Oct. 2025 to Present
Beijing, China
• Developed cross-domain fine-ranking models, improving cart-attached GMV by about 7pp and
Live-avatar GMV by about 5pp
• Extended cross-domain modeling to coarse ranking, delivering about 20pp and 2pp gains on the two
GMV metrics, respectively
• Built an A/B Agent-based score-fusion approach for mixed ranking, yielding about 4pp and 2pp GMV gains
• Built AIR (KDD'26 ADS Track, CCF-A), an
offline-to-online LLM system that distills cross-domain user behavior into atomic intents for
efficient serving; and A/B Agent, a self-evolving
LLM agent for closed-loop strategy optimization