Zhuohang Jiang

I am currently pursuing a PhD degree at the Hong Kong Polytechnic University. My supervisors are Qing Li and Wenqi Fan. My current Cumulative GPA is 3.60/4.00.

I studied at Sichuan University (SCU) from 2020 to 2024, where I majored in Computer Science & Technology. My Major GPA (CS courses): 3.79/4, 89.39/100; Overall GPA: 3.78/4, 89.25/100

During my time at Sichuan University, I worked as a research assistant at MachineILab from 2022 to 2024, advised by Prof. JiZhe Zhou. I participated in one National Natural Science Foundation of China project and one National Key R&D Program of China.

Email  /  CV  /  Google Scholar  /  Github 

profile photo
🤝 Seeking Internship Opportunities
I am currently seeking internship opportunities in large language models (LLMs) and AI agents. Please feel free to reach out via Email or WeChat.
Please note: search, advertising, and recommendation opportunities are not a fit.

Research Topics

My research centers on large language models (LLMs), AI agents, and retrieval-augmented generation (RAG), with a particular interest in turning LLM reasoning into reliable, efficient systems for large-scale decision-making and personalized user experiences.

My previous research was primarily focused on topics within computer vision, such as tampering detection and object recognition tasks. I have contributed to the design of high-impact benchmarks, such as HiBench, and comprehensive surveys like WebAgents. My work has been published in top-tier conferences, including NeurIPS 2024 (spotlight) , KDD 2025, KDD 2026 and AAAI 2025, and has accumulated 500+ citations, with an h-index of 6.


News

🏆 2026-09-08 - Our paper "Do You Truly Love Me?" Benchmarking LLM Capability on Hierarchical Pragmatic Tactic Conversations was accepted to AACL-IJCNLP 2026 Main! 🎉

📜 2026-08-14 - Our new survey A Survey on AI Smart Glasses for Wearable Intelligence: From Egocentric Sensing to Agentic Personalization is now available on Preprints.org! 🎉

📜 2026-08-05 - Our new paper A/B Agent: A Self-Evolving Agent for Strategy Iteration in Industrial A/B Testing is now available on arXiv! 🎉

🏆 2026-05-16 - Our paper Atomic Intent Reasoning: Bringing LLM Semantics to Industrial Cross-Domain Recommendations was accepted as a presentation at the KDD 2026 Main Conference ADS Track! 🎉

🏆2026-04 - Our paper SUPERGLASSES was accepted by CVPR 2026 Findings! 🎉

💼 2025-10 - Will join Kuaishou E-commerce Team as a research intern.

📝 2025-09 - Appointed as Topic Coordinator for Frontiers in Artificial Intelligence (Impact Factor: 4.7, CiteScore: 7.3, Logic and Reasoning Section) and Frontiers in Big Data (Impact Factor: 2.3, CiteScore: 6.1).

🏆 2025-08-20 - Our work QA-Dragon was accepted by KDD 2025 Workshop for Multimodal Retrieval Augmented Generation.

🎤 2025-08-07 - Invited to give a talk "Understanding Hierarchical Data with Large Language Models: RAG, Structural Reasoning, and Future Directions" at KDD 2025 Reasoning Day in Toronto, Canada! 🎙️

🏅 2025-06-18 - Achieved 12th place globally in KDD Cup 2025 - Meta CRAG-MM Multimodal Retrieval Challenge among hundreds of international teams! 🌍

🏆 2025-05-16 - Our benchmark paper HiBench was accepted to KDD Benchmark Track! 🎉

🎉 2025-05-07 - Our survey paper A Survey of WebAgents: Towards Next-Generation AI Agents for Web Automation with Large Foundation Models was accepted to KDD Tutorial Track! 🎊

📘 2025-03-01 - Completed the HiBench paper and released the code and dataset on GitHub and Hugging Face.

🌟 2025-01-15 - Mesoscopic Insights: Orchestrating Multi-Scale & Hybrid Architecture for Image Manipulation Localization was published in AAAI 2025.

🏆 2024-12-01 - IMDL-BenCo was published in NeurIPS 2024 Benchmark Tracks and received a Spotlight award.

🎓 2024-09-01 - Beginning my pursuit of a PhD degree in Hong Kong PolyU.

🎓 2024-06-26 - Got Outstanding Graduate Award from Sichuan University and Sichuan Province! 🎉

🎓 2024-06-26 - Graduated from Sichuan University with a bachelor's degree.

🛠️ 2024-06-12 - Completed the co-work project IMDLBenCo and finished a paper IMDL-BenCo: A Comprehensive Benchmark and Codebase for Image Manipulation Detection & Localization

🔍 2024-05-24 - Finished a paper Beyond Visual Appearances: Privacy-sensitive Objects Identification via Hybrid Graph Reasoning

📚 2023-08-01 - Participated in NUS Summer School research program at National University of Singapore, completed face recognition project! 🇸🇬


Selected Publications
[KDD'26 ADS Track] Atomic Intent Reasoning: Bringing LLM Semantics to Industrial Cross-Domain Recommendations
Zhuohang Jiang, Yuxin Chen, Shijie Wang, Haohao Qu, Zhou Jindong, Wenqi Fan, Li Qing, Dongxu Liang, Jun Wang
CCF-A CORE A* (KDD) arXiv
LLM Reasoning · Atomic Intents · Cross-Domain Recommendation · Industrial Deployment
[CVPR'26 Findings] SUPERGLASSES: Benchmarking Vision Language Models as Intelligent Agents for AI Smart Glasses
Zhuohang Jiang, Xu Yuan, Haohao Qu, Shanru Lin, Kanglong Liu, Wenqi Fan, Qing Li
CCF-A arXiv
AI Smart Glasses · VQA Benchmark · Vision-Language Models · Multimodal RAG
[KDD'25 Workshop] QA‑Dragon: Query‑Aware Dynamic RAG System for Knowledge‑Intensive Visual Question Answering
Zhuohang Jiang, Pangjing Wu, Xu Yuan, Wenqi Fan, Qing Li
CCF-A arXiv
Query-Aware RAG · Knowledge-Intensive VQA · Multi-Hop Reasoning
[KDD'25] HiBench: Benchmarking LLMs Capability on Hierarchical Structure Reasoning
Zhuohang Jiang, Pangjing Wu, Ziran Liang, Peter Q. Chen, Xu Yuan, Ye Jia, Jiancheng Tu, Chen Li, Peter H.F. Ng, Qing Li
CCF-A CORE A* (KDD) arXiv GitHub Hugging Face
Hierarchical Reasoning · LLM Benchmark · 30 Tasks · 39,519 Queries
[AACL-IJCNLP'26 Main] "Do You Truly Love Me?" Benchmarking LLM Capability on Hierarchical Pragmatic Tactic Conversations
Peter Q. Chen, Pangjing Wu, Zhuohang Jiang, Xiaodong Li, Peter H. F. Ng
AACL-IJCNLP 2026 Main CORE B (IJCNLP)
LLM Evaluation · Hierarchical Pragmatics · Conversational Reasoning
[KDD'25] A Survey of WebAgents: Towards Next-Generation AI Agents for Web Automation with Large Foundation Models
Liangbo Ning, Ziran Liang, Zhuohang Jiang, Haohao Qu, Yujuan Ding, Wenqi Fan, Xiao-yong Wei, Shanru Lin, Hui Liu, Philip S. Yu, Qing Li
CCF-A arXiv
Web Agents · Foundation Models · Web Automation · Trustworthiness
[AAAI'25] Mesoscopic Insights: Orchestrating Multi-Scale & Hybrid Architecture for Image Manipulation Localization
Xuekang Zhu, Xiaochen Ma, Lei Su, Zhuohang Jiang, Bo Du, Xiwen Wang, Zeyu Lei, Wentao Feng, Chi-Man Pun, Jizhe Zhou
CCF-A CORE A* (AAAI) arXiv
Image Manipulation Localization · Multi-Scale Features · Hybrid Architecture
[NIPS'24] IMDL-BenCo: A Comprehensive Benchmark and Codebase for Image Manipulation Detection & Localization
Xiaochen Ma, Xuekang Zhu, Lei Su, Bo Du, Zhuohang Jiang, Bingkui Tong, Zeyu Lei, Xinyu Yang, Chi-Man Pun, Jiancheng Lv, Jizhe Zhou
CCF-A CORE A* (NeurIPS) arXiv GitHub
Image Forensics · Modular Benchmark · Robustness Evaluation · Reproducibility
[ICONIP'23] TPTGAN: Two-Path Transformer-Based Generative Adversarial Network Using Joint Magnitude Masking and Complex Spectral Mapping for Speech Enhancement
Zhaoyi Liu, Zhuohang Jiang, Wendian Luo, Zhuoyao Fan, Haoda Di, Yufan Long, Haizhou Wang
CCF-C CORE B (ICONIP) Paper
Speech Enhancement · Two-Path Transformer · GAN · Spectral Mapping

Selected Projects
HiBench: Benchmark for Hierarchical Reasoning
First Author & Team LeaderKDD 2025 Benchmark Track • 2025
arXiv GitHub Hugging Face
Hierarchical Reasoning · 39,519 Queries · Open-Source Toolkit · KDD 2025 Oral
Meta CRAG-MM: Multimodal Retrieval Challenge
Team LeaderKDD Cup 2025 • 2025
Multimodal Retrieval · RAG · Team Leadership · Global Rank #12
IMDL-BenCo: Benchmark for Image Manipulation Detection & Localization
Co-First AuthorNeurIPS 2024 Benchmark Track — Spotlight • 2024
arXiv GitHub
Image Forensics · 8 Baselines · GPU-Accelerated Evaluation · NeurIPS 2024 Spotlight

Invited Talks
Understanding Hierarchical Data with Large Language Models: RAG, Structural Reasoning, and Future Directions
Invited TalksReasoning Day @ KDD 2025 • Toronto, ON, Canada • Aug 2025
Invited Talk · Hierarchical Reasoning · RAG · KDD 2025 Reasoning Day

Tutorials
WebAgents: Towards Next-Generation AI Agents for Web Automation with Large Foundation Models
Yujuan Ding, Liangbo Ning, Ziran Liang, Zhuohang Jiang, Haohao Qu, Wenqi Fan, Qing Li, Hui Liu, Xiaoyong Wei, Philip S. Yu
KDD 2025 TutorialThe 31st ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD 2025) • 2025
arXiv
KDD 2025 Tutorial · Web Agents · Foundation Models · Web Automation

Education
Sichuan University, Chengdu, Sichuan, China
B.E. in Computer Science and Technology • Sep. 2020 to Jun. 2024

Hong Kong Polytechnic University, Hongkong, China
PHD. in Computer Science and Technology • Sep. 2024 to Present

Experience
National University of Singapore (NUS)
Summer School Participant • Aug. 2023
• Participated in intensive research program at School of Computing
• Completed face recognition project using CNN-based feature extraction and similarity matching
• Gained international research experience and cross-cultural collaboration skills

DICALab, Sichuan University
Research Assistant • Sep. 2022 to Jun. 2024
Advisor: Prof. JiZhe Zhou
• Developed graph-based frameworks for privacy-sensitive object detection
• Participated in National Natural Science Foundation of China project
• Contributed to National Key R&D Program of China
• Co-authored multiple publications in top-tier conferences and journals

Kuaishou E-commerce Team, Kuaishou Inc.
Research Intern • Oct. 2025 to Present
Beijing, China
• Developed cross-domain fine-ranking models, improving cart-attached GMV by about 7pp and Live-avatar GMV by about 5pp
• Extended cross-domain modeling to coarse ranking, delivering about 20pp and 2pp gains on the two GMV metrics, respectively
• Built an A/B Agent-based score-fusion approach for mixed ranking, yielding about 4pp and 2pp GMV gains
• Built AIR (KDD'26 ADS Track, CCF-A), an offline-to-online LLM system that distills cross-domain user behavior into atomic intents for efficient serving; and A/B Agent, a self-evolving LLM agent for closed-loop strategy optimization


Selected Awards
KDD Cup 2025 — Meta CRAG-MM
Toronto, Canada, 2025
12th Place (Global)
NeurIPS 2024 — IMDL‑BenCo (Co‑first Author)
Vancouver, Canada, 2024
Spotlight Award
Outstanding Graduate
Sichuan University & Sichuan Province, 2024
Top Achievement
Tencent Scholarship
Sichuan University, China, 2023
Top 2%
A-Level Certificate
Comprehensive Quality Evaluation, China, 2023
Excellence
Comprehensive First Class Scholarship
Sichuan University, Sichuan, China, 2022
Top 1%
Outstanding Students of Sichuan University
Sichuan University, Sichuan, China, 2022
Top 5%

Professional Service
Conference & Journal Reviewer
2023-2025
TIP, ECCV, NeurIPS, KDD, AAAI, IoTJ,
Conference & Journal Topic Coordinator
2025-2026
Frontiers in Artificial Intelligence, Frontiers in Big Data
Teaching Assistant
The Hong Kong Polytechnic University (PolyU)
Artificial Intelligence (COMP4431)
NLP Practicum (COMP5423)
DataBase System (COMP2411)

Skills
Research Topics Large Language Models (LLMs), AI Agents, Retrieval‑Augmented Generation (RAG), LLM-powered Personalization
Frameworks & Tools PyTorch, Hugging Face, NumPy, Docker, Git, Anaconda
Languages Mandarin (native), English (fluent, IELTS 6.5)

Friends
Pangjing Wu  /  Xu Yuan  /  Changao Shen  /  Xiaochen Ma  /  Yuchuan Deng  /  Jiaye  /  Haohao Qu  /  Wendian Luo

Updated at Sep. 2025 · Template inspired by Jon Barron