I am an MPhil student in Artificial Intelligence at the Hong Kong University of Science and Technology (Guangzhou), where I am supervised by Prof. Xuming Hu. I have also had the opportunity to work with Prof. Mingxun Zhou at HKUST and Prof. Philip S. Yu at the University of Illinois Chicago. Previously, I earned my B.S. in Data Science from Tongji University.

My research focuses on LLM post-training, multimodal learning, agent memory, and trustworthy AI. I am particularly interested in understanding why learning methods work and translating those insights into reliable, practical systems. I am grateful to have collaborated with Weize Liu, Yibo Yan, Kaichen Huang, Na Min An, Wenjie Qu, Kening Zheng, and Weiwei Sun. I also sincerely appreciate the guidance and support I have received from all my mentors and collaborators.

You can find my work on Google Scholar, GitHub, and OpenReview.

🔥 News

  • 2026 02: Released CausalEmbed, an autoregressive multi-vector approach to visual document embedding.
  • 2026.01: PMark was accepted by ICLR 2026 through a direct submission.
  • 2025.09: One paper submited to ICLR 2026.
  • 2025 04: Joined MemoryOS as a core developer.
  • 2025.05: MMUnlearner was accepted by ACL 2025 Findings.
  • 2025.05: MathAgent was selected as an ACL 2025 Industry Track oral paper.
  • 2025.02: Two paper submited to ACL 2025.
  • 2024.09: MMNeuron was accepted by EMNLP 2024 Main.
  • 2024.06: One paper submited to EMNLP 2024.

🔬 Research Interests

  • Post-training and distillation: on-policy distillation, reasoning models, and efficient adaptation.
  • Multimodal representation learning: compact visual-document representations and generative embeddings.
  • Agent memory: persistent, retrievable, and multimodal memory for AI agents.
  • Trustworthy AI: text watermarking, machine unlearning, and model interpretability.

📝 Selected Publications

Selected publications are listed below; see Google Scholar for the full list.

Under review · 2026
SAMark overview

SAMark: A Self-Anchored Text Watermarking with Paragraph-Level Paraphrase Robustness

Jiahao Huo, Wenjie Qu, Yibo Yan, Kening Zheng, Jiaheng Zhang, Xuming Hu, Philip S. Yu, Mingxun Zhou

  • A self-anchored semantic watermarking framework designed to remain detectable under paragraph-level paraphrase attacks while preserving generation quality.

Paper | Citations: 0 · Code GitHub stars for SAMark · Project

ICLR 2026
PMark framework

PMark: Towards Robust and Distortion-Free Semantic-Level Watermarking with Channel Constraints

Jiahao Huo, Shuliang Liu, Bin Wang, Junyan Zhang, Yibo Yan, Aiwei Liu, Xuming Hu, Mingxun Zhou

  • A semantic watermarking method that encodes detectable structure through jointly constrained channels while preserving the language model’s sampling distribution.

Paper | Citations: 15 · Code GitHub stars for PMark · Project

Under review · 2026
CausalEmbed framework

CausalEmbed: Auto-Regressive Multi-Vector Generation in Latent Space for Visual Document Embedding

Jiahao Huo, Yu Huang, Yibo Yan, Ye Pan, Yi Cao, Mingdong Ou, Philip S. Yu, Xuming Hu

  • An autoregressive latent-space embedding model that represents visual documents with compact multi-vector sequences and supports controllable late-interaction retrieval.

Paper | Citations: 1 · Code GitHub stars for CausalEmbed · Models · Project

ACL 2025 Findings
MMUnlearner framework

MMUnlearner: Reformulating Multimodal Machine Unlearning in the Era of Multimodal Large Language Models

Jiahao Huo, Yibo Yan, Xu Zheng, Yuanhuiyi Lyu, Xin Zou, Zhihua Wei, Xuming Hu

  • A modality-aware unlearning framework that suppresses target visual concepts through saliency-guided, geometry-constrained updates while retaining textual knowledge and general visual capabilities.

Paper | Citations: 57 · Code GitHub stars for MMUnlearner · Project

ACL 2025 Industry Track · Oral
MathAgent paper overview

MathAgent: Leveraging a Mixture-of-Math-Agent Framework for Real-World Multimodal Mathematical Error Detection

Yibo Yan, Shen Wang, Jiahao Huo, Philip S. Yu, Xuming Hu, Qingsong Wen

  • A mixture-of-agents framework for multimodal mathematical error detection, combining image–text consistency validation, visual semantic interpretation, and integrative error analysis.

Paper | Citations: 39

Technical report · 2025
MemOS 2.0 Stardust overview

MemOS: A Memory OS for AI System

Zhiyu Li, Shichao Song, Chenyang Xi, Hanyu Wang, Chen Tang, Simin Niu, Ding Chen, Jiawei Yang, Chunyu Li, Qingchen Yu, Jihao Zhao, Yezhaohui Wang, Peng Liu, Zehao Lin, Pengyuan Wang, Jiahao Huo, Tianyi Chen, Kai Chen, Kehang Li, Zhen Tao, Huayi Lai, Hao Wu, Bo Tang, Zhenren Wang, Zhaoxin Fan, Ningyu Zhang, Linfeng Zhang, Junchi Yan, Mingchuan Yang, Tong Xu, Wei Xu, Huajun Chen, Haofen Wang, Hongkang Yang, Wentao Zhang, Zhi-Qin John Xu, Siheng Chen, Feiyu Xiong

  • An open-source memory operating system that provides persistent memory production, retrieval, and lifecycle management for AI agents and multimodal applications.

Paper | Citations: 136 · Code GitHub stars for MemOS · Project

EMNLP 2024 Main
MMNeuron framework

MMNeuron: Discovering Neuron-Level Domain-Specific Interpretation in Multimodal Large Language Models

Jiahao Huo, Yibo Yan, Boren Hu, Yutao Yue, Xuming Hu

  • An interpretability framework that identifies domain-specific neurons, traces their contribution to multimodal predictions, and validates their causal effect through controlled intervention.

Paper | Citations: 62 · Code GitHub stars for MMNeuron · Project

Additional publications:

  • Unveiling Language Routing Isolation in Multilingual MoE Models for Interpretable Subnetwork Adaptation - Kening Zheng, Wei-Chieh Huang, Jiahao Huo, Zhonghao Li, Henry Peng Zou, Yibo Yan, Xin Zou, Jungang Li, Junzhuo Li, Hanrong Zhang, Xuming Hu, Philip S. Yu. EMNLP 2026 Findings. Paper
  • ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models via Error Detection — Yibo Yan, Shen Wang, Jiahao Huo, Hang Li, Boyan Li, Jiamin Su, Xiong Gao, Yi-Fan Zhang, Tianlong Xu, Zhendong Chu, Aoxiao Zhong, Kun Wang, Hui Xiong, Philip S. Yu, Xuming Hu, Qingsong Wen. ACL 2026 Findings. Paper
  • Pierce the Mists, Greet the Sky: Decipher Knowledge Overshadowing via Knowledge Circuit Analysis — Haoming Huang, Yibo Yan, Jiahao Huo, Xin Zou, Xinfeng Li, Kun Wang, Xuming Hu. EMNLP 2025. Paper · Code GitHub stars for PhantomCircuit
  • EssayJudge: A Multi-Granular Benchmark for Assessing Automated Essay Scoring Capabilities of MLLMs — Jiamin Su, Yibo Yan, Fangteng Fu, Han Zhang, Jingheng Ye, Xiang Liu, Jiahao Huo, Huiyu Zhou, Xuming Hu. ACL 2025 Findings. Paper
  • Explainable and Interpretable Multimodal Large Language Models: A Comprehensive Survey — Yunkai Dang, Kaichen Huang, Jiahao Huo, Yibo Yan, Sirui Huang, Dongrui Liu, Mengxi Gao, Jie Zhang, Chen Qian, Kun Wang, Yong Liu, Jing Shao, Hui Xiong, Xuming Hu. Preprint. Paper
  • Memory in the Age of AI Agents — Yuyang Hu, Shichun Liu, Yanwei Yue, Guibin Zhang, Boyang Liu, Fangyi Zhu, Jiahang Lin, Honglin Guo, Shihan Dou, Zhiheng Xi, Senjie Jin, Jiejun Tan, Yanbin Yin, Jiongnan Liu, Zeyu Zhang, Zhongxiang Sun, Yutao Zhu, Hao Sun, Boci Peng, Zhenrong Cheng, Xuanbo Fan, Jiaxin Guo, Xinlei Yu, Zhenhong Zhou, Zewen Hu, Jiahao Huo, Junhao Wang, Yuwei Niu, Yu Wang, Zhenfei Yin, Xiaobin Hu, Yue Liao, Qiankun Li, Kun Wang, Wangchunshu Zhou, Yixin Liu, Dawei Cheng, Qi Zhang, Tao Gui, Shirui Pan, Yan Zhang, Philip Torr, Zhicheng Dou, Ji-Rong Wen, Xuanjing Huang, Yu-Gang Jiang, Shuicheng Yan. Preprint. Paper

📖 Education

💼 Work Experience

🧑‍⚖️ Academic Service

  • Conference reviewer: ARR (ACL/EMNLP 2026), AAAI 2025, SIGIR 2026, NeurIPS 2026, ICLR 2027.
  • Journal reviewer: IEEE TNNLS.