Hi, I’m Mingyi Deng (邓明熠), an MStat student at the University of Hong Kong. I received my B.S. in Statistics from Renmin University of China.
I work on LLM-based agents, especially interactive and user-centric agents: how agents should recognize ambiguity, ask effective clarifying questions, and learn from multi-turn user feedback. I’m currently an LLM Algorithm Research Intern at Alibaba.
Reach me at dengmingyi1219@163.com.
Publications
- InteractComp: Evaluating Search Agents With Ambiguous Queries · ICML 2026 · first author A benchmark for evaluating whether search agents can recognize query ambiguity and interact to resolve it.
- InfoPO: Information-Driven Policy Optimization for User-Centric Agents · ICML 2026 Turns per-turn information gain into a dense signal for multi-turn RL on user-centric agents.
- Foundation Protocol: A Coordination Layer for Agentic Society · arXiv 2026 A graph-first coordination layer for organizing agents, humans, tools, resources, and institutions in an open and accountable agentic society.
- ReCode: Unify Plan and Action for Universal Granularity Control · arXiv 2025 Unifies planning and action as executable code with universal granularity control.
Education
- MStat, The University of Hong Kong · Aug 2025 – Dec 2026
- B.S. in Statistics, Renmin University of China · Sep 2020 – Jul 2024
Internships
- Alibaba · LLM Algorithm Research Intern · Jul 2026 – Present
- Meituan Beam · LLM Research Algorithm Intern · Apr 2026 – Jun 2026
- DeepWisdom (深度赋智) · Research Intern · Jul 2025 – Feb 2026 LLM agent research: InteractComp, InfoPO, ReCode.
- Nanjing Xuming Private Fund · Quant Research Intern · Feb 2025 – May 2025 Tree-model / NN composite quant models; agent-driven strategy exploration.
