Ph.D. Student @ CAS

Ming Ma

LLM & Agents 🤖 Brain-Inspired AI 🧠 Photography 📷 Skateboarding 🛹
Portrait of Ming Ma
Internship Experience
Alibaba Tongyi Lab 2026.02 - Present
Research Intern
Qwen-Intelligence Team: Agentic Post train
Ant Group 2025.10 - 2026.01
Research Intern
Ling Team: Pre-training Quality
Microsoft Research Asia (MSRA) 2025.07 - 2025.10
Research Intern
Multi-Agent & Debugging
Papers
  1. A closed-loop intervention–validation framework that auto-debugs LLM multi-agent systems beyond passive failure-log analysis.
  2. A co-evolving dual-graph architecture (outline + knowledge) for open-ended deep-research agents that detects gaps and steers retrieval.
  3. Reveals that in-context learning relies on distributed local task vectors carried by label words rather than a single global encoding.
  4. A bidirectional question–answer coherence filter for selecting high-value synthetic code-instruction data.
  5. A training-free intermediate-layer decoding method enabling accurate early exit on single-token LLM tasks.
  6. Uses environment checks to assign progress credit to intermediate actions in long-horizon agentic reinforcement learning without an additional reward model.
  7. Explores the planning capabilities of multimodal models and evaluates their performance on agentic tasks.
  8. OmniMemBench: Towards Scalable Evaluation of Long-Term Omni-Modal Agent Memory (NeurIPS 2026)
Education
Chinese Academy of Sciences (CAS)
Ph.D. in Computational Neuroscience
2022.09 - Present
Shandong University
B.S. in Automation
2019.03 - 2022.06
Beijing Institute of Technology
Exchange Student in Automation
2019.09 - 2020.06
Naval Aviation University
Undergraduate in Mechanical and Electronic Engineering
2017.08 - 2019.03
Life & Photography
Writing & Media
Gaming
VALORANT VALORANT Rank Profile
Steam Steam Steam Profile