Skip to content

Research

Reliable agents for open-ended environments.

I investigate the evaluation principles and system mechanisms that make intelligent agents auditable, adaptive, and grounded across long-horizon interactions. My work spans action execution, persistent memory, multi-agent simulation, and multimodal embodied reasoning.

01

Reliable & Auditable LLM Agents

Evaluation and control for tool-using and multi-agent systems, with an emphasis on action semantics, order sensitivity, progress attribution, replayability, and failure diagnosis.

02

Long-Term Memory & Personalization

Mechanisms for agents to acquire, update, retrieve, and forget long-term memories while preserving user preferences, temporal consistency, and controllable behavior.

03

Multi-Agent & Social Simulation

Executable environments for studying coordination, interaction, and emergent behavior through traceable world-state transitions and reproducible counterfactual experiments.

04

Embodied & 3D Intelligence

Multimodal models that connect vision, language, and geometry for tiny-object perception, pose understanding, 3D scene reasoning, and embodied decision-making.

Publications

Research outputs

Paper links lead to the corresponding arXiv records or publisher pages.