From ideas to reusable research systems.
The portfolio spans reasoning, reinforcement learning, interactive agents, environments, multimodal diffusion language models, and post-training infrastructure.
ReasonFlux
A family of systems for hierarchical reasoning, thought templates, trajectory-aware process reward models, and co-evolving coders.
GitHub ↗RLAnything
Dynamic environment, policy, and reward learning for general reinforcement learning and agent training.
GitHub ↗OpenClaw-RL
Fully asynchronous reinforcement learning for interactive agents operating across conversation, terminal, GUI, SWE, and tool-use settings.
GitHub ↗GenEnv
Environment construction and evolution for agent learning, exploration, and environment–policy co-evolution.
GitHub ↗MMaDA
Multimodal large diffusion language models and systems for unified multimodal generative modeling.
GitHub ↗dLLM-RL
Trajectory-aware post-training and reinforcement learning infrastructure for diffusion language models.
GitHub ↗LatentMAS
Efficient multi-agent collaboration through reasoning and communication in latent space.
GitHub ↗Gen-Verse
The broader open research community connecting models, environments, datasets, evaluation frameworks, and algorithms.
Organization ↗DIG
The academic research group connecting learning, evolution, and discovery into a unified long-term research program.
Research agenda →