I received my Ph.D. from the University of Science and Technology of China (USTC) in November 2025. My research interests include shortcut learning, robustness and generalization, post-training, reward hacking, and legal intelligence.
In December 2025, I joined the LongCat-Interaction team, where I work on post-training, agentic reinforcement learning, and self-evolving agents. We are actively recruiting — feel free to reach out if you are interested in joining us!
Publications
-
2026
-
2025
-
2025
-
2025
-
2025
-
2025
-
2024
-
2024
-
2024
-
2024
-
2023
-
2023
-
2022
-
2021
* Equal contribution. Full list on Google Scholar.