I am Ke Xu, an Algorithm Engineer II at Huawei 2012 Laboratory. I started with RLHF training infrastructure (GRPO and DAPO on Ascend NPU clusters), and my focus has since moved to agent systems: multi-agent scheduling for scientific discovery and, most recently, recursive self-improvement (RSI).
I completed my M.S. in Computer Science at the Computer Science and Engineering Department, University of California San Diego (UCSD), where my concentration was Machine Learning.
Prior to that, I graduated from a dual-degree program jointly held by Zhejiang University and the University of Illinois at Urbana-Champaign, with major in Electrical Engineering and minor in Computer Science, where I was a member of Engineering National Graduate Institutional Name Exchange (ENGINE), a participant of Research Experiences for Undergraduate (REU) supported by National Science Foundation (NSF) 1947135, advised by Prof. Hanghang Tong.
Previously, I was an Algorithm Engineer Intern at Ant Group, where I worked on the application of Large Language Models (LLM) in security and risk management domains such as Anti-Money Laundering (AML).
Research Interests: Reinforcement learning and agent systems, especially recursive self-improvement (RSI).
Designed an active-learning strategy for KGQA that fuses graph centrality, information density, and node uncertainty.
Investigated reflective Abstract State Machines (ASMs) for security.
Built blacklist/whitelist safety classification and knowledge-graph-based intent and emotion recognition for counseling dialogue.
Away from the terminal, I love exploring the world — 0 countries and counting. visited planned not yet
Photos from my travels.
Under maintenance