- 👋 Hi, I’m Kai Yang(杨恺).
- 👀 Research Focus: Large Language Models (LLM) & Reinforcement Learning (RL).
- 🎓 M.S. in Artificial Intelligence, Tsinghua University.
- 💼 Tencent Hunyuan X Team | Research Engineer, specializing in RL for LLMs.
- 📫 Contact: yangkaisigsrl@gmail.com
Hi, I'm Kai Yang. I earned my master's degree from Tsinghua in 2025. I am currently working at Tencent Hunyuan, specializing in LLM model fusion and OPD.
-
Hunyuan R3, Tencent
- Shenzhen, China
-
12:37
(UTC +08:00) - https://yk7333.github.io/
Pinned Loading
-
-
-
TaskAllocation
TaskAllocation Public[EAAI] A two-stage reinforcement learning-based approach for multi-entity task allocation.
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.
