- 🔭 I’m currently working on RL-based fine-tuning for LLMs.
- 🏫 I’m currently studying at the School of Automation Science and Engineering, Xi’an Jiaotong University.
- 📫 Reach me at: wanghm@stu.xjtu.edu.cn / whm568019240@gmail.com
- 🖥️ Research interests: Reinforcement Learning, LLM
PhD student in Xi'an Jiaotong University
-
Xi'an Jiaotong University
-
07:54
(UTC +08:00) - https://scholar.google.com/citations?user=ODrcO0wAAAAJ&hl=en
Highlights
- Pro
Pinned Loading
-
ComplexOvercooked
ComplexOvercooked PublicA multi-agent cooperative benchmark environment with intra-episode task switching
Python 2
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.


