Minbeom Kim
Hi! I am a Research Scientist at SK hynix Research, working on Recursive Self-Improvement (RSI) of AI scientists — ensuring that RSI remains safe and controllable, and preventing harm from misalignment in the increasingly capable agents that emerge from it.
I received my Ph.D. from Seoul National University, advised by Prof. Kyomin Jung. During my Ph.D., I developed Distributional Matching RL — preserving diversity and mitigating reward hacking in language model post-training, thereby enabling scalable test-time scaling of AI agents. I was also a student researcher at Google, and interned at LawZero and NAVER Labs Europe.
Please feel free to reach out to me for a discussion here research@minbeom.kim!
news
| Aug 28, 2026 | It is a great honor to complete my Ph.D.🎓 and receive the department’s 1st dissertation honor, the Best Ph.D. Dissertation Award🥇. |
|---|---|
| Jul 20, 2026 | I have joined SK hynix Research as a Senior Research Scientist. I will work on Recursive Self-Improvement of AI Scientists beyond Memory Constraints. |
| May 01, 2026 | Our 🛡️ CausalArmor and 🕹️ PACED-RL are accepted to ICML 2026! See you in Seoul. 🇰🇷 |
| Oct 02, 2025 | I have joined Google as a Student Researcher. |
| Sep 18, 2025 | Our Syntra is accepted to NeurIPS 2025!!! See you in San Diego, USA. 🇺🇸 |
| Aug 20, 2025 | Our 🏎️ Drift and 🤖 ReflAct are accepted to EMNLP 2025!!! See you in China. 🇨🇳 |
| Jun 01, 2025 | I have joined 🧪LawZero founded by Yoshua Bengio as a research intern!! I will contribute to the Scientist AI project. Watch this wonderful journey [🍿TED Link]! |
| Jan 23, 2025 | Our 🛡️GUARD is accepted to ICLR 2025!! See you in Singapore. 🇸🇬 |







