Minbeom Kim
Hi! I am a Research Scientist at SK hynix Research, working on Recursive Self-Improvement (RSI) of AI scientists โ ensuring that RSI remains safe and controllable, and preventing harm from misalignment in the increasingly capable agents that emerge from it.
I received my Ph.D. from Seoul National University, advised by Prof. Kyomin Jung. During my Ph.D., I developed Distributional Matching RL โ preserving diversity and mitigating reward hacking in language model post-training, thereby enabling scalable test-time scaling of AI agents. I was also a student researcher at Google, and interned at LawZero and NAVER Labs Europe.
Please feel free to reach out to me for a discussion here research@minbeom.kim!
news
| Jul 20, 2026 | I have joined SK hynix Research as a Senior Research Scientist. I will work on Recursive Self-Improvement of AI Scientists beyond Memory Constraints. |
|---|---|
| May 01, 2026 | Our ๐ก๏ธ CausalArmor and ๐น๏ธ PACED-RL are accepted to ICML 2026! See you in Seoul. ๐ฐ๐ท |
| Oct 02, 2025 | I have joined Google as a Student Researcher. |
| Sep 18, 2025 | Our Syntra is accepted to NeurIPS 2025!!! See you in San Diego, USA. ๐บ๐ธ |
| Aug 20, 2025 | Our ๐๏ธ Drift and ๐ค ReflAct are accepted to EMNLP 2025!!! See you in China. ๐จ๐ณ |
| Jun 01, 2025 | I have joined ๐งชLawZero founded by Yoshua Bengio as a research intern!! I will contribute to the Scientist AI project. Watch this wonderful journey [๐ฟTED Link]! |
| Jan 23, 2025 | Our ๐ก๏ธGUARD is accepted to ICLR 2025!! See you in Singapore. ๐ธ๐ฌ |
| Jan 22, 2025 | Our 3 papers are accepted to NAACL 2025!!! See you in Albuquerque, USA. ๐บ๐ธ |







