Minbeom Kim

face2.jpg

Hi! I am a Research Scientist at SK hynix Research, working on Recursive Self-Improvement (RSI) of AI scientists โ€” ensuring that RSI remains safe and controllable, and preventing harm from misalignment in the increasingly capable agents that emerge from it.

I received my Ph.D. from Seoul National University, advised by Prof. Kyomin Jung. During my Ph.D., I developed Distributional Matching RL โ€” preserving diversity and mitigating reward hacking in language model post-training, thereby enabling scalable test-time scaling of AI agents. I was also a student researcher at Google, and interned at LawZero and NAVER Labs Europe.

Please feel free to reach out to me for a discussion here research@minbeom.kim!

news

Jul 20, 2026 I have joined SK hynix Research as a Senior Research Scientist. I will work on Recursive Self-Improvement of AI Scientists beyond Memory Constraints.
May 01, 2026 Our ๐Ÿ›ก๏ธ CausalArmor and ๐Ÿ•น๏ธ PACED-RL are accepted to ICML 2026! See you in Seoul. ๐Ÿ‡ฐ๐Ÿ‡ท
Oct 02, 2025 I have joined Google as a Student Researcher.
Sep 18, 2025 Our Syntra is accepted to NeurIPS 2025!!! See you in San Diego, USA. ๐Ÿ‡บ๐Ÿ‡ธ
Aug 20, 2025 Our ๐ŸŽ๏ธ Drift and ๐Ÿค– ReflAct are accepted to EMNLP 2025!!! See you in China. ๐Ÿ‡จ๐Ÿ‡ณ
Jun 01, 2025 I have joined ๐ŸงชLawZero founded by Yoshua Bengio as a research intern!! I will contribute to the Scientist AI project. Watch this wonderful journey [๐ŸฟTED Link]!
Jan 23, 2025 Our ๐Ÿ›ก๏ธGUARD is accepted to ICLR 2025!! See you in Singapore. ๐Ÿ‡ธ๐Ÿ‡ฌ
Jan 22, 2025 Our 3 papers are accepted to NAACL 2025!!! See you in Albuquerque, USA. ๐Ÿ‡บ๐Ÿ‡ธ

education

work experience