Jihwan Kim
Hi 👋, I am Jihwan. I am currently an integrated M.S./Ph.D. student at KAIST GSAI, advised by Professor Jaegul Choo. My research interests lie in the evaluation of agents, speech and audio processing, and multilingual LLMs.
I have worked on:
- Building an evaluation suite for web browsing agents;
- Building effective audio detection and contributing an open-source full-duplex audio pipeline;
- Synthesizing data that converts unverifiable multilingual RLVR tasks into verifiable ones.
Fun fact: I finally built something with my day6.ai.kr domain 🙂 — come and have a look!
Please feel free to contact me via email at jihvvan.kim@kaist.ac.kr.
News
| Sep 08, 2026 | Two papers accepted to EMNLP 2026 — cross-lingual RLVR via code-switched task synthesis to the main conference, and a next-scene-prediction benchmark for omni-models to Findings. 🎉 |
|---|---|
| May 20, 2026 | Two papers accepted to ACL 2026 — Adaptive Table Retrieval to Findings, and Sommelier to the Industry Track. |
| Mar 26, 2026 | New preprint on speech: Sommelier — scalable, open multi-turn audio pre-processing for full-duplex speech language models. |
| Jan 22, 2026 | LiveWeb-IE was accepted to ICLR 2026 as a poster. 🎉 |
| Jan 06, 2026 | The HyperCLOVA X 8B Omni and 32B Think technical reports are out on arXiv. |
Selected Publications
- Can Omni-models Imagine the Future? A Next Scene Prediction Benchmark for Multimodal World ModelingIn Findings of the Association for Computational Linguistics: EMNLP 2026, 2026
-