Dongwon Kim
I am a postdoctoral researcher at KAIST, working with Prof. Jeany Son. I previously completed my BS and PhD at POSTECH CVLab where I worked with Prof. Suha Kwak. My research centers on whether machines can learn representations with the right abstraction and hierarchy, spanning work in compositional representation (SelfMod), multi-modalities (DivE, SaG, MaskGen), and world model (ongoing, CompACT).
24 Sep 2026 A paper on generation-friendly image tokenization is accepted at NeurIPS 2026 (arXiv).
all news- 24 SepA paper on generation-friendly image tokenization is accepted at NeurIPS 2026 (arXiv).
- 20 AprI gave invited talks at Google DeepMind and Yonsei University, about CompACT and abstraction for world model.
- 22 FebA paper on world model is accepted at CVPR 2026 (CompACT).
- 01 FebI won POSTECH CSE Best Research Award 2025, recognizing the best research among PhD graduates.
- 14 JulA paper on 1-dimensinal tokenization and text-to-image generation is accepted at ICCV 2025 (MaskGen).
- 14 JulI have completed my defense - now I’m officially a Ph.D (thesis title: “Learning Compositional Visual Representations for Vision-Language Understanding and Generation”).
- 26 SepA paper about object-centric learning is accepted at NeurIPS 2024.
- 10 JunOne paper is accepted at ECCV 2024.
- 10 JunJoined Bytedance (San Jose, CA) as a research intern
- 14 MarOur paper about referring image segmentation has been accepted at NAACL 2024 main track.
- 07 DecI am really grateful to have been granted the POSTECHIAN fellowship!
- 15 JulA paper about weakly-supervised referring image segmentation is accepted at ICCV 2023.
- 21 MarA paper about cross-modal retrieval is selected as a highlight at CVPR 2023!
Publications
Highlighted: first, co-first, or co-corresponding author · * equal contribution · † co-corresponding
-
arXiv 2026 V+LEventCoT: Event-centric Video Chain-of-thought for Reasoning Temporal Localization
-
-
NeurIPS 2026 GENStructured State-Space Regularization for Generation-Friendly Image Tokenization
-
-
ECCV 2024 V+LPLOT: Text-based Person Search with Part Slot Attention for Corresponding Part Discovery
-
Experience
| 2025 – Now |
Postdoctoral researcher · KAIST, Daejeon, KR
InnoCORE-LLM, PI: Jeany Son
|
|---|---|
| 2024 |
Research Intern · Fundamental Research Team, ByteDance SEED, San Jose, US Developed efficient text-to-image generative model using 1D tokens (MaskGen)
|
Honors and Awards
|
POSTECH CSE Best Research Award, POSTECH, 2025 POSTECHIAN Fellowship, POSTECH, 2023
BK21 Best Paper Award, POSTECH GSAI, 2023
Qualcomm Innovation Fellowship Winner, Qualcomm Korea Corp., 2022
NAVER x POSTECH AI DAY The 2nd and 3rd Prize, 2022
Qualcomm Innovation Fellowship Winner, Qualcomm Korea Corp., 2021
IPIU Best Paper Award, 2021
|
Professional Services
Reviewer
|