Jinyoung Park

me.jpg

CV

Welcome to my research profile!

I am a Ph.D. candidate in CILab at KAIST EE, advised by Professor Changick Kim. Previously, I was a research intern at Microsoft Research Asia.

My research focuses on multimodal AI systems that perceive, understand, and act in complex real-world environments. My work spans multi-sensor fusion, vision-language and video understanding, and efficient sequence modeling with state space models (SSMs). More recently, I have been developing benchmarks and evaluation methods for multimodal agents that create and edit real-world documents in applications such as Microsoft Word.

If you are interested in collaborating, feel free to reach out :)

news

Apr 29, 2026 One paper got accepted to ICIP 2026! See you in Tampere :finland:
Dec 15, 2025 My internship at Microsoft Research Asia has been extended for another 6 months! :sparkles:
Jul 01, 2025 Started a research internship at Microsoft Research Asia! :wave:
Dec 09, 2024 One paper got accepted to AAAI!
Jul 01, 2024 Two papers(including VideoMamba) got accepted to ECCV! :sparkles: :smile:

selected publications

  1. Towards Efficient Vision State Space Models via Token Merging
    Jinyoung Park, Minseok Son, and Changick Kim
    ICIP, 2026
  2. Difficulty-aware Balancing Margin Loss for Long-tailed Recognition
    Minseok Son, Inyong Koo, Jinyoung Park, and Changick Kim
    AAAI, 2025
  3. VideoMamba: Spatio-Temporal Selective State Space Model
    Jinyoung Park, Hee-Seon Kim, Kangwook Ko, Minbeom Kim, and 1 more author
    ECCV, 2024
  4. Flow-Assisted Motion Learning Network for Weakly-Supervised Group Activity Recognition
    Muhammad Adi Nugroho, Sangmin Woo, Sumin Lee, Jinyoung Park, and 3 more authors
    ECCV, 2024
  5. Sketch-based Video Object Localization
    Sangmin Woo, So-Yeong Jeon, Jinyoung Park, Minji Son, and 2 more authors
    WACV, 2024
  6. Anchoring Vision and Language Knowledge for Weakly Supervised Group Activity Recognition
    Muhammad Adi Nugroho, Jinyoung Park, Donguk Kim, and Changick Kim
    VCIP, 2024
  7. Multi-modal Social Group Activity Recognition in Panoramic Scene
    Donguk Kim, Sumin Lee, Sangmin Woo, Jinyoung Park, and 2 more authors
    VCIP, 2023
  8. Rainunet for super-resolution rain movie prediction under spatio-temporal shifts
    Jinyoung Park, Minseok Son, Seungju Cho, Inyoung Lee, and 1 more author
    NeurIPS Weather4Cast Challenge, 2022
  9. Nowformer: A Locally Enhanced Temporal Learner for Precipitation Nowcasting
    Jinyoung Park, Inyoung Lee, Minseok Son, Seungju Cho, and 1 more author
    NeurIPS Tackling Climate Change with Machine Learning Workshop, 2022
  10. Dat: Domain adaptive transformer for domain adaptive semantic segmentation
    Jinyoung Park, Minseok Son, Sumin Lee, and Changick Kim
    ICIP, 2022
  11. Explore-and-match: Bridging proposal-based and proposal-free with transformer for sentence grounding in videos
    Sangmin Woo, Jinyoung Park, Inyong Koo, Sumin Lee, and 2 more authors
    arXiv, 2022