Yoojin Jang
I’m Yoojin Jang (장유진), an Integrated M.S./Ph.D. student at UNIST AIGS, advised by Prof. Jaejun Yoo at LAIT. My research focuses on multimodal generation and evaluation, particularly for audio and video. I am interested in developing multimodal generation and editing models, as well as benchmarks for evaluating their quality and cross-modal consistency.
Research Interests
- Multimodal Generation & Editing: Audio–video–text consistent, modality-aware generation and editing
- Data-Centric Learning & Robustness: Benchmark redesign, data imbalance, synthetic-based data, and robust model evaluation
- Multimodal LLMs: Multi-event semantic understanding, temporal reasoning, and real-world multimodal comprehension
News
| Apr 15, 2026 | AVENUE (Audio-Video EditiNg Understanding and Evaluation) — co-first-authored with Hayeon Kim — accepted to the Learning to Listen workshop at ICML 2026. |
|---|---|
| Mar 20, 2026 | “Rethinking Video-and-Text-to-Audio Generation through Multimodal Coverage” (2nd author) accepted to the Sight and Sound workshop at CVPR 2026. |
International conference
- CVPRWRethinking Video-and-Text-to-Audio Generation through Multimodal CoverageIn Sight and Sound Workshop, CVPR, 2026