Publications

Author names marked with * denote equal contribution. Hyojun Go is shown in bold. Author order follows each official publication.

Topic
Thumbnail for Understanding, Accelerating, and Improving MeanFlow Training

Understanding, Accelerating, and Improving MeanFlow Training

Jin-Young Kim*, Hyojun Go*, Lea Bogensperger, Julius Erbach, Nikolai Kalischek, Federico Tombari, Konrad Schindler, Dominik Narnhofer

IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2026

Analyzed MeanFlow's training dynamics and developed a training scheme that improves one-step generation with 2.5× faster convergence.

Thumbnail for Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs

Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs

Hyungjin Chung, Hyelin Nam, Jiyeon Kim, Hyojun Go, Byeongjun Park, Junho Kim, Joonseok Lee, Seongsu Ha, Byung-Hoon Kim

IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Findings 2026

Thumbnail for VIST3A: Text-to-3D by Stitching a Multi-view Reconstruction Network to a Video Generator

VIST3A: Text-to-3D by Stitching a Multi-view Reconstruction Network to a Video Generator

Hyojun Go, Dominik Narnhofer, Goutam Bhat, Prune Truong, Federico Tombari, Konrad Schindler

International Conference on Learning Representations (ICLR) 2026 Oral

Unified a pretrained video diffusion model and a feed-forward 3D reconstruction model into a single end-to-end latent diffusion model that directly generates 3D worlds from text.

Thumbnail for SteerX: Creating Any Camera-Free 3D and 4D Scenes with Geometric Steering

SteerX: Creating Any Camera-Free 3D and 4D Scenes with Geometric Steering

Byeongjun Park*, Hyojun Go*, Hyelin Nam, Byung-Hoon Kim, Hyungjin Chung, Changick Kim

International Conference on Computer Vision (ICCV) 2025

Introduced zero-shot inference-time geometric steering with reconstruction-based rewards and particle filtering for camera-free 3D/4D generation.

Thumbnail for VideoRFSplat: Direct Scene-Level Text-to-3D Gaussian Splatting Generation with Flexible Pose and Multi-View Joint Modeling

VideoRFSplat: Direct Scene-Level Text-to-3D Gaussian Splatting Generation with Flexible Pose and Multi-View Joint Modeling

Hyojun Go*, Byeongjun Park*, Hyelin Nam, Byung-Hoon Kim, Hyungjin Chung, Changick Kim

International Conference on Computer Vision (ICCV) 2025

Thumbnail for SplatFlow: Multi-View Rectified Flow Model for 3D Gaussian Splatting Synthesis

SplatFlow: Multi-View Rectified Flow Model for 3D Gaussian Splatting Synthesis

Hyojun Go*, Byeongjun Park*, Jiho Jang, Jin-Young Kim, Soonwoo Kwon, Changick Kim

IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025

Built a unified rectified-flow model that generates and edits 3D Gaussian Splats by jointly modeling multi-view images, depths, and camera poses.

Thumbnail for Denoising Task Difficulty-based Curriculum for Training Diffusion Models

Denoising Task Difficulty-based Curriculum for Training Diffusion Models

Jin-Young Kim*, Hyojun Go*, Soonwoo Kwon, Hyun-Gyoon Kim

International Conference on Learning Representations (ICLR) 2025

Thumbnail for Switch Diffusion Transformer: Synergizing Denoising Tasks with Sparse Mixture-of-Experts

Switch Diffusion Transformer: Synergizing Denoising Tasks with Sparse Mixture-of-Experts

Byeongjun Park, Hyojun Go, Jin-Young Kim, Sangmin Woo, Seokil Ham, Changick Kim

European Conference on Computer Vision (ECCV) 2024

Thumbnail for TWLV-I: Analysis and Insights from Holistic Evaluation on Video Foundation Models

TWLV-I: Analysis and Insights from Holistic Evaluation on Video Foundation Models

Twelvelabs Team

Technical Report 2024

Thumbnail for HarmonyView: Harmonizing Consistency and Diversity in One-Image-to-3D

HarmonyView: Harmonizing Consistency and Diversity in One-Image-to-3D

Sangmin Woo*, Byeongjun Park*, Hyojun Go, Jin-Young Kim, Changick Kim

IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024

Thumbnail for Pegasus-v1 Technical Report

Pegasus-v1 Technical Report

Twelvelabs Team

Technical Report 2024

Thumbnail for Bridging Implicit and Explicit Geometric Transformation for Single-Image View Synthesis

Bridging Implicit and Explicit Geometric Transformation for Single-Image View Synthesis

Byeongjun Park*, Hyojun Go*, Changick Kim

IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) 2024

Thumbnail for Multi-Architecture Multi-Expert Diffusion Models

Multi-Architecture Multi-Expert Diffusion Models

Yunsung Lee*, JinYoung Kim*, Hyojun Go*, Myeongho Jeong, Shinhyeok Oh, Seungtaek Choi

AAAI Conference on Artificial Intelligence (AAAI) 2024

Thumbnail for Addressing Negative Transfer in Diffusion Models

Addressing Negative Transfer in Diffusion Models

Hyojun Go*, JinYoung Kim*, Yunsung Lee*, Seunghyun Lee*, Shinhyeok Oh, Hyeongdon Moon, Seungtaek Choi

Advances in Neural Information Processing Systems (NeurIPS) 2023

Reframed diffusion pre-training as multi-task learning and showed that harmonizing training across denoising tasks improves generation quality and convergence.

Thumbnail for Evaluation of Question Generation Needs More References

Evaluation of Question Generation Needs More References

Shinhyeok Oh*, Hyojun Go*, Hyeongdon Moon, Yunsung Lee, Myeongho Jeong, Hyun Seung Lee, Seungtaek Choi

Findings of the Association for Computational Linguistics (ACL) 2023

Thumbnail for Cross Encoding as Augmentation: Towards Effective Educational Text Classification

Cross Encoding as Augmentation: Towards Effective Educational Text Classification

Hyun Seung Lee, Seungtaek Choi, Yunsung Lee, Hyeongdon Moon, Shinhyeok Oh, Myeongho Jeong, Hyojun Go, Christian Wallraven

Findings of the Association for Computational Linguistics (ACL) 2023

Thumbnail for Towards Practical Plug-and-Play Diffusion Models

Towards Practical Plug-and-Play Diffusion Models

Hyojun Go*, Yunsung Lee*, Jin-Young Kim*, Seunghyun Lee, Myeongho Jeong, Hyun Seung Lee, Seungtaek Choi

IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2023

Thumbnail for Towards Flexible Inductive Bias via Progressive Reparameterization Scheduling

Towards Flexible Inductive Bias via Progressive Reparameterization Scheduling

Yunsung Lee*, Gyuseong Lee*, Kwangrok Ryoo*, Hyojun Go*, Jihye Park*, Seungryong Kim

European Conference on Computer Vision Workshop (ECCVW) 2022

Thumbnail for On the Effectiveness of Small Input Noise for Defending Against Query-based Black-box Attacks

On the Effectiveness of Small Input Noise for Defending Against Query-based Black-box Attacks

Junyoung Byun, Hyojun Go, Changick Kim

IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2022

Thumbnail for Geometrically Adaptive Dictionary Attack on Face Recognition

Geometrically Adaptive Dictionary Attack on Face Recognition

Junyoung Byun, Hyojun Go, Changick Kim

IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2022

Thumbnail for Fine-grained Multi-class Object Counting

Fine-grained Multi-class Object Counting

Hyojun Go, Junyoung Byun, Byeongjun Park, Myung-Ae Choi, Seunghwa Yoo, Changick Kim

IEEE International Conference on Image Processing (ICIP) 2021