Youngmin
Kim.

M.S. Student, KAIST AI · BISPL · Advisor: Prof. Jong Chul Ye

I am an M.S. student at KAIST AI (BISPL) advised by Prof. Jong Chul Ye. I work on 3D vision and generative modeling — camera-controlled video generation, diffusion-based image enhancement, and vision-language-action models — with an eye toward robotics and embodied AI.

Youngmin Kim

latest

What's new

  1. Two more preprints are on arXiv, both under review: FastOPD, an on-policy distillation framework for lightweight VLA deployment, and DriftOPD, sequence-level reverse-KL distillation for one-step VLA policies.

  2. Three new preprints are on arXiv, all under review: Dynamic Predictive Planning for dynamic manipulation, Action Upcycling for policy acceleration, and Adjoint Guidance Flow for VLA critic guidance.

  3. CRePE was accepted as a poster at NeurIPS 2026.

  4. Starting my M.S. at the Kim Jaechul Graduate School of AI, KAIST, joining BISPL under Prof. Jong Chul Ye.

  5. CRePE was accepted to the ECCV 2026 Workshop on 3D in the Era of World Models, in Malmö this September.

  6. Received my B.S. in Computer Science and Engineering from Korea University — early graduation in 3.5 years.

  7. Preprint on DMD-augmented unpaired neural Schrödinger bridges for ultra-low-field MRI enhancement is on arXiv.

  8. Our ultra-low-field brain MRI enhancement method was presented at the MICCAI ULF-EnC Challenge Workshop.

  9. Received the Academic Excellence Award from Korea University.

background

Where I come from, and what I care about.

I am a first-year M.S. student at the Kim Jaechul Graduate School of AI, KAIST, advised by Prof. Jong Chul Ye at BISPL. I received my B.S. in Computer Science and Engineering from Korea University in August 2026, graduating early in 3.5 years.

Before joining BISPL full-time, I interned there for over a year working on ultra-low-field MRI enhancement and camera-controlled video generation, and earlier studied Kubernetes scheduling at the Distributed and Cloud Computing Lab at Korea University.

Languages
Python, C++
Frameworks
PyTorch
Simulation
MuJoCo, Isaac Sim, LIBERO
Models
VLA, diffusion models, video generation models
  1. 01

    3D Vision & World Models

    Geometry-aware representations and camera-controlled video generation as a step toward world action models.

  2. 02

    Robotics & VLA

    Vision-language-action policies and simulation (MuJoCo, Isaac Sim, LIBERO) for embodied agents.

  3. 03

    Diffusion & Generative Modeling

    Distribution matching, Schrödinger bridges, and diffusion models for image and video synthesis.

  4. 04

    Medical Imaging

    Ultra-low-field MRI enhancement and medical foundation models that work under real scanner constraints.

research

Selected papers

  1. fig. 1 — robotics
    arXiv 20262026

    Dynamic Manipulation with World-Action Models via Counterfactual Planning

    Sunwoo Park*, Wonbin Lee*, Seonghyun Jin*, Youngmin Kim*, Jangho Park, Jong Chul Ye

    Dynamic Predictive Planning treats manipulation of moving targets as counterfactual planning: the world-action model's own rollout estimates when an interaction will happen and where the target will be, and a counterfactual observation lets the policy invoke a skill it already has instead of improvising a recovery. Runs in real time on a single consumer GPU with no training on dynamic data.

  2. fig. 2 — video generation
    NeurIPS 20262026

    CRePE: Curved Ray Expectation Positional Encoding for Unified-Camera-Controlled Video Generation

    Seonghyun Jin*, Youngmin Kim*, Sunwoo Park*, Jong Chul Ye

    A positional encoding that models curved rays so a single video diffusion model can be controlled by pinhole, fisheye and panoramic cameras alike. An earlier version appeared at the ECCV 2026 Workshop on 3D in the Era of World Models.

timeline

Education & experience

education

  1. Sep. 2026 – PresentM.S.

    KAIST, Kim Jaechul Graduate School of AI

    M.S. in Artificial Intelligence · BISPL · Advisor: Prof. Jong Chul Ye

    Seoul, Korea · Concentration: 3D Vision and Robotics

  2. Feb. 2023 – Aug. 2026B.S.

    Korea University

    B.S. in Computer Science and Engineering · Advisor: Prof. Jaehoon Lee

    Seoul, Korea · GPA 4.3 / 4.5 · Early graduation in 3.5 years · Academic Excellence Award

experience

  1. Sep. 2026 – PresentResearch

    BISPL, KAIST · Graduate Researcher

    Advisor: Prof. Jong Chul Ye

    Seoul, Korea

    • 3D vision and generative modeling for robotics.
  2. Jun. 2025 – Aug. 2026Research

    BISPL, KAIST · Research Intern

    Advisor: Prof. Jong Chul Ye

    Seoul, Korea

    • Camera-controlled video generation, ultra-low-field MRI enhancement, medical foundation models.
  3. Jun. 2024 – Jun. 2025Research

    Distributed and Cloud Computing Lab, Korea University · Research Intern

    Advisor: Prof. Heonchang Yu

    Seoul, Korea

    • Kubernetes scheduler performance analysis (KCC 2025).

projects

Things I've built

RoboticsJul. 2026 – Present

DPP — Dynamic Manipulation with World-Action Models

Counterfactual planning for moving targets: the world-action model's own rollout predicts when contact happens and where the target will be, so the policy reuses a skill it already has. Real time on one consumer GPU.

BISPL, KAISTproject page ↗
Video GenerationJun. 2026 – Aug. 2026

CRePE — Unified-Camera-Controlled Video Generation

Curved Ray Expectation Positional Encoding so one video model follows pinhole, fisheye and panoramic camera trajectories.

BISPL, KAISTpaper ↗
Medical ImagingJun. 2025 – Present

ULF to High-field MRI Enhancement

Diffusion, residual-artifact suppression and unpaired Schrödinger-bridge models that lift 64 mT scans toward 3 T quality.

BISPL, KAISTpreprint ↗
Foundation ModelsAug. 2025 – Mar. 2026

Medical Foundational Model

Large-scale pretraining and adaptation of a foundation model for medical imaging tasks.

BISPL, KAIST

05 — get in touch

Let's talk about 3D vision,
video models, or robots.