Author: 이 재찬
[CVPR 2026] Back to Basics: Let Denoising Generative Models Denoise
안녕하세요. 이번 x-review는 최근 이미지 생성 분야에서 큰 주목을 받고 있는 논문인 “Back to Basics: Let Denoising Generative Models Denoise”입니다. MIT의 Tianhong Li와 Kaiming He가…
[arxiv 2026′] VLA-JEPA Enhancing Vision-Language-Action Model with Latent World Model
안녕하세요. 이번 x-review는 기존 VLA 아키텍쳐에 JEPA 기반 video world model 과의 결합을 다룬 논문입니다. 요즘 제가 다루는 LeJEPA기반 LeWM이 real-world robotic task나 흔히 사용되던…
[RSS 2026] Mimic Intent, Not Just Trajectories
안녕하세요. 이번 논문 리뷰는 RSS 2026′ MINT (Mimic Intent, Not Just Trajectories) 인데요, action chunk를 주파수 도메인에서 분해해서 intent(전역적인 행동 의도)와 execution(세부 실행 디테일)을 명시적으로…
[NeurIPS 2025] FIPER: Failure Prediction at Runtime for Generative Robot Policies
안녕하세요. 이번 논문 리뷰는 DP나 Flow Matching policy같은 generative IL policy가 runtime에서 task failure를 일으킬 때, 이를 failure data 없이 사전에 예측하는 방법론인 FIPER(Failure Prediction…
[arxiv 2026] LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
논문 정보 저자:Lucas Maes*¹, Quentin Le Lidec*², Damien Scieur¹·³, Yann LeCun², Randall Balestriero⁴1: Mila & Université de Montréal, 2: New York University, 3: Samsung SAIL,…
[CoRL 2025] Steering Your Diffusion Policy with Latent Space Reinforcement Learning
논문 정보 저자: Andrew Wagenmaker1, Mitsuhiko Nakamoto1, Yunchu Zhang2, Seohong Park1, Waleed Yagoub2, Anusha Nagabandi3, Abhishek Gupta2, Sergey Levine1* 1: UC Berkeley, 2: University of Washington, 3: Amazon 링크: https://arxiv.org/abs/2506.15799 프로젝트페이지: https://diffusion-steering.github.io/ 안녕하세요. 이번 논문…
[RSS 2025 Workshop] From Foresight to Forethought VLM-In-the-Loop Policy Steering via Latent Alignment
안녕하세요. 이번 논문 리뷰는 DP같은 generative robot policy가 deployment-time 에 다양한 실패를 보이는 문제를 해결하기 위한 runtime policy steering 방법론입니다. 특히 해당 실패를 DreamerV3 기반…
[NeurIPS 2025] Chain-of-Action: Trajectory Autoregressive Modeling for Robotic Manipulation
안녕하세요. 저번 세미나 시간에 발표로 들고 왔던 Chain of Action 논문을 리뷰로 남기기 위해 가져왔습니다. ByteDance Seed에서 제안한 액션 역방향 생성의 새로운 패러다임인데요. 기존의 액션을…
[arxiv 2025] Motus: A Unified Latent Action World Model
이번 리뷰는 논문 작업이 끝난 후 다음 연구 주제인 Long-horizon Task와 Failure Detection 분야를 서칭하던 중, 자극적인 제목에 끌려 보게되었습니다. Latent Action, World Model 을…
[이재찬] 2025년을 보내며
이번 회고글은 일요일 밤 자전거 길 위에서 가다서다 하며 핸드폰 메모장에 조각글처럼 모아놓은 생각들에서 시작되네요. 막상 회고글을 써볼까~하고 각 잡고 카페나 집에 죽치고 노트북 앞에만…
Q1. k-means 실험의 목적? >> DINOv2+k-means와 DINOv2+CA는 DINOSAURv2의 장점이 단순히 token을 줄였기 때문인지 확인하는 비교 실험입니다. 아마 리뷰어가 "그냥 토큰…