Skip to content

Robotics and Computer Vision Lab

AI in Sensing, AI in Perception, AI in Action

  • About
    • History
    • Photo
    • Admission
  • Members
  • Publications
    • Patents
  • X-Review
  • X-Diary
  • Peer Review

Profile

김 영규

About Posts
2026년 상반기 회고 – 김영규
  • Posted on: 09/07/2026 –
  • Comments: 1 Comment
[ICRA 2026] VITRA : Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos
  • Posted on: 06/22/2026 –
  • Comments: 2 Comments
[arXiv 2026] ActiveMimic: Egocentric Video Pretraining with Active Perception
  • Posted on: 06/08/2026 –
  • Comments: 8 Comments
[arxiv 2025] Is Diversity All You Need for Scalable Robotic Manipulation?
  • Posted on: 06/01/2026 –
  • Comments: 8 Comments
[ICLR 2026 Workshop] World Action Models are Zero-shot Policies
  • Posted on: 05/26/2026 –
  • Comments: 2 Comments
[ICML 2026] DECO: Decoupled Multimodal Diffusion Transformer for Bimanual Dexterous Manipulation with a Plugin Tactile Adapter
  • Posted on: 05/18/2026 –
  • Comments: 4 Comments
GR00T : An Open Foundation Model for Generalist Humanoid Robots
  • Posted on: 05/04/2026 –
  • Comments: 6 Comments
[arXiv 2026] PokeVLA: Empowering Pocket-Sized Vision-Language-Action Model with Comprehensive World Knowledge Guidance
  • Posted on: 04/27/2026 –
  • Comments: 2 Comments
[arXiv 2026] Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models
  • Posted on: 04/12/2026 –
  • Comments: 4 Comments
[ICLR 2026] Emergent Dexterity via Diverse Resets and Large-Scale Reinforcement Learning
  • Posted on: 04/06/2026 –
  • Comments: 2 Comments
1 2 … 6 7 Older Posts

Conference Deadline

NEW POST

  • [ArXiv 2026]AORCHESTRA: Automating Sub-Agent Creation for Agentic Orchestration
  • [CVPR 2026] SEASON: Mitigating Temporal Hallucination in Video Large Language Models via Self-Diagnostic Contrastive Decoding
  • [COLM 2025] Training Large Language Model to Reason in a Continuous Latent Space
  • [arXiv 2026] ABot-N0: Technical Report on the VLA Foundation Model for Versatile Embodied Navigation
  • [arxiv 2025]visual-tactile 6D tracking

New Comment

  1. 신 인택 on [ICCV 2025] Perspective-Aware Reasoning in Vision-Language Models via Mental Imagery Simulation09/14/2026

    안녕하세요 주연님 좋은 댓글 감사합니다. 우선 생각하신 분석이 제가 논문 읽으면서 느낀것과 같습니다. 즉 저자가 주장하는 camera 시점의 bias 가…

  2. 신 인택 on [ICCV 2025] Perspective-Aware Reasoning in Vision-Language Models via Mental Imagery Simulation09/14/2026

    안녕하세요 우현님 좋은 댓글 감사합니다. 우선 제가 데이터셋을 확인해본건 아니지만 객체가 큐브인 경우가 perspective인 경우는 없을 것 같습니다. 그리고 그렇게…

  3. 신 인택 on [COLM 2025] Training Large Language Model to Reason in a Continuous Latent Space09/14/2026

    안녕하세요 우진님 좋은 댓글 감사합니다. CoT와 다르게 추가 학습이 필요합니다. (기존의 last hidden state가 reasoning 용도가 아니었기에) 분석은 중간의 정성적…

  4. 신 인택 on [COLM 2025] Training Large Language Model to Reason in a Continuous Latent Space09/14/2026

    안녕하세요 희승님 좋은 댓글 감사합니다. 각 질문에 대해 답변드리자면 1. 정확히 그렇게 이해하시면 될 것 같습니다. 2. continuous thought 을…

  5. 신 인택 on [COLM 2025] Training Large Language Model to Reason in a Continuous Latent Space09/14/2026

    안녕하세요 유진님 좋은 댓글 감사합니다. 저자의 제안 방법론이 사실 모든 일반적인 상황에서 통하는건 아니라고 저도 생각합니다. 실제로 벤치마크별로 성능이 많이…

  • Sign-in
  • RCV-Calendar
  • RCV-Github
  • Paper R/W
    • Arxiv
    • Deadline
    • Overleaf
  • Coding
    • OnlineJudge
    • Kaggle

포기하지 않는 강한 집념 만이 작은 차이를 만든다.

Design by SejongRCV