Skip to content

Robotics and Computer Vision Lab

AI in Sensing, AI in Perception, AI in Action

  • About
    • History
    • Photo
    • Admission
  • Members
  • Publications
    • Patents
  • X-Review
  • X-Diary
  • Peer Review

Profile

임 근택

About Posts
[arXiv 2022] InternVideo : General Video Foundation Models via Generative and Discriminative Learning
  • Posted on: 01/16/2023 –
  • Comments: No Comments
[arXiv 2022] Movie2Scenes : Learning Scene Representations Using Movie Similarities
  • Posted on: 01/14/2023 –
  • Comments: 4 Comments
[ICLR 2022] Uniformer : Unified Transformer For Efficient Spatiotemporal Representation Learning
  • Posted on: 01/12/2023 –
  • Comments: No Comments
[ICML 2021] An Image is Worth 16×16 Words : Transformer for image recognition at scale
  • Posted on: 01/06/2023 –
  • Comments: 5 Comments
[NIPS 2022] VideoMAE : Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training
  • Posted on: 01/04/2023 –
  • Comments: 2 Comments
[CVPR 2022] Masked Autoencoders Are Scalable Vision Learners
  • Posted on: 01/04/2023 –
  • Comments: 7 Comments
<2022년 RCV 연구실 생활을 마무리 하며>
  • Posted on: 01/01/2023 –
  • Comments: No Comments
[CVPR 2022] Probabilistic Representations for Video Contrastive Learning (Part.1)
  • Posted on: 12/05/2022 –
  • Comments: No Comments
CVPR 2023 논문 작성기
  • Posted on: 11/20/2022 –
  • Comments: 3 Comments
Optimization Theory (Gradient Descent – Convergence Analysis)
  • Posted on: 11/14/2022 –
  • Comments: No Comments
Newer Posts 1 2 … 4 5 6 … 10 11 Older Posts

Conference Deadline

NEW POST

  • [ArXiv 2026]AORCHESTRA: Automating Sub-Agent Creation for Agentic Orchestration
  • [CVPR 2026] SEASON: Mitigating Temporal Hallucination in Video Large Language Models via Self-Diagnostic Contrastive Decoding
  • [COLM 2025] Training Large Language Model to Reason in a Continuous Latent Space
  • [arXiv 2026] ABot-N0: Technical Report on the VLA Foundation Model for Versatile Embodied Navigation
  • [arxiv 2025]visual-tactile 6D tracking

New Comment

  1. 신 인택 on [ICCV 2025] Perspective-Aware Reasoning in Vision-Language Models via Mental Imagery Simulation09/14/2026

    안녕하세요 주연님 좋은 댓글 감사합니다. 우선 생각하신 분석이 제가 논문 읽으면서 느낀것과 같습니다. 즉 저자가 주장하는 camera 시점의 bias 가…

  2. 신 인택 on [ICCV 2025] Perspective-Aware Reasoning in Vision-Language Models via Mental Imagery Simulation09/14/2026

    안녕하세요 우현님 좋은 댓글 감사합니다. 우선 제가 데이터셋을 확인해본건 아니지만 객체가 큐브인 경우가 perspective인 경우는 없을 것 같습니다. 그리고 그렇게…

  3. 신 인택 on [COLM 2025] Training Large Language Model to Reason in a Continuous Latent Space09/14/2026

    안녕하세요 우진님 좋은 댓글 감사합니다. CoT와 다르게 추가 학습이 필요합니다. (기존의 last hidden state가 reasoning 용도가 아니었기에) 분석은 중간의 정성적…

  4. 신 인택 on [COLM 2025] Training Large Language Model to Reason in a Continuous Latent Space09/14/2026

    안녕하세요 희승님 좋은 댓글 감사합니다. 각 질문에 대해 답변드리자면 1. 정확히 그렇게 이해하시면 될 것 같습니다. 2. continuous thought 을…

  5. 신 인택 on [COLM 2025] Training Large Language Model to Reason in a Continuous Latent Space09/14/2026

    안녕하세요 유진님 좋은 댓글 감사합니다. 저자의 제안 방법론이 사실 모든 일반적인 상황에서 통하는건 아니라고 저도 생각합니다. 실제로 벤치마크별로 성능이 많이…

  • Sign-in
  • RCV-Calendar
  • RCV-Github
  • Paper R/W
    • Arxiv
    • Deadline
    • Overleaf
  • Coding
    • OnlineJudge
    • Kaggle

포기하지 않는 강한 집념 만이 작은 차이를 만든다.

Design by SejongRCV