Skip to content

Robotics and Computer Vision Lab

AI in Sensing, AI in Perception, AI in Action

  • About
    • History
    • Photo
    • Admission
  • Members
  • Publications
    • Patents
  • X-Review
  • X-Diary
  • Peer Review

Profile

홍 주영

About Posts
[Arxiv 2026] RANKVIDEO: Reasoning Reranking for Text-to-Video Retrieval
  • Posted on: 03/28/2026 –
  • Comments: 2 Comments
[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
  • Posted on: 03/21/2026 –
  • Comments: No Comments
EV-5, VLM2Vec, VLM2Vec-V2: Generative MLLMs as Embedding Models
  • Posted on: 03/15/2026 –
  • Comments: 5 Comments
[ICLR 2023] CLIP-ViP: Adapting Pre-trained Image-Text Model to Video-Language Alignment
  • Posted on: 03/08/2026 –
  • Comments: 2 Comments
[ECCV 2024] InternVideo2: Scaling Foundation Models for Multimodal Video Understanding
  • Posted on: 03/01/2026 –
  • Comments: 4 Comments
[CVPR 2025] LamRA: Large Multimodal Model as Your Advanced Retrieval Assistant
  • Posted on: 02/22/2026 –
  • Comments: 2 Comments
[Arxiv 2026] Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking
  • Posted on: 02/08/2026 –
  • Comments: 10 Comments
[Arxiv 2026] DeepSeek-OCR 2: Visual Causal Flow
  • Posted on: 02/03/2026 –
  • Comments: 4 Comments
[ICCV 2025] Bidirectional Likelihood Estimation withMulti-Modal Large Language Models for Text-Video Retrieval
  • Posted on: 01/25/2026 –
  • Comments: 4 Comments
[Arxiv 2026] Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models
  • Posted on: 01/18/2026 –
  • Comments: 4 Comments
Newer Posts 1 2 3 … 14 15 Older Posts

Conference Deadline

NEW POST

  • [ICML 2026] Causal-JEPA: Learning World Models through Object-Level Latent Masking
  • [ICML 2026] Zooming without Zooming: Region-to-Image Distillation for Fine-Grained Multimodal Perception
  • [ICLR 2021] AN IMAGE IS WORTH 16X16 WORDS: TRANSFORMERS FOR IMAGE RECOGNITION AT SCALE
  • [ICRA 2025] Dual-BEV Nav: Dual-layer BEV-based Heuristic Path Planning for Robotic Navigation in Unstructured Outdoor Environments
  • [CVPR2025] ProAPO: Progressively Automatic Prompt Optimization for Visual Classification

New Comment

  1. 홍 주영 on [ICML 2026] Are Object-Centric Representations Better At Compositional Generalization?07/27/2026

    Q1. k-means 실험의 목적? >> DINOv2+k-means와 DINOv2+CA는 DINOSAURv2의 장점이 단순히 token을 줄였기 때문인지 확인하는 비교 실험입니다. 아마 리뷰어가 "그냥 토큰…

  2. 이 예은 on [arXiv 2025] UniFGVC: Universal Training-Free Few-Shot Fine-Grained Visual Classification via Attribute-Aware Multimodal Retrieval07/27/2026

    안녕하세요 주영님! 댓글 감사드립니다. 질문주신 것과 관련된 실험들이 있었는데, 리뷰에는 담지 못한 것 같아 댓글로 설명드리고자 합니다. 잘못된(판별력 있는 description…

  3. 이 예은 on [arXiv 2025] UniFGVC: Universal Training-Free Few-Shot Fine-Grained Visual Classification via Attribute-Aware Multimodal Retrieval07/27/2026

    안녕하세요 주연님 댓글 감사합니다! 저도 학습없이 fine-grained class에 대해 discriminative한 캡션을 잘~뽑아서 classification 성능을 올린다는 점에서 이 논문을 재미있게 읽은…

  4. 이 예은 on [arXiv 2025] UniFGVC: Universal Training-Free Few-Shot Fine-Grained Visual Classification via Attribute-Aware Multimodal Retrieval07/27/2026

    안녕하세요 승현님 댓글 감사합니다! Q1) Category-Discriminative Visual Captioner과정에서, 타겟 이미지와 시각적으로 유사한 t개의 샘플을 선택한다고 하셨는데, 클래스당 K개의 이미지로 구성된다면…

  5. 홍 주영 on ICML 2026 참관기07/27/2026

    특정 대화 하나보다는, 저자들이 후속 연구와 현재 방법의 한계를 솔직하게 설명해주셨던 순간들이 기억에 남습니다. 논문에 적힌 결과보다 저자들이 실제로 중요하게…

  • Sign-in
  • RCV-Calendar
  • RCV-Github
  • Paper R/W
    • Arxiv
    • Deadline
    • Overleaf
  • Coding
    • OnlineJudge
    • Kaggle

포기하지 않는 강한 집념 만이 작은 차이를 만든다.

Design by SejongRCV