본문 바로가기 메뉴 바로가기

ainote

프로필사진
  • 글쓰기
  • 관리
  • 태그
  • 방명록
  • RSS

ainote

검색하기 폼
  • 분류 전체보기 (64)
    • 파이썬 머신러닝 완벽 가이드 (38)
    • 알고리즘 (3)
    • 논문 리뷰 (10)
      • 멀티모달 (7)
    • 선형대수학 (2)
    • 밑바닥부터 시작하는 딥러닝1 (9)
    • 알고리즘 문제풀이 (2)
    • 컴퓨터 공학 전공 (0)
      • 신호처리 (0)
      • 컴퓨터 구조 (0)
      • 데이터베이스 (0)
      • 운영체제 (0)
  • 방명록

전체 글 (64)
r

보호되어 있는 글입니다.

보호글 2026. 9. 26. 12:19
r

보호되어 있는 글입니다.

보호글 2025. 6. 4. 18:14
[ICML 2022] VLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts

Introduction Two mainstream architectures are widely used in previous work. Dual-stream : encode images and text separately. Modality interaction is handled by the cosine similarity of the image and text feature vectors. This architecture is effective for retrieval tasks, especially for masses of images and text Representative model : CLIP, ALIGN Limitation : Its shallow interaction is not enoug..

논문 리뷰/멀티모달 2024. 4. 6. 18:57
이전 1 2 3 4 ··· 22 다음
이전 다음
«   2026/09   »
일 월 화 수 목 금 토
1 2 3 4 5
6 7 8 9 10 11 12
13 14 15 16 17 18 19
20 21 22 23 24 25 26
27 28 29 30

Blog is powered by Tistory / Designed by Tistory

티스토리툴바