UCF CRCV
Publishes 5 feeds
CAP5415 Computer Vision - Fall 2023
15 posts · theirs
CAP6412 Advanced Computer Vision - Spring 2024
15 posts · theirs
CAP6412 Advanced Computer Vision - Spring 2018
15 posts · theirs
CAP5415 Fall 2014
14 posts · theirs
CAP5415 Computer Vision - Fall 2020
10 posts · theirs
Lately
Lecture 15 - Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Lecture 14 - Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic
Lecture 4 - Visual-Language Models Introduction Part-I: CoCA, PALI
Lecture 3 - CLIP
Lecture 2 - Transformers Introduction
Lecture 1 - Introduction
Lecture 5 - Visual-Language Models Introduction Part-II: FLAMINGO, FLAVA, PAINTER, BLIP-2
Lecture 6 - Visual-Language Models Introduction Part-III: Image-Bind, Language-Bind, LLaVA
Lecture 7 - Visual-Language Models Introduction Part-IV: Video ChatGPT, PG-Video LLaVA
Lecture 13 - MERLOT RESERVE: Neural Script Knowledge through Vision and Language and Sound
Lecture 12 - MaMMUT: A Simple Architecture for Joint Learning for MultiModal Tasks
Lecture 3 - Image Filtering
Everything on this page was read from markup UCF CRCV published — a rel="me" link, an h-card, or the feed’s own author element. Nothing was inferred from anywhere else. To correct or remove it, get in touch. Machine-readable: JSON
