← 开源
SkalskiP

vlms-zero-to-hero

This series will take you on a journey from the fundamentals of NLP and Computer Vision to the cutting edge of Vision-Language Models.

TutorialsML/AI fundamentalsJupyter Notebook
在 GitHub 打开
增长势头
+124 小时新增 Star+0.1%
1.18k
Star
105
Fork
+0
本周
1
贡献者
创建于 2024-12-20 · 更新于 2026-10-07 · 今日第 5706 名
主要开发者
README

VLMs zero-to-hero

coming: january 2025...

hello

Welcome to VLMs Zero to Hero! This series will take you on a journey from the fundamentals of NLP and Computer Vision to the cutting edge of Vision-Language Models.

tutorials

notebook open in colab video paper
01.01. Word2Veq: Distributed Representations of Words and Phrases and their Compositionality link soon link

roadmap

natural language processing (NLP) fundamentals

computer vision (CV) fundamentals

early vision-language models

scale and efficiency

modern vision-language models

extra

contribute and suggest more papers

Are there important papers, models, or techniques we missed? Do you have a favorite breakthrough in vision-language research that isn't listed here? We’d love to hear your suggestions!