V
VLM_survey
jingyi0000/VLM_survey
Collection of AWESOME vision-language models for vision tasks
★3.1kstars
Updated: Today
📋 Project at a Glance
Tap to expand
What's this?A clip/computer-vision tool in the Learning & Resources category, open-source
Who made it?Maintained by jingyi0000 team, 3.1K⭐ on GitHub, #684 out of 3640 in Learning & Resources
Why does it exist?In the Learning & Resources space, clip workflows faced efficiency bottlenecks. VLM_survey was built by jingyi0000 to address these computer-vision challenges.
What can it do?Key use cases: deep-learning, knowledge-distillation, multi-modal-model
How to install with AI?Use an AI coding assistant to follow the README and automatically handle the install and environment setup.
🔗 github.com/jingyi0000/VLM_survey
🔗 github.com/jingyi0000/VLM_survey
Topics
clipcomputer-visiondeep-learningknowledge-distillationmulti-modal-modelsurveytransfer-learningvision-language-model