M
Multi-Modality-Arena
OpenGVLab/Multi-Modality-Arena
Chatbot Arena meets multi-modality! Multi-Modality Arena allows you to benchmark vision-language models side-by-side while providing images as inputs. Supports MiniGPT-4, LLaMA-Adapter V2, LLaVA, BLIP-2, and many more!
★565stars
Python
Updated: 1w ago
📋 Project at a Glance
Tap to expand
What's this?A , built with Python open-source project in the Models category, core strengths: chat/chatbot
Who made it?Maintained by OpenGVLab team, 565⭐ on GitHub, #1687 out of 3201 in Models
Why does it exist?In the Models space, chat workflows faced efficiency bottlenecks. Multi-Modality-Arena was built by OpenGVLab to address these chatbot challenges.
What can it do?Key use cases: chatgpt, gradio, large-language-models
How to install with AI?Use an AI coding assistant (Claude Code, Cursor, Copilot) to automatically set up pip dependencies and virtual env. Follow the README — the AI handles the rest.
🔗 github.com/OpenGVLab/Multi-Modality-Arena
🔗 github.com/OpenGVLab/Multi-Modality-Arena
Topics
chatchatbotchatgptgradiolarge-language-modelsllmsmulti-modalityvision-language-modelvqa