Vision Understanding
Models for image analysis, visual question answering, and scene understanding
Last Updated
Nov 9, 2025
Total Votes
85,580
Total Models
10
Single Column
| Rank (UB) | Model | Score | Win% [?] | Votes | Organization | License | |||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 🥇 | GPT-4 Vision | 94 | 90 | 100,535 | OpenAI | Proprietary | 🥇 GPT-4 VisionOpenAI Score94 Win%90 Votes100,535 | ||||||||
| 🥈 | Claude 3 Opus | 93 | 78 | 100,236 | Anthropic | Proprietary | 🥈 Claude 3 OpusAnthropic Score93 Win%78 Votes100,236 | ||||||||
| 🥉 | Gemini Ultra Vision | 93 | 79 | 95,571 | Proprietary | 🥉 Gemini Ultra VisionScore93 Win%79 Votes95,571 | |||||||||
| 4 | LLaVA 1.6 | 87 | 70 | 88,152 | Open Source | Proprietary | 4 LLaVA 1.6Open Source Score87 Win%70 Votes88,152 | ||||||||
| 5 | Qwen-VL-Plus | 86 | 80 | 93,270 | Alibaba | Proprietary | 5 Qwen-VL-PlusAlibaba Score86 Win%80 Votes93,270 | ||||||||
| 6 | CogVLM | 84 | 80 | 92,807 | Tsinghua | Proprietary | 6 CogVLMTsinghua Score84 Win%80 Votes92,807 | ||||||||
| 7 | BLIP-2 | 82 | 84 | 86,824 | Salesforce | Proprietary | 7 BLIP-2Salesforce Score82 Win%84 Votes86,824 | ||||||||
| 8 | InstructBLIP | 80 | 81 | 89,857 | Salesforce | Proprietary | 8 InstructBLIPSalesforce Score80 Win%81 Votes89,857 | ||||||||
| 9 | MiniGPT-4 | 79 | 63 | 82,603 | Open Source | Proprietary | 9 MiniGPT-4Open Source Score79 Win%63 Votes82,603 | ||||||||
| 10 | Flamingo | 77 | 72 | 84,334 | DeepMind | Proprietary | 10 FlamingoDeepMind Score77 Win%72 Votes84,334 | ||||||||