← Flash Papers
Visual Instruction Tuning: Demystifying LLaVA
Share
Share this Flash Paper
Visual Instruction Tuning: Demystifying LLaVA
https://flashpapers.ai/p/5bvqdwjqc6
Copy link
Post to X
Post to Bluesky
Share on LinkedIn
Submit to Hacker News
Submit to Reddit
Email a link
Source PDF ↗
Open standalone ↗
Make your own
Make your own paper
Haotian Liu, Chunyuan Li +2 more
cs.CV
2023
arXiv:2304.08485
Code
· 25k★
· updated 2y ago
Details ▸
Details ▴
ScienceQA Accuracy:
92.53%
· ScienceQA
Relative Score vs GPT-4:
85.1%
· LLaVA-Bench (COCO)
Relative Score vs GPT-4 (Complex Reasoning):
96.5%
· LLaVA-Bench (COCO)
ScienceQA Accuracy:
92.53%
· ScienceQA
Relative Score vs GPT-4:
85.1%
· LLaVA-Bench (COCO)
Relative Score vs GPT-4 (Complex Reasoning):
96.5%
· LLaVA-Bench (COCO)