document understanding ByteDance/Dolphin Image-Text-to-Text • 0.4B • Updated Jul 16, 2025 • 922 • 519
Multimodal LLM Datasets A collection of the multimodal LLM datasets haonan3/V1-33K-Old Viewer • Updated Mar 22, 2025 • 31.8k • 366 • 4
Multimodal Evaluation MMInstruction/ArxivQA Viewer • Updated Mar 5, 2024 • 100k • 477 • 38 lmms-lab-encoder/DocVQA Viewer • Updated Apr 18, 2024 • 16.6k • 34.7k • 86 vidore/shiftproject_test_captioning Viewer • Updated Jun 20, 2025 • 2.05k • 14 vidore/syntheticDocQA_government_reports_test Viewer • Updated Jun 20, 2025 • 1k • 361 • 1
ml-question-corpus joey234/mmlu-machine_learning-neg-prepend Viewer • Updated Aug 23, 2023 • 117 • 13 • 1 joey234/mmlu-machine_learning-verbal-neg-prepend Viewer • Updated Apr 27, 2023 • 112 • 17 • 1 win-wang/Machine_Learning_QA_Collection Viewer • Updated Sep 25, 2024 • 12.4k • 30 • 8 efeno/colpali_training_machine_learning Viewer • Updated Aug 16, 2024 • 723 • 10
Document Embeddings openbmb/VisRAG-Ret Feature Extraction • 3B • Updated Nov 4, 2024 • 528 • 73 vidore/colpali Visual Document Retrieval • Updated 19 days ago • 3.51k • 487
Document Embedding Datasets & Models bevaya/ScreenSpot Viewer • Updated Apr 10, 2024 • 1.27k • 1.93k • 52 osunlp/Multimodal-Mind2Web Viewer • Updated Jun 5, 2024 • 14.2k • 8.26k • 97 cjfcsjt/AITW_General Viewer • Updated May 4, 2024 • 100k • 969 • 2 microsoft/OmniParser Image-Text-to-Text • Updated Dec 2, 2024 • 468 • 1.71k
document understanding ByteDance/Dolphin Image-Text-to-Text • 0.4B • Updated Jul 16, 2025 • 922 • 519
ml-question-corpus joey234/mmlu-machine_learning-neg-prepend Viewer • Updated Aug 23, 2023 • 117 • 13 • 1 joey234/mmlu-machine_learning-verbal-neg-prepend Viewer • Updated Apr 27, 2023 • 112 • 17 • 1 win-wang/Machine_Learning_QA_Collection Viewer • Updated Sep 25, 2024 • 12.4k • 30 • 8 efeno/colpali_training_machine_learning Viewer • Updated Aug 16, 2024 • 723 • 10
Multimodal LLM Datasets A collection of the multimodal LLM datasets haonan3/V1-33K-Old Viewer • Updated Mar 22, 2025 • 31.8k • 366 • 4
Document Embeddings openbmb/VisRAG-Ret Feature Extraction • 3B • Updated Nov 4, 2024 • 528 • 73 vidore/colpali Visual Document Retrieval • Updated 19 days ago • 3.51k • 487
Multimodal Evaluation MMInstruction/ArxivQA Viewer • Updated Mar 5, 2024 • 100k • 477 • 38 lmms-lab-encoder/DocVQA Viewer • Updated Apr 18, 2024 • 16.6k • 34.7k • 86 vidore/shiftproject_test_captioning Viewer • Updated Jun 20, 2025 • 2.05k • 14 vidore/syntheticDocQA_government_reports_test Viewer • Updated Jun 20, 2025 • 1k • 361 • 1
Document Embedding Datasets & Models bevaya/ScreenSpot Viewer • Updated Apr 10, 2024 • 1.27k • 1.93k • 52 osunlp/Multimodal-Mind2Web Viewer • Updated Jun 5, 2024 • 14.2k • 8.26k • 97 cjfcsjt/AITW_General Viewer • Updated May 4, 2024 • 100k • 969 • 2 microsoft/OmniParser Image-Text-to-Text • Updated Dec 2, 2024 • 468 • 1.71k