document understanding ByteDance/Dolphin Image-Text-to-Text • 0.4B • Updated Jul 16, 2025 • 856 • 519
Multimodal LLM Datasets A collection of the multimodal LLM datasets haonan3/V1-33K-Old Viewer • Updated Mar 22, 2025 • 31.8k • 539 • 4
Multimodal Evaluation MMInstruction/ArxivQA Viewer • Updated Mar 5, 2024 • 100k • 483 • 38 lmms-lab-encoder/DocVQA Viewer • Updated Apr 18, 2024 • 16.6k • 29.1k • 86 vidore/shiftproject_test_captioning Viewer • Updated Jun 20, 2025 • 2.05k • 27 vidore/syntheticDocQA_government_reports_test Viewer • Updated Jun 20, 2025 • 1k • 357 • 1
ml-question-corpus joey234/mmlu-machine_learning-neg-prepend Viewer • Updated Aug 23, 2023 • 117 • 15 • 1 joey234/mmlu-machine_learning-verbal-neg-prepend Viewer • Updated Apr 27, 2023 • 112 • 11 • 1 win-wang/Machine_Learning_QA_Collection Viewer • Updated Sep 25, 2024 • 12.4k • 39 • 8 efeno/colpali_training_machine_learning Viewer • Updated Aug 16, 2024 • 723 • 19
Document Embeddings openbmb/VisRAG-Ret Feature Extraction • 3B • Updated Nov 4, 2024 • 683 • 73 vidore/colpali Visual Document Retrieval • Updated 1 day ago • 8.31k • 486
Document Embedding Datasets & Models bevaya/ScreenSpot Viewer • Updated Apr 10, 2024 • 1.27k • 2.09k • 51 osunlp/Multimodal-Mind2Web Viewer • Updated Jun 5, 2024 • 14.2k • 5.42k • 97 cjfcsjt/AITW_General Viewer • Updated May 4, 2024 • 100k • 795 • 2 microsoft/OmniParser Image-Text-to-Text • Updated Dec 2, 2024 • 493 • 1.71k
document understanding ByteDance/Dolphin Image-Text-to-Text • 0.4B • Updated Jul 16, 2025 • 856 • 519
ml-question-corpus joey234/mmlu-machine_learning-neg-prepend Viewer • Updated Aug 23, 2023 • 117 • 15 • 1 joey234/mmlu-machine_learning-verbal-neg-prepend Viewer • Updated Apr 27, 2023 • 112 • 11 • 1 win-wang/Machine_Learning_QA_Collection Viewer • Updated Sep 25, 2024 • 12.4k • 39 • 8 efeno/colpali_training_machine_learning Viewer • Updated Aug 16, 2024 • 723 • 19
Multimodal LLM Datasets A collection of the multimodal LLM datasets haonan3/V1-33K-Old Viewer • Updated Mar 22, 2025 • 31.8k • 539 • 4
Document Embeddings openbmb/VisRAG-Ret Feature Extraction • 3B • Updated Nov 4, 2024 • 683 • 73 vidore/colpali Visual Document Retrieval • Updated 1 day ago • 8.31k • 486
Multimodal Evaluation MMInstruction/ArxivQA Viewer • Updated Mar 5, 2024 • 100k • 483 • 38 lmms-lab-encoder/DocVQA Viewer • Updated Apr 18, 2024 • 16.6k • 29.1k • 86 vidore/shiftproject_test_captioning Viewer • Updated Jun 20, 2025 • 2.05k • 27 vidore/syntheticDocQA_government_reports_test Viewer • Updated Jun 20, 2025 • 1k • 357 • 1
Document Embedding Datasets & Models bevaya/ScreenSpot Viewer • Updated Apr 10, 2024 • 1.27k • 2.09k • 51 osunlp/Multimodal-Mind2Web Viewer • Updated Jun 5, 2024 • 14.2k • 5.42k • 97 cjfcsjt/AITW_General Viewer • Updated May 4, 2024 • 100k • 795 • 2 microsoft/OmniParser Image-Text-to-Text • Updated Dec 2, 2024 • 493 • 1.71k