Document to Markdown This collection contains models which convert text or multimodal documents to markdown format for various downstream tasks. dots-studio/dots.ocr Image-Text-to-Text • 3B • Updated Oct 31, 2025 • 257k • 1.33k numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 357k • 494 zai-org/GLM-4.5V Image-Text-to-Text • 108B • Updated Oct 25, 2025 • 46.5k • • 721 microsoft/kosmos-2.5 Image-Text-to-Text • 1B • Updated Aug 28, 2025 • 17.8k • 271
Document Datasets docling-project/DocLayNet Updated Jan 25, 2023 • 663 • 147 common-pile/caselaw_access_project Viewer • Updated Jun 6, 2025 • 5.52M • 2.99k • 219 llamaindex/vdr-multilingual-test Viewer • Updated Jan 10, 2025 • 15k • 94 • 3 PleIAs/common_corpus Viewer • Updated May 6 • 69.9k • 117k • 417
Document to Markdown This collection contains models which convert text or multimodal documents to markdown format for various downstream tasks. dots-studio/dots.ocr Image-Text-to-Text • 3B • Updated Oct 31, 2025 • 257k • 1.33k numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 357k • 494 zai-org/GLM-4.5V Image-Text-to-Text • 108B • Updated Oct 25, 2025 • 46.5k • • 721 microsoft/kosmos-2.5 Image-Text-to-Text • 1B • Updated Aug 28, 2025 • 17.8k • 271
Document Datasets docling-project/DocLayNet Updated Jan 25, 2023 • 663 • 147 common-pile/caselaw_access_project Viewer • Updated Jun 6, 2025 • 5.52M • 2.99k • 219 llamaindex/vdr-multilingual-test Viewer • Updated Jan 10, 2025 • 15k • 94 • 3 PleIAs/common_corpus Viewer • Updated May 6 • 69.9k • 117k • 417