Document to Markdown This collection contains models which convert text or multimodal documents to markdown format for various downstream tasks. rednote-hilab/dots.ocr Image-Text-to-Text • Updated Oct 31, 2025 • 254k • 1.24k numind/NuMarkdown-8B-Thinking Image-to-Text • Updated Nov 13, 2025 • 703k • 441 zai-org/GLM-4.5V Image-Text-to-Text • 108B • Updated Oct 25, 2025 • 39.1k • • 707 microsoft/kosmos-2.5 Image-Text-to-Text • Updated Aug 28, 2025 • 78.1k • 270
Document Datasets docling-project/DocLayNet Updated Jan 25, 2023 • 741 • 126 common-pile/caselaw_access_project Viewer • Updated Jun 6, 2025 • 5.52M • 762 • 206 llamaindex/vdr-multilingual-test Viewer • Updated Jan 10, 2025 • 15k • 233 • 3 PleIAs/common_corpus Viewer • Updated 5 days ago • 517M • 55.4k • 351
Document to Markdown This collection contains models which convert text or multimodal documents to markdown format for various downstream tasks. rednote-hilab/dots.ocr Image-Text-to-Text • Updated Oct 31, 2025 • 254k • 1.24k numind/NuMarkdown-8B-Thinking Image-to-Text • Updated Nov 13, 2025 • 703k • 441 zai-org/GLM-4.5V Image-Text-to-Text • 108B • Updated Oct 25, 2025 • 39.1k • • 707 microsoft/kosmos-2.5 Image-Text-to-Text • Updated Aug 28, 2025 • 78.1k • 270
Document Datasets docling-project/DocLayNet Updated Jan 25, 2023 • 741 • 126 common-pile/caselaw_access_project Viewer • Updated Jun 6, 2025 • 5.52M • 762 • 206 llamaindex/vdr-multilingual-test Viewer • Updated Jan 10, 2025 • 15k • 233 • 3 PleIAs/common_corpus Viewer • Updated 5 days ago • 517M • 55.4k • 351