Dolphin (ByteDance) vs dots.ocr (dots.mocr)

Dolphin (ByteDance) is free, with no paid plan attached. dots.ocr (dots.mocr) is free, with no paid plan attached. Both are listed under AI OCR & Document Extraction, so this is a like-for-like comparison. Neither is ranked above the other — Flocci carries no sponsored placement.

Dolphin (ByteDance) vs dots.ocr (dots.mocr) — straight answers

Dolphin (ByteDance) vs dots.ocr (dots.mocr): what is the difference?

Dolphin (ByteDance) is free, with no paid plan attached and is listed for two-stage document-type and layout analysis then element parsing. dots.ocr (dots.mocr) is free, with no paid plan attached and is listed for 3B vision-language model for layout parsing across many scripts. Both sit in AI OCR & Document Extraction.

Is Dolphin (ByteDance) or dots.ocr (dots.mocr) cheaper to start with?

Neither — Dolphin (ByteDance) and dots.ocr (dots.mocr) are both free, so the choice comes down to capability rather than cost. Both entries list what they uniquely do above.

Which should I choose, Dolphin (ByteDance) or dots.ocr (dots.mocr)?

Choose Dolphin (ByteDance) if you need two-stage document-type and layout analysis then element parsing; choose dots.ocr (dots.mocr) if you need 3B vision-language model for layout parsing across many scripts. Flocci AI Tools does not rank one above the other — it shows both feature sets side by side and lets the requirement decide.

Dolphin (ByteDance) compared with dots.ocr (dots.mocr): pricing tier, category, listed capabilities and links.
 Dolphin (ByteDance)dots.ocr (dots.mocr)
Pricing tierFreeFree
Free to startYesYes
CategoryResearch & LearningResearch & Learning
TypeAI OCR & Document ExtractionAI OCR & Document Extraction
Listed capabilities
  • Two-stage document-type and layout analysis then element parsing
  • Page-level and element-level parsing modes
  • Hugging Face Transformers integration
  • MIT licensed; Dolphin-v2 released Dec 2025
  • 3B vision-language model for layout parsing across many scripts
  • Converts charts and diagrams into SVG code
  • Scores 83.9 on olmOCR-bench
  • MIT licensed with free live demo
Tagsdolphin ocr, bytedance dolphin, dolphin-v2 document parsing, open source document parser, free document image parsingdots.ocr, dots ocr, multilingual ocr model, open source document layout parsing, rednote ocr, xiaohongshu ocr
Websitegithub.comgithub.com
Full pageDolphin (ByteDance) details →dots.ocr (dots.mocr) details →
AlternativesDolphin (ByteDance) alternatives →dots.ocr (dots.mocr) alternatives →

DeepSeek-OCR

AI OCR & Document Extraction
freeNew
  • Vision-language OCR that compresses document context optically
  • Document-to-Markdown, layout detection and figure parsing
  • Multiple resolution modes up to 1280x1280
  • MIT license; DeepSeek-OCR2 released Jan 2026

Docling

AI OCR & Document Extraction
free
  • MIT-licensed, 64.7k GitHub stars, IBM Research-originated, now under LF AI & Data Foundation
  • Advanced PDF layout + table structure recognition, local execution
  • Exports to Markdown/HTML/JSON, integrates with LangChain/LlamaIndex

LangExtract (Google)

AI OCR & Document Extraction
freeNew
  • Python library extracting structured data from unstructured text with LLMs
  • Maps each extraction to its exact source span
  • Works with Gemini, OpenAI and local Ollama models
  • Interactive HTML visualization of results

Marker

AI OCR & Document Extraction
free
  • Apache-2.0 code, 38.7k GitHub stars, converts PDFs/DOCX/PPTX/XLSX to markdown/JSON
  • 76.0% accuracy on the olmocr-bench benchmark
  • Runs on GPU, CPU, or Apple Silicon; optional LLM-boosted accuracy mode

MarkItDown

AI OCR & Document Extraction
freeNew
  • Converts PDF, Office files, images, audio and HTML into LLM-friendly Markdown
  • Ships an MCP server for agents
  • Open-source Python package and CLI
  • Optional LLM-based image description

MinerU

AI OCR & Document Extraction
freeNew
  • Parses PDFs, images and Office files into Markdown, JSON, HTML or LaTeX
  • Runs fully locally with OCR, table and formula recognition
  • Free hosted web app at mineru.net plus agent-ready outputs
  • Version 4.0 adds four parsing quality tiers