Docling

MIT-licensed, 64.7k GitHub stars, IBM Research-originated, now under LF AI & Data Foundation

What Docling does

  • MIT-licensed, 64.7k GitHub stars, IBM Research-originated, now under LF AI & Data Foundation
  • Advanced PDF layout + table structure recognition, local execution
  • Exports to Markdown/HTML/JSON, integrates with LangChain/LlamaIndex

Docling — straight answers

What is Docling?

Docling is listed under AI OCR & Document Extraction, in the Research & Learning category on Flocci AI Tools. MIT-licensed, 64.7k GitHub stars, IBM Research-originated, now under LF AI & Data Foundation. It is free, with no paid plan attached, and it lives at github.com.

Is Docling free?

Docling is listed as fully free — there is no paid tier attached to it in the catalog. That makes it one of the 115 entries on Flocci AI Tools with no upgrade path built in.

What can Docling do?

Docling does 3 things the catalog singles out: Mit-licensed, 64.7k github stars, ibm research-originated, now under lf ai & data foundation; advanced pdf layout + table structure recognition, local execution; exports to markdown/html/json, integrates with langchain/llamaindex.

What is the best free alternative to Docling?

Marker is the closest free alternative: it sits in the same AI OCR & Document Extraction sub-category and is free. olmOCR, AskYourPDF and Humata also start free. The full list is on the alternatives page.

See the full list →

Docling alternatives

Compare all alternatives →

AskYourPDF

AI OCR & Document Extraction
freemium
  • Chat with any PDF plus auto-fill PDF forms
  • Embeddable document chatbot for websites
  • 5M+ users including MIT/Harvard/Oxford/Stanford

PDF.ai

AI OCR & Document Extraction
freemium
  • Chat-with-PDF product plus a separate parse/extract/split PDF API
  • Popular ChatPDF-category alternative with API access for developers

Humata

AI OCR & Document Extraction
freemium
  • Free tier: 60 pages, 10 answers, no card
  • Cited answers across many uploaded files at once
  • Paid tiers priced per extra page rather than per seat

Mistral OCR 4

AI OCR & Document Extraction
freemium
  • Self-hostable single-container OCR model for regulated industries
  • 170 languages, bounding boxes and per-field confidence scores
  • Low per-thousand-page API pricing; free to try inside Le Chat

LlamaParse

AI OCR & Document Extraction
freemium
  • Parses 90+ formats incl. complex tables/charts/handwriting into clean markdown
  • 300,000+ users, 1B+ documents processed
  • Free trial credits via LlamaCloud, then usage-based

Reducto

AI OCR & Document Extraction
freemium
  • 15,000 free starter credits, no card required
  • Citation-ready structured JSON across 30+ file types
  • Edits extracted data back into original PDF/DOCX layout

Head-to-head comparisons