Docling
FreeGreat library for ingesting any kind of document for RAG ⭐
About Docling
Docling simplifies document processing by parsing diverse formats — including advanced PDF understanding — and providing seamless integrations with the generative AI ecosystem. It supports a wide range of formats such as PDF, DOCX, PPTX, XLSX, HTML, EPUB, WAV, MP3, WebVTT, images, email formats (EML, MSG), LaTeX, plain text, and more. Advanced PDF capabilities include layout analysis, reading order, table structure, code, formulas, and image classification. Docling offers local execution for sensitive data, extensive OCR, support for visual language models (e.g., GraniteDocling), audio automatic speech recognition, and video parsing. It exports to Markdown, HTML, WebVTT, DocLang, DocTags, and JSON, and integrates with LangChain, LlamaIndex, Crew AI, and Haystack. A CLI and API server (docling-serve) are also available.
Key Features
Pros & Cons
- Open-source and free to use
- Supports an extensive variety of document formats (text, audio, video, images)
- Advanced PDF layout and structure analysis
- Local execution for data privacy and offline scenarios
- Seamless integration with popular AI frameworks (LangChain, LlamaIndex, etc.)
- Active development with regular updates and new feature releases
- Requires Python 3.10 or higher (dropped Python 3.9 support)
- Some features (e.g., metadata extraction, complex chemistry) are still in development