# What is Docling ? Docling simplifies document processing, parsing diverse formats โ€” including advanced PDF understanding โ€” and providing seamless integrations with the gen AI ecosystem. # Features - ๐Ÿ—‚๏ธ Parsing of multiple document formats incl. PDF, DOCX, PPTX, XLSX, HTML, WAV, MP3, WebVTT, images (PNG, TIFF, JPEG, ...), LaTeX, plain text, and more - ๐Ÿ“‘ Advanced PDF understanding incl. page layout, reading order, table structure, code, formulas, image classification, and more - ๐Ÿงฌ Unified, expressive DoclingDocument representation format - โ†ช๏ธ Various export formats and options, including Markdown, HTML, WebVTT, DocTags and lossless JSON - ๐Ÿ“œ Support of several application-specifc XML schemas incl. USPTO patents, JATS articles, and XBRL financial reports. - ๐Ÿ”’ Local execution capabilities for sensitive data and air-gapped environments - ๐Ÿค– Plug-and-play integrations incl. LangChain, LlamaIndex, Crew AI & Haystack for agentic AI - ๐Ÿ” Extensive OCR support for scanned PDFs and images - ๐Ÿ‘“ Support of several Visual Language Models (GraniteDocling) - ๐ŸŽ™๏ธ Audio support with Automatic Speech Recognition (ASR) models - ๐Ÿ”Œ Connect to any agent using the MCP server - ๐Ÿ’ป Simple and convenient CLI # What's new - ๐Ÿ“ค Structured information extraction [๐Ÿงช beta] - ๐Ÿ“‘ New layout model (Heron) by default, for faster PDF parsing - ๐Ÿ”Œ MCP server for agentic applications - ๐Ÿ’ผ Parsing of XBRL (eXtensible Business Reporting Language) documents for financial reports - ๐Ÿ’ฌ Parsing of WebVTT (Web Video Text Tracks) files and export to WebVTT format - ๐Ÿ’ฌ Parsing of LaTeX files - ๐Ÿ“ Parsing of plain-text files (.txt, .text) and Markdown supersets (.qmd, .Rmd) - ๐Ÿ“ Chart understanding (Barchart, Piechart, LinePlot): converting them into tables, code or adding detailed descriptions Coming soon - ๐Ÿ“ Metadata extraction, including title, authors, references & language - ๐Ÿ“ Complex chemistry understanding (Molecular structures)