mirror of
https://github.com/DS4SD/docling.git
synced 2025-12-08 12:48:28 +00:00
* docs: Fix broken homepage links
Signed-off-by: Robyn J <robynjohnson@us.ibm.com>
* docs: Remediate sign-off
DCO Remediation Commit for Robyn J <bobbinrobyn@users.noreply.github.com>
I, Robyn J <bobbinrobyn@users.noreply.github.com>, hereby add my Signed-off-by to this commit: e873e24c11
Signed-off-by: Robyn J <bobbinrobyn@users.noreply.github.com>
---------
Signed-off-by: Robyn J <robynjohnson@us.ibm.com>
Signed-off-by: Robyn J <bobbinrobyn@users.noreply.github.com>
6.0 KiB
Vendored
6.0 KiB
Vendored
Docling simplifies document processing, parsing diverse formats — including advanced PDF understanding — and providing seamless integrations with the gen AI ecosystem.
Getting started
🐣 Ready to kick off your Docling journey? Let's dive right into it!
⬇️ Installation
Quickly install Docling in your environment ▶️ Quickstart
Get a jumpstart on basic Docling usage 🧩 Concepts
Learn Docling fundamentals and get a glimpse under the hood 🧑🏽🍳 Examples
Try out recipes for various use cases, including conversion, RAG, and more 🤖 Integrations
Check out integrations with popular AI tools and frameworks 📖 Reference
See more API details
Quickly install Docling in your environment ▶️ Quickstart
Get a jumpstart on basic Docling usage 🧩 Concepts
Learn Docling fundamentals and get a glimpse under the hood 🧑🏽🍳 Examples
Try out recipes for various use cases, including conversion, RAG, and more 🤖 Integrations
Check out integrations with popular AI tools and frameworks 📖 Reference
See more API details
Features
- 🗂️ Parsing of multiple document formats incl. PDF, DOCX, PPTX, XLSX, HTML, WAV, MP3, VTT, images (PNG, TIFF, JPEG, ...), and more
- 📑 Advanced PDF understanding incl. page layout, reading order, table structure, code, formulas, image classification, and more
- 🧬 Unified, expressive DoclingDocument representation format
- ↪️ Various export formats and options, including Markdown, HTML, DocTags and lossless JSON
- 🔒 Local execution capabilities for sensitive data and air-gapped environments
- 🤖 Plug-and-play integrations incl. LangChain, LlamaIndex, Crew AI & Haystack for agentic AI
- 🔍 Extensive OCR support for scanned PDFs and images
- 👓 Support of several Visual Language Models (GraniteDocling)
- 🎙️ Support for Audio with Automatic Speech Recognition (ASR) models
- 🔌 Connect to any agent using the Docling MCP server
- 💻 Simple and convenient CLI
What's new
- 📤 Structured [information extraction][extraction] [🧪 beta]
- 📑 New layout model (Heron) by default, for faster PDF parsing
- 🔌 MCP server for agentic applications
- 💬 Parsing of Web Video Text Tracks (WebVTT) files
Coming soon
- 📝 Metadata extraction, including title, authors, references & language
- 📝 Chart understanding (Barchart, Piechart, LinePlot, etc)
- 📝 Complex chemistry understanding (Molecular structures)
What's next
🚀 The journey has just begun! Join us and become a part of the growing Docling community.
- :fontawesome-brands-github: GitHub
- :fontawesome-brands-discord: Discord
- :fontawesome-brands-linkedin: LinkedIn
Live assistant
Do you want to leverage the power of AI and get live support on Docling? Try out the Chat with Dosu functionalities provided by our friends at Dosu.
LF AI & Data
Docling is hosted as a project in the LF AI & Data Foundation.
IBM ❤️ Open Source AI
The project was started by the AI for knowledge team at IBM Research Zurich.
