docling/tests/data
Panos Vagenas 7c5614a37a
fix(markdown): fix single-formatted headings & list items (#1820)
* fix(markdown): fix formatting & inline edge cases (show behavior before change)

Signed-off-by: Panos Vagenas <pva@zurich.ibm.com>

* add change and updated test data

Signed-off-by: Panos Vagenas <pva@zurich.ibm.com>

* update lock

Signed-off-by: Panos Vagenas <pva@zurich.ibm.com>

* improve test case

Signed-off-by: Panos Vagenas <pva@zurich.ibm.com>

---------

Signed-off-by: Panos Vagenas <pva@zurich.ibm.com>
2025-06-25 13:05:06 +02:00
..
asciidoc fix(asciidoc): set default size when missing in image directive (#1769) 2025-06-16 10:38:46 +02:00
audio feat: Support audio input (#1763) 2025-06-23 14:47:26 +02:00
csv feat: Add support for CSV input with new backend to transform CSV files to DoclingDocument (#945) 2025-02-14 08:55:09 +01:00
docx fix(msword_backend): Identify text in the same line after an image #1425 (#1610) 2025-06-20 10:55:30 +02:00
groundtruth fix(markdown): fix single-formatted headings & list items (#1820) 2025-06-25 13:05:06 +02:00
html test: add missing ground truth files (#1667) 2025-05-28 13:26:49 +02:00
jats feat(xml-jats): parse XML JATS documents (#967) 2025-02-17 10:43:31 +01:00
md fix(markdown): fix single-formatted headings & list items (#1820) 2025-06-25 13:05:06 +02:00
pdf fix(pypdfium): resolve overlapping text when merging bounding boxes (#1549) 2025-05-19 15:26:00 +02:00
pptx fix: pptx line break and space handling (#1664) 2025-06-16 10:44:30 +02:00
uspto feat: create a backend to parse USPTO patents into DoclingDocument (#606) 2024-12-17 16:35:23 +01:00
webp fix(markdown): fix single-formatted headings & list items (#1820) 2025-06-25 13:05:06 +02:00
xlsx feat: support xlsm files (#1520) 2025-06-10 16:55:59 +02:00
2305.03393v1-pg9-img.png feat!: Docling v2 (#117) 2024-10-16 21:02:03 +02:00