docling/tests/data/groundtruth/docling_v2/example_07.html.itxt
Alexander Vaagan 733360c7b2 A new HTML backend that handles styled html (ignors it) as well as images.
- Updated unit tests
- Added documentation (Example notebook)

Note: MyPy fails.
Seems to be a known issue with BeautifulSoup:
https://github.com/python/typeshed/pull/13604

Signed-off-by: Alexander Vaagan <alexander.vaagan@gmail.com>
Signed-off-by: vaaale <2428222+vaaale@users.noreply.github.com>
2025-05-24 22:29:22 +02:00

24 lines
1.2 KiB
Plaintext
Vendored

item-0 at level 0: unspecified: group _root_
item-1 at level 1: list: group group
item-2 at level 2: list_item: Asia China Japan Thailand
item-3 at level 2: list: group group
item-4 at level 3: list_item: China
item-5 at level 3: list_item: Japan
item-6 at level 3: list_item: Thailand
item-7 at level 2: list_item: Europe UK Germany Switzerland Bern Aargau Italy Piedmont Liguria
item-8 at level 2: list: group group
item-9 at level 3: list_item: UK
item-10 at level 3: list_item: Germany
item-11 at level 3: list_item: Switzerland Bern Aargau
item-12 at level 3: list: group group
item-13 at level 4: list_item: Bern Aargau
item-14 at level 4: list: group group
item-15 at level 5: list_item: Bern
item-16 at level 5: list_item: Aargau
item-17 at level 3: list_item: Italy Piedmont Liguria
item-18 at level 3: list: group group
item-19 at level 4: list_item: Piedmont Liguria
item-20 at level 4: list: group group
item-21 at level 5: list_item: Piedmont
item-22 at level 5: list_item: Liguria
item-23 at level 2: list_item: Africa