docling/tests/data/groundtruth/docling_v2/example_05.html.md
Cesar Berrospi Ramis a112d7a035
fix: parse html with omitted body tag (#818)
* fix: parse HTML files without body tag

Parse HTML files without 'body' tag, since it is optional in HTML5 specification.

Signed-off-by: Cesar Berrospi Ramis <75900930+ceberam@users.noreply.github.com>

* test: ensure docling converts HTML without body tag

Signed-off-by: Cesar Berrospi Ramis <75900930+ceberam@users.noreply.github.com>

---------

Signed-off-by: Cesar Berrospi Ramis <75900930+ceberam@users.noreply.github.com>
2025-01-27 16:59:00 +01:00

474 B

Omitted html and body tags

Header 1 Header 2 & 3 (colspan) Header 2 & 3 (colspan)
Row 1 & 2, Col 1 (rowspan) Row 1, Col 2 Row 1, Col 3
Row 1 & 2, Col 1 (rowspan) Row 2, Col 2 & 3 (colspan) Row 2, Col 2 & 3 (colspan)
Row 3, Col 1 Row 3, Col 2 Row 3, Col 3