docling/tests/data/groundtruth/docling_v2/word_sample.docx.md
Peter W. J. Staar f542460af3
fix: fix duplicate title and heading + add e2e tests for html and docx (#186)
* add real e2e tests for html and docx

Signed-off-by: Peter Staar <taa@zurich.ibm.com>

* updated the output of itxt

Signed-off-by: Peter Staar <taa@zurich.ibm.com>

* reformatted the text

Signed-off-by: Peter Staar <taa@zurich.ibm.com>

* fixed the tests

Signed-off-by: Peter Staar <taa@zurich.ibm.com>

* fixed the tests (2)

Signed-off-by: Peter Staar <taa@zurich.ibm.com>

* fixed the examples (1)

Signed-off-by: Peter Staar <taa@zurich.ibm.com>

* fixed the output of the test

Signed-off-by: Peter Staar <taa@zurich.ibm.com>

* updated the tests, moved the ground-truth

Signed-off-by: Peter Staar <taa@zurich.ibm.com>

* moved the ground-truth data

Signed-off-by: Peter Staar <taa@zurich.ibm.com>

* fixed the html tests

Signed-off-by: Peter Staar <taa@zurich.ibm.com>

* restructure title fix (#187)

Signed-off-by: Panos Vagenas <35837085+vagenas@users.noreply.github.com>

---------

Signed-off-by: Peter Staar <taa@zurich.ibm.com>
Signed-off-by: Panos Vagenas <35837085+vagenas@users.noreply.github.com>
Co-authored-by: Panos Vagenas <35837085+vagenas@users.noreply.github.com>
2024-10-30 13:14:56 +01:00

980 B
Raw Blame History

Summer activities

Swimming in the lake

Duck

Figure 1: This is a cute duckling

Lets swim!

To get started with swimming, first lay down in a water and try not to drown:

  • You can relax and look around
  • Paddle about
  • Enjoy summer warmth

Also, dont forget:

  • Wear sunglasses
  • Dont forget to drink water
  • Use sun cream

Hmm, what else…

Lets eat

After we had a good day of swimming in the lake, its important to eat something nice

I like to eat leaves

Here are some interesting things a respectful duck could eat:

Food Calories per portion
Leaves Ash, Elm, Maple 50
Berries Blueberry, Strawberry, Cranberry 150
Grain Corn, Buckwheat, Barley 200

And lets add another list in the end:

  • Leaves
  • Berries
  • Grain