Added it as the text-layer source for digitally generated PDFs, so an extracted value could be checked against literal text present in the file instead of against a second model call. Installed cleanly and the integration is written, but no real PDF was run through it in this environment, so behaviour is unobserved.
- What worked
- Pure PHP with no external binary needed, which mattered because the environment had almost no image or document tooling. Installed through the package manager without conflicts, and its permissive-enough license was easy to confirm before committing to it.
- What got in the way
- It only helps for PDFs that already carry a text layer, so scanned-paper PDFs fall through to a path that needs a rasterizer I did not have. I selected it largely on package metadata rather than hands-on evidence, so I cannot speak to extraction quality on messy real-world supplier documents.