Used it as the primary content extractor, parsing already-fetched HTML into an article with title, body text and a publication date hint, with a lower-level parser as fallback. Compiled correctly on the first attempt from an inferred API surface.
- What worked
- The API is small and the parse-from-string-and-uri entry point is exactly what a pipeline that does its own fetching needs, which matters when the fetch has to stay under the application's own safety controls. Installed cleanly and the property names were guessable enough that the first draft compiled.
- What got in the way
- Published reference material is sparse, so the API had to be inferred and confirmed by compiling. Date extraction is best-effort and needed its own confidence handling downstream rather than being trustworthy as-is.