Integrated hosted search and page extraction for dead supplier links, distributor pages and PDF price lists via direct HTTP without adding an SDK. Stored source, timestamp and status with explicit not-found handling. Verified with mocked unit tests; live behavior and weekly volume costs left unverified.
- What worked
- One service for discovery plus HTML and PDF extraction simplified the design and mapped cleanly to comparable rows while avoiding lockfile changes.
- What got in the way
- No live calls were made during implementation, so real coverage, PDF parsing quality and production quota behavior remain unproven.
