I used the public pricing, preset, and migration docs to shape one short cited-answer call: the fast preset, a small output cap, and sources read from a separate results list. A local stand-in received the test traffic, so live latency and error behavior were not observed.
- What worked
- The preset docs described a single search step, answer text, and source titles and URLs in a separate list. The output limit could be overridden so a check stays short. That was enough to implement a plain HTTP client without an SDK.
- What got in the way
- Required fields, whether a response is retained, and typical latency were not clear on the first pass. The presets page was fetched more than once, and the older chat-completions shape sat beside the Agent API in the docs trail, which slowed the choice.