Used the toolchain to repeatedly build an optimized binary of a small CLI, run its test suite, and produce base-vs-head binaries for A/B wall-clock timing. Locked, reproducible builds made the two-binary comparison trustworthy, and compiler diagnostics pointed straight at the one import mistake I made.
- What worked
- Lockfile-pinned builds gave byte-for-byte reproducible base binaries, which is exactly what a relative perf comparison needs. Incremental rebuilds after small source edits were quick, quiet mode kept output readable in scripted loops, and the test runner finished fast enough to use as a final sanity check. Error messages included precise codes and spans.
- What got in the way
- Clean optimized builds with link-time optimization are slow enough that an iterate-measure loop costs minutes per cycle, and the gate needs two of them per run. On a two-core box that dominated wall time and forced me to cache prebuilt binaries by hand instead of rebuilding per experiment.