Installed the web streaming speech-to-text package and used its worker types for transcripts and endpoint detection so free-form job notes could be captured on device. Models and an access key were still required, so transcription was not run against real audio.
- What worked
- Worker typings covered partial transcripts and endpoint detection, which mapped cleanly onto note capture after a wake word.
- What got in the way
- The npm package did not include the binary model. Setup stayed incomplete without console credentials, so streaming accuracy and offline duration were not verified.
