How transcription works, and how fast it is
Parakeet TDT turns speech into private, offline-ready text after targeted pre-transcription enhancement.
Your words become text on your device with NVIDIA Parakeet TDT speech recognition, so the recording does not need to be uploaded to an Obsidian Ridge Labs server for transcription.
Before Parakeet receives the audio, Echo Chamber applies a targeted, speech-focused filter designed for transcription. This is not generic normalization. The complete enhanced Echo Chamber pipeline has produced an internal observed word error rate of approximately 4.5% under tested conditions. Results vary with speakers, accents, acoustics, crosstalk, vocabulary, and source quality.
Speed depends on your device and the length of the recording, but short notes are usually ready the moment you stop, and longer recordings finish processing quickly in the background.

The speech model downloads a single time over the internet. After that, recording and transcription keep working fully offline.
Related guides
Email support@obsidianridgelabs.com, a real person will reply.