Offline Speech to Text (OSTT)

app.offlinespeechtotext
by Offline Speech to Text

Private, fully offline speech to text for Android

Offline Speech to Text provides private speech recognition that runs entirely on your Android device. It can serve compatible keyboards and apps, provide live subtitles, transcribe audio files, and act as an optional voice keyboard. Audio is processed on-device and is never uploaded.

First release: Sep 11, 2026, 17 total releases.

Most recent release: Sep 24, 2026.

Website Repo

Appears in 1 app stack.

0 sats / 0 zaps received in the past year.

Sats Received

Underlying data available via MCP: app_zaps, app_releases.

Zap Count

Underlying data available via MCP: app_zaps, app_releases.

Releases

  • Sep 24, 2026 0.11.0
    English dictation now removes a standalone “uh” after recognition while preserving nearby words, sentence capitalization, and quoted speech. Other languages and filler-only results stay unchanged. For multilingual Parakeet, choose English under the model language setting to enable this English-only cleanup; Auto mode cannot reliably identify the language. Models, settings, and transcripts remain on your device.
  • Sep 24, 2026 0.10.1
    Voice input now waits through Android's device-wide microphone unblock prompt instead of treating the blocked interval as a pause in speech. Fixed automatic stopping in the voice keyboard, improved the first-install message when no speech model is present, and made preferred-word spelling cleanup available across supported model families. Ordinary words are not automatically replaced with similar-sounding names.
  • Sep 23, 2026 0.10.0
    Preferred phrases now work with Parakeet and other non-Whisper models: exact matches use your chosen spelling and capitalization, including names such as White Noise. The preferred-word list also handles close versions of Nostr and a few common spoken forms of npub. This stays private and on-device. Larger or ambiguous recognition errors still need an optional exact correction; OSTT does not rewrite text based on broad sound-alike guesses.
  • Sep 23, 2026 0.9.0
    Add a simpler personal dictionary: enter preferred words once, without listing every misrecognition. Whisper uses bounded recognition hints; Parakeet and other models conservatively correct very close spellings of single words after transcription. Exact correction pairs remain available for phrases and larger errors. Make the dictionary screen denser and clarify what each model can do. Use warm, high-contrast confirmation colors in light and dark themes.
  • Sep 22, 2026 0.8.0
    Keep long dictation smooth across natural pauses. OSTT now preserves each ordinary recording as one continuous utterance and lets the speech runtime maintain its own long-audio context, preventing short hesitations from creating false sentence boundaries. Very long input to other offline model families uses overlapping, reconciled windows. Diagnostic logging remains off by default and content-free. When enabled, Settings can save the bounded log as a text file, share a temporary read-only snapshot through Android's Sharesheet, or clear it immediately.
  • Sep 22, 2026 0.7.0
    Faster long dictation with model-aware background processing while you speak. Long recordings now finish sooner without dropping audio: OSTT keeps the original recording until the final result and automatically retries the complete message if background processing cannot finish. Rapid app or editor changes also release an old recording immediately, avoiding false microphone errors when the next transcription starts.
  • Sep 21, 2026 0.6.3
    Keep completed voice-keyboard transcriptions in the active composer when Compose or another reactive UI recreates its input connection. OSTT now retries brief connection replacement and inserts into Android's currently focused safe composer instead of opening its recovery screen. If no writable composer exists, the exact transcript remains privately recoverable without interrupting the current app.
  • Sep 21, 2026 0.6.2
    Fix automatic voice-keyboard recording when switching between apps or text fields. An editor transition before the first audio buffer no longer produces a false microphone-permission error. When Android keeps the voice keyboard visible for a new editor, automatic recording starts there correctly, while any real captured result still uses private recovery if its original field is gone.
  • Sep 21, 2026 0.6.1
    Fix a startup crash in compact installs that do not yet have a speech model. The home screen now waits to initialize transcription until a downloaded, imported, or retained model is actually available, while the native loader handles a missing packaged model as a normal setup state instead of allowing an Android file exception to escape.
  • Sep 21, 2026 0.6.0
    Upgrade the offline transcription engine to transcribe.cpp 0.2.3 for current decoder fixes and more efficient resource handling. German-primary phones now recommend the German-tuned Parakeet primeLine model, while the multilingual Parakeet model remains the balanced default elsewhere. Manual model imports are more resilient: copying, verification, and atomic installation now continue safely across screen rotation, with clearer progress and failure cleanup.
  • Sep 20, 2026 0.5.0
    Never lose a finished voice-keyboard transcription when Android refreshes or closes the original text field. Same-editor connection refreshes now insert normally; otherwise OSTT saves one private recovery result with copy, keep, and delete controls. The APK is now small: first setup recommends and downloads a verified model once, and ordinary app updates keep it. Models are grouped and explained by family, with the largest comfortable option for this phone highlighted. Download guidance uses phone resources and system language, with an App Info shortcut for optional Network removal afterward.
  • Sep 20, 2026 0.4.1
    Keep final words by giving offline decoders a short inference-only silence tail after recording stops. Use simple original/replacement fields for exact personal-dictionary corrections, with no rule syntax to learn. Render every voice-keyboard icon in white for a consistent high-contrast layout. Keep optional diagnostic logging responsive with bounded append-and-compact storage.
  • Sep 19, 2026 0.4.0
    Add a private personal dictionary with Whisper hints and safe exact corrections for every model. Show an amber progress ring while transcription is processing, improve optimized ARM backend detection, and document custom voice UI integration. Keep keyboard switching and live subtitles safe on older Android versions.
  • Sep 18, 2026 0.3.1
    Model selection is now the first section on the Speech models screen. Recommendations and verified downloads follow underneath, with manual import and source links kept as secondary options. Normal ready-state status and retry controls stay hidden until they are needed.
  • Sep 17, 2026 0.3.0
    Offline-first model setup and a faster, denser voice workflow. - Includes the recommended Parakeet model so transcription works without a download. - Adds guided model recommendations, verified in-app downloads, and practical device-fit estimates. - Keeps download progress and cancellation on the model card that started the download. - Adds compact pause, cancel, insert, and insert-and-switch keyboard controls. - Improves hold-drag Backspace selection and release-to-delete behavior. - Adds foreground recording controls and opt-in, content-free diagnostic logging.
  • Sep 12, 2026 0.1.20
    - Keeps the APK compact with clear, verified on-demand Parakeet model setup - Improves the true-black voice-input interface and model guidance - Replaces the launcher and store artwork with a circular black icon and white microphone
  • Sep 11, 2026 0.1.19
    - Introduces the Offline Speech to Text identity and logo - Enables private, caller-supplied audio transcription for app integrations