Skip to content

Roadmap

  • Real-time transcription
  • Voice Activity Detection (Silero VAD v5)
  • Multi-platform (Android, iOS, Desktop, Web)
  • Spanish support + punctuation
  • Target language selection
  • Audio recording (WAV/OGG)
  • Session restart without crashes
  • Docker image on Hub
  • Landing + docs (this site) on GitHub Pages
  • Model switching UI (client)
  • Export transcriptions (TXT, SRT)
  • Authentication for remote deployments
  • HTTPS/WSS reverse proxy guide (Caddy/Nginx + Tailscale)
  • PWA packaging for Web App
  • i18n for docs (en/es)

Contributions welcome β€” see GitHub.