Skip to content

feat: support 60db speech-to-text for voice commands - #301

Open
uditgoenka wants to merge 1 commit into
Robitx:mainfrom
uditgoenka:feat/60db-stt
Open

uditgoenka wants to merge 1 commit into
Robitx:mainfrom
uditgoenka:feat/60db-stt

Conversation

@uditgoenka

@uditgoenka uditgoenka commented Oct 1, 2026 •

Copy link
Copy Markdown

60db provides a hosted speech-to-text API. This change adds it as an optional provider for the GpWhisper* commands.

Configure whisper = { provider = "60db", language = "auto" } and SIXTYDB_API_KEY, or use the existing secret resolver through whisper.secret. OpenAI remains the default.

Recording, SoX processing, secret resolution, and task management are reused. The adapter passes request arguments directly to curl, rejects empty or oversized recordings, and handles HTTP/JSON failures without inserting error bodies into the buffer. README setup includes the audio-upload behavior. No plugin dependencies are added.

Validation: headless Neovim checks passed for routing, multipart arguments, response handling, size limits, and cancellation. A real curl/local HTTP fixture checked authentication and quoted/spaced filenames. Whitespace checks passed.

Live transcription and physical microphone/SoX recording were not tested.

Reuse the Whisper recording flow with independent transcription credentials and provider-specific multipart fields.

Confidence: high
Scope-risk: narrow
Tested: Headless Whisper flow checks and real curl local multipart fixture
Not-tested: Live 60db transcription and physical microphone recording
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant