Private by architecture
Microphone audio, model artifacts, and transcripts remain on your machine. No account, API key, or remote inference endpoint.
VoicePad turns speech into text on your own NVIDIA GPU. Durable audio, resident inference, and a global shortcut—without sending a word to the cloud.
Dictation should feel immediate, private, and dependable. Your audio stays here while the model turns every phrase into text.
VoicePad keeps capture, inference, and presentation separate so a slow model—or a failed job—never gets to decide whether your recording survives.
Microphone audio, model artifacts, and transcripts remain on your machine. No account, API key, or remote inference endpoint.
Canonical WAV audio is written continuously. Existing recordings are never overwritten, modified, or deleted automatically.
Let your Wayland compositor own the global key. VoicePad receives a secure local toggle and reports state through native notifications.
Parakeet loads and warms once, then stays resident while the TUI is open. CPU-side VAD finds natural boundaries; CUDA handles transcription.
Explore the architecture →Disk-backed from the first sample.
Silero VAD on lightweight CPU frames.
Official Parakeet FP16 safetensors.
Conservative overlap reconciliation.
VoicePad currently prioritizes one reliable deployment over a matrix of silent fallbacks. Unsupported hardware fails clearly before recording starts.
Prepare the verified deployment, keep VoicePad resident, and dictate from anywhere on your desktop.