๐Ÿง Linux Native

RocketWhisper Linux Edition

Fully Offline, One-Time Purchase
AI Speech Recognition & Transcription

๐Ÿ”’ Privacy Protected
โšก CUDA Accelerated
๐ŸŒ Works Offline
Supported Platforms:
๐ŸŸ  Ubuntu ๐Ÿ”ด Debian ๐Ÿ”ต Fedora ๐ŸŸข Arch
RocketWhisper
Converting speech to text...

๐Ÿ’ญ Sound Familiar?

๐Ÿ”

Worried About Confidential Data

Cloud-based speech recognition is convenient, but sending sensitive data to external servers is a concern...

๐ŸŒ

No Internet Available

Need voice input on air-gapped environments or development servers with network restrictions...

โŒจ๏ธ

Tired of Typing

Writing documentation while working in the terminal. If only you could just speak and type...

๐Ÿง

Few Linux Options

Plenty of options for Windows/Mac, but hard to find a quality speech recognition app for Linux...

โœจ RocketWhisper Solves It All!

๐Ÿš€

Fully Local Processing

Runs OpenAI Whisper locally.
Your voice data never leaves your machine.
Safe to use even in air-gapped environments.

โœ… No internet connection required
โœ… Confidential data stays safe
โœ… No server outage impact
โœ… Zero monthly fees
CUDA Accelerated

Blazing Fast Recognition with GPU

On machines with NVIDIA GPUs, CUDA hardware acceleration delivers incredible processing speed.

DGX Spark โœ… Optimized
Jetson AGX Orin โœ… Supported
Jetson Orin NX/Nano โœ… Supported

๐Ÿ’ก Works on CPU too when no GPU is available (auto-detected)

๐ŸŽฎ
10x Faster*
128GB VRAM Support

*Compared to CPU with large-v3-turbo model

๐Ÿ†• NEW in v1.2.2

๐Ÿค– Even More Powerful with AI

Works with local LLMs (Ollama, etc.).
Completely offline and free AI features.
No API key required. No cloud communication. Full privacy protection.

๐ŸŽฏ

AI Command Mode

Simply give voice instructions on selected text. "Translate to Japanese", "Summarize this", "Make it formal", etc.
Powered by local LLM — works offline

1 Select text
โ†’
2 Press hotkey
โ†’
3 Speak command
โ†’
4 AI processes it!
๐Ÿ”

Voice Search

Just say "Search for ..." and it automatically opens a browser search. Research made effortless.

"Search for Docker Compose"
"Look up Rust ownership"
"What is systemd?"

โญ 17 Premium Features

๐ŸŽ™๏ธ

High-Accuracy Recognition

State-of-the-art speech recognition powered by OpenAI Whisper. Excellent for both Japanese and English.

โŒจ๏ธ

Global Hotkey

Start recording with a single hotkey from any app. No need to interrupt your workflow.

๐Ÿ“‹

Auto Paste

Recognition results are automatically typed into the active window. No copy-paste needed.

๐ŸŽฏ

Per-App Processing Modes

Automatically apply different settings for each app. Separate configs for VSCode, Slack, and more.

๐Ÿค–

AI Processing

Auto-format recognized text with local LLMs (Ollama, etc.). Grammar correction, formalization, and more. Completely offline and free.

๐Ÿ“

Correction Rules

Automatically fix common misrecognitions. Regular expressions supported.

๐Ÿ’ฌ

Voice Commands

Edit text with voice commands like "new line", "period", and more.

๐Ÿ”

Voice Search

Say "Search for ..." to automatically open a browser search.

๐Ÿš€

Voice Launcher

Launch apps by voice. Say "Open VS Code" and it starts.

๐ŸŽฏ

AI Command Mode

Give voice instructions on selected text. Translation, summarization, and more via local AI. No API key required.

๐Ÿ“–

Custom Dictionary

Register technical terms and proper nouns to improve recognition accuracy.

๐Ÿ”Š

Notification Sounds

Audio cues for recording start/stop. Stay informed even when looking away.

๐Ÿ“

Batch Processing

Transcribe multiple audio and video files at once. Supports MP4, MKV, and more. SRT/VTT subtitle export available.

โœจ

Custom Instructions

Register frequently used AI tasks to dedicated hotkeys. Translation, formal language, summarization โ€” up to 20 custom instructions, each triggered with a single key.

๐Ÿ“œ

Recognition History

All recognition results are automatically saved. Search, copy, and reuse anytime.

๐Ÿ”ด

Recording Indicator

Floating indicator at the bottom of the screen during recording. Real-time audio level waveform display.

๐ŸŒ™

Dark Theme

Easy-on-the-eyes dark mode built in. Comfortable for extended use.

๐Ÿ“š Get Started in 3 Steps

1

Install Dependencies

sudo apt install pulseaudio-utils xdotool xclip ffmpeg

For Ubuntu/Debian. Use dnf for Fedora, pacman for Arch.

2

Download & Run the AppImage

chmod +x RocketWhisper-*.AppImage
./RocketWhisper-*.AppImage

No installation required. Just download and run.

3

Start Recording with a Hotkey!

F8 Start/Stop Recording

Press to talk. Press again to recognize. That's it.

๐Ÿ–ฅ๏ธ System Requirements

๐Ÿ“ฆ Supported Distributions

Ubuntu 20.04 LTS or later
Debian 11 (Bullseye) or later
Fedora 35 or later
Linux Mint 20 or later
Pop!_OS 20.04 or later
Arch Linux Rolling

๐Ÿ’ป Architecture

ARM64 (aarch64) โœ… Recommended (DGX Spark, etc.)
x86_64 (AMD64) โœ… Supported

๐Ÿ“‹ Required Packages

pulseaudio-utils Microphone recording (parec)
xdotool Keyboard automation
xclip Clipboard access
ffmpeg Audio conversion (optional)

๐ŸŽฎ CUDA-Compatible GPUs

DGX Spark Blackwell โœ… Optimized
Jetson AGX Orin Ampere โœ…
Jetson Orin NX/Nano Ampere โœ…
Falls back to CPU when no GPU is present

โš ๏ธ Display Server

X11 โœ… Fully supported (recommended)
Wayland โš ๏ธ Limited support
Some features are limited under Wayland

๐Ÿ’พ Memory Requirements

small/medium model 8GB or more
large-v3-turbo 8GB or more recommended
large-v3 16GB or more recommended
โš ๏ธ

For Wayland Users

The following features are limited under Wayland:

  • Global hotkeys (ydotool required, may need root privileges)
  • Auto paste (wl-clipboard required)
  • AI Command Mode (affected by clipboard restrictions)
  • Per-app processing (limited window detection)
Additional packages: sudo apt install ydotool wl-clipboard

๐Ÿ’ก Log in with an X11 session for full functionality

๐Ÿ“ฅ Download

๐Ÿ’ป

x86_64

AMD64 / x86_64

For standard Linux PCs
and servers

~196MB
Download
โš ๏ธ

Before You Download

Running RocketWhisper requires installing dependency packages.
Please install the required system packages for microphone recording and keyboard automation beforehand.

sudo apt install pulseaudio-utils xdotool xclip ffmpeg

For detailed installation instructions, see Help: Installation.
Package lists and commands for each distribution are available at Help: Required Packages.

๐Ÿ“Œ Latest version: v1.2.2
๐Ÿ“‹ AI model is automatically downloaded on first launch
๐Ÿ”’ License: Commercial (30-day free trial)

๐Ÿ’ฐ Licensing

Personal License

ยฅ4,800 (excl. tax)
  • โœ… All features included
  • โœ… Valid for 1 PC
  • โœ… Perpetual license
  • โœ… Free updates
Purchase

Free Trial

ยฅ0 30 days
  • โœ… All features included
  • โœ… No credit card required
  • โœ… No automatic billing
Try Now

๐Ÿ‘ฅ Perfect For

๐Ÿ‘จโ€๐Ÿ’ป

Linux Developers

Write documentation while working in the terminal. Input text without leaving the keyboard.

๐Ÿ”’

Security-Conscious Users

Safe to use even in environments handling confidential data. No data is ever sent externally.

๐Ÿ–ฅ๏ธ

System Administrators

Works in air-gapped and network-restricted environments.

๐ŸŽฎ

GPU Developers

Leverage DGX Spark and Jetson hardware. Blazing-fast recognition with CUDA acceleration.

โ“ Frequently Asked Questions

Q Is the license shared with the Windows version?
A Yes, the license is universal. A license purchased for the Windows version works on the Linux version as well.
Q Can I use it without a GPU?
A Yes, it works on CPU as well. CUDA acceleration is automatically enabled when a GPU is available, but the app works fine without one.
Q Does everything work on Wayland?
A Some features are limited under Wayland. For full functionality, please log in with an X11 session.
Q Do I need to extract the AppImage?
A No, you can run it directly. Just grant execute permission with chmod +x. On systems without FUSE, you can extract it with --appimage-extract.
Q Can I use it for meeting transcription?
A Yes, you can transcribe meetings in real-time or from recorded audio/video files via batch processing. AI processing can automatically summarize and format the results.
Q Can it create subtitles from video files?
A Yes, RocketWhisper can transcribe video files (MP4, MKV, AVI, MOV, WebM) and export subtitles in SRT and VTT formats.