Speak anywhere you type.

Viska turns your voice into finished text at the cursor — in email, chat, documents or AI prompts. It is the private alternative to Wispr Flow: your speech becomes text on your own computer, offline, not in someone’s cloud. Free and open source for Mac, Windows and Linux.

For Apple Silicon Macs (0.7.0), Windows 11 (0.6.0) and Linux (0.7.1).

Private by default. Works offline. No account, no subscription.

New message

To Anna Lind

Subject Board documents

Hi Anna, thanks for the meeting yesterday. I have sent the documents to the board, and we can go through the budget on Wednesday before it goes out.

hi anna thanks for the meeting yesterday I have sent the the documents to the board and we can go through the budget on wednesday before it goes out

Proofread · Intuitive Example text. The grey line is what was said; above it, what Viska typed.

02What Wispr Flow does in the cloud, Viska does on hardware you own.

Private dictation means your speech becomes text on your own devices, not on someone else’s servers. With host mode, your own GPU becomes a private dictation service for your laptop and your Android phone.

Viska and Wispr Flow compared, including where Wispr Flow is ahead →

03Stay in your app. Viska comes to you.

Viska is a speech-to-text app that works system-wide. It waits in the background, so you never switch windows to dictate, and it types only where you put the cursor.

  1. Put the cursor where the words should go

    An email, a chat, a document, a terminal, a prompt box. Any app that takes text.

  2. Double-press a key and speak

    A small overlay shows that Viska is listening. Pause, think, carry on; it keeps listening until you stop.

  3. Stop, and the text is there

    Whisper transcribes on your computer, proofreads if you want it to, and types the result at the cursor. If an app blocks typing, the text waits on the clipboard and Viska tells you.

Three gestures, the same on every system

PressMacWindows & Linux
Start or stop recordingleft ⌘ twiceleft Ctrl twice or CtrlAltD
Record with the other proofreading choicehold ⇧, left ⌘ twicehold Shift, left Ctrl twice or CtrlAltShiftD
Switch dictation languageright ⌘ twiceright Ctrl twice or CtrlAltShiftL

Double-press gestures need macOS Input Monitoring, an opt-in keyboard hook on Windows, or opt-in input access on Linux. The key combinations work everywhere without it. macOS setup guide

04Spoken language in, written language out.

Speech is full of “um”, restarts and second thoughts. Turn on proofreading and every recording is cleaned up before it reaches the cursor — or keep it off and use the second gesture when a message matters.

What was said

um so I was thinking we could uh meet on Tuesday no Wednesday actually and go over the the budget before we send it to

Grammar

Um, so I was thinking we could, uh, meet on Tuesday—no, Wednesday, actually—and go over the budget before we send it to

Fixes spelling, punctuation and grammar. Keeps your words and your voice.

Intuitive

So I was thinking we could meet on Wednesday actually and go over the budget before we send it to

Also drops filler words, restarts and self-corrections, and finds the word you meant.

I’ve been testing cloud code from anthropic with the new opus model I’ve been testing Claude Code from Anthropic with the new Opus model.

Real outputs from DeepSeek V4.1 Flash with Viska’s bundled prompts, 25 September 2026. When a name was misheard, Intuitive writes the one you meant if the context makes it clear, and otherwise leaves your words alone. A spoken question such as “can you explain what a vector database is” comes back corrected — “Can you explain what a vector database is?” — not answered.

Viska settings: proofread automatically after each recording, mode Grammar or Intuitive, on this computer with Ollama or in the cloud with Ollama Cloud and DeepSeek V4.1 Flash

Local or cloud — your call

Proofreading runs on your computer through Ollama, or in the cloud with your own key. Nothing is sent anywhere until you choose a cloud provider.

Our cloud pick: Ollama Cloud

Ollama states that prompts and responses on its cloud are never logged or trained on, for every model it hosts (ollama.com/pricing). Viska defaults to DeepSeek V4.1 Flash there: capable enough to understand what you meant, answered in 0.3 to 0.9 seconds per request in our tests on 25 September 2026, and costs $0.30 per million input tokens and $1.20 per million output tokens at peak rates — a typical dictation is a fraction of a cent.

Using OpenRouter instead?

Turn on Zero Data Retention in your OpenRouter privacy settings. OpenRouter then routes your requests only to providers that keep nothing, and OpenRouter itself does not retain prompts unless you opt in to logging. OpenAI and Anthropic are available too, under their own terms.

Checked 25 September 2026.

05Your voice stays in the room.

Dictation is personal: names, health, money, half-finished thoughts. Viska is built so that none of it has to leave your computer.

By default

  • Audio stays in memory and is never written to disk unless you save a recording yourself.
  • Speech-to-text runs locally with Whisper models on your CPU or GPU.
  • No account, no telemetry, no analytics in the app.
  • History is off unless you turn it on, and then it stays on your disk.

Only if you choose

  • Cloud speech engines (Groq, OpenAI, Deepgram, ElevenLabs, Mistral, xAI, Gemini, OpenRouter) with your own API key.
  • Cloud proofreading, labelled in the app with the provider’s name.
  • API keys are kept in your system keychain, never in a settings file.

06Viska and Wispr Flow, side by side.

Both put dictated text where your cursor is. They differ in where your voice goes and what it costs. Facts about Wispr Flow checked 30 September 2026.

ViskaWispr Flow
Where your voice goes Stays on your devices. With host mode it goes only to your own computer over your private Tailscale network. Works offline. To Wispr’s cloud: “Transcription always occurs on the cloud.” Data is processed and stored in the United States; you can switch off cloud storage of your dictations. source source
Price Free, open source (GPL-3.0). No account, no word limit. Free plan capped at 2,000 words a week on desktop; Pro $15 per user per month, or $12 billed yearly source
Computers Mac, Windows and Linux Mac and Windows; no Linux source
Phone Android, with Viska Voice, transcribed on your own computer. No iPhone app. iOS and Android apps, transcribed in the cloud source
Use a desktop GPU from a laptop or phone Yes, privately over Tailscale (host mode) Not applicable: transcription happens in Wispr’s cloud

The full comparison, including where Wispr Flow is ahead →

07Nine languages, and a specialist for Swedish.

English, Swedish, German, French, Spanish, Italian, Russian, Chinese and Japanese, or let Viska detect the language. Switch with a double-press of the right modifier.

Viska Models window listing English and Swedish specialist models with speed labels, sizes and Apple Metal acceleration

The right model for your machine

Pick a model by speed and accuracy, from 30 MB to large multilingual Whisper. Swedish uses KB-Whisper from the National Library of Sweden.

Apple Silicon uses Metal. Compatible AMD and Intel graphics use Vulkan on Linux, and NVIDIA cards use CUDA. Long dictations are split at natural pauses, so the seams do not add stray words.

08Share your GPU.

One computer with a strong graphics card runs Viska in host mode and does the transcribing for your other devices, privately over your own Tailscale network. In the Swedish interface it is called Värdläge – dela GPU. It is off until you switch it on.

  1. Turn on host mode

    On the computer with the GPU, open Settings and choose Host mode. Pick a Whisper model for each language, for example KB-Whisper for Swedish and a Whisper model for English. Viska suggests models that fit your graphics card and downloads what is missing.

  2. Publish it privately

    The host itself only listens on localhost. One click runs tailscale serve, which makes it reachable from your own tailnet over HTTPS and nowhere else.

  3. Pair a device

    Show a QR code or copy a pairing link. It carries the host address and a bearer token, and it is a secret: treat it like a password. Renew the token if a device is lost.

Who can use the host

  • Another computer. Choose the engine “Remote GPU (Viska host)” (Fjärr-GPU) and paste the pairing link, for example a Mac (Apple Silicon) or a Linux laptop without a big GPU. If the host cannot be reached, Viska falls back to the model on that computer.
  • An Android phone. Viska Voice is a voice keyboard that sends your speech to the host and types the text into the app you are in.
  • The host can also run without a window (viska --host) or as a background service.

What stays private

  • Nothing goes to the cloud. Audio travels only between your own devices over Tailscale.
  • Every request except a health check needs the bearer token, which the host keeps in the system keyring. The host does not log audio, text or the token.
  • Models are loaded when they are needed and unloaded after 20 idle minutes by default, so the GPU is free again when nobody is dictating. Dictation on the host computer itself always goes first.
  • Proofreading through the host is optional and uses the host’s own provider and settings.

Host mode and the Remote GPU client are in Viska 0.7.1 for Linux and macOS, available now. Windows gets them in a later build; it is still at 0.6.0. How fast it is depends on your GPU and the model you choose; we publish no benchmark for it. Linux setup notes

09Download Viska.

Free, no account. Linux and macOS are at version 0.7.1, which adds host mode and the Remote GPU client; Windows is still at 0.6.0 and gets them in a later build. The Mac app is signed by Aeon Media Group AB and notarised by Apple. The Windows and Linux builds are not signed yet, so Windows asks you to confirm the first launch; the guides walk you through it.

macOS

Apple Silicon (M1 and later). Open the disk image and drag Viska to Applications. Grant Input Monitoring for the gestures.

Download for Mac
Version
0.7.0
Size
132.3 MiB
SHA-256
4af6d0d497c947cc81acc5b79ea7dcd9c03ad655cb2793ece9c184f631619e9e

Signed with a Developer ID and notarised by Apple, so it opens without a warning and should keep its permissions when you update. Intel Macs are not supported.

Windows

Windows 11 on Intel or AMD (x64). Installs for your user account, no administrator rights needed. Experimental preview.

Version
0.6.0
Size
214.1 MiB
SHA-256
78b7f641694a188de2ba28ee2538c4796ea9b7be46268b26c21054c65e799698

Unsigned: SmartScreen may warn. Assembled from four parts and checked in your browser before saving. Microphone and paste still need more real-desktop testing; Windows on ARM is untested.

Debian & Ubuntu

x86_64 Debian 12+, Ubuntu 24.04+ and derivatives, including Kubuntu. Includes Vulkan acceleration for compatible AMD graphics and optional NVIDIA CUDA.

Download .deb
Version
0.7.1
Size
25.6 MiB
SHA-256
842f9adc9e9e3cc9459ff45be5eb060d12dcc0b65cbebcb59870a3988302007c

Arch & Omarchy

For x86_64 Arch Linux and Arch-based Omarchy/Hyprland. A native package with its own Qt, Python and whisper.cpp runtime.

Version
0.7.1-1
Size
286.7 MiB
SHA-256
117d9e09b0e6e652b06000edf681c417569ed002c78a7a65dbc396bab08774ca

Desktop integration on Hyprland is still experimental.

All checksums: SHA256SUMS. Tested most on KDE Plasma; GNOME 48+ uses the same desktop portal. Other Linux distributions: Flatpak is planned.

Viska Brev proofreads your email before you send it.

A Thunderbird extension for Swedish and English drafts. Check spelling, grammar or expand a rough draft, review the changes as a diff, and keep quotes and signatures untouched. Runs through your own Ollama server or Ollama Cloud, OpenRouter or an SGLang server.

About Viska Brev →

Viska Brev in Thunderbird: a proofreading panel with a diff of the changes and Replace and Discard buttons

Viska Voice types with your voice on Android.

A voice input method for your phone that sends speech over Tailscale to your own computer running Viska 0.7.1 in host mode, and types the transcript into the field you are in. Works with a keyboard that has a voice key, such as FUTO Keyboard. No cloud service, no account, no telemetry. Free software (GPL-3.0-or-later).

About Viska Voice →

Viska Voice icon: a white V with a microphone and send-and-receive arrows on a green disc

Recent changes.

The three latest updates to Viska, Viska Brev and this site. Read the full changelog.

  1. Viska App 0.7.1

    Viska 0.7.1 for Debian/Ubuntu and Arch/Omarchy improves the intuitive proofreading mode: it ignores the pauses Whisper marks with periods, joins fragments into natural sentences, keeps short messages on one line and uses paragraphs only for longer messages. macOS stays at 0.7.0 and Windows at 0.6.0 for now.

  2. Website

    Published Viska 0.7.1 for Linux: the Debian and Arch packages, checksums and the Linux setup guide. The macOS and Windows downloads are unchanged (0.7.0 and 0.6.0).

  3. Viska App 0.7.0

    Viska 0.7.0 adds host mode: one computer with a strong GPU can transcribe for your other devices, privately over Tailscale. Published for Debian/Ubuntu and Arch/Omarchy; the macOS build followed the same day (see below), and the Windows download stays at 0.6.0 until its 0.7.0 build is ready.

10Questions.

Does Viska work offline?

Yes. Once a speech model is downloaded, dictation runs entirely on your computer with no internet connection. Cloud engines and cloud proofreading are optional and only used when you choose them.

Does my voice leave my computer?

Not by default. Viska transcribes on your own computer and keeps audio in memory. Your voice leaves the computer only if you choose it. With host mode, audio goes only to your own host over your private Tailscale network. If you pick a cloud speech engine such as Groq or OpenAI, it goes to that provider. Cloud proofreading, if you turn it on, sends the transcribed text, not the audio.

Is it really free?

Yes. Viska is open source under the GPL-3.0 licence: no account, no subscription, no word limit. If you add a cloud provider, that provider bills you directly for what you use, and the app can track the estimated cost.

Is Viska a private alternative to Wispr Flow?

For dictation on a computer, yes. Both type where your cursor is. Wispr Flow transcribes in the cloud, and its documentation says data is processed and stored in the United States; Viska transcribes on your own machine by default, is free and open source, needs no account, and can share your own GPU with your laptop and Android phone over Tailscale. Wispr Flow has iPhone and Android apps, more than 100 languages and a personal dictionary that learns names; Viska runs on Linux too. Read the full comparison.

Can I use Viska for voice typing in ChatGPT, Claude or my code editor?

Yes. Viska types into whatever app has the cursor, so it works for AI prompts in ChatGPT, Claude or Gemini, for email and chat, for documents and for code editors and terminals. In a terminal, paste with CtrlShiftV if typing is blocked.

Which speech recognition does Viska use?

OpenAI’s open Whisper models, run locally through whisper.cpp or faster-whisper, and KB-Whisper from the National Library of Sweden for Swedish. You can also plug in cloud engines such as Groq, Deepgram or ElevenLabs with your own key.

How accurate is it?

Viska uses OpenAI’s Whisper models and, for Swedish, KB-Whisper. Larger models are more accurate and slower; the Models window labels each one by speed and size so you can choose for your computer. Proofreading can then fix what was misheard.

Why does Windows warn me when I open the installer?

The Windows preview is not yet signed with a Microsoft code-signing certificate, so SmartScreen asks you to confirm. The Windows guide shows how to do that safely, and every download has a SHA-256 checksum you can verify. The Mac app is signed and notarised by Apple and opens normally; see the macOS guide for the permissions it needs.

Can I dictate on my phone?

On Android, yes, with Viska Voice: a voice keyboard that sends your speech over Tailscale to your own computer running Viska in host mode. There is no iPhone app, and nothing is transcribed by a cloud service.

Can one computer with a GPU serve my other devices?

Yes. Run Viska 0.7.1 in host mode, publish it with Tailscale, and pair your other computers or your Android phone. How host mode works.

Does it work on Wayland?

Yes. On KDE Plasma and GNOME 48+ the global shortcuts use the XDG GlobalShortcuts portal; Hyprland uses xdg-desktop-portal-hyprland. On older desktops, bind a system shortcut to viska --toggle.