CLASSEVE
RouteLearn
Learn / On-device speech to text

What is on-device speech to text?

On-device speech to text is voice-to-text that runs entirely on your own hardware: a speech model on your CPU or GPU converts microphone audio into text locally, so the audio is never uploaded to a server. After the model is set up, transcription itself can run without a network connection.

Also called:offline speech to textprivate speech to textlocal speech to textspeech to text without internetoffline dictation

How it differs from cloud speech-to-text

Cloud dictation streams your microphone audio to a server, which returns text. That can offer strong accuracy on weak hardware, but every word you speak transits and is processed off your machine, and nothing works when the network doesn't.

On-device flips the trade: the model runs where you are. Privacy becomes structural — audio has nowhere to go — and dictation keeps working on a plane, on hostile Wi-Fi, or in a security-conscious workplace. Your hardware does the work: Lven Instant's engine needs about 600 MB of memory and no GPU.

What 'offline' means here

In any credible on-device product, the transcription path is local. Setup still downloads the model once, and products with accounts periodically check your access over the network. The distinction that matters: none of those requests carry your audio or your transcript text.

Where it's heading

Speech models keep shrinking while consumer chips keep gaining AI throughput, so the accuracy gap between local and cloud keeps narrowing. Dictation is following the same arc storage and photography followed: from a service back into the device.

From ClassEve

Lven Instant is ClassEve's on-device voice-to-text for Windows, Linux, and Android: unlimited dictation, audio never uploaded, one account across up to five devices, free.

On-device speech to text · FAQ

Common questions.

Does on-device dictation work with no internet at all?
The transcription itself does, once the model is installed. Products with accounts still need occasional connectivity for setup, updates, and access checks — but those requests don't include your audio or transcripts.
Is on-device transcription accurate enough to replace cloud dictation?
Yes, on modern hardware: Lven Instant's engine measures a 3.76% word error rate at about six times faster than you speak, on consumer-grade hardware with no GPU.
Which ClassEve product does on-device transcription?
Lven Instant — on-device voice-to-text for Windows, Linux, and Android, with unlimited dictation and no audio uploaded, free. Lven Cloud is the server-transcribed option, for machines too small for the local model.