Frequently Asked Questions

Is ScriptFlow really free?

Yes. Every transcription runs on your own device, so there is no server cost per file and no reason to charge or limit usage.

Are my files uploaded anywhere?

No. Files are processed locally in your browser using WebAssembly and an on-device AI model. They never leave your device.

What file formats are supported?

Any audio or video format your browser can decode — including MP3, WAV, M4A, MP4, MOV, and WebM.

Can I export the transcript?

Yes — copy it with one click, download it as a plain .txt file, or download an .srt or .vtt subtitle file with timestamps for use in video editors, players, and web video captions. You can also toggle per-segment timestamps on or off in the transcript view.

Why is the first transcription slow?

The first time you use ScriptFlow, your browser downloads a small speech-recognition model (roughly 40–75MB). After that it is cached, and every transcription after the first is fast and works offline.

How accurate is the transcript?

ScriptFlow uses a compact Whisper model tuned for speed on everyday hardware. It handles clear speech well; heavy background noise, overlapping speakers, or strong accents can reduce accuracy.