Licenses & credits
This site runs entirely in your browser using the following open components.
Speech recognition models
A different model is used per language. All of them are redistributed unmodified apart from packaging for WebAssembly.
- Kroko ASR Community models — © Banafo. Used for English, Italian, German, Spanish, French, Dutch, Portuguese, Swedish, Turkish and Hebrew. Licensed under CC BY-SA 4.0. Source: huggingface.co/Banafo/Kroko-ASR
- Korean streaming zipformer — Apache License 2.0. Trained on KsponSpeech with icefall and published as johnBamma/icefall-asr-ksponspeech-pruned-transducer-stateless7-streaming; converted to ONNX by k2-fsa.
- Russian streaming zipformer — Apache License 2.0. © Alpha Cephei (Vosk). Source: alphacep/vosk-model-small-streaming-ru
- Silero VAD — MIT License. github.com/snakers4/silero-vad
Sample audio
- The sample files on the English and Italian pages are excerpts from FLEURS (Google), licensed under CC BY 4.0. Source: huggingface.co/datasets/google/fleurs
Runtime
- sherpa-onnx — Apache License 2.0. github.com/k2-fsa/sherpa-onnx
- ONNX Runtime — MIT License. github.com/microsoft/onnxruntime
Interface
- Poppins — SIL Open Font License 1.1, served by Google Fonts.
Privacy
Audio and video files are processed locally and are never uploaded. Anonymous usage counters (page views, processing time, cache hits) are collected without cookies and contain no file names, audio, or transcripts.