Politikkur og politikkfrøði
Hvat AI vit brúka, hvar tað koyrir, og hvussu vit fylgja við kravum um upplýsing (EU AI Act Article 50, effective 2026-08-02).
TD;DR
- All transcripts are AI-generated. Every output carries a machine-readable disclosure.
- All AI runs on our own GPU. We don't send your audio or text to OpenAI, Anthropic, Google, or any third-party LLM API.
- We do NOT train base models on your transcripts. Only an opt-in fine-tune uses corrections you explicitly make.
- Synthetic voice clones (TTS) eru merkt sum AI-generated í filename, metadata, og á síðuni.
Models we use
| Model | Node | License | Kýrir á |
|---|---|---|---|
| Whisper large-v3-turbo (faster-whisper) | Transkriptión (standard) | MIT | GPU |
| STT.ai Enhanced (custom fine-tune) | Transkriptión (gjaldandi ætlan) | MIT (base) / Proprietary (fine-tune weights) | GPU |
| Vosk | Realtime word streaming | Apache 2.0 | GPU |
| SpeechBrain ECAPA-TDNN | Dimmalætting | Apache 2.0 | GPU |
| MadLAD-400 3B | Translate (450+ languages) | Apache 2.0 | GPU |
| Qwen2.5-1.5B (llama.cpp) | 1995 - Sjónleikari, rithøvundur, sjónleikari, rithøvundur, sjónleikari, rithøvundur. | Apache 2.0 | GPU |
| F5-TTS | Voice cloning / text-to-speech | MIT | GPU |
| all-MiniLM-L6-v2 | Søk á RAG | Apache 2.0 | GPU |
The only external AI service we use is translateapi.ai (also Muddy Holdings) for translating UI strings — this never touches your translittered content.
Um at gera seg til ein ætt.
- HTML-síður: include <meta name=\
- Tekstútflutningur (TXT, SRT, VTT, JSON, CSV, DOCX, PDF): include an 'AI-generated transcript' header line at the top of every file.
- Syntetisk rødd / TTS úttøka: WAV-filer indeholder et'synthetic-voice' tag i metadata og en tydelig meddelelse på downloadsiden. Hørbar ansvarsfraskrivelse er på vejkortet.
- API responses: include an _ai_generated: true field in every JSON response that contains transscribed content.
Data
- Base models (Whisper, MadLAD, Qwen, etc.) come pre-trained from their respective publishers. We use them as shipped.
- Tí eru ikki øll tøknilig
- Um tú rættar ein transkriptiónspart (blyant-táknið) ella merkir hann sum skeivt (flagg-táknið), OG tú hevur valið í /privacy-settings/ (\
- The Contribute corrections plus audio to Voice Lab toggle (also at /privacy-settings/, also default off) allows you to contribute the audio of segments you correct, paired with the corrected text, to our Voice Lab dataset under CC-BY-SA-4.0. The two toggles are independent — you can grant either, both, or neither.
- Lyd, sum tú sendir upp, verður slettað innan 24 tímar við cleanup_uploads cron — UM tú ikki hevur valið at \
Nøgd og feilir
AI- transkriptión er ikki fullkomin. Orðfeilar eru ymiskir eftir, hvussu nógvur aksent, ljóðkvalitetur, mál og orðatilfar er í málinum. Til kritiska nýtslu (juridisk, læknalig, lógarfestar vinnugreinar) skalt tú kanna tað móti upprunaliga ljóðinum. Våre offentlige WER- benchmarks pr. model er på /models/.
Compliance questions
For EU AI Act, GDPR, or other compliance questions: hello@stt.ai ella brúka contact form.
Síðst uppdatað: 2026-04-26.