Politikkur og politikkfrøði

Hvat AI vit brúka, hvar tað koyrir, og hvussu vit fylgja við kravum um upplýsing (EU AI Act Article 50, effective 2026-08-02).

TD;DR

  • All transcripts are AI-generated. Every output carries a machine-readable disclosure.
  • All AI runs on our own GPU. We don't send your audio or text to OpenAI, Anthropic, Google, or any third-party LLM API.
  • We do NOT train base models on your transcripts. Only an opt-in fine-tune uses corrections you explicitly make.
  • Synthetic voice clones (TTS) eru merkt sum AI-generated í filename, metadata, og á síðuni.

Models we use

ModelNodeLicenseKýrir á
Whisper large-v3-turbo (faster-whisper)Transkriptión (standard)MITGPU
STT.ai Enhanced (custom fine-tune)Transkriptión (gjaldandi ætlan)MIT (base) / Proprietary (fine-tune weights)GPU
VoskRealtime word streamingApache 2.0GPU
SpeechBrain ECAPA-TDNNDimmalættingApache 2.0GPU
MadLAD-400 3BTranslate (450+ languages)Apache 2.0GPU
Qwen2.5-1.5B (llama.cpp)1995 - Sjónleikari, rithøvundur, sjónleikari, rithøvundur, sjónleikari, rithøvundur.Apache 2.0GPU
F5-TTSVoice cloning / text-to-speechMITGPU
all-MiniLM-L6-v2Søk á RAGApache 2.0GPU

The only external AI service we use is translateapi.ai (also Muddy Holdings) for translating UI strings — this never touches your translittered content.

Um at gera seg til ein ætt.

  • HTML-síður: include <meta name=\
  • Tekstútflutningur (TXT, SRT, VTT, JSON, CSV, DOCX, PDF): include an 'AI-generated transcript' header line at the top of every file.
  • Syntetisk rødd / TTS úttøka: WAV-filer indeholder et'synthetic-voice' tag i metadata og en tydelig meddelelse på downloadsiden. Hørbar ansvarsfraskrivelse er på vejkortet.
  • API responses: include an _ai_generated: true field in every JSON response that contains transscribed content.

Data

  • Base models (Whisper, MadLAD, Qwen, etc.) come pre-trained from their respective publishers. We use them as shipped.
  • Tí eru ikki øll tøknilig
  • Um tú rættar ein transkriptiónspart (blyant-táknið) ella merkir hann sum skeivt (flagg-táknið), OG tú hevur valið í /privacy-settings/ (\
  • The Contribute corrections plus audio to Voice Lab toggle (also at /privacy-settings/, also default off) allows you to contribute the audio of segments you correct, paired with the corrected text, to our Voice Lab dataset under CC-BY-SA-4.0. The two toggles are independent — you can grant either, both, or neither.
  • Lyd, sum tú sendir upp, verður slettað innan 24 tímar við cleanup_uploads cron — UM tú ikki hevur valið at \

Nøgd og feilir

AI- transkriptión er ikki fullkomin. Orðfeilar eru ymiskir eftir, hvussu nógvur aksent, ljóðkvalitetur, mál og orðatilfar er í málinum. Til kritiska nýtslu (juridisk, læknalig, lógarfestar vinnugreinar) skalt tú kanna tað móti upprunaliga ljóðinum. Våre offentlige WER- benchmarks pr. model er på /models/.

Compliance questions

For EU AI Act, GDPR, or other compliance questions: hello@stt.ai ella brúka contact form.

Síðst uppdatað: 2026-04-26.