Export Formats

Download your transcripts in any format you need. STT.ai supports six export formats, each optimized for different workflows.

ทำงานกับวิดีโอและเสียงที่เปิดให้ใช้โดยทั่วไป ไม่รองรับเนื้อหาที่ได้รับการปกป้องด้วย DRM

ปรับปรุงสำหรับ Enhanced
ส่วนตัว
คุยกับแปล
เปิดล็อคด้วยโปร →
วางแฟ้มที่นี่ หรือคลิกเพื่อค้นหา
MP3, WAV, M4A, FLAC, MP4, MKV, MOV, WebM - สูงสุด 2GB
ปรับปรุงสำหรับ Enhanced
ส่วนตัว
คุยกับแปล
เปิดล็อคด้วยโปร →
ปรับปรุงสำหรับ Enhanced
บันทึก: 0:00
ตามเวลาจริง ขี้ผึ้ง (ชั่วคราว)
เพิ่มประสิทธิภาพ กระซิบ (แม่นยำ)
ลิงค์สาธารณะ: 24 ชั่วโมง, ข้อความเท่านั้น · ลงทะเบียน สำหรับ 7d + เสียง · โปร สำหรับลิงก์ส่วนตัว

คำพูดเป็นข้อความแบบเรียลไทม์ AI ปรับปรุงอัตโนมัติเมื่อคุณพูด - ความแม่นยำจะดีขึ้นเมื่อคุณพูดนานขึ้น

ทดสอบไมโครโฟนก่อน
❤️ รัก STT.ai บอกเพื่อนๆ
คุณใช้การแปลภาษาฟรีของคุณ

ลงทะเบียนฟรีเพื่อรับ 600 นาที หรือปรับปรุงจาก $5/เดือน สำหรับอีกหลายพัน

10 นาทีฟรี/ วัน 600 นาทีฟรี กับการสมัคร ไม่มีบัตรเครดิต เข้ารหัสไว้
ลงทะเบียนฟรี →

Supported Export Formats

After transcribing your audio or video, you can download the transcript in any of the following formats. All formats include the full transcript text, and timed formats include word-level or segment-level timestamps.

TXT (Plain Text)

.txt

Simple plain text transcript without formatting. Best for copying into documents, emails, or other applications. Includes speaker labels when speaker detection is enabled.

Free plan

SRT (SubRip Subtitle)

.srt

The most widely supported subtitle format. Includes sequential numbering, timestamps, and text. Compatible with YouTube, Vimeo, VLC, Premiere Pro, Final Cut, and virtually every video player and editor.

Free plan

VTT (WebVTT)

.vtt

Web Video Text Tracks format, the standard for HTML5 video captions. Supports styling, positioning, and metadata. Used by web browsers, streaming platforms, and modern video players.

Basic plan+

DOCX (Word Document)

.docx

Formatted Word document with proper headings, timestamps, and speaker labels. Ideal for meeting minutes, reports, and documents that need further editing in Microsoft Word or Google Docs.

Basic plan+

JSON (Structured Data)

.json

Machine-readable structured format with word-level timestamps, confidence scores, speaker IDs, and segment data. Perfect for developers building on top of STT.ai or feeding data into other systems.

Basic plan+

PDF (Portable Document)

.pdf

Professional formatted PDF with timestamps, speaker labels, and STT.ai branding. Ideal for sharing with clients, archiving records, or printing. Layout is optimized for readability.

Basic plan+

Format Comparison

ตัวเลือก TXT SRT VTT DOCX JSON PDF
Plain text
Timestamps
Speaker labels
Word-level timing
Confidence scores
Video player compatible
Editable
Machine-readable

Which Format Should You Choose?

For subtitles and captions

Use SRT for maximum compatibility or VTT for web-based video players. SRT works with YouTube, Vimeo, Premiere Pro, Final Cut, and DaVinci Resolve.

For documents and reports

Use DOCX for editable documents or PDF for sharing and archiving. Both include formatted timestamps and speaker labels.

For developers and integrations

Use JSON for the richest data including word-level timestamps, confidence scores, and speaker IDs. Ideal for building custom applications.

For quick copy-paste

Use TXT for a simple plain text transcript you can paste anywhere -- emails, notes, chat, or any text field.

Batch Export

Need to export multiple transcripts at once? STT.ai supports batch export from your transcript library. Select multiple transcripts, choose your format, and download them all in a single ZIP file. Available on all paid plans.

API Export

Developers can retrieve transcripts in any format via the STT.ai API. Simply specify the desired format in your API request and receive the formatted output directly. The JSON format includes the most detailed data including word-level timestamps and confidence scores.

Transcribe and export in any format

Upload audio or video. Choose your export format. Download instantly.

Start Transcribing Free

คำถามที่พบบ่อย

ส่งออกรูปแบบ ทำงานในเบราว์เซอร์ของคุณ: ปักหมุด URL, โหลดแฟ้ม, หรือบันทึกจากไมโครโฟนของคุณ STT.ai เลือกโมเดล AI และส่งผลลัพธ์กลับมาเป็นข้อความในเวลาไม่ถึง5นาที ส่งออกเป็น TXT, SRT, VTT, DOCX, JSON หรือ PDF

ใช่ — ผู้เข้าชมทุกคนจะได้รับ 600 นาทีฟรี เพื่อเริ่มต้นบน STT.ai, ใช้ได้สำหรับ ส่งออกรูปแบบ เหมือนกับกระบวนการทำงานอื่น ๆ ค่าเริ่มต้นของแผนการจ่ายเริ่มที่ $5/ เดือน เปิดใช้งานแฟ้มที่ยาวกว่า, ส่วนตัวตีความและคิวความสำคัญ

ส่งออกรูปแบบ ทำงานบนโมเดล AI เดียวกันกับ STT.ai ส่วนอื่น ๆ - โมเดลที่ดีที่สุดของเรามีค่าความแม่นยำ 95- 97% ในการพูดอย่างชัดเจน (อัตราคำผิดพลาด 3- 5% ตามการทดสอบ) เปลี่ยนโมเดลโดยทันที หากการผ่านครั้งแรกต่ำกว่าเป้าหมายของคุณ

ส่งออกรูปแบบ สามารถทำงานบนเครื่อง STT.ai รุ่น 10+ รุ่นใดก็ได้ - STT.ai Enhanced (แม่นยำที่สุด), Whisper Large V3 (ภาษา 99 ภาษา), NVIDIA Canary (#1 WER on supported langs), Whisper Turbo (เร็วที่สุด), Moonshine (น้ำหนักเบาที่สุด) และอื่นๆ

ใช่ ทุกๆ ส่วนที่แปลออกมาจะถูกส่งออกเป็นรูปแบบ SRT หรือ VTT ทำงานกับ YouTube, Vimeo, TikTok, VLC และเครื่องเล่นวิดีโอหลักๆ ทุกเครื่อง เครื่องมือเขียนคำอธิบายจะนำมันมาวางบนวิดีโอเป็นคำอธิบายแบบ Hardsub

ใช่ การจัดเรียงเสียงให้เป็นแผ่น จะทำการตั้งชื่อเสียงแต่ละเสียง (ผู้พูด 1, ผู้พูด 2,...) โดยอัตโนมัติ และคุณสามารถเปลี่ยนชื่อเสียงได้ในตัวแก้ไขที่ติดตั้งไว้ ทำงานได้กับทุกรุ่นและภาษา

งาน ส่งออกรูปแบบ ส่วนใหญ่จะเสร็จสมบูรณ์ภายในเวลาไม่ถึง5นาที แฟ้มเสียง 1 ชั่วโมง จะเสร็จสมบูรณ์ภายในเวลา 2-3 นาที ด้วยโมเดลที่เร็วที่สุดของเรา ความเร็วขึ้นอยู่กับโมเดลที่เลือกและค่าแรงของ GPU ปัจจุบัน

ส่งออกรูปแบบ รองรับรูปแบบมากกว่า 20 รูปแบบ - MP3, WAV, M4A, FLAC, OGG, MP4, MKV, MOV, WebM, AVI และอื่น ๆ อีก นำออกมาเป็น TXT, SRT, VTT, DOCX, JSON หรือ PDF

ใช่ แฟ้มเสียงที่ส่งไปยัง ส่งออกรูปแบบ จะถูกประมวลผลและลบโดยปริยาย แผน Pro เพิ่มการเข้ารหัสด้านคลาวด์ - แม้ว่าฐานข้อมูลของ STT.ai จะถูกทำลาย ส่วนที่คุณเขียนจะอ่านไม่ได้หากไม่มีกุญแจของคุณ ข้อมูลจะไม่ถูกใช้ในการฝึกโมเดลโดยไม่ต้องเลือกเข้าร่วมอย่างชัดเจน

ใช่ STT.ai เสนอ API REST กับ Python และ Node.js SDKs, รวมถึงเซิร์ฟเวอร์ MCP สำหรับ Claude และ Cursor - ทั้งหมดใช้ได้กับ ส่งออกรูปแบบ workflows ระดับ API ฟรี รวมถึง 100 นาที/เดือน

ใช่ ทุกๆ ข้อความจะเปิดในเครื่องมือแก้ไขที่ติดตั้งไว้ เพื่อให้คุณสามารถแก้ไขคำ เปลี่ยนชื่อผู้พูด ปรับเวลา และเพิ่มข้อความ ทุกๆ การเปลี่ยนแปลงจะถูกบันทึกอัตโนมัติ

ทุกๆ ส่วนของการแปลจะได้รับ URL ที่สามารถแบ่งปันได้ แบบเอกสาร DOCX หรือ PDF เพื่อใช้ส่งอีเมล์ แบบโปรเพิ่มการป้องกันด้วยรหัสผ่าน และลิงก์ถาวร - เหมาะสำหรับงานของลูกค้า

STT.ai จัดการกับแพลตฟอร์มมากกว่า 1,300 แพลตฟอร์ม รวมถึง YouTube, Vimeo, TikTok, SoundCloud, Zoom, Google Meet, เจ้าของพอดคาสต์ และอื่น ๆ การแปล URL ทำงานกับเนื้อหาที่เปิดเผยเท่านั้น - ต้นกำเนิดที่ป้องกัน DRM ไม่สามารถแปลได้