SRT Subtitle Generator for Audio and Video
The SRT Subtitle Generator turns audio or video into timed cues you can review against playback and download without leaving this page. Your task and saved cue edits remain available in My records.
- audio and video upload
- numbered timed cues
- source playback
- cue text and timing edits
Sign in when you are ready to transcribe on this page.
- Upload audio or video
- Correct cue text and timing
- Download the saved SRT file
Sign in to create a saved task. If your file is unavailable after signing in, select it again.
Dashboard
How do you want to transcribe?
Free minutes are included. Upload a file or record audio to start.
What does SRT Subtitle Generator do?
SRT Subtitle Generator turns a supported audio or video upload into numbered, time-coded subtitle cues. The task remains on this page for playback and owner-only review; you can correct cue text and timing, save the cue revision, and download a UTF-8 SRT file whose sequence and comma-millisecond timestamps come from that saved version.
| Capability | Evidence in the workflow |
|---|---|
| Timed SRT result | The successful task opens on numbered subtitle cues and the only configured download format is SRT. |
| Cue-level review | Each start time, end time, and Unicode text value can be edited, validated, and saved as a versioned cue set. |
| Source verification | Owner-only playback remains beside the cues so wording and timing can be checked against the uploaded recording. |
| Durable export | Refresh and My records reopen the same task; downloads are generated from the saved cue revision rather than an unsaved browser draft. |

The deliverable: a separate, reviewable SRT file
The download is a caption file, not a burned-in video. Cue numbers start at 1, timestamps use HH:MM:SS,mmm, and the text comes from the cue revision you saved on this page.
How to use SRT Subtitle Generator
The SRT Subtitle Generator turns audio or video into timed cues you can review against playback and download without leaving this page. Your task and saved cue edits remain available in My records.
How it works: upload, review, and export
Choose a supported recording or capture audio in your browser, select the source language when that helps recognition, and submit the task. Sign-in happens in place: after authentication you return to this route, and WhisperWeb preserves the selected file when the browser allows it. If the browser cannot retain that local File object, the page asks you to select it again instead of claiming an upload was saved.
The task stays on this page through upload, queueing, processing, success, and recoverable errors. A task ID in the address lets the owner reopen the server-side result after a refresh. The same task also appears in My records with its SRT-generator source, so generating subtitles does not create a disconnected second history.
When transcription succeeds, the Subtitles tab opens first. Listen to the source, check each numbered cue, save deliberate edits, and download the saved SRT version. Repeated export or cue editing reuses the completed transcription rather than creating or charging for another base transcription.
- Cue numbers are emitted in order, starting at 1.
- Start and end times use HH:MM:SS,mmm with a comma before milliseconds.
- Saved Unicode cue text is used by the downloaded SRT file.
Review subtitle text and timing
An automatic transcript is raw material, not a finished caption track. Replay names, numbers, speaker changes, and fast passages, then adjust cue text or start and end seconds. WhisperWeb rejects negative, empty, reversed, or out-of-order cues rather than exporting a file that only looks valid in the browser.
Cue editing is intentionally separate from the free-form transcript editor. Changing prose does not silently retime captions; when the transcript and cue versions differ, the workspace marks the subtitles for review. Save the cue set before downloading so the file you receive matches the version shown in your account history.
See an SRT example
This short, original spoken sample and its two-cue subtitle file show the exact structure produced by the download: an integer sequence, a start-to-end time range, subtitle text, and a blank line before the next cue. The comma in 00:00:02,100 is part of SRT timing; WebVTT uses a different header and timestamp convention.
Download the sample and import it as a caption or subtitle file in your editor. Match it to a four-second practice clip on the editor timeline, confirm both cues appear in order, and inspect whether your editor interprets UTF-8 text correctly. For a real project, export from the reviewed task above so the timing stays tied to your uploaded recording.
- 1 · 00:00:00,000 → 00:00:02,100 · Welcome to the subtitle review.
- 2 · 00:00:02,400 → 00:00:04,800 · Check every cue before you publish.
Supported files and plan limits
The chooser lists the media extensions accepted by the shared transcription service, including common audio and video containers. Browser-reported type and duration are only an early check; the server makes the final decision and verifies the current plan, duration, file size, and credits. The tool shows your available balance before submission and links to current plans and usage limits rather than hard-coding an allowance that may change.
Processing is cloud based, and private media, task results, and downloads require the task owner. Do not upload confidential or regulated recordings unless your own consent, retention, and organizational requirements allow it. Generated subtitles can still contain recognition and timing errors, especially with noise, overlap, accents, music, or specialized terms, so review remains part of the workflow.
Choose the output your next editor needs
SRT is a portable timed-text deliverable, but it is not the right endpoint for every transcription job.
- Use the SRT Subtitle Generatorwhen you need numbered, timed cues that a video platform or editing application can import.
- Use the general transcription workspacewhen searchable prose, document formats, summaries, or broader transcript editing are the main deliverable.
- Use a video editorwhen you need captions rendered into the picture, custom fonts, animation, positioning, or frame-level finishing.
How SRT output is produced and checked
WhisperWeb validates the source, creates one durable transcription task, derives timed cues, and keeps those cues separate from free-form transcript edits. The server rejects invalid cue ranges and the export endpoint generates the SRT download from the owner's saved revision, using sequential indexes and comma-separated milliseconds.
Frequently asked questions
Are subtitles burned into my video?
No. This tool downloads a separate .srt subtitle file. Import that file into a compatible editor, player, or publishing platform; use a video editor if you need open captions rendered into the picture.
Can I edit timestamps before downloading?
Yes. Open the Subtitles tab, change a cue's non-negative start or end time, review its text, and save. The download is disabled while edits are unsaved, and the server validates cue order and timing before persistence.
Does generating SRT create another transcription charge?
The initial media task follows the account's current credit rules. Editing cues and downloading SRT reuse that successful task, so those actions do not start a second base transcription.
Can I reopen my subtitles later?
Yes, after a task has been created under your account. Its task ID, source page, cue revision, and output remain in My records; only the owner can reopen the private media, cues, or download.
Is processing local or cloud based?
Transcription is processed in the cloud. The page validates the source in the browser for early feedback, while the service performs its own authorization, media, plan, and usage checks before accepting a task.