Skip to content
CutConvert
Post-production file converter

Convert JSON to SRT subtitles

Turn timed transcripts or YouTube chat replay JSON/JSONL into SRT. Preview the detected captions, or map text and timestamp fields in a custom JSON export.

Convert your files

Drop files here

browseWhisper or transcript JSONConverted in your browser. Files never leave your device.About an hour of footage is free. Larger files: $3 each, or Pro $9/month for everything up to 50 MB.

up to 2 files · 2 MB each

free to start · project to project: $5 · every other conversion: $3 · Pro $9/month · pricing

Converting as a guest. A free account has no monthly cap on files within the free size limit.

Queue
0/2
No files queued yet — drop something above to begin.

Cómo convertir JSON a SRT

Convierte JSON a SRT en línea. CutConvert lee transcripciones de Whisper — verbose_json de OpenAI, salida de whisper.cpp — y JSON genérico de segmentos con tiempos, y escribe subtítulos SRT limpios y numerados con los hablantes preservados.

  1. Añade tus archivos JSON

    Arrastra tus archivos .json de Whisper o de transcripción a la zona de carga, o haz clic para buscarlos.

  2. Convierte JSON a SRT

    Pulsa Convertir. Cada segmento con tiempos se convierte en un cue SRT con precisión de milisegundos; los campos de hablante se convierten en etiquetas.

  3. Descarga tu resultado

    Descarga el archivo convertido al instante — o un único archivo ZIP cuando conviertes un lote de archivos.

JSON de transcripción y Subtítulos SRT, explicados

La transcripción con IA produce JSON; los reproductores y editores quieren SRT. Este convertidor acepta las estructuras que existen realmente en la práctica: verbose_json de OpenAI Whisper (segmentos con inicio/fin en segundos), whisper.cpp (arrays de transcripción con desplazamientos en milisegundos) y arrays genéricos de segmentos {start, end, text}. Los tiempos se leen del archivo — nunca se estiman.

Convierte verbose_json de Whisper, la salida de whisper.cpp o cualquier JSON de transcripción con tiempos en subtítulos SRT limpios. Es gratis para empezar, va cifrado en tránsito y convierte un lote completo en un solo ZIP; inicia sesión cuando necesites trabajos grandes o de gran volumen.

Preguntas frecuentes sobre JSON a SRT

¿Qué formatos JSON se admiten?

verbose_json de OpenAI Whisper (un array de segmentos con inicio/fin en segundos), la salida de whisper.cpp (un array de transcripción con desplazamientos o marcas de tiempo) y arrays genéricos de objetos {start, end, text} — con campos de hablante opcionales, en segundos o milisegundos.

¿Cómo obtengo verbose_json de Whisper?

Con la API de OpenAI, establece response_format en verbose_json. Si ejecutas Whisper en local, la salida .json por defecto funciona; el archivo generado con --output-json de whisper.cpp también.

Mi JSON solo tiene un campo de texto plano — ¿se puede convertir?

No. SRT necesita tiempos de inicio y fin por segmento, y una transcripción de texto sin más no los tiene. Vuelve a exportar con tiempos por segmento (verbose_json en el caso de Whisper) y se convertirá.

¿Se conservan las etiquetas de hablante de la diarización?

Sí — los segmentos que llevan un campo speaker salen como cues SRT etiquetados, de modo que las transcripciones diarizadas mantienen su atribución.

¿Y si mi archivo es un JSON de transcripción de Premiere Pro?

Usa el convertidor de transcripción de Premiere a SRT en /prtranscript/srt — lee directamente el formato de transcripción de Premiere (incluidos los archivos .prtranscript binarios).

Guías para este flujo de trabajo

Why editors trust it

Free to start

Convert everyday batches for free. Create an account when you need large, high-volume jobs.

Batch in, one ZIP out

Convert a batch within your allowance and download a single ZIP, or one file if that is all you need.

Made for editors

Whole timelines from Premiere, Final Cut and Avid with transitions, markers, titles, keyframes, levels, nests and bins; captions that keep their italics, colour, placement and speakers across SRT, VTT, ASS, TTML, iTT, SCC and STL, any of them in, any of them out; transcripts into Premiere with the words the engine timed. Frame rates and timecode handled the way post expects.

Private by design

Files are converted in your browser. Queued files can be saved locally to restore your job after checkout. Conversion metadata helps us diagnose errors; file contents are not uploaded.

Questions, answered

Where are my files processed?

In your browser. The converter runs on your own machine and the file never leaves it; the site only records that a conversion happened. CutConvert never sees, stores, sells or shares your media.

Is it free? Do I need an account?

About an hour of footage per file is free, on every converter: 100 KB of subtitles or captions, 500 KB of Word transcript, 2 MB of transcript JSON, 1 MB of Resolve project, 5 MB of Final Cut XML, with Premiere projects and Avid AAF a little under the hour. Guests get five conversions a month, a free account has no monthly cap within those sizes, and above that project-to-project transfers are $5 per file. Every other conversion is $3, including project files to captions, reports or cut lists, and captions or transcripts into an editor. Pro is $9 a month ($90 a year) for every file type up to 50 MB; Studio is $19 a month ($190 a year) for files up to 500 MB. Paid-file access lasts 24 hours.

What do I get when I convert multiple files?

Drop several files and you get back one ZIP archive containing every converted file. Convert a single file and you get that one file back, no ZIP. Tick extra formats under the output chips and the same ZIP carries every one of them, and a file counts as one conversion whatever it is written to.

Which formats are supported?

Timelines: Premiere .prproj, DaVinci Resolve .drp (as exported, or gzip- or Zstandard-wrapped), Avid AAF and Final Cut FCPXML in; FCP7 XML, FCPXML, Resolve .drp, CMX 3600 EDL and shot-list CSV out. A Premiere or Final Cut edit travels with its transitions, markers, titles, opacity and scale keyframes, clip volume, speed changes, nested sequences, bins, audio channel routing and each file's own frame rate, wherever the output has a place for them; an Avid AAF brings its transitions, markers, clip levels with their keyframes, muted tracks, nested stacks and each clip's own rate. Captions and transcripts, read and written: SRT, WebVTT, SBV, ASS, TTML, DFXP, iTT, SCC, MCC (a second 608 channel comes out as its own file), EBU-STL, Avid caption TXT, CapCut drafts of any size, timed CSV, Word, plain text, Whisper, WhisperX, AssemblyAI, Deepgram, Rev.ai and CapCut JSON, TwitchDownloader chat logs as a chat replay, Premiere transcript JSON and .prtranscript, Rev, Otter and Word-style transcripts, with italics, colour, placement and speakers kept where the output can hold them and measured word timing into Premiere. EDLs become event CSV reports. The capability table lists what every pair keeps.

Can I pick a frame rate?

Yes. For EDL and Avid caption output you can choose 23.976, 24, 25, 29.97, or 30 fps, or let CutConvert auto-detect it from your source.