Source
Audio, Video · Required
Return to the catalogue and choose an available model.
Back to the catalogueTranscribe multilingual audio locally with the open Whisper model.
Writes down what is said in a clip, word by word with a time on each, and fills the clip with the captions. Runs on this device.
The model defines every slot it reads and every result it returns. Required inputs must be present before a run can begin.
Prompts and source material
Audio, Video · Required
Files and structured results
Metadata
| Control | Available values | Requirement |
|---|---|---|
| Language | en, zh, de, es, ru, ko, fr, ja, pt, tr, pl, ca, nl, ar, sv, it, id, hi, fi, vi, he, uk, el, ms, cs, ro, da, hu, ta, no, th, ur, hr, bg, lt, la, mi, ml, cy, sk, te, fa, lv, bn, sr, az, sl, kn, et, mk, br, eu, is, hy, ne, mn, bs, kk, sq, sw, gl, mr, pa, si, km, sn, yo, so, af, oc, ka, be, tg, sd, gu, am, yi, lo, uz, fo, ht, ps, tk, nn, mt, sa, lb, my, bo, tl, mg, as, tt, haw, ln, ha, ba, jw, su | Optional |
Works especially well for
Know the trade-offs
The generation returns to the scene, asset folder and timeline it belongs to.