The transcribe tool already exports as text, subtitles, PDF, plain text, Markdown, SRT, and VTT — that covers most of what I need. The one format I'd love to see added is a word-level JSON export, where every single word has its own start/end timestamp, rather than grouping multiple words or sentences into one timestamp block (like SRT/VTT do).
Adobe Premiere Pro already supports this kind of export. Having it in Vowel would help in two main cases:
YouTube chapter/timestamp generation. Feeding an AI the current subtitle export gives it far fewer timestamp anchors, since each block often spans multiple sentences. A word-level JSON gives the AI enough granularity to place chapter markers exactly when a topic starts — instead of chapters that start or end mid-sentence because the underlying block was too coarse.
Motion graphics / animation timing. When converting a design into a motion graphic to sync with a video, having every word's exact timestamp means I know precisely when to trigger each animation, instead of estimating from a block-level subtitle.
This is achievable today by exporting from Premiere Pro instead, but since I'm already using Vowel's transcription for the whole session, having this JSON export available directly from Vowel would save a full extra step.
Please authenticate to join the conversation.
In Review
Feature Request
14 days ago

Dio
Get notified by email when there are changes.
In Review
Feature Request
14 days ago

Dio
Get notified by email when there are changes.