Add precise timings to transcript documents
Send a known transcript and its recording into a Notis-triggered alignment workflow, then use the timings in your Docmosis package.
Trigger
Webhook received
Notis starts this workflow when an external tool or custom backend sends an HTTP request.
Action
Forced Alignment
Align a known transcript to its audio recording, returning precise word-level timestamps (subtitles, audiobook timings, karaoke).
Why this helps
Matching transcript text to exact audio moments by hand is tedious and makes caption preparation easy to postpone.
- Returns precise word-level timestamps for a supplied transcript and recording.
- Helps prepare caption and audiobook timing documents.
- Avoids manually scrubbing audio to locate every phrase.
Setup
Build it in a few focused steps.
- 1Connect Docmosis and ElevenLabs to Notis once in the portal.
- 2Create an automation by asking Notis or using Automations -> New Automation.
- 3Write one plain-language instruction asking Notis to align the supplied transcript to its recording and make the returned timings available to the Docmosis package.
- 4Choose an incoming webhook trigger and a channel for run reports.
- 5Test with a real transcript and its matching audio.
Questions about this workflow
Does this action create a transcript?
No. Forced Alignment uses a known transcript and aligns it to the matching audio.
What can I use the result for?
The word-level timings can support subtitle, audiobook or karaoke timing documents.
When this happens · Trigger
Do this · Action
Supported Triggers and Actions
Notis builds workflows that link Docmosis to ElevenLabs. A trigger fires from one place; an action lands in another.
Docmosis triggers
ElevenLabs actions
Recurring trigger
Notis starts this workflow on a schedule, such as daily, weekly, or during business hours.
Voice Isolator
Remove background noise, music and ambient sounds from a recording, returning clean studio-quality speech (MP3, base64).
Webhook trigger
Notis starts this workflow when an external tool or custom backend sends an HTTP request.
Voice Changer
Transform a recording into a different ElevenLabs voice while preserving the original emotion, timing and delivery (MP3, base64 output).
Forced Alignment
Align a known transcript to its audio recording, returning precise word-level timestamps (subtitles, audiobook timings, karaoke).
List Voices
List the ElevenLabs voices available to this account (voice_id, name, category, labels, preview_url).
Speech to Text
Transcribe an audio or video file from a URL with ElevenLabs Scribe — 90+ languages, word timestamps, speaker diarization, audio-event tagging, optional entity detection and keyterm biasing.
Music Generation
Generate studio-grade music (MP3) in any style from a natural language prompt with Eleven Music — vocals or instrumental, commercial-use cleared.
Connect any two apps with Notis in the middle.
Docmosis and ElevenLabs, or any other combination from 1,000+ integrations.
When this happens · Trigger
Do this · Action
Save your first hour today.
7-day trial of any paid plan, with 20$ of usage included.
No card. Works with personal or business ElevenLabs.