
Upload your script, narration audio, and sequential images. Video Editor aligns every line to the audio with open-source speech recognition, then builds a fully editable timeline you refine and export β all running on your own machine.
Script, narration, and images enter the same automated path: alignment analysis, a verification pass, then a fully editable timeline.
Once alignment is verified, your script lines and sequential images land on an editable timeline beside the original narration. Trim, split, reorder, add text and subtitle overlays, layer extra audio tracks, and preview it all smoothly β with undo and redo always within reach.
Video Editor is designed for local or self-hosted operation using free and open-source technology end to end β from narration alignment to the final rendered MP4.
Runs on your own machine or server β no hosted account required to operate.
Local speech-to-text models such as Whisper / faster-whisper analyze narration.
Word-level timestamps derive precise start and end points for each script line.
FFmpeg or another free, open-source engine composes the final render locally.
1080p and 1440p H.264 MP4 export with the original narration audio.

Upload your script, narration audio, and sequential images. Video Editor aligns every line to the audio with open-source speech recognition, then builds a fully editable timeline you refine and export β all running on your own machine.
Script, narration, and images enter the same automated path: alignment analysis, a verification pass, then a fully editable timeline.
Once alignment is verified, your script lines and sequential images land on an editable timeline beside the original narration. Trim, split, reorder, add text and subtitle overlays, layer extra audio tracks, and preview it all smoothly β with undo and redo always within reach.
Video Editor is designed for local or self-hosted operation using free and open-source technology end to end β from narration alignment to the final rendered MP4.
Runs on your own machine or server β no hosted account required to operate.
Local speech-to-text models such as Whisper / faster-whisper analyze narration.
Word-level timestamps derive precise start and end points for each script line.
FFmpeg or another free, open-source engine composes the final render locally.
1080p and 1440p H.264 MP4 export with the original narration audio.
No comments yet. Be the first!