Three separate video workflow improvements shipped this month. Each one solves a specific friction point for different parts of the video generation and editing flow.
Auto-Subtitles
Upload any video to the video editing tools and Kolbo can transcribe it and burn styled subtitles directly into the file - automatically, without any manual timing or captioning work.
The subtitles are synced to the speech in the video. Style options control the appearance: font weight, size, placement, and color. The output is a single video file with the subtitles baked in and ready to post.
Useful for social content, educational clips, and any video where you want captions without the usual SRT-syncing workflow.
Active-Speaker Selection for Lipsync
When you run a lipsync generation with multiple people visible on screen, Kolbo now lets you pick which person gets the lip-sync treatment.
A visual selector appears over the video frame. Click the face you want to target. The generation runs on that person specifically, leaving the other faces unchanged. Before this update, the model would pick the most prominent face automatically - which was often the right person, but not always.
This matters for interview-style footage, multi-character scenes, and any clip where you know exactly which speaker should be animated.
Video Trim on Upload
Most video generation models have a maximum input length. If your source clip is longer than the model accepts, the generation would previously fail outright - leaving you to manually trim the video in a separate tool and re-upload.
Now, when you upload a video that is too long for the selected model, Kolbo surfaces a trim control directly in the upload flow. Set the in and out points for the range you want, and only that segment is sent to the model. No external editing step needed.
The same trim control also works proactively: even if your video is within limits, you can trim it to the specific range you want before the generation runs.


