True First and Last Frame Control
Both ends of the shot are locked to your images. The AI generates only the motion between them, so the clip starts and ends exactly where you decided.
Upload a starting image and an ending image, describe what happens in between, and AI generates the full transition as a real video clip. No editing skills needed.
Veo 3.1 Lite · 8s · about 24 credits. Your first clip is free.
Sample video. Your result will vary based on the style, voice, and settings you choose.
No editing skills. No complex software. Just describe what you want.
Add your first frame (where the clip starts) and your last frame (where it ends). Matching aspect ratios and subject placement give the smoothest results.
Write what happens during the transition: the morph, the camera move, the lighting shift. The AI uses your description to plan the motion between both frames.
Pick a length (4, 6, or 8 seconds) and aspect ratio, then generate. Download the MP4 with no watermark, or drop the clip into a full AITuber video.
Professional tools, zero learning curve.
Both ends of the shot are locked to your images. The AI generates only the motion between them, so the clip starts and ends exactly where you decided.
Frames mode runs on models built for dual-frame input, including Veo 3.1 and its Fast and Lite variants. Use Lite for cheap drafts and full Veo 3.1 for final quality.
Renovation reveals, fitness progress, makeup transformations, aging morphs. Any two photos of the same subject become one continuous transformation shot.
Start on the closed box, end on the hero shot. Start on the front view, end on the back. The AI generates the reveal or rotation between your two product photos.
Generate 4, 6, or 8 second clips in 9:16 for Shorts and TikTok, 16:9 for YouTube, or 1:1 for feeds. Pick the frame that matches where the clip will live.
Every clip downloads as a standard MP4 with no watermark. Use it in ads, social posts, client work, or as raw footage in any editor.
Veo 3.1 Lite starts at 3 credits per second, Fast at 10, and full Veo 3.1 at 20; generating sound with the clip adds a little more per second. The cost shows before you generate, and signup includes 50 free credits with no card.
Any clip can be used as a scene inside a complete AITuber video, with AI voiceover, word-synced captions, and direct publishing to YouTube and TikTok.
First frame to last frame AI video is a technique where you give the AI two images, a starting frame and an ending frame, and it generates the motion between them as a real video clip. Instead of hoping a text prompt lands where you want, you lock both ends of the shot and let the model fill in the journey: a morph, a transformation, a reveal, or a camera move that connects point A to point B.
AITuber's clip studio has a dedicated Frames mode for exactly this. Upload your first frame, upload your last frame, describe what happens in between, and generate. It runs on models that support both frames, like Google's Veo 3.1, plus its Fast and Lite variants for cheaper drafts. Clips come out at 4, 6, or 8 seconds in 9:16, 16:9, or 1:1, exported as clean MP4 with no watermark.
The format shines for before and after content: renovations, fitness progress, makeup transformations, product reveals, character turnarounds, day-to-night scene shifts, aging morphs, and logo transforms. Signup includes 50 free credits with no card required, and Veo 3.1 Lite starts at 3 credits per second, so an 8-second first-to-last-frame clip costs 24 credits with sound off or 40 with sound. Your first clip is covered free either way.
Use two images with the same aspect ratio and keep the subject in roughly the same spot in both. Big jumps in framing force the AI to invent camera moves you did not ask for.
The AI can already see both images. Spend your prompt on what happens between them: "the paint peels away to reveal polished wood" beats re-describing the start and end.
A 4-second clip gives the model less room to drift, so morphs and transformations stay tighter. Save 8 seconds for slow reveals and camera moves that need the time.
At 3 credits per second, Veo 3.1 Lite is the cheap way to test whether your two frames connect well. Once the motion looks right, rerun the same setup on full Veo 3.1.
If your first frame is daylight and your last frame is candlelight, make that shift part of the prompt ("the light fades to warm candlelight") so the transition feels intentional.
It is a generation technique where you give the AI a starting image and an ending image, and it creates the video motion between them. You control both ends of the shot; the AI fills in the morph, transformation, or camera move that connects them.
Yes to start. Signup includes 50 free credits with no card required. Veo 3.1 Lite starts at 3 credits per second, so an 8-second first-to-last-frame clip costs 24 credits with sound off or 40 with sound, which the free credits cover.
Open the clip studio and switch to Frames mode. Upload your first frame, upload your last frame, describe what happens in between, pick a length and aspect ratio, and generate. The finished clip downloads as an MP4.
They work best when they match. Use the same aspect ratio for both frames and keep the subject in a similar position. Mismatched sizes get cropped or padded, which can make the transition feel like a jump instead of a smooth move.
AITuber's Frames mode runs on models built for dual-frame input, like Google's Veo 3.1, plus its Fast and Lite variants. Lite is the cheapest, starting at 3 credits per second; Fast starts at 10 and full Veo 3.1 at 20.
You can generate 4, 6, or 8 second clips. Shorter clips tend to morph more cleanly, while 8 seconds suits slow reveals and bigger camera moves. For longer sequences, generate several clips and chain them inside a full AITuber video.
The AI will still connect them, but the transition gets more surreal the further apart they are. A photo of a dog and a photo of a skyscraper will produce a dramatic morph. For realistic results, keep the subject, framing, and lighting related; for creative effects, big differences can be the point.
Yes, that is one of the most popular uses. Renovation reveals, fitness progress, makeup transformations, and product glow-ups all work: upload the "before" photo as the first frame, the "after" photo as the last frame, and describe the change happening.
Yes. Use the clip as a scene inside a full AITuber video, which adds AI voiceover, word-synced captions, and music, then publish straight to YouTube or TikTok. The raw clip export itself is video only.
No. Clips export as clean MP4 files with no watermark, in 9:16, 16:9, or 1:1. They are ready to post directly or to drop into any video editor as footage.
Describe the transition, not the images. The AI already sees both frames, so tell it how to get from one to the other: "the walls repaint themselves and furniture fades in" or "the camera pushes in as day turns to night." Specific motion verbs give the best results.
Yes. Upload any two images you have the rights to use: phone photos, product shots, AI-generated images, sketches, or renders. Many creators generate both frames with an AI image tool first, then animate between them here.
Create videos for other popular niches
Join 43,699+ creators using AITuber to make professional first & last frame videos with AI.
No credit card required