Upload Two Photos
Add one clear, front-facing photo for the maiden and one for the little man. Even lighting and one person per frame is all the template needs.
Upload two photos and render the spinning-room clip with both faces swapped.
Upload one photo of the little man, then render the spinning-room clip with his face swapped.
Your photo replaces the dancing man while the straw room, the candlelight and the camera move stay exactly as they are.
No prompt writing, no filming, and no timeline editing. The whole workflow stays on this page.
Add one clear, front-facing photo for the maiden and one for the little man. Even lighting and one person per frame is all the template needs.
The spinning room is a fixed scene, so the only choices left are the output ratio and the resolution. Both are priced before you submit.
Confirm the credits, submit the paid task, preview the finished clip, and download the file you want to keep.
The template is a single continuous shot. A woman sits in the straw while a small man dances in front of her, and both of them are replaced by your uploads. The room, the bales, the candlelit grade and the camera move all stay intact.
The maiden and the little man are replaced together, so the two halves of the joke land on the same beat instead of being stitched from separate renders.
The template audio stays attached during generation, so the tiptoe rhythm lands on the same beat as the source clip.
Widescreen matches the source clip. Other ratios stay available when the render needs to travel further.
The selected ratio controls the output canvas while The Spinning Room stays the motion reference.
One source template and one finished swap. Pick either clip to watch it, then scroll back up to render your own.
The product keeps one clear goal in view: turn two photos into the spinning-room clip without a camera crew or a costume department.
A dedicated Rumpelstiltskin AI generator removes the usual setup work around prompts, camera instructions, timing and wardrobe. You provide two photos and the built-in scene provides everything else: the low candlelight, the stacked bales, the dance, and the exact moment the little man turns toward the camera.
Keeping the scene fixed is the important difference. A general video tool asks you to describe a scene and hope the model interprets it well. This workflow keeps the scene and the camera move fixed, and narrows your decisions to the two uploads and the ratio the clip needs, which makes the result predictable and easy to repeat across a batch of posts.
The generator is also a simple way to join the trend. Instead of recreating the room, the dance, and the timing yourself, upload two photos and the shared template places both people in the scene. The same workflow covers a one-off meme, a short reaction clip, or a batch of variations, and the 5-video and 15-video packs cover several Pro renders.
A front-facing photo with even lighting gives the system reliable information about each face. Keep one person in frame per upload and avoid heavy filters, sunglasses, or anything covering the face.
The straw room, the candlelight, the wide framing and the dance are already prepared, so you never write a prompt or direct a timeline. The only creative decision is who stands in for each figure.
Render widescreen for YouTube and embedded posts, a vertical frame for short-form feeds, or a square canvas for mixed timelines. The ratio is fixed before the task begins.
The cost is shown before submission. A 720P Standard video uses 100 credits, and providers that fail to complete a task are handled by the refund flow already built into the product.
The best results usually come from two focused photos, the correct output frame, and a resolution that matches the final channel.
Use a recent, clear, well-lit portrait with one person in frame for each subject. A face that is easy to read produces the most consistent identity across the clip.
The seated woman and the dancing figure are separate slots, so a single photo of two people will not work. Upload one portrait per slot, even when the two people are the same person.
480P is the cheapest way to test a pairing. 720P Standard is the practical default for finished social content, and 1080P Pro costs more credits for higher-detail results.
Rumpelstiltskin AI turns the well-known spinning-room scene into a reusable video template. Upload one portrait for the maiden and one for the little man, then render the clip with both identities swapped while the straw room, candlelight, camera movement, and original tiptoe audio stay in place. The result works as a short-form meme, a reaction clip, or green-screen footage for your own edit. Because the scene is fixed and paid credits are shown before submission, there is no prompt to write and no timeline to rebuild.
Background on the meme, a full walkthrough, and the complete pricing breakdown.
Common questions about the generator, photo requirements, credits, and the original clip.
Need help with credits, failed generations, or a paid order? Contact support.
Upload two photos, then choose a subscription or a one-time credit pack.