MimicMotion and Florence-2: One to All workflow tutorial

Learn to rewire pose tracking with MimicMotion, keep DW OpenPose as a fallback, and add Florence-2 captions plus a shared 30 FPS control.

Player not loading? Watch on YouTube

This tutorial updates the One to All extended AI video workflow to address pose mismatches, motion blur and sparse manual prompts. The speaker replaces its animation preprocessor with MimicMotion, which extracts poses from the source video relative to the reference image's proportions. The resized reference image feeds the reference output directly.

For the reference pose, the tutorial provides two paths. MimicMotion takes the reference image in both inputs, but the speaker notes that it can miss a person in static frames. The DW OpenPose fallback uses InspireNet RemBg and image compositing to isolate and position the subject. Its square output passes through Resize Image V2 with the intended width and height and a bottom crop. A switch and group muter let the user choose either path.

A shared float control changes FPS inputs from 15 to 30. Florence-2 supplies a reference-image caption; text replacement changes "painting" or "picture" to "video", and added text specifies dancing. The speaker attributes clearer motion and richer clothing detail to these changes. The tutorial shows a comparison after these workflow changes.