Stitch-a-Demo: Creating Video Demonstrations from Multistep Descriptions
When obtaining visual illustrations from text descriptions, today's methods take a description with a single text context--a caption, or an action description--and retrieve or generate the matching visual context. However, prior work does not permit visual illustration of multistep descriptions, e.g