Assignment 1 · 15% of the course grade
From Images to Film
A Generative AI Workflow
Make a 30–60-second video with sound. Use generative AI to create images, turn them into video clips, and create at least one audio element. Edit these into a finished piece that tells a story, expresses an idea, or creates a mood.
Due
Plan Generate images Animate Add sound Edit
Need a starting point? Explore workflow tools and creation websites ↓ · Workflow figures · Technical resources
Follow these five steps
Plan your idea
Choose a simple idea, such as a short story, fictional advertisement, or visual experiment. Describe who it is for and what you want viewers to feel. Sketch or describe 4–6 shots in order.
Generate your images
Create a starting image for each shot. Refine the images so that the style, characters, and setting fit together. Save your key prompts and selected images.
Turn images into video
Use image-to-video generation to animate your images. Decide what should move in each shot, including the subject or camera. Select the clips that best support your idea.
Add sound
Generate at least one audio element: music, sound effects, background sound, or narration. Match the sound to the action or mood. Dialogue is optional.
Edit and export
Arrange your clips, trim the timing, and balance the sound. Add titles or captions if helpful. Export a 30–60-second MP4 with audio and watch it once from beginning to end.
Save your progress: Keep two before-and-after examples as you work. These can show a change to an image, motion, sound, or edit. For each, explain the problem, your change, and the result.
Submit two files
- 1. Final video (MP4)
30–60 seconds, with audio. - 2. Process report (PDF, 6–10 slides)
Include all the items below in this one report.
- Idea and storyboard: Your concept, intended audience, and plan for 4–6 shots.
- Workflow: A simple diagram showing how you went from images to video and audio. Include tool names, key images, and selected prompts.
- Two improvements: Before-and-after examples with a short explanation of what changed and why. Use screenshots or still frames; describe sound changes in words.
- Reflection (200–300 words): What decisions did you make? What was difficult? What did you learn?
- Credits: State what AI generated, what you edited yourself, and where any external images, music, or other assets came from.
| What we look for | Weight |
|---|---|
| Idea: A clear story, message, or mood. | 25% |
| Visuals: Consistent images, purposeful motion, and clear editing. | 25% |
| Sound: Audio supports the video and is balanced. | 15% |
| Process: Explain your workflow, improvements, and learning. | 25% |
| Documentation: Complete report and clear credits. | 10% |
Suggested workflow tools and websites
Choose one workspace to organize your images, video, and sound. The options below link to actual creation tools, with a suggested route for this assignment. A node is simply one step in a visual workflow, such as generating an image or animating it.
Where to start: For Chinese-language tools, try RunningHub for ComfyUI workflows, TapNow for a creative canvas, Jimeng for a simple image-to-video route, or LiblibAI for image workflows. International options are listed below too. You only need one route.
Chinese-language platforms
RunningHubRun and adapt ComfyUI workflows online
Open RunningHub ↗ · rhTV canvas and templates ↗
- Browse ComfyUI workflows or AI applications. Start with an image-generation template and make your shot images.
- Choose an image-to-video workflow, upload one selected image, and describe the movement. Run a short test before repeating it for the other shots.
- Use a supported audio workflow or generate audio separately. Download the clips and sound, then assemble your final edit.
Search terms to try: 文生图, 图生视频, 分镜, 音效. Save a screenshot of the workflow and the settings you changed.
LiblibAI / 哩布哩布Find image models and reusable community workflows
Open LiblibAI ↗ · Image-editing workflow tutorial (中文) ↗
- Find an image-generation or editing workflow. Choose one that explicitly supports online running.
- Open “在ComfyUI中运行” where available, upload your reference image, and edit the prompt. Create a consistent set of shot images.
- Save the images and animate them in Jimeng, Vidu, RunningHub, or another accessible image-to-video tool. Add sound and finish the edit.
Some community workflows require local installation or additional models. Read the workflow description before choosing it.
Jimeng / 即梦Generate images, refine them on a canvas, and animate them
Open Jimeng ↗ · Official reference and audio-video examples (中文) ↗
- Use image generation to create your shot images. Use the smart canvas to refine composition or edit selected areas.
- Open video generation, upload a starting image, and describe the subject and camera movement. Use an ending frame if that mode is available and useful.
- Review each clip. Keep generated sound if suitable, or add generated audio separately. Assemble the shots in your video editor.
Keep the character description, visual style, and aspect ratio consistent across shots.
TapNowConnect text, images, video, and audio on one canvas
Open TapNow ↗ · Chinese user guide ↗ · Nodes and connections ↗
- Create a canvas or copy a template. Add your idea and references as nodes.
- Generate a shot image, then connect it as a reference for a video node. Repeat for your other shots.
- Generate audio using an available audio node. Download the selected outputs, assemble the final sequence, and save a screenshot of your canvas.
A connection passes a reference to the next step. Check the selected model’s supported inputs before running it.
More workflow platforms
Freepik SpacesConnect image, video, and audio steps in your browser
Open Spaces ↗ · Official getting-started guide ↗
- Create a Space or start from a template. Add your idea and reference images.
- Connect a prompt to an image-generation node. Review the image, then feed your chosen result into a video-generation node.
- Repeat for your planned shots. Add an audio step, such as voiceover, where available.
- Export your clips and audio, then assemble the final 30–60-second video in your editor. Save a screenshot of the connected steps.
LTX Studio / FlowsDevelop a storyboard, generate shots, and edit a sequence
Open LTX Studio ↗ · Build your first Flow ↗
- Start with your concept or script and develop a 4–6-shot storyboard.
- Generate and refine the shot images, then animate the selected images. Reuse character and setting references across shots.
- Arrange the clips in the timeline and work on pacing and sound. If a feature is unavailable on your plan, finish that step in your usual editor.
- For a visible workflow diagram, explore Flows: connect prompts, image generation, and video generation, and save a screenshot for your report.
ComfyUI / Comfy CloudBuild or adapt a reusable node workflow
Open Comfy Cloud ↗ · ComfyUI and workflow gallery ↗ · Beginner tutorial ↗
- Open Comfy Cloud, or install ComfyUI on a suitable computer. Start with an image-generation template.
- Generate and save your shot images. Load an image-to-video template and use your chosen image as its input.
- Run one shot, inspect the result, and adjust the prompt or settings. Generate sound with a supported workflow or add it separately.
- Save the workflow and selected outputs. Assemble the clips and sound in your video editor.
Choose templates supported by your setup. Local workflows need suitable hardware and model files; Cloud and hosted Partner Nodes may require credits. Partner Node requirements ↗
Figma Weave (formerly Weavy)Combine generation and visual editing on a node canvas
Explore Weave and workflow examples ↗ · Open the workspace ↗
- Start from a workflow example and add your reference image or prompt.
- Generate image variations and refine the selected image with editing steps such as cropping, masking, or relighting.
- Use an available video-generation step to animate it, or export the image to one of the video websites below.
- Repeat for your shots, then add generated audio and finish the sequence in a video editor. Capture the workflow for your report.
Other creation websitesUseful when you prefer a guided interface for individual steps
- Lovart ↗ — Develop visual ideas, edit images, and generate video in a design workspace. Suggested route: brief → shot images → revisions → animation → final edit.
- Hailuo AI / 海螺 ↗ — MiniMax’s video creation website. Upload your generated image, describe the motion, and review the clip. Explore its Video Agent workflow if available in your account.
- Vidu ↗ — Use Image to Video or Reference to Video to animate your shot images. Download your selected clips and assemble them with sound.
For every route: Keep your generated images, save two before-and-after examples, and show the tools and steps in your report. Audio may be generated with the video or separately.
Check access before starting. Features, supported models, export options, and credits vary by account and region. These tools are optional; use an accessible alternative if needed. A more expensive platform does not earn higher marks.
See how the pieces fit together
Use the generated clip and its audio.
Generate music, effects, or narration and align it in the edit.
A boat rests in a rainy alley.
Sound: Rain + distant street ambienceIt drifts forward; the camera follows.
Sound: Water + light rainA warm doorway appears ahead.
Sound: Water + a soft musical cueThe boat settles beside the doorstep.
Sound: Rain fades; music resolvesThis example totals 40 seconds. Keep the same boat, palette, and lighting; change the framing and action. You may use any subject or style.
Technical resources · use what you need
These are optional references. Start with the guide for your chosen tool; you do not need to read everything.
- ComfyUI: your first image ↗Load a template, run it, save an image, and export the workflow.
- ComfyUI: image-to-video with Wan 2.2 ↗Workflow examples, model requirements, input images, and settings. Choose a workflow supported by your machine or cloud service.
- LiblibAI: extend an image (中文) ↗Learn outpainting when a shot needs more space around its subject.
- TapNow: connect nodes (中文) ↗Understand how text, image, video, and audio references move between steps.
- TapNow: image, video, and audio guides (中文) ↗In “在画布中制作,” choose the media type you want to generate or edit.
- Seedance in ComfyUI: reference workflows ↗Examples for using references and first/last frames, including audio-video generation.
Quick tips for your first experimentSmall changes make results easier to understand
- Image prompt: Describe the subject, setting, composition, lighting, and style.
- Motion prompt: Describe what moves and how the camera moves. Start with one clear action.
- Consistency: Reuse references and keep the aspect ratio the same across shots.
- Comparison: Change one thing at a time. Keep the model and available settings fixed; if seeds cannot be fixed, note that randomness can affect the result.
- Sound: Listen to the full edit. Keep narration clear and avoid abrupt changes in volume.
Keep it simple. Use any accessible tools. Coding and paid subscriptions are not required. We value thoughtful choices and what you learn; expensive tools or high resolution do not automatically earn higher marks.