Video

AI Image to Video

You have a still you trust — a hero product shot, a key concept frame, a finished render — and you need it to move. Astorie takes that image into a video node, fans it out across Sora 2, Seedance 2, Kling 3, Runway Gen4, Veo, Hailuo, and Luma Ray on one canvas, and exports the take you pick straight to your NLE.

What this feature solves

You uploaded a product shot and hit "animate." One tool gave you a clean dolly that froze the label; the next moved the bottle but warped the typography; the third nailed the motion but blew the color. Three tabs, three subscriptions, no hero shot. That is the trap of single-model image-to-video tools: each platform locked itself to one engine, so you fit your shot to whatever their team picked. There is no "best i2v model" — only the best model for this shot, and a single-engine tool can never show you which one that is.

The other half of the problem is reference fidelity. A polished still has details — fabric, logo placement, talent likeness, lighting direction — that a single-prompt video model will quietly drop the moment motion starts. Without a way to pin the reference, push it into the right model for the shot, and compare takes side by side, every clip becomes a guess. Multiply that by a 30-second ad with eight cuts and the budget evaporates before you have a usable sequence.

Image-to-video is also where workflow breaks meet export breaks. You generate, you download, you transcode, you re-link in your NLE, and the codec or frame rate fights your timeline. Production teams need an animation tool that respects the reference, fans the choice out across the strongest models for the shot type, and lands a file the editor can actually cut with — without a chain of converters in between.

Why Astorie is different

Astorie puts the still inside an image node, then connects it to as many video nodes as you want — each driving a different model. Run Sora 2, Seedance 2, Kling 3, Runway Gen4, and Hailuo from the same reference in parallel and pick the take that holds the brand, the talent, or the camera move best. Where single-model platforms lock you to one engine and format tools rotate a couple of in-house models, the canvas runs the fanout natively — one source, parallel outputs, model chosen per shot, switchable mid-canvas.

Reference handling is first-class, and the canvas finishes what the generation starts. Plug the same image into multiple downstream nodes, add a prompt, set duration, and the model receives both the visual anchor and the motion brief. When a take lands, chain it forward — trim dead frames, crop to 9:16 for Reels, pipe into a follow-up shot for continuity, a lip-sync node for dialogue, or an upscale pass for delivery. Once a fanout works, save the canvas as a Recipe: the next product shot drops into the source node and the same pipeline runs with your signature treatment baked in.

Export is engineered for editors, not creators stuck in a sandbox. Astorie renders frame-rate-clean, timeline-friendly files that drop into Premiere Pro, DaVinci Resolve, and Final Cut Pro without re-encoding. The export is sequence-aware, so the eight image-to-video shots that make up your spot land as a real edit, not a folder of orphan MP4s. That is the difference between a creative toy and a production tool.

Common use cases

Animate a hero product shot for a launch ad

Take the brand-approved still, fan it across Seedance 2, Kling 3, and Runway Gen4 for camera moves, and export the winning take to your editor for the cut.

Turn a concept frame into a cinematic shot

Upload a Midjourney still, push it through Sora 2 or Veo for film-grade motion, and chain the result into a follow-up shot for continuity.

Spin still imagery into social posts

Animate a brand photo, a real-estate listing shot, or a single artwork into vertical and square formats for Reels, TikTok, and Shorts without re-shooting on set.

Bring talent stills to life for a spokesperson cut

Use a portrait reference, generate motion with Hailuo or Kling Avatar, then chain into lip-sync for dialogue without a re-shoot.

Generate b-roll from photo libraries

Send archive stills into Luma Ray for motion variants and assemble a b-roll bin without flying a crew or licensing stock footage.

Stress-test storyboard frames before greenlight

Animate every frame in a storyboard so the client sees real motion, not a stack of stills, before approving the shoot budget.

Recommended model stack

How the workflow works in Astorie

  1. 1

    1. Drop your still into an image node

    Drag a PNG or JPG onto the canvas. The image node holds the reference and feeds it downstream — keep the source resolution as high as you can, since downstream models will use it as the anchor.

  2. 2

    2. Connect a video node and pick a model

    Wire the image node into a video node. Select Seedance 2, Sora 2, Kling 3, Runway Gen4, Veo, Hailuo, or Luma Ray from the model dropdown. Each has different strengths — Seedance for fidelity, Sora for long takes, Kling for camera moves, Runway for product realism.

  3. 3

    3. Write the motion brief

    In the prompt field, describe what should move — camera path, subject action, atmosphere. Avoid restating the image; the model already sees it. Set duration and aspect ratio for the deliverable.

  4. 4

    4. Fan out across models in parallel

    Duplicate the video node, swap the model, and run them simultaneously. The same image drives every branch so you can compare takes against an identical reference, not against drift.

  5. 5

    5. Chain the winning take forward

    Connect the best clip into a follow-up shot, a lip-sync node, an upscale pass, or a sequence builder — or trim and crop to 9:16 right on the canvas. The lineage is preserved, so you can swap the source and re-render downstream automatically.

  6. 6

    6. Export to your NLE

    Use NLE export to push the sequence into Premiere Pro, DaVinci Resolve, or Final Cut Pro. Files land at clean frame rates and codecs your editor opens natively — no transcode round-trip.

Example workflow

A DTC skincare brand has a hero product photo of their new serum and needs a 6-second loop for an Instagram Reel. Drop the still into an image node. Wire it into four video nodes — Seedance 2, Kling 3, Runway Gen4, and Hailuo — each with the same prompt: "slow push-in, soft window light, droplets gently catching the rim." Run all four. Seedance holds the bottle label cleanest; Kling gives the most cinematic push; Runway keeps the typography crisp; Hailuo nails the droplet motion. Pick Seedance, trim and crop to 9:16 on the canvas, chain it into an audio node for ambient score, and export to Premiere as a vertical Reel. The runner-up takes stay on the canvas as alternates for the ad set. The whole pass — concept to NLE-ready file — fits inside one canvas and one afternoon.

Tips and common mistakes

Tips

  • Use the highest-resolution reference you have. Models down-sample, but they cannot up-rez detail that was never in the source.
  • Write motion-only prompts. Restating what the image already shows wastes tokens and confuses the model on what to change.
  • Always run two or three models in parallel for hero shots. The one that wins the camera move is rarely the one that wins the subject fidelity.
  • Lock duration before you fan out — short clips iterate faster, and most social cuts only need 3-6 seconds per shot.
  • Save the canvas as a template once a model+prompt combo works. Reuse it for the next shot in the same campaign.

Common mistakes

  • Uploading a low-res or compressed still — the video output inherits every artifact and amplifies it across 24 frames per second.
  • Writing a prompt that re-describes the image instead of the motion. The model sees the image; tell it what should change.
  • Running one model and re-rolling it ten times. Different models solve different shots — fan out instead of grinding one branch.
  • Skipping reference re-anchoring on long takes. Past 5 seconds, most models drift; chain shorter clips with the same source instead.
  • Exporting as a generic MP4 and round-tripping through HandBrake. Use NLE export so frame rate and codec match your timeline from the first import.

Related how-to guides

Related models and tools

Related features

Multi-Shot AI Video — Build Connected Scenes, Not Isolated Clips

Plan, generate, and sequence multi-shot AI video on Astorie — keep characters, style, and motion consistent across shots.

AI Video Workflow — Node-Based Production From Concept to Final Sequence

Build node-based AI video production pipelines on Astorie's canvas — from concept and storyboard to final NLE-ready sequence.

AI Character Consistency Across Images and Video

Keep a subject consistent across image and video generations on Astorie using reference workflows.

AI Video NLE Export — From Generation to Premiere, DaVinci, Final Cut

Move AI-generated sequences from Astorie into Premiere Pro, DaVinci Resolve, and Final Cut Pro.

AI Product Video Generator — From Product Image to Ad Video

Create product ads and demos from product images on Astorie's canvas — chain product photo to multi-shot video across Seedance, Runway Gen-4, and GPT Image.

AI Influencer Video Generator — Repeatable Character Pipeline

Design, generate, and scale AI influencer videos on Astorie — character library, voice cloning, lip-synced video, all on one canvas.

AI Talking Head Video — Spokesperson, Course, and Narration

Produce spokesperson, course, and narration videos on Astorie's canvas — Kling Avatar, OmniHuman, ElevenLabs, Fish Audio, locked identity end to end.

AI Video Reference Images — Preserve Subject and Style

Lock subject, character, and style across every video generation on Astorie's canvas — Vidu, Kling O3, Seedance 2, Nano Banana 2 reference workflows.

Video to Video AI — Restyle, Edit, Transform Source Footage

Restyle, transform, and edit source video on Astorie's canvas — Runway Aleph, Kling O3, Wan chained into multi-shot pipelines.

AI Video Generator — Multi-Model AI Video Production on Astorie

Multi-model AI video generation with text, image, reference, and editing workflows on Astorie's canvas.

Text to Video AI — Generate Video From Prompts on Astorie

Generate video from prompts and chain outputs into scenes on Astorie's multi-model canvas.

AI Explainer Video — Educational and B2B Demo Videos

Generate explainer videos, B2B demos, and educational content on Astorie's canvas.

AI Video Generator for Social Media

Build TikTok, Reels & YouTube content at scale with reusable Recipes and 50+ models — Sora 2, Kling 3.0 4K, Veo 3.1, Runway Gen4 — on one canvas.

Related docs

Related reading

Comparisons

Frequently asked questions

Which model should I start with for product shots?

Start with Seedance 2 for reference adherence or Runway Gen4 Turbo for clean realism that holds logos, labels, and packaging. Fan out to Kling 3 for cinematic push-ins and Luma Ray 2 when you need smoother motion — the winner is per-shot, and the canvas shows you the answer instead of making you guess.

How is this different from Runway or Krea?

Runway and Krea give you one model in one tab. Astorie lets you fan a single reference image across every major i2v model in parallel on one canvas, so you compare takes against an identical source and pick the strongest. The chosen clip then chains into follow-up shots, lip-sync, or audio without re-uploading.

Can I keep characters consistent across shots?

Yes — that is what reference-image i2v is for. Generate or import your character image once, then wire it into every i2v node on the canvas. Hailuo 2, Kling O3 reference mode, Vidu reference modes, and Seedance 2 Omni all preserve the character across shots because every shot starts from the same anchor.

How long can a single image-to-video clip be?

Durations vary per engine. As examples on Astorie today: Veo 3.1 defaults to 8-second clips with an extend option, Hailuo 2 commonly runs 6 seconds, Sora 2 Pro runs 8 to 15 seconds depending on the variant, and Sora 2 Pro Storyboard supports up to 25 seconds across multiple scenes. For longer sequences, chain clips on the canvas with the same reference so the subject stays consistent instead of drifting.

Can I animate a logo or text without distortion?

For logos and stylized graphics, Seedance 2, Runway Gen4, and Veo handle hard edges and typographic detail better than the others. Keep prompts focused on subtle camera moves rather than transforming the logo itself, and use a high-resolution source PNG rather than a flattened JPG.

Image to video vs text to video — when should I use which?

Use image to video when you have a specific look to preserve — a product, character, location, or your own uploaded photo or artwork. Use text to video for scenes from scratch. On Astorie both live on the same canvas: text-to-video the establishing shot in Sora 2, image-to-video the product close-up in Runway Gen4, all in one workflow.

Is image to video free on Astorie?

The Free tier covers basic i2v on entry-tier models — enough to try the fanout. Production-grade i2v across the flagship engines lives on the paid tiers; see /pricing for per-model durations, resolutions, and credit details.

What does it cost to run multiple models in parallel?

Each generation deducts credits from the model that runs it — fanning out costs the sum of the branches you launch. In practice, three parallel takes on a 6-second clip is faster and cheaper than re-rolling a single model six times to find the winner.

Build it on the canvas

Open Astorie and wire this workflow up in minutes. Free to start — no card required.