The complete research: every model, in depth
Three kinds of evidence: [UI] read from the AutoHDR beta interface, [WEBINAR] said on tape by Matt or JT (timecodes point into the full webinar video above), [WEB] public sources listed at the end.
Google: Veo 3.1 family and Gemini Omni Flash
In our UI: Veo 3.1, Veo 3.1 First-Last, Veo 3.1 Fast, Veo 3.1 Fast First-Last, Gemini Omni Flash Video, Gemini Omni Flash 1.1. Resolutions 720p / 1080p / 4K. [UI]
- [UI] Veo 3.1: "High-quality, cinematic videos with realistic motion." First-Last adds "precise first and last frame control." Fast is "at half the cost per second." Observed cost: 8 credits for a 4s 1080p generation.
- [WEB] Launched October 2025 by Google DeepMind: native audio, up to 4K output, first/last-frame control on both standard and Fast tiers. Public pricing runs roughly 0.35 to 0.40 dollars per second for standard and about 0.10 per second for Fast.
- [WEBINAR 00:12:59] Matt on Google's older video models: mostly replaceable by Kling. Omni Flash is the keeper: "gemini-omni-1.1-flash"-class, strong at editing existing footage and reference-to-video, 4K output, scenes extendable to about 40 seconds (the 1.1 entry in our UI adds the end frame and 4K tier).
- [WEBINAR 00:12:04] JT: "I honestly don't use Omni hardly at all. Most of my workflow is between H3, Seedance, and Kling." Use Omni for quick drafts and video edits, not hero shots.
Kuaishou: Kling 3.0 Pro and Turbo
In our UI: Kling 3.0 Pro ("Frontier Kling, with an optional end frame and clip length") and Kling 3.0 Turbo. 3 credits per generation. [UI]
- [WEBINAR 00:08:56] JT: "a trustworthy workhorse that I would use no matter what. It's had the best track record."
- [WEBINAR 00:11:09] Matt: cheap and fast with exact start/end adherence ("always starts on that start frame, always ends on that end frame"), but "the movements look a bit more fake, a little bit more AI."
- [WEBINAR 00:28:10] JT: movement can read clay-like ("Kling kind of feels like a Crayola..."), while other models put more realistic life into people in frame.
- [WEBINAR 01:01:08] Matt's rule of thumb: "Kling really shines when you give a start and end frame... quick, cheap transformations" (day-to-night, snow-to-spring). In the day-to-night shootout [01:00:41], Kling "didn't even fully get to nighttime" while H3 Max nailed it.
- [WEB] Kling 3.0 released February 2026: 3 to 15 second clips, first and last frame on both tiers, an All-in-One Reference mode (3-8s character video), claims of native 4K/60fps, and Turbo roughly 20x faster than standard.
MiniMax: the H3 family
In our UI: MiniMax H3 ("Image-to-video with an optional end frame"), H3 Max ("Sharper MiniMax H3, at 480P and 768P"), H3 Reference, and H3 Max Reference. [UI]
- [WEBINAR 00:55:47] Matt breaking down the family: H3 is the base, Max is the better-quality tier ("the most natural-looking is H3 Max" per JT [00:57:10], whose base H3 can have a stop-motion feel).
- [WEBINAR 01:00:25] The day-to-night shootout went to H3 Max; Matt: "the movement is so much more natural" (Kling "is not trained on, clearly, like, actual enough real video").
- [WEBINAR 01:02:49] JT on construction and build sequences: "H3 is obviously the most construction-like" (the drywall and tape-lines detail held up).
- [WEBINAR 00:09:27] Matt on cost: top models ran him about 20 dollars for high-res generations; H3 is the cheaper, faster tier.
- [WEB] H3 is natively multimodal with synced stereo audio, 5-15 second clips, 768p and 2K tiers, keyframe and reference inputs. Our Reference variants add reference-image conditioning.
ByteDance: Seedance 2.5
In our UI: Seedance 2.5 ("Next-generation image-to-video, up to 30 seconds") and the Reference variant for longer, reference-driven clips. [UI]
- [WEBINAR 00:12:04] Part of JT's daily trio (H3, Seedance, Kling).
- [WEBINAR 01:01:08] Matt: use H3 and Seedance for movement shots, Kling for the exact A-to-B transforms.
- [WEB] Released July 2026: one-take 30-second audio-video clips, roughly 30 to 50 multimodal references (images, video, audio), 1080p to 4K output, region editing, identity/portrait reference support.
Alibaba: Wan 3.0 and Wan 3.0 Prime
In our UI: Wan 3.0 ("Clips up to 30 seconds, with an optional end frame") and Wan 3.0 Prime ("tuned for identity and scene continuity"). [UI]
- [WEB] Announced August 2026: 2 to 30 second single-pass generation with native audio and multimodal reference inputs (turning even text-heavy material into video); Prime is the premium tier tuned for identity and scene continuity, strongest in reference-to-video workflows.
- No on-tape verdict from Matt or JT in our webinars yet. Try it where a shot needs to hold a face or a room's identity across a longer clip, and report back through the feedback form.
Luma AI: Ray 3.2
In our UI: Luma Ray 3.2 ("Cinematic five-second motion from a start frame"). 540p / 720p / 1080p, every aspect ratio. 5 credits. [UI]
- [WEBINAR 00:24:29] JT: "I like the movement from Luma, but it's a little shaky... Kling 3 is nice and cinematic and stable."
- [WEBINAR 00:11:28] Matt: "for advanced users only," but it can reference exact keyframes along the timeline, which no other model here does.
- [WEBINAR 00:37:11] Matt's production use: re-cropping real horizontal footage into vertical without hallucinating new pixels.
- [WEB] 5 or 10 second clips, aspect ratios from 9:16 up, start and end frame interpolation, and a Reframe feature for changing aspect after generation.
Adjacent tools from the webinar (not in our picker)
- Astra (GPT with vision) [WEBINAR 00:31:39]: JT's orchestration assistant; it reads footage frame by frame and writes the prompts. His biggest cheat code [01:39:23]: "have AI read the documentation... You just focus on being creative."
- Astra 2 upscaler [WEBINAR 01:39:01]: Matt calls it "the best upscaler for video I've found right now."
- Topaz [WEBINAR 01:38:32]: JT: longest time in market, enterprise-grade, movement-preserving upscaling.
- Nano Banana (Google image model) [WEBINAR 01:43:01]: for wild re-angles on the photo side; AutoHDR's own models are fine-tuned on real estate to not hallucinate.
- LTX 2.5 [WEBINAR 00:07:24]: named as a recent major release; not in our picker yet.
Sources
Google: Veo 3.1 announcement, Gemini API pricing, DeepMind.
Kling: kling.ai, Pexo on 3.0 Turbo.
MiniMax: minimax.io, Hailuo.
Seedance: ByteDance Seed announcement, fal.ai.
Wan: Alibaba Cloud, Runway.
Luma: lumalabs.ai, Scenario.
Webinar quotes: the full JT and Matt conversation embedded at the top of this page.