Source: raw/Minimax_H3_is_a_Local_AI_Video_BEAST_for_Anyone_Everyone..md (MattVidPro AI, hands-on local test) and raw/Seedance_2.5_The_New_Age_of_AI_Video_w_Theoretically_Media.md (AI For Humans, with guest Tim Simmons / Theoretically Media). Both auto-caption transcripts fetched 2026-08-07.

A cluster of video-model releases landed together in early August 2026 — Seedance 2.5, MiniMax/Hailuo H3, and Wan 3.0 — and the wiki had no coverage of any of them. The one that changes the most is H3: a fully open-weights video model with native audio that the community immediately got running on consumer hardware far below the official spec. That moves local video generation from “possible on a workstation” to “possible on a laptop.”

Both sources are creator-channel hands-on reactions, not benchmarks. Every number below is a single operator’s report.

MiniMax / Hailuo H3 — the open-weights one

The headline is the hardware floor collapsing.

ClaimDetail
Official specRTX 5090 or an RTX 6000-series workstation card
Community low-endA user reported running it on a 2060 with 6GB VRAM — “nearly usable footage,” visibly mushy
Mac path4-bit quantized build + Diff Studio → runs on Apple M-series with 8GB VRAM
Pinocchio installerNVIDIA/CUDA only
Output720p, with native audio generated
DistributionDownloadable via ModelScope and Hugging Face

The gap between the official floor (a $2,000+ GPU) and the reported floor (a six-year-old 6GB card, or an M-series Mac) is the substantive finding. Treat the 2060 result as a single relayed Twitter report, not a verified benchmark — but the Mac 4-bit path is described from the presenter’s own setup.

Native audio at 720p. The presenter’s assessment of the generated audio is measured — “it does not sound bad,” with orchestral music swelling in the background — and he is explicit that the resolution is not 1080p. Native audio in an open-weights model is the notable part, not the fidelity.

The architecture appears to be omni. It reportedly accepts multiple reference types — not only images but up to 15 seconds of video — allowing extension while preserving character consistency and voice. The presenter notes this “is very much the Seedance vibe,” i.e. the reference-conditioning approach Seedance popularized, now in an open model.

It ships uncensored. The transcript reports copyrighted characters generating directly (Dragon Ball, Mickey Mouse are named) and “reports of nudity and gore generating right out of the model as it comes open source.” The presenter also flags likeness generation of real people as something viewers should not do. This is a meaningful operational difference from every hosted model in this topic — the safety layer is not in the weights, so it is entirely on the operator. For any commercial workflow, that is a legal exposure question, not a feature.

A workflow note worth stealing: rather than following a README by hand, the presenter used Codex to do the local setup — “the days of downloading everything from Hugging Face or GitHub, setting it all up manually, carefully following a readme are over.” Same pattern as the agent-assisted install flows documented in Forage’s install prompt. ComfyUI was run headless in the background.

Seedance 2.5 — hosted, longer, and priced accordingly

From the AI For Humans discussion with a working AI filmmaker:

  • Generate at 720. The guest’s direct advice: “you don’t want to be trying to do 4K or even 1080 really.” Around 10 seconds at 720 is the practical sweet spot.
  • Up to 15 seconds per the documentation.
  • 30-second output exists and is expensive — reported at ~1,100 credits per generation. Context for that number: the guest is in ByteDance’s Dreamina creative-partners program receiving roughly 15,000–20,000 credits/month, so a 30-second clip is on the order of 5–7% of a partner-tier monthly allowance.
  • The 30-second length is genuinely useful for dialogue work — the guest’s film is “very two guys talking at a bar,” a shape that short clip limits break.
  • An API is expected shortly, with rumors of “aggressively saving” pricing. The guest’s read: ByteDance is aware MiniMax is now competitive, “so I expect the price to go down… I do expect there to be crazy sales left and right.”

The competitive dynamic is the useful takeaway: an open-weights competitor appearing is what is expected to move hosted pricing, and the API launch is the event to watch.

The “all Chinese” framing, corrected

The AI For Humans episode opens on the question “nearly all of these models are Chinese — does that mean America has given up on the AI video frontier?” That framing is broadly right for the video models and wrong for one item in its own list:

  • Seedance 2.5 — ByteDance (China)
  • MiniMax / Hailuo H3 — MiniMax (China)
  • Wan 3.0 — Alibaba (China)
  • Flux 3Black Forest Labs, which is German, and is an image model rather than a video model. See Flux 3.

So the accurate version is narrower and still striking: the three video models in this wave are all Chinese, and the open-weights one is too. ^[the correction is this wiki’s, not the source’s — the episode groups Flux 3 into the same sentence]

Wan 3.0 is named in both sources and described in neither. It is recorded here as a pointer only.

Key Takeaways

  • The local floor for video generation dropped sharply. Official spec is an RTX 5090; community reports put H3 on a 6GB 2060 and on 8GB Apple M-series via a 4-bit build. If you have any modern machine, local video generation is now worth testing.
  • H3 is open weights with native audio at 720p — the first model in this topic combining all three.
  • Open weights means no safety layer. Copyrighted characters, likenesses, and NSFW output generate directly. That is an operator liability, and it disqualifies the model from client work without a filtering layer of your own.
  • Seedance 2.5’s practical envelope is ~10 seconds at 720, not 4K, with a 30-second option that costs roughly 5–7% of a partner-tier monthly credit allowance.
  • Watch for the Seedance API launch — the pricing signal for the whole hosted tier, and expected to be pressured downward by H3’s existence.
  • Reference conditioning is converging. H3 reportedly accepts up to 15s of video as a reference for extension with character and voice consistency — the Seedance-style approach, now in an open model.
  • These are creator reactions, not evaluations. No side-by-side, no fixed prompt set, no repeated trials. Directional only.

Try It

  1. Check your actual hardware floor before believing the spec sheet. If you have an NVIDIA card, try the Pinocchio installer (CUDA only). On Apple silicon, the 4-bit quantized H3 + Diff Studio path is the one reported working at 8GB.
  2. Delegate the local setup to an agent. The presenter used Codex to handle the download-and-configure step end to end; the same works with Claude Code.
  3. Generate at 720, not 1080 or 4K, on both models — this is the one piece of advice both sources agree on.
  4. Before any client use of H3, decide your content-filtering story. The model has none.
  5. If you need 30-second single-take dialogue, Seedance 2.5 is currently the option — budget ~1,100 credits per generation and wait for the API pricing.

Open Questions

  • No license is named for H3 anywhere in the source. “Fully open” and “open source” are used loosely; whether that means Apache-2.0, a custom weights license, or something with commercial restrictions is unverified and matters a great deal for any professional use. Confirm on the model card before use.
  • No parameter count, VRAM figure at full precision, or generation-time benchmark is given for H3.
  • Wan 3.0 is entirely uncovered — named in both sources, described in neither.
  • The 2060/6GB result is a relayed third-party claim seen on Twitter, not reproduced by the presenter.
  • Seedance 2.5 credit pricing is partner-program-relative. 1,100 credits has no stated dollar conversion, and retail pricing may differ from the Dreamina partner tier.
  • No quality comparison between H3 and Seedance 2.5 exists in either source — the open-vs-hosted trade-off is asserted on availability grounds, not measured output quality.