← All models

Bernini vs Wan

Bernini
Bernini
ByteDance

ByteDance's Bernini-R is a unified model for text-to-video and instruction-based video editing, from adding or removing objects to changing weather, background or art style. A semantic planner reads the instruction before a diffusion renderer produces the video.

Strengths

  • One interface across generation and editing
  • Strong instruction-following via a planning stage
  • Open-source 1.3B renderer weights released

Weaknesses

  • Open renderer is small, so fidelity trails larger models
  • Very new with limited independent benchmarks
  • Two-stage pipeline adds complexity
Wan
Wan
Alibaba

Alibaba's video generation family (Wan 2.x), an open-weight model series that punches above its weight class for self-hosted deployments. Strong multilingual prompt support and competitive quality make it popular for cost-conscious production.

Strengths

  • Open-weight models available for self-hosting
  • Strong multilingual and Chinese-language support
  • Competitive quality relative to inference cost

Weaknesses

  • Self-hosting requires significant compute resources
  • Cloud API availability less consistent than Western providers
  • Style range narrower than top closed-source models
See Bernini vs Wan in the full pricing comparison
Wan
Wan 3.0
Alibaba
Not Available
Not Available
Wan
Wan 3.0 Prime
Alibaba
Not Available
Not Available
Not Available
Not Available
Wan
Wan 2.7
Alibaba
Bernini
Bernini R
Not Available
Not Available
Not Available
Not Available
Wan
Wan v2.6
Alibaba
Not Available
CreditCrunch logo. a purple magnifying glass observing a gold coin

You could be overpaying by

$6.00/min

81+ models · 333+ price points · from $3.99/week