Model profile

Vidu AI video generator

ShengShu Technology's video model family for native audio-video storytelling, image-to-video, and scene-oriented short-form production.

Decision Snapshot

Score: 79

Score source: Vidu Q3 Pro is represented in Arena snapshot among top I2V systems, and official Q3 pages document native audio-video positioning.. Evidence confidence is high.

Best For

  • Native Audio-Video Stories
  • Short Drama Clips
  • Narrative Ads
  • Multi-Speaker Scenes
  • Comic And Manga-Style Video

Supported Workflows

  • T2v
  • I2v
  • Native Audio
  • Multi Speaker
  • Scene Cuts
  • Industry Templates

Access And Company

Built by ShengShu Technology. Public access is listed as Closed Web, Closed Api, Aggregator.

Entry Points

Score Breakdown

Weighted evidence dimensions separate model quality from access, web authority, ecosystem, and source confidence.

DimensionScoreWeightEvidenceConfidence
Model Quality Proxy 89 24% Vidu Q3 Pro is represented in Arena snapshot among top I2V systems, and official Q3 pages document native audio-video positioning. high
Missing: live Artificial Analysis rank/Elo by modality, live Arena AI rank/score by modality
Capability Depth 78 16% Supported workflow tags: t2v, i2v, native_audio, multi_speaker, scene_cuts, industry_templates. Depth rewards T2V/I2V/V2V, references, native audio, editing, API, and open weights where present. high
Missing: hands-on feature verification, mode-specific limits by duration/resolution
Access And Pricing 88 12% Access types: closed_web, closed_api, aggregator. Pricing note: Consumer and API pricing varies by model variant and generation mode. The platform model map should be paired with the current pricing page before publicat... high
Missing: current free tier, normalized price per video minute
Version Maturity 81 12% 4 version/history entries in the profile; latest public date is 2026. high
Missing: automated release-note monitor, model ID/version mapping across providers
Ecosystem Popularity 68 10% GitHub stars are not applicable to this closed or research-only product unless an official public model repository is confirmed.

No public source link is attached yet.

medium
Missing: GitHub: n/a for closed-source model
Web Authority 59 10% SEO snapshot found product domain vidu.com and company domain genspi.com; sampled sitemap URL count is 0. medium
Missing: Similarweb/API traffic, Tranco rank
Company Distribution 62 10% ShengShu Technology distribution profile: Vidu matters for users who want a China-origin video generator with reference-to-video, sound-and-vision generation, marketing tools, and a clear roadmap from creative media int... medium
Missing: app store ratings/review volume, product MAU/traffic
Source Confidence 100 6% Profile source confidence is confirmed with 2 official URL(s), 4 feature evidence item(s), SEO present, GitHub n/a. high
Missing: paid/API evidence refresh, manual output review artifacts

What It Is

Primary content is rendered into static HTML.

Vidu is a Chinese AI video generation product from ShengShu Technology, known for Q-series releases that focus on quality, consistency, and narrative generation. The Q3 generation shifts the product toward native audio-video output and story-first use cases.

For consumers, Vidu is especially relevant when the desired output is a complete short scene rather than a silent visual clip. Q3 can generate audio, dialogue, voiceover, sound effects, and music together with the video, and it exposes variants for drama, advertising, balanced general use, and faster modes.

Vidu should be tested on dialogue timing, multi-speaker scenes, scene cuts, and promptable rhythm. It is also worth comparing against Kling, Seedance, and Veo for native audio and character consistency.

Key Features

  • Native audio-video outputVidu Q3 generates visuals, dialogue, voiceover, sound effects, and music together.Source
  • 16-second clipsThe Q3 product page says users can generate up to 16 seconds in one generation.Source
  • Q3 model variantsThe model map lists Q3 Pro, Mix, Drama, Ad, and Turbo variants for different quality and scenario needs.Source
  • Multilingual outputVidu Q3 supports English, Japanese, and Chinese output according to the product page.Source

Version Progress

  • 2026-04-13 / current_confirmed / confirmedVidu Q3 Reference-to-VideoExpanded reference-based generation, synchronized audio-video, visual effects, scene composition, and consistency controls. Vidu Q3 supports up to 16 seconds of synchronized audio and video generation.
  • 2026-04-13 / Vidu Platform Docs / confirmedVidu Q3 reference-to-video modelsReference-to-video adds viduq3-mix, viduq3-turbo, and viduq3 models with audio-video generation support.
  • 2026 / Vidu / confirmedVidu Q3Native audio-video direct output. Up to 16-second generation. Q3 Pro, Mix, Drama, Ad, and Turbo variants.
  • 2025-01-15 / historical / confirmedVidu 2.0Vidu 2.0 update announced as an earlier flagship video generation release.
  • 2024 / Vidu Platform Docs / partialVidu 1.x / Vidu 2.0Early public Vidu product and model generations; Vidu2.0 remains listed in the platform model map.

Latest News

Model-specific news entries link directly to the original publisher; article bodies stay off-site.

Limitations And Pricing

Consumer and API pricing varies by model variant and generation mode. The platform model map should be paired with the current pricing page before publication.

  • Vidu S1 Needs A Dedicated First-Party Profile Before It Is Merged Into The Current Product Page.
  • Exact API Prices And Regional Availability Are Not Yet Normalized.