The 2026 video timeline
6 January 2026
5 models · 1 providerLTX shipped 5 video models on 6 January 2026. All 5 have a spec page. Cheapest complete job in the window: 24 credits on LTX 2. Highest output in it: 4K on LTX 2.
- New name on the roster: LTX.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| LTX 2 Retake VideoRegenerate and improve video segments with LTX 2 | LTX | 30-100 credits | — | 3s, 5s, 10s |
| LTX 2Latest video generation with enhanced control | LTX | 24-320 credits | 4K | 6s, 8s, 10s, 12s, 14s, 16s |
| LTX 2 ProProfessional image-to-video with LTXV2 | LTX | 36-240 credits | 4K | 6s, 8s, 10s |
| LTX 2 Text to Video FastFast text-to-video generation with LTXV2 | LTX | 24-320 credits | 4K | 6s, 8s, 10s, 12s, 14s, 16s |
| LTX 2 Text to Video ProProfessional text-to-video with LTXV2 | LTX | 36-240 credits | 4K | 6s, 8s, 10s |
26–30 January 2026
4 models · 2 providersPixverse and Vidu shipped 4 video models between 26 January 2026 and 30 January 2026. All 4 have a spec page. Cheapest complete job in the window: 18 credits on Vidu Q3 Image to Video. Highest output in it: 1080p on Pixverse 5.6 Image to Video.
- New name on the roster: Vidu.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| Pixverse 5.6 Image to VideoLatest Pixverse model for converting images to high-quality animated videos | Pixverse | 35-300 credits | 1080p | 5s, 8s, 10s |
| Pixverse 5.6 Text to VideoGenerate videos directly from text prompts with Pixverse 5.6 advanced capabilities | Pixverse | 35-300 credits | 1080p | 5s, 8s, 10s |
| Vidu Q3 Image to VideoTransform static images into dynamic videos with Vidu Q3 technology | Vidu | 18-116 credits | 1080p | 5s, 10s, 15s |
| Vidu Q3 VideoGenerate high-quality videos from text prompts with Vidu Q3 engine | Vidu | 35-231 credits | 1080p | 5s, 10s, 15s |
1–5 February 2026
19 models · 2 providersByteDance and Kling shipped 19 video models between 1 February 2026 and 5 February 2026. Of the 19, 9 have a spec page and 10 are folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 25 credits on DreamActor V2. Highest output in it: 4K on Kling Video V3 4K Image to Video.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| DreamActor V2AI-powered face reenactment model that transfers facial expressions and head movements from a driving video to a reference image | ByteDance | 25-600 credits | — | — |
| Kling O3 Pro Image to VideoKling O3 Pro image-to-video with advanced reasoning-enhanced generation and camera controls | Kling | 70-210 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling O3 Pro Reference to VideovariantKling O3 Pro reference-to-video for character-consistent generation using reference images | Kling | 70-210 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling O3 Pro Text to VideoKling O3 Pro text-to-video with reasoning-enhanced generation, camera controls, and audio | Kling | 70-210 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling O3 Pro Video to Video EditvariantKling O3 Pro video-to-video editing with reasoning-enhanced prompt-based modifications | Kling | 84 credits | — | — |
| Kling O3 Pro Video to Video ReferencevariantKling O3 Pro video-to-video reference-based editing for style and character transfer | Kling | 84-252 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling O3 Standard Image to VideoKling O3 Standard image-to-video with reasoning-enhanced generation | Kling | 56-168 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling O3 Standard Reference to VideoKling O3 Standard reference-to-video for character-consistent generation | Kling | 56-168 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling O3 Standard Text to VideovariantKling O3 Standard text-to-video with reasoning-enhanced generation | Kling | 56-168 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling O3 Standard Video to Video EditvariantKling O3 Standard video-to-video editing with prompt-based modifications | Kling | 63-1512 credits | — | — |
| Kling O3 Standard Video to Video ReferencevariantKling O3 Standard video-to-video reference-based editing for style transfer | Kling | 63-189 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling Video V3 4K Image to VideovariantKling V3 4K image-to-video - cinema-grade resolution with advanced camera controls, audio generation, and element composition | Kling | 210-630 credits | 4K | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling Video V3 4K Text to VideoKling V3 4K text-to-video - cinema-grade resolution with advanced camera controls, audio generation, and multi-prompt support | Kling | 210-630 credits | 4K | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling Video V3 Pro Image to VideovariantKling V3 Pro image-to-video generation with camera controls, audio generation, and element composition | Kling | 84-252 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling Video V3 Pro Motion ControlKling V3 Pro motion control - transfer motion from a driving video to a reference image at pro-tier quality | Kling | 84-2016 credits | 1080p | — |
| Kling Video V3 Pro Text to VideovariantKling V3 Pro text-to-video generation with advanced camera controls, audio generation, and multi-prompt support | Kling | 84-252 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling Video V3 Standard Image to VideoKling V3 Standard image-to-video generation with camera controls and element composition | Kling | 63-189 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling Video V3 Standard Motion ControlKling V3 Standard motion control - transfer motion from a driving video to a reference image | Kling | 63-1512 credits | 720p | — |
| Kling Video V3 Standard Text to VideovariantKling V3 Standard text-to-video generation with camera controls and multi-prompt support | Kling | 63-189 credits | — | 3s, 4s, 5s, 6s, 7s, 8s |
12 February 2026
6 models · 1 providerByteDance shipped 6 video models on 12 February 2026. Of the 6, 2 have a spec page and 4 are folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 31 credits on Seedance 2.0 Fast. Highest output in it: 4k on Seedance 2.0.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| Seedance 2.0Seedance 2.0 generates Hollywood-grade cinematic videos from text prompts with native audio-visual synchronization,… | ByteDance | 54-4096 credits | 4k | 4s, 5s, 6s, 7s, 8s, 9s |
| Seedance 2.0 FastvariantSeedance 2.0 Fast generates cinematic videos from text prompts with native audio-visual synchronization, director-level camera… | ByteDance | 31-255 credits | 720p | 4s, 5s, 6s, 7s, 8s, 9s |
| Seedance 2.0 Fast Image to VideovariantSeedance 2.0 Fast generates cinematic videos from reference images and text prompts with native audio-visual synchronization and… | ByteDance | 68-255 credits | 720p | 4s, 5s, 6s, 7s, 8s, 9s |
| Seedance 2.0 Fast Reference to VideoSeedance 2.0 Fast Reference to Video generates cinematic videos from reference images, videos, and audio inputs with native… | ByteDance | 68-255 credits | 720p | 4s, 5s, 6s, 7s, 8s, 9s |
| Seedance 2.0 Image to VideovariantSeedance 2.0 generates Hollywood-grade cinematic videos from reference images and text prompts with native audio-visual… | ByteDance | 121-454 credits | 4k | 4s, 5s, 6s, 7s, 8s, 9s |
| Seedance 2.0 Reference to VideovariantSeedance 2.0 Reference to Video generates cinematic videos from reference images, videos, and audio inputs with native… | ByteDance | 121-454 credits | 4k | 4s, 5s, 6s, 7s, 8s, 9s |
5 March 2026
5 models · 1 providerLTX shipped 5 video models on 5 March 2026. Of the 5, 3 have a spec page and 2 are folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 20 credits on LTX 2.3 Extend Video. Highest output in it: 4K on LTX 2.3 Image to Video Pro.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| LTX 2.3 Extend VideovariantExtend an existing video — the model continues motion/action past the current end (or before the current start). Takes a source… | LTX | 20-200 credits | — | 2s, 4s, 6s, 8s, 10s, 12s |
| LTX 2.3 Image to Video ProPro-quality image-to-video generation. Animates a still image with a prompt. Supports 1080p/1440p/2160p, 6-20s duration… | LTX | 36-240 credits | 4K | 6s, 8s, 10s |
| LTX 2.3 Retake VideoRetake a segment of an existing video. Takes a source video_url + prompt describing desired changes (e.g. "change flower to red… | LTX | 20-200 credits | — | 2s, 4s, 6s, 8s, 10s, 12s |
| LTX 2.3 Text to Video FastFast text-to-video generation. Generates a video directly from a text prompt at 1080p/1440p/2160p, 6-20s duration, with optional… | LTX | 24-320 credits | 4K | 6s, 8s, 10s, 12s, 14s, 16s |
| LTX 2.3 Text to Video ProvariantPro-quality text-to-video generation. Generates a video directly from a text prompt at 1080p/1440p/2160p, 6-20s duration, with… | LTX | 36-240 credits | 4K | 6s, 8s, 10s |
30 March 2026
6 models · 1 providerPixverse shipped 6 video models on 30 March 2026. Of the 6, 4 have a spec page and 2 are folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 13 credits on Pixverse V6 Image to Video. Highest output in it: 1080p on Pixverse Effects.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| Pixverse EffectsAdd special effects to videos with Pixverse | Pixverse | 15-80 credits | 1080p | 5s, 8s, 10s |
| Pixverse Image to VideovariantConvert images to videos with Pixverse V5.5 | Pixverse | 15-80 credits | 1080p | 5s, 8s, 10s |
| Pixverse Text to VideoGenerate videos from text with Pixverse V5.5 | Pixverse | 15-80 credits | 1080p | 5s, 8s, 10s |
| Pixverse TransitionCreate smooth transitions between video clips | Pixverse | 15-80 credits | 1080p | 5s, 8s, 10s |
| Pixverse V6 Image to VideovariantPixVerse V6 image-to-video. Animates a still image with strong motion and prompt control, optional audio and multi-clip, up to… | Pixverse | 13-72 credits | 1080p | 5s, 8s |
| Pixverse V6 Text to VideoPixVerse V6 text-to-video. Generates video from a text prompt with strong motion and prompt control, optional audio and… | Pixverse | 13-72 credits | 1080p | 5s, 8s |
31 March – 6 April 2026
5 models · 2 providersGoogle and Wan shipped 5 video models between 31 March 2026 and 6 April 2026. Of the 5, 4 have a spec page and 1 is folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 20 credits on VEO 3.1 Lite. Highest output in it: 4K on VEO 3.1 Lite.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| VEO 3.1 LitevariantVEO 3.1 Lite is the most cost-effective model in the VEO 3.1 family for high-volume video generation, supporting text-to-video… | 20-64 credits | 4K | 4s, 6s, 8s | |
| Wan 2.7 Image to VideoWan 2.7 animates images with three modes: first-frame to video, first-and-last-frame interpolation, or video continuation, with… | Wan | 20-225 credits | 1080p | 2s, 3s, 4s, 5s, 6s, 7s |
| Wan 2.7 Reference to VideoWan 2.7 reference-to-video — generate videos from up to 5 reference images and/or videos, with optional first frame and voice… | Wan | 24-180 credits | 1080p | 2s, 3s, 4s, 5s, 6s, 7s |
| Wan 2.7 Text to VideoWan 2.7 generates high-fidelity videos from text prompts with strong motion consistency, optional custom audio input, and… | Wan | 20-225 credits | 1080p | 2s, 3s, 4s, 5s, 6s, 7s |
| Wan 2.7 Video EditWan 2.7 video editing — modify a source video using prompts and an optional reference image for character, clothing, or style… | Wan | 24-180 credits | 1080p | 2s, 3s, 4s, 5s, 6s, 7s |
26 April 2026
4 models · 1 providerAlibaba shipped 4 video models on 26 April 2026. Of the 4, 3 have a spec page and 1 is folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 42 credits on Happy Horse 1.0 Image to Video. Highest output in it: 1080p on Happy Horse 1.0 Image to Video.
- New name on the roster: Alibaba.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| Happy Horse 1.0 Image to VideovariantHappy Horse 1.0 animates a first-frame image into video with native synchronized audio, Foley sound effects, and multilingual… | Alibaba | 42-420 credits | 1080p | 3s, 4s, 5s, 6s, 7s, 8s |
| Happy Horse 1.0 Reference to VideoHappy Horse 1.0 generates videos from up to 9 reference images using character1–character9 placeholders in the prompt, with… | Alibaba | 42-420 credits | 1080p | 3s, 4s, 5s, 6s, 7s, 8s |
| Happy Horse 1.0 Text to VideoHappy Horse 1.0 generates expressive videos from text prompts with native synchronized audio, Foley sound effects, and… | Alibaba | 42-420 credits | 1080p | 3s, 4s, 5s, 6s, 7s, 8s |
| Happy Horse 1.0 Video EditHappy Horse 1.0 edits a source video using a prompt and up to 5 reference images, with auto/origin audio handling — output capped… | Alibaba | 70-3360 credits | 1080p | — |
19 May – 1 June 2026
4 models · 2 providersByteDance and Google shipped 4 video models between 19 May 2026 and 1 June 2026. Of the 4, 1 has a spec page and 3 are folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 27 credits on Seedance 2.0 Mini. Highest output in it: 4K on Gemini Omni Video.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| Gemini Omni VideoGoogle Gemini Omni multimodal video generation. Accepts a prompt plus optional reference images, source video clips, character… | 36-90 credits | 4K | 4s, 6s, 8s, 10s | |
| Seedance 2.0 MinivariantSeedance 2.0 Mini is a faster, lower-cost tier of Seedance 2.0 — cinematic text-to-video with native audio-visual sync at high… | ByteDance | 27-227 credits | 720p | 4s, 5s, 6s, 7s, 8s, 9s |
| Seedance 2.0 Mini Image to VideovariantSeedance 2.0 Mini animates a first-frame image into cinematic video with native audio-visual sync — faster and cheaper than full… | ByteDance | 61-227 credits | 720p | 4s, 5s, 6s, 7s, 8s, 9s |
| Seedance 2.0 Mini Reference to VideovariantSeedance 2.0 Mini generates cinematic video from up to 9 reference images, 3 reference videos, and 3 reference audios… | ByteDance | 61-227 credits | 720p | 4s, 5s, 6s, 7s, 8s, 9s |
10–17 June 2026
3 models · 2 providersBria and Kling shipped 3 video models between 10 June 2026 and 17 June 2026. Of the 3, 2 have a spec page and 1 is folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 2 credits on Bria Video Background Removal. Highest output in it: 1080p on Kling 3 Turbo Image to Video.
- New name on the roster: Bria.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| Bria Video Background RemovalAI-powered video background removal and replacement | Bria | 2 credits | — | — |
| Kling 3 Turbo Image to VideovariantKling 3 Turbo fast image-to-video generation. | Kling | 28-135 credits | 1080p | 3s, 4s, 5s, 6s, 7s, 8s |
| Kling 3 Turbo Text to VideoKling 3 Turbo fast text-to-video generation. | Kling | 28-135 credits | 1080p | 3s, 4s, 5s, 6s, 7s, 8s |
23 June 2026
3 models · 1 providerAlibaba shipped 3 video models on 23 June 2026. Of the 3, 1 has a spec page and 2 are folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 42 credits on Happy Horse 1.1 Image to Video. Highest output in it: 1080p on Happy Horse 1.1 Image to Video.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| Happy Horse 1.1 Image to VideoHappy Horse 1.1 animates a first-frame image into 1080p video with synchronized native audio and multilingual lip-sync (aspect… | Alibaba | 42-270 credits | 1080p | 3s, 4s, 5s, 6s, 7s, 8s |
| Happy Horse 1.1 Reference to VideovariantHappy Horse 1.1 turns up to 9 reference images (character1–character9 placeholders) into 1080p video with synchronized native… | Alibaba | 42-270 credits | 1080p | 3s, 4s, 5s, 6s, 7s, 8s |
| Happy Horse 1.1 Text to VideovariantHappy Horse 1.1 is Alibaba's #1-ranked video model — generates 1080p video with synchronized native audio and multilingual… | Alibaba | 42-270 credits | 1080p | 3s, 4s, 5s, 6s, 7s, 8s |
30 June 2026
4 models · 1 providerGoogle shipped 4 video models on 30 June 2026. Of the 4, 1 has a spec page and 3 are folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 50 credits on Gemini Omni Flash Image to Video.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| Gemini Omni Flash EditGoogle Gemini Omni Flash - conversational video editing (video-to-video) | 63-1500 credits | — | — | |
| Gemini Omni Flash Image to VideovariantGoogle Gemini Omni Flash - animates a still image into video with audio | 50-125 credits | — | 4s, 6s, 8s, 10s | |
| Gemini Omni Flash Reference to VideovariantGoogle Gemini Omni Flash - video with audio from multimodal reference images | 50-125 credits | — | 4s, 6s, 8s, 10s | |
| Gemini Omni Flash Text to VideovariantGoogle Gemini Omni Flash - text-to-video with synchronized native audio | 50-125 credits | — | 4s, 6s, 8s, 10s |
5 August 2026
7 models · 2 providersBlack Forest Labs and MiniMax shipped 7 video models on 5 August 2026. Of the 7, 4 have a spec page and 3 are folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 43 credits on Flux 3 First Last Frame to Video. Highest output in it: 4K on Minimax H3 Image to Video.
- New names on the roster: Black Forest Labs, MiniMax.
| Model | Built by | Credits per job | Max output | Clip lengths |
|---|---|---|---|---|
| Flux 3 Extend VideoFLUX.3 continues an existing clip beyond its final frame, generating footage consistent with the original motion and scene… | Black Forest Labs | 103-410 credits | 1080p | 5s, 6s, 7s, 8s, 9s, 10s |
| Flux 3 First Last Frame to VideoFLUX.3 interpolates a smooth, coherent transition between a defined start frame and end frame, with native audio — up to 1080p… | Black Forest Labs | 43-170 credits | 1080p | 5s, 6s, 7s, 8s, 9s, 10s |
| Flux 3 Image to VideovariantFLUX.3 animates a single still image into coherent, natural motion with native audio — up to 1080p and 5–20 seconds | Black Forest Labs | 43-170 credits | 1080p | 5s, 6s, 7s, 8s, 9s, 10s |
| Flux 3 Text to VideoFLUX.3 is Black Forest Labs' frontier video model — generates video with native audio directly from a text prompt, at up to 1080p… | Black Forest Labs | 43-170 credits | 1080p | 5s, 6s, 7s, 8s, 9s, 10s |
| Minimax H3 Image to VideovariantMiniMax H3 animates a supplied image into 2K video as the opening frame, or pairs a first and last frame to control a transition… | MiniMax | 130-390 credits | 4K | 5s, 6s, 7s, 8s, 9s, 10s |
| Minimax H3 Reference to VideovariantMiniMax H3 generates 2K video from multimodal references — up to 9 images for subject and style, 3 video clips for motion, and 3… | MiniMax | 130-390 credits | 4K | 5s, 6s, 7s, 8s, 9s, 10s |
| Minimax H3 Text to VideoMiniMax H3 is a frontier video model — generates 2K video from a text prompt alone, in durations from 5 to 15 seconds across six… | MiniMax | 130-390 credits | 4K | 5s, 6s, 7s, 8s, 9s, 10s |
When 2026 was busy
7 of the twelve months carried a video release; the busiest window was 1–5 February 2026, with 19.
- January 2026
- 9
- February 2026
- 25
- March 2026
- 12
- April 2026
- 8
- May 2026
- 1
- June 2026
- 13
- August 2026
- 7
Who shipped video in 2026
Kling (20), ByteDance (10) and LTX (10) led on volume. SKUs, not quality — four tiers of one model count four times.
| Provider | Models | Spec pages | First | Latest |
|---|---|---|---|---|
| Kling | 20 | 9 | 5 February 2026 | 17 June 2026 |
| ByteDance | 10 | 3 | 1 February 2026 | 1 June 2026 |
| LTX | 10 | 8 | 6 January 2026 | 5 March 2026 |
| Pixverse | 8 | 6 | 26 January 2026 | 30 March 2026 |
| Alibaba | 7 | 4 | 26 April 2026 | 23 June 2026 |
| 6 | 3 | 31 March 2026 | 30 June 2026 | |
| Black Forest Labs | 4 | 3 | 5 August 2026 | 5 August 2026 |
| Wan | 4 | 4 | 6 April 2026 | 6 April 2026 |
| MiniMax | 3 | 1 | 5 August 2026 | 5 August 2026 |
| Vidu | 2 | 2 | 30 January 2026 | 30 January 2026 |
| Bria | 1 | 1 | 10 June 2026 | 10 June 2026 |
What 2026 moved
Firsts, measured against every video model released before them. “New name on the roster” is the provider label in the catalog, not the company — one lab can hold several labels.
- 6 January 2026
- New name on the roster: LTX.
- 26–30 January 2026
- New name on the roster: Vidu.
- 26 April 2026
- New name on the roster: Alibaba.
- 10–17 June 2026
- New name on the roster: Bria.
- 5 August 2026
- New names on the roster: Black Forest Labs, MiniMax.
Dates are each model’s released_at value, walked in order and grouped until a window held 3 or more. Credits are what one complete generation costs per the model’s own price matrix; where it states none, the headline rate is shown and the model sits out the cheapest-in-window line.
Other release years
Frequently asked questions
How many AI video models were released in 2026?+
Versely's catalog carries 75 video models with a 2026 release date, from 11 providers, arriving in 13 launch windows between 6 January 2026 and 5 August 2026.
What was the biggest AI video launch of 2026?+
1–5 February 2026, with 19 video models from ByteDance and Kling. Of the 19, 9 have a spec page and 10 are folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 25 credits on DreamActor V2. Highest output in it: 4K on Kling Video V3 4K Image to Video.
Which company released the most AI video models in 2026?+
Kling, with 20 of the 75 video models dated 2026 — first on 5 February 2026, most recently on 17 June 2026. ByteDance shipped 10, LTX shipped 10, Pixverse shipped 8.
What changed in AI video generation in 2026?+
Measured against everything the catalog carried before it: New name on the roster: LTX; New name on the roster: Vidu; New name on the roster: Alibaba.
Run any 2026 video model in Versely
The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync in a single app.