Eleven v4 Turbo, Gemini 3.8 Flash Live, Kling 3.0, Muse Video, Suno v5, Higgsfield, Nano Banana 2.
| Rank | Release Date | Model & Lab | Category | Architecture | Score | Action |
|---|---|---|---|---|---|---|
| #69 | Sep 24, 2026 |
Eleven v4 Turbo Low-Latency Voice Engine
ElevenLabs
|
Multimodal, Video & Speech | Streaming Neural Audio Codec & Acoustic Transformer | 99.0 | Try β |
| #70 | Sep 23, 2026 |
Gemini 3.8 Flash Live Voice-to-Voice Stream
Google DeepMind
|
Multimodal, Video & Speech | Multimodal Streaming Speech-to-Speech Transformer | 98.5 | Try β |
| #71 | Jul 19, 2026 |
Muse Video Physical Consistency Engine
Meta Superintelligence Labs
|
Multimodal, Video & Speech | 3D Spatial-Temporal Diffusion Transformer | 97.8 | Try β |
| #72 | Aug 28, 2026 |
Kling 3.0 Cinematic Multi-Shot Video
Kuaishou
|
Multimodal, Video & Speech | Diffusion Transformer with 3D VAE Latent Compression | 97.3 | Try β |
| #73 | Aug 30, 2026 |
Suno v5 Multi-Track Stem Music Engine
Suno AI
|
Multimodal, Video & Speech | Multi-Track Acoustic Diffusion Transformer | 96.8 | Try β |
| #74 | Sep 16, 2026 |
Higgsfield Cinematic Ads Video Engine
Higgsfield AI
|
Multimodal, Video & Speech | Motion-Conditioned Video Diffusion Transformer | 96.3 | Try β |
| #75 | Sep 19, 2026 |
Nano Banana 2 Hyper-Aesthetic Visual Diffusion
Midjourney Research
|
Multimodal, Video & Speech | Rectified Flow High-Density Latent Diffusion | 95.8 | Try β |
| #76 | Sep 21, 2026 |
ChatGPT Image 2.0 In-Context Generation
OpenAI
|
Multimodal, Video & Speech | Autoregressive Multimodal Token Diffusion | 95.5 | Try β |
| #77 | Sep 25, 2026 |
KlingAI Pro Video Production Suite
Kling AI
|
Multimodal, Video & Speech | Video Generation Studio & Workflow Runner | 95.0 | Try β |
| #78 | Sep 5, 2026 |
Udio v2.0 4-Minute Full Song Synthesis
Udio Music
|
Multimodal, Video & Speech | Autoregressive Audio Codec & Diffusion Refiner | 94.5 | Try β |
| #79 | Sep 12, 2026 |
FLUX.1 Pro Ultra 2K Resolution Flow
Black Forest Labs
|
Multimodal, Video & Speech | Multimodal Diffusion Transformer (MMDiT) | 94.1 | Try β |
| #80 | Aug 18, 2026 |
Runway Gen-4 Alpha Interactive Physics
Runway
|
Multimodal, Video & Speech | World Simulation Spatial Diffusion Transformer | 93.5 | Try β |
| #81 | Sep 9, 2026 |
Pika 2.5 Physics Effects Video Generator
Pika Labs
|
Multimodal, Video & Speech | Material-Aware Video Diffusion Model | 93.0 | Try β |
| #82 | Sep 20, 2026 |
Whisper large-v3-turbo Multilingual Speech
OpenAI / Hugging Face
|
Multimodal, Video & Speech | Encoder-Decoder Speech Transformer (4 decoder layers) | 92.6 | Try β |
| #83 | Sep 14, 2026 |
Cartesia Sonic 2.0 Streaming Voice
Cartesia AI
|
Multimodal, Video & Speech | State Space Acoustic Streaming Model | 92.1 | Try β |
| #84 | Sep 17, 2026 |
Mistral Pixtral Large 124B Multimodal
Mistral AI
|
Multimodal, Video & Speech | Multimodal Decoder Transformer with 400M Vision Encoder | 91.7 | Try β |