LLM Models
Explore and compare the latest large language models from leading AI companies.
179 LLM Models found
Sora
OpenAI
DISCONTINUED. OpenAI's original text-to-video diffusion model, announced February 2024 (public release December 2024 as 'Sora Turbo'). Generated videos up to one minute long using a transformer architecture that represents video as patches. Superseded by Sora 2; the entire Sora product is being discontinued (app/web ended April 26, 2026; API ends September 24, 2026).
Seedance 2.0
ByteDance
ByteDance's video generation model (2026), powering the Dreamina/Doubao platforms. Seedance 2.0 accepts quad-modal inputs (text, image, video, audio) and features a 'Universal Reference' system that locks composition, camera movement, and character actions across shots for consistent multi-shot sequences.
Hunyuan Video
Tencent
Tencent's open-weight text-to-video model (December 2024), a 13B-parameter model released with public weights that brought open video generation close to closed-source quality at launch. It produces high-motion, coherent clips and anchors a growing open video ecosystem.
Wan 2.2
Alibaba Cloud
Alibaba's open-weight cinematic video model (2025), supporting text-to-video and image-to-video at 720p and 1080p with various aspect ratios. Wan 2.x weights are downloadable from Hugging Face under Apache 2.0, making it a leading open option for self-hosted video generation.
MiniMax Hailuo 02
MiniMax
MiniMax's Hailuo video model, known for fast render speed and surprisingly high quality, including a capable free tier. It delivers natural motion and decent prompt adherence for short-form content, making it popular for rapid ideation and experimental clips.
Luma Ray 3
Luma AI
Luma AI's current-generation video model (Ray3), regarded as one of the best image-to-video tools in 2026 and strong for cinematic mood and environment-heavy shots. It supersedes the earlier Dream Machine naming and targets high-fidelity, physically coherent motion.
Runway Gen-4
Runway
Runway's professional video generation model (March 2025), the pro favorite for granular creative control. Gen-4 emphasizes reference-driven character and world consistency, camera moves, and motion brush, making it a production pick for teams that need editing control alongside model quality.
Kling 3.0
Kuaishou
Kuaishou's flagship AI video model, consistently ranked among the best overall for cinematic quality in 2026. Kling v3 improved temporal consistency so objects and people keep their shape across a clip, generating up to 10 seconds at 1080p with complex camera movement, from text and image prompts.
Pika 1.0
Pika Labs
Pika Labs' AI video generation platform enabling text-to-video, image-to-video, and video-to-video transformations. Released November 2023 with $55M funding. Generates 3-second clips in seconds to tens of seconds, outputs up to 1080p. Supports diverse styles: 3D animation, anime, cartoon, cinematic. Features: custom aspect ratios (square, portrait, widescreen for Instagram/TikTok), camera controls (pan, zoom, rotate), AI upscaling, video outpainting/inpainting for seamless editing. Web and Discord platform access. Free downloads, no watermark. Represents major milestone in accessible AI video generation.
Ideogram 2.0
Ideogram
Ideogram's groundbreaking text-to-image model with exceptional text rendering accuracy. Released August 2024, trained from scratch to significantly outperform competitors on image-text alignment, subjective preference, and text rendering. Achieves 85-90% text rendering accuracy vs competitors producing garbled text (DALL-E, Midjourney). Five style options: regular, realistic (photo-quality), design (premium graphics), 3D, anime. Features color palette control for brand consistency. Human evaluations consistently rate it as significant improvement over Flux Pro and DALL-E 3. iOS app and API (beta) available. API pricing lower than DALL-E 3, on par with Flux Pro.
Reve 2.0
Reve AI
Reve AI's June 2026 image model, introducing a planning-then-rendering architecture that builds code-based intermediate layouts for native 4K output, lossless iterative editing, and precise composition control. Reve's typography engine (trained on 50M font samples) targets ~98% in-image text accuracy.
Qwen-Image
Alibaba Cloud
Alibaba's open-source 20B-parameter image foundation model (August 2025), an MMDiT architecture that ranks first across nine public benchmarks including GenEval, DPG, and OneIG-Bench. It leads on complex text rendering (Chinese and English at commercial quality) and precise image editing.