swap_horiz Make-A-Video Alternatives
Looking for alternatives to Make-A-Video? Compare the top Model options ranked by our AI scoring system.
Make-A-Video
Make-A-Video is a text-to-video generation artificial intelligence system developed by Meta AI and introduced in 2022. The system works by extending the capabilities of text-to-image generation models, learning how the visual world changes over time to synthesize novel videos. By leveraging existing...
apps Top Make-A-Video Alternatives
The top alternative to Make-A-Video in 2026 is Stable Diffusion 1.5 with a score of 9.2/10, followed by Whisper (9.1) and SAM (9.1).
Stable Diffusion 1.5
Stable Diffusion 1.5 is an open-weight text-to-image artificial intelligence model released in October 2022 by Stability...
Whisper
Whisper is an automatic speech recognition model released by OpenAI in September 2022 as open-source software. It was tr...
SAM
SAM (Segment Anything Model) is a vision foundation model developed by Meta AI and released in 2023. It was trained on t...
Chinchilla
Chinchilla is a large language model developed by DeepMind and detailed in a 2022 research paper. It features 70 billion...
Llama 3.1 405B
Llama 3.1 405B is a large language model released by Meta in 2024, serving as the flagship of the Llama 3.1 collection....
Sora
Sora is a text-to-video generative artificial intelligence model developed by OpenAI and announced in February 2024. The...
DINOv2
DINOv2 is a self-supervised vision foundation model developed by Meta AI and released in 2023. It was trained on a highl...
SAM 2
SAM 2 (Segment Anything Model 2) is an artificial intelligence model developed by Meta and released in 2024. It extends...
Midjourney v4
Midjourney v4 is a version of the Midjourney text-to-image artificial intelligence model released in 2022. It represente...
InstructGPT
InstructGPT is a family of large language models introduced by OpenAI in 2022, designed to align artificial intelligence...
Llama 3.3
Llama 3.3 is an instruction-tuned large language model developed by Meta and released in December 2024. Built with 70 bi...
Wan 2.1
Wan 2.1 is a text-to-video generation model developed by Alibaba and released as open-source in early 2025. The model ut...
Imagen
Imagen is a text-to-image diffusion model introduced by Google Research in 2022. The model is distinguished by its archi...
Flamingo
Flamingo is a multimodal visual language model introduced by DeepMind in 2022. The architecture is designed to process a...
Kling 1.5
Kling 1.5 is an artificial intelligence video generation model developed by the Chinese technology company Kuaishou, rel...
Runway Gen-2
Runway Gen-2 is an AI model released by the company Runway in 2023, designed for text-to-video and image-to-video genera...
Emu Video
Emu Video is a text-to-video generation model introduced by Meta in 2023. It utilizes a factorized, or cascaded, diffusi...
VideoPoet
VideoPoet is a large language model developed by Google Research and presented in 2023 as a system for zero-shot video g...
Stable Diffusion 2.1
Stable Diffusion 2.1 is an open-weight text-to-image latent diffusion model released by Stability AI in late 2022. As a...
OPT
The Open Pre-trained Transformers (OPT) are a suite of large language models released by Meta AI in 2022. The suite incl...
summarize Quick Comparison Summary
See all Model ranked by score
emoji_events View Full Model Rankings