Text-to-Video

Text-to-video is a category of AI models that generate video sequences directly from a text prompt, including motion, scene composition, and, in some cases, audio. Examples of this category include Sora, Runway, and Google Veo, which enable rapid prototyping of video content without the need for a camera, actors, or expensive production.