Muse Video
Muse Video is Meta’s upcoming video generation model from Meta Superintelligence Labs, previewed alongside the launch of Muse Image. The model is built on the same pretraining foundation as Muse Image and is designed to generate high-fidelity videos with native audio support. Muse Video focuses on prompt adherence, visual realism, temporal consistency, and the ability to create short scenes with clear motion, continuity, and audio context. It can generate a wide range of video styles, including cinematic footage, UGC-style ads, animal scenes, product commercials, handheld point-of-view clips, and realistic moments with sound effects, voices, and music. Meta is continuing to improve areas such as audio-video synchronization and physically accurate fast motion before broader release. Coming soon to creators and Meta AI, Muse Video is positioned as a powerful tool for generating dynamic media across Meta’s creative ecosystem.
Learn more
MiniMax H3
MiniMax H3 is a general-purpose omni-modal generation model that jointly understands multimodal contexts spanning text, images, video, and audio. It generates videos with native stereo sound at up to 2K resolution and 15 seconds in length, delivering content for advertising, branding, ecommerce, product design, UI/UX, gaming, and creative workflows. Users can combine reference types in one instruction, for example, transferring camera movement from a video, placing a character from an image into the scene, and matching vocals from an audio clip, while describing the relationships in natural language. H3 supports text-to-image, text-to-video with jointly generated audio, multi-shot modeling, text-to-audio, and generalized reference and editing across images, videos, and audio. Voice, sound effects, and music are modeled together. The model excels at instruction following, accurate text and brand presentation, and video-to-video motion transfer.
Learn more
Lyria 3 Clip
Lyria 3 Clip is a lightweight AI music generation capability within Google’s Lyria 3 ecosystem that focuses on creating short-form audio tracks from prompts. It enables users to generate brief music clips, typically around 30 seconds, using text, images, or video inputs. The model transforms creative ideas into complete soundtracks with vocals, lyrics, and instrumentals automatically. It is designed for fast, iterative creation, allowing users to experiment with different styles, moods, and genres. Lyria 3 Clip is integrated into platforms like the Gemini app and developer tools, making it accessible for both creators and developers. The tool emphasizes ease of use, requiring no musical expertise to produce polished audio outputs. Overall, it provides a quick and intuitive way to generate short, high-quality music clips for creative projects.
Learn more
AzurBeat
AzurBeat is an AI music generation platform that turns text prompts into complete, original tracks – full vocals, instrumentation, and song structure – in seconds. Beyond audio, AzurBeat generates matching AI cover art and one-click music videos, so a single idea becomes a release-ready package. Flexible tiers scale from free experimentation to a full Label plan with commercial-grade output, mastering, and video generation. Designed for content creators, indie musicians, and marketers who need original music without a studio. If you're comparing AI music tools, AzurBeat is a fast, affordable alternative to Suno and Udio, with production and video features built in.
Learn more