Stable Diffusion
Stable Diffusion is Stability AI’s professional image generation model family built for creating high-quality visuals from text prompts. The models support a wide range of styles, including photography, 3D, painting, illustration, line art, and other creative formats. Stable Diffusion is designed for strong prompt adherence, diverse visual outputs, and flexible use across professional, creative, and technical workflows. Users can deploy the models through self-hosted licensing, the Stability AI API, cloud partner ecosystems, or web-based creative applications. Stability AI also provides image editing tools for inpainting, outpainting, object removal, upscaling, sketch control, structure control, and style transformation. Built for creators, developers, brands, and enterprises, Stable Diffusion helps teams generate, edit, customize, and scale visual content production.
Learn more
Stable Diffusion 3.5
Stable Diffusion 3.5 is Stability AI’s image generation and editing model suite, built for professional-grade creative production across self-hosted deployment, API integration, cloud partner ecosystems, and web-based creation. Its flagship Stable Diffusion 3.5 family is described as Stability AI’s most powerful image model yet, designed to generate a wide range of image styles, including 3D, photography, painting, line art, and more, with market-leading prompt adherence, diverse outputs, and flexible options for different use cases. Stable Diffusion 3.5 Large is the most powerful model in the Stable Diffusion family, with superior quality and prompt adherence for professional use cases at 1 megapixel resolution. Stable Diffusion 3.5 Large Turbo is designed to run faster than Large while generating high-quality images with exceptional prompt adherence in just four steps. Stable Diffusion 3.5 Medium balances quality and customization with improved architecture and training methods.
Learn more
Grok Imagine
Grok Imagine is an AI-powered creative platform designed to generate both images and videos from simple text prompts. Built within the Grok AI ecosystem, it enables users to transform ideas into high-quality visual and motion content in seconds. Grok Imagine supports a wide range of creative use cases, including concept art, short-form videos, marketing visuals, and social media content. The platform leverages advanced generative AI models to interpret prompts with strong visual consistency and stylistic control across images and video outputs. Users can experiment with different styles, scenes, and compositions without traditional design or video editing tools. Its intuitive interface makes visual and video creation accessible to both technical and non-technical users. Grok Imagine helps creators move from imagination to polished visual content faster than ever.
Learn more
Sora
Sora is an AI model that can create realistic and imaginative scenes from text instructions.
We’re teaching AI to understand and simulate the physical world in motion, with the goal of training models that help people solve problems that require real-world interaction.
Introducing Sora, our text-to-video model. Sora can generate videos up to a minute long while maintaining visual quality and adherence to the user’s prompt.
Sora is able to generate complex scenes with multiple characters, specific types of motion, and accurate details of the subject and background. The model understands not only what the user has asked for in the prompt, but also how those things exist in the physical world.
Learn more