Wan2.2: Open and Advanced Large-Scale Video Generative Model
Infinite Canvas Workbench for AI creation integrates AI generation
Stable Diffusion web UI
Official MiniMax Model Context Protocol (MCP) server
Open-source image generative foundation model
Guiding Instruction-based Image Editing via Multimodal Large Language
Collection of Gemma 3 variants that are trained for performance
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Implementation of Imagen, Google's Text-to-Image Neural Network
Text and image to video generation: CogVideoX and CogVideo
Multimodal-Driven Architecture for Customized Video Generation
JavaScript OCR and text extraction for images and PDFs
Stable Diffusion web UI
Easily compute clip embeddings and build a clip retrieval system
Readest is a modern, feature-rich ebook reader
Generating Immersive, Explorable, and Interactive 3D Worlds
Contexts Optical Compression
Capable of understanding text, audio, vision, video
AI PPT Track Terminator, the strongest PPT Skill ever
Flexible Photo Recrafting While Preserving Your Identity
OpenAI swift async text to image for SwiftUI app using OpenAI
[NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences
State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX
The free, Open Source alternative to OpenAI, Claude and others
CogView4, CogView3-Plus and CogView3(ECCV 2024)