RGBD video generation model conditioned on camera input
Netease Youdao's open-source embedding and reranker models
An Efficient Agentic Model for Computer Use
Revolutionizing Database Interactions with Private LLM Technology
Pokee Deep Research Model Open Source Repo
Stable Diffusion WebUI Forge is a platform on top of Stable Diffusion
DeepMind model for tracking arbitrary points across videos & robotics
Tooling for the Common Objects In 3D dataset
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
Renderer for the harmony response format to be used with gpt-oss
Generating Immersive, Explorable, and Interactive 3D Worlds
Run GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine
Miso TTS is an 8 billion, highly emotive text-to-speech model
Qwen3-omni is a natively end-to-end, omni-modal LLM
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
The official PyTorch implementation of Google's Gemma models
Inference code for scalable emulation of protein equilibrium ensembles
Tongyi Deep Research, the Leading Open-source Deep Research Agent
Audio Language Models are Few-Shot Learners
Open Source Speech Language Model
Open-source industrial-grade ASR models
Foundation model for image generation
Hunyuan Translation Model Version 1.5
Block Diffusion for Ultra-Fast Speculative Decoding