Revolutionizing Database Interactions with Private LLM Technology
Run GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine
Pokee Deep Research Model Open Source Repo
Stable Diffusion WebUI Forge is a platform on top of Stable Diffusion
DeepMind model for tracking arbitrary points across videos & robotics
Tooling for the Common Objects In 3D dataset
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
Renderer for the harmony response format to be used with gpt-oss
Generating Immersive, Explorable, and Interactive 3D Worlds
Miso TTS is an 8 billion, highly emotive text-to-speech model
Qwen3-omni is a natively end-to-end, omni-modal LLM
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
The official PyTorch implementation of Google's Gemma models
Inference code for scalable emulation of protein equilibrium ensembles
Tongyi Deep Research, the Leading Open-source Deep Research Agent
Audio Language Models are Few-Shot Learners
Open Source Speech Language Model
Open-source industrial-grade ASR models
Foundation model for image generation
Hunyuan Translation Model Version 1.5
Block Diffusion for Ultra-Fast Speculative Decoding
Multimodal embedding and reranking models built on Qwen3-VL
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning