State-of-the-art TTS model under 25MB
Code for running inference with the SAM 3D Body Model 3DB
Repo for SeedVR2 & SeedVR
A Unified Framework for Text-to-3D and Image-to-3D Generation
Qwen2.5-VL is the multimodal large language model series
Python bindings for llama.cpp
Contexts Optical Compression
Provides convenient access to the Anthropic REST API from any Python 3
DeepSeek Coder: Let the Code Write Itself
AI PPT Track Terminator, the strongest PPT Skill ever
Open image model at the forefront of design
Python SDK for Claude Agent
Open-Source Financial Large Language Models
tiktoken is a fast BPE tokeniser for use with OpenAI's models
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
Global weather forecasting model using graph neural networks and JAX
GLM-4-Voice | End-to-End Chinese-English Conversational Model
GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
A SOTA open-source image editing model
Audio foundation model excelling in audio understanding
Revolutionizing Database Interactions with Private LLM Technology
Convert Google Gemini web into OpenAI-compatible API
26m function call model that runs on incredibly small devices