Lets make video diffusion practical
Qwen3-TTS is an open-source series of TTS models
The official repo of Qwen chat & pretrained large language model
Python SDK for Claude Agent
OCR expert VLM powered by Hunyuan's native multimodal architecture
Provides convenient access to the Anthropic REST API from any Python 3
Official implementation of DreamCraft3D
Robust Speech Recognition Across Languages, Dialects
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Generate Any 3D Scene in Seconds
Global weather forecasting model using graph neural networks and JAX
Language modeling in a sentence representation space
Large Multimodal Models for Video Understanding and Editing
A Conversational Speech Generation Model
Powerful open source image generation model
Pushing the Limits of Mathematical Reasoning in Open Language Models
Chat & pretrained large vision language model
Example Discord bot written in Python that uses the completions API
Official code for Style Aligned Image Generation via Shared Attention
Tencent’s 36-language state-of-the-art translation model