Hunyuan Translation Model Version 1.5
Qwen3-ASR is an open-source series of ASR models
26m function call model that runs on incredibly small devices
tiktoken is a fast BPE tokeniser for use with OpenAI's models
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
AI cognitive-enhancement Skills based on Anthropic's J-space
GLM-4-Voice | End-to-End Chinese-English Conversational Model
GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning
Sharp Monocular Metric Depth in Less Than a Second
scikit-learn compatible tabular foundation model
1B text generation model based on the HRM architecture
Foundation model for image generation
A Pragmatic VLA Foundation Model
Qwen3-omni is a natively end-to-end, omni-modal LLM
Video Object and Interaction Deletion
CogView4, CogView3-Plus and CogView3(ECCV 2024)
Renderer for the harmony response format to be used with gpt-oss
Open image model at the forefront of design
Genome modeling and design across all domains of life
Achieving 3+ generation speedup on reasoning tasks
Ultra-Efficient LLMs on End Device
Project Lyra: Open Generative 3D World Models
A Unified Framework for Text-to-3D and Image-to-3D Generation
Provides convenient access to the Anthropic REST API from any Python 3
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding