Strong, Economical, and Efficient Mixture-of-Experts Language Model
Codex plugin that turns attached object images into code-only
Qwen's most powerful open-source image generation model
Accurate × Fast × Comprehensive
Open-source multi-speaker long-form text-to-speech model
Let the Xiaoai speaker "hear your voice"
GLM-Image: Auto-regressive for Dense-knowledge and High-fidelity Image
Tiny vision language model
AI cognitive-enhancement Skills based on Anthropic's J-space
Instructions on how to use the Realtime API on Microcontrollers
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
New family of code large language models (LLMs)
Di♪♪Rhythm: Blazingly Fast & Simple End-to-End Song Generation
Open Multilingual Multimodal Chat LMs
Pushing the Limits of Mathematical Reasoning in Open Language Models
800,000 step-level correctness labels on LLM solutions to MATH problem
Learning to Act by Watching Unlabeled Online Videos
Stable fine-tuned Gemma model for structured, clear responses
Hermes 4 FP8: hybrid reasoning Llama-3.1-405B model by Nous Research
Efficient 250B MoE model for agents, coding, and long-context work
Lightweight MoE model for local reasoning, coding, and AI agents
Vision-language-action model for robot control via images and text