SOTA on-device LLMs, small yet powerful
GLM-5: From Vibe Coding to Agentic Engineering
Convert Google Gemini web into OpenAI-compatible API
New set of lightweight state-of-the-art, open foundation models
Open Frontier Intelligence
High-Resolution 3D Assets Generation with Large Scale Diffusion Models
Contexts Optical Compression
MiniMax M2.1, a SOTA model for real-world dev & agents.
Moonshot's most powerful AI model
Strong, Economical, and Efficient Mixture-of-Experts Language Model
MiMo-V2-Flash: Efficient Reasoning, Coding, and Agentic Foundation
State of the art LLM and coding model
gpt-oss-120b and gpt-oss-20b are two open-weight language models
Qwen3 is the large language model series developed by Qwen team
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Fast, Sharp & Reliable Agentic Intelligence
New family of code large language models (LLMs)
MiniMax-M2, a model built for Max coding & agentic workflows
DeepSeek LLM: Let there be answers
Efficient 13B MoE language model with long context and reasoning modes
Pruned GLM-5.3 model for self-hosted cybersecurity AI and coding
Open agentic coding model optimized for local deployment
685B model with improved agents and consistency
Efficient 14B multimodal instruct model with edge deployment and FP8
Efficient MoE model for reasoning, coding, and AI agent workflows