Trillion-parameter MoE model for coding and million-token reasoning
Lightweight 24B agentic coding model with vision and long context
Powerful 14B LLM with strong instruction and long-text handling
770B MoE model for coding, research, reasoning, and long-context work
Efficient multimodal MoE model for coding, reasoning, and AI agents
Efficient 320B multimodal MoE model for coding and autonomous agents
Efficient 30B MoE model for long-running agents and local inference
NVFP4 DiffusionGemma model for fast multimodal text generation
Unified multimodal Gemma model for local coding and reasoning
FP8 Qwen model for efficient multimodal coding and agent tasks
Efficient bilingual MoE model for reasoning, coding, RAG, and agents
Coding-focused Kimi model for long-horizon agent workflows
Frontier-scale 675B multimodal base model for custom AI training
Quantized 675B multimodal instruct model optimized for NVFP4
Frontier-scale 675B multimodal instruct MoE model for enterprise AIMis
Compact 3B-param multimodal model for efficient on-device reasoning
Statelets is a coordination language.
Open VLA model for autonomous driving reasoning and planning
Llama 3.2–1B: Multilingual, instruction-tuned model for mobile AI
Rule-based information extraction.
Pruned GLM-5.3 model for self-hosted cybersecurity AI and coding
Dense 27B multimodal model for coding, agents, and visual reasoning
Local multimodal 30B model for autonomous agents, coding, and tools
Frontier multimodal MoE model for coding and AI agent workflows