Inkling-SmallThinking Machines Lab
|
Qwen CodeQwen
|
|||||
Related Products
|
||||||
About
Inkling-Small is an efficient model that offers performance comparable to Inkling at a quarter of its size. It is a Mixture-of-Experts transformer with 276 billion total parameters and 12 billion active parameters, trained on NVIDIA GB300 NVL72 systems. It supports native reasoning across text, images, and audio, variable thinking effort, and context windows of up to one million tokens. Users adjust reasoning effort from minimal to extra high to balance performance and compute. Improved pre-training data, post-training with on-policy distillation from Inkling, and extended agentic coding reinforcement learning helped Inkling-Small surpass its larger counterpart on reasoning and coding benchmarks. It performs well in coding and tool-use harnesses, exceeds 80% on SWE-bench Verified, and combines strong reasoning with efficient output. Its encoder-free multimodal architecture processes audio as dMel spectrograms and images as 40-by-40-pixel patches alongside text tokens.
|
About
Qwen3‑Coder is an agentic code model available in multiple sizes, led by the 480B‑parameter Mixture‑of‑Experts variant (35B active) that natively supports 256K‑token contexts (extendable to 1M) and achieves state‑of‑the‑art results on Agentic Coding, Browser‑Use, and Tool‑Use tasks comparable to Claude Sonnet 4. Pre‑training on 7.5T tokens (70 % code) and synthetic data cleaned via Qwen2.5‑Coder optimized both coding proficiency and general abilities, while post‑training employs large‑scale, execution‑driven reinforcement learning and long‑horizon RL across 20,000 parallel environments to excel on multi‑turn software‑engineering benchmarks like SWE‑Bench Verified without test‑time scaling. Alongside the model, the open source Qwen Code CLI (forked from Gemini Code) unleashes Qwen3‑Coder in agentic workflows with customized prompts, function calling protocols, and seamless integration with Node.js, OpenAI SDKs, and more.
|
|||||
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
|||||
Audience
Developers, AI agent builders, software engineering teams, research teams, enterprise AI teams, multimodal application developers, coding assistant builders, tool-use workflow teams, and organizations that need efficient reasoning, long-context processing, text-image-audio understanding, adjustable thinking effort, coding performance, and scalable Mixture-of-Experts inference
|
Audience
AI researchers and software engineers interested in a tool providing an agentic coding model with large‑context support and turnkey CLI tools for real‑world, and automated code generation
|
|||||
Support
Phone Support
24/7 Live Support
Online
|
Support
Phone Support
24/7 Live Support
Online
|
|||||
API
Offers API
|
API
Offers API
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
$0.30 per million input tokens
$0.30 per million input tokens and $1.20 per million output tokens
Free Version
Free Trial
|
Pricing
Free
Free Version
Free Trial
|
|||||
Reviews/
|
Reviews/
|
|||||
Pros & Cons from Real UsersPros
Cons
|
||||||
Training
Documentation
Webinars
Live Online
In Person
|
Training
Documentation
Webinars
Live Online
In Person
|
|||||
Company InformationThinking Machines Lab
Founded: 2025
United States
thinkingmachines.ai/news/inkling-small/
|
Company InformationQwen
Founded: 2023
China
github.com/QwenLM/qwen-code
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
|
|
||||||
Categories |
Categories |
|||||
Integrations
Alibaba AI Coding Plan
Claude Fable 5
Claude Mythos 5
Claude Opus 4.1
Claude Opus 4.5
Claude Opus 4.6
Claude Opus 4.7
Claude Opus 4.8
Claude Opus 5
Claude Sonnet 4
|
Integrations
Alibaba AI Coding Plan
Claude Fable 5
Claude Mythos 5
Claude Opus 4.1
Claude Opus 4.5
Claude Opus 4.6
Claude Opus 4.7
Claude Opus 4.8
Claude Opus 5
Claude Sonnet 4
|
|||||
|
|
|