Inkling-SmallThinking Machines Lab
|
Llama 4 ScoutMeta
|
|||||
Related Products
|
||||||
About
Inkling-Small is an efficient model that offers performance comparable to Inkling at a quarter of its size. It is a Mixture-of-Experts transformer with 276 billion total parameters and 12 billion active parameters, trained on NVIDIA GB300 NVL72 systems. It supports native reasoning across text, images, and audio, variable thinking effort, and context windows of up to one million tokens. Users adjust reasoning effort from minimal to extra high to balance performance and compute. Improved pre-training data, post-training with on-policy distillation from Inkling, and extended agentic coding reinforcement learning helped Inkling-Small surpass its larger counterpart on reasoning and coding benchmarks. It performs well in coding and tool-use harnesses, exceeds 80% on SWE-bench Verified, and combines strong reasoning with efficient output. Its encoder-free multimodal architecture processes audio as dMel spectrograms and images as 40-by-40-pixel patches alongside text tokens.
|
About
Llama 4 Scout is a powerful 17 billion active parameter multimodal AI model that excels in both text and image processing. With an industry-leading context length of 10 million tokens, it outperforms its predecessors, including Llama 3, in tasks such as multi-document summarization and parsing large codebases. Llama 4 Scout is designed to handle complex reasoning tasks while maintaining high efficiency, making it perfect for use cases requiring long-context comprehension and image grounding. It offers cutting-edge performance in image-related tasks and is particularly well-suited for applications requiring both text and visual understanding.
|
|||||
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
|||||
Audience
Developers, AI agent builders, software engineering teams, research teams, enterprise AI teams, multimodal application developers, coding assistant builders, tool-use workflow teams, and organizations that need efficient reasoning, long-context processing, text-image-audio understanding, adjustable thinking effort, coding performance, and scalable Mixture-of-Experts inference
|
Audience
Llama 4 Scout is perfect for developers, researchers, and businesses seeking an efficient, high-performance multimodal AI model for tasks involving both text and image data, including complex reasoning, summarization, and image understanding
|
|||||
Support
Phone Support
24/7 Live Support
Online
|
Support
Phone Support
24/7 Live Support
Online
|
|||||
API
Offers API
|
API
Offers API
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
$0.30 per million input tokens
$0.30 per million input tokens and $1.20 per million output tokens
Free Version
Free Trial
|
Pricing
Free
Open source
Free Version
Free Trial
|
|||||
Reviews/
|
Reviews/
|
|||||
Pros & Cons from Real UsersPros
Cons
|
||||||
Training
Documentation
Webinars
Live Online
In Person
|
Training
Documentation
Webinars
Live Online
In Person
|
|||||
Company InformationThinking Machines Lab
Founded: 2025
United States
thinkingmachines.ai/news/inkling-small/
|
Company InformationMeta
Founded: 2004
United States
ai.meta.com
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
|
|
||||||
Categories |
Categories |
|||||
Integrations
BLACKBOX AI
Baseten
C++
CSS
CometAPI
Elixir
F#
Groq
Kotlin
LLM Council
|
Integrations
BLACKBOX AI
Baseten
C++
CSS
CometAPI
Elixir
F#
Groq
Kotlin
LLM Council
|
|||||
|
|
|