GLM-4.1V

GLM-4.1V

Zhipu AI
Kimi K2

Kimi K2

Moonshot AI
+
+

Related Products

  • LM-Kit.NET
    23 Ratings
    Visit Website
  • Vertex AI
    783 Ratings
    Visit Website
  • Google AI Studio
    11 Ratings
    Visit Website
  • Ango Hub
    15 Ratings
    Visit Website
  • LTX
    141 Ratings
    Visit Website
  • LogicalDOC
    123 Ratings
    Visit Website
  • Google Cloud Speech-to-Text
    373 Ratings
    Visit Website
  • Interfacing Integrated Management System (IMS)
    71 Ratings
    Visit Website
  • Rise Vision
    1,281 Ratings
    Visit Website
  • Picsart Enterprise
    26 Ratings
    Visit Website

About

GLM-4.1V is a vision-language model, providing a powerful, compact multimodal model designed for reasoning and perception across images, text, and documents. The 9-billion-parameter variant (GLM-4.1V-9B-Thinking) is built on the GLM-4-9B foundation and enhanced through a specialized training paradigm using Reinforcement Learning with Curriculum Sampling (RLCS). It supports a 64k-token context window and accepts high-resolution inputs (up to 4K images, any aspect ratio), enabling it to handle complex tasks such as optical character recognition, image captioning, chart and document parsing, video and scene understanding, GUI-agent workflows (e.g., interpreting screenshots, recognizing UI elements), and general vision-language reasoning. In benchmark evaluations at the 10 B-parameter scale, GLM-4.1V-9B-Thinking achieved top performance on 23 of 28 tasks.

About

Kimi K2 is a state-of-the-art open source large language model series built on a mixture-of-experts (MoE) architecture, featuring 1 trillion total parameters and 32 billion activated parameters for task-specific efficiency. Trained with the Muon optimizer on over 15.5 trillion tokens and stabilized by MuonClip’s attention-logit clamping, it delivers exceptional performance in frontier knowledge, reasoning, mathematics, coding, and general agentic workflows. Moonshot AI provides two variants, Kimi-K2-Base for research-level fine-tuning and Kimi-K2-Instruct pre-trained for immediate chat and tool-driven interactions, enabling both custom development and drop-in agentic capabilities. Benchmarks show it outperforms leading open source peers and rivals top proprietary models in coding tasks and complex task breakdowns, while its 128 K-token context length, tool-calling API compatibility, and support for industry-standard inference engines.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Developers and AI researchers seeking a solution offering a vision-language model that balances size and capability, ideal for building multimodal agents, document/image analysis tools, or GUI-based automation workflows

Audience

AI researchers, developers and organizations seeking an open source, high-performance agentic intelligence model for advanced reasoning, coding and autonomous task execution

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Free Version
Free Trial

Pricing

Free
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Zhipu AI
Founded: 2023
China
chat.z.ai/

Company Information

Moonshot AI
Founded: 2023
China
moonshotai.github.io/Kimi-K2/

Alternatives

GLM-4.6V

GLM-4.6V

Zhipu AI

Alternatives

Claude Code

Claude Code

Anthropic
HunyuanOCR

HunyuanOCR

Tencent
Claude Opus 4.5

Claude Opus 4.5

Anthropic
GLM-4.5V-Flash

GLM-4.5V-Flash

Zhipu AI
Kimi K2 Thinking

Kimi K2 Thinking

Moonshot AI
DeepSeek-V2

DeepSeek-V2

DeepSeek

Categories

Categories

Integrations

AiAssistWorks
Brokk
Claude Code
Cline
EaseMate AI
Kilo Code
Kimi
NVIDIA TensorRT
Nebius Token Factory
Okara
OpenCode
OpenRouter
Roo Code
SiliconFlow
Sup AI

Integrations

AiAssistWorks
Brokk
Claude Code
Cline
EaseMate AI
Kilo Code
Kimi
NVIDIA TensorRT
Nebius Token Factory
Okara
OpenCode
OpenRouter
Roo Code
SiliconFlow
Sup AI
Claim GLM-4.1V and update features and information
Claim GLM-4.1V and update features and information
Claim Kimi K2 and update features and information
Claim Kimi K2 and update features and information