Aya Vision

Aya Vision

Cohere AI
GLM-OCR

GLM-OCR

Z.ai
+
+

Related Products

  • Vertex AI
    961 Ratings
    Visit Website
  • Google AI Studio
    11 Ratings
    Visit Website
  • LM-Kit.NET
    25 Ratings
    Visit Website
  • LTX
    181 Ratings
    Visit Website
  • Imorgon
    5 Ratings
    Visit Website
  • PackageX OCR Scanning
    46 Ratings
    Visit Website
  • RaimaDB
    12 Ratings
    Visit Website
  • Windocks
    7 Ratings
    Visit Website
  • CompUp
    66 Ratings
    Visit Website
  • TeleRay
    6 Ratings
    Visit Website

About

Aya Vision is a research model advancing in multilingual multimodal AI through innovative synthetic data generation, cross-modal model merging, and a comprehensive benchmark suite. It achieves state-of-the-art performance across 23 languages, surpassing larger models while efficiently addressing data scarcity and catastrophic forgetting by reducing computational overhead up to 40% via optimized training techniques.

About

GLM-OCR is a multimodal optical character recognition model and open source repository that provides accurate, efficient, and comprehensive document understanding by combining text and visual modalities into a unified encoder–decoder architecture derived from the GLM-V family. Built with a visual encoder pre-trained on large-scale image–text data and a lightweight cross-modal connector feeding into a GLM-0.5B language decoder, the model supports layout detection, parallel region recognition, and structured output for text, tables, formulas, and complicated real-world document formats. It introduces Multi-Token Prediction (MTP) loss and stable full-task reinforcement learning to improve training efficiency, recognition accuracy, and generalization, achieving state-of-the-art benchmarks on major document understanding tasks.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Researchers and developers building multilingual AI applications that require understanding and generating content from both text and images

Audience

Developers, researchers, and engineers wanting a tool to accurately parse and understand complex documents, layouts, and visual-text content at scale

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Free Version
Free Trial

Pricing

Free
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Cohere AI
Founded: 2019
Canada
cohere.com/research/aya

Company Information

Z.ai
Founded: 2019
China
github.com/zai-org/GLM-OCR

Alternatives

Pixtral Large

Pixtral Large

Mistral AI

Alternatives

HunyuanOCR

HunyuanOCR

Tencent
CodeT5

CodeT5

Salesforce
Falcon 2

Falcon 2

Technology Innovation Institute (TII)
GLM-OCR

GLM-OCR

Z.ai
Mu

Mu

Microsoft
Qwen3.5

Qwen3.5

Alibaba
Qwen3-VL

Qwen3-VL

Alibaba

Categories

Categories

Integrations

No info available.

Integrations

No info available.
Claim Aya Vision and update features and information
Claim Aya Vision and update features and information
Claim GLM-OCR and update features and information
Claim GLM-OCR and update features and information