GLM-OCR

GLM-OCR

Z.ai
Ray2

Ray2

Luma AI
+
+

Related Products

  • PackageX OCR Scanning
    46 Ratings
    Visit Website
  • LogicalDOC
    125 Ratings
    Visit Website
  • Nutrient SDK
    104 Ratings
    Visit Website
  • Square 9
    403 Ratings
    Visit Website
  • Apryse PDF SDK
    149 Ratings
    Visit Website
  • MyQ
    179 Ratings
    Visit Website
  • onPhase
    216 Ratings
    Visit Website
  • LM-Kit.NET
    23 Ratings
    Visit Website
  • Google AI Studio
    11 Ratings
    Visit Website
  • Google Cloud Speech-to-Text
    374 Ratings
    Visit Website

About

GLM-OCR is a multimodal optical character recognition model and open source repository that provides accurate, efficient, and comprehensive document understanding by combining text and visual modalities into a unified encoder–decoder architecture derived from the GLM-V family. Built with a visual encoder pre-trained on large-scale image–text data and a lightweight cross-modal connector feeding into a GLM-0.5B language decoder, the model supports layout detection, parallel region recognition, and structured output for text, tables, formulas, and complicated real-world document formats. It introduces Multi-Token Prediction (MTP) loss and stable full-task reinforcement learning to improve training efficiency, recognition accuracy, and generalization, achieving state-of-the-art benchmarks on major document understanding tasks.

About

Ray2 is a large-scale video generative model capable of creating realistic visuals with natural, coherent motion. It has a strong understanding of text instructions and can take images and video as input. Ray2 exhibits advanced capabilities as a result of being trained on Luma’s new multi-modal architecture scaled to 10x compute of Ray1. Ray2 marks the beginning of a new generation of video models capable of producing fast coherent motion, ultra-realistic details, and logical event sequences. This increases the success rate of usable generations and makes videos generated by Ray2 substantially more production-ready. Text-to-video generation is available in Ray2 now, with image-to-video, video-to-video, and editing capabilities coming soon. Ray2 brings a whole new level of motion fidelity. Smooth, cinematic, and jaw-dropping, transform your vision into reality. Tell your story with stunning, cinematic visuals. Ray2 lets you craft breathtaking scenes with precise camera movements.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Developers, researchers, and engineers wanting a tool to accurately parse and understand complex documents, layouts, and visual-text content at scale

Audience

Content creators seeking a tool to generate high-quality, realistic videos efficiently

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Free Version
Free Trial

Pricing

$9.99 per month
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Z.ai
Founded: 2019
China
github.com/zai-org/GLM-OCR

Company Information

Luma AI
Founded: 2021
United States
lumalabs.ai/ray

Alternatives

CodeT5

CodeT5

Salesforce

Alternatives

Ray3

Ray3

Luma AI
HunyuanOCR

HunyuanOCR

Tencent
Ray3.14

Ray3.14

Luma AI
Mu

Mu

Microsoft
Qwen3-VL

Qwen3-VL

Alibaba
Kling 2.5

Kling 2.5

Kuaishou Technology
Kling O1

Kling O1

Kling AI

Categories

Categories

Integrations

AIVideo.com
CinemaDrop
Fuser
KomikoAI
Weavy

Integrations

AIVideo.com
CinemaDrop
Fuser
KomikoAI
Weavy
Claim GLM-OCR and update features and information
Claim GLM-OCR and update features and information
Claim Ray2 and update features and information
Claim Ray2 and update features and information