DeepSeek-OCR

DeepSeek-OCR

DeepSeek
+
+

Related Products

  • RunPod
    211 Ratings
    Visit Website
  • Servers.com
    15 Ratings
    Visit Website
  • Google Compute Engine
    1,168 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Gemini Enterprise Agent Platform
    967 Ratings
    Visit Website
  • QEval
    30 Ratings
    Visit Website
  • Nexcess Managed Solutions
    210 Ratings
    Visit Website
  • Google Cloud Speech-to-Text
    365 Ratings
    Visit Website
  • ManageEngine ServiceDesk Plus
    1,947 Ratings
    Visit Website
  • RaimaDB
    12 Ratings
    Visit Website

About

DeepInfra is an AI inference cloud that makes it simple to run the latest machine learning models at scale, including LLMs, vision models, embeddings, image generation, video generation, speech, and more. It provides serverless inference through simple APIs, allowing developers to integrate production-ready AI models without managing GPU infrastructure, autoscaling, deployment complexity, or model hosting operations. DeepInfra supports OpenAI-compatible APIs for LLMs and embeddings, making it easier to switch from existing OpenAI-style integrations while accessing a broad catalog of open and commercial models. Its Native API gives access to every model type available on the platform, including image generation, speech recognition, object detection, token classification, fill-mask, image classification, zero-shot image classification, and text classification. DeepInfra is optimized for scalable, low-latency inference and runs models on high-performance GPU infrastructure.

About

DeepSeek-OCR is an open source model for Contexts Optical Compression, built to explore the boundaries of visual-text compression and investigate the role of vision encoders from an LLM-centric viewpoint. It is designed to compress long contexts through optical 2D mapping, using DeepEncoder as the core engine and DeepSeek3B-MoE-A570M as the decoder. DeepEncoder maintains low activations under high-resolution input while achieving high compression ratios, keeping the number of vision tokens manageable for document understanding. The model supports OCR and document parsing workflows for images and PDFs, with inference through vLLM or Transformers. Users can run image OCR with streaming output, process PDFs with high concurrency, or run batch evaluation for benchmarks. DeepSeek-OCR can convert documents to Markdown, perform free OCR without layouts, parse figures, describe images in detail, and locate referenced text inside an image.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

AI developers and product teams that need scalable serverless inference, OpenAI-compatible APIs, and hosted access to LLM, vision, embedding, speech, and image models

Audience

AI researchers and document-processing engineers who need an open OCR model for efficient document parsing, Markdown conversion, and vision-text compression experiments

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

$1.98 per hour
Free Version
Free Trial

Pricing

Free
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

DeepInfra
Founded: 2022
United States
deepinfra.com

Company Information

DeepSeek
Founded: 2023
China
github.com/deepseek-ai/DeepSeek-OCR

Alternatives

Alternatives

GLM-OCR

GLM-OCR

Z.ai
fal

fal

fal.ai
DeepSeek-VL

DeepSeek-VL

DeepSeek
DeepSeek-V2

DeepSeek-V2

DeepSeek
DeepSeek-V4

DeepSeek-V4

DeepSeek

Categories

Categories

Integrations

DeepSeek
Anthropic
Claude
Gemini
Markdown
Mistral AI
OpenAI
Qwen

Integrations

DeepSeek
Anthropic
Claude
Gemini
Markdown
Mistral AI
OpenAI
Qwen
Claim DeepInfra and update features and information
Claim DeepInfra and update features and information
Claim DeepSeek-OCR and update features and information
Claim DeepSeek-OCR and update features and information