deepseek-ocr
DeepSeek logo

DeepSeek Ocr

deepseek-ocr
DeepSeek
DeepSeek-OCR is a vision-language model launched by DeepSeek AI, focusing on optical character recognition (OCR) and “contextual optical compression.” The model is designed to explore the limits of compressing contextual information from images, efficiently processing documents and converting them into structured text formats such as Markdown. The model requires an image as input.

Pricing

  • Input Tokens: $0.020 /M tokens
  • Output Tokens: $0.020 /M tokens

Input Modalities

  • Text
  • Vision

Output Modalities

  • Text

Frequently asked questions

What is deepseek-ocr?

DeepSeek-OCR is a vision-language model launched by DeepSeek AI, focusing on optical character recognition (OCR) and “contextual optical compression.” The model is designed to explore the limits of compressing contextual information from images, efficiently processing documents and converting them into structured text formats such as Markdown. The model requires an image as input.