DeepSeek-OCR
DeepSeekDeepSeek-OCR is a vision-language model launched by DeepSeek AI, focusing on optical character recognition (OCR) and “contextual optical compression.” The model is designed to explore the limits of compressing contextual information from images, efficiently processing documents and converting them into structured text formats such as Markdown. The model requires an image as input.
Pricing
- Input Tokens: $0.020 /M tokens
- Output Tokens: $0.020 /M tokens
Input Modalities
- Text
- Vision
Output Modalities
- Text
Try this model
Python
