deepseek-ocr · DeepSeek
DeepSeek-OCR is a vision-language model launched by DeepSeek AI, focusing on optical character recognition (OCR) and “contextual optical compression.” The model is designed to explore the limits of compressing contextual information from images, efficiently processing documents and converting them into structured text formats such as Markdown. The model requires an image as input.
DeepSeek-OCR is a vision-language model launched by DeepSeek AI, focusing on optical character recognition (OCR) and “contextual optical compression.” The model is designed to explore the limits of compressing contextual information from images, efficiently processing documents and converting them into structured text formats such as Markdown. The model requires an image as input.
DeepSeek Ocr has a 8,000 token context window.
On AIHubMix, DeepSeek Ocr costs $0.02 per million input tokens and $0.02 per million output tokens.
DeepSeek Ocr accepts text and image input.
DeepSeek Ocr is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to deepseek-ocr — no other code changes needed.
DeepSeek Ocr is developed by DeepSeek. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model…
DeepSeek’s officially released new multimodal visual-understanding model…
DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent…
DeepSeek V4 Flash 0731 Fast is a high-speed deployment of DeepSeek’s agentic model…
(This model currently points to the older 0423 version; if you need to request the latest…
(This model currently points to the older 0423 version; if you need to request the latest…
Use DeepSeek Ocr via the AIHubMix unified API — one interface for every major LLM.