AI AGENT ADDONS
AI & Agent Building
1,428installs

DeepSeek-OCR is a smart tool that reads text from pictures, PDFs, and documents. It uses a special kind of AI that understands both images and words. This makes it great for turning scanned documents into editable text.

The tool can handle many images at once. It outputs clean markdown or plain text. You can also choose different modes for OCR, figure parsing, or grounding.

You run it on your own computer using tools like vLLM or HuggingFace. It works best with CUDA and PyTorch installed.

Add Deepseek Ocr skill to your workflow

Global

mkdir -p ~/.claude/skills/deepseek-ocr

Project

mkdir -p .claude/skills/deepseek-ocr

Source Repository

Stars
49
Forks
11
Watchers
49
Last Push
2 months ago
Created
2 months ago