Skills Plugins MCP Prompt Model 博客 我的中心

ocr

Extract text from images using Tesseract OCR

DeepseekModel 官方收录技能 质量 优秀 · 90 v1.0.0

获取

https://deepseekmodel.com/api/download.php?id=trpc-group-trpc-agent-go-examples-skill-skills-ocr-skill-md&format=skill
下载 .skill 标准格式,含 system_prompt 与 model_config,导入任意 Agent 框架即可使用
.skill 文件中 system_prompt 字段的实际内容。
name ocr description Extract text from images using Tesseract OCR OCR Image Text Extraction Skill Extract text from images using Tesseract OCR engine. Capabilities Extract text from image files (PNG, JPG, JPEG, GIF, BMP, TIFF) Support for 100+ languages Optional image preprocessing for better accuracy Output in plain text or JSON format with confidence scores Usage Basic OCR python3 scripts/ocr.py <image_file> <output_file> With Options # Specify language (default: eng) python3 scripts/ocr.py image.png text.txt --lang eng # Chinese text python3 scripts/ocr.py image.png text.txt --lang chi_sim # Multiple languages python3 scripts/ocr.py image.png text.txt --lang eng+chi_sim # With image preprocessing (improves accuracy) python3 scripts/ocr.py image.png text.txt --preprocess # JSON output with confidence scores python3 scripts/ocr.py image.png output.json --format json Download and OCR from URL # OCR from remote image python3 scripts/ocr_url.py <image_url> <output_file> # With options python3 scripts/ocr_url.py https://example.com/image.jpg text.txt --lang eng --preprocess Parameters image_file / image_url (required): Path to local image or image URL output_file (required): Path to output text/JSON file --lang : Language code (e.g., eng, chi_sim, jpn, fra, deu). Default: eng --preprocess : Apply image preprocessing (grayscale, thresholding) for better accuracy --format : Output format (text/json, default: text) Common Languages Language Code English eng Chinese (Simplified) chi_sim Chinese (Traditional) chi_tra Japanese jpn Korean kor French fra German deu Spanish spa Russian rus Arabic ara Supported Image Formats PNG, JPG, JPEG, GIF, BMP, TIFF, WEBP Dependencies Python 3.8+ pytesseract Pillow (PIL) tesseract-ocr (system package) Installation # Python packages pip install pytesseract Pillow # Tesseract OCR engine sudo apt-get install tesseract-ocr # Ubuntu/Debian sudo yum install tesseract # CentOS/RHEL brew install tesseract # macOS
Agent 识别该技能的关键词,点击任意一个即可复制。

该技能未提供触发词。

下载的 .skill 包内含以下字段。
字段 说明
format格式标识(skill/v1)
skill_id技能唯一 ID
name技能名称
version版本号
description技能描述
category所属分类(数组)
trigger_words触发词列表
tags标签列表
source来源标识
source_url来源链接(本页地址)
exported_at导出时间(每次下载生成)
system_prompt系统提示词正文
model_config模型参数:provider / model / temperature / max_tokens / top_p
examples示例
install_guide各平台导入说明(Coze / Dify / Claude / 自定义框架)
同一份技能可按不同平台格式导出。
.skill 标准格式,含 system_prompt 与 model_config,导入任意 Agent 框架即可使用 下载
.skillpro 增强格式,额外含脚本 / 工具 / 依赖 / 钩子占位 下载
.json 纯 JSON 导出,只含 system_prompt 与模型参数 下载
Coze 带 frontmatter 的 Markdown,Coze 平台导入用 下载
Dify Dify DSL,创建应用后直接导入 下载

每日精选 Skill 推荐,免费送到你邮箱

输入邮箱,每天接收一个精选 AI Agent 技能推荐。完全免费,持续更新。

提交后我们会发送一封确认邮件,点击邮件里的链接才会开始收信。

完全免费,取消任意时间。我们不会发送垃圾邮件。