Skill · Design
ocr-image-to-markdown
鉴于本地 OCR 工具的缺失,本技能利用 Agent 的多模态能力来查看图像(PNG, JPG 等)并将内容(文本、表格、逻辑图)转录为格式化的 Markdown。
Install
$
npx skills add hugohe3/ppt-master --skill ocr-image-to-markdownGeneric steps — the original listing is authoritative. Copy the command, run it where you keep agent skills, then confirm version and requirements in the original listing.
At a glance
- 18 installs reported on mcpdirectory.
- Published by hugohe3 — listed by mcpdirectory. Setup details live in the original listing.
Capabilities
MarkdownImages/OCR
Auto-detected from the listing text — confirm in the original listing.
How to use
- Copy the Install command above.
- Run it where you keep agent skills.
- Confirm version and requirements in the original listing.
More in Design
- skeleton-svelte — Use this skill when working with Skeleton UI components in Svelte projects.
- baoyu-article-illustrator — Analyzes article structure, identifies positions requiring visual aids, generates illustrations with Type × St
- office-to-md — Convert Office documents (Word, Excel, PowerPoint, PDF) to Markdown using Microsoft's markitdown
- pdf-ocr-extraction — Extract text from scanned PDFs using optical character recognition
- interaction-design — Design and implement microinteractions, motion design, transitions, and user feedback patterns. Use when addin
- generate-image — Generate or edit images using AI models (FLUX, Gemini). Use for scientific illustrations, diagrams, schematics
← Back to Skills catalog · View original listing ↗
Outdated or wrong? Report this listing.