Skill · Design

ocr-image-to-markdown

Designmcpdirectory↓ 18

鉴于本地 OCR 工具的缺失,本技能利用 Agent 的多模态能力来查看图像(PNG, JPG 等)并将内容(文本、表格、逻辑图)转录为格式化的 Markdown。

Install
$ npx skills add hugohe3/ppt-master --skill ocr-image-to-markdown

Generic steps — the original listing is authoritative. Copy the command, run it where you keep agent skills, then confirm version and requirements in the original listing.

At a glance

  • 18 installs reported on mcpdirectory.
  • Published by hugohe3 — listed by mcpdirectory. Setup details live in the original listing.

Capabilities

MarkdownImages/OCR

Auto-detected from the listing text — confirm in the original listing.

How to use

  1. Copy the Install command above.
  2. Run it where you keep agent skills.
  3. Confirm version and requirements in the original listing.

More in Design

  • skeleton-svelte — Use this skill when working with Skeleton UI components in Svelte projects.
  • baoyu-article-illustrator — Analyzes article structure, identifies positions requiring visual aids, generates illustrations with Type × St
  • office-to-md — Convert Office documents (Word, Excel, PowerPoint, PDF) to Markdown using Microsoft's markitdown
  • pdf-ocr-extraction — Extract text from scanned PDFs using optical character recognition
  • interaction-design — Design and implement microinteractions, motion design, transitions, and user feedback patterns. Use when addin
  • generate-image — Generate or edit images using AI models (FLUX, Gemini). Use for scientific illustrations, diagrams, schematics

← Back to Skills catalog · View original listing ↗

Outdated or wrong? Report this listing.