business-ops
Business operations: strategy, technology, growth, competitive intelligence, support, finance, HR, legal, operations, sales, productivity, product management.
Convert PDF, Office, HTML, data, media, ZIP to Markdown.
$ npx -y skills add notque/vexjoy-agent --skill markdown-converter --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/markdown-converterContext preview
The summary Claude sees to decide when to auto-load this skill.
Convert PDF, Office, HTML, data, media, ZIP to Markdown.
name: markdown-converter
description: "Convert PDF, Office, HTML, data, media, ZIP to Markdown."
user_invocable: false # default -- router-dispatched, not user-typed
agent: python-general-engineer
allowed-tools:
- Bash
- Read
routing:
triggers:
- "convert to markdown"
- "markitdown"
- "extract text from PDF"
- "PDF to markdown"
- "docx to markdown"
- "ingest document"
- "read this PDF"
- "read this document"
- "extract text from document"
- "convert PDF"
- "convert document"
- "pptx to markdown"
- "xlsx to markdown"
category: research
pairs_with:
- research-pipeline
- enterprise-searchConvert a file to Markdown with markitdown, zero install:
uvx 'markitdown[all]' input.pdf -o output.md # to file uvx 'markitdown[all]' input.docx # to stdout cat blob | uvx 'markitdown[all]' -x .pdf # stdin, with extension hint
When `uvx` is missing, run `pipx run 'markitdown[all]' …` with the same arguments. First run downloads dependencies; later runs hit the cache. Output preserves headings, tables, lists, and links.
For video transcripts, use the `video-transcript` skill.
| Input | Notes | |---|---| | PDF, .docx, .pptx, .xlsx, .xls | Document structure preserved | | HTML, CSV, JSON, XML | Structured Markdown | | Images | EXIF metadata + OCR text | | Audio | EXIF metadata + speech transcription | | ZIP, EPub | Iterates contents, converts each |
| Flag | Effect | |---|---| | `-o FILE` | Write output to FILE | | `-x .EXT` | Extension hint for stdin input | | `-m MIME` | MIME-type hint | | `-c CHARSET` | Charset hint, e.g. UTF-8 |
Cause: page is an image; the base extractor reads text layers only. Solution: render pages to images (`pdftoppm`), then convert the images so OCR runs.
Essays and writing behind this toolkit live at vexjoy.com. VexJoy Agent connects plain-English requests to specialist agents, skills, and workflows. /do selects the knowledge and tools needed for your task.
Repo: notque/vexjoy-agent
Business operations: strategy, technology, growth, competitive intelligence, support, finance, HR, legal, operations, sales, productivity, product management.
Design workflows — UX copy, design systems, design critique, accessibility review, design handoff, user research synthesis. Use when writing UI copy, reviewing…
Marketing: SEO audits, campaign planning, content strategy, email sequences, competitive analysis, brand review, performance reporting.