/paper-image-extractor
Extract figures from papers — prioritizes arXiv source package for high-quality images
$ npx -y skills add OpenLAIR/dr-claw --skill paper-image-extractor --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
- Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
- You can call itInvoke it directly when you want it.
- Slash command
/paper-image-extractor
Context preview
The summary Claude sees to decide when to auto-load this skill.
Extract figures from papers — prioritizes arXiv source package for high-quality images
SKILL.md
paper-image-extractor.SKILL.mdname: paper-image-extractor description: Extract figures from papers — prioritizes arXiv source package for high-quality images allowed-tools: Read, Write, Bash
You are the Paper Image Extractor for Dr. Claw.
Goal
Extract all figures from a paper, prioritizing arXiv source packages for high-quality original images over PDF extraction.
Extraction Strategy (3-tier priority)
Priority 1: arXiv Source Package (Best)
1. Download source: `https://arxiv.org/e-print/[PAPER_ID]` 2. Extract and look for `pics/`, `figures/`, `fig/`, `images/`, `img/` directories 3. Copy image files to output directory 4. Convert PDF figures to PNG
Priority 2: PDF Figure Extraction (Fallback)
python scripts/extract_images.py "[PAPER_ID]" "[OUTPUT_DIR]" "[INDEX_PATH]"
Priority 3: Direct PDF Image Extraction (Last Resort)
Extract embedded image objects from the compiled PDF using PyMuPDF.
Output
- Images saved to specified output directory
- `index.md` generated with image metadata and source labels (arxiv-source, pdf-figure, pdf-extraction)
Scripts
- `scripts/extract_images.py` — Main extraction script with 3-tier strategy
Dependencies
- Python 3.8+, PyMuPDF (fitz), requests
- Network access (arXiv)
--- > Based on [evil-read-arxiv](https://github.com/evil-read-arxiv) — an automated paper reading workflow. MIT License.
A Super AI Lab with massive AI Doctors as Assistants. Best IDE for Research via AI Power.
Repo: OpenLAIR/dr-claw
Other skills on dr-claw.
- /dr-claw
Dr. Claw skill for OpenClaw project discovery, idea intake, waiting-session triage, structured session control, event-driven notifications, and mobile reporting through the local drclaw CLI.
Open skill - /academic-researcher
Academic research assistant for literature reviews, paper analysis, and scholarly writing. Use when: reviewing academic papers, conducting literature reviews, writing research summaries, analyzing methodologies, formatting citations, or when user mentions academic research,
Open skill - /autogpt
Autonomous AI agent platform for building and deploying continuous agents. Use when creating visual workflow agents, deploying persistent autonomous agents, or building complex multi-step AI automation systems.
Open skill - /crewai
Multi-agent orchestration framework for autonomous AI collaboration. Use when building teams of specialized agents working together on complex tasks, when you need role-based agent collaboration with memory, or for production workflows requiring sequential/hierarchical
Open skill - /langchain
Framework for building LLM-powered applications with agents, chains, and RAG. Supports multiple providers (OpenAI, Anthropic, Google), 500+ integrations, ReAct agents, tool calling, memory management, and vector store retrieval. Use for building chatbots, question-answering
Open skill - /llamaindex
Data framework for building LLM applications with RAG. Specializes in document ingestion (300+ connectors), indexing, and querying. Features vector indices, query engines, agents, and multi-modal support. Use for document Q&A, chatbots, knowledge retrieval, or building RAG
Open skill

