Skip to content
Agent Orchestration
Skill

/omh-long-document-reading

[omh] Huge PDF or document to read in full: read a very large PDF, contract, manual, or report through Hermes in page-anchored ranges with a coverage ledger. Use when the user says: long-document-reading, long document reading, summarize this pdf, read this pdf, process this

BOOST
From plugin
oh-my-hermes
3.2k145 skills
Install
$ npx -y skills add rlaope/oh-my-hermes --skill omh-long-document-reading --agent claude-code

How it fires

How this skill gets triggered: by you, by Claude, or both.

  • Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
  • You can call itInvoke it directly when you want it.
  • Slash command/omh-long-document-reading

Context preview

The summary Claude sees to decide when to auto-load this skill.

[omh] Huge PDF or document to read in full: read a very large PDF, contract, manual, or report through Hermes in page-anchored ranges with a coverage ledger. Use when the user says: long-document-reading, long document reading, summarize this pdf, read this pdf, process this

SKILL.md

omh-long-document-reading.SKILL.md
name: "omh-long-document-reading"
description: "[omh] Huge PDF or document to read in full: read a very large PDF, contract, manual, or report through Hermes in page-anchored ranges with a coverage ledger. Use when the user says: long-document-reading, long document reading, summarize this pdf, read this pdf, process this pdf, go through this pdf, summarize this document, read this document."
metadata:
  hermes:
    tags: [workflow, oh-my-hermes, research]
    category: research
    phase: long-document-reading
    role: researcher
    quality_tier: long-document-gated

Long Document Reading

This is a Hermes-native `long-document-reading` workflow skill.

Why This Exists

`long-document-reading` exists because a 300-page PDF is about 500,000 characters and Hermes' `read_file` returns 100,000 per call with no page numbers, re-converting the whole file each time; five unanchored reads then sit in the conversation until the ratio-based compressor summarizes them without a page number, so without a ledger the session either truncates, loses the early ranges to compaction, or claims a summary of pages it never read.

Do Not Use When

  • The document is a research paper and the user wants it explained by level; use `paper-learning`.
  • The request asks to convert, export, split into a new file, compare two PDFs, or extract tables into CSV; use `materials-package`.
  • The input is an image, screenshot, receipt, audio, or video rather than a document; use `media-input-operator`.
  • The user is still looking for the document or its download link; use `source-finder`.
  • The document fits one read (under about 60 pages of prose); read it directly and answer.

Examples

Good example:

  • Prompt: summarize this 300-page vendor contract pdf and list every obligation with a deadline
  • Expected behavior: Prepare long_document_card/v1: record the page count and scanned flags, plan five 60-page ranges, delegate them with the per-range brief, merge obligations with page anchors, and close with covered / next / missing.
  • Why: The document is far past one read budget and the goal needs page-anchored claims from every range.

Bad example:

  • Prompt: turn this 300-page pdf into a slide deck
  • Expected behavior: Route to `materials-package`: the user wants a produced file, not a page-anchored reading of the document; the page count alone does not make it a reading request.
  • Why: Reading and producing are different lanes; a deck request is file output work.

Completion Checklist

  • The page count is observed or the card says it is not.
  • Every ledger range is covered, or the missing ranges are listed with a reason.
  • Every claim in the merged answer carries a page anchor.
  • Scanned ranges are read, declined with a reason, or listed as missing.
  • Not-observed boundaries remain visible: page_count, text_extraction, scanned_page_ocr, range_delegation, hosted_ocr, cross_range_consistency.

Recovery Notes

  • If a script reports a missing dependency, install the one it names once with `pip install` (`pdfplumber`, `pypdf`, `pymupdf`, or `pypdfium2`; poppler `pdftoppm` is the system alternative for rendering), rerun, and record the install.
  • If a range read truncates, halve the range, record the observed characters per page, and re-plan the remaining ranges from that measurement.
  • If the context was compacted or the session resumed, reread the ledger and continue from the `next` range; do not restart from page 1.
  • If the document is encrypted, ask for the password or stop; `pdf_read.py` and `pdf_split.py` accept `--password`.
  • If most pages are scanned and the goal needs them all, stop and get approval for the per-page OCR job before spending one vision call per page.

Workflow Lane

  • Current lane: **Research and company ops** (`product-docs`, `source-finder`, `web-research`, `research`, `model-optimization`, `inference-serving`, `model-finetuning`, `research-brief`, `+20 more`) - research, signals, ops, and briefings.
  • If intent belongs to another lane, hand back to `oh-my-hermes` or name the adjacent workflow.
  • Shared product, routing, compatibility, and evidence rules: `omh-routing/references/skill-common-rail.md`.

Use When

Use when Hermes must read a supplied document that does not fit one read: a contract, manual, annual report, specification, or any PDF past about 60 pages. The skill plans page ranges sized to the `read_file` budget, keeps a page-anchored chunk ledger with covered / next / missing state, and delegates ranges when there are more than 4, so a compacted or resumed session continues instead of restarting.

Strong routing signals: `long-document-reading`, `long document reading`, `summarize this pdf`, `read this pdf`, `process this pdf`, `go through this pdf`, `summarize this document`, `read this document`, `process this document`, `read this whole document`, `summarize this manual`, `read this manual`, `summarize this contract`, `read this contract`, `summarize this annual report`, `read this annual report`, `read the whole pdf`, `chunk this pdf`, `pdf in chunks`, `pdf too big`, `pdf too large`, `このpdfを要約`, `この文書を要約`, `この契約書を要約`, `マニュアルを要約`, `긴 문서 읽기`, `이 pdf 요약해줘`, `이 pdf 읽어줘`, `이 문서 요약해줘`, `이 문서 읽어줘`, `계약서 요약해줘`, `매뉴얼 요약해줘`, `연간 보고서 요약해줘`, `pdf 전체 읽어`, `문서 전체 읽어`, `总结这个pdf`, `总结这份文档`, `总结这份合同`, `总结这本手册`

Catalog Metadata

Category: `research` Phase: `long-document-reading` Hermes role: `researcher` Quality tier: `long-document-gated` Reasoning demand: `standard`

Quality bar:

  • Get the page count and scanned flags first with `pdf_read.py --meta`; each script names its own missing dependency (`pdfplumber` for `pdf_read.py`, `pypdf` for `pdf_split.py`, `pymupdf` for `extract_pymupdf.py`, `pypdfium2` or poppler `pdftoppm` for `pdf_page_image.py`); install it once, and say so.
  • Size ranges to the read budget: about 60 pages per 100,000-character call at typical density; halve the range when a probe read truncates.
  • Extract each range with page selection (
Read more
Ships withoh-my-hermes

English | 한국어 | 日本語 | 中文 Install once. Keep Hermes. Add a stronger operating layer. Planning, research, creation, coding handoffs, operations, and project memory with explicit evidence boundaries.

Get the whole plugin
Stats
3,206
Stars
244
Forks
Active
Maintenance
Python
Language
MIT
License
4h ago
Last commit
4mo ago
Created
9h ago
Added

Repo: rlaope/oh-my-hermes

Other skills on oh-my-hermes.