bug-reproduce
Turn a known bug into a tight, red-capable reproducer, then prove the reproducer locks that…
Search arXiv papers by keyword, author, category, or ID.
$ npx -y skills add Prismer-AI/PrismerCloud --skill research-arxiv --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/research-arxivContext preview
The summary Claude sees to decide when to auto-load this skill.
Search arXiv papers by keyword, author, category, or ID.
name: research-arxiv
scope: common
category: research
description: "Search arXiv papers by keyword, author, category, or ID."
version: 1.0.0
author: Hermes Agent
license: MIT
platforms: [linux, macos, windows]
metadata:
nativeReplaces: [arxiv]
upstream: Hermes Agent skills/research/arxiv at 1a1f4a59e2
hermes:
tags: [Research, Arxiv, Papers, Academic, Science, API]
related_skills: [office-artifacts, evidence-citations]Search and retrieve academic papers from arXiv via their REST API. No API key. Use the shipped Python 3.9+ standard-library script relative to this installed skill directory, not a hardcoded Hermes repository path:
python3 "<skill-dir>/scripts/search_arxiv.py" "GRPO reinforcement learning" --max 5 --json python3 "<skill-dir>/scripts/search_arxiv.py" --id 1706.03762v1 --json
The script preserves versioned identifiers and full abstracts, bounds response size and retries transient HTTP errors, and rejects malformed feeds/options. `unknown` metadata is a coverage gap, not permission to invent dates. Use `evidence-citations` for quotations and `office-artifacts` for PDFs. Attach only requested file outputs through `cloud task attach`/`cloud deliver`; honor PKF carrier instructions for inline answers. Abstract retrieval is not full-paper reading.
| Action | Command | |--------|---------| | Search papers | `curl "https://export.arxiv.org/api/query?search_query=all:QUERY&max_results=5"` | | Get specific paper | `curl "https://export.arxiv.org/api/query?id_list=2402.03300"` | | Read abstract (web) | `web_extract(urls=["https://arxiv.org/abs/2402.03300"])` | | Read full paper (PDF) | `web_extract(urls=["https://arxiv.org/pdf/2402.03300"])` |
The API returns Atom XML. Use an XML parser (the shipped script), not grep/sed.
curl --fail --silent --show-error --max-time 30 "https://export.arxiv.org/api/query?search_query=all:GRPO+reinforcement+learning&max_results=5"
python3 "<skill-dir>/scripts/search_arxiv.py" "GRPO reinforcement learning" --max 5 --sort date
| Prefix | Searches | Example | |--------|----------|---------| | `all:` | All fields | `all:transformer+attention` | | `ti:` | Title | `ti:large+language+models` | | `au:` | Author | `au:vaswani` | | `abs:` | Abstract | `abs:reinforcement+learning` | | `cat:` | Category | `cat:cs.AI` | | `co:` | Comment | `co:accepted+NeurIPS` |
# AND (default when using +) search_query=all:transformer+attention # OR search_query=all:GPT+OR+all:BERT # AND NOT search_query=all:language+model+ANDNOT+all:vision # Exact phrase search_query=ti:"chain+of+thought" # Combined search_query=au:hinton+AND+cat:cs.LG
| Parameter | Options | |-----------|---------| | `sortBy` | `relevance`, `lastUpdatedDate`, `submittedDate` | | `sortOrder` | `ascending`, `descending` | | `start` | Result offset (0-based) | | `max_results` | Number of results (default 10, max 30000) |
# Latest 10 papers in cs.AI curl --fail --silent --show-error --max-time 30 "https://export.arxiv.org/api/query?search_query=cat:cs.AI&sortBy=submittedDate&sortOrder=descending&max_results=10"
# By arXiv ID curl --fail --silent --show-error --max-time 30 "https://export.arxiv.org/api/query?id_list=2402.03300" # Multiple papers curl --fail --silent --show-error --max-time 30 "https://export.arxiv.org/api/query?id_list=2402.03300,2401.12345,2403.00001"
Retrieve structured metadata using `--json`, retaining the complete versioned ID. Register the inspected abstract/PDF URL and title with `evidence-citations`, then export its escaped `--style bibtex` entries. For article-specific author/year/ primaryClass fields, use only returned metadata and a BibTeX serializer; leave missing fields absent rather than inventing a category or author. Compile the actual bibliography and verify rendered URLs and unresolved-citation warnings.
After finding a paper, read it:
# Abstract page (fast, metadata + abstract) web_extract(urls=["https://arxiv.org/abs/2402.03300"]) # Full paper (PDF → markdown via Firecrawl) web_extract(urls=["https://arxiv.org/pdf/2402.03300"])
For local PDF processing, see `office-artifacts`. Use an available extraction tool; these examples do not imply a Firecrawl account or full-text entitlement.
| Category | Field | |----------|-------| | `cs.AI` | Artificial Intelligence | | `cs.CL` | Computation and Language (NLP) | | `cs.CV` | Computer Vision | | `cs.LG` | Machine Learning | | `cs.CR` | Cryptography and Security | | `stat.ML` | Machine Learning (Statistics) | | `math.OC` | Optimization and Control | | `physics.comp-ph` | Computational Physics |
Full list: https://arxiv.org/category_taxonomy
The `scripts/search_arxiv.py` script handles XML parsing and provides clean output:
python3 "<skill-dir>/scripts/search_arxiv.py" "GRPO reinforcement learning" python3 "<skill-dir>/scripts/search_arxiv.py" "transformer attention" --max 10 --sort date python3 "<skill-dir>/scripts/search_arxiv.py" --author "Yann LeCun" --max 5 python3 "<skill-dir>/scripts/search_arxiv.py" --category cs.AI --sort date python3 "<skill-dir>/scripts/search_arxiv.py" --id 2402.03300 python3 "<skill-dir>/scripts/search_arxiv.py" --id 2402.03300,2401.12345
No dependencies — uses only Python stdlib.
---
arXiv doesn't provide citation data or recommendations. The optional **Semantic Scholar API** provides JSON. Check current endpoint access, authentication and rate-limit responses; do not assume a fixed public or API-key quota.
# By arXiv ID curl --fail --silent --show-error --max-time 30 "https://api.semantic
Repo: Prismer-AI/PrismerCloud
Turn a known bug into a tight, red-capable reproducer, then prove the reproducer locks that…
Review a diff against its acceptance criteria in four segments (convention adherence, bug…
Five-dimension design audit (frontend UI/UX · server data-model & flow · endpoint spec ·…
Before merge, mechanize Documentation-First — derive the code delta from git diff, then…
Diagnose the local dev machine before any APC loop step — run apc env doctor, classify each…
Close out a local coding task on the bound daemon — stage, commit, branch, merge, push via…