๐Ÿ“š
DataPython

ArXiv Search

by WILLOSCAR

ArXiv Search is a Data skill for Claude Code, published by WILLOSCAR in research-units-pipeline-skills.

499 stars39 forkson WILLOSCAR/research-units-pipeline-skillsAdded 2026/08/26Repository updated 2026/08/26
claudeclaude-codecodexgptpipelineresearchresearch-paperresearch-projectresearch-toolskillstoolsunitsvibevibe-codingvibecoding
Install in seconds
Install ArXiv Search
Copy ArXiv Search into your Claude Code skills folder. Run the command in your terminal, or review the source on GitHub before installing.
terminal
npx degit https://github.com/WILLOSCAR/research-units-pipeline-skills/tree/main/.codex/skills/arxiv-search ~/.claude/skills/arxiv-search

Requires Node.js. Downloads this skill only โ€” not the rest of the repository โ€” into your Claude Code skills folder.

Without Node.js

git clone https://github.com/WILLOSCAR/research-units-pipeline-skills.git

Clones the whole repository, then copy the skillโ€™s own directory into your skills folder yourself.

In this catalog

Source file
.codex/skills/arxiv-search/SKILL.md in WILLOSCAR/research-units-pipeline-skills
Installs to
~/.claude/skills/arxiv-search
Collection
One of 25 skills cataloged from this repository
Category
Data โ€” 668 skills

What ArXiv Search does

ArXiv Search retrieves paper metadata from arXiv keyword queries and writes normalized results to JSONL. Use it to build an initial paper set for ranking, taxonomy, or citation workflows, or to import and enrich offline exports.

ArXiv Search is cataloged under Data on DirSkills. ArXiv Search comes from a repository tagged claude, claude-code, codex, gpt and pipeline.

Documentation

README

arXiv Search (metadata-first)

Collect an initial paper set with enough metadata to support downstream ranking, taxonomy building, and citation generation.

When online, prefer rich arXiv metadata (categories, arxiv_id, pdf_url, published/updated, etc.). When offline, accept an export and convert it cleanly.

Load Order

Always read:

  • references/domain_pack_overview.md โ€” how domain packs drive topic-specific behavior

Domain packs (loaded by topic match):

  • assets/domain_packs/llm_agents.json โ€” pinned IDs, query rewrite rules for LLM agent topics

This is the opening of the README. Read the full README on GitHub.

Frequently asked about ArXiv Search

  • What else does WILLOSCAR publish alongside ArXiv Search?

    ArXiv Search is one of 25 skills that DirSkills catalogs from WILLOSCAR/research-units-pipeline-skills, the repository it ships in. Its siblings there include Agent Survey Corpus, Anchor Sheet and Appendix Table Writer. Each one is a separate skill with its own page in this directory, installs the same way ArXiv Search does, and is maintained by WILLOSCAR in that same repository. The rest of the collection is listed on the WILLOSCAR/research-units-pipeline-skills page.

  • How does ArXiv Search compare to other Data skills?

    ArXiv Search ranks #605 by stars among the 668 Data skills in this catalog. The most-starred ones next to it are Benchmark Methodology, Jupyter Notebook and Solana. DirSkills ranks by the star count of the repository each skill ships in, so that order reflects how popular those repositories are rather than any review of ArXiv Search against them. Open each page to compare what they document and how they install.

More from WILLOSCAR/research-units-pipeline-skills

ArXiv Search is one of 25 skills cataloged on DirSkills from WILLOSCAR/research-units-pipeline-skills.

See all 25 skills โ†’
๐Ÿ“š
5d ago

Agent Survey Corpus

Agent Survey Corpus downloads open-access arXiv survey PDFs about agentic systems and extracts text for local reference. Use it to study how real surveys organize sections, subsections, and evidence-backed comparisons.
Writing
49939
๐Ÿช
5d ago

Anchor Sheet

Anchor Sheet extracts subsection-specific evidence hooks from draft packs, focusing on numbers, benchmarks, metrics, and limitations. Use it when you want sections to be grounded in cited facts instead of generic summaries.
Writing
49939
๐Ÿ“‹
5d ago

Appendix Table Writer

Appendix Table Writer curates reader-facing survey tables for paper appendices using only in-scope evidence and existing citation keys. Use it when you have evidence packs, anchor sheets, and citations and need publishable tables instead of internal logs.
Writing
49939
๐Ÿงพ
5d ago

Argument Selfloop

Argument Selfloop builds a final machine-readable snapshot from current H3 sections, refreshes section hashes, and checks consistency before merge. Use it after section edits when you need an updated argument ledger and manifest.
Writing
49939
๐Ÿงพ
5d ago

Artifact Contract Auditor

Artifact Contract Auditor checks whether completed Units and the locked Pipeline have all required output files present, then writes output/CONTRACT_REPORT.md. Use it for mid-run coverage snapshots or final delivery completeness, not provenance checks.
Quality
49939
๐Ÿ“„
5d ago

Beamer Compile QA

Beamer Compile QA compiles `latex/slides/main.tex` into `latex/slides/main.pdf` and writes `output/SLIDES_BUILD_REPORT.md`. Use it when a Beamer tutorial deck needs a build check and a readable success or failure report.
Quality
49939