
Arxiv Automation
- 38 installs
- 82 repo stars
- Updated August 2, 2026
- aaaaqwq/claude-code-skills
arxiv-automation is a Claude Code skill that searches and monitors arXiv papers, downloads PDFs, and summarizes abstracts for research workflows.
About
arxiv-automation is a Claude Code skill for searching, monitoring, and analyzing academic papers on arXiv. A developer or researcher uses it to query papers by keyword, author, or category, track new submissions via RSS, and download PDFs to summarize abstracts. It includes a ready Python search function and notes arXiv's rate limits.
- Searches and monitors arXiv papers by topic, author, or category via the arXiv API
- Downloads PDFs and extracts/summarizes abstracts for research workflows
- Provides a Python query helper, CS category table, RSS monitoring feeds, and rate-limit guidance
Arxiv Automation by the numbers
- 38 all-time installs (skills.sh)
- Ranked #1,163 of 2,719 Automation & Workflows skills by installs in the Skillselion catalog
- Data as of Aug 3, 2026 (Skillselion catalog sync)
arxiv-automation capabilities & compatibility
free (arXiv API is public)
- Capabilities
- paper search · paper monitoring · pdf download · abstract summarization
- Use cases
- research · web search · pdf parsing
- Pricing
- Free
What arxiv-automation says it does
Search and monitor arXiv papers. Query by topic, author, or category. Track new papers, download PDFs, and summarize abstracts for research workflows.
arXiv API: max 1 request per 3 seconds
npx skills add https://github.com/aaaaqwq/claude-code-skills --skill arxiv-automationAdd your badge
Show developers this skill is listed on Skillselion. Paste this into your README.
| Installs | 38 |
|---|---|
| repo stars | ★ 82 |
| Last updated | August 2, 2026 |
| Repository | aaaaqwq/claude-code-skills ↗ |
What it does
A researcher uses it to search, monitor, and summarize arXiv papers for research workflows.
Who is it for?
Searching, monitoring, and summarizing arXiv papers in research workflows
Skip if: Non-arXiv literature sources or tasks unrelated to academic papers
When should I use this skill?
Searching or monitoring arXiv papers, tracking new submissions, or summarizing abstracts
What you get
- arXiv search results
- monitored new submissions
- downloaded PDFs
By the numbers
- arXiv API rate limit: max 1 request per 3 seconds
- CS category table with 5 categories (cs.AI, cs.CL, cs.LG, cs.CV, cs.SE)
Files
arXiv Automation
Search, monitor, and analyze academic papers from arXiv.
Capabilities
- Search papers by keyword, author, category
- Monitor new submissions in specific categories
- Download PDFs for analysis
- Extract and summarize abstracts
- Track citation-worthy papers
Usage
Search Papers (arXiv API)
import urllib.request, urllib.parse, xml.etree.ElementTree as ET
def search_arxiv(query, max_results=10):
base_url = "http://export.arxiv.org/api/query?"
params = urllib.parse.urlencode({
"search_query": query,
"start": 0,
"max_results": max_results,
"sortBy": "submittedDate",
"sortOrder": "descending"
})
url = base_url + params
response = urllib.request.urlopen(url).read()
root = ET.fromstring(response)
ns = {"atom": "http://www.w3.org/2005/Atom"}
papers = []
for entry in root.findall("atom:entry", ns):
papers.append({
"title": entry.find("atom:title", ns).text.strip(),
"summary": entry.find("atom:summary", ns).text.strip()[:200],
"link": entry.find("atom:id", ns).text,
"published": entry.find("atom:published", ns).text,
"authors": [a.find("atom:name", ns).text for a in entry.findall("atom:author", ns)]
})
return papers
# Example: search for LLM agent papers
papers = search_arxiv("all:LLM AND all:agent", max_results=5)
for p in papers:
print(f"{p['title']}\n {p['link']}\n {', '.join(p['authors'][:3])}\n")Monitor Categories
Common CS categories:
| Category | Description |
|---|---|
| cs.AI | Artificial Intelligence |
| cs.CL | Computation and Language (NLP) |
| cs.LG | Machine Learning |
| cs.CV | Computer Vision |
| cs.SE | Software Engineering |
RSS feeds: http://arxiv.org/rss/{category} (e.g., http://arxiv.org/rss/cs.AI)
Download PDF
# arXiv ID format: 2401.12345
arxiv_id = "2401.12345"
pdf_url = f"https://arxiv.org/pdf/{arxiv_id}.pdf"Rate Limits
- arXiv API: max 1 request per 3 seconds
- Be respectful of arXiv's resources
- Use RSS feeds for monitoring (less load than API queries)
Integration
Combine with pdf skill for PDF text extraction and analysis. Combine with rss-automation for periodic monitoring of new papers.
Related skills
FAQ
How does it search arXiv?
It uses the arXiv API (export.arxiv.org/api/query) with a Python helper that queries by search terms and sorts by submitted date, returning titles, summaries, links, and authors.
What are the rate limits?
The arXiv API allows at most 1 request per 3 seconds; the skill recommends using RSS feeds for monitoring to reduce load.