
Arxiv Automation
- 7 installs
- 82 repo stars
- Updated August 2, 2026
- aaaaqwq/agi-super-skills
arxiv-automation is a Claude Code skill that searches, monitors, and summarizes academic papers from arXiv via the arXiv API.
About
arxiv-automation is a Claude Code skill for searching, monitoring, and analyzing academic papers on arXiv. It queries the arXiv API by keyword, author, or category, monitors new submissions via category RSS feeds, downloads PDFs, and extracts and summarizes abstracts. The SKILL.md provides a Python search function and notes the arXiv rate limit of one request per three seconds. It suggests pairing with the pdf and rss-automation skills.
- Searches arXiv by keyword, author, or category via the arXiv API
- Monitors new submissions through category RSS feeds and downloads PDFs
- Documents the arXiv rate limit of 1 request per 3 seconds
Arxiv Automation by the numbers
- 7 all-time installs (skills.sh)
- Ranked #1,590 of 2,719 Automation & Workflows skills by installs in the Skillselion catalog
- Data as of Aug 3, 2026 (Skillselion catalog sync)
arxiv-automation capabilities & compatibility
Free; uses the public arXiv API with a 1-request-per-3-seconds limit.
- Capabilities
- research · paper search · pdf parsing
- Use cases
- research · web search · pdf parsing
- Pricing
- Free
What arxiv-automation says it does
Search, monitor, and analyze academic papers from arXiv.
arXiv API: max 1 request per 3 seconds
npx skills add https://github.com/aaaaqwq/agi-super-skills --skill arxiv-automationAdd your badge
Show developers this skill is listed on Skillselion. Paste this into your README.
| Installs | 7 |
|---|---|
| repo stars | ★ 82 |
| Last updated | August 2, 2026 |
| Repository | aaaaqwq/agi-super-skills ↗ |
What it does
A developer or researcher uses this skill when they want to search, monitor, and summarize arXiv papers programmatically.
Who is it for?
Searching and monitoring arXiv papers and summarizing their abstracts for research.
Skip if: High-frequency scraping, since arXiv limits requests to one per three seconds.
When should I use this skill?
The user wants to find or track academic papers on a topic, author, or category.
What you get
Ranked paper results with titles, authors, links, and summarized abstracts.
By the numbers
- rate limit 1 request per 3 seconds
- 5 common CS categories documented (cs.AI, cs.CL, cs.LG, cs.CV, cs.SE)
Files
arXiv Automation
Search, monitor, and analyze academic papers from arXiv.
Capabilities
- Search papers by keyword, author, category
- Monitor new submissions in specific categories
- Download PDFs for analysis
- Extract and summarize abstracts
- Track citation-worthy papers
Usage
Search Papers (arXiv API)
import urllib.request, urllib.parse, xml.etree.ElementTree as ET
def search_arxiv(query, max_results=10):
base_url = "http://export.arxiv.org/api/query?"
params = urllib.parse.urlencode({
"search_query": query,
"start": 0,
"max_results": max_results,
"sortBy": "submittedDate",
"sortOrder": "descending"
})
url = base_url + params
response = urllib.request.urlopen(url).read()
root = ET.fromstring(response)
ns = {"atom": "http://www.w3.org/2005/Atom"}
papers = []
for entry in root.findall("atom:entry", ns):
papers.append({
"title": entry.find("atom:title", ns).text.strip(),
"summary": entry.find("atom:summary", ns).text.strip()[:200],
"link": entry.find("atom:id", ns).text,
"published": entry.find("atom:published", ns).text,
"authors": [a.find("atom:name", ns).text for a in entry.findall("atom:author", ns)]
})
return papers
# Example: search for LLM agent papers
papers = search_arxiv("all:LLM AND all:agent", max_results=5)
for p in papers:
print(f"{p['title']}\n {p['link']}\n {', '.join(p['authors'][:3])}\n")Monitor Categories
Common CS categories:
| Category | Description |
|---|---|
| cs.AI | Artificial Intelligence |
| cs.CL | Computation and Language (NLP) |
| cs.LG | Machine Learning |
| cs.CV | Computer Vision |
| cs.SE | Software Engineering |
RSS feeds: http://arxiv.org/rss/{category} (e.g., http://arxiv.org/rss/cs.AI)
Download PDF
# arXiv ID format: 2401.12345
arxiv_id = "2401.12345"
pdf_url = f"https://arxiv.org/pdf/{arxiv_id}.pdf"Rate Limits
- arXiv API: max 1 request per 3 seconds
- Be respectful of arXiv's resources
- Use RSS feeds for monitoring (less load than API queries)
Integration
Combine with pdf skill for PDF text extraction and analysis. Combine with rss-automation for periodic monitoring of new papers.
Related skills
FAQ
What is the rate limit?
The arXiv API allows at most one request per three seconds; RSS feeds are recommended for monitoring.
Can it summarize papers?
Yes, it extracts and summarizes abstracts and can download PDFs, pairing with the pdf skill for extraction.