Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
aaaaqwq avatar

Arxiv Automation

  • 7 installs
  • 82 repo stars
  • Updated August 2, 2026
  • aaaaqwq/agi-super-skills

arxiv-automation is a Claude Code skill that searches, monitors, and summarizes academic papers from arXiv via the arXiv API.

About

arxiv-automation is a Claude Code skill for searching, monitoring, and analyzing academic papers on arXiv. It queries the arXiv API by keyword, author, or category, monitors new submissions via category RSS feeds, downloads PDFs, and extracts and summarizes abstracts. The SKILL.md provides a Python search function and notes the arXiv rate limit of one request per three seconds. It suggests pairing with the pdf and rss-automation skills.

  • Searches arXiv by keyword, author, or category via the arXiv API
  • Monitors new submissions through category RSS feeds and downloads PDFs
  • Documents the arXiv rate limit of 1 request per 3 seconds

Arxiv Automation by the numbers

  • 7 all-time installs (skills.sh)
  • Ranked #1,590 of 2,719 Automation & Workflows skills by installs in the Skillselion catalog
  • Data as of Aug 3, 2026 (Skillselion catalog sync)
At a glance

arxiv-automation capabilities & compatibility

Free; uses the public arXiv API with a 1-request-per-3-seconds limit.

Capabilities
research · paper search · pdf parsing
Use cases
research · web search · pdf parsing
Pricing
Free
From the docs

What arxiv-automation says it does

Search, monitor, and analyze academic papers from arXiv.
SKILL.md
arXiv API: max 1 request per 3 seconds
SKILL.md
npx skills add https://github.com/aaaaqwq/agi-super-skills --skill arxiv-automation

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs7
repo stars82
Last updatedAugust 2, 2026
Repositoryaaaaqwq/agi-super-skills

What it does

A developer or researcher uses this skill when they want to search, monitor, and summarize arXiv papers programmatically.

Who is it for?

Searching and monitoring arXiv papers and summarizing their abstracts for research.

Skip if: High-frequency scraping, since arXiv limits requests to one per three seconds.

When should I use this skill?

The user wants to find or track academic papers on a topic, author, or category.

What you get

Ranked paper results with titles, authors, links, and summarized abstracts.

By the numbers

  • rate limit 1 request per 3 seconds
  • 5 common CS categories documented (cs.AI, cs.CL, cs.LG, cs.CV, cs.SE)

Files

SKILL.mdMarkdownGitHub ↗

arXiv Automation

Search, monitor, and analyze academic papers from arXiv.

Capabilities

  • Search papers by keyword, author, category
  • Monitor new submissions in specific categories
  • Download PDFs for analysis
  • Extract and summarize abstracts
  • Track citation-worthy papers

Usage

Search Papers (arXiv API)

import urllib.request, urllib.parse, xml.etree.ElementTree as ET

def search_arxiv(query, max_results=10):
    base_url = "http://export.arxiv.org/api/query?"
    params = urllib.parse.urlencode({
        "search_query": query,
        "start": 0,
        "max_results": max_results,
        "sortBy": "submittedDate",
        "sortOrder": "descending"
    })
    url = base_url + params
    response = urllib.request.urlopen(url).read()
    root = ET.fromstring(response)
    ns = {"atom": "http://www.w3.org/2005/Atom"}
    papers = []
    for entry in root.findall("atom:entry", ns):
        papers.append({
            "title": entry.find("atom:title", ns).text.strip(),
            "summary": entry.find("atom:summary", ns).text.strip()[:200],
            "link": entry.find("atom:id", ns).text,
            "published": entry.find("atom:published", ns).text,
            "authors": [a.find("atom:name", ns).text for a in entry.findall("atom:author", ns)]
        })
    return papers

# Example: search for LLM agent papers
papers = search_arxiv("all:LLM AND all:agent", max_results=5)
for p in papers:
    print(f"{p['title']}\n  {p['link']}\n  {', '.join(p['authors'][:3])}\n")

Monitor Categories

Common CS categories:

CategoryDescription
cs.AIArtificial Intelligence
cs.CLComputation and Language (NLP)
cs.LGMachine Learning
cs.CVComputer Vision
cs.SESoftware Engineering

RSS feeds: http://arxiv.org/rss/{category} (e.g., http://arxiv.org/rss/cs.AI)

Download PDF

# arXiv ID format: 2401.12345
arxiv_id = "2401.12345"
pdf_url = f"https://arxiv.org/pdf/{arxiv_id}.pdf"

Rate Limits

  • arXiv API: max 1 request per 3 seconds
  • Be respectful of arXiv's resources
  • Use RSS feeds for monitoring (less load than API queries)

Integration

Combine with pdf skill for PDF text extraction and analysis. Combine with rss-automation for periodic monitoring of new papers.

Related skills

FAQ

What is the rate limit?

The arXiv API allows at most one request per three seconds; RSS feeds are recommended for monitoring.

Can it summarize papers?

Yes, it extracts and summarizes abstracts and can download PDFs, pairing with the pdf skill for extraction.

Automation & Workflowsresearchautomation

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.