Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
aaaaqwq avatar

Arxiv Automation

  • 38 installs
  • 82 repo stars
  • Updated August 2, 2026
  • aaaaqwq/claude-code-skills

arxiv-automation is a Claude Code skill that searches and monitors arXiv papers, downloads PDFs, and summarizes abstracts for research workflows.

About

arxiv-automation is a Claude Code skill for searching, monitoring, and analyzing academic papers on arXiv. A developer or researcher uses it to query papers by keyword, author, or category, track new submissions via RSS, and download PDFs to summarize abstracts. It includes a ready Python search function and notes arXiv's rate limits.

  • Searches and monitors arXiv papers by topic, author, or category via the arXiv API
  • Downloads PDFs and extracts/summarizes abstracts for research workflows
  • Provides a Python query helper, CS category table, RSS monitoring feeds, and rate-limit guidance

Arxiv Automation by the numbers

  • 38 all-time installs (skills.sh)
  • Ranked #1,163 of 2,719 Automation & Workflows skills by installs in the Skillselion catalog
  • Data as of Aug 3, 2026 (Skillselion catalog sync)
At a glance

arxiv-automation capabilities & compatibility

free (arXiv API is public)

Capabilities
paper search · paper monitoring · pdf download · abstract summarization
Use cases
research · web search · pdf parsing
Pricing
Free
From the docs

What arxiv-automation says it does

Search and monitor arXiv papers. Query by topic, author, or category. Track new papers, download PDFs, and summarize abstracts for research workflows.
SKILL.md
arXiv API: max 1 request per 3 seconds
SKILL.md
npx skills add https://github.com/aaaaqwq/claude-code-skills --skill arxiv-automation

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs38
repo stars82
Last updatedAugust 2, 2026
Repositoryaaaaqwq/claude-code-skills

What it does

A researcher uses it to search, monitor, and summarize arXiv papers for research workflows.

Who is it for?

Searching, monitoring, and summarizing arXiv papers in research workflows

Skip if: Non-arXiv literature sources or tasks unrelated to academic papers

When should I use this skill?

Searching or monitoring arXiv papers, tracking new submissions, or summarizing abstracts

What you get

  • arXiv search results
  • monitored new submissions
  • downloaded PDFs

By the numbers

  • arXiv API rate limit: max 1 request per 3 seconds
  • CS category table with 5 categories (cs.AI, cs.CL, cs.LG, cs.CV, cs.SE)

Files

SKILL.mdMarkdownGitHub ↗

arXiv Automation

Search, monitor, and analyze academic papers from arXiv.

Capabilities

  • Search papers by keyword, author, category
  • Monitor new submissions in specific categories
  • Download PDFs for analysis
  • Extract and summarize abstracts
  • Track citation-worthy papers

Usage

Search Papers (arXiv API)

import urllib.request, urllib.parse, xml.etree.ElementTree as ET

def search_arxiv(query, max_results=10):
    base_url = "http://export.arxiv.org/api/query?"
    params = urllib.parse.urlencode({
        "search_query": query,
        "start": 0,
        "max_results": max_results,
        "sortBy": "submittedDate",
        "sortOrder": "descending"
    })
    url = base_url + params
    response = urllib.request.urlopen(url).read()
    root = ET.fromstring(response)
    ns = {"atom": "http://www.w3.org/2005/Atom"}
    papers = []
    for entry in root.findall("atom:entry", ns):
        papers.append({
            "title": entry.find("atom:title", ns).text.strip(),
            "summary": entry.find("atom:summary", ns).text.strip()[:200],
            "link": entry.find("atom:id", ns).text,
            "published": entry.find("atom:published", ns).text,
            "authors": [a.find("atom:name", ns).text for a in entry.findall("atom:author", ns)]
        })
    return papers

# Example: search for LLM agent papers
papers = search_arxiv("all:LLM AND all:agent", max_results=5)
for p in papers:
    print(f"{p['title']}\n  {p['link']}\n  {', '.join(p['authors'][:3])}\n")

Monitor Categories

Common CS categories:

CategoryDescription
cs.AIArtificial Intelligence
cs.CLComputation and Language (NLP)
cs.LGMachine Learning
cs.CVComputer Vision
cs.SESoftware Engineering

RSS feeds: http://arxiv.org/rss/{category} (e.g., http://arxiv.org/rss/cs.AI)

Download PDF

# arXiv ID format: 2401.12345
arxiv_id = "2401.12345"
pdf_url = f"https://arxiv.org/pdf/{arxiv_id}.pdf"

Rate Limits

  • arXiv API: max 1 request per 3 seconds
  • Be respectful of arXiv's resources
  • Use RSS feeds for monitoring (less load than API queries)

Integration

Combine with pdf skill for PDF text extraction and analysis. Combine with rss-automation for periodic monitoring of new papers.

Related skills

FAQ

How does it search arXiv?

It uses the arXiv API (export.arxiv.org/api/query) with a Python helper that queries by search terms and sorts by submitted date, returning titles, summaries, links, and authors.

What are the rate limits?

The arXiv API allows at most 1 request per 3 seconds; the skill recommends using RSS feeds for monitoring to reduce load.

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.