arXiv Research
SkillSearchSearch arXiv papers by keyword, author, category, or ID.
Available today. Use it from your connected AI after setup.
No other account needed.
Connect ahel once, and every AI you use reads what you have installed.
Then ask your AI: use the arXiv Research skill
What this skill tells your AI
The instructions your AI receives, as published by hezaohezao/poirot in poirot/backend/agents/skill/builtin_skills/research/arxiv/SKILL.md and read by ahel’s review.
Search and retrieve academic papers from arXiv via their free REST API. No API
key, no dependencies — just bash with curl.
Quick Reference
| Action | Command |
|---|---|
| Search papers | bash("curl -s 'https://export.arxiv.org/api/query?search_query=all:QUERY&max_results=5'") |
| Get specific paper | bash("curl -s 'https://export.arxiv.org/api/query?id_list=2402.03300'") |
| Read abstract | browse_page(url="https://arxiv.org/abs/2402.03300") |
| Read full paper (PDF) | browse_page(url="https://arxiv.org/pdf/2402.03300") |
Searching Papers
The API returns Atom XML. Parse with python3 for clean output.
Basic search
curl -s "https://export.arxiv.org/api/query?search_query=all:GRPO+reinforcement+learning&max_results=5"
Clean output (parse XML to readable format)
curl -s "https://export.arxiv.org/api/query?search_query=all:GRPO+reinforcement+learning&max_results=5&sortBy=submittedDate&sortOrder=descending" | python3 -c "
import sys, xml.etree.ElementTree as ET
ns = {'a': 'http://www.w3.org/2005/Atom'}
root = ET.fromstring(sys.stdin.read())
for entry in root.findall('a:entry', ns):
title = entry.find('a:title', ns).text.strip().replace('\n', ' ')
published = entry.find('a:published', ns).text[:10]
summary = entry.find('a:summary', ns).text.strip()[:200]
link = entry.find('a:id', ns).text
print(f'{published} | {title}')
print(f' {link}')
print(f' {summary}...')
print()
"
Search Query Syntax
| Field | Example |
|---|---|
| All fields | all:transformer |
| Title | ti:attention |
| Author | au:vaswani |
| Abstract | abs:reinforcement |
| Category | cat:cs.CL |
| Combine (AND) | all:transformer+AND+ti:attention |
| Combine (OR) | all:LLM+OR+all:large+language+model |
Common Categories
cs.CL— Computation and Language (NLP)cs.CV— Computer Visioncs.LG— Machine Learningcs.AI— Artificial Intelligencestat.ML— Machine Learning (Stats)physics— Physics
Workflow
- Search by keyword/author/category to find relevant papers
- Read abstracts from search results
- Fetch full abstract page with
browse_pagefor promising papers - Read PDF with
browse_pagefor full content (if needed)
Pitfalls
- Rate limiting: arXiv API asks for 3-second间隔 between requests. Don't hammer it.
- Query too specific:
"diffusion models in computer vision for medical imaging"returns 0 results. Use 2-3 core keywords + category filter. - sortBy=relevance is usually better than
submittedDatefor topical searches; usesubmittedDateonly when user wants chronological order. - PDF parsing:
browse_pageon PDF URLs may return raw text or fail on some papers. Prefer abstract pages for reliable content.
Signals
- GitHub stars
- 220
- Forks
- 19
- Last commit
- Jul 2026
Advanced
- Catalog kind
- skill
- Gateway key
arxiv-hezaohezao- Source
- github.com/hezaohezao/poirot