Skill metadata
Reference: full SKILL.md
The following is the complete skill definition that Mibyan loads when this skill is triggered. This is what the agent sees as instructions when the skill is active.
arXiv Research
Search and retrieve academic papers from arXiv via their free REST API. No API key, no dependencies — just curl.Quick Reference
Searching Papers
The API returns Atom XML. Parse withgrep/sed or pipe through python for clean output.
Basic search
Clean output (parse XML to readable format)
Search Query Syntax
Boolean operators
Sort and Pagination
Fetching Specific Papers
BibTeX Generation
After fetching metadata for a paper, generate a BibTeX entry: {% raw %}Reading Paper Content
After finding a paper, read it:ocr-and-documents skill.
Common Categories
Full list: https://arxiv.org/category_taxonomy
Helper Script
Thescripts/search_arxiv.py script handles XML parsing and provides clean output:
Semantic Scholar (Citations, Related Papers, Author Profiles)
arXiv doesn’t provide citation data or recommendations. Use the Semantic Scholar API for that — free, no key needed for basic use (1 req/sec), returns JSON.Get paper details + citations
Get citations OF a paper (who cited it)
Get references FROM a paper (what it cites)
Search papers (alternative to arXiv search, returns JSON)
Get paper recommendations
Author profile
Useful Semantic Scholar fields
title, authors, year, abstract, citationCount, referenceCount, influentialCitationCount, isOpenAccess, openAccessPdf, fieldsOfStudy, publicationVenue, externalIds (contains arXiv ID, DOI, etc.)
Complete Research Workflow
- Discover:
python scripts/search_arxiv.py "your topic" --sort date --max 10 - Assess impact:
curl -s "https://api.semanticscholar.org/graph/v1/paper/arXiv:ID?fields=citationCount,influentialCitationCount" - Read abstract:
web_extract(urls=["https://arxiv.org/abs/ID"]) - Read full paper:
web_extract(urls=["https://arxiv.org/pdf/ID"]) - Find related work:
curl -s "https://api.semanticscholar.org/graph/v1/paper/arXiv:ID/references?fields=title,citationCount&limit=20" - Get recommendations: POST to Semantic Scholar recommendations endpoint
- Track authors:
curl -s "https://api.semanticscholar.org/graph/v1/author/search?query=NAME"
Rate Limits
Notes
- arXiv returns Atom XML — use the helper script or parsing snippet for clean output
- Semantic Scholar returns JSON — pipe through
python -m json.toolfor readability - arXiv IDs: old format (
hep-th/0601001) vs new (2402.03300) - PDF:
https://arxiv.org/pdf/{id}— Abstract:https://arxiv.org/abs/{id} - HTML (when available):
https://arxiv.org/html/{id} - For local PDF processing, see the
ocr-and-documentsskill
ID Versioning
arxiv.org/abs/1706.03762always resolves to the latest versionarxiv.org/abs/1706.03762v1points to a specific immutable version- When generating citations, preserve the version suffix you actually read to prevent citation drift (a later version may substantially change content)
- The API
<id>field returns the versioned URL (e.g.,http://arxiv.org/abs/1706.03762v7)
Withdrawn Papers
Papers can be withdrawn after submission. When this happens:- The
<summary>field contains a withdrawal notice (look for “withdrawn” or “retracted”) - Metadata fields may be incomplete
- Always check the summary before treating a result as a valid paper

