skills/43-wentorai-research-plugins/skills/literature/search/biorxiv-api/SKILL.md
Preprint server API for biology and medicine papers
npx skillsauth add brycewang-stanford/Awesome-Agent-Skills-for-Empirical-Research biorxiv-apiInstall this skill globally with one command. Works with Claude Code, Cursor, and Windsurf.
3 of 9 scanners reported clean
Some scanners were skipped, did not run, or reported a non-clean status. Review each row below.
bioRxiv (pronounced "bio-archive") is a free online archive and distribution service for unpublished preprints in the life sciences. Operated by Cold Spring Harbor Laboratory, it provides researchers with immediate access to the latest findings before formal peer review. The bioRxiv API enables programmatic access to preprint metadata, content details, and publication linkage data across biology and medical sciences.
The API serves researchers who need to track emerging research trends, monitor preprint activity in specific subfields, or build automated literature surveillance pipelines. It is particularly valuable for systematic reviewers who want to capture the latest evidence before journal publication, and for bibliometric analysts studying the preprint-to-publication pipeline.
bioRxiv hosts preprints across more than 25 subject areas including neuroscience, genomics, bioinformatics, cell biology, and many more. The API returns structured metadata including titles, authors, abstracts, DOIs, publication dates, and links to corresponding published journal articles when available.
No authentication required. The bioRxiv API is fully open and does not require any API key, token, or registration. All endpoints are publicly accessible without rate limiting restrictions.
Fetch detailed metadata for preprints posted within a specified date range or for a specific server (bioRxiv or medRxiv).
GET https://api.biorxiv.org/details/{server}/{interval}/{cursor}| Parameter | Type | Required | Description |
|------------|--------|----------|--------------------------------------------------|
| server | string | Yes | Server name: biorxiv or medrxiv |
| interval | string | Yes | Date range in YYYY-MM-DD/YYYY-MM-DD format |
| cursor | int | No | Pagination cursor (default 0, increments of 100) |
curl "https://api.biorxiv.org/details/biorxiv/2024-01-01/2024-01-31/0"
doi, title, authors, author_corresponding, date, category, abstract, published (journal DOI if available), and jatsxml link.Look up which preprints have been published in peer-reviewed journals, providing the mapping between preprint DOIs and journal article DOIs.
GET https://api.biorxiv.org/pubs/{server}/{interval}/{cursor}| Parameter | Type | Required | Description |
|------------|--------|----------|--------------------------------------------------|
| server | string | Yes | Server name: biorxiv or medrxiv |
| interval | string | Yes | Date range in YYYY-MM-DD/YYYY-MM-DD format |
| cursor | int | No | Pagination cursor (default 0, increments of 100) |
curl "https://api.biorxiv.org/pubs/biorxiv/2024-01-01/2024-06-30/0"
preprint_doi, published_doi, preprint_title, published_journal, published_date, and preprint_date.No formal rate limits are documented for the bioRxiv API. However, responsible use is expected. Results are paginated at 100 records per request, and the cursor parameter should be incremented to retrieve additional pages. Avoid excessive concurrent requests to ensure availability for all users.
Retrieve the latest preprints and filter by category to track new submissions in your field:
# Fetch recent neuroscience preprints
curl "https://api.biorxiv.org/details/biorxiv/2024-06-01/2024-06-07/0" \
| jq '.collection[] | select(.category == "neuroscience")'
Monitor which preprints in your area have been formally published:
# Check publication status for recent preprints
curl "https://api.biorxiv.org/pubs/biorxiv/2024-01-01/2024-06-30/0" \
| jq '.collection[] | select(.published_doi != "")'
Paginate through all results for a given date range to build a comprehensive alert feed:
import requests
base = "https://api.biorxiv.org/details/biorxiv/2024-06-01/2024-06-07"
cursor = 0
all_preprints = []
while True:
resp = requests.get(f"{base}/{cursor}").json()
records = resp.get("collection", [])
if not records:
break
all_preprints.extend(records)
cursor += 100
print(f"Total preprints retrieved: {len(all_preprints)}")
medrxiv as server parameter)tools
Recommend AND run open-source AI tools, agents, Claude Code / Codex skills, and MCP servers for any stage of a literature review — searching, reading, extracting, synthesizing, screening, citation-checking, and paper writing. Use when the user asks "what tool should I use to..." OR "install/run/use <tool> to ..." for research/lit-review work: automating a survey or related-work section, PDF→Markdown extraction for LLMs (MinerU/marker/docling), PRISMA / systematic review (ASReview), citation-backed Q&A over PDFs (PaperQA2), wiring papers into Claude/Cursor via MCP (arxiv/paper-search/zotero servers), or chatting with a Zotero library. Ships a launcher (scripts/litrun.py) that installs each tool in an isolated venv and runs it. Curated catalog of 70+ vetted projects. 支持中英文(用于「文献综述工具选型」与「一键安装/运行」)。
development
Route empirical-research requests through the Auto-Empirical Research Skills catalog when this whole repository is installed as one skill in Codex, CodeBuddy, Claude Code, or another IDE. Use to choose and load the right vendored AERS skill for causal inference, econometrics, replication, data acquisition, manuscript writing, peer review and referee responses, citation checking, de-AIGC editing, or full empirical-paper workflows without reading the entire repository at once.
documentation
Use when the project collects primary data or runs a field, lab, or survey experiment, before the intervention begins — write the pre-analysis plan, size the sample from a power calculation, and register with the AEA RCT Registry. Apply after the design is chosen in aer-identification and before any outcome data are seen.
tools
Guide economists to authoritative data sources with explicit, confirmed data specifications before retrieval; interfaces with Playwright MCP to navigate portals and extract real data, not articles about data.