skills/43-wentorai-research-plugins/skills/research/funding/figshare-api/SKILL.md
Research data sharing and repository
npx skillsauth add brycewang-stanford/Awesome-Agent-Skills-for-Empirical-Research figshare-apiInstall this skill globally with one command. Works with Claude Code, Cursor, and Windsurf.
3 of 9 scanners reported clean
Some scanners were skipped, did not run, or reported a non-clean status. Review each row below.
Figshare is a cloud-based research data management platform that allows researchers to store, share, and discover research outputs including datasets, figures, media, papers, posters, and fileset collections. Every item uploaded to Figshare receives a citable DOI and is stored in a FAIR-compliant manner, making research outputs findable, accessible, interoperable, and reusable.
The Figshare API provides comprehensive programmatic access to the repository, enabling researchers and institutions to automate data publishing, integrate with research workflows, and build custom discovery interfaces. The platform supports versioning, embargo periods, and flexible access controls for both public and private research data.
Researchers, data managers, institutional repository administrators, and research infrastructure developers use the Figshare API to automate deposit workflows, harvest metadata for institutional dashboards, build data discovery tools, and integrate research data management into existing laboratory information management systems. Figshare serves over 150 institutions worldwide and hosts millions of research outputs.
Authentication via personal access token is required for write operations and accessing private content. Read access to public content is available without authentication but has lower rate limits.
Authorization headercurl -H "Authorization: token YOUR_FIGSHARE_TOKEN" "https://api.figshare.com/v2/account/articles"
Public endpoints can be accessed without a token for browsing published content.
Search the public Figshare repository for published articles (datasets, figures, papers, media, and other item types).
GET https://api.figshare.com/v2/articles| Parameter | Type | Required | Description |
|----------------|--------|----------|--------------------------------------------------------|
| search_for | string | No | Free-text search query |
| item_type | int | No | Item type filter (1=figure, 2=media, 3=dataset, etc.) |
| published_since| string | No | Filter by date (YYYY-MM-DD format) |
| order | string | No | Sort: published_date, modified_date, views |
| order_direction| string | No | asc or desc |
| page | int | No | Page number (default 1) |
| page_size | int | No | Results per page (default 10, max 1000) |
curl -X POST "https://api.figshare.com/v2/articles/search" \
-H "Content-Type: application/json" \
-d '{"search_for": "genomics CRISPR", "item_type": 3, "page_size": 5}'
id, title, doi, url, published_date, description, defined_type_name, categories, tags, authors, files (with download URLs), and citation.Retrieve and manage dataset-specific content in Figshare. Datasets are a specialized article type with additional support for large file collections.
GET https://api.figshare.com/v2/articles/{article_id}| Parameter | Type | Required | Description | |------------|------|----------|------------------------------------| | article_id | int | Yes | The Figshare article/dataset ID |
# Get a specific dataset by ID
curl "https://api.figshare.com/v2/articles/12345678"
# List files in a dataset
curl "https://api.figshare.com/v2/articles/12345678/files"
id, title, doi, description, authors, categories, tags, files (array with name, size, download_url, computed_md5), license, version, is_embargoed, and custom_fields.Rate limits vary based on authentication status and endpoint. Authenticated requests generally allow up to 100 requests per minute. Unauthenticated requests are limited to approximately 10 requests per minute. The API returns HTTP 429 with a Retry-After header when limits are exceeded. For bulk data harvesting, Figshare provides OAI-PMH endpoints at https://api.figshare.com/v2/oai which are more suitable for large-scale metadata collection.
Find publicly available datasets matching specific research topics:
import requests
payload = {
"search_for": "single cell RNA-seq",
"item_type": 3, # datasets only
"page_size": 20,
"order": "published_date",
"order_direction": "desc"
}
resp = requests.post("https://api.figshare.com/v2/articles/search", json=payload)
results = resp.json()
for item in results:
print(f"{item['title']}")
print(f" DOI: {item['doi']}")
print(f" Published: {item['published_date']}")
print()
Automate data deposit for reproducible research workflows:
import requests
TOKEN = os.environ["FIGSHARE_API_TOKEN"]
headers = {"Authorization": f"token {TOKEN}", "Content-Type": "application/json"}
# Step 1: Create a new article
article_data = {
"title": "Supplementary Data for Analysis of Gene Expression",
"defined_type": "dataset",
"description": "RNA-seq counts and metadata for the analysis.",
"tags": ["RNA-seq", "gene expression"],
"categories": [69] # Genetics category
}
resp = requests.post("https://api.figshare.com/v2/account/articles",
headers=headers, json=article_data)
article_url = resp.json()["location"]
# Step 2: Upload file
article = requests.get(article_url, headers=headers).json()
print(f"Created article ID: {article['id']}, DOI will be assigned on publish")
Collect metadata from all Figshare items in an institution's repository:
curl "https://api.figshare.com/v2/oai?verb=ListRecords&metadataPrefix=oai_dc&set=institution_123"
tools
Recommend AND run open-source AI tools, agents, Claude Code / Codex skills, and MCP servers for any stage of a literature review — searching, reading, extracting, synthesizing, screening, citation-checking, and paper writing. Use when the user asks "what tool should I use to..." OR "install/run/use <tool> to ..." for research/lit-review work: automating a survey or related-work section, PDF→Markdown extraction for LLMs (MinerU/marker/docling), PRISMA / systematic review (ASReview), citation-backed Q&A over PDFs (PaperQA2), wiring papers into Claude/Cursor via MCP (arxiv/paper-search/zotero servers), or chatting with a Zotero library. Ships a launcher (scripts/litrun.py) that installs each tool in an isolated venv and runs it. Curated catalog of 70+ vetted projects. 支持中英文(用于「文献综述工具选型」与「一键安装/运行」)。
development
Route empirical-research requests through the Auto-Empirical Research Skills catalog when this whole repository is installed as one skill in Codex, CodeBuddy, Claude Code, or another IDE. Use to choose and load the right vendored AERS skill for causal inference, econometrics, replication, data acquisition, manuscript writing, peer review and referee responses, citation checking, de-AIGC editing, or full empirical-paper workflows without reading the entire repository at once.
documentation
Use when the project collects primary data or runs a field, lab, or survey experiment, before the intervention begins — write the pre-analysis plan, size the sample from a power calculation, and register with the AEA RCT Registry. Apply after the design is chosen in aer-identification and before any outcome data are seen.
tools
Guide economists to authoritative data sources with explicit, confirmed data specifications before retrieval; interfaces with Playwright MCP to navigate portals and extract real data, not articles about data.