firecrawl-research-papers

作者: firecrawl

使用 Firecrawl 查找並綜合研究論文、白皮書、PDF、技術報告及學術來源。適用於用戶需要文獻回顧、論文摘要、研究現狀分析,或從 PDF 及學術/行業出版物中獲取有來源的綜合資訊時。

npx skills add https://github.com/firecrawl/firecrawl-workflows --skill firecrawl-research-papers

Firecrawl Research Papers

Use this to create a sourced literature review.

Onboarding Interview

Infer the topic, source constraints, target count, and output format from context. If the topic is clear, proceed immediately.

Ask at most 1-3 concise questions only if blocked, such as the topic, target paper count, or required venue/date/method constraints.

Firecrawl Collection Plan

Use Firecrawl Research through the CLI, MCP, or equivalent Firecrawl tool surface as the primary path for paper discovery and verification. Fall back to general Firecrawl search and scrape for whitepapers, technical reports, research blogs, leaderboards, or facts outside the paper corpus.

What the paper index holds: paper abstracts, with full text reachable per paper. Its largest share is biomedical and life-science literature — PubMed journal articles plus bioRxiv and medRxiv preprints — so clinical, drug, gene, disease, epidemiology, and public-health questions are in scope. arXiv preprints cover computer science, physics, and mathematics. Coverage outside those sources is thinner, and the web tools below are the fallback there.

Core tools:

  • MCP: firecrawl_research_search_papers(query, k?) CLI: firecrawl research search-papers <query> [--k <number>] Semantic search over paper abstracts. Start here for most paper-finding queries, and retry with alternate framing when results are thin or too narrow.
  • MCP: firecrawl_research_related_papers(seed_ids, intent, mode?, k?) CLI: firecrawl research related-papers <seedIds...> --intent <intent> [--mode <similar|citers|references>] [--k <number>] Expand from strong seed papers into similar work, citing papers, or references. Use this to find the relevant paper family, not just the first matching result.
  • MCP: firecrawl_research_inspect_paper(id) CLI: firecrawl research inspect-paper <id> Fetch canonical metadata for a candidate paper: title, abstract, authors, categories, source ids, and dates.
  • MCP: firecrawl_research_read_paper(id, question) CLI: firecrawl research read-paper <id> --question <question> Verify a specific claim or constraint inside one paper, such as method, reported score, benchmark, affiliation, comparison, or limitation.
  • MCP: firecrawl_search(query) / firecrawl_scrape(url) CLI: firecrawl search <query> / firecrawl scrape <url> Use for web-only context: benchmark leaderboards, rankings, reports, whitepapers, research blogs, and source pages outside the paper index.

Not the paper index, despite the name: passing categories: ["research"] to firecrawl_search (CLI firecrawl search <query> --categories research) filters an ordinary web search to research-affiliated websites — the list includes PubMed, bioRxiv, medRxiv, arXiv, and publisher sites — and returns page results from them. It reaches those sites' web pages; what it does not do is query their paper records in the index above, so there is no abstract search, no related-paper or citation-graph expansion, no canonical paper metadata, and no in-body passages. Use it when a web search is what you want and those sites should be weighted in the same call; use the firecrawl_research_* tools for paper work.

Match the approach to the query:

  • Single named paper: run one paper search, then inspect or read the paper if metadata or body verification is needed.
  • Paper by description, method, or topic family: search for strong anchors, then expand with related papers and keep close neighbors.
  • Enumeration queries, such as papers that do a task or benchmark a method: search multiple framings, expand several strong anchors, and re-seed from newly found relevant papers.
  • Papers that use or exhibit a property: start from the defining paper or strongest anchor, expand via similar, citers, or references, and use read-paper to verify the property.
  • Superlatives and leaderboards: use general web search or scrape to find the ranking, then map top entries back to papers with paper search.
  • Author, organization, venue, date, or methodology constraints: verify with inspect-paper metadata or read-paper before keeping a candidate.

Target source types:

  • biomedical and life-science literature from PubMed, with bioRxiv and medRxiv preprints for work that has not appeared in a journal yet
  • arXiv preprints in computer science, physics, and mathematics
  • academic papers from university sites and ACM/IEEE pages where accessible
  • industry reports and whitepapers
  • company research blogs
  • technical articles and conference summaries

Principles:

  • When in doubt, include the relevant paper family rather than only the single best result.
  • Use related-paper expansion to avoid stopping at one strong hit.
  • Use read-paper to verify load-bearing constraints, not to summarize every candidate.
  • Drop only clearly off-topic papers.

Parallel Work

If appropriate, use sub-agents or equivalent parallel task runners:

  • Academic Papers researcher
  • Biomedical and Life Sciences researcher, for PubMed journal articles and bioRxiv/medRxiv preprints on a clinical, drug, gene, disease, epidemiology, or public-health topic
  • Industry Reports researcher
  • Technical Articles researcher
  • Synthesis and citation reviewer

Split by source or sub-topic, not by tool. Give each researcher the same paper tools and let the topic decide which part of the corpus answers.

Final Deliverable

# Literature Review: [Topic]

## Abstract
[2-3 paragraph summary]

## Key Papers
[Title, authors, source URL, key findings, methodology, relevance]

## Themes And Consensus
[What sources agree on]

## Open Questions And Debates
[Disagreements and unresolved questions]

## Emerging Trends
[Recent developments]

## Sources
[Organized by paper/report/article]

## Rerun Inputs
workflow: firecrawl-research-papers
topic: [topic]
target_count: [number]
output: [markdown/brief]

Quality Bar

  • Every major claim should trace to a source.
  • Note inaccessible or failed PDFs.
  • Distinguish peer-reviewed work from blogs and vendor reports.

來自 firecrawl 的更多技能

firecrawl-research-index
firecrawl
使用 Firecrawl Research 進行語義搜尋、語義與結構擴展以及內文驗證,找出能回答研究查詢的論文。對於任何文獻查找或論文檢索任務(無論是單篇論文查詢還是完整的多論文集合),一律使用此技能。
data-analysisresearchweb-scraping
oracle
firecrawl
使用 oracle CLI 的最佳實踐(提示與檔案捆綁、引擎、會話及檔案附加模式)。
pinecone
firecrawl
專為生產級AI應用設計的受管向量資料庫。全受管、自動擴展,具備混合搜尋(密集+稀疏)、元資料過濾與命名空間功能。…
wpds
firecrawl
在構建利用WordPress設計系統(WPDS)及其組件、標記、模式等的用戶界面時使用。
audiocraft-audio-generation
firecrawl
用於音訊生成的 PyTorch 函式庫,包含文字轉音樂(MusicGen)和文字轉音效(AudioGen)。當需要從文字生成音樂時使用…
skypilot-multi-cloud-orchestration
firecrawl
針對機器學習工作負載的多雲編排,具備自動成本優化功能。當您需要跨多個雲端執行訓練或批次作業、利用…時使用。
firecrawl-seo-audit
firecrawl
使用 Firecrawl 審查網站的 SEO。適用於使用者要求進行 SEO 審查、中繼資料與標題檢視、網站地圖/網站結構分析、關鍵字機會、競爭對手 SERP 比較,或優先搜尋最佳化建議時。
data-analysisresearchweb-scraping
gh-issues
firecrawl
擷取 GitHub 問題,生成子代理來實作修復並開啟 PR,然後監控並處理 PR 審查意見。用法:/gh-issues [owner/repo] [--label…