跳到主内容
智客 ZICQ

技能库 智客分类:文档办公 firecrawl-knowledge-base

Firecrawl Knowledge Base

与 Firecrawl 从网络内容建立知识库. 用于本地参考文件、 RAG 已准备好的块、 微调数据集、 文档镜像、 主题 Corpora 或 LLM 已准备好的从网络来源整理的标记.

31617 安装量

官方网址:作者主页

技能介绍

先看中文介绍;官方 description 原文单独保留,不改写 SKILL.md。

做什么

与 Firecrawl 从网络内容建立知识库. 用于本地参考文件、 RAG 已准备好的块、 微调数据集、 文档镜像、 主题 Corpora 或 LLM 已准备好的从网络来源整理的标记.

何时用

官方 description 未单独写出 Use when。按规范,代理会在用户任务与这段 description 的关键词匹配时激活本技能。

代理如何加载

按 Agent Skills 渐进披露:启动时只加载 name 与 description(约 100 token);任务匹配后才读入整份 SKILL.md 正文;scripts/、references/、assets/ 仅在需要时再读。 本文件正文结构:Firecrawl Knowledge Base、Onboarding Interview、Firecrawl Collection Plan、Parallel Work、Output Modes、Final Deliverable。 其中含规范建议的小节:边界情况。

文件分析

文件分析:这是一份仅含 SKILL.md 的指令型技能,代理激活后整份正文进入上下文。

官方 description(原文)

Build a knowledge base from web content with Firecrawl. Use for local reference docs, RAG-ready chunks, fine-tuning datasets, documentation mirrors, topic corpora, or LLM-ready markdown organized from web sources.

Firecrawl Knowledge BaseOnboarding InterviewFirecrawl Collection PlanParallel WorkOutput ModesFinal DeliverableKnowledge Base: [Source]SummaryOutput StructureCoverageUsage NotesSources

· 许可:ISC

来源分类:skills.sh agent-skill

SKILL.md 与 Agent 调用

官方规范 ↗
name
firecrawl-knowledge-base
description
Build a knowledge base from web content with Firecrawl. Use for local reference docs, RAG-ready chunks, fine-tuning datasets, documentation mirrors, topic corpora, or LLM-ready markdown organized from web sources.
许可
ISC
  1. 发现技能客户端向 Agent 提供名称与描述目录。
  2. 匹配与调用用户指定或任务匹配后,载入 SKILL.md 指令。
  3. 按需加载按步骤读取参考文档、使用脚本与素材。

具体调用语法与可用工具以目标 Agent 客户端为准。 查看调用机制说明 ↗

安装这个技能

Skills CLI ↗

先选择目标 Agent 和安装范围,保留技能包的附属文件,安装后检查客户端能否发现该技能。

交给 Agent 安装

复制安装指令给支持 Agent Skills 的代理,确认其中的目标目录与客户端匹配。

把 Agent Skill「firecrawl-knowledge-base」安装到我的项目:SKILL.md 原文与官方 description 见 https://zicq.com/zh/skills/skl-f1e90271af314afb-Firecrawl-Knowledge-Base.html
请存为 .cursor/skills/firecrawl-knowledge-base/SKILL.md 或 .claude/skills/firecrawl-knowledge-base/SKILL.md,frontmatter 的 name 与 description 保持原样,不要改写。

GitHub 完整包 ↗

终端安装 · Skills CLI

需要 Node.js 与 npx。先查看仓库技能列表,确认实际名称。

npx skills add 'https://github.com/firecrawl/firecrawl-workflows' --list

npx skills add 'https://github.com/firecrawl/firecrawl-workflows' --skill 'firecrawl-knowledge-base'

CLI 会交互选择目标 Agent,默认安装到项目;用户级安装使用 -g。先通过查看命令核对仓库内容,再用 npx skills list 检查已安装技能。

阅读排版
--- name: firecrawl-knowledge-base description: Build a knowledge base from web content with Firecrawl. Use for local reference docs, RAG-ready chunks, fine-tuning datasets, documentation mirrors, topic corpora, or LLM-ready markdown organized from web sources. license: ISC metadata: author: firecrawl version: "0.1.0" homepage: https://www.firecrawl.dev source: https://github.com/firecrawl/firecrawl-workflows inputs: - name: FIRECRAWL_API_KEY description: Firecrawl API key for hosted Firecrawl requests. required: true --- # Firecrawl Knowledge Base Use this to turn URLs or topics into organized LLM-ready content. ## Onboarding Interview Infer the source, goal, depth, and output location from context. If the source and goal are clear, proceed immediately. Ask at most 1-3 concise questions only if blocked, such as the source URL/topic, whether the output is reference/RAG/training/docs, or training format if training is requested. ## Firecrawl Collection Plan Use Firecrawl map for documentation sites, search for topic-based corpora, scrape pages into markdown, and preserve code examples and tables. For files, follow the Firecrawl download-style convention: ```text .firecrawl/ / / index.md ``` ## Parallel Work If appropriate, use sub-agents or equivalent parallel task runners: - one docs section per researcher - official docs, tutorials, community discussions, and references by source type - source scraping vs chunk generation vs manifest generation ## Output Modes - Reference: markdown files, `index.md`, and `sources.json`. - RAG: markdown files plus chunk files and `manifest.json`. - Training: scraped source files plus `training-data.jsonl` and `training-metadata.json`. - Docs mirror: complete markdown mirror with a table of contents. ## Final Deliverable ```markdown # Knowledge Base: [Source] ## Summary [What was collected and why] ## Output Structure [Files/directories created] ## Coverage [Sections, source types, counts] ## Usage Notes [How to use in RAG, docs, training, or agent context] ## Sources [URLs collected] ## Rerun Inputs workflow: firecrawl-knowledge-base source: [url/topic] goal: [reference/rag/train/docs] depth: [quick/thorough/exhaustive] output_dir: [.firecrawl/] ``` ## Quality Bar - Preserve code examples and formatting. - Remove boilerplate navigation where possible. - Include source URLs in frontmatter or metadata.

相关技能

文档办公

Ontology

为结构化的代理内存和可堆肥技能所打入的知识图. 在创建/征服实体(Person, project, Task, Evention, Document)时使用,链接相关对象,强制约束,规划多步动作作为图变,或技能需要共享状态时使用. 触发到"记住","我知道什么","链接X到Y",…

文档办公

Nano Pdf

使用纳米-pdf CLI编辑带有自然语言指令的PDF.

文档办公

Word / DOCX

创建,检查,并编辑有可靠样式的Microsoft Word文档和DOCX文件,编号,跟踪更改,表格,章节,并进行相容性检查. 当 (1) 任务涉及 Word 或 ".docx " 时使用; (2) 文件包括跟踪的更改,评论,字段,表格,模板,或页面布局限制; (3) 文档必须在不…

文档办公

Baidu Search

使用Baidu AI搜索引擎(BDSE)搜索网页. 用于实时信息、文件或研究专题.