
Skill
firecrawl-knowledge-ingest
ingest knowledge bases with Firecrawl
Description
Ingest public or authenticated knowledge bases and docs portals with Firecrawl browser. Use for JS-heavy docs, login-gated portals, paginated help centers, support knowledge bases, or structured JSON/markdown extraction from documentation sites.
SKILL.md
Firecrawl Knowledge Ingest
Use this when a docs portal needs browser navigation, auth, pagination, or JS rendering.
Onboarding Interview
Infer the portal URL, output format, auth needs, and page limit from context. If the portal is clear, proceed immediately.
Ask at most 1-3 concise questions only if blocked, such as the portal URL, whether authentication is required, or the desired output format.
Firecrawl Collection Plan
Use Firecrawl browser to:
- open the portal and inspect navigation
- identify sections, categories, sidebar links, and article URLs
- follow sidebar navigation, next links, pagination, load-more controls, or search
- scrape article content as markdown
- extract metadata such as title, section, last updated date, author, and tags
Try Firecrawl map as a supplement for public URLs, but use browser navigation for auth-gated or JS-heavy content.
Final Deliverable
# Knowledge Ingest: [Portal]
## Summary
[Pages extracted, sections covered, limitations]
## Output
[JSON/markdown/merged file path or content]
## Sections
[Section names and article counts]
## Failed Or Restricted Pages
[Any access/loading issues]
## Sources
[URLs extracted]
## Rerun Inputs
workflow: firecrawl-knowledge-ingest
url: [portal url]
format: [json/markdown/merged]
max_pages: [number]
JSON Shape
Use source, url, extractedAt, totalArticles, and sections[] with article title, url, section, content, and metadata.
Quality Bar
- Preserve code examples, tables, and formatting.
- Strip nav chrome, headers, and footers.
- Track extraction progress and page failures.
- Respect authentication boundaries.
More skills from the firecrawl-workflows repository
View all 16 skillsfirecrawl-company-directories
extract company directories with Firecrawl
Jun 30AutomationData EngineeringFirecrawlWeb Scrapingfirecrawl-competitive-intel
monitor competitor product changes with Firecrawl
Jun 30Competitive IntelligenceFirecrawlMarketingWeb Scrapingfirecrawl-dashboard-reporting
pull metrics from analytics dashboards
Jun 30AnalyticsDashboardsFirecrawlReporting +1firecrawl-deep-research
conduct deep research with Firecrawl
Jun 30FirecrawlFirecrawl ResearchResearchSummarization +1firecrawl-demo-walkthrough
generate product walkthroughs with Firecrawl
Jun 30AutomationFirecrawlUX DesignWeb Scrapingfirecrawl-knowledge-base
build knowledge bases from web content
Jun 30Data EngineeringFirecrawlKnowledge ManagementSearch
More from Firecrawl
View publishercompetitor-analysis
analyze competitors across features and pricing
web-agent
Apr 17Competitive IntelligenceFirecrawlResearchdeep-research
conduct multi-source deep research
web-agent
Apr 17FirecrawlKnowledge ManagementResearche-commerce
extract product data from e-commerce sites
web-agent
May 15E-commerceFirecrawlWeb Scrapingfinancial-research
pull financial data for public companies
web-agent
May 15FinanceFirecrawlResearchSEC Filingspricing-tracker
track and compare vendor pricing tiers
web-agent
Apr 17Competitive IntelligenceFirecrawlResearchSaaSstructured-extraction
extract structured data from websites
web-agent
May 15Data EngineeringFirecrawlWeb Scraping