docsforge — Crawler and Adaptive Harvest for Technical Documents like Code Language , API , Technologies. Search -> Identify URL -> Crawl -> Harvest Documentation
-
Updated
Aug 31, 2026 - Python
docsforge — Crawler and Adaptive Harvest for Technical Documents like Code Language , API , Technologies. Search -> Identify URL -> Crawl -> Harvest Documentation
Turn any documentation portal (GitBook, Mintlify, Docusaurus, ReadTheDocs) into LLM-ready Markdown, vector RAG JSONL, llms.txt, and styled offline PDFs. Zero-config CLI, desktop GUI & FastMCP server.
A documentation scraper for developers which deep crawls any website and creates a markdown context file for that site
通用高可用离线文档抓取与转换系统 (Generic High-Availability Documentation Scraper & Converter)
Python web scraper for BetterDocs documentation sites. Auto-discovers content and exports to JSON, Markdown & CSV. Easy to configure.
A reusable Python scraper framework for efficiently crawling documentation websites and preparing content for AI training, with async support and structured output.
To associate your repository with the documentation-scraper topic, visit your repo's landing page and select "manage topics."