html_docs_crawler is an open-source command-line tool that crawls HTML documentation sites and converts their content to Markdown format, correcting internal links in the process. It is designed for developers and technical writers who need to migrate or archive documentation efficiently. The tool is distributed via PyPI and is fully open source under the MIT license.
In the CLI tools & terminal space, html_docs_crawler takes a focused approach. It focuses on automating the conversion of HTML documentation to Markdown with correct internal links for developers and technical writers. html_docs_crawler is an open-source project aimed at developers. html_docs_crawler is open source under the Open Source license. It runs on the command line.
html_docs_crawler first shipped in 2026. Development happens publicly on GitHub with 26 commits in the last 90 days. Among its 5 catalogued features are HTML to Markdown, internal link correction, and documentation crawling.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do