v1.0.0: URL-to-PDF + Doc Site Crawler

- Single URL to PDF via Puppeteer (Chrome headless)
- Doc Site Crawler: crawl entire documentation sites
  - Auto-extract all pages from TOC navigation
  - 20-page batch processing with browser restart (anti-OOM)
  - Output: PDF + Markdown
- Flask-style HTTP server on port 8099
- Deployed at https://pdf.donton.cloud/url2pdf
This commit is contained in:
IT Dog
2026-05-22 14:24:05 +08:00
commit 0d3e40f797
11 changed files with 1789 additions and 0 deletions
+3
View File
@@ -0,0 +1,3 @@
node_modules/
*.pdf
/tmp/