Project overview
PDF Bookmarks reads a printed table of contents from scanned or OCR’d books and writes matching PDF bookmarks and page labels. It corrects front-matter offsets between printed and PDF pages, and can fall back to font-based detection or a readable manual list when parsing is uncertain. OCR options include Tesseract, macOS Vision, and Windows OCR, but the Windows backend is still beta and unverified on real Windows hardware.
Repository facts
- Primary language
- Python
- License
- MIT
- Repository updated
- Jul 7, 2026
- Default branch
- main
Resource types
CLI appGeneral skill
Use cases
Document productivity
Runtime
Command line