figure-extractor
Extracts complete paper figures with caption-aware, high-DPI crops and contact-sheet checks.
- Stars
- 2
- Forks
- 0
- Updated
- Updated Jun 29, 2026
Extracts complete paper figures with caption-aware, high-DPI crops and contact-sheet checks.
figure-extractor finds complete figures and images in PDFs, arXiv papers, or HTML, using captions and bounding boxes to avoid cutting off multi-panel content. It crops with PyMuPDF at a chosen DPI and can build a contact sheet for quick visual checking. The project is a shell CLI, not a hosted or in-agent service, so Python, PyMuPDF, and a pip-capable runtime are required.
Resource types
Use cases
Platforms
Back up Evernote or Yinxiang locally and export notes as ENEX files.
Let agents read, edit, comment on, and redline Word files without losing formatting.
Runtime
Public GitHub facts last synced Jul 10, 2026.
Shrink PDFs and Office documents to email-safe sizes while protecting visual quality.
A PubMed reference resolver for citation extraction, lookup, and journal auditing.