18F/doc_processing_toolkit
Python library to extract text from PDF, and default to OCR when text extraction fails.
Generates and cleans up temp files for automated tests
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
Python library to extract text from PDF, and default to OCR when text extraction fails.
18F's contributions to the GSA enterprise data inventory and public data listing
WIP: a gem to cycle through environment variables
A complete agency API program.