Built something? We create video reels & spotlights for GitHub projects.Promote your project →
Public GitHub Catalog Discovered Sep 27, 2026

rossmounce / ocr

Batch process old PDF files into text using Tesseract OCR

View repository on GitHub ↗ View creator profile Browse directory

About this discovery

This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.

#9776170GitHub System ID
rossmounceOrganization / User
PublicVisibility
ActiveCatalog Status

More from rossmounce

↗

rossmounce/elife-flickr

An adaption of my workflow for BMC journals. eLife figure numbering is *really* annoying. Can't simply iterate. Need to fix that at some point if I'm to embed the correct caption in the correct figure image.

Discovered
↗

rossmounce/Trying-beautiful-soup

Trying out beautiful soup in an IPython notebook, on BMC journal article HTML with a view to parsing out figure captions

Discovered
FOR MAINTAINERS

Built something? Put it in front of millions of developers.

We make a short reel about your project and post it across YouTube, Instagram, Threads, and X. Send a link, we do the rest.