fiveoceans-dev/pdf2htmlEX
Convert PDF to HTML without losing text or format.
Scrapy, a fast high-level screen scraping and web crawling framework for Python.
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
Convert PDF to HTML without losing text or format.
PDF Reader in JavaScript
A JavaScript visualization library for HTML and SVG.
Crawl-Anywhere - Web Crawler and document processing pipeline with Solr integration.