PDFText
JSON →pdftext is a Python library designed for fast and accurate extraction of structured text from PDF documents. It focuses on efficiently parsing text, detecting elements like tables and links, and handling complex layouts. The current version is 0.6.3, and it's actively maintained with frequent minor releases addressing bug fixes and introducing new features.
Traffic · last 30 days stale · no recent hits
total hits 18
actors 6 distinct systems
last hit 18d ago ByteDance
top countries 🇺🇸 United States · 🇨🇦 Canada · 🇩🇪 Germany · 🇸🇬 Singapore
API endpoints
full doc /v1/registry/pdftext
install /v1/registry/pdftext/install
compatibility /v1/registry/pdftext/compatibility