Libraries tagged by text-extract
xberg-io/xberg
227 Downloads
High-performance document intelligence library
mayaram/laravel-ocr
4397 Downloads
Laravel OCR & Document Data Extractor - A powerful OCR and document parsing engine for Laravel
iamgerwin/php-pdf-to-markdown-parser
16157 Downloads
A lightweight PHP library to convert PDF documents into clean, structured Markdown. Supports text extraction, headings, lists, tables, diagrams and code blocks for easier content reuse and publishing.
oxide/pdf-oxide
30 Downloads
The fastest PHP PDF library — 0.8ms mean, 5× faster than the industry leaders, 100% pass rate on 3,830 real-world PDFs. Text extraction, Markdown/HTML conversion, and PDF creation over a Rust core via FFI.
silverstripe/textextraction
189688 Downloads
Text Extraction API for SilverStripe CMS (mostly used with 'fulltextsearch' module)
cryde/json-text-extractor
9141 Downloads
Helper that will extract JSON from plain text
tamirrental/laravel-text-extractor
17900 Downloads
A Laravel package for extracting structured data from documents via OCR APIs. Ships with Koncile AI provider.
daniel-jorg-schuppelius/php-pdf-toolkit
3884 Downloads
PHP 8.2+ library for PDF text extraction with automatic reader selection. Supports embedded text and scanned documents via OCR.
jcfrane/pdf-text-extractor
349 Downloads
A Laravel PDF text extraction package with multiple strategies (PdfParser, XObject, AWS Textract, Tesseract OCR). Handles Canva-generated PDFs, scanned documents, and other edge cases with automatic fallback.
moinul/laravel-pdf-to-html
202 Downloads
A Laravel package to convert PDF files to HTML using poppler-utils
keyvan/german-ocr
0 Downloads
High-performance German document OCR - Local & Cloud API
carrooi/docx-extractor
2263 Downloads
DOCX text extractor
bpmore/readability-core
50 Downloads
Readability and plain-language analysis for English text: extraction, formulas, rules and findings. Framework-agnostic, so one engine can serve more than one product.
metolabs/text-extractor-php
127 Downloads
PHP wrapper for extracting text from PDF documents and images
manofstrong/sitescrapper
71 Downloads
A Package to Scrape Websites from their Sitemaps and Extract Relevant Content from the Webpage and Upload to a Database