Search results for document extraction

xberg-io/xberg

9329 Favers
227 Downloads

High-performance document intelligence library

Go to Download


causal/extractor

16 Favers
277384 Downloads

Metadata and content analysis service - This extension detects and extracts metadata (EXIF / IPTC / XMP / ...) from potentially thousand different file types (such as MS Word/Powerpoint/Excel documents, PDF and images) and bring them automatically and natively to TYPO3 when uploading assets. Works with built-in PHP functions but takes advantage of Apache Tika and other external tools for enhanced metadata extraction.

Go to Download


iamgerwin/php-pdf-to-markdown-parser

8 Favers
16509 Downloads

A lightweight PHP library to convert PDF documents into clean, structured Markdown. Supports text extraction, headings, lists, tables, diagrams and code blocks for easier content reuse and publishing.

Go to Download


subhashladumor1/laravel-ai-docs

9 Favers
437 Downloads

Laravel AI Document Intelligence & OCR package for Laravel 12 AI SDK. Convert PDF to JSON, extract tables, image to text, Ask PDF with AI, audio transcription and multi-language support using GPT-5.2, Claude and Gemini.

Go to Download


daniel-jorg-schuppelius/php-pdf-toolkit

0 Favers
4139 Downloads

PHP 8.2+ library for PDF text extraction with automatic reader selection. Supports embedded text and scanned documents via OCR.

Go to Download


laravelsmartocr/laravel-smart-ocr

6 Favers
20 Downloads

Laravel OCR package — extract text from images, PDFs, invoices, receipts using Tesseract, Claude Vision, OpenAI GPT-4o. Supports barcode, QR code, table extraction, multi-language OCR, AI cleanup, and document templates. Zero Guzzle dependency.

Go to Download


ges/ocr

0 Favers
551 Downloads

Core document processing services for OCR, classification, extraction, and normalization.

Go to Download


keyvan/german-ocr

114 Favers
0 Downloads

High-performance German document OCR - Local & Cloud API

Go to Download


jcfrane/pdf-text-extractor

2 Favers
355 Downloads

A Laravel PDF text extraction package with multiple strategies (PdfParser, XObject, AWS Textract, Tesseract OCR). Handles Canva-generated PDFs, scanned documents, and other edge cases with automatic fallback.

Go to Download


kreuzberg/kreuzberg

13 Favers
1 Downloads

High-performance document intelligence for PHP. Extract text, metadata, and structured information from PDFs, Office documents, images, and 75 formats. Powered by Rust core for 10-50x speed improvements.

Go to Download


docxtract/php-sdk

0 Favers
11 Downloads

Official PHP SDK for the DocXtract document extraction API

Go to Download


aspose/pdf

2 Favers
868 Downloads

A powerful library for manipulating and converting PDF files.

Go to Download


solution-forest/ai-kit-core

3 Favers
34 Downloads

Core of ai-kit: document ingestion, chunking, extraction, vector retrieval (sqlite-vec + pgvector), and RAG agents built on the official Laravel AI SDK.

Go to Download


avvertix/liteparse-php

0 Favers
40 Downloads

PHP FFI bindings for LiteParse, fast local PDF and document parsing with spatial text extraction

Go to Download


jkudish/laravel-document-extraction

0 Favers
1 Downloads

Laravel-native document text and schema extraction with provenance and AI cost tracking.

Go to Download


Next >>