Libraries tagged by crawlora
laurentvw/scrapher
2537 Downloads
A web scraper for PHP to easily extract data from web pages
justinholtweb/craft-sanka
26 Downloads
Instant indexing and GEO for Craft CMS — push URLs to Google's Indexing API and IndexNow the moment they change, then manage, log and audit how AI search engines read your site.
justinholtweb/craft-blackhole
23 Downloads
A honeypot for bad bots — a hidden trap link that robots.txt tells crawlers to stay away from, and a permanent block for the ones that follow it anyway.
jayanta/laravel-ai-guard
15 Downloads
Protect your Laravel app from AI scrapers, LLM crawlers, and prompt injection attacks
itgalaxy/webcrawler-verifier
2992 Downloads
PHP library providing functionality to verify that user-agents are who they claim to be.
gstjohn/thumbsnag
29 Downloads
Thumbsnag crawls an HTML document and finds imagery that best represents the given page.
ghalambaz/googleplay-spider
235 Downloads
Another GooglePlay Website Scraper based on goutte
dimtrovich/user-agent
178 Downloads
A PHP desktop/mobile user agent parser with bot detection, based on Mobiledetect and CrawlerDetect.
diggin/diggin-robotrules
187 Downloads
parser/handler for Robots Exclusion Protocol (robots.txt and more)
cobaia/marsvin
113 Downloads
Marsvin it's a framework to write crawler applications
cable8mm/water-melon
2767 Downloads
Water Melon is simple melon.com api sdk for php
brunodebarros/http-request
207 Downloads
A web crawler, built to be easily dropped into PHP applications, which behaves just like a regular web browser, interpreting location redirects and storing cookies automatically.
blogdaren/webman-phpcreeper
358 Downloads
PHPCreeper plugin for webman
bleuren/agent
109 Downloads
Enhanced user agent parser for Laravel
angeo/module-robots-txt-aeo
19 Downloads
Magento 2 module for AI Engine Optimization (AEO). Manages AI crawler rules in robots.txt (OAI-SearchBot, GPTBot, ChatGPT-User, PerplexityBot, Google-Extended, ClaudeBot, Claude-SearchBot, Applebot-Extended, CCBot and more) without overwriting your existing configuration, emits IETF Content-Usage, Cloudflare Content-Signal and RSL License directives, and verifies that a crawler is genuine via Web Bot Auth request signatures (RFC 9421) or vendor-published IP ranges.