🔥 Official Firecrawl MCP Server - Adds powerful web scraping and search to Cursor, Claude and any other LLM clients.
- 
            Updated
            Oct 19, 2025 
- JavaScript
🔥 Official Firecrawl MCP Server - Adds powerful web scraping and search to Cursor, Claude and any other LLM clients.
Model Context Protocol (MCP) Server for Graphlit Platform
A fork of Dragnet that also extract author, headline, date, keywords from context, as well as built in metadata extraction all in one package
Readability2 converts HTML to plain text.
Next.js template for seamless PDF parsing using pdf2json and FilePond. Ideal for developers seeking a ready-to-use solution for PDF content extraction in Next.js projects.
Pure ruby implementation of the Boilerpipe content extraction algorithm tuned for online articles
DOM Based Content Extraction via Text Density
Web content extraction using machine learning
🔍 Model Context Protocol (MCP) tool for parsing websites using the Jina.ai Reader
Tool to extracts the text from a web article urls and get frequency words, entities recognition, automatic summary and more
Make PDF Files Accessible, Extract Data from PDF, Convert PDF to HTML, Fill-in PDF Form, Stamp PDF and more...
Benson turns a list of URLs into mp3s of the contents of each web page - take control over your reading backlog!
This repository houses a Python application for extracting YouTube video transcripts and summarizing its content.
Via Text Density Simple Web Crawler With Go
Seize is light Node or Browser web-page content extractor inspired by arc90 readability and Safari Reader
A userscript that adds a button to YouTube video pages for copying the transcript with or without timestamps.
📸 Crawell – 网页图片/正文一键提取、Markdown 转换与批量下载的浏览器扩展,本地化,免费 Crawell browser extension for one-click image & article extraction, Markdown conversion and bulk download – 100 % local processing.
Mobile First Indexing Tool
The Ultimate Web Content Extraction & Conversion Tool for AI/LLM Applications. Convert almost any web content into clean Markdown with intelligent AI processing.
Chrome extension to copy YouTube transcripts with AI-friendly features
Add a description, image, and links to the content-extraction topic page so that developers can more easily learn about it.
To associate your repository with the content-extraction topic, visit your repo's landing page and select "manage topics."