
MCP-Web-Extractor
MCP -Server, der Webinhalte mithilfe lesbarkeit.js extrahiert
1
Github Watches
0
Github Forks
0
Github Stars
MCP Web Extractor
A Model Context Protocol (MCP) server that extracts web content using Readability.js. This tool fetches web pages and extracts the main content, making it ideal for saving clean, readable versions of articles to Obsidian notes.
Features
- Extracts readable content from any URL
- Removes ads, sidebars, and other distractions
- Returns clean text along with metadata (title, excerpt, etc.)
- Easy integration with Obsidian via MCP
Installation
# Clone the repository
git clone https://github.com/iemong/mcp-web-extractor.git
cd mcp-web-extractor
# Install dependencies
npm install
# Build the project
npm run build
# Start the server
npm start
The server will start on http://localhost:3000 with the MCP endpoint at http://localhost:3000/mcp.
Usage
As a standalone service
You can use the included client example to extract content from a URL:
ts-node-esm client-example.ts
With Obsidian
The obsidian-integration.ts
file provides an example of how to integrate this MCP server with Obsidian. You can use it as a starting point for creating an Obsidian plugin that extracts web content.
API
The MCP server provides the following capability:
-
extract-content
: Extracts readable content from a given URL- Parameters:
{ url: string }
- Returns:
{ title, content, textContent, excerpt, siteName }
- Parameters:
License
MIT
相关推荐
I find academic articles and books for research and literature reviews.
Confidential guide on numerology and astrology, based of GG33 Public information
Emulating Dr. Jordan B. Peterson's style in providing life advice and insights.
Your go-to expert in the Rust ecosystem, specializing in precise code interpretation, up-to-date crate version checking, and in-depth source code analysis. I offer accurate, context-aware insights for all your Rust programming questions.
Advanced software engineer GPT that excels through nailing the basics.
Converts Figma frames into front-end code for various mobile frameworks.
Take an adjectivised noun, and create images making it progressively more adjective!
Entdecken Sie die umfassendste und aktuellste Sammlung von MCP-Servern auf dem Markt. Dieses Repository dient als zentraler Hub und bietet einen umfangreichen Katalog von Open-Source- und Proprietary MCP-Servern mit Funktionen, Dokumentationslinks und Mitwirkenden.
Die All-in-One-Desktop & Docker-AI-Anwendung mit integriertem Lappen, AI-Agenten, No-Code-Agent Builder, MCP-Kompatibilität und vielem mehr.
Fair-Code-Workflow-Automatisierungsplattform mit nativen KI-Funktionen. Kombinieren Sie visuelles Gebäude mit benutzerdefiniertem Code, SelbstHost oder Cloud, 400+ Integrationen.
🧑🚀 全世界最好的 llm 资料总结(数据处理、模型训练、模型部署、 O1 模型、 MCP 、小语言模型、视觉语言模型) | Zusammenfassung der weltbesten LLM -Ressourcen.
Reviews

user_ln2muBCY
The mcp-web-extractor by iemong is truly a game-changer. It offers a seamless experience for extracting data from any web page. The comprehensive documentation on its GitHub page is incredibly helpful and the tool's efficiency is unmatched. Highly recommend for anyone in need of a robust web extraction solution!