video-sum-mcp
Extracts video and live streaming content from Bilibili, enabling knowledge graph generation.
Extracts short video content from Douyin (TikTok China) with context-aware processing.
Extracts social media posts from Xiaohongshu with OCR support for image text recognition.
Extracts Q&A content from Zhihu platform for structured analysis.
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@video-sum-mcpsummarize and extract knowledge graph from this Bilibili video BV1xx411c7mD"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
Video Content Summarization MCP Server
A Model Context Protocol (MCP) server that extracts content from multiple video platforms and generates intelligent knowledge graphs.
Features
š Multi-Platform Support
Douyin (TikTok China) - Short video content extraction
Bilibili - Video and live streaming content
Xiaohongshu (Little Red Book) - Social media posts with OCR support
Zhihu - Q&A platform content
⨠Advanced Capabilities
OCR Text Recognition - Extract text from images using PaddleOCR
Knowledge Graph Generation - Intelligent content structuring
Chinese Content Optimization - Specialized processing for Chinese text
Context-Aware Extraction - Smart content understanding and quality control
Related MCP server: MediaCrawler MCP Server
Installation
Prerequisites
Python 3.8 or higher
Anaconda (recommended for dependency management)
Setup
Clone the repository:
git clone https://github.com/fakad/video-sum-mcp.git
cd video-sum-mcpCreate and activate conda environment:
conda create -n vsc python=3.8
conda activate vscInstall dependencies:
pip install -r requirements.txtConfiguration
For Claude Desktop
Add this configuration to your Claude Desktop config file:
macOS: ~/Library/Application Support/Claude/claude_desktop_config.json
Windows: %APPDATA%/Claude/claude_desktop_config.json
{
"mcpServers": {
"video-sum-mcp": {
"command": "python",
"args": ["/path/to/video-sum-mcp/main.py"],
"cwd": "/path/to/video-sum-mcp",
"env": {
"CONDA_DEFAULT_ENV": "vsc"
}
}
}
}For Other MCP Clients
The server can be started directly:
python main.pyUsage
Basic Video Processing
# Example: Process a Bilibili video
result = process_video(
url="https://www.bilibili.com/video/BV1234567890",
output_format="markdown"
)Supported URL Formats
Douyin:
https://v.douyin.com/...or full URLsBilibili:
https://www.bilibili.com/video/...Xiaohongshu:
https://www.xiaohongshu.com/discovery/item/...Zhihu:
https://www.zhihu.com/question/...
Context-Enhanced Processing
For platforms with anti-crawling measures, you can provide context:
result = process_video(
url="https://...",
context_text="Additional context information..."
)Features in Detail
OCR Integration
Automatic image text extraction from Xiaohongshu posts
PaddleOCR for accurate Chinese character recognition
Batch processing for multiple images
Knowledge Graph Generation
Structured content analysis
Intelligent relationship mapping
Quality control and validation
Anti-Crawling Strategies
Smart fallback mechanisms
Context-based extraction
User guidance for optimal results
Development
Project Structure
video-sum-mcp/
āāā core/ # Core functionality modules
ā āāā extractors/ # Platform-specific extractors
ā āāā processors/ # Content processing logic
ā āāā knowledge_graph/ # Knowledge graph generation
ā āāā managers/ # Resource management
āāā scripts/ # MCP server implementation
āāā main.py # Main entry point
āāā requirements.txt # Python dependencies
āāā pyproject.toml # Project configurationRunning Tests
python -m pytestDependencies
Key dependencies include:
bilibili-api-python- Bilibili API integrationyt-dlp- Video downloading capabilitiesPaddleOCR- OCR text recognitionbeautifulsoup4- Web scrapingrequests- HTTP requests
See requirements.txt for complete list.
Contributing
Fork the repository
Create a feature branch (
git checkout -b feature/amazing-feature)Commit your changes (
git commit -m 'Add some amazing feature')Push to the branch (
git push origin feature/amazing-feature)Open a Pull Request
License
This project is licensed under the MIT License - see the LICENSE file for details.
Acknowledgments
Built using the Model Context Protocol
OCR powered by PaddleOCR
Platform integrations using various open-source APIs
This server cannot be deployed
Maintenance
Related MCP Connectors
Video analytics for TikTok, Instagram, and YouTube. Track, analyze, and discover content.
Social media data: 85 tools across 11 platforms (YouTube, TikTok, Instagram, X & more), one key.
Xiaohongshu, Douyin, TikTok, YouTube, X links to text: transcript, on-screen text, images described
Find viral outlier posts on TikTok, Instagram and YouTube, pull creator stats, and crawl on demand.
Related MCP Servers
- AlicenseBqualityFmaintenanceEnables users to search and retrieve content from Xiaohongshu (Red Book) platform with smart search capabilities and rich data extraction including note content, author information, and images.152 npm28MIT
- FlicenseNot gradedqualityDmaintenanceEnables AI assistants to crawl and extract data from Chinese social media platforms like Bilibili, Xiaohongshu, and Douyin. Provides search, content detail retrieval, and creator information tools with persistent browser sessions and QR code login support.77-
- AlicenseAqualityCmaintenanceEnables AI clients to search Xiaohongshu notes by brand and category, batch extract comments, and perform keyword/sentiment/heat analysis, with results exported as Excel and JSON reports.4MIT
- AlicenseNot gradedqualityCmaintenanceEnables local-first ingestion of Douyin links, favorites, and public Bilibili videos into Markdown, timestamped transcripts, OCR, visual assets, and deduplicated indexes for downstream RAG workflows.1MIT