NanmiCoder/MediaCrawler
#206 this week
小红书笔记 | 评论爬虫、抖音视频 | 评论爬虫、快手视频 | 评论爬虫、B 站视频 | 评论爬虫、微博帖子 | 评论爬虫、百度贴吧帖子 | 百度贴吧评论回复爬虫 | 知乎问答文章|评论爬虫
56.8K stars
11.4K forks
56.8K GitHub watchers
Updated 7/18/2026
Backblaze Generative Media Hackathon
Build the next generation of AI media apps with Genblaze, stored on Backblaze B2. $10,000 in prizes.
Loading star history...
Use Cases & Benefits
- Provides a multi-platform media data crawler for collecting public posts and comments from major Chinese social media platforms using browser automation.
- Eliminates complex JS reverse engineering by leveraging logged-in browser contexts and JS expressions to obtain signature parameters, lowering technical barriers.
- Use for extracting keyword-based posts and comments from platforms like Xiaohongshu, Douyin, and Bilibili for social media analysis.
- Use for archiving specified post IDs and their comment threads including nested replies across multiple platforms for research or monitoring.
- Use for saving crawled data into CSV, JSON, or lightweight SQLite/MySQL databases to facilitate data management and further processing.