NanmiCoder / NanmiCoder/MediaCrawler
[问题] 爬取失败
Open
Nobody has claimed this yet.
question
- Dominant language
- Python
- Stars
- 65.3k
- Forks
- 12.6k
- PR merge metrics
- No merged PRs in 30d
Description
⚠️ 提交前确认
- 我已经仔细阅读了项目使用过程中的常见问题汇总
- 我已经搜索并查看了已关闭的issues
- 我确认这不是由于滑块验证码、Cookie过期、Cookie提取错误、平台风控等常见原因导致的问题
❓ 问题描述
标准模式在最低限制下触发xhs安全机制,cdp静默模式扫码登录爬取成功但本地没有data文件夹,cdp关闭静默模式浏览器扫码登录但显示浏览器请求超时。🔍 使用场景
关键词搜索- 目标平台: (如:小红书/抖音/微博等) xhs
- 使用功能: (如:关键词搜索/用户主页爬取等) 关键词搜索
💻 环境信息
- 操作系统: win11
- Python版本: 3.13.5
- 是否使用IP代理: 否
- 是否使用VPN翻墙软件:否
- 目标平台(抖音/小红书/微博等):小红书
📋 错误日志
在此粘贴完整的错误日志
📷 错误截图
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
No source file, test, or complete error log is identified. Reproduce the xhs keyword-search flow on Windows 11 with Python 3.13.5, comparing standard mode with CDP silent and non-silent login, then capture complete logs and verify where output is written. Done means the failure is isolated and the expected local data output or an environmental blocker is established.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- backend
- Issue type
- Bug
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 18/100