geekcomputers / geekcomputers/Python
Google search result scraper
未关闭
还没有人认领这个 Issue。
- 主要语言
- Python
- 星标
- 35.4k
- 派生
- 12.9k
- 平均合并
- 2 小时 37 分钟
- 30 天内合并 PR
- 1
描述
I've noticed I can't seem to scrape the results anymore.
On my debug console:
soup.select('.r a') = []
An empty list. Despite the fact the results are under the class .r with links in the a element.
Anyone else having trouble?
贡献指南
从这里开始
- 先读完整个 Issue,再读项目的贡献指南。
- 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
- Fork 仓库,在一个分支上完成修改。
- 提交 Pull Request,并在描述里引用这个 Issue 编号。
调研方向
首先定位调用 soup.select('.r a') 的 scraper 代码,并针对 Google 搜索结果复现空结果行为。确定当前的结果链接选择器,并验证 scraper 能够成功提取链接;issue 中未提供文件路径或测试路径。
由索引模型根据 Issue 内容生成。
评估
- 技术栈
- python
- 领域
- search
- Issue 类型
- 缺陷
- 难度
- 3/5
- 预计耗时
- 1-2 天
- 活跃度
- 停滞
- 描述清晰度
- 需要澄清
- 新手友好度
- 25/100