wistbean / wistbean/learn_python3_spider

爬取 20w 表情包

Open
#57 1 comment 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
22.1k
Forks
3.9k
PR merge metrics
No merged PRs in 30d

Description

无法爬取所有页面的表情包,下载几百个表情包后程序停止。代码用的是博主的源代码,爬取的页码为1-200页。已加请求头

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

No file or test is named. Start by locating and running the crawler entry point for pages 1–200, then reproduce the stop after several hundred downloads and inspect the request and pagination behavior. Done means the crawler completes the requested page range or clearly reports any pages it cannot fetch.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
data-engineering
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.