wistbean / wistbean/learn_python3_spider
爬取 20w 表情包
Open
Nobody has claimed this yet.
- Dominant language
- Python
- Stars
- 22.1k
- Forks
- 3.9k
- PR merge metrics
- No merged PRs in 30d
Description
无法爬取所有页面的表情包,下载几百个表情包后程序停止。代码用的是博主的源代码,爬取的页码为1-200页。已加请求头
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
No file or test is named. Start by locating and running the crawler entry point for pages 1–200, then reproduce the stop after several hundred downloads and inspect the request and pagination behavior. Done means the crawler completes the requested page range or clearly reports any pages it cannot fetch.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- data-engineering
- Issue type
- Bug
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100