codestates / codestates/ds-TIL
[TIL] 최근후_201214
Open
DSFT01
- Dominant language
- No language data
- Stars
- 2
- Forks
- 1
- PR merge metrics
- No merged PRs in 30d
Description
### **키워드**:
자연어처리, 텍스트 토큰화
### **배운 것**:
자연어 처리로 가는 첫번째 단계인 토큰화를 배웠다. 토큰화는 문장을 단어 단위로 쪼개는 것을 말한다. 토큰화는 구두점, 스페이스 등 분석에 필요없는 것들을 떼어 낸 뒤 단어들로 쪼갬!
### **느낀 점**:
주피터 노트북을 써서 해보고있는데 역시 처음 배우는 것이라 과제하는데 좀 오래걸린다!
Contributor guide
No contributing guide indexed for this repository
Research direction
The issue contains a Korean TIL entry about natural-language processing and text tokenization, but names no file, test, or entry point. It does not specify a requested change or what completion would look like, so review the repository's contribution guidance before determining whether any documentation work remains.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- jupyter-notebook
- Domain
- documentation, machine-learning
- Issue type
- Documentation
- Difficulty
- 1/5
- Estimated time
- Under an hour
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 15/100