codestates / codestates/ds-TIL

[TIL] 최근후_201214

Open
#1,164 0 comments 0 reactions 0 assignees View on GitHub
DSFT01
Dominant language
No language data
Stars
2
Forks
1
PR merge metrics
No merged PRs in 30d

Description

### **키워드**:

자연어처리, 텍스트 토큰화

### **배운 것**:

자연어 처리로 가는 첫번째 단계인 토큰화를 배웠다. 토큰화는 문장을 단어 단위로 쪼개는 것을 말한다. 토큰화는 구두점, 스페이스 등 분석에 필요없는 것들을 떼어 낸 뒤 단어들로 쪼갬!

### **느낀 점**:

주피터 노트북을 써서 해보고있는데 역시 처음 배우는 것이라 과제하는데 좀 오래걸린다!

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue contains a Korean TIL entry about natural-language processing and text tokenization, but names no file, test, or entry point. It does not specify a requested change or what completion would look like, so review the repository's contribution guidance before determining whether any documentation work remains.

Written by the indexing model from the issue text.

Assessment

Tech stack
jupyter-notebook
Domain
documentation, machine-learning
Issue type
Documentation
Difficulty
1/5
Estimated time
Under an hour
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
15/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.