Splitting with hashlib
未關閉
- 主要語言
- Jupyter Notebook
- 星號
- 25.6k
- 分支
- 12.7k
- PR 合併指標
- PR 指標待擷取
描述
Good time a day ageron, you have a very nice book.
Anyway I have one question from Chapter 2. Which of algorithm given in book is more faster:
**- 1) def split_train_test(data, test_ratio)
- 2) def test_set_check(identifier, test_ratio, id_column, hash=hashlib.md5)
- 3) train_test_split from sklearn**
?
Which function is more save memory?
I would be gratefull if you give comments about this. Thanks for your attention
貢獻指南
這個儲存庫沒有索引到貢獻指南
評估
這個 Issue 還沒有評估資料。