alteryx / alteryx/woodwork

Improve performance of reading data from CSV files

未关闭
#1,309 0 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看
主要语言
Python
星标
155
派生
24
PR 合并指标
30 天内没有已合并 PR

描述

Once #1307 is complete, we should investigate different approaches for speeding up the process of deserializing data from CSV files.

This could possibly be used for deserialization with the `read_woodwork_table` function as well as in the `read_file` function.

Some resources that may be helpful:
- https://pythonspeed.com/articles/pandas-read-csv-fast/
- https://ursalabs.org/blog/fast-pandas-loadi
- https://towardsdatascience.com/apache-arrow-read-dataframe-with-zero-memory-69634092b1a

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。