microsoft / microsoft/SandDance
Optimize performance for large datasets
Nobody has claimed this yet.
- Dominant language
- TypeScript
- Stars
- 7.1k
- Forks
- 573
- Avg merge
- 8d 9h
- Merged PRs (30d)
- 4
Description
I am trying to analyze a 500MB csv file and rendering time takes minutes on a strong Windows machine. Sometimes the plot will render the first time in a few minutes, but never again (if I change the x-axis for example). Is it possible to launch this tool without analyzing/rendering by default? That way I can choose my options and then render once? Also, is anyone working on improving the performance of this tool?
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
The issue names no files, tests, or entry points. Reproduce the 500MB CSV case on Windows and profile the initial and repeated rendering paths; done should include a way to configure options before rendering and measurable improvement for large datasets.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- typescript
- Domain
- data-visualization, performance
- Issue type
- Feature
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Mostly clear
- Newbie friendliness
- 25/100