Investigate sum_3_azip's performance

Open
#561 1 comment 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Assessment

Difficulty
4/5
Estimated time
3-5 days
Newbie friendliness
35/100
Issue type
Bug
Clarity
Needs clarification
Activity status
Stale
Tech stack
rust
Domain
performance

Research direction

Start with the benchmark comparing sum_3_azip and sum_3_azip_fold, then inspect their generated behavior to determine why only the latter autovectorizes. Done means the performance difference and autovectorization cause are established, with a clear direction for addressing them.

Written by the indexing model from the issue text.

Description

performance

benchmark sum_3_azip seems to perform abysmally compared with the equivalent sum_3_azip_fold, investigate why, and why the former doesn't autovectorize like the latter.

Dominant language
Rust
Stars
4.3k
Forks
391
PR merge metrics
No merged PRs in 30d

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

More from rust-ndarray/ndarray

All issues in rust-ndarray/ndarray

Similar issues

More Rust issues

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.