DReichLab / DReichLab/waldo

Splitting SQ runs so that they run through the pipeline better

Open
#125 0 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
3
Forks
0
PR merge metrics
No merged PRs in 30d

Description

Adam mentioned that when we ask for more than 1 lane of sequencing from the NovaSeq, it would be better that we pool our plates so that samples only run on 1 lane instead of over all the lanes requested. We can do this in the lab, but WALDO needs to be able to assign libraries on a plate differently then we are assuming so that the pipeline knows which group of libraries to assign to each lane

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

The issue does not name files or tests. Start by locating WALDO's library-assignment and plate-handling entry points, then trace how NovaSeq lane groups are passed to the sequencing pipeline. Done means libraries can be assigned so each pooled sample group runs on one requested lane, with coverage for the new assignment behavior.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
backend, data-engineering
Issue type
Feature
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
30/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.