AnswerDotAI / AnswerDotAI/fsdp_qlora
Bigger context size?
Open
- Dominant language
- Jupyter Notebook
- Stars
- 1.6k
- Forks
- 201
- PR merge metrics
- No merged PRs in 30d
Description
Is training with 1024 or 2048 sequence length feasible using this method?
Contributor guide
No contributing guide indexed for this repository
Research direction
The issue names no file, test, or entry point. Start by inspecting the repository's training notebook and determine whether sequence lengths 1024 or 2048 are supported by the method. Done requires a documented feasibility result or a maintainer-defined implementation scope.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- jupyter-notebook
- Domain
- machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100