AnswerDotAI / AnswerDotAI/fsdp_qlora

Bigger context size?

Open
#38 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
1.6k
Forks
201
PR merge metrics
No merged PRs in 30d

Description

Is training with 1024 or 2048 sequence length feasible using this method?

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue names no file, test, or entry point. Start by inspecting the repository's training notebook and determine whether sequence lengths 1024 or 2048 are supported by the method. Done requires a documented feasibility result or a maintainer-defined implementation scope.

Written by the indexing model from the issue text.

Assessment

Tech stack
jupyter-notebook
Domain
machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.