Lightning-AI / Lightning-AI/lit-llama

getting random tokens after finetuning

Open
#403 0 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
6.1k
Forks
517
PR merge metrics
No merged PRs in 30d

Description

Hi Everyone,
I followed the steps for finetuning the llama model from https://github.com/Lightning-AI/lit-llama and did the fine-tuning.
when getting the results from the model on validation dataset while using the model to generate different inputs with same model object it is generating some random tokens can anyone know how can i control the random tokens and get the refined result?
has anyone tried for batch generation for different inputs if yes how can we use the lit-llama functionality for batch prediction ?

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Start with the linked lit-llama fine-tuning instructions and the validation-generation path described in the issue. Reproduce generation with the same model object and inputs, then investigate whether the token variation is expected or indicates a reproducibility problem. Also determine whether batch prediction is supported; done means the behavior and any required usage are clearly established.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.