AnswerDotAI / AnswerDotAI/fsdp_qlora

Request for Scripts to Merge QDoRA Adapters with Base Model for vLLM Inference

Open
#60 4 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
1.6k
Forks
201
PR merge metrics
No merged PRs in 30d

Description

Hello,

I've successfully finetuned Llama-3 8B with QDoRA and am now looking to perform inference using vLLM. Could you provide guidance or scripts on how to merge the QDoRA adapters with the original base model? Additionally, does this process involve quantization and dequantization of the base model?

Thank you!

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue names no files, tests, or entry points. Start by reviewing the QDoRA adapter output and the requirements for vLLM inference, including whether merging requires quantization or dequantization. Done would be a clear, agreed-upon script or documented procedure for preparing the fine-tuned model.

Written by the indexing model from the issue text.

Assessment

Domain
ai, machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.