deepspeedai / deepspeedai/DeepSpeed
AttributeError: partially initialized module 'deepspeed' has no attribute 'init_inference'
Nobody has claimed this yet.
- Dominant language
- Python
- Stars
- 43.1k
- Forks
- 5k
- Avg merge
- 4d 15h
- Merged PRs (30d)
- 112
Description
To Reproduce
inference script:
`import torch
from transformers import AutoModelForCausalLM, AutoTokenizer
import deepspeed
model_name = "/home/pzl/models/Qwen2.5-3B"
tokenizer = AutoTokenizer.from_pretrained(model_name)
model = AutoModelForCausalLM.from_pretrained(model_name)
ds_engine = deepspeed.init_inference(model,
mp_size=1,
dtype=torch.half,
replace_with_kernel_inject=True)
input_text = "DeepSpeed is?"
inputs = tokenizer(input_text, return_tensors="pt")
with torch.no_grad():
outputs = ds_engine.module.generate(**inputs, max_length=10)
output_text = tokenizer.decode(outputs[0], skip_special_tokens=True)
print(output_text)
`
Expected behavior
get the correct result
ds_report output
Screenshots
System info (please complete the following information):
- OS: Ubuntu 22.04
- GPU count and types 0
- Python 3.10
Contributor guide
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
Start by reproducing the provided inference script, focusing on the deepspeed.init_inference entry point and capturing the full traceback and installed environment details. Done means the cause of the AttributeError is identified and the script completes with the expected generated output, with any required configuration documented.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python, pytorch
- Domain
- machine-learning
- Issue type
- Bug
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100