deepspeedai / deepspeedai/DeepSpeed
Adding support for custom operator
Nobody has claimed this yet.
- Dominant language
- Python
- Stars
- 43.1k
- Forks
- 5k
- Avg merge
- 4d 15h
- Merged PRs (30d)
- 112
Description
I am not entirely sure whether this has been implemented. However, it is very often the case that one might want to speed up the training by adding some custom operator written in CUDA.
Is this currently supported? If so, is there such an example?
Please let me know.
Contributor guide
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
The issue does not name any files, tests, or entry points. Start by checking DeepSpeed's existing documentation and examples for custom operators and CUDA extensions, then determine whether support already exists; done would be a documented answer with a working example or a clearly scoped implementation plan.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100