AI-Hypercomputer / AI-Hypercomputer/JetStream
Question about producing the logits from Jetstream (context and generated output) like gather_all_token_logits in TRTLLM
Open
- Dominant language
- Python
- Stars
- 457
- Forks
- 67
- PR merge metrics
- No merged PRs in 30d
Description
Any example to set it up? Thanks
Contributor guide
Assessment
This issue has not been assessed yet.