AlibabaResearch/flash-llm
在 GitHub 查看Flash-LLM: Enabling Cost-Effective and Highly-Efficient Large Generative Model Inference with Unstructured Sparsity
- 星标
- 247
- 派生
- 24
- 开放的新手 issue
- 0
- 已索引 issue
- 7
- 主要语言
- Cuda
- 许可证
- Apache-2.0
- 最近 GitHub push
- 2023年9月22日
- 最近索引
- 2026年9月13日
- 贡献指南
- 没有贡献指南
- 行为准则
- 没有行为准则
- 新手标签
- 没有已索引的新手标签
- PR 合并指标
- 30 天内没有已合并 PR
-
AlibabaResearch/flash-llm#8 · 5 条评论 · 0 个 reaction · 已指派 0 人 ·
-
AlibabaResearch/flash-llm#9 · 3 条评论 · 0 个 reaction · 已指派 0 人 ·
-
smaller OPT? 未关闭
AlibabaResearch/flash-llm#10 · 0 条评论 · 0 个 reaction · 已指派 0 人 ·
-
AlibabaResearch/flash-llm#12 · 0 条评论 · 0 个 reaction · 已指派 0 人 ·
-
AlibabaResearch/flash-llm#13 · 1 条评论 · 0 个 reaction · 已指派 0 人 ·
-
AlibabaResearch/flash-llm#14 · 0 条评论 · 0 个 reaction · 已指派 0 人 ·
-
AlibabaResearch/flash-llm#15 · 0 条评论 · 0 个 reaction · 已指派 0 人 ·