[test] suppport for modeling / [test] 建模支持
Nobody has claimed this yet.
Assessment
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Newbie friendliness
- 30/100
- Issue type
- Feature
- Clarity
- Mostly clear
- Activity status
- Stale
- Domain
- infrastructure, testing-qa
Research direction
Start with runners/launch_b200-nb.sh and the requested benchmarks/glm5_fp8_b200.sh entry points, then inspect how the existing B200 benchmark workflows are structured. Confirm the specified HF_HUB_CACHE_MOUNT path and run the workflow through e2e-tests. Done means GLM-5 FP8 launches on Nvidia B200 using the new benchmark file and the end-to-end workflow passes.
Written by the indexing model from the issue text.
Description
@claude
GLM5 just released. https://cookbook.sglang.io/autoregressive/GLM/GLM-5
Create a new config for Nvidia B200 that runs GLM-5 FP8. The new benchmark file should be benchmarks/glm5_fp8_b200.sh, and should be launched from runners/launch_b200-nb.sh, but HF_HUB_CACHE_MOUNT="/mnt/data/hf_cache/hub/ . The model is already downloaded.
Then test the workflow using e2e-tests.
中文说明
测试建模支持功能。
- Dominant language
- Python
- Stars
- 1.7k
- Forks
- 303
- Avg merge
- 1d 13h
- Merged PRs (30d)
- 284
Contributor guide
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
More from SemiAnalysisAI/InferenceX
-
Difficulty 2/5 1-3 hours Newbie friendliness 68/100
SemiAnalysisAI/InferenceX#2125 ·
-
Difficulty 2/5 1-3 hours Newbie friendliness 68/100
SemiAnalysisAI/InferenceX#1587 ·
-
Difficulty 1/5 Under an hour Newbie friendliness 78/100
SemiAnalysisAI/InferenceX#1369 · 3 comments ·
-
Difficulty 1/5 1-3 hours Newbie friendliness 76/100
SemiAnalysisAI/InferenceX#1359 · 1 comment ·
-
Difficulty 5/5 Over a week Newbie friendliness 30/100
SemiAnalysisAI/InferenceX#3122 · 3 comments ·
All issues in SemiAnalysisAI/InferenceX
Similar issues
-
Difficulty 2/5 1-3 hours Newbie friendliness 74/100
bancolombia/sentinel#23 ·
-
test md OpenCI
Difficulty 2/5 1-3 hours Newbie friendliness 74/100
-
integration:quickjs org:external priority:backlog topic:code-interpreter topic:middleware type:feature
Difficulty 2/5 1-3 hours Newbie friendliness 74/100
langchain-ai/deepagents#6450 ·
-
bug client
Difficulty 2/5 1-3 hours Newbie friendliness 88/100
-
Difficulty 2/5 1-3 hours Newbie friendliness 74/100