Clouditera / Clouditera/SecGPT
能否开源SFT数据集
Open
- Dominant language
- Python
- Stars
- 3.1k
- Forks
- 370
- PR merge metrics
- No merged PRs in 30d
Description
有在其他issue看到说sft数据比较少,那sft是混合了其他数据一起训的?可否讲下sft的数据情况?
Contributor guide
No contributing guide indexed for this repository
Research direction
The issue asks whether the SFT dataset can be open sourced and how the training data is composed, but it names no files, tests, or entry points. Start by reviewing the project documentation and any existing dataset descriptions. Done means the data sources, composition, and release status are clearly documented or the maintainer has answered the availability question.
Written by the indexing model from the issue text.
Assessment
- Domain
- documentation, machine-learning
- Issue type
- Documentation
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100