🦙 Add advanced features for LLM datasets
- Dominant language
- Python
- Stars
- 2
- Forks
- 1
- PR merge metrics
- No merged PRs in 30d
Description
Turkish LLM Fine-Tune Datasets:
- https://huggingface.co/datasets/umarigan/openhermespreference_tr
- https://huggingface.co/datasets/umarigan/openhermes_tr
- https://huggingface.co/datasets/umarigan/tinystories_tr
- https://huggingface.co/datasets/umarigan/rm_instruct_helpful_preferences_tr
- https://huggingface.co/datasets/umarigan/turkiye_finance_qa
- https://huggingface.co/datasets/umarigan/oo-gpt4-filtered-tr
- https://huggingface.co/datasets/umarigan/GPTeacher-General-Instruct-tr
Random:
- https://huggingface.co/datasets/turkish-nlp-suite/InstrucTurca
- https://huggingface.co/datasets/ftuncc/new-turkish-for-llama
Llama-Datasets:
- https://github.com/mlabonne/llm-datasets
Contributor guide
No contributing guide indexed for this repository
Research direction
The issue names no files, tests, or entry points. Start by reviewing the linked Hugging Face datasets and llm-datasets repository, then inspect llmrush to define the feature scope and acceptance criteria; done cannot be determined from the current description.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- huggingface, python
- Domain
- data, machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100