argosopentech / argosopentech/argos-train

Data priority, incremental training?

Open
#22 2 comments 2 reactions 0 assignees View on GitHub
enhancement good first issue help wanted
Dominant language
Python
Stars
158
Forks
29
PR merge metrics
No merged PRs in 30d

Description

Hi there!

1. I would like to use the data currently provided in data-index.json, but at the same time, I would like to use my custom data. Can I tell the script to generate a model considering my custom data is more relevant / has a bigger priority?

2. Let's say I have one large dataset I am using all the time, and then I have multiple smaller datasets which I would like to train different models for each. Is something like an incremental build possible, so I would reuse some previous output and just "append" my custom data to save some training time and resources?

Thanks!

Contributor guide

No contributing guide indexed for this repository

Research direction

No specific source file, test, or entry point is named. Start by reviewing the training scripts and data-index.json to determine whether custom-data priority and incremental model building are supported, and define done as a clear implementation or documented answer for both workflows.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.