Batch Size Issue in Maissi Generative Model Configuration
Dieses Issue hat noch niemand übernommen.
Bewertung
- Schwierigkeit
- 3/5
- Geschätzter Aufwand
- 1-2 Tage
- Anfängerfreundlichkeit
- 38/100
- Issue-Typ
- Bug
- Klarheit
- Muss geklärt werden
- Aktivitätsstatus
- Veraltet
- Tech-Stack
- python, pytorch
- Bereich
- data, machine-learning
Rechercherichtung
Beginne in scripts/diff_model_train.py bei prepare_data und prüfe, wie die konfigurierte Batch-Größe den ThreadDataLoader-Aufruf erreicht. Vergleiche die protokollierte len(train_loader) mit der Anzahl der Trainingsdateien und verifiziere die effektive Loader-Konfiguration; das Issue ist erledigt, wenn sich die konfigurierte Batch-Größe tatsächlich im Verhalten des Loaders widerspiegelt oder die Abweichung erklärt ist.
Vom Indexierungsmodell aus dem Issue-Text verfasst.
Beschreibung
Dear Dong Yang (@dongyang0122),
I hope this message finds you well. Thank you in advance for your time and support.
I am currently working with the Maissi generative model and planning to accelerate the training process by increasing the batch size. However, I encountered an issue where, despite modifying the batch size in the configuration file, the DataLoader batch size remains set to 1.
Could you kindly advise on how to resolve this issue?
The log file is as bellow:
wherein the log is recorded base on the code:
if local_rank == 0:
logger.info(
"[{0}] epoch {1}, iter {2}/{3}, loss: {4:.4f}, lr: {5:.12f}.".format(
str(datetime.now())[:19], epoch + 1, _iter, len(train_loader), loss.item(), current_lr
)
)
Note that the number of itereation is equal to the length of train_loader and the number of training set is 1000. In my understanding, the enlarged batch size should decrease the length of train_loader. However, the length of train_loader is still equal to 1000 (the number of training set), which seems that the batch size is 1.
Additionaly, the corresponding code for data loader is in the scripts.diff_model_train.py:
def prepare_data(
train_files: list, device: torch.device, cache_rate: float, num_workers: int = 2, batch_size: int = 1
) -> ThreadDataLoader:
"""
Prepare training data.
Args:
train_files (list): List of training files.
device (torch.device): Device to use for training.
cache_rate (float): Cache rate for dataset.
num_workers (int): Number of workers for data loading.
batch_size (int): Mini-batch size.
Returns:
ThreadDataLoader: Data loader for training.
"""
train_transforms = Compose(
[
monai.transforms.LoadImaged(keys=["image"]),
monai.transforms.EnsureChannelFirstd(keys=["image"]),
monai.transforms.Lambdad(
keys="top_region_index", func=lambda x: torch.FloatTensor(json.load(open(x))["top_region_index"])
),
monai.transforms.Lambdad(
keys="bottom_region_index", func=lambda x: torch.FloatTensor(json.load(open(x))["bottom_region_index"])
),
monai.transforms.Lambdad(keys="spacing", func=lambda x: torch.FloatTensor(json.load(open(x))["spacing"])),
monai.transforms.Lambdad(keys="top_region_index", func=lambda x: x * 1e2),
monai.transforms.Lambdad(keys="bottom_region_index", func=lambda x: x * 1e2),
monai.transforms.Lambdad(keys="spacing", func=lambda x: x * 1e2),
]
)
train_ds = monai.data.CacheDataset(
data=train_files, transform=train_transforms, cache_rate=cache_rate, num_workers=num_workers
)
return ThreadDataLoader(train_ds, num_workers=6, batch_size=batch_size, shuffle=True)
- Vorherrschende Sprache
- Jupyter Notebook
- Sterne
- 2.5k
- Forks
- 803
- Ø Merge
- 6 T. 22 Std.
- Gemergte PRs (30 T.)
- 3
Beitragsleitfaden
Erste Schritte
- Lesen Sie das ganze Issue und danach den Beitragsleitfaden des Projekts.
- Schreiben Sie ins Issue, dass Sie es übernehmen — das erspart doppelte Arbeit.
- Forken Sie das Repository und arbeiten Sie in einem Branch.
- Öffnen Sie einen Pull Request, der die Issue-Nummer nennt.
Mehr aus Project-MONAI/tutorials
-
Schwierigkeit 1/5 Unter einer Stunde Anfängerfreundlichkeit 92/100
Project-MONAI/tutorials#2076 ·
-
Schwierigkeit 1/5 1-3 Stunden Anfängerfreundlichkeit 76/100
Project-MONAI/tutorials#2069 ·
-
Schwierigkeit 1/5 Unter einer Stunde Anfängerfreundlichkeit 85/100
Project-MONAI/tutorials#1704 ·
-
Schwierigkeit 4/5 3-5 Tage Anfängerfreundlichkeit 62/100
Project-MONAI/tutorials#2071 ·
-
Schwierigkeit 3/5 1-2 Tage Anfängerfreundlichkeit 35/100
Project-MONAI/tutorials#2067 ·
Alle Issues in Project-MONAI/tutorials
Ähnliche Issues
-
bug
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 88/100
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 76/100
-
Schema-level dtype cannot be serialized: to_yaml raises RepresenterError, to_json raises TypeError Offen
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 76/100
unionai-oss/pandera#2511 ·
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 74/100
unitaryfoundation/qldpc-challenge#1651 ·
-
Schwierigkeit 1/5 Unter einer Stunde Anfängerfreundlichkeit 90/100
statsmodels/statsmodels#10271 ·