tinygrad

mirror of https://github.com/tinygrad/tinygrad.git synced 2026-06-24 02:14:17 +00:00

History

chenyu 994944920b simpler batch_load_train_bert [pr] (#8582 ) don't think that buffer is really beneficial. 5% faster data_time and 1ms faster per step. https://wandb.ai/chenyuxyz/MLPerf-BERT/runs/69c9lx8y/overview		2025-01-12 20:25:05 -05:00
..
scripts	UNet3D MLPerf (#3470 )	2024-09-10 04:37:28 -04:00
training_submission_v4.0/tinycorp	copy mlperf 4.0 to mlperf 4.1 (#5614 )	2024-07-20 16:12:00 -04:00
training_submission_v4.1/tinycorp	update mlperf systems and copy 4.1 to 5.0 (#7004 )	2024-10-11 16:20:34 -04:00
training_submission_v5.0/tinycorp	EVAL_BS=36 for bert [pr] (#8576 )	2025-01-12 09:43:56 -05:00
dataloader.py	simpler batch_load_train_bert [pr] (#8582 )	2025-01-12 20:25:05 -05:00
helpers.py	no load in INITMLPERF (#5957 )	2024-08-08 11:28:24 -04:00
initializers.py	Float in scaled dot product attention (#4985 )	2024-06-18 08:16:41 -04:00
losses.py	[MLPerf][UNet3D] Add DICE loss + metrics (#4204 )	2024-04-17 20:09:33 -04:00
lr_schedulers.py	fp16 resnet (without expand backwards sum in float, doesn't work) (#3816 )	2024-03-28 01:25:37 -04:00
metrics.py	[MLPerf][UNet3D] Add DICE loss + metrics (#4204 )	2024-04-17 20:09:33 -04:00
model_eval.py	[MLPerf] Prepare openimages dataset script (#6747 )	2024-09-27 11:13:56 -04:00
model_spec.py	move globalcounters to ops (#2960 )	2024-01-01 14:21:02 -08:00
model_train.py	use tqdm tqdm in mlperf training (#7929 )	2024-11-27 21:57:05 -05:00
README	start on mlperf models	2023-05-10 16:30:49 -07:00

README

Each model should be a clean single file.
They are imported from the top level `models` directory

It should be capable of loading weights from the reference imp.

We will focus on these 5 models:

# Resnet50-v1.5 (classic) -- 8.2 GOPS/input
# Retinanet
# 3D UNET (upconvs)
# RNNT
# BERT-large (transformer)

They are used in both the training and inference benchmark:
https://mlcommons.org/en/training-normal-21/
https://mlcommons.org/en/inference-edge-30/
And we will submit to both.

NOTE: we are Edge since we don't have ECC RAM