Files
axolotl/examples
Wing Lian 328d598114 gemma3 packing fixes (#2449)
* make gemma3 work with packing

* multi-gpu e2e for ci

* update gemma3 model namespace to use mirror

* add gradient checkpointing to multigpu e2e ci

* update gemma3 examples for use_reentrant and fix ddp find unused params

* fix tests for gemma3

* fix import for test utils

* set correct train loss for gemma3 e2e
2025-03-31 17:15:23 -04:00
..
2025-01-30 11:45:56 -05:00
2025-03-31 17:15:23 -04:00
2025-01-30 11:45:56 -05:00
2025-01-30 11:45:56 -05:00