axolotl

Author	SHA1	Message	Date
Wing Lian	87e073d0de	fix lora target module, require explicit flash attention, fix min logging steps, don't use adam8bit for int4, hash prepared datasets, support hf hub datasets	2023-04-17 18:01:12 -04:00
Wing Lian	77fca25f1b	4bit quantized support (wip)	2023-04-17 11:37:39 -04:00
Wing Lian	d1aed4c8e5	deepspeed doesn't work with flash-attn, and the gpu savings w flash attn are better than the deepspeed headaches	2023-04-16 06:59:47 -04:00
Wing Lian	d060c803ce	add llama 7b config and fiz lora_fan_in_fan_out for llama (copy pasta bug)	2023-04-15 14:26:52 -04:00