axolotl

Author	SHA1	Message	Date
Wing Lian	469e15607d	basic llama multipack	2024-06-20 14:39:55 -04:00
DavidFarago	559562d790	Allow "weight: 0" in messages to mask them (#1703 ) Allow in message objects the additional key `weight`, which can be set to 0 (or 1) to cause that message to be masked out (or left unmasked) for training (similar to [1]). This is helpful for training the model to be robust and capable of error recovery upon a bad assistant message. A missing `weight` key defaults to weight 1, to guarantee downward compatibility. [1]: https://github.com/mistralai/mistral-finetune	2024-06-20 10:05:16 -04:00
Wing Lian	4de4b4089f	add support for multipack for deepseek_v2 (#1712 )	2024-06-20 10:02:55 -04:00
Wing Lian	3f1f5e3312	drop length column for issues with eval without packing (#1711 )	2024-06-18 23:32:29 -04:00
Wing Lian	5783839c6e	download model weights on preprocess step (#1693 )	2024-06-09 20:10:17 -04:00
Wing Lian	cbbf039a46	verbose failure message (#1694 )	2024-06-09 20:09:36 -04:00
Wing Lian	851ccb1237	bump deepspeed for fix for grad norm compute putting tensors on different devices (#1699 )	2024-06-09 17:13:28 -04:00
Wing Lian	18cabc0c46	fix for when sample_packing and eval_sample_packing are different (#1695 )	2024-06-08 09:48:30 -04:00
Wing Lian	ed8ef65371	add back packing efficiency estimate so epochs and multi-gpu works properly (#1697 )	2024-06-08 09:48:10 -04:00
Wing Lian	00ac3022a1	add qwen2-72b fsdp example (#1696 )	2024-06-07 16:38:29 -04:00
Wing Lian	9c1af1a9c0	ensure explicit eval_sample_packing to avoid mismatch issues (#1692 )	2024-06-07 11:28:43 -04:00
Aaditya Ura (looking for PhD Fall’24)	a82a711522	Create phi3-ft-fsdp.yml (#1580 ) rename to be fsdp specific and tweak settings a bit	2024-06-04 16:20:25 -04:00
Brian Fitzgerald	cf64284a04	Phi-3 conversation format, example training script and perplexity metric (#1582 ) * phi-3 support and perplexity metric * phi-3 chat template * metrics updates * chore: lint * fix assertion on Tensor * fix tests since tokenization happens in the metric * fix perplexity value of shorter passage --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-06-04 16:11:56 -04:00
Wing Lian	c996881ec2	add support for rpo_alpha (#1681 ) * add support for rpo_alpha * Add smoke test for dpo + nll loss	2024-06-04 16:09:51 -04:00
Wing Lian	1f151c0d52	re-enable DPO for tests in modal ci (#1374 ) * re-enable DPO for tests in modal ci * workaround for training args * don't mixin AxolotlTrainingArguments * fix mixin order so MRO doesn't result in TypeError: non-default argument follows default argument error * use smaller datasets for dpo tests	2024-06-03 12:50:44 -04:00
Saeed Esmaili	5cde06587a	Fix the broken link in README (#1678 ) [skip ci]	2024-06-03 09:38:44 -04:00
Wing Lian	05b0bd08d2	need to add back drop_last for sampler (#1676 )	2024-05-31 13:13:13 -04:00
Wing Lian	d4f6c65e4c	cleanup the deepspeed proxy model at the end of training (#1675 )	2024-05-30 13:40:35 -04:00
Wing Lian	a944f7b32b	load explicit splits on datasets (#1652 )	2024-05-29 22:27:59 -04:00
Wing Lian	9d4225a058	set chat_template in datasets config automatically (#1664 ) * set chat_template in datasets config automatically * dynamic chat_template, not jsut chatml	2024-05-29 22:27:26 -04:00
Wing Lian	f7332ac449	use mixins for orpo and kto configs so they work with axolotl customizations (#1674 )	2024-05-29 22:27:00 -04:00
Wing Lian	16d46b74e4	re-enable phi for tests in modal ci (#1373 )	2024-05-29 15:41:46 -04:00
Wing Lian	a6b37bdeb4	revert multipack batch sampler changes (#1672 ) * revert multipack batch sampler changes * fix default val for drop_last	2024-05-29 11:51:18 -04:00
Wing Lian	b7520801a3	handle the system role too for chat templates (#1671 )	2024-05-29 10:21:11 -04:00
Wing Lian	fe650dd326	make sure the CI fails when pytest script fails (#1669 ) * make sure the pytest script fails * make sure the defaults come through for tests * make sure tensorboard is loaded for test assertion	2024-05-29 10:12:11 -04:00
Abe Voelker	49b967b62f	Fix README quick start example usage model dirs (#1668 )	2024-05-28 18:10:40 -04:00
Seungduk Kim	65db903714	Correct name of MixtralBlockSparseTop2MLP (L -> l) (#1667 )	2024-05-28 18:10:29 -04:00
Davide Caroselli	6a5a725f10	Fix: ensure correct handling of `val_set_size` as `float` or `int` (#1655 ) * Fix: ensure correct handling of val_set_size as float or int * chore: lint --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-28 12:00:32 -04:00
Wing Lian	f5febc729a	fix lint issue that snuck through (#1665 )	2024-05-28 11:36:50 -04:00
Faria Huq	230e0ac363	Fix Lora config error for Llama3 (#1659 ) The current yml code throws an error: ValueError: Please set lora_modules_to_save to [`embed_tokens`, `lm_head`] when using an adapter and changing the special tokens. I added the required changes to resolve it	2024-05-28 11:25:08 -04:00
Keith Stevens	cc11c6bce2	Generalizing the chat_template prompt strategy (#1660 ) [skip ci] The strategy now supports configuring several fields: * The data field holding message arrays * the role and content fields for each message * role mapping from source to target types additionally this adds a sample llama3-8b instruct template using the chat template	2024-05-28 11:24:13 -04:00
Maciek	5f91064040	Fix Google Colab notebook 2024-05 (#1662 ) [skip ci] * include mlflow installation in the colab notebook Without explicitly installing mlflow the `accelerate launch` command fails. * update the colab noteboko to use the latest tinyllama config	2024-05-28 11:23:52 -04:00
Wing Lian	ef223519c9	update deps (#1663 ) [skip ci] * update deps and tweak logic so axolotl is pip installable * use vcs url format * using dependency_links isn't supported per docs)	2024-05-28 11:23:34 -04:00
Charles Frye	8a20a7b711	document how to use `share_strategy="no"` (#1653 ) [skip ci] The literal value `no` is parsed in some YAML parsers to the boolean `False`, which fails Pydantic validation. To be sure that the value is parsed to the string `"no"`, the value should be enclosed in quotes. [Discussion on StackOverflow](https://stackoverflow.com/questions/53648244/specifying-the-string-value-yes-in-yaml).	2024-05-24 14:15:44 -04:00
Wing Lian	367b2e879b	Switch to parallel FFD bin packing algorithm. (#1619 ) * Switch to parallel FFD bin packing algorithm. Add support for packing in a distributed context. Add packing efficiency estimate back. * revert changes to distributed code * chore: lint * fix config w new params for packing test * add sample_packing_group_size and sample_packing_bin_size to cfg schema * fix lamdbda function * fix sampler/dataloader calculations for packing --------- Co-authored-by: dsesclei <dave@sescleifer.com>	2024-05-23 17:32:14 -04:00
Wing Lian	bbfed318bc	support for custom messages field in sharegpt (#1651 )	2024-05-23 13:03:22 -04:00
Jaydeep Thik	84bb8061ba	Update tiny-llama qlora.yml addressing eval packing error (#1638 )	2024-05-22 08:34:06 -04:00
George Grigorev	a27d5e1f4e	enable loraplus setting for dpo trainer (#1646 )	2024-05-22 08:29:06 -04:00
Wing Lian	6299eb5919	allow report_to for multiple providers (#1647 )	2024-05-22 08:27:44 -04:00
Leonard	7c2bf3091f	Fix llama3 chat_template (extra <\|eot_id\|> on last turn) (#1635 ) * Fix llama3 chat_template (the {{eos_token}} leads to an extra <\|eot_id\|> being added in the last turn). Output now matches official Llama 3 Instruct model * add tests * chore: lint --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-21 09:08:53 -04:00
Ben Redmond	22ae21a6c2	Add KTO support (#1640 ) * add kto support * test cleanup * fix outdated comment * fix llama3 ultra * chore: lint * update to use rl_beta instead of dpo_beta --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-20 16:05:16 -04:00
Wing Lian	ba45531802	fixes to save on fractional save_steps (#1643 )	2024-05-20 14:24:45 -04:00
Wing Lian	8a1572a831	Unsloth optims for Llama (#1609 ) * WIP for unsloth integrations * import the unsloth code in the right context * add unsloth mlp, qkv, o lora optimizations * apply unsloth mlp and qkv kernels	2024-05-20 09:55:06 -04:00
Jeffrey Quesnelle	702a669cad	add save_only_model option (#1634 )	2024-05-17 00:23:18 -04:00
Wing Lian	891ae8aa13	fix ray install (#1630 )	2024-05-16 01:25:42 -04:00
Wing Lian	0c49ecc429	more fixes to work with runpod + skypilot (#1629 )	2024-05-16 00:05:56 -04:00
Wing Lian	60113437e4	cloud image w/o tmux (#1628 )	2024-05-15 22:27:40 -04:00
Wing Lian	419b2a6a98	install rsync too (#1627 )	2024-05-15 21:36:00 -04:00
Wing Lian	2501a371c6	fix setting the authorized keys when there are more than one in the env var (#1626 )	2024-05-15 20:48:56 -04:00
Wing Lian	e6937e884b	fix symlinks for axolotl outputs (#1625 )	2024-05-15 19:41:45 -04:00

1 2 3 4 5 ...

1491 Commits