axolotl

Author	SHA1	Message	Date
Wing Lian	5783839c6e	download model weights on preprocess step (#1693 )	2024-06-09 20:10:17 -04:00
Wing Lian	cbbf039a46	verbose failure message (#1694 )	2024-06-09 20:09:36 -04:00
Wing Lian	18cabc0c46	fix for when sample_packing and eval_sample_packing are different (#1695 )	2024-06-08 09:48:30 -04:00
Wing Lian	ed8ef65371	add back packing efficiency estimate so epochs and multi-gpu works properly (#1697 )	2024-06-08 09:48:10 -04:00
Wing Lian	9c1af1a9c0	ensure explicit eval_sample_packing to avoid mismatch issues (#1692 )	2024-06-07 11:28:43 -04:00
Brian Fitzgerald	cf64284a04	Phi-3 conversation format, example training script and perplexity metric (#1582 ) * phi-3 support and perplexity metric * phi-3 chat template * metrics updates * chore: lint * fix assertion on Tensor * fix tests since tokenization happens in the metric * fix perplexity value of shorter passage --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-06-04 16:11:56 -04:00
Wing Lian	c996881ec2	add support for rpo_alpha (#1681 ) * add support for rpo_alpha * Add smoke test for dpo + nll loss	2024-06-04 16:09:51 -04:00
Wing Lian	1f151c0d52	re-enable DPO for tests in modal ci (#1374 ) * re-enable DPO for tests in modal ci * workaround for training args * don't mixin AxolotlTrainingArguments * fix mixin order so MRO doesn't result in TypeError: non-default argument follows default argument error * use smaller datasets for dpo tests	2024-06-03 12:50:44 -04:00
Wing Lian	05b0bd08d2	need to add back drop_last for sampler (#1676 )	2024-05-31 13:13:13 -04:00
Wing Lian	d4f6c65e4c	cleanup the deepspeed proxy model at the end of training (#1675 )	2024-05-30 13:40:35 -04:00
Wing Lian	a944f7b32b	load explicit splits on datasets (#1652 )	2024-05-29 22:27:59 -04:00
Wing Lian	9d4225a058	set chat_template in datasets config automatically (#1664 ) * set chat_template in datasets config automatically * dynamic chat_template, not jsut chatml	2024-05-29 22:27:26 -04:00
Wing Lian	f7332ac449	use mixins for orpo and kto configs so they work with axolotl customizations (#1674 )	2024-05-29 22:27:00 -04:00
Wing Lian	a6b37bdeb4	revert multipack batch sampler changes (#1672 ) * revert multipack batch sampler changes * fix default val for drop_last	2024-05-29 11:51:18 -04:00
Wing Lian	b7520801a3	handle the system role too for chat templates (#1671 )	2024-05-29 10:21:11 -04:00
Wing Lian	fe650dd326	make sure the CI fails when pytest script fails (#1669 ) * make sure the pytest script fails * make sure the defaults come through for tests * make sure tensorboard is loaded for test assertion	2024-05-29 10:12:11 -04:00
Seungduk Kim	65db903714	Correct name of MixtralBlockSparseTop2MLP (L -> l) (#1667 )	2024-05-28 18:10:29 -04:00
Davide Caroselli	6a5a725f10	Fix: ensure correct handling of `val_set_size` as `float` or `int` (#1655 ) * Fix: ensure correct handling of val_set_size as float or int * chore: lint --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-28 12:00:32 -04:00
Keith Stevens	cc11c6bce2	Generalizing the chat_template prompt strategy (#1660 ) [skip ci] The strategy now supports configuring several fields: * The data field holding message arrays * the role and content fields for each message * role mapping from source to target types additionally this adds a sample llama3-8b instruct template using the chat template	2024-05-28 11:24:13 -04:00
Wing Lian	367b2e879b	Switch to parallel FFD bin packing algorithm. (#1619 ) * Switch to parallel FFD bin packing algorithm. Add support for packing in a distributed context. Add packing efficiency estimate back. * revert changes to distributed code * chore: lint * fix config w new params for packing test * add sample_packing_group_size and sample_packing_bin_size to cfg schema * fix lamdbda function * fix sampler/dataloader calculations for packing --------- Co-authored-by: dsesclei <dave@sescleifer.com>	2024-05-23 17:32:14 -04:00
Wing Lian	bbfed318bc	support for custom messages field in sharegpt (#1651 )	2024-05-23 13:03:22 -04:00
George Grigorev	a27d5e1f4e	enable loraplus setting for dpo trainer (#1646 )	2024-05-22 08:29:06 -04:00
Wing Lian	6299eb5919	allow report_to for multiple providers (#1647 )	2024-05-22 08:27:44 -04:00
Leonard	7c2bf3091f	Fix llama3 chat_template (extra <\|eot_id\|> on last turn) (#1635 ) * Fix llama3 chat_template (the {{eos_token}} leads to an extra <\|eot_id\|> being added in the last turn). Output now matches official Llama 3 Instruct model * add tests * chore: lint --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-21 09:08:53 -04:00
Ben Redmond	22ae21a6c2	Add KTO support (#1640 ) * add kto support * test cleanup * fix outdated comment * fix llama3 ultra * chore: lint * update to use rl_beta instead of dpo_beta --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-20 16:05:16 -04:00
Wing Lian	ba45531802	fixes to save on fractional save_steps (#1643 )	2024-05-20 14:24:45 -04:00
Wing Lian	8a1572a831	Unsloth optims for Llama (#1609 ) * WIP for unsloth integrations * import the unsloth code in the right context * add unsloth mlp, qkv, o lora optimizations * apply unsloth mlp and qkv kernels	2024-05-20 09:55:06 -04:00
Jeffrey Quesnelle	702a669cad	add save_only_model option (#1634 )	2024-05-17 00:23:18 -04:00
bofeng huang	81da7d2531	Fix `total_num_steps` (#1566 ) * Fix `total_num_steps` * Fix total_num_steps * lint	2024-05-14 20:10:37 -04:00
Ali Mosavian	1e1921b794	FIX: max_length and max_prompt_length was not being sent to ORPOTrainer (#1584 ) * FIX: TRL trainer preprocessing step was running in one process * FIX: max_length and max_prompt_length was not being sent to ORPOTrainer * FIX: Change ORPO max prompt length to 1/4 of max length, otherwise we get strange behaviour * FIX: Removed change from a different PR * FIX: Black fix * explicitly set max prompt len for orpo config --------- Co-authored-by: Ali Mosavian <ali.mosavian@kry.se> Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-14 08:51:17 -04:00
Wing Lian	1634ac82e0	make sure to save on the last step (#1615 )	2024-05-14 08:48:39 -04:00
Wing Lian	02982733ec	fix attention mask collation (#1603 )	2024-05-14 08:17:30 -04:00
Wing Lian	2147cf6837	Llama3 dpo (#1610 ) * add dpo llama3 * fix dpo bos and eos * bos token gets added automatically by the tokenizer * explicit <\|end_of_text\|> not needed, as eot_id is sufficient --------- Co-authored-by: Nero10578 <owenarliawan@gmail.com>	2024-05-11 18:29:03 -04:00
Ram	50421c8b1d	feat: Add LLaMA-3 instruct prompt strategies for fine-tuning (#1553 ) * Add prompt strategies * Update modified URL * Update modified URL * Update fastchat_conversation_turns.py * Update register function * Remove extra /n for system prompt * Fix return * Fix BOS * Update requirements, pylint * Linting * Linting * fix tuples, make sure to set system message in template * tests for llama3 tokenization * fix conditionals for loading chat template --------- Co-authored-by: Ram <ram@Rams-MacBook-Pro.local> Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-11 00:08:04 -04:00
Antoni-Joan Solergibert	b32c08f8cc	adding llama3 fastchat conversation monkeypatch (#1539 ) * adding llama3 fastchat conversation monkeypatch * Updated conversation turns to work with PR3259 of FastChat * fixed bos token * bump fastchat version --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-10 10:40:05 -04:00
Wing Lian	fff06af8d0	ignore the fsdp_config section too (#1606 ) [skip ci]	2024-05-09 13:30:39 -04:00
Wing Lian	796a085b2f	make sure to save the lora adapter at the end of RL/dpo training (#1573 )	2024-05-08 10:39:33 -04:00
Wing Lian	cb78a36374	improve tool handling roles (#1587 )	2024-05-07 11:30:40 -04:00
NanoCode012	8b9c15b17f	feat: exclude mamba blocks for jamba (#1578 )	2024-05-07 22:52:57 +09:00
Chirag Jain	9e1480e9ca	Pass deepspeed and fsdp as None explicitly when merging adapters to allow custom device_map (#1575 )	2024-05-07 22:47:55 +09:00
marijnfs	3367fca732	Gradio configuration parameters (#1591 ) * Gradio Configuration Settings * Making various Gradio variables configurable instead of hardcoded * Remove overwriting behavour of 'default tokens' that breaks tokenizer for llama3 * Fix type of gradio_temperature * revert un-necessary change and lint --------- Co-authored-by: Marijn Stollenga <stollenga@imfusion.de> Co-authored-by: Marijn Stollenga <stollenga@imfusion.com> Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-06 15:43:42 -04:00
Wing Lian	29cf15a28c	improve save callbacks (#1592 )	2024-05-04 23:19:18 -04:00
Chirag Jain	dde02fcb94	Pass weakref to model in the SIGINT handler to free up model post train function (#1581 ) * Pass weakref to model in the SIGINT handler to free up model post train() * Fix lint issues * chore: lint --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-05-03 11:05:28 -04:00
Ali Mosavian	b9bb169602	FIX: TRL trainer preprocessing step was running in one process (#1583 ) * FIX: TRL trainer preprocessing step was running in one process * FIX: Changed so that dataset_num_proc is sent to CPO, KTO and ORPO trainer args and directly to the trainer when DPO * FIX: Changed back to only support ORPO for now, since KTO is handled in another way --------- Co-authored-by: Ali Mosavian <ali.mosavian@kry.se>	2024-05-03 11:02:59 -04:00
JohanWork	601c08b4c2	ADD: warning hub model (#1301 ) * update warning for save_strategy * update * clean up * update * Update test_validation.py * fix validation step * update * test_validation * update * fix * fix --------- Co-authored-by: NanoCode012 <kevinvong@rocketmail.com>	2024-05-01 01:05:12 +09:00
Abhinand	cc5d31e0d9	Add debug option for RL dataset preprocessing (#1404 ) * adding debug option for RL dataset preprocessing * Refine formatting of debugging code in RL dataset preprocessing * Update __init__.py * chore: fix lint --------- Co-authored-by: NanoCode012 <kevinvong@rocketmail.com>	2024-05-01 00:36:04 +09:00
Wing Lian	5294653a2d	PoSE context length ext (#1567 ) * PoSE wip * fixes for pose splitting * set pose context len so we can pick that up seperately from the usable training context len * support min sample len and define num chunks * fix chunk splitting * support for curriculum/ordered learning with pose * fix sequence len sort * add curriculum_sampling to pydantic	2024-04-27 12:28:20 -04:00
Wing Lian	68601ec6ad	make sure everything stays in the same dtype when using dpo + FSDP (#1559 )	2024-04-22 16:00:05 -04:00
Haoxiang Wang	60f5ce0569	Add support for Gemma chat template (#1530 ) * Add support for Gemma chat template * Update fschat version to include its newest support for Gemma chat style * pin fastchat to current HEAD --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-04-21 19:55:40 -04:00
Frank Ruis	7477a53287	wrap prepared_ds_path in str() to avoid TypeError in fsspec package (#1548 ) * wrap prepared_ds_path in str() to avoid TypeError in fsspec package `fsspec` calls `if "::" in path` on `prepared_ds_path`, which will throw an error if it is a `PosixPath` object. * update test too --------- Co-authored-by: Wing Lian <wing.lian@gmail.com>	2024-04-21 19:55:20 -04:00

... 6 7 8 9 10 ...

1085 Commits