Commit History

drop empty tokenized rows too (#509)
c56b450
unverified

winglian commited on

set zero3 optimizer betas to auto so they inherit from HF trainer config (#507)
1e07c16
unverified

tmm1 commited on

add eval benchmark callback (#441)
7657632
unverified

winglian commited on

customizable ascii art (#506)
548787d
unverified

winglian commited on

support for datasets with multiple names (#480)
5ac3392
unverified

winglian commited on

remove --force-reinstall from Dockerfile to ensure correct pytorch version (#492)
e356b29
unverified

tmm1 commited on

Fix(doc): Clarify no amp to full yaml docs (#496)
48c5647
unverified

Nanobit commited on

tweak: use default config file when only one file is present (#501)
36b2e1c
unverified

Maxime commited on

Refactor train cfg cli (#499)
125cccb
unverified

winglian commited on

use math.ceil instead of round /cc #498
fd55bc8

tmm1 commited on

pad_to_worst_case_seq_len boolean, for testing memory limits (#498)
8e197f6
unverified

Birch-san tmm1 commited on

simplify linear layer locator
267b7b2

tmm1 commited on

fsdp requires params be the same type too (#493)
98bf76e
unverified

winglian commited on

Fix(tokenizer): Make sure to add pad for CodeLlamaTokenizer (#489)
4c37bd0
unverified

Nanobit commited on

Merge pull request #485 from maximegmd/patch-4
f144e98
unverified

tmm1 commited on

fix condition and add logging
3a011ea

tmm1 commited on

Merge branch 'main' into patch-4
1f613e5

tmm1 commited on

rename var and reformat
f319b0b

tmm1 commited on

Update src/axolotl/utils/models.py
7fd662d
unverified

Maxime tmm1 commited on

Update src/axolotl/utils/models.py
9e69968
unverified

Maxime tmm1 commited on

Feat(cfg): Add code-llama configs for all sizes (#479)
3513071
unverified

mhenrichsen mhenrichsen commited on

Feat(deepspeed): Add zero2 config (#476)
3fc9006
unverified

mhenrichsen mhenrichsen commited on

Feat(doc): Update eval_steps doc (#487)
ad8be43
unverified

Nanobit commited on

Add example Llama 2 ReLoRA config (#471)
fe4d6ba
unverified

chargoddard commited on

Merge pull request #486 from OpenAccess-AI-Collective/adam-bnb-simpler
f313010
unverified

tmm1 commited on

let transformers handle adamw_bnb_8bit
868530c

tmm1 commited on

ignore: address pr review
d03887f
unverified

Maxime commited on

fix: inference did not move the model to the correct device (#483)
17605b8
unverified

Maxime commited on

ignore: linter
a184549
unverified

Maxime commited on

fix: finetune model inference needs the dtype fix to work with flash-attn
f311df9
unverified

Maxime commited on

Fix missing 'packaging' wheel (#482)
c500d02
unverified

Maxime commited on

fix checkpints on multigpu (#481)
31f3e71
unverified

winglian commited on

Merge pull request #484 from OpenAccess-AI-Collective/reqs
56c4a94
unverified

tmm1 commited on

allow newer deps
c29117a

tmm1 commited on

fix types w lora (#478)
0b7ba57
unverified

winglian commited on

Fix(tokenizer): Fix condition to add pad token (#477)
71bd062
unverified

Nanobit commited on

improve llama pad token handling (#475)
cb9797e
unverified

winglian commited on

ReLoRA implementation (with quantization) (#322)
bde3c5a
unverified

chargoddard winglian commited on

Fix(doc): Clarify config (#466)
55c23c7
unverified

Nanobit commited on

workaround so training doesn't hang when packed dataloader batches aren't even (#461)
c69faee
unverified

winglian commited on

fix test fixture b/c hf trainer tokenization changed (#464)
d5dcf9c
unverified

winglian commited on

feat: add Metharme prompt strategy (#446)
f474650
unverified

TearGosling Nanobit commited on

recast loralayer, norm, lmhead + embed token weights per original qlora (#393)
96deb6b
unverified

winglian commited on

always drop samples that are too long (#452)
50682a3
unverified

winglian commited on

set env var for FSDP layer to wrap (#453)
5a1985b
unverified

winglian commited on

Merge pull request #451 from OpenAccess-AI-Collective/eval-is-causal
5e9c6af
unverified

tmm1 commited on

is_causal fix for evals?
fbf49a4

winglian commited on

add missing positional arg (#450)
58cf7e7
unverified

winglian commited on

feat(docs): improve user customized prompts (#443)
04a42b6
unverified

Nanobit commited on