Bug fixes by danielhanchen · Pull Request #1195 · unslothai/unsloth

danielhanchen · 2024-10-26T08:20:54Z

No description provided.

orginal -> original

* Fix DPO, ORPO (#1177) * Fix TRL * Update mistral.py * Patch processing_class * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Installation guide (#1165) * chore: update chat_templates.py (#1166) orginal -> original * Disable Flex Attention * Update tokenizer_utils.py * Update _utils.py * n_items * Update cross_entropy_loss.py * Fix DPO, ORPO * Update _utils.py --------- Co-authored-by: timothelaborie <97834767+timothelaborie@users.noreply.github.com> Co-authored-by: Ikko Eltociear Ashimine <eltociear@gmail.com> * Add warning for missing Unpack and KwargsForCausalLM in older Transformers versions --------- Co-authored-by: Daniel Han <danielhanchen@gmail.com> Co-authored-by: timothelaborie <97834767+timothelaborie@users.noreply.github.com> Co-authored-by: Ikko Eltociear Ashimine <eltociear@gmail.com>

* Enhance rotary embedding handling in LlamaAttention and LongRopeRotaryEmbedding * Typo * Improve rotary embedding handling in LlamaAttention to prevent errors with short KV cache * Update llama.py * Update llama.py --------- Co-authored-by: Daniel Han <danielhanchen@gmail.com>

WizKnight · 2024-11-03T11:32:07Z

Hi @danielhanchen 🤗, I'm excited about the idea of adding float8 + QLoRA finetuning support via Torch AO into Unsloth !
I'd like to contribute and work on this. Do you have any specific points to consider before I start?

* Fix TRL * Update mistral.py * Patch processing_class * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Installation guide (unslothai#1165) * chore: update chat_templates.py (unslothai#1166) orginal -> original * Disable Flex Attention * Update tokenizer_utils.py * Update _utils.py * n_items * Update cross_entropy_loss.py * Fix DPO, ORPO * Update _utils.py * Update _utils.py * fix/transformers-unpack (unslothai#1180) * Fix DPO, ORPO (unslothai#1177) * Fix TRL * Update mistral.py * Patch processing_class * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Update tokenizer_utils.py * Installation guide (unslothai#1165) * chore: update chat_templates.py (unslothai#1166) orginal -> original * Disable Flex Attention * Update tokenizer_utils.py * Update _utils.py * n_items * Update cross_entropy_loss.py * Fix DPO, ORPO * Update _utils.py --------- Co-authored-by: timothelaborie <97834767+timothelaborie@users.noreply.github.com> Co-authored-by: Ikko Eltociear Ashimine <eltociear@gmail.com> * Add warning for missing Unpack and KwargsForCausalLM in older Transformers versions --------- Co-authored-by: Daniel Han <danielhanchen@gmail.com> Co-authored-by: timothelaborie <97834767+timothelaborie@users.noreply.github.com> Co-authored-by: Ikko Eltociear Ashimine <eltociear@gmail.com> * Update cross_entropy_loss.py * Update _utils.py * Update _utils.py * donot upcast lm_head and embeddings to float32 (unslothai#1186) * Cleanup upcast logs (unslothai#1188) * Fix/phi-longrope (unslothai#1193) * Enhance rotary embedding handling in LlamaAttention and LongRopeRotaryEmbedding * Typo * Improve rotary embedding handling in LlamaAttention to prevent errors with short KV cache * Update llama.py * Update llama.py --------- Co-authored-by: Daniel Han <danielhanchen@gmail.com> * Update transformers --------- Co-authored-by: timothelaborie <97834767+timothelaborie@users.noreply.github.com> Co-authored-by: Ikko Eltociear Ashimine <eltociear@gmail.com> Co-authored-by: Edd <68678137+Erland366@users.noreply.github.com> Co-authored-by: Datta Nimmaturi <datta.nimmaturi@nutanix.com>

danielhanchen and others added 29 commits October 21, 2024 01:02

Fix TRL

f0aca90

Update mistral.py

f4ae585

Patch processing_class

106f213

Update tokenizer_utils.py

ef84212

Update tokenizer_utils.py

4f7c527

Update tokenizer_utils.py

aa2b207

Update tokenizer_utils.py

101389d

Update tokenizer_utils.py

c0f0fc9

Update tokenizer_utils.py

b3e0033

Installation guide (#1165)

aabb5ff

chore: update chat_templates.py (#1166)

30bf339

orginal -> original

Disable Flex Attention

2895839

Update tokenizer_utils.py

06f5d75

Update _utils.py

28e6eea

n_items

b821f20

Update cross_entropy_loss.py

e561366

Fix DPO, ORPO

4ff247a

Merge branch 'main' into nightly

2b858a5

Update _utils.py

1c063b4

Update _utils.py

f195ee1

Update cross_entropy_loss.py

5961c34

Update _utils.py

7308bb8

Update _utils.py

0096e5b

Merge branch 'main' into nightly

44b480f

donot upcast lm_head and embeddings to float32 (#1186)

6776055

Cleanup upcast logs (#1188)

625209e

Update transformers

6f28d16

danielhanchen merged commit d76eda4 into main Oct 26, 2024

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

Bug fixes#1195

Bug fixes#1195
danielhanchen merged 29 commits into
mainfrom
nightly

danielhanchen commented Oct 26, 2024

Uh oh!

WizKnight commented Nov 3, 2024

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

6 participants

Uh oh!

Conversation

danielhanchen commented Oct 26, 2024

Uh oh!

WizKnight commented Nov 3, 2024

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

6 participants