mirror of https://github.com/rasbt/LLMs-from-scratch.git synced 2025-10-06 05:27:01 +00:00

History

Sebastian Raschka a08d7aaa84

* Uv workflow improvements

* Uv workflow improvements

* linter improvements

* pytproject.toml fixes

* pytproject.toml fixes

* pytproject.toml fixes

* pytproject.toml fixes

* pytproject.toml fixes

* pytproject.toml fixes

* windows fixes

* windows fixes

* windows fixes

* windows fixes

* windows fixes

* windows fixes

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

* win32 fix

2025-02-16 13:16:51 -06:00

01_main-chapter-code

Uv workflow improvements (#531 )

2025-02-16 13:16:51 -06:00

02_dataset-utilities

fix typos, add codespell pre-commit hook (#264 )

2024-07-16 07:07:04 -05:00

03_model-evaluation

Uv workflow improvements (#531 )

2025-02-16 13:16:51 -06:00

04_preference-tuning-with-dpo

fix reward margins plot label in dpo nb

2025-01-12 14:04:05 -06:00

05_dataset-generation

Uv workflow improvements (#531 )

2025-02-16 13:16:51 -06:00

06_user_interface

Add user interface to ch06 and ch07 (#366 )

2024-09-21 20:33:00 -05:00

README.md

Update bonus section formatting (#400 )

2024-10-12 10:26:08 -05:00

README.md

Chapter 7: Finetuning to Follow Instructions

Main Chapter Code

01_main-chapter-code contains the main chapter code and exercise solutions

Bonus Materials

02_dataset-utilities contains utility code that can be used for preparing an instruction dataset
03_model-evaluation contains utility code for evaluating instruction responses using a local Llama 3 model and the GPT-4 API
04_preference-tuning-with-dpo implements code for preference finetuning with Direct Preference Optimization (DPO)
05_dataset-generation contains code to generate and improve synthetic datasets for instruction finetuning
06_user_interface implements an interactive user interface to interact with the pretrained LLM