mirror of https://github.com/rasbt/LLMs-from-scratch.git synced 2025-12-20 11:42:10 +00:00

History

Sebastian Raschka 992f3068d1 Auto download DPO dataset if not already available in path (#479 )

* Auto download DPO dataset if not already available in path

* update tests to account for latest HF transformers release in unit tests

* pep 8

2025-01-12 12:27:28 -06:00

fix ch07 unit test (#470 )

2025-01-05 17:40:57 -06:00

2024-07-16 07:07:04 -05:00

2024-07-28 10:48:56 -05:00

2025-01-12 12:27:28 -06:00

2024-09-15 08:05:04 -05:00

2024-09-21 20:33:00 -05:00

README.md

2024-10-12 10:26:08 -05:00

Chapter 7: Finetuning to Follow Instructions

Main Chapter Code

02_dataset-utilities contains utility code that can be used for preparing an instruction dataset
03_model-evaluation contains utility code for evaluating instruction responses using a local Llama 3 model and the GPT-4 API
04_preference-tuning-with-dpo implements code for preference finetuning with Direct Preference Optimization (DPO)
05_dataset-generation contains code to generate and improve synthetic datasets for instruction finetuning
06_user_interface implements an interactive user interface to interact with the pretrained LLM