mirror of https://github.com/rasbt/LLMs-from-scratch.git synced 2025-10-21 13:00:42 +00:00

History

* minor dpo fixes

* Update dpo-from-scratch.ipynb

metadata diff

2025-04-16 12:56:49 -05:00

2025-04-12 10:29:53 -05:00

2024-07-16 07:07:04 -05:00

2025-02-16 13:16:51 -06:00

Minor DPO fixes (#617 )

2025-04-16 12:56:49 -05:00

2025-02-16 13:16:51 -06:00

Add PyPI package (#576 )

2025-03-23 19:28:49 -05:00

README.md

2025-04-12 14:51:02 -05:00

Chapter 7: Finetuning to Follow Instructions

Main Chapter Code

02_dataset-utilities contains utility code that can be used for preparing an instruction dataset
03_model-evaluation contains utility code for evaluating instruction responses using a local Llama 3 model and the GPT-4 API
04_preference-tuning-with-dpo implements code for preference finetuning with Direct Preference Optimization (DPO)
05_dataset-generation contains code to generate and improve synthetic datasets for instruction finetuning
06_user_interface implements an interactive user interface to interact with the pretrained LLM