mirror of https://github.com/rasbt/LLMs-from-scratch.git synced 2026-04-11 14:21:41 +08:00

History

Sebastian Raschka be5e2a3331 Readability and code quality improvements (#959 ) * Consistent dataset naming * consistent section headers		2026-02-17 18:44:56 -06:00
..
01_main-chapter-code	Readability and code quality improvements (#959 )	2026-02-17 18:44:56 -06:00
02_dataset-utilities	Readability and code quality improvements (#959 )	2026-02-17 18:44:56 -06:00
03_model-evaluation	Switch from urllib to requests to improve reliability (#867 )	2025-10-07 15:22:59 -05:00
04_preference-tuning-with-dpo	Readability and code quality improvements (#959 )	2026-02-17 18:44:56 -06:00
05_dataset-generation	Readability and code quality improvements (#959 )	2026-02-17 18:44:56 -06:00
06_user_interface	Add PyPI package (#576 )	2025-03-23 19:28:49 -05:00
README.md	Update pixi (#661 )	2025-06-13 10:50:17 -05:00

Chapter 7: Finetuning to Follow Instructions

Main Chapter Code

02_dataset-utilities contains utility code that can be used for preparing an instruction dataset
03_model-evaluation contains utility code for evaluating instruction responses using a local Llama 3 model and the GPT-4 API
04_preference-tuning-with-dpo implements code for preference finetuning with Direct Preference Optimization (DPO)
05_dataset-generation contains code to generate and improve synthetic datasets for instruction finetuning
06_user_interface implements an interactive user interface to interact with the pretrained LLM