Daniel Kleine 
							
						 
					 
					
						
						
						
						
							
						
						
							7e6f8ce020 
							
						 
					 
					
						
						
							
							updated RoPE statement ( #423 )  
						
						... 
						
						
						
						* updated RoPE statement
* updated .gitignore
* Update ch05/07_gpt_to_llama/converting-gpt-to-llama2.ipynb
---------
Co-authored-by: Sebastian Raschka <mail@sebastianraschka.com> 
						
						
					 
					
						2024-10-30 08:00:08 -05:00 
						 
				 
			
				
					
						
							
							
								ROHAN WINSOR 
							
						 
					 
					
						
						
						
						
							
						
						
							e85d154522 
							
						 
					 
					
						
						
							
							Fix argument name in LlamaTokenizer constructor ( #421 )  
						
						... 
						
						
						
						This PR addresses an oversight in the LlamaTokenizer class by changing the constructor argument from filepath to tokenizer_file. 
						
						
					 
					
						2024-10-29 18:01:36 -05:00 
						 
				 
			
				
					
						
							
							
								Daniel Kleine 
							
						 
					 
					
						
						
						
						
							
						
						
							2b24a7ef30 
							
						 
					 
					
						
						
							
							minor fixes: Llama 3.2 standalone ( #420 )  
						
						... 
						
						
						
						* minor fixes
* reformat rope base as float
---------
Co-authored-by: rasbt <mail@sebastianraschka.com> 
						
						
					 
					
						2024-10-25 21:08:06 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							75ede3e340 
							
						 
					 
					
						
						
							
							RoPE theta rescaling ( #419 )  
						
						... 
						
						
						
						* rope fixes
* update
* update
* cleanup 
						
						
					 
					
						2024-10-25 15:27:23 -05:00 
						 
				 
			
				
					
						
							
							
								Daniel Kleine 
							
						 
					 
					
						
						
						
						
							
						
						
							0ed1e0d099 
							
						 
					 
					
						
						
							
							fixed typos ( #414 )  
						
						... 
						
						
						
						* fixed typos
* fixed formatting
* Update ch03/02_bonus_efficient-multihead-attention/mha-implementations.ipynb
* del weights after load into model
---------
Co-authored-by: Sebastian Raschka <mail@sebastianraschka.com> 
						
						
					 
					
						2024-10-24 18:23:53 -05:00 
						 
				 
			
				
					
						
							
							
								Daniel Kleine 
							
						 
					 
					
						
						
						
						
							
						
						
							8b60460319 
							
						 
					 
					
						
						
							
							Updated Llama 2 to 3 paths ( #413 )  
						
						... 
						
						
						
						* llama 2 and 3 path fixes
* updated llama 3, 3.1 and 3.2 paths
* updated .gitignore
* Typo fix
---------
Co-authored-by: Sebastian Raschka <mail@sebastianraschka.com> 
						
						
					 
					
						2024-10-24 07:40:08 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							632d7772b2 
							
						 
					 
					
						
						
							
							Update test-requirements-extra.txt  
						
						
						
						
					 
					
						2024-10-23 19:19:58 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							f8bdfe12e1 
							
						 
					 
					
						
						
							
							RoPE updates ( #412 )  
						
						... 
						
						
						
						* RoPE updates
* Apply suggestions from code review
* updates
* updates
* updates 
						
						
					 
					
						2024-10-23 18:07:49 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							6dd3fbd79d 
							
						 
					 
					
						
						
							
							Update tests.py  
						
						
						
						
					 
					
						2024-10-23 07:48:33 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							9726ca6546 
							
						 
					 
					
						
						
							
							RoPE increase ( #407 )  
						
						
						
						
					 
					
						2024-10-21 19:58:38 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							37db3f0913 
							
						 
					 
					
						
						
							
							Add Llama 3.2 RoPE to CI ( #391 )  
						
						... 
						
						
						
						* add Llama 3.2 RoPE to CI
* update 
						
						
					 
					
						2024-10-08 08:28:34 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							06604f4b84 
							
						 
					 
					
						
						
							
							Introduce buffers to improve Llama 3.2 efficiency ( #389 )  
						
						... 
						
						
						
						* Introduce buffers to improve Llama 3.2 efficiency
* update
* update 
						
						
					 
					
						2024-10-06 12:49:04 -05:00 
						 
				 
			
				
					
						
							
							
								Daniel Kleine 
							
						 
					 
					
						
						
						
						
							
						
						
							4f9775d91c 
							
						 
					 
					
						
						
							
							fixed Llama 2 to 3.2 NBs ( #388 )  
						
						... 
						
						
						
						* updated requirements
* fixes llama2 to llama3
* fixed llama 3.2 standalone
* fixed typo
* fixed rope formula
* Update requirements-extra.txt
* Update ch05/07_gpt_to_llama/converting-llama2-to-llama3.ipynb
* Update ch05/07_gpt_to_llama/converting-llama2-to-llama3.ipynb
* Update ch05/07_gpt_to_llama/standalone-llama32.ipynb
---------
Co-authored-by: Sebastian Raschka <mail@sebastianraschka.com> 
						
						
					 
					
						2024-10-06 09:56:55 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							81053ccadd 
							
						 
					 
					
						
						
							
							Add a note about weight tying in Llama 3.2 ( #386 )  
						
						
						
						
					 
					
						2024-10-05 09:20:54 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							8d6b25785d 
							
						 
					 
					
						
						
							
							Llama 3.2 requirements file  
						
						
						
						
					 
					
						2024-10-05 07:32:43 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							6f86c78763 
							
						 
					 
					
						
						
							
							Implement Llama 3.2 ( #383 )  
						
						
						
						
					 
					
						2024-10-05 07:30:47 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							d313f61c86 
							
						 
					 
					
						
						
							
							Cos-sin fix in Llama 2 bonus notebook ( #381 )  
						
						
						
						
					 
					
						2024-10-03 20:45:40 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							feb0647c79 
							
						 
					 
					
						
						
							
							Improve rope settings for llama3 ( #380 )  
						
						
						
						
					 
					
						2024-10-03 08:29:54 -05:00 
						 
				 
			
				
					
						
							
							
								rasbt 
							
						 
					 
					
						
						
						
						
							
						
						
							2ae4ad15ba 
							
						 
					 
					
						
						
							
							add section numbers  
						
						
						
						
					 
					
						2024-09-30 08:42:22 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							b8497c1bf5 
							
						 
					 
					
						
						
							
							Add llama2 unit tests ( #372 )  
						
						... 
						
						
						
						* add llama2 unit tests
* update
* updates
* updates
* update file path
* update requirements file
* rmsnorm test
* update 
						
						
					 
					
						2024-09-25 19:40:36 -05:00 
						 
				 
			
				
					
						
							
							
								rasbt 
							
						 
					 
					
						
						
						
						
							
						
						
							a23fca84d5 
							
						 
					 
					
						
						
							
							improve formatting  
						
						
						
						
					 
					
						2024-09-24 18:49:17 -05:00 
						 
				 
			
				
					
						
							
							
								Daniel Kleine 
							
						 
					 
					
						
						
						
						
							
						
						
							4541177063 
							
						 
					 
					
						
						
							
							ch05/07 gpt_to_llama text improvements ( #369 )  
						
						... 
						
						
						
						* fixed typo
* fixed RMSnorm formula
* fixed SwiGLU formula
* temperature=0 for untrained model for reproducibility
* added extra info hf token 
						
						
					 
					
						2024-09-24 18:45:49 -05:00 
						 
				 
			
				
					
						
							
							
								rasbt 
							
						 
					 
					
						
						
						
						
							
						
						
							941629d2c7 
							
						 
					 
					
						
						
							
							add json import  
						
						
						
						
					 
					
						2024-09-23 09:12:35 -05:00 
						 
				 
			
				
					
						
							
							
								rasbt 
							
						 
					 
					
						
						
						
						
							
						
						
							835832a0f9 
							
						 
					 
					
						
						
							
							move access token to config.json  
						
						
						
						
					 
					
						2024-09-23 08:56:16 -05:00 
						 
				 
			
				
					
						
							
							
								rasbt 
							
						 
					 
					
						
						
						
						
							
						
						
							5e6c7230ac 
							
						 
					 
					
						
						
							
							add llama3 comparison  
						
						
						
						
					 
					
						2024-09-23 08:17:10 -05:00 
						 
				 
			
				
					
						
							
							
								Sebastian Raschka 
							
						 
					 
					
						
						
						
						
							
						
						
							c38b003aa9 
							
						 
					 
					
						
						
							
							GPT to Llama ( #368 )  
						
						... 
						
						
						
						* GPT to Llama
* fix urls 
						
						
					 
					
						2024-09-23 07:34:06 -05:00