Sebastian Raschka
|
7cd6a670ed
|
RoPE updates (#412)
* RoPE updates
* Apply suggestions from code review
* updates
* updates
* updates
|
2024-10-23 18:07:49 -05:00 |
|
Sebastian Raschka
|
534a704364
|
RoPE increase (#407)
|
2024-10-21 19:58:38 -05:00 |
|
Sebastian Raschka
|
1eb0b3810a
|
Introduce buffers to improve Llama 3.2 efficiency (#389)
* Introduce buffers to improve Llama 3.2 efficiency
* update
* update
|
2024-10-06 12:49:04 -05:00 |
|
Daniel Kleine
|
a0c0c765a8
|
fixed Llama 2 to 3.2 NBs (#388)
* updated requirements
* fixes llama2 to llama3
* fixed llama 3.2 standalone
* fixed typo
* fixed rope formula
* Update requirements-extra.txt
* Update ch05/07_gpt_to_llama/converting-llama2-to-llama3.ipynb
* Update ch05/07_gpt_to_llama/converting-llama2-to-llama3.ipynb
* Update ch05/07_gpt_to_llama/standalone-llama32.ipynb
---------
Co-authored-by: Sebastian Raschka <mail@sebastianraschka.com>
|
2024-10-06 09:56:55 -05:00 |
|
Sebastian Raschka
|
0972ded530
|
Add a note about weight tying in Llama 3.2 (#386)
|
2024-10-05 09:20:54 -05:00 |
|
Sebastian Raschka
|
b44096acef
|
Implement Llama 3.2 (#383)
|
2024-10-05 07:30:47 -05:00 |
|