LLMs: A Journey Through Time and Architecture
September 24, 2024
11,719
420
50
4.01%
Search the Record
IndexedEvery word spoken in this episode is indexed. Type any phrase to jump straight to the moment it was said.
Type any word or phrase that may have been spoken. Click a result to seek the player to that exact moment.
Try a name, a topic, or a quoted line
Sebastian Raschka Episodes Around September 24, 2024
See what was published immediately before and after this episode.
19:44Now PlayingLLMs: A Journey Through Time and Architecture
Chapters
YouTube Description
as posted by the channelREFERENCES:
Stepbystep guide converting GPT to Llama
Build a Large Language Model (From Scratch)
The Llama 3 Herd of Models (31 July 2024),
Qwen2 Technical Report (15 July 2024),
Apple Intelligence Foundation Language Models (29 July 2024),
Gemma 2Improving Open Language Models at a Practical Size (31 July 2024),
DESCRIPTION:
In this video, you'll learn about the architectural difference between the original GPT model and the various Llama models. Moreover, you'll also learn about new pre-training recipes used for Qwen 2, Gemma 2, Apple's Foundation Models, and Llama 3, as well as some of the efficiency tweaks introduced by Mixtral, Llama 3, and Gemma 2.
---
To support this channel, please consider purchasing a copy of my books
---
---
OUTLINE:
Guests & Subjects Covered
Sentinel Indexing in Progress
Metadata and chapters are available. Claim extraction for this episode is pending.
All video content is delivered via YouTube embedded players in accordance with the YouTube Terms of Service. Sentinel provides research tools that promote discovery and accountability across political media.









