In this video, we dive deep into the world of Large Language Models (LLMs) by learning how to fine-tune one from scratch on a text summarization task! We will be using the TinyStories 19 Million parameter base model and fine-tuning it on the TinyStories instruct dataset for summarizing stories.
We will walk through the initial base model and how it generates tiny stories. We will further train that model on summarization tasks. The lecture involves dataset preparation and fine-tuning from scratch with the Pytorch framework. We will also understand how to shift the model parameters and training dataset on GPU for speeding up the training process.
Access the Notebook: https://github.com/SauravP97/llm-fine...
Github Repo: https://github.com/SauravP97/llm-fine...
Fine-tuned TinyStories 19M model: https://huggingface.co/SauravP97/tiny...
TinyStories Paper: https://arxiv.org/pdf/2305.07759
My Socials 🚀
🙋♂️ Linkedin: / saurav-prateek-7b2096140
🧑💻 Github: https://github.com/SauravP97
☀️ Instagram: / saurav_prateek
⚡️ Book a 1:1 session with me for Interview Preparation and Career guidance, Mock Interviews and Resume Review on Topmate: https://topmate.io/saurav_prateek
На этой странице сайта вы можете посмотреть видео онлайн Fine-Tune LLM on Text Summarization task from scratch | Pytorch | Python длительностью часов минут секунд в хорошем качестве, которое загрузил пользователь Saurav Prateek 10 Апрель 2026, поделитесь ссылкой с друзьями и знакомыми, на youtube это видео уже посмотрели 532 раз и оно понравилось 28 зрителям. Приятного просмотра!