In this video, we dive deep into the world of Large Language Models (LLMs) by learning how to fine-tune one from scratch on a text summarization task! We will be using the TinyStories 19 Million parameter base model and fine-tuning it on the TinyStories instruct dataset for summarizing stories.
We will walk through the initial base model and how it generates tiny stories. We will further train that model on summarization tasks. The lecture involves dataset preparation and fine-tuning from scratch with the Pytorch framework. We will also understand how to shift the model parameters and training dataset on GPU for speeding up the training process.
Access the Notebook: https://github.com/SauravP97/llm-fine...
Github Repo: https://github.com/SauravP97/llm-fine...
Fine-tuned TinyStories 19M model: https://huggingface.co/SauravP97/tiny...
TinyStories Paper: https://arxiv.org/pdf/2305.07759
My Socials 🚀
🙋♂️ Linkedin: / saurav-prateek-7b2096140
🧑💻 Github: https://github.com/SauravP97
☀️ Instagram: / saurav_prateek
⚡️ Book a 1:1 session with me for Interview Preparation and Career guidance, Mock Interviews and Resume Review on Topmate: https://topmate.io/saurav_prateek
En esta página del sitio puede ver el video en línea Fine-Tune LLM on Text Summarization task from scratch | Pytorch | Python de Duración hora minuto segunda en buena calidad , que subió el usuario Saurav Prateek 10 abril 2026, comparta el enlace con amigos y conocidos, en youtube este video ya ha sido visto 532 veces y le gustó 28 a los espectadores. Disfruta viendo!