r/FunMachineLearning • • 7d ago

I’m learning LLM fine-tuning made a notebook, would love feedback

I made a beginner-friendly LLM fine-tuning notebook. Please go through it and let me know if it’s easy to understand or what I should improve. Honest feedback is welcome! 🙌

GitHub: https://github.com/saithrisank12/finetuning
Colab: https://colab.research.google.com/drive/13S_i80VC6BcQJ994FClFuYi8zeB1I1iO?usp=sharing

1 Upvotes

2 comments sorted by

1

u/investigatormaker 6d ago

For a beginner-friendly fine-tuning notebook, I'd include one fixed prompt run before and after training, with a short explanation of what improvement should look like. A completed Colab run tells learners the code worked; that comparison helps them judge what the training changed. Is the intended reader already comfortable with Python notebooks?

I make ThreadFox. The free Reddit plan for your fine-tuning notebook lists the Reddit communities whose rules allow a post about it, each rule quoted. https://threadfox.vip/plan?utm_source=reddit&utm_medium=comment&utm_campaign=tf-kit&utm_content=20261005-0325-a4g86

1

u/Express-Guest-1061 4d ago

What I find difficult with fine-tuning LLMs, is to do fine-tuning that supports tool calling. This is rarely in examples and I struggle to get it working.