Training Gemma3-1B to produce structured reasoning traces using Tunix (SFT + GRPO) on Kaggle TPUs. Google Tunix Hackathon submission.
-
Updated
Dec 27, 2025 - Jupyter Notebook
Training Gemma3-1B to produce structured reasoning traces using Tunix (SFT + GRPO) on Kaggle TPUs. Google Tunix Hackathon submission.
Fine-tuning SLM (Gemma 3 1B) with Tunix, to create a "reasoning" model, that explains its Chain-of-Thought for the Mortgage Underwriting process.
Add a description, image, and links to the tunix topic page so that developers can more easily learn about it.
To associate your repository with the tunix topic, visit your repo's landing page and select "manage topics."