QLoRA fine-tuning of Llama 3.2 3B for reliable structured tool-call generation. Investigating whether ~3,000 curated examples can close the reliability gap between small open models and hosted APIs (GPT-4o, Claude) at <5% of the cost. Runs on free Colab. Data quality > quantity.