TLDRocket
Sign in

Gemma

11 summarised stories about Gemma, each linking back to the original source. Browse all topics →

+ Follow this topic

Friday, 23 February 2024

Fine-Tuning Gemma Models in Hugging Face

Hugging Face 2 years ago 50 2 sources

Google Deepmind's Gemma language models are now available via Hugging Face in 2 billion and 7 billion parameter sizes, optimized for fine-tuning using parameter-efficient techniques on both GPUs and TPUs. The article demonstrates Low-Rank Adaptation (LoRA) fine-tuning, which reduces memory requirements by training only adapter layers rather than all model weights, enabling users to adapt Gemma models on platforms like Colab or Kaggle. Users can now customize Gemma's responses for specific tasks—such as formatting quote completion—without the computational cost of full model retraining.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.