DEV Community

#finetuning

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
From adapter to deployment: merging LoRA weights and serving with vLLM or a Space

From adapter to deployment: merging LoRA weights and serving with vLLM or a Space

Comments
3 min read
Small language models (1B–3B): what they're good for after fine-tuning

Small language models (1B–3B): what they're good for after fine-tuning

Comments
3 min read
Fine-tuning on a laptop or a free GPU: a VRAM budget walkthrough

Fine-tuning on a laptop or a free GPU: a VRAM budget walkthrough

Comments
3 min read
Ten fine-tuning mistakes I see students make (and made myself)

Ten fine-tuning mistakes I see students make (and made myself)

Comments
3 min read
Instruction tuning vs domain adaptation: two different fine-tuning goals

Instruction tuning vs domain adaptation: two different fine-tuning goals

Comments
3 min read
Validate your dataset before you burn a GPU hour: a checklist

Validate your dataset before you burn a GPU hour: a checklist

Comments
3 min read
Fine-tuning datasets: Alpaca, ShareGPT and chat templates without the confusion

Fine-tuning datasets: Alpaca, ShareGPT and chat templates without the confusion

Comments
3 min read
QLoRA hyperparameters that actually matter (rank, alpha, learning rate, epochs)

QLoRA hyperparameters that actually matter (rank, alpha, learning rate, epochs)

Comments
3 min read
How to read a loss curve during fine-tuning (and when to stop)

How to read a loss curve during fine-tuning (and when to stop)

Comments
3 min read
RAG vs fine-tuning: which one your problem actually needs

RAG vs fine-tuning: which one your problem actually needs

Comments
3 min read
LoRA vs QLoRA vs full fine-tuning: cost, quality and when each makes sense

LoRA vs QLoRA vs full fine-tuning: cost, quality and when each makes sense

Comments
3 min read
Fine-tuning a 1.7B model at 3.2 GB VRAM — building FineTune Studio

Fine-tuning a 1.7B model at 3.2 GB VRAM — building FineTune Studio

Comments
2 min read
How a Dedup Pass Deleted My Training Curriculum

How a Dedup Pass Deleted My Training Curriculum

Comments
6 min read
A generic fine-tuning playbook, written after doing it wrong several times

A generic fine-tuning playbook, written after doing it wrong several times

1
Comments 1
6 min read
Teaching a local coding agent from its own mistakes: DPO on a 30B model

Teaching a local coding agent from its own mistakes: DPO on a 30B model

1
Comments 2
7 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.