Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
finetuning
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
From adapter to deployment: merging LoRA weights and serving with vLLM or a Space
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
From adapter to deployment: merging LoRA weights and serving with vLLM or a Space
#
finetuning
#
deployment
#
vllm
#
serving
Comments
Add Comment
3 min read
Small language models (1B–3B): what they're good for after fine-tuning
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
Small language models (1B–3B): what they're good for after fine-tuning
#
smallmodels
#
finetuning
#
llm
#
deployment
Comments
Add Comment
3 min read
Fine-tuning on a laptop or a free GPU: a VRAM budget walkthrough
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
Fine-tuning on a laptop or a free GPU: a VRAM budget walkthrough
#
finetuning
#
vram
#
qlora
#
budget
Comments
Add Comment
3 min read
Ten fine-tuning mistakes I see students make (and made myself)
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
Ten fine-tuning mistakes I see students make (and made myself)
#
finetuning
#
mistakes
#
students
#
guide
Comments
Add Comment
3 min read
Instruction tuning vs domain adaptation: two different fine-tuning goals
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
Instruction tuning vs domain adaptation: two different fine-tuning goals
#
finetuning
#
instructiontuning
#
domainadaptation
#
llm
Comments
Add Comment
3 min read
Validate your dataset before you burn a GPU hour: a checklist
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
Validate your dataset before you burn a GPU hour: a checklist
#
finetuning
#
datasets
#
validation
#
checklist
Comments
Add Comment
3 min read
Fine-tuning datasets: Alpaca, ShareGPT and chat templates without the confusion
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
Fine-tuning datasets: Alpaca, ShareGPT and chat templates without the confusion
#
finetuning
#
datasets
#
chattemplates
#
guide
Comments
Add Comment
3 min read
QLoRA hyperparameters that actually matter (rank, alpha, learning rate, epochs)
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
QLoRA hyperparameters that actually matter (rank, alpha, learning rate, epochs)
#
qlora
#
hyperparameters
#
finetuning
#
llm
Comments
Add Comment
3 min read
How to read a loss curve during fine-tuning (and when to stop)
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
How to read a loss curve during fine-tuning (and when to stop)
#
finetuning
#
losscurve
#
training
#
guide
Comments
Add Comment
3 min read
RAG vs fine-tuning: which one your problem actually needs
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
RAG vs fine-tuning: which one your problem actually needs
#
rag
#
finetuning
#
architecture
#
decisionguide
Comments
Add Comment
3 min read
LoRA vs QLoRA vs full fine-tuning: cost, quality and when each makes sense
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
LoRA vs QLoRA vs full fine-tuning: cost, quality and when each makes sense
#
finetuning
#
lora
#
qlora
#
llm
Comments
Add Comment
3 min read
Fine-tuning a 1.7B model at 3.2 GB VRAM — building FineTune Studio
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 5
Fine-tuning a 1.7B model at 3.2 GB VRAM — building FineTune Studio
#
finetuning
#
qlora
#
llm
#
opensource
Comments
Add Comment
2 min read
How a Dedup Pass Deleted My Training Curriculum
Seth Wheeler
Seth Wheeler
Seth Wheeler
Follow
Aug 23
How a Dedup Pass Deleted My Training Curriculum
#
llm
#
measurement
#
finetuning
#
security
Comments
Add Comment
6 min read
A generic fine-tuning playbook, written after doing it wrong several times
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Aug 24
A generic fine-tuning playbook, written after doing it wrong several times
#
llm
#
finetuning
#
machinelearning
#
mlops
1
 reaction
Comments
1
 comment
6 min read
Teaching a local coding agent from its own mistakes: DPO on a 30B model
Wuic Framework
Wuic Framework
Wuic Framework
Follow
Aug 31
Teaching a local coding agent from its own mistakes: DPO on a 30B model
#
dpo
#
qlora
#
finetuning
#
llm
1
 reaction
Comments
2
 comments
7 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account