Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
training
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
How to read a loss curve during fine-tuning (and when to stop)
PRANJUL RATHOUR
PRANJUL RATHOUR
PRANJUL RATHOUR
Follow
Sep 6
How to read a loss curve during fine-tuning (and when to stop)
#
finetuning
#
losscurve
#
training
#
guide
Comments
Add Comment
3 min read
A Better FP4 Gradient Quantizer That Training Couldn't Notice
Seth Wheeler
Seth Wheeler
Seth Wheeler
Follow
Aug 25
A Better FP4 Gradient Quantizer That Training Couldn't Notice
#
llm
#
measurement
#
quantization
#
training
Comments
Add Comment
7 min read
Training vs Inference: Why Building Costs Millions and Asking Costs Cents
Internals Decoded
Internals Decoded
Internals Decoded
Follow
Aug 25
Training vs Inference: Why Building Costs Millions and Asking Costs Cents
#
ai
#
training
#
inference
#
compute
Comments
Add Comment
9 min read
Error Feedback, Gradient Compression, and Why Adam Breaks It
Seth Wheeler
Seth Wheeler
Seth Wheeler
Follow
Aug 21
Error Feedback, Gradient Compression, and Why Adam Breaks It
#
llm
#
measurement
#
quantization
#
training
5
 reactions
Comments