Ask HN: When is fine-tuning a small LLM worth it?
5
kkooldeep7 about 4 hours ago 6 comments
I'd be interested in hearing about your experiences. What kind of task did you use it for, what model did you train, and what were the results?
Feel free to share examples of what you've tried.

Discussion (6 Comments)Read Original on HackerNews
https://artreviewgenerator.com/
I was using Qwen3.5:2b models for both, running on Dell Pro Max GB10 Cuda,128GB.
Notably the latter is more of the bottleneck, particularly with the price race-to-zero with models such as GPT-6 Luna.
It's faster to make iterate when you're toying around with a 1B model than a 27B one.