A $500 RL fine-tune of a 9B open model beat frontier models on catalog review(fermisense.com)
324 points by ilreb 1 day ago | 125 comments
tl;dr: Summary not available
HN Discussion:
  • Small fine-tuned models are sufficient for most use cases, undermining frontier model economics
  • Frontier models improve fast enough that fine-tunes become obsolete, and hidden costs exceed the $500 figure
  • The comparison is flawed or a post hoc fallacy, questioning methodology like how scoring works
  • The real insight is that reward function design and problem understanding matter more than model size
  • Curious practitioners asking about accessibility and resources for fine-tuning without taking a stance