Hacker News
A $500 RL Fine-Tune of a 9B Open Model Reportedly Beat Frontier Models on Catalog Review
A post titled "A $500 RL fine-tune of a 9B open model beat frontier models on catalog review" reached 223 points and 64 comments on HN within the last day, landing in the middle of the open-weights policy fight as a concrete economic data point for task-specific small models. The source site returned 403 to automated fetches, so the underlying methodology, base model, RL technique and evaluation set could not be independently verified for this report — the claim is currently attested only by the HN thread and its discussion. Flagged low confidence pending a readable primary source.
Source
↳ Follow the thread