Bridgewater fine-tuning with expert data beats prompting-only approaches
Original titlePeople sometimes ask why fine-tune when general-purpose models keep getting better. Bridgewater's work is a good reminder that with the r...
AISummary
John Schulman argues that fine-tuning with the right data, such as expert judgments, can substantially outperform prompting-only approaches even as general-purpose models improve. He cites Bridgewater's work, where an expert-labeled dataset and on-policy distillation were used to fine-tune a model to triage financial documents reliably and cheaply.
Source: John Schulman · x.comPublished · added here