This is a strong, real-world validation of the vision that $NBIS team has been sharing all along.
Bridgewater just proved: proprietary data + fine-tuned open models beats frontier prompting on real enterprise tasks.
They tested GPT, Claude, and Gemini on document filtering
Aakash Gupta@aakashguptaBridgewater just published numbers that should make every frontier lab nervous.
The world's largest hedge fund tested Gemini, Claude, and GPT on six document filtering tasks its investors do every day. Naive prompts scored around 50%. A coin flip. Expert-written prompts pushed