$ cat /etc/cookies.conf
We use cookies to understand how people use this site.
Analytics cookies help us improve your experience.
They are off by default. Nothing tracks you until you say so.
$ select cookie_preferences
Members-Only
Recent Talks & Demos are for members only
You must be an AI Tinkerers active member to view these talks and demos.
The talk explains how self‑play and DPO fine‑tuning on ranked Polymarket predictions improves 14‑billion‑parameter models to GPT‑4o level by 7‑10% accuracy using real‑world outcomes.
In our first paper from Lightning Rod Labs (https://lightningrod.ai), we explore if AI can improve its forecasts via self-play and real-world outcomes:
Result: +7-10% accuracy over control, bringing two small (14B) models on par with GPT-4o (over 10x larger).
Arxiv Link: https://arxiv.org/abs/2502.05253
Loading recent emails...