← FOG·CITY

“ Talks

Hill-climbing Alignment using Humans | Sampura Research

When
Wednesday, October 21 · 7:00 PM
Where
San Francisco
Listed by
Mox
​As AI systems grow more capable, the judges we use to train them increasingly fall short. Most of the field responds by replacing humans with LLM judges; we think humans and AIs have complementary strengths, and that human oversight is an uncorrelated layer of defence that AI-only judges can't provide. In this talk we'll share how we're building better judges: a leaderboard of 50+ alignment-relevant datasets, and methods that route, decompose and assist across humans and AIs to hill-climb it. Sampura Research is a London-based nonprofit, founded by Google DeepMind alumni, building Human-AI Complementarity for AI alignment.... more

More talks soon