Bounded Regret
  • Home

Updates and Lessons from AI Forecasting

Film Study for Research

Measurement, Optimization, and Take-off Speed

Advice for Authors

Foundation Models for Oversight

19 hours ago 34 min read
Cross-posted from the Transluce blog. To oversee an AI model, we'd ideally like to ask questions such as: * What are important situations where the model sandbags? * Does the model have
Read Now Read Later
Jacob Steinhardt
By: Jacob Steinhardt

Building Technology to Drive AI Governance

5 months ago 7 min read
Technically skilled people who care about AI going well often ask me: how should I spend my time if I think AI governance is important? By governance, I mean the constraints, incentives, and
Read Now Read Later
Jacob Steinhardt
By: Jacob Steinhardt

Oversight Assistants: Turning Compute into Understanding

7 months ago 10 min read
Currently, we primarily oversee AI with human supervision and human-run experiments, possibly augmented by off-the-shelf AI assistants like ChatGPT or Claude. At training time, we run RLHF, where humans (and/
Read Now Read Later
Jacob Steinhardt
By: Jacob Steinhardt

Analyzing long agent transcripts (Docent)

a year ago 1 min read
This is a brief overview of a recent release by Transluce. You can see the full write-up on the Transluce website. AI systems are increasingly being used as agents: scaffolded systems in
Read Now Read Later
Jacob Steinhardt
By: Jacob Steinhardt

Introducing Transluce — A Letter from the Founders

2 years ago 3 min read
We are launching an independent research lab that builds open, scalable technology for understanding AI systems and steering them in the public interest. Transluce means to shine light through something to reveal its
Read Now Read Later
Jacob Steinhardt
By: Jacob SteinhardtSarah Schwettmann
Page 1 of 18
Older Posts
Powered by Ghost
Bounded Regret