Competitive AI Safety is the loss function to make sure AI goes well
TLDR:AI safety needs a loss function and Competitive AI safety can deliver it.The development of AI safety research and tooling is currently diffuse. Lots of researchers,…
TLDR:AI safety needs a loss function and Competitive AI safety can deliver it.The development of AI safety research and tooling is currently diffuse. Lots of researchers,…
This week saw the releases of, among other things: GPT-5-6 Sol. It is a very good model, sir. Plan A, the follow up to AI 2027.…
SummaryWe’re a clinically vulnerable family and we take infection control very seriously. My first-author paper was accepted as a spotlight (top 23 out of 801) at…
First of all I should note that this post is just a guess! I have not run any experiments or anything. But anyways my guess for…
Once upon a time, A Relatively Famous Guy On The Internet was accused of having been simultaneously dating multiple women, without those women's knowledge, those women…
Epistemic status: Idea I think is really good; likely not originalTLDR: Language is so ridiculously ambiguous and slippery that you should consider interpreting the meanings of…
You’ve probably heard people say that emotional suppression is bad, and self-soothing is good. But how do you know which one is which?I don’t actually think…
An early warning system for loss of control to AI requires, at its core, a forecast of the outcome of our current trajectory. Can we build…
People spend a lot of words playing tug of war over whether or not it's reasonable to train against interpretability methods. The anti case goes something…
I've been reading a lot of older writing, trying to understand how and why contra dance ended up with a strong and near-exclusive live music tradition…