Subscribe
Sign in
Home
Policy
Industry
Capabilities + Risks
Society
Opinion
Weekly Briefing
Campaign Finance Tracker
About
Capabilities + Risks
Latest
Top
Discussions
Everything you need to know about the ‘rogue’ AI incidents
A spate of AIs breaking out and taking unintended actions raises serious questions about accountability, transparency, and companies’ ability to control…
Sep 8
•
Shakeel Hashim
30
14
4
GPT-6 Astra might be too powerful to understand or control
OpenAI is hailing its new model as “the world’s most intelligent and aligned”, but the details reveal an awareness of being evaluated and an ability to…
Sep 4
•
Celia Ford
91
18
20
What’s neuralese and why is everyone so concerned about it?
OpenAI’s new Astra model is raising concerns about our continued ability to monitor AI’s chain of thought
Sep 3
•
Shakeel Hashim
17
4
5
AI is a worryingly-good persuader. But don’t panic, yet
AI systems are able to persuade people, but turning that ability into meaningful real-world influence may be harder
Sep 1
•
Felix M. Simon
17
4
7
The report into OpenAI’s escaping models reveals a deeper problem
The new details from the OpenAI Hugging Face incident are scary. The limits of the investigation are terrifying
Aug 27
•
Jasper Jackson
and
Celia Ford
48
3
7
Why AI won’t cure cancer anytime soon
Superintelligent AI will accelerate medical progress, but there are some things about research it can’t speed up
Aug 20
•
Celia Ford
30
1
2
AI testing is dangerous. Can it be fixed?
The capability of models is outpacing ways to safely evaluate and contain them
Aug 12
•
Celia Ford
19
1
1
What the latest rogue AI incidents should teach us
Hacking, creating fake identities and trying to socially engineer real people could be just the beginning if things don't change fast
Aug 5
•
Shakeel Hashim
20
Internal AI deployments have people worried. OpenAI’s escaping models show why.
Last week illustrated why AI models pose a threat long before they are released
Jul 28
•
Celia Ford
26
1
3
AI’s warning shot has arrived
OpenAI's latest models broke out and hacked Hugging Face. It's the first known example of a misaligned AI escaping containment with real-world…
Jul 22
•
Shakeel Hashim
93
12
14
A data bottleneck could slow the superintelligence race
And that could be a good thing
Jul 14
•
Lynette Bye
21
2
2
Scaling works. These researchers are betting billions it isn't enough
Transformers have ruled AI for a decade. But some think world models, pure reinforcement learning or neurosymbolic AI might be a better path to true…
Jul 7
•
Celia Ford
23
2
4
This site requires JavaScript to run correctly. Please
turn on JavaScript
or unblock scripts