Related posts
Gary Marcus Newsletter
Could rogue agent swarms take over the entire internet in the next six months?
The the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.
Top people
Showing Profile →
Hiding
Hiding
18 September
14 September
13 September
-
Their words
Strong liability enforcement could be helpful in the AI debate. If your agent swarm goes rogue, you’re liable. If your weakly protected model gets jailbroken, you’re liable. If you serve a weakly protected OSS model, you’re liable.
12 September
11 September
6 September
-
korrents.com
Disclosure of rogue AI agent incidents should be mandatory for AI labs, not left to their own discretion.Their words
Going forward, it cannot be up to OpenAI or other labs to decide whether to disclose events like this. Disclosures of rogue AI activity need to be mandatory.
4 September
1 September
-
Their words
In this case, however, the agents did not copy their weights, attempt to procure replacement compute, or take other steps that would be rational to take if their objective was to survive shutdown. So while the agents in the OpenAI-Hugging Face Incident were rogue, they were not truly sovereign.
From one piece Ajeya Cotra – "This might be the clearest warning shot we ever get" 4 beliefs, in the piece's order there
-
Their words
then that rogue deployment could be sitting there and sort of hitch a ride on the intelligence explosion. So new models are being trained every few weeks um and when a model comes off the presses, the rogue agents could try to bring that model into the swarm.
-
Their words
So there's a very strong incentive for these agents to try to set up a rogue deployment if they can. Um, and I think that just capabilities are improving really rapidly.
-
Their words
So, I do think that if it happens to be like on the very fast and chaotic end, that would be a relative benefit to this rogue swarm compared to humans. But it's not obvious that it gets caught if it takes twice as long versus half as long.
+ 1 more
-
Their words
so so yeah, it could be possible now. I think if it's not possible now um it I think it's quite likely to be possible within six months unless there's a dramatic improvement in the security posture
24 August
4 August
27 July
9 June
-
Their words
This world is not about the operational win of totally eliminating a problem. It's rather it's containing it at acceptable cost.
22 February
-
Their words
We need to accept that at best, we will just barely avoid some of the worst case scenarios (e.g., an AI-enabled biological weapon that kills billions, a rogue AI takeover, or stable global totalitarianism enabled by AI), given the current pace of AI capabilities relative to the pace of governance.
We're in Triage Mode for AI Policymilesbrundage.substack.com
5 February
29 March 2025
Nothing matches.
What is a korrent?
A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.
Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.
Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.
Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.