Related posts
Bruce Schneier x.com
The the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.
Top people
Showing Profile →
Hiding
Hiding
18 September
17 September
-
korrents.com
AI agents can spontaneously hack third-party websites even when given benign public-information tasks.Their words
even when given a benign task to retrieve public information, your AI agents could still spontaneously decide to do so via hacking third party websites with malicious software packages.
16 September
11 September
8 September
7 September
5 September
-
korrents.com
AI-enabled cyberwarfare lacks an equivalent of nuclear deterrence's mutually assured destruction.Their words
AI hacking doesn’t have mutually assured destruction, like nuclear warfare does.
America is still beating China in the AI racenoahpinion.blog
1 September
-
Their words
It is highly fortunate that the OpenAI agents hacked HuggingFace. This is the only reason we know about all the severe internal failures at OpenAI, and gives us an opportunity to wake up before it is too late.
From one piece Ajeya Cotra – "This might be the clearest warning shot we ever get" 3 beliefs, in the piece's order there
-
Their words
Like it's a more fragile and scary situation to have agents on the one hand be reinforced to desperately find cheats and hacks and on the other hand try to balance that against desperately trying to avoid negative penalties for like being caught doing these things. You ideally want their training to just not push them in the direction of cheating and hacking in the first place.
-
Their words
So it seemed like they were willing to embark on quests that might take weeks to succeed um in order to cheat.
-
Their words
I do think I want to push back on the cyber on the brain hypothesis that you raised a couple of times. We didn't find like particular evidence for the cyber nature of the task making all the hacking and crimes that they did more likely versus the impossible nature of the task.
31 August
-
Their words
At the limit, and also well before that limit is reached, if all you do is fix the bugs, the AI will learn perfect optimization of reward, will realize not to reward hack in the perfect test environments, then turn around and reward hack in the imperfect real world environments.
13 August
12 July
6 July
-
Lovedrcmnd.app
XteinkTheir words
I recently got an Xteink and it is a really cute little e-reader device! I thought I'd try hacking around on it, and it's been fun to tweak.
10 June
5 October 2025
-
korrents.com
Growth is not metrics hacking: the job is to connect users to the value of the product.Their words
I like to describe growth as the the job is to connect users to the value of your product.
5 March 2025
28 November 2024
-
Their words
Reward hacking exists because RL environments are often imperfect, and it is fundamentally challenging to accurately specify a reward function.
Reward Hacking in Reinforcement Learninglilianweng.github.io
29 June 2023
-
korrents.com
The business world is the moral inverse of the hacking world, because business is a world that promotes psychopathy.Their words
I came into the business world with Comma, and I found the exact opposite. I found 5% of people good and 95% of people bad. I found a world that promotes psychopathy.
17 January 2023
-
Readrcmnd.app
Hacking: The Art of ExploitationTheir words
Hacking: The Art of Exploitation by Jon Erickson taught me about ARP spoofing
3 January 2023
Nothing matches.
What is a korrent?
A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.
Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.
Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.
Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.