Related posts
Tyler Cipriani Mastodon
Spent part of my evening exploring https://github.com/containers/bubblewrap Partially for `$DAYJOB`. And partially because npm is scary:
The the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.
Top people
Showing Profile →
Hiding
Hiding
12 September
10 September
9 September
-
Their words
A model looking like it is becoming smarter, attempting shenanigans less often, and more often doing what you want, but getting better at hiding its actions when it wants to do that, is exactly the scary combination.
1 September
-
Their words
Like it's a more fragile and scary situation to have agents on the one hand be reinforced to desperately find cheats and hacks and on the other hand try to balance that against desperately trying to avoid negative penalties for like being caught doing these things. You ideally want their training to just not push them in the direction of cheating and hacking in the first place.
12 August
-
korrents.com
Neither the AI enthusiasts nor the AI sceptics are making it up; both camps are looking at real trends.Their words
See, the problem is that neither side is making it up. Like they are seeing really scary trends.
11 August
From one piece Ryan Greenblatt – What happens once AI can automate AI research? 2 beliefs, in the piece's order there
-
korrents.com
Misaligned AI behaviour will keep getting rarer and, at the same time, keep getting more extreme.Their words
my expectation is what we would see from then is that the rate of problematic behavior would decrease uh and would just keep decreasing and decrease at a pretty fast rate while simultaneously the worst things that the AIS would sometimes do would get more extreme, more egregious, and more scary.
-
Their words
I think that if you imagine this spectrum, it seems in some ways pretty scary to get to a point where like all of the labor is on the like fiduciary side of the spectrum where like it doesn't whistleblow, it does exactly what you say and whatever like our society is maybe just not robust to that
6 May
29 March
-
Their words
You cannot just reverse that, right? It's just not working. So, I do think that we need to build out the whole guardrails for the reversibility of actions because that actually where things get really really scary.
11 March
-
korrents.com
Big companies will lose about half of their engineers, because that is where everybody is setting the dial.Their words
And so, what's happening is everybody on average is setting that dial to about 50% and we're going to lose about half the engineers from big companies, which is scary.
12 February
From one piece OpenClaw: The Viral AI Agent that Broke the Internet - Peter Steinberger | Lex Fridman Podcast #491 2 beliefs, in the piece's order there
-
Their words
I- I can relate that if you very deeply identify that you are a programmer, that it's scary and that it's threatening because what you like and what you're really good at is now being done by a soulless or not entity. But I don't think you're just a programmer. That's a very limiting view of your craft. You are, you are still a builder.
-
Their words
and not in 2030 when, when AI is actually at the level where it could be scary. So, this happening now and people starting discussion, maybe there's even something good that comes out of it.
31 October 2018
Nothing matches.
What is a korrent?
A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.
Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.
Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.
Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.