ppll

Related posts

Alex Tabarrok x.com

Monitoring COT and communication between agents is also getting more difficult as networks get bigger and models become more advanced, according to @ConnorTabarrok, policy lead at Equistamp, a startup that provides safety evaluations for AI projects.

startupsevaluationsaccording

The subject this post names, from the same vocabulary the directory files beliefs under, and the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.

Sources

Top people

19 September

Zvi MowshowitzNewsletter

6 September

1 September

Ajeya CotraKorrents

From one piece Ajeya Cotra – "This might be the clearest warning shot we ever get" 2 beliefs, in the piece's order there

11 August

Ryan GreenblattKorrents

  • Their words

    when AIs are extremely extremely capable my view is that those AIs will be harder to align than current systems. So for current systems, we have this feedback loop where we basically like we create an AI. We do some evaluations on it. We see that it has some kind of messed up behavior that we can kind of quickly understand. Then we like can like go look in training and be like, "Oh, the these training environments led to this problematic behavior. Let's like tweak that training data. Let's introduce some additional training data to like correct this other issue and then move forward from there." But in a regime where the AIs are extremely situationally aware, very very very very capable and um you know uh we don't necessarily understand what they're doing, this feedback loop breaks down.

    youtube.com

26 June

Noam BrownKorrents

3 June

Matt KepnesBlog

1 October 2025

Hamel HusainRecommends

  • Mixed onrcmnd.app

    Braintrust

    Their words

    The panel had a generally positive view of Braintrust, highlighting its clean UI and structured approach to evaluations. The tool’s emphasis on human-in-the-loop workflows was a significant strength.

    hamel.dev

11 September 2025

Jason LiuSite

7 May 2024

Buck ShlegerisKorrents

From one piece The case for ensuring that powerful AIs are controlled 2 beliefs, in the piece's order there

26 October 2023

Yoshua BengioKorrents

10 March 2010

Paul BloomPapers

What is a korrent?

A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.

Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.

Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.

About the English under a post

Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.

The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.

Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.