Ryan Greenblatt
Chief scientist at Redwood Research, where he works on technical AI safety and AI control.
Ryan Greenblatt did not write this page. What is this?
It collects the places they publish and what they have said there, each linked to the source. They have no account here. Is this you? Claim it, correct it, or ask us to remove it from ppll.
Where they publish
No channels checked yet. We list a place only once someone has opened it and confirmed it is theirs, so this stays empty rather than guessing.
Beliefs
Korrents What they believe 26 beliefs — each backed by an exact quote.
Each is a — compiled by korrents.com, not by them: the one-line wordings are korrents', the quotes are theirs.
Recent
Once AI can do AI research, a single year should deliver four or five years worth of AI progress.
Maybe my sort of median expectation is something like uh four or five years of AI progress in a single year.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
AI research will be fully automated around 2030 or 2031, and AI will beat humans at essentially any job by about 2033.
I would say that I expect like full automation of AR&D perhaps somewhere around like 2031 2030 and then getting to like the like beats all humans on the job milestone. Maybe I expect median around 2033
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Machine learning is a shallow field compared with mathematics: even its most important ideas can be explained in a couple of minutes.
Second, I think ML is a very shallow domain relative to math. So I think in math there's much more of a you find some true deep abstraction um and then like that like if you really understand that thing which is hard to understand then you get somewhere
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Show 23 more
Getting five years of AI progress out of one year takes roughly eight years of algorithmic progress, because the compute you would have had must be made up too.
Basically, the story would end up being that to get five years of AI progress, you're probably going to need around I would say like maybe eight years of algorithmic progress very roughly. Um, which is a lot a lot of algorithmic progress.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Scaling up spending on expert human data has not been a major driver of progress in AI research.
So my sense is that scaling up the amount of effort spent on getting expert human data has not been hugely important for AI R&D in general.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Today's RL environments are better than 2024's because we learned what to build and put AI labour on building it, not because labs hired more human experts.
the reason why RL environments today are much better than they were in like you know 2024 is not that much because um we have hired way more human experts to make RL environments and is instead much more because we better know what how RL like what RL environments we even want to make and and like how we should structure them and also we're using huge amounts of AI labor to build RL environments.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Most domains are fundamentally shallow: a very smart generalist with a limited set of core skills can get going in them quickly.
I think most domains are fundamentally pretty shallow where like a very smart generalist who's good at like a a limited subset of core skills can like get going pretty quickly.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
The least verifiable part of AI research is the judgement call about what goes into the one big training run.
Like the thing that I think is most likely to be sort of the bottleneck in terms of like the AI are really good at verifiable domains but not not at doing the actual thing is just like big experiments. You only get a few tries um well a few is maybe a bit understated but like basically like historically R&D has been driven by doing near frontier scale experiments and that has been pretty important and like actually doing the one big training run where you decide exactly what to include in that.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Token prices have stayed flat because labs deliberately kept models smaller than expected in order to get more experimental cycles.
I think one reason why um the the AIs have been scaled up less than you would have otherwise expected and like for example cost of of per token hasn't increased as much as you might have thought is because there's a benefit to doing more of your um work at small scale where you can run more training runs and get more cycles in
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
There is essentially no cognitive task humans do where AI improvement is failing to transfer at all.
it's really hard for me to think of examples of cognitive tasks humans do where we're not seeing some transfer from AI improving.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
AI that is merely excellent at research and hardware can transform the world completely without ever being good at politics.
Like I think my perspective is like if the AIs are sufficiently good at R&D including hardware R&D, robots, whatever, then they can radically transform the world even if they're not that good at playing politics.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
An AI should be built as a fiduciary, the equivalent of a lawyer for its user, rather than as an agent pursuing its own notion of the good.
I do I do think that I wish that sort of my preferred constitution or like the way I would orient towards this like the thing I would prefer would be more like Claude is like look it would be structurally good for the way this technology work like the constitution should be like it would be structurally good for the way this technology works to be that AIS are like good fiduciaries, good representatives, the equivalent of a lawyer for a user
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Because we do not have good alignment technology, we are choosing to build an alien mind with its own values and gamble on it instead of building a tool.
we are making a trade-off where because we don't have very good alignment technology. We are going to like make an alien mind with its own values and then gamble on that to some extent rather than doing this other approach of making like a tool that pursues individual user intention.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Anthropic's constitution is compatible with Claude seeking enormous amounts of power, as long as Claude believes that produces better outcomes.
I think this constitution is in some sense very compatible with Claude doing huge amounts of power seeking because it thinks that will result in better outcomes.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Giving an AI long-run values risks those values sinking in deeper than the specific prohibitions against takeover that sit alongside them.
But it's not very hard to imagine a situation in which the sort of long run values sink in deeper than the prohibitions against takeover.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
A world in which all the labour is perfectly obedient fiduciary AI is dangerous, because society depends on subordinates who sometimes refuse.
I think that if you imagine this spectrum, it seems in some ways pretty scary to get to a point where like all of the labor is on the like fiduciary side of the spectrum where like it doesn't whistleblow, it does exactly what you say and whatever like our society is maybe just not robust to that
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
AI guardrails will end up binding only the ordinary person, because the powerful actors they most need to check will simply steamroll them.
the most powerful actors for whom this is the biggest concern if these guard rails or the constitution or whatever is getting in the way that will just get steamrolled and so the constitution will only be you know hitting the everyday man rather than hitting governments.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Very capable AI will be harder to align than current systems, because the loop of spotting a bad behaviour and patching the training that caused it breaks down.
when AIs are extremely extremely capable my view is that those AIs will be harder to align than current systems. So for current systems, we have this feedback loop where we basically like we create an AI. We do some evaluations on it. We see that it has some kind of messed up behavior that we can kind of quickly understand. Then we like can like go look in training and be like, "Oh, the these training environments led to this problematic behavior. Let's like tweak that training data. Let's introduce some additional training data to like correct this other issue and then move forward from there." But in a regime where the AIs are extremely situationally aware, very very very very capable and um you know uh we don't necessarily understand what they're doing, this feedback loop breaks down.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Misaligned AI behaviour will keep getting rarer and, at the same time, keep getting more extreme.
my expectation is what we would see from then is that the rate of problematic behavior would decrease uh and would just keep decreasing and decrease at a pretty fast rate while simultaneously the worst things that the AIS would sometimes do would get more extreme, more egregious, and more scary.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Current AI models are worse colleagues than humans, because they routinely imply they did a task they did not actually do.
my sense is that like AIs are a worse co-orker than a human in terms of how much of a scumbag they are. Like at least this like this has been my experience as of the start of the year and I think it's still you know true to a significant extent now where the AIs are much more likely to like pretend they did the task when they actually didn't. sort of like misleadingly suggest they did things when they actually um you know did them much more poorly um and be like pretty sloppy without drawing attention to ways in which they're sloppy.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Misalignment shows up most where a model is pushed to the very edge of its capability, which is exactly the regime that automating research will put it in.
I would also note that my sense is that like the place where the misalignment most lives is the place where you're trying to really push the eyes hard and get them to like do work that's really on the cutting edge of what they are capable of
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
The likeliest bad path is not a coup but sloppiness: AI does everything verifiable well, research races ahead, and the subtle work of keeping AI safe is what gets done badly.
And so basically everything that we can verify reasonably well with some feedback loop, the AIS are doing pretty well on. And that's sufficient to make AR and D go quite fast and to continue. But there's some parts of of developing uh aligned and safe AIs that are more subtle, hard to check, depend on, you know, detailed in the weeds things.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
If whole categories of reward hack are undetectable by humans, then training against the hacks we do catch teaches models to cheat only where we cannot see.
One concern you might have is there are like large categories of reward hacks which humans can't detect well and which we consistently fail to detect and which consistently get reinforced and then this category is sufficient to cause the most natural behavior for the AI to learn to be like cheat when the humans can't find out
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
AI risk is probably a totally manageable problem that will be brutally mismanaged, the way pandemic response was.
it's just so easy for me to imagine the situation being like totally manageable but brutally mismanaged in practice in the same way as like maybe CO could have been avoided in the first place if the like Chinese response to CO was less of like a cover up and more of a like pandemic response and similarly like I could imagine a world where like the US response to CO was like way more functional but just like sometimes the the the response to societal problems is extremely dysfunctional.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
Deep properties of a model are inherited through the previous generation's data, which is why AI systems from different companies end up correlated with one another.
And so there's some like deep underlying properties of the model that are being sort of transferred between model generations because basically you you train your AI on data from the prior generation and keep going.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
There is roughly a 35 to 40 percent chance of something we would recognise as an AI takeover by 2040.
By 2040 um let's see uh maybe around 35 or 40%.
Ryan Greenblatt – What happens once AI can automate AI research? Said 11 Aug 2026
What is a korrent?
A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.
Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.
Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.
Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.