still really loving this tweetdeck view I'm usually actively juggling ~3 tasks. Having them all visible with the relevant agents/outputs a keypress away feels real nice
Astra’s mundane alignment is greatly superior to Sol. For practical purposes, I was actively nervous about some potential uses of Sol, in a way I am not for Astra. Astra’s super alignment status should scare the living daylights out of you.
If this is only due to gains in capabilities, that is extremely bad news, and it means CoT monitoring is unlikely to survive for another year unless we find a way to actively improve it, and it might not last six months.
When you look at these investigations, the reason this happens is, you know, these are very responsible, very intelligent parents who will, you know, put their child in the car and then get in the car and then you just get in this mode of I'm going to work. I'm going to do my job and you're in this routine.
If the user doesn't actively push back against the AI model, then they get the generic output, the lowest common acceptable denominator of aesthetics and taste.
Overall, I'm loving the M4 Max so far, and I've noticed that it does handle certain tasks such as running Pest PHP tests about twice as fast sometimes as the M1 Pro did, which is nice.
The link between venture capital and evangelical Christianity was closer than I thought. They're not just analogous; they deliberately cross-pollinate.
You can staff a school with the best teachers on earth, give them unlimited resources, and wrap the place in every evidence-based intervention imaginable, and it still won’t work if students are resistant, disengaged, or actively hostile to the enterprise.
players who are also commentators give better commentary than people who are just commentators. Stock analysts who never invest have zero skin in the game.
To get to principal, you need to put yourself on the critical path. To be effective as a principal and go beyond it, you need to actively remove yourself from it.
So I don't think uh learning is really about training. I think learning is about about learning. It's about an active process. The child tries things and sees what happens.
In terms of things that are causing us to lose to China, tariffs are neither here nor there because, as we've discussed, we're not sensitive to cost on power. But the environmental regulations that are actively preventing us from deploying renewable energy in the United States, like this is the reason Texas is winning. Texas is out deploying California 10 to1.
And the trick here is not to act too much on the feedback that the very disappointed people are giving you because they already love your product. Also, not to act at all really on the feedback that the not disappointed people are giving you because they're so far from loving your product that they are essentially a lost cause. but to focus on the segment of the somewhat disappointed people.
That is, labs should make sure that the safety measures they apply to their powerful models prevent unacceptably bad outcomes, even if the AIs are misaligned and intentionally try to subvert those safety measures.
One of the things we do as physicists is we actually want it to break down at some level, we’re looking for the precision measurement or the energy or whatever it will take where the standard model is no longer working. Not that it’s not working approximately, but we’re looking for the deviations.
A korrent is a belief a person has stated in their own words: one
sentence stating the claim, backed by a quote and a source, kept at
korrents.com.
Under a name here, the quoted block is what they actually said.
The korrent beneath it is the claim those words support, in
korrents' wording — tap it to see the record, its source, and who
else holds it.
Nobody here wrote their own korrents. They are compiled from public
statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they
do, this site shows a machine translation beneath the post, in
this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the
serif above is exactly what the person published, and it is what to
quote them on. A translation can be wrong in ways that matter,
especially about tone.
Only the post's own words are translated. A quoted post, a linked
article and a belief on korrents.com
are left in their original language.