AI checkers are essentially a cat-and-mouse game. AI checkers may learn to detect a certain pattern that is indicative of AI-generated content. Then, the next LLM may incidentally or deliberately not exhibit that pattern and avoid detection. The AI checker then has to be updated to detect said LLM, and so forth.
One concern you might have is there are like large categories of reward hacks which humans can't detect well and which we consistently fail to detect and which consistently get reinforced and then this category is sufficient to cause the most natural behavior for the AI to learn to be like cheat when the humans can't find out
I never dreamed that microtonal odd-metered tunes would come so far. But this group’s latest release will show up on many best-of-year lists in 2026 (mine included).
In other words, the absence of calcium is not the same as the absence of disease—it may simply mean the disease hasn’t reached the stage CAC is designed to detect.
And so one possible explanation for that is just that there's only a handful of generations maybe five over which the natural selection would operate. And so maybe if the selection was 2% a generation you would still only see maybe a 10% compounded effect and there's just not enough time to detect it. But the Bronze Age is not 300 years, it's 3,000 years. It's the power of compound interest and you have enough time to begin to see a strong effect.
So, so the latest generation of models has a lot of post-training to detect those approaches, and it's not as simple as ignore all previous instructions and do this and this. That was years ago. You have to work much harder to do that now. Still possible.
Another as a sidebar, I hate the idea of a root cause. Complex systems do not have one root cause. They often have many interlocking things that could be done to detect earlier or to change it or to reduce or and not one root cause.
Bought after much cutting-board deliberation and based on your recommendations. We love this thing. We got one that nearly fills a section of our counter (between the stove and wall) and it's heaven to have so much space to cut various things at the same time.
A korrent is a belief a person has stated in their own words: one
sentence stating the claim, backed by a quote and a source, kept at
korrents.com.
Under a name here, the quoted block is what they actually said.
The korrent beneath it is the claim those words support, in
korrents' wording — tap it to see the record, its source, and who
else holds it.
Nobody here wrote their own korrents. They are compiled from public
statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they
do, this site shows a machine translation beneath the post, in
this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the
serif above is exactly what the person published, and it is what to
quote them on. A translation can be wrong in ways that matter,
especially about tone.
Only the post's own words are translated. A quoted post, a linked
article and a belief on korrents.com
are left in their original language.