Related posts
Ramez Naam x.com
This matches my expectations. Cheaper models (open and closed) will capture most of the token volume. But in a great many domains (and anything adversarial, zero sum, or first past the post), people will pay for the very most advanced model.
adversarialmatchesexpectations
The the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.
Top people
Showing Profile →
Hiding
Hiding
28 August
17 August
24 July
1 May 2025
-
korrents.com
Longer test-time thinking improves an AI model's robustness to adversarial or unusual inputs.Their words
thinking for longer should be especially useful when the model is presented with an unusual input, such as an adversarial example or jailbreak attempt
18 April 2024
12 April 2024
30 November 2023
25 October 2023
From one piece Adversarial Attacks on LLMs 2 beliefs, in the piece's order there
-
korrents.com
Universal adversarial trigger attacks are easy to detect because the learned trigger tokens tend to be nonsensical.Their words
One drawback with UAT (Universal Adversarial Trigger) attacks is that it is easy to detect them because the learned triggers are often nonsensical.
-
Their words
High perplexity makes an attack more vulnerable to be detected and mitigated.
Nothing matches.
What is a korrent?
A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.
Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.
Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.
Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.