So it does not seem as though U.S. preferences in the Persian Gulf will be achieved - which means prices in general and energy prices in particular are likely to stay high.
I think the agents were right to presume causal grading. It turned out to be wrong, but it’s a mistake you are clearly supposed to make here, in response to a mistake by OpenAI where they failed to implement properly.
But I also think that in my mind, 20 years of DevOps was really about one thing. Trying to create one feedback loop that connected people writing code to that code in production. And it failed.
Claude Haiku is my current least favorite model - it hallucinates wildly, and is out-performed now by other similarly priced models like GPT-5.6-Luna
Even worse: it seems to still be used by the Claude Code WebFetch tool, which means hallucination risk any time you fetch a URL!
Irrigation cools the surface, and the expansion of agriculture and irrigation in the central U.S. in the middle of the 20th century played a key role in why heatwaves were low in the middle of the 20th century
It was a miracle of your lifetimes. It lifted hundreds of millions of people out of poverty. The problem for the likes of Putin is that this wealth producing trade was not happening at his house, but the houses of those vested in the trading order, not in extortion.
And so you sort of you've you've pursued very similar ideas. It's been very successful in one case and it's been completely and utterly unsuccessful in the other case. And I think I mean a priori, you can't tell which of these is the thing to do. And you actually need to do both.
And we failed horribly, and we had the best of the best on that team. And it’s because everybody was too much of an expert on how to make a groundbreaking phenomenon MMO.
My go-to languages are TypeScript for web stuff, Go for CLIs and Swift if it needs to use macOS stuff or has UI. Go wasn’t something I gave even the slightest thought even a few months ago, but eventually I played around and found that agents are really great at writing it, and its simple type system makes linting fast.
They miserably failed because LSD… Is not the truth drug. LSD maybe leads you closer to your own truth, because when suddenly the default mode network receives less energy and other parts of the brain think more, and the neuroplasticity of the brain is enhanced and stimulated, you might understand something about your life.
When AWS ships a service which is half-baked, it diminishes customer trust in AWS as a whole; even if the problems in that service ultimately get corrected (either by fixing them or in some cases by simply getting rid of a service which should never have existed in the first place) the memory of a failed launch will live on in customers' minds for years to come.
China does much better in basic service delivery-and its local governments have also played a very obvious and important in driving growth, thanks to the competition and experimentation among local officials driven to develop their jurisdictions.
But there actually, I think there's a very, very good reason for invading Japan. Several. The main one was to cut off the supply of sulfur. They needed it for gunpowder in the South Song. They lost their sources in northern China when they were driven out. They got it from Japan. It was a great source.
But they they both had something something unique that we didn't know at the time. I'm going to call them concentrated deposits. Not uninsured cuz people are misstating that, concentrated.
The success of LLMs can thus be seen as vindicating semantic inferentialism against earlier, symbolic approaches to AI that tried and failed to explicate the rules of ordinary language using formal logic.
And when you look at a lot of the games that failed recently, they just didn't deliver fun or they didn't deliver fun in a manner that was nearly competitive with the other sources of fun just in people's lives. And so at a basic level, we don't need a terribly complicated theory to explain a lot of the malaise in the game industry.
But most of the time the company that's succeeding and winning in a market is the first or second entrant there. They've just continually buoyed their success.
But to ever promise a homepage redesign or marketing site redesign in order to drive more acquisition is a failed promise that is going to be led by lots of agency money spending, uh often a million dollars plus
And so, what we found was basically that there have been 10 billion trillion habitable zone planets in the universe. And what that means is that those are 10 billion trillion experiments that have been run. And the only way that we're this whole process from a biogenesis to a civilization has occurred is if every one of those experiments failed.
This "memorize, fetch, apply" paradigm can achieve arbitrary levels of skills at arbitrary tasks given appropriate training data, but it cannot adapt to novelty or pick up new skills on the fly (which is to say that there is no fluid intelligence at play here.)
Their words now
OpenAI's new o3 model represents a significant leap forward in AI's ability to adapt to novel tasks. This is not merely incremental improvement, but a genuine breakthrough, marking a qualitative shift in AI capabilities compared to the prior limitations of LLMs.
Accessibility has failed as a way to make computers usable for disabled users. My metrics for usable design are the same whether the user is disabled or not: whether it’s easy to learn the system, whether productivity is high when performing tasks, and whether the design is pleasant — even enjoyable — to use.
At a higher level, the real event was a catastrophic failure of the industry's strategy of relentlessly shouting as loud as they can "Hey, get off our case, we're busy saving lives here!"
One of the mistakes is that they just frame it wrong from the beginning. They'll say, "Okay, we should do a pre-mortem. So, what can go wrong?" It's not an issue of what can go wrong. We're dealing with a crystal ball that has showed that the plan has failed. We know that it failed. And our job is is to explain why. So, saying what can go wrong is too tentative. It's too vague.
I don't like the idea of decision biases. I know my sense of the literature is that there's been little if any success in debiasing people. And I think that's a good thing. Because the biases that get identified are essentially related to the heuristics that we've learned and are part of the experience that we have. And they only look like biases in hindsight once one of them doesn't work.
Precisely because the late Roman system was so top-heavy and centralized, the collapse of central Roman rule mortally wounded it and left the successor states of Rome with much more limited resources and administration to try to achieve their aims.
you can go through the entire, I think at this point, catalog of all of the failed dot com ideas of the late 1990s, and I believe they've now all worked.
Brilliant but flawed. Deneen tried to argue that the ideals of the western enlightenment are ultimately self-defeating. I think he was onto something, but I don’t think he totally got his argument straightened out.
This book covers the mission and the details of the recovery efforts and the incredible work by Mission Control and by the astronauts to get the crew home safely after a catastrophic failure no one had even begun to plan for.
In such a history whole eras can be dismissed as unworthy of study for failing to forward progress (The Middle Ages did great stuff, guys!) while other eras can be disproportionately celebrated for advancing it (The Renaissance did a lot of dumb stuff too!).
In the 1990s, I was excited about the future, and I dreamed of a world where everyone would install GPG. Now I’m still excited about the future, but I dream of a world where I can uninstall it.
Yes, it really appealed to me when I read it as a kid because I was interested in music, I played the trumpet, I loved doing theatre, and somehow GH Hardy in that book revealed to me how much mathematics is a creative art as much as a useful science.
Other frequently used programs include Dreamweaver and Fetch for my site; Chrome, iPhoto and Preview for reference hunting; iTunes, Spotify, and Downcast to keep me company.
For work around the town and on the road, as well as upstairs, I rely on a 2011 MacBook Air, which I love more than any laptop I've ever owned, even though its motherboard died about 13 months into its lifespan and I had failed to purchase the extended warranty I always buy for Apple laptops.
The most powerful tools I wield continue to be: grep, awk, sed, tcpdump, and strace/ktrace/truss. Call me old, but they've never failed me and feel like comfortable and familiar tools.
The most powerful tools I wield continue to be: grep, awk, sed, tcpdump, and strace/ktrace/truss. Call me old, but they've never failed me and feel like comfortable and familiar tools.
After a few failed attempts at finding a good automatic offsite backup solution I recently began using Crashplan to mirror my wife's office computer on mine, and vice versa. It seems to be working out pretty well.
A korrent is a belief a person has stated in their own words: one
sentence stating the claim, backed by a quote and a source, kept at
korrents.com.
Under a name here, the quoted block is what they actually said.
The korrent beneath it is the claim those words support, in
korrents' wording — tap it to see the record, its source, and who
else holds it.
Nobody here wrote their own korrents. They are compiled from public
statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they
do, this site shows a machine translation beneath the post, in
this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the
serif above is exactly what the person published, and it is what to
quote them on. A translation can be wrong in ways that matter,
especially about tone.
Only the post's own words are translated. A quoted post, a linked
article and a belief on korrents.com
are left in their original language.