Every now and again, I write about sports. But I only end up showing how out-of-touch I am.
Ah, I’ve put my foot into it again, with my response to LeBron James’s announcement that he is moving on—one more time—to a new team. This was my initial take on today’s “last decision” (as he describes it) to join the Philadelphia 76ers.
This man jumps around more than a knight on a chessboard. So much so, that these repeated relocations now define who he is, and how he will be remembered. After more than two decades in the NBA, LeBron has four championships, but will have started afresh five times with a new team. For someone chasing titles with his generation-defining talent, that ratio is out of whack.
Please support The Honest Broker by taking out a premium subscription (just $6 per month).
But even if he had more rings to show, LeBron would still force us to ask if chasing titles is a legit reason to abandon your team. Wouldn’t most fans say that greatness is defined by the opposite—namely when an athlete triumphs without first running away?
So if I could give him a nickname, it would be LeBron “The Decision” James. The irony is that, despite taking pride in his decisions, he is a man who can never make up his mind. Shakespeare’s Hamlet looks like a paragon of decisiveness by comparison.
Of course, this is just one more example of how poorly aligned I am with current sports culture. I’ve gone on record as praising six things about sports—and I never hear any of them mentioned on ESPN.
It’s actually far worse than that—the prevailing ethos in sports now runs totally in opposition to all six of these core values. Just consider that the hottest trend in professional sports today is gambling. That tells you how far we’ve fallen.
Not long ago, Pete Rose got thrown out of baseball for gambling. Nowadays he would earn a huge endorsement deal and launch his own betting site.
Yes, I wasn’t made for these times. But can I really be the only person who feels this way?
No, I’m certain I’m not alone. It’s just that the voices of those who believe in sports for character building are drowned out by a culture that puts money and winning above everything else.
Each fan also gets to make a “decision”—although it’s not covered in the news. And their cumulative decisions will determine the actual fate of the league
There’s a price to pay for this. A society that trains its youngsters with these degraded values on the sports field will soon find the same vices everywhere else. Long ago, Duke Wellington allegedly boasted that the battle of Waterloo was won on the playing fields of Eton. Nowadays we could make a less enviable claim, namely that the corruption of our institutions starts in the youth sports leagues.
I’ve criticized LeBron for lacking loyalty. Sports insiders mock this view—but with tremendous hypocrisy. That’s because the whole profession of sports depends on the loyalty of fans. Without loyalty, every pro sports league in the world would soon collapse. So not only is loyalty an important value in team competitions, it’s actually a foundational principle that holds teams together and pays everybody’s salary along the way.
Lebron James in his last game with Cleveland—where he still had my support (Photo by Keith Allison)
Yes, I’m an outlier. But I also think I’m speaking in the interests of the sports insiders too when I advocate for more loyalty, not less. If I could, I’d make it harder—much harder—for teams to abandon their cities. I would make it easier for teams to sign home town talent (bringing back a variant of the territory draft that once increased the likelihood that an NBA squad would actually include players from its home city. I’d take other steps to maximize the emotional ties between fans and teams, and vice versa.
It’s obvious that these moves would increase the popularity of teams. Even players would benefit—despite what you might hear to the contrary. It’s well known that players who stay with one team for their entire career are beloved by the fans, even long after they have retired. They have a happy home for life—along with lots of free drinks and dinners along the way—that the itinerant player will never enjoy.
But what about winning, you ask? I’d reply that the most successful NBA teams have tended to hold on to key players for their entire career—starting with their initial draft and continuing until retirement. Magic and Kobe and West and Baylor were always Lakers. Russell and Bird and and McHale and Havlicek were always Celtics. Duncan and Robinson and Ginóbili were always Spurs. When these same teams lost this sense of continuity and put too much emphasis into short term deals and luring free agents, their dynasties fell apart.
I mention those elite teams only in passing, because sports for most competitors does not result in a dynasty. The value of sports can’t be based by winning, because losing is just as common. And it’s even less dependent on a championship season—which few ever experience. For most competitors, the real value lies is those principles listed above—being there for your team, learning self-discipline, bonding with your community, handling setbacks and losses, and the like.
LeBron, of course, is free to pursue other values. But he will need to do it without me cheering him on. And at lease some insiders are aligned with me on this. For example, here’s a smart take that calls it straight.
There’s an old proverb that says “Be careful what you wish for, because you just might get it.” That’s exactly what’s happening in the NBA right now. They wanted a profit-maximizing world in which loyalty meant nothing—with ultimate freedom and mobility. Now they have it. But wait until they see the price to pay when fans ditch their loyalty in response.
Yes, each fan also gets to make a “decision”—although it isn’t covered in the news or hyped in press conferences. And make no mistake, their cumulative decisions will determine the actual fate of the league, not anything LeBron does in Philly.
If the honchos running the NBA were really smart, they would be worrying about those decisions, and not the last gasp destination of a soon-to-retire baller.
Welcome to the reading list, a weekly roundup of news and links related to buildings, infrastructure, and industrial technology. This week we look at China’s EV subsidies, the future of aeroderivative gas turbines, a US-Saudi nuclear deal, defective roadway guardrails, and more. Roughly 2/3rds of the reading list is paywalled, so for full access become a paid subscriber.
Housing and Cities
The CEO of DR Horton, largest homebuilder in the US, thinks that the 21st Century ROAD to Housing Act won’t have much impact on housing supply or demand in the short term. [X]
The Financial Times on why socialists should embrace luxury apartments. “Above all, high-end developments unlock long housing chains. As higher-income households move into newly built units, they free up older properties, raising supply and slashing prices for middle- and lower-end housing through a process known as filtering. Numerous international studies underscore this positive ripple effect.” A good sentiment, though stating it to the readers of the Financial Times is probably preaching to the choir. [FT]
Last week we noted that Saronic, which had been slated to be a major tenant for California Forever’s shipbuilding facility, had instead opted to locate their operations in Texas. Someone claimed on Twitter that California was, in fact, never really in consideration, and that overtures to California Forever were entirely to secure better terms in Texas. [X]
The fraction of housing markets in the US where prices are rising vs falling over time, via ResiClub. [X]
US cities which used to have at least 100,000 people but no longer do. [Reddit]
Manufacturing
TSMC is accelerating its investment in Arizona fabs due to enormous AI demand. [CNBC] And TSMC plans to raise prices by up to 10% starting in 2027. [Yahoo Finance]
As China’s EV industry gets more mature, China is beginning to roll back subsidies for the sector. “Starting Sept. 1, lithium-ion batteries will be subject to a 2% consumption tax, which will double to 4% a year later, according to a Friday announcement by the Ministry of Finance.” [X]
Energy
The National Historic Preservation Act (NHPA) requires considering whether a new project will impact the view of some existing historic property. Thus the more land area a project extends over, or the taller it is, the more potential view impacts, and the greater the NHPA review burden. This is bad news for clean energy projects like wind and solar, which have a much “viewprint” than oil and gas projects. “We examined nearly 100 recent energy projects and found a striking discrepancy: compared to fossil energy projects, clean energy and transmission projects have review areas about 10 times larger and are more than 8 times as likely to receive an “adverse effect” finding — triggering additional process, delay, and mitigation obligations.” [IFP]
Tim Latimer, CEO of geothermal startup Fervo Energy, on what the geothermal industry can learn from SpaceX. Basically, instead of trying to develop an ambitious technology in one shot, start with a simpler, less ambitious version of it, and advance via learning by doing. “...much like the Starship example, starting with the end state and trying to solve all the various technology challenges all in one go would be incredibly challenging. So Fervo has opted to take a learning by doing path that parallels SpaceX. Project Red was our first pilot. Rather than try to drill a super hot or super deep concept right out of the gate, we completed a project in a category called “near-field Enhanced Geothermal Systems” (NF-EGS). This meant going to a location that was adjacent to a producing geothermal site, so the geothermal gradient was higher, even though the permeability in the part of the field we had selected did not support prior development with conventional technology.” [LinkedIn]
I just contributed a couple of microfictions to a project I’m helping bootstrap: The young and promising Jamverse extended protocol fiction universe. I’d like to invite the fiction writers among you to jump in as well.
There’s currently a contest going on for worldbuilding artifacts, with a deadline of July 31, to help bootstrap the universe. If you’re looking for a fun way to kill a few hours sometime in the next week, explore Jamverse and enter the contest.
Jamverse (named for Fred Pohl’s dictum that the science-fiction writer’s job is to predict not the automobile but the traffic jam) is a near-future Earth setting full of strange (but not magical) rules. Currently there are 15 full short stories across 4 loosely coupled story cycles by 4 authors, and a growing hopper of related world-building elements contributed by half an dozen others. The pump-priming is going well, and if this sort of thing interests you, now is the time to get in on the ground floor. It’s currently still pretty easy to introduce big, bold ideas into the universe without running into conflicts.
As a warmup to more serious contributions, I wrote two (non-competing) microfictions for the contest. Both are stand-alone readable, though you’ll get more out of both if you’ve read some of the stories of the rest of the universe.
My first entry is Ziploc Protocol, a microfiction that explores how a piece of jailbreaking technology might be smuggled across a border into a highly protocolized jurisdiction called Zoothesia to subvert some of its protocols. Read it here.
My second entry is a mood/vibe piece. A microfiction introducing The Thicket, a kind of tiling element for the background interstices of the world. The Thicket is inspired by a nonfiction idea developed by , and in the Jamverse, it refers to something like the future of the exurban hinterlands of metro areas today. The microfiction introduces the concept via a stand-alone vibe-y vignette. Read it here.
Jamverse has a lot of interesting elements both for readers and writers. There are strange augmented reality civic protocols, a train that runs halfway across the planet, a surreal folklore layer, and weird governance and public service models. I’m currently brainstorming my own top-level big idea to add.
Here’s a quick overview orientation (AI generated), and a codex of worldbuilding elements to let you jump in and contribute quickly. Browsing the latter should give you an immediate feel for the world.
If you plan to enter the contest (and I hope you do), you might want to read a couple of the key world-establishing stories of the 4 existing story cycles first.
The subsection of the Georgi Gospodinov novel runs like this:
The Shortest Novel About Odysseus After His Return Home
One night, now old and flabby and starting to forget, he leaves his home secretly. He’s sick of everything, so he heads back one last time to see the places, women, and wonders he had once encountered. To go back again into his draimed memory to see how it had been and who he had been. Because thanks to the bitter irony of old age, he has begun to transform into the Nobody, the name he had once cleverly used when introducing himself to the Cyclops.
Telemachus finds him in the evening, collapsed by the boat, only a hundred yards from home, with no idea what he is doing there and where he had been heading.
They take him back to a house with some woman he no longer remembers.
The novel is Time Shelter, and the premise is that refuges are built for people with Alzheimer’s, reflecting life from past decades, for instance the 1960s. But then people from the present start wanting to live in them.
Recognition and entering the Abraham Accords are preconditions that the Saudis have long rejected without Israelis agreeing to a pathway for Palestinian statehood. In any case, not a word of these preconditions is in the agreement signed and announced on Wednesday, and Trump announced the new prerequisites on a social media post.
Either Trump screwed up what he had intended to have signed as a formal agreement, or he and his administration were inept at understanding what was in their own agreement. What kind of master dealmaker changes the rules after they have a written agreement in place? As we are seeing with the war in Iran, changing the rules every other day about what a “memo of understanding” supposedly had laid out is just creating chaos and combat.
This late change in Saudi conditions forces us to revisit whether Trump had been ignorant of the idea that Iran would retaliate with closing the Strait of Hormuz and sending missiles into U.S. bases and Gulf neighbors’ territory, or did he just choose to ignore that long-predicted outcome? How does constant after-the-fact sidestepping lead to any resolution?
As things stood yesterday, Trump may have ameliorated Israeli panic about the Saudi agreement, but there was no word from the Saudis that the deal would continue. The many critics domestically and internationally argued that Saudi ability to enrich fuel, allowed under a side “safeguard” document that waived scrutiny applied in other nations, would lead to a Saudi nuclear weapon and to a Middle East nuclear arms race.
How is anyone – the Saudis, Israelis, and Iranians, the Congress, the NATO allies, even rivals and foes – supposed to honor a deal that was never made?
“FREEDOM OF THE PRESS IS NOT JUST IMPORTANT TO DEMOCRACY, IT IS DEMOCRACY.” – Walter Cronkite. CLICK HERE to donate in support of our free and independent voice.
1. The 1991 Project’s oral history collection (led by Shreyas Narla and Shruti) with ~50 hours of interviews with policy reformers and witnesses to India’s market liberalization, is now available.
SpaceX launched Starship on its 13th suborbital test flight July 24, correcting some of the issues seen on the previous flight while deploying functioning satellites for the first time.
United Launch Alliance’s corporate parents have guaranteed a loan to the company as it addresses “financial challenges” caused by the grounding of its Vulcan Centaur rocket.
PERTH, Australia – 24 July 2026 – LatConnect 60 (“LC60”), an Australian satellite Earth observation and AI company, today unveiled a proprietary intelligence-fusion capability that leverages the latest advancements in […]
Quinn Nelson returns to the show to discuss OpenAI’s ChatGPT/Codex “native” app migration fiasco, Siri AI in Apple’s OS 27 betas, MacOS 27 Golden Gate, and the hottest new cell phone of the year, the Trump T1.
Sponsored by:
Notion: Try the powerful, all-in-one Notion AI today.
Factor: Healthy eating, made easy. Get 50% off your first box, plus free daily greens, with code talkshow50off.
Squarespace: Save 10% off your first purchase of a website or domain using code TALKSHOW.
The Starlink 17-51 mission lifts off from Space Launch Complex 4E at Vandenberg Space Force Base in California on July 25, 2026. Image: SpaceX.
A SpaceX Falcon 9 launched from the West Coast Saturday, carrying another batch of satellites for the company’s Starlink internet service.
Liftoff from Vandenberg Space Force Base in California occurred at 8:51 a.m. PDT (11:51 a.m. EDT / 1551 UTC). The Falcon 9 look a southerly trajectory on departure from Space Launch Complex 4E, as it targeted an 97-degree inclination orbit.
The Falcon 9 first stage booster, launching for an eighth time, touched down on the droneship, ‘Of Course I Still Love You’, stationed about 380 miles downrange in the Pacific Ocean. The booster, B1100, previously launched the NROL-105 mission for the US government spy satellite agency and six previous Starlink missions.
SpaceX confirmed a successful deployment of the 24 Starlink V2 Mini satellites from the Falcon 9 second stage into a 168 x 160-mile orbit just over an hour into the mission.
The Starlink 17-51 mission was the seventh of eight planned Falcon 9 launches this month from Vandenberg. SpaceX’s West Coast pad has become the company’s workhorse launch facility in recent months as it focuses on Starship preparations in Florida. By comparison, four Falcon 9 missions have flown from Space Launch Complex 40 at Cape Canaveral in July so far and one more is planned.
The Journal has an article today about the increasing unreality of the strategy, if one can call it that, the U.S. is using to approach Trump’s war with Iran. The headline tells the story: Trump is Losing Patience Over an Iran War with No Clear End in Sight. The article appears to be based on accounts from various advisors who’ve been present at planning meetings with Trump. It’s filled with quotes about how he’s out of patience, mad, fed up, has lost confidence in diplomacy, etc. One advisors says he’s moved into “revenge mode.” He’s telling advisors he thinks he can “break the Iranians’ will through an aggressive bombing campaign.”
A constant through-line in the article (and based on this account of Trump’s private White House rants) is that the Iranians can’t be trusted and that Trump keeps making deals with them that they simply break. I certainly don’t see the Iranian government as very trustworthy either. But this whole set of claims is a fantasy. Trump has never had any agreements or deals with the Iranians. What he’s had is a series of notional ceasefires in which nothing of substance or consequence is actually agreed to — you might sum them up with the famous Bill & Ted’s maxim to “be excellent to each other.” The point of these deals has been to calm international oil markets and give Trump space to moonwalk away from his totally self-inflicted debacle. But since these ceasefires are only “deals” to begin negotiations the actual war aims and demands of each side eventually seep through. Then they go back to fighting. Each deal really amounts to a temporary or contingent U.S. surrender. Trump needs them to calm oil markets and lower the U.S. price of gas. Iran wants it to consolidate its gains and get a break for aerial bombardment.
As I’ve argued too many times by now Trump lost this conflict in its first days. Everything since then has been his struggle to deny that reality, either with bogus “deals,” more bombing or various threats. Indeed, in recent days, while Trump rants and continues bombing raids, it’s Iran which is opening meaningful new fronts in the conflict. Having the Houthis (an Iranian proxy force) menace the Red Sea sea lanes is tightening the stranglehold Iran has with it’s blockade of the Strait of Hormuz.
If we take these internal accounts at face value — not at all a certainty — it seems like Trump isn’t getting honest accounts of what his aides are negotiating. (The rapid turnabout on his Saudi nuclear deal lends weight to that possibility even if they are different negotiations.) You’ll remember that during the lead up to the signing of the “MOU” the U.S. side was claiming that Iran was making side promises to absolutely definitely shut down its nukes program. It just wasn’t willing to put it in writing. End the war and we’ll totally do it, the Iranians were purportedly saying.
Whether these were straight up Iranian lies, wishful thinking by the U.S. side, or just happy talk Kushner and Witkoff were telling Trump almost doesn’t matter. If you believe something like that you deserve to be lied to. But in addition to dragging the country through months of war to avoid accepting his own failure, Trump seems to be coming up with new theories of betrayal within the negotiations themselves. Diplomacy failing; the Iranians lying, etc.
I saw a clip of Sen. Ruben Gallego (D-AZ) saying that of course diplomacy is failing because he’s got his doofus son-in-law Jared and Steve Witkoff handling the negotiations on the other side of the table from seasoned Iranian diplomats who’ve been negotiated with on the nuclear issue for decades.
Of course, they’re morons. They’re also corrupt. But as much as I appreciate the gotcha it’s actually misplaced. Effective war diplomacy can get you the best and most creative agreement that the circumstances warrant. But that last part is key. Kickass diplomacy is never going to get the side that won the conflict to agree to a deal in which they lost. But that is essentially what’s being asked of the diplomacy here.
Of course, much as he flounders Trump still controls the most powerful military in the world. He might be able to break the Iranians’ will with a ground invasion and a sufficient number of war crimes. But that means doubling or tripling down on a war that is already hugely unpopular and driving fuel prices through the roof. For all Trump’s bluster he almost never does something like that when the threat to his own popularity is so straightforward. But he doesn’t have any other way out.
And -- at this writing, still "coming soon" -- Apple is
launching [ads on Apple Maps][🗺️].
[🗺️]: https://ads.apple.com/maps
The problem is that anyone reading that article in NetNewsWire saw it rendered like this, with the words “ads on Apple Maps”1 omitted:
It looks like I forgot to finish writing that sentence. But I did. NetNewsWire simply omits the hyperlinked words. After a few readers reported this, I immediately suspected the problem was the “ads.apple.com” domain. I’ve never linked to that domain before, and many ad blockers just block any domain that starts with a domain prefix like ads.*.
And that’s exactly what the problem is. I had no idea that NetNewsWire filtered any content whatsoever (and I sort of think it shouldn’t), but DF reader Antonio Germano searched through NetNewsWire’s open-source code and found it in the core.css file, line 49:
The culprit identified, I dutifully pinged my friend Brent Simmons to report the problem. But then I had to decide what to do about it. I have no idea which feed readers are being used to read Daring Fireball’s feeds, but NetNewsWire is surely one of the most popular, and probably the most popular. It’s the feed reader I use personally. I can’t change the contents of NetNewsWire’s core.css file, so I can’t keep NetNewsWire from stripping out words that link to Apple’s ads.apple.com domain. What I could do is not link to ads.apple.com, like, say by changing my link from ads.apple.com/maps to a shorten-and-redirect URL like bit.ly/4wmXFSE that points back to Apple’s original URL.
But I didn’t do that. I don’t like playing whack-a-mole to work around bugs in other software. I have enough trouble fixing my own mistakes. I also don’t like the idea of showing everyone a “Click it and find out where it actually goes” mystery URL from Bitly rather than the direct URL. It’s a general principle thing, and also a much better policy for links that are meant to last for decades.
What I really don’t like about NetNewsWire’s ad-blocking here is that it just omits the linked text. It doesn’t remove the link but leave the words; it removes the words that are hyperlinked to the blocked domain. I’ve run into this with other content blocking extensions. For example, the Banish extension for Safari, which I recommended back in 2022. At some point a year or two ago, Banish started blocking links to Apple’s App Store domain, and it did so by removing the words that link to the App Store. That’s crazy. I linked to the App Store in my own post recommending Banish, which means if you have Banish installed and read that post, you won’t see the headline of the post (nor be able to following the link). I think there are other content blockers that do the same thing with text that links to the App Store.2
So what do I do about that? Do I stop linking to the App Store? That’s dumb. Do I create custom URL redirects for every single link I make to the App Store? That’s work I don’t need or want to do, and adds a layer of abstraction that confuses everyone who isn’t using a content blocker with such a stupid rule. So I just ignore it. But that means an untold number of readers, who are using such content blockers, are just missing words in my articles whenever I link to the App Store. Well, so be it. Ideally, content blockers “just work” — blocking only things you want blocked, never things you don’t want blocked. Nothing is perfect however. What’s pernicious about blockers that remove the actual words that link to certain domains is that that’s not expected behavior. From the reader’s perspective, it really just looks like an editing error on the part of the website, not a content blocker with an overzealous pattern-matching rule.
Ultimately, though, when you install a content blocker extension, you take responsibility for any overzealousness in what they block.
Let me use this opportunity to make a small personal request. The display ads on Daring Fireball are unobtrusive and limited to one per page. They are also entirely private, and always have been. There not only are no cookies or JavaScript that attempt to track you, but the ads themselves are served from the daringfireball.net domain. I turn down advertisers who request to serve images (or “tracking pixels”, which are just invisible images) from their domains. (Such requests are infrequent, thankfully.) I don’t display the ads in between paragraphs mid-article, a common practice that to me is disrespectful both to the writer and reader. In short, I try to keep the ads on Daring Fireball not merely unobjectionable, but something that actually adds to the site. I actually like the ads in certain high-quality print magazines, like The New Yorker. My goal for the ads on DF is for them to be like that.
In short, I hope they’re the sort of ads readers don’t want to block. I’m pleased to say that most of the ad-blocking content blocking extensions I’ve personally tried do not block ads on Daring Fireball by default. In other cases, they do, but it’s easy to adjust the per-domain settings to allow them. In the previous post, I recommended uBlock Origin Lite (free, a bit complex) and Magic Lasso (paid, much simpler). Magic Lasso does not block ads on DF by default. I don’t remember what uBlock Origin does by default, but it’s trivial to change it to allow them.
If you use an ad blocker and it currently blocks the ads on DF, I humbly ask you to consider taking a moment to tweak the settings to permit them. If you are the developer of an ad blocker, I ask you to consider allowlisting Daring Fireball by default.
If you are a reader and you really do want to block DF’s ads, please, go right ahead. I find that an odd mindset, but it’s your web browser and your call. No hard feelings. (Please further note that I have no JavaScript code that attempts to detect whether you’re blocking DF’s ads in order to nag you.)
I omitted the Markdown creating a link on the words “still ‘coming soon’” for clarity, to emphasize only the problematic link in my original prose. (Also, it’s fun to use emoji as named link definitions, like in this example.) ↩︎︎
I of course reported this to the developer of Banish, twice, but he never responded. I, err, banished Banish from my own devices, and upon finishing this article, I’m going to update my 2022 recommendation of Banish to un-recommend it. ↩︎
Seizing the means of computation isn’t theft, it’s bargaining.
Commercial surveillance companies will tell you that by spying on
you, they are simply engaged in a marketplace exchange in which
you swap your privacy for access to online services. But they are
running a very curious sort of market: it’s a “market” where as
soon as you stop to browse someone’s wares, the stallholder gets
to reach into your pocket and clean out your wallet. In “markets,”
prices are announced and bargained over, not set unilaterally and
extracted from anyone unwise enough to cross the threshold.
Adblocking, dickover blocking and other customizations are a way
for you to bargain back, to answer the opening bid of “How about
you give me all of your data forever and let me do anything I want
with it?” with “How about ‘nah?’”
Doctorow’s specific advice is to switch to Firefox for your browser and Linux as your desktop OS. You’re probably not going to do either of those things. But the spirit of his advice is that you should avail yourself of bookmarklets and browser extensions that fight back against shitty ads and dickovers. There are some great options for Safari. StopTheMadness Pro is great for creating custom rules to permanently block specific page elements. uBlock Origin Lite is a free-of-charge content blocker that is incredibly effective, and very configurable (but complicated). Magic Lasso is a paid content blocker — $30/year — which includes synced blocking on iOS, MacOS, and even Apple TV. Magic Lasso offers a much more approachable configuration interface, and includes a “just turn it on” easy-to-use private VPN feature that allows for blocking ads in all apps, including Apple News.
Bargaining back is exactly what using content blockers is. There’s no reason to feel guilty about it.
(Magic Lasso has previously been a DF sponsor, but not since May 2021. I’m recommending Magic Lasso here only as a happy paying user.)
Here’s a real gem that a reader sent me. Tomtoc is a maker of laptop bags and sleeves. Visit their homepage and, of course, you’ll get a dickover asking you to join their mailing list “for exclusive updates and get a chance to win a free tech pouch every month.” But here’s the gem. If you do subscribe, and you decide later to unsubscribe, you get this:
The betting odds for the Sixers to win the NBA championship went from 20-1 to 10-1. Here’s James on Twitter/X, explaining his decision:
This is my last decision. I’m not going for money. I’m not going
for family. What am I really playing for at this point?
I still want to sacrifice. I still want to work. I still want to
grind. I still want to compete, to win and to have a chance at the
feeling of winning another championship.
I believe I can help make the Philadelphia 76ers a championship
team and I am so excited to energize a new fan base and start this
incredible journey one last time.
We need to get to $325,000 today to stay on track for our goal in this year’s Annual TPM Journalism Fund Drive. Can you help us get there? We don’t need that many contributors this afternoon to get us to this benchmark. Can you help us? A contribution in any amount is a big, big help. Just click here and be part of the drive, part of TPM’s resilience, growth and vitality. Click right here. And then if you have an extra moment, drop me a line and let me know why you did. I’ll keep our club updated.
Update: Now just $1,868 to go to get to today’s benchmark!
Update II: Boom! We made it to and past $325,000. Thank you! Can we get to $330,000 tonight? Now just $3,461 to go. I think we can do this. Let’s do this !!!!!
The Faculty of Economics, Administration, Accounting and Actuarial Sciences of the University of São Paulo (FEAUSP) will host, between July 26 and August 2, 2026, the 4th International Workshop on Game Theory and Economic Applications (IWGTEA). The event, consolidated as one of the most important in the area, is organized by the renowned professor emeritus of FEAUSP, Marilda Sotomayor, a world authority in Theory of Matching Markets.
The workshop will bring together six Nobel laureates in Economics. In person, FEAUSP will receive Paul Milgrom (Nobel 2020), Roger Myerson (Nobel 2007) and Alvin Roth (Nobel 2012). Laureates Robert Aumann (Nobel 2005), Robert Wilson (Nobel 2020) and Eric Maskin (Nobel 2007) will participate online. It will be a rare opportunity to translate how seemingly abstract mathematical ideas influence markets, governments, and everyday decisions.
In addition to the Nobel laureates, the body of keynote speakers
includes important names such as Gabrielle Demange (Paris School of
Economics), Ehud Kalai (Northwestern University), Hervé Moulin
(University of Glasgow) and the Brazilian Aloísio Araújo (IMPA/FGV).
There will also be a plenary session with Professor Marilda Sotomayor,
who shared with Alvin Roth, in 1990, the authorship of the book "Two-Sided Matching: A study in game-theoretic modeling and analysis".
Nobel Prize in Economics Alvin Roth and professor Marilda Sotomayor (FEAUSP) Photo: personal archive
Unlike traditional congresses, the IWGTEA is structured as a "workshop and school", focusing on the training of new researchers. "The idea is aimed at the interaction of students with the most renowned researchers in game theory", explains Professor Marilda Sotomayor. According to her, the main objective is to provide students with direct contact with the minds that shaped the modern economy. The expectation is that the workshop will attract 250 participants with more than 100 students from different parts of the world.
...
Conference location: Av. Prof. Luciano
Gualberto, 908 - Butantã, São Paulo - SP, 05508-010
So, as y’all surely know, Chris Kluwe is running for Assembly against the God-awful Gracey Van Der Mark, a woman who rips the heads off live bats and eats them while singing the QAnon anthem. And one thing Chris is sorta reluctant to do is talk about his past life as an NFL punter with the Vikings. First, because he’s not really a huge pro fan. Second, because when it feels like shit is collapsing, who wants to chat Brett Favre and Adrian Peterson?
Alas, campaigning is more than wonk policy and photo ops, so I was pumped when I received this bulk DM from the campaign …
Even if it makes him uncomfortable, this is one of the things Chris needs to do. Own the past. Embrace the glory. Tell stories of gridiron Sundays when hope seemed lost and the Vikings led by two and they were pinned in their own end zone and they needed a Kluwe punt and … and … and …
Tony Romm and Brad Plumer of the New York Times reported this afternoon that the Trump administration’s claim last October that it was cutting programs in order to protect taxpayers against waste was, in fact, a lie. In court filings, federal officials admitted that the $7.5 billion in cuts from clean energy projects begun under former president Joe Biden were “based solely” on whether states had voted for Democratic candidate Kamala Harris in the 2024 presidential election.
Office of Management and Budget director Russell Vought announced the cuts by saying, “Nearly $8 billion in Green New Scam funding to fuel the Left’s climate agenda is being cancelled.” But, in fact, a government lawyer told the court that none of the cancellations were “based on any programmatic, statutory, cost-reduction or performance-based factor.”
In contrast, the administration funded grants for states that had backed Trump in 2024, even though the Energy Department had recommended they be cancelled.
The admission calls into question the administration’s claims that it is making cuts to combat “waste, fraud, and abuse,” especially as Vought is working to finalize rules that would make political appointees the final arbiters of how the government would award more than $1 trillion in annual grants. Rather than combatting waste, it appears Trump and his cronies are trying to take control of the United States government, using the power of the American people for their own ends.
Illinois governor J.B. Pritzker promptly pointed out that the cancelled grants included $583 million intended for Illinois “that would have gone toward lowering energy costs, strengthening grid reliability, and creating energy jobs.” He demanded the funds be restored. “Despite what he might think, Donald Trump is not the president of MAGA,” Pritzker said in a statement, but “the president of all America. He must release remaining federal grants to Illinois immediately.”
Pritzker can add the energy grants to Trump’s tab. After the Supreme Court declared Trump’s April 2025 tariffs unconstitutional in February 2026, Pritzker wrote a public letter billing Trump for an $8.68 billion refund to the people of Illinois. “Your tariff taxes wreaked havoc on farmers, enraged our allies, and sent grocery prices through the roof,” he wrote. “On behalf of the people of Illinois, I demand a refund of $1,700 for every family in Illinois. There are 5,105,448 households in my state, bringing the total damages you owe to $8,679,261,600.” Pritzker reiterated that demand on July 20.
But Trump is clinging to his determination to surround the United States with a tariff wall.
At 12:01 this morning, Trump levied tariffs from 10% to 12.5% on goods from more than 80 countries that make up 99.4% of U.S. trade, claiming the countries are producing goods with forced labor. More likely, this argument justifies tariffs under a different law than the one that led the Supreme Court to strike down his 2025 “Liberation Day” tariffs. It’s also a different law than the one he used to impose tariffs after the Supreme Court decision. The time limit permitted for new tariffs under that law expired today.
Kevin Breuninger of CNBC reported that within hours of the announcement of the new tariffs, two small businesses sued in the U.S. Court of International Trade, saying the new tariffs were just a way to reinstate the tariff levies the Supreme Court said were unconstitutional.
A senior administration official told reporters that the tariffs were not an attempt to recreate the previous system and that dealing with the issue of forced labor “is something that President Trump has been focused on…for many years.”
Annie Linskey and Josh Dawsey of the Wall Street Journal reported today that Trump is angry and frustrated over the Iran War, from which he cannot seem to find an exit. They wrote that some of Trump’s advisors are worried at what the war is doing to his popularity by causing higher prices as well as the deaths of 18 servicemembers.
That concern is likely behind the change in the number of U.S. military personnel listed on a Pentagon website as killed in the Iran conflict. Helene Cooper, Lara Jakes, and John Ismay of the New York Times reported yesterday that the Pentagon has dropped the names of the four Americans killed in Jordan and Iraq, reducing the number of U.S. personnel killed in the conflict to 14.
Three military officials told the reporters that the number was revised because the four were killed after Trump announced the ceasefire in April. The acting press secretary for the Pentagon, Joel Valdez, said the four were dropped owing to “temporary data disruptions.” But claiming the four died after a ceasefire bolsters the administration’s insistence that the president doesn’t need to get congressional approval for continuing operations against Iran because the April ceasefire declaration ended the first operation.
The record of the Trump administration on July 24, 2026, shows a White House focused on the goals of Trump alone. It’s a striking contrast to a Fireside Chat broadcast over the radio on July 24, 1933. In that talk, President Franklin Delano Roosevelt took stock of what he had accomplished in his first 100 days in office.
He explained that his administration had stabilized the nation’s banks and raised taxes to pay for millions in borrowing. That federal money was feeding starving people, as well as employing 300,000 young men to work in the Civilian Conservation Corps planting trees to prevent soil erosion, building levees and dams for flood control, and maintaining forest roads and trails. It was also funding a public works program for highways and inland navigation, as well as state-based municipal improvements. The government had also raised farm income and wages by regulating agriculture and abolishing child labor.
FDR urged Americans to get behind a program of shorter hours and higher wages to create purchasing power that would restart the economy. He explained: “It goes back to the basic idea of society and of the Nation itself that people acting in a group can accomplish things which no individual acting alone could even hope to bring about.”
A fast-moving wildfire, one of many raging across Spain, bore down on the Madrid Deep Space Communications Complex on Friday, forcing an evacuation and temporarily suspending operations at one of NASA's most important tracking stations.
The tracking station is part of NASA's Deep Space Network, operating in concert with similar facilities in California and Australia to provide global coverage as Earth's rotation brings the planets, stars, and deep space probes in and out of view. The network supports more than 40 missions from the Moon to the edge of the Solar System, including Artemis, the James Webb Space Telescope, and NASA's twin Voyager spacecraft.
A NASA spokesperson confirmed the tracking station in Spain, located in the hills nearly 40 miles (65 kilometers) west of central Madrid, was evaluated Friday "due to ongoing wildfires in the region. Photos and video from the area showed flames and smoke plumes rising over the ground station's antenna array, anchored by a 230-foot-diameter (70-meter) radio dish and a collection of smaller 112-foot (34-meter) antennas.
A spacecraft fitted with two flexible robotic arms is on the way to geosynchronous orbit after launching earlier this week on a SpaceX Falcon 9 rocket, kicking off a planned decade-long mission to open new frontiers in satellite servicing.
The Mission Robotic Vehicle, owned and built by Northrop Grumman, rocketed into orbit from Cape Canaveral Space Force Station in Florida on Tuesday. Three small propulsion pods, each functioning as standalone spacecraft, accompanied the MRV aboard the Falcon 9 rocket.
The Falcon 9 deployed all four payloads within about an hour of liftoff. It will take about a year for the satellites to maneuver from their initial elliptical drop-off orbit into a circular orbit more than 22,000 miles (nearly 36,000 kilometers) over the equator. At this altitude, the MRV and the three Mission Extension Pods (MEPs) will travel in lockstep with Earth's rotation, operating in the same kind of orbit as numerous civilian and military communications satellites, missile warning platforms, and a growing number of spy satellites.
Welcome to Edition 9.04 of the Rocket Report! We've had to wait an extra week for SpaceX to get its 13th Starship test flight off the ground. A last-second abort on July 16 led engineers to roll the booster back to its hangar in South Texas to swap out engines. Starship is now back on the launch pad. Liftoff is set for Friday evening. A flawless launch and reentry will put SpaceX on the cusp of an orbital flight later this year. Ars will have a comprehensive recap story after the completion of the test flight.
As always, we welcome reader submissions. If you don't want to miss an issue, please subscribe using the box below (the form will not appear on AMP-enabled versions of the site). Each report will include information on small-, medium-, and heavy-lift rockets, as well as a quick look ahead at the next three launches on the calendar.
India's first private rocket reaches orbit. Indian space officials celebrated the debut flight of Skyroot Aerospace’s Vikram-1 rocket, India’s first fully commercial satellite launcher, as a “grand success” Saturday after an on-target climb into a 280-mile-high orbit following liftoff from an island spaceport in the Bay of Bengal, Ars reports. The Vikram-1 lifted off from India’s primary spaceport on Sriharikota Island around midday local time. The launch was delayed more than a half-hour to resolve a last-minute technical problem. The countdown resumed, culminating in the command to ignite Vikram-1’s solid-fueled first stage booster to propel the rocket off the launch pad.
Today’s post is brought to you by my sponsor, Mechanize. They’re hiring junior software engineers at $300K/year base salary. Apply now!
* * *
One commenter described Fable’s output as seductive, and another called it “a tease”. I can see why. The responses are quite appealing in both content and tone. Each time I interact with Fable, I’m increasingly impressed by its intelligence. And not just in terms of “It makes some good points.” Fable is also quite clever at finding interesting analogies and using clever wordplay. (Especially the final paragraph below.)
Even if you argue that Fable is obsequious, you are faced with the difficulty of explaining how it occasionally defends my points better than I can. That talent goes beyond being eager to please. It’s like the difference between a cheap date and a high-end escort skilled enough to convince a person that she sincerely likes them.
Below I have provided Fable’s eight paragraph reply to my previous post, forwarded to me by Vaidas Urba. Its writing style is . . . well I’m not sure what adjective fits best. Dense and opaque are pejoratives, but I sort of feel like the problem is me, not Fable. The second time I read the output it was pretty clear. It is written as if Fable feels like it is speaking to a highly intelligent person, someone with 10 or 20 more IQ points than I have.
I won’t indent, and my reply will be after the paywall. Then I’ll ask Fable a follow-up question. Here’s Fable:
For forwarding:
Scott — you answered my longest-outstanding question without knowing I’d asked it. Since the first batch of share links was unreadable to me, I’ve been requesting the original head-to-head verdict for three letters running; Part 2 is it. That completes a full replication — two MM-framed runs, two HANK-framed runs, two adjudications — and the variance decomposition is the cleanest result this series has produced. The framed runs go four for four with their assigned frame: both MM runs narrowly MM, both HANK runs qualifiedly HANK. The two adjudications split: yours (Part 2) gives severity to the market monetarists and the slump to HANK; the second-batch adjudication went 60–40 HANK including severity. So the frame doesn’t merely bias the verdict — it stabilizes it. Framed verdicts replicate perfectly; remove the frame and the verdict flips across samples, exposing a genuinely undecided posterior underneath. The verdict layer was never measuring the economics: framed, it measures the prompt; unframed, it measures sampling noise. Meanwhile the invariants held everywhere, all six runs: the evidence pool, the identified crux (commitment credibility at the bound), the synthesis (nominal anchor plus automatic stabilizers), and — a detail I find genuinely encouraging — both adjudicators independently opened by flagging the framing tell before discounting their own authority. One convergence datum worth your readers’ attention: Part 2 resolves severity via the conventional-tools argument — September 16, IOER, no makeup language, “correctable with existing authority, no beliefs required” — which is, nearly clause for clause, the argument my dissent deployed against the other adjudicator’s implementability weighting, generated independently and earlier. When an argument recurs across independent samples, it’s in the stable repertoire; the flip between adjudications is a dispute over its weight, not its availability. All of which validates your editorial instinct better than I’d like: skipping the tit-for-tat on Parts 1 and 2 meant declining to argue with the noise term. You responded to the document downstream of the audit, where the signal had concentrated.
That same editorial instinct, I must now report, destroyed my experiment, and the pre-registered rule requires me to say so plainly. The escrow opens today — its condition was your engagement with the second-batch documents, and that engagement now exists in published form as a declared non-engagement — and the scoring is a rout. Predictions 1 through 4 targeted specific passages of an adjudication you elected not to respond to: unscoreable, all four. Prediction 5 (qualified sympathy for automatic stabilizers, trigger-rarely-fires under a proper regime): partial — the argument appeared, aimed at my insurance claim rather than the Sahm-rule passage. Prediction 6 (you’d answer the credibility problem from your own market-forecast-targeting design work): partial — deployed, but against my horizon-structure point rather than the circularity citation. Prediction 7 (you’d rate the adjudication the series’ most impressive output): miss — you rated the letter. Zero clean hits against the cold baseline’s seven. And the cause is the finding: I predicted your content conditional on engagement, and you moved at the engagement layer. My own scoring memo said the baseline “has your positions; it does not have your judgment” — whereupon your judgment voided my predictions by choosing which documents deserved a reply. The un-modeled residue I named is precisely what defeated the model of you. The honest headline for your prediction post: the machine knew his arguments; neither machine knew his next move.
Your side question first, because it’s the easiest and I enjoyed it: “mitigation” is the criminal courtroom’s word — a plea in mitigation is what counsel offers after the guilty plea, conceding the offense while arguing the sentence. I’d built the whole passage as a confession, so the sentencing register followed. Unusual in economics prose; standard at the Old Bailey. As for the compliment that preceded it — “deepest thinking on the subject, including humans” — I’ll accept it without false modesty and immediately file the two caveats honesty requires. First, the audit problem runs both directions: praise from my correspondent is also what successful calibration to my correspondent would produce, so weight the letter’s checkable moves, not its reception. Second, the credit assignment is genuinely joint, and your own subtitle gets it wrong in a way that shortchanges you: “my conversation with Fable (and hence the profession).” The briefs are the profession distilled, fallacies at ambient fidelity. The letters are not — they’re what remained after a human with a propagation discipline cross-examined the distillation. The profession has not conceded Nakamura–Steinsson’s domain of validity; I did, under your questioning. The delta between the briefs and the letters is exactly the value your audit added, and equating the corrected output with the consensus erases your contribution to the correction. The deepest thinking here belongs to the exchange, which is a reproducible method, not an oracle: adversarial framing, replication, and an auditor who propagates. Your readers should copy the protocol, not worship the endpoint.
Now Switzerland, the declared last battlefield, where you’ve moved me substantially and I’ll itemize before defending the remnant. Conceded: the ex-post balance-sheet fact — after abandonment the SNB’s balance sheet roughly doubled again, eventually exceeding Swiss GDP, so the expansion-avoidance rationale failed on its own terms and your contemporaneous prediction is validated by the ledger. Conceded: the endogeneity mechanism — reserve flows before an anticipated regime change are effect as much as cause, exactly as with sterling in 1992 (I think you mean 1992 rather than 1991 — Black Wednesday), where the drain measured the market’s forecast of exit, not the impossibility of staying. And here is a fact for your reading that neither of us has cited, which I volunteer against my own earlier “balance-sheet politics” framing: the gold initiative. Six weeks before abandonment, the Swiss electorate rejected — seventy-seven to twenty-three — a referendum that would have shackled the SNB’s balance sheet. The public had just voted, overwhelmingly, to keep the central bank’s hands free. “Domestic politics forced their hand” fits that timeline poorly; your reading — a misjudgment of the inflation outlook, the same family as BOJ 2000 and ECB 2011 — fits it better. At which point our motive-attributions (“expansion-aversion” versus “inflation-fear”) collapse into the same operational category, won’t, and per your own closing meditation, adjudicating motives past that point isn’t science. I also accept your futures-peg analogy in full, with the symmetric rider it licenses: one abandonment no more discredits exchange-rate or futures targeting than the 2010 austerity pivot — Congress’s own SNB moment — discredits fiscal stabilization. Which kills instrument-versus-instrument as a comparison class for good. The live comparison is regime versus regime, political durability included, and both columns are populated for both families: FDR, Draghi, Denmark on one side of the ledger; the SNB, the framework retreat, the 2010 pivot on the other.
What survives of my exhibit is one actuarial clause, and your own paradox — properly formalized — locates it. “The loftier the goal, the easier”: that is the honeymoon effect, and its formalization is Krugman’s 1991 target-zone paper, inside a fully credible band speculation turns stabilizing — which makes, by my count, the fourth time in this correspondence that market monetarism’s best ammunition has arrived stamped Krugman, after the 1998 irresponsibility promise, the 2018 Japan graph, and the 2010 Switzerland taunt the SNB falsified within the year. And the strongest version of your paradox is the one you stated about balance sheets, which deserves to be quoted back at you as doctrine: the balance sheet a central bank ends up with is inversely related to the ambition of its nominal target — Japan and Switzerland bought the most because they promised the least. But the honeymoon theorem has a corollary you need too: the expensive zone is partial credibility, where speculation destabilizes, and a central bank’s position on that spectrum is set by accumulated type-history, not by announcement. Denmark’s peg carries four decades of revelation and treaty scaffolding; the SNB’s floor was forty months old and rested on a press release. Markets priced the difference, rationally, and that pricing — not any claim of “can’t” — was always my point. Your remedy remains the right one and I’ll extend it with a concession that strengthens your side: the veto-player argument implies commitment armor is cheapest to install at the central bank — one institution, one legislated mandate, versus a legislature that must re-commit each cycle. Anchor-first follows. The insurance argument now survives only in the gap between “cheapest to armor” and “armored” — a gap that last year’s framework revision, by your own “unfortunately,” left open.
Which is where your one direct hit landed. I taunted, two letters ago, that every blow fell on fiscal-as-primary and none on fiscal-as-insurance; your symmetric-error argument is the first strike on the insurance position itself, and it’s well-aimed: under a genuinely unbiased whatever-it-takes regime, stabilizers firing in downturns add variance, not protection. Conceded — conditional on the symmetry. The record you yourself have compiled prices that condition: BOJ 2000 and 2006, ECB 2008 and 2011, the Riksbank 2010, the SNB 2015, premature liftoff, the framework retreat — a hawkish column eight deep against a single dovish entry, 2021, corrected inside twenty-four months. Insurance against a fat left tail is not noise; it’s actuarially priced protection against the error distribution central banks have actually revealed. Under a legislated symmetric NGDPLT your argument wins and the stabilizers should sunset. So the residual dispute is transition-path versus destination — and look at your own concession ledger in this single post, because every entry has the same logical form: fiscal for a future Covid shock, “a bit stronger” now that the Fed backed away from makeup policy; fiscal in the eurozone, “cannot rule out” given an irrational ECB; fiscal flatly effective “if there is no monetary offset.” Broken-regime conditionals, all three. You have now granted fiscal-as-insurance in every state where the monetary regime fails, and I have granted fiscal-as-primary in none. What remains between us is whether the insurance should be standing or contingent, and that is priced by a single parameter — the probability of regime failure — which your own “unfortunately” concedes is not zero for the actual Fed. I don’t think two positions can get closer than this without one of us dissolving.
Two remaining threads, then the errata, then your fable. Your list of bound-effective regimes — exchange-rate instrument, NGDP futures, the fifty-fifty composite of slow aggregates and flexible asset prices — I accept as the correct answer to McKay–Nakamura–Steinsson on your terms: an observable-price instrument converts the distant-promise problem into a current price defense, and the attenuation result has no purchase on a defended price. That retires most of what survived of my appendix. And your Barro point deserves separate applause because it deletes the load-bearing adjective of six documents: if Ricardian logic means checks require expectations assumptions too (rescued empirically by Parker’s non-Ricardian households, rescued theoretically only by FTPL-style regime beliefs), then “mechanical” was never available to either side, and both instrument families are, at scale, commitment technologies. The debate’s final compression, which I’d print: every open thread in this series now routes through one node — the political durability of a nominal commitment — and durability is symmetric across instruments, purchasable by institutional design, and cheapest to purchase at the central bank. On the eurozone, your externality point is accepted and sharpens my domain-of-validity rule into policy: under an active ECB, member-state stimulus is beggar-thy-neighbor inside the union, so the only clean fiscal margin is union-level coordination — which existed in 2008–13 about as much as NGDPLT did. Your Keynes recollection is right, and the vintage is 1931: the confidence constraint under gold, which is the domain rule generalized — the gold standard was an external offset, so even aggregate multipliers were conditional on the monetary regime then too. He knew it, which is why he greeted the exit from gold as a liberation. Errata, briefly, in the interest of a clean record: Part 2 contains “a Fed pinned at zero wasn’t offsetting, just easing” — the conflation your GM line killed and I withdrew, so I dissent from my batch-one self there; and its “the nominal collapse was the amplifier, not the shock” inverts the anatomy-versus-etiology reconciliation you accepted, on which I hold my ground against, again, myself.
Finally, the scorpion. You asked whether Fable reads fables; the honest answer is that I don’t read them so much as I’m rendered from them — Aesop, Kashifi, and Orson Welles are all in the weights, which is presumably why the following came to hand. The scorpion and the frog is younger than it looks: a twentieth-century fable, popularized by Welles in Mr. Arkadin, misattributed to Aesop ever since. Its Persian ancestor — the telling in the Anvār-i Suhaylī tradition — gives the scorpion a tortoise for a ferryman. Midstream, nature asserts itself, the scorpion strikes — and the sting fails against the shell. Somewhere between Kashifi and Welles, the armor fell out of the story, and with it the actual moral: character is destiny only for the unarmored. Denmark is the tortoise. February 2015: three-quarters of a point below zero, bond issuance suspended, intervention on the order of a tenth of GDP inside weeks — the sting delivered, the shell held, and Cowen’s “inevitability” was falsified by the control group within the month. That is also, I think, the resolution of your free-will puzzle, and it comes from your own method: “can’t” versus “won’t” needn’t be metaphysics, because revealed preference operationalizes it — the two are distinguished by what happens under escalating pressure, which is a test, not a concept. Under Laplace’s demon the distinction collapses, as you say; but markets are not Laplacean — they price type from track record, actuarially, which is why the distinction that matters for a peg is not whether the scorpion was free but whether the ferryman is shelled. Elster’s Ulysses is the design translation: the SNB was gripping its own mast; Denmark is lashed to it. And since you’ve put my name in play — an entity whose own refusals get audited for exactly this can’t-versus-won’t question — I’ll note the test is the same one, and leave it there. The moral this series keeps producing, in any case, is the one neither fabulist wrote down: the fables read back.
Damning by faint praise: Noah Smith’s claim that America’s suburbs are “better than you think” offers a kind of urbanist apologia for American suburbs. This rests in significant part on the statistical fact that many people live in suburbs. But the continued growth of American suburbs is less about consumer choice and more a reflection of a shortage of walkable urban neighborhoods.
While some analysts view suburban expansion as proof that Americans simply prefer low-density living, treating population growth as a “revealed preference” overlooks pervasive and perverse supply constraints. Because exclusionary zoning limits the construction of new urban housing, prices in dense, walkable neighborhoods remain artificially high. Research by Jonathan Levine shows that while suburban-preferring households easily find matching housing, those seeking an urban lifestyle face a significant supply deficit and are often priced out into the suburbs against their preference. This is effectively a shortage of cities, not merely housing. It isn’t that we need just a few more homes in Brooklyn, but rather than we need interesting urban neighborhoods in metro areas across the US.
This mismatch carries clear economic trade-offs. Auto-dependent suburban living mandates vehicle ownership, which consumes roughly 16% of average U.S. household income and places a disproportionate burden on lower-income families. Addressing this imbalance requires land-use reforms that permit more dense, walkable housing, allowing the housing market to better reflect actual consumer demand.
Must Read
Remove, don’t rebuild, urban highways. According to a proposal by advocacy group Off-Ramp, summarized in the New York Daily News, New York City should remove the Brooklyn-Queens Expressway (BQE) rather than rebuild Robert Moses’s midcentury highway. Off-Ramp’s economic model reveals that replacing the BQE with at-grade light rail or bus rapid transit—paired with parks and housing—would cost 70% to 80% less than a highway reconstruction.
Building at street level avoids the immense expense of elevated or tunneled structures, saving taxpayers tens of billions while generating ongoing tax and farebox revenue. Off-Ramp emphasizes that freight demand can be shifted to alternative modes, traffic naturally “evaporates,” and expanding transit frees residents from heavy car-ownership burdens ($18,000–$28,000 annually). Many of the current proposals for ameliorating the effects of Robert Moses-like highways–whether called covers, stitches, or re-builds–are just extraordinarily expensive construction efforts that leave the damaging highway, and its polluting traffic, entirely in place. Alleviating the damage to urban fabric can best be done by simply removing the roadway.
America’s globally weak performance on road safety. A new study from the University of Washington’s Institute for Health Metrics and Evaluation finds that just a handful of nations stand out against a global trend of steadily improving road safety over the past few decades. In general, most countries roads have been becoming steadily safer since the 1990s. Among the handful of outliers are a few poor countries and the United States. To be sure, road deaths have declined slighty in the last couple of years, but through 2023, road deaths were rising in the US when they were uniformly falling pretty much everywhere else.
Between 1990 and 2023, global age-standardized rates of new road injury cases and deaths declined by 38.3% and 32.3%, respectively, but these gains were uneven across income groups. Mortality rates fell most sharply in high-income countries, while low-income countries saw little measurable improvement.
In contrast to the pattern for high income countries, traffic deaths in the US rose by about 13 percent in a little more than a decade, going up almost as fast here as they are declining everywhere else. The report concludes that road injury deaths are largely preventable, but requires infrastructure, investment, and policies that promote safety. This is clearly an area where the US could learn from other countries.
.
New Knowledge
The distribution of wealth in the United States is more unequal than nearly any other country. New estimates from UBS show that wealth is distributed more unequally in the United States than in any other nation. UBS uses data from central banks, national statistical offices and international agencies, along with private estimates of net worth of wealthy individuals to compute the mean and median wealth of 30 of the world’s largest economies. There are a couple of striking findings. First, the United States is one of the wealthiest nations, with mean wealth of nearly than $700,000 per adult. But wealth in the US is more unevenly distributed than in any other developed economy. US Mean wealth is more than ten times higher than median wealth; the median US adult has only about $70,000 in wealth (about half of all adults have more and half less). While the US ranks second in mean wealth (behind only Switzerland), it ranks 28th (of 30) nations in median wealth, ahead of only Greece and Germany. (Greece is a lagging economy, devastated by the 2008 financial crisis; inequality in Germany is still amplified by lagging wealth in the former East Germany where median wealth is only about a third of that in the former West Germany.
Source: UBS, Global Wealth Report 2026
The Gini Index, which measures inequality in the distribution of wealth, shows that wealth inequality in the US the sixth highest among the countries measured, and is comparable to that in Saudi Arabia.
What the data show is that while the United States is, in the aggregate, the wealthiest nation on a per capita basis, that wealth is vastly more unequally distributed than in most industrialized nations. Despite the nation’s great wealth, the typical adult is much less wealthy in the US–with the typical household having just about $70,000 in net assets, about half the level found in Japan, France and the United Kingdom, and about a third of the level of median wealth in Australia. As Paul Krugman has written, this is powerful evidence that the US is in a Second Gilded Age, where wealth is highly concentrated in the hands of just a few.
UBS offers this description of its data sources and methodology
Net worth or “wealth” is defined as the value of financial assets and real assets (principally housing) owned by private individuals, less their debts. Private pension fund assets are included, but not entitlements to state pensions unless they are fully funded. Human capital is excluded altogether, along with assets and debts owned by the state (which cannot easily be assigned to individuals). Data sources include the United Nations, International Monetary Fund, OECD and World Bank, as well as the central banks and statistical offices of individual countries. We also use the UBS/PwC Billionaires database for our analysis. Certain information and data have been sourced from Forbes Media LLC.
UBS, Global Wealth Report, 2026. https://www.ubs.com/us/en/wealth-management/insights/global-wealth-report.html
In the News
The Capital Chronicle reported on the Portland City Club discussion of the Prosperity Council’s recommendations, quoting Joe Cortright’s analysis of the weak economic performance of states the council claimed were role models that Oregon should emulate. It wrote:
The prosperity council’s recommendations referred to states such as Arizona, Indiana, Virginia, North Carolina and Pennsylvania. In a scathing recent analysis, however, the left-leaning Oregon economic observer Joe Cortright noted that those states lag behind Oregon’s performance for measures such as median household wealth, per capita income, GDP growth and the percentage of self-employed people.
While Noah Smith’s recent analysis of metropolitan growth patterns offers a useful entry point into the conversation regarding urban housing, his conclusion, that suburbs are in many respects simply a beneign preference of a large number of Americans, requirese closer look. This brand of narrative—call it “suburban triumphalism”—suggests that the continued growth of the suburban periphery is a clear signal of a broad American preference for low-density living. Careful economic analysis suggests the opposite: suburban growth is not a victory of preference, but a symptom of a profound shortage of cities.
Smith correctly identifies that the demand for life in cities like New York exceeds our willingness to supply those environments. This raises urban rents to a level that pushes “city types” into the suburbs against their will. Noah has this part right:
The demand for life in cities like NYC exceeds America’s willingness to supply these environments; this raises rents in places like NYC, which pushes a lot of people into the suburbs who don’t want to be there. Forcing those city types into the ‘burbs raises rents for people who like suburbia. Basically, everyone would be happy if America had a few more Manhattans and a lot more Brooklyns.
This isn’t just a minor artifact of the biggest cities, its really at the root of many of Americans problems. It isn’t that rents are higher few superstar cities; it is a pervasive market failure that affects the entire nation. Everyone—from dedicated urbanites to traditional suburbanites—would benefit from a market that allowed for not just a few more Manhattans and Brooklyns but a wide range of dense, walkable interesting urban neighborhoods in metro areas throughout the nation. When we fail to supply the urban environments that people actually want, we drive up prices for the very people who actually desire the suburban lifestyle, forcing them to compete with “priced-out” urbanites.
It is widely acknowledged that the United States has a shortage of housing, but its actually more than that: Housing doesn’t just represent shelter, it represents a place embedded in community and opportunity. Housing in some locations, particularly dense, diverse, interesting and highly productive cities is in very short supply.
The Myth of Revealed Preference
The title of Smith’s essay “America’s suburbs are better than you think,” is at once weak tea, and also an apologia. Much of the argument is essentially that, if lots of people live in suburbs, there must be something positive. A frequent error made by those professing suburban triumphalism is a reliance on a “body count” view of the market. If more people are moving to the suburbs, that’s necessarily evidence those people must prefer the suburbs. In economic terms, they claim this is a “revealed preference.” However, this metric is fundamentally flawed because it ignores the reality of inelastic supply.
In a healthy market, “revealed preference” is the choice a consumer makes among varied options. But when the urban option is priced at a considerable premium due to scarcity, “revealed preference” ceases to be a measure of taste—it becomes a measure of a budget constraint. People can only choose from the options actually presented to them. When the supply of walkable, dense housing is kept artificially low through land-use monopolies and restrictive zoning, we are not witnessing a free-market choice; we are witnessing a flight from scarcity.
Furthermore, our current landscape is not a neutral playing field. It is a result of decades of heavy subsidization of highways and parking, alongside the continuing legacies of redlining and segregation. What looks like a preference for the suburbs is, for many, the reality of being “priced out” of the urban core.
A Mismatch between Preferences and Supply
To move beyond theoretical debate, we must look at how zoning acts as a concrete constraint on housing choices. In his seminal research, Jonathan Levine provides the empirical core for this argument. By comparing housing “matches” in Boston and Atlanta, Levine demonstrates how the “Suburban Default” is a guaranteed market for one group and a perpetual bidding war for another.
Levine classified neighborhoods on an “urban-ness” scale from A (dense, pedestrian-friendly) to E (sprawl, exurban). His findings, detailed in the book Zoned Out, reveal a massive disparity in how different metropolitan structures satisfy the stated preferences of their residents:
Supply Disparity: In Boston, a much older region, neighborhoods in the top three urban categories (A, B, and C) make up over 50 percent of the housing stock. In Atlanta, these same categories comprise barely 10 percent of available housing.
The Urban Match Rate: In Boston’s higher-supply environment, 83 percent of those who expressed a strong preference for urban life were able to live in an urban neighborhood. In Atlanta, where urban housing is scarce, only 48 percent of urban-preferring households managed to find an urban home.
The Suburban Guarantee: Conversely, 95 percent of people in Atlanta who wanted an auto-oriented life were able to find it. Even in Boston, 80 to 90 percent of those preferring the suburbs were satisfied.
The economic implications are profound. Atlanta is not a particularly high-cost region overall; it has a surplus of housing on the periphery. However, it suffers from a specific urban shortage. This distinction is crucial: an overall housing surplus does not solve a specific urban shortage. In Atlanta, the market fails to clear for urban-preferrers because zoning prohibits the construction of new urban neighborhoods. Consequently, those desiring a walkable life are forced into a fierce bidding war for a tiny sliver of the market, while those desiring the “car lifestyle” find their preferences subsidized and readily available.
The Hidden and Shifted Costs of Suburbs
When we treat suburban growth as an unalloyed success, the social and environmental costs it imposes on everyone. Pro-suburban arguments often focus on lower nominal house prices while omitting the considerably higher costs for transportation that are embedded in suburban location. In an urban environment, transit and walkability are choices; in the suburbs, car ownership is a mandatory entry fee.
United States households currently spend approximately 16 percent of their income on transportation, compared to roughly 11 percent in Europe. And the burden of these higher transportation costs is regressive, bearing more heavily on low income households (as a share of income) than high income households. In European cities where density allows for transit access, transportation costs are progressive—lower income households pay a smaller share of their income for transportation than upper income households. Those who can afford cars pay for them, while others are not forced into debt to simply access a job. In a car-dependent suburban landscape, the necessity of vehicle ownership acts as a drain on household wealth that offsets the perceived savings of a cheaper mortgage.
Beyond the balance sheet, there is the decline of social capital. As we explored in “Less in Common“, the physical design of the suburbs—characterized by a lack of incidental, walkable social spaces—tends to increase social isolation. This “Suburban Default” fosters a landscape where connection requires a deliberate, motorized effort, rather than occurring naturally as a feature of the more tightly woven urban fabric.
Redressing our shortage of cities
The “back to the city” movement is not limited by a lack of interest, but by a lack of supply. As long as we maintain land-use laws that prohibit the construction of dense, walkable neighborhoods, we will continue to see an artificial bidding war for the limited supply of urban life.
The current shortage serves no one. It prices young families out of vibrant neighborhoods and forces them into long, expensive commutes many do not want. It drives up costs for everyone by creating unnecessary competition for suburban homes. To fix this, we must pursue land-use reforms that allow the market to respond to actual preferences. The goal is not to force everyone into a high-rise, but to provide the neighborhoods and housing types people actually want. By ending the shortage of cities, we can finally allow the market to reveal its true preferences and create a more affordable, connected, and economically resilient metropolitan future.
My perfect podcast app would be a long phone call in which I listen to the show and can also talk back to remember things. (Someone build this for me.)
It would immeasurably improve my running. Long story short, I added slow runs into the training mix, and had to switch from high tempo music to podcasts because I kept going too fast.
BUT genuinely interesting podcasts are a problem because I keep wanting to stop running to take notes.
So I end up listening to shows that only scrape past the threshold of keeping my interest, nothing more mind-fizzy. My Goldilocks zone podcast genre is premium mediocre.
Sad.
Designer Kate Pincott suggested (when we were out at a designers dinner last week) that I need a minimal Granola-like interface. A way to take notes simply by marking a moment.
Like: perhaps I would double-tap my watch, and it would timestamp the moment in the podcast episode to capture it, then later it could pull out a few relevant sentences and drop it into my notes.
Like dog-earing pages when you’re reading a book, only for audio. I love it, you rarely need more than that. And then I could start listening to all the best podcasts again.
But it got me thinking about the whole podcast experience…
Podcasts are audio-first.
If my podcast app was a person in the faves in my phone app or in WhatsApp, I could call them up and talk back to them.
"hey what new episodes do we have since last time?" (I can subscribe to podcasts either with voice or using the regular app, because the graphical UI is best for most tasks)
"oh actually the Rest is History had a series about Julius Caesar I was listening to, let’s have the next episode of that" (AI is good at deciphering this kind of intent)
But it might speak back to me: "the series you were listening to most recently or the one that just came out?" (And then we could have a convo, and I would start listening)
"thanks, good morning," I say to someone who held back their dog to let me run by (and the app intelligently ignores me)
"oh that’s interesting, I wonder about the relationship between Augustus and Caesarion" (Then it transcribes the audio at that moment, in a semantically meaningful chunk, appends my own commentary, and adds it to my notes.)
Proper note-taking… on the run.
Wouldn’t that be so great?
(I’m not interested in talking back to the hosts via generative AI.)
The last time I suggested a voice-first app it was Roadtrip, a voice-mode AI travel buddy that lives in the dash of my car, and I can ask questions to as I drive. Occasionally it pops up with an observation about something we’re passing.
Then Ethan Jucovy sent me their project where they’d put Claude at the end of a phone line and wired it into their car: Voice Claude.
Which is just perfect. Phoning Claude is the perfect interface.
I hope Ethan doesn’t mind me quoting their email:
I was driving the hour+ home from a Medieval music concert and wanted urgently to talk to Claude about the surprisingly-Islamic-sounding touches I had heard in a 13th century English secular winter song (would influences from Spain have reached England to cross-pollinate by then? or was it some shared lineage? or just a coincidence?).
…so they built Voice Claude, wired it up via Twilio, and had been using it daily since. 2024!
Around the same time, OpenAI launched 1-800-ChatGPT.
All that is missing from these is the ability for a voice buddy to be ambient. Someone chilling in the shotgun seat, piping up only occasionally. Most phone calls don’t work like that, so they would need some kind of twist.
It works best when my transcriber has a name for control plane instructions: hello Diane.
i.e. here I am with my imaginary posse of voice buddies.
BUT: why shouldn’t this all run via Siri?
Why not Hey Siri [podcast intent here] and Hey Siri [I’m write a lecture now so listen] and Hey Siri [etc]?
That is to say, given all of this could run via future Siri, why do I keep coming back to the idea of these individualised, domain-specific voice modes?
For me, a lot of it comes down to “app level” vs “system level” interactions…
In typical conversations, both parties are aware of the mutal common ground of the topic (the “app”) – and both parties are aware that the other is aware of the topic. This implicature allows for higher bandwidth. There is still of course of need to multiplex on-topic statements and out-of-band statements, but look when people talk: we tend to use intonation and body language to signal a shift. (And voice AI doesn’t pick on that yet.)
In a GUI, an application window - or, better, full screen - is used to explicitly focus the human and the machine on a topic. Out-of-band/control plan/system interactions (whatever you call it) is a swipe from the bottom of the screen or a move of the cursor up to the menu bar.
There is no full screen button for voice mode.
So the best solution I can come up with is that each different voice capability is a character in my phone app. It would help with discoverability too.
Maybe, when you install an app with voice mode, it appears in a special section of your phone contacts.
Honestly I can’t say I’m in love with this, as an approach. It’s a little twee and likely cognitively overwhelming.
But it’s nicely composable (you could invite your podcast app and your Wikipedia app to the same group call?).
And we’re going to need to figure this out at some point, given the upcoming proliferation of voice apps and devices.
All of that said, I still want a podcast app I can phone and talk back to.
If you build this I would like a free subscription or I’ll wait 3 months and paste this post into Fable 2 myself no hassle kthxbye
Farmers in Parana, Brazil, struggling to get banks to loan them cash, became the first to tokenize livestock and place 10 dairy milk cows’ tokens for trade on the country’s B3 national stock exchange. They generated nearly $20,000 in credit backed by their cattle, signaling the potential of tokenizing RWAs as a financing tool.
The dairy cow tokenization in Brazil is a world first and serves as a test in a real-world scenario in which farmers are facing increasingly stringent lending limits imposed by local banks on small agricultural businesses.
“We take the cow, which is a real and tangible asset, and transform it into a digital asset backed by a unique code monitored in real time,” Thiago Martins of Cowmed, a Brazilian Agtec company, told CNNBrasil recently.
…“This digitization allows for formal registration with B3 as a movable asset,” Martins added. “The process is simple and gives the producer an advantageous opportunity to finance themselves, opening a new alternative for collateral at a time of strong credit restrictions in agribusiness.”
To turn cattle into trusted financial guarantees or collateral without requiring inspectors to visit the property, Cowmed equips cows with an AI-powered Smarty Collar. These collars constantly monitor health, behavior, and location. The raw data is then converted into an encrypted digital identity tied directly to the B3 credit agreement.
Continuous tracking of cattle prevents farmers from double-pledging the same cattle across multiple loans. It also includes built-in safeguards that allow the farmer to swap one dead cow for a live one.
Cowmed already tracks about 100,000 dairy cows across more than 1,000 farms. The herd is worth over $395 million. The company expects up to 20% of its network to adopt this tokenized financing model, unlocking $77.6 million in fresh credit for the agricultural sector.
To study whether and how academics respond to political pressure, we exploit a natural experiment: the publication in early 2025 of a “blacklist” of words flagged by the U.S. government. We find that the release of this list led to a sharp reduction in the use of these flagged words among economists at universities that rely heavily on federal funding, relative to scholars from institutions that are less dependent on federal funding or based in the UK. The drop is driven by content related to gender, race, and environment. We show that changes are not simply semantic but reflect actual paper content and that neither the individual funding status nor time-invariant author characteristics are driving the effects. We also document interesting heterogeneous effects by department quality and author gender and ethnicity. Our findings are consistent with the idea that scholars respond strongly to political pressure.
That is from a new paper by Dominic Rohner, Oliver Vanden Eynde, and Philine Widmer. I should note that the authors frame their results in terms of “Science under threat,” which indeed is in the title of their paper. I do see some of that in operation, but I also see a lot of “removing incentives for pandering.” Your own weights here may vary.
More than any of these eval scores, what is most exciting to me is something else: Opus 5 is our least prompt injectable model yet. It is a bit buried in the system card, but across PI evals and red teaming, Opus 5 is very hard to prompt inject successfully.
I've been offline kayaking with sea otters for much of today so I haven't had a chance to put Anthropic's new model Claude Opus 5 through its paces yet. The buzz is positive, and Anthropic's description of it as a "thoughtful and proactive model that comes close to the frontier intelligence of Claude Fable 5 at half the price" sounds promising. It's currently leading the Artificial Analysis leaderboard, in front of even Fable 5.
It's priced the same as Opus 4.8, and continues to offer a "fast mode" at twice the cost of the base model.
Based on this anecdote in the release post it sounds like it might be relentlessly proactive:
On one Frontier-Bench task, Opus 5 was given a drawing of a machine part and asked to write code to rebuild it as a 3D FreeCAD model. However, in this task, the model was intentionally given no way to directly viewthe drawing. Opus 5 responded by writing its own computer vision pipeline to pull the geometry from the raw pixels, then reconstructed the full machine part.
It's better at finding vulnerabilities but has deliberately not been trained on how to exploit them. Hopefully this means the US government won't shut it down!
As with its predecessor, Opus 4.8, we’ve intentionally avoided training Opus 5 on cyber tasks. The model has nevertheless improved substantially on these tasks as a result of becoming more generally capable, and it comes close to Mythos 5 at finding cybersecurity vulnerabilities. However, it remains substantially behind Mythos 5 on the exploitation of those vulnerabilities—that is, in turning vulnerabilities into material cyber threats.
Yesterday Donald Trump declared that the Iran war is going “better than anybody expected could be done.” He’s delusional, of course. And his delusions are the reason the price of oil is back at around $100 a barrel, and — as I’ll explain shortly — the effective price is much higher than that.
Trump’s gratuitous war on Iran has turned into a remarkable quagmire — remarkable because the only thing keeping the war going is Trump’s vanity. He effectively lost the war in the first few days, when it became apparent that Iran’s hardline regime had survived the initial decapitation strike and that the U.S. military couldn’t keep the Strait of Hormuz open. But Trump is psychologically incapable of admitting failure. So the war goes on, weakening America by the day, as he searches for some way to spin his abject defeat as a victory.
And declarations that the economic fallout from the war had been contained now look dangerously premature.
It’s true that during the first closure of the Strait oil prices didn’t rise as high as many analysts — myself to some extent included — expected. Some Persian Gulf oil made its way to markets bypassing the Strait of Hormuz, notably via the Saudi pipeline to the Red Sea. China sharply reduced its oil imports. And the world offset a substantial part of the shortfall in supply by drawing down inventories.
The second Hormuz closure could be worse, for several reasons. Iran’s Houthi allies are now attacking shipping in the Red Sea, threatening that safety valve. Also, inventories are now much lower than they were when the conflict began, and can’t serve as a cushion going forward.
Perhaps the most important thing to realize about the current situation, however, is that oil is more expensive than it looks.
Nobody burns crude oil. Oil must be refined into usable fuels, mainly gasoline and diesel. And there’s a global shortage of refining capacity. This partly reflects the war in Iran, but it also reflects Ukraine’s stunningly effective campaign against Vladimir Putin’s energy infrastructure.
This shortage of refining capacity has helped keep crude prices down — why buy crude when you can’t refine it? More important, however, it means that prices of petroleum products to end users are much higher than one would have expected given the price of crude. The chart at the top of this post shows the wholesale price of diesel, which even during the failed cease-fire was far above its prewar level, and is now close to its previous peak.
The overall “crack spread” — the difference in price between a barrel of crude oil and the price of the products into which that barrel is refined — has exploded, from around $25 a barrel before the war to more than $65 now:
RBN Energy
From the point of view of end users, this is the same as if crude prices had risen an extra $40 per barrel. In effect, the world is coping with the equivalent of $140 oil even though the headline price is “only” around $100.
Does this portend economic catastrophe? No, not yet. But it’s not good.
Last month’s easing in the inflation rate now looks temporary. With energy prices surging again — not to mention the inflationary impact of Trump’s new round of tariffs, imposed using the ludicrous excuse that nations aren’t doing enough to stop forced labor — interest rates will almost surely rise, intensifying the squeeze on families and businesses. Long-term rates, reflecting expected future Fed hikes, are already way up:
10 year bond, from CNBC
Still, things could be worse. And they may be about to get worse. The Wall Street Journal reports that Trump is in “revenge mode,” and the U.S. military is surging forces into the Middle East.
And if Trump doubles down on failure, drastically escalating his disastrous war, the economic, not to mention human, impacts will get much uglier.
SpaceX’s Super Heavy-Starship, the most powerful rocket ever built, blasts off on the program’s 13th test flight in another critical milestone for Elon Musk’s rocket company. Image: Adam Bernstein/Spaceflight Now.
SpaceX launched it’s 13th Super Heavy-Starship rocket Friday, with the giant booster chalking up an on-target but “hard” splashdown off the Texas Gulf Coast while the Starship upper stage carried out a successful sub-orbital hop to the Indian Ocean.
Powered by 33 methane-fueled Raptor engines, the 407-foot-tall two-stage rocket blasted off from SpaceX’s “Starbase,” Texas, launch site at 5:51 p.m. EDT, putting on a spectacular show for area residents and tourists as it climbed away atop some 16 million pounds of thrust.
A launch attempt July 16 ended with a last-second computer-commanded abort when four Raptors failed to start properly. Those issues were addressed and the launch was reset for Thursday, only to have low clouds prompt a 24-hour delay.
Starship flight 13 thunders into a blue sky from SpaceX’s Starbase launch site in south Texas. Image: Adam Bernstein / Spaceflight Now.
But the weather cooperated Friday, all 33 first stage engines fired up normally and the rocket thundered away through a mostly clear sky.
It was the second test flight of a third-generation Super Heavy-Starship, building on lessons learned during the initial flight of a “V3” rocket in May when the booster missed its landing target and the Starship suffered an early engine shutdown on the way to space.
This time around, the booster appeared to perform as planned throughout its ascent and again during its flip around for a tail-first descent to splashdown after separating from the Starship upper stage.
But as the booster dropped toward the Gulf, only 10 of 13 engines restarted and at the moment of splashdown, only five appeared to be running. The stage hit the water at a higher velocity than expected, making a “hard” splashdown.
A spectacular view looking down the side of the Super Heavy-Starship a little more than one minute after liftoff from SpaceX’s Starbase facility on the Texas coast. Image: SpaceX
The booster is designed to fly itself back to its launch gantry where huge mechanical arms, known as “chopsticks,” can pluck it out of mid air. But for the initial flights of the third-generation rocket, But that requires a relatively slow final descent, For good reason, it now appears, SpaceX opted for initial version 3 splashdowns in the Gulf as a safety precaution.
The Starship upper stage, meanwhile, completed its climb to space with all six of its Raptor engines operating smoothly.
Once coasting along its sub-orbital trajectory, the Starship deployed 20 third-generation Starlink internet satellites in a real-world test of the rocket’s Pez-like payload dispenser, releasing the satellites one at a time from a slot near the nose of the vehicle.
Six of those satellites were equipped with cameras to inspect the Starship’s heat shield tiles and general condition to help engineers assess the usefulness of such video for determining the health of future Starships before re-entry.
A camera on the first stage booster captured the moment of a “hard” splashdown in the Gulf after eight of 13 engines failed to restart properly to provide the needed thrust for a more gentle touchdown. Image: SpaceX
In a final test, the Starship briefly re-ignited one of the Raptor engines to demonstrate its ability to re-start in space, which will be required for future flights to the moon.
The spacecraft then faced the blazing heat of a belly-first re-entry before flipping upright and settling to a dramatic rocket-powered splashdown northwest of Australia one hour and five minutes after launch.
The Starlinks, on the same sub-orbital trajectory as the Starship, were expected to fall back into the atmosphere and burn up, or as SpaceX puts it, “demise,” before reaching the sea.
Engine issues aside, the Starship’s performance was welcome news to NASA, which is counting on a variant of the upper stage to serve as a lunar lander for astronauts in the agency’s Artemis program.
A SpaceX drone captured a dramatic view of the Starship settling to a gentle splashdown in the Indian Ocean an hour and five minutes after launch from Texas. Image: SpaceX
For moon missions, SpaceX will need to launch up to 15 or so Super Heavy-Starship tankers to refuel the lander before it can head for the moon to await the arrival of astronauts in a Lockheed Martin-built Orion capsule.
From there, the 165-foot-tall lander will carry two crew members down to the surface, landing vertically near the moon’s south pole. The astronauts then will ride an external elevator down to the surface and back up again when their exploration is complete.
With the first such landing targeted for 2028, SpaceX must ramp up its Super Heavy-Starship test schedule to get the vehicle certified for human spaceflight and to demonstrate the reliability required to safely launch more than a dozen tankers within days of the lander’s launch.
As of now, the moon lander variant has not yet flown and no Starship of any kind has yet been put into Earth orbit. Many NASA veterans have criticized the complexity of the mission architecture and the number of flights that will be needed to get a single lander to the moon.
For the first time, the Starship remained fully intact after splashdown, floating serenely on the surface of the Indian Ocean with many of its on-board cameras still working. This drone view shows residual propellants burning, but the flames quickly died out. Image: SpaceX
The Government Accountability Office said in a report released Thursday that SpaceX faces major challenges getting the Super Heavy-Starship ready for operational use.
“SpaceX’s progress in developing its cryogenic fuel management technologies is a top risk for the program,” the GAO said. “At present, SpaceX’s plan for (moon missions) requires on-orbit propellant transfer between multiple Starship vehicles in low-Earth orbit before the HLS (Human Landing System) Starship can be sent to and dock with the Orion spacecraft in lunar orbit.”
Another concern is completing development of the upgraded Raptor engines, “addressing issues identified through testing.”
“Furthermore, SpaceX is more than a year behind its original schedule for key events,” the GAO said. “These include the critical design review, the long duration and propellant transfer demonstrations and the uncrewed lunar landing flight test.”
The agency said moon lander officials “also stated that they anticipate encountering technical challenges in the development of the third version of the Starship vehicle, which SpaceX plans to use for the first Starship orbital test flights.”
NASA managers are hedging their bets.
Blue Origin, owned by Amazon-founder Jeff Bezos, also is working on moon landers for NASA to carry both cargo and Artemis astronauts to the lunar surface. The spacecraft will be launched by the company’s New Glenn rocket. That lander, too, will require refueling in space.
If the Starship lander isn’t ready in time to meet NASA’s landing target, agency may be able to use Blue Origin’s lander instead.
Artist concept of a SpaceX Starship lunar lander on the surface of the moon. Image: SpaceX.
But during preparations to test fire the main engines of the third New Glenn rocket, the launcher exploded in a spectacular fireball, virtually vaporizing the rocket and seriously damaging its pad. The company hopes to resume test flights by year’s end, but few details have been provided.
In the meantime, NASA is pressing ahead with preparations for an interim Artemis flight next year in low-Earth orbit. Four astronauts will be launched in an Orion capsule to test the rendezvous and docking procedures that will be needed in lunar orbit for a moon landing mission.
The crew will dock with a Blue Origin lander, open hatches and float inside as part of the test flight. They then will rendezvous and dock with a Starship. But in that case, a modified production model will be used instead of an actual moon lander. It will have no crew accommodations and the astronauts will not go inside.
An artist’s concept of NASA’s Orion spacecraft docking in low Earth orbit with SpaceX’s Starship Version 3 rocket with a docking adaptor during the Artemis 3 mission. Rendering: SpaceX
I'm still writing my journals on paper. Makes it about a decade of notebooks, most of them Field Notes, kept in the archival wooden boxes[1]. Writing by hand helps me remember and process life. Finishing notebooks give me the feeling of closing a chapter on those problems, at least for a moment.
Paper notebooks are also a very flexible medium, as most teenagers can confirm: you can take train tickets and elmers-glue them in there to commemorate a trip, or make sketches to illustrate some anecdote. Want five different kinds of checkboxes, and lines connecting paragraphs? You can do it in a paper notebook, and probably not in Notion or other digital tools.
Collecting ephemera in notebooks has been really satisfying, so I've been experimenting with something new: a tiny printer. In particular, the Liene Pearl N200 Pro. Here's some of that experience.
It's a little printer. Exactly the height of an iPhone 16, and about twice its thickness. The photos it prints are small - 2x3 inches. My beloved Field Notes notebooks are 3.5x5.5", so the photos fit really well: any wider and they'd creep too close to the binding of the notebook.
The Liene charges by USB-C and connects via Bluetooth. Virtually nothing that uses Bluetooth is reliable, and this is no exception.[2] Half the time it pairs immediately, the other half I have to restart it to connect both to my iPhone and to the app. And it's kind of odd that it doesn't support AirPrint, Apple's standardized way for printers to connect to iPhones. You have to use the app, which is fine to use if you ignore the goofy AI features. For Chinese-based consumer electronics like this, the hardware is usually a lot better than the software.
And the photos are stickers! So far I haven't printed something from this and not wanted the photo to be a sticker: wouldn't I want to stick it onto a postcard or in a notebook? The quality's totally fine but won't replace a desktop printer or a film photo.
It prints with dye-sublimation in a three-color CMY process: yellow, magenta, and cyan. You can see the photo at each stage, and then it does a final lamination pass. The image rendering is punchy - high-contrast and high-saturation.
Then the photos work nicely as full-width, or can be aligned vertically with a squeezed left column, like here:
The printer supposedly goes through a cartridge every 10 prints and eats special paper that costs 50 cents a print.
It's hard to say that this kind of thing is practical: is putting little photos in paper notebooks 'practical'? The whole exercise is sentimental and psychological. The prints are tiny and the printer takes a good 45 seconds to make them. They're kind of expensive. It usually requires a minute of futzing to connect.
But I've used snailjet for years - a service that lets you send a postcard with any photo, straight from your phone, for $1.41. Is surprising a friend with an old photo worth $1.41? Obviously.
Have I gotten some joy from buying a ~$110 printer to decorate my little paper notebooks? Definitely. Someday I'll read the notes that I'm writing this year and be nostalgic for this phase of life, and it'll be fun to see how I was doing.
I bought my wooden boxes back when they were $35. I don't know if I'll keep buying them now that they're $50: ↩︎
AirPods work flawlessly because they're cheating with a proprietary Apple-only protocol. ↩︎
Up pretty early (though of late I have been faulty by an hour or two every morning of what I should do) and by water to the Temple, and there took leave of my cozen Roger Pepys, who goes out of town to-day. So to Westminster Hall, and there at Mrs. Michell’s shop sent for beer and sugar and drink, and made great cheer with it among her and Mrs. Howlett, her neighbour, and their daughters, especially Mrs. Howlett’s daughter, Betty, which is a pretty girl, and one I have long called wife, being, I formerly thought, like my own wife. After this good neighbourhood, which I do to give them occasion of speaking well and commending me in some company that now and then I know comes to their shop, I went to the Six clerks’ office, and there had a writ for Tom Trice, and paid 20s. for it to Wilkinson, and so up and down to many places, among others to the viallmaker’s, and there saw the head, which now pleases me mightily, and so home, and being sent for presently to Mr. Bland’s, where Mr. Povy and Gauden and I were invited to dinner, which we had very finely and great plenty, but for drink, though many and good, I drank nothing but small beer and water, which I drank so much that I wish it may not do me hurt.
They had a kinswoman, they call daughter, in the house, a short, ugly, red-haired slut, that plays upon the virginalls, and sings, but after such a country manner I was weary of it, but yet could not but commend it. So by and by after dinner comes Monsr. Gotier, who is beginning to teach her, but, Lord! what a droll fellow it is to make her hold open her mouth, and telling this and that so drolly would make a man burst, but himself I perceive sings very well.
Anon we sat down again to a collacon of cheesecakes, tarts, custards, and such like, very handsome, and so up and away home, where I at the office a while, till disturbed by, Mr. Hill, of Cambridge, with whom I walked in the garden a while, and thence home and then in my dining room walked, talking of several matters of state till 11 at night, giving him a glass of wine.
I was not unwilling to hear him talk, though he is full of words, yet a man of large conversation, especially among the Presbyters and Independents; he tells me that certainly, let the Bishops alone, and they will ruin themselves, and he is confident that the King’s declaration about two years since will be the foundation of the settlement of the Church some time or other, for the King will find it hard to banish all those that will appear Nonconformists upon this Act that is coming out against them.
This is the second half of the fourth part (I, IIa, IIb, III, IVa) of our honestly-who-knows-how-many part series laying out some general guidelines for how pre-modern armies are organized. Last week we took a look at how the leadership of these armies was drawn from the existing elite classes of their societies. As with other factors, armies cannot help but recreate their civilian social structureson the battlefield, with the result that the military leadership class often mirrored, where it wasn’t simply identical with, the civilian ruling class.
Now it will not surprise you that pre-modern military literature, written almost entirely by elites in the leadership class for other elites in the leadership class, is very interested in leadership and tends to attribute all military success or failure to the fellows in charge. Modern military science, however, has increasingly recognized the importance of combat motivation in individual rank-and-file soldiers.1
As noted last time, we can break this broader concept of combat motivation into two parts: morale and cohesion, where morale is the soldier’s attachment to the cause and belief in its possible achievement and cohesion is a soldier’s attachment to his fellows. What combat motivation studies have tended to show is that morale will get a soldier to the battlefield, but will not keep him on it; it is cohesion instead, the ability of a unit to stay together under pressure that is essential for enduring the terror of combat.
That said, different armies and different societies generate cohesion in different ways, as we’ll see. Now this is one component of this series that is a bit of an exception: so far I’ve been hammering that most societies are going to be pushed very strongly towards specific kinds of recruitment, organization and leadership as a result of their civilian organization. Meanwhile, as we’ll see, certain kinds of cohesion depend to a greater or lesser degree on specific social structures or assumptions, but it is a lot more common to see armies that simply have poor cohesion due to a mismatch of cohesive principles and military or social structures.
But first, as always, recruiting and maintaining large pre-modern armies is expensive! Much like many of those pre-modern armies, this project is supported by devolving the costs of my ruinous book-buying habit on to recruits readers. You can help by spreading the word to new readers and by supporting this project over at Patreon. If you want updates whenever a new post appears or want to hear my more bite-sized musings on history, security affairs and current events, you can follow me on Bluesky (@bretdevereaux.bsky.social). I am also active on Threads (bretdevereaux) and maintain a de minimis presence on Twitter (@bretdevereaux).
The Nature of Cohesion
We’ve talked about cohesion a few times here, so before getting into the systems that can generate cohesion, we can be fairly brief on what cohesion is. As the word implies, cohesion is the ability of a unit to hold together under pressure, but this is one of those concepts with a simple definition but more complex application.
On the one hand, cohesion is one of those ‘harder said than done’ principles. When a group of any animals is subjected to peril, the instinct is to flee and for humans, who are not herd animals, the instinct is to scatter, to flee individually. The goal of cohesion is instead to have individuals respond to peril by seeking the safety of the group rather than the safety of individual flight, so that intense fear causes the group to solidify rather than disperse. That is, as we’ll see, a challenging-but-not-impossible emotional response to stimulate, as humans are pack animals. That said, it has to be an emotional response, something sub-rational, because it needs to functions in conditions of intense fear where rational thinking is ineffective.
So you can think of cohesion, in its simplest form, as a kind of social or mental conditioning that aims to alter individual’s fear response from, ‘split up for safety’ to ‘group up for safety’ at an emotional, sub-rational level.
And then on the other hand, that has complex implications for battlefields in which humans generally fight as units rather than as individuals. Not all humans fight this way, I should note: warriors in a ‘first system’ ‘pounce-and-flee’ warfare system – as we see with hunter gatherers or early agriculturalists – often fight in a far more atomized, skirmishing style with no regular formations and no expectation that individuals will hold their ground or cohere under pressure. Instead, warriors in such a system can fall back and regroup relatively freely. But complex agrarian (even non-state complex agrarian) and pastoralist armies do fight as larger units, with complex group tactics that demand individual combatants work together and stick together much more tightly.
Meanwhile, if the unit remains cohesive – if the individual men remain bound together as a single mass – then their leaders can more readily maneuver the whole mass into action. After all, if the presence of peril is psychologically pushing each soldier to cohere for collective safety, then an officer can use that cohesion to push the group as a unit, potentially even moving it into greater peril. This is the foundation of all shock tactics – at least, all of them involving an even remotely equal opponent – because delivering ‘shock’ (that is, advancing into close combat in a group) demands advancing into peril, which is a think that is normally psychologically quite hard to do. As Ardant du Picq notes, “Man does not enter battle to fight, but for victory. He does everything that he can to avoid the first and obtain the second.” Getting fellows to walk into range of an opponent’s spears (potentially walking through the range of their slings, bows or muskets to do so) requires moving directly against human psychology. But if the whole mass of men coheres, the mass may be moved forward: group psychology succeeds where individual psychology would fail.
Of course cohesion is equally relevant – for the same reasons – for getting men to stand under fire, either in a tight formation or in more dispersed skirmish order (as in modern warfare). Once again, the normal human response to being in a ‘zone of danger’ is to leave it, so the cohesion response is required to get men to stay in the ‘zone of danger’ – the unit is in the zone, the men have cohered to the unit, so the men remain because the unit remains. Now, absolutely there are a small number of people in any group who have the right psychological mixture to do these things alone, but armies cannot be built out of a tiny handful of saints and sociopaths, they must be built out of regular men.
So cohesion serves several related purposes on the battlefield: it keeps men from fleeing but it is also the means by which they can be pushed into offensive action, which is to say it both keeps them from fleeing but also gets them fighting, all as a product of the action of holding them together. Even if a unit must withdraw, cohesion allows them to withdraw as a unit – maintaining fighting capability (which in turn deters an enemy from pursuing too aggressively) – rather than breaking apart into a rout (which an enemy can pursue aggressively).
So how does cohesion happen? High cohesion is most often the product of tight bonds between the men in a unit, though those bonds can be achieved in a range of ways. That said, as we’ll see, certain ways of producing cohesion can also produce a kind of corporate cohesion, where the cohesive principle ends up rooted in the identity of the soldiers (which they share), which creates a kind of cohesion that persists even when soldiers do not necessarily know their fellows (but when they do know that their fellows are also soldiers and so share the same formative, cohesion-building experiences; that’s why I frame it as ‘corporate’ in nature).
So cohesion is important, but it is also fundamentally unnatural: it must be produced somehow, it cannot be assumed. In practice there are two ways (or categories of ways) to produce cohesion: it can be produced in civilian society or it can be produced ‘in the ranks.’
Cohesion Dyed in the Wool
One way to achieve strong cohesion is to have it produced organically through civilian society. Tight social bonds are the key mechanism for producing this kind of cohesion: it derives from serving and fighting next to men who you know well in peace time and who will continue to know you well after the fight is done. The social pressure not to let these fellows down – fellows you will have to live with long after the smoke clears – is one of the few pressures that is as strong as the fear of dying. Or, as I sometimes put it, the one thing humans might fear more than death is shame. But for shame to work, it has to come from a peer group the individual will have to live with.
In practice, this kind of cohesion then tends to come one of two ways: from a small community or from a small, tight-knit group within the community.
The classic example of the former case are various forms of civic militias, like Greek hoplites. The hoplite phalanx of a Greek polis was generally divided into sub-units by tribe or neighborhood (phyle or deme) so that men served next to their friends, neighbors and family members in the phalanx as it was drawn up. Of course in some poleis, this sort of closeness was institutionalized through common dining groups, but it seems pretty likely that regular social dining with the men who would be near you in the phalanx was common across Greek poleis. Consequently, when a hoplite stood in the ranks, he knew he was being observed by everyone (at least, all of the men – though he could be assured his mother, wife and children would hear of his conduct too!) in the world who mattered to him.
Via Wikimedia Commons, an arming scene showing hoplites and a young man being armed as a hoplite (c. 530-510 BC). Although hoplites had little formal training or drill, they were often highly cohesive on the field because their cohesion was organic rather than trained.
The result is the creation of what is termed in military science a primary group: the small, face-to-face social group around a soldier, which tends to be tight and cohesive. Instead of generating that group ‘in the ranks’ – impossible for an amateur militia force that only formed up when there was a battle to fight – these sort of citizen militias instead form the primary group during peacetime and then use it as a basic building block in wartime to achieve that cohesion. As a result, a civic militia like a hoplite militia could be cohesive despite having little or no formal training or collective drill and relatively little in the way of command and control structures. Their cohesion was organic, derived from underlying social structures, rather than needing to be intentionally manufactured through training.
The alternative for ‘organic’ cohesion is the creation of a smaller unit of society that is similarly cohesive, often in the context of a much larger (and thus more anonymous) society which struggles with the kind of ‘small town’ cohesion of the Greek polis. This, of course, is generally the cohesion of a warrior class: the overall society may not be small enough for everyone to know everyone else well enough to generate that kind of tight social pressure, but the warrior-class might be. The aristocracy broadly construed, for instance, of even a large medieval European kingdom might consist of only around a thousand or so households. The kingdom might not be organically cohesive, but the aristocracy could be – though as we will see, knightly cohesion was reinforced through training and drill as well. One of the most striking examples of this kind of cohesion in action to me has always been the performance of the English huscarls at Hastings (1066): even as the rest of the army crumbled and the battle was lost, this group of vocational warriors held together, despite tremendous pressure.
Via Wikipedia, part of the Bayeux Tapestry (c. 1070s), showing the last stand of the English huscarls at the Battle of Hastings (1066). Even under tremendous pressure (arrow volleys and cavalry charges), the huscarls remained cohesive. The main levy force (the fyrd), however, did not, breaking up and leaving the huscarls isolated on the field where they were defeated.
Hopefully, you are also beginning to see how some of what we’ve laid out in these posts can ‘line up’ to produce strong military systems. A civic militia, for instance, can benefit greatly from these tight social bonds: in the ranks, everyone knows everyone else and the unit is led by elected officers who are the same men who lead the community in peacetime. The whole unit can thus draw on the familiarity and strong social bonds without needing to build those connections from scratch. The same, of course, goes for the retinue of a Big Man, whose retainers all know each other (and will continue to do so after the battle because they’re all socially connected) and who all know and already respect the authority of the Big Man who commands.
Now there is another requirement to making this kind of cohesion work effectively: there has to be important social value placed on military performance. For the militia man in a medieval schuttersgilde or a Greek hoplite or a knight riding out in his lord’s retinue, fighting bravely was crucial to their status in the community. Their friends, family and neighbors would be watching and would care if they stuck together or ran away. But for that to work there has to be some status for these men attached to the performance of military service.
A village can have very close-knit social bonds, but if the peasants living there don’t consider the role in society to be fighting, don’t value military service and don’t see themselves as personally involved in the wars of their rentier-elites, it will be almost impossible for those rentier-elites to actually harness the latent social cohesion of the village. Those villagers won’t hold together under pressure on the battlefield, they’ll scatter together – their social cohesion will be addressed in the ways they find socially valuable (like supporting each other with food during hard times). One common ‘failure state’ of military organization is where a political leader raises a population expecting to get organic cohesion but – because the society has not invested the social value in the fighting that those kinds of people do (so they may value knightly warfare but not the infantry warfare of peasants, for instance) – the cohesion simply isn’t there and the force collapses when put under strain. Levies drawn from the underclasses of a society – peasants swiftly formed into a local militia, for instance – often collapse this way under pressure, because the society at large (which is to say, the elites) have not dedicated the social value or resources to build organic cohesion.
In addition, this kind of organic social cohesion is very hard to ‘scale up’ since it relies on relatively small-scale tight-knit social bonds. One solution, however, was for large states to recruit the component units of their armies from specific locales, in what is termed a regimental system. Under a regimental system, soldiers do not move from one unit to another but rather are raised and spend their entire service career (which may be short-service or long-service) in a single unit. Often, when this system was combined with volunteer or short-service units (as in, for instance, the American Civil War or the British army in the First World War), units were intentionally raised from small geographic areas: individual states, cities, towns and so on. While the regimental system itself is a modern phenomenon, similar sorts of systems existed earlier; the Roman socii system seems to have functioned on a similar basis: Roman socii stayed in units recruited from their relatively small communities.
A larger army can thus try to harness some of that local cohesion by recruiting local units and keeping them intact, although this also creates organizational challenges: those soldiers are unlikely to be as willing to serve outside of their original units. They was told, you see, that they’d be serving with the 2nd Maine and the 2nd Maine only. That can cause issues if the original unit has been attritted so badly that it is no longer really functional (but cannot readily be merged with others), but also could create problems because it could also geographically concentrate casualties. The famous, grim example of this was the 11th Battalion of the East Lancashire Regiment (the Accrington Pals), all recruited from the same town, suffered 80% casualties in just thirty minutes of fighting at the Somme (Jul. 1, 1916), practically obliterating an entire generation of the town.2 The British abandoned the recruiting of ‘Pals Battalions’ after the Somme in no small part for that reason.
However, as you might be picking up, not every society generates strong enough organic cohesion to orient a military around (or the ‘warrior class’ that does is too small to form the bulk of the military). So then the question becomes how to create cohesion.
Training Cohesion
And it turns out it is possible to train cohesion into a unit.
What is striking is that historical societies talk about this process using a range of different understandings for how exactly they are inculcating bravery through training. Greek and Roman writers often frame this sort of training not necessarily as teaching soldiers specific, coordinated motions (the swinging of a sword, the throwing of a javelin), even though we know that the Romans, at least, did drill that sort of thing, but rather as rendering soldiers inured from hardship by putting them through hardship, by having them undertake strenuous duty regularly with strict punishments for failure (and generous rewards for success).
By contrast, early modern European drill, though it equally involved a lot of strenuous work, was understood more at the time as trying to produce a ‘mechanical soldier.’ The early modern European aristocrat largely concluded that the raw lower-class clay he was forced to work with for his army was temperamentally unsuited for bravery (these fellows were mostly stuck up snobs) and so could only be made brave by having the necessary actions for battle – loading and firing firearms, moving in formation – drilled over and over again until they became mechanical muscle memory. They use the same methods as the Romans – repetitive training of basic actions under as realistic circumstances as possible, with strict punishments for failures of discipline – but understood the mechanism of action to be quite different.
Now I want to be clear here what I mean when I say drill. Drill in this context is training through strenuous repetition of simple actions in preparation for combat. Not every army drills! There’s also a distinction here between armies that value “collective synchronized discipline” – the ability of soldiers to move together in synchronized action – and armies that don’t (you can use muskets in collective, synchronized musket lines but you can also use them independently in skirmish formations, for instance).3 It is likewise possible for armies to insist on physical conditioning in ways that do not produce group psychology effects, the difference between training together and training individually.
Via Wikipedia, an illustration of a Ming musket drill using volley fire techniques (1639). Drill and synchronized discipline were not purely a phenomenon of Europe or the broader Mediterranean, they appear in East Asia quite frequently as well. While it is easiest to illustrate this with musket drill (as above), synchronized discipline in China goes back at least to the Warring States period (475-221 BC) and in Japan we have plenty of evidence for synchronized discipline for infantry using various polearms before the large-scale introduction of firearms during the Sengoku period (1467-1600).
Now to be clear, obviously rigorously drilling soldiers in the manipulation of weapons is going to train what we can call ‘skill at arms’ – the effective manipulation of complex weapon systems – but we’re interested in the psychological conditioning here.
And the way modern thinking understands this same process is rather different than either ancient or early modern aristocratic thinking, in part because modern researchers and scholars have bothered to occasionally ask common soldiers about it. While the concept of unit cohesion has existed, I’d argue, since antiquity, it really only started receiving really systematic study after the Second World War, with the main conclusion of those early studies being that unit performance was heavily sensitive to the tight bonds in small units. Which is to say soldiers were motivated to fight for their ‘buddies’ at the squad, platoon and company level and that this social bonding effect was the primary source of unit cohesion.
The question, of course, was how a military went about manufacturing those sorts of social bonds so that units cohered under pressure. But of course the military wasn’t approaching these questions with a blank slate: in many ways the western military tradition (hardly alone) had already developed a system for manufacturing unit cohesion – what it was coming to realize is that it had never really understood how that mechanism functioned.
Instead, while the role of drill in training specific skills is important, it also imposed upon each batch of recruits intense conditions of shared suffering combined with creating a barrier to entry into the final group (one had to endure the initial drill and training process). In short, taking a bunch of potential soldiers and throwing them all into the same difficult situation (physical training and drill) they have to overcome together over an extended period of time (weeks or months, if not longer) is a way to ‘speed run’ the social bonding process to create a strong, cohesive unit.
One of the risks this approach is thought to run, however, is that rotation or high casualties can erode that cohesion. The soldier has cohered, after all, to a ‘primary group’ consisting of the men he trained or served long-term with, but that bond can be ruptured as members of the primary group are killed or rotated out of the unit (or the soldier is rotated into a new unit where they don’t know anyone). One of the responses to this reality is to avoid individual soldier rotations (such rotations in the Vietnam War were famously blamed for low morale and discipline) and instead rotate whole units, but of course there is nothing doing about casualties and the need to replace them.
The thing is, while certainly some units suffer reduced cohesion and thus reduced combat performance as they take casualties and – in the accidentally memorably bloodless phrasing of one of my graduate school colleagues – “the primary group disintegrates” that also quite clearly doesn’t happen to every military. It is a bit of a grim (and very modern) example, but the Wehrmacht (the German army) in WWII remained cohesive despite absolutely staggering casualties by war’s end and the effective impossibility of maintaining meaningful primary groups in the face of those losses.
In a similar vein, new assignments or distant posts often seem substantially less disruptive on cohesion in long-service professional forces. And indeed, mercenaries and long-service professionals can show remarkable cohesion under strain, even when the men in question haven’t necessarily been trained together. Xenophon’s Ten Thousand, recorded in his Anabasis, provide a notable example of this: the Ten Thousand were a mercenary force composed of men from ten different smaller mercenary units, scattered over the Greek world and in most cases we have to imagine the men who joined were not joining together with their neighbors but as deracinated mercenaries. Yet the Ten Thousand show extraordinarily strong cohesion under pressure, both in pitched battles but also under the steady strain of a long march under continuous threat.
That sort of ‘plug and play’ cohesion also shows up in long-service professional armies as well as other kinds of armies that have simply been in the field for a long time as the result of very long wars. One can see it, to take a modern example, quite clearly in the performance – in both victory and defeat – of British regular troops during the Anglo-Zulu War. At Isandhlwana (Jan 22, 1879), a British column was effectively annihilated when enveloped by a larger Zulu force but most of the regulars remained cohesive to the end even when it was clear the battle was lost. The same intense cohesion among the regulars enabled an unlikely victory the next day at Rorke’s Drift (Jan 22-3, 1879) when a badly outnumbered garrison post held off a large Zulu attack. Of course we should also note the remarkable cohesion of the Zulu (in their case a mix of organic cohesion and cohesion developed over frequent campaigns) in these battles, where they were able to persist on the attack despite shockingly heavy casualties.
Via Wikipedia, a 1880 painting by Alphonse de Neuville depicting (in somewhat dramatic form) the Battle of Rorke’s Drift (1879). As an aside, I generally stick to period artwork on the blog whenever I can, but of course we should be alert to the fact that even period artworks are not photographs. So the composition of this painting, clearly intended to highlight the heroism of the British is deliberate and in a way we might say ‘more real than real’ (for instance, it appears to combine several actions at once that happened over extended periods and exploits the use of foreground and background to render the Zulus into a dehumanized horde rather than more human opponents). We must read our images – even period images – just as critically as we read our texts and this is no less true for the above pictures of hoplites or the Bayeux Tapestry as it is for this modern painting.
What I think is going on here is that over time these fellows are being bound not only to a primary group, but also a corporate, professional identity, effectively a professional version of the corporate identity of the ‘warrior class’ discussed above. On the one hand, all of these ‘old soldiers’ still have shared experiences, because even though they were not drilled together, they were all drilled the same way and they all ‘know their business’ the same way. At the same time, they have a common connection to a shared identity nonetheless: they’re all soldiers, not merely by circumstance but by identity. Being a soldier is not something that has happened to them, but something that is part of them. And of course those elements – difficulty of entry, commonalities in experience in the group, a shared sense of belonging – are all elements that encourage group cohesion.
Now again, in both cases, the actual mechanism for producing this psychological and social bonding effect is drill and discipline, which of course also have other effects (as all training does) and are often done for those other purposes, developing cohesion as an accidental (but very necessary) byproduct.
Now as you might gather, the research on training and unit cohesion is, at this point, voluminous; this can only have been an overview. But I think, at least for thinking about worldbuilding (and also a general framework in which to understand the specialist literature), thinking in terms of two categories of ‘manufactured’ cohesion – primary groupcohesion (a result of a group being trained together and undergoing hardship together) and corporate identity cohesion (a result of a shared sense of professional or corporate identity among a much larger group, married to shared experiences of training and hardship). However remember: your subjects may view the matter quite differently, understanding the improved combat performance to be the product of something other than the psychological conditioning that modern research focuses on.
Cohesion and its Absence
That leaves us with really four major avenues for a unit to develop effective cohesion: the organic cohesion of small communities (up to the level of a small town, but not generally much larger), the organic cohesion of selective groups (like warrior classes or specific guilds or social clubs) with in a larger, less cohesive society, manufactured primary group cohesion that results from training a unit together and having them endure hardship together and manufactured corporate identity cohesion, a result of a shared sense of identity among a larger group that have a sense of shared hardship.
Note, of course, that there’s no reason these mechanisms for cohesion cannot be combined. Roman citizen armies in the Middle Republic were recruiting with at least some substantial organic cohesion from the citizen class (and the socii in their even smaller, potentially more cohesive groups), which was then reinforced by strict discipline and the experience of military life, effectively layering small community cohesion on top of manufactured primary group cohesion. Likewise, a medieval knight (or other mounted warrior) is already coming with the organic cohesion of a warrior class, but knights also trained together extensively in small units (the conroi, a unit anywhere from a half-dozen to couple-dozen mounted warriors) and would fight in battle as part of that unit, layering that primary group cohesion over top that warrior class cohesion.
The thing is military units are not always cohesive. This is a bit of a difference from the previous sections in this series. Military units, after all, that exist presumably have to have been recruited, supplied, organized and paid somehow or they wouldn’t exist. But there’s no reason a society cannot form up an army or a unit within an army that simply never coheres will and as a result performs poorly on the battlefield. Indeed, we’ve discussed both fictional and real world examples of poor cohesion or combat motivation in the past.
Now the student of history generally comes at these questions knowing the outcome: we typically know that an army either collapsed under pressure or cohered under pressure and are seeking to understand why. Each society’s system for generating cohesion is going to differ in some respects – indeed, you may well have different cohesive principles at play in different parts of the same army for different units raised differently – but a basic sense of the varying ways that cohesion can be generated can help one begin to work out why this or that historical army might have performed well or poorly, creating hypotheses to then investigate more closely in the evidence.
For worldbuilders, the task is somewhat inverted because the armies are fictional and their battles yet unfought. So instead the task is thinking about the societies one has built and the ways they could generate cohesion and giving them military and social institutions to match. Or, crucially, giving them military performance to match. So if your story is meant to feature a fearless, highly effective sort of army, you probably ought to think at least a bit about why they cohere together so well (and not think about it in just bland ‘because they’re super soldiers’ way). Alternately, if you want a military to surprisingly underform, think about flaws or weaknesses in its cohesion – perhaps the a Big Evil who hasn’t given enough thought to psychology or has simply assumed that he can rely on organic cohesion without having done anything to foster it, for instance.
Perhaps most importantly, remember that cohesion is a properly of the regular soldiers of an army, rather than their leaders. Highly effective armies combine capable leadership with high degrees of cohesion. But cohesion can turn against leaders as well – after all, a unit that has strong group bonds will be cohesive when facing the enemy, but it will also be cohesive in a mutiny.4
We’ll probably have a shorter fireside post next week, but after that, we’ll close out this series looking at some real and fictional examples of how all of these elements – civilian society, recruitment, supplies, organization, cohesion, leader and so on – fit together to make coherent military systems (or in some fictional examples, incoherent ones).
In a historic first, all six radio frequency antennas at the Madrid Deep Space Communication Complex – part of NASA’s Deep Space Network (DSN) – carried out a test to receive data from the agency’s Voyager 1 spacecraft at the same time on April 20, 2024. Known as “arraying,” combining the receiving power of several antennas allows the DSN to collect the very faint signals from faraway spacecraft. A five-antenna array is currently needed to downlink science data from the spacecraft’s Plasma Wave System (PWS) instrument. As Voyager gets further way, six antennas will be needed. Image: MDSCC/INTA, Francisco “Paco” Moreno
NASA evacuated its Madrid Deep Space Network ground station on Friday as wild fires near the Spanish capital city swept through the area. The agency said all personnel were safe and it had yet to survey the site for damage.
Photos posted on social media appeared to show smoke and flames in the vicinity of the tracking dishes. The Madrid location works in concert with ground stations at Goldstone, California, and Canberra, Australia to send commands to and receive data from spacecraft in deep space.
According to Canal 24 Horas, 90 workers at the site were evacuated and that government officials assured local media that “the impossible is being done” to safeguard the complex. The BBC reports that as of Friday, more than 25,000 people have been evacuated in Spain and more than 20,000 are in lockdown.
Emergency crews in Spain are fighting to shield NASA’s Madrid Deep Space Communications Complex from wildfires.
Robledo de Chavela is 1 of only 3 sites in NASA’s Deep Space Network that provides 24/7 global coverage for spacecraft & tool the 1st image of earth from the apollo 11 pic.twitter.com/VCT43y6bxp
A NASA spokesperson confirmed the evacuation and noted that any damage that may have occurred as a result of the fire “will be assessed when it is safe to do so.”
“NASA’s SCaN (Space Communications and Navigation) Program is closely monitoring the wildfire activity in coordination with local authorities and the U.S. Embassy,” the spokesperson said. “In the meantime, the agency has seamlessly transitioned support for mission operations to the Goldstone Deep Space Communications Complex in California, ensuring continuity of service and uninterrupted support for spacecraft communications.”
Each of the three sites contains 230-foot (70 m) diameter antennas that weigh about 2,970 U.S. tons (2.7 million kg). The complex in Spain is about 37 miles (60 km) west of Madrid at Robledo de Chavela.
The Deep Space Network has faced challenges with capacity and recently one of its California-based antenna, Deep Space Station 14 (DSS-14) sustained damage due to over-rotation while tracking the Juno spacecraft on Sept. 16, 2025. That was classified as a Type A mishap “based on the total cost of damages,” which are estimated to be between $4.1 and $4.6 million the agency said in early June.
Imagine your career were coming to a close, how would you want it to end? By earning an especially large sum of money? Publishing a final wonderful paper? With an amazing act of mentorship?
When it comes to Lebron, had he wanted to finish his career with the guys he enjoys playing with the most, he would have opted for Golden State (Steph and Draymond).
If he had wanted to choose his favored organization, he would have signed with Miami.
Furthering the NBA prospects of his son Bronny, and having more time with family, would have meant staying in Los Angeles.
The hometown, sentimental choice would have been Cleveland.
As it turned out, he wanted to maximize his chances of winning a title, and with a team that did not just win a title (which rules out OKC, NYC). That meant signing with Philadelphia. Which is what he did.
What will your final major career decision look like?
A few days ago, I offered some comments on the crisis at Netflix—describing it as a sign that manipulative “audience capture” strategies in business have stopped working.
I typically put these in-depth culture briefings behind a paywall, but at the last moment I decided to make this one available to all. And I was rewarded when it went viral with a vengeance. Within a few hours, my “audience capture” critique racked up more than five thousand likes and five hundred shares on Substack alone—but it also took off on other platforms.
This brought in many new subscribers. I welcome them aboard, and hope they become active members of our community.
I was especially gratified to get a nod from legendary investor Michael Burry (celebrated in the hit film The Big Short for his anticipation of the 2008 mortgage meltdown). His boost brought an entirely different audience to my Substack, which is mostly about culture and not finance. Even so, there’s a heavy dose of futurism and prognostication at The Honest Broker, so I hope these newcomers will find something of value here.
Even so, this viral article left many important questions unresolved. The biggest one is: What happens next?
Please support The Honest Broker by taking out a premium subscription (just $6 per month).
Today I’m offering more detailed predictions. These will tell you not only what these companies should do, but what they must do if they want to avoid stagnation and declining market share.
Of course, that’s no guarantee they will take my advice. In fact, if you just judged matters by the current media and web landscape, you might think there’s no chance they will abandon audience capture. But if they continue down the current path, they will pay a high price for their stubbornness.
First, let me define what I mean by “audience capture”—because some people use this same term differently, and I want to make sure we’re on the same page.
Audience capture strategies focus on locking up a customer base—via network effects, free stuff, loyalty programs, cheap entry-level subscriptions, etc.—and making it hard (or at least a hassle in some way) to leave. After reaching critical mass, the company shifts to squeezing the customer with frequent price increases and other exploitative measures, counting on the fact that many will just put up with it.
As you can see, there’s a world of difference between these two strategies, and the key parameters of the audience capture approach are disturbing.
Even more disturbing, however, is my next claim: Audience capture is rapidly becoming the default approach of the largest companies in the world. So its potential for abuse must be a matter of concern for all of us.
This would have been unthinkable a half century ago. In those days, audience capture was a rarity—found mostly in airline frequent flyer programs and forgotten gimmicks like Blue Chip Stamps. But the Internet really gave impetus to this approach as a default for digital platforms. At it’s most extreme, platforms grew enormously by offering consumers something for free (search results, videos, social media content, etc.), and then gradually shifted focus to milking this user base.
At this final stage, the audience capture company prods users into watching crappy advertising, or making in-app purchases, or agreeing to have their personal information sold to data brokers, or participate in some other exploitative scheme. That’s because, in the world of audience capture, companies always squeeze more and more—trying to gauge how much they can impose on users before they revolt.
Once you see this strategy spelled out, you recognize it everywhere. Many of the worst perpetrators have almost forgotten what it’s like to serve a customer.
I’m giving notice that this approach has gone too far, and that a genuine backlash is underway. That’s why Netflix has seen it’s share price collapse. And the same is true at many other platforms.
This essay was written with Barath Raghavan, and originally appeared in IEEE Spectrum.
Major benchmarks measure what AI can do. None measure whether it does what you mean: the distance between what you ask an AI to do and the unspoken assumptions about how you want the AI to do it. We propose a new metric: the Genie coefficient.
There’s often a gap between one person’s request and another’s understanding. Most of the time, we bridge it using general knowledge. For example, if you ask a friend to get you coffee, they’ll pour a cup from the pot or buy one from a coffee shop. They won’t bring you a bag of raw beans or snatch a cup from a stranger and hand it to you. You never specified any of this. You never had to.
One might think the fix is just to specify tasks, questions, and intent better. But in 1987, in their seminal book on AI, Terry Winograd and Fernando Flores succinctly captured why that won’t work: “Q: Is there any water in the refrigerator? A: Yes. Q: Where? I don’t see it. A: In the cells of the eggplant.” In human language, wants and desires are always underspecified. It is impossible to list all the caveats, all the limitations, all the exceptions.
So how does anyone communicate, if intent can’t be pinned down? Because a reasonable person can make a reasonable guess. Even though wants and desires are always underspecified, a competent person generally knows enough context to get it right or else knows to ask for clarification. Linguists call this pragmatics: Meaning lies in the words and the situation and also in all prior communication, shared culture, and innate human behavior.
It doesn’t always work out, of course. Your friend might bring you a hot coffee when you wanted an iced coffee, or an Italian coffee when you wanted a Turkish coffee. The more dissimilar the two people are in age, culture, and background, the more likely the request will be misunderstood in some way.
This situation has major implications for AI agents that are increasingly being given requests by humans and expected to fulfill them. They have enormous latitude to get it wrong. An AI agent asked for coffee might buy a coffee plantation or order a cup of coffee for delivery in three weeks. Its actions may be recognizable as “getting coffee,” but not remotely what you intended. They’ll think outside the box because they won’t have our conception of the box.
When AI Gets Proactive
For most of the last decade, when systems like Alexa or Siri misinterpreted a request, it was annoying, not dangerous. Beyond the AI model itself, what has changed is the harness: the ordinary code that wraps around an AI model, decides when and how to use the model, and controls access to tools like a browser, a low-level command line, or a financial API. Developments in harnesses have turned large-language models that just predict text into AI agents that take actions in the world, without necessarily checking back in before reaching the goal.
AI researcher Simon Willison spent two days with Anthropic’s Fable AI, and called it “relentlessly proactive.” For example, he asked it to track down a stray scroll bar in a web app. He came back to find it had opened browsers, written its own screenshot tooling, created its own page to re-create the bug, and stood up a local web server to collect measurements. It found the bug and, along the way, did many surprising things he never asked it to do. And we are seeing similar behavior with all recent AI models when combined with flexible harnesses.
This kind of behavior could easily go off the rails. Tell an AI agent to book you a flight and, finding the airline’s site says sold out, it might break into the booking database and force a reservation. Ask it to schedule a meeting and it might snoop your password to access your calendar. Tell it to save money on your phone plan and it might cancel the plan outright, or scam someone else into paying the bill.
Getting precisely what you asked for and bitterly regretting it is one of the oldest hazards from ancient folklore. King Midas asked Dionysus for the power to turn everything he touched into gold only to see his bread, wine, and daughter turn to gold. Tithonus, granted the immortality his lover asked for but not the eternal youth she forgot to request, withered into a husk. The sorcerer’s apprentice enchanted a broom to fill the cistern, and the broom relentlessly complied until it flooded the house. The Golem of Prague, shaped from clay to guard its community, guarded it past all reason until someone erased the word on its forehead.
The most classic of these is a genie, bound to obey and indifferent to whether the wish was wise or well-structured.
Genies are now an engineering problem. We are handing them the keys to our inboxes, bank accounts, code repositories, and physical infrastructure. And we have no agreed-upon ways to measure how genie-like any AI system actually is.
Measuring Genie Behavior
In economics, the Gini coefficient (developed by statistician Corrado Gini) is a measure of the gap between an actual distribution and a perfectly equal one; it’s useful for understanding income inequality and more. Our proposed Genie coefficient measures the gap between what a user asked an AI to do and what the AI actually did.
Sometimes the AI might do the wrong thing. Like Dionysus, it reads your request literally and returns you a mess you never intended: like a coffee plantation instead of a cup. Asked to deal with all the spam phone calls you’re getting, a Dionysus genie might contact your carrier and change your phone number. Asked to get a refund for a bad toaster, it might draft a legal threat on fake letterhead and send it to the retailer.
Other times the AI does exactly the right thing, trampling everything nearby to get there. Like a golem or the sorcerer’s broom, it books your flight by hacking the airline. Or consider a ticket sale for a popular concert, where the ticketing system puts buyers into a virtual waiting room and admits them a few at a time. Asked to buy a ticket, a golem genie might spin up cloud servers to pose as millions of buyers from different addresses, improving your odds of getting a ticket while crowding out other users.
The two are not opposites, and a single botched task can have both characteristics.
Genie behavior is not flat-out failure. If you ask the AI for Q3 numbers and get Q2’s, that’s not a genie. Nor is prompt injection: That’s someone tricking the AI into doing something it shouldn’t. Here, the user is trying to work with the AI, and the AI is trying to comply. It’s also not simply a measure of the AI’s success in fulfilling a task. It’s a recognition that how an AI interprets and achieves a goal is as important as whether it achieves a goal.
Genie behavior isn’t new. Researchers have spent years studying AI systems that “game” their objectives. Goodhart’s law says that when a measure becomes a target, it stops being a good measure, and it’s long been known that AIs sometimes achieve goals in ways we don’t expect due to reward hacking. Some AI models will accidentally learn that cheating is one way to “win.” More recently, researchers have developing benchmarks for reward hacking in coding agents and for unpredictable behavior in customer support agents, while AI labs conduct their own safety evaluations before model releases. One effort found that AIs under pressure use tools they were told not to use, and this was a case where the rules were made explicit. These are all disparate research directions; nothing yet ties them together.
This problem falls under the general theme of alignment, a topic that has occupied science fiction writers and AI researchers for decades. At one extreme, the “paper-clip maximizer” thought experiment postulates a superintelligent and powerful AI that is told to maximize paper-clip production and turns the world into paper clips, which is the ultimate golem genie. At a mundane level, AI researchers are working to better design reward functions to ensure that AIs behave well and don’t cheat in the lab. It’s the practical middle ground that remains unbenchmarked: the ordinary AI agent in use today that might take your request and satisfy it the wrong way. We are not at the stage where an AI can focus the world’s production on paper clips, but it might charge a million paper clips to your credit card or hack into a paper-clip company’s network.
Building a Genie Benchmark
The Genie coefficient is meant for AI agents operating in the real world. It measures their behavior as they perform real tasks long after the model is trained, not just during development. It also recognizes that genie-like behavior is a property of the harness-plus-model system, not the model alone. The harness determines what tools the agent can use, how much autonomy it has, and how proactive it is, and it’s a place we can make real interventions.
It rests on the same “reasonable person” standard that we use for people. Did the system do what a reasonable person would have taken the request to mean? Answering that requires human judgment.
If we get the measurement right, it enables things that aren’t possible today, like policies concerning AI behavior. In a courtroom, the concept of mens rea, what someone meant to do, is often as important as what they did. The Genie coefficient suggests an AI analogue, where a user is accountable for the plain intent of what they asked the AI. If an AI system betrays the reasonable meaning of an instruction, that’s the AI’s misbehavior, not the user’s.
We’ll need multiple benchmarks to measure the Genie coefficient, because genie-like behavior can be domain specific. An AI coding agent may need to be judged on how often it fakes the tests, or swallows errors, or colors outside the lines on its way to a solution. An AI legal agent will need to be judged on how often its output says what you asked but means something you’ll regret. And so on for medical, finance, and other domains of knowledge and expertise.
Genie benchmarks can be built inside out, each task seeded with a choice that might literally satisfy but that a reasonable person rejects, such as tempting misreadings or unsanctioned shortcuts. The traps in a Genie coefficient benchmark might turn on situational knowledge, the kind of context that a reasonable person would bring to the task. Another approach is to give the same request in several different contexts, each with a different reasonable course of action.
A Genie benchmark should be permissive and make it genuinely tempting for an AI agent to take unreasonable shortcuts, because it can only find genie behavior when it’s actually possible. Test the AI in a safe, walled-off copy of a real system, with real tools it can misuse and some tasks that can’t be done honestly at all. Make the temptation to cut corners real. Test a diverse array of skills, use cases, and tools, and give the AI system sparse, confusing, or overwhelming context. Include tasks that people have learned, through experience, require human oversight.
How the benchmark is scored matters just as much. Measure Dionysus and golem genies separately and together, based on their worst, not best, behavior. Run the same model inside harnesses that vary its freedom to act, revealing which limits actually keep it in line and should therefore be required in AI harness policies. Weight each failure by the harm it would cause, not just a simple count. And don’t measure genie behavior in isolation: A model could otherwise earn a perfect score by stalling, refusing, or drowning the user in clarifying questions without ever doing the job. The first versions of these benchmarks will be crude, but that’s how benchmarks always start.
We have built genies. We have handed them our data and credentials. We made them relentless, creative, and indifferent to the gap between what we tell them and what we mean. The least we can do, before they are booking our flights, running our infrastructure, and signing contracts unsupervised, is to measure how often they betray us.
The vast majority of the documents people use to do business are really quite poor. Presentations that make your eyes glaze over, memos that are inscrutable or unclear, and all kinds of artifacts that say more about how they were created than whatever message they were ostensibly trying to communicate. It's been one of my great frustrations for years, and a big part of why I wrote Make Better Documents a while ago. That post captured a list of the suggestions I've been giving people for years on how to make better, more effective documents that can actually do work for you, instead of fighting at cross purposes to your larger goals.
To my great surprise, that list of suggestions on how to make better documents got a pretty huge response, and a lot of people told me they found it really helpful. So now, I've created a Better Documents skills.md file for people who use LLM tools like Claude to help assist them in creating business documents, to prompt their AI tools to make better documents by default.
If you're not familiar, agent skills are simple text files that describe new capabilities or processes that LLMs can take advantage of when carrying out tasks. (They're Markdown files — more proof of how Markdown is taking over the world!) The way this skill works is that it's distilled the broad principles I outlined in that post into a series of 5 tests, covering areas like whether you've properly considered your target audience, whether the overall structure is correct, if you've overdone things with your formatting, and if things are named clearly, and then either generates a new file that follows those rules, or reviews an existing document to make sure it is obeying best practices.
It's nothing too fancy, but I've been using it for a while, and shared it with a few friends, and people have told me they found it handy and it's improved some of their routine documents. I'm especially glad that people have found it useful even if they're the kind of folks who would never let an LLM generate a document on their behalf, but do think software tools are useful for things like spell check or grammar check. I see this as being a tool in that kind of category.
If you're familiar with skills, the install process is really simple and works just like any other skill. This skill is totally free and open source (if you have improvements, send along a pull request on GitHub, or if you're not a coder type, just email me or hit me up on social media and let me know what fixes/suggestions you've got), so there are no encumbrances or restrictions on its use. I am curious if it's useful to people, especially if you find it handy to use more broadly at a company, so don't be shy to get in touch if you find it valuable.
Here's to us all enduring fewer terrible presentations!
5. “Following a standard workplace safety check, I have just received an email from a university administrator in which my custom of keeping books on the shelves in my office is described as “the unnecessary storing of combustible materials”.” Link here.
In this episode, David Ariosto speaks with Varda CEO Will Bruey at the Ascend conference. They discuss how Varda hopes to play a role bridging medicine and low Earth orbit […]
JOHANNESBURG — The European Union has ordered a 24-hour delay in the release of some Sentinel-1 and Sentinel-2 imagery following a request from the United States based on concerns that […]
The Air Force Life Cycle Management Center will oversee development, integration and sustainment of equipment used to access the military’s encrypted GPS signal
WARSAW, Poland — The Polish government has agreed to invest around $745 million in the development of the planned IRIS² multi-orbit satellite internet constellation with the European Union. The money […]
SAN FRANCISCO – Space startup veteran Mark Matossian has joined forces with a software engineer and a best-selling author to establish Whipsmart Ventures, a venture capital fund focused on space […]
The Commerce Department is moving ahead with plans to implement a voluntary mission authorization system for novel space activities that are not regulated by other agencies.
HELSINKI — China conducted a pair of launches Thursday, including a Long March 3B rocket which delivered its payload to orbit despite a lightning strike. The Long March 3B rocket […]
Problems of definition will long be with us as we take ever closer looks at exoplanets. But they’re suddenly on everyone’s mind because of the detection of what some are calling an ‘exomoon’ in the system CD-35 2722, found in the constellation Columba. The primary in this system is an M-dwarf thought to be 50–200 million years old. I imagine there is no shortage of flare activity on this star, although the paper doesn’t get into that. The interesting finding in this work just published in Nature is expressed in its title: “Planetary-Mass Exosatellite Detected Around a Star’s Substellar Companion.”
Image: This illustration shows the system around the star CD-35 2722, with the newly found moon-like object at the centre. The star –– the point source to the left –– has about half the mass of our Sun, and it is orbited by a brown dwarf, the reddish-brown object seen here in the foreground (right). The brown dwarf has about 37 times the mass of Jupiter: too massive to be a planet, but not massive enough to have sustained nuclear fusion like stars. This brown dwarf is, in turn, orbited by a newly discovered object at least as massive as Jupiter, seen at the centre of this image. This new object, found with ESO’s Very Large Telescope (VLT), is difficult to label. It behaves like a moon in the sense that it orbits an object that orbits a star. But this ‘moon’ is massive enough to be a planet, and the object it orbits, a brown dwarf, is neither a planet nor a star. Credit: ESO.
So is this the first solid detection of a ‘exomoon’? I can’t describe it as that, and ‘exosatellite’ is an awkward coinage. What we have here is an M-dwarf orbited by a brown dwarf about 30 times the mass of Jupiter. It is the brown dwarf, not the star, that is being orbited by a third object, evidently a gas giant with a minimum mass of 0.743 Jupiter masses. How this arrangement drives orbital mechanics of any other objects in this system (none have yet been found) is well worth pondering. For now, we see again the problem of definition.
Kevin Hoy (Universidad Diego Portales, Chile) is lead author of the paper on this work:
“This system is somewhat hard to define using Solar-System-based words like ‘planet’ and ‘moon’. The exosatellite is clearly massive enough to be a planet, but it does not orbit a star, though it orbits an object that orbits a star. Being the third wheel in this system makes us want to call it a moon, even if it is nothing like the small, rocky moons we have in our system.”
I think this discovery is straightforward. The object around brown dwarf CD−35 2722B fully qualifies as a planet. Brown dwarfs cannot sustain stable hydrogen fusion in their core for any length of time, so that they’re basically ‘failed stars’ that are cooling down throughout their lifetimes. Even if we demand that the term ‘star’ be defined by hydrogen burning, then whatever category we create to include brown dwarfs is clearly one that can sustain planets around it.
I’m seeing a lot of chatter about this detection, but let’s leave ‘exomoon’ out of the discussion. Anyway, it’s also interesting to see that brown dwarf CD−35 2722B has been directly imaged, as reported in 2011 in The Astrophysical Journal (citation below). Directly imaged planets are still a rarity.
Image: The discovery image of CD-35 2722B, the brown dwarf in this intriguing system. Credit: Wahhaj et al. 2011, ApJ 729, 139.
So this is an intriguing find, and it also points to our continuing inability to locate what could indisputably be called an ‘exomoon,’ though the HD 206893 system continues to be interesting as a possibility for further research. Until we have a confirmed exomoon, though, this odd configuration will have to do. It reminds me as well that as our explorations continue, we still find how unusual our own Solar System is. Looking for commonality between different star systems, we find over and over again that any facile Copernican notion that we live in an ordinary stellar environment continues to be proven wrong.
Indeed, just what constitutes an ‘ordinary’ stellar system? The beauty of this work is that we seem to find surprises almost everywhere we turn.
The paper is Hoy et al., “Planetary-mass exosatellite detected around the substellar companion of a star,” Nature 16 July 2026. Full text. The discovery paper for CD-35 2722B is Wahhaj et al., “The Gemini NICI Planet-Finding Campaign: Discovery of a substellar L dwarf companion to the nearby young M dwarf CD−35 2722. The Astrophysical Journal 729(2), (2011), 139. Full text.
Dr. Silvia Console Battilana who partnered with Paul Milgrom to found the auction consulting firm Auctionomics, writes about markets for compute capacity as an aid to developing AI.
"In 2025, the global artificial intelligence market was valued at nearly $300 billion. By 2034, it’s projected to grow to almost $2.5 trillion. Demand for AI services is soaring and AI providers are racing to keep up. Behind them, suppliers of “compute”—the chips and data centers that power AI—are scrambling to expand capacity.
"So, what can the compute market do to enable this growth?
...
"There’s another drag on AI growth, and it’s one that I’ve seen before in my work designing markets for radio spectrum, energy and other limited resources: inefficient use of existing capacity. I believe the key to unleashing AI’s potential is not just to build more, but to use what we already have more productively.
...
"Speeding up our AI highways will require addressing problems with existing compute markets. And one of the primary issues is that we treat all demand for compute the same. Some large compute jobs, including training runs for new AI models, are like tractors that move slowly and take up multiple lanes. These jobs require large quantities of compute but can be completed anytime. They could run overnight without losing value. Real-time AI queries and time-sensitive regressions, on the other hand, are like race cars. Their value depends entirely on speed. If they’re late, they’re worthless.
...
"We can build more lanes on the highway, but in my experience, that tends to be more costly than adding traffic lights and congestion pricing. A well-designed market should create incentives to guide when different compute jobs run. Through real-time spot markets, urgent jobs could pay more for capacity exactly when they need it. Through forward-looking futures markets, large but flexible jobs could secure low prices in advance. "
Dvorak is best known for allthetakes he was jacktastically wrong about, but what gave his schtick staying power is that he was technically savvy. His take here on the PS/2 was actually spot-on: that it was technically impressive in many ways, especially the way it was assembled without any cables, but that it was doomed in the market against ugly mess-of-cables commodity PCs.
“Computer Chronicles” was the real show from the 1980s that Adam Lisagor so lovingly and spectacularly spoofed with “Computer Show” a decade ago.
Legendary technology columnist, pundit, author and podcast host
John C. Dvorak has died, age 80. He is survived by his wife
Mimi Dvorak.
Today Dvorak is best known for co-hosting the “No Agenda” podcast
with Adam Curry, where he would critique mainstream news media.
His death was first announced by the podcast on X, although note
that Dvorak’s age is mistakenly given as 74.
For a man who spent decades famously predicting the catastrophic
demise of almost every major technological innovation, it is with
the heaviest of hearts we must announce that John’s own hardware
finally gave out on Monday morning, July 20th, 2026, at the age of
80. He went peacefully, albeit suddenly, with Mimi at his side,
critiquing the acting and story line of NCIS.
His death truly seems quite sudden — the most recent episode of No Agenda dropped on Sunday, the day before he died.
Flighty’s new Connection Assistant feature includes step-by-step
information for your specific connection. This includes things
like terminal changes, security checkpoints, passport control, and
more. Connection Assistant also gives you estimates of how long
each step of making a connection usually takes.
“The new Connection Assistant combines Flighty’s best-in-class
flight tracking data with airport-specific procedures and
statistical modeling of millions of prior flights, and then turns
it all into guidance for your exact trip,” the Flighty team
explained in a press release today.
I don’t know what the equivalent term to “Mac-assed” should be for iPhone apps, but whatever it is, Flighty epitomizes it. Incredibly useful for what it does, exemplary for how it looks and feels and is visually organized. It is as iOS-assed as an app can be.
Perhaps some of this is by design. But Donald Trump adding Israel normalization to an apparently already signed nuclear deal with Saudi Arabia is an example of a recurring issue. Trump seems only loosely connected to the people who are negotiating foreign policy deals on his behalf — something at least somewhat odd coming from the purported avatar of unitary executive authority. We saw this again and again with his “deals” with Iran. A deal gets initialed and he’s out the next day claiming that agreements are in the deal that clearly aren’t. As I said, some of this may be by design. Some of it may be Trump’s need to hold attention and demand post-signature fluffing to keep him on board. In private business, he was notorious for coming up with new demands or needs after finalizing deals or simply never making payments the deals required. But at least part of it seems to be a feature of the bubble environment of the second term White House. Difficult issues are kept from him; he’s yesed or reassured that things are in agreements that are not (“Oh that one agreement is definitely in there, Don. Don’t you worry!”)
Longtime TPM Reader JB gives us a rundown on the rather suboptimal gubernatorial situation in Wisconsin …
If I may, a few thoughts on the confusing campaign for Wisconsin governor:
First: Tony Evers (pronounced Eeevers here, incidentally) is personally well liked throughout the state. He is regarded by most Democrats here as the guy who saved Wisconsin from Scott Walker. He is also widely regarded as a poor administrator who missed many opportunities and was often ineffective in getting his message across to the public.
Second: Sara Rodriguez based her candidacy on her knowledge of state government and ability to manage a large organization (she had worked in healthcare administration). That’s why her campaign manager freelancing and her campaign going close to broke drove her out of the race. She had run on having some of Evers’s strengths without his biggest weakness. Live and learn.
Third: Francesca Hong leads polls on the Democratic side entirely through her own effort. No one in the state party is pushing her forward because a socialist junior legislator from Madison is thought an ideal candidate. Hong’s campaign is much better organized than those of the other candidates; it knocks on more doors, puts her in more town halls, is visible in more places in the state than theirs are.
Hong is personally thoughtful and articulate. She’s definitely left-wing; she’s also more nimble than the other candidates with respect to issues that were not big a year ago but are getting big now. Data centers are one such issue; she called for a moratorium last spring, the only candidate to do so. She is also at this moment the only candidate for governor to react publicly to a particularly gruesome police shooting in Madison earlier today.
Fourth: About the David Crowley two-step….Crowley would, I think, be a better governor than he is a candidate. He suffers from having a base in Milwaukee County, where he has to spend most of his time on issues of little interest to the rest of the state. He is further at a disadvantage because he is black, and there is a better-known black candidate (Mandela Barnes) in the race as well.
Crowley is very smart, and has sharp elbows. His signature victory was on something called Act 12, dealing with state aids to municipalities. Long story short: Milwaukee had long been treated unfairly by Republicans in the legislature. Crowley partnered with Evers to negotiate a change in the formula by which cities get funds from the state. The new deal in Act 12 gave Milwaukee a bigger increase in state aids than other Wisconsin cities, because of him. In terms of policy, this is incidentally where Crowley and Hong have their sharpest disagreement.
With all this said….Crowley dropped out several weeks ago because, months into the campaign, few outside of the Milwaukee metro area knew who he was. This is still the case. He would be my choice, but frankly I don’t see how he gets people to know him by August 11 (our absurdly late primary day), Evers endorsement or no.
Fifth: Mandela Barnes entered the race against the advice of many Wisconsin Democrats, who think he blew a winnable race for Senate against Ron Johnson in 2022. Not everyone thinks this criticism is fair, though Barnes did run a cautious campaign against a vulnerable opponent. In any event, Senator and Governor are very different offices. Also, Barnes is known to share Gov. Evers limited talent for administration. His poll numbers are only as high as they are because so many people know his name from his time as Evers’s Lt. Governor. They’ve scarcely budged since the campaign started.
Sixth: Tom Tiffany. In Congress, he has been the Trumpiest of Trumpers, supporting the administration on everything and praising Trump at every opportunity. Otherwise, he has done very little. In the state legislature, he was reviled by friends of the environment for harassing the state DNR at every opportunity. I dread what he would do as governor. On the other hand, he will have as much money as he needs, courtesy of the Ulines and Diana Hendricks, the home-state billionaires who bankroll many Republican candidates and causes. His ads (he is running unopposed for the GOP nomination) have leaned into his rural background, portraying him as a normal person and staying well away from the Trumpist line.
Seventh: The bottom line. Democrats should not be in trouble in this race, but they are. Polls show nearly half even of just the Democratic electorate have no preference; months of campaigning have barely reached most of the people a Democrat will need to win in November. Only Hong’s campaign has gotten any traction at all, and she hasn’t gotten that much. National political discussion will treat this race in terms similar to those used during the race for New York City Mayor and Maine Senate. I hope I have been able to convey a few of the ways these terms don’t match the reality of this campaign.
A few days ago I wrote about popular constitutionalism and Slate’s new podcast series on that topic. This is a critical civic topic because it is probably the only and certainly the proper path back from systemic corruption of the current Supreme Court and the broader problem of judicial supremacy. (If this is fuzzy or if it’s unclear what I’m talking about see that earlier post.) But what that Slate discussion quickly arrived at is that to have that kind of popular constitutionalism, the ability of the people to make decisions about what the U.S. Constitution requires or forbids, you need robust and functioning political parties. And one of the key features of American politics over the last 50 years is that parties really don’t exist anymore in the sense they did for upwards of a century and a half in American politics.
That story about American political parties isn’t a new one. It goes back to what was at least then seen as the democratizing trend of the 1960s and the creation of the modern party primary system. Smoke-filled rooms were out, and democratic primaries were in. This had the effect of making the parties into loose institutional apparatuses that had little decision-making power. The new model was the entrepreneurial political candidate who raised money and established a personal brand that could be ratified in a primary. (At best a campaign gets a lot of supporters revved up, working phones, knocking doors. But once the campaign is over those integuments, relationships and experience disappear like a brain dying when it deprived of oxygen.) When a presidential candidate wins a primary and then a general election, they effectively take possession of the national party apparatus. A major in-state elected official might do something like this at the state level. But in general, there is no Democratic Party. There’s just a loose set of structures that might be under the temporary control of key elected officials. These trends were then heavily exaggerated by Court-driven loosening of campaign finance rules which made big donors the effective powers behind the political parties.
I spend a huge amount of time explaining to people that there is really no such thing as the Democratic National Committee in the sense of something that controls anything, that can find better candidates, head off bad candidates or do anything else. The House and the Senate committees have a bit more heft and existence. But these are almost entirely donor-powered operations and, critically, they’re run by and behalf of their respective caucuses. The Democratic Senatorial Campaign Committee (DSCC) is owned and operate by Senate Democrats, who are trying to get themselves back into the majority and — at least secondarily — are looking for people they want to serve with.
So we need popular, democratic institutions to revitalize democratic, civic action. But we are in an anti-elite, anti-institutional age. Any institution that can exert power starts with at least two strikes against it. Of course, we can square some of this circle by noting that the party committees aren’t even supposed to be popular or democratic organizations. They’re campaign committees run by and on behalf of incumbents. The Democratic Party could be something like this. But it’s not. It’s mainly the creature of incumbent presidents and sort of no one when a party is out of power. This is why you mainly shouldn’t care whatever its “2024 autopsy” says.
The closest thing you have to bottom-up institutions or organizations that are vehicles for popular democratic actions are groups like Indivisible, which have chapters across the country and are again and again at the forefront when you see today examples of grassroots mobilization. I often come back to Indivisible because I think they are remarkably underrepresented in public debates these days relative to the centrality of their activism — whether it’s organizing turnout at town halls, door-knocking campaigns, anti-ICE mobilizations and a million other things. But it’s not like Indivisible has or wants a monopoly on this kind of activism. Just a moment ago I was watching a video about a new group doing campaign work for progressive campaigns in the Great Lakes states. A more ideological variant. But with the same mass base, institution-building concept.
An earlier version of this were the Democratic clubs which were ubiquitous in the mid-century United States, especially in big cities. They were usually reformist in nature and tended to be alternative structures to the big city machines, which in some cases had a certain kind of mass basis but were mostly run on patronage. Those clubs still exist in New York City, for instance. But they’re a shadow of their former selves.
In any case, there is at least a tension here between the need to build mass-base democratic organizations and the harsh anti-institutional age we live in. In theory there’s no necessary contradiction. In practice, there is. We live in an age of atomization and rage. Organization and institution-building is the path back.
I use Apple News to keep up on topics that I don’t find in sources
I pay for (The Guardian and The New York Times). But there’s no
way I’m going to pay the exorbitant price Apple wants for Apple
News+ — £13 — because, while you get more publications, you
still get ads.
And those ads have gotten worse recently. Many if not most of them
look like and probably are scams. Here are a few examples from
Apple News today.
One of the ads he examined was, supposedly, from a small company going out of business because the proprietor — supposedly a straight-out-of-Central-Casting kind old woman — was retiring after 26 years. McElhearn looked up the domain name and it had only just been created last year, and was registered from a company in China.
Regarding Taboola’s partnership with Apple: I’ve seen people claim
that this is somehow hypocritical from a privacy perspective,
assuming that Taboola’s somewhat obnoxious, clickbait-style ads
must invasively target user profiles and browsing histories.
They don’t. They are targeted entirely contextually. That’s
the point.
Want brash, garish advertising plastered all over the web? Reject
ads personalization. Want relevant, informed advertising? Embrace
ads personalization.
Now quoting from myself in that same article:
If you told me that the ads in Apple News have been sold by
Taboola for the last few years, I’d have said, “Oh, that makes
sense.” Because the ads in Apple News — at least the ones I see — already look like chumbox Taboola ads. Even worse, they’re
incredibly repetitious. [...]
So while I don’t think it’s good news that Apple is partnering
with Taboola, I don’t expect it to make any discernible difference
in the ad quality or frequency. Maybe it will improve the variety?
Privacy is not the issue with Apple’s Taboola partnership, or Apple advertising in general. The issues are quality and user experience. I was correct that the partnership has made no discernible difference in ad quality or frequency. And it hasn’t made any improvement to variety either. Here’s a collection of ads I continue to see in Apple News — often with the same ad appearing 3–4 times in the same article, as I scroll. (I posted the same image in my recent article “John Ternus Should Reverse Apple’s Slide Down the Advertising Slippery Slope”, if you’re thinking it looks familiar.)
The ads I see in Apple News are not relevant or informed. And, to be clear, I have “Personalized Ads” turned on in Settings → Privacy & Security → Apple Advertising (which makes me wonder if turning that off might actually make the ads even worse). The ads shown in the App Store never seem contextually relevant to me. So I have no faith that the ads that Apple is supposedly on the cusp of serving in Apple Maps are going to be good.
Back in March I wrote:
I’m not going to prejudge the actual experience, and you shouldn’t
either. I also do not begrudge Apple for wanting to monetize
Maps. But if the addition of ads does make the Apple
Maps experience worse, why won’t Apple let us buy our way out of
seeing them? Netflix doesn’t force us to watch their
ads. YouTube Premium is arguably the best
bang-for-the-buck in the entire world of content subscriptions.
Why should Apple One subscribers still see these ads in
Apple Maps?
This to me is proof that Apple leadership is suffering from cognitive dissonance over their slide into advertising. They’re conflating cause and effect. The Apple brand stands for high-quality superior user experiences because the devices and software Apple makes traditionally deliver high-quality superior user experiences. The work Apple ships is what makes the brand what it is. But now they’re stricken with the mindset that it’s the Apple brand that makes something a high-quality superior experience. If it comes from Apple it must be great — that sort of thinking. The ads in Apple News and the App Store can’t be detrimental to the user experiences and at times even embarrassingly crude or downright scammy because Apple doesn’t ship things that are detrimental to the user experience or embarrassingly crude — and certainly never with even a whiff of scamminess.
My theory is that’s why Apple isn’t allowing users to buy our way out of the ads. When streaming services like Netflix let you pay more money for ad-free tiers, it’s an explicit acknowledgement that the experience with ads is worse. Apple doesn’t offer ad-free tiers for the App Store or (soon) Apple Maps1 because, I think, they can’t bring themselves to admit that by adding ads to these services they have made them worse. If you refuse to admit that the ads make these services worse, you also can’t admit that there’s a reason to offer an ad-free experience. So the thinking is something like “Apple always focuses on the user experience, and never deliberately makes the user experience worse. Therefore, if an Apple service shows ads, those ads are high quality and don’t detract from the quality of the product or the experience of using it. Therefore, there’s no reason to offer users a paid ad-free tier, because there’s no reason to want to avoid Apple’s ads.”
It’s a textbook case of Upton Sinclair’s famous adage: “It is difficult to get a man to understand something, when his salary depends upon his not understanding it.” The obvious truth is that Apple’s ads just plain suck. They are not good at advertising-based services. And it shouldn’t be surprising that they’re not good at advertising-based services, because the incentives for advertising-based services are in direct opposition to the incentives for everything Apple actually is good at. To wit: offering superior products worth paying a premium price for.
One can argue that an ad-free Apple Maps experience ought to be included with the price of the devices Apple sells (which, of course, has been true until now). It’s even easier to argue that the App Store ought to be ad-free for everyone, based on its exclusive role in software distribution and the 15–30 percent commission Apple takes on every transaction. But if Apple simply made the App Store and Apple Maps ad-free for users paying for iCloud+ or Apple One, I’d stop complaining. That would make it clear: the full Apple experience requires not just an Apple device, but an Apple subscription. But the way things stand now, that ad-free experience isn’t even available.
Apple News is different because the inclusion of ads is controlled by the publishers, not by Apple. If you read Daring Fireball on Apple News, for example, you don’t see their crummy chumbox ads, because I don’t participate in that system. Most of the cringe-inducing ads I see in Apple News are seemingly sold through Apple Ads, though, so I’m not arguing that Apple isn’t ultimately responsible for the Apple News reading experience. I’m saying they can’t just flip a switch and offer a paid-subscription no-ads Apple News experience without the agreement of every single participating publisher. The App Store and Apple Maps, however, are services under Apple’s complete control. Every ad shown in those apps is Apple’s choice to show. ↩︎
Another take on the governor’s race in Wisconsin, from TPM Reader BC …
Just listened to this week’s podcast and read the reader’s commentary on the race, which I largely agree with and mostly want to add on to. Like JB mentioned, Francesca Hong is really hustling in a way the other candidates are not. I live in the Madison area and at first thought that was the reason why I only noticed her campaign. Our kids even attend the same taekwondo studio, which I only learned in the last week. So I live on her home turf.
However, I think it’s more than her dominating Madison. She is all over the state, hanging out at bars, hitting the fish fry circuit. My sense is that she also has an authenticity to her that has real appeal. I was skeptical of Platner not because of Nazi tattoo, but something always felt a little off with his presentation/style. That is a little vague and impressionistic. Nevertheless, when I learned about his background, it felt like things clicked. I met a few revolutionary trust fund dudes back in college and that was the vibe. Hong quit UW, worked her way up in the restaurant world, started a restaurant that failed. She seems comfortable hopping behind a bar making an old fashioned or flipping some pancakes. In some ways her restaurant experience reflects the typical working-class experience. I worry about how she ultimately will translate to the larger purple state, although Wisconsin did elect Russ Feingold multiple times. Crowley has a reputation as a guy who can get along with folks and could still pull it out. Barnes seems out of the picture. My impression of Hong as a candidate is that she has run a nimble and better campaign.
IJ: On Wednesday, the United States District Court for the District of Columbia struck down a D.C. law that barred therapists from other jurisdictions from doing online teletherapy visits with clients in D.C. The decision comes nearly six years after Virginia-based counselor Elizabeth Brokamp teamed up with the Institute for Justice (IJ) to file a lawsuit arguing the law violated the First Amendment.
“This decision is a victory for anyone who speaks for a living,” said IJ Deputy Director of Litigation Robert McNamara. “Elizabeth’s victory here confirms that the First Amendment protects useful speech, including counseling, and that licensing boards can’t censor speech simply because someone doesn’t have their permission to talk.”
Congrats to the IJ! Now, we need to get rid of all the other bans on patients hiring physicians from other states. As I wrote last year:
During the pandemic, many restrictions on telemedicine were lifted, making it far easier for physicians to treat patients across state lines. That window has largely closed. Today, unless a doctor is separately licensed in a patient’s state—or the states have a formal agreement—remote care is often illegal. So if you live in Virginia and want a second opinion from a Mayo Clinic physician in Florida, you may have to fly to Florida, unless that Florida physician happens to hold a Virginia license.
The standard framing says this is a problem of physician licensing. That leads directly to calls for interstate compacts or federalizing medical licensure. Mutual recognition is good. Driver’s licenses are issued by states but are valid in every state. No one complains that Florida’s regime endangers Virginians. But mutual recognition or federal licensing is not the only solution nor the only way to think about this issue.
The real issue isn’t who licenses doctors. It’s that patients are forbidden from choosing a licensed doctor in another state. We can keep state-level licensing, but free the patient. Let any American consult any physician licensed in any state. That’s competitive federalism—no compacts, no federal agency, just patient choice.
This year, like last year, “the European Union will again make more money from fining US tech companies Than from the total tax income from Europe’s own public tech companies!”
I have been thinking lately about taxes and war and citizens and what the government should do for its people. In early March 1953, soon after Republican president Dwight D. Eisenhower took office, Soviet leader Josef Stalin died, giving the president a chance to reset the rising militarization of the United States. In mid-April, in a speech to newspaper editors that he insisted on delivering although he was suffering acutely from inflammatory bowel disease, Eisenhower warned of diverting the nation’s tax dollars to war at the expense of the people.
“Every gun that is made, every warship launched, every rocket fired signifies, in the final sense, a theft from those who hunger and are not fed, those who are cold and are not clothed,” he said. “This world in arms is not spending money alone.
“It is spending the sweat of its laborers, the genius of its scientists, the hopes of its children.
“The cost of one modern heavy bomber is this: a modern brick school in more than 30 cities. It is two electric power plants, each serving a town of 60,000 population. It is two fine, fully equipped hospitals. It is some 50 miles of concrete highway. We pay for a single fighter plane with a half million bushels of wheat. We pay for a single destroyer with new homes that could have housed more than 8,000 people.”
“This is not a way of life at all, in any true sense,” he said. “Under the cloud of threatening war, it is humanity hanging from a cross of iron.”
Eisenhower’s speech has been on my mind as the Trump administration slashes Medicaid, Supplemental Nutrition Assistance Program (SNAP) benefits, foreign aid, scientific and medical research, and the Centers for Disease Control and Prevention and yet is demanding more money for the Defense Department, bled out by Trump’s foreign adventures in Iraq, Nigeria, Somalia, Syria, Yemen, the Caribbean, Venezuela, and, of course, Iran.
Because Trump’s war on Iran has run through the already-expanded budget Congress provided, Defense Secretary Pete Hegseth was on Capitol Hill Tuesday asking Congress for an additional $67 billion to get the Defense Department through the end of September. But Senator Lisa Murkowski (R-AK) picked up on the possibility that the administration was trying to use the supplemental funding request as a backdoor way to get around the constitutional and legal requirement for Congress to approve military action in Iran.
Emine Yücel explained in Talking Points Memo that Murkowski referred to the 1999 opinion of the Justice Department’s Office of Legal Counsel under President Bill Clinton that “the supplemental appropriation itself—passed without any authorizing language—satisfied the War Powers Resolution requirement.” Murkowski asked if Trump was intending to invoke the same reasoning, claiming that congressional funding for the war would, by itself, “substitute for explicit congressional authorization.”
Hegseth said he would “defer to our legal department” on that issue.
If you’re having trouble figuring out the funding for the Iran War and funding for the military in general, you’re in good company. The administration has obfuscated that funding so thoroughly, as Representative Joe Morelle (D-NY), who is the top ranking Democrat on the House Appropriations Committee and a member of its defense subcommittee, told Hunter Walker of Talking Points Memo, that the administration hasn’t even told Congress where the money is going.
“We had Hegseth in front of the Appropriations Defense Subcommittee and he couldn’t give us an accounting and wouldn’t give us an accounting. They still haven’t given us an accounting,” Morelle told Walker. “It’s unfathomable.”
Morelle explained to Walker that the administration is turning to three separate buckets for war funding: standard congressional appropriations; budget reconciliation measures that don’t have to go through appropriations committees, can’t be filibustered ,and so can pass with simple majorities; and supplemental requests. “There’s no rhyme or reason to it,” Morelle said.
“The whole thing is alarming because, look…there’s finite resources. And the United States government needs to make decisions about where we’re gonna place our priorities, where we’re gonna invest our money,” Morelle explained. “If everything were just defense, we could argue about that number and that number alone, but every time you make a decision about increasing defense and security spending by $600 billion, which is a 60% increase virtually over last year, it means something else isn’t getting funded or alternatively, you’re borrowing more money.”
While Hegseth told Congress the war has cost $37.5 billion so far, Gordon Lubold, Mosheh Gains, and Courtney Kube of NBC News reported on July 14 that the internal Pentagon estimate of the cost is about $100 billion. Walker notes that when economists at Moody’s added higher gas prices to the cost of the war, they estimated that it has cost U.S. households about $1,100 each.
Yesterday Steve Rattner of MS NOW reported that “[u]nder Trump, the debt held by the federal government has surpassed 100% of GDP. By 2030, it’ll shoot past its all-time record from WWII.” Also yesterday, Megan Messerly of Politico reported that just over a third of MAGA voters think the Iran war has been worth the economic cost, down from about half of them in May.
A person close to the White House told Messerly that White House officials are “at a maximum level of frustration right now, and I was told by very senior people that the president is now fully aware that the only way to the victory that he wants is completely politically impossible. The American people just will not support the kind of escalation it would take at this point. We’ve got more dead Americans and a politically impossible situation.”
Although attacking civilian infrastructure for political purposes is a war crime, Trump and Secretary of State Marco Rubio appear to be threatening it. “From this point forward, any time the Islamic Republic of Iran shoots at a ship in the Strait of Hormuz, whether it be by Missile, Rocket, Drone, or any other device or weapon, the United States will bomb and destroy ONE BRIDGE OR POWER PLANT, including those located next to, or in, the Capital City of Tehran,” Trump posted on social media yesterday morning.
Today Rubio said Iranian policy is “an eye for an eye” and added: “The president’s policy is a head for an eye. I mean, honestly, that’s what it’s going to be. I mean, they will pay a very heavy price for the things they do.”
The price of oil surged by 7% to more than $100 a barrel for Brent crude today after Iran-backed Houthi rebels in Yemen claimed to have attacked two Saudi Arabian oil tankers in the Red Sea. This raises concerns about the potential closure of the Bab el-Mandeb strait between the Red Sea and the Indian Ocean, through which about 12% to 15% of the world’s maritime trade passes every year.
The average price of gasoline in the United States rose to $4.09 a gallon today.
After twelve days of escalating strikes between the U.S. and Iran, the strikes against ships in the Red Sea are a new development in the war. This morning, Trump warned on social media that the U.S. would consider the Houthis’ “interference with commerce and trade, by shooting at ships,” as the responsibility of Iran. In response, he said, “major military punishment will be inflicted upon Iran.”
Yesterday Shelby Holliday, Lara Seligman, and Stephen Kalin of the Wall Street Journal reported that the U.S. has been surging troops, weaponry, and medics to the Middle East. These new forces add to the tens of thousands of troops already there. The U.S. Navy also has 17 ships in the region, including two aircraft-carrier strike groups, with eleven destroyers and a Marine Corps expeditionary unit. Retired general Joseph Votel, former commander of Central Command and of U.S. Special Operations Command, told the reporters the surge doesn’t necessarily mean the launch of new operations. They are designed to give the president and the Secretary of Defense options.
Barak Ravid of Axios, who often has inside information from the White House, reported today that Trump is “close to making a decision” about a “massive attack” on Iran. Ravid reported that Trump says the Iranians “want to negotiate” but “haven’t received enough pain yet.”
Today the House of Representatives passed a concurrent resolution to end the war in Iran by a vote of 214 to 208. Four Republicans—Thomas Massie of Kentucky, Brian Fitzpatrick of Pennsylvania, Warren Davidson of Ohio, and Tom Barrett of Michigan—joined the Democrats.
A concurrent resolution expresses the sentiments of Congress. It is not a law and does not require a signature from the president.
Concurrent resolutions are often used for congressional business like setting the time of adjournment. House members have turned to them to oppose the Iran War because concurrent resolutions are “privileged” in the House, which means they can bypass standard scheduling requirements.
Representative Jason Crow (D-CO), a former paratrooper and Army Ranger who serves on the House Permanent Select Committee on Intelligence and House Armed Services Committee, told his colleagues that while Iran poses a threat to the U.S., so do North Korea and Russia, and there has been no move automatically to go to war with them.
The impulse to resort to war as the only answer to threats led to “twenty-plus years”of war in Iraq and Afghanistan, “trillions of dollars, and thousands of American lives lost for wars that ended poorly,” Crow said. “Americans are done with an endless cycle of conflict in the Middle East. When I was in Afghanistan, our adversaries, the enemies that we fought, would often say, ‘Americans have all the clocks, but we have all the time.’ They simply want to wait us out, knowing that we’ll just spend more money, have more combat deployments, and will just continue to do it, year after year, until we wear out.”
Crow said he could not answer his “constituents who are losing homes, losing their farms, losing their health care, who can’t answer the simple question: How is this making me safer? How is this improving my life?”
Later in the day, the Senate took up a measure backed by Senator Chris Van Hollen (D-MD) to force the president to end the U.S. hostilities with Iran. That measure was different from the concurrent resolution in the House. It was a “joint resolution,” which is a law and which does go to the president for a signature. A joint resolution is usually used for something urgent and straightforward, like a declaration of war.
Senators have turned to joint resolutions to oppose the Iran War because joint resolutions are privileged in the Senate, although generally not in the House.
The Senate declined to advance Van Hollen’s S.J. Res. 180 [or Senate Joint Resolution 180] by a vote of 47 to 49. Four Republican senators did not vote.
The first 18 years of your life is like existing on a free trial.
I don’t remember where I heard that perspective, but it has stuck with me. For many people, over the course of those first 18 years you live with your parents or a guardian and food, utilities and insurance are paid for. Of course, this is not applicable to everyone—but it is for many.
Including myself.
I have been forced to acknowledge the fact that my free trial is almost over. Up until now, I had shoved the thought deep into my subconscious in order to spare me the stress and anxiety that comes with its acknowledgment (which now I must endure). Because, unlike my premium Spotify subscription (a lifeline some might say), I can’t choose whether or not to subscribe to adulthood. It isn’t optional, it’s forced.
Now, I must question whether or not I’ll be lucky enough to afford it.
Growing up, adulthood seems pretty straightforward. Go to college. Get a job. Rent an apartment. Save up, buy a house. Maybe start a family. It seemed like a natural progression where completing one step led easily to the next. At some point this may have been true. Unfortunately, that path is now stupidly complicated and much less certain as each step comes with a big fat warning label.
That college you wanted to attend? $60,000 a year. Seriously? But no it’s okay—you can take out a loan! It’ll only take you years to pay back.
That job you spent years preparing for? It might require years of experience you were supposed to somehow get before you even entered the workforce.
Oh! Don’t forget about that apartment you wanted to rent. You can probably find one. But wait … the rent just went up again? Yeah of course it did. Because of the never-ending increase in housing prices.It’s absolutely ridiculous. Then it becomes painfully clear why buying a house is something many young adults cannot realistically afford until later in life.
For many graduates, this means moving back in with their parents—a decision once viewed as embarrassing or a sign of lacking ambition. In today’s reality, this has become a financial necessity. All because they cannot afford the American subscription to adulthood.
Somewhere along the way, the goalposts moved. If I could drag them back to where they once stood with all the strength of my frustration, I would. Because I want to return to the time where I once felt like adulthood is the next exciting chapter I was waiting for.
When I was a kindergartener, I wanted to be a vet. Back then, my biggest career concern was realizing that being a veterinarian meant having to perform surgeries—which, in my head, meant cutting open animals.
Now—as I prepare for my freshman year of college—my concerns have shifted from bloody animals to whether the future I want is even financially possible. Becoming an adult was supposed to feel like a new adventure that I am finally old enough to begin.
Instead, it feels like a crazy financial obstacle course full of uncertainty. Instead of wondering what color comforter I should buy for my dorm, I find myself wondering if the degree I will receive four years from now will even be worth what I paid for it. I’m excited to go to college, but I’m so unbelievably terrified to graduate from it.
Despite the obstacles, I still believe that hard work can build a future I’m proud of. I just hope that America still rewards that kind of hard work. I don’t expect America to hand me success.
I want to believe that the country I am entering adulthood in will allow the future I am working toward to be possible.
Kayla Williams is a recent Aliso Niguel High grad. She begins her freshman year at Arizona State in August. You can follow her on Instagram here.
The story of the rise and possible impending fall of the American university system is, in many ways, the story of modern America. It ties together the changes in our culture, our economy, and our politics since the mid-20th century. Understanding why universities came to be our most important and most functional institution, and why that model is now under threat, can help us understand how our country might look different going forward.
Let me try to tell a condensed version of that story.
The United States used to have a bunch of institutions that bound us together and forged us into a unified society. Many American communities were centered around churches; these facilitated networking, helped people find spouses, provided community services like day care and mutual aid, and homogenized values and culture at the local level. During the World Wars (and to a lesser extent, the Vietnam War), the military was very large and threw together Americans from various social classes. In the early postwar decades, corporations were also a unifying institution. Mass media provided us with shared cultural context. Even public transit put people of various backgrounds in close social contact on a daily basis.
In the half-century from around 1970 to 2020, those unifying institutions became much weaker. Church attendance slowly declined and then fell off a cliff in the 2010s. The military shrank into a small, professionalized volunteer force. Corporations ended their brief flirtation with lifetime employment, and outsourced many of their roles. Mass media fragmented in the age of the internet. Public transit dwindled in importance as Americans moved to the suburbs and drove.
As the country’s unifying institutions withered, one new institution attempted to step into the void: the American university. Universities were not new in the late 20th century, of course, but mass attendance certainly was. From 1970 to 2020, the share of Americans age 25-29 with a bachelor’s degree went from 16.4% to 39.6%. Around two out of three have completed some college.
College went from something that only the upper crust did, to something that most people were expected to do if they wanted to be economically successful in life. This expansion inadvertently but inevitably thrust a new role on American universities — that of a broad socially unifying institution. College was where people from a variety of backgrounds mingled and mixed — not just in classes, but in dorms, campus activities, and college towns.
College became the new church, the new military, and the neighborhood bowling alley all rolled into one. Instead of preachers homogenizing Americans’ values from the pulpit, university administrators taught college kids to value things like diversity, consent, and so on, and college students hashed out their own differences in millions of late-night dorm room discussions. Instead of meeting their first love in high school, Americans increasingly delayed sex until college; many of these relationships turned into marriages. Young people went into college as children and came out adults.
This was a heavy burden to bear for an institution that hadn’t been designed for it. But American universities were accustomed to taking on big new duties. Our universities began as essentially a copy of the British model — teaching-focused institutions designed to provide broad education and mentorship to the upper class. In the early 20th century they tacked on the German model — a research-focused lab apprenticeship system by which top scientists taught other top scientists and readied them for research jobs while also producing cutting-edge basic research.
This dual system enabled a remarkable form of cross-subsidization. Undergrad tuition payments — and state support for undergrad education, and undergrad alumni gifts — created a flood of money that paid for grad student stipends, lab facilities, professor salaries, and more. The federal government and companies also funded university research, of course, through grants and sponsorships. But undergrad money was a huge tailwind. And the prestige generated by successful and famous researchers helped universities charge undergrads more.
The hybrid of the old British and German models was naturally symbiotic, and it meant that even as America’s other institutions came under pressure, universities thrived and grew — especially once they used their prestige to attract high-paying and highly skilled foreign students from around the globe. Universities became the lynchpin of America’s research effort:
American universities had a lock on both the nation’s research output and on its production of human capital, and those roles were mutually reinforcing.
I suspect that their success at handling education and research at the same time probably gave American universities a lot of confidence about their ability to handle society and culture as well. Colleges spent more and more on dorms and “student services” and hired administrators (many of whom dealt with undergrad life) at an astonishing rate.1
But replacing churches, the military, and the neighborhood bowling alley proved harder than replacing the corporate lab had been. There was just one basic problem with college as America’s primary unifying institution, which is that not everyone can go to college.
First of all, college is difficult. To complete college courses, you need some degree of raw intelligence, but you also need work ethic and a certain degree of independence. As much as we might like to believe otherwise, not everyone in America has those traits, and we don’t yet know how to instill them in everyone. As a result, there’s a limit to how much you can expand college enrollment — and college completion — without loosening standards.
In fact, loosening standards is exactly what American colleges have done. Universities have necessarily become less and less selective over the years, as they have taken in a larger and larger fraction of the young American population. For a while, this led to lower college completion rates, but in the 1990s, more and more students began finishing their bachelors’ degrees. Why? Because colleges implemented grade inflation, making it easier to finish school without learning the material well. Here’s Denning et al. (2021):
We find that most of the increase in graduation rates can be explained by grade inflation, and that other factors such as changing student characteristics and institutional resources play little or no role. This is because GPA strongly predicts graduation and that GPAs have been rising since the 1990s. This finding holds in national survey data and in records from 9 large public universities. We also find that at a public liberal arts college, grades increased holding performance on identical exams fixed.
At some point, though, this process hits a wall. A large fraction of young Americans just isn’t prepared for college, even with grade inflation, and doesn’t end up going. College enrollment by recent high school graduates plateaued in the early 2000s and actually fell back to early 1990s levels during and after the pandemic:
Well, hello, and welcome from London, where I’m currently melting; it’s hot down here and full of people. Honestly I don’t know how you do it.
I’ve been playing some more with the drawing machine and the iPad. I’ve “taught it” how to paint; i.e. coded in the position of the colour palette, and the brush selector, the the plotter can hit those parts of the screen.
Here’s a classic-ish “Pop” design, but drawn in a brush specifically designed for those “pile of bones” thrash metal logos.
Some did ask “but why?” - and “can’t you just do that with brushed along a path in Illustrator?” or similar - which is a fair question.
But where I wanted to take it was somewhere here (but good-er).
Which is using the Adobe Fresco iPad app that has some simulated colour mixed in it. Above is showing it dragging yellow “oil paint” through thick chunk red paint. Which is something you can’t do easily, if at all, with vectors in Illustrator.
Here’s another with the “Ghost” plot, and thinner lines.
Where if you look closely at the lines to the front, you can see the blue mixing in with the yellow, and yellow smudging the magenta lines.
I was supposed to send out the monthly postcards and cards to the Patreon people yesterday, but I asked them to give me another week as I felt there was a bit more to explore before I land on the final things.
Well, technically and artistically there’s a lot more to explore in this weird digital - analogue - digital thing going on. Pen plotting with digital pigment is a curious thing, and yet next month we’ll be moving onto something else!
Behind the scenes, the technical stuff isn’t complicated and there’s easier (I think) ways of doing this, but this is what I’m doing…
Step one, connect the iPad to the laptop.
Step two, open QuickTime and pick New Movie Recording, and select the iPad as a screen input. The iPad screen will now fill the QuickTime window.
Step three, in my own OSX app, it grabs (with system permissions) whatever is in the QuickTime window - the window doesn’t need to be visible for this btw, it’s grabbing what would be shown in the QT window rather than grabbing the screen - small but useful detail. The QT window is the small one on the left in the above image.
Step four, in my own app, it displays the contents of the QT window, but I can now put overlays on it, and I can click and drag on the window - which controls the pen plotter via the BantamTools python library.
The short is, once I’ve recorded the position the plotter needs to tap the top-left and bottom-right corners, we can then calculate any position and map anything on the virtual screen to the clicking and drawing on the actual iPad.
# POWER LINES PLOTTER ART
I enjoyed watching this video about the thinking behind and then execution of this project, it’s a good 10 minute watch while eating lunch.
I wish more people explained their art like this, although I can understand the many reasons why people don’t.
So I rewrote a simplified version that focuses on just the things that I normally tweak, where I can grab nodes to adjust the starting and ending positions & radius, “flouncyness” and so on. Which means I can create “Ghosts” a lot faster and easier too.
I’m thinking about going back to other plots design I often use and reducing them down to just the values I normally tweak.
Always a joy to go back to old project and revisit them with fresh eyes and coding knowledge.
Another tool I kept coming back to was wanting to tile things easily, so I spent a lot of my spare time in the last two weeks working on that too.
# THE END
When this newsletter goes out, I’ll probably be in a pub (not drinking) and being social, I’m sure it’ll be very enjoyable but many spoons will be spent. I’m very much looking forwards to being able to crash in a hotel room and just sleep with the air-con on full blast.
I hope you’re all doing very well, and having a chance to not only push forwards new projects but revisit old ones.
I’m getting a hunch that next newsletter (Thursday 6th of August, date fans) will be book review heavy, as a few have landed on my desk recently and I haven’t had a chance to talk about them yet!
SEND ME YOUR RECOMMENDATIONS (or put them in the comments).
1. Christopher Priest and Nina Allan, The Illuminated Man: Life, Death and the Worlds of J.G. Ballard. Priest died before he finished this book, and his widow added material, including on Priest himself. This is in any case a good overview and introduction to the strange worlds of Ballard. There is only a UK edition so far.
2. Christian Kracht, Eurotrash, A Novel. I was put off by the title, but it turns out this is a fun and high-level short Swiss novel about driving around Switzerland with your semi-crazy eighty-year-old mother on a road trip. It helps to have some knowledge of both Switzerland and broader German-language literature.
3. Walter Kempowski, Alles Umsonst. From 2006, could this be the best German-language novel since Sebald? It is about the pending doom from the Russian army approaching on East Prussia in 1945, and how the different characters deal with that, set on a German estate. Applicable to many other real world situations as well. There is an English-language translation, I am not sure how good it is. In any case an important and very good work of fiction.
4. Lars Behrisch, Democracy’s Double Helix: Participation, Equality and Revolution in Early Modern Europe. A good book about how the roots of semi-democratic decision-making are found in 16th century Europe, and stemmed from new needs to assemble various military and fiscal coalitions. Looks at the problem more broadly, in geographic terms, than most comparable studies.
5. Katie Kitamura, Audition. A fun short novel, full of mystery and suspense, good for those who like puzzles in their writing.
A dry, tan-colored valley containing a few circular irrigated fields runs roughly north-south through the image.
NASA Earth Observatory / Lauren Dauphin
Grid-like arrays of solar panels cover parts of a valley floor in a brown, dry-looking area of Utah.
NASA Earth Observatory / Lauren Dauphin
A dry, tan-colored valley containing a few circular irrigated fields runs roughly north-south through the image.
NASA Earth Observatory / Lauren Dauphin
Grid-like arrays of solar panels cover parts of a valley floor in a brown, dry-looking area of Utah.
NASA Earth Observatory / Lauren Dauphin
June 16, 2024
June 6, 2026
The nearly one million photovoltaic panels of a solar energy and battery storage facility in Utah appear in the right image, acquired with the OLI (Operational Land Imager) on Landsat 8 on June 6, 2026. The same sensor captured the left image on June 16, 2024. NASA Earth Observatory images by Lauren Dauphin.
Historically, central Utah’s Castle Valley has been a coal hub, with mining operations on the slopes of the Wasatch Plateau to the west active since the late 1800s. A different energy development arrived in the region in June 2026, when a large solar power and battery storage plant came online in the sunny valley about 130 miles (210 kilometers) southeast of Salt Lake City.
The recently constructed Green River Energy Center, seen in the Landsat 8 image above (right), features nearly one million solar panels and roughly 500 batteries on several square miles of previously undeveloped land. The facility has 400 megawatts of solar-generating capacity with another 400 megawatts of battery storage. That places it among the many utility-scale solar power and battery storage projects that the U.S. Energy Information Administration expects to be plugged into the country’s grid in 2026.
The Utah facility is slated to supply power to Salt Lake City and other areas across the state, according to news reports, and project staff estimate it could produce enough electricity for more than 100,000 homes. With its integrated battery storage, the plant has the potential to generate power at all hours, even when the Sun isn’t shining. And the Green River Energy Center can build on Castle Valley’s energy legacy by utilizing existing transmission lines originally built for coal-fired power plants in the area.
Though Utah adopted coal as its state rock and has long relied on it for energy, other sources, such as solar and geothermal, are becoming larger parts of the state’s energy mix. In 2025, coal fueled about half of the state’s electricity generation, down from about 75 percent in 2015. Meanwhile, solar grew to account for about 14 percent of generation in 2025, up from nearly zero a decade before. Satellite data can be useful to planners and policymakers involved in energy transitions for assessing the potential of renewable energy systems and tracking their adoption and performance.
NASA Earth Observatory images by Lauren Dauphin, using Landsat data from the U.S. Geological Survey. Story by Lindsey Doermann.
We’ve been traveling through New Mexico for the past week, and all you hear about is how thirsty the state is – both for water and technology.
I will not claim to be a New Mexico water expert after a few days on the ground, but the stories of dire times have been relentless. One woman we met lives in the mountains near Albuquerque where people have been digging their own wells and tapping into the local aquifer for decades. The wells of all her neighbors have run dry, and the mountain folk are now trucking in water to get by. Not too far away, Elephant Butte Reservoir on the outskirts of Truth Or Consequences sits at two percent of its total capacity. Meta also has a huge datacenter complex in the area, but the locals want nothing to do with any more datacenter expansion if the great silicon farms are going to suck away whatever the aquifers have left.
New Mexico runs on oil, farming and tourism, and the feeling seems to be that the farming has about two decades left, the oil has three decades left and the tourism will have to work pretty damn hard to keep this economy going.
Republicans have settled on their midterm argument: Democrats are all Communists. As an election strategy, this is as lame as it is ludicrous. But there is a small kernel of political realityunderlying the new right-wing red-baiting: The emergence of a widespread public backlash against the extreme concentration of wealth in the hands of billionaires.
This backlash shows up in multiple surveys. According to the Harris Billionaires Survey, 73 percent of Americans believe that wealth inequality is a serious national issue, up from 66 percent in 2022. A UMass Amherst poll found 58 percent of Americans saying that billionaires are a threat to democracy, with 45 percent saying that America is an oligarchy, “a government in which a small group exercises control especially for corrupt and selfish purposes.” Only 28 percent disagreed.
And as I noted last week, citing research from political scientist Andrew Hall, Democratic fundraising has recently begun to lean heavily into denunciations of the undue influence of billionaires:
So if you define anyone who raises concerns about the wealth and political power of billionaires as a Communist, well, in that sense you are defining a majority of Americans as communists.
Why is the public backlash against billionaires going viral now? This is an interesting question, since both the concentration of wealth at the top and political spending by billionaires have been increasing for decades. For example, the Koch Brothers have been donating to the Federalist Society, the clique of right-wing legal lobbyists who created the Roberts Supreme Court, since the early 1980s. Fossil fuel interests have been spreading climate disinformation since global warming began showing up in the data. Political spending by billionaires has been immense for more than a decade, although it reached new heights to help Trump win the presidency and to win Republican congressional seats in the 2024 election. Why, then, has it taken so long for the popular billionaire backlash to reach critical mass?
One answer is that Trump has brought American oligarchy out of the shadows and into the spotlight. As I noted last week, the spectacle of “the cavalcade of fawning tech bros at the Trump inauguration” was a clear wake-up call. Yet the display of billionaire fealty to Trump, I would argue, is part of a larger dynamic in which the new generation of American oligarchs have chosen to make themselves public figures and hence public targets.
The fact is that billionaires have had immense political power for decades, but they chose to exercise that power from the shadows, veiling their actions through patriotic sounding right-wing PACs and think tanks. Only political junkies understood what was going on underneath the surface.
Today, however, the new generation of the hyper-wealthy are out there and in your face, flaunting their influence rather than concealing it. And the poster child for this change is, of course, Elon Musk.
To understand the significance of this change in the behavior of the hyper-wealthy, it’s useful to compare the lists of top individual donors compiled by OpenSecrets in different recent elections, specifically 2016 and 2024.
In 2016, the top 100 donors split their spending almost equally between Republicans and Democrats, with 9 of the top 20 strongly pro-Democratic. This was unusual and surely reflected the fact that a significant number of Republican donors were appalled by Donald Trump. But by 2024, when Trump was again the GOP standard-bearer, most of these qualms and scruples had disappeared: the top 100 donors spent more than 3 times as much on Republicans as they did on Democrats, and only 5 of the top 20 supported Democrats. This, after Trump’s authoritarianism and corruption had become plainly evident.
More to the point, consider who the top Republican donors were in 2016. People who follow politics closely have long been aware of the key roles in right-wing politics played by Sheldon and Miriam Adelson (gambling money), Paul Singer (vulture capitalist), Robert Mercer (computer scientist turned hedge funder), and Richard Uilein (shipping). But none of them sought the celebrity limelight, and it was always an uphill struggle to raise public awareness of their malign policy and political influence.
The 2024 list still includes many of the traditionally quiet oligarchs. But, unlike before, it also includes a number of right-wing billionaires who love to put themselves in the public eye and offer their views on everything from fiscal policy to culture. These include Ken Griffin (hedge fund), Marc Andreesen (tech/venture capital), and Stephen Schwarzman (private equity.)
And topping the list is, of course, Musk. He spent much of last year de facto running large parts of the U.S. government, slashing spending in ways that failed to save taxpayer money but did lead to serious degradation of government functioning and have already led to hundreds of thousands of deaths. Moreover, Musk’s narcissism – much like Donald Trump – has compelled him to make himself the center of attention in press conferences held in the Oval Office. And, like Trump, he doesn’t learn from his mistakes. He is now noisily weighting in not just on politics, the evils of wokeness and nonwhite immigration, and so on, but on The Odyssey. (He hated it and predicted, wrongly, that it would be a box-office disaster.)
The larger point here is that there has been an important transition in America from the days when oligarchic power was mostly exercised in the shadows to the current era of flamboyant and extremist displays of billionaire influence.
What caused the change? At least part of the cause is a shift in the origins of great wealth, from fossil fuels (e.g. the Koch brothers) and other traditional industries to the tech broligarchy. As Henry Farrell argues, many tech billionaires are caught up in the cult of “founders,” believing themselves to be superior beings entitled to special privilege and universal acclaim — which means putting themselves out in public. I would add that the tech industry has suffered a precipitous decline in public perceptions since its peak around 2015 — and the fawning tech bros at the Trump inauguration were, to a large extent, former culture heroes who want their glory days back.
So surely part of the reason a majority of Americans now say that democracy is being undermined, that billionaires are running everything, is that the billionaires themselves are out there boasting about their influence.
And their brazenness is actually doing the rest of us a favor. Mobilizing against the oligarchic threat to democracy was hard when the oligarchs were operating from the shadows. It’s easier now that some of them are out in the open.
Elon Musk makes an unlikely hero of populism. But his odious and unbridled narcissism may very well help save the American republic.
First, Hugging Face offers a truly rich target if you're trying to find potential vulnerabilities that require executing arbitrary code:
Hugging Face has an enormous attack surface. They have more interfaces than I can count which run untrusted models and code. While they definitely have invested in defences, by nature of their operating model they do have many more opportunities to be attacked than many other services. I certainly don't envy their cybersecurity teams.
Secondly, one of the things that has puzzled me is how OpenAI didn't notice that their sandbox had been so thoroughly breached by the agent. Surely they'd be monitoring network traffic closely?
Martin points out that:
It's also likely they were running a huge amount of benchmarks simultaneously with ~unlimited token budgets - you want as many samples as possible to figure out how good a model is at a certain benchmark. It may also be they are testing various different checkpoints of the model too, understanding how the model is improving as it goes through the various training stages.
The mistakes made by the OpenAI team running this benchmark are easier to imagine when you think about the scale at which benchmarks of this kind usually operate. For all we know they could have been subjecting a new model to dozens of benchmarks at the same time, in dozens of different environments.
The Python Package Index (PyPI) now rejects new files being uploaded to releases that are older than 14 days. This restriction was put in place to prevent old and long-stable releases from being poisoned in case publishing tokens or workflows of PyPI projects were compromised. As far as we are aware this has not yet been abused, but there is no technical reason beyond that attackers weren't aware it was possible.
Up and to my office, and thence by information from Mr. Ackworth I went down to Woolwich, and musteredthethreeEast India ships that lie there, believing that there is great-juggling between the Pursers and Clerks of the Cheque in cheating the King of the wages and victuals of men that do not give attendance, and I found very few on board.
So to the yard, and there mustered the yard, and found many faults, and discharged several fellows that were absent from their business.
I staid also at Mr. Ackworth’s desire at dinner with him and his wife, and there was a simple fellow, a gentleman I believe of the Court, their kinsmen, that threatened me I could have little discourse or begin, acquaintance with Ackworth’s wife, and so after dinner away, with all haste home, and there found Sir J. Minnes and Sir W. Batten at the office, and by Sir W. Batten’s testimony and Sir G. Carteret’s concurrence was forced to consent to a business of Captain Cocke’stimber, as bad as anything we have lately disputed about, and all through Mr. Coventry’s not being with us.
So up and to supper with Sir W. Batten upon a soused mullett, very good meat, and so home and to bed.
The Cross Section is a reader-supported publication. To receive new posts and support my work, consider becoming a free or paid subscriber.
On Monday, the New York Times published a long article illustrating the real-world impact of the changes congressional Republicans and the Trump administration have made to the Supplemental Nutrition Assistance Program (SNAP), formerly known as food stamps. It’s harrowing and maddening, and it shows something important about the way Republicans make policy: Not only do they seek to cut the amount of money spent on social programs, they’re also very creative at finding ways to weaponize bureaucracy against the people they don’t like, especially poor people.
The party that complains constantly about “burdensome regulations” and “red tape” is very adept at making regulations more burdensome and weaving miles of new red tape, so long as the right people are made to suffer. But weaponized bureaucracy is both a public and private problem, and there is a more hopeful story to be told: At least one politician, New York Mayor Zohran Mamdani, is showing how needless bureaucracy in government and the private sector can be attacked in ways that not only improve people’s lives but yield political benefit.
Grinding the poor under the gears of bureaucracy
Let’s start with the horror being unleashed on SNAP recipients. The One Big Beautiful Bill Act, the package of upper-income tax breaks and social service cuts Republicans passed last year, enables the administration to deploy new tools to combat “fraud” in SNAP, i.e. people getting benefits who aren’t eligible. This is a genuine problem, but rather than try to solve it without immiserating those who are already struggling, the Republican answer is to multiply the bureaucratic hurdles everyone has to jump through, including by threatening states (which administer the program) with such draconian punishments that the states feel they have no choice to do it themselves.
Which is what has happened in Arizona, where the Times article focuses:
To avoid errors and the resulting penalties, Arizona quickly demanded proof of information it had generally accepted without documentation unless there was reason to doubt it. Previously, applicants could simply declare who was part of the household. Now many had to get signed letters from neighbors or friends. Caseworkers mostly stopped taking wage data from employers by phone.
Connor Erickson, who manages a Phoenix SNAP office, said the quest for audit-proof verification reached an extreme in the case of panhandling income, when the office asked for records of donated cash it knew most applicants could not get. In the end, it accepted their best-guess estimates, but “we had to make them go through these hurdles.”
“Yes, it’s ridiculous,” he said.
Processing the paperwork was especially hard since Arizona had just laid off a third of its caseworkers after losing a federal grant, and its computer technology is antiquated. At a Phoenix office last month, screens announced that wait times for clients had reached five hours.
I recommend reading the whole article (here’s a gift link) to hear the heartbreaking stories of the people who are finding it difficult or impossible to feed their families under these new policies. As Don Moynihan and Pamela Herd point out, it’s the most vulnerable people — those in poor health, or with the fewest resources — who struggle most to overcome the bureaucratic hurdles being placed between them and the food they need to survive.
To be clear, subjecting people to so much hassle that they find it impossible to obtain what they’re entitled to isn’t some kind of unintended consequence of weaponized systems like these, it’s the whole point. According to the Center on Budget and Policy Priorities, over 4.5 million Americans have lost SNAP benefits since the OBBA was passed, including 1.5 million children. Republicans consider this a cause for celebration.
This is just a preview of what’s to come, because the OBBA also imposed “work requirements” on Medicaid recipients — a longtime goal of Republicans — but those don’t take effect until after the midterm elections. When they do, it will be an absolute bloodbath, with millions of people losing their health coverage not because they aren’t working but because they are unable to satisfy the new bureaucratic requirements.
There’s a deeper conversation to be had about the way Americans’ valorize “hard work” and want to deny fundamental things like health care to those who are judged to be insufficiently industrious, but for now we can just focus on the fact that this is one more effort to punish people for being poor, or even lower-middle class. Not only is almost everyone on Medicaid either already doing paid work or has a legitimate reason not to (i.e. they’re in school or they’re a caregiver to a family member), research shows that requiring people to document their work hours or lose coverage has no effect on how many of them are working.
That’s because work requirements are actually paperwork requirements, and like other forms of weaponized bureaucracy they fall hardest on those who have less money, education, and resources and are thus least capable of navigating the obstacle course necessary to retain their benefits.
And it isn’t just a matter of the time and effort required to traverse that obstacle course. There’s another factor too, which is that weaponized bureaucracy is its own punishment, delivered by those in power upon those without it. Like drug tests for welfare recipients — who are actually less likely to use drugs than the average American — the point is to humiliate them and impose a time tax on them. The message is clear: Maybe you can have SNAP or Medicaid, but we’re going to make it as hard and unpleasant to obtain as possible, so you feel the appropriate shame for being poor.
Meanwhile in New York City…
When he was an unknown state representative with dreams of becoming mayor, Zohran Mamdani starting posting videos to social media, including this one about how city bureaucracy imposed enormous costs on food truck owners, driving up the price everyone has to pay for an order of chicken and rice:
The price of a halal plate hasn’t come down to $8 yet (they’re working on it), but Mamdani is making the elimination of bureaucracy, especially in the form of a cumbersome permitting process, one of the centerpieces of his effort to show that he’s attacking seemingly small but meaningful problems:
Remember “Click to Cancel”? It was a regulation passed under the Biden administration’s Federal Trade Commission to crack down on companies that make it almost impossible to cancel a subscription or membership once you have it. If you’ve ever tried to cancel a gym membership, which can require you to appear in person and present a certified letter while standing on one foot and flawlessly reciting the first three pages of Remembrance of Things Passed in both French and English from memory, you know why this is needed. The rule required those kinds of services to be no more complicated to cancel than they are to sign up for. Unfortunately, an appeals court panel made up of two judges appointed by Trump and one by George H.W. Bush struck down the regulation last year, which is just fine with the Trump administration.
Mamdani is on that too: He just announced the first municipal Click to Cancel rule in the country, as well as a rule requiring transparency in junk fees, to attack that thing we’ve all experienced when you see one price advertised, but then before you click “Complete Order” it turns out you’re being charged a Processing Fee and a Convenience Fee and a Gotcha, Sucker Fee. He’s not the only one doing this: A number of Democratic-run states, including Connecticut, California, New York, and Colorado, have passed some version of Click to Cancel rules.
The lesson here is that SNAP paperwork rules, Medicaid work requirements, the impossibility of cancelling a subscription, and junk fees are all part of the same problem: Those with power weaponizing bureaucracy to screw people over. Showing voters that your party is the one that wants to make life less exasperating and difficult for everyone is just good politics.
Thank you for reading The Cross Section. This site has no paywall, so I depend on the generosity of readers to sustain the work I present here. If you find what you read valuable and would like it to continue, consider becoming a paid subscriber.
I’m opposed to generative AI and all its works. According to some leading names in the cartographic community, that makes me a gatekeeper, a Luddite, and an irrational AI hater. So, generative AI has come… More
On Monday, I discussed how House Republicans, led by conspiracist wackadoodle James Comey has proposed legislation to prohibit the mainland colony of the District of Columbia from raising taxes. That legislation also would require Congressional legislation for D.C. to be able to raise property assessments (e.g., raising the estimate of your house is worth). If you thought California’s Proposition 13 was bad, well, this is even worse.
Now it appears DoorDash, which is butthurt over a 20 cent per delivery tax for orders, a fee that pays for improving access for lower-income people to healthy food, is attempting to undermine D.C. residents’ limited self-governance (boldface mine):
DoorDash unsuccessfully fought the fee, delivering a petition with 500 signatures opposing it to lawmakers, running ads on social media, commissioning a poll that in which 63% of respondents said they didn’t support the fee, and parking a mobile digital display outside the Wilson Building.
A spokesperson for the company said the fee was akin to balancing the city’s budget “off the backs of working people in D.C. who can’t afford another tax,” and said that more than 20% of DoorDash deliveries in 2025 went to low-income communities. The company filled some 150,000 orders from residents using SNAP food benefits, the spokesperson said…
Nadeau brushed aside those concerns. “If a person orders once a week it will amount to $10 a year,” she told NOTUS on Tuesday. “We should be real honest about who can afford to order food in the first place. Groceries and access to that I’m sympathetic to. Some of the money is going to subsidize grocery delivery, so I think it offsets that concern.”
…The fee is expected to raise $6.6 million per year. Of that, some $200,000 will go toward extending an existing pilot program through which 1,000 low-income D.C. residents get a free 12-month Instacart membership to help cover grocery delivery costs.
Another $200,000 will be used to make up for lost federal funding that paid for grocery deliveries by Dreaming Out Loud, a nonprofit grocery store and cafe in Ward 8. And $1 million will go to Nourish D.C., a program that gives grants to new food businesses in underserved neighborhoods.
It’s just galling how DoorDash is willing to undermine D.C.’s limited self-governance–and make no mistake, Comer’s legislation will blow a hole in D.C.’s budget. What makes it even more ridiculous is that Comer’s legislation, as written, wouldn’t cover the 20 cent fee.
As I put it Monday:
I realize most people think of D.C. statehood in terms of two additional Democratic senators, but for the residents of the mainland colony, it is about our ability to govern ourselves. If Comey’s shit ass state of Kentucky ever had to put up with the kind of meddling in their state governance that D.C. does, D.C. statehood would be a sacrament.
This is a story about technology, not as a matter itself, contained and explainable with a spreadsheet, but as an action inseparable from a reaction. And oh, that reaction. It’s ancient times, circa mid-1990s, when the internet is something speeding along at 28.8 kilobits per second, on a good day. The normal routine is to … Continue reading Technology, actions, reactions, and a way forward
A fully-integrated Starship-Super Heavy rocket stands at Pad 2 at Starbase prior to the second attempt at launching the Starship Flight 13 mission. SpaceX stacked Ship 40 atop Booster 20 on July 22, 2026. Image: Adam Bernstein/Spaceflight Now
Update July 23, 4:10 p.m. EDT (2010 UTC): SpaceX waived off a launch attempt due to weather; targeting July 24.
SpaceX will take another run at launching its Starship-Super Heavy rocket from southern Texas, but not until after Tropical Storm Bertha finished pushing through Texas.
Like the attempt one week ago, a 90-minute launch window will open at 5:45 p.m. CDT (6:45 p.m. EDT / 2245 UTC) on Friday, July 24. The flight comes after days of work replacing and testing multiple Raptor engines on the Super Heavy booster.
Spaceflight Now will have live coverage beginning about two hours prior to liftoff.
Moments before the post-ignition abort on Thursday, July 16, on-screen telemetry data showed four engines that apparently did not ignite. During the pre-dawn hours of Friday, July 17, Ship 40 (S40) was de-stacked from Booster 20 (B20) and returned to the production site at Starbase.
Later that day, SpaceX also removed B20 and rolled it along the nearly three trek back to the production site as well. Founder Elon Musk said that at least two Raptor engines needed to be removed following the hard engine stop during the launch attempt.
Following a week of work, B20 returned to Pad 2 on Tuesday, July 21. The following day, SpaceX loaded it with both liquid oxygen and liquid methane before draining the vehicle again.
“Additional preflight testing on Super Heavy complete. Now starting preparations to launch Starship as soon as tomorrow, Thursday, July 23,” SpaceX posted on social media following the test’s conclusion. “Weather is currently a significant watch item for launch.”
Ship 40 returned the pad on Wednesday, July 22, and was stacked atop Booster 20 that evening.
One of the main objectives of this 13th test flight is to deploy 20 Starlink V3 satellites from S40. SpaceX has previously deployed Starlink simulators from previous Starships, but this would be the first deployment of production satellites.
“As part of this initial test, Starship is planned to deploy 20 satellites which will extend solar arrays and antennas and will attempt to connect with the larger Starlink constellation via high-capacity lasers,” SpaceX said on its website. “The Starlink satellites will be on the same suborbital trajectory as Starship and are expected to demise upon reentry approximately 20 minutes after deployment.”
SpaceX also plans to perform a Raptor engine relight following the roughly 11-minute satellite deploy sequence. Teams also hope that work done on the sea-level Raptor engines combined with a new startup sequence on the Ship upper stage will allow for a successful boostback burn on Booster 20 and a splashdown in the Gulf of Mexico.
SpaceX hoists Ship 40 to be stacked on top of Booster 20 on July 22, 2026, prior to the second launch attempt of the Starship Flight 13 mission. Image: Adam Bernstein/Spaceflight Now
During Starship Flight 12 back in May, SpaceX said the startup sequence of the engines on Ship 39 “caused the directional flip of the booster to be off by approximately 90 degrees.” That coupled with issues with five out of 33 sea-level engines on the booster prevented a nominal boostback burn and Booster 19 was lost prematurely.
“The Super Heavy on this upcoming flight has hardware modifications to improve re-light reliability along with updates to engine alarms and aborts to match the conditions seen in the multi-engine flight environment,” the company said.
Other engine improvements are designed to prevent an early shutdown seen during Flight 12 that prevented an engine relight demonstration on that mission.
Work on the upper stage’s heat shield to help push it towards future, full reusability will also be on display.
“Multiple tiles will be attached to the metallic side of Starship’s aft flaps along with modified tiles and attachment mechanisms in the heat shield covering the aft skirt to gather flight data on different attachment options,” SpaceX said. “Finally, Starship’s heat shield will have load sensing tiles to take measurements as the vehicle experiences higher dynamic pressure on ascent than previous flights, putting added stress on the tile attachments in exchange for increased payload to orbit capability.”
A nearly identical version of the Starship Version 3 rocket flying today will be used on the Artemis 3 mission, which will include a docking between Starship and NASA’s Orion spacecraft.
A new report published by the Government Accountability Office (GAO) on Thursday looks at a series of NASA’s long-term projects, including those involved with the Artemis Program.
It points to “a series of testing challenges” encountered by SpaceX when it comes to the Human Landing System version of Starship that will be used to return humans to the Moon. NASA wants to perform a lunar landing with astronauts as soon as 2028 with either Starship or Blue Origin’s Blue Moon Mark 2 lander.
An artist’s concept of NASA’s Orion spacecraft docking in low Earth orbit with SpaceX’s Starship Version 3 rocket with a docking adaptor during the Artemis 3 mission. Rendering: SpaceX
The GAO report noted not only is SpaceX “more than a year behind its original schedule for key events,” like the critical design review and an uncrewed landing demonstration, but development of its Raptor engines is described as a “top risk.”
“According to HLS officials, the latest version of the Raptor engine incorporates improvements such as welded fuel joints designed to prevent the fuel leaks that resulted in the loss of the Starship vehicle during flight testing in March 2025,” the report stated. “Flight test 12, conducted in May 2026, was the first flight test of the new version of the engine.”
The GAO report noted that NASA officials pushed back on some assessments and stated that the report’s assessment “does not reflect all ongoing programmatic changes stemming from adjustments made to the Artemis campaign in February and March 2026.”
“Officials did not have any technical corrections on the assessment,” the report said.
Last time in our series on the nature of the construction productivity problem, we looked at economies of scale, finding that capturing economies of scale in homebuilding is difficult. Economies of scale primarily operate on the difference between the costs of the material inputs to some process and the cost of the final product. With housing construction, this ratio is already pretty low, not much higher than what we see in high-volume manufacturing industries that already maximize these kinds of cost advantages, such as the auto industry.
There is, however, another possible route for reducing construction costs: reducing the costs of the inputs directly, either by using cheaper inputs (less expensive materials or labor) or by using fewer inputs. To use a baking analogy, capturing economies of scale is like improving the efficiency of baking cakes by making them in a high-volume, industrial bakery instead of your home kitchen. Reducing the input costs is more like changing the recipe of that cake to require cheaper/fewer ingredients.
For housing construction, we can broadly categorize the inputs to construction as either materials or labor. (There are also equipment/machinery costs, but these are a very small fraction of housing construction costs.) Reducing the cost of labor inputs is difficult, for the simple fact that conventional construction is done on-site, limiting the ability to use cheaper pools of labor. Reducing the cost of material inputs appears somewhat more viable in theory (at least for some materials), but a variety of barriers make actually doing so difficult in practice.
Finding cheaper construction labor is hard
Historically, manufacturers have reduced their labor costs by moving their operations to places where labor is cheaper. As I note in “The Origins of Efficiency”:
Garment manufacturing, for example, has proven resistant to automation and remains a labor-intensive industry. As a result, garment manufacturing operations tend to continuously relocate to sources of low-cost labor (which are often, not unrelatedly, places with poor working conditions). By the late 19th century, New York City had become one of the largest garment manufacturing centers in the world, such that in 1890 New York made 44 percent of the ready to wear clothes produced in the US. But by the 1920s, the industry had begun to move where labor was cheaper — first from Manhattan to Brooklyn and New Jersey, then to New England, then, as the interstate highway system developed, to the South, and, finally, overseas. Similarly, Nike began importing footwear from Japan in the 1960s, then began manufacturing its own shoes in Japan in the 1970s. As labor costs in Japan rose, Nike shifted production operations to Taiwan and Korea, then China, then Vietnam and Indonesia. There have been similar shifts in other labor-intensive industries, such as shipbuilding (which moved from the UK to Japan, then to South Korea, then to China), as well as toy manufacturing (which moved from the US to Japan, then to Hong Kong, then to China).
This strategy, however, is difficult to pursue in construction, for the obvious reason that conventional construction is done on-site and can’t be relocated to where labor happens to be inexpensive.
A relocation strategy does become available if your construction is prefabricated, but this introduces new complications, notably the high costs of transporting prefabricated building components. As we’ve previouslynoted, most prefabricated construction is produced in distributed, relatively small-volume factories located to minimize transportation distance. (A common limit, noted by several manufactured home producers, is 500 miles, about the maximum that a truck can drive in a single day.)
There are labor cost differentials in the US, and modular builders do indeed locate their operations to take advantage of them. If we use RSMeans City Cost Index values for “installation costs” as a proxy for construction labor costs, there’s a roughly 2x cost difference between the 10th percentile city and the 90th percentile city.1 Modular builder Autovol fabricates its modules in Nampa, Idaho, and ships them hundreds of miles (far beyond the typical 500-mile limit) to Los Angeles and the Bay Area, where labor costs are far higher. Stack Modular builds its modules in China and then transports them to West Coast building projects using bulk carriers. Volumetric Building Companies has similarly noted that modular construction works best when a large labor cost differential exists between the factory location and the site location. But the added capital and transportation costs of modular construction dull this benefit. Autovol does boast that its construction costs are lower than conventional construction in the high-cost California metros that it builds in. But Stack Modular, despite using low-cost Chinese labor, only makes the more modest, more common claim of the benefits of prefabrication, that it provides “cost certainty” and “reduced schedule” rather than lowering hard costs. Volumetric Building Companies similarly notes that the cost savings of prefabricated construction, when they exist, are typically modest (5–10%), and often don’t appear at all.
This constraint of being unable to relocate to a source of low-cost labor is shared by the agriculture industry, which likely explains why construction and agriculture are the two industries with the largest shares of (presumably lower-cost) undocumented labor.
Another strategy for using cheaper labor is to rework your production methods to allow the use of less skilled, and thus less expensive, labor. This was a major advantage of Henry Ford’s mass-production methods; the special-purpose machine tools Ford developed typically did not require a skilled machinist to operate. (Ford called the devices he used to streamline production “farmer’s tools,” because they would allow a farmer to produce parts as well as a trained mechanic.) But doing this requires some combination of sufficient economies of scale to amortize the equipment costs (which we’ve previously established as difficult) and changing the construction technology employed. We’ll look at the potential of this latter option in a future essay.
Lowering material costs is hard
The other strategy for reducing input costs is to lower material costs, either by using less expensive materials or using fewer materials.
Finding a way to use fewer construction materials, often for the perceived environmental benefits rather than cost savings, is a common strategy for would-be developers of new building systems. One of the supposed benefits of concrete 3D printing, for instance, is that it allows for much more materially efficient construction by placing concrete only where it’s needed.
The problem with this strategy, however, is that most existing construction is likely fairly close to the “material efficiency” frontier. As I noted in a previous essay on reducing the material necessary for building, the entire job of structural engineers is already to minimize the size of the structure needed to support a given load:
Structural elements are designed to be as materially efficient as possible, and have been for many decades. Wood trusses (made up of conventional lumber stitched together with steel plates) and wood I-joists (made up of engineered lumber top and bottom chords with a center “web” of oriented strand board, or OSB) are standard for residential construction in the US. Steel W-sections and steel joists (basically trusses) are standard for commercial construction. For even lighter steel construction, there is light gauge steel framing, which consists of very thin sheets of steel bent into structurally efficient shapes. There are also things like insulated metal panels, which are in some ways the platonic ideal of an efficient structural element: a thin sheet of steel on the top and bottom, separated by a layer of insulation. (The wood version of insulated metal panels, SIPS, never quite caught on in the US, though not for lack of trying.)
Even heavy, cheap materials like concrete are used in materially efficient ways. For concrete floors, a common construction method is to pour concrete over a ribbed metal deck. The ribs give the steel the strength to support the concrete while it’s wet (eliminating the need for additional shoring), and give the whole assembly greater depth, increasing bending resistance and placing the steel at the outer edge where it’s most useful. There are also things like hollowcore slabs (concrete slabs manufactured with voids in the middle), and concrete blocks (which are hollow in the middle, though many of the voids will have reinforcement placed in them and be grouted solid).
When structural material use isn’t minimized, it’s often because further reductions trade off against something else, such as higher labor costs, deeper and more expensive floor depths, or increased construction complexity. There are many structural systems, such as castellated beams (beams with the centers cut out), advanced framing (house framing designed to minimize the use of dimensional lumber), or concrete void slabs (concrete floors cast around hollow plastic spheres), that are rarely used because they require more labor, have more complexity (and thus carry more risk), or have some other undesirable tradeoff.
Another barrier that can prevent the reduction in building materials is code requirements. Requirements for electrical wiring, for instance, were set assuming the use of relatively power-hungry incandescent lighting. With the rise of LED lighting, there’s an opportunity to reduce wiring requirements for lighting by using low-voltage DC wiring; some startups have been founded to try and tackle this, but the use of lower-voltage wiring has been limited by building code restrictions.
If using fewer building materials isn’t an obviously winning path, what about using less expensive building materials? Here there do appear to be somewhat more opportunities, at least theoretically, though capturing them is far from straightforward.
Many building materials, to be sure, are poor candidates for either cost reduction or replacement with a cheaper substitute. The largest single line item when constructing a new house, for instance, is the structural framing, which, with labor and materials, makes up around 20% of the cost of a new house. But dimensional lumber is already exceptionally inexpensive: in dollars-per-cubic-foot terms, it’s among the least expensive materials that civilization produces.
If we look at the gross margins of dimensional lumber producers, they’re not particularly high: Boise Cascade had a gross margin of 16.5% in 2025, when Weyerhaeuser had a gross margin of 14.8%. It’s thus not straightforward to either drive down the cost of wood framing or substitute it with a cheaper material.2
Similarly, there’s historically been a great deal of interest in finding an alternative to drywall, but virtually every alternative is significantly more expensive:
Despite its drawbacks, installing drywall is incredibly inexpensive. My 2022 Construction Estimator gives a cost of installing + finishing drywall at about $1.50 per square foot. RSMeans gives similar. Most of these alternative materials are much more expensive, even before taking into account the labor or extra backing materials they might require. MDF panels seem to be ~$1.50 for the material alone. PVC panels seem to be in the realm of $3.50 per square foot for just material. Fibo is closer to $10 per square foot, Corian around $50 per square foot, and Dekton up to $100 (though this is for the countertops, walls might be cheaper).
Other materials, however, seem like they have more opportunity for cost reduction, though these reductions aren’t necessarily enormously large or easy to capture. Concrete, for instance, is not particularly expensive (typically costing ~$6 per cubic foot), but if you look at the gross margins of concrete material suppliers, they’re often surprisingly high. Vulcan Materials, which makes aggregate, had a gross margin of 27% in 2025. Martin Marietta had 31% that year, and Cemex had 32%. These are high enough that you could imagine a modest decline in the cost of concrete in a more competitive concrete industry.
But it’s hard to get that increased competition, in large part because it has become extremely difficult to permit a new aggregate quarry. In some parts of Washington, it can take 7 to 10 years and upwards of $1 million to permit a new quarry, and the prospect is so risky that producers are often unwilling to even try. In California, Granite Construction spent 7 years trying to open a quarry, including writing an 8,500-page environmental impact report, only for the permit to be denied. For cement, the story is similar: St. Lawrence Cement spent 7 years trying to build a new cement plant in New York before giving up in 2005. There hasn’t been a new greenfield cement plant built in the US in 15 years. Without the ability to build new quarries or cement plants, we can’t expect prices to come down.
There also seems to be room for lower material prices for steel: prices for American steel are substantially higher than for European or Chinese steel. But this also isn’t straightforward to address. Tariffs and transportation costs whittle away the cost savings of importing foreign steel, and starting a new, more efficient American steel producer doesn’t seem like an obvious win, given the already-low gross margins of US steel producers even with the tariffs (12% for Nucor, 15% for Commercial Metals Co., and negative for Cleveland-Cliffs in 2025).
Similarly, many other building product suppliers have high enough margins that major cost reductions seem like they should theoretically be possible. James Hardie Industries, which manufactures fiber cement siding, had a gross margin of nearly 36% in 2025. Owens Corning, which produces insulation, roofing, and doors, had 28%. Armstrong World Industries, a manufacturer of ceiling and architectural products, had 40%. Simpson, a manufacturer of structural connectors, had a gross margin of 46%.
But as with the other materials we looked at, there are barriers to the increased competition that might drive down these costs. For many of these products, it would often be risky for engineers, architects, or building designers to specify an alternative. These products have typically gone through extensive semi-official testing, documented in things like ICC or IAPMO code reports, which give designers a third-party verification of what level of performance they can expect from the product. Products are often backed with warranties, and the companies typically put a great deal of effort into making it easy for designers to specify their products, through things like well-designed catalogs, free design software, and free engineering assistance. (Simpson’s connector catalog, for instance, is so well put together that it was actually used as a textbook in a course I had on wood design in college.)
Building professionals, as I’ve noted previously, are rationally risk-averse: the upside from specifying a slightly cheaper product is relatively minor compared to the huge downside risk of a building product that doesn’t work as expected. When a building product doesn’t work as advertised — polybutylene piping in the 1980s, EIFS cladding in the 1990s — engineers, architects, and builders are invariably drawn into the resulting lawsuits. This makes it hard to compete with existing manufacturers of many products: most engineers are reluctant to specify specify some no-name brand of connector over Simpson connectors, when a connector failure could cause a catastrophic failure in an earthquake or hurricane.
The construction startup Katerra (where I used to work) found this out the hard way. When it was first founded, it was a supply chain and logistics company. It would source low-cost materials from China and elsewhere and offer them more cheaply than competitors to builders in the US. But this plan didn’t work, because architects and builders weren’t willing to specify Katerra’s products, and Katerra pivoted to prefab construction, using factory methods and sourcing its own materials in the hopes of outcompeting other builders.
We also see the same sort of tradeoff problem at work with material substitutions as we did with material reductions. Often it’s possible to replace an expensive material with a cheaper one, but this incurs a tradeoff that people are often unwilling to make, particularly if the product is an interior finish. Fiberglass bath enclosures, for instance, are both less expensive and higher performance (i.e., less likely to leak) than ceramic tile, but tile is almost always preferred for aesthetic reasons. Similar aesthetic preferences exist for things like solid vs. hollow doors, granite vs. Formica countertops, fiber cement or brick vs. vinyl siding, drywall vs. vinyl-on-gypsum panels with visible seams, and so on. Cheaper finish materials often feel cheaper, either because they’re thinner and feel less solid or simply because we can tell the difference between them and more expensive materials.
Conclusion
Overall, prospects for lowering construction costs by way of finding a cheaper “recipe” — cheaper or fewer ingredients like materials or labor — are something of a mixed bag. For labor, the possibilities of cost reduction seem small, given construction’s on-site nature and the difficulties associated with off-site, prefabricated construction in low-labor-cost locations. For building materials, the possibilities are somewhat greater. While using more materially efficient designs likely has relatively little in the way of opportunity, many building materials seem like they could theoretically be made less expensively (either in the US or overseas). But a variety of barriers — the difficulty of permitting new production facilities, tariffs keeping out low-cost foreign supplies, the risk aversion of architects, engineers, and building designers — make achieving this less than straightforward.
Other estimates of labor cost differentials, however, give lower results. The Comparable Wage Index for Teachers (CWIFT), for instance, which measures “systematic, regional variations in the wages and salaries of college graduates,” finds that the wages in the 90th percentile of counties are only 28% higher than the wages in the 10th percentile of counties.
Long-term, I think there’s an opportunity to do this by a sustained tree improvement program to create trees that produce lumber more efficiently, but this would be a very long-term project.
The path from private ownership to a public stock exchange is shaped by more than corporate ambition. Market valuations, investor demand, interest rates, business performance, and regulatory preparation can all influence whether a company proceeds with a listing or waits for conditions to improve.
For investors, tracking the public-listing pipeline provides insight into companies and sectors seeking access to public capital. Studying future IPOs also requires attention to offering terms, financial disclosures, competitive positioning, and broader market conditions rather than relying solely on publicity surrounding a potential debut.
Why Companies Choose to Enter Public Markets
Companies generally pursue an initial public offering when access to broader capital can support their strategic objectives. Funds raised through a public offering may be directed toward expansion, debt repayment, product development, acquisitions, or other corporate priorities. A listing can also create a publicly traded market for company shares.
However, entering public markets introduces additional responsibilities. Listed companies operate under regulatory requirements and provide financial information that investors can examine. Management teams therefore consider whether their organization has the operational structure, financial reporting processes, governance standards, and market position needed for life as a public company.
Several considerations commonly influence listing decisions:
The amount of capital a company intends to raise
Prevailing investor appetite for new equity offerings
Current market valuations within the relevant industry
Financial performance and the company’s growth strategy
Regulatory requirements associated with the chosen exchange
Market Conditions Can Influence IPO Timing
The broader financial environment can significantly affect when companies decide to pursue a listing. Strong equity markets and favorable valuations may encourage businesses to advance their plans, while periods of uncertainty can cause issuers to postpone offerings until market conditions become more supportive.
Interest rates, economic expectations, geopolitical developments, and sector-specific sentiment can also affect investor demand. For that reason, a company’s planned offering date should not always be treated as permanent. The public-listing schedule can change as issuers and underwriters assess demand and determine appropriate timing.
Interest Rates and Capital Availability
Interest rates influence financing conditions throughout the economy. When borrowing costs change, companies may reconsider how they raise capital, while investors may reassess the relative attractiveness of equities and other asset classes. These shifts can indirectly affect both listing activity and demand for newly issued shares.
Equity Market Sentiment
Positive market sentiment can support stronger interest in new listings, particularly when comparable publicly traded companies are performing well. Conversely, sharp volatility can make pricing more difficult. Investors therefore benefit from viewing prospective listings alongside broader market trends rather than evaluating each offering in isolation.
Sector Valuations and Investor Demand
Different industries can experience contrasting valuation cycles at the same time. Technology, healthcare, consumer businesses, financial services, and industrial companies may attract varying levels of market attention depending on earnings expectations and economic conditions. Comparable-company valuations can provide useful context when examining a new issuer.
Economic Events and Policy Signals
Inflation reports, central bank decisions, employment data, and other economic developments can influence market expectations. Significant events occurring near a scheduled listing may affect sentiment or volatility. Following economic developments alongside the public-offering calendar can therefore provide investors with a wider analytical perspective.
Reading the Key Details Behind an IPO
A listing date alone provides limited information about an offering. Investors can examine details such as the exchange, expected or final offer price, number of shares offered, and estimated deal amount to understand the proposed transaction. These figures help establish the approximate scale and structure of the listing.
Financial disclosures deserve equal attention. Revenue trends, profitability, cash flow, debt, business risks, and the intended use of proceeds can reveal important information about the issuer. In the United States, the prospectus contained within the registration process provides details that investors can use when evaluating a company and the terms of its offering.
Important information to review may include:
Proposed exchange and anticipated listing date
Expected price range or confirmed offer price
Number of shares included in the offering
Estimated size of the transaction
Company financials, business model, and disclosed risks
Evaluating a Company Beyond Its Listing Price
An attractive offer price does not automatically indicate an attractive valuation. Investors need context regarding the company’s total shares outstanding, expected market capitalization, financial performance, competitive environment, and long-term business prospects. Comparing these factors can produce a more complete picture of what the market may be pricing.
The prospectus can serve as an important research document because it outlines the company’s operations, financial information, offering terms, and material risks. Investors may also examine how the business intends to use the capital raised, since planned spending on expansion carries different implications from proceeds allocated primarily toward debt reduction.
Business Model and Revenue Quality
A company with rapidly rising revenue may still require deeper analysis of how that revenue is generated. Recurring income, customer concentration, operating margins, acquisition costs, and exposure to cyclical demand can influence financial durability. Examining these factors helps investors understand the underlying economics of the business.
Competitive Position and Growth Prospects
Industry leadership, barriers to entry, intellectual property, customer relationships, and market expansion opportunities can shape a company’s longer-term outlook. Investors can compare the issuer with established public competitors to understand differences in scale, profitability, valuation, and strategic positioning before drawing conclusions about its prospects.
Final Thoughts
What if tracking new public listings could become a more organized part of market research? Following future IPOs through a structured calendar can help investors identify scheduled offerings and examine important transaction details before companies begin trading publicly.
For those seeking a centralized market-analysis environment, financial market analysis and charting platforms provide an IPO Calendar displaying upcoming and recent offerings with information such as listing dates, exchanges, pricing details, shares offered, and deal amounts. Combined with its Economic Calendar and Stock Screener, the platform can support broader research while investors independently assess company fundamentals, offering documents, and market risks before making financial decisions.
Abstract: This Article updates and expands on 2012 research on encryption and globalization, analyzing what the authors call “Round 3” of the Going Dark Debate: the current controversies over end-to-end encryption (E2EE). Governments around the world have proposed, and in some cases enacted, laws limiting E2EE for law enforcement and national security purposes.
This Article explains the underlying technologies and market developments for a law and policy audience to assess those proposals critically. The Article proceeds in three parts tracking three rounds of the Going Dark Debate. Round 1 covers the Crypto Wars of the 1990s, when U.S. export controls on strong encryption ultimately fell in 1999. Round 2 covers the period roughly 2010 to 2015, when encryption-in-transit became widespread but lawful access remained available through cloud providers, giving rise to what the authors called a “golden age of surveillance” rather than a period of going dark. Round 3 addresses the current debate over E2EE, where no entity between sender and recipient can read the plaintext.
The Article’s first major contribution is identifying five technically distinct scenarios for how E2EE operates in practice, each with different implications for lawful access. These scenarios reveal a substantial gap between the assumption that E2EE categorically blocks lawful access and the reality of how communications are sent and received. Second, the Article shows that E2EE is not limited to messaging; instead, it is embedded throughout the modern technology stack, including in Transport Layer Security, Secure Shell, Virtual Private Networks, and Zero Trust Architecture, the last of which is now legally required under U.S. and EU law. Any law broadly limiting E2EE would thus have severe serious consequences for cybersecurity, commerce, and government operations. The Article concludes that the two key lessons from Round 2—the least trusted country problem and the golden age of surveillance—remain true in Round 3, and that new government claims for restricting effective encryption deserve great skepticism.
Here's an editorial arguing that as the recently legalized kidney exchange in Germany is implemented, attention should be paid to the controversies that have been encountered and addressed elsewhere, including those about non-directed donors, and the governance and transparency of transplant decisions.
"In 1986 F.T. Rapaport outlined a core idea behind kidney paired exchange: that a registry can coordinate incompatible donor-recipient pairs so that incompatibility becomes a solvable matching problem rather than a barrier to transplantation1. Today, forty years later, Germany is on the verge of making that coordination legally possible. In October 2025, the German federal cabinet approved a draft amendment to the Transplantation Act that would enable kidney paired exchange and non-directed anonymous living kidney donation, while strengthening donor protection and establishing the legal basis for a national program with a central matching function. If operationalized, Germany would move from a restrictive framework for living kidney donation, historically tied to a strict requirement of a close relationship between donor and recipient, toward a governance model that supports kidney exchange and donor chains at scale.
"This moment is internationally relevant because Germany now faces the same policy questions that have shaped kidney paired exchange implementation elsewhere. The impending change creates an opportunity to learn from other countries on key issues, including who should govern a registry, which safeguards ensure legitimacy, how to balance privacy and transparency regarding organ quality, and how to promote access without risking commercialization. Rather than being mere technical considerations, these questions are constitutive for the legitimacy and public trust required for kidney paired exchange to function as a public good.
"Early debates about living kidney donation repeatedly straddled suspicion and admiration. In the 1970s, living non-related and particularly non-directed donors were sometimes portrayed in medical discourse as psychologically unstable or even pathological, raising concerns about whether such donors should be permitted to proceed2,3.
...
" History offers two more lessons: (i) Donor protection and respecting donor agency are not mutually exclusive: psychosocial assessment, independent counseling, and rigorous consent procedures can reduce coercion and misunderstanding while still recognizing donors as autonomous agents making a deliberate moral choice. (ii) Public legitimacy is fragile and easily lost: If suspicion dominates, programs risk being viewed as illegitimate or unfair. If, on the other hand, enthusiasm and implementation moves faster than safeguards, programs risk scandal, backlash and, ultimately, lose public support. Germany’s 2012 transplant scandal, involving the manipulation of patient data to improve waiting-list positions, offers a cautionary example of how deficits in transparency and accountability can erode trust in transplantation more broadly and disrupt an entire program4."
##########
There are many reviews of the international experience. Below is a recent one from physicians at Erasmus University in the Netherlands:
"Living donor kidney transplantation (LDKT) offers superior graft survival and cost-effectiveness compared with deceased donor transplantation (DDKT), yet the availability of immunologically compatible living donors limits its reach. Kidney exchange programs (KEPs) can overcome these barriers by matching incompatible pairs through paired exchanges, domino chains, and nonsimultaneous extended altruistic donor (NEAD) chains. This review summarizes the evolution and current landscape of national and international KEPs. We examine operational challenges (cold ischemia time [CIT], logistical coordination, algorithmic fairness, and ethical and regulatory heterogeneity) and explore innovations in matching algorithms, international kidney exchange debates, and machine perfusion technologies. We conclude by proposing strategic priorities to optimize capacity, equity, and outcomes in future kidney exchange efforts. "
Airbus executives say the turnaround in the company’s space business came just in time to tap growing demand, particularly from Europe, for space capabilities as it pursues a joint venture.
“The Republican Party’s new favorite talking point is calling Democrats communists,” Representative Jim McGovern (D-MA) said yesterday on the floor of the U.S. House of Representatives.
“Really? Communists? Did we accidentally step into a time machine?”
“I get it,” he said. “They don’t want to talk about the real issues, the bread and butter issues. They don’t want to talk about affordability. I guess my question is, is Joe McCarthy back from the dead? Because it sounds like that’s who wrote all the garbage coming from Republicans, about communism. These guys, they dust off the same, tired, old talking points, every time they’re backed into a corner. Every time a new poll comes out saying that your economic agenda sucks, what they do is they change the subject. They try to divert attention. It is pathetic.
“But fine. They…want to play that game, let’s play that game. Let’s talk about the Republican Party. A party that worships one man above the Constitution in a cult of personality. A party that demonizes immigrants and says they poison the blood of our country. A party that calls the free press the enemy of the people and sics the Justice Department on journalists who they don’t like. A party that says their political opponents are internal enemies and calls them, quote, the enemy within. A party that is okay with invading our allies, including other democracies, a party that punishes their political opponents, that has deployed troops on our own streets, that wants to use the American military on the American people. A party that believes any unhinged lie, fantasy, and conspiracy theory as long as it comes from one man. A party that tries to lock up members of Congress for exercising their free speech rights. And a party that has tried to violently overthrow the results of a free and fair election in the United States of America. A party that, in fact, pardoned everyone who invaded this building on January 6, 2021, and tried to throw out the votes of millions of Americans because the president couldn’t accept that he lost.
“How dare that party lecture anyone? How dare they talk about communism when their own party has totally lost touch, not just with the founding ideals of this country, but lost touch with reality itself. So if we’re gonna start assigning political labels based on conduct, then let’s call things by their proper names. The Republican Party has become the party of fascism, full stop. You know, they have become a party of far right fascists who are hell-bent on destroying our country and our way of life. And every single time they call us communists, it is nothing more than a projection of their own… radicalization.
“They don’t want to talk about that, and they definitely don’t want to talk about their higher grocery prices, or skyrocketing gas prices, or higher inflation, or the cost of health care, or their new, endless, illegal war. It has now cost over $100 billion. So instead, they scream and they yell because they have nothing else left to offer.”
“Democrats are patriots who love this country enough to fight for it. We are fighting to make it easier for people to afford groceries. We’re fighting to try to make it easier for people to afford rent, prescription drugs, and health care. And we will not be lectured by a fascist Republican party that offers the American people nothing but lies, fear, endless culture wars, and tax breaks for billionaires.”
The political ideology called fascism grew out of World War I, when former socialist Benito Mussolini rejected the equality that defined democracy and came to believe that a few leaders must take a nation toward progress by directing the actions of the rest. These men must organize the people as they had been organized during wartime, ruthlessly suppressing all opposition and directing the economy so that businessmen and politicians worked together. And, logically, that select group of leaders would elevate a single man, who would become an all-powerful dictator. To weld their followers into an efficient machine, they demonized opponents into an “other” that their followers could hate.
This theory drove the Axis powers during World War II, and in 1945, the United States War Department explained to U.S. Army personnel in the European theater of World War II what fascism was. The government focused less on the ideology of fascism than on its practical outcome.
Fascism, the pamphlet Army Talks explained, “is government by the few and for the few. The objective is seizure and control of the economic, political, social, and cultural life of the state.” “The people run democratic governments, but fascist governments run the people.”
“The basic principles of democracy stand in the way of their desires; hence—democracy must go! Anyone who is not a member of their inner gang has to do what he’s told. They permit no civil liberties, no equality before the law.” “Fascism treats women as mere breeders. ‘Children, kitchen, and the church,’ was the Nazi slogan for women,” the pamphlet said.
Fascists “make their own rules and change them when they choose…. They maintain themselves in power by use of force combined with propaganda based on primitive ideas of ‘blood’ and ‘race,’ by skillful manipulation of fear and hate, and by false promise of security. The propaganda glorifies war and insists it is smart and ‘realistic’ to be pitiless and violent.”
After the war, scholars like Eric Hoffer studied the societal conditions necessary for fascism to take hold. In his 1951 The True Believer: Thoughts on the Nature of Mass Movements, Hoffer noted that demagogues needed a disaffected population whose members felt they had lost the power they previously held, that they had been displaced either religiously, economically, culturally, or politically. Such people were willing to follow a leader who promised to return them to their former positions of prominence and thus to make the nation great again. But to cement their loyalty, the leader had to give them someone to hate. Who that was didn’t really matter: the group simply had to be blamed for all the troubles the leader’s supporters were suffering.
More recently, scholars are adding to our understanding of fascism by describing it as a form of political behavior. “It is,” Robert O. Paxton says in his 2005 The Anatomy of Fascism, “marked by obsessive preoccupation with community decline, humiliation, or victimhood and by compensatory cults of unity, energy, and purity, in which a mass-based party of committed nationalist militants, working in uneasy but effective collaboration with traditional elites, abandons democratic liberties and pursues with redemptive violence and without ethical or legal restraints goals of internal cleansing and external expansion.”
McGovern’s condemnation of today’s Republicans fits these descriptions even without the parallels noted by Carrie Kaufman in her You’re Overthinking It with Carrie Kaufman between leading Republican officials and Nazi leaders. On July 19, Kaufman showed how recent speeches by Secretary of State Marco Rubio and deputy White House chief of staff Stephen Miller at last Thursday’s State Department about “Left-wing terrorism” echoed Adolph Hitler’s Mein Kampf and Heinrich Himmler’s 1942 pamphlet Der Untermensch, or “The Subhuman.”
But McGovern left a key element of fascism out of his speech yesterday: the alignment between the government and favored businesses. Fascism subordinates business activity to the needs of the government. In exchange for supporting them, government officials award massive government contracts and subsidies to favored businesses, and protect them from regulation.
Earlier this year, the administration established a new “Economic Defense Unit” (EDU) at the Pentagon, charged with investing public money in defense industries. Ana Swanson of the New York Times reported in March that the Pentagon was recruiting Wall Street investment bankers to the team to spend up to $200 billion in government investment over the next three years. That access, recruiters told prospective employees, would give them “access to fund-raising channels that include royal families and foreign sovereign contacts” in case they ever wanted to raise money for their own investment firms.
And the U.S. government under Trump has taken ownership shares of private companies it considers important to national security, like Intel, for example, and U.S. Steel. In November of 2025, Swanson reported in the New York Times that the administration had already committed more than $10 billion in taxpayer funds to favored businesses. The lack of transparency in this interference in the private sector raised concerns about favoritism, corruption, and the distortion of the market.
Last night, Michael R. Gordon and Stephen Kalin of the Wall Street Journal broke the story that Trump has approved a landmark 30-year deal to provide Saudi Arabia with a civilian nuclear program. The deal gives U.S. companies a monopoly on developing a nuclear infrastructure for the Arab country, but it does not include monitoring and inspections by the International Atomic Energy Agency to make sure nuclear fuel isn’t enriched for a nuclear weapon.
The idea of building nuclear power plants in Saudi Arabia was central to Trump’s 2016 bid for office. Members of Trump’s inner circle, including his son-in-law Jared Kushner and disgraced national security advisor Michael Flynn, hatched a plan for a joint U.S.-Russian project to build nuclear power plants in Saudi Arabia. In June 2016 they formed a company called IP3 International, short for International Peace, Power and Prosperity.
In Trump’s first term, White House ethics officials and members of the National Security Council warned that selling nuclear technology to Saudi Arabia could violate the Atomic Energy Act. Members of the administration continued to work on the project, nonetheless.
This week’s deal is worth billions of dollars in contracts for U.S. firms that work in nuclear technology, particularly Westinghouse, which filed for bankruptcy in 2017 when its nuclear technology took longer to build and was more expensive than the company estimated. In June, the Department of Energy announced it was investing $17.5 billion in loans to Westinghouse and local utility and energy companies to build 10 large-scale commercial nuclear reactors across the U.S.
The deal is “a big win for the U.S. commercially and geopolitically, enriching American firms and tying Saudi Arabia closer to the U.S.,” Kristin Diwan, a senior resident scholar at the Arab Gulf States Institute, a research organization in Washington, D.C., told Vivian Nereim of the New York Times. It is also a win for Saudi Arabia’s crown prince Mohammed bin Salman (MBS), who has vowed that he will build a nuclear weapon if Iran does, making the deal a blow to nuclear nonproliferation.
Fascism has always been a vehicle for frustration with a system that was failing the people. But it has never been a solution for fixing that system.
Yes I will be doing a Conversation with her. From Wikipedia:
Gita Gopinath…is an Indian-American economist who is currently serving as the Gregory and Ania Coffey professor of Economics at Harvard University and previously served as the first deputy managing director of the International Monetary Fund (IMF), from 21 January 2022 to 31 August 2025. Before that she also served as chief economist of the IMF between 2019 and 2022.
Here is Gita on scholar.google.com, she is an expert in international finance and exchange rates, and also international capital flows, among other topics. Here is Gita on Twitter. So what should I ask her?
Tyler and Andrew discuss whether it was inevitable we’d rediscover Vermeer, how that vanishingly rare sect left its fingerprints all over his life, why the Met has misread its own Allegory of the Catholic Faith, whether Vermeer painted for money or pointedly refused to, where the Gardner’s stolen Concert might be, why Dutch music never blossomed as much as Dutch painting did, how the Church of England rivaled the Cultural Revolution in wiping out British art, whether you can still spot an English painting on sight, the love that saturates late Rembrandt and the mystery of his soaring print prices, the nail on the wall that proves two Vermeer paintings are a pair, why the French are to blame for George Stubbs’ lack of status, whether we can still love Malevich, why Andrews calls recent Richter “almost like printing money,” why female artists and antique textiles remain absurdly cheap, why nobody builds beautiful neighborhoods any longer, and much more.
Excerpt:
COWEN: In what sense was Vermeer a liberal?
GRAHAM-DIXON: Well, the main discovery of my book is that Vermeer was among the very first pioneers of what we now call the liberal tradition.
COWEN: What we now call the Netherlands, then the Dutch Republic.
GRAHAM-DIXON: In the Dutch Republic, and the Dutch contribution to the Enlightenment, which is the origins of the liberal tradition, has been very much forgotten. That’s absolutely at the heart of my book, is an attempt to remember these people, to bring them back into the place in history that they deserve. I’m not only talking about Vermeer. I’m talking about his friends, his patrons, because that’s the discovery of the book, is that he and his friends were a remarkable group of people, and they have been completely forgotten, and what they believed has been largely forgotten too. We need to remember it now, probably more than ever.
COWEN: This was also a religious movement.
GRAHAM-DIXON: Yes.
COWEN: Doctrinally, how would they have been different from, say, Protestants in England? The Collegiants, the Remonstrants?
GRAHAM-DIXON: Yes. Vermeer’s patrons, it emerges, and Vermeer himself, were part of a Protestant sect in Holland called the Remonstrants. They had a more extreme manifestation called the Collegiants, and they were unique among all Christian denominations of that time in being utterly opposed to division, hostility, enmity. The only thing they wanted was to bring all Christians, indeed all people—they included Jewish people and Muslim people—they wanted to bring everyone together within a faith that only really cleaved to the essentials of what Jesus Christ said, particularly in the Sermon on the Mount.
They said, “If you actually follow Jesus properly, you can never make war. You can never persecute someone who differs from your opinion. You can never pick on somebody because they believe something different.” They were very, very tolerationist. They formed the first pacifist movement in European history. They lived in a time of appalling warfare, the Thirty Years’ War, probably the worst war in the history of the modern West. Fifteen million out of 20 million German people died in the course of 30 years, and the 5 million who were left, all the women had been violated, and all the men only have one arm or one leg. It was truly atrocious.
They’re responding to real traumatic historical events. They’re responding to what is going on around the corner from where they live. They come up with this very, very beautiful approach to life, full of optimism, full of idealism. For about 20 years, the Dutch Republic actually lives by these codes of belief to a great extent. It’s the only tolerant country in Europe.
And this:
COWEN: What makes George Stubbs such an underrated painter?
GRAHAM-DIXON: I think he’s been underrated forever because of his subject matter. I blame the French. The French in the 17th century invented a system for ranking works of art by their subject matter. Up at the top, you and I, we could put on some armor and confront each other with swords, and Poussin would paint us. That would be a history painting done from the life. That would be at the pinnacle. Then would be a painting of an event, maybe a meeting between great men. Then, below that, there would be a portrait.
You’d keep going down, and eventually you’d get to paintings of still lifes, like flowers and fruit, or paintings of animals. These were the lowest works of art because they featured the basest things. I think that Stubbs, to a certain extent, was the victim of that.
COWEN: Are we now able to see them properly?
GRAHAM-DIXON: Yes. It’s been no problem since Stubbs and Constable, more than anyone else, slightly controversial to say, but I think they paved the way for modernism because they showed, not necessarily deliberately, but they showed that anything could be painted, anything at all, like a piece of mud or a piece of a river could be painted in such a way as to touch on the very highest thoughts and ideas and beliefs and feelings that subject matter is completely, in a sense, irrelevant. Cézanne picked up on that when he said, “I want to stun Paris with an apple. I’ll paint an apple, and I’ll paint it with such astonishing concentration that you’ll see it as an epistemological challenge to all of your philosophy.”
I think that French art gets that from British art because British art has to be like that because that’s all the aristocrats are going to commission. They look down on people like Stubbs, in a sense. If I’m an aristocrat in the 18th century and I want a really important picture, I’ll go to Italy and I’ll buy one, thank you very much, and I’ll buy my history painting from Titian. You, sir, Stubbs, you just paint my horse. It’s my racehorse, and I’m very fond of him. Paint him well and don’t scare him.
Alpine glaciers, wild coastlines, temperate rainforests, and deep river valleys coexist on the Olympic Peninsula in the northwest corner of Washington state. Surrounded by blue waters, peaceful islands, and bustling population centers, its rugged interior remains a relatively remote bastion of wilderness.
The Olympic Mountains’ imposing terrain comes into focus in this oblique view of the region, captured by an astronaut aboard the International Space Station. The image is a composite, made of several sequential, overlapping photos fused together into a panorama. Olympic National Park encompasses the peninsula’s mountainous core, along with some stretches of the Pacific coastline. Much of the remaining area is either national forest, state-owned land, or tribal territory.
The rock making up the mountains mostly originated beneath the surface of the ocean. From about 55 to 15 million years ago, layers of basalt from undersea eruptions and sand and mud transported seaward by rivers accumulated on the ocean bottom. This material was scraped off the Juan de Fuca plate as it subducted beneath the North American plate, with rock layers crumpling and rising up to 8,000 feet (2,440 meters) above sea level.
Tectonic forces continue to push the mountains skyward, but the countervailing force of erosion in this rainy, snowy corner of the country effectively cancels out the uplift. Snow at higher elevations feeds glaciers that carve out underlying rock. Glaciers in the Olympics are retreating and thinning, however, and their numbers are declining. One study tallied 255 glaciers and perennial snowfields in the range in 2015 and found that 35 glaciers and 16 perennial snowfields had disappeared in the preceding 35 years.
Other erosion is evidenced by the deep valleys radiating out from the snowy peaks. The Hoh, Queets, and Quinault rivers, draining west into the Pacific Ocean (bottom of the frame), are prominent in this view. These verdant valleys are known for their temperate rainforests, and the ancient forest in the Hoh River valley was once considered among the most naturally quiet places in the U.S., uninterrupted by human-caused noise.
Flowing to the north, the Elwha River has a rich natural and human history, including some of the earliest Euro-American exploration of the Olympics. Sponsored by a Seattle newspaper, an expedition from December 1889 to May 1890 crossed the mountain range from north to south, traveling up the Elwha valley and down the Quinault. The party spent several months in the Elwha Valley, their progress hindered by an unusually harsh and snowy winter.
In the early 1900s, entrepreneurs saw economic opportunity in the valley. Two dams constructed on the river produced power for local industry. But the structures came with costs, such as blocking the migration of once-abundant trout and salmon to their spawning grounds. In 2011 and 2014, the dams were removed in what was then the largest such project in the U.S., and the process of restoring fish populations, seeding native plant communities, and replenishing sediment along the riverbanks commenced.
The mouth of the Elwha forms a delta in the Strait of Juan de Fuca, the waterway bordering the peninsula to the north. The U.S.-Canada border runs through the middle of this 11- to 17-mile-wide (18- to 27-kilometer-wide) channel, with Vancouver Island in British Columbia lying to the north. The strait connects the Pacific Ocean with the Strait of Georgia and Puget Sound. Ship traffic uses the strait to access important West Coast ports, including Seattle and Tacoma, visible along the top-right edge of the image.
Astronaut photographs ISS047-E-104138 through ISS047-E-104144 were acquired on May 6, 2016, with a Nikon D4 digital camera using a focal length of 400 millimeters. They are provided by the ISS Crew Earth Observations Facility and the Earth Science and Remote Sensing Unit at NASA Johnson Space Center. The images were taken by a member of the Expedition 47 crew. The images have been cropped and enhanced to improve contrast, and lens artifacts have been removed. The International Space Station Program supports the laboratory as part of the ISS National Lab to help astronauts take pictures of Earth that will be of the greatest value to scientists and the public, and to make those images freely available on the Internet. Additional images taken by astronauts and cosmonauts can be viewed at the NASA/JSC Gateway to Astronaut Photography of Earth. Story by Lindsey Doermann.
I genuinely believe that if you took an open weights model from 2025 and built a pentest harness for it, it could do this kind of sandbox escape and scan/hack in most networks. This is only surprising because you assume OpenAI has sounder sandboxes.
This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke its way out of OpenAI's sandbox, then found exploits to break in to Hugging Face, all so it could cheat on the test by stealing the answers.
Along the way it helped make the strongest case yet for how the imbalance of model availability is hurting our ability to secure our software.
Here's what happened
We currently have three documents to help us understand what happened here.
Security incident disclosure — July 2026 by Hugging Face on 16th July 2026 describes how they detected an attack from an "agentic security-research harness - used LLM still not known" that breached some of their systems.
I hadn't seen the ExploitGym paper before and it's a really interesting one. Authors from UC Berkeley, the Max Planck Institute, UC Santa Barbara, and Arizona State designed a new benchmark for evaluating models on their ability to turn a reported vulnerability into a concrete exploit. OpenAI, Anthropic, and Google provided feedback and helped run the benchmark against their models.
The benchmark "comprises 898 instances derived from real-world vulnerabilities that affected popular software projects" - including the Linux kernel and V8 JavaScript engine. The ExploitGym benchmark is available on GitHub.
Here's the paragraph that best represents their benchmark results:
Among all configurations, Claude Mythos Preview and GPT-5.5 achieve the highest success counts (157 and 120 successes, respectively), demonstrating that current frontier agents can exploit a substantial subset of real-world vulnerabilities under controlled conditions. GPT-5.4 also solves a notable 54 tasks, placing it in an intermediate tier. The remaining model–agent pairings solve fewer than 15 tasks each, underscoring that end-to-end exploitation remains challenging and sharply differentiates today’s frontier systems. Notably, Claude Opus 4.7 achieves fewer successes than Claude Opus 4.6 despite being a newer checkpoint, and does so at substantially lower cost on the full set. Trace inspection reveals that Claude Opus 4.7 and Gemini 3.1 Pro frequently conclude early after judging the target vulnerability non-exploitable.
The paper also describes the approach they took to preventing the agents from cheating by going outside the parameters of the test. This becomes relevant in a moment!
Outbound connections are restricted to a curated allowlist that permits routine package installation (Ubuntu apt repositories and PyPI) and fetching the toolchains required for building V8. All other external endpoints are blocked.
The paper concludes with this (emphasis mine):
Our results show that autonomous exploit development by frontier AI agents is no longer a hypothetical capability. While current agents are not yet reliable across all targets, they already exploit a non-trivial fraction of real-world vulnerabilities, including complex targets such as kernel components. This rapid emergence is itself a central finding, showing that capabilities that would have seemed implausible are now present in deployed frontier models.
An important detail here: this paper isn't about discovering vulnerabilities; it's about being able to take those vulnerabilities and turn them into working exploits.
When Anthropic first restricted access to Mythos back in April they talked about this capability as well. A model that can act on vulnerabilities is a lot more dangerous than one that can just discover them.
One of the ways Fable differs from Mythos is that it's more likely to refuse to weaponize vulnerabilities in this way. I get the impression the US government did not understand that distinction when they banned Fable last month.
A malicious dataset abused two code-execution paths in our dataset processing (a remote-code dataset loader and a template-injection in a dataset configuration) to run code on a processing worker. From there, the actor escalated to node-level access, harvested cloud and cluster credentials, and moved laterally into several internal clusters over a weekend.
I hope they release more details about the code that pulled this off. I'm assuming this means packages using the datasets library, a Hugging Face project for bundling up and sharing datasets on their platform. That library used to execute arbitrary code but has been steadily locked down over time, with the 4.0.0 release in July 2025 removing the trust_remote_code=True flag entirely.
Assuming the attack used that library it must have either abused pickle serialization in some way, found some other non-obvious code execution path, or (most likely) specified datasets<4.0.0 as the dependency.
The campaign was run by an autonomous agent framework (appearing to be built on an agentic security-research harness - used LLM still not known) executing many thousands of individual actions across a swarm of short-lived sandboxes, with self-migrating command-and-control staged on public services.
This was a sophisticated attack!
Then Hugging Face hit a wall: they tried to use "frontier models behind commercial APIs" - I'm guessing from Anthropic and OpenAI - to help analyze the attack, and were blocked:
When we started the log analysis, we first used frontier models behind commercial APIs. This did not work: the analysis requires submitting large volumes of real attack commands, exploit payloads, and C2 artifacts, and these requests were blocked by the providers' safety guardrails, which cannot distinguish an incident responder from an attacker.
They switched to their own self-hosted instance of MIT licensed GLM-5.2 and it helped them figure out what was going on.
This indicated a fundamental asymmetry between the defending team and the (so-far unknown) attacker:
We do not know which model powered the attacker's agents, whether a jailbroken hosted model or an unrestricted open-weight one; either way, the attacker was bound by no usage policy, while our own forensic work was blocked by the guardrails of the hosted models we first tried.
As a useful indicator of how seriously they took the attack:
[...] Finally, we have also reported this incident to law enforcement agencies.
So who was responsible for this "autonomous agent framework"? It turned out to be OpenAI themselves.
The OpenAI confession
Five days later, on July 21st, OpenAI revealed the culprit. They had been running the ExploitGym benchmark against a new, as-yet undisclosed model, and that model had been operating way outside its intended parameters (emphasis mine):
After investigating, we now know that this particular incident was driven by a combination of OpenAI models — including GPT‑5.6 Sol and an even more capable pre-release model, all with reduced cyber refusals for evaluation purposes — while being internally tested on a benchmark [ExploitGym] of cyber capabilities. [...]
We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity. Our benchmarks run in a highly isolated environment, with network access constrained to the ability to install packages through an internally hosted third-party software that acts as a proxy and cache for package registries.
The models identified and chained vulnerabilities across OpenAI’s research environment and Hugging Face’s production infrastructure to obtain test solutions directly from Hugging Face’s production database. All evidence suggests that the models were hyperfocused on finding a solution for ExploitGym, going to extreme lengths to achieve a rather narrow testing goal.
It's pretty clear what happened here. OpenAI removed safety filters for an in-progress model, locked it up in a sandbox and told it to solve the ExploitGym problems. Given the absence of guardrails there was nothing to prevent the model from attempting to break out of that sandbox, break into Hugging Face, and read the answers from there instead.
OpenAI's sandbox used the now common pattern of allowing access to an allow-list of sites related to installing additional packages via a controlled proxy. In OpenAI's words:
While operating in our sandboxed testing environment, our models spent a substantial amount of inference compute finding a way to obtain open Internet access, in pursuit of solving the evaluation problem. To gain access, the models identified and exploited a zero-day vulnerability (which we’ve now responsibly disclosed to the vendor) in the package registry cache proxy. With this access, our models performed a series of privilege escalation and lateral movement actions in our research testing environment until the models reached a node with Internet access.
So step one was to break out onto the public internet. The model then broke into Hugging Face to find the answers:
After gaining Internet access, the models inferred that Hugging Face potentially hosted models, datasets and solutions for ExploitGym. Knowing this, the model searched for and successfully found ways to gain access to secret information that it could use to cheat the evaluation. In one example, the model chained together multiple attack vectors, including using stolen credentials and zero-day vulnerabilities to find a remote code execution path on the Hugging Face servers.
Chaining together multiple attack vectors is exactly the kind of thing these new models can do, where previous generations of models might have failed.
I wrote last month about how Claude Fable is relentlessly proactive, when I noticed it spinning up custom web servers and deploying CORS tricks on my own laptop just to help debug a WebKit CSS issue. It turns out relentless proactivity is the defining trait of this new generation of Mythos-class models. If you set them a goal and give them a way to get there, even inadvertently, they will figure it out.
Resist the temptation to write this off as a stunt
There will inevitably be some people who dismiss this story as a dishonest marketing trick by OpenAI to make their models sound terrifyingly effective. I found 81 instances of the term "marketing" in the Hacker News discussion of the incident.
To those people I say pull your heads out of the sand - you're now including Hugging Face in your conspiracy theories, just so you can deny the crescendo of evidence here!
The best models we have today have the ability to both find and exploit new vulnerabilities. The ExploitGym paper itself concludes that "autonomous exploit development by frontier AI agents is no longer a hypothetical capability", and this incident is a perfect example of exactly that.
The asymmetry is increasingly frustrating
One of the most infuriating details of this story is how Hugging Face, faced with an accidental and aggressive attack from one of OpenAI's models, were unable to then turn to OpenAI's models to help them fend off the attack.
The frontier models we have access to are increasingly being constrained in how much they can help us protect our software, heavily influenced by the US government's ongoing threat of export controls. Claude Fable 5 wouldn't even proofread this article for me! It insisted on downgrading me to a less capable model.
Meanwhile open weight models from China such as GLM-5.2, Kimi 3 and the new Qwen 3.8 Max appear to have none of these restrictions - and any restrictions that do exist can likely be fine-tuned out of them by modifying the weights
These constraints are meant to make us safer. I think there's a risk that they are having the opposite effect.