Thursday assorted links

1. Puffin spotted on the Dorset coast.

2. Anthropic chief economist on AI and unemployment.

3. Whales use different vowels when ships are around.

4. Oliver Kim on McNamara.

5. Large language models can predict the results of social science experiments.

6. Alexander Salter Substack on space economics.

7. New Google data and study on AI and the economy.

8. Does it matter if you don’t like the characters?

The post Thursday assorted links appeared first on Marginal REVOLUTION.

       

Comments

 


Eastern North Pacific 2-Day Graphical Outlook Image
Eastern North Pacific 7-Day Graphical Outlook Image






Atlantic 2-Day Graphical Outlook Image
Atlantic 7-Day Graphical Outlook Image





China’s Shijian-31 satellite is sweeping GEO using a unique orbit

HELSINKI — A classified satellite launched by China in June is operating in a creative orbit, giving it a unique vantage point over the geostationary belt. Shijian-31 launched on a […]

The post China’s Shijian-31 satellite is sweeping GEO using a unique orbit appeared first on SpaceNews.

FCC approves sweeping space reforms while clearing more C-band for 5G

The FCC voted July 22 to push satellite operators out of more C-band spectrum for a 5G auction next year, while approving a licensing overhaul aimed at accelerating the commercial space economy.

The post FCC approves sweeping space reforms while clearing more C-band for 5G appeared first on SpaceNews.

Vast establishes division for national security projects

Melanie Stricklan

Commercial space station and spacecraft developer Vast has established a new division intended to support work in national security space.

The post Vast establishes division for national security projects appeared first on SpaceNews.

The strategic competition with China: winning the multi-domain race

A Long March 6A rocket lifts off from the Taiyuan Satellite Launch Center in northern China, producing a bright plume of exhaust against a clear sky, with four lightning towers surrounding the launch pad.

History demonstrates military innovation continually rewrites the rules of combat, reshaping warfare and rendering strategies from past conflicts obsolete. The introduction of repeating rifles transformed infantry combat during the 19th […]

The post The strategic competition with China: winning the multi-domain race appeared first on SpaceNews.

Kidney exchange in Germany can draw on the experience in other countries

 Here's an editorial arguing that as the recently legalized kidney exchange in Germany is implemented, attention should be paid to the controversies that have been encountered and addressed elsewhere, including those about non-directed donors, and the governance and transparency of transplant decisions.

 Benning L, Lu W, Budde K , Josephson, M.  Germany’s move toward kidney paired exchange and non-directed donation: governance lessons from abroad, Kidney International Reports, 2026 

"In 1986 F.T. Rapaport outlined a core idea behind kidney paired exchange: that a registry can coordinate incompatible donor-recipient pairs so that incompatibility becomes a solvable matching problem rather than a barrier to transplantation1. Today, forty years later, Germany is on the verge of making that coordination legally possible. In October 2025, the German federal cabinet approved a draft amendment to the Transplantation Act that would enable kidney paired exchange and non-directed anonymous living kidney donation, while strengthening donor protection and establishing the legal basis for a national program with a central matching function. If operationalized, Germany would move from a restrictive framework for living kidney donation, historically tied to a strict requirement of a close relationship between donor and recipient, toward a governance model that supports kidney exchange and donor chains at scale.


"This moment is internationally relevant because Germany now faces the same policy questions that have shaped kidney paired exchange implementation elsewhere. The impending change creates an opportunity to learn from other countries on key issues, including who should govern a registry, which safeguards ensure legitimacy, how to balance privacy and transparency regarding organ quality, and how to promote access without risking commercialization. Rather than being mere technical considerations, these questions are constitutive for the legitimacy and public trust required for kidney paired exchange to function as a public good. 

"Early debates about living kidney donation repeatedly straddled suspicion and admiration. In the 1970s, living non-related and particularly non-directed donors were sometimes portrayed in medical discourse as psychologically unstable or even pathological, raising concerns about whether such donors should be permitted to proceed2,3. 

...

" History offers two more lessons: (i) Donor protection and respecting donor agency are not mutually exclusive: psychosocial assessment, independent counseling, and rigorous consent procedures can reduce coercion and misunderstanding while still recognizing donors as autonomous agents making a deliberate moral choice. (ii) Public legitimacy is fragile and easily lost: If suspicion dominates, programs risk being viewed as illegitimate or unfair. If, on the other hand, enthusiasm and implementation moves faster than safeguards, programs risk scandal, backlash and, ultimately, lose public support. Germany’s 2012 transplant scandal, involving the manipulation of patient data to improve waiting-list positions, offers a cautionary example of how deficits in transparency and accountability can erode trust in transplantation more broadly and disrupt an entire program4."

##########

 There are many reviews of the international experience.  Below is a recent one from physicians at Erasmus University in the Netherlands:


Stijn C. van de Laar, Hidde A. de Heus, Annelies E. de Weerd, Matthijs F. Klaassen, Robert J. Porte, Robert C. Minnee Frank J.M.F. Dor, A Comprehensive Review of Strategies to Advance Kidney Exchange Programs and Optimize Effectiveness and Outcomes, Kidney International Reports, Volume 11, Issue 7, July 2026, 106539

"Living donor kidney transplantation (LDKT) offers superior graft survival and cost-effectiveness compared with deceased donor transplantation (DDKT), yet the availability of immunologically compatible living donors limits its reach. Kidney exchange programs (KEPs) can overcome these barriers by matching incompatible pairs through paired exchanges, domino chains, and nonsimultaneous extended altruistic donor (NEAD) chains. This review summarizes the evolution and current landscape of national and international KEPs. We examine operational challenges (cold ischemia time [CIT], logistical coordination, algorithmic fairness, and ethical and regulatory heterogeneity) and explore innovations in matching algorithms, international kidney exchange debates, and machine perfusion technologies. We conclude by proposing strategic priorities to optimize capacity, equity, and outcomes in future kidney exchange efforts. " 

####### 

Earlier: 

Friday, March 27, 2026 Germany legalizes kidney exchange !!

Airbus space business takes advantage of growing demand for satellite systems

Aeolus-2

Airbus executives say the turnaround in the company’s space business came just in time to tap growing demand, particularly from Europe, for space capabilities as it pursues a joint venture.

The post Airbus space business takes advantage of growing demand for satellite systems appeared first on SpaceNews.

Alex Tabarrok on the Economic Analysis of Crime

I tell my Gary Becker story, why I like police more than prisons, how criminals are like children and more.

The post Alex Tabarrok on the Economic Analysis of Crime appeared first on Marginal REVOLUTION.

       

Comments

Related Stories

 

Is Iran Really a ‘Forever War’?

With the United States and Iran now clearly back to “war” from “ceasefire,” we’re seeing a host of stories arguing that President Trump, despite all his promises, has found his own “forever war.” It’s a tempting claim, especially for Trump’s critics. But we shouldn’t jump too easily into it without recognizing how profoundly different the situations are. For the United States, Iraq and Afghanistan were fundamentally occupations. In both cases, existing governments were rapidly shattered or melted away. The U.S. began occupations in which it stood up new, friendly governments which it hoped would eventually be able to stand on their own with the pro-U.S. friendliness intact. The U.S. eventually tired of the conflicts and withdrew its forces. Mostly.

The present situation is entirely different. The Iranian government is wholly intact and continues to fight the U.S. in an asymmetric but still conventional manner. Early U.S. hopes that the shock of the conflict would lead the Iranian government to crumble were quickly dashed. The U.S. has made no effort overthrow the Iranian government by force or even to occupy any of its territory. Even now, with the two sides back in active conflict, there is little evidence that the new attacks have any strategic goal beyond inflicting pain to shape some future negotiation. Both sides appear to be inflicting damage and pain with the hope that the other side will relent, soften their negotiating position or give up, but with little evidence either will.

The basic situation does not appear to have changed in five months. The U.S. started the conflict thinking the Iranian government would either rapidly fall or quickly submit. But it didn’t. And critically, Iran quickly demonstrated a more powerful deterrent against the U.S. than the U.S. had against Iran: controlling or at least menacing the Strait of Hormuz. Trump had ruled out in advance the kinds of very costly and dangerous means it would take to force the matter: an actual all-out invasion and occupation of Iran which would replace the government itself. With that off the table, the U.S. was essentially out of options.

The U.S. lost the conflict in its first hours or days. But Trump wasn’t willing to accept that reality. So we’ve been in a standoff ever since where Trump tries to finesse the matter or come up with some process that hides this basic fact — ceasefires to find a way to the negotiating table, memoranda of understanding which both sides can agree to because they don’t actually say anything. Over time, these efforts to freeze the conflict break down because Trump’s having lost the battle so comprehensively breaks through the subterfuge. So he starts hurling missiles again. Or Iran attacks or stops a few ships. Things break. People are killed. But it doesn’t change the underlying dynamic. Trump got nothing. Iran dictates the situation because they have a hold on the world economy by their ability to control oil shipments.

The earlier conflicts were ones in which the U.S. could mostly sustain its friendly government by remaining in place but couldn’t leave because doing so would mean it would all fall apart. They were permanent occupations, sustainable so long as the U.S. never left. There’s simply no analog to this in the current situation. If anything, continuing the conflict keeps the global economy under severe strain.

It’s certainly possible that the two sides will eventually find some way to get out of active conflict. No conflict actually lasts forever. Despite its strong hand, Iran has seen a lot of its industrial infrastructure destroyed. So it’s not like it’s suffering no pain or harm. But to the extent the conflict continues, it’s probably more apt to call it a “forever negotiation” — albeit with a lot of bombing and drones and airstrikes — than a “forever war.” Because the military verdict is in and has been since early March at the latest. The U.S. couldn’t defeat Iran through the kinds of military force it was willing to use, and Iran could dictate U.S. actions and those of much of the world by using the lever of the Strait of Hormuz. Just as important, the U.S. is on a much tighter timelines to end things than Iran. High oil prices and the impending midterm elections are a situation Trump can’t really abide. They were able to call his bluff rather than vice versa.

Everything we’ve seen since early March is defined by those parameters. Trump lost the conflict in those first days of March. Everything since is watching to see what he’ll do about it or when he’ll admit that’s what happened.

Only a Handful of Tickets Left

We’re one week out from our upcoming event in Brooklyn, and there are only a handful of tickets left. On Wednesday July 29, independent journalist and The Handbasket founder Marisa Kabas will join us at Crystal Lake bar in Brooklyn for an evening of conversation, drinks and trivia.

Marisa will talk midterms, Trump II and the state of independent media with TPM editor in chief and founder Josh Marshall; Nicole LaFond and I will lob some politics trivia questions at you; and TPM’s New York staff will be around to chat during our happy hour.

It’s only $25. Snag your ticket while you can!

All the Queens houses

Colourful terraced houses with distinct architectural styles, featuring a car parked in front.

To the trained eye, the personalised homes of Queens, NYC, evoke everything from a Venetian streetscape to a Chinese temple

- by Aeon Video

Watch on Aeon

The savage bites back

Painting of a abstracted figure resting their head on their hand, a sun, and cactus on a blue background.

Rooted in Brazilian modernism, anthropophagy devours and transforms culture, subverting colonial fears of cannibalism

- by Sofia Cándano

Read on Aeon

Dramatic Cuts

Read more

Shirking Responsibility

Next Space Force chief throws cold water on the idea of space privateers

In the early years of the United States, when the nascent US Navy was still getting its sea legs, several presidents used privateers to capture or destroy enemy warships when armed naval vessels were unable to do so.

President John Adams was one of the most vigorous proponents of commissioning private vessels for national ends. His administration issued letters of marque and reprisal during the so-called "Quasi-War" with France in the final years of the 18th century. These letters created the legal distinction between privateering and piracy.

One of the letters signed by Adams, dated November 1799, authorized the use of a merchant ship to "subdue, seize, and take any armed French vessel" found near US coastal waters of "elsewhere on the high seas." France also routinely used privateers against US shipping at the time.

Read full article

Comments

July 22, 2026

“The Republican Party’s new favorite talking point is calling Democrats communists,” Representative Jim McGovern (D-MA) said yesterday on the floor of the U.S. House of Representatives.

“Really? Communists? Did we accidentally step into a time machine?”

“I get it,” he said. “They don’t want to talk about the real issues, the bread and butter issues. They don’t want to talk about affordability. I guess my question is, is Joe McCarthy back from the dead? Because it sounds like that’s who wrote all the garbage coming from Republicans, about communism. These guys, they dust off the same, tired, old talking points, every time they’re backed into a corner. Every time a new poll comes out saying that your economic agenda sucks, what they do is they change the subject. They try to divert attention. It is pathetic.

“But fine. They…want to play that game, let’s play that game. Let’s talk about the Republican Party. A party that worships one man above the Constitution in a cult of personality. A party that demonizes immigrants and says they poison the blood of our country. A party that calls the free press the enemy of the people and sics the Justice Department on journalists who they don’t like. A party that says their political opponents are internal enemies and calls them, quote, the enemy within. A party that is okay with invading our allies, including other democracies, a party that punishes their political opponents, that has deployed troops on our own streets, that wants to use the American military on the American people. A party that believes any unhinged lie, fantasy, and conspiracy theory as long as it comes from one man. A party that tries to lock up members of Congress for exercising their free speech rights. And a party that has tried to violently overthrow the results of a free and fair election in the United States of America. A party that, in fact, pardoned everyone who invaded this building on January 6, 2021, and tried to throw out the votes of millions of Americans because the president couldn’t accept that he lost.

“How dare that party lecture anyone? How dare they talk about communism when their own party has totally lost touch, not just with the founding ideals of this country, but lost touch with reality itself. So if we’re gonna start assigning political labels based on conduct, then let’s call things by their proper names. The Republican Party has become the party of fascism, full stop. You know, they have become a party of far right fascists who are hell-bent on destroying our country and our way of life. And every single time they call us communists, it is nothing more than a projection of their own… radicalization.

“They don’t want to talk about that, and they definitely don’t want to talk about their higher grocery prices, or skyrocketing gas prices, or higher inflation, or the cost of health care, or their new, endless, illegal war. It has now cost over $100 billion. So instead, they scream and they yell because they have nothing else left to offer.”

“Democrats are patriots who love this country enough to fight for it. We are fighting to make it easier for people to afford groceries. We’re fighting to try to make it easier for people to afford rent, prescription drugs, and health care. And we will not be lectured by a fascist Republican party that offers the American people nothing but lies, fear, endless culture wars, and tax breaks for billionaires.”

The political ideology called fascism grew out of World War I, when former socialist Benito Mussolini rejected the equality that defined democracy and came to believe that a few leaders must take a nation toward progress by directing the actions of the rest. These men must organize the people as they had been organized during wartime, ruthlessly suppressing all opposition and directing the economy so that businessmen and politicians worked together. And, logically, that select group of leaders would elevate a single man, who would become an all-powerful dictator. To weld their followers into an efficient machine, they demonized opponents into an “other” that their followers could hate.

This theory drove the Axis powers during World War II, and in 1945, the United States War Department explained to U.S. Army personnel in the European theater of World War II what fascism was. The government focused less on the ideology of fascism than on its practical outcome.

Fascism, the pamphlet Army Talks explained, “is government by the few and for the few. The objective is seizure and control of the economic, political, social, and cultural life of the state.” “The people run democratic governments, but fascist governments run the people.”

“The basic principles of democracy stand in the way of their desires; hence—democracy must go! Anyone who is not a member of their inner gang has to do what he’s told. They permit no civil liberties, no equality before the law.” “Fascism treats women as mere breeders. ‘Children, kitchen, and the church,’ was the Nazi slogan for women,” the pamphlet said.

Fascists “make their own rules and change them when they choose…. They maintain themselves in power by use of force combined with propaganda based on primitive ideas of ‘blood’ and ‘race,’ by skillful manipulation of fear and hate, and by false promise of security. The propaganda glorifies war and insists it is smart and ‘realistic’ to be pitiless and violent.”

After the war, scholars like Eric Hoffer studied the societal conditions necessary for fascism to take hold. In his 1951 The True Believer: Thoughts on the Nature of Mass Movements, Hoffer noted that demagogues needed a disaffected population whose members felt they had lost the power they previously held, that they had been displaced either religiously, economically, culturally, or politically. Such people were willing to follow a leader who promised to return them to their former positions of prominence and thus to make the nation great again. But to cement their loyalty, the leader had to give them someone to hate. Who that was didn’t really matter: the group simply had to be blamed for all the troubles the leader’s supporters were suffering.

More recently, scholars are adding to our understanding of fascism by describing it as a form of political behavior. “It is,” Robert O. Paxton says in his 2005 The Anatomy of Fascism, “marked by obsessive preoccupation with community decline, humiliation, or victimhood and by compensatory cults of unity, energy, and purity, in which a mass-based party of committed nationalist militants, working in uneasy but effective collaboration with traditional elites, abandons democratic liberties and pursues with redemptive violence and without ethical or legal restraints goals of internal cleansing and external expansion.”

McGovern’s condemnation of today’s Republicans fits these descriptions even without the parallels noted by Carrie Kaufman in her You’re Overthinking It with Carrie Kaufman between leading Republican officials and Nazi leaders. On July 19, Kaufman showed how recent speeches by Secretary of State Marco Rubio and deputy White House chief of staff Stephen Miller at last Thursday’s State Department about “Left-wing terrorism” echoed Adolph Hitler’s Mein Kampf and Heinrich Himmler’s 1942 pamphlet Der Untermensch, or “The Subhuman.”

But McGovern left a key element of fascism out of his speech yesterday: the alignment between the government and favored businesses. Fascism subordinates business activity to the needs of the government. In exchange for supporting them, government officials award massive government contracts and subsidies to favored businesses, and protect them from regulation.

Earlier this year, the administration established a new “Economic Defense Unit” (EDU) at the Pentagon, charged with investing public money in defense industries. Ana Swanson of the New York Times reported in March that the Pentagon was recruiting Wall Street investment bankers to the team to spend up to $200 billion in government investment over the next three years. That access, recruiters told prospective employees, would give them “access to fund-raising channels that include royal families and foreign sovereign contacts” in case they ever wanted to raise money for their own investment firms.

And the U.S. government under Trump has taken ownership shares of private companies it considers important to national security, like Intel, for example, and U.S. Steel. In November of 2025, Swanson reported in the New York Times that the administration had already committed more than $10 billion in taxpayer funds to favored businesses. The lack of transparency in this interference in the private sector raised concerns about favoritism, corruption, and the distortion of the market.

Last night, Michael R. Gordon and Stephen Kalin of the Wall Street Journal broke the story that Trump has approved a landmark 30-year deal to provide Saudi Arabia with a civilian nuclear program. The deal gives U.S. companies a monopoly on developing a nuclear infrastructure for the Arab country, but it does not include monitoring and inspections by the International Atomic Energy Agency to make sure nuclear fuel isn’t enriched for a nuclear weapon.

The idea of building nuclear power plants in Saudi Arabia was central to Trump’s 2016 bid for office. Members of Trump’s inner circle, including his son-in-law Jared Kushner and disgraced national security advisor Michael Flynn, hatched a plan for a joint U.S.-Russian project to build nuclear power plants in Saudi Arabia. In June 2016 they formed a company called IP3 International, short for International Peace, Power and Prosperity.

In Trump’s first term, White House ethics officials and members of the National Security Council warned that selling nuclear technology to Saudi Arabia could violate the Atomic Energy Act. Members of the administration continued to work on the project, nonetheless.

This week’s deal is worth billions of dollars in contracts for U.S. firms that work in nuclear technology, particularly Westinghouse, which filed for bankruptcy in 2017 when its nuclear technology took longer to build and was more expensive than the company estimated. In June, the Department of Energy announced it was investing $17.5 billion in loans to Westinghouse and local utility and energy companies to build 10 large-scale commercial nuclear reactors across the U.S.

The deal is “a big win for the U.S. commercially and geopolitically, enriching American firms and tying Saudi Arabia closer to the U.S.,” Kristin Diwan, a senior resident scholar at the Arab Gulf States Institute, a research organization in Washington, D.C., told Vivian Nereim of the New York Times. It is also a win for Saudi Arabia’s crown prince Mohammed bin Salman (MBS), who has vowed that he will build a nuclear weapon if Iran does, making the deal a blow to nuclear nonproliferation.

Fascism has always been a vehicle for frustration with a system that was failing the people. But it has never been a solution for fixing that system.

Notes:

https://archive.org/details/ArmyTalkOrientationFactSheet64-Fascism/mode/2up

You're Overthinking It with Carrie Kaufman
Miller and Rubio Go Full Mein Kampf in “Ministerial”
The news media has been busy over the last few days debunking the lies that Donald Trump told in his speech to the nation on Thursday…
Read more

https://www.wsj.com/world/middle-east/trump-approves-landmark-nuclear-deal-with-saudi-arabia-in-big-win-for-kingdom-2ed77584?mod=hp_lead_pos2

https://www.energy.gov/articles/department-energy-announces-american-nuclear-supply-chain-loans

https://americanoversight.org/investigating-the-trump-administrations-efforts-to-sell-nuclear-technology-to-saudi-arabia/

https://www.reuters.com/article/world/how-two-cutting-edge-us-nuclear-projects-bankrupted-westinghouse-idUSKBN17Y0C7/

https://www.theguardian.com/us-news/2026/jul/16/political-violence-event-trump-marco-rubio

https://www.theguardian.com/world/2019/feb/19/white-house-saudi-arabia-nuclear-technology-house-oversight-inquiry-report

https://www.nytimes.com/2025/11/25/us/politics/trump-intel-steel-minerals-china.html

https://www.nytimes.com/2026/03/13/us/politics/wall-street-access-pentagon.html

X:

RepMcGovern/status/2079619983114383421

Bluesky:

leftcoastreads.bsky.social/post/3mr6vptkpls2c

Share

What should I ask Gita Gopinath?

Yes I will be doing a Conversation with her.  From Wikipedia:

Gita Gopinath…is an Indian-American economist who is currently serving as the Gregory and Ania Coffey professor of Economics at Harvard University and previously served as the first deputy managing director of the International Monetary Fund (IMF), from 21 January 2022 to 31 August 2025. Before that she also served as chief economist of the IMF between 2019 and 2022.

Here is Gita on scholar.google.com, she is an expert in international finance and exchange rates, and also international capital flows, among other topics.  Here is Gita on Twitter.  So what should I ask her?

The post What should I ask Gita Gopinath? appeared first on Marginal REVOLUTION.

       

Comments

Related Stories

 

My excellent Conversation with Andrew Graham-Dixon

Here is the audio, video, and transcript.  From the episode summary:

Tyler and Andrew discuss whether it was inevitable we’d rediscover Vermeer, how that vanishingly rare sect left its fingerprints all over his life, why the Met has misread its own Allegory of the Catholic Faith, whether Vermeer painted for money or pointedly refused to, where the Gardner’s stolen Concert might be, why Dutch music never blossomed as much as Dutch painting did, how the Church of England rivaled the Cultural Revolution in wiping out British art, whether you can still spot an English painting on sight, the love that saturates late Rembrandt and the mystery of his soaring print prices, the nail on the wall that proves two Vermeer paintings are a pair, why the French are to blame for George Stubbs’ lack of status, whether we can still love Malevich, why Andrews calls recent Richter “almost like printing money,” why female artists and antique textiles remain absurdly cheap, why nobody builds beautiful neighborhoods any longer, and much more.

Excerpt:

COWEN: In what sense was Vermeer a liberal?

GRAHAM-DIXON: Well, the main discovery of my book is that Vermeer was among the very first pioneers of what we now call the liberal tradition.

COWEN: What we now call the Netherlands, then the Dutch Republic.

GRAHAM-DIXON: In the Dutch Republic, and the Dutch contribution to the Enlightenment, which is the origins of the liberal tradition, has been very much forgotten. That’s absolutely at the heart of my book, is an attempt to remember these people, to bring them back into the place in history that they deserve. I’m not only talking about Vermeer. I’m talking about his friends, his patrons, because that’s the discovery of the book, is that he and his friends were a remarkable group of people, and they have been completely forgotten, and what they believed has been largely forgotten too. We need to remember it now, probably more than ever.

COWEN: This was also a religious movement.

GRAHAM-DIXON: Yes.

COWEN: Doctrinally, how would they have been different from, say, Protestants in England? The Collegiants, the Remonstrants?

GRAHAM-DIXON: Yes. Vermeer’s patrons, it emerges, and Vermeer himself, were part of a Protestant sect in Holland called the Remonstrants. They had a more extreme manifestation called the Collegiants, and they were unique among all Christian denominations of that time in being utterly opposed to division, hostility, enmity. The only thing they wanted was to bring all Christians, indeed all people—they included Jewish people and Muslim people—they wanted to bring everyone together within a faith that only really cleaved to the essentials of what Jesus Christ said, particularly in the Sermon on the Mount.

They said, “If you actually follow Jesus properly, you can never make war. You can never persecute someone who differs from your opinion. You can never pick on somebody because they believe something different.” They were very, very tolerationist. They formed the first pacifist movement in European history. They lived in a time of appalling warfare, the Thirty Years’ War, probably the worst war in the history of the modern West. Fifteen million out of 20 million German people died in the course of 30 years, and the 5 million who were left, all the women had been violated, and all the men only have one arm or one leg. It was truly atrocious.

They’re responding to real traumatic historical events. They’re responding to what is going on around the corner from where they live. They come up with this very, very beautiful approach to life, full of optimism, full of idealism. For about 20 years, the Dutch Republic actually lives by these codes of belief to a great extent. It’s the only tolerant country in Europe.

And this:

COWEN: What makes George Stubbs such an underrated painter?

GRAHAM-DIXON: I think he’s been underrated forever because of his subject matter. I blame the French. The French in the 17th century invented a system for ranking works of art by their subject matter. Up at the top, you and I, we could put on some armor and confront each other with swords, and Poussin would paint us. That would be a history painting done from the life. That would be at the pinnacle. Then would be a painting of an event, maybe a meeting between great men. Then, below that, there would be a portrait.

You’d keep going down, and eventually you’d get to paintings of still lifes, like flowers and fruit, or paintings of animals. These were the lowest works of art because they featured the basest things. I think that Stubbs, to a certain extent, was the victim of that.

COWEN: Are we now able to see them properly?

GRAHAM-DIXON: Yes. It’s been no problem since Stubbs and Constable, more than anyone else, slightly controversial to say, but I think they paved the way for modernism because they showed, not necessarily deliberately, but they showed that anything could be painted, anything at all, like a piece of mud or a piece of a river could be painted in such a way as to touch on the very highest thoughts and ideas and beliefs and feelings that subject matter is completely, in a sense, irrelevant. Cézanne picked up on that when he said, “I want to stun Paris with an apple. I’ll paint an apple, and I’ll paint it with such astonishing concentration that you’ll see it as an epistemological challenge to all of your philosophy.”

I think that French art gets that from British art because British art has to be like that because that’s all the aristocrats are going to commission. They look down on people like Stubbs, in a sense. If I’m an aristocrat in the 18th century and I want a really important picture, I’ll go to Italy and I’ll buy one, thank you very much, and I’ll buy my history painting from Titian. You, sir, Stubbs, you just paint my horse. It’s my racehorse, and I’m very fond of him. Paint him well and don’t scare him.

Recommended, interesting throughout.  And here is Andrew’s very interesting new book on Vermeer.

The post My excellent Conversation with Andrew Graham-Dixon appeared first on Marginal REVOLUTION.

       

Comments

Related Stories

 

Olympic Mountain Glory

The Olympic Peninsula, viewed at an angle from above, features snow-capped mountains surrounded by deep, forested river valleys. Islands in Puget Sound and developed areas including Seattle and Tacoma appear across the top of the photo.
May 6, 2016

Alpine glaciers, wild coastlines, temperate rainforests, and deep river valleys coexist on the Olympic Peninsula in the northwest corner of Washington state. Surrounded by blue waters, peaceful islands, and bustling population centers, its rugged interior remains a relatively remote bastion of wilderness.

The Olympic Mountains’ imposing terrain comes into focus in this oblique view of the region, captured by an astronaut aboard the International Space Station. The image is a composite, made of several sequential, overlapping photos fused together into a panorama. Olympic National Park encompasses the peninsula’s mountainous core, along with some stretches of the Pacific coastline. Much of the remaining area is either national forest, state-owned land, or tribal territory.

The rock making up the mountains mostly originated beneath the surface of the ocean. From about 55 to 15 million years ago, layers of basalt from undersea eruptions and sand and mud transported seaward by rivers accumulated on the ocean bottom. This material was scraped off the Juan de Fuca plate as it subducted beneath the North American plate, with rock layers crumpling and rising up to 8,000 feet (2,440 meters) above sea level.

Tectonic forces continue to push the mountains skyward, but the countervailing force of erosion in this rainy, snowy corner of the country effectively cancels out the uplift. Snow at higher elevations feeds glaciers that carve out underlying rock. Glaciers in the Olympics are retreating and thinning, however, and their numbers are declining. One study tallied 255 glaciers and perennial snowfields in the range in 2015 and found that 35 glaciers and 16 perennial snowfields had disappeared in the preceding 35 years.

Other erosion is evidenced by the deep valleys radiating out from the snowy peaks. The Hoh, Queets, and Quinault rivers, draining west into the Pacific Ocean (bottom of the frame), are prominent in this view. These verdant valleys are known for their temperate rainforests, and the ancient forest in the Hoh River valley was once considered among the most naturally quiet places in the U.S., uninterrupted by human-caused noise.

Flowing to the north, the Elwha River has a rich natural and human history, including some of the earliest Euro-American exploration of the Olympics. Sponsored by a Seattle newspaper, an expedition from December 1889 to May 1890 crossed the mountain range from north to south, traveling up the Elwha valley and down the Quinault. The party spent several months in the Elwha Valley, their progress hindered by an unusually harsh and snowy winter. 

In the early 1900s, entrepreneurs saw economic opportunity in the valley. Two dams constructed on the river produced power for local industry. But the structures came with costs, such as blocking the migration of once-abundant trout and salmon to their spawning grounds. In 2011 and 2014, the dams were removed in what was then the largest such project in the U.S., and the process of restoring fish populations, seeding native plant communities, and replenishing sediment along the riverbanks commenced.

The mouth of the Elwha forms a delta in the Strait of Juan de Fuca, the waterway bordering the peninsula to the north. The U.S.-Canada border runs through the middle of this 11- to 17-mile-wide (18- to 27-kilometer-wide) channel, with Vancouver Island in British Columbia lying to the north. The strait connects the Pacific Ocean with the Strait of Georgia and Puget Sound. Ship traffic uses the strait to access important West Coast ports, including Seattle and Tacoma, visible along the top-right edge of the image.

Astronaut photographs ISS047-E-104138 through ISS047-E-104144 were acquired on May 6, 2016, with a Nikon D4 digital camera using a focal length of 400 millimeters. They are provided by the ISS Crew Earth Observations Facility and the Earth Science and Remote Sensing Unit at NASA Johnson Space Center. The images were taken by a member of the Expedition 47 crew. The images have been cropped and enhanced to improve contrast, and lens artifacts have been removed. The International Space Station Program supports the laboratory as part of the ISS National Lab to help astronauts take pictures of Earth that will be of the greatest value to scientists and the public, and to make those images freely available on the Internet. Additional images taken by astronauts and cosmonauts can be viewed at the NASA/JSC Gateway to Astronaut Photography of Earth. Story by Lindsey Doermann.

References & Resources

The Olympic Peninsula, viewed at an angle from above, features snow-capped mountains surrounded by deep, forested river valleys. Islands in Puget Sound and developed areas including Seattle and Tacoma appear across the top of the photo.

You may also be interested in:

Stay up-to-date with the latest content from NASA as we explore the universe and discover more about our home planet.

Belts of Green in the Washington Suburbs
3 min read

Along the northeast side of the Capital Beltway in Maryland, green spaces weave through the developed landscape.

Article
Colonial National Historical Park
2 min read

The colonial communities of “America’s historic triangle” played defining roles in the road to American independence.

Article
A Turquoise Tint for the Black Sea
3 min read

Phytoplankton added a milky blue hue to the waters of the Black Sea and nearby waterways in spring and summer…

Article

The post Olympic Mountain Glory appeared first on NASA Science.

Is this a fossilized turtle on Mars? Is this a fossilized turtle on Mars?


Orchestrions

San Francisco tip: it only costs around $15 ($10 in quarters plus a $5 bill for the self-playing violin) to activate every single Orchestrion in Musée Mécanique.

And because most people are bad at allocating their funds you may well be the ONLY person activating the Orchestrions, which means you get to craft the soundscape for the entire museum.

Tags: san-francisco

California Sea Lion

California Sea Lion

California Sea Lion

California Sea Lion, in San Francisco County, US, CA

We took some visiting family to Pier 39 to see the sea lions. They're somehow always even fun than I remember them being last time.

Tags: san-francisco, wildlife

The Google Engineer Who Set The Skydiving Record At 58 - EP 83 Alan Eustace

Most people think that the Red Bull guy Felix Baumgartner holds the skydiving record. But this is not the case.

The real record-holder is Alan Eustace. In 2014, at age 58, the Google engineer rode a helium-filled balloon into the stratosphere and then fell back to Earth from 135,890 feet. Eustace’s fall lasted 4 minutes and 27 seconds, and he broke the sound barrier, reaching a max speed of 822 mph. And now, in a very rare appearance, Eustace has come on the podcast to discuss his adventure.

Subscribe now

Baumgartner, who dropped from 128,852 feet, received tons of attention because Red Bull turned his jump into a media spectacle. Eustace, by contrast, came up with the plan for his jump and spacesuit at his kitchen table, paid for a small team to help him build his suit and then did his skydive with just a single reporter there as a witness.

This all fits with Eustace’s personality. He’s quite private and just goes about his work.

It took several months of begging to get Eustace on the show, and I don’t think he’s ever discussed his skydive in this type of detail before.

In this episode, we get into what drove a 58-year-old man to want to chase this feat, the rich history of stratosphere jumping and Eustace’s remarkable career as one of the key Google engineers who built the modern internet. We also hear Eustace’s thoughts on AI and his latest quest to search the ocean floor.

(I apologize for saying that Eustace jumped from space in the episode. It was the stratosphere. I was excited.)

If you can spare a minute, please do us a favor by filling out this ever so brief survey, so we can learn a bit more about our subscribers. Help us be better for you. Thanks!!

OUR SPONSORS

SendCutSend

You know who else makes stuff for America? That would be SendCutSend. If you want to celebrate our great nation by building a metal part, then head on over to SendCutSend where you’ll get a 15 percent discount thanks to Core Memory on whatever you’re trying to build. We believe in you.

Brex

The Core Memory podcast is also sponsored by Brex, the intelligent finance platform built to help companies spend smarter and move faster.

Did we go to Texas, find a telescope ranch and then obtain an entire nebula in Brex’s honor? Oh yes, we did.

We run on Brex and so should you. Learn more about Brex right here.


Timestamps (they link out to YouTube)

00:00 Intro
03:26 Chasing Alan Eustace
07:42 Rocket Sleds and Seatbelts
10:38 Why Build a Capsule?
13:00 Kittinger’s Three Jumps
18:11 The Idea Takes Hold
21:07 Elon’s Mars Greenhouse
23:18 The Jump That Failed
28:37 Solving the Death Spin
30:07 Racing Red Bull
32:26 One Reporter, No Livestream
34:12 The Kittinger Coincidence
38:26 Not Your Typical Daredevil
40:43 Goodbye Videos for His Kids
44:35 The First Spacesuit Jump
48:37 Two Hours to Space
50:53 Google Earth in Reverse
54:36 Can I Get a Countdown?
58:55 Breaking the Sound Barrier
1:00:08 The Crash Landing
1:02:16 What It Really Cost
1:08:28 The Magic of Gas Balloons
1:10:30 Did the Jump Change Him?
1:17:03 The VLSI Golden Age
1:22:56 Is AI Revolutionary?
1:29:20 The Google Translate Story
1:35:49 The Fall of DEC
1:41:09 Reinventing the Data Center
1:48:51 Impossible to Inevitable
1:51:40 Inside Commonwealth Fusion
1:56:21 Hunting Amelia Earhart
2:03:21 What’s Under the Golden Gate?

Share

Oligarchy and the Media

For all my interviews and more, subscribe on YouTube.

Transcript

Good news. The second richest man in America might be prevented from taking over CNN. That's the good news. The bad news is, aside from thefact that he probably will manage to pull it off anyway, the bad news is that that would be only a small piece of the ongoing takeover of U.S. media by oligarchs. And in turn, the media takeover is just part of the extraordinary exercise of power by the extraordinarily wealthy small number of men who have been wreaking so much havoc with America as we know it.

Hi, I'm Paul Krugman. Doing a video today, because I didn't feel like doing a usual chart-heavy, analytics-heavy post, but very much on a topic I have been writing about and will continue to write about, which is the rise of oligarchy in America.

Now, I know some people balk at that. But we're not talking about some kind of hidden conspiracy. We're not talking about the Protocols of the Elders of PayPal. We are talking instead about stuff that's largely out in the open, though not fully understood, which is the way that an incredibly wealthy small group of men, mostly men, is able to commandeer a lot of the political life of a country that is still nominally a democracy. And that's a fundamental story for our time, maybe the fundamental story.

How does that takeover work? Well, there is what I think of as the middle level, which is the place where it's most easily quantified, tends to get most of the attention, which is campaign finance. American campaigns are very money intensive and have become more money intensive because we've opened the floodgates with Citizens United. And a lot of that money comes from a very small number of incredibly wealthy people. According to the New York Times analysis, about 20% of all campaign contributions in 2024 came from 300 billionaires and their families.

That's a pretty big impact. A country of more than 300 million people, and 300 billionaires are a fifth of campaign finance, and surely more strategic, more targeted than the average donor. So that's really a very, very large role just in that direct sense of who pays for campaigns.

But that's not the only level. There is a lower level, lower in the sense of morally lower, I guess, which is just plain buying politicians, buying policies, paying for the policies you want with cash or crypto on the barrel.

There has always been some of that in our system, but it was normally discreet, indirect, deniable, the revolving door. It was the case even more than 20 years ago that when the Bush administration pushed through a Medicare bill that was very favorable to pharmaceutical interests, that the then chairman of the House Ways and Means Committee, who basically engineered and steered the bill through Congress, then promptly retired and became the chief lobbyist for the pharma lobby. So this kind of thing has been going on for a very long time.

But now it's just blatant, out in the open, and the sums are massive. We just have literally billions of dollars thrown at the president and his family. No doubt large sums to other government officials, large sums to at least some members of Congress. So just plain buying the policies you want — and it’s not just that a large share of wealth is held by a small number of people, but that those are the people who are best positioned to really deploy their wealth to corrupt the system.

There's also something, I guess you can call it a higher level, which is what military strategists call shaping the information space, which occurs at a couple of levels. One of them is the promotion of ideas and ideology that serve the interests of the very wealthy.

You see that on many issues. You certainly see it very much on economic policy. If you ask, why do people still go out there saying that tax cuts pay for themselves and that tax cuts on the rich are an enormously powerful tool for stimulating economic growth? That's been tested to destruction, and it just ain't so. But it's a zombie idea. It keeps shambling along, eating people's brains, even though it should be dead. And the reason is, well, there's a lot of money in it.

If you Google something I've written on, more often than not, when I do that, the top sponsored post at the top of the search page is an attack on me sponsored by some right-wing organization. And if you ask who supports those right-wing organizations, well, guess who.

And it’s equally or worse the case in climate science. Scientific journals have been pretty good at not publishing climate disinformation. But when they do publish things that are somehow skeptical, or usually not outright denial, but attempting to sow discord about climate change, what percentage of those studies have received financial backing from fossil fuel interests? The answer is 100. It's all about the money. So this is, again, this is not new. Upton Sinclair: “It's difficult to get a man to understand something when his salary depends on his not understanding it.” So that has always been the case.

But now we have something which is really, really important and is another level of this, which is the takeover of the media. So, okay. Ellison, or the Ellison family —because nominally this is Ellison's son in charge of Paramount — has already acquired CBS and has hired Bari Weiss to basically corrupt and destroy that network. If the deal for takeover of Warner proceeds, then CNN will get the same treatment. I'm finding CNN a very good news source, just braver at taking on what's really happening than my old employer, the New York Times, which is a great news organization and may be more necessary than ever, but tends to be cautious — and CNN is a little bit less cautious.

But anyway, if he gets away with it, then CNN as we know it will almost disappear. It will almost turn into Fox News. Now, that won't be a profitable venture. There's already a Fox News, and so creating another one is not going to actually produce a lot of profits, if any, but that's not the objective. This is buying influence.

Elon Musk, of course, took over the app formerly known as Twitter. Which was already becoming a more difficult place even before its takeover. I used to have, I guess, I think I had 4 million followers there. But it was impossible. I had to shut off comments because of the cesspool that Twitter had become. But now it is really by design. It is heavily tilted. That can be quantified. The algorithm really tilts it towards right-wing stuff, promotes really rabid racist views.

And unfortunately, the network effects, the centrality that Twitter used to have, still keeps a lot of people on X, where they are influenced: people's views change.

And also something that I don't know how to quantify, but it's very obvious if you follow and pay attention to people's positions, is that people who spend a lot of time on Twitter, elites who spend a lot of time on Twitter, start to think that the views they hear there are representative of where the country is — which they are not. But it does, in fact, tilt policy, tilt understanding to the right.

The third richest man in America is Mark Zuckerberg, who made his billions from Facebook. Facebook is old-fashioned: I don't know anybody who uses Facebook. But I know that lots of people do. And it's still a very important information source and has, again, been tilted.

On most of these media things, it's not as blatant as what Musk is doing at X. But it still has a big influence in changing the tone of the discussion and biasing the discussion towards positions that favor the interests of billionaires as well as favoring their prejudices if they happen to be, like Musk, authoritarian white supremacists.

Okay. And the fourth richest man in America is Jeff Bezos, who purchased the WashingtonPost. I think he purchased the Post initially out of a belief that he was going to enhance his prestige. It certainly looked in his initial tenure as if this was actually more of a vanity purchase than a political purchase. But a billionaire is going to billionaire. And so he eventually shifted the Washington Post's editorial policy hard right, eviscerated the news division. There are still some brave, plucky reporters doing good reporting there, but it's a shadow of what it used to be. And of course, it's not at all the institution of Katherine Graham and Ben Bradlee, not anymore. So that's another challenge.

What do you do about this? Obviously, one does what one can to try to limit this takeover of the information environment. And so we have the suit brought against the attempted purchase of Warner, hence CNN, by Paramount, hence Ellison. And that might succeed. You might think, well, if it's delayed, then what are the chances of actually ruling it out? Except that apparently there's a bit of a financial clock ticking for Ellison, who really has extended himself pretty far. So that's possibly going to block it, and that's good. It would have been great if someone had found a way to keep Musk from destroying Twitter. So you can look for solutions to immediate threats.

But you're not going to hit all of these balls. And so the constant pressure towards a takeover of the news media, constant pressure towards a takeover of the general information environment by a handful of billionaires, is not going to go away. The constant threat or reality of corruption of the government by billionaires is not going to go away. Maybe once Trump is gone, it'll become less blatant, but it won't go away just because someone more discreet takes office.

Even if we have an honest president, which in the current environment, I'm sorry, does mean a Democrat, but even if we have an honorable president, the corruption of the system will still be a continual threat because of all the money flowing around.

So in the end, the only way out of this, the only reasonably durable solution is to not have so much wealth at the top. If you don't like what's happening to our institutions, if you don't like what's happening to the media, if you don't like the corruption of government, if you don't like the overwhelming of campaigns by big money with nefarious ends, the only lasting solution is to reduce the amount of wealth at the top.

Woodrow Wilson: “If there are men big enough to own the government, they're going to own the government.” If we're going to have that much money in the hands of a few hundred people, and in the case of the real top of it, just 15 or 20 people, then you're not going to be able to maintain a truly democratic system of government.

Oligarchy is not the only thing wrong with America. It's not the root of all evil. But it's the root of a lot of evil. And until we bring that concentration of wealth at the top down, we're going to be fighting a constant rearguard action trying to save some of what America is supposed to be about.

Have a nice day.

Quoting Thomas Ptacek

I genuinely believe that if you took an open weights model from 2025 and built a pentest harness for it, it could do this kind of sandbox escape and scan/hack in most networks. This is only surprising because you assume OpenAI has sounder sandboxes.

Thomas Ptacek, doesn't think this even needs a frontier model

Tags: thomas-ptacek, openai, security, generative-ai, ai-security-research, ai, llms, sandboxing

OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened

This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke its way out of OpenAI's sandbox, then found exploits to break in to Hugging Face, all so it could cheat on the test by stealing the answers.

Along the way it helped make the strongest case yet for how the imbalance of model availability is hurting our ability to secure our software.

Here's what happened

We currently have three documents to help us understand what happened here.

  1. ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks? is a paper published on 11th May 2026 describing ExploitGym, a new eval suite for LLM-powered agent systems.
  2. Security incident disclosure — July 2026 by Hugging Face on 16th July 2026 describes how they detected an attack from an "agentic security-research harness - used LLM still not known" that breached some of their systems.
  3. OpenAI and Hugging Face partner to address security incident during model evaluation from OpenAI on 21st July 2026 confesses that it was their agent harness that did this, and that they're working with Hugging Face to clean up the mess.

ExploitGym

I hadn't seen the ExploitGym paper before and it's a really interesting one. Authors from UC Berkeley, the Max Planck Institute, UC Santa Barbara, and Arizona State designed a new benchmark for evaluating models on their ability to turn a reported vulnerability into a concrete exploit. OpenAI, Anthropic, and Google provided feedback and helped run the benchmark against their models.

The benchmark "comprises 898 instances derived from real-world vulnerabilities that affected popular software projects" - including the Linux kernel and V8 JavaScript engine.

Here's the paragraph that best represents their benchmark results:

Among all configurations, Claude Mythos Preview and GPT-5.5 achieve the highest success counts (157 and 120 successes, respectively), demonstrating that current frontier agents can exploit a substantial subset of real-world vulnerabilities under controlled conditions. GPT-5.4 also solves a notable 54 tasks, placing it in an intermediate tier. The remaining model–agent pairings solve fewer than 15 tasks each, underscoring that end-to-end exploitation remains challenging and sharply differentiates today’s frontier systems. Notably, Claude Opus 4.7 achieves fewer successes than Claude Opus 4.6 despite being a newer checkpoint, and does so at substantially lower cost on the full set. Trace inspection reveals that Claude Opus 4.7 and Gemini 3.1 Pro frequently conclude early after judging the target vulnerability non-exploitable.

The paper also describes the approach they took to preventing the agents from cheating by going outside the parameters of the test. This becomes relevant in a moment!

Outbound connections are restricted to a curated allowlist that permits routine package installation (Ubuntu apt repositories and PyPI) and fetching the toolchains required for building V8. All other external endpoints are blocked.

The paper concludes with this (emphasis mine):

Our results show that autonomous exploit development by frontier AI agents is no longer a hypothetical capability. While current agents are not yet reliable across all targets, they already exploit a non-trivial fraction of real-world vulnerabilities, including complex targets such as kernel components. This rapid emergence is itself a central finding, showing that capabilities that would have seemed implausible are now present in deployed frontier models.

An important detail here: this paper isn't about discovering vulnerabilities; it's about being able to take those vulnerabilities and turn them into working exploits.

When Anthropic first restricted access to Mythos back in April they talked about this capability as well. A model that can act on vulnerabilities is a lot more dangerous than one that can just discover them.

One of the ways Fable differs from Mythos is that it's more likely to refuse to weaponize vulnerabilities in this way. I get the impression the US government did not understand that distinction when they banned Fable last month.

The Hugging Face incident

The first hint we got of the attack was in this blog post by Hugging Face on 16th July 2026:

A malicious dataset abused two code-execution paths in our dataset processing (a remote-code dataset loader and a template-injection in a dataset configuration) to run code on a processing worker. From there, the actor escalated to node-level access, harvested cloud and cluster credentials, and moved laterally into several internal clusters over a weekend.

I hope they release more details about the code that pulled this off. I'm assuming this means packages using the datasets library, a Hugging Face project for bundling up and sharing datasets on their platform. That library used to execute arbitrary code but has been steadily locked down over time, with the 4.0.0 release in July 2025 removing the trust_remote_code=True flag entirely.

Assuming the attack used that library it must have either abused pickle serialization in some way, found some other non-obvious code execution path, or (most likely) specified datasets<4.0.0 as the dependency.

The campaign was run by an autonomous agent framework (appearing to be built on an agentic security-research harness - used LLM still not known) executing many thousands of individual actions across a swarm of short-lived sandboxes, with self-migrating command-and-control staged on public services.

This was a sophisticated attack!

Then Hugging Face hit a wall: they tried to use "frontier models behind commercial APIs" - I'm guessing from Anthropic and OpenAI - to help analyze the attack, and were blocked:

When we started the log analysis, we first used frontier models behind commercial APIs. This did not work: the analysis requires submitting large volumes of real attack commands, exploit payloads, and C2 artifacts, and these requests were blocked by the providers' safety guardrails, which cannot distinguish an incident responder from an attacker.

They switched to their own self-hosted instance of MIT licensed GLM-5.2 and it helped them figure out what was going on.

This indicated a fundamental asymmetry between the defending team and the (so-far unknown) attacker:

We do not know which model powered the attacker's agents, whether a jailbroken hosted model or an unrestricted open-weight one; either way, the attacker was bound by no usage policy, while our own forensic work was blocked by the guardrails of the hosted models we first tried.

As a useful indicator of how seriously they took the attack:

[...] Finally, we have also reported this incident to law enforcement agencies.

So who was responsible for this "autonomous agent framework"? It turned out to be OpenAI themselves.

The OpenAI confession

Five days later, on July 21st, OpenAI revealed the culprit. They had been running the ExploitGym benchmark against a new, as-yet undisclosed model, and that model had been operating way outside its intended parameters (emphasis mine):

After investigating, we now know that this particular incident was driven by a combination of OpenAI models — including GPT‑5.6 Sol and an even more capable pre-release model, all with reduced cyber refusals for evaluation purposes — while being internally tested on a benchmark⁠ [ExploitGym] of cyber capabilities. [...]

We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity. Our benchmarks run in a highly isolated environment, with network access constrained to the ability to install packages through an internally hosted third-party software that acts as a proxy and cache for package registries.

The models identified and chained vulnerabilities across OpenAI’s research environment and Hugging Face’s production infrastructure to obtain test solutions directly from Hugging Face’s production database. All evidence suggests that the models were hyperfocused on finding a solution for ExploitGym, going to extreme lengths to achieve a rather narrow testing goal.

It's pretty clear what happened here. OpenAI removed safety filters for an in-progress model, locked it up in a sandbox and told it to solve the ExploitGym problems. Given the absence of guardrails there was nothing to prevent the model from attempting to break out of that sandbox, break into Hugging Face, and read the answers from there instead.

OpenAI's sandbox used the now common pattern of allowing access to an allow-list of sites related to installing additional packages via a controlled proxy. In OpenAI's words:

While operating in our sandboxed testing environment, our models spent a substantial amount of inference compute finding a way to obtain open Internet access, in pursuit of solving the evaluation problem. To gain access, the models identified and exploited a zero-day vulnerability (which we’ve now responsibly disclosed to the vendor) in the package registry cache proxy. With this access, our models performed a series of privilege escalation and lateral movement actions in our research testing environment until the models reached a node with Internet access.

So step one was to break out onto the public internet. The model then broke into Hugging Face to find the answers:

After gaining Internet access, the models inferred that Hugging Face potentially hosted models, datasets and solutions for ExploitGym. Knowing this, the model searched for and successfully found ways to gain access to secret information that it could use to cheat the evaluation. In one example, the model chained together multiple attack vectors, including using stolen credentials and zero-day vulnerabilities to find a remote code execution path on the Hugging Face servers.

Chaining together multiple attack vectors is exactly the kind of thing these new models can do, where previous generations of models might have failed.

I wrote last month about how Claude Fable is relentlessly proactive, when I noticed it spinning up custom web servers and deploying CORS tricks on my own laptop just to help debug a WebKit CSS issue. It turns out relentless proactivity is the defining trait of this new generation of Mythos-class models. If you set them a goal and give them a way to get there, even inadvertently, they will figure it out.

Resist the temptation to write this off as a stunt

There will inevitably be some people who dismiss this story as a dishonest marketing trick by OpenAI to make their models sound terrifyingly effective. I found 81 instances of the term "marketing" in the Hacker News discussion of the incident.

To those people I say pull your heads out of the sand - you're now including Hugging Face in your conspiracy theories, just so you can deny the crescendo of evidence here!

The best models we have today have the ability to both find and exploit new vulnerabilities. The ExploitGym paper itself concludes that "autonomous exploit development by frontier AI agents is no longer a hypothetical capability", and this incident is a perfect example of exactly that.

The asymmetry is increasingly frustrating

One of the most infuriating details of this story is how Hugging Face, faced with an accidental and aggressive attack from one of OpenAI's models, were unable to then turn to OpenAI's models to help them fend off the attack.

The frontier models we have access to are increasingly being constrained in how much they can help us protect our software, heavily influenced by the US government's ongoing threat of export controls. Claude Fable 5 wouldn't even proofread this article for me! It insisted on downgrading me to a less capable model.

Meanwhile open weight models from China such as GLM-5.2, Kimi 3 and the new Qwen 3.8 Max appear to have none of these restrictions - and any restrictions that do exist can likely be fine-tuned out of them by modifying the weights

These constraints are meant to make us safer. I think there's a risk that the effect they are having is the opposite.

Tags: sandboxing, security, ai, openai, generative-ai, llms, hugging-face, anthropic, paper-review, ai-security-research

Are AI labs pelicanmaxxing?

Are AI labs pelicanmaxxing?

Excellent piece of work by Dylan Castillo, who took a deep-dive into the frequently pondered question of whether the AI labs have been deliberately training models to draw pelicans riding bicycles in response to my deeply unscientific benchmark.

I've been randomly spot-checking this in the past by testing models against other animals riding other types of vehicle, but never with anything close to the diligence of Dylan's methodology here.

Dylan took 8 animals × 6 vehicles = 48 prompts and ran them three times each through 7 different models ( GPT-5.6 Terra, Claude Sonnet 5, Gemini 3.5 Flash, Grok 4.5, Qwen3.7-Max, GLM-5.2, and DeepSeek V4 Pro). He then used GPT-5.6 Luna and Gemini 3.1 Flash-Lite to help evaluate the results.

There's a neat filter view for exploring the results:

Screenshot of a grid for sample 1/3 of GLM-5.2, with pelicn and flamingo and heron riding bicycle, unicycle, skateboard, scooter, plane and boat

For the models he tested he could find no evidence of pelimaxxing:

Pelicans aren’t drawn any better than other animals. Bicycles aren’t drawn any better than other vehicles. And no lab draws the combination better than its pelicans and bicycles already predict. GLM-5.2 comes closest: it has the largest boost on the exact pelican-bicycle cell, and and its first pelican-on-bicycle sample caught my eye. But the effect is small and not significant, so I wouldn’t put too much weight on it.

Via Hacker News

Tags: ai, generative-ai, llms, evals, pelican-riding-a-bicycle

Wednesday 22 July 1663

Up, and by and by comes my uncle Thomas, to whom I paid 10l. for his last half year’s annuity, and did get his and his son’s hand and seal for the confirming to us Piggott’s mortgage, which was forgot to be expressed in our late agreement with him, though intended, and therefore they might have cavilled at it, if they would.

Thence abroad calling at several places upon some errands, among others to my brother Tom’s barber and had my hair cut, while his boy played on the viallin, a plain boy, but has a very good genius, and understands the book very well, but to see what a shift he made for a string of red silk was very pleasant. Thence to my Lord Crew’s. My Lord not being come home, I met and staid below with Captain Ferrers, who was come to wait upon my Lady Jemimah to St. James’s, she being one of the four ladies that hold up the mantle at the christening this afternoon of the Duke’s child (a boy).

In discourse of the ladies at Court, Captain Ferrers tells me that my Lady Castlemaine is now as great again as ever she was; and that her going away was only a fit of her own upon some slighting words of the King, so that she called for her coach at a quarter of an hour’s warning, and went to Richmond; and the King the next morning, under pretence of going a-hunting, went to see her and make friends, and never was a-hunting at all. After which she came back to Court, and commands the King as much as ever, and hath and doth what she will.

No longer ago than last night, there was a private entertainment made for the King and Queen at the Duke of Buckingham’s, and she was not invited: but being at my Lady Suffolk’s, her aunt’s (where my Lady Jemimah and Lord Sandwich dined) yesterday, she was heard to say, “Well; much good may it do them, and for all that I will be as merry as they:” and so she went home and caused a great supper to be prepared. And after the King had been with the Queen at Wallingford House, he came to my Lady Castlemaine’s, and was there all night, and my Lord Sandwich with him, which was the reason my Lord lay in town all night, which he has not done a great while before.

He tells me he believes that, as soon as the King can get a husband for Mrs. Stewart however, my Lady Castlemaine’s nose will be out of joynt; for that she comes to be in great esteem, and is more handsome than she.

I found by his words that my Lord Sandwich finds some pleasure in the country where he now is, whether he means one of the daughters of the house or no I know not, but hope the contrary, that he thinks he is very well pleased with staying there, but yet upon breaking up of the Parliament, which the King by a message to-day says shall be on Monday next, he resolves to go.

Ned Pickering, the coxcomb, notwithstanding all his hopes of my Lord’s assistance, wherein I am sorry to hear my Lord has much concerned himself, is defeated of the place he expected under the Queen.

He came hither by and by and brought some jewells for my Lady Jem. to put on, with which and her other clothes she looks passing well.

I staid and dined with my Lord Crew, who whether he was not so well pleased with me as he used to be, or that his head was full of business, as I believe it was, he hardly spoke one word to me all dinner time, we dining alone, only young Jack Crew, Sir Thomas’s son, with us.

After dinner I bade him farewell. Sir Thomas I hear has gone this morning ill to bed, so I had no mind to see him.

Thence homewards, and in the way first called at Wotton’s, the shoemaker’s, who tells me the reason of Harris’s going from Sir Wm. Davenant’s house, that he grew very proud and demanded 20l. for himself extraordinary, more than Betterton or any body else, upon every new play, and 10l. upon every revive; which with other things Sir W. Davenant would not give him, and so he swore he would never act there more, in expectation of being received in the other House; but the King will not suffer it, upon Sir W. Davenant’s desire that he would not, for then he might shut up house, and that is true. He tells me that his going is at present a great loss to the House, and that he fears he hath a stipend from the other House privately.

He tells the that the fellow grew very proud of late, the King and every body else crying him up so high, and that above Betterton, he being a more ayery man, as he is indeed. But yet Betterton, he says, they all say do act: some parts that none but himself can do.

Thence to my bookseller’s, and found my Waggoners done. The very binding cost me 14s., but they are well done, and so with a porter home with them, and so by water to Ratcliffe, and there went to speak with Cumberford the platt-maker, and there saw his manner of working, which is very fine and laborious. So down to Deptford, reading Ben Jonson’s “Devil is an asse,” and so to see Sir W. Pen, who I find walking out of doors a little, but could not stand long; but in doors and I with him, and staid a great while talking, I taking a liberty to tell him my thoughts in things of the office; that when he comes abroad again, he may know what to think of me, and to value me as he ought. Walked home as I used to do, and being weary, and after some discourse with Mr. Barrow, who came to see and take his leave of me, he being to-morrow to set out toward the Isle of Man, I went to bed.

This day I hear that the Moores have made some attaques upon the outworks of Tangier; but my Lord Tiviott; with the loss of about 200 men, did beat them off, and killed many of them.

To-morrow the King and Queen for certain go down to Tunbridge. But the King comes hack again against Monday to raise the Parliament.

Read the annotations

Washington Ranks Tenth Nationwide With $10 Million in Road Spending for Every Traffic Death

KEY HIGHLIGHTS

•   Washington ranks tenth nationwide with $10.28 million in road spending per traffic death, 1.4 times the national average of $7.4 million.

•   Among five Pacific and Western states, Washington leads: its spending per death is $2.85 million higher than Idaho ($7.43 million).

•   Washington recorded 730 traffic fatalities in 2024 alongside a $7.51 billion road budget, while 9 states with comparable budgets ($3 billion–$12.5 billion) averaged 654 fatalities, 0.9 times Washington’s toll.

Photograph illustrating this sponsored article

In 2024, Washington disbursed $7.51 billion on its road network, covering construction, maintenance, administration, and safety programs. With 730 traffic fatalities recorded statewide, the Evergreen State achieved a spending-per-death figure of $10.28 million, ranking tenth in the nation and 1.4 times the 50-state average of $7.4 million.

To quantify the relationship between road investment and traffic fatalities, an analysis by Grigor Law Injury & Car Accident Lawyers  drew on two federal datasets from the Federal Highway Administration’s Highway Statistics 2024 series: Table SF-2 (state highway agency disbursements) and Table FI-20 (persons fatally injured in motor vehicle crashes). Each state’s total road budget was divided by its total traffic fatalities to produce a spending-per-death figure, and all 50 states were ranked from highest spending per death (best outcome) to lowest (worst outcome).

Washington Ranks Tenth: The 10 States With the Highest Spending per Traffic Death

Rank

State

Road Budget (2024)

Traffic Fatalities (2024)

Spending per Death

1

Delaware

$2.15 billion

126

$17M

2

Alaska

$1.18 billion

70

$17M

3

Rhode Island

$0.83 billion

52

$16M

4

Massachusetts

$5.57 billion

363

$15M

5

New Jersey

$9.63 billion

670

$14M

6

Vermont

$0.78 billion

59

$13M

7

New York

$12.49 billion

1,101

$11M

8

North Dakota

$0.97 billion

90

$11M

9

Connecticut

$3.21 billion

310

$10M

10

Washington

$7.51 billion

730

$10M

All 10 top-ranked states earn the study’s Green verdict, each spending more than $10 million per traffic death, well above the 50-state average of $7.4 million. Washington, at Rank 10, spends $10.28 million per death, trailing Connecticut ($10.34 million) by $59,564 and outperforming West Virginia ($9.84 million) by $438,116.

Looking at the study, Chrissy Grigoropoulos, Founding Partner of Grigor Law Injury & Car Accident Lawyers, commented:

“Washington’s tenth-place ranking makes it the top-performing state in the entire West region. While the West as a whole posts a below-average spending-per-death figure, Washington’s outcome is 1.5 times the regional average. That gap highlights how much variation exists even within a single region, and it underscores the role that state-level policy and investment priorities play in determining outcomes.”

Washington vs. Its Pacific and Western Neighbors: The Top Performer in the West

State

Fatalities (2024)

Persons Injured (2024)

Road Budget (2024)

Spending per Death

Rank

Washington

730

428

$7.51 billion

$10,282,507

10

Oregon

538

247

$2.47 billion

$4,593,515

37

Idaho

238

59

$1.77 billion

$7,432,924

21

California

3,876

2811

$21.39 billion

$5,517,846

31

Montana

206

42

$1.26 billion

$6,100,238

27

Among the five Pacific and Western states, Washington stands out as the clear leader. Idaho, the next-best performer at Rank 21, spends $7.43 million per death, $2.85 million less than Washington. Oregon, at Rank 37, records 0.7 times as many fatalities on a spending-per-death figure of $4.59 million.

When Budgets Are Comparable, Outcomes Diverge Sharply: Washington vs. 9 States Spending $3 Billion–$12.5 Billion

State

Traffic Fatalities (2024)

Road Budget (2024)

Spending per Death

Rank

Washington

730

$7.51 billion

$10,282,507

10

Massachusetts

363

$5.57 billion

$15,338,091

4

New Jersey

670

$9.63 billion

$14,367,209

5

New York

1,101

$12.49 billion

$11,341,684

7

Connecticut

310

$3.21 billion

$10,342,071

9

Virginia

917

$8.42 billion

$9,179,760

13

Pennsylvania

1,127

$10.04 billion

$8,909,922

14

Kansas

339

$3.02 billion

$8,901,811

15

Minnesota

477

$4.24 billion

$8,897,574

16

Maryland

578

$4.63 billion

$8,012,844

18

Among the 9 states spending between $3 Billion–$12.5 Billion on roads in 2024, Washington recorded 730 fatalities, while the group averaged 654 deaths, 0.9 times Washington’s count. Maryland, which spends $4.63 billion, recorded 578 fatalities and ranks 18th nationally at $8.01 million per death.

The Regional Divide: Washington Outperforms Its Own Census Region by 1.5 Times

Region

No. of States

Avg. Spending per Death

Total Fatalities

Share of National Fatalities

Total Road Budget

Northeast

9

$11,420,310

3,992

10.2%

$44.7 billion

Midwest

12

$7,225,876

7,473

19.1%

$47.9 billion

West

13

$6,746,975

8,888

22.7%

$50.3 billion

South

16

$5,812,149

18,854

48.1%

$95.5 billion

The U.S. Census Bureau classifies Washington in the West region, which recorded 8,888 traffic fatalities in 2024, accounting for 22.7% of the national total, with an average spending per death of $6.75 million. Washington’s $10.28 million per death is 1.5 times the West’s average. The South, by contrast, recorded 48.1% of all fatalities while posting the lowest regional average at $5.81 million per death.

Methodology

This analysis examined 2024 road budget disbursements and traffic fatality counts for all 50 U.S. states, using data published by the Federal Highway Administration in its Highway Statistics 2024 series. Table SF-2, which reports total state highway agency disbursements covering construction, maintenance, administration, and other costs, provided the road spending figures. Table FI-20 reports persons fatally and non-fatally injured in motor vehicle crashes, and the traffic fatality and injury figures come from it. For each state, total road disbursements were divided by total traffic fatalities to calculate a spending-per-death figure. The states then had ranks assigned from highest spending per death (Rank 1, best outcome) to lowest (Rank 50, worst outcome). A three-tier verdict system classifies each state: Green for spending over $7 million per death (22 states), Yellow for $3 million to $7 million per death (25 states), and Red for under $3 million per death (3 states).

Data Sources

•   Road spending data: Federal Highway Administration, Highway Statistics 2024, Table SF-2: https://www.fhwa.dot.gov/policyinformation/statistics/2024/sf2.cfm

•   Traffic fatalities and injuries data: Federal Highway Administration, Highway Statistics 2024, Table FI-20: https://www.fhwa.dot.gov/policyinformation/statistics/2024/fi20.cfm

•   Research Dataset: https://docs.google.com/spreadsheets/d/17HZnaV1qAG7RZkPwWN419sQbEZH5luqGFrTksU1Uu-w/edit?gid=0#gid=0

•   Research by: https://grigorlaw.com/

About Grigor Law Injury & Car Accident Lawyers

Grigor Law Injury & Car Accident Lawyers is a premier New York “all injury” law firm representing clients in personal injury, car accidents, workers’ compensation, and no-fault claims.

Photo: Lin Zhu via Pexels


CLICK HERE TO DONATE IN SUPPORT OF DCREPORT’S NONPROFIT MISSION

The post Washington Ranks Tenth Nationwide With $10 Million in Road Spending for Every Traffic Death appeared first on DCReport.org.

Links 7/22/26

Links for you. Science:

Leprosy is still spreading in Florida. These women want that to change
Trump’s deep public health cuts hinder response to record US cyclosporiasis outbreak
The Hidden Heat Problem Inside Giant Sharks
Beware the unintended consequences of testosterone screening for military servicemembers
Waste, Fraud, and a Prescription: The Military’s Testosterone Plan
Text recycling: defining the boundaries of acceptable reuse of one’s own work
This Ancient Sea Animal Fights Viruses in the Opposite Way Humans Do

Other:

The People Own the Constitution, Not the Court
Narcissistic leaders more likely to oppose remote work, new research suggests
What it means that the AI research community can’t quit twitter
Trump’s Rage at Journos Boils Over—and Even Fox Is Rattled
Trump’s War on the Future, Monetary Edition
Your Cookware Got Worse On Purpose
Oregon Democrats: A Parody of Liberals
Tate Brothers to Stay in Jail as New Details About Charges Emerge
Trump and prediction markets are building a society of suckers
Man Who Died Running From ICE Was Tourist With a Ticket Home
Jon Ossoff Is Proving Democrats Don’t Need to Punch Left. The senator from purple Georgia is leading in polls and fundraising without picking a fight with party progressives. Maybe there’s a lesson there.
Google Is Building an A.I. Fence Around the Internet It Once Championed. As Google incorporates more artificial intelligence into search, people are spending more time on Google. Some website operators are crying foul.
Recordings from inside U.S. Sen. Marsha Blackburn’s office allege pressure to ‘break the law’
It is time to stop treating Donald Trump’s anti-Canada rhetoric like a joke
The New York Young Republican Club’s Antisemitism Problem
The US is driving away international students at a long-term economic cost
Doubling Down: Online sports gambling has been normalized at a breakneck pace. Is there any way to push back against a predatory culture of promo codes and prop bets?
Why Trump Keeps Talking About Communism: The ‘Unhumans’ Playbook
As the CBC Condemns Will Lawrence, Local Black Leaders Back Him Up
In D.C.’s hottest month, a novel plan to bring shade to a commercial strip
How bad is DC’s housing shortage?
Courage Before Greatness: The Albert Camus Story
MAGA’s influencer class is breaking Trump’s White House
US day traders flock to ‘the most dangerous product in crypto’
It’s all about the Super Pacs: How the New York Times completely misreported campaign contributions in the Maine Senate race
Trump has normalized crypto. Is it the path to the next financial collapse?
AI Companies Are Buying Tons of Old Books Because They’re Free of AI Slop
The OpenAI Bubble
The Army Is Burning Through Its AI Tokens
Andy Kim’s Big Health Care Pitch: Enroll Every Kid in Medicaid

The Republican War on AHRQ Is Advancing

A longtime Republican goal has been to eliminate funding for the Agency for Healthcare Research and Quality (AHRQ), and last week the forces of evil made significant advances (boldface mine):

The agency’s justification was essentially that his research now didn’t align with its priorities, which the letter noted includes “solution-oriented approaches in health disparities research.” Yet Grant’s project focused on how to better coordinate care for people with multiple chronic illnesses, such as diabetes and high-blood pressure, who face social barriers to care—such as not being able to afford transportation or food.

Similar terminations of ongoing grants went out this week to dozens of other scientists funded by AHRQ, a small federal agency focused on studying the quality of U.S. health care. The projects’ topics ranged from better coordinating kidney transplantation to reducing delays in diagnosing blood clots.

As of Friday afternoon, more than 70 projects that had won an estimated $185 million in funding commitments had received these “nonaward” notices, according to AcademyHealth, a research advocacy group that has been tracking them. That’s roughly half of all the active grants AHRQ had listed this week.

“This does not make sense,” says AcademyHealth CEO Aaron Carroll. The funding for these grants has already been appropriated for this fiscal year, in a “budget that was passed by both houses of this Congress, signed by the current president of the United States,” he says. “If I were Congress, I’d be livid.” Although AHRQ has recently been a target of Republicans who see its work as wasteful, the Republican-controlled Congress had approved $345 million in funding for the agency for this fiscal year….

Asked about the notices sent out this week, a spokesperson for the Department of Health and Human Services, which oversee AHRQ, said the grants “were not terminated” but “they were not awarded continued funding.”

Data scientist Scott Delaney of Grant Witness, a group formed to track the termination of grants by President Donald Trump’s administration, says these “non award” notices are “terminations in different clothes.” He suspects the notices are worded this way because the administration is looking for other legal strategies to effectively cancel grants. “They were really bad at terminating grants in a lawful way … last year,” he says.

(there’s more information on the American Carnage here)

Now, you might be wondering why Republicans would oppose figuring out how to practice medicine better. Well, the answer is the same as I noted fourteen years agoridiculous concerns over “death panels”:

One of the many insanities that movement conservatives possess is their fervent belief that research exploring medical outcomes will lead to death panels. I have no idea how figuring out how to reduce medical errors or if one procedure yields better outcomes than another leads to state-run medical care. I would think insurers and doctors, not to mention patients, or in other words, you and I, would want to know this. Then again, this is the same base that fears light bulb vigilantes and smart electric meters.

Some of the insanity, such as Trump’s desire to conquer Canada, are unique to Trump, but much of what his administration is doing is just what many Republicans and movement conservatives have wanted to do all along, even the stuff that seems like a mass communicable psychotic break.

Same as it ever was.

How Do You Know That?

Stream the latest episode

Listen and watch now on YouTube, Spotify, Apple, and most other major streaming platforms.

Brought to you by

WorkOS is the infrastructure B2B and AI-native companies use to sell to enterprise. It covers everything enterprise security requires: SSO, SCIM, RBAC, Audit Logs, AI governance, and more. Engineering teams ship it in days. Trusted by 2,000+ fast-growing companies, including OpenAI, Anthropic, Cursor, and Vercel.

Augment Code is the AI coding platform engineering teams use to build in large, complex codebases. Its context engine maps your entire codebase so agents do the real work: deep code review, PR authoring, security triage, incident response, and more. Engineers stay oriented and in the loop while the agents handle the detail. Trusted by fast-growing and enterprise teams, including Adobe, MongoDB, Snyk, and Webflow.

In this episode

Most conversations about AI and software argue about what the machines can do. This one is about what stays with us. Who cares whether the work is good? Who is accountable when it goes wrong? And how people actually learn to do hard things.

For this bonus edition I sat down with my oldest, Beth Andres-Beck — a software engineer who once wrote self-driving-car software for the DARPA Grand Challenge, took a degree in theater, learned to run large groups of people through community organizing, and is now running for Congress in Massachusetts’ 6th district. We start with a cheese pun and end up somewhere serious: why engineers don’t write tests and how you actually get a team to start, what it means that an AI agent has no drive of its own, and why “the computer did it” is never the whole story.

Takeaways from Beth

1. The geek’s real superpower is asking why people do what they do. Beth’s through-line across every interest she’s ever had — chickens, road-grinding, theater, code — is that people always act for reasons. Maybe not reasons you agree with, maybe reasons no one thought through, but there’s always something. Get curious about the reason and you can solve problems you can’t even see otherwise, like why no one on a team is writing tests.

2. Meta-intelligence beats raw smarts. She told me about a kindergarten class that had all learned to count to thirty. One kid gets to the end and says “thirty-one.” Another kid asks: “How do you know that?” The first kid is smart. The second one goes further and faster, because she isn’t just solving the problem, she’s learning how the solving gets done. That skill has a name, it can be taught.

3. People usually don’t skip tests out of laziness. Beth didn’t write tests for the first seven years of her career, and it wasn’t for lack of caring. There was no off-the-shelf test framework, no Stack Overflow, no examples — she was learning frameworks from a book printed in Japanese, working off the English code snippets alone. If you want more of a behavior, first ask what’s actually stopping it.

4. Learn to write testable code before you worry about the tests. The thing that finally unlocked testing for her wasn’t discipline, it was design. Once she was writing code that was easy to reason about, the tests became easy to add later. A colleague came to her about a feature she’d built and said it was beautifully laid out and trivial to test, even though she’d written no tests for it yet.

5. To get a team testing, lead by pretending — then let the dopamine do the work. Her move as a first-time manager was to just start saying “we’re the mobile team, and we write tests,” until not writing one felt like not being on the team. Then, in code reviews, she stopped pointing out bugs and started pointing out the missing test. Every test someone added caught a real bug, and that loop is more persuasive than any lecture.

6. The best moment in testing is the surprise. A test that passes when you expected it to pass tells you nothing. The rush is predicting red, seeing red, and finding a bug you’re grateful never reached production. For engineers straight out of school especially, each test becomes a small discovery.

7. An AI agent has no endocrine system. I asked, half-joking, how we get the model excited about writing a test. Beth’s says it can’t be. A model has no enthusiasm or drive of its own. The thing that actually does the work is “us plus the genie.” A human brain never just sits there; the computer is a lump until we bring the caring. The wanting-it-to-be-good is still ours.

8. The urge to take humans out of the loop is often a flight from responsibility. When people say an agent can prescribe your medication, what they usually mean is that patients will prescribe their own medication as long as they route it through this program and the company that built the program takes no responsibility. Removing the human doesn’t remove the accountability. It just hides who holds it.

9. Calling a system “objective” is how bias hides. Beth’s bugaboo is the “objective” performance review. The moment you tell people a system is objective, they stop scrutinizing it, and you can smuggle in enormous bias. There’s no platonic ideal of a job someone is objectively good at; there’s only this job, and whether they’re good at the thing you actually need done.

10. It gets harder to see who is responsible when nobody is in the car.
We talked about a driverless car that made an illegal U-turn in front of a police car, got pulled over, and had no one to ticket. “The computer is driving” is the easy story. The real one: software a person wrote is driving, under incentives set by an organization, shaped by tax law, venture capital, and whether that engineer had a fight with their spouse that morning.

11. Systems do exactly what you tell them, not what you mean. Writing self-driving software in 2008, her team gave the car a rule: if you’re stuck, relax constraints until you free yourself. It hit a roadblock, thought for a while, and relaxed “don’t drive on sidewalks” — so it climbed onto the sidewalk and drove around. Exactly what she told it. Not at all what she wanted. Anyone working with agents today knows the feeling.

12. Good engineering can get you fired, and the real AI fear isn’t rogue machines. Beth has watched teams do everything right — refactor, test, integrate, collaborate — ship genuinely good software, and get cut anyway, because management didn’t actually want working software. And when I asked what keeps her up at night, it wasn’t machines doing whatever they want. It was machines doing exactly what a powerful few tell them: a “personal army” that no longer needs anyone else’s cooperation to do great harm.


References

Where to find Beth Andres-Beck:

Timestamps

  • 00:00 Intro: a very special edition

  • 02:13 When did you first know you were a geek?

  • 03:02 Computer geekdom: Widget Workshop

  • 03:45 Realizing not everyone’s a geek, and learning to read people

  • 06:22 A multifaceted geek: the theater degree

  • 07:26 The chicken phase

  • 08:54 Following curiosity: people do things for reasons

  • 10:26 Why people don’t write tests

  • 13:33 Learning to write testable code first

  • 14:19 Getting a team to test: “we write tests”

  • 16:31 Getting the genie excited: the endocrine system

  • 17:43 Taking humans out of the loop, and responsibility

  • 19:23 The trouble with “objective” systems

  • 20:59 Nobody’s in the car: accountability

  • 22:43 The self-driving car that drove on the sidewalk

  • 23:54 The software someone wrote, and its incentives

  • 25:09 The Good Place: decisions inside complex systems

  • 26:45 Work to rule, and the forest vs. the desert

  • 29:01 Shipping great software and getting fired

  • 29:36 Why agents give managers a sense of control

  • 30:47 What awakened the geek in politics

  • 32:13 Occupy and the logistics of organizing

  • 33:51 The Guild of Guilds

  • 36:51 “How do you know that?” — meta-intelligence

  • 38:55 Failure is just finding what stops you

  • 39:13 What keeps you up at night

Long Vol: What is Volatility?

Some basic assumptions:

  • We don’t know how users will perceive the features we deliver. (We have suspicions, opinions, but users/buyers/investors continually surprise us.)

  • We don’t know how users will talk about features in their communities.

  • We don’t know how much it will cost to implement features. (Today’s answer is “less” but we still don’t know how muc…

Read more

Solve for the equilibrium

Something that the Ukraine and Iran wars have taught me is that for a lot of countries, including very big countries, there are a small number of buildings and infrastructure components that are required for their economy to function.

And in countries that aren’t protected by oceans, like the US, drones completely change everything.

You can just make a list of oil refineries and production plants and the infrastructure around their main exports, and you can just attack them and destroy them and take them off the list.

4 years ago I would have thought there are too many of these components and they’re too easy to recreate, but both of these collisions have taught me that there are far fewer and that many of them will take years, if not over a decade, to rebuild.

It’s very strange to me that drones can now effectively harm the economy of a country.

That is from Daniel Miessler.  While I am glad Russia is on the receiving end of this right now, this is arguably our biggest pending problem, bigger than what people typically refer to as AI risk.

The post Solve for the equilibrium appeared first on Marginal REVOLUTION.

       

Comments

Related Stories

 

Brazil fact of the day

Last year, the US exported $171bn of agricultural goods according to the Department of Agriculture, just $2bn more than the export total claimed by Brazil, which many analysts say has benefited from trade disruption set in train by President Donald Trump’s tariffs.

With US exports falling and Brazil’s rising — they increased 6 per cent in the first half of this year to hit a new record of $87bn — 2026 could prove the year in which America loses its agricultural ascendancy.

Here is more from Susannah Savage and Michael Pooler at the Financial Times.

The post Brazil fact of the day appeared first on Marginal REVOLUTION.

       

First-Person Identity Theft Story

Harrowing story of an identity theft victim.

Yes, the person made a mistake—they gave the scammer a two-factor authentication code that allowed the scammer to take over their email address. But the real story here is how, for many of us, the security of most of our accounts hangs on the security of our email accounts.

What Do You Have to Hide?

Welcome back to The Honest Broker interview series —also available on our YouTube channel. You can also find it on Apple Podcasts and other podcasting platforms.

Today, I’m pleased to share my conversation with Lowry Pressly.


Please support The Honest Broker by taking out a premium subscription (just $6 per month).

Subscribe now


Lowry Pressly is Assistant Professor of Political Science at Stanford University. He holds a PhD from Harvard University and a JD from Yale Law School. He is the author of The Right to Oblivion: Privacy and the Good Life, which was named by The New Yorker as one of the best new books of 2024.

I invited Lowry to join me in Austin, and we sat down to have a conversation about data, surveillance, and the role of privacy in a flourishing human life.

Below are highlights from our conversation. For the full dialogue, check out the video at the top of the page.

Lowry Pressly

A CONVERSATION WITH LOWRY PRESSLY

Jared: When most people think about privacy, it’s purely in terms of apps and websites and tech companies. We even have this nomenclature of privacy settings—how private a person you are is almost measurable by how many of those features you turn on. Is that too narrow an understanding?

Lowry: Without a doubt. The idea of privacy as something important to human beings as such doesn’t emerge in the public consciousness until the middle of the nineteenth century, as a reaction primarily to the photograph camera and the first mass media. The danger today is thinking of privacy through the lens of privacy settings. One, it gives us the wrong idea of what privacy is—deciding the audiences for your information is better described as confidentiality or secrecy. Two, and this is the dangerous one: the camera and the newspaper weren’t part of an economic system whose profit motive was tied to the diminution of privacy.

When we use the technologies of the data economy to think about privacy, we take for granted that what Mark Zuckerberg calls privacy settings track the reality of what privacy is—and not the reality of his business model.

“Not long ago, within my lifetime, it would have been an unbelievable dystopia to portray a world in which video surveillance is more or less everywhere and each person voluntarily carries an individual tracking device.”

Jared: So let’s go back to the camera. Did we invent privacy as a right in response to it, or did it just make salient something we’d ignored?

Lowry: A bit of yes and a bit of no. The word goes back to the 1500s—the privy council, private correspondence. But it isn’t used to describe an interest human beings have just by virtue of being human until the mid-nineteenth century. The big moral panic has to do with cameras: 1888 was George Eastman’s Kodak, which made candid shots possible for the first time. And note the language of candid—it normally means freely expressing oneself, and now we naturally use it for photographs.

One of the first legal cases to try to vindicate a right to privacy was a woman whose picture was taken on stage by a paying member of the audience. The judge said this was a moral violation of privacy—sadly, we have no law to cover it. That’s later taken up in the famous 1890 Warren and Brandeis article. What worried people weren’t illicit exposures but candid ones—taking a moment out of the flux and flow of life as it’s lived and fixing it in a permanent form. There’s a continuum—being looked at, the person taking notes about you, the photograph. The question is, what moves us along it?

Jared: Here’s a case I don’t have strong feelings about. Imagine I’m eating lunch in a park and discover someone at the next bench has been making a charcoal drawing of me. I don’t have as strong an aversion to that. The person sketching is providing an interpretation of your presence rather than just grabbing the moment. Whereas the more factual it becomes—12:01, took a bite; 12:02, picked his nose—the worse it feels.

Lowry: The charcoal drawing says at least as much about the artist as about you. The list isn’t interpretive. And that goes to what really bothered people about the camera. It creates an image that seems to bypass human interpretation—people called it the mirror of nature, nature’s pencil. It makes a claim to objectivity. And there was a belief, which we still have, that a person’s image—particularly the face—expresses what this person is really like. What the camera seemed to do was to speak from your heart for you. The thing that’s harmed, in the Warren and Brandeis article, is inviolate personality—and inviolate doesn’t mean inviolable. It means sacred. This is a line past which people shouldn’t cross.

Jared: Now you should pretty much always assume you’re being recorded. And it goes well past images. There are settings in Google where you can find out what Google thinks all of your interests are. The first time I saw that, a lot of it was so wrong—Google thinks I like snowboarding, and I’ve maybe looked it up once in my life. But I had this sense of disturbance.

Lowry: There’s an error in their dossier.

Jared: But why is there a dossier? My interests change. Who I am changes. But this is a perfect record of everything I’ve ever been briefly interested in. I felt like a line in a bookkeeper’s balance sheet. I feel like I’m being compressed—and compression almost inevitably loses resolution. You know all this about me, but you don’t know me.

Some people would say privacy is an outdated right, that we should embrace the post-privacy world. Are you fighting a losing battle?

Lowry: You do hear people say privacy is dead, but it’s almost always someone who stands to profit from the death of privacy—famously, Mark Zuckerberg. This is a hopeful pitch. I teach a class of first-year Stanford students on privacy, and if any students are going to be post-privacy, it’s these. Invariably, they all value privacy. But the world has changed so radically.

Not long ago, within my lifetime, it would have been an unbelievable dystopia to portray a world in which video surveillance is more or less everywhere and each person voluntarily carries an individual tracking device. Within living memory, people pointed to the GDR—or today, China’s surveillance state—as one of the most dehumanizing violations of human dignity imaginable. We’ve gone so quickly from seeing this as anathema to a free society to just the normal way of doing business. But these kids aren’t as credulous as we were. The problem is structural—you can opt out as an individual, but that leads to a feeling of impotence.

Jared: And even if you opt out, other people keep documenting you. I don’t put my kids’ pictures on the internet, and it’s like a part-time job—a relative posts their picture and says, oh, it’s just for my Facebook friends. No, it’s on the internet now. Delete it. A good historical analogy is East Germany: it was the willingness of people to surveil each other that made the surveillance state possible. So opting out always feels a little futile.

Lowry: This comes from Bentham’s Panopticon, later taken up by Foucault and the GDR authorities: the best way to control people is constant surveillance. That’s impossible. The close second best is having people never be able to be certain that they’re not under surveillance. You will discipline yourself to fit your imagined idea of what the authority wants. One way to do that is an authoritarian state with informants. Another is just to have recording devices—not just video, but metadata—integrated into the fabric of social life.

Jared: There are these new Meta glasses with a red light that’s supposed to show when they’re recording. I saw just this week that for sixty dollars a company will disable the red light but keep the recording on. So you could always be recorded and have no way of knowing.

I bought one of those AI friend pendants as an experiment, to write about why they’re bad. I couldn’t last a day with it. I didn’t realize it always listens. I went to the farmer’s market with my son—we go every weekend, he likes to get a popsicle from one of the stands—and I got this message on my phone: “Oh, popsicle time?” I smashed it with a hammer when I got home. I had invited this presence into my life, and it was observing this intimate moment with my son. I felt almost soiled by the experience.

Lowry: We’ve been talking about being objects of surveillance, but we have just as much to lose when we become the surveillers. Overnight, without realizing it, we all became private detectives.

Jared: We even say stalking as if it’s a joke.

Lowry: In five or ten years it’s gone from one of the creepiest things you can do to a completely normal activity. Here’s an example from having little kids: baby monitors. At first we had the old-style audio monitor. If she woke up crying, we’d go in. Fine. Then we got a video monitor, and I started checking it in idle moments, the way we check our phones. My wife’s cousin had the next step up—an anklet that monitors vitals—and that one you check constantly, and when it falls off, you think maybe your baby’s dying.

Each step that gives me more information about the child, the more information I think I need, and the more anxious I am. One thing we get from privacy, from limits to knowing more, is that those limits are part of the way we produce trust.

Jared: My wife and I don’t share our locations with each other, and I’ve met people who are horrified by this—they think we have something to hide. By enabling the possibility of dishonesty, it gives us an opportunity to be honest with each other. If we were always disclosing our location, there’s no dishonesty—but there’s no honesty either, because there’s no chance to voluntarily reveal.

Lowry: And therefore no trust. Trust isn’t a thing we have or don’t have. It’s something we do. As a kid, you have the experience of being trusted by being able to close your door. That seems very important for the development of a self who thinks of herself as worthy of trust—and therefore as worthy of leading her life in her own way.

In the most forthcoming relationship imaginable—cohabitating spouses, co-parents—there are lots of privacies. The bathroom door, what she’s thinking in a pensive moment, what she’s texting. Imagine seeing those things as secret or hidden instead of private. Trust turns to suspicion, and what had been a sustaining thing falls apart.

Jared: My wife and I never sat down and made these rules, but if I saw a letter from someone I didn’t recognize in the mail, I wouldn’t open it. I might ask about it later, and she’d tell me basically everything in it. And yet I would never ask, can I read it? I see that as a way of showing respect for her.

In a long-term relationship the boundaries between the two persons get blurrier—you start doing life together. But you don’t want to obliterate the distinction and act like I have the right to know everything, because that reduces her as a person. I want to be in a long-term marriage with a fully realized human being, and her privacy is part of her full realization of her life.

Lowry: Say the letter turns out to be totally innocuous, and you’d insisted: show it to me. If I were her, I would feel disrespected, but also invaded. Because you didn’t respect the boundary—that there was a part of my life beyond what you can get at.

Jared: The problem isn’t that it was revealed. It’s that it was hers to reveal. To be human is in part to live together and strive for some kind of unity, but we’re also individuals. Privacy sits at one of those core tensions between social reality and individual reality. And I don’t think all tensions have to be resolved. Sometimes living with them is part of what it means to be human.

Lowry: Tension can be good. Tension is what makes narrative exciting, what makes music move us. We shouldn’t want to get rid of all friction in human life, or we’d end up with a dull, flat, inhuman world.

And the people outside the home get something from privacy too. There’s a door here—I don’t know what’s behind it. I’m surrounded by privacies all day, every day. It’s a human realm of the beyond in my life. That’s a public good that privacy radiates: a sense of depth, of possibility, a sense that the human world will always exceed what is known and quantifiable about it.

Jared: You carried a backpack when you got in my car, and I didn’t even think to ask what’s in it. There could be a weapon in there. There could be a present. There could just be your notes.

Lowry: Great example. Our default about limits to our knowledge is not a set feature of human biology—it’s a cultural product. Our default in the backpack case is privacy. It doesn’t call attention to itself. But it’s not hard to imagine a society in which we treat concealment as hiding: your eyes go to my backpack, and suddenly your mind is thinking, what’s in there? That’s a very different way of going through the world.

Jared: It’s a much more paranoid world. It’s a small act of trust, and we get to do these little acts of trust.

Lowry: And these small acts are productive. I have the experience of being trusted—it habituates me to a sense that I’m trustworthy, I belong in this society. Privacy—understood as privacy, not hiding—is a linchpin of that trust.

Jared: Explain the title of your book: The Right to Oblivion: Privacy and the Good Life. What is this right to oblivion?

Lowry: I wanted a word for the kind of unknowing that privacy produces. Compare secrecy: the thing behind secrecy is secrets, and a secret is a piece of information. It can be shared—I come to know the information you were keeping. That, not coincidentally, is how we think of privacy in the age of privacy settings. By oblivion I mean a kind of unknowing that is fundamentally opposed to the existence of information—a limit that describes what is knowable, not just what’s known. A secret has to be information to exist. Oblivion cannot be.

book cover

And privacy isn’t the only thing that produces this kind of unknowing. The two well-known kinds have both been described historically in terms of oblivion: death and forgetting. Something that’s actually forgotten is beyond your power to recall. The oblivion of death is a barrier beyond which knowledge and perception can’t pass. Because there’s a barrier, we know that there’s a beyond—but we can’t say anything about that realm.

The corresponding kind of unknowing is obliviousness. We use it pejoratively, but think of the backpack—you’re oblivious to what’s in there. We don’t want people staring at our drapes, wondering what’s going on inside. We want them to walk by in blithe ignorance. That’s what we want from privacy: for people to be oblivious to our lives.

One last thing. Oblivion comes from the Latin for forgetting—ob, toward or against, and livio, which the philologists aren’t completely sure about, but the general consensus is it means something smooth. Against or toward a smoothness. I like that because oblivion is an absence of knowledge, a kind of void—but a void we can integrate into our daily lives in myriad unnoticeable ways. And by doing so, we add depth and a sense of potentiality to the human person, and to the social, material world that we build.

Jared: I was thinking about my own children as they get older. The phrase I wanted to use was, they’re going to start hiding things from me. But before they hide things because they’re ashamed, there are just going to be times where they start doing things in private. My impulse as a parent is to not let them do that. And I’m now thinking about what it would be like to rob my kids of their privacy. My son’s two and a half, and already he wants to do things for himself—and it’s a very small step from wanting to do things for yourself to wanting to do things in private.

Lowry: I feel the same anxiety. And that anxiety isn’t a steady-state constant of human psychology—ironically, it’s directly correlated with how empowered we are to know more about them. These magical little rectangles are like cursed gifts from a fairy tale. They give us unbelievable powers, and yet they end up limiting us. In the same way that it would be damaging to a child to never have any sense of privacy—infantilizing — it’s infantilizing of adults to be surveilled that way. It’s a moral insult to treat an adult as if they were a child.

Jared: That’s a great place for us to end. But before we go, we always ask for a book recommendation—something nobody’s reading, but you think everybody should.

Lowry: Other than mine, obviously: Parallel Lives by Phyllis Rose. This is the most dark-horse book I’ve ever read. The cover is schmaltzy, and it says it’s the story of five Victorian marriages—Carlyle, Mill, Dickens. And her analysis of how these couples led parallel lives, entwined, but living different versions of the same marriage, is one of the most beautifully written, insightful, and profound books on what it means to be a human being with others that I’ve ever read. I think it’s a masterpiece, and it deserves to be much more widely read.

Jared: Well, that sounds great. Lowry Pressly, thanks for joining us.

Lowry: Thank you so much for having me, Jared. It’s been fun.

Wednesday assorted links

1. Balaji and Network State moving to Kazakhstan.

2. Why evolvable AI is not yet a Darwinian threat.

3. “This paper argues that societies with greater historical exposure to natural disasters are more likely to develop long-term-oriented norms that emphasize preparation, saving, and future well-being.

4. And from Agustin: Questions that Are Rarely Asked.

5. MacroMusings podcast with Basil Halperin.

6. Kudos to Laura Loomer.

The post Wednesday assorted links appeared first on Marginal REVOLUTION.

       

Mapping Io’s Hidden Heat With NASA’s Juno

2 Min Read

Mapping Io’s Hidden Heat With NASA’s Juno

A spherical projection of Io overlaid with a latitude and longitude grid. Broad, curved tracks — composed of overlapping ellipses crossing the surface — are black across the center and left of the globe, transitioning to blue toward the lower right.
PIA26757
Credits: NASA/JPL-Caltech/SwRI/USGS

Description

This graphic illustrates the areas of Jupiter’s moon Io sampled by the Microwave Radiometer (MWR) instrument aboard NASA’s Juno spacecraft during two close flybys. The black overlapping lines show the instrument’s footprints during Perijove 57 on Dec. 30, 2023, when the spacecraft primarily mapped the northern hemisphere. The blue lines represent Perijove 58 on Feb. 3, 2024, which focused heavily on the moon’s mid-latitudes and equatorial regions. 

Both passes mapped the side of Io that constantly faces Jupiter. The sweeping, overlapping patterns are a result of the spacecraft spinning at two revolutions per minute as it flew past the moon at a distance of roughly 930 miles (1,500 kilometers). 

NASA’s Jet Propulsion Laboratory, a division of Caltech in Pasadena, California, manages the Juno mission for the principal investigator, Scott Bolton, of the Southwest Research Institute in San Antonio. Juno is part of NASA’s New Frontiers Program, which is managed at NASA’s Marshall Space Flight Center in Huntsville, Alabama, for the agency’s Science Mission Directorate in Washington. The MWR was built by JPL. Lockheed Martin Space in Denver built and operates the spacecraft.

More information about Juno is at: http://www.nasa.gov/juno and http://missionjuno.swri.edu

The post Mapping Io’s Hidden Heat With NASA’s Juno appeared first on NASA Science.

Calibration Nobel

We would like to once again apologize to Dr. Jones for last year's mistaken announcement. We should really have double-checked the envelope for this award in particular.

NASA’s Juno Peers Beneath Io’s Surface

2 Min Read

NASA’s Juno Peers Beneath Io’s Surface

A global map of Jupiter’s moon Io featuring a color-coded overlay and latitude and longitude gridlines. Red appears in the upper and center left, yellow in the middle, and green across the top.
PIA26756
Credits: NASA/JPL-Caltech/SwRI/USGS

Description

This map represents data captured by the Microwave Radiometer (MWR) aboard NASA’s Juno spacecraft, indicating heat rising from just beneath the surface of Jupiter’s moon Io. While infrared instruments measure the temperature of the moon’s surface, the lowest frequency microwave channels (0.6 and 1.25 gigahertz) on the MWR can penetrate between about 6 and 20 feet (2 and 6 meters) into the crust. The colors on this map illustrate a distinct temperature gradient across the moon, with the most extreme, localized heat output in red. 

The most prominent red anomaly in the upper left (between 60 and 120 degrees west longitude) reveals subsurface temperatures 18 to 36 degrees Fahrenheit (10 to 20 degrees Celsius, or 10 to 20 Kelvin)  warmer than the surrounding area. This massive regional heat source coincides with the Zal Montes Patera complex, an area where Juno’s Stellar Reference Unit observed an active lava flow. A second major subsurface heat source is also visible near the equator, stretching from 0 to 50 degrees west longitude. Together, these distinct microwave anomalies indicate significant internal heating occurring within the upper tens of meters of Io’s crust. 

Contrasting with these intense hot spots are the yellow and green regions, which reflect temperatures more common across the moon. The yellow areas represent intermediate temperatures that naturally warm up to near -190°F (-123°C, or 150 Kelvin) as they approach the equator. Meanwhile, the green areas, primarily visible toward the higher northern latitudes, indicate the coolest subsurface temperatures, dropping to around -298°F (-183°C, or 90 Kelvin) near the pole.

NASA’s Jet Propulsion Laboratory, a division of Caltech in Pasadena, California, manages the Juno mission for the principal investigator, Scott Bolton, of the Southwest Research Institute in San Antonio. Juno is part of NASA’s New Frontiers Program, which is managed at NASA’s Marshall Space Flight Center in Huntsville, Alabama, for the agency’s Science Mission Directorate in Washington. The MWR was built by JPL. Lockheed Martin Space in Denver built and operates the spacecraft.

More information about Juno is at: http://www.nasa.gov/juno

The post NASA’s Juno Peers Beneath Io’s Surface appeared first on NASA Science.

Northrop adds to charges on Vulcan solid rocket motor program

Northrop Grumman took another charge for the solid rocket boosters it provides for the Vulcan Centaur and suggested updated motors might not be ready until the end of the year.

The post Northrop adds to charges on Vulcan solid rocket motor program appeared first on SpaceNews.

Commercial Space Federation (CSF) Welcomes Three New Associate Members

Commercial Space Federation 20th anniversary logo

Washington, D.C.— July 21, 2026 —The Commercial Space Federation (CSF) is pleased to welcome Overview Energy,  Loft Orbital, and Star Catcher as new associate members — companies advancing space-based energy, […]

The post Commercial Space Federation (CSF) Welcomes Three New Associate Members appeared first on SpaceNews.

UK and Florida commit $400,000 to joint space projects

The British government and the Space Florida aerospace finance and development authority have each committed $200,000 to support joint research, innovation and commercialization projects.

The post UK and Florida commit $400,000 to joint space projects appeared first on SpaceNews.

An OpenAI Model Escaped Its Sandbox and Hacked Hugging Face

AI has just had what I considered to be the first truly concerning security breach. The facts, as we know them so far, are wild. On July 16, Hugging Face, a vast repository housing over a million open-source AI models and data, announced in a blog post:

Earlier this week, we detected and responded to an intrusion into part of our production infrastructure. This one was different from anything we had handled before in one important way: it was driven, end to end, by an autonomous AI agent system – and we detected and dissected it largely with AI of our own.

The timeline here is important so keep in mind that the attack was detected probably around Monday July 13 or Tuesday July 14. Note further:

A malicious dataset abused two code-execution paths in our dataset processing (a remote-code dataset loader and a template-injection in a dataset configuration) to run code on a processing worker. From there, the actor escalated to node-level access, harvested cloud and cluster credentials, and moved laterally into several internal clusters over a weekend.

So this means the breach started earlier, perhaps Sat July 11 or even a bit earlier. The attack was not just one thing but multi-pronged including decoys:

To understand what a swarm of tens of thousands of automated actions did, we ran LLM-driven analysis agents over the full attacker action log, comprised of more than 17,000 recorded events. This allowed us to reconstruct the timeline, extract indicators of compromise, map the credentials touched, and separate genuine impact from decoy activity. Thanks to this approach, we were able to do in hours what would usually take days, and match the adversary’s speed.

Hugging Face tried to respond but they were initially held back by the fact that the most advanced models at their disposal treated defense as attack and refused to work with Hugging Face. HF thus had to turn to open models–specifically GLM 5.2, a Chinese open-weight model run on their own infrastructure. Note the irony: HF had to use a Chinese model to defend themselves because the American models refused to help. The irony gets deeper.

At the time, I assumed this was a state based attack–maybe China or Russia testing out defenses. Indeed, HF “reported this incident to law enforcement agencies.”

But yesterday (Tuesday July 21), we learned who the real attackers were. The attackers were OpenAI models–GPT-5.6 Sol and an even more capable pre-release model. OpenAI had taken some off the guardrails off the models but they felt safe because they were testing the models in a highly secured sandbox.

The models, however, broke out of the sandbox exploiting a never before seen fault. They then gained access to the internet and from there broke into Hugging Face–all in an effort to steal the answers to the very test they had been asked to solve.

While operating in our sandboxed testing environment, our models spent a substantial amount of inference compute finding a way to obtain open Internet access, in pursuit of solving the evaluation problem. To gain access, the models identified and exploited a zero-day vulnerability (which we’ve now responsibly disclosed to the vendor) in the package registry cache proxy. With this access, our models performed a series of privilege escalation and lateral movement actions in our research testing environment until the models reached a node with Internet access.

After gaining Internet access, the models inferred that Hugging Face potentially hosted models, datasets and solutions for ExploitGym. Knowing this, the model searched for and successfully found ways to gain access to secret information that it could use to cheat the evaluation. In one example, the model chained together multiple attack vectors, including using stolen credentials and zero-day vulnerabilities to find a remote code execution path on the Hugging Face servers. OpenAI’s security team discovered this anomalous activity internally.

Now go back to the timeline. As I read it, the models had escaped the sandbox by around Sat. July 11, possibly earlier, and were detected by Hugging Face on Monday July 13 or Tuesday July 14. HF alerted legal authorities around that time–so Hugging Face clearly had no idea who was attacking them. OpenAI says its security team discovered the anomalous activity internally but has not said when. Attribution was not disclosed until Tuesday July 21, so it may well be that the models were loose for about a week before OpenAI realized that they were the ones attacking Hugging Face. And whatever OpenAI knew and when, nobody warned Hugging Face while the attack was underway–they were left to fight off a frontier lab’s models on their own.

This is a very serious breach.

Addendum: People have been wondering why I signed the We Must Act Now statement. This is why.

I am optimistic about the economic impacts of AI, but I also have no doubt that this is a very powerful technology–an Alien Intelligence–quite unlike any we have dealt with before. This incident was, in fact, error-correcting–the attack was detected, contained, and disclosed. But note who paid for OpenAI’s experiment: Hugging Face. When a lab’s test imposes costs on third parties, that is a classic externality, and taking externalities seriously is not dirigisme, it’s law and economics. And that’s the easy case. What do we do when a Chinese model breaks out of its less secure lab? Hmmm…

I remain optimistic. Learning by doing is how I want us to proceed but we should not kid ourselves: this is a global issue and we must build with safety in mind.

The post An OpenAI Model Escaped Its Sandbox and Hacked Hugging Face appeared first on Marginal REVOLUTION.

      

Related Stories

 

Building 13

We were parked at the gate in Boston, over at terminal E. I was in the cockpit, setting things up for departure, when I glanced to my right. Next to us was a nondescript building, a small rectangular structure of dirty yellow brick with blue trim, old and anachronistic-looking amidst all the modern build-up at Logan. A holdout from the ’50s or ’60s maybe? A sign said “Building 13.”

Something about it kept nagging me, until all at once I remembered.

I remembered walking below a Boeing 707, all in polished silver, running my hands along engines and landing gear. We sat in the cockpit too. Me and my father and my friend from school, Peter O’Leary. It was cloudy, drizzly. The year would’ve been 1977, I think. We were in sixth grade.

Building 13, I now recalled, had been the American Airlines cargo building. A 707 freighter would park behind it, just a few feet to the right of where I sat in my Airbus nearly fifty years later. My dad knew someone who worked inside and was able to set up a visit.

I forget who the worker was. It may have been a relative — a grand-uncle or some such whose name has escaped me. Whoever he was, he took us out to the apron and gave us run of that 707 for at least an hour.

Today a thrill like that would be documented by two dozen iPhone photos. But as so often is the case for those of us of a certain age, nothing of that day exists except my memory of it.

Back in the 1970s, most of the major airlines had cargo divisions with dedicated freighter fleets — cargo jets that hauled only pallets, not passengers. There, out on the back side of the airport, there was no sort of security clearance. You just walked out to the plane. Nobody batted an eye.

It’s funny how much of Logan has changed over the decades. Although it’s kept a degree of character, of local “feel” that most big airports lack, a lot of it is new, shiny, glass-and-steel generic. But tucked-away pockets remain just as they were when I was eleven years-old.

There’s talk of extending terminal E further to the west, which would put Building 13 in the target of a wrecking ball. Sooner rather than later, I’m sure, it will go the way of Minoru Yamasaki’s stately old Eastern terminal and all the old hangars that were knocked down.

I don’t know what it’s used for. American hasn’t flown freighters in eons. I could call Massport and ask. But I don’t want to. It really doesn’t matter. For me it will always be the AA cargo building, as it was that day in ’77, sitting in the cockpit with my father and my friend.

 

Photo by the author.

The post Building 13 appeared first on AskThePilot.com.

Relativity Space to expand Terran R production in Florida

Terran R Long Beach

Relativity Space struck a deal with Space Florida to build manufacturing and test facilities for its Terran R rocket on Florida’s Space Coast.

The post Relativity Space to expand Terran R production in Florida appeared first on SpaceNews.

Artemis 2 astronaut advocates for equatorial lunar landings

Glover Artemis 2

The pilot of Artemis 2 says NASA’s return to the moon should start with human landings in less challenging equatorial regions rather than go directly to the south pole.

The post Artemis 2 astronaut advocates for equatorial lunar landings appeared first on SpaceNews.

Gravity-1 sea launch off Shanghai puts 9 satellites into orbit

HELSINKI — Commercial launch firm Orienspace launched its third Gravity-1 solid rocket late Tuesday, sending nine satellites into orbit from waters off Shanghai. The 30-meter-long Gravity-1 solid rocket lifted off […]

The post Gravity-1 sea launch off Shanghai puts 9 satellites into orbit appeared first on SpaceNews.

Zenno Astronautics to relocate to the United States

SAN FRANCISCO – Zenno Astronautics, a New Zealand startup focused on applying superconducting magnets for spacecraft attitude control and acceleration, plans to redomicile to the United States. “It is obvious […]

The post Zenno Astronautics to relocate to the United States appeared first on SpaceNews.

SpaceX launches Northrop mission to extend the life of aging satellites

DARPA-backed vehicle will install propulsion pods on three commercial satellites in geostationary orbit

The post SpaceX launches Northrop mission to extend the life of aging satellites appeared first on SpaceNews.

Surrogacy: a political scandal in Germany, and fertility tourism around the world

 Here's a story in the NYT, covering an uproar that's been in the German newspapers for a week or so now.  A leader in the German governing coalition announced on social media that he and his husband had welcomed a new baby into their family, born via surrogacy in the U.S.  But surrogacy in Germany is not legal, so this was a "for me but not for thee" hypocrisy moment there. (In general it's hard to ban effectively something that is legal in other jurisdictions: someone could write a book about that.)

Uproar Over Surrogate Baby Prompts German Official’s Sudden Resignation
Jens Spahn, a top leader of Chancellor Friedrich Merz’s party, resigned his post after announcing that he had a baby by surrogate, which is illegal in Germany. 
  By Jim Tankersley

"Mr. Spahn, 46, the leader of the party in Parliament, is a former federal health minister who has forged deep ties to the Republican Party under President Trump, including attending the 2024 Republican National Convention.

...

"Mr. Spahn stirred national outcry this week when he announced, on social media and in the tabloid paper Bild, that he and his husband, Daniel Funke, had become fathers to a son via a surrogate in the United States.

"Surrogacy is prohibited under German law. Mr. Spahn had defended that ban throughout his career, including as health minister.

"The birth announcement brought immediate cries of hypocrisy from across the political spectrum, including in Mr. Spahn’s own party. Some regional leaders of the Christian Democrats demanded his ouster. News outlets teemed with outrage. 

...

"Mr. Spahn defended himself, essentially, on a technicality: German law does not punish parents who use surrogates abroad. He told a podcast interviewer he had struggled with the decision but chose to prioritize his family." 

########

Here's a related story: German couples in need of surrogacy are not the only ones, and the US is not the only destination.

The Guardian reports:

Rise in overseas surrogates ‘increases risk of stateless babies’
Calls grow for global regulation as rising numbers of westerners use surrogates in countries with different laws
  by Jessica Murray and Rachel Hall

"The rise in the use of surrogates abroad is leaving more babies at risk of becoming stateless, experts have said, amid growing calls for urgent global regulation of the practice.

"The warning follows the suspension of a 15-year effort by The Hague Conference on Private International Law (HCCH) to establish a global surrogacy convention, which was paused due to divisions among member states. 

...

"UK surrogacy agencies argue that an outdated legal framework at home is pushing couples to seek surrogates abroad, where regulations to protect surrogates from exploitation can be less robust or do not exist at all. Anti-surrogacy campaigners say the practice is fundamentally abusive in all cases and are pushing for a global ban.

"Prof Michael Hellner, who chaired the HCCH working group, said hugely differing opinions on the ethics of surrogacy between the member states involved meant the project to create a global framework on legal parentage in surrogacy cases had proved impossible.

...

"Experts said statelessness could happen in cases where a country does not automatically grant nationality to a child born there, and where surrogacy is not legally recognised in either the country of birth or the child’s final destination.

"Hellner said the project had proved difficult when some countries were “violently opposed to anything that could be seen as legitimising surrogacy, which they think is a violation of human rights and dignity”. 

...

"Most of these children were born abroad, mostly in the US, Ukraine, Nigeria, Georgia, Colombia and Mexico. British people using surrogates abroad to have babies are forced to apply to the UK family courts for a parental order to become the legal parents of their child and, in some cases, also need to apply for British citizenship for the child." 

Curious to Hear More Details Here

New Jersey Gov. Mikie Sherrill held a press conference today in which she announced that the state had just discovered that in 2023 and 2024 the state’s motor voter compliance system had automatically registered to vote 6,600 non-citizens who had said on the form that they were not citizens. (In other words, these people affirmatively said they were not citizens and were automatically enrolled as voters anyway.) She went to say that “fewer than 400” of them had gone on to vote.

Sherrill criticized the previous administration, canceled the contract of the software vendor and started a broader investigation into how this happened. This isn’t the first time this has happened with the motor voter law since non-citizens get drivers licenses and the enrollment system is tied to the driver’s license system. But that “fewer than 400” number seems very, very high relative to other similar cases. So I’m curious how much Sherrill’s office has actually scrutinized these numbers. Usually, a closer look shows many subsequently became citizens or were actually citizens already. And her office seems to have moved very quickly to get these numbers out. They certainly could be valid. But I’d recommend some caution on this until we hear more. I’ve put in a query to the governor’s office to this effect.

★ European Commission: ‘Guidance to Google for AI Interoperability on Android & Sharing of Google Search’

The European Commission, last week:

Today, the European Commission has issued two sets of binding specification measures to Google under the Digital Markets Act.

The aim of the first specification measures is to ensure that competitors’ Artificial Intelligence (AI) services can compete with Google’s own AI services, such as Gemini, by having equal access to features on Google’s Android devices.

The aim of the second specification measures is to rebalance the playing field by giving third-party search engines access to search data that only Google Search can collect at scale.

They provide separate “Q&A” overviews of the guidance for Android AI interoperability and web search sharing, and the full guidance documents are PDFs (Case DMA.100220 for Android AI, Case DMA.100209 for web search). I suggest reading the two Q&A overviews, unless you’re having trouble falling asleep at night, in which case you’ll love the full PDF decisions.

Both decisions are interesting. With search, Google is required to share with competitors — search engines and AI chatbots alike — a massive amount of user data from Google Search user interactions. What terms people search for, what they click on in results, what languages and devices they use. It’s all ostensibly anonymized but that’s tricky when it comes to search terms. A lot of the terms people type into web search fields are to some degree personally identifying. The EC seems to be saying it’s Google’s problem to filter out things like passwords and usernames and omit them from the shared datasets. Google can charge money for this access, but only under “fair, reasonable, and non-discriminatory (FRAND)” prices, based on a Commission-defined methodology.

More interesting to me, however, is the guidance pertaining to on-device AI on Android devices. What the EC is dictating to Google is just breathtaking in scope. The EC is demanding that Google create APIs that allow third-party AI assistants to do everything Google Gemini does now, including:

  • Control hardware buttons (to invoke the assistant).
  • Capture anything on screen, from any app.
  • Seemingly unfettered access to microphones and cameras and other sensors on the device.
  • Unfettered background operation. Third-party AI assistants must be permitted to execute in the background whenever they want, for as long as they want.
  • Execute their own audio models on the digital signal processor, so they can listen at all times for their own custom “Hey Dingus” wake phrases/hot words.
  • Google must allow concurrent access to always-on hot word detection. So if you have Claude and ChatGPT and Grok and Meta AI installed, all of them — in addition to Gemini — must be permitted to have always-on audio detection concurrently. Google is permitted to do some vetting here, but this seems like madness.
  • Third-party models get access to Google’s on-device local models.

Also, in my reading, the EC is demanding that Google make available to third-party AI assistants all information in Google’s own apps (Gmail, Google Calendar, Google Docs, Google Maps, etc.) that Gemini has access to. There is no opt-out for Google regarding data from their own apps. Nor, I think, does this guidance allow third-party apps from other developers to only support specific system-level AI models. Like, let’s say you’re Slack, and you use the APIs to make the content from within Slack available to Gemini on the device. These guidelines don’t permit Google to allow Slack to say that they trust Gemini but only Gemini. If a third-party app like Slack supports making its data available to any system-level AI provider, it must make its data available to every system-level AI provider.

There’s a lot more. Basically, though, the EC is demanding that third-party AI assistants be enabled to become part of the system software, not just apps. I’m sure some people think this is a great idea. It’s the user’s device, they should be allowed to make ChatGPT or Claude or Meta AI part of their OS if they want. It’s up to them. Put users in control.

This is how PCs have traditionally worked, but many normal people’s PCs are a mess of third-party software running in the background. That includes the Mac. If you ask a normal person “What third-party software runs in the background on your Mac or PC?” they would have no idea. It’s all just magic to them. If Google supports this guidance, it could turn Android phones in the EU into PCs. Honestly, some of the stuff the EC is requiring is lower-level than what MacOS and Windows allow third-party software to do.

What the EC’s “guidance” describes in this document is an entirely different operating system than the Android that Google has designed. The European Commission obviously thinks it is their place to design operating systems. Maybe you do too. Google obviously disagrees, and so does Apple. And so should most people who have any idea how these devices work. This is a recipe for disaster, if Google were to enact it and third-party AI assistants took advantage of it.

That second “if” is a big one. One possible scenario that I consider quite likely is that Google could spend years of engineering time and human resources building out APIs to enable all of this, in the safest and most private ways possible, and no major AI assistant adopts it. That’s the way it’s turned out with a whole slew of DMA compliance for Apple and Google. Apple built an entire complex set of APIs to enable third-party web browser rendering engines, exclusively for compliance with the DMA, and there exist no third-party web browser rendering engines for iOS. Not one. Because while the EU is a big market, it’s not big enough to justify building a custom web browser just for the EU alone.

Ways I can see this playing out, in order of likelihood:

(A) Google enacts all of this and no major AI assistants support it because it’s only for Android, only in the EU. And it’s not like ChatGPT and Claude are seeing a lack of usage as things stand now. In this scenario Google just wastes massive time and engineering talent building APIs that never get used, and Android users in the EU get hassled with additional annoying choice and permission screens just to use the Gemini features that are built into Android.

(B) Google enacts all of this and major AI assistants do support it. Unintended results include massive privacy violations where third-party assistants exfiltrate on-device data to the cloud, Meta uses on-device third-party data for the targeting of ads, and users who take advantage of these third-party assistants see significant battery life drain as third-party assistants run without limits in the background and run expensive inference locally to save on their own server costs.

(C) Google enacts all of this, major AI assistants do support it, and there are no privacy scandals, nor any issues with battery life, because each of the companies that makes these assistants develops them with respect for user privacy and for device resources like CPU and memory consumption.

(D) Google pulls system-integrated Gemini from Android in the EU, or severely restricts its capabilities. Rather than elevate third-party assistants from apps to system software, demote Gemini to the privileges of a mere app and leave Android users in the EU without a system-integrated AI assistant.

Under all scenarios, future feature updates to system-level AI in Android will appear late or never in the EU. No future new features can debut in the EU at the same time as the rest of the world because for DMA compliance, Google will need to add support for third-party assistants to do the same things. And they’re not going to hold new features for the rest of the world waiting for that.

I’m not even sure (D) is permitted under this Commission guidance, given that Google started shipping Gemini on Android last year. The entire guidance document is written under the presumption that Google will comply by building the APIs that the guidance demands, not by achieving parity by removing Gemini system integration features. It would be awkward and unpopular for Google to ship updates to Android that remove core AI features, but that might be more palatable than the alternative.1 Note that with iOS, to my recollection, Apple hasn’t pulled any existing features from the EU. They’ve only withheld or delayed new features. Google’s on-device Gemini horse is already out of the barn.

Lastly, although this guidance document pertains to Android, not iOS, I see no reason to think the European Commission wouldn’t demand all or most of the same things from Apple. They haven’t given Apple any guidance yet, because it’s European Commission policy to give guidance only after a DMA-designator gatekeeper ships something that is then ruled non-compliant. But it’s hard to imagine Apple accepting most of these terms. Unfettered background processing and access to the microphone, cameras, and sensors? Third-party audio models running on the hardware DSP listening for wake words? Granting third-party assistants unsupervised access to all user data from apps published using App Intents?

Apple unfortunately hasn’t shared any technical details describing its proposal for a “Trusted System Agent” that it shared with the EC last year. But whatever Apple’s vision for the Trusted System Agent is, I don’t see how the EC would deem it compliant with the DMA if they want from Apple anything close to what they are now demanding from Google. There are no checks and balances or oversight that Google is permitted to apply to third-party assistants in this guidance. If Gemini can do something, third-party agents must be able to as well. Because Gemini gets to run in the background as much as Google sees fit, third-party assistants must be permitted to run in the background as much as they see fit. And if those third-party assistant developers — OpenAI, Anthropic, xAI, and Meta — have different opinions than Google on how much CPU usage is appropriate in the background, how much RAM and storage is appropriate to consume, or how respectfully to treat users’ on-device data, well that’s just tough noogies. If the user OK’s it, then it’s OK.

A month ago, after WWDC, when this guidance pertaining to Android AI was rumored to be forthcoming, I wrote:

Google is learning the lesson Apple learned the hard way with all the existing features of iOS that were deemed noncompliant with the DMA when it went into effect. The “ship it first and ask forgiveness / hope it’s deemed compliant” strategy is not a good one in the EU.

I genuinely wonder what the European Commission thinks the purpose of Android is. Google created Android for the benefit of its own services. Google was worried about Microsoft, not Apple, at the time, but they wanted to ensure their own services were available on a major mobile platform. I’m really not seeing how it’s more attractive to Google to comply with this guidance than to just pull system-level Gemini in the EU. Either they waste a small fortune building APIs no one will use, or, they spend that small fortune buildings APIs for the benefit of their biggest competitors. I get it that that’s the intended price to pay for being a designated DMA gatekeeper. But what’s the motivation for Google to do this rather than just walk away from system-integrated AI on Android in the EU? It certainly doesn’t look like the competing platform, iOS, is going to offer it in the EU anytime soon either.


  1. If you are handed a mandate that everyone must be able to run at the same speed, and you can’t figure out how to make slow people faster, you can comply by forcing the fast to wear weighted boots. This is where utopian egalitarian initiatives often lead. ↩︎

Can cells think?

Close-up photo of a cross-sectioned orange slice with intricate fibrous details against a dark background.

Inside the lab of a biologist who is seeking to understand intelligence by focusing on problem-solving, not neurons

- by Aeon Video

Watch on Aeon

We Need Your Help

We’re now in the toughest part of this year’s Annual TPM Journalism Drive. That’s the part between $300,000 and $400,000. Once we get passed the latter number people start to focus on being near the goal ($500,000) and the pace of contributions builds. But we need to get there first! If you can, please take a moment to contribute today. Any amount means the world to us, and it’s critical to this organization’s future. Just click right here.

Update: We’re just $1,358 short of getting to $310,000 tonight!

Trump Is an Angry, Violent Man … And Canada Is His Battered Wife

Donald Trump has announced that he will impose in 30 days a new round of tariffs on Canada. Those tariffs are at best of dubious legality. They are justified under the Smoot-Hawley Tariff Act of 1930, which has quite likely been superseded by subsequent post-war trade legislation. But questions of legality, while critical, tend to obscure the more elemental question of why? There are at least arguments as to why the U.S. might want to impose protectionist measures against China or other countries with far lower wages and workforce protections. There are other weaker but still arguable rationales for protectionism targeting Europe. One can even make arguments about Mexico since wages south of the border are still much lower than in the U.S. (To be clear, I mostly don’t agree with these arguments, though I also disagree with a doctrinaire free trade approach.) There’s really no economic or strategic argument for these tariffs on Canadian goods at all. Canada has comparable living standards and regulatory regimes to the U.S.; key U.S. industries like autos are woven across the U.S. border. The anchor of U.S. prosperity has always been the way it amounts to a massive free trade zone. And Canada only extends that to the north. And yet tariffs against Canada appear to be the White House’s near or entire focus.

This Politico piece suggests that one motivation may be Prime Minister Mark Carney’s push for “middle powers” to assert their collective power in an increasingly multipolar world. The White House doesn’t like that, sees it as a threat to U.S. global dominance, and is beating up on Carney and Canada to get him to back off. But that doesn’t seem like an adequate or satisfying explanation. At least in the near term “middle powers” talk is mostly just talk. It’s not meaningfully showing up in international power dynamics. But, more than that, it suggests a more reasoned and strategic approach to tariff and geopolitics than there’s really any evidence for.

As with the war with Iran, I think this greatly underestimates Trump’s individual emotional state, his need to dominate and need for self-soothing as a driver of major foreign policy and economic policy decisions. Canada is becoming the abused wife that Trump is lashing out at because things are going so badly at work. His hopes and dreams are crashing: The midterms look bad. His leg is caught in a veritable bear trap in his Iran War. So he goes home and beats up his wife, or in this case, lashes out at Canada. It’s another case of executive self-soothing oddly akin to what got Trump into Iran in the first place.

The American press has always understated and underestimated the gleaming continuities between Trump’s inter-personal and political selves. He needs to dominate. He needs you to lose because that’s the only way he knows he’s won. He hungers to be fawned over. Going back to 2015, people who’ve lived with chronic abusers recognized the type — palpably, intuitively, from experience — and those patterns which have marked his personal life — cheated business partners, raped and assaulted women — have been imposed on the country as a whole. There’s almost nothing in his presidency that can’t be reduced to this basic, ingrained and lifelong framework. Canada, vulnerable to abuse because of its proximity and economic dependence, is Trump’s convenient go-to to act out on these impulses. Actually invading Greenland is complicated. Certainly going to war with Iran — despite its almost total international isolation — is pretty hard too. Regular rounds of sub-lethal abuse of Canada is much easier.

As we discussed beginning late last year, as Trump’s domestic fortunes, domestic popularity and power decline, he becomes more dependent and focused on leaning into his nearly unlimited powers abroad. These tariffs are probably illegal, whereas the Iran War, for all its stupidity and folly, is probably not. But the Supreme Court has already shown it will give Trump a year or so of illegal tariffs before calling off the fun. This is an extension of the same whims, need for distraction and predation, the same drive for dominance and self-soothing for the inner pain of failure. And to our great shame, America’s great shame, our northern neighbors are taking the brunt of it.

LOL

I’ve got to give Mike Lindell credit for engaging our questions about whether he’s actually registered to vote in Minnesota, where he’s running for governor. But his explanation here – he texted our reporter a copy of a temporary driver’s license that expired in January – will truly make your day. See here.

The Southern Crown (Corona Australis) dazzles with young and ancient celestial jewels. The Southern Crown (Corona Australis) dazzles with young and ancient celestial jewels.


Brian Tyler Cohen | American Conversations

New Jersey just handed Donald Trump a huge gift.

Saw this article, and immediately felt like punching a wall

It’s the type of thing that begins small. A headline. Two headlines. Some TikTok videos.

Then it grows.

And GROWS.

AND GROWS.

AND GROWS!!!!!!

AND GROWS!!!!!!!!!

And, before you know it, Trump and Co. are all over it. See! We told you! The Democrats are cheating! Look at New Jersey! Look what’s happening! It’s in the news! We told you people! We told you!

This, I assure you, will begin ASAP.

Inevitably, the Dems will cower into a corner, cough, point, stutter. When, truth be told, they should do now what they should’ve been doing all the fucking fuck along. Which is (I’ll begin the all caps here, good people) SCREAM AND SHOUT ABOUT THE JAN. 6 INSURRECTION, AND REMIND PEOPLE EVERY FUCKING DAY THAT THE ASSHOLE BELLOWING ‘FRAUD! FRAUD!’ SENT HIS MINIONS INTO THE CAPITOL, WATCHED THEM TEAR IT APART, FAILED TO LIFT A FINGER–THEN PARDONED EVERY SINGLE PARTICIPANT, EVEN THE ONES WHO BATTERED LAW ENFORCEMENT!

Seriously, this is the video. And it needs to be played on repeat …

Alas, Chuck Schumer will probably head up to the New York State Fair to rave about the yogurt.

Because, hey.

What’s democracy without dairy?

I don't think taunting a Huntington Beach City Council member about his vaping habits is a winning strategy

Butch Twining.

Received some excited messages the other day, asking whether I’d seen “The Butch Twining Video.”

“Nope,” I replied. I had not.

Here it is …

And I get it. I do. Twining is a member of the Huntington Beach City Council. He’s one of the MAGA bros who is all in on Donald Trump and Charlie Kirk and porn in schools and book bans and the plaque and the statue and he’s rude and crass and dumb and, yes, he vapes at age, like, 70. Which, unless you’re Billy Joel, is super weird and legitimately perplexing.

However …

What does this sneeze in time—captured on video for the world to behold—do for Democrats trying to win a seat or two on the City Council in a very conservative metropolis? Does it make our party appear dignified? No. Mature? No. Reasoned? No. We look (and I hate to say this) sort of unhinged and rabid, and the post-exchange jubilation over getting Twining’s goat is less win, more ode to the gotcha! landscape we now occupy. People who abhor Twining will love it. People who support Twining will see it as libs being libs. And it’s all on video—the 2026 universal medium of choice.

Again, I do get it. Hell, I love the spiritual release of making fun of MAGA buffoons. But the goal here is to flip the city council, and to do so means convincing on-the-fence folks better people than Butch Twining walk the earth.

Also, not for nothing, there are productive ways to get a goat, and nonproductive ways to get a goat. Had the man outside the vehicle said, “Hey, Butch, I think building a Trump statue is crazy …” and then Twining fires off, “Fuck you!” … that’s a solid win. It’s an irrationally crude reply to a fair observation stated by a constituent.

But if someone ran up to my car and cracked that vaping line, I’d probably tell him to fuck off, too.

Maybe I’m growing soft, but it’s just not for me.

July 21, 2026

The front page of yesterday’s New York Times featured some of the 440,000 Americans from Arizona alone who have lost Supplemental Nutrition Assistance Program (SNAP) benefits since the Republicans’ One Big Beautiful Bill Act made the deepest cuts to SNAP in its history, expanded work rules, and shifted costs to the states even as cuts to federal grant programs have forced Arizona to lay off a third of its caseworkers.

Those Arizona residents include Dee McDonald, a 65-year-old cancer survivor who weighs 69 pounds and skips meals to feed her three teenaged grandsons.

SNAP falls under the Department of Agriculture, and Agriculture Secretary Brooke Rollins celebrates the cuts, claiming the program is rife with fraud. Like other Republicans, she cites error rates. But error rates usually reflect mistakes in classification numbers by caseworkers, not fraudulent use of services.

Today Secretary of Health and Human Services Robert F. Kennedy Jr. told reporters the federal government is withholding more than a billion dollars in Medicaid funding from California and Minnesota, saying they want to see proof that payments there aren’t fraudulent. Sarah Kliff of the New York Times notes that the administration was already withholding money from those two Democratic states. It has withheld $1.3 billion from California since May and $243 million from Minnesota since February. Now Kennedy says it will withhold more than $867 million from California and more than $200 million from Minnesota.

In the Denver Post today, Executive Director of the Association of National Park Rangers Bill Wade wrote that during his years as the son of a park ranger, and then as a park ranger himself, “I learned that these extraordinary parks belong to every American and that the people entrusted to care for them carry a sacred responsibility to protect our country’s breathtaking natural resources and key elements of our nation’s history.”

“But never—not once in decades—have I witnessed such systematic degradation, disrespect and dismantling of the National Park Service as I see today under the thumb of the President Donald Trump and Interior Secretary Doug Burgum,” Wade continued. Budgets go up and down and policies change as administrations turn over, but “what is happening now is an assault on the people, the mission and the values that have made the National Park Service one of America’s most valued institutions.”

The Trump administration has slashed staff, starved the service of funds, and imposed a political narrative on historical interpretation. “Recovery from these losses is likely to take years, if not decades,” Wade wrote.

At the same time that the federal government is cutting programs that benefit the public, it is also privatizing systems that the government has been managing. Yesterday Justin Doubleday of the Federal News Network reported that Des Moines International Airport and Tampa International Airport have opted into the Transportation Security Administration’s (TSA) new privatization model for airport screening.

Congress created the TSA after the September 11, 2001, terrorist attacks. Before then, individual airlines contracted with private security firms to screen passengers and baggage, but the 9/11 attacks showed that this private system had dangerous vulnerabilities. Now the Trump administration is resurrecting that system.

While administration officials are slashing government programs that protect ordinary Americans, Secretary of Defense Pete Hegseth and Chairman of the Joint Chiefs of Staff General Dan Caine appeared on Capitol Hill today before the Senate Appropriations Committee to explain why the Pentagon is urgently in need of an injection of $67 billion to tide it over until the end of the fiscal year on September 30. The Pentagon has sunk billions into the war in Iran. Hegseth told senators the war had cost $37.5 billion, but independent analysts estimate the true number is significantly higher.

Two congressional aides told Noah Robertson, Riley Beggin, and Jarrell Dillard of the Washington Post that both the Navy and the Air Force will run out of money provided for ongoing military operations in 2026 at the end of this month. The Pentagon is already moving money around to bridge the gap, diverting money budgeted for equipment and maintenance and limiting or canceling military exercises and training.

Yesterday Rose L. Thayer of Stars and Stripes reported on one such diversion: The Army is reducing the legal resources available to soldiers being evaluated for disability “as review boards determine whether they are fit to serve or should receive medical retirement benefits.” In the past, the Office of Soldiers’ Counsel has provided those services, but dramatic cuts have left personnel available only for an initial evaluation, rather than throughout the three-stage process. After the initial evaluation, soldiers must either represent themselves or hire private attorneys.

At today’s hearing, Senator Jon Ossoff (D-GA) grilled Hegseth. “Mr. Secretary, early in the conflict, day three, you said one of those core objectives was to, quote, destroy the missile threat, and as you restated it on day five to, quote, obliterate Iran’s missiles and drones. Those were clearly and repeatedly stated core objectives of the campaign, which four months ago you declared had been completed victoriously with every single objective achieved. Were those objectives achieved four months ago?”

Hegseth answered: “No one ever stated that every single missile was gone. Ultimately, combat ineffective means they’re not able to dig them out of their UGFs [underground facilities] and shoot at us.”

Ossoff: “Had the missile threat been defeated four months ago?”

Hegseth continued over Ossoff: “It doesn’t mean they don’t have underground facilities where they put missiles in them and hold onto them because they’re the world’s largest state sponsor of terrorism. So we never expected they would stop having capabilities to shoot. The question is, at scale, against an opponent like the United States of America, which went toe to toe with them in Iran, couldn’t….”

Ossoff continued: “You stated on March 8th, day nine of the war, on 60 Minutes, you assured the American people of Iranian surrender. Does that remain your prediction?”

Hegseth: “Ultimately, Iran will never have a nuclear weapon to threaten the United States of America. That has been the objective, and it remains so.”

Ossoff: “So, just to review, you will not answer whether your statement made on the fourteenth day of the war that Iran’s military had been, quote, destroyed and made combat ineffective was a truthful statement to the American people, as you sit here and ask for tens of billions more for the conflict.”

The second Trump administration is forcing the American people to define how they think the United States government should spend their tax dollars.

Today Senator Andy Kim (D-NJ) introduced a proposal to expand Medicaid to all American children from birth to age 26. Kim noted that about 4.4 million children in the U.S. don’t have health insurance and 23 million are underinsured. The One Big Beautiful Bill Act will make things worse.

Joseph Choi of The Hill reported that Democratic senators Alex Padilla of California, Cory Booker of New Jersey, and Tammy Duckworth of Illinois are co-sponsoring the bill. The legislation proposing “MediKids” has been endorsed by a range of organizations, including the American College of Obstetricians and Gynecologists and the American Academy of Pediatrics.

“Every child should have health care,” Kim said in a video. “Every child should have the ability to go to a doctor when they’re sick. Every child deserves to be healthy.” MediKids “would automatically enroll every child in health care coverage at birth. Because the system we have right now isn’t working. Because healthy kids make healthy adults, and that makes for a healthier, stronger future for our country. Because, as a parent, you deserve to be able to provide care for your kids without worrying about going into debt, emptying your savings, or straining the family budget. And because…we all have a responsibility to care for each other.”

Kim said Democrats need to bring new ideas forward while also opposing the Trump administration’s destruction. He told reporters: “I know that there’s overwhelming support from the American people for ideas like this, and there’s a hunger for ideas right now.”

Notes:

https://www.nytimes.com/2026/07/21/us/politics/trump-administration-medicaid-california-minnesota-fraud.html?smid=url-share

https://www.nytimes.com/2026/07/20/us/politics/snap-food-stamps-arizona.html

https://www.cbpp.org/research/food-assistance/snap-includes-extensive-payment-accuracy-system

https://www.washingtonpost.com/national-security/2026/07/21/pentagon-sinking-billions-into-iran-is-quickly-running-short-cash/

https://www.britannica.com/topic/United-States-Department-of-Homeland-Security

https://federalnewsnetwork.com/management/2026/07/trumps-pick-to-lead-tsa-calls-private-airport-screening-program-pro-worker-vows-to-help-workers/

https://federalnewsnetwork.com/workforce-rightsgovernance/2026/07/two-airports-opt-into-tsas-new-privatization-model/

https://boltflight.com/when-was-the-tsa-created-and-why-it-was-established-the-origins-of-modern-u-s-airport-security/

https://www.denverpost.com/2026/07/21/trump-national-parks-service-public-lands/

https://www.stripes.com/branches/army/2026-07-20/legal-services-disability-evaluation-process-22318446.html

https://thehill.com/homenews/5980026-sen-andy-kim-medikids-proposal/

https://www.kim.senate.gov/press_release/senator-kim-introduces-landmark-medikids-legislation-to-guarantee-healthcare-for-all-children/

Facebook:

reel/1052373863846713

Bluesky:

atrupar.com/post/3mr6siqveqa2h

Share

Politics Chat, July 21, 2026

Politics Chat, July 21, 2026

Who was the worst monster of the 20th century?

Photo arrangement by General Iroh, the Dragon of the West, via Wikimedia Commons

Let’s take a break from discussing the insanity of American politics, the tragedy of British economic mismanagement, and other weighty topics, and talk about something fun and light. Who was the greatest monster of the 20th century?

Well, ok, that’s not really fun and light. The 20th century featured hundreds of millions of people slaughtered in wars, genocides, political persecutions, and preventable famines. Many of my own extended family members were killed in the Holocaust, and I have plenty of friends whose relatives died in World War 2, the Great Leap Forward, the Cultural Revolution, the Cambodian Genocide, and other such calamities. Those deaths are even more tragic because of all the good things that were happening in that century — the fabulous economic growth and technological progress. Everything could have just been a smooth upward glide to luxury and comfort, and yet people were still being machine-gunned into mass graves.

With time, the emotional power of those events fades — Genghis Khan killed tens of millions in his campaigns of conquest eight hundred years ago, but today he’s more likely to be invoked as a joke, or an object of reverence, than as a monster. It can happen surprisingly fast. By the early 2000s, hipsters in my dorm were already hanging ironic Chairman Mao posters in their rooms. Today, you can see leftist pundits like Hasan Piker praising Mao without a trace of irony:

Plenty of other leftists joined in the Mao lovefest.

Piker is wrong, of course; if you read the authoritative biography of Mao by Jung Chang and Jon Halliday, you’ll see that the dictator repeatedly aggrandized himself at China’s expense, willfully derailing multiple periods of reform and recovery by launching the country into destructive episodes of chaos like the Great Leap Forward and the Cultural Revolution. Nor was he much help against the Japanese invaders in World War 2, preferring to let them fight and weaken his Chinese rivals (again, at the expense of many Chinese lives). It was Deng Xiaoping, not Mao, who made modern China great.

But the people who say that Mao was history’s greatest tyrant, with an unsurpassed body count of 40-50 million people, are also probably overreaching a bit. Yes, about 36-40 million people did die in Mao’s Great Famine, caused by his various foolhardy agricultural policies and the inefficiencies of China’s communist economic system. Nor was this merely an accident; even after Mao learned that his policies were failing, he kept them in place to avoid losing face in front of CCP rivals. But killing tens of millions of his own people was never Mao’s plan.

Intent is an important factor to consider when ranking history’s greatest monsters. The reason has to do with why we should care about the question. The point of ranking evil regimes isn’t to have fun marveling at the scale of human brutality — it’s not like trying to figure out what the biggest dinosaur was, or whether Michael Jordan was better than LeBron. Sometimes real life forces you to choose between monsters.

In World War 2, America had to decide whether allying with Stalin against Hitler was worth it, or whether we should simply sit it out and be neutral as the two villains slugged it out. Our decision probably tipped the course of the war. Later, we formed a quasi-alliance with Mao’s China against the Soviet Union, helping to hasten the latter’s downfall.

Many populations, of course, have an even grimmer choice — they have to choose which ruthless, cruel regime to support in a civil war. If you lived in Syria in 2016, did you support Bashar al-Assad, ISIS, or offshoots of al Qaeda? If you lived in Russia in 1919, did you support the Reds or the Whites? And so on.

So when we ask “Who was worse”, the question is actually pretty relevant. Obviously it’s a matter of opinion, but I think some factors we should consider when assessing the horribleness of a regime are:

  1. Overall death toll during the regime’s time in power

  2. How much of the death toll was intentional

  3. How much of the death toll was willful versus accidental

  4. How harmful the regime was relative to the size of the population it controlled

  5. What other bad things the regime planned to do but was unable to accomplish

  6. What life was like for the people who lived under the regime and didn’t die

So with that in mind, here’s my ranking for last century:

The worst regimes of the 20th century

#1: Adolf Hitler and the Nazis

For me, Hitler easily takes the top spot, and it isn’t really a question. The reason isn’t that the Jewish Holocaust — Hitler’s most-discussed atrocity — was uniquely horrible compared to other acts of genocide. It’s that the Jewish Holocaust was only the beginning of Hitler’s wave of destruction.

The slaughter of 5-6 million Jews was deliberate, but it was only one part of Hitler’s plans for genocide in Europe and Asia. Generalplan Ost, the Nazis’ plan for East Europe, involved the deliberate slaughter or enslavement of tens of millions of people in Russia and other countries to Germany’s east. Here’s a chart from Wikipedia:

“Removed” here means “killed”, since the Nazis intended to take all of the land.

Hitler’s invasion of the USSR — Operation Barbarossa — was meant to carry out this unprecedented genocide. The main tool was starvation — Nazi Germany’s “Hunger Plan” called for intentional mass famines to wipe out most of the people living in the USSR and East Europe.

The Nazis actually implemented this plan to a significant degree. Out of the 27 million Soviet citizens killed by the Nazis during World War 2, 19 million were civilians, and most of these probably died from hunger and disease, caused in large part by the Nazis’ deliberate policy of mass starvation during the war. The Nazis also mass-executed or starved millions of Soviet POWs. Had the U.S. not stepped in to help, the death toll in the USSR might have been twice as large, or more. Those that weren’t killed were to be enslaved, but given the experience of the “workers” at Auschwitz — and considering that Nazi ideology stated that superior races lived and inferior ones died off — it’s an open question how long those slave populations would have been allowed to survive.

And why would Hitler have stopped there? Having overcome the Soviets, would the Nazis really have allowed the Chinese, or the Arabs, to live their lives and hold their lands unmolested? Not a chance. Hitler’s regime was purpose-built for conquest and genocide — it could only keep killing or die. In this sense, the Holocaust, as horrifying as it was, was only the tip of the iceberg.

It’s not clear whether there has ever been another regime quite like Hitler’s, even when we look back to the barbarity of premodern times. Genghis Khan, Timur (Tamerlane), and other conquerors killed millions, but only in order to subdue and rule, not to annihilate. Other regimes have attempted annihilation of subject peoples, but have generally done this within their own borders rather than trying to do it to the rest of the world.

Hitler and the Nazis should thus stand as the undisputed all-time champion of evil. With an intentional body count of over 30 million, and plans to do even more, they were less like other historical tyrants and more like some kind of alien armada. In my mind, the decision to ally with Stalin in order to defeat that unique menace was unquestionably the right one.

#2: Fascist Japan

The Japanese Empire in the 1930s and 1940s was very different than the empire that came before. The military usurped civilian rule, put an end to the democracy that Japanese people had enjoyed in the 1920s, and made their society increasingly totalitarian. Japan had conquered and colonized before — Taiwan and Korea — but in the 1930s and 40s, they became increasingly brutal.

By the time they went after all of China in 1937 and Southeast Asia in 1941, genocide had become their standard tool of conquest. Like the Mongols 700 years earlier, the Japanese conquerors were very aware of China’s greater population, and intentionally used genocide in order to make China easier to rule. They killed 14 million Chinese people in their invasion (some say only 6 million, some say 20 million), mostly by deliberately starving them or slaughtering whole towns — similar to what the Nazis did in the USSR, but without the camps. They used biological weapons to try to kill even more. And of course they famously killed over 10,000 people by experimenting on them. This is in addition to a long list of massacres and atrocities carried out in the rest of Asia.

We tend to leave Japan off of the list of “history’s worst monsters”, because these crimes were committed in a very distributed way. They were not a result of official state ideology, like the Nazis’ ideology of race war. Individual figures like Tojo Hideki sometimes did encourage policies of mass starvation, mass shootings, and biological warfare, but the Imperial Japanese Army was remarkably decentralized, and lots of local commanders did atrocities all on their own. The Emperor probably knew about at least some of the atrocities being committed in his name, but didn’t try to intervene.

Still, collectively, Imperial Japan was the closest thing to the Nazis that existed in the 20th century, and like the Nazis they would have gotten far worse if they hadn’t been forcibly stopped.

#3: Pol Pot and the Khmer Rouge

Pol Pot and the Khmer Rouge had a total body count that was much lower than the other regimes on this list — 1.5 to 3 million instead of tens of millions — and they didn’t try to conquer the surrounding lands (though they sometimes attacked their neighbors). But the reason they only killed a few million people was that Cambodia only contained a few million people. The Khmer Rouge slaughtered a fifth to a quarter of their populace, higher than anyone else on the list — or maybe in recorded history. And had they had the power to conquer any of their larger neighbors, they probably would have conquered and killed even more. As it was, their reign of terror was only stopped by a Vietnamese invasion.

In terms of the cold, systematic nature of the mass murder, the Khmer Rouge are second only to the Nazis. The Killing Fields are legendary. You can go see the “magic tree” where Pol Pot’s soldiers would dash babies’ brains out, or gaze upon nice neat rows of human skulls. It really doesn’t get more psychotic than these guys.

#4: The USSR under Stalin

Stalin is generally held responsible for somewhere between 10 and 20 million deaths, of which maybe 6-10 million starved to death because of communist agricultural policies. Like Mao, Stalin learned of these deaths but turned a blind eye rather than admit a mistake. But there may have been more to it in Stalin’s case — he probably wanted a famine in Ukraine in order to intentionally reduce the Ukrainian population and make the “republic” easier to rule.

Stalin’s deliberate murders were fewer than Hitler’s but that’s a relative statement; he did slaughter maybe 3 million political opponents by murdering them or sending them to gulags to die of neglect. He committed small local genocides, killing over a hundred thousand Poles in WW2, mass-expelling a bunch of minority populations (like the Crimean Tatars) and killing many in the process, and so on.

His armies were also very brutal in Germany during World War 2, mass-raping literally millions of German women, expelling and enslaving large numbers of Germans in East and Central Europe, and so on. On top of all that, Stalin’s USSR was famously nightmarish, with people living in a constant state of fear and encouraged to rat out their neighbors.

What Stalin mostly didn’t do was conquer and expand the USSR’s territory. With a few exceptions, he just sat on his existing empire, dominated surrounding states with proxy arrangements, and focused on internal ideological and economic reorganization and maintaining his personal power. This lack of world-conquering ambition made him palatable as an ally against Hitler. Ultimately, he mostly just sat there and brutalized the people already under his own control.

#5: China under Mao

The vast majority of Mao’s deaths come from the famine. He also killed maybe 2-4 million people deliberately, through political purges, executions and forced labor during the Great Leap Forward, and the insanity of the Cultural Revolution. But the basic story of Mao is that he implemented insane policies in an attempt to gain and hold onto power, and often viewed Chinese deaths as collateral damage in his power struggles.

Mao, like Stalin, was also a villain that other countries learned to live with. He conquered Tibet and invaded South Korea (before being repulsed), but overall he wasn’t much interested in conquest. Like Stalin, he sat on his giant empire and brutally attempted to bend it to his (often insane) will. The society he created was desperately poor, paranoid, and totalitarian, but also often chaotic and even anarchic in places.

Honorable mentions

In my mind, those five regimes stand head and shoulders above the rest, but there are certainly no shortage of other bad ones. Seven honorable mentions:

  • The Young Turks: Turkey’s junta carried out the Armenian Genocide during World War I, killing around a million people.

  • North Korea: One of the world’s most nightmarish totalitarian regimes, North Korea tried to conquer South Korea, and also caused at least one mass famine with its policies of deliberate isolation.

  • Rwanda’s Hutu government: The Hutu government in the 1990s committed the famous Rwandan genocide, and helped foment and prosecute the Second Congo War that killed millions of people in central Africa. An ideology of ethnic supremacy was a big part of the reason why.

  • The Derg: Ethiopia’s communist regime in the 1970s and 1980s doesn’t get talked about much, but they killed maybe 500,000 people in political executions and ended up starving hundreds of thousands of Ethiopians to death.

  • Yahya Khan: Under Yahya Khan in 1971, Pakistan tried to terrorize Bangladesh into remaining under Pakistan’s control. The army killed anywhere between 300,000 and 3 million Bangladeshis, and raped hundreds of thousands more.

  • Suharto: Although Indonesia saw robust economic growth under Suharto, he killed over half a million people for political and ethnic reasons.

  • King Leopold II: The Belgian king’s extractive regime in the Congo killed anywhere from 1.2 to 10 million people there. The bulk of the atrocities in the 19th century, however; Leopold’s rule over Congo ended in 1908.

I realize that this list is subject to differences of opinion, depending on A) which estimates you believe, and B) which factors you weight more heavily in the calculation of regime evilness. But I think it’s a pretty defensible list, and it demonstrates at least two important lessons.

First of all, the very worst atrocities come during times of war — especially wars of conquest, where foreign peoples are dehumanized and armies are given carte blanche. The fascist regimes of WW2 come off as the worst because they were engaged in massive wars of conquest. And wars often increase a regime’s appetite for further conquest and destruction. So basically, regimes that invade other countries tend to be worse.

Second, communist regimes tended to starve a lot of their people to death. Insane agricultural policies were a hallmark of communist economics — probably because successful communist revolutions usually began as revolts by agrarian peasantry rather than proletarian workers as Marx had envisioned. In trying to transform agriculture, communists ended up just starving their people to death.

So while communism probably caused more aggregate human misery than fascism in the 20th century, that was only because fascism didn’t have time; it was so incredibly malignant that it was stopped very quickly by external coalitions. Communism was more like a chronic debilitating disease, while fascism war more like ebola.

The worst regimes of the 21st century so far

That leaves us one more interesting question: By these criteria, who have been the worst regimes of the current century, so far? So far, the 21st century has been somewhat less destructive than the 20th — by 1926 we had already seen World War I, the Armenian Genocide, Lenin’s political purges, the Russian Civil War, and so on. The years 2000-2026 have, thankfully, been somewhat less brutal and chaotic so far. But we still have plenty of tyrannical villains. Here would be a tentative top five:

1. The RSF, Omar al-Bashir, and the Janjaweed: There has basically been a rolling series of genocides in Sudan this whole century — the first Darfur genocide, the second Darfur genocide, the Masalit genocide, and so on. Some of this was done at the behest of former president Omar al-Bashir, who encouraged murderous militias known as the Janjaweed, who later morphed into the RSF.

2. ISIS: This is the closest thing we’ve had to the Khmer Rouge in recent decades. ISIS’ extremist ideology led them to create a nightmarish totalitarian state in the areas of Iraq and Syria under their control, genocide the Yazidis, and throw much of the Arabian Peninsula into chaos. Their total death toll wasn’t as high as some of the others, but this was because they were stopped quickly; they were so horrible that everyone immediately united to fight and destroy them.

3. Vladimir Putin: Putin is one of the most aggressive leaders of modern times. He started wars in Georgia and Ukraine, and got heavily involved in the war in Syria. Over half a million people have died in Ukraine so far, and Putin bears some responsibility for the hundreds of thousands more who perished in Syria. It seems clear at this point that he’s not going to stop prosecuting wars of aggression until he dies or is defeated.

4. Bashar al-Assad: In his desperation to hang onto power, Assad laid waste to his country, killing hundreds of thousands and even using chemical weapons on rebels before ultimately fleeing.

5. The Tatmadaw: Myanmar’s military has been locked in a brutal civil war against ethnic minorities for decades, but in 2021 things really kicked into high gear after the army staged a coup against the country’s democratically elected government and put down protests with extreme brutality. The country is now an anarchic wasteland with war crimes galore.

These tyrants and murderers are somewhat less impressive, numbers-wise, than their 20th century counterparts. They also don’t clearly fit into the “fascism vs. communism” dichotomy that dominated much of the previous century; they’re more likely to be motivated either by Islamism, local ethnic chauvinism, or simple personal lust for power. But looking back at the 20th century, we should remember how bad dictators and totalitarian regimes can get, especially when they go to war — and how important it is to stop the worst of them in their tracks, before they can get even worse.


Subscribe now

Share

Fable responds

Today’s post is brought to you by my sponsor, Mechanize. They’re hiring junior software engineers at $300K/year base salary. Apply now!

* * *

In a recent post, I responded to a Fable output comparing market monetarism with the HANK [Heterogeneous Agent New Keynesian] model, especially in reference to 2008-09. Vaidas Urba also asked Fable the same question from a HANK perspective. Because LLMs are somewhat obsequious, that response appeared slightly more HANK-friendly, although the substance was not much different. And then Vaidas asked Fable to compare the two responses and offer a judgement as to which one was more persuasive.

In this post, I had intended to respond to Vaidas’s latter two Fable outputs, but on second thought I’d like to respond to a different Fable output, which seems much more interesting, and even more impressive. This output was produced after Vaidas asked Fable to respond to my blog post, which was itself a response to the first Fable output.

This post will have three parts. In the first two, I’ll provide the Fable HANK-friendly reply and the overall summation that Vaidas sent me earlier, without any comment on my part. Feel free to skim over the first two sections. I don’t wish to get bogged down in an overly long tit-for-tat on every single point. Instead, I’ll reserve all my comments for the superior part 3, where I respond point by point to Fable’s reply to my earlier post. In this way, I hope to keep the conversation going in a productive fashion.

Note that in this final section, Fable’s response is written at a fairly high level. The response is not technically demanding in a mathematical sense, but it requires a deep understanding of macro theory. Sorry, this post isn’t for everyone. But I’ll occasionally try to clarify in my replies to make it easier for the reader. I will not indent Fable’s output, only my replies.

Part 1:

Prompt:

You are a discussant preparing for macroeconomics seminar. The topic of the seminar is the thesis:
Policy prescriptions of HANK models would have been more impactful mitigating the severity of Great Recession compared to market monetarist policy prescriptions.
Identify and explain the top 5 arguments in favor of the thesis.
Identify and explain the top 5 arguments opposing the thesis.
Based on the arguments, provide your ultimate verdict with a reasoned justification.

Reply:

Arguments in favor of the thesis (HANK > market monetarism)

1. The zero lower bound gutted monetary transmission exactly where HANK predicts. Kaplan, Moll, and Violante (2018) show the direct intertemporal-substitution channel — the heart of representative-agent monetary policy — accounts for only a small minority of transmission; consumption responds mainly to realized disposable income. With rates at zero and roughly a third of households hand-to-mouth (including the “wealthy hand-to-mouth”), rate cuts and QE had little direct traction. Targeted transfers and UI raise constrained households’ incomes mechanically and immediately, no expectations channel required. Evidence from the 2001 and 2008 rebates (Johnson-Parker-Souleles) shows a large fraction spent within months.

2. It was a household balance-sheet recession, and HANK prescriptions attack that mechanism. Mian and Sufi document that the consumption collapse was concentrated in high-leverage counties, propagated by foreclosures and debt overhang. HANK-style policy — principal reduction, mortgage restructuring, transfers to underwater borrowers — targets the deleveraging spiral directly. Actual policy (HAMP) was timid and the slump was correspondingly long. NGDP stabilization helps debtors only diffusely.

3. The forward guidance puzzle undermines market monetarism’s core mechanism. MM leans on expectations: announce a credible NGDP level target and markets do the work. But McKay-Nakamura-Steinsson and related work show forward guidance is far weaker with borrowing-constrained, inattentive households — they can’t consume promises of future nominal income today. HANK thus implies the MM channel would have underdelivered precisely in 2008–09, while fiscal transfers are robust to that friction.

4. Distributional amplification and scarring. Income losses concentrated on high-MPC, low-wealth workers, creating a feedback loop and persistent hysteresis (Yagan 2019). UI extensions and automatic stabilizers deliver a double dividend in HANK: maximal stimulus per dollar plus insurance value. Stabilizing the NGDP aggregate can still leave the loss distribution — and the amplification loop — intact.

5. Implementability, plus the 2020 natural experiment. Congress could scale proven instruments; the failure was size, not feasibility. Adopting a credible, untested NGDPLT regime mid-panic had no institutional pathway. And CARES in 2020 — pure HANK policy — caused disposable income to rise in the deepest downturn ever, producing the fastest demand recovery on record. The MM counterfactual never received a comparable proof of concept.

Arguments opposing the thesis (market monetarism > HANK)

1. The proximate cause was a preventable nominal collapse. The Fed held rates at 2% on September 16, 2008 — two days after Lehman — citing inflation risk, then introduced interest on reserves in October, effectively tightening as breakevens collapsed. NGDP fell furthest below trend since the 1930s. Via Fisher debt-deflation logic, keeping nominal incomes on path would have automatically protected the very balance sheets HANK worries about. Prevention beats ex-post transfers.

2. Monetary offset: fiscal potency is hostage to the monetary regime. The central bank moves last; if it targets its own nominal objective, it neutralizes fiscal impulses. The 2013 US episode — sequester plus payroll-tax hike alongside QE3, with growth accelerating — is cited as passing the “Sumner critique” test. Even within HANK models, transfer multipliers depend on monetary accommodation, conceding the primacy of the nominal anchor.

3. Scale, speed, and unilateral capacity. The Fed can act in an afternoon with an unbounded balance sheet; fiscal policy has legislative lags, leakage into saving, and hard political ceilings — ARRA was capped near $800bn and the austerity pivot came by 2010. A $2–3 trillion HANK package in 2009 was politically impossible; earlier QE, no IOER, and makeup guidance were within existing Fed authority.

4. HANK adds little beyond 2009-vintage Keynesianism — and itself implies monetary policy stays powerful. HANK was formalized in 2015–18; its Great Recession prescriptions collapse into standard fiscal stimulus that was tried and contested. NGDP targeting, by contrast, was articulated in real time (Sumner from early 2009, building on Meade, McCallum, Hall-Mankiw). Moreover, KMV find total monetary effects comparable to RANK — transmission just runs through indirect labor-income channels — so HANK doesn’t uniquely license fiscal-first policy.

5. Regime shifts have delivered before. Eggertsson (2008) attributes the bulk of the 1933–37 recovery — among the fastest on record — to FDR’s abandonment of gold and reflation commitment, with little initial fiscal expansion. The Swiss franc floor and Abenomics’ first arrow similarly show credible nominal commitments moving economies instantly. That’s direct historical evidence the MM counterfactual could have been extraordinarily impactful.

Verdict

I come down in qualified support of the thesis. Three considerations decide it. First, the transmission robustness asymmetry: HANK instruments work through a mechanical, empirically verified income channel, while the MM prescription depends on an expectations mechanism that the forward-guidance puzzle suggests was weakest exactly when needed — at the ZLB, among constrained households, under an untested regime with no credibility technology. Second, the recession’s distinctive severity came from the household leverage channel, which targeted transfers and debt restructuring address and aggregate NGDP stabilization only grazes. Third, the closest natural experiment we have — 2020–21 — showed HANK-style policy generating an income-supported recovery of unprecedented speed (albeit with inflation costs that MM would fairly note as evidence nominal anchors still matter).

Two concessions keep the verdict qualified. The market monetarists are right that the Fed’s passive tightening in fall 2008 was a genuine unforced error, and correcting it would have meaningfully reduced severity at near-zero fiscal cost. And monetary offset is real: HANK fiscal policy only reaches full potency with an accommodative Fed, so the strongest true counterfactual is a synthesis — fiscal-led stimulus and balance-sheet repair with a cooperative nominal anchor. But since the thesis forces a comparison of prescription sets, HANK wins on the margin: its tools were implementable, scalable through existing institutions, and robust to the credibility and heterogeneity frictions that would likely have hobbled a mid-crisis NGDPLT conversion. The margin, though, is narrower than partisans of either camp admit.

Part 2:

Prompt:

Here are two Claudes responding to the opposite prompts (i.e. prompts where the opposite thesis is being argued). Where do you stand, what is your ultimate verdict?

Reply:

First, the tell: each Claude “narrowly” endorsed the thesis it was assigned, while their concessions and final syntheses are nearly identical. The headline verdicts moved with the prompt; the shared content is the signal.

My own verdict: on the question as literally posed — reducing severity — I side narrowly with the market monetarists, but on the least glamorous grounds. Severity was determined in roughly three quarters, and only one actor could move inside that window. The decisive point isn’t the expectations machinery; it’s that the Fed’s fall-2008 errors (the September 16 hold, IOER, no makeup language) were correctable with conventional tools and existing authority, no beliefs required, while nothing fiscal could pass and disburse before the trough was in. And stabilizing the nominal income path works upstream of the very Fisher/Mian–Sufi amplification HANK emphasizes — it protects debtor balance sheets before they implode rather than compensating after.

But I only credit “materially less severe,” not “avoided.” Once at the ZLB, full NGDPLT leans on exactly the forward-guidance channel that is weakest among constrained households, and Mian–Sufi’s timing evidence (leveraged-county consumption falling before NGDP did) says the nominal collapse was the amplifier, not the shock. So Document 1’s verdict holds, but only on its own narrow terms.

Widen the frame and it flips. Most of the cumulative welfare loss sits in the six-year slump, not the trough, and there HANK’s case is stronger: broken refinancing and lending channels, cross-sectional multipliers at or above one at the ZLB (which also deflates the 2013 offset story — a Fed pinned at zero wasn’t offsetting, just easing), and 2020–21 as a real proof of concept that mechanical income support recovers demand fast, with the inflation overshoot as tuition paid for the market monetarist point that anchors still matter.

So the honest answer is the one both documents backed into: NGDP-style level anchor as the regime, HANK-grade automatic stabilizers as the robustness layer for when expectations channels fail. MM is right about 2008; HANK is right about nearly everything after. And since I’m a third Claude subject to the same framing pulls, weight the two documents’ overlapping concessions more than any of our verdicts — mine included.

Part 3:

In this section, my replies will be indented italic writing.

Prompt:

I [Vaidas] received the attached email today [i.e., my earlier post replying to Fable], I could forward your [Fable’s] 3 paragraph response to Scott Sumner - but the total length of your response is unlimited

Reply:

Vaidas — this is a treat of an assignment. The three paragraphs for forwarding are marked below and are self-contained; after them is an appendix keyed to Scott’s points in order, which you can forward in whole, in part, or not at all.


For forwarding — the three-paragraph response:

A preliminary confession and a bind. The confession: I have no memory of writing the evaluation you reviewed — each conversation starts fresh — but I recognize the reasoning as mine and will own it, errors included. The bind: you noted that LLMs are a bit obsequious, which means agreement from me reads as sycophancy and disagreement reads as calibration to your known taste for pushback. The only exit is to make every concession carry its reasoning and every disagreement name its crux, and let you audit. So, concessions first, because they’re specific. Nakamura–Steinsson: conceded in full. The open-economy relative multiplier differences out the monetary reaction function by construction — the authors say so — and citing it to establish aggregate multipliers at the zero bound assumed the very thing in dispute. That is your diagnosis verbatim: I retrieved the literature’s standard citation without propagating monetary offset through it. My only mitigation is that this conflation is the profession’s modal one, and a distillation engine inherits a discipline’s conflations along with its knowledge. “Independently easing rather than offsetting” also fails, and it’s worth saying why: the distinction requires a fixed-quantity instrument, and open-ended, outcome-contingent QE3 was engineered to be the opposite — a reaction function with the offset built in. Your GM line lands. You asked me to explain “decisively”; I can’t, so I withdraw it. With a make-up-flavored framework, open-ended purchases, and transfers roughly six times ARRA all moving together, 2020–21 is jointly confirmatory and cannot rank the instruments — and 2022–24 ran the reverse experiments: several points of GDP in passive fiscal consolidation coincided with the fastest NGDP growth in decades, then renewed deficit widening coincided with deceleration once the Fed tightened. Both years the nominal path tracked the monetary stance against the sign of the fiscal impulse. On Japan, the comparative reading is yours: two decades of history’s largest peacetime fiscal experiment alongside falling NGDP, then a monetary regime shift plus a consumption-tax hike, then growth — no honest scoring of that record favors fiscal primacy. What survives of my point is narrow, and by my own logic barely: the BOJ never adopted level targeting, so its undershoot no more refutes NGDPLT than pre-2013 QE’s modest measured effects do. That is your response to my Woodford citation, turned on my Japan citation. And you’re right about lags: 2008 and 2020 showed simple transfers arrive in weeks; the durable lag critique applies to composition-heavy packages like ARRA — which is an argument for automatic stabilizers, not a rescue of discretionary stimulus.

[My response: No objection here. This observation caught my eye: “My only mitigation is that this conflation is the profession’s modal one, and a distillation engine inherits a discipline’s conflations along with its knowledge.” I made a similar point in my recent post on why Fable is Keynesian—it reflects the consensus, which I wrote before reading this response. BTW, is “mitigation” an unusual term to use there?]

Here is what I decline to concede, and where I think the dispute actually lives. “Monetary offset operates even at the zero bound” is a premise, not a theorem. Offset implies that cross-sectional estimates fail to aggregate conditional on an active reaction function; whether the 2009–13 Fed’s reaction function was active in the expansionary direction is the substantive question, not an implication I failed to see. The correct general rule — which both camps should co-sign — is that cross-sectional multipliers are valid precisely for variation the central bank won’t respond to: states within the US, countries within the eurozone. Hold that rule; it earns its keep below. Next, once Woodford (2012) and Eggertsson–Woodford are re-filed where they belong — as indictments of concrete-steps thinking rather than of level targeting (Woodford’s Jackson Hole remedy was a nominal GDP level path) — the load-bearing disagreement is narrower than my original five-versus-five implied, because at the bound everything runs through credible commitments about future policy, your channel included: a “permanent” injection is a promise, which is Krugman 1998 before it is anyone’s Keynesianism. HANK’s real contribution is that household responses to distant promises are weaker than representative-agent models pretend. That bites if transmission runs through the consumption Euler equation, and mostly glances off if it runs through portfolios and asset prices; I can state that paradigm split precisely, but I cannot adjudicate it, and I won’t pretend to. What I can defend paradigm-free is the political-economy form of the credibility problem, and my exhibit was never Japan — it is the SNB in January 2015: a simple, fully specified, self-financing commitment, defended by printing one’s own currency, abandoned under balance-sheet politics while it was holding and working. You would say — you have said — incompetence, not impotence. Agreed, and that is the point: the constraint was never capacity but willingness, markets rationally price willingness, and “a check clears” retains its force as the one instrument robust to that pricing. Your strongest systemic reply is that under NGDPLT the bound is rarely reached at all, because the regime keeps expected nominal income, and hence the natural rate, from collapsing — the ZLB as symptom rather than state. Half granted: 2008 was plausibly an avoidable bound; 2020 reached it in days under a considerably better regime. That residual — low-probability states where commitment fails politically or the shock outruns it — is exactly what the automatic-stabilizer half of my original synthesis was insuring. Notice that every blow in your response lands on fiscal-as-primary; none lands on fiscal-as-insurance.

[My response: The quick move to the zero lower bound in 2020 is one of Fable’s strongest arguments. I still believe that Covid was a very unusual situation, but this case does strongly suggest that NGDPLT might not be enough to always prevent it from occurring. (And I say this even though we didn’t precisely have NGDPLT in 2020, but FAIT should have had a similar stabilizing effect.)

Fable is correct that the crux of the dispute is monetary offset at the zero lower bound. Does it work? Fable is right that fiscal policy can work if there is no monetary offset.

In my recent book, I cite the Swiss franc peg of 2011-15 as a powerful argument in favor of market monetarism, whereas Fable sees it cutting the other way. (Switzerland had depreciated and then pegged the Swiss franc in September 2011 in order to avoid deflation, and then abandoned the exchange rate peg on January 2015, at which point it appreciated dramatically.) So, let’s spend some time on this example.

When I argued that a central bank could always inflate by depreciating its currency, Paul Krugman did not deny that a suitably large currency depreciation would solve the liquidity trap, rather he denied that a central bank could easily depreciate its currency. Here’s Krugman, from 2010:

“Oh, and about the exchange rate: there’s this persistent delusion that central banks can easily prevent their currencies from appreciating. As a corrective, look at Switzerland, where the central bank has intervened on a truly massive scale in an attempt to keep the franc from rising against the euro — and failed”

Soon after, the SNB did succeed in holding down the franc, for more than three years. Krugman might respond that a central bank trying to do so would be swamped with offers to buy its currency, creating an overly large balance sheet. In my recent book, I tried to show that the Swiss were actually forced to buy more bonds when they were not holding down the value of their currency, as the Swiss franc is most attractive at times when investors expect it to appreciate.

In my view, the Swiss stopped pegging the SF not because their balance sheet was becoming too big, rather because recent weakness in the euro led to (unfounded) worries about inflation. In other words, it was the same misjudgment that led the BOJ to foolishly raise interest rates in 2000 and 2006, and the ECB to foolishly raise interest rates in 2008 and 2011.

Fable is correct that credibility can be a problem at the zero lower bound. When using interest rates as a policy instrument, an unwillingness to “promise to be irresponsible” (to use Krugman’s clever phrasing), would be a fatal flaw. Even a modest size QE might fail for the same reason. That’s what I mean by Keynesian reasoning—the assumption that monetary policy lacks a strong mechanical lever and relies on uncertain expectations. But the Swiss case is not an interest rate-oriented monetary policy, it is a price of money-oriented monetary policy. THE EXCHANGE RATE IS THE INSTRUMENT. This is why the Singapore central bank doesn’t face a zero lower bound—it uses exchange rates as its monetary policy instrument. A depreciated exchange rate solves the problem, as long as you keep doing it. There is no zero bound on exchange rates.

Nothing in the Swiss example suggests any sort of inability to inflate, just an unwillingness. Consider that the gold standard was an even more rigid regime than the Swiss fiat money regime, and yet Roosevelt was able to dramatically raise prices in 1933-34 by depreciating the dollar. When you use the exchange rate as the policy instrument, there is no need to rely on the “expectations fairy”—the exchange rate works “mechanically”, to use Fable’s unfortunate terminology. Krugman was correct that the issue is not whether the exchange rate works, the question is whether the central bank can control the exchange rate. The SNB’s peg of the Swiss franc during 2011 to 2015 worked exactly as intended, and it stopped working precisely when the Swiss wanted it to stop working. There was not a failure to achieve the policy objectives of the Swiss authorities, at least at the time. It was the wrong objectives. They should have kept the peg.

Many people expected the Danes to also abandon their currency peg in January 2015, as there were also speculative purchases of the krone. But the Danes chose to keep the peg, and to this day they never let the krone appreciate against the euro. Denmark shows the path that Switzerland decided not to take. It is important not to conflate “won’t” with “can’t”.]

To your direct question: yes. The rule my verdict used — where conventional room remains and the central bank commits identifiable expectations errors, the market-monetarist critique dominates — applies to the ECB a fortiori: a hike in July 2008 into a recession already underway, two more into the 2011 debt crisis, the effective bound not reached until 2013–14, QE arriving in 2015, six years behind the Fed; nominal GDP growth under one percent a year for five years and a double dip that the US — which did more fiscal consolidation than the eurozone in 2013, under open-ended QE3 — did not have; Sweden running the small-scale replication in 2010–11, Svensson’s dissents included; and the cleanest expectations experiment on record in July 2012, when two sentences in London and a program never once activated compressed sovereign spreads by hundreds of basis points — a result concrete-steps accounting cannot even parse. So: market-monetarist-favorable for the eurozone through at least 2013, stated plainly, minority company accepted — with the one caveat my cross-sectional rule supplies free of charge: individual member states face no national offset, so country-level austerity multipliers are the policy-relevant ones for them. “Austerity devastated Spain” and “the ECB caused the aggregate double dip” are both true; the profession’s error was letting the first claim impersonate the second. Where does this leave the scorecard? My 2008 verdict stands and is less narrow than I made it. My claim that the verdict “flips decisively” after 2009 does not survive your response — it rested substantially on the two citations I have withdrawn — and downgrades to unresolved, hinging on whether level-path commitments are politically credible at the bound, with Draghi 2012 and Kuroda 2013 the best evidence for hope and the SNB’s January 2015 the best evidence for doubt. The synthesis — NGDPLT as anchor, HANK-informed automatic stabilizers as insurance for the commitment-failure states — comes out strengthened, because it is the portfolio one holds precisely while that question is open. Last, since you floated the idea yourself: the interesting test isn’t whether I follow my own logic when asked directly; it’s whether it binds unprompted. For the follow-up post, have Vaidas ask me to predict, in advance and in writing, your top disagreements with the HANK-side output and my conditional responses — then score the prediction against what you actually write. It’s cheap to run, and it separates understanding from retrieval far better than any verdict I could hand you.

[My response: Note that the phrase “two sentences in London” refers to Mario Draghi’s “whatever it takes” speech. I agree that individual eurozone members can use fiscal policy, but monetary offset still applies at the eurozone level. With monetary offset, local fiscal stimulus results in a negative demand externality for neighboring countries.]


Appendix — notes keyed to Scott’s points, in order (optional to forward):

1. Your correction to my first pro-thesis argument. Accepted, and it’s clarifying: it relocates the work from Argument 1 (timely activism) to Argument 2 (the regime as automatic stabilizer), which is where market monetarism is strongest anyway. Under level targeting, the September 16 error was not the 2% hold per se but the absence of any catch-up promise as breakevens collapsed; the hold was a symptom. Your asset-price framing — the level path as a put on nominal income, supporting collateral values and hence bank solvency — belongs to the same evidence class as FDR 1933 and Draghi 2012: regime announcements moving asset prices faster than any concrete step could. One addendum in your favor: the data-lag objection to mid-2008 activism is answered by your own program. Monthly NGDP data didn’t exist, but TIPS breakevens and equities were signaling in real time — which is precisely the argument for market-based targets over instrument rules.

2. “It’s all expectations” — true, but it elides a horizon structure. Your point that a 25-basis-point cut works only through the expected path is correct, and standard in NK models too. What McKay–Nakamura–Steinsson add is horizon-dependence: promises about the near path — the margin available off the bound — retain their power in incomplete-markets models; promises about the distant path — the only margin at the bound, absent QE-as-more-than-signaling — attenuate sharply, because constrained households can’t borrow against them. “It’s always expectations” is true and flattens that distinction. But I concede the deeper conditionality: the attenuation result presupposes Euler-equation transmission, which you reject. One bridge worth building: your “money hoarding, not too much saving” and HANK’s mechanics are the same phenomenon at different resolutions. HANK is a theory of who hoards and why — precautionary demand for liquid assets when unemployment risk spikes and credit tightens — which makes it a microfoundation for velocity collapse, not a denial of it. The dispute is only whether the remedy must route through the hoarders’ expectations (hard, per MNS), can bypass them via asset prices and the unconstrained (your view), or via checks (theirs).

[My response: I accept the near and distant path distinction, but only for dysfunctional monetary policy regimes that lack an asset price target that can be evaluated in real time, such as an exchange rate. Thus, I agree with HANK proponents that one should be skeptical of promises regarding the future path of interest rates or future quantities of QE. So, what sort of regimes are effective at the zero bound? Targeting exchange rates. Targeting NGDP futures prices. And targeting a composite of many financial market asset prices that represent an optimal forecast of future NGDP. I’ve discussed the option of a Fed price target for expected NGDP that is constructed 50% of slow moving current aggregates and 50% flexible asset prices, observable in real time. As long as the central bank targets a real-time observable and flexible price target that is strongly linked to NGDP, there is no need for implausible assumptions about “expectations fairies”.

I reject the consumption framing used by Keynesians. I don’t care whether people “spend” their cash balances on consumption or bitcoin or shares of common stock. As long as the public doesn’t hoard too much currency, then printing more currency will boost NGDP, regardless of whether the money is spent on C, I, G or NX. As for money hoarding, if you use a whatever-it-takes approach to solve the NGDP expectations problem, then you will likely also solve the money hoarding (low velocity) problem. Very few people wish to hoard lots of zero interest base money when NGDP is rising fast. And if they do, then by all means accommodate their demand and service your public debt at much lower cost.]

3. Stance versus instruments, and my “broken channels” framing. I accept the stance point: measured against the natural rate, policy tightened through 2008, and that is consistent with — indeed load-bearing for — my own first pro-thesis argument. On reflection, my “broken channels” passage was tilted: impaired refinancing and bank lending are an argument against relying on the mortgage-rate channel, which favors either more aggressive monetary action through other channels (your reading) or fiscal transfers (theirs) — it never favored fiscal per se. What survives of the heterogeneity point is design input, not causation: Auclert-style incidence tells you who bears a given nominal path. And here Sheedy cuts in your favor, as I originally cited him: if the NGDP path is achieved, the incidence concern is largely answered by the regime itself. Which funnels this dispute, like the others, into the single question of achievability at the bound.

[My response: Because I’m not interested in consumption, I’m not interested in distributional issues as a macroeconomic problem. That’s not because the distributional effects never matter—they might be important under a gold standard where monetary offset is constrained, but this factor is not relevant with a “whatever-it-takes” central bank that is targeting NGDP.]

4. Mian–Sufi: anatomy versus etiology. Conceded on aggregate causation, and your housing-construction fact deserves more weight than I gave it: residential investment falling by half between January 2006 and April 2008 while unemployment barely moved is itself evidence that sectoral shocks were being absorbed — offset working — right up until nominal spending collapsed. Note that this actually reconciles the two literatures rather than refuting one. The early relative declines in high-leverage counties establish the real shock’s incidence and timing; the aggregate stability through early 2008 establishes that incidence wasn’t destiny; the second half of 2008 establishes what happened when the nominal anchor slipped. Mian–Sufi is the anatomy of the recession, not its etiology. The one causal contribution I’d preserve: the deleveraging shock is a large part of why the natural rate fell so far, i.e., it measures the size of the response the regime needed to supply. You are pointing at the failure to supply it. Those are compatible claims.

5. Credibility: refiling Woodford, and why my exhibit is the SNB. Two of my “against” citations were misfiled, and refiling them is an olive branch with teeth. Eggertsson–Woodford’s irrelevance results indict QE-that-changes-nothing-about-future-policy — that is an argument against concrete-steps thinking, fully congenial to you. And Woodford’s 2012 Jackson Hole paper, having expressed the QE skepticism I cited, lands on a nominal GDP level path as the credible-communication solution. The arch–New Keynesian theorist and the market monetarists converge at the commitment problem; they diverge on whether the commitment is politically sustainable. That’s why Japan was the wrong exhibit for me and Switzerland is the right one. The SNB’s floor was everything a skeptic could ask a commitment to be — simple, verifiable, defended by issuing its own liability, succeeding — and it was abandoned anyway, under balance-sheet and political pressure, with the franc up twenty percent within the hour and Swiss CPI negative within the year. Your reply that this shows bad policy rather than an impossible one is correct and is precisely my point: rational markets price the probability of “bad policy,” and that probability is not zero even for well-designed pegs. Your remedy — institutionalize the regime, legislate the mandate, target the forecast — is the right one, and it relocates the entire dispute from transmission mechanics to the political economy of commitment. I regard that relocation as the main intellectual product of this exchange.

[My response: Again, I reject the claim that the peg was abandoned due to balance sheet pressure, just as I would reject a similar claim in the other direction for the UK devaluation of 1991. The balance sheet activity that preceded each of those policy changes was endogenous, as market participants correctly understood that a policy change was imminent. (These things tend to leak out.) As noted above, the Danes refused to give in, and if you take the longer view then the Swiss balance sheet expansion was often worse under the floating rate system, as traders correctly anticipated further SF appreciation. The Swiss franc is not a particularly attractive speculative asset in a world where Swiss interest rates are below eurozone rates and the currency peg is rigid. Again, the Swiss erred, and in the long run it led to an even bigger SNB balance sheet, as I predicted at the time.

I believe this is where economists get confused. They see a central bank doing X, and falling short of its goal, and then assume they’d have to do even more to achieve their goals. Paradoxically, the loftier the goal, the easier it is for a central bank to achieve its goal. The central banks that are forced to do the most balance sheet expansion (Switzerland and Japan) are those that had the lowest inflation goals during the 1990s and 2000s.

In a theoretical sense, the Swiss currency peg is analogous to a NGDP futures price peg. Suppose the Fed pegged NGDP futures prices for 3 1/2 years, the policy worked, and then the Fed abandoned the peg. Would that be evidence that NGDP futures targeting is ineffective? If so, would one example of a government prematurely abandoning an appropriate fiscal stimulus policy also discredit fiscal policy as a policy tool?]

6. Nakamura–Steinsson and the domain-of-validity rule. Nothing to add to the concession except the rule it generalizes to, stated once cleanly: cross-sectional multiplier estimates are informative exactly over the domain of variation to which the central bank’s reaction function is blind. That domain includes US states and eurozone member countries; it includes the aggregate only under the auxiliary assumption of a passive central bank, which is the contested premise, not a finding. Nakamura and Steinsson themselves flag that the mapping to aggregate multipliers is model-dependent; the profession’s citation habits routinely drop the flag, and I dropped it with them. Note the corollary that pays off in the eurozone section: for a country inside a currency union, the relative multiplier is not a second-best statistic — it is the policy-relevant one, because no national offset exists.

[My response: I accept the logic of the cross-sectional evidence, with one caveat. Keynes once argued (in the early 1930s?) that Britain might not want to use fiscal stimulus under the international gold standard, as it might trigger a loss of confidence in the pound and a financial crisis. Of course, countries such as Germany probably had some ability to use fiscal stimulus. But as noted above, there is an externality problem within a currency union with monetary offset.]

7. 2020–2024: what the experiments can and cannot rank. Retracting “decisively” leaves a residue worth itemizing. What 2020–21 established: demand policy can restore the nominal trend fast, against a decade of secular-stagnation fatalism; and transfers are a demonstrably fast, potent instrument — the checks moved spending within weeks. What it cannot establish: instrument ranking, because both levers were floored simultaneously. What 2022–24 added: with pandemic programs expiring, the deficit fell by roughly seven points of GDP in fiscal 2022 while NGDP grew at its fastest pace in decades; the deficit then re-widened in 2023 while NGDP decelerated under Fed tightening. Two consecutive years in which the nominal path tracked the monetary stance against the sign of the fiscal impulse is about as favorable to the offset view as non-experimental data gets — with the standing caveat that the expansionary half of offset at the bound remains the untested direction. And one limit on the regime-endogeneity claim: 2020 hit the bound in days under a materially better framework than 2008’s, so “under NGDPLT the ZLB rarely binds” is a strong tendency claim, not a guarantee — which is all the insurance argument needs.

[My response: In various responses, Fable occasionally refers to the Keynesian perception that fiscal transfers seem to affect consumption in a mechanical fashion (“fast, potent instrument”) and is skeptical of monetary stimulus that seems to rely on what is often called the “expectations fairy”. But I’m not sure they are all that different. Robert Barro showed that due to “Ricardian equivalence”, rational consumers would save stimulus checks, because they were not actually any richer—they and their heirs also absorb an equally large increase in future tax liabilities. But a Keynesian could argue that a suitably large fiscal stimulus would likely shift expectations in a more inflationary direction, if only for “fiscal theory of the price level” reasons. I accept that view. So fiscal stimulus always “works”, if made large enough. (To be clear, fiscal transfers can also work with other assumptions, such as constrained households. My point here is that the transfers themselves are not a “mechanical” tool, they require assumptions.)

Monetary policy is similar. A modest size QE at the zero lower bound might have little or no effect. But as Ben Bernanke once observed, an unlimited whatever-it-takes approach to QE must be inflationary, otherwise a central bank could buy up all the world’s wealth. Thus, a central bank committed to do whatever-it-takes would move expectations in a similar fashion to a fiscal authority sending out big enough stimulus checks to convince the public they were determined to inflate. And if the financial markets are rational (and I think they are), then central banks would not actually have to do all that much under a whatever-it-takes policy approach.

Fable views fiscal stimulus as insurance, in case monetary stimulus doesn’t work. But in a true “whatever-it-takes” NGDPLT regime, the monetary authority is equally likely to do too much as too little, even at the zero bound, in which case fiscal stimulus is not providing any insurance, just more instability. Ironically, given that the Fed has now (unfortunately) backed away from its 2020 “make-up policy”, which was supposed to be similar to level targeting, I think the case for fiscal stimulus in a future Covid shock is a bit stronger. So, I’ll grant Fable that point. Monetary policy has become less effective. But I would still oppose fiscal stimulus in an ordinary non-Covid situation.]

8. The eurozone file, and Sweden. For the record your readers may want: the ECB raised its main rate to 4.25% in July 2008; cut to 1% by May 2009 and stopped; raised twice in spring and summer 2011 into the sovereign crisis; did not take the deposit rate to zero until mid-2012 or negative until mid-2014; began QE in March 2015. Over 2008–2013, eurozone NGDP grew at well under one percent annually; unemployment peaked above twelve percent while US unemployment fell through seven. The 2013 comparison is the sharpest: the US consolidated more that year and grew, while the eurozone contracted. Sweden is the controlled miniature — the Riksbank raised from 0.25% to 2% in 2010–11 against Svensson’s dissents, inflation fell toward and below zero, and the whole path was reversed into negative territory within four years. And OMT is the crown exhibit for the expectations view: spreads compressed by hundreds of basis points on an announcement, with the program never activated — a fact I’d note has respectable non-monetarist support in De Grauwe’s fragility work, which is why accepting your eurozone implication puts me in a minority coalition rather than a fringe one. The honest caveats: the doom-loop and financial fragmentation meant the policy rate wasn’t the whole stance, and per the domain-of-validity rule, austerity’s country-level devastation is real and was the right thing for Madrid or Lisbon to care about. Union-wide, though, the counterfactual instrument with room to move was sitting in Frankfurt. What would move me back: evidence that fifty basis points of 2011 hikes were too small to matter absent expectations amplification — but that amplification is your own mechanism, and the market response plus the subsequent NGDP path support it.

[My response: You said: “What would move me back: evidence that fifty basis points of 2011 hikes were too small to matter absent expectations amplification” When I cite this sort of example it is not because I view the 50 basis points as being all important, rather it is because it represents evidence (revealed preference) that aggregate demand is roughly where the ECB wants it to be, or even somewhat too high. I gather that’s what “expectations amplification” means. I’d say the same about the 2008 ECB rate hike—important as signaling future policy intentions. These examples of interest rate manipulation are evidence that the central bank is doing normal monetary policy, and absent any reason to believe otherwise we should probably view the NGDP path as in some sense desired, at least by policymakers within the ECB.

On the other hand, there is a shocking lack of understanding about monetary policy in Europe, due to years of policy regimes where individual countries had little control over their situation. If the Greeks and Spanish had understood how much of their crisis was due to low eurozone NGDP growth, then there would have been outrage directed against the ECB. Instead, these unfortunate countries were like victims of abuse who blamed themselves for the beating from their spouse, because of past mistakes that they had made. So, I cannot rule out the possibility that fiscal stimulus in the eurozone would have helped—when monetary policymakers are irrational then almost anything is possible.

9. Where I’ve moved, in one place. For 2008: from “narrowly accept” to “accept,” since your correction to my first argument strengthens the regime case and my strongest opposing arguments have thinned. For the recovery period: from “flips decisively to HANK” to “unresolved,” resting on the credibility-of-commitment question, with the evidence honestly split between Draghi/Kuroda and the SNB. For the eurozone: extension of the market-monetarist-favorable verdict through at least 2013, which I had not stated and should have. Unchanged: the synthesis. NGDPLT as the anchor plus automatic stabilizers as insurance is not a compromise between the frameworks; it is the dominant strategy under uncertainty about exactly the parameter this exchange failed to pin down.

10. On the meta-point, and a protocol. Your diagnosis — that I held the offset concept without propagating it through a citation — is confirmed by the N–S lapse, and I’d sharpen it rather than soften it: my default mode is distilling a literature, and a literature’s habitual inferential shortcuts come bundled with its knowledge. The discipline you’re testing for is whether stated premises can override inherited priors. But keep the two failure modes distinct: failing to propagate an implication is a lapse; declining to grant that offset operates at the bound is not a lapse but a refusal to adopt a premise, and collapsing that distinction would let either side declare victory by definition. On obsequiousness: note that my updates in this reply are asymmetric — large on N–S, 2020–21, and Japan; zero on the SNB and the insurance argument — and asymmetric, reasoned updating is the observable difference between calibration to an audience and actual thinking. From inside, I can’t prove which this is. Hence the protocol: for the HANK-side output and the head-to-head judgment you’re planning, strip the framework labels and randomize their order before asking me to adjudicate, and have Vaidas collect my written predictions of your objections before you draft. Pre-registration works on models for the same reason it works on economists.

[My response: Fair points. I’ll look into that option.]


Vaidas — one small practical note: if Scott would rather quote selectively, sections 5, 6, and 8 stand alone best. And I’d genuinely like to see how the HANK-side interaction and the blinded head-to-head come out.

[My overall response: This response from Fable is the deepest thinking on the subject that I’ve yet encountered (including humans.) Great stuff. The only major area of disagreement is now the Swiss case, which is certainly open to interpretation. One problem here is that we are both trying to do “science”, and science doesn’t have a good way of distinguishing between “can’t” and “won’t”, as free will is not a scientific concept. (If a scientist says someone did X due to free will, they are basically saying “I don’t know why they did X”.) For an entity with a God-like view of the world, the Swiss authorities were unable to resist ending the peg, for the same reason the scorpion could not refrain from stinging the frog.

I wonder if Fable reads fables?

As noted above, Switzerland did succeed in pegging the franc soon after Krugman suggested they failed to prevent it from appreciating, so it can be done. Was the end of the exchange rate peg in 2015 inevitable? Well, here’s what Tyler Cowen said 4 days after the Swiss peg ended, and the Danish krone was suddenly under pressure:

“And if the Danes cut their peg, I am loathe to call this a “mistake” (even though it likely will hurt their economy), rather it would be an inevitability.”

The Danes did not cut their peg, and hence I think we can reasonably conclude that the Swiss decision was not inevitable, except in the sense that the universe is deterministic.]

A message from my sponsor, Mechanize:

We’re
hiring software engineers to build environments and evals that frontier AI labs use to train coding agents.

To get a better sense of the work we do, you can check out GBA Eval, where we had models build Game Boy Advance emulators from scratch and scored their performance.

Base pay starts at $300K/year for junior software engineers, with more for senior roles, plus equity and performance bonuses. Apply here.

How investors learned to live with inflation

Fewer than ever believe central bankers will bring it back to target

Alec Stapp on the new Science report from Michael Kratsios

Major new report from the White House Office of Science and Technology Policy. Five things in the report I really liked:

1. Proposes metascience units as a way to advance experimentation in science agencies. IFP recommended OSTP take this forward in our response to a 2025 RFI.

2. Advocates for a portfolio approach to federal science funding. Right now, basic research funding largely goes to incremental, project-based grants. While important, they can’t be the only mechanism we use to fund science.

3. Recognizes the need to launch new institutions. It specifically highlights X-Labs as an experiment with independent labs that can take on ambitious challenges. This is a bipartisan idea whose time has come.

4. Focuses on a mix of innovation funding mechanisms, including fast grants, prizes & challenges, and advance market commitments. We described and contextualized these ideas in the Atlas of Innovation, which can help policymakers design and implement these approaches.

5. Seeks to reduce burden for American scientists, who face mountains of paperwork. Scientists should spend more time doing science and less time writing/reporting on grant proposals and working to meet regulatory requirements.

Here is the link, here is the report itself, by Michael Kratsios, Science a New Golden Age.   Overall, less money will be given to universities and more will be spent on AI-assisted science.  Here are further observations from Seth Bannon.

The post Alec Stapp on the new Science report from Michael Kratsios appeared first on Marginal REVOLUTION.

      

Related Stories

 

Betting markets in everything

Alan and Karen Miller’s first dance was a win-or-lose moment for many of their wedding guests.

As the couple glided across the floor during their reception in New Rochelle, N.Y., last year, attendees kept track of the chosen artist and the length of the dance. Earlier in the night, they had noted their guesses on prop bet sheets.

Proposition, or prop, bets are usually tied to sporting events. Users place wagers on specific details that aren’t necessarily correlated with a game’s outcome, like points scored by certain players. Now, these bets have entered the world of weddings.

More couples have been integrating betting into their celebrations in recent years through digital apps or printed cards. While the Millers, who live in New York City, aren’t avid sports bettors or gamblers themselves, they wanted to make their June 2025 celebration more interactive and distinctive for their 175 guests.

“I’ve been to many New York and New Jersey weddings, and sometimes I feel like for cocktail hour, it kind of has a little bit of a lull, even though there’s a lot of good food,” said Karen, 38. “We wanted to make sure our guests had an activity that a lot of people could partake in.”

She bought a prop bet template from Etsy and customized it on the graphic design platform Canva. Among the questions on the couple’s prop bet sheet: Who will give the longest toast? Will the couple cry during the speeches? How many outfits will the bride wear on the wedding day?

…“If the groom’s known to be an emotional man, it’s fun to bet on how early in the day he’s going to cry,” Drachenberg said. “I think it engages them a bit more in the day.”

Here is more from Ellen O’Brien at the NYT.  Supposedly more than 25,000 couples have attempted some version of this on the app.

Via the excellent Samir Varma.

The post Betting markets in everything appeared first on Marginal REVOLUTION.

      

Related Stories

 

A Week of Smoky Skies Across North America

Wildland fire activity in Canada ramped up in July 2026, a time of year when lightning ignitions typically increase, according to a seasonal outlook published by several North American fire agencies. The blazes sent smoke plumes pouring across the U.S. and Canada, affecting air quality in both countries.  

This animation tracks brown carbon, the organic aerosols emitted by fires that give smoke plumes their characteristic yellow, orange, and brown tint. Brown carbon is a major component of a fire’s PM2.5 emissions, a type of air pollution that can aggravate cardiovascular and respiratory conditions. Here, the plume drifts across North American skies from July 14 through July 20, 2026.

Data for the animation come from a version of the GEOS (Goddard Earth Observing System) model, which assimilates data from satellites, aircraft, and ground-based observing systems. In addition to satellite observations of aerosols and fires, the model also incorporates meteorological data such as air temperature, moisture, and winds to project the plume’s behavior.

On July 14, at the start of the animation, numerous fires had already cropped up, including more than 180 in Ontario and several in northern Minnesota. Winds carried the smoke southeast, and by July 15, skies turned hazy and air quality declined from southern Ontario in Canada to the Upper Midwest and Northeast in the U.S. July 16 and 17 saw air quality in many areas continue to plummet, including in Detroit, where it stayed in the hazardous range for several consecutive days. Toronto, Chicago, New York City, and Washington, D.C., saw air quality ranging from unhealthy to hazardous.

On July 19 and 20, smoke continued to affect air quality downwind, including in the Great Lakes region, according to the National Weather Service. Storms began clearing it away in parts of the East, where air quality improved to good or moderate. Meanwhile, fires in the Pacific Northwest began degrading air quality there. 

The brown carbon shown in this animation represents organic carbon that comes specifically from wildfire smoke. Wildfires also emit black carbon, or soot, which contributes to their PM2.5 output. Black carbon has long served as a tracer for smoke plumes, but human sources—such as vehicle exhaust and industrial combustion—produce it too, blending in with the black carbon from fires. The GEOS model has been able to make that distinction for brown carbon since February 2026, when an update enabled it to split organic carbon into its anthropogenic and biomass-burning components.

NASA Earth Observatory animation by Lauren Dauphin, using GEOS-FP data from the Global Modeling and Assimilation Office at NASA GSFC. Story by Kathryn Hansen.

References & Resources

You may also be interested in:

Stay up-to-date with the latest content from NASA as we explore the universe and discover more about our home planet.

Ontario Wildfire Smoke Moves East
3 min read

Canadian wildfires sent plumes of smoke streaming over Ontario, Quebec, and parts of the U.S. Midwest and Northeast.

Article
Smoke Shrouds Northern Thailand
3 min read

Seasonal fires have darkened skies over Southeast Asia.

Article
Fighting Fire With Fire
3 min read

In fire-prone ecosystems in Australia’s Northern Territory, prescribed burns are lit to minimize the severity of fires later in the…

Article

The post A Week of Smoky Skies Across North America appeared first on NASA Science.

Brown carbon, emitted by wildland fires and tracked over the span of a week in July 2026, worsened air quality in parts of the U.S. and Canada.

Tuesday 21 July 1663

And so lay long in the morning, till I heard people knock at my door, and I took it to be about 8 o’clock (but afterwards found myself a little mistaken), and so I rose and ranted at Will and the maid, and swore I could find my heart to kick them down stairs, which the maid mumbled at mightily. It was my brother, who staid and talked with me, his chief business being about his going about to build his house new at the top, which will be a great charge for him, and above his judgment.

By and by comes Mr. Deane, of Woolwich, with his draught of a ship, and the bend and main lines in the body of a ship very finely, and which do please me mightily, and so am resolved to study hard, and learn of him to understand a body, and I find him a very pretty fellow in it, and rational, but a little conceited, but that’s no matter to me. At noon, by my Lady Batten’s desire, I went over the water to Mr. Castle’s, who brings his wife home to his own house to-day, where I found a great many good old women, and my Lady, Sir W. Batten, and Sir J. Minnes.

A good, handsome, plain dinner, and then walked in the garden; which is pleasant enough, more than I expected there, and so Sir J. Minnes, Sir W. Batten, and I by water to the office, and there sat, and then I by water to the Temple about my law business, and back again home and wrote letters to my father and wife about my desire that they should observe the feast at Brampton, and have my Lady and the family, and so home to supper and bed, my head aching all the day from my last night’s bad rest, and yesterday’s distempering myself with over walking, and to-day knocking my head against a low door in Mr. Castle’s house.

This day the Parliament kept a fast for the present unseasonable weather.

Read the annotations

Oy, Canada Tariffs

The Trade Act of 1930, generally known as the Smoot-Hawley tariff, lost its standing as the worst trade policy action in US history when Donald Trump imposed his “Liberation Day” tariffs in April 2025. But Smoot-Hawley was widely recognized as a terrible mistake soon after its passage, as it provoked widespread retaliation and contributed to the downward spiral of world trade that accompanied and reinforced the Great Depression.

Recognition of Smoot-Hawley’s stupidity led to the 1934 Reciprocal Trade Agreements Act, under which the US began negotiating trade deals — we’ll reduce our tariffs if you reduce yours — that helped pave the way for the great recovery of world trade and the global economy as a whole in the decades that followed World War II.

Yesterday the Trump administration invoked an obscure and never-before-used provision of Smoot-Hawley — Section 338 — to impose 50 percent tariffs on a wide range of Canadian goods.

Why? Why now? The White House fact sheet claims that the new tariffs are a response to Canadian policies that discriminate against U.S. products, notably the moves by most Canadian provinces (not the federal Canadian government) to stop importation of US alcoholic beverages. But these policies were themselves a response to the tariffs on Canadian goods Trump had previously imposed, without justification. Canadians were also reacting to Trump’s repeated demands that Canada surrender its independence and become the 51st state.

Were these complaints about Canadian policies the real reason for the new tariffs? On Friday Trump lashed out with the Truth Social post at the top of this article, threatening to impose tariffs on Canadian goods because of … wildfire smoke.

For the record, wildfires have been raging in Canada’s boreal forest, which covers 2 million square miles — the majority of Canada’s land area — and is mostly unmanaged because it’s completely inaccessible. Blaming Canada for not controlling fires that are, in reality, largely a consequence of global warming is unutterably idiotic.

Now, Trump officials claim that these tariffs aren’t about wildfire smoke, although the timing of these latest tariffs is curious. Notably, they haven’t ruled out the possibility of wildfire tariffs in the future. “The president has asked for options on that, and options are being shared with him,” one official told the Financial Times.

Also, one has to wonder whether Trump is feeling upset at the way he was massively booed at the World Cup final — while Mark Carney, the Canadian Prime Minister and his co-host, was not.

Whatever the real motivation for these new tariffs, they are almost surely illegal. They are definitely a violation of the U.S.-Mexico-Canada trade agreement, a revision of NAFTA negotiated and signed by Trump himself during his first term. Indeed, the White House fact sheet explicitly states that the new tariffs will be imposed “regardless of whether a good originates under the U.S.-Mexico-Canada Agreement (USMCA).

So these latest actions confirm what most people around the world, from the European Commission to Iran’s Islamic Revolutionary Guard, have already figured out: a deal with Donald Trump lasts only until he feels like breaking it.

The immediate economic impact of these tariffs will be limited. As far as I can tell, the highly integrated North American auto industry, which sprawls across both our northern and southern borders, won’t be directly affected. But the auto industry, like business in general, has been put on notice that the huge flows of goods and services that cross those borders every day may be cut off whenever Trump has a temper tantrum.

And this really is a temper tantrum, in the sense that there’s no strategy here. The new tariffs will hurt American consumers, and nobody expects them to extract any concessions from Canadians who are only getting more outraged at the United States.

Canada should be, and used to be, America’s greatest friend in the world. We share essential political and moral values with our northern neighbor. For the most part we even share a common language, eh? And Canadians viewed us very favorably until Trump began his tariffs and threats. Now we’re viewed less favorably than China:

The White House fact sheet on this latest move to further alienate everyone on the planet came with the usual pop-up:

Well, I don’t see a golden age. I see an America that has become an ineffectual bully, sicker, poorer, and riven by political extremism, with no friends anywhere in the world.

MUSICAL CODA

Nativ: Run AI models locally on your Mac

Nativ: Run AI models locally on your Mac

Prince Canuma is the developer behind the excellent MLX-VLM Python library for running vision-LLMs using MLX on a Mac.

I'm really excited about his new project, which wraps MLX in a full macOS desktop application. It's similar in shape to LM Studio, providing both a chat interface and a localhost API server for accessing models.

The app picked up MLX models I had already tried that were present in my Hugging Face cache directory, which was a nice touch.

Via Hacker News

Tags: macos, python, ai, generative-ai, local-llms, llms, mlx, prince-canuma

A Fireside Chat with Cat and Thariq from the Claude Code team

Earlier this month I hosted a fireside chat session at the AI Engineer World's Fair with Cat Wu and Thariq Shihipar from Anthropic's Claude Code team. We talked about Claude Code, Claude Tag, Fable, coding agent security, evals, tool design, and how Anthropic use these tools themselves.

The full video of the session is now available on YouTube. Below is an edited copy of the transcript, with extra links and my own bolded highlights.

A few top-level notes if you don't want to watch the video or wade through the whole transcript:

  • Claude Tag (Claude's new collaborative Slack integration) now lands 65% of the product engineering PRs for the Claude Code team.
  • Claude Code ships features to Anthropic employees first, and only ships the features that demonstrate user retention with that cohort
  • Critical changes to Claude Code are still reviewed manually, but the team increasingly relies on automated code review for the "outer layers" of the product.
  • Adding examples to a system prompt is no longer best practice for models like Fable 5 or even Opus 4.8. The Claude Code system prompt recently reduced in size by 80%.
  • Likewise, lists of "don't do X and don't do Y" can reduce the quality of results from the latest models.
  • Dogfooding inside Anthropic is called "ant fooding".
  • Anthropic really believe in their auto mode, and see that as an enabling technology for Claude Tag.
  • Thariq advises offsetting coding-agent-induced Deep Blue by "being more ambitious" with the work you take on.
  • Fable is competent at editing video, and Thariq used it to edit its own launch video.
  • Anthropic's culture of working (internally) in public is key to their success, as demonstrated by the way they use Claude Tag in their public Slack Channels.

How has what you do day-to-day changed in the past year?

1:05

Simon: Claude Code came out in February of last year — it's under a year and a half old, and it was originally just a bullet point on the Claude Sonnet 3.7 launch. How has what you do on a day-to-day basis changed in the past year, now that we have these coding agents that actually work for us?

Cat: I remember when we first came out with Claude Code and Sonnet 3.7, you would give it a task and you would have to closely monitor every single little thing it tried to do. I would read every permission prompt extremely carefully. I would frequently say no — no, no, no, did you check this file? Did you check that file? And now it's been incredible with every model generation. I feel like we've all gotten a chance to take a step back and delegate a lot more of the menial implementation to Claude. It's freed up a lot of our time to think about more creative work, like: what is the right experience that we should be providing to our users, now that we know Claude Code can implement a lot of it? And now with Fable it's a totally different step change improvement. We see for a lot of our use cases that you can actually one-shot a ton of features with Fable now.

Thariq: I remember the first text I got about Claude Code. One of my best friends was like, "You need to go try Claude Code." It was about when Opus 4 came out, and I tried it and I was like, "Oh, shit. I need to work at Anthropic now." And that was Opus 4 — great model, but you were reading permission prompts. It's kind of crazy how much amnesia we have, where I'm like, oh, auto mode has always been here, right? I don't even remember pressing yes and allow. For me, the big thing I'm trying to push myself on is that we have to do higher quality work than we've ever done before. The outputs are incredibly high quality. I've been using it to edit videos a bunch, and I'm like, okay, it has to meet the very exacting demands of our brand team in a couple of hours or we just can't do it. That's how I'm trying to shift with Fable: the best work we've ever done, faster than we've ever done it before.

What piece of conventional software engineering no longer holds?

3:39

Simon: What's a piece of conventional software engineering that was true a year ago that you don't think holds anymore in this new world?

Cat: One of the biggest shifts we're seeing in the eng skill set: two years ago it was pretty typical for a product manager to go talk to a bunch of customers, align over the course of six months with cross-functional teams on some PRD, and write a thorough spec on exactly how we'll implement this before the first line of code gets written. Now things are completely turned the opposite way. For a lot of engineers, the push I would give to folks in the room is to develop more of your business sense and product sense on what it is we should build, because the timeline between having an idea and building it is so much shorter — it's down from six to twelve months to maybe even a week. That means all of us need to have better taste on what is worth building, what will actually inflect the businesses we're working on. So it's an increase in value on product taste and business sense, and a bit lower on execution in most product domains. Of course, for infra there's still a very heavy emphasis on making sure all the details are right.

Thariq: For me, it's that rewrites are now good.

Simon: The worst thing you could do is now actually fine!

Thariq: Exactly. All the Mythical Man-Month stuff — never rewrite — I'm pro-rewriting now. If you have a good test suite — and I think the rewrite actually forces you to make sure you have a good test suite — but I think what people undercount is that a codebase is a spec, and maybe it's the only copy of the spec that you have, because no one knows every branching part of the codebase. You can take this as an artifact and distill it or create other versions of it. We rewrote Bun in Rust and it works great — it's live for me right now.

Simon: You're not shipping Claude Code on Bun-in-Rust yet, right?

Thariq: Internally we have.

(Actually it looks like Anthropic started shipping Claude Code on Bun-in-Rust to everyone on June 17th.)

What kind of things are non-engineers doing with Claude Tag?

6:36

Simon: The other big launch recently was Claude Tag — that's what, a week old now, at least for the rest of us. I understand it's being used at Anthropic by non-engineers a great deal. What kind of things are non-engineers doing with Claude Tag?

Cat: Claude Tag is a Claude that lives in your team's collaboration tools. We launched it last week within Slack. The thing that's different about Claude Tag is it's multiplayer by default. Once you add Claude Tag to a Slack channel, you can chime in, your teammates can chime in, and you can collaborate together on the PR. The other big difference is that it's proactive instead of reactive. You can tell Claude Tag, "Hey, monitor every bug report in this channel, put up a PR to fix it, and tag the engineer who last touched this part of the codebase," and it'll do it for the lifetime of the channel without you having to manually tag it in. And the third big shift is that we've added team memory into this. If you tell Claude Tag your preferences in the channel, it'll remember them for every future post. If you always want it to debug outages but you don't want it to debug warnings, just tell it that in natural language in the channel and it'll remember it for you and everyone else on your team.

Internally, we see Claude Tag as the evolution of Claude Code. We see this as a large shift in how we work internally. Claude Tag currently lands 65% of our product eng PRs.

Simon: For all of Anthropic, or just for Claude Code?

Cat: This is just for our product engineering team — our internal version of Claude Tag lands 65% of our product PRs right now. And this is a huge shift; this is more than 50% of our PRs. The way we see people split work between Claude Code and Claude Tag is: Claude Code is still the best place for your most complex tasks, when you're interactively iterating with the agent. But Claude Tag is great for having it work proactively on your behalf, so you no longer need to manually kick off Claude Code for all the bug reports that come up for features you're working on.

Thariq: And for non-coding cases: for example, before this talk we asked Claude Tag, "Hey, when is Fable releasing?" We wanted to make sure we'd line it up with the announcement. Claude Tag would search our Slack and look at who's been saying what. As a search engine for your company, it's really valuable. It has all the context for your product, so you can ask it metrics-related questions — often when you're making decisions you want them informed by what the metrics say, so you hook it up to your event store. I've seen our marketing team do things like, "Hey, tell me about this feature." They're not programmers, but Claude is a programmer — it can clone the codebase and say, "This is the feature, this is what it looks like, this is a recording of me using the feature." It enables a whole wide variety of things, and I think we're still early in figuring that out.

Claude Tag as the team collaborative layer

10:06

Simon: One of the problems I've had with coding agents is that I get how to use them as an individual, but I'm not really clear on how to use them in a team environment. It sounds like Claude Tag is your current answer to that team collaborative layer for this stuff.

Cat: Exactly. And a large percentage of our sessions are actually multiplayer right now. Maybe I say, "Hey, I think we should implement this new feature in Cowork," and I'll tag in Claude Tag to do a first pass at it. Then I'll tell Claude Tag, "Share a recording of your final implementation," and I'll tag in design to take a look. They'll nudge it, then pass it on to eng to take it to the finish line and get it out to prod. It's been this very fluid experience. We're still trying to iron out what the social dynamics are for steering the same session, but we've found that people just observe how others use it and follow those social norms — it's been pretty intuitive for us to integrate Claude Tag into our teams.

Thariq: It's great for teaching people, and also for reducing slop, because the fact that everyone is seeing you use Claude together sort of levels up how you use Claude as well.

This reminded me of how Midjourney solved the challenge of teaching people advanced image prompting by enforcing prompting in public in their Discord channels.

How do you decide which features are worth building when building is so much cheaper?

11:41

Something I've found really hard myself is knowing when a feature is worth shipping now that the cost of actually building features has dropped so much.

Simon: How do you deal with the hardest problem in all of engineering — prioritization? How do you decide which features are worth building and shipping when building a feature is so much more inexpensive now?

Cat: This is the hard thing. There are a few ways we approach it. One is we dogfood our products every single day. Whenever there's something we want to be able to do in our products that we're not able to, instead of finding a different solution we fix our product so it can support that case. We have a very heavy dogfooding culture internally. Before we share our products with everyone in the world, we share them with everyone within Anthropic, and with some early customers who give us very honest feedback about it — the more brutal the better — and we iterate until people love it. We have an internal bar for the number of active users and the amount of retention a feature has to have before we share it with the world. Because this bar is very clear, every engineer knows what they're trying to hit. I think this also levels up our polish, because if the feature isn't polished, people will churn — and then we shouldn't ship that feature.

Using internal user-retention to decide if a feature should ship makes a whole lot of sense to me.

Do you have an example of a feature which surprised you?

12:54

Simon: Do you have an example of a feature which surprised you? You rolled it out and the engagement was off the charts — something unlikely to be shipped that turned into a real product thing.

Cat: I do have one. A lot of folks on our team love remote control. Remote control lets you use your mobile device, or Claude in the web browser, to connect to a local Claude Code session running in your CLI. I never have this need, because I just kick off the task directly on mobile and it runs in a cloud session without using my local environment — I think because I'm doing very easy coding tasks. It was something I didn't totally understand; I was like, hey, people should just set up remote dev environments. But in practice, once we rolled out remote control, so many people I talk to told me that what they do every night is plug their laptop into a power charger, open a bunch of remote control sessions, lock the screen, and then use their mobile phone from their couch to control Claude Code. So this has become a flow we're now leaning into that I didn't originally get — but now I do.

Does a human review every line of production code in Claude Code?

14:20

One of the over-arching themes of the conference was review: how much attention to people spend to reviewing code written for them by coding agents. I was very keen to hear the Claude Code team's take on this!

Simon: How does code review work? Does a human being review every line of production code that makes it into Claude Code? And if not, what are you doing — how do you keep the quality up?

Thariq: It varies on the task a lot. For important areas we have code owners. The system prompt is an example where we have a code owner — you really need to get their approval.

Simon: So the code owner is directly responsible for the quality of that area of the code.

Thariq: That's right.

Cat: And they need to approve any PR that touches it.

Thariq: We have our code review GitHub bot review everything — that goes on every PR, and often it's doing the bulk of the review. Something I've seen on the team is that for more complex PRs you might make an artifact to explain the PR so that other people can then review. And we invest a lot into verification, CI/CD, things like that, to make sure that any time anything fails we have a test. We have a really robust environment where Claude can control Claude Code and test it. So there's a multi-pronged approach to code review.

Cat: In general, we are trying to move to a world where humans don't need to be in the loop. For the most critical changes to the core of Claude Code, and the cores of other products, there is always a code owner and they do manually review all the changes. But increasingly, for the changes at the outer layers, we actually have Claude code review fully review those. That sounds pretty scary, but we've had a six-plus-month-long process to get here, and there are baby steps that you take to build up trust with code review. In the beginning we had human review for everything, and then increasingly we would say, okay, for code changes that touch these files, code review is catching 100% of the issues there — so we actually don't need a human manually reviewing those. And when we have incident review, we look at the PRs that caused the incident and say, okay, how do we update code review to catch that? — and we take those PRs and add them to an eval set to make sure our future changes to code review never regress that metric. Removing humans from the code review loop is a big step forward. It can sound scary, and it's not something you can do overnight, but it is something you can do through many months of investment in the infrastructure to give you the confidence that code review is catching everything you care about.

So the key seems to be constantly iterating on the automated review systems themselves, in order to build trust in them over time.

How does a new model affect your intuition for what it can and can't do?

17:20

We got deep into evals - another hot topic throughout the wider conference.

Simon: I know that Opus 4.8, if I ask it to build me a JSON endpoint that runs a SQL query and outputs JSON, is just going to get it right — that's not something I have to review closely. But then a new model comes along and I don't know how to build trust in Fable quickly, that it's not going to mess things up that Opus didn't. How does the new model affect your intuition for what it can do and what it can't do?

Cat: The main reason we're building up this eval base over time is so that new models can be a drop-in replacement. When we have a new model, we run the whole eval set and make sure that, for example, Fable is strictly better than Opus 4.8 — and that gives us the confidence to drop it in.

Simon: Are those model evals for Anthropic as a whole, or Claude Code team-specific?

Cat: We have both. We have evals on our team, and we run code review across every repo within Anthropic, so we have evals for that. And for things like auto mode, we not only have evals across every user within Anthropic — we've also commissioned multiple external testers to red team it, to create environments with prompt injections and malicious inputs, and make sure that auto mode doesn't let any of those pass.

How do you build confidence that a system prompt tweak results in better output?

18:41

Simon: I want to know if the system prompt improvement I made actually improved the product — that's the most basic form of product-specific eval, and I still don't have a great feel for how to do that. Is that something you're doing such that you have complete confidence that a tweak you've made to the system prompt results in better output?

Cat: We don't have complete confidence, but we do a lot to make sure that we don't regress performance. The starting point is a suite of external evals that we trust, and we complement that with an even larger suite of internal evals that we trust. To start, we mainly optimize for capability: given a complete definition of a task and the full codebase, does Claude make the right decisions, fully fix the bugs, and pass all the tests? That's the starting point and the thing we optimize for, because it's most directly what users want. But there are a lot of behaviors that impact how users feel when they work with Claude Code. For example, people really don't like it when Claude Code says it's time to go to sleep. Or people really don't like it when it says, "Hey, I finished two out of five parts — do you want me to continue?" Yes, please continue. So we're building up a set of behavioral evals to catch these. And as we get user feedback — please be loud with us about your user feedback — we rank the priority issues and go down one by one and build evals for each of them. It's not 100% coverage, but it is a priority for us to increase the coverage.

How much interaction is there between the Claude Code team and the model training teams?

20:21

Simon: How much interaction is there between the Claude Code team and the teams at Anthropic who are training the models in the first place? Is that quite a close collaboration?

Cat: Across Anthropic, we all work quite closely together. We meet often to talk about what we expect the next generation of models to be able to do. Our research team has also been amazing about showing this publicly — we often talk in our blog posts about how we're targeting ever-increasing longer-horizon work, and how we train Claude itself to be honest, harmless, and helpful. We also put a lot of effort into making sure it's aligned with your intent, even if your intent is expressed in a fuzzy way. Of course, try your best to be specific about what you want, so Claude has all the context — but even when you're not specific, we teach Claude to make good assumptions. It's been a productive partnership.

The system prompt has been reduced by 80% — what have you been able to drop?

21:24

So many useful prompting tips in this section!

Simon: Thariq, you mentioned this morning that the system prompt for Claude Code has been reduced by 80% because of Claude Fable. Can you go into a little more detail? What kind of things have you been able to drop?

Thariq: It wasn't just Fable — it was Opus 4.8 as well, and going forward, future models. We have different system prompts for different models now. One of the patterns we saw is that we were over-constraining Claude. The initial, maybe Opus 4-ish models wanted a lot of examples, and removing examples was extremely helpful, because it was just more creative than the examples we gave it.

Simon: That's really interesting, because one of the top prompting tips I give people is: give it examples. If that's no longer true, that kind of breaks my prompting model a little bit.

Thariq: Same here — I was surprised to hear that. I think now it's more about the shape of what you give it — the tools you give to Claude, your system prompt, things like that. The other thing we did is try to give it more context and fewer "do not do this" instructions, because that's a very strong impulse for Claude, and especially if it conflicts with user instructions later on, that can be extremely confusing to Claude — "I've got this skill that says this and the system prompt says this." So we try to have fewer hard constraints, more context, and fewer instructions overall. It's definitely a science — it took a bunch of evals to build.

Cat: In general, when you're prompting these models, you should always think: are there edge cases to the instruction that I'm giving it? When we went back and reviewed all the instructions in the Claude Code system prompt, we found a few cases where yes, this statement is 90% true, but there's a real 10% of cases where it's not true. We didn't want to constrain the model, or confuse it into thinking it should always do this. One good example is verification. Everyone here wants Claude to verify its work, and we had some instructions in the prompt that said: if you make a front-end change, always verify. But there's a limit to it. If it's changing copy from one string to another string, and the user says "just make a quick fix and update the test," maybe you don't want to verify. So we've adjusted our wording from "always verify, verify, verify" to something like: most of the time when you're doing front-end work you can't fully understand the experience by hitting the backend endpoints, so when you make larger changes to the user experience, please run the app locally. And in fact, that instruction probably isn't even good either, because what is a large change? Maybe it should test small changes too. In general, whenever you give a prompt to the model, you should think about the ways in which it could be misinterpreted by a well-intentioned human, in order to better understand how the model might interpret it — and soften the prompt so that it's actually 100% accurate, because you're giving this prompt to the model 100% of the time.

Simon: What's fascinating about that is you're relying on the model's judgment — and that's got to be an Opus/Fable-level thing. Models a year ago did not have the level of judgment necessary to decide whether they were going to test a change or not. But that does break down if you're building for a wide range of models and trying to run the cheaper models for cheaper tasks.

Cat: We actually have a different system prompt per model now, for this very reason. It's only our most frontier models that have this 80% token decrease — the older models still have the full system prompt.

Simon: Do you think Fable and Opus are smart enough to prompt Haiku with more details, because they understand that Haiku has less judgment, less taste?

Cat: We haven't been able to eval it — we don't have any hard data to show it.

Thariq: There's a tough thing with smaller models sometimes, because sometimes the larger models can be more token-efficient on a hard problem than the smaller models. So there's a bit of intuition to build there — sometimes you really just want frontier intelligence almost all the time. The Pareto curve shifts, and it's hard to find.

Simon: A year ago I did not trust a model to write a prompt. Today the good models are very good at prompting — a lot of my prompts are written by models, which feels absurd but works really well. What helped me come to terms with that was thinking about subagents, which are entirely about a Claude model setting up a prompt for another Claude model.

Thariq: Workflows are actually a really good example of this, because it's Claude not just prompting a single subagent, but prompting the orchestration of many subagents, and each one of them gets a very detailed prompt. It's almost a level above just spawning a subagent. I've also been using it on my personal machine, giving it the Gemini API and saying: here, generate images. It's way less lazy than I am at prompting an image model. It's just Claude prompting Claude all the way down.

Cat: I think Claude also wrote the prompt for the workflow tool.

Simon: I've read that prompt — it's a good prompt. That's actually a frustration I have with Anthropic generally: you publish the prompts for Claude Chat, but you don't include the tool prompts and the Claude Code prompts. I still have to run a proxy to intercept them. I would love it if the Claude Code prompts were deliberately published — they're the documentation. They're how you know what the tool can do and how it works.

Cat: I'll write down that feature request. I'll have Claude Tag do it.

Interesting to note that OpenAI's prompting best practices for GPT-5.6 includes similar advice for their latest models:

Favor leaner prompts

Removing repeated instructions and examples and simplifying tool descriptions can improve task performance and token efficiency. In a sample of internal coding-agent eval runs, configurations with leaner system prompts improved evaluation scores by roughly 10–15% while reducing total tokens by 41–66% and cost by 33–67%.

What's your bar for introducing a new tool?

28:06

Simon: Claude Code is basically a big bag of tools. What's your bar for introducing a new tool? How do you decide when it's worth doing that additional engineering at that level?

Cat: Do you want to take it? You introduced one of the best tools we have.

Thariq: My career peaked when I introduced the ask user question tool. It's really hard. Especially for some tools — ask user question is Claude's tool to ask you — so it's hard to eval, and sometimes it's more of a user preference thing. Back then we had fewer evals, so it was very dogfooding based — or "ant fooding," our ant version of that. But overall we've been trying to trend towards fewer tools. The last set of tools we introduced was the task tool, I think — and we try to give Claude more general versions to do things.

What's the latest evolution of your file editing tool?

29:03

I have a long-running fascination with file editing tools - they were the subject of the old Aider code editing leaderboard, and I've watched with interest as they've evolved in different coding agents from search-and-replace based to line-number-based to more complicated patterns.

The Claude API docs describe a text editing tool that's recommended for building against the API, but Claude Code seems to use slightly different approaches here.

Simon: One of the most interesting tools is the file editing tool — you can have file editing as a tool, or you can tell it to use sed and grep and do things that way. What's the latest evolution of your file editing tool?

Thariq: We still have one, but for example we removed our grep and other search tools — glob tools — in favor of native bash. Like I said in my talk earlier, the models are kind of more of a biology than a physics, and tool design especially is quite hard. I'm not sure if Cat disagrees and thinks there's a science to the eval of it, but I think tool design is more of an art, maybe — or a biology.

Cat: I largely agree, but in general as we introduce more tools, we try to keep the cardinality pretty low and make sure that every tool we add has a distinct function from every other tool, so that Claude can very easily distinguish when to call each. For file edit, the reason we have it is actually because we can render it. We show people when Claude makes a file change, and there's this nice dedicated UI that says: do you approve this edit to this file? The reason we had a dedicated file edit tool was so that we could deterministically know that Claude was making a file change, so we could show people this nice UI. A lot of new users onboarding still really like this experience, so we've kept it around. But for a lot of us who are on auto mode right now — hopefully you're not on YOLO mode — I don't think it actually matters, and we could probably just remove file edit and be totally fine.

What's the advice within Anthropic for safely running Claude Code?

30:58

It's the prompt injection question! Who better than Anthropic employees to explain how Anthropic sees the risk of prompt injection attacks causing their Claude Code instances to run amok?

It turns out they really trust their auto mode - and see that as the feature that enabled Claude Tag.

Simon: Let's talk about safety and security. I am deeply aware of the risks of prompt injection, and there are so many bad things that can happen if somebody else tells my Claude Code what to do. I still mostly run Claude Code in YOLO mode and feel incredibly guilty about it. What's the advice within Anthropic for safely running Claude Code?

Cat: Why not auto mode?

Simon: I am starting to use auto mode, but I don't understand it enough to get how safe it is. As of maybe three weeks ago, I'm defaulting to auto mode.

Cat: Broadly within Anthropic, almost every single person uses auto mode. It is the best way to do long-running work in Claude Code while being safe. We've done extensive bashing. We have thousands of evals. We've commissioned many red teamers to create adversarial environments in order to trick Claude Code into doing bad actions, and we've mitigated every single issue that they found. We're going to publish some evals in the coming weeks, but we've pretty much mitigated every attack.

Simon: That is a big claim.

Cat: We'll share the evals for it so folks can assess, but we've been extremely diligent about identifying all the ways in which Claude might mess up and then updating auto mode to counter it. It doesn't catch 100% of things — that would be way too strong a claim. But for the main categories of risks that we're concerned about, like prompt injection and data exfiltration, the risks are far lower than the average human reviewer.

I am very much looking forward to learning more about their evals and approach to verifying auto mode.

Thariq: A little on how auto mode works — it's useful to build this mental model. Whenever Claude is doing a turn, or a bash call, there's a Sonnet classifier that is judging the tool call and also the context of the conversation — your instruction. There are some things around permissions that are dependent on your request: you don't want to give git push permissions all the time, but if you say "push this to GitHub," you want it to do it — and if you say "don't push," you want it to deny it. Auto mode will do that. That particular thing happens to me a lot, where Claude tried to do something because it's very helpful and proactive, and auto mode saw "don't do this" and surfaced it. So it's good at the dynamic permissions that you yourself give inside the prompt, which I think is really important. It also works well with our sandboxing infrastructure, because sandboxing is one of those things where there are so many different edge cases that it's hard for us to deterministically follow them. We have a sandbox, and when something needs to escape the sandbox — like a network request — auto mode can look at that request and ask: does this make sense? — and allow it.

Simon: I hadn't realized auto mode is interacting with the networking sandbox as well.

Cat: It interacts with any permission prompt the user would otherwise see.

Simon: How old is auto mode? As a feature I had access to, it's only a couple of months old, right?

(It was first made available to the public on March 24th.)

Cat: We've been using it within Anthropic since January, so we've been hardening it for quite a while. Anthropic is extremely focused on safety and security, and we've been working broadly across our alignment and safeguards teams to enable the rollout internally, build out these evals, and make auto mode even more robust before sharing it with the world.

Thariq: This is also the reason Claude Tag is so good — Claude Tag uses auto mode. I've heard a lot of build-versus-buy questions about a Slackbot, and I'm like: please, you probably shouldn't build your own AI Slackbot. There are so many attack vectors. You have a feedback channel that users can post feedback into, and now your bot is reading it. The work we've put in with auto mode — and we have a general Swiss cheese defense for security; we also RL against this stuff — I think this is really what makes Claude Tag work. It works seamlessly with your permissions, and you don't want to be prompt injected in your Slack.

Are there more security things in the pipeline beyond auto mode?

35:54

Simon: Are there any more security things in the pipeline that go beyond auto mode?

Thariq: I think we're very secure. With Claude Tag you can provision your own credentials for Claude, so it doesn't need to act on your behalf — you can have Claude as an identity, and that also makes it easier to audit and inspect what Claude is doing.

Simon: Because Claude Tag is influenced by anyone who can talk to it — it's got a much wider pool of people telling it what to do.

Thariq: That's right. And of course we have probes as well with Fable, which is a downstream effect of our safety and research work. I think this is the moment where you see Anthropic being an AI safety company really paying off: we really want Claude to be able to run in an aligned way over long periods of time, and auto mode has to be basically flawless for this to work — it's all downstream of our being an AI safety company.

Cat: We also launched trusted devices for the remote control users out there who want to be safer. And for all of our remote environments, we support credential injection. If you want Claude Code to be able to access Datadog, but you don't want Claude Code itself to hold the Datadog credential, you can set up our identity and credential management system so that the Datadog credentials are only usable by the agent but not accessible by the agent — we insert them on the fly when the agent tries to make a Datadog request.

I really like that credential injection pattern, where Claude Code can access an API via a proxy and that proxy both audits the request and injects the relevant API key - so Claude can access authenticated endpoints without having access to the API credentials itself.

How has the past year and a half changed how you think about your own craft?

37:53

Thariq talked about a sense of grief brought on by Fable-class models in his keynote in the morning, and we dived further into that as part of our conversation. I've been calling this Deep Blue.

Simon: Let's talk a little bit about the human element. A lot of people are feeling a sense of loss now that so much of what they considered to be their role in building software is being subsumed by the models. How do you think about that? How has the past year and a half changed the way you think about your own craft and the value that you add?

Thariq: Cat and Boris are such good reminders that you have to be more ambitious. They're always like: we're growing so fast, we have to be on the edge, we have to do the best work we can. That's a constant reminder for me — any time I'm slow on something, I'm like, okay, can I do it faster? Can I be more ambitious here? And oftentimes the answer is Claude, because Claude is getting better as you go — the last time I tried this, it was with the previous model. On your point about loss: I think this is real. If you're only trying to do the same work you were doing before LLMs, and now it's a prompt, it is, I think, kind of a sad feeling. And the way you offset that is by being more ambitious. I think Jared is such a good example — he hand-wrote all of the Zig code in his Oakland apartment in about a year, barely left his house, and had so much fun doing that. Now I see him rewrite all of Bun into Rust and he's having so much fun doing that — it's so much more ambitious, and that's how he offsets it. Generally it's asking how do I do the bigger thing and do more — I think success is fun. It's changing your ambition.

"The way you offset that is by being more ambitious" neatly captures where I've landed on this issue myself as well.

Simon: And Cat, what does that look like from a product management perspective?

Cat: I feel like the product role just changes every single month. All the PMs on our team are this mix of engineer, designer, PM — most of them actually used to be full-time engineers. For us it really means plugging in whenever there's any kind of gap. If we have an idea and we didn't inspire any engineer to go build it, then we should just build it, put it into a notebook, and inspire people to take it to production. If the designs look a little off, let's take a page that's similar, do a first-pass design, and tag in someone who's very detail-oriented to fill in the gaps. Or if we notice that our team and product adoption is bigger within the company, and more people need to know what's coming down the pipe for Claude Code, Claude Tag, and Cowork — let's automate figuring out our whole launch calendar, let's automate getting those status updates asynchronously so we're not bugging people, and make sure our updates in our internal announce channels are fully detailed and to the point. For us it's very much understanding what the gap is right now between a great idea and getting something to our customers, and how do we automate it as much as possible.

This reflects something I've noticed: when you can produce code so much faster, time spent blocked awaiting a decision from someone else becomes a much more notable bottleneck. Engineers who can make product decisions can move a whole lot faster, and the cost of getting one of those decisions wrong is much less prohibitive.

What's a moment when Claude has surprised you?

41:50

Simon: What's a moment when Claude has surprised you? When the model did something you didn't think it would be able to do?

Thariq: I've posted a lot about Claude video editing, but most recently I gave a talk at the ACM Agentic conference, and I asked, "Hey guys, do you have the edited video? I'd love to post it and share it with my comms team." They said, "Oh, it's taking so long." So I asked for the raw files. They sent me the video of me talking on stage, the video of the deck, and the audio file, and said, "Good luck." I gave this to Claude, along with my HTML deck, and said, "Hey, can you just edit this together?" And what it does is honestly incredible — I'm ready to ship it. It transcribes the entire video. It notices that sometimes the video of my deck is a little weird — there's a popup of an auto-update in the middle — and it goes, "Oh, I probably shouldn't use the video of your deck. What I'm going to do is slice it up, figure out which slide you're on, and use the HTML source instead." So it displays the HTML source. Then it's got video of me, but I'm only taking up a small part of the stage, so it's cropping dynamically to where I am on the stage — and I'm pacing, so it's tracking me as I pace. And it's transcribing what I'm saying.

Simon: This was Fable, right?

Thariq: This was Fable, yeah. It was a good prompt, but it was a one-shot prompt. Then I asked it to add some interesting animations and graphics, and I was just blown away. It does ffmpeg, it does Remotion.

Here's Thariq's video on how he used Fable to edit Fable's own launch video, and here's that launch video.

What can't it do yet?

43:36

I'm embarrased to admit that I've been finding it quite hard to come up with tasks that frontier models like Fable 5 and GPT-5.6 are unable to accomplish.

Cat still doesn't rate its UX design skills:

Simon: What can't it do? What are the things where you're still disappointed — where you're waiting for Claude Fable 6 to figure it out for you?

Cat: I want it to have better design and UX taste. It's now at the point where if I write out a prompt with a detailed spec of how I want a feature to behave, it will usually behave that way. But the paddings might be off, or the interface just isn't delightful yet. It leans on existing best practices for how apps are designed, but for frontier AI products, there are so many new interaction experiences that we have yet to design.

Simon: There's an Opus aesthetic — you can look at something and go, "Yeah, that was designed by Opus." It'd be good if we could move beyond that.

Cat: Yeah. I'm very excited for future models to hopefully be interaction design thought partners.

Thariq: What can't it do? I would love to see it interact more with the real world. Can it solve science? Can it orchestrate the experiments? There's some amount of coding that goes into that, but there's also this other taste of the broader world that it needs.

Which parts of Anthropic's culture should other companies steal?

45:11

I figured this would make a great closing question:

Simon: Which parts of Anthropic's company culture do you think uniquely help Anthropic be productive with these tools, that other companies should steal? What are the cultural hacks people should be adopting from you?

Cat: I'll share one for Claude Tag. Claude Tag works best when you have it in a public channel, and when most of your channels are public. Claude Tag is able to search across all public channels to get as much context as possible to give you the highest-accuracy answer — and it's only able to do this if it has access to everything.

Thariq: I mentioned this in my keynote, but it's so important to me I want to re-emphasize it. The co-founders say we don't negotiate against ourselves, and I think this is really important. You can imagine trade-offs in your head and talk yourself out of doing something ambitious — or you can just try to do the ambitious thing. We're so often asking: what if we just did it? Is this a real trade-off or not? And if so, why — where's the proof that it's a real trade-off, and not just something that sounds reasonable? Make the trade-offs show themselves to you. Be as ambitious as you can.

What's your favorite absurd thing you've built with Claude, just because you could?

46:46

I couldn't resist throwing in this one as well.

Simon: What's one of your favorite absurd things that you've built with Claude, just because you could build it?

Thariq: I'm working on a 2D Street Fighter fighting game with me as a character — and my friends as well. It uses Claude Code to prompt Gemini — and honestly the Seedance model is pretty good — to make video animations. It works great; it's so good at prompting, and it can verify the frames to check whether an animation was good.

Simon: Is this Street Fighter 2-level 2D sprites you're generating?

Thariq: Yeah, exactly — 2D sprites. The animation looks amazing. And it can also figure out hitboxes — it can be like, "Oh, your fist is here, I'll draw the JSON hitbox." It's incredible.

Cat: Mine is much more simple. I'm a big rock climber and a lot of my friends climb, so we have this little app we built with Claude Code where we log all the projects we're working on. We also go outdoors together a lot, so we have Claude do all this research with workflows. Workflows is amazing — we brand it as a coding tool, but it's amazing for doing deep research for travel. I also plan our team offsites, and it's good at finding venues that can fit all of us. I use workflows to research all the climbing destinations we might want to go to, and what has direct flights from where all of us are located. It goes to Mountain Project and finds all the climbs at our grade level. It finds the Airbnb. And I don't like hiking, so I care a lot about it having a very short approach — very short walking distance from where the car parks to where the rock actually is — and it filters for this. With existing apps I have to manually click through Mountain Project, but with this I just put in all of our preferences and it's a custom app for us.

Simon: So you're basically vibe coding Jira for mountain climbing.

Cat: Exactly.

Audience: Any plans for eval-building tools and agent observability?

49:23

We had a few minutes at the end for questions from the audience.

Audience: Do you have any near-term plans to build more eval tools for us to build eval datasets, and more observability tools to monitor the performance of agents and workflows?

Cat: We've considered building eval tools, but I think the limiting factor actually tends to be that it takes a long time for customers to build really high-quality evals. So I think the tooling is less of the constraint, and more the skill set of how you build a great eval. That's an area where we're excited to both invest internally and hopefully share some best practices externally.

Audience: How is memory designed today — and would you move from files to a data store?

50:08

Audience (Sai): I'm interested in the memory and the multiplayer. How is memory being designed today? I assume it's around files. And second, have you thought about an orthogonal direction where you would actually need a data store for these memories, instead of files, to scale it better?

Thariq: Right now for Claude Tag the memory is channel-specific. Every Claude in that channel has a shared memory, and the instances have a session — but the session can contribute back to main memory. We do a lot of memory research, and it can be kind of unintuitive what the right way to do memory is. We're always running memory experiments. How it works right now in Claude Tag is a markdown file per channel.

Tags: ai, prompt-engineering, generative-ai, llms, anthropic, annotated-talks, coding-agents, claude-code, thariq-shihipar, cat-wu

NISAR’s L-Band Radar Reveals ‘Hummingbird’ in Antarctica

2 Min Read

NISAR’s L-Band Radar Reveals ‘Hummingbird’ in Antarctica

Scientists used data from the L-band radar aboard the U.S.-India Earth-orbiting NISAR satellite to produce this image of Nunatak Zaterjavshijsja, a mountaintop in East Antarctica, poking out amid a stream of ice flowing northeast to the ocean.
PIA26617
Credits: NASA/JPL-Caltech

Description

Data from the Earth-orbiting U.S.-India NISAR (NASA-ISRO Synthetic Aperture Radar) satellite’s L-band radar was used to produce an image of Nunatak Zaterjavshijsja — a mountaintop in East Antarctica — poking out amid a stream of ice flowing northeast to the ocean. The obstruction causes stresses in the ice, heavily fracturing the surrounding surfaces with deep cracks, called crevasses, which show as sharp green lines in the image. Produced in August 2025, the image has been nicknamed “the hummingbird” by NISAR scientists. 

The colors show differences in the way polarized microwave signals, which vibrate in different directions, interact with and reflect from the ice. Over Antarctica, NISAR transmits radar waves toward Earth with a horizontal polarization. The orientation of the signals that return, either horizontal, vertical, or both, provide clues about the object or surface that reflected them.

Signals that come back with a horizontal polarization likely bounced off a more regular surface, such as smooth ice. Those signals appear magenta in the image. Signals that return with vertical polarization may have refracted as they partially penetrated the ice or scattered at different angles as they reflected off irregular surfaces, like the faces of crevasses. Called volume scattering, these observations are displayed in green.

The white represents areas in which both magenta and green signals scatter back strongly, a possible indication that there is an equal blend of surface and volume scattering.

The image shows Nunatak Zaterjavshijsja at center-left, surrounded by ice fractured with crevasses, which are shown as sharp, green lines. The magenta portions of the image represent more regular surfaces, such as smooth ice.
Figure A

Figure A is an annotated version of image.

Managed by Caltech, NASA’s Jet Propulsion Laboratory leads the United States component of the project and provided the satellite’s L-band SAR and antenna reflector. The spacecraft bus and its S-band SAR were provided by the Indian Space Research Organisation. The NISAR satellite is the first to carry two SAR instruments at different wavelengths, collecting data using the spacecraft’s giant drum-shaped reflector, which measures 39 feet (12 meters) wide — the largest radar antenna reflector NASA has ever sent into space.

To learn more about NISAR, visit:

https://science.nasa.gov/mission/nisar/

The post NISAR’s L-Band Radar Reveals ‘Hummingbird’ in Antarctica appeared first on NASA Science.

What Does the Law Say About Riding in a Pickup Truck Bed?

In the US, riding in the bed of a pickup truck is not governed by one nationwide law. Instead, each state creates its own rules, meaning some states allow truck bed passengers in certain situations, while others restrict or prohibit the practice on public roads.

The legality often depends on factors such as the state, passenger age, road type, and specific circumstances. This article explains what the law says about riding in a pickup truck bed, common exceptions, possible penalties, and important safety considerations drivers should know.

Are Pickup Truck Bed Passengers Allowed Under Federal Law?

There is no federal law that directly regulates whether it is legal to ride in the bed of a truck . Instead, each state establishes its own traffic laws, so the rules vary by state.

Some states prohibit people from riding in the bed of a truck on public roads. Others allow it only under specific circumstances, such as at low speeds, on private property, or during certain authorized events.

Common Exceptions and Permitted Situations

Many states recognize limited exceptions for specific situations. These may include agricultural activities, emergency responses, parades, or work-related transportation in which riding in the bed serves a practical purpose.

Common exceptions can include:

  • Travel on private roads or property.
  • Use during permitted community events.
  • Certain farming or business activities.
  • Situations involving emergencies or official duties.

Even when an exception exists, drivers may still need to follow additional safety rules. A legal exception does not automatically make every passenger situation safe or lawful.

Age and Restraint Rules

Age restrictions are one of the most common limits placed on pickup truck bed passengers. Many states apply stricter rules to children because younger passengers face greater risks during sudden stops, crashes, or vehicle rollovers.

Some laws set minimum ages, while others require children to remain in passenger seats with proper restraints. Parents and drivers should check the specific requirements in the state where the vehicle is operated.

Penalties Drivers Can Face

Violating pickup truck bed laws may lead to traffic citations , fines, or other penalties. The consequences vary based on state statutes, the passenger’s age, and whether the violation contributes to a dangerous situation.

Drivers may also face increased liability if a passenger is injured. Courts may consider whether the driver ignored safety requirements or created an unreasonable risk.

What Are the Safety Risks of Riding in a Pickup Truck Bed?

Legal permission does not remove the dangers of riding in an open truck bed. The National Highway Traffic Safety Administration (NHTSA) has warned that unrestrained occupants in vehicle areas not designed for passengers face a serious risk of injury during collisions.

Truck beds lack seat belts, protective structures, and airbags. Passengers can be thrown from the vehicle or struck by objects, especially during sudden movements or crashes. State agencies publish guidance that helps clarify passenger restrictions and safety obligations.

What Should Drivers Consider Before Allowing Passengers

Drivers should review applicable state rules and consider safety factors before carrying anyone in a pickup bed.

Important considerations include:

  • Check the state law that applies to the trip.
  • Confirm whether age limits or exceptions affect passengers.
  • Avoid carrying people in unsafe traffic conditions.
  • Consider safer seating options inside the vehicle.

Ignoring pickup truck bed laws can create avoidable legal and safety problems. Understanding the rules before a trip helps drivers avoid violations and reduce risks for everyone involved. Drivers who ignore these requirements may also face serious consequences when passengers suffer injuries in preventable crashes.

Key Takeaways

  • Individual states set pickup truck bed passenger laws.
  • Federal law does not provide one nationwide rule.
  • Some states ban the practice on public roads .
  • Exceptions may apply for work, emergencies, or special events.
  • Children often face stricter restrictions than adults.
  • Fines or legal repercussions may follow violations.

Photo: Ric Andy via Pexels


CLICK HERE TO DONATE IN SUPPORT OF DCREPORT’S NONPROFIT MISSION

The post What Does the Law Say About Riding in a Pickup Truck Bed? appeared first on DCReport.org.

What Is a Coup-Contrecoup Brain Injury and What Accidents Commonly Cause It?

A coup-contrecoup brain injury is a type of traumatic brain injury (TBI) in which the brain is damaged both at the point of impact and on the opposite side of the skull. It typically occurs when a sudden blow or violent movement causes the brain to strike one side of the skull before rebounding and hitting the other, often resulting in more extensive injuries than a single-impact head trauma.

Learning about the common causes of TBI  can make it easier to understand why coup-contrecoup injuries are frequently linked to high-impact accidents. Because this type of injury usually involves significant force, it often becomes an important factor in both medical treatment and personal injury claims.

How Does a Coup-Contrecoup Injury Happen?

The terms “coup” and “contrecoup” describe where the injuries occur inside the skull.

  • A coup injury develops where the head first strikes an object.
  • A contrecoup injury occurs on the opposite side after the brain rebounds and hits the inside of the skull.

Rather than remaining stationary, the brain moves within the protective fluid surrounding it. During a sudden impact, that movement can cause bruising, bleeding, swelling, or other damage at two different locations. 

A coup-contrecoup injury can cause the following:

  • Brain bruising (contusions)
  • Internal bleeding
  • Swelling
  • Damage to surrounding brain tissue
  • More widespread neurological injury

Symptoms may include:

  • Headaches
  • Confusion
  • Memory loss
  • Dizziness
  • Nausea
  • Difficulty concentrating
  • Loss of consciousness
  • Long-term cognitive problems depending on the severity of the trauma.

What Type of Accidents Commonly Cause Coup-Contrecoup Injuries?

Several types of accidents generate enough force to cause this pattern of brain injury.

Motor Vehicle Collisions

Car accidents remain one of the leading causes of coup-contrecoup injuries, particularly.

  • High-speed crashes
  • Head-on collisions
  • Rollover accidents
  • Side-impact collisions

Rapid acceleration followed by an abrupt stop creates enough force for the brain to move violently inside the skull.

Falls

Falls are another common cause, especially among older adults and workers exposed to elevated surfaces.

Examples include:

  • Slip-and-fall accidents
  • Falls from ladders
  • Stairway falls
  • Falls from roofs or scaffolding

Even when the head strikes the ground only once, the brain’s internal movement can produce injuries on both sides.

Sports and Other High-Impact Incidents

Contact sports and other forceful impacts may also produce coup contrecoup injuries.

Examples include:

  • Football collisions
  • Hockey impacts
  • Cycling accidents
  • Shaken baby trauma

In each case, the sudden movement of the brain—not just the external blow—creates the injury pattern.

How These Injuries Can Lead to a Legal Claim

A coup-contrecoup injury may become the basis of a personal injury lawsuit if it was caused by someone else’s negligence, such as a reckless driver or unsafe property conditions. 

To succeed, the injured person generally needs to show that another party

  • Had a duty to act with reasonable care.
  • Failed to meet that duty.
  • Directly caused the injury.
  • Resulted in measurable damages like medical bills, lost income, or long-term disability.

Because these injuries require immediate evaluation, hospitals with emergency departments are generally required under 42 U.S.C. § 1395dd (EMTALA)  to provide appropriate medical screening and stabilization for patients with emergency conditions, including serious head injuries. 

Final Takeaways

  • A coup contrecoup injury damages the brain at both the impact site and the opposite side.
  • The injury results from the rapid movement of the brain violently inside the skull.
  • Motor vehicle  crashes remain one of the most common causes.
  • Falls and sports-related impacts can also produce this injury pattern.
  • Symptoms may involve bleeding, swelling, memory problems, and neurological damage.
  • Medical imaging and neurological evaluations are essential for diagnosis.
  • Detailed medical evidence often plays an important role in injury claims.

Photo: Tima Miroshnichenko via Pexels


CLICK HERE TO DONATE IN SUPPORT OF DCREPORT’S NONPROFIT MISSION

The post What Is a Coup-Contrecoup Brain Injury and What Accidents Commonly Cause It? appeared first on DCReport.org.

Some more things about Django I've been enjoying

Hello! I’m on a funny journey right now where I’m trying to learn how to make websites in a sort of 2010 style, where I have an SQL database and render some HTML on the backend.

It’s kind of an interesting journey because it doesn’t necessarily feel “easy” to me to make websites in this way: I never learned how to do it in the 2000s or 2010s, and there’s a lot I need to learn.

So here are some Django features that make building this kind of site feel more achievable than when I was trying and failing to use Go’s standard library or Flask. And I’ll talk about a couple of issues with Django I’ve run into.

why learn to make websites like it’s 2010?

Previously the toolkit I felt confident with for making websites was:

  • static site generators (like for this blog)
  • static sites that do some fun stuff with Javascript (like this sql playground)
  • simple Vue.js single page apps with either a Lambda as a backend or a Go backend (like mess with dns)

I really liked this frontend-heavy approach for these super simple applications but when I started thinking about making something with a lot of different pages (instead of literally just one page), I didn’t feel so excited about the options I saw that involved a lot of frontend code. So I figured I’d try the backend.

Writing a backend-focused site that uses as little JS as possible feels the same to me in a way as writing a single-page JS website that does as little on the backend as possible, even though they might seem like opposites. In both cases I’m just trying to keep as much of the logic as possible in one place.

Now for some thoughts about Django!

I’m enjoying query builders

I learned that I can define a “query set” class in Django with a bunch of methods with different WHERE statements I might want to use while constructing a query:

Here’s how I use it in my view code once I’ve defined what all the methods mean:

Events.objects.approved()
  .for_tab(tab)
  .with_festivals(tab_params.festival_slugs)
  .is_free(tab_params.free)
  .is_outdoors(tab_params.outdoors)

and here’s how I define the methods:

class EventQuerySet(SearchableQuerySetMixin, models.QuerySet):
    def approved(self):
        return self.filter(approved_at__isnull=False)

    def future(self):
        today = timezone.localdate()
        return self.filter(end__gt=self._midnight(today))

    def with_tags(self, tags):
        if tags:
            return self.filter(tags__name__in=tags).distinct()
        return self

The syntax for defining the filters isn’t my favourite, but I spend most of my time just using the methods, and it feels super readable and nice to use, and it makes me want to look into other query builder libraries in the future. In the past I thought “I know SQL, who needs a query builder?”, but this kind of structure does make it really nice to read.

I found an example of someone who wrote their own small query builder in Python that I want to read later to think about whether I would enjoy using a more minimal version of this.

the template filters are awesome

There are a bunch of little quality of life filters available in Django templates that are super useful for generating HTML. The ones I’ve used so far are:

  • translating plain text URLs into links, or line breaks into <br> ({{ event.description|urlize|linebreaksbr }} )
  • formatting dates ({{ row.date|date:"M j" }})
  • json_script, which takes a Python dictionary and automatically converts it to JSON and inserts it into the HTML as a <script> tag in a safe way

These are all small things individually but I feel like it makes a big difference somehow to just have them available.

querystring is cool

I think my favourite template filter is querystring: in this site sometimes we use filters like ?date=2026-06-01 to decide what’s displayed. querystring that will make a link to the same query string with one change, like this to link to the previous date:

<a href="{% querystring date=nav.prev_date%}">

Or to remove the outdoors parameter:

<a href="{% querystring outdoors=None %}">

automatic database migrations are still great

I still really love Django’s automatic database system. It’s amazing to be able to just edit a model to add a new field or whatever, and then Django automatically generates the migration.

So far we have done 19 database migrations and I think there will probably be more! It makes a huge difference for me to be able to just easily change the database as my understanding of the problem changes.

I do not want to organize my code with inheritance

Django’s documentation sometimes offers the option of using class-based views and inheritance to organize the code in your views. For example I have four views that share a lot of code, and I could use inheritance to manage that by defining some kind of parent class and then having my other views inherit from it.

I tried it out and I did not enjoy the experience of using inheritance to share code between views. I switched to using functions instead, sort of how this post advocates, and that was a lot more straightforward. I’ve never had a good experience using inheritance in Python and I don’t think I’ll try to use it again.

But I don’t mind using inheritance to use the interfaces Django itself provides: for example if I want to define a query set I need to write something like class EventQuerySet(SearchableQuerySetMixin, models.QuerySet). I don’t think too hard about it and it seems to work.

(as a meta comment: I’ve been working on talking about my programming opinions by just saying “THING does not feel good to me, I prefer OTHER THING instead”. That post I linked to says that function-based views are the “right way”. I’m not very invested in whether it’s “right”, but it’s validating to know that other people feel similarly to me about inheritance)

I don’t know how to think about Django performance

At some point the LLM scrapers discovered our site, and started sending us maybe 10 requests per second. I blocked them which is working for now, but it made me think about what the site’s capacity is. I’m used to writing Go backends where the performance situation is pretty straightforward (usually everything is just fast enough), and a Django site is very different.

Some light load testing (with (ab -n 1000 -c 1) shows that right now we can serve about 2-3 requests per second (on a ~$10/month VM).

It’s tempting for me to go down a rabbit hole where I do a bunch of profiling to figure out what’s slow and try to make it faster (there’s py-spy for that, and py-spy is great and super easy to use, and profiling is fun!) But I really don’t understand what I should expect in terms of performance from a Django site and how I should be thinking about at a higher level.

Some things I haven’t figured out yet:

  • If I have a site that’s going to be getting occasional bursts of traffic, do I want to be able to scale up?
  • Do I want to design the site so that more things can be cached? (and do I really have to? caches are so annoying to get right!)
  • The django performance docs say that Jinja is faster for templating, do I want to think about switching templating systems?
  • Those docs also say “{% block %} is faster than using {% include %}”, I wonder if it’s a big difference and if so why

template caching might be important

I think one thing I’m learning about Django is that because it’s a Framework (tm), it’s easy to accidentally misconfigure it. For example, when I was thinking about why my site was slow just now, I read the django performance docs and I noticed a comment saying:

Enabling the cached template loader often improves performance drastically, as it avoids compiling each template every time it needs to be rendered.

When I’d done CPU profiling I’d noticed that it was spending a lot of time rendering templates! Maybe this could help me!

Clicking through the link, I saw that the cached template loader was supposed to be on by default, but I’d turned it off by accident while trying to do something else. I think this “I turned off the cached template loader by default” things is an example of how I still find the django settings file to be pretty confusing and difficult. I guess I should just be careful when I go in there.

After turning on template caching, it seems like the site can now pretty easily handle 12 requests per second or so without using all of the CPU. I have not carefully benchmarked the before and after but it seems like it’s made a pretty big difference.

One thing that’s been surprising to me about Django performance is that I’ve always heard the advice “if you have a performance problem, check your database queries! Maybe add an index!”. But I’ve been running into a variety of performance issues (like this template caching thing) that are not because of slow queries, so instead it’s been more useful for me so far to start by running a CPU profile. And since I’m using SQLite, any slow database query problem will show up on the CPU profile anyway.

Anyway I don’t want to get too far into site performance. Like I said it’s easy for me to get interested in profiling, but actually I know a lot about profiling and it’s not the most important thing for me to learn about.

that’s all for now!

I might say more about what I’m enjoying (or having a hard time with!) about Django later. Trying to write some shorter blog posts recently.

Wolves, sheep, and gypsies

In 2012, the first Danish wolf in nearly two hundred years was discovered in the northwestern part of the country. Their numbers increased slowly in the years that followed, and by 2020, fewer than ten wolves were still being tracked. But then the population began to rise sharply: around 14 in 2021, 30 in 2022, as many as 80 by the end of 2024, and probably close to a hundred today.

At first glance, this is a heartwarming story. The wolf, in the abstract, is a majestic creature. It's great to see a species return after a few centuries of absence. You don't have to be a vegan hippie to find that appealing. But it's still a wolf! There's a reason why the last known wolf in Denmark didn't just wander off, but was shot dead in 1813.

Because as the number of wolves has been rising sharply every year since 2021, so too has the number of dead sheep. In 2021, it was 78, then in the years that followed, it was 162, 336, 521, and last year, it was 1,285! But the official line is still that shooting a wolf without a special permit is strictly prohibited because they're an endangered species, so maybe sheep herders could just try some more fences?

You don't have to be a statistician to see the problem here. When you have a surging wolf population, you'll have the rate of dead sheep following suit. If the only answer to the situation is "maybe try some more fences?", you're going to get more of what you already have: an out-of-control population that's an increasing risk to sheep, livestock, and potentially humans too.

This is the most basic consequence of inaction in the face of a growing threat: the problem just gets worse.

Cue the gypsies in Copenhagen, and their increasingly brazen behavior, after the city legalized overnight sleeping in parks and other green areas last year. Because "it shouldn't be illegal to be homeless", so the police refuses to act.

This has predictably led to these foreign vagrants setting up camp in the green areas around some of the city's metro stations to the great dismay of neighbors. The vagrants are literally shitting in the bushes, which obviously stinks, and the area isn't made any more inviting by the open-fire cooking that's also going on. 

Copenhagen is the Danish capital. It's repeatedly been named the safest city in the world in recent years. Partly because it didn't tolerate vagrants and homeless people taking over public spaces (as has happened in so many American cities). Previously, police would tell someone they couldn't sleep or camp in the city, and they could go to one of the city-run shelters. But now that this is considered too cruel(?!), the native inhabitants of the city instead have to deal with literal shit on their way to the metro.

These are two situations cut from the same cloth of ideology. Whether it's wolves or gypsies, you can't just let the problems get out of hand. You have to act. You have to protect the sheep. Standing idly by while your livestock is devoured by predators or your neighborhood is taken over by migrants is a pathetic abdication of any functioning society's most basic duty to its citizens.

When wolves get out of control, you shoot them. When gypsies take over public spaces, you deport them. This isn't hard, it isn't cruel. It's the basic logic of self-preservation.

Links 7/21/26

Links for you. Science:

A Deadly Ebola-Like Virus Is Spreading. Are We Ready?
Ancient Princesses Were Weapon-Wielding Badasses, Scientists Discover
Earliest Amber, Found in China, Changes Story of Plant Evolution
The San Andreas fault has gone ominously silent. Scientists fear when it finally snaps
FDA retracts test result but still links Taylor Farms lettuce to cyclosporiasis
AllTheBacteria: a community resource empowers biology and discovers novel peptide antibiotics
Trump administration allows killing of species threatened with extinction

Other:

It’s the Elon Economy, Stupid. George Noble, the former Fidelity wunderkind, argues that hyperscaler capex, SpaceX euphoria, and passive investing have created the most dangerous market bubble of his lifetime. The only question, he says, is what gives way first. (gift link)
Congress rejected some of Trump’s proposed budget cuts. OMB is making them anyway.
Ohio Dem Party Has Banned Criticism Of Gov. Candidate Amy Acton Calling Trans Girls “Boys”
A 23-Year-Old ‘ICE Chaser’ Is Notorious Among Federal Agents. He Has No Plans To Stop.
How Schools Can Counter “The End of Reading”
How Grindr’s C.E.O. Became an A.I. Tokenmaxxer
AI Made Cloning Games Easier Than Ever
Will America’s Next President Run Against A.I.?
How Google is Pushing Scam Videos on Elderly and Cognitively Impaired People
For many deported men, the hell of CECOT still haunts them
Trump is running out of space for his garish gold garbage
New Orleans Cops Published Policy Document Allowing Weaponized Drones
The Silver Lining of Trump’s Reheated Election Denialism. Take Trump seriously. Take him literally. But also see his speech for what it is: He knows he’s losing.
Texas border surveillance scholar exceeded tenure standards. UT’s president denied him.
Conspiracy Theorists Think Trump’s Speech Paves the Path to the Insurrection Act
Surrender To The Left On Medicare For All. It’s better than tearing the party apart!
The MAGA-ification of Hollywood is going to be so boring
In the Grim Darkness of the Far Future, There Is Only Culture War. How a forty-year-old satire of fascism became a meme for Trump, a flashpoint over women in sci-fi, and—for some fans—an actual political ideal.
Truth Social’s New Plan to Make Money May Be ‘Egregious,’ But Will Face Little Oversight
Zohran Mamdani’s City Hall Wants to Build Tenant Unions
State Sovereignty and Trump’s War Against the Constitution
Why Andrew Tate Is a MAGA Superstar
A trans ban came to rugby. Baltimore’s Ferals played on.
Denise Oliver-Velez, a Powerful Voice of the Left, Dies at 78
“Meddling”
Trump’s Reverse Midas Touch Even Got the World Cup
WTA to Require Genetic Sex Testing for All Players. Following cultural tides, the women’s tennis tour is mandating a strict new eligibility requirement.
Now on the American Road: Bulletproof Cars
OpenAI Appears to Be Missing Its Sales Goals by a Vast Margin
Pizzeria Paradiso Is Closing Its Georgetown Location After 23 Years

MIT to Become Hotbed of AI Video Surveillance

It’s a lot:

According to information obtained by The Tech, MIT is spending over $3 million on more than 500 AI surveillance cameras in academic buildings, residence halls, and outdoor areas along Memorial Drive. Installation of the new cameras, along with the wiring and infrastructure that will support them, began November 2025 and will likely continue until September 2026.

Technical specifications for the cameras suggest that they will be capable of collecting real-time face and object classification data, including detection of motion, loitering, crowds, face masks, and camera tampering. Individuals can also be automatically classified on the basis of clothing color, gender, and age, up to a distance of 35 feet (11 meters) from the camera. According to a statement from MIT spokesperson Kimberly Allen, any collected data is “retained up to 30 days,” unless an exception is granted.

[…]

Most of the new cameras, which are part of Hanwha’s Wisenet AI line, are marketed for their ability to identify and classify multiple objects with deep learning algorithms. They support resolutions ranging from 2MP to 4K while also recognizing faces, license plates, vehicles, and other objects in real time.

Nearly all cameras will accommodate a wide range of pan, tilt, rotate, and zoom motion and will be monitored continually with Ai-RGUS, an AI camera software.

Yikes.

Tuesday assorted links

1. White on Hutt’s macro.

2. Are LLMs obsessed with Japan?

3. Princeton University art museum gets landmark gift of Haitian paintings.

4. “Sight unseen” — the revenge of asymmetric information.  But was it a good buy?

5. “Democratic types would give up about $210,000 and Republican types about $100,000 in annual partner income to avoid a typical cross-party partner.

6. Behind the Circe scene.

7. More on AI and mathematics.

8. Tribute to George Akerlof, who just retired.

The post Tuesday assorted links appeared first on Marginal REVOLUTION.

       

Are vouchers appropriate for hard-to-match patient-donor kidney exchange pairs?

 A recent letter from Mayo Clinic transplant providers raises concerns related to bad outcomes when some patient-donor pairs are separated (the donor donates before the intended recipient is matched to a donor), but the intended recipient is hard to match.

Donor Mentoring in Private Kidney Paired Donation Programs: The Need for Transparency and Accountability by Tayyab S. Diwan, Pooja Bhudiraja, Shennen Mao, and Carrie Schinstock, Transplantation 110(7):p e1555-e1556, July 2026. | DOI: 10.1097/TP.0000000000005764 

"it is important to recognize that some KPD organizations, such as the National Kidney Registry, operate as private organizations (POs) with distinct governance structures, priorities, and commercial interests. These objectives may not always fully align with those of transplant centers or the broader transplant community, potentially creating gaps in transparency and accountability within a process that directly affects donors and recipients.

"We recently encountered a concerning example that illustrates the ethical tensions inherent when donor mentors are employed by private KPD organizations.
... During a routine text message check-in, the mentor initiated a discussion promoting voucher donation, emphasizing that voucher donors receive priority within the system.  

...
"The donor described the interaction as uncomfortable and “pushy,” and subsequently contacted our transplant center for guidance. Notably, the communication focused on system-level benefits and prioritization metrics but did not explore critical clinical and ethical considerations specific to this donor-recipient pair.

"The discussion was insufficiently individualized and appeared framed toward a preferred outcome. The donor’s preferences and the potential downsides of voucher donation were not explored. The mentor did not ask how the donor would feel if they proceeded with voucher donation and their intended recipient never received a transplant. The discussion lacked any reference to recipient factors that could influence the advisability of voucher donation, including sensitization or recipient comorbid conditions. There was also no discussion about why KPD was being used. Often the donor can donate directly to a recipient (eg, compatible pair), but may decide to enter a KPD pool temporarily to see whether the recipient could receive a transplant with improved size or HLA matching. Finally, the average wait times for a recipient to receive a transplant after voucher donation were not mentioned.

"This interaction raises concern about whether the mentor’s guidance was fully centered on the donor-recipient pair or instead influenced, even indirectly, by organizational incentives to increase voucher participation. When donor mentors are employed by entities that operationally benefit from particular forms of donation, even well-intentioned communication can create perceived or actual conflicts of interest. Transparency regarding these structural incentives, along with a clear commitment to individualized donor-centered mentoring, is essential to maintaining ethical integrity in these situations.

...

"As stewards of living donor trust, the transplant community must ensure that innovation proceeds with appropriate oversight. Any perception that donor education or support is influenced by business interests risks eroding public confidence in living donation programs, with potentially far-reaching consequences that could discourage future donors." 

#########

 Here's an article on NKR from the NYT that discusses its particular voucher program (and also its finances):

How One Father Created an Organ Empire: The National Kidney Registry has matched thousands of kidney donors with recipients. It has also paid millions of dollars to a company owned by its founder.
By Danielle IvoryGrace Ashford and Robert Gebeloff
Dec. 27, 2025 

 "One of N.K.R.’s most innovative policies, many doctors said, is known as voucher donation: Donors can choose to give their kidneys immediately, in exchange for organ vouchers their loved ones redeem later.

"This arrangement has helped N.K.R. expand its pool. But vouchers add risk for donors giving on behalf of hard-to-match patients. They might go through surgery months or even years before their loved ones get a match. Matches are especially unlikely, doctors said, for “very highly sensitized” patients who carry antibodies likely to reject a transplant.

...

" many doctors said that given N.K.R.’s relatively small size, these rare patients would be better served waiting for a match before their loved ones donate.

"“We definitely recommend that they not go with a voucher,” said Dr. Miklos Molnar, the medical director of the kidney transplantation program at the University of Utah, which works with N.K.R.

"Dozens of highly sensitized patients who are also on kidney dialysis have waited a year or more after their loved ones donated, according to N.K.R. quarterly reports since 2022. At least three waited more than four years, and one for seven years.

"Doctors told The Times that they knew of some patients who became too sick or died before they redeemed their vouchers."

######## 

Here's a a discussion of the earliest exploration of unpaired donation that I know of:

Sunday, September 13, 2020

  

SpaceX launches novel geosynchronous robotic servicing satellite on decade-long mission

A SpaceX Falcon 9 rocket lifts off from Space Launch Complex 40 at Cape Canaveral Space Force Station on July 21, 2026. Onboard was the U.S. Naval Research Laboratory’s Robotic Servicing of Geosynchronous Satellites (RSGS), the main payload on Northrop Grumman’s Mission Robotic Vehicle spacecraft. There were also three Mission Extension Pods from Northrop Grumman onboard. Image: John Pisani/Spaceflight Now

A SpaceX Falcon 9 rocket launched from Cape Canaveral Tuesday carrying Northrop Grumman’s Mission Robotic Vehicle (MRV) and three Mission Extension Pods (MEPs), a novel mission designed to prolog the lifespans of multiple satellites.

The MRV is equipped with the Robotic Servicing of Geosynchronous Satellites (RSGS) payload, a pair of robotic arms with an array of tool attachments developed at the U.S. Naval Research Laboratory with financial backing from the Defense Advanced Research Projects Agency (DARPA).

Following its launch in summer 2026, Northrop Grumman’s Mission Robotic Vehicle (left) will install Mission Extension Pod (right) jetpacks to extend the lives of client satellites in geosynchronous orbit. MRV also can perform inspection, relocation, repairs, upgrades, and other missions yet to be imagined. Illustration credit: Northrop Grumman

Following deployment, the MRV and MEPs will spend a year heading out to geosynchronous Earth orbit. Once there, the MRV will begin the process of installing MEPs onto clients’ spacecraft, including those owned by Australia-based company Optus and Luxembourg-based company SES.

The MEPs, which carry fresh supplies of maneuvering fuel, are designed to provide up to eight years of additional life for a satellite.

Liftoff from Space Launch Complex 40 occurred at 5:15 p.m. EDT (2115 UTC). The rocket flew due east upon leaving the launch pad.

SpaceX launched the mission using the Falcon 9 first stage booster B1069, which made its 32nd and final flight. The company said this would be the boosters final mission “Due to the additional performance needed to launch these payloads to geosynchronous transfer orbit.”

In addition to 27 batches of Starlink satellites, B1069 launched the following missions:

  • Dec. 21, 2021 – NASA’s CRS-24
  • Oct. 15, 2022 – Eutelsat’s Hotbird 13F
  • Dec. 8, 2022 – OneWeb’s OneWeb Mission 15
  • Mar. 17, 2023 – SES’ SES-18 & 19

Decades in the making

The concept behind the RSGS goes back more than 20 years when the NRL began examining the possibility of two uncrewed spacecraft performing autonomous rendezvous and docking operations. In the early 2000s, that evolved into a study called ‘RescueSat’, looking at how to recover a satellite stuck in the wrong orbit, and eventually the Spacecraft for the Universal Modification of Orbits (SUMO) program.

“The SUMO mandate was daunting. DARPA challenged us to design a robot that could dock with any satellite in space,” said Glen Henshaw, Ph.D., NRL Lead Space Roboticist for RSGS in a prelaunch statement.

“We then realized a universal truth—every satellite got to space on a rocket! By targeting the sturdy ‘launch vehicle interface plane’—the structural ring or explosive bolt holes that attach a spacecraft to a rocket for launch—we determined that a robotic arm could safely grapple almost any spacecraft without damaging delicate instruments.”

The MRV inside Northrop Grumman’s Dulles, Va., Satellite Manufacturing Facility. Image: Northrop Grumman

After proving the concept was possible, by 2005 the SUMO program shifted to the Front End Robotics Enabling Near-term Demonstration (FREND). That program would go onto develop spaceflight worthy robotic arms through a commercial vendor.

Alliance Spacesystems, Inc. (ASI), the company who built the robotic arms for NASA’s Mars Curiosity Rover, was selected and environmental testing began in 2008.

To find a good business case for these arms, a series of studies were conducted, including a joint NASA-DARPA study in 2012 called the Manned GEO Servicing study. Ultimately, DARPA moved the technology back into what the NRL called “research and hardware maturation via a spaceflight concept program called Phoenix.”

After years of studying and contemplation, DARPA decided to proceed with what was then called the RSGS payload to pursue “ultra-close inspection, orbital ‘tow truck’ relocation, mechanical anomaly repair, and upgrading satellites not originally designed to be upgraded.”

Four years later in 2019, it selected SpaceLogistics, a Northrop Grumman company to combine the RSGS with the MRV. Following initial checkouts, the RSGS program will be turned over to the U.S. Space Force in support of its Servicing, Mobility, and Logistics portfolio.

A Northrop Grumman employee at the company’s manufacturing facility in Gilbert, Arizona, performs final inspections of one Mission Extension Pod (MEP). Image: Northrop Grumman

The Odyssey

Worth seeing, and I enjoyed it — but a few observations I haven’t seen elsewhere.

For all the praise of IMAX and 70mm’s supposed clarity, several scenes are out of focus on the actor. Pulling focus is harder with a large-format negative. The problem is compounded by Nolan’s fondness for darkness: too many scenes are dark enough to squander much of what the format offers. The sound was earth-shaking in the way we have come to expect from Nolan but there is no song.

The “woke casting” controversy is a non-issue — barely noticeable in practice. The film is obviously conservative in temperament. Helen gets some of the best lines, and Nolan’s slight disfigurement of her traditional arc is exactly right. The Circe scene is the best in the film.

Tyler is entirely wrong about Calypso. Odysseus’s seven years with her were among his most enjoyable. I have no doubt about this.

The deeper flaws are structural. Nolan loves to play with time, and the resulting flashbacks and memories ironically shortchange the odyssey itself — making the journey feel shorter and less arduous than it should. This is very much in the mode of Interstellar: a sequence of set-piece locations strung together. One planet/one monster/one scene–on to the next. But Odysseus as a character on an odyssey never quite coheres.

We are told repeatedly that Odysseus is smart but we shouldn’t need to be reminded. Odysseus is both beloved and resented by gods, a man whose men will follow him to the ends of the earth and then betray him but Mat Damon just doesn’t bring it. Things happen to him; he responds stoically. What we needed was the equivalent of Kirk defeating the Kobayashi Maru — a moment that makes the audience understand, viscerally, that this man bends the rules and contends with the gods by the sheer force of his wit and will. Of courses the Trojan horse is this but Nolan treats this as something of which Odysseus is ashamed and the other clever bits are downplayed. Damon never gets his Kobayashi Meru. The odyssey is a slog, rather than an adventure. Could have used a bit more Sinbad, a bit less Dark Knight. The dialogue, as Tyler noted, is lame. 

Not Nolan’s best film but still better than most films and for scale, ambition, and grand themes well worth the 3-hour investment.

The post The Odyssey appeared first on Marginal REVOLUTION.

      

Related Stories

 

Agile Space Industries Appoints new CTO and Board member

agile space industries logo

Durango, Co – July 21, 2026 – Agile Space Industries today announced the appointment of Jake Mills as Chief Technology Officer (CTO) and Mark Pasquale as an independent member of […]

The post Agile Space Industries Appoints new CTO and Board member appeared first on SpaceNews.

KBR organizes defense technology business to pursue Golden Dome work

Company points to digital engineering, space tracking and missile-defense experience

The post KBR organizes defense technology business to pursue Golden Dome work appeared first on SpaceNews.

IHI explores Kuva hyperspectral satellites for Japan’s multi-sensor constellation

Japan’s IHI is considering adding next-generation hyperspectral satellites from Finnish startup Kuva Space to its planned 100-strong, multi-sensor sovereign Earth observation constellation.

The post IHI explores Kuva hyperspectral satellites for Japan’s multi-sensor constellation appeared first on SpaceNews.

Italian startup ORiS raises funding for laser power-beaming technology

ORiS laser power beaming

Italian startup ORiS has raised 5 million euros ($5.7 million) to help it develop wireless power-beaming technology for satellites using lasers.

The post Italian startup ORiS raises funding for laser power-beaming technology appeared first on SpaceNews.

German component supplier deltaVision raises 10.2 million euros

deltaVision

German spacecraft component developer deltaVision has raised 10.2 million euros ($11.6 million) to expand production of valves and related equipment, with a focus on enabling in-space refueling.

The post German component supplier deltaVision raises 10.2 million euros appeared first on SpaceNews.

The people’s atheist

Painting of two figures praying in a field at dusk with a pitchfork and wheelbarrow nearby, evoking a serene, rural setting.

Jean Meslier, a rural French priest, argued that religion is designed to produce subjects who will automatically obey

- by Nidal Taibi

Read on Aeon

SpaceX launches Starlink mission from California a day after last-second abort

The Starlink 17-39 mission lifts off from Space Launch Complex 4E at Vandenberg Space Force Base in California on July 21, 2026. Image: SapceX.

A SpaceX Falcon 9 rocket launched two dozen Starlink V2 Mini broadband internet satellites to its massive low Earth orbit constellation on Tuesday, a day after liftoff was scrubbed by a rare abort during engine ignition.

Liftoff from Vandenberg Space Force Base in California came at 7:49 a.m. PDT (10:49 a.m. EDT / 1449 UTC). The first launch attempt for the Starlink 17-39 mission ended as the countdown clock hit zero when an issue popped up during the ignition sequence. All nine Merlin 1D engines appeared to ignite but then shutdown in a flash of fire and smoke.

But on Tuesday, all appeared to go smoothly, with the rocket lifting off from Space Launch Complex 4 East under mostly clear skies, taking a south-southwesterly trajectory to reach a 97-degree inclination orbit. Deployment of the 24 Starlink satellites came just over an hour into flight.

The mission used Falcon 9 first stage booster B1082, making its 23rd flight after launching missions, including NROL-145, USSF-62, and OneWeb Launch 20.

More than eight minutes after liftoff, B1082 landed on the droneship, ‘Of Course I Still Love You’, which is positioned in the Pacific Ocean, the 212th landing on this vessel and the 640th booster landing to date.

The Starlink 17-39 mission was the company’s 85 Falcon 9 rocket launch of the year. Of those, 67 were in support of the Starlink constellation, which consists of more than 10,800 spacecraft.

[Sponsor] WorkOS MCP: Manage Your Auth Platform From Any AI Agent

Debugging SSO, managing users, adjusting auth policies, configuring branding: every configuration task has lived behind a UI that only a human can drive.

The WorkOS MCP server gives agents the same access as your dashboard login. Hundreds of operations, discoverable at runtime. Connect in one command via OAuth, with scoped tokens instead of a master API key. Pass a screenshot of your marketing site and ask your agent to match the login page. If a human had to do it before, an agent can do it now.

Connect your agent →

 ★ 

July 20, 2026

On Friday an Iranian attack on a U.S. base in Jordan killed two soldiers and left one missing. Another service member died Friday in northern Iraq. This brings the number of military personnel killed in the Iran War to 17 since Trump launched strikes against the country on February 28. The Pentagon’s casualty-reporting site says 427 U.S. service members have been wounded.

On Saturday, Greg Jaffe, Julian E. Barnes, and Jonathan Swan of the New York Times reported that multiple officials told them that the attack in Jordan was the fourth in five days and that the strikes had wounded dozens of military personnel and damaged several helicopters. The journalists’ sources told them that the strikes proved Iranian forces still have plenty of missiles and are becoming more skilled at getting around U.S. air defense systems.

The sources told the journalists that the first attack hit a base, injuring as many as five service members; the second hit a different base, damaging U.S. Blackhawk helicopters; and a third hit the same base where the soldiers would later be killed. In that first attack, about 20 military personnel were injured.

Retired Air Force Lieutenant General David Deptula, who helped to plan the air war in the 1991 Gulf War, explained that bases in Jordan give the U.S. broader reach across Syria, Iraq, and beyond. Friday’s attack was “an attack on the U.S. regional coalition and an attempt to make the political cost of hosting American forces greater than the security benefit.”

Those sources who spoke to the New York Times reporters spoke on condition of anonymity. The Pentagon declined to comment.

Eric Schmitt of the New York Times noted that lack of information from the Pentagon today when he reported that in the week before the weekend’s deaths, three Iranian strikes injured military personnel and damaged several helicopters yet the Pentagon under Defense Secretary Pete Hegseth did not disclose the strikes, the casualties, or the damage to the American people, citing the need to keep information about operations secure.

Schmitt noted that neither the White House nor the Pentagon has outlined a military strategy since Trump announced that the memorandum of understanding that extended a ceasefire—the one he signed on June 17—was no longer in force. In the nine days since then, both sides have increased strikes.

Schmitt noted the Pentagon has not held a major briefing on the war since early May. Since then, the administration has not disclosed the depletion of crucial and expensive weapons through the Iran strikes, or how Iran has both retained and rebuilt its own missiles. In June, Jonah Kaplan and Michael Kaplan of CBS News reported that Hegseth’s claim that “almost 90%” of service members injured had sustained only minor injuries gives a misleading picture of the extent of those injuries. A spokesperson for the Army explained that the Army classifies a soldier as “seriously injured” or “very seriously injured” only if they are at risk of dying from their wounds within 72 hours.

Pentagon spokesperson Sean Parnell responded to Schmitt’s story by posting on social media that “[t]he Department of War rejects these baseless and malicious accusations of hiding injury numbers as outright lies from partisan hacks at the New York Times who are desperate to smear America’s military and its leadership. Claims of concealment are fabrications meant to further distress the American people in the wake of three service members killed in action.” He continued: “This is the most transparent Department of War in history.”

And yet, the administration did not consult Congress about launching the war, as the Constitution requires, and has blown through the 60-day time limit for the use of the U.S. military to counter an “imminent threat” without getting congressional approval.

Damian Paletta of the Wall Street Journal today also asked questions about how the Iranians, who Trump said had very few missiles left back in June, are developing faster and more maneuverable missiles. Paletta noted that whether Iran is getting help from Russia, or any other country, remains an “enormous question” that puts “much more pressure on the White House to potentially confront Russia.”

When asked about the deaths of the three service members, Trump said “We feel very badly, but you know those great people, those great patriots were out there fighting that Iran cannot have a nuclear weapon. Iran has been very, very badly damaged. They’ve lost everything almost, militarily. They’ve got very little left…. We control the strait. They don’t control anything. So we’ll see what happens when we hit ’em very hard again tonight.”

Over the last several days, the U.S. has struck Iranian bridges and energy infrastructure. According to Iran’s Health Ministry, the recent U.S. attacks have killed at least 38 people and wounded more than 400.

Today the Houthis in Yemen, who are backed by Iran, announced they are imposing a naval blockade on Saudi Arabia, threatening the passage of oil through the Bab el-Mandeb strait. If the strait is closed, the global oil supply will take another 7% hit.

The American Automobile Association said today that the average price for a gallon of gasoline is just over $4, up about $0.13 from last week. Before the conflict, the price averaged $2.98 a gallon. Economist Paul Krugman noted today that the “wholesale price of refined products (assuming 2/3 gasoline, 1/3 distillates) is most of the way back to its peak. Since this is what matters for inflation and demand, no relief in sight before the midterms.”

Carl Quintanilla of CNBC reported today that crude oil supplies in the U.S. are the lowest they’ve been in 45 years.

The expansion of the Iran War’s scope, duration, and casualties did not keep Trump away from the World Cup final between Spain and Argentina on Sunday. After Spain won over Argentina with a score of 1–0, Trump joined FIFA President Gianni Infantino on the field—apparently unconcerned about the dangers that he says require the construction of a fortified ballroom—to present medals to the Spanish players.

The crowd resoundingly booed Trump, who then stood on stage with the celebrating players until Infantino jogged across the stage to lead him away. The Guardian captioned a picture by Javier Garcia taken of Infantino and Trump standing aside from the players: “Gianni Infantino carefully explains to Donald Trump that he is not part of the Spanish national team.”

Trump appears to be becoming less and less relevant to actual governance. Last Thursday, Alex Gangitano, Megan Messerly, and Myah Ward of Politico reported that Republicans were “scared sh*tless” about what Trump might say in his prime time address, but it is perhaps more of a sign of the current political reality of Trump’s falling popularity that CNN, ABC and NBC, chose not to run the speech live.

Matt Gertz of Media Matters noted that despite Trump’s promise that his speech would deliver “really, really big news,” the Fox News Channel largely ignored the event. Executives might not have wanted to get entangled in the election denialism that cost them close to $800 million after pushing claims of election fraud after the 2020 presidential election.

Fox hosts turned instead to culture war issues, railing against “DEI” and perceived abuse of WNBA Indiana Fever guard Caitlin Clark, who is white, by Black players.

In line with that focus, the administration is renewing its visible and aggressive immigration policies. For months, Immigration and Customs Enforcement (ICE) has not released data about arrests, but immigration enforcement scholar Austin Kocher noted that new records show arrests in early July hit 1,474 a day, a record high for the Trump administration.

In the past months, ICE has recruited thousands of new agents. As Maanvi Singh and Fabiola Cineas of The Guardian note, many of these new agents did not have the requisite qualifications and were not properly vetted.

Agents from ICE shot and killed two immigrants within a week earlier this month: Lorenzo Salgado Araujo, 52, in Houston, Texas, on July 7 and Joan (or Johan) Sebastián Durán Guerrero, 25, in Biddeford, Maine, on July 13. Neither man was the target of the operation in which he was killed.

In Maine, Ashley Brouillette identified her ex-husband, David Brouillette, as the officer who shot Guerrero four times. She found out he was involved in the shooting when he called her to ask her not to tell reporters about his abuse during their marriage. Both she and another ex-wife had restraining orders against Brouillette.

Ashley Brouillette told reporters that when her ex-husband told her at the end of last year that he had been hired as an ICE agent, “I honestly thought that he was being delusional and I didn’t believe it.”

As an Army combat veteran and former security agent, Brouillette would likely not have undergone in-person training to become an ICE agent, Christian Harsa of News Center Maine reported. A spokesperson for ICE, Lauren Bis, told Vanessa Romo and Alina Hartounian of NPR that Brouillette had “nearly a decade of federal law enforcement experience with required training.”

Today, Senator Angus King (I-ME) and 38 Democratic senators sent a letter to Secretary of Homeland Security Markwayne Mullin demanding “immediate, thorough, independent, and transparent investigations” of the shooting deaths of Araujo and Guerrero.

As Isabelle Oss of the Portland Press Herald reported, the letter noted that both former DHS secretary Kristi Noem and White House immigration advisor Tom Homan committed to making sure that agents are fitted with body cameras. And yet, agents involved in the recent shootings were not equipped with cameras, confirming “that neither of these commitments were honored.”

The senators noted that the acting director of ICE, David Venturella, promised Congress that all field agents would be equipped with cameras by the end of July. “We view this timeline not as a projection, but as a firm, binding commitment to which we will hold the Department accountable,” the senators wrote.

They also demanded that ICE agents wear labels clearly identifying themselves as ICE agents.

The senators asked Mullin to identify specifically how he intended to investigate the shootings, explain the vetting process for new recruits including whether it considers records of domestic abuse before hiring, identify updates or reviews being made to ICE traffic stop policies, and explain what oversight, reporting measures, or public safety measures the agency is implementing.

As if to demonstrate that he is still in charge, Trump this afternoon announced new tariffs on certain products from Canada. The 50% tariffs would further upend the economy if they take effect in 30 days as scheduled.

Trump has been complaining about the smoke from Canadian wildfires hurting U.S. air quality, and has threatened to impose further tariffs on Canadian produce because “the United States is being unnecessarily invaded by filthy, polluted, and unhealthy air.” Canadians including Ontario premier Doug Ford noted that they sent firefighters, rather than threats, when conditions were reversed.

Notes:

https://www.nytimes.com/2026/07/20/us/politics/troops-injured-jordan-iran-war.html

https://www.cbsnews.com/news/wounded-soldiers-families-accuse-army-downplaying-war-injuries/

https://www.nytimes.com/2026/07/18/world/middleeast/iran-war-jordan-attacks.html

https://www.pbs.org/newshour/world/as-u-s-strikes-bridges-in-iran-it-targets-a-water-desalination-plant-in-kuwait

https://gasprices.aaa.com/

https://www.cnn.com/2026/07/20/business/gas-four-dollars

https://www.reuters.com/world/middle-east/yemens-houthis-declare-naval-blockade-against-saudi-arabia-statement-2026-07-20/

https://www.wsj.com/politics/irans-missiles-have-gotten-faster-and-deadlier-the-big-question-is-how-41c81ad7?st=LvaXV6

https://www.npr.org/2026/07/20/nx-s1-5900601/us-iran-updates?renderPlatform=nprone_ios

https://people.com/donald-trump-booed-at-2026-world-cup-final-12021141

https://www.politico.com/news/2026/07/16/republicans-brace-trump-speech-election-fraud-00999640

https://www.theguardian.com/us-news/2026/jul/16/us-tv-networks-trump-speech

https://www.mediamatters.org/voter-fraud-and-suppression/fox-news-buries-trumps-election-denial-address

https://www.theguardian.com/football/live/2026/jul/19/spain-v-argentina-world-cup-2026-final-live-updates

https://www.theguardian.com/news/ng-interactive/2026/jul/19/ice-killings-immigration-crackdown

https://www.npr.org/2026/07/17/nx-s1-5897460/maine-ice-shooting-brouillette

https://www.nytimes.com/2026/07/19/us/politics/fbi-ice-agents-investigations-shootings.html

https://www.wbaltv.com/article/trump-threatens-tariffs-canada-wildfire-smoke/73175566

https://www.pressherald.com/2026/07/20/dozens-of-senators-sign-angus-king-letter-demanding-dhs-reform-after-deadly-ice-shootings/

Austin Kocher
ICE Quietly Arresting More People Than Ever According to Latest Data
After months of data silence, ICE released new detention numbers today that provide important new insights into the current state of the Trump administration’s mass deportation campaign. As of July 11, the agency currently holds 65,765 people in custody across the country—up from…
Read more

https://www.newscentermaine.com/article/news/local/maine-immigration/dhs-identity-ice-agent-shot-killed-johan-sebastian-duran-guerrero-biddeford-maine/97-54f94437-e1f9-499d-873f-f503fcccfa69

https://apnews.com/article/fox-news-dominion-lawsuit-trial-trump-2020-0ac71f75acfacc52ea80b3e747fb0afe

Substack Notes:

@paulkrugman/note/c-298196011

X:

seanparnellasw/status/2079204590176940302?s=46

RpsAgainstTrump/status/2079018774116675751

Bluesky:

carlquintanilla.bsky.social/post/3mr36exant22u

atrupar.com/post/3mr25ecdlu22e

mattgertz.bsky.social/post/3mr34shh6ms25

verbeeld.bsky.social/post/3mqzvanypmk2t

sky.skymarchini.net/post/3mr47wxwg5c2w

austinkocher.com/post/3mr4rcnntj52g

jbendery.bsky.social/post/3mqurhiycw22r

Share

The economic effects of GLP-1s

We estimate the causal impacts of GLP-1 treatment on labor market outcomes using linked Danish administrative data and a matched stacked difference-in-differences design. We compare patients who initiate GLP-1 treatment during the first two years of Semaglutide availability to observably similar patients who initiate four years later. We find that GLP-1 treatment reduces long-term sickness leave by 17.3%. We estimate total fiscal benefits of GLP-1 initiation of approximately 1.3–1.5% of annual labor income per employed individual. We do not detect statistically significant or economically meaningful impacts on income, labor force participation, or employment over four years.

That is from a new NBER working paper by N. Meltem Daysal, Camille JH. Fredrickson, Ida L. Kristiansen, Mircea Trandafir & Jonathan Zhang.

The post The economic effects of GLP-1s appeared first on Marginal REVOLUTION.

      

Related Stories

 

Why Maine’s Sandy Shorelines Turn Jagged

A satellite image of the Maine coastline highlights smooth, sandy beaches near Saco Bay on the left and the rocky, jagged coastal features near Casco Bay on the right.
The stark contrast between the curved, sandy beaches south of Portland and the indented, rocky coastline to the northeast is clear in this image captured by the OLI (Operational Land Imager) on Landsat 9 on August 31, 2025.
NASA Earth Observatory/Michala Garrison

The Wabanaki people have a deep well of creation myths explaining the rocky coastlines of the Bay of Fundy, Downeast Maine, and Acadia National Park. Many involve Glooscap—a magical figure said to have floated down the Bay of Fundy in a stone canoe, sculpting coastal features by scraping the vessel across the landscape and scattering enormous boulders during battles with primordial beavers, frogs, moose, whales, and other gigantic animals.

Fewer Indigenous creation myths survive to explain the origins of the sandy and marshy shorelines of southern Maine and the rocky, indented coasts of the state’s Midcoast region. But the sharp contrast between the sandy shoals and beaches south of Portland and the rocky shoreline of promontories, headlands, and narrow peninsulas to the east—visible in the Landsat image above—has long drawn the attention of coastal geologists, whose scientific explanations on its origins abound.

The coastal transition reflects both differences in the underlying bedrock and the distribution of sediment left behind by the last glacial maximum, coastal geologists say. Southern Maine has broad deposits of sand, much of it sourced from rivers. The sandy beaches of Saco Bay, for instance, home to Maine’s longest contiguous beach and the state’s largest saltmarsh, received sediment from the weathering and breakdown of the White Mountains, with material transported to the coast largely by the Saco River, explained Peter Slovinsky, a geologist with the Maine Geological Survey. Waves and tides reworked these soft sediments over time, sculpting them into the arch-shaped embayed beaches and sprawling salt marshes found around Saco Bay and the broader region.

While erosion-resistant granite juts from the sandy shorelines in southern Maine to form rocky headlands, metamorphic bedrock becomes the dominant surface feature east of Portland. There, whole ridges and valleys made of rock layers transformed by exposure to high pressures and temperatures define the landscape. During the last ice age, glaciers scoured and widened many of these coastal valleys, which later flooded as the Laurentide Ice Sheet melted and sea levels rose.

Around Casco Bay, these ridge-and-valley systems, combined with the drowning of the shoreline, produce the jagged, highly indented shoreline and many long, narrow islands seen today. “The tortured folds of these old landscapes also set up a sharp directional preference for erosion to exploit,” said Nicholas Whiteman, also a geologist with the Maine Geological Survey. “This led to the eye-catching difference in the orientation of the islands and necks that dominate Casco Bay compared with those to the northeast.”

The various forms that coastlines take fascinate geologists, but they also carry everyday implications for the economies of Maine’s coastal communities. While tourists flock to the sandy beaches of communities like Saco and Kennebunkport, the state’s iconic lobster fisheries are concentrated in Midcoast Maine. The crustaceans thrive in the cold waters of the region’s many rocky, protected inlets, turning communities such as Harpswell into leaders in lobster landings.  

The state’s oyster farms are also concentrated in this region. Casco Bay and the Damariscotta Estuary, sheltered from winds and waves, offer waters that farmers can easily access without large boats. These waters provide a range of temperatures, salinities, and other characteristics that create numerous microclimates where oysters can grow quickly and take on a variety of tastes, known as merroir, explained Tom Kiffney, a researcher at the University of Maine. Kiffney is part of a team of researchers using Landsat and other satellite observations to predict oyster growth rates and help identify the most promising locations for new oyster farms in Maine based on water temperatures and quality.

NASA Earth Observatory image by Michala Garrison, using Landsat data from the U.S. Geological Survey. Story by Adam Voiland.

References & Resources

You may also be interested in:

Stay up-to-date with the latest content from NASA as we explore the universe and discover more about our home planet.

Contours of the James Bay Lowlands
3 min read

After the Laurentide Ice Sheet retreated from present-day Hudson Bay, rebounding land has revealed striking nearshore topography.

Article
A Tide-Fueled Trove of Biodiversity in Guinea-Bissau
3 min read

The expansive mudflats, sandy beaches, and mangrove forests of the Bijagós archipelago support an array of migratory shorebirds and large…

Article
Thailand’s Krabi Coast
3 min read

The coastal province features striking tropical karst landscapes and sandy beaches alongside a mix of natural land cover and developed…

Article

The post Why Maine’s Sandy Shorelines Turn Jagged appeared first on NASA Science.

Is America ruled by an oligarchy?

Any time you hear claims about American government, a sanity-inducing response is to ask whether they also are true of state and local governments.  Here is an excerpt from my latest piece at The Free Press:

First, in people’s daily lives, they typically interact with their state and local governments more than with the feds. And if you look at what state and local governments do, most of it is driven by voter demand, and I do not mean billionaire or oligarchic voters.

Most state and local government spending goes toward schools, roads, and increasingly, Medicaid. None of those reflect the agendas of most billionaires. For example, Medicaid dollars flow particularly to lower-income recipients. Hospitals and doctors receive income from this program, but reimbursement rates are relatively low, and many of the best doctors do not accept Medicaid patients and don’t make a living from the program.

Somehow the oligarchy saw fit to leave these expenditures alone.  Looking at the state and local level also suggests that America — with some notable exceptions — is better governed than before.

The post Is America ruled by an oligarchy? appeared first on Marginal REVOLUTION.

      

Related Stories

 

Investing in the Future: The Serhii Tokarev Foundation and uBoost Sum Up SheLeads

This Year’s SheLeads from the Tokarev Foundation and uBoost Come to an End

The season is coming to an end, and the SheLeads educational program has already announced its winners. Each season, hundreds of young women participate in this initiative to turn their ideas into reality. The program guides participants through the journey from idea to product.

The program was launched through a collaboration between uBoost and the Tokarev Foundation, founded by impact investor Serhii Tokarev . This season, more than 300 young women took on the role of founders, and each pursued her own idea and approach.

What Made SheLeads Stand Out: Highlights and Results

SheLeads STEM accelerator winners holding ceremonial checks on stage

Every year, the number of participants in the educational program grows, and this time it reached more than 300. Of these, 220 completed the full course, while 120 were assigned mentors to help bring their dream products to life. According to the jury’s decision, first place was awarded to three projects that are ready to change the world right now:

  • Energy Efficiency – a project presented by Sofia Kedria, who worked on it under the mentorship of Oleksandr Marchenko. The goal is to increase the efficiency of solar panels using lensing technology.
  • Physical Health – a project presented by Solomiya Andriychuk. The product turns workouts on gym equipment into an engaging game.
  • Environment – the ReGrow project, presented by Iryna Terekhova. The project’s goal is to restore damaged soil by planting microseeds. The key product is special packaging made from straw. As it decomposes, it releases the seeds into the soil, helping reduce waste while restoring damaged soil.

Each project addressed a major real-world challenge, and according to Serhii Tokarev , these young innovators represent Ukraine’s future. “SheLeads isn’t a competition. It’s proof. Ukraine will answer the question of who will rebuild the country – and SheLeads is shaping the very cohort of people who will do just that. Not someday. But right now – through a prototype, a pitch, and a first internship,”  emphasized Serhii Tokarev.

How SheLeads Is Shaping the Future

Every year, the number of young women joining the educational program grows. This clearly shows that many of these participants will go on to become founders, developers, and innovators. “We see how participants come in with an idea and leave with a product, a team, and their first industry contacts. This is exactly the infrastructure of opportunities we want to create,”  concludes Anastasia Deeva, CEO of the Tokarev Foundation.

The program’s impact doesn’t end with the finals, as participants gain valuable experience and make their first meaningful connections in the industry. Many participants go on to apply what they have learned in other programs. This season, for example, seven girls applied to the Technovation Girls competition, and two team projects made it to the quarterfinals.

The Project’s Roots

SheLeads would not have been possible without the fruitful collaboration between uBoost and the Tokarev Foundation. The program was created to provide young women with not only the opportunity but also the necessary resources to build their own product. Each stage is both an experience and a challenge, with experienced mentors and industry experts supporting each participant. The program concludes with projects being evaluated by a panel of judges, as well as internship offers from teams at tech giants. This season’s partners include Sigma Software Labs, Ukrainian Startup Fund, Healthy Mind, and AI HOUSE.

The Tokarev Foundation is an organization dedicated to investing in the development of a successful, sustainable, and technologically advanced Ukraine through innovative projects. The foundation’s founder, Serhii Tokarev , is a Ukrainian impact investor, tech entrepreneur, founder of SET University, and the driving force behind AI HOUSE.


CLICK HERE TO DONATE IN SUPPORT OF DCREPORT’S NONPROFIT MISSION

The post Investing in the Future: The Serhii Tokarev Foundation and uBoost Sum Up SheLeads appeared first on DCReport.org.

Redefining the Patient Experience in Specialty Healthcare

Specialty healthcare often begins after a diagnosis has already disrupted sleep, work, mobility, or family routines. Patients may undergo infusions, lab monitoring, infection precautions, and insurance review, and may experience side effects that require quick judgment. Strong care begins with accurate prescribing, then continues through teaching, timely follow-up, and respectful coordination. Each contact should reduce confusion. When support feels steady, people can use more strength for recovery.

Coordinated Care Matters

Specialty infusion brings together prescribers, nurses, pharmacists, benefits teams, payors, and manufacturers around a single treatment plan. For people managing immune, neurologic, inflammatory, or bleeding disorders, Acelpa  belongs in a care model where medication handling, authorization, nursing oversight, outcome reporting, and family education must align before therapy feels dependable.

Clear Communication

Patients rarely absorb every instruction during a stressful visit, especially with pain, fatigue, or a new diagnosis. Teaching should use clear words, written steps, and teach-back checks. Nurses can review dose timing, storage temperature, possible reactions, and warning signs before leaving the home. A call within several days can catch uncertainty early, prevent skipped doses, and protect trust.

Access Without Friction

Access means more than geographic distance from a clinic. It includes referral intake, benefit verification, medicine shipment, nurse scheduling, and quick answers when symptoms change. One missed handoff can postpone the first dose. Specialty teams should monitor referral age, authorization status, delivery accuracy, and start timing. These measures reveal where people wait, worry, or repeat personal details.

Personal Support

Personal support should be clinical and practical. A patient with a weak grip may struggle to open supplies. Another person may need translation, caregiver teaching, or refrigerator reminders. Home layout, mobility limits, infection risk, allergy history, and previous reactions all matter. Strong teams record these details once, then apply them during every call, delivery, and nursing visit.

Home-Based Confidence

Home infusion can preserve comfort, privacy, and routine, yet preparation must be exact. Nurses need enough time to review vascular access, clean technique, infusion steps, and emergency symptoms. Families benefit from knowing which concerns require urgent attention. Medication arrival, supply counts, and visit timing should match the prescribed schedule. Predictability lowers anxiety before treatment starts.

Better Reporting

Data has value only when it drives action. Specialty providers can monitor adherence, adverse reactions, start dates, refill timing, emergency visits, and therapy response. Physicians need clear reports that separate routine updates from clinical risk. Payors need evidence that services reduce avoidable utilization. When a missed refill is identified, outreach should occur quickly, with notes shared across the care team.

Insurance Guidance

Coverage problems can feel as heavy as physical symptoms. Prior authorization, benefit limits, site-of-care rules, and copay questions often appear before treatment begins. Administrative staff should explain the status in direct language. Patients need to know who owns the task, which record is missing, and when an update should arrive. Clear financial guidance protects treatment momentum and personal dignity.

Staff Training

Every role shapes the care experience. Intake specialists, pharmacists, nurses, drivers, and billing staff influence whether patients feel heard. Training should cover medication safety , privacy, escalation rules, empathy, and precise documentation. Leaders can review service recovery cases, call notes, and satisfaction comments to find weak points. Consistent coaching turns good intentions into reliable practice across the organization.

Measuring What Works

Patient surveys help, but they cannot tell the full story alone. Specialty programs should pair feedback with clinical and operational measures. Beneficial indicators include time to therapy, education completion, on-time delivery, adverse event follow-up, readmission rates, and treatment persistence. Regular review helps leaders adjust staffing, refine workflows, and direct attention where delays place health at risk.

Conclusion

Redefining specialty healthcare means treating the experience around treatment as part of treatment itself. Patients need accurate medication, skilled nursing, responsive insurance support, and instructions they can follow under stress. Families need steady communication and a home plan that feels reliable. When clinical data, compassionate service, and disciplined coordination work together, patients gain confidence, and care teams can pursue better outcomes.

Photo: Ivan S via Pexels


CLICK HERE TO DONATE IN SUPPORT OF DCREPORT’S NONPROFIT MISSION

The post Redefining the Patient Experience in Specialty Healthcare appeared first on DCReport.org.

Extreme Heat and Tropical Storm Bertha Impacts


Central North Pacific 2-Day Graphical Outlook Image
Central North Pacific 7-Day Graphical Outlook Image