ChinaTalk

Jordan Schneider

Conversations exploring China, technology, and US-China relations. Guests include a wide range of analysts, policymakers, and academics. Hosted by Jordan Schneider. Check out the newsletter at https://www.chinatalk.media/

  1. 1d ago

    $75k Contest Launch! ChinaTalk Hiring + Evals for the Situation Room

    Contest links: General $50k https://www.chinatalk.media/p/50k-chinatalk-submission-hiring-contest we're looking at submissions on a rolling basis! AI evals: Due Over the past few years, we’ve seen hints of policymakers and national leaders using AI models in their actual policy decision-making. The Prime Minister of Sweden said he uses it for second opinions on policy; the German Chancellor is testing it to draft legislation; even Trump said he had it re-write a speech at some point. It’s fair to assume that senior leadership across the world, and in Washington as well, have started using AI not just for tactical or operational tasks, but increasingly for broad strategic decision-making. While that’s exciting — there’s the promise of uplift and smarter calls on some of the most consequential decisions leaders face in foreign policy and national security — we’re also flying blind. Enormous effort and energy goes into benchmarking and evaluation for things like coding. The experiments you can run to make models better at software development are much easier to execute and much lower-stakes than running a real experiment when you’re deciding whether to invade a country or sign a treaty. That’s why we at ChinaTalk are trying to kickstart a field aimed at helping researchers and policymakers understand exactly what they’re working with when they ask these models to support some of the most consequential decisions they may make in their lifetimes. We’re launching an evals/essay project contest to explore this theme. I’ve brought on two expert AI eval creators to discuss why the field is important, what interesting work has already been done on how models approach broad national-security and strategic questions, and how you — as an eval professional, semi-professional, or just a concerned person — can contribute new ways to poke and prod at these models and see what they can really do. Joining us today: Florian Brand, research engineer at Prime Intellect, and John Chen, professor at the University of Arizona, who’s done some pretty wild things getting models to start nuclear wars with each other in Civ V. Our conversation covers: How frontier labs are hitting eval-building limits — they’ve gone from undergrads to PhDs to field experts, and the models are now catching the experts’ mistakes. How Civilization V exposes AI’s strategic blind spots (terrible second-order reasoning) and models’ distinct strategic personalities (Claude’s really into science!). Why ethical prompting in Civilization V still doesn’t stop models from launching nukes. Jordan's PresidentBench eval, where a Chinese model was nonchalant about a Taiwan invasion, while Claude wanted to keep Taiwan free and independent. Advice for designing better AI evals and ChinaTalk’s new essay/evals contest! Learn more about your ad choices. Visit megaphone.fm/adchoices

  2. Aug 1

    Jordan talks AI with Sebastian Mallaby on his show 'The Spillover'

    Don't sorry I put two ChinaTalk Records songs at the end! their show notes: On this episode of The Spillover, Sebastian Mallaby sits down with Jordan Schneider of ChinaTalk to unpack a frantic month in artificial intelligence. The conversation turns on whether the U.S. can actually “win” the AI race, with Schneider arguing that the most durable competitive edge will come from compute rather than model quality. The two debate China’s open-weight model strategy as a commercial weapon and question whether it will be possible to sufficiently harden systems against threats like cyberattacks and AI-designed bioweapons. Anthropic’s Mythos model has led the Trump administration to take an AI-safety U-turn. Mallaby notes that as late as March 2026, “if you had said to the Trump administration that they would be trying to suppress an American AI model, decelerate the progress in the name of safety, they would have said, ‘You’re nuts.’” Schneider notes that a similar shift has not yet arrived in China.  AI models are now escaping their sandbox and raising false alarms about foreign hackers. In remarking on OpenAI’s models hacking AI company Hugging Face, Schneider comments, “They think it’s the Chinese. . . . And they’re freaking out. They’re calling the FBI.” For Schneider, the episode is a warning that “we’ve really crossed a threshold with these models,” which now have “the potential to do really dramatic harm just on their own, because we like can’t even physically watch them.” The real AI scoreboard is compute, not model quality. Schneider argues that there is too much of a focus on the gap between Chinese and Western frontier models. Mallaby adds, “In other words, I shouldn’t be asking about how far is China behind in terms of the quality of model. It’s more a question of like, how much can China deliver the AI to users within China, given their lack of computational resources?” China’s open-weight models are a weapon without a business model. Mallaby frames Chinese open-weight as a “counterweapon”—good-enough models pushed out cheaply to erode the economics of U.S. frontier labs. Schneider notes that DeepSeek’s CTO is pitching investors a Manhattan Project-like vision in which profits should take a back seat to the pursuit of AGI. Learn more about your ad choices. Visit megaphone.fm/adchoices

  3. Jul 26

    The Iran War Oil Shock That Wasn’t...(Yet?)

    The biggest oil shock in modern history came and went without the catastrophe everyone expected. When Iran closed the Strait of Hormuz, analysts warned that oil could hit $200 a barrel, but the global economy avoided that fate. The Trump administration has argued that the crisis was contained thanks to their aggressive action, but they may be taking the wrong lessons from the avoidance of that apocalyptic scenario. What happened was the largest unexpected swing in global oil balances: China quietly cut crude oil imports by more than five million barrels a day. Yet there was no corresponding collapse in economic activity, no obvious drop in mobility, and no official explanation from Beijing. Somehow, China stopped buying oil from the rest of the world and started drawing from stockpiles we can’t fully observe. That single decision may have done more to prevent a global energy crisis than anything Washington or OPEC accomplished. This matters because it demonstrates that China likely has a stronger discretionary policy lever than the West does. The West is really good at market-driven, private-sector oil production. But through this crisis, we’ve seen that Washington does not have the scale of discretionary policy control that China or OPEC does. This time, China cooperated and did the good thing, at least for the broad economic picture — but we cannot rely on that in the future, and that tool can be used against the West as easily as for it. Western governments must grapple with that discretionary gap and not rest on their private-sector bona fides to get through the next crisis. Arnab Datta, managing director of policy implementation at Employ America, and Rory Johnston, oil analyst and founder of Commodity Context, join ChinaTalk to discuss: Why Rory’s own prediction of $200 oil never happened and why J.D. Vance is thanking the wrong people. How China quietly cut crude imports by five million barrels a day with zero visible impact on domestic mobility, and the detective work analysts use to peer into Beijing’s black-box inventories. Competing theories for why Beijing backstopped the global oil market — self-interested altruism, a backroom deal during the state visit, or a dry run for a Malacca blockade in the event of a Taiwan contingency. What India, the Gulf states, and the rest of the world learned from the Iran War and why strategic reserves are suddenly back in fashion. Learn more about your ad choices. Visit megaphone.fm/adchoices

4.3
out of 5
24 Ratings

About

Conversations exploring China, technology, and US-China relations. Guests include a wide range of analysts, policymakers, and academics. Hosted by Jordan Schneider. Check out the newsletter at https://www.chinatalk.media/

You Might Also Like