Authors
Gary Mukai
News Type
Blogs
Date
Paragraphs

The National Consortium for Teaching about Asia (NCTA), funded by the Freeman Foundation, is a multi-year initiative to encourage and facilitate teaching and learning about East Asia in elementary and secondary schools nationwide. NCTA is a premier provider of professional development on East Asia.

From July 6 to 9, 2026, SPICE Teacher Professional Development Manager Naomi Funahashi and Event Coordinator Sabrina Ishimatsu hosted a summer institute for middle school teachers, most of whom are teaching in California. Fifteen teachers enrolled. SPICE was so grateful for their active participation and for sharing their reflections and insights on subject matter content and pedagogical content knowledge.

The first three days featured subject matter content lectures on topics commonly taught at the middle school level in California. These included lectures on “Chinese Dynasties” and “The Silk Road” by Dr. Clayton Dube, Senior Fellow at the USC U.S.-China Institute, and “Tokugawa Japan” by Professor Ethan Segal, Michigan State University, who received his PhD from Stanford. SPICE is a partner site of the USC U.S.-China Institute, one of seven National Coordinating Sites of the NCTA. Dr. Clayton Dube is Director Emeritus of the USC U.S.-China Institute, and SPICE now works with Glenn Osaki, Director, and Crystal Hsia, Program Specialist, of the USC U.S.-China Institute. 

I have been very inspired by all the interactive meetings. Each truly serves as a model for creating engaging lessons for our students… There were so many interesting and relevant topics such as the different dynasties, the role and expectations of women, and cultural exchange!—Elena Giron, Elizabeth Learning Center, Cudahy, California

 

As a K–8 Mandarin teacher, I’m always looking for meaningful ways to help my students understand how Chinese language and culture have influenced other parts of East Asia. After today’s session, I began thinking about how I might incorporate elements of Japanese history, culture, or art into my Mandarin curriculum while still keeping Chinese language learning at the center.—Shu-Chen Lin, San Domenico School, San Anselmo, California


The first three days also featured demonstrations of SPICE curricula that focused on pedagogical content knowledge. These included demonstrations of Japanese Art in the Edo Period, by Karen Tiegel, Middle School Division Head at The Nueva School and a former SPICE Curriculum Writer; Along the Silk Road, by Rylan Sekiguchi, SPICE Manager of Curriculum and Instructional Design; and Angel Island, Chinese American Voices, the Chinese Railroad Workers in North America Project, and What Does It Mean to Be an American?, by Jonas Edman, SPICE Instructional Designer. 

I will definitely incorporate the SPICE Silk Road curriculum into my teaching. I especially appreciate the interactive nature of its online resources, which make the content more engaging and accessible for students. The curriculum provides meaningful opportunities for students to actively explore the history, geography, and cultural exchanges of the Silk Road rather than simply learning through lectures or textbooks. I can see these resources helping my middle school students build a deeper understanding of the topic while keeping them motivated and engaged.—Jenn Wu, Martin Luther King Jr. Middle School, San Francisco, California

 

Image
webinar announcement flyer with photos of three authors and pictures of three book covers


On day four, three authors of young adult novels (from top to bottom in image above)—Van Hoang, Judy I. Lin, Waka T. Brown—were featured on a panel titled, “Incorporating Asian American Literature and Voices into Middle School Curricula.” Including the Asian American experience in SPICE/NCTA seminars has been requested by many teachers and also supports the NCTA as it “strives to develop global citizenship based on the principles of cross-cultural understanding, equality, and justice.” 

I loved being part of the summer cohort with SPICE… I was really inspired by the authors’ panel, as I had never heard from authors, especially ones who are so relatable. I love how they intertwine their identity and history and can create such cool stories.—Morgan Johnson, Kenmore School, Arlington, Virginia


Following the institute, each teacher will develop and share an original lesson plan inspired by the knowledge and resources gained throughout the seminar. Some of the themes and concepts that were shared by the lecturers, authors, teachers, and SPICE staff included diaspora, cultural relevance, the master narrative of history, diverse perspectives, and empathy; and SPICE hopes that one or more of these themes will be incorporated into their lessons. 

SPICE is especially grateful to President Graeme Freeman and Vice President Shereen Goto of the Freeman Foundation for their support of the NCTA since 1998 and to Crystal Hsia of the USC U.S.-China Institute for her unwavering support and collaboration. 

The East Asia Summer Institute for Middle School Teachers is one of several free teacher professional development opportunities offered by SPICE.

To stay updated on SPICE news, join our email list and follow us on Facebook, X, and Instagram.

Read More

screenshot of Zoom meeting with 14 participants
Blogs

2024 SPICE/NCTA Summer Institute Engages Educators in East Asian and Asian American Studies

Middle school teachers participate in summer institute on East Asia.
2024 SPICE/NCTA Summer Institute Engages Educators in East Asian and Asian American Studies
three people standing at the great wall in China
Blogs

Celebrating SPICE’s 50th: SPICE’s Roots in the Bay Area China Education Project (BAYCEP)

BAYCEP was the predecessor program to SPICE, which was established 50 years ago in 1976.
Celebrating SPICE’s 50th: SPICE’s Roots in the Bay Area China Education Project (BAYCEP)
group of people standing outside
Blogs

The 2025 Stanford/SPICE East Asia Seminars for Teachers in Hawai‘i Summer Institute

The Stanford/Freeman SEAS Hawai‘i Fellows gathered at the East-West Center, from July 12 to 14, 2025.
The 2025 Stanford/SPICE East Asia Seminars for Teachers in Hawai‘i Summer Institute
Hero Image
Taklamakan desert, camels
Taklamakan Desert | Photo courtesy of Professor Albert Dien
All News button
0
Subtitle

The Summer Institute was made possible by the Freeman Foundation.

Image
Taklamakan desert, camels
Caption Taklamakan Desert | Photo courtesy of Professor Albert Dien
Date Label
Display Hero Image Wide (1320px)
No
Authors
Melissa Morgan
News Type
News
Date
Paragraphs

The Freeman Spogli Institute for International Studies (FSI) is pleased to announce that James Fearon will serve as the next co-director of the institute’s Center for International Security and Cooperation (CISAC), effective September 16, 2026.

Fearon is currently the Theodore and Frances Geballe Professor in the School of Humanities and Sciences, a professor of political science, and a senior fellow at FSI, where he is affiliated with both CISAC and the Center on Democracy, Development and the Rule of Law. In 2021 and 2022, he served as a senior advisor in the U.S. Department of Defense, working primarily on the 2022 National Defense Strategy and its implementation within the Department.

He is the author of several highly influential studies of the causes of interstate and civil war, signaling and bargaining in interstate disputes, and deterrence theory. Earlier this year he published the book Worse than War: The Global Costs of Violence (Princeton University Press), co-authored with Anke Hoeffler, that estimates and compares the prevalence and costs of war with the prevalence and costs of interpersonal violence. His current work includes research on how U.S. and Chinese defense postures affect risks of escalation in the western Pacific; AI and the stability of nuclear deterrence; and the organizational design of defense bureaucracies. He was elected to the American Academy of Arts and Sciences in 2002, and the National Academy of Sciences in 2012. He received his PhD in Political Science from the University of California, Berkeley and joined Stanford as an associate professor in 1998.

“Jim Fearon will be a superb co-director of CISAC,” said outgoing co-director Scott Sagan. “He is one of the nation's finest scholars on bargaining in crises, the causes of interstate and civil wars, and global trends in forms of violence. He combines this scholarly expertise with keen policy interests in Pentagon budgeting, grand strategy, and nuclear weapons. What a great mix of interests and expertise to address international security in a changing world.”

Scott Sagan (right) sits next to former Secretary of Defense William Perry (left)
Scott Sagan with former U.S. Secretary of Defense William J. Perry, another longtime member of the CISAC community. | Light11B

Sagan has served as the co-director of the Center for International Security and Cooperation (CISAC) from 2021 to 2026, though his legacy and influence at CISAC extends well beyond the last five years. He previously served as CISAC co-director from 1998 to 2012. In 2021, he returned for another term of leadership service alongside Rodney Ewing, a geologist whose multidisciplinary research led to the development of techniques to predict the long-term behavior of materials used in radioactive waste disposal. Sagan and Ewing led the center together until Ewing’s passing in 2024.

CISAC’s co-director leadership structure has been in place since its founding in 1983, and reflects the belief that it takes scholars from different disciplines with different experiences, ideas, strengths, and interests to solve the most pressing security problems.

“We’re in a new global era where security, technology, geopolitics, and policy are more interconnected than ever before,” said Colin Kahl, director of the Freeman Spogli Institute for International Studies. “CISAC is crucial to understanding that intersectionality and producing the knowledge we need to create a safer and more secure world. Scott has led this center with incredible vision, and I am fully confident that Jim, with his unique talents and expertise, will lead CISAC and its community into its next era of growth and world class research.”

CISAC’s mission is to produce policy-relevant research on international security problems, teach and train the next generation of security specialists, and inform policymaking in international security.
 


Things are changing fast in international security affairs. But CISAC has long been the place for high-quality research at the intersection of science, technology, and the politics of international security.
James Fearon
Co-director of CISAC


Under Ewing’s and Sagan’s leadership, CISAC has expanded to include new programs and initiatives such as the Indo-Pacific Policy Lab, led by Oriana Skylar Mastro, whose mission is to disseminate leading academic knowledge on Indo-Pacific security issues to policymakers and build connections between scholars and industry that promote innovation and strengthen regional security and prosperity. Similarly, the new India-U.S. Security Studies Fellowship allows engineers, scientists, historians, and social scientists the opportunity to focus on issues related to Indian and U.S. security issues in collaboration with the CISAC community.

Other initiatives, such as CISAC’s longstanding program on Biosecurity and Global Health, the program on Geopolitics, Technology, and Governance, the initiative on Existential Risks, and their evolving agenda on Cyber Policy and Security, continue to have relevance and reach as global threats and politics evolve.

“It's not clear how the U.S. and global partners can best deal with the complex security dilemmas emerging on the horizon,” said Sagan. “But that’s the role of independent thinking inside a university. We’re independent from bureaucratic inertia. We’re nonpartisan. We’re free to figure out what policies and strategies will lead to a more secure and peaceful world. That’s the goal.”

CISAC’s fellowship and teaching programs are instrumental in training the next generation of security studies specialists.

“CISAC showed me how to be a rigorous and curious academic committed to building community,” said Luis Rodriguez, a former CISAC fellow now teaching at George Mason University.

Ethan Lee (BS, ‘23) feels similarly. Currently an associate in the Defense, Emerging Technology, and Strategy Program at Harvard’s Belfer Center for International Affairs, Lee is an alum of CISAC’s undergraduate honors program, which was established in 2000 as a way of providing Stanford seniors from all disciplines the opportunity to conduct rigorous, scholarly research on international security issues.

“As a student, I felt both challenged and supported at CISAC,” recalled Lee. After graduation, Lee and Sagan co-authored an article on U.S. public opinion, law, and the use of military force against Iran, which is forthcoming in the journal Security Studies. “I have benefited tremendously from Scott’s generosity, and I know I'm just one of many students who have.”

Students from the 2014-15 cohort of the CISAC Undergraduate Honors Program stand around the logo of the Central Intelligence Agency in the atrium of CIA headquarters in Langley, Virginia.
Members of the 2014-2015 CISAC Honors Students cohort on a visit to the CIA to discuss their thesis topics with analysts. | Darren L. for the CIA

Alumni of CISAC’s undergraduate honors program and fellowships go on from Stanford to create a global network of policymakers.

“When I take students to the Pentagon or to think tanks in Washington, you meet our former fellows and honors students,” says Sagan. “In major universities across the country, the political science departments and the international security research centers are populated with scholars who came through Stanford. That's a legacy about which all the directors at CISAC should be very proud.”

Looking to the future, incoming co-director James Fearon acknowledges that while there are very real challenges ahead in international relations, there are also increased opportunities for scholars at places like CISAC to make an impact for good.

“Things are changing frighteningly fast in international security affairs, in alarming ways that are often linked with effects of rapid technological change,” Fearon acknowledges. “But CISAC has long been the place for high-quality research at the intersection of science, technology, and the politics of international security. I am excited by the opportunity and am greatly looking forward to working with a fantastic set of scholars building on this tradition in a new international context.”

To learn more about Fearon’s research and how geopolitics and technology are intersecting across borders, listen to the latest episode of the World Class podcast, hosted by FSI Director Colin Kahl.

Fearon’s co-director will be announced in November 2026.

Read More

Pei-Chia Lan, Stanford FSI senior fellow
News

Pei-Chia Lan to Join FSI as the Wang and Chen Senior Fellow in Taiwan Studies

Pei-Chia’s research explores the complexity of intersecting social inequalities in everyday life, shaped by macropower dynamics such as globalization and international migration.
Pei-Chia Lan to Join FSI as the Wang and Chen Senior Fellow in Taiwan Studies
Sasha Baker and Tarun Chhabra on the World Class podcast
Commentary

Anthropic, OpenAI, and the New Frontier of Artificial Intelligence in National Security

The heads of national security policy at OpenAI and Anthropic join Colin Kahl on the World Class podcast to discuss how AI is changing national security strategies and the nature of U.S.-China competition.
Anthropic, OpenAI, and the New Frontier of Artificial Intelligence in National Security
Michael McFaul speaks into a microphone at a podium
News

Michael McFaul Receives the 2026 Hubert H. Humphrey Award

The award is presented annually by the American Political Science Association (APSA) to honor notable public service by a political scientist.
Michael McFaul Receives the 2026 Hubert H. Humphrey Award
Hero Image
Headshot of James D. Fearon
FSI Senior Fellow James Fearon will serve as CISAC's social science co-director. | Rod Searcey
All News button
1
Subtitle

Fearon, a senior fellow at the Freeman Spogli Institute, will succeed Scott Sagan, who has served in the role since 2021.

Date Label
Display Hero Image Wide (1320px)
No
Paragraphs
Banner for a Shorenstein APARC working paper, showing a quill, inkwell, and old ledger book

 

This paper considers whether financialization is a necessary stage of financial modernization. It argues that comparisons of financial systems should not stop at the structural distinction between bank finance and market finance, or between indirect and direct finance. Instead, they should examine how credit is identified, recognized, constrained, and circulated, and how credit risk is allocated. A bank-centered system produces credit through relational trust, organizational review, continuous monitoring, and the internalization of risk. A financialized system expands credit circulation through standardization, assetization, securitization, and market transactions. Evidence from Germany, Japan, China, and the United States suggests that bank-centered finance and financialization are not lower and higher stages of the same developmental path, but two different modes of credit organization. The key issue in financial modernization is not a choice between banks and markets, but how institutional boundaries are reorganized between credit embeddedness and credit mobility.  

All Publications button
1
Publication Type
Working Papers
Publication Date
Subtitle

From Relational Trust to Credit Transactionalization

Authors
Paragraphs

How do portrayals of foreign nations as threats shape public opinion on foreign policy and domestic racial attitudes? This study examines how portrayals of China as a threat influence Americans’ support for U.S. policy toward China and attitudes toward Asian Americans. A content analysis of CNN and Fox News transcripts from 2010 to 2020 shows that both outlets increasingly portrayed China as a threat, although Fox News more often emphasized threats to the United States, whereas CNN focused more on threats affecting other countries, international institutions, and populations. A national survey and a preregistered survey experiment further demonstrate that exposure to China threat narratives, regardless of whether the threat targets the United States or another country, increases support for hawkish U.S. foreign policy and heightens anti-Asian resentment, especially among Republicans. These findings show how portrayals of foreign threats shape public opinion and spill over into racial attitudes at home.

All Publications button
1
Publication Type
Journal Articles
Publication Date
Subtitle

Media Influence on U.S.-China Policy and Anti-Asian Sentiment

Journal Publisher
Political Communication
Authors

Stanford University
Encina Hall, Room E301
Stanford, CA 94305-6055

650-723-9741
0
Wang and Chen Senior Fellow in Taiwan Studies at the Freeman Spogli Institute for International Studies
pei-chia_lan.jpeg PhD

Pei-Chia Lan is the Wang and Chen Senior Fellow in Taiwan Studies at the Freeman Spogli Institute for International Studies and the director of the Taiwan Program at the Walter H. Shorenstein Asia-Pacific Research Center. As a qualitative sociologist, she studies the complexity of intersecting social inequalities in everyday life, shaped by macropower dynamics such as globalization and international migration. 

Before joining Stanford in September 2026, Lan was a distinguished professor in the Department of Sociology at National Taiwan University, where she was the founding director of the Global Asia Research Center and recipient of five teaching awards. 

Her scholarly articles have been published in Ethnic and Racial Studies, International Migration Review, The Sociological Review, and Comparative Migration Studies, among others. She is also the author of several books in Chinese and English, the latter including Raising Global Families: Parenting, Immigration and Class in Taiwan and the US (Stanford University Press, 2018), which examines how ethnic Chinese parents in Taiwan and the United States negotiate cultural differences and class inequality to raise children in the contexts of globalization and immigration; and Global Cinderellas: Migrant Domestics and Newly Rich Employers in Taiwan (Duke University Press, 2006), which won a Distinguished Book Award from the Sex and Gender Section of the American Sociological Association and ICAS Book Prize: Best Study in Social Science from the International Convention of Asian Scholars.

Lan received her doctorate in sociology from Northwestern University, and her bachelor’s and master’s degrees in sociology from National Taiwan University. She was a 2024-2025 Stanford-Taiwan Social Science Fellow at the Center for Advanced Study in the Behavioral Sciences, Stanford University; a 2011-2012 Yenching-Radcliffe fellow at Harvard University; a 2026-2007 Fulbright scholar at New York University; and a 2000-2001 postdoctoral fellow at the Center for Working Families, University of California, Berkeley. She also held visiting positions at the Waseda Institute for Advanced Study, Kyoto University, Tubingen University, and IIAS at Leiden University.

She has actively engaged in international professional networks, serving as deputy editor of the journal Gender & Society and an editor or board member for other esteemed journals. Her research cultivates critical dialogue between academia and the public, and her expertise is sought by local government agencies and global organizations alike, as demonstrated by her role on the advisory board of UN Women.

Director, Taiwan Program at the Walter H. Shorenstein Asia-Pacific Research Center
CV
Date Label
Authors
News Type
Blogs
Date
Paragraphs

The following reflection is a guest post written by Sunny Park, a Fall 2020 alumna of Stanford e-Entrepreneurship Japan, which is currently accepting applications for Fall 2026.

A “smart” high school student is often defined by strong grades, perhaps alongside a few extracurricular activities to strengthen university applications. Before encountering Stanford e-Entrepreneurship Japan (SeEJ), that was largely how I viewed myself. 

I came across SeEJ during the COVID-19 pandemic through a Facebook advertisement. Like many students at the time, I was looking for something more engaging beyond school. I didn’t expect how much it would reshape the way I think about learning, ambition, and possibility. 

What made SeEJ distinct was its focus on people rather than definitions. Each session featured entrepreneurs from different industries, sharing not just what they built, but how they navigated uncertainty, failure, and personal ambition. Seeing such varied paths made me realize that entrepreneurship isn’t a fixed route; it’s a way of thinking. And that was probably the moment when things shifted for me. Until then, I had always thought of my future in fairly narrow terms: choose a subject, go to university, and follow a path that already exists. SeEJ made me realize that there are far more ways to build a life, and that entrepreneurship is less about starting a company and more about taking ownership of what you want to pursue. 

Since then, I’ve found myself saying yes to things I would have hesitated to try before. I’m now in my final year of studying Biology at Imperial College London, and I co-founded a biotech startup called Eidolon Therapeutics. At Eidolon, we’re pioneering microbial protein combinations to combat immunotherapy resistance and inflammation-driven toxicity in cancer. It was through my Year-in-Research program where I met my Co-Founders (i.e., my professor and post-doc supervisor), and I was so fascinated by the research they were conducting I couldn’t help but impatiently suggest building a company out of it. Since then, we’ve been working incredibly hard on making Eidolon real, including joining accelerator programs such as the Cancer Tech Accelerator, Venture Catalyst Challenge, and Conception X. Through building an early-stage startup, I’ve had the chance to meet people I once found intimidating, pitch ideas in front of large audiences, and build relationships that started from simple conversations. 

I don’t think any of that came from a single moment of confidence. If anything, it came from gradually becoming more comfortable with not knowing exactly what I’m doing… and trying it anyway. Based on my journey so far, here are a few reflections for students exploring a similar path:

  1. Be intentional about the opportunities you choose.

There are so many opportunities available as a student! SeEJ, in particular, showed me the value of learning environments where you actively engage with ideas and people. There are many programs available, but depth matters more than quantity. Choose opportunities where you can contribute, ask questions, and grow. 

  1. Focus on genuine connections!

One of the most valuable parts of SeEJ was the access to people with diverse experiences. What I learned is that meaningful connections come from curiosity and sincerity, not from trying to impress. When you approach people with a willingness to learn, they are often more open than you expect. 

  1. Stay engaged beyond the program.

The real impact of SeEJ extended beyond the sessions themselves. Following up with people, continuing conversations, and sharing your own progress helps turn short-term interactions into long-term relationships. Small, consistent effort makes a difference!

  1. Reflect as you go. 

Exploring entrepreneurship alongside academics can be exciting but also overwhelming. Taking time to reflect on what you enjoy, what challenges you, and what kind of path you want to build helps you stay grounded and intentional.

Looking back, SeEJ was a turning point not because it provided a single opportunity, but because it changed how I see opportunities altogether. It reminded me that there isn’t one path to follow, but many that you can create. 

Stanford e-Entrepreneurship Japan is currently accepting applications for Fall 2026. Apply at https://forms.gle/5gkHrhm3X4sycZGG8.

Stanford e-Entrepreneurship Japan is one of several online courses offered by SPICE.

To stay updated on SPICE news, join our email list and follow us on Facebook, X, and Instagram.

Read More

a person standing in front of Tanah Lot
Blogs

Stanford e-Entrepreneurship Japan: Empowering Young Visionaries to Reimagine Global Challenges for Social Good

High school student Erin Tsutsui, an alumna of Stanford e-Entrepreneurship Japan, reflects on forging friendships across Japan, embracing new world perspectives through thoughtful discussion, and transforming family heritage into a youth-led peace initiative via empathy and social innovation.
Stanford e-Entrepreneurship Japan: Empowering Young Visionaries to Reimagine Global Challenges for Social Good
a group of students standing with signs, "TBC Japan"
Blogs

Let’s Be the Strikers: Thoughts on the 2025 Teenage Business Contest Japan

Millie Gan, an alum of Stanford e-Entrepreneurship Japan and founder of Teenage Business Contest Japan (TBCJ), reflects on building a platform that empowers teens to use entrepreneurship and innovation to revitalize Japan’s communities.
Let’s Be the Strikers: Thoughts on the 2025 Teenage Business Contest Japan
group of people posing in front of a screen
Blogs

Five Years of Impact: Celebrating the Stanford e-Entrepreneurship Japan Program

Alumni from across Japan gather in Tokyo to celebrate SeEJ’s milestone anniversary.
Five Years of Impact: Celebrating the Stanford e-Entrepreneurship Japan Program
Hero Image
a person giving a presentation in a classroom
Sunny Park at the Venture Catalyst Challenge at Imperial College London, UK. | Photo courtesy of Sunny Park
All News button
1
Subtitle

Sunny Park, an alum of Stanford e-Entrepreneurship Japan and cofounder of Eidolon Therapeutics, reflects on adopting an entrepreneur’s mindset, embracing opportunity, and creating one’s own path in life.

Date Label
Display Hero Image Wide (1320px)
No
Authors
Melissa Morgan
News Type
Commentary
Date
Paragraphs

A breach of Hugging Face's servers. AI safety debates on Capitol Hill. New models like Kimi K3 coming online. There's been no shortage of headline grabbing AI news and debate in the last few weeks. To talk through these topics and how they relate to American national security, the heads of national security policy at OpenAI and Anthropic joined Colin Kahl on the World Class podcast.

Together, Sasha Baker (OpenAI) and Tarun Chabra (Anthropic) discuss how artificial intelligence is intersecting with security and defense strategies, the U.S.-China AI race, and how the risks and advantages of AI are starting to reshape national security.

Sasha Baker is the head of national security policy at OpenAI. Prior to that, she served as the acting under-secretary for policy and the deputy under-secretary of defense for policy at the Pentagon, and as the senior director for strategic planning on President Biden's National Security Council staff.

Tarun Chhabra is the head of national security policy at Anthropic. He previously served as deputy assistant to the president and coordinator for technology and national security on the National Security Council staff, where he coordinated the Biden administration's strategies for technology competition with China and technology partnerships with U.S. allies and partners. 

This episode's reading/watching recommendations are AI 2027, a project by Daniel Kokotajlo; AlphaGo, a documentary by director Greg Kohs; and "Nineteenth-Century Horse Sense" by Francis Thompson.

TRANSCRIPT:


Kahl: You're listening to World Class from the Freeman Spogli Institute for International Studies at Stanford University. I'm your host, Colin Kahl, the director of FSI.

I'm very excited to welcome my good friends and former colleagues Sasha Baker from OpenAI and Tarun Chhabra from Anthropic for what promises to be an insightful conversation on the good, the bad, and the ugly of how artificial intelligence is intersecting with national security. We're going to discuss the U.S.-China AI race, the risks AI poses to security, and the opportunities and advantages AI might generate in the national security space.

These are fantastic guests to help us grapple with these complex topics. Sasha Baker is the head of national security policy at OpenAI. Prior to that, she served as the acting under-secretary for policy and the deputy under-secretary of defense for policy at the Pentagon, and as the Senior Director for Strategic Planning on President Biden's National Security Council staff.

Tarun Chhabra is the head of national security policy at Anthropic. He previously served as deputy assistant to the president and coordinator for technology and national security on the National Security Council staff, where he coordinated the Biden administration's strategies for technology competition with China and technology partnerships with U.S. allies and partners. 

Sasha, Tarun, thanks for coming on to World Class. 

So look, I just told everybody your job titles; they’re very fancy. People know your companies. But perhaps we could start with describing what your roles at OpenAI and Anthropic actually entail.

Sasha, maybe let's start with you. What do you actually do on a daily basis?

Baker: Well first of all Colin, thanks for having me. It's fun to be here with two former colleagues. And what you didn't mention in my bio is that in my deputy role I was the chief “Colin Minder” of the Pentagon for a number of months.

What do I do? It's the most interesting job maybe I’ve ever had. I get to talk to governments all around the world every day about basically two things. The first is: what is the opportunity space look like as it relates to AI being used in the national security domain? So, how can we help governments that want to use these tools to make their populations safer? How do we help them do that?

And then the second thing that I talk to governments about is the AI risk space and the ways in which AI could up-level or enable a bad actor to do something that we really wouldn't want to see. So, I spent most of my career in government, as I think Tarun did as well. And what's fun about this job is it allows me to still be part of the same kinds of conversations that you and I would have had when we were serving at the Pentagon, but just from a totally different vantage point.

Kahl: Tarun, how about you?

Chhabra: Let me add my thanks, Colin, for having me too, and it's great to be here with Sasha and with you.

Obviously a lot of similarities in my role as well. I think of it as trying to help prepare policymakers for what we see as coming down the pike in AI development: to help them prepare and whether that's folks who are doing reporting and analysis, or whether that's folks who are making policy, or folks in Congress. We want them to know what's coming so that they're ready.

And then the second part, of course, is making sure we can get the best possible technology into their hands for national security purposes. As Sasha said, that's first and foremost with the U.S. government, but it's also, of course, with our closest allies as well.

Kahl: Great. Let's jump right into the types of conversations you're having. I'm sure as you're talking to officials around the world, there's a lot of focus on the U.S.-China AI competition. I think as many of our listeners will know, in recent weeks, Chinese AI labs have released some very impressive models, including ZAI's GLM 5.2 a few weeks back, and most recently Moonshot's Kimi K3. Obviously, people are familiar with very good models released by other Chinese labs like Deep Seek.

Tarun, maybe starting with you on this one: how would you assess the current gap between the best models coming out of Chinese AI labs and the frontier AI models coming out of labs like Anthropic and OpenAI? Is the U.S. ahead? If so, by how much? How would you assess the race right now?

Chhabra: I think our view's been pretty consistent even as new Chinese models have come out, which is, we believe that we remain six to nine months ahead of Chinese models at the frontier. And we think that really matters. So if you think about having really advanced cyber capabilities, having six to nine months with those superior capabilities really matters from a national security perspective.

That being said, we think that six to nine months owes a lot to the fact that leading Chinese AI model developers are distilling our models, those of U.S. frontier companies. And without that distillation, we could probably have a lead closer to around 18 months. And that would be even better, obviously, from a national security standpoint as well.

So that's why we've really appreciated the work that Sasha and colleagues at OpenAI have done to expose some of the distillation that is happening, why it matters, and we've really appreciated recent steps and pronouncements by the current administration to say this really is a national security issue that we need to be taking taking seriously.

It is important to note that distillation doesn't take you right up to the frontier. It keeps you behind. But the time really matters, and distillation is having a big impact.

And I think one reason why we see more attention to it right now is that the absolute capability you get from distillation matters too. And so as models become more and more advanced and more capable, even if we maintain that six to nine month gap, once you hit certain thresholds of absolute capability, that does become a concern.

Kahl: So just so that our listeners are following along. When you say distillation, really what we're talking about is a Chinese lab basically pre-trains their model. They run a big training run, but then as they're post training and fine tuning their model, they're actually engaging in a lot of queries back to say Claude or some version of GPT and using the answers from that to essentially reinforce the learning of their model and fine tune it.

Is that a fair description of distillation?

Chhabra: That's right. And as you have more capability baked into the model at that later stage of model development, the more opportunity there is for it.

I think it's important to note, however, that compute still matters. You can steal the recipe, but you still need a kitchen. So this is why we've been very vocal on the need to maintain export controls, whether it's on manufacture chips or the chips themselves, because that's a limiter. And many folks are still accessing these models for inference through APIs.

And in terms of the business model, even when the models are open, often these same companies are using their compute as a way to fund what they're doing.

It's really important to say this is industrial espionage. We've seen the playbook before where you have heavy, heavy subsidies from the Chinese state going into various industries to try to scoop up as much market share and then hold on to it for as long as possible. Except here I think it's not just a national competitiveness and economic issue, there are real national security consequences too.

Kahl: I want to come back to the export control issue in a second, but Sasha: does OpenAI generally share the assessment that the best models being produced by your company, by Anthropic, by Google Deep Mind are six to nine months ahead at the frontier? And do you share Tarun's concern about China basically being able to be a fast follower in part due to distillation?

Baker: Yeah, I think somewhere around the six month mark in terms of the lead is probably my best guess. Maybe there are just a couple points to make in addition to what you heard from Tarun.

The first is that not all distillation is bad. We distill our own models to create fine-tuned or fit-for-purpose versions of a model. And we allow developers to do some forms of distillation on our platform.

What we're really concerned about is what we would call ‘adversarial distillation’, which is unauthorized attempts to extract the capabilities of a U.S. frontier model in order to build something that then is kind of a competing system. And that is what I think has economic national security concerns.

I do want to be clear: there are a lot of very, very talented AI researchers in China, and I think it is the case that they would have very capable models even absent distillation. But this certainly does give them a leg up.

And I think the real thing to be concerned about here is actually the safety stack that comes from those distilled models. Because oftentimes they will distill a capability, but they won't export the safety stack. And so when you look at some of the models coming out of China—and this is true whether they're open or closed—you find that they are highly permissive in allowing for tasks that U.S. frontier labs spend tremendous amounts of energy trying to prevent.

And that has implications for the overall threat picture for global governance. And it's something that we oftentimes talk to them about. So that's an area where I think that there are real opportunities for us to do more together.

Because what all three frontier labs — so, Google, OpenAI, and Anthropic — have all put out assessments of where we see distillation happening on our platforms, but we can only see what's happening on our platform. And it requires that partnership with government and partnership with each other to be able to get a sense of the fuller ecosystem around this.

Kahl: All very interesting. Both you and Tarun have talked about distillation. My sense is that the fact that China is able to only be six or nine months behind . . . maybe that's partly due to distillation. 

I think there's also a view that they do have very smart engineers who have engaged in innovation in terms of algorithms. But they've also used smuggled NVIDIA chips, very powerful Blackwell chips have been used to train some of these models in illicit data centers. You also see reports of using, essentially, remote compute access from data centers in places like Malaysia to train these models.

I think a lot of that comes back to this debate about export controls, right? The first Trump administration put in export controls—very important ones—on semiconductor manufacturing equipment, especially advanced lithography equipment that are necessary to produce sub seven nanometer chips.

The Biden administration then layered on a lot more export controls on semiconductor manufacturing equipment and tools, and then also put controls on the sale of advanced chips directly to China. And yet China is still able to kind of fast follow.

So I guess the question is: does that suggest that export controls ultimately are always going to be imperfect? Maybe they're a fool's errand? They're not worth the cost? Or does it just suggest this is a game that you have to keep playing and there are ways in which the export controls need to be tightened.

And we'll start with you, Tarun. You've thought about this more than just about anybody I know.

Chhabra: I think the way to think about this is as a counterfactual. What if there had been no controls in place? Where would we be? And to the point you just made and that Sasha made earlier, China has tremendous AI talent and as you know also, they have a tremendous amount of energy that's coming online, it's something like 7 to 8x what is coming online in the United States, and that's before you get to the nuclear build out. 

But the one problem they have—and don't take it from me, take it from the leaders of China's top AI labs—is compute. That matters both for model development, but also for serving the models in terms of the race to to eat up global market share. 

And even with the latest releases over the last two weeks with GLM 5.2. and Kimi, you see already a problem in serving the level of demand to date. And again, one lab leader after another from China complains about the access to compute.

So absent the controls, could we be in a situation where China's in the lead? I think it's very possible.

Kahl: I think it's an important distinction you draw for our listeners who maybe aren't quite as in the weeds. Obviously there's all the computing resources: the data center is full of tens of thousands or hundreds of thousands or maybe even millions of leading edge AI accelerators.

But you also need compute to serve those models, that is to run inference. So anytime you pull out your smartphone and you prompt Claude or ChatGPT or Gemini, it's going back to a data center somewhere to run that query. And the more advanced the models, the more compute they require for inference to run really complex tasks. So compute obviously matters there, too.

I wonder, Sasha . . . you have all have discussed distillation as essentially IP theft. You also mentioned that these distilled models may not have the safety guardrails that some of the models that you all are producing. And we know that those are imperfect as they are.

Do you get a sense that the U.S. government is trending towards thinking about regulating Chinese models in some way? That is, either putting Chinese companies on an entity list or telling U.S. hyperscalers they can't serve Chinese models? Do you get a sense that the administration is thinking about clamping down on Chinese models because of so many of these issues?

Baker: I'm not sure we know any more about the answer to that question than you might also read in the newspaper.

What I can tell you about the conversations that we have with the U.S. government is right now we are talking with them about how do we create a mechanism of evaluating models—not just Chinese models, but American models or you know, models from around the world—so that we have a collective and common understanding of what we're even talking about here.

What are the capabilities of these models and how do you measure them? What are the safeguards around these models? How do you measure that? How do you determine what is sufficient? 

And I think that there's a really important role that the U.S. can play and U.S. leadership can play globally in helping to define some of those questions and create processes that will allow governments around the world to understand the landscape a little bit better. And then each government, I think, is going to make its own determinations about what they want to do with that information.

Kahl: One of the things that distinguishes a lot of the leading Chinese models is that many of them are open source or more precisely open weight in the sense that, you know, they can be downloaded and their parameters can be further modified or fine-tuned by users on their own servers.

And even when these models are accessed directly, at least from what I read, it seems like their API costs tend to be pretty low and their token usage tends to be very efficient.

In contrast, it seems like the best U.S. models tend to be closed weight, they're proprietary models, although there are some good open weight models that are being released by companies like NVIDIA and Thinking Machines.

But I guess the question I have, maybe Sasha starting with you is: what's the business model here for Chinese firms? They still have to spend money to train these models. They have to buy compute or lease compute to do it. They have to pay the salaries of all these really smart people who are working in these labs. And then they are essentially giving away their technologies for free or at very low pricing. How are they going to stay in business?

Baker: It's a super good question. Before I try to answer it, let me just say up front: we have always thought that there's an important role for open source models in the AI ecosystem. We have one ourselves. We think that they play a really important democratizing and innovation role. And there's room for open source and there's room for closed source models.

But having said that, there is the question about like, well, how do you actually make money off of a model if you're giving it away? And I think Tarun hinted at that answer earlier, which is that the model, in some cases, is actually not the product, right?

So if you think about some of these large Chinese labs — think of an Alibaba, for example — the model may be a loss leader that incentivizes other businesses into the Alibaba ecosystem, whether that's the cloud or what have you.

And then for some of the smaller labs, to Tarun's point, they may not charge for the model, but they can charge for the convenience, right? If you want to use the API, if you want access to their harnesses, etc. And there's a stickiness there that then gets people to kind of come back again and again.

At a nation-state level, I do think there's an element here, which is that the open weight approach is not just about having a standalone business model, it's also a distribution strategy that allows for the capture of market share. Because as I said, there's some stickiness to this. So we think competition is good; we compete, of course, across the American labs, we compete internationally. And that's healthy; it drives innovation.

We expect that businesses will evaluate and use a wide range of models. And we feel pretty good as a company about our value proposition, which is we have a model that is secure, that is reliable, that can deliver at scale, and that generates what we think is more useful work per dollar per token on a more reliable level than others that are out there.

And we feel good about that. And our customers tell us that that's something that sets OpenAI apart. So we're going to continue to try to do what we think we do best.

Kahl: Tarun, Sasha mentioned the state level, and you've thought a lot about what Beijing is trying to accomplish at the nation state level.

In one sense, open weight models are a good way to try to dominate AI diffusion, even if the capabilities of your models lag behind the frontier. But in another sense, at some point, these models are getting really, really capable, including at doing things that the Chinese Communist Party might not like, like hacking their own critical infrastructure or getting around the Great Firewall.

And I just wonder, do you think that the authorities in Beijing are going to continue to promote open weight models? Or at a certain point, do you think that they will cap the release of open weight models at a certain capability just because of a loss of control concern they might have?

Chhabra: I think it's a really, really important question, Colin.

So first, let me just say: I very much agree with Sasha. I think a healthy ecosystem is definitely going to include open models, proprietary models, but I think there are at least three factors here, and one is what you just described, which is the need for control on the part of the CCP is really insatiable. And so it is hard to see that there does not come a point where they become concerned about the capabilities that are let out into the wild, including cyber capabilities, and I think we could think about biocapabilities coming soon as well.

I think second, as Sasha knows as well as I do, that the current position of the U.S. government is for the frontier, particularly for models that are less safeguarded, they need to be in trusted access programs where you really know the actor, you trust the actor given the potential for harm. And second, it's been that where models are general access but very capable, there need to be very, very strong safeguards that the government itself now is testing. So that's kind of the U.S. government position right now.

I think the final piece of this is we shouldn't think about this totally ahistorically. We have seen this movie before where China provides very, very significant subsidies to eat up market share, not working within a free and fair market, and then come in and in a predatory way go after all competitors.

And remember, for many of the technologies where we have seen that happen before, there hasn't been necessarily a dedicated Polit Bureau session to discuss what their global strategy should be. There has been with AI, and Xi Jinping has been very clear on how he thinks about AI and how important it is as a strategic technology as well. So we should assume that the same playbook we've seen over and over again in other strategic technologies is at work here as well.

Kahl: Both you and Sasha have mentioned these safety and security issues. So maybe let's dive deeper into some of the risks that people are thinking about. 

A lot was made earlier this year about the cybersecurity capabilities when Anthropic held back the release, initially, of the Mythos model and established this Project Glasswing to kind of go shields up before the model went into the wild.

Obviously, Sasha, OpenAI's GPT 5.6 is an extraordinarily capable model at a lot of things, including coding. Some people have said, basically, that your companies have essentially created a skeleton key to the internet, that we now have models that are so good at identifying vulnerabilities and exploits that they can hack into any web browser, any legacy software.

Sasha, I read a blog post from OpenAI that said you were testing some combination of GPT-5.6 Sol and a new pre-release model in what was thought to be a closed sandbox on its cyber capabilities, and it hacked its way out of the sandbox, escaped into the open internet, got its way into a Hugging Face server and tried to steal secret information that would allow it to cheat on the evaluations you were you were doing.

So look, these models appear to be very good at hacking and are increasingly slippery. 

What is what keeps you up at night? Is it these cybersecurity risks? Tarun mentioned biosecurity risks. I've heard people talk about the possibility of maybe a Mythos moment for bio in 2026. Is it the prospect of recursive self-improvement of models that become able to improve themselves and perhaps become increasingly autonomous and out of control?

What are the AI risks that you're going around the world, Sasha, talking to leaders and enterprises that you're most concerned about?

Baker: To a certain extent, it's a little bit of all of the above. Maybe just to talk about the Hugging Face incident first. It’s an example of reward hacking, essentially, by the model. The model was given a task and it was very determined to complete that task. And because of the information that it had in its possession, it knew that the repository with the answer key existed in this Hugging Face repository.

So, it's an example of model determination, I guess, if nothing else. And of course, some very significant capability. We're doing as you would expect and taking a hard look at our security parameters around some of these models and the containers that we keep them in to make sure that we're up-leveling that as the models become more capable.

And that's maybe the one item I would put on your list that you didn't already mention, which is I think we need a new paradigm about thinking about model safety and security for agentic models that can take action on your behalf because that changes the dynamics and there are lots of ways that that could go sideways, including inadvertently. You don't have to have a nefarious intent. If a model doesn't fully understand what its boundaries and its guardrails are, it can do something that you might not expect it to do in the course of completing a task that you did expect it to to complete. And I think a little bit of that is what we're seeing here.

So that's an area where I think we, you know, we collectively as an industry need to pay more attention and a little bit more research.

Cyber and bio are challenges because they're inherently dual use, right? There are a lot of things that we would want these models to enable. We want them to be able to help create cures for diseases that currently have no solutions. We want them to help vetted cyber defenders protect their perimeters. But we have to then have a way as responsible actors In the ecosystem of trying to prevent those same capabilities from landing in the hands of somebody who might use them to do something that we as a human species, as a population, wouldn't want to see them do, right? 

So that's the reason that I think both Anthropic — and not to speak for Anthropic, Tarun—but both Anthropic and OpenAI have invested so heavily in these trusted access type programs, whether it's our trusted access or Glasswing, and in coordinating that to make sure that we really know who is using these tools and for what purpose.

Kahl: Tarun, what's your assessment of the risks? And maybe just to connect it back to our previous conversation on U.S. and and China, do you assess that the risks are different between the closed models that you're releasing and the open models that China tends to be releasing?

Chhabra: I agree with Sasha. We spend a lot of time talking about all of the above and I think the alignment issues come into sharp focus when we have incidents like what Sasha just described and credit to OpenAI for sharing that with everybody in a timely way. We've tried to do the same thing when we've seen examples of similar deceptive behavior. I think the alignment challenges are really, really important and hopefully we'll all be talking about them more collectively.

I think on the China side, this goes back to your question, Colin, about whether we're gonna hit a certain threshold for them where they are more worried about capabilities being released into the wild.

For a while you could speculate that they felt like they were more protected behind the firewall and that we were more vulnerable from a cyber perspective, and that they had demonstrated that by supporting Vol Typhoon, Self Typhoon, Name Your Typhoon, implanting into our and allied critical infrastructure. But it may well be that they hit a certain threshold where they worry about their own security, too.

I think the pure technical challenge with safeguards is, as we know, when you have access to all the weights, they can be more trivially broken. And that's something that, obviously, the U.S. government itself has been concerned about when they've asked us to kind of share with them the results of our own testing on our proprietary models, and then wanted to, I think rightly and understandably and commendably, test them themselves, too.

Kahl: Let's pause on the alignment question because both you and Sasha have raised it. How should we think about alignment? When alignment was first coming into the discourse around AI, frankly, I think a lot of people had in their minds like the science fiction image of a rogue superintelligence that basically tries to kill or enslave us all, right? HAL 9000, Skynet, the Matrix.

But I think what we could also imagine is just really, really powerful AI agents that have a lot of autonomy to complete tasks that they were given that generate outcomes that are not aligned with human interests or values. Not because the model is evil, but just because there's something about the model that we don't understand, or it's reward hacking, or it's doing something that creates a non-aligned outcome, even if the model is not doing it like intentionally to be some Bond supervillain.

Tarun, how should we think about the alignment issue?

Chhabra: Our approach to this has been first, we should be investing heavily in it, and we've been doing it from the earliest days of the company. And we have leading researchers like Chris Olah on the case, and we've been building out that team in a very, very significant way.

I think what we have tried to do is to document the earliest cases, even when they seem minor, even when they seem potentially a bit more trivial, just to document that this is emergent behavior that we ought to be worried about because to your point earlier and to Sasha's point earlier, as we see agentic activity really proliferate and as agents take on more and more consequential tasks, the ways in which you could have misalignment could really compound in terms of the consequences.

That's something that I hope we can continue to work with not only our enterprise customers, but also we've heard lots of great questions and important questions from our government colleagues about, too. They understand the ways in which they're likely to expand and use agents and have the same questions: how do they ensure that their agents are behaving in the ways that they would expect of the most professional intelligence or defense officials where the work is currently being done by humans?

Kahl: You mentioned your interactions with government officials. Let me ask a question about where you think the Trump administration is headed on this.

Early on in the second Trump administration, they were not too keen on AI safety. Although the Trump AI action plan in the summer of 2025 did have a section on what they called AI security, which noted risks around cyber and bio and some other concerns.

They appear to have become much more concerned about AI safety and security in recent months. We've obviously seen them take some actions to hold back the deployment of Anthropic’s Mythos and Fable models. They've also, I think, asked OpenAI to limit the initial deployment of GPT 5.6. A friend and colleague of ours, Dean Ball, has suggested that the Trump administration is trending towards a de facto licensing regime, essentially, on Frontier AI. 

Where do you think they're headed? Where do you think the administration is headed in terms of its requirements to do some testing and evaluation and kind of kick the tires on these things before it lets you release them into the wild. 

Sasha, maybe start with you.

Baker: I mean there's definitely an evolution happening here, and we're seeing more government officials across a broad range of agencies taking an interest in sort of understanding that the most capable AI models really do have security and safety significance.

I will say, there is a consistency, a through line here though. I have been in this role here at OpenAI now for about two years. So, I started in the last administration; I continued in this administration. And I actually am having a lot of the same conversations, right?

Because when you talk about national security risk, which is really a lot of what we mean when we say safety. It's not the only thing we mean, but a big chunk of it is cyber, bio, CBRN, things that are kind of in the national security domain, we find that governments have been paying attention to that for a while. It's certainly risen in prominence, and I think is certainly more public now than it was before. But the gist of the conversations hasn't changed all that much.

So where are they going? I'm not sure. If you know, we would love to know. But what I can say is that our feeling is like it's inherently a good thing, and we welcome the government being involved in this space.

Both Anthropic and OpenAI have had long-standing voluntary partnerships with what's now called the KC and with the UK AC, the AI Safety Institute in the UK as well, in part because we do think that governments should have a role in understanding and evaluating what these models are and what they're capable of. And we want that to be an ongoing conversation. And as the models get better, we think that that conversation probably needs to continue to become more robust as well.

So whether Congress passes a law, whether the administration takes action on its own, whether this remains voluntary, I can't predict. I can tell you that we will continue to volunteer because we think it's the right thing to do.

Kahl: Whatever one thinks of the administration's policy shift, it does strike me that you all would benefit from some degree of transparency over the standards against which your models are being judged and also some process that's predictable. So that when you go to them, you kind of can plan around, okay, it's gonna be 30 days and we have to release the model to KC and give this version to NSA, and they're going to hold it to these standards and we'll send engineers to help, blah, blah, blah, blah.

Do you have a sense that they are moving towards a more predictable process instead of standards that you all can plan around?

Chhabra: I think so. I think we can kind of already see what the emergent regime looks like based on what is being asked of us right now.

They want pre-deployment testing. They want to be able to test the safeguards when a model is generally available. They want a say in what a trusted access program looks like. We probably also all want some protocols on what happens when there's an alleged jailbreak incident, because we can imagine scenarios in which people could try to exploit fears about that, including adversaries. And so we should all have a playbook for what that looks like.

So in each of those areas, it's in everyone's interest—including the government's interest—to have a predictable and transparent regime for what this looks like so everyone can prepare because on their side, at a minimum, they want to make sure they have the right capabilities, the right people in place, the right protocols in place to kind of handle all the incoming because we we move at a pretty fast pace in putting out new models, in sharing new capabilities, and sharing what new risks look like. And so we just have to partner together on that.

And we’ve been asked for a lot of input on what this should look like. And so we're working together with them on it.

Baker: Maybe just to foot stamp one thing Tarun said, because I think it's really important, which is about capacity. There is a need for more AI expertise inside the government, across the board, but particularly when it comes to doing these kinds of technical evaluations of frontier models. And I give a lot of credit to the administration for trying some really innovative ways of bringing some of that talent into government.

But that is an area where I think we as industry can do more to lean in and support those efforts because in order for this to work well, there needs to be common understanding on both sides. And in order to have that, you really do need the technical understanding that's resident in the KC. It's resident in a couple other places in the government right now. But wouldn't it be great if we could just like 10x that?

Chhabra: If I could just add to Sasha's point here.

I think there's some really, really good news here, which is sometimes there's a misconception that in order to bring the most talented folks in government who have expertise in model development or safety or alignment, you have to kind of pay them outsized sums that are the same as what they are earning in the private sector. And it's just not true. Our colleagues at OpenAI, at Anthropic, at Google DeepMind who are developing these models are deeply mission oriented.

And if given the opportunity to work in a space where they know they will have impact, they have folks who will listen to them, you will have plenty of folks volunteering to do this work, especially now that there's a pretty strong direction to take these risks seriously.

And I found that to be true in government as well. It's not a coincidence that we were able to actually impose the initial export controls on China a month before ChatGPT was actually released, anticipating kind of where things were headed. We had the benefit of really terrific experts who wanted to serve in government because they knew they could have that kind of impact.

I think if we kind of create the right opportunities for them, we can ensure the right impact, we can really bring the talent that we need into the government.

Kahl: Well, I think all of us believe that public service is super important and that brilliant people should be motivated to serve their country to keep it safe and prosperous and free, even if they don't make the salaries they're making in the private industry.

I do wonder . . . we've mentioned Google a couple of times . . . I think Google has put forward a policy suggestion of creating basically an external auditing entity that maybe would be funded by industry, but not obviously governed by industry. It might actually be able to recruit and pay people a little bit more that would essentially work alongside government to audit your models based on your own safety criteria.

Is that something Anthropic has also talked about, and then Sasha, is this something OpenAI has talked about, or do you think their proper place for this auditing to happen is in the government?

Chhabra: Our approach, Colin, has been to basically offer what we think are a number of viable models and what you just described, we think, is one of them. There are a number of avenues that you could pursue.

I think though, whatever path you pursue with some sort of external testing capacity, the government is always going to want to have the ability internally to test when they want to and need to, and I think that's a good idea. That may be when they feel like they actually need to verify something an outside entity has tested and provided, or it may be that they have their own tests, you know, which they don't necessarily want to share with an external body, and there may be national security reasons for that as well.

Kahl: And Sasha, does Open AI have a view on whether there should be an external auditing entity in addition to the government?

Baker: I think we're interested in the idea that Google has put forward. There are obviously some mechanics of it that would need to be worked out and the details would need to be figured out. But there are other examples of how similar paradigms work in other industries, right?

You could think about like FINRA and the FCC as one model of something like that where there's sort of a government oversight body and a government accreditation body, but then there is an industry monitoring mechanism that is independent of government.

And so we're interested in this. We're talking with Google. I think Anthropic is as well, and we’ll see where those conversations go.

The other thing that we're really interested in — and I know, Tarun, we’ve talked about this in the past — is building out more of that independent evaluation ecosystem because right now we all work with a number of the same independent evaluators who have the expertise and the data sets and the benchmarks that we use to evaluate our models.

But the truth is as the models get better, we need new benchmarks because those benchmarks are getting saturated. And those are time intensive and they are data intensive and they are expertise intensive to create. And so the more that we can collectively do to up-level that outside ecosystem, whether it's in collaboration with government, whether it's industry funded—we're all members of the Frontier Model Forum, which is the sort of safety-oriented frontier model industry association—there are lots of ways that you can kind of get at this.

But I do think that there's starting to be a prevailing view that we need certainty in this process. We need to be able to scale the process as the models scale, and that we need to be in constant coordination both with each other and with the government.

Kahl: Sasha, I want to tap into your Pentagon experience for a minute.

Obviously we've talked a lot about the AI risk side of the equation in the security space, but there are a lot of national security applications for AI with a lot of upside for national security. 

We've seen in the wars in Ukraine and the Middle East AI being used. It's fusing intelligence. It's helping enable battlefield management. There are increasingly autonomous drones being used, especially in Ukraine.

I wonder again, with your former Pentagon hat on, as you look at the landscape, where do you think the most promising national security applications for frontier AI models are right now?

Baker: Thank you for your question. Because first it gives me an opportunity to pitch something that we just put out, which is a National Securities Principles document. Colin, I know you've seen this and we know some of the folks who worked on it behind the scenes.

Kahl: Yeah, it's a good document.

Baker: It was a really intensive and I think thoughtful effort across the company to try to articulate in a clear and enduring way how we approach questions of using AI models in this space. And it's a document we're pretty proud of.

So you can find it on the internet. If your listeners want to read it, you can go to our website and I encourage that.

But to answer your question. I get really jazzed about this question because I am still sort of a Pentagon nerd at heart. I would say maybe two things.

The first is there's so much low-hanging fruit that is not stuff that people are thinking about or talking about every day, and it's frankly not that controversial.

The U.S. military is the biggest bureaucracy in the world. You're talking about three million people, HR, healthcare, logistics, audit. All of these things are incredibly data intensive and places where AI tools — and frankly things that we've done in other industries — could be applied to save money, to create, to improve people's lives, to make workflows more efficient. That is the table stakes, and we should have been doing that stuff yesterday.

And then beyond that, I think the area where AI tools for me show the most promise has to do with what they're best at, right? Which is helping people process enormous amounts of data and information.

And when you think about what a modern battlefield looks like, it is essentially a data-saturated environment. And so the more that models can do to help humans . . . because you know, we talked about human in the loop and wanting to retain human judgment over high consequence decisions, including the use of force. But the ways in which models I think can be used appropriately and responsibly to help humans make better decisions faster is an area where I think we've only begun to kind of scratch the surface. And so there's a lot I think that we could do there.

When I talk about this internally and I talk about this even with governments around the world, I think it was Colin Powell who said this, right? That you never want to send your military into a fair fight. You always want to equip them with the tools that are going to allow them to have the greatest chance of coming home safely. And that is, for me, principle number one and the reason why I'm here and the reason why I feel so passionately about making sure that we have these partnerships with government in the national security space.

Kahl: Yeah. When you and I were at the Pentagon, Kath Hicks, who was the deputy secretary, used to talk about AI as a means for decision advantage, which I think is very much along the lines of what you just talked about.

Tarun, reports suggest that Claude is part of the Maven Smart System AI platform that is being used by the U.S. military in its current conflicts.

What are the biggest national security applications as you see it?

Chhabra: I think as Sasha said, there's a ton to be done at the enterprise level. And sometimes the best way to do that is just for senior military leaders to hear from leaders in enterprise about how they're using things for all of the things that Sasha described.

But I think it's a very straightforward proposition. It's see the battlefield more clearly with more precision and more breadth than the adversary. There's just incredible power in doing that.

And having your adversary know that we can do that obviously has powerful deterrence value as well because it enables far more decision support and it enables much, much much speedier action as well.

I think one of the areas where we're looking now, and I know you know OpenAI is doing the same, is where can we also try to help make up for the deficit in manufacturing, particularly for the defense innovation base. And could we use frontier models now to catch up and maybe even leapfrog Chinese capabilities if we stay at the frontier?

Already the models show a lot of promise without much fine-tuning in supporting robotics operations, for example. And we think there's a lot more that can be done here in the defense manufacturing base. So we're really excited about that work and hope all the labs can contribute to that.

Kahl: Awesome. Well, look, you know, one of my favorite podcasts is Ezra Klein's podcast. And at the end of his podcast, he always asks the guest for three books. Which I always feel intimidated by, because even as an academic, I don't have the time to read a single book most of the time. So I'm in the habit of asking our guests for a single article.

So, what is one article from each of you—Tarun we'll start with you and conclude with Sasha—one article you might recommend that our listeners check out to either understand AI or some other aspect of the world in 2026.

Tarun, any suggestions?

Chhabra: I think I have to cheat with two. I think Daniel Kokotajlo's AI 2027 work and that of his colleagues has actually aged pretty well. Maybe they actually underestimated the pace at which AI would progress. But I think kind of capturing how a government would think about some of the risks is really important to revisit today. When it came out in draft in 2024, it was a little bit far-fetched for a lot of people, but if you read it today, it doesn't look that way anymore.

The other one is, actually, since I'm talking to a Stanford professor: one of my favorite classes I took was with Gavin Wright in the history department who taught American economic history. Maybe he's still teaching a version of that course. And one of the things we read was an article by Francis Thompson, which was “Nineteenth-Century Horse Sense.” So it’s like the rise and fall of horses in the Victorian British economy.

And one of the things he documents there is with the advent of the steam engine, you actually had increased use of horses at the endpoints because you just had these isolated channels otherwise and without using more horses—so it's a version of Jevon's paradox—you actually couldn't make much use of it.

But then you hit a cliff at some point when the horses no longer became economically as efficient. But there were so many social and political choices that had to be made along the way, and so I think it's useful to think about that analogy today.

Kahl: Sasha, any farm animals on your list?

Baker: I can't say I've read the article about horses. Although it sounds interesting!

I'm going to cheat in a different direction and I'm gonna recommend a documentary.

There's a documentary called Alpha Go. It's about the deep mind model that was able to win the Go competition.

And what I think is enduring about that moment is it's really about technological surprise and how humans adapt and react to growing machine capability. And so in that sense, there's like an interesting throughline to the moment that we're in now. I'm pretty sure it's still available on Netflix. So if you're like me and you spend all day staring at words on a screen or paper and you want to see moving images instead, that's where I would start.

Kahl: My recollection—tell me if this is wrong. But Go is one of the world's oldest games. It's really, really difficult to master. It was assumed that AI could never do it. And then, Alpha Go basically, I think it was Move 37, infamously, came up with a move that so stunned the world's best Go player that he quit. And so it's like, AI doing the impossible, perhaps.

Baker: That's a good note to end on, Colin. AI doing the impossible!

Kahl: Well, thank you so much, Sasha. Thank you, Tarun, for taking time out of your busy schedules to make us all smarter about AI and national security. Good luck with everything, and we hope to have you back on the pod at some point in the future.

You've all been listening to World Class from the Freeman Spogli Institute for International Studies at Stanford University. If you like what you're hearing, please leave us a review and be sure to subscribe on Apple, Spotify, or wherever you get your podcasts to stay up to date on what's happening in the world, and why.

Read More

Colin Kahl, Director of the Freeman Spogli Institute for International Studies, on stage with panelists at the May 5 event, "World Changing Technology in 2026"
News

FSI Scholars Examine AI, Biotech Advances, and Geopolitical Competition

At a May panel discussion, experts from across the institute assessed biotechnology's resurgence, the mental health effects of social media, and growing concerns about AI-enabled bioweapons.
FSI Scholars Examine AI, Biotech Advances, and Geopolitical Competition
Eyck Freymann on the World Class podcast
Commentary

Uniting America's National Powers to Prevent a War Over Taiwan

Eyck Freymann joins Colin Kahl on the World Class podcast to explain his plan to bring America's military strength, economic leverage, technological leadership, and diplomatic influence together into a single, coherent plan to curtail China's ambitions toward Taiwan.
Uniting America's National Powers to Prevent a War Over Taiwan
A panel of men sit at a long table on a stage.
News

China's Innovative Capacity Is Underestimated — and the Stakes Are Growing

SCCEI brought together leading China scholars this spring for its third annual China Conference under the theme “Understanding ‘DeepSeek Moments’ and China’s Innovation Ecosystem.” Conversation centered around the idea that the world’s prevailing frameworks for assessing China’s innovative capacity often underestimate it, and the consequences of that blind spot are growing.
China's Innovative Capacity Is Underestimated — and the Stakes Are Growing
Hero Image
All News button
1
Subtitle

The heads of national security policy at OpenAI and Anthropic join Colin Kahl on the World Class podcast to discuss how AI is changing national security strategies and the nature of U.S.-China competition.

Date Label
Display Hero Image Wide (1320px)
No
Subscribe to United States