Center for AI Safety

AI Safety Newsletter

Narrations of the AI Safety Newsletter by the Center for AI Safety. We discuss developments in AI and AI safety. No technical background required. This podcast also contains narrations of some of our publications. ABOUT USThe Center for AI Safety (CAIS) is a San Francisco-based research and field-building nonprofit. We believe that artificial intelligence has the potential to profoundly benefit the world, provided that we can develop and use it safely. However, in contrast to the dramatic progress in AI, many basic problems in AI safety have yet to be solved. Our mission is to reduce societal-...

Author

Center for AI Safety

Category

Technology

Podcast website

newsletter.safe.ai

Latest episode

Jul 6, 2026

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

AISN #7: Disinformation, recommendations for AI labs, and Senate hearings on AI. 23.05.2023

How AI enables disinformation Yesterday, a fake photo generated by an AI tool showed an explosion at the Pentagon. The photo was falsely attributed to Bloomberg News and circulated quickly online. Within minutes, the stock market declined sharply, only to recover after it was discovered that the picture was a hoax. This story is part of a broader trend. AIs can now generate text, audio, and images...

AISN #6: Examples of AI safety progress, Yoshua Bengio proposes a ban on AI agents, and lessons from nuclear arms control . 16.05.2023

Examples of AI safety progress Training AIs to behave safely and beneficially is difficult. They might learn to game their reward function, deceive human oversight, or seek power. Some argue that researchers have not made much progress in addressing these problems, but here we offer a few examples of progress on AI safety. Detecting lies in AI outputs. Language models often output false text, but...

AISN #5: Geoffrey Hinton speaks out on AI risk, the White House meets with AI labs, and Trojan attacks on language models. 09.05.2023

Geoffrey Hinton is concerned about existential risks from AI Geoffrey Hinton won the Turing Award for his work on AI. Now he says that part of him regrets his life’s work, as he believes that AI poses an existential threat to humanity. As Hinton puts it, “it’s quite conceivable that humanity is just a passing phase in the evolution of intelligence.” AI is developing more rapidly than Hinton expect...

AISN #4: AI and cybersecurity, persuasive AIs, weaponization, and Hinton talks AI risks. 02.05.2023

Cybersecurity Challenges in AI Safety Meta accidentally leaks a language model to the public. Meta’s newest language model, LLaMa, was publicly leaked online against the intentions of its developers. Gradual rollout is a popular goal with new AI models, opening access to academic researchers and government officials before sharing models with anonymous internet users. Meta intended to use this str...

AISN #3: AI policy proposals and a new challenger approaches. 25.04.2023

Policy Proposals for AI Safety Critical industries rely on the government to protect consumer safety. The FAA approves new airplane designs, the FDA tests new drugs, and the SEC and CFPB regulate risky financial instruments. Currently, there is no analogous set of regulations for AI safety. This could soon change. President Biden and other members of Congress have recently been vocal about the ris...

AISN #2: ChaosGPT and the rise of language model agents, evolutionary pressures and AI, AI safety in the media. 18.04.2023

ChaosGPT and the Rise of Language Agents Chatbots like ChatGPT usually only respond to one prompt at a time, and a human user must provide a new prompt to get a new response. But an extremely popular new framework called AutoGPT automates that process. With AutoGPT, the user provides only a high-level goal, and the language model will create and execute a step-by-step plan to accomplish the goal....

AISN #1: Public opinion on AI, plugging ChatGPT into the internet, and the economic impacts of language models.. 10.04.2023

Growing concerns about rapid AI progress Recent advancements in AI have thrust it into the center of attention. What do people think about the risks of AI? The American public is worried. 46% of Americans are concerned that AI will cause “the end of the human race on Earth,” according to a recent poll by YouGov. Young people are more likely to express such concerns, while there are no significant...

Listen to the AI Safety Newsletter podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.