Center for AI Safety
AI Safety Newsletter
Narrations of the AI Safety Newsletter by the Center for AI Safety. We discuss developments in AI and AI safety. No technical background required. This podcast also contains narrations of some of our publications. ABOUT USThe Center for AI Safety (CAIS) is a San Francisco-based research and field-building nonprofit. We believe that artificial intelligence has the potential to profoundly benefit the world, provided that we can develop and use it safely. However, in contrast to the dramatic progress in AI, many basic problems in AI safety have yet to be solved. Our mission is to reduce societal-...
Author
Center for AI Safety
Category
Podcast website
Latest episode
Jul 6, 2026
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
AISN #7: Disinformation, recommendations for AI labs, and Senate hearings on AI. 23.05.2023 12:43
How AI enables disinformation Yesterday, a fake photo generated by an AI tool showed an explosion at the Pentagon. The photo was falsely attributed to Bloomberg News and circulated quickly online. Within minutes, the stock market declined sharply, only to recover after it was discovered that the picture was a hoax. This story is part of a broader trend. AIs can now generate text, audio, and images...
AISN #6: Examples of AI safety progress, Yoshua Bengio proposes a ban on AI agents, and lessons from nuclear arms control . 16.05.2023 11:30
Examples of AI safety progress Training AIs to behave safely and beneficially is difficult. They might learn to game their reward function, deceive human oversight, or seek power. Some argue that researchers have not made much progress in addressing these problems, but here we offer a few examples of progress on AI safety. Detecting lies in AI outputs. Language models often output false text, but...
AISN #5: Geoffrey Hinton speaks out on AI risk, the White House meets with AI labs, and Trojan attacks on language models. 09.05.2023 8:00
Geoffrey Hinton is concerned about existential risks from AI Geoffrey Hinton won the Turing Award for his work on AI. Now he says that part of him regrets his life’s work, as he believes that AI poses an existential threat to humanity. As Hinton puts it, “it’s quite conceivable that humanity is just a passing phase in the evolution of intelligence.” AI is developing more rapidly than Hinton expect...
AISN #4: AI and cybersecurity, persuasive AIs, weaponization, and Hinton talks AI risks. 02.05.2023 9:30
Cybersecurity Challenges in AI Safety Meta accidentally leaks a language model to the public. Meta’s newest language model, LLaMa, was publicly leaked online against the intentions of its developers. Gradual rollout is a popular goal with new AI models, opening access to academic researchers and government officials before sharing models with anonymous internet users. Meta intended to use this str...
AISN #3: AI policy proposals and a new challenger approaches. 25.04.2023 7:50
Policy Proposals for AI Safety Critical industries rely on the government to protect consumer safety. The FAA approves new airplane designs, the FDA tests new drugs, and the SEC and CFPB regulate risky financial instruments. Currently, there is no analogous set of regulations for AI safety. This could soon change. President Biden and other members of Congress have recently been vocal about the ris...
AISN #2: ChaosGPT and the rise of language model agents, evolutionary pressures and AI, AI safety in the media. 18.04.2023 7:28
ChaosGPT and the Rise of Language Agents Chatbots like ChatGPT usually only respond to one prompt at a time, and a human user must provide a new prompt to get a new response. But an extremely popular new framework called AutoGPT automates that process. With AutoGPT, the user provides only a high-level goal, and the language model will create and execute a step-by-step plan to accomplish the goal....
AISN #1: Public opinion on AI, plugging ChatGPT into the internet, and the economic impacts of language models.. 10.04.2023 8:29
Growing concerns about rapid AI progress Recent advancements in AI have thrust it into the center of attention. What do people think about the risks of AI? The American public is worried. 46% of Americans are concerned that AI will cause “the end of the human race on Earth,” according to a recent poll by YouGov. Young people are more likely to express such concerns, while there are no significant...
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.