AI Safety Newsletter
AI Safety Newsletter

AI Safety Newsletter

Narrations of the AI Safety Newsletter by the Center for AI Safety. We discuss developments in AI and AI safety. No technical background required. This podcast also contains narrations of some of our publications. ABOUT US The Center for AI Safety (CAIS) is a San Francisco-based research and field-building nonprofit. We believe that artificial intelligence has the potential to profoundly benefit the world, provided that we can develop and use it safely. However, in contrast to the dramatic progress in AI, many basic problems in AI safety have yet to be solved. Our mission is to reduce societal-scale risks associated with AI by conducting safety research, building the field of AI safety researchers, and advocating for safety standards. Learn more at https://safe.ai

配信元ページ

エピソード

表示 12 / 全 87

AISN #25: White House Executive Order on AI, UK AI Safety Summit, and Progress on Voluntary Evaluations of AI Risks.

Tue, 31 Oct 2023 00:00:00 GMT

AISN #24: Kissinger Urges US-China Cooperation on AI, China’s New AI Law, US Export Controls, International Institutions, and Open Source AI.

Wed, 18 Oct 2023 00:00:00 GMT

AISN #23: New OpenAI Models, News from Anthropic, and Representation Engineering.

Wed, 04 Oct 2023 00:00:00 GMT

AISN #21: Google DeepMind’s GPT-4 Competitor, Military Investments in Autonomous Drones, The UK AI Safety Summit, and Case Studies in AI Policy.

Tue, 05 Sep 2023 00:00:00 GMT

AISN #20: LLM Proliferation, AI Deception, and Continuing Drivers of AI Capabilities.

Tue, 29 Aug 2023 00:00:00 GMT

[Paper] “An Overview of Catastrophic AI Risks” by Dan Hendrycks, Mantas Mazeika and Thomas Woodside

Mon, 21 Aug 2023 09:00:01 GMT

[Paper] “X-Risk Analysis for AI Research” by Dan Hendrycks and Mantas Mazeika

Mon, 21 Aug 2023 09:00:00 GMT

[Paper] “Unsolved Problems in ML Safety” by Dan Hendrycks, Nicholas Carlini, John Schulman and Jacob Steinhardt

Mon, 21 Aug 2023 09:00:00 GMT

AISN #19: US-China Competition on AI Chips, Measuring Language Agent Developments, Economic Analysis of Language Model Propaganda, and White House AI Cyber Challenge.

Tue, 15 Aug 2023 00:00:00 GMT

AISN #18: Challenges of Reinforcement Learning from Human Feedback, Microsoft’s Security Breach, and Conceptual Research on AI Safety.

Tue, 08 Aug 2023 00:00:00 GMT

AISN #17: Automatically Circumventing LLM Guardrails, the Frontier Model Forum, and Senate Hearing on AI Oversight.

Tue, 01 Aug 2023 00:00:00 GMT

AISN #16: White House Secures Voluntary Commitments from Leading AI Labs, and Lessons from Oppenheimer .

Tue, 25 Jul 2023 00:00:00 GMT