AI Safety Newsletter
AI Safety Newsletter

AI Safety Newsletter

Narrations of the AI Safety Newsletter by the Center for AI Safety. We discuss developments in AI and AI safety. No technical background required. This podcast also contains narrations of some of our publications. ABOUT US The Center for AI Safety (CAIS) is a San Francisco-based research and field-building nonprofit. We believe that artificial intelligence has the potential to profoundly benefit the world, provided that we can develop and use it safely. However, in contrast to the dramatic progress in AI, many basic problems in AI safety have yet to be solved. Our mission is to reduce societal-scale risks associated with AI by conducting safety research, building the field of AI safety researchers, and advocating for safety standards. Learn more at https://safe.ai

配信元ページ

エピソード

表示 12 / 全 83

AISN #20: LLM Proliferation, AI Deception, and Continuing Drivers of AI Capabilities.

Tue, 29 Aug 2023 00:00:00 GMT

[Paper] “An Overview of Catastrophic AI Risks” by Dan Hendrycks, Mantas Mazeika and Thomas Woodside

Mon, 21 Aug 2023 09:00:01 GMT

[Paper] “X-Risk Analysis for AI Research” by Dan Hendrycks and Mantas Mazeika

Mon, 21 Aug 2023 09:00:00 GMT

[Paper] “Unsolved Problems in ML Safety” by Dan Hendrycks, Nicholas Carlini, John Schulman and Jacob Steinhardt

Mon, 21 Aug 2023 09:00:00 GMT

AISN #19: US-China Competition on AI Chips, Measuring Language Agent Developments, Economic Analysis of Language Model Propaganda, and White House AI Cyber Challenge.

Tue, 15 Aug 2023 00:00:00 GMT

AISN #18: Challenges of Reinforcement Learning from Human Feedback, Microsoft’s Security Breach, and Conceptual Research on AI Safety.

Tue, 08 Aug 2023 00:00:00 GMT

AISN #17: Automatically Circumventing LLM Guardrails, the Frontier Model Forum, and Senate Hearing on AI Oversight.

Tue, 01 Aug 2023 00:00:00 GMT

AISN #16: White House Secures Voluntary Commitments from Leading AI Labs, and Lessons from Oppenheimer .

Tue, 25 Jul 2023 00:00:00 GMT

AISN #15: China and the US take action to regulate AI, results from a tournament forecasting AI risk, updates on xAI’s plan, and Meta releases its open-source and commercially available Llama 2.

Wed, 19 Jul 2023 00:00:00 GMT

AISN #14: OpenAI’s ‘Superalignment’ team, Musk’s xAI launches, and developments in military AI use .

Wed, 12 Jul 2023 00:00:00 GMT

AISN #13: An interdisciplinary perspective on AI proxy failures, new competitors to ChatGPT, and prompting language models to misbehave.

Wed, 05 Jul 2023 00:00:00 GMT

AISN #12: Policy Proposals from NTIA’s Request for Comment, and Reconsidering Instrumental Convergence.

Tue, 27 Jun 2023 00:00:00 GMT