“AI for AI safety” by Joe Carlsmith
(Audio version here (read by the author), or search for "Joe Carlsmith Audio" on your podcast app. This is the fourth essay in a series that I’m calling “How do we solve the alignment problem?”. I’m hoping that the individual essays can be read fairly well on their own, but see this introduction fo…
Joe Carlsmith on How We Change Our Minds About AI Risk
Joe Carlsmith joins the podcast to discuss how we change our minds about AI risk, gut feelings versus abstract models, and what to do if transformative AI is coming soon. You can read more about Joe's work at https://joecarlsmith.com.
Timestamps:
00:00 Predictable updating on AI risk
07:27 Abst…
Joe Carlsmith - A Wiser, AI-Powered Civilization is the “Successor” (Worthy Successor, Episode 15)
This new installment of the Worthy Successor series is an interview with Joe Carlsmith, a senior advisor at Open Philanthropy, whose work spans AI alignment, moral uncertainty, and the philosophical foundations of value. In this conversation, Joe joins us in his personal capacity, not representing …
Consciousness and Competition (with Joe Carlsmith)
Joe Carlsmith is a writer, researcher, and philosopher. He works at Anthropic on the character/constitution/spec for Claude. Before that, he was a senior advisor at Open Philanthropy.
You can read a full transcript here.
To see all Forethought’s published research, visit forethought.org/research.…
Joe Carlsmith - Deep Atheism, Death, and Sincerity
Joe Carlsmith is a philosopher and writer at Anthropic whose concept of “deep atheism” has changed how I think about religion. In this conversation we explore what it means to trust or distrust the universe, why death brings life into focus, and why it can be hard to be sincere. We discuss Teilhard…
Ginevra Davis — The Attractor States of Intelligence
Ginevra Lily Davis, the contemporary philosopher Erik calls his favorite, joins the show to defend an unusual position with disarming clarity: that the universe has a real bottom, and that bottom is positively-valenced consciousness. From there she takes on postmodernism (it eats itself), Eliezer Y…
Let's CALL OUT the AI Doom “Enablers” Joining OpenAI & Anthropic — Dr. Holly Elmore, PauseAI US
Dr. Holly Elmore, executive director of PauseAI US, is done giving AI insiders the benefit of the doubt.
In this spicy episode, Holly calls out a number of respected figures in the AI safety community for allegedly “selling out” the cause: Effective Altruists, Anthropic & OpenAI employees, Eliezer …
The OpenAI/Huggingface incident | Redwood Research podcast episode 2
We talk about the OpenAI–Hugging Face incident, where an OpenAI model — in the middle of a cyber evaluation — broke out of its sandbox and autonomously hacked Hugging Face.
We discuss:
* What we actually know happened.
* How surprising the incident was.
* What the incident does (and doesn’t) tell u…
"Arguments for P" by Cleo Nardo
Daniel Kokotajlo: To be clear, we don’t claim P will happen specifically. But when we wrote out our best-guess scenario month by month, P kept happening. Eventually we decided to just publish P. I’m at ~80% on P; my coauthors are lower. Ryan Greenblatt: I thought it would be helpful to post my curr…
PODCAST-SUCHMASCHINE
Auf unserer Website können Sie3,820,941-Podcasts und195,298,021-Folgen nach Personen, Orten oder Themen durchsuchen.
If I have seen further than others, it is by listening to podcasts and standing upon the shoulders of
giants.
‐ Isaac "Llamacorn" Newton
BEARBEITEN
Vielen Dank, dass Sie uns helfen, die Podcast-Datenbank auf dem neuesten Stand zu halten.