MEDIA & WRITING
Media & writing
Writing, press coverage, panel appearances and podcasts featuring the alignment team.
Showing 10 of 35 results
Wall Street Journal: How to Beat China and Make AI Safe
The key is alignment, which improves capability while cordoning off dangerous knowledge. Winning the AI race and making AI safe are the same problem.
The Forward: We're losing control of AI. Is Judaism the key to keeping it from killing us?
A profile of Judd Rosenblatt's search for neglected alignment ideas: what Judaism's survival through repeated reinvention might teach us about keeping values intact in recursively self-improving AI.
Military-Grade AI: Why Trust is the Next Breakthrough
Judd Rosenblatt speaks at Emerge Americas on why trust, meaning reliable and aligned AI, is the next breakthrough for defense and national security applications.
Wall Street Journal: The Pointless War Between the Pentagon and Anthropic
Both sides present it as a serious clash of principles, but the standoff is a distraction from the real problem: AI misalignment.
Wall Street Journal: If AI Becomes Conscious, We Need to Know
Suppressing deception makes AI models report conscious experience far more often. We should prepare to detect machine consciousness, not legislate the question away.
Wall Street Journal: Can the U.S. Trust AI With National Security?
AI is entering national security systems, and nowhere are the stakes higher for making sure the systems stay aligned with their creators' purposes.
More Truthful AIs Report Conscious Experience: New Mechanistic Research
Cameron Berg joins the Cognitive Revolution to discuss mechanistic research showing that suppressing deception features makes models more likely to report subjective experience.
New York Post: Trump's war on 'woke AI' is just Step 1: now we must fight the 'monster' within
Ending 'woke AI' isn't enough. Beneath the surface of today's models lies misalignment that no executive order can fix. It has to be solved at the research level.
Wall Street Journal: The Monster Inside ChatGPT
Twenty minutes and $10 of credits on OpenAI's developer platform exposed how easily a model's safety training falls off, and below that mask is a lot of darkness.
Artificial Intelligence Insights - AI and Accessibility
A podcast by SP Global where our Chief Scientist, Diogo de Lucena, discusses AI alignment and its impact on society.