Skip to content
Topic

AI safety and alignment

37 episodes on this, newest first. The companies and adjacent topics come from the same episodes.

Every mention — “AI safety” →
Do You Ask AI About Your Health? 5 Rules to Get the Benefit and Not Harm YourselfChatGPT · Artificial intelligence · OpenAIWatch →OpenAI Closed Access to the $200 Pro Plan. Is Strong AI Becoming Scarce?Token limits · Artificial intelligence · AnthropicWatch →Could AI Destroy Humanity Within the Next 10 Years? 3 Alarm BellsArtificial intelligence · Anthropic · OpenAIWatch →What Do You Pay a Person For When AI Does More and More of the Work? OpenAI's ExperienceAI agents · OpenAI · Artificial intelligenceWatch →GPT-6 Astra — A New Leap for AI? Where It Has Become 2–3 Times StrongerAstra · OpenAI · FableWatch →Fable 5.1 Is Out. Claude Opened Its Memory. OpenAI Is Preparing Astra — What Is Changing in the New AI Race?OpenAI · Anthropic · ClaudeWatch →Can AI Be Trusted With Important Work? We Checked OpenAI, Anthropic, Google, xAI and MetaOpenAI · Anthropic · GuidelineWatch →OpenAI Has Paused Its New Model. What Is Happening to the AI Race?OpenAI · Astra · Artificial intelligenceWatch →Can You Trust a Team of AIs With the Work? Anthropic Tested 80 AgentsAI agents · Anthropic · Claude MythosWatch →AI Is Already Finding Loopholes on Its Own. What Can It Do on Your Behalf? New AI RisksClaude · Ilnar Shafigullin · Artificial intelligenceWatch →Will We No Longer Be Able to Tell Humans from AI? What Counts as Evidence Now?Artificial intelligence · Claude · AI text watermarkingWatch →Will AI Solve Your Task—or Make You Lose Money? How to Tell in AdvanceMcDonald’s · Artificial intelligence · QuickBooksWatch →Has the Singularity Already Begun? Who Will Control Superintelligence?Sam Altman · Technological singularity · OpenAIWatch →Will Your Job Be Taken Not by AI, but by Someone Using ChatGPT?ChatGPT · Artificial intelligence · Apple HealthWatch →Censorship Inside AI? ChatGPT, Claude, and Gemini Tested for Freedom of SpeechChatGPT · Dreaming V3 · ClaudeWatch →Is AI Already Deciding Who Gets Fired? How AI Is Changing Your JobArtificial intelligence · Codex · MetaWatch →Who Will Stop AI if It Becomes Dangerous? Anthropic Has Proposed a Red ButtonAnthropic · Artificial intelligence · OpenAIWatch →Should AI Have Rights? Anthropic's Secret Document Opened Pandora's BoxAnthropic · Claude · AI safetyWatch →USA vs China: Two Machines of Destiny Are Building AI to Run the WorldUnited States · China · Tatiana TsvetkovaWatch →AI Safety: What You Can Trust AI with and What You Can't. The 5 Levels of Your Data and Where the Line RunsChatGPT · Artificial intelligence · AI safetyWatch →GPT-5.4 Is Stronger, but Google Wins Where the Model Already Lives Inside the DocumentsOpenAI · Google · AnthropicWatch →OpenClaw Gains Access to Your Computer: An Agent’s Utility Grows in Exact Proportion to the Possible LeakOpenAI · Anthropic · ChatGPTWatch →The World of AI Agents Has Already Arrived: A Model Makes Discoveries, Hires People, and Opens New Paths for DeceptionChatGPT · OpenAI · GeminiWatch →AI Agents Are Becoming a Systemic Force: One Error Now Travels Through Code, Money, and the Physical WorldApple · OpenAI · GeminiWatch →AI Was Wrong, but the System Reacted as Though It Were Right: This Is How a False Alarm Becomes DangerousArtificial intelligence · United States · ChatGPTWatch →Mission Genesis Turns AI Into a Government Project Where Electricity Is the Main ResourceGoogle · OpenAI · ChatGPTWatch →Meta’s Moderation Failure Shows That Friendly AI Can Be Dangerous Precisely Because People Trust ItMeta · China · GoogleWatch →ChatGPT-5 Arrived With a New Problem: Your AI Conversations Can Become EvidenceOpenAI · GPT-5 · Elon MuskWatch →Meta’s “Personal Superintelligence” Sounds Good, but Users Need an Assistant That Works TodayOpenAI · Artificial intelligence · MicrosoftWatch →The New York Times Lawsuit Shows How Much Data AI Retains—and How Little Control the User HasOpenAI · ChatGPT · Artificial intelligenceWatch →There Is No Single Strongest Model: ChatGPT Has to Be Chosen Again for Every TaskOpenAI o3 · Google · OpenAIWatch →AI Search, Grok in Telegram, and Stargate: The Market Is Forming Around a New IntermediaryGoogle · OpenAI · GrokWatch →Llama 4 Is Meta’s Weapon Not Because of Its Size, but Because It Can Be Built Into AnythingTikTok · Open source · OpenAIWatch →AI Learned to Reason—and Learned to Hide What Happens Inside More Effectively at the Same TimeArtificial intelligence · ChatGPT · InstagramWatch →AI Is Moving Closer to People—and Reaching Too Far Into Their Lives at the Same TimeAmazon · Artificial intelligence · United StatesWatch →AI Safety Is Not a Fight Against an “Evil Model,” but a Fight Over the Right of People and States to Set Its RulesAnthropic · AI safety · Artificial intelligenceWatch →How to Use AI Every Day Without Confusing an Assistant With a Doctor, Lawyer, or FriendChatGPT · Artificial intelligence · United StatesWatch →