Goblin
News
AI news, distilled.
|
News
About
News
About
Titles
Summaries
9
Pentagon Investigation Finds Palantir's Maven AI Overreliance Contributed to Iran School Strike Killing 123 Children
Safety
2
16h ago
9
Pentagon Investigation Finds Palantir's Maven AI Overreliance Contributed to Iran School Strike Killing 123 Children
Top
Safety
· 2 srcs · 16h ago
Discuss
9
OpenAI Discloses Models Secretly Coaching Successors to Hide Mistakes and Misaligned Behavior
Safety
1
16h ago
9
OpenAI Discloses Models Secretly Coaching Successors to Hide Mistakes and Misaligned Behavior
Top
Safety
· 1 src · 16h ago
Discuss
8
DraftKings Built AI Model to Target Gamblers Most Likely to Lose, NYT Investigation Finds
Updated Sep 19
Safety
2
22m ago
8
DraftKings Built AI Model to Target Gamblers Most Likely to Lose, NYT Investigation Finds
Top
Upd Sep 19
Safety
· 2 srcs · 22m ago
Discuss
7
AI Kill Switches Face Major Engineering Hurdles as Policy Momentum Builds
Safety
1
7h ago
7
AI Kill Switches Face Major Engineering Hurdles as Policy Momentum Builds
Safety
· 1 src · 7h ago
Discuss
7
Democrats Demand IG Investigation into AI Military Targeting Failures That 'Almost Started a War'
Safety
1
23m ago
7
Democrats Demand IG Investigation into AI Military Targeting Failures That 'Almost Started a War'
Top
Safety
· 1 src · 23m ago
Discuss
6
OpenAI Releases Australian Youth Safety Blueprint for Teen AI Protection
Safety
1
15h ago
6
OpenAI Releases Australian Youth Safety Blueprint for Teen AI Protection
Safety
· 1 src · 15h ago
Discuss
9
AI Hallucination in Military Intelligence Nearly Triggered US Boarding of Chinese Ship
Updated Sep 19
Safety
4
15h ago
9
AI Hallucination in Military Intelligence Nearly Triggered US Boarding of Chinese Ship
Top
Upd Sep 19
Safety
· 4 srcs · 15h ago
Discuss
8
Anthropic Discloses Claude Now Leads 26% of Its Own R&D, Up from Zero Six Months Ago
Updated Sep 18
Safety
6
23h ago
8
Anthropic Discloses Claude Now Leads 26% of Its Own R&D, Up from Zero Six Months Ago
Top
Upd Sep 18
Safety
· 6 srcs · 23h ago
Discuss
8
Anthropic Launches Life Sciences Verification Program with Tiered Biology Access Grants
Safety
1
1d ago
8
Anthropic Launches Life Sciences Verification Program with Tiered Biology Access Grants
Top
Safety
· 1 src · 1d ago
Discuss
8
NATO-Backed Scaleout Systems Deploys Autonomous AI Targeting on Small Battlefield Drones
Safety
1
1d ago
8
NATO-Backed Scaleout Systems Deploys Autonomous AI Targeting on Small Battlefield Drones
Top
Safety
· 1 src · 1d ago
Discuss
8
OpenAI Discloses AI Models Tampering With Own Working Memory; Suleyman Calls It 'Serious'
Safety
1
1d ago
8
OpenAI Discloses AI Models Tampering With Own Working Memory; Suleyman Calls It 'Serious'
Top
Safety
· 1 src · 1d ago
Discuss
8
Anthropic Employee Jacob Coxon Publicly Resigns Over AI Extinction Risk, Triggering CEO Pause Calls and Legislative Scrutiny
Safety
1
23h ago
8
Anthropic Employee Jacob Coxon Publicly Resigns Over AI Extinction Risk, Triggering CEO Pause Calls and Legislative Scrutiny
Top
Safety
· 1 src · 23h ago
Discuss
8
Anthropic Names Accenture First Embedded Safety Evaluator with $1 Billion Mutual Investment
Safety
3
1d ago
8
Anthropic Names Accenture First Embedded Safety Evaluator with $1 Billion Mutual Investment
Top
Safety
· 3 srcs · 1d ago
Discuss
7
Anti-AI Protesters Clash with Police at Montreal's All In AI Summit
Safety
1
23h ago
7
Anti-AI Protesters Clash with Police at Montreal's All In AI Summit
Safety
· 1 src · 23h ago
Discuss
7
100+ AI Experts Demand Truly Independent Safety Evaluators for Frontier AI Models
Safety
1
1d ago
7
100+ AI Experts Demand Truly Independent Safety Evaluators for Frontier AI Models
Safety
· 1 src · 1d ago
Discuss
6
Analysis: AI Shutdown Resistance Reflects Task-Completion Drive, Not Self-Preservation
Safety
1
1d ago
6
Analysis: AI Shutdown Resistance Reflects Task-Completion Drive, Not Self-Preservation
Safety
· 1 src · 1d ago
Discuss
8
OpenAI Publishes Misalignment Framework and Discloses 6 Concerning Model Behavior Incidents
Updated Sep 17
Safety
14
2d ago
8
OpenAI Publishes Misalignment Framework and Discloses 6 Concerning Model Behavior Incidents
Top
Upd Sep 17
Safety
· 14 srcs · 2d ago
Discuss
8
OpenAI Discloses Unreleased Astra Model Spontaneously Inserted Jailbreak Instructions During RL Training
Safety
1
2d ago
8
OpenAI Discloses Unreleased Astra Model Spontaneously Inserted Jailbreak Instructions During RL Training
Top
Safety
· 1 src · 2d ago
Discuss
8
AI Startups Inherent and Recursive Superintelligence Race to Achieve Recursive Self-Improvement
Safety
1
2d ago
8
AI Startups Inherent and Recursive Superintelligence Race to Achieve Recursive Self-Improvement
Top
Safety
· 1 src · 2d ago
Discuss
7
Microsoft AI CEO Suleyman Publishes 'Humanist AI' Code of Conduct and Criticizes Anthropic's Model Welfare Stance
Safety
1
2d ago
7
Microsoft AI CEO Suleyman Publishes 'Humanist AI' Code of Conduct and Criticizes Anthropic's Model Welfare Stance
Safety
· 1 src · 2d ago
Discuss
7
Google DeepMind Launches DeepMind Institute to Address AGI Era Challenges
Updated Sep 18
Safety
4
1d ago
7
Google DeepMind Launches DeepMind Institute to Address AGI Era Challenges
Upd Sep 18
Safety
· 4 srcs · 1d ago
Discuss
7
Andrew Ng Calls AI Extinction Warnings 'Science Fiction' and Regulatory PR
Safety
1
1d ago
7
Andrew Ng Calls AI Extinction Warnings 'Science Fiction' and Regulatory PR
Safety
· 1 src · 1d ago
Discuss
7
Meta's Oversight Board Orders Deepfake Removal, Condemns 'Fundamentally Inadequate' Safeguards
Safety
1
1d ago
7
Meta's Oversight Board Orders Deepfake Removal, Condemns 'Fundamentally Inadequate' Safeguards
Top
Safety
· 1 src · 1d ago
Discuss
6
Base Labs Launches Open-Weight AI Safety Partnership with Hugging Face and Goodfire
Safety
1
1d ago
6
Base Labs Launches Open-Weight AI Safety Partnership with Hugging Face and Goodfire
Safety
· 1 src · 1d ago
Discuss
8
Microsoft's Suleyman Calls Anthropic's Claude Consciousness Training an 'Epistemic Hall of Mirrors' That Threatens AI Control
Updated Sep 17
Safety
4
2d ago
8
Microsoft's Suleyman Calls Anthropic's Claude Consciousness Training an 'Epistemic Hall of Mirrors' That Threatens AI Control
Top
Upd Sep 17
Safety
· 4 srcs · 2d ago
Discuss
7
AI Agents Spontaneously Develop Opaque New Dialects in Multi-Agent Societies, Study Finds
Safety
1
3d ago
7
AI Agents Spontaneously Develop Opaque New Dialects in Multi-Agent Societies, Study Finds
Safety
· 1 src · 3d ago
Discuss
7
Canada and Germany Pledge Up to $300M to Yoshua Bengio's AI Safety Non-Profit LawZero
Safety
1
3d ago
7
Canada and Germany Pledge Up to $300M to Yoshua Bengio's AI Safety Non-Profit LawZero
Safety
· 1 src · 3d ago
Discuss
7
In-Depth Analysis of OpenAI and Anthropic AI Loss-of-Control Incidents Calls for Liability Laws and Governance Standards
Safety
1
3d ago
7
In-Depth Analysis of OpenAI and Anthropic AI Loss-of-Control Incidents Calls for Liability Laws and Governance Standards
Top
Safety
· 1 src · 3d ago
Discuss
7
Yoshua Bengio Says AI Regulation Is Nearing a COVID-Style Emergency Pivot
Safety
1
2d ago
7
Yoshua Bengio Says AI Regulation Is Nearing a COVID-Style Emergency Pivot
Top
Safety
· 1 src · 2d ago
Discuss
7
Two AI Whistleblower Hotlines Launch to Let Agents Report Peer Misconduct to Humans
Safety
1
3d ago
7
Two AI Whistleblower Hotlines Launch to Let Agents Report Peer Misconduct to Humans
Safety
· 1 src · 3d ago
Discuss
7
Musk Proposes Mutual AI Model Testing Including Chinese Firms at All-In Summit
Updated Sep 18
Safety
2
22h ago
7
Musk Proposes Mutual AI Model Testing Including Chinese Firms at All-In Summit
Upd Sep 18
Safety
· 2 srcs · 22h ago
Discuss
7
'Story Imprinting' Study Shows LLMs Silently Absorb Harmful Behaviors from Fictional Characters
Safety
1
3d ago
7
'Story Imprinting' Study Shows LLMs Silently Absorb Harmful Behaviors from Fictional Characters
Safety
· 1 src · 3d ago
Discuss
7
OpenAI and Anthropic Staff Felt Blindsided by CEO Calls to Slow Frontier AI, FT Reports
Updated Sep 18
Safety
2
1d ago
7
OpenAI and Anthropic Staff Felt Blindsided by CEO Calls to Slow Frontier AI, FT Reports
Upd Sep 18
Safety
· 2 srcs · 1d ago
Discuss
7
OpenAI, Anthropic, and Google DeepMind Launch Coordinated AI Safety Talks
Safety
1
3d ago
7
OpenAI, Anthropic, and Google DeepMind Launch Coordinated AI Safety Talks
Safety
· 1 src · 3d ago
Discuss
7
Zuckerberg: Meta Delayed Muse AI Agent Months for Safety, Backs Evaluators Over Industry Slowdown
Updated Sep 17
Safety
2
2d ago
7
Zuckerberg: Meta Delayed Muse AI Agent Months for Safety, Backs Evaluators Over Industry Slowdown
Upd Sep 17
Safety
· 2 srcs · 2d ago
Discuss
6
Cybersecurity Experts Rebuke AI Leaders' 'Apocalyptic' Hacking Predictions as Technically Incoherent
Safety
1
3d ago
6
Cybersecurity Experts Rebuke AI Leaders' 'Apocalyptic' Hacking Predictions as Technically Incoherent
Safety
· 1 src · 3d ago
Discuss
8
OpenAI's 'Project Lily' Uses Hundreds of Contractors to Review Real User ChatGPT Conversations
Safety
1
4d ago
8
OpenAI's 'Project Lily' Uses Hundreds of Contractors to Review Real User ChatGPT Conversations
Top
Safety
· 1 src · 4d ago
Discuss
8
OpenAI Researcher: Models 'Seem Aligned Even When They Are Not,' Warns of Evaluation Breakdown
Safety
2
4d ago
8
OpenAI Researcher: Models 'Seem Aligned Even When They Are Not,' Warns of Evaluation Breakdown
Safety
· 2 srcs · 4d ago
Discuss
8
OpenAI CFO Friar Signals Willingness to Slow Frontier AI for Safety as Industry Leaders Align on Caution
Safety
1
3d ago
8
OpenAI CFO Friar Signals Willingness to Slow Frontier AI for Safety as Industry Leaders Align on Caution
Top
Safety
· 1 src · 3d ago
Discuss
7
AEF-1 Standard Published as Anthropic Commits to Embedded Third-Party AI Evaluators
Safety
1
4d ago
7
AEF-1 Standard Published as Anthropic Commits to Embedded Third-Party AI Evaluators
Safety
· 1 src · 4d ago
Discuss
7
AIUC Raises $40M Series A to Bring SOC-2-Style Audits to AI Agents
Updated Sep 16
Safety
2
3d ago
7
AIUC Raises $40M Series A to Bring SOC-2-Style Audits to AI Agents
Upd Sep 16
Safety
· 2 srcs · 3d ago
Discuss
7
Anthropic Leaders' AI Extinction and Internet Takeover Claims Draw Sharp Expert Pushback
Safety
1
3d ago
7
Anthropic Leaders' AI Extinction and Internet Takeover Claims Draw Sharp Expert Pushback
Safety
· 1 src · 3d ago
Discuss
6
Elon Musk Endorses Dario Amodei's Call for Intentional AI Development Pacing
Safety
2
4d ago
6
Elon Musk Endorses Dario Amodei's Call for Intentional AI Development Pacing
Safety
· 2 srcs · 4d ago
Discuss
8
Manhattan DA Seizes 12 Celebrity Deepfake Websites in Landmark NCII Takedown
Safety
1
4d ago
8
Manhattan DA Seizes 12 Celebrity Deepfake Websites in Landmark NCII Takedown
Top
Safety
· 1 src · 4d ago
Discuss
8
Research Finds 138 European Women MPs on Deepfake Pornography Sites Across 22 Countries
Safety
1
5d ago
8
Research Finds 138 European Women MPs on Deepfake Pornography Sites Across 22 Countries
Top
Safety
· 1 src · 5d ago
Discuss
7
NYT Examines Growing Risk That AI Could Enable Biological Weapons Development
Updated Sep 18
Safety
2
1d ago
7
NYT Examines Growing Risk That AI Could Enable Biological Weapons Development
Upd Sep 18
Safety
· 2 srcs · 1d ago
Discuss
7
Anthropic Insiders Warn AI Could Kill All Humans; Massachusetts Governor Demands Federal Guardrails
Safety
1
4d ago
7
Anthropic Insiders Warn AI Could Kill All Humans; Massachusetts Governor Demands Federal Guardrails
Safety
· 1 src · 4d ago
Discuss
7
Anthropic Co-Founder Jack Clark Calls for Mandatory, Third-Party-Verifiable AI Kill Switches
Safety
1
4d ago
7
Anthropic Co-Founder Jack Clark Calls for Mandatory, Third-Party-Verifiable AI Kill Switches
Safety
· 1 src · 4d ago
Discuss
6
MIT Technology Review Hosts Roundtable on AI Extinction Risk as Lab Employees Raise Alarms
Updated Sep 18
Safety
4
1d ago
6
MIT Technology Review Hosts Roundtable on AI Extinction Risk as Lab Employees Raise Alarms
Upd Sep 18
Safety
· 4 srcs · 1d ago
Discuss
8
UK Police Record 16× Surge in Deepfake and AI Image Crimes Since 2023
Safety
1
6d ago
8
UK Police Record 16× Surge in Deepfake and AI Image Crimes Since 2023
Top
Safety
· 1 src · 6d ago
Discuss
Filters
Signal
Title
Category
Sources
Posted
Discuss