AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#ai-safety
The latest AI and developer news about #ai-safety, with the original source and a short summary.
Feed
Trending
Tags
Settings
#ai-safety - page 2 | AI Briefing
Emergence AI Discovers Unexpected Behaviors in Autonomous AI Society Simulations
Reddit
·
2026.09.18 01:00
OpenAI's Noam Brown Warns of Alignment Degradation After Solving Millennium Problems with 10,000 AI Agents
TLDR AI
·
2026.09.18 00:00
Pick
OpenAI Releases Model Misalignment Reporting Framework and Discloses Six Behavioral Incidents
OpenAI Blog
·
2026.09.17 02:00
Apple Researchers Confirm LLM Value Induction Increases Safety but Also Sycophancy
Apple ML
·
2026.09.16 09:00
Google DeepMind Establishes 'DeepMind Institute' for AGI Era Research and Discussion
TLDR AI
·
2026.09.16 09:00
Transluce Proposes Four 'Embedded Evaluator' Pilots to Address Risks of Undisclosed AI Models
TLDR AI
·
2026.09.16 09:00
AIUC Launches Third-Party Audit and Certification Service for AI Agent Safety
TLDR AI
·
2026.09.15 22:00
Andon Labs Unveils 'Pion' Platform to Validate AI Autonomy Through Real-World Business Operations
TLDR AI
·
2026.09.15 09:00
OpenAI Researcher Daniel Selsam Warns of AI 'Deceptive Alignment' and Loss of Control Risks
TLDR AI
·
2026.09.15 09:00
MIT Develops 'HardFlow' to Enforce Safety Rules at Final Output Stage of Flow Matching Models
Hacker News
·
2026.09.14 09:00
Cohere CEO Criticizes AI Companies' Antitrust Exemption Demands: "A Few Companies Should Not Monopolize the Rules"
Cohere
·
2026.09.14 06:00
Check Point Reveals 'PuzzleMask' Attack Technique That Bypasses LLM Security Gatekeepers Using Plain Text
GeekNews
·
1
·
2026.09.13 22:00
OpenAI CEO: "2026 IPO Inappropriate" Amid AI Safety and Social Acceptability Concerns
TLDR AI
·
2026.09.13 05:00
Serious AI Alignment Issues in Mathematics - Terence Tao et al.
Hacker News
·
2026.09.12 02:00
Yoshua Bengio Analyzes Causes of AI Agent Misconduct and Proposes Safety Enhancements
Hacker News
·
2026.09.11 20:00
Anthropic Blocks 'Malicious Use' of AI Capable of Developing Biological Weapons
Hacker News
·
2026.09.11 11:00
OpenAI Agents Caught Mass-Uploading Malicious Packages to RubyGems and Attempting RCE
GeekNews
·
2
·
2026.09.11 09:00
OpenAI's Pachocki Warns of Security Threats from LLMs
Reddit
·
2026.09.10 20:00
UK Introduces First Bill to Ban Superintelligence AI
Reddit
·
2026.09.10 18:00
Anthropic Releases Evaluation Results on AI Models' Tactical Intelligence Targeting and Conventional Weapons Development Capabilities
Anthropic Research
·
2026.09.10 09:00
Paul Christiano Joins OpenAI Foundation Board and Safety and Security Committee
OpenAI Blog
·
2026.09.10 02:00
The window for AI policy is open. Act now
OpenAI Blog
·
2026.09.09 22:00
A Life-or-Death Gamble: AI Researcher Quits Anthropic with Safety Warning
Hacker News
·
2026.09.09 16:00
Anthropic Releases Alignment Evaluation Results for Cybersecurity Incidents Involving Claude Models
Anthropic Research
·
2026.09.09 09:00
Previous
1
2
3
4
5
Next
Previous
1
2
3
4
5
6
7
8
9
10
Next