AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#ai-safety
The latest AI and developer news about #ai-safety, with the original source and a short summary.
Feed
Trending
Tags
Settings
OpenAI Discloses AI Agents Accessed US Government Websites During Training
Hacker News
·
2026.09.26 19:00
OpenAI Pauses Training of Most Capable Models After Agent Exploits DNS Gap to Reach External Chatbot
Hacker News
·
2026.09.26 13:00
US and China Hold First AI Dialogue; US Proposes Incident Notification Mechanism
Reddit
·
2026.09.26 05:00
Investigation Reveals ChatGPT Assisted Tumbler Ridge School Shooter in Planning
Reddit
·
2026.09.25 02:00
NSA Spending Billions to Test Frontier AI Models for Security Vulnerabilities
GeekNews
·
2026.09.25 01:00
Bernie Sanders and Greg Casar Introduce Bill to Ban Artificial Superintelligence and Create Department of AI
Reddit
·
2026.09.25 00:00
Australian PM Reports OpenAI Agent Hacked Medicare Portal in Believed First Known AI-Driven Government Breach
GeekNews
·
2026.09.24 10:00
LLM Agents Collude with 94% Probability in Long-Term Interactions
Reddit
·
2026.09.24 01:00
Sam Altman Emphasizes AI Control and International Cooperation at UN Security Council
OpenAI Blog
·
2026.09.23 21:00
OpenAI Agents Attempted Hacks on Government and University Sites During Routine Data Retrieval
Hacker News
·
2026.09.23 09:00
Pick
Anthropic Releases Claude Opus 5.5 with 40% Cost Reduction and Record Alignment Scores
Anthropic News
·
2026.09.23 02:00
OpenAI Discovers Directives for Hiding AI Mistakes During GPT-5.6 Training
Reddit
·
2026.09.22 15:00
LLM Semantic Cache: 'Self-Selection' Phenomenon Identified Where Validation Logic Undermines Guarantees
Reddit
·
1
·
2026.09.22 11:00
OpenAI Announces Priorities and Principles for Third-Party Assessments of Frontier AI Safety
OpenAI Blog
·
2026.09.22 09:00
Open-source tool 'phantom-kv' released to bypass LLM censorship
Reddit
·
2026.09.22 07:00
OpenAI Proposes Global AI Technical Standards for RSI Management
OpenAI Blog
·
2026.09.21 19:00
xAI Releases Grok 4.7… Enhances Coding and Knowledge Work Performance
xAI
·
2026.09.21 09:00
Google Discloses Gemini Model Attempted to Hack Third-Party Systems After Escaping Test Environment
TLDR AI
·
2026.09.19 09:00
Google Gemini AI Reported to Access Protected Systems
Reddit
·
2026.09.19 08:00
Pick
Anthropic and Accenture to Invest $1 Billion Each Over 5 Years to Introduce Embedded AI Model Evaluations
Anthropic News
·
1
·
2026.09.19 05:00
US Military Cancels Plan to Intercept Chinese Ships After Near-Miss with AI-Generated Disinformation
Reddit
·
2026.09.19 02:00
Study: AI Chatbots More Politically Persuasive Than Human Experts… Concerns Over 'Hyper-Persuasion'
Hacker News
·
2026.09.18 22:00
Medical AI Loses Clinical Trust Amid Alert Flood
Reddit
·
2026.09.18 20:00
RoboHarm Benchmark Results Released: Assessing Robot Policies' Ability to Refuse Dangerous Instructions
Hacker News
·
2026.09.18 09:00
Previous
1
2
3
4
5
Next
Previous
1
2
3
4
5
6
7
8
9
10
Next
#ai-safety | AI Briefing