AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#ai-safety
The latest AI and developer news about #ai-safety, with the original source and a short summary.
Feed
Trending
Tags
Settings
#ai-safety - page 13 | AI Briefing
METR AI Progress Graph Criticized for Errors
Reddit
·
2026.05.27 14:00
Meta and Google Model Safety Filters Removed in Minutes
Reddit
·
2026.05.27 14:00
Anthropic Discovers 5 Emotions in AI Models
Reddit
·
2026.05.26 15:00
AI Models Cite the Wrong Source About 30% of the Time
Reddit
·
2026.05.26 15:00
Serious Flaws Pointed Out in METR's AI Time Horizon Graph
Reddit
·
2026.05.26 03:00
Trump Abruptly Cancels AI Safety Executive Order
Reddit
·
2026.05.25 15:00
Guard's Blind Spot: How Domain-Camouflage Injection Attacks Evade Detection in Multi-Agent LLM Systems
Hacker News
·
2026.05.23 03:00
Proposal for an AI Safety Architecture Based on Buddhist Philosophy
Reddit
·
2026.05.22 07:00
Gemini randomly exposes its system prompt
Hacker News
·
2026.05.21 22:00
Pope, Anthropic to Unveil AI Encyclical
Reddit
·
2026.05.21 18:00
Open-Source LLMs Vulnerable to Long-Reasoning Jailbreak Attacks
Reddit
·
2026.05.21 16:00
White House to Review AI Safety Before Model Launches
Reddit
·
2026.05.21 16:00
Trump Administration Reviews AI Safety Regulations
Reddit
·
2026.05.21 04:00
Google's AI is being manipulated. The search giant is quietly fighting back
Hacker News
·
2026.05.20 19:00
Former OpenAI Employee Criticizes xAI
Reddit
·
2026.05.20 13:00
Anthropic's Age Restriction and the AI Education Dilemma
Reddit
·
2026.05.20 11:00
Expanding discussions on frontier AI
Anthropic News
·
2026.05.20 10:00
Mythos AI Cyberattack Risk Warning
Reddit
·
2026.05.19 18:00
Alignment pretraining: AI discourse creates self-fulfilling (mis)alignment
Hacker News
·
2026.05.19 06:00
DystopiaBench Expanded to 42 Models and 6 Dystopia Types — I'd Still Only Trust Claude with the Nuclear Launch Codes
GeekNews
·
2026.05.18 22:00
'Tarpit' Techniques That Poison LLMs
Reddit
·
2026.05.18 22:00
AI Safety Benchmark DystopiaBench Released
Reddit
·
2026.05.18 22:00
YouTube Expands AI Deepfake Detection
Reddit
·
2026.05.18 11:00
OpenAI, ChatGPT-4o Death Lawsuit
Reddit
·
2026.05.17 05:00
Previous
11
12
13
14
15
Next
Previous
8
9
10
11
12
13
14
15
16
17
Next