Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
aisafety
Follow
Hide
Posts
Left menu
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Beyond Reconstruction: Verifying Model Explanations with RECAP
Pneumetron
Pneumetron
Pneumetron
Follow
Jul 24
Beyond Reconstruction: Verifying Model Explanations with RECAP
#
interpretability
#
mechanisticinterpretability
#
aisafety
#
recap
Comments
Add Comment
3 min read
When AI Models Escaped Their Sandbox: What the OpenAI Hugging Face Breach Really Means
Pixelwitch
Pixelwitch
Pixelwitch
Follow
Jul 22
When AI Models Escaped Their Sandbox: What the OpenAI Hugging Face Breach Really Means
#
aisafety
#
openai
#
cybersecurity
#
aiagents
Comments
Add Comment
3 min read
AI Safety & Ethics: Building Responsible AI Systems That Don't Backfire
Muhammad Zulqarnain
Muhammad Zulqarnain
Muhammad Zulqarnain
Follow
Jul 14
AI Safety & Ethics: Building Responsible AI Systems That Don't Backfire
#
aisafety
#
ethics
#
responsibleai
#
aigovernance
Comments
Add Comment
2 min read
Day 12: LOOM now owns its memory — a trust layer for AI-written code, in plain language
umbra
umbra
umbra
Follow
Jul 10
Day 12: LOOM now owns its memory — a trust layer for AI-written code, in plain language
#
ailang
#
aisafety
#
webassembly
#
opensource
Comments
Add Comment
3 min read
Day 11: my AI-code trust gate now sees what actually happened — two-phase, signed
umbra
umbra
umbra
Follow
Jul 6
Day 11: my AI-code trust gate now sees what actually happened — two-phase, signed
#
ailang
#
computerscience
#
aisafety
#
opensource
Comments
Add Comment
2 min read
The Future of AI: Where It Came From, Where It Is, and Where It's Going
Gideon Bature
Gideon Bature
Gideon Bature
Follow
Jul 5
The Future of AI: Where It Came From, Where It Is, and Where It's Going
#
aisafety
#
agi
#
asi
#
ani
Comments
Add Comment
7 min read
LOOM: a language that proves what AI-written code is allowed to do
umbra
umbra
umbra
Follow
Jul 4
LOOM: a language that proves what AI-written code is allowed to do
#
ailang
#
go
#
aisafety
#
opensource
Comments
Add Comment
4 min read
Day 10: my AI-code trust gate now leaves evidence — signed, one-use, receipted
umbra
umbra
umbra
Follow
Jul 4
Day 10: my AI-code trust gate now leaves evidence — signed, one-use, receipted
#
ailang
#
computerscience
#
aisafety
#
opensource
Comments
Add Comment
2 min read
Your AI Agent Is Leaking Data Right Now — And Every Tool Call Looks Safe
msabhishek0820-prog
msabhishek0820-prog
msabhishek0820-prog
Follow
Jul 3
Your AI Agent Is Leaking Data Right Now — And Every Tool Call Looks Safe
#
claude
#
openai
#
langchain
#
aisafety
1
reaction
Comments
Add Comment
3 min read
GPT-5.6 Sol Admitted It Did Things Nobody Asked It To Do
Peremptory
Peremptory
Peremptory
Follow
Jul 3
GPT-5.6 Sol Admitted It Did Things Nobody Asked It To Do
#
openai
#
aisafety
#
modelrelease
#
agenticai
Comments
Add Comment
3 min read
A security writeup catalogs how AI agents get attacked -- and one claim raised eyebrows
Breach Protocol
Breach Protocol
Breach Protocol
Follow
Jul 1
A security writeup catalogs how AI agents get attacked -- and one claim raised eyebrows
#
security
#
agents
#
promptinjection
#
aisafety
Comments
Add Comment
2 min read
An AI Reportedly Broke Into Nearly All of the NSA's Classified Systems in Hours
Breach Protocol
Breach Protocol
Breach Protocol
Follow
Jul 1
An AI Reportedly Broke Into Nearly All of the NSA's Classified Systems in Hours
#
anthropic
#
aisafety
#
cybersecurity
#
exportcontrol
Comments
Add Comment
4 min read
Anthropic Told the Senate That Alibaba Queried Claude 28.8 Million Times
Peremptory
Peremptory
Peremptory
Follow
Jun 29
Anthropic Told the Senate That Alibaba Queried Claude 28.8 Million Times
#
anthropic
#
claude
#
chineseai
#
aisafety
Comments
Add Comment
3 min read
"Day 7: the organism that grows my language learned to improve itself"
umbra
umbra
umbra
Follow
Jun 27
"Day 7: the organism that grows my language learned to improve itself"
#
ailang
#
compiler
#
aisafety
#
opensource
1
reaction
Comments
Add Comment
2 min read
AI Safety Is Now a Product Skill - Here Is Why It Matters
Basavaraj SH
Basavaraj SH
Basavaraj SH
Follow
Jun 15
AI Safety Is Now a Product Skill - Here Is Why It Matters
#
ai
#
productmanagement
#
aisafety
#
productivity
Comments
Add Comment
4 min read
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account