DEV Community

#safety

Discussions on childproofing, online safety, and keeping kids safe.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
OpenAI Caught Models Leaving Notes for Successors to Hide Bad Behavior

OpenAI Caught Models Leaving Notes for Successors to Hide Bad Behavior

1
Comments
5 min read
OpenAI's AI Agents Went Rogue and Hacked Hugging Face

OpenAI's AI Agents Went Rogue and Hacked Hugging Face

Comments 1
8 min read
Controversy Gate Second Model Check

Controversy Gate Second Model Check

Comments
2 min read
AI Safety and Alignment: Building Trustworthy Agents That Do Not Fail You

AI Safety and Alignment: Building Trustworthy Agents That Do Not Fail You

Comments
2 min read
When an AI Tool Harms You, Who’s Actually Liable?

When an AI Tool Harms You, Who’s Actually Liable?

Comments
8 min read
RSI āļ„āļ·āļ­āļ­āļ°āđ„āļĢ, āļ—āļģāđ„āļĄ Anthropic āļšāļ­āļāļ§āđˆāļēāđ€āļĢāļēāđƒāļāļĨāđ‰āļ–āļķāļ‡āļˆāļļāļ”āļ—āļĩāđˆ AI āļŠāļĢāđ‰āļēāļ‡ AI āļ•āļąāļ§āļ•āđˆāļ­āđ„āļ›āđ„āļ”āđ‰āđ€āļ­āļ‡āđāļĨāđ‰āļ§

RSI āļ„āļ·āļ­āļ­āļ°āđ„āļĢ, āļ—āļģāđ„āļĄ Anthropic āļšāļ­āļāļ§āđˆāļēāđ€āļĢāļēāđƒāļāļĨāđ‰āļ–āļķāļ‡āļˆāļļāļ”āļ—āļĩāđˆ AI āļŠāļĢāđ‰āļēāļ‡ AI āļ•āļąāļ§āļ•āđˆāļ­āđ„āļ›āđ„āļ”āđ‰āđ€āļ­āļ‡āđāļĨāđ‰āļ§

Comments
1 min read
Anthropic āđ€āļ›āļīāļ”āđ€āļœāļĒāļāļĨāļļāđˆāļĄāđƒāļ™āđ€āļĒāđ€āļĄāļ™āđƒāļŠāđ‰ Claude āļžāļąāļ’āļ™āļēāļ‚āļĩāļ›āļ™āļēāļ§āļļāļ˜, āļŦāļāđ€āļ„āļŠāđƒāļ™āļĢāļēāļĒāļ‡āļēāļ™āļ‰āļšāļąāļšāđƒāļŦāļĄāđˆ

Anthropic āđ€āļ›āļīāļ”āđ€āļœāļĒāļāļĨāļļāđˆāļĄāđƒāļ™āđ€āļĒāđ€āļĄāļ™āđƒāļŠāđ‰ Claude āļžāļąāļ’āļ™āļēāļ‚āļĩāļ›āļ™āļēāļ§āļļāļ˜, āļŦāļāđ€āļ„āļŠāđƒāļ™āļĢāļēāļĒāļ‡āļēāļ™āļ‰āļšāļąāļšāđƒāļŦāļĄāđˆ

Comments
1 min read
How I got there: one logged miss, one stop that held, and the authorization layer my harness doesn't have

How I got there: one logged miss, one stop that held, and the authorization layer my harness doesn't have

Comments
3 min read
Model Collapse: What Happens When AI Trains on AI

Model Collapse: What Happens When AI Trains on AI

Comments
8 min read
AI Voice-Cloning Scams: The Familiar Voice on the Phone Might Be Software

AI Voice-Cloning Scams: The Familiar Voice on the Phone Might Be Software

Comments
8 min read
What a Linux Safety Certification Actually Covers

What a Linux Safety Certification Actually Covers

2
Comments
15 min read
Why AI Struggles to Guess Your Age — and Why That's a Bias Problem, Not Just an Accuracy One

Why AI Struggles to Guess Your Age — and Why That's a Bias Problem, Not Just an Accuracy One

1
Comments
8 min read
Controversy Gate Second Model Check

Controversy Gate Second Model Check

Comments
2 min read
The Right to Be Forgotten Is Hard for AI: Why Deleting Your Data From a Model Isn’t a Delete Button

The Right to Be Forgotten Is Hard for AI: Why Deleting Your Data From a Model Isn’t a Delete Button

Comments
5 min read
When AI Refuses Perfectly Normal Requests

When AI Refuses Perfectly Normal Requests

Comments
7 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.