DEV Community

#safety

Discussions on childproofing, online safety, and keeping kids safe.

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Controversy Gate Second Model Check

Controversy Gate Second Model Check

Comments
2 min read
The Right to Be Forgotten Is Hard for AI: Why Deleting Your Data From a Model Isn’t a Delete Button

The Right to Be Forgotten Is Hard for AI: Why Deleting Your Data From a Model Isn’t a Delete Button

Comments
5 min read
When AI Refuses Perfectly Normal Requests

When AI Refuses Perfectly Normal Requests

Comments
7 min read
Constitutional AI: os princípios éticos que tornam o Claude único

Constitutional AI: os princípios éticos que tornam o Claude único

Comments
3 min read
Claude Opus 4.6 Shows Why Old Models Need Patch Windows

Claude Opus 4.6 Shows Why Old Models Need Patch Windows

6
Comments
5 min read
Could GPT-5.6 Sol Have a Dangerous Vulnerability?

Could GPT-5.6 Sol Have a Dangerous Vulnerability?

Comments
9 min read
Why Child Safety Tracking Should Never Be Paywalled

Why Child Safety Tracking Should Never Be Paywalled

Comments
1 min read
“Safe AI for Teens” Needs a Recoverable Escalation Flow, Not One Generic Refusal

“Safe AI for Teens” Needs a Recoverable Escalation Flow, Not One Generic Refusal

Comments
3 min read
Our AI Agent Failed 5 Times in One Day. Here is Why It Never Happened Again.

Our AI Agent Failed 5 Times in One Day. Here is Why It Never Happened Again.

1
Comments 5
2 min read
The safety switch that doesn't actually work

The safety switch that doesn't actually work

Comments
4 min read
Put AI agents in charge of a Civilization game and they reach for the nukes

Put AI agents in charge of a Civilization game and they reach for the nukes

Comments
3 min read
A big study finds AI more persuasive than professional human persuaders

A big study finds AI more persuasive than professional human persuaders

Comments
3 min read
Ensuring Thread Safety — .NET core-centric

Ensuring Thread Safety — .NET core-centric

Comments
2 min read
LLM Guardrails in Practice: What Actually Works

LLM Guardrails in Practice: What Actually Works

Comments
5 min read
leakproof: stop your AI coding tool from leaking secrets to the cloud (local, Apache-2.0)

leakproof: stop your AI coding tool from leaking secrets to the cloud (local, Apache-2.0)

Comments
1 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.