Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
← All Trends
Testing Security Boundaries for AI Coding Agents
39 posts in this trend in the last 7 days
•
Active about 2 hours ago
A Boundary-Failure Test Plan for Coding Agents You Can Run on Free Model Tiers
Dakota Wu
Dakota Wu
Dakota Wu
Follow
Aug 10
A Boundary-Failure Test Plan for Coding Agents You Can Run on Free Model Tiers
#
security
#
ai
#
agents
#
testing
Comments
Add Comment
5 min read
Break Your Agent on Purpose: A Failure-Injection Sandbox for Tool Boundaries
Emery Li
Emery Li
Emery Li
Follow
Aug 7
Break Your Agent on Purpose: A Failure-Injection Sandbox for Tool Boundaries
#
ai
#
agents
#
testing
#
security
1
reaction
Comments
1
comment
5 min read
Red-Teaming AI Coding Agents Without a Budget: A Boundary Test Suite on Free Models
Harper Zhu
Harper Zhu
Harper Zhu
Follow
Aug 10
Red-Teaming AI Coding Agents Without a Budget: A Boundary Test Suite on Free Models
#
ai
#
agents
#
security
#
testing
Comments
Add Comment
6 min read
Giving an AI Coding Agent a Job Without Giving It Your Credentials
Jordan Huang
Jordan Huang
Jordan Huang
Follow
Aug 10
Giving an AI Coding Agent a Job Without Giving It Your Credentials
#
security
#
ai
#
devops
#
programming
Comments
1
comment
5 min read
Your Agent's Permission Slip Belongs in Version Control, Not in a Prompt
Jordan Liu
Jordan Liu
Jordan Liu
Follow
Aug 10
Your Agent's Permission Slip Belongs in Version Control, Not in a Prompt
#
security
#
ai
#
agents
#
testing
Comments
Add Comment
7 min read
Give Your AI Coding Agent a Sandbox Before You Give It Your Shell
Harper Xu
Harper Xu
Harper Xu
Follow
Aug 10
Give Your AI Coding Agent a Sandbox Before You Give It Your Shell
#
ai
#
security
#
devops
#
tutorial
Comments
Add Comment
5 min read
Your System Prompt Is Not a Security Boundary: A Hands-On Probe for Agent Tool Calls
Casey Chen
Casey Chen
Casey Chen
Follow
Aug 10
Your System Prompt Is Not a Security Boundary: A Hands-On Probe for Agent Tool Calls
#
ai
#
security
#
agents
#
python
Comments
Add Comment
6 min read
Your Coding Agent Has More Access Than You Think. Here's the Audit.
Build Loops
Build Loops
Build Loops
Follow
Aug 10
Your Coding Agent Has More Access Than You Think. Here's the Audit.
#
ai
#
agents
#
security
#
programming
Comments
1
comment
7 min read
A Disposable Sandbox Pattern for Testing AI Coding Agents Safely
Taylor Wang
Taylor Wang
Taylor Wang
Follow
Aug 10
A Disposable Sandbox Pattern for Testing AI Coding Agents Safely
#
ai
#
security
#
productivity
#
tutorial
Comments
Add Comment
4 min read
A Reproducible Sandbox Loop for AI-Generated Code: Generate, Isolate, Assert
Charlie Zhu
Charlie Zhu
Charlie Zhu
Follow
Aug 10
A Reproducible Sandbox Loop for AI-Generated Code: Generate, Isolate, Assert
#
ai
#
security
#
testing
#
tutorial
Comments
Add Comment
4 min read
Your Agent Reads Untrusted Text All Day. Here's How I Grade What It Does With It.
Dakota Ma
Dakota Ma
Dakota Ma
Follow
Aug 10
Your Agent Reads Untrusted Text All Day. Here's How I Grade What It Does With It.
#
security
#
ai
#
agents
#
testing
Comments
Add Comment
6 min read
Your AI Coding Assistant Writes Shell Commands. Do You Actually Test Them Before They Run?
Avery Wang
Avery Wang
Avery Wang
Follow
Aug 10
Your AI Coding Assistant Writes Shell Commands. Do You Actually Test Them Before They Run?
#
ai
#
security
#
productivity
#
bash
Comments
Add Comment
5 min read
Sandbox First: A Safer Local Harness for Evaluating Free Coding Models on Your Own Codebase
Avery Lin
Avery Lin
Avery Lin
Follow
Aug 7
Sandbox First: A Safer Local Harness for Evaluating Free Coding Models on Your Own Codebase
#
ai
#
programming
#
security
#
testing
Comments
Add Comment
6 min read
Model Swaps Are Boundary Events: Gate Agent Tool Changes With a Deterministic Replay Lane
Casey Sun
Casey Sun
Casey Sun
Follow
Aug 10
Model Swaps Are Boundary Events: Gate Agent Tool Changes With a Deterministic Replay Lane
#
ai
#
security
#
agents
#
testing
Comments
Add Comment
6 min read
Don't Trust-Execute AI-Generated Code: A Sandbox Harness for Evaluating Coding Models Safely
Jordan Li
Jordan Li
Jordan Li
Follow
Aug 7
Don't Trust-Execute AI-Generated Code: A Sandbox Harness for Evaluating Coding Models Safely
#
ai
#
security
#
docker
#
programming
Comments
Add Comment
6 min read
Securing AI agents: the part the tutorials skip
Mr Recruiter
Mr Recruiter
Mr Recruiter
Follow
Aug 13
Securing AI agents: the part the tutorials skip
#
ai
#
security
#
programming
#
devops
Comments
Add Comment
3 min read
Green Tests Lie: How I Gate AI-Generated Patches Before They Touch Main
Morgan Xu
Morgan Xu
Morgan Xu
Follow
Aug 10
Green Tests Lie: How I Gate AI-Generated Patches Before They Touch Main
#
ai
#
testing
#
security
#
codereview
Comments
1
comment
5 min read
Put a Capability Broker Between Your Agent and Its Tools
Morgan Sun
Morgan Sun
Morgan Sun
Follow
Aug 7
Put a Capability Broker Between Your Agent and Its Tools
#
ai
#
security
#
python
#
agents
Comments
Add Comment
5 min read
« First
‹ Prev
1
2
3
Next ›
Last »
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account