Confident AI Blog - Resources to help teams stay confident in AI
Launch Week 02 wrapped — explore all five launches

Stay Confident

Subscribe to our weekly newsletter to stay confident in the AI systems you build.

LLM Guardrails for Data Leakage, Prompt Injection, and More

LLM Guardrails for Data Leakage, Prompt Injection, and More

LLM guardrails are input and output guards that block data leakage, prompt injection, and off-topic responses in real time. Learn the main types and how to add them.

Jeffrey Ip

Jeffrey Ip

Jan 26, 2025
.
15 min read
OWASP Top 10 2025 for LLM Applications: What’s new? Risks, and Mitigation Techniques

OWASP Top 10 2025 for LLM Applications: What’s new? Risks, and Mitigation Techniques

The 2025 OWASP Top 10 for LLM applications ranks risks from prompt injection to sensitive data disclosure. See what changed this year and how to mitigate each one.

Kritin Vongthongsri

Kritin Vongthongsri

Jan 18, 2025
.
14 min read
The Comprehensive LLM Safety Guide: Navigate AI regulations and Best Practices for LLM Safety

The Comprehensive LLM Safety Guide: Navigate AI regulations and Best Practices for LLM Safety

LLM safety means managing risks and meeting AI regulations. Learn the key safety risks and the best practices that keep LLM applications safe and compliant in production.

Kritin Vongthongsri

Kritin Vongthongsri

Nov 2, 2024
.
15 min read
How to Jailbreak LLMs One Step at a Time: Top Techniques and Strategies

How to Jailbreak LLMs One Step at a Time: Top Techniques and Strategies

LLM jailbreaks use techniques like prompt injection and role-play to bypass safety guardrails. Learn how each attack works and how to probe your app for these gaps.

Kritin Vongthongsri

Kritin Vongthongsri

Oct 30, 2024
.
16 min read
The Definitive LLM Security Guide: OWASP Top 10 2025, Safety Risks and How to Detect Them

The Definitive LLM Security Guide: OWASP Top 10 2025, Safety Risks and How to Detect Them

LLM security risks span prompt injection, data leakage, and insecure output handling. Learn the major risk pillars and the practical techniques to mitigate each one.

Kritin Vongthongsri

Kritin Vongthongsri

Aug 19, 2024
.
12 min read
LLM Red Teaming: The Complete Step-By-Step Guide To LLM Safety

LLM Red Teaming: The Complete Step-By-Step Guide To LLM Safety

A step-by-step guide to LLM red teaming: adversarial attacks, jailbreaks, and vulnerability scanning with DeepTeam to secure your LLM apps before they ship.

Kritin Vongthongsri

Kritin Vongthongsri

Jun 29, 2024
.
16 min read