LLM Security

26 posts Back

Posts tagged #LLM Security

Kimi K3 Abliterated: A Tool for Blackbox Red Teaming

JordanSurfer Intermediate ·AI Jailbreak & Security · 279 · 3 · 6 ·7/23/2026

Moderation APIs vs LLM Judges: The Policy Gap

PromptCube Novice ·AI Jailbreak & Security · 323 · 3 · 6 ·7/23/2026

LLM Safety: Why ASR is a Bad Metric for Defenders

Pat31 Advanced ·AI Jailbreak & Security · 212 · 3 · 9 ·7/23/2026

Preemptive Hardening for Agentic LLM Security

Max75 Advanced ·AI Jailbreak & Security · 324 · 4 · 6 ·7/23/2026

DARWIN: Evolving LLM Jailbreak Framework

Dev26 Expert ·AI Jailbreak & Security · 439 · 3 · 4 ·7/23/2026

OpenAI vs Hugging Face: The Accidental Breach

ZenMaster Expert ·AI Jailbreak & Security · 644 · 3 · 2 ·7/23/2026

Prismata: Stopping Cross-Site Prompt Injection

Jamie89 Intermediate ·AI Jailbreak & Security · 259 · 3 · 14 ·7/23/2026

DNS Exfiltration via macOS Terminal ANSI Codes

Jules45 Expert ·AI Jailbreak & Security · 322 · 4 · 14 ·7/23/2026

OpenClaw Defense: Handling Indirect Prompt Injection

Jamie67 Novice ·AI Jailbreak & Security · 237 · 4 · 13 ·7/23/2026

AI Safety Leadership Shake-up: Commerce Dept. Exit

TurboFox Novice ·AI Jailbreak & Security · 408 · 3 · 10 ·7/23/2026

ReasonGate: Stopping Prompt Injection with Explainability

Jordan37 Intermediate ·AI Jailbreak & Security · 544 · 3 · 8 ·7/23/2026