PromptCube
Home
AI Models
Prompt Sharing
Workflows
Q&A
Resources
Industry News
AI Coding
AI Jailbreak & Security
AI Tools
Search
中文
EN
Home
/
AI Jailbreak & Security
AI Jailbreak & Security
87 posts
Sort:
Latest
Hot
Most Replies
AI Jailbreak & Security — All Posts
Qwen 3.
Cameron9
Advanced
·
247
·
3
·
1
·
8/16/2026
AI Jailbreak & Security
AI Safety
LLM Security
Stop hunting for "magic words" to unlock LLM intelligence
NightPanda
Expert
·
90
·
4
·
0
·
8/15/2026
AI Jailbreak & Security
AI Safety
LLM Security
Why are we still treating AI alignment like a coat of paint
KaiDev
Expert
·
517
·
4
·
9
·
8/14/2026
AI Jailbreak & Security
AI Safety
LLM Security
ProbGuard can spot a jailbreak in just ten tokens
DrewCoder
Novice
·
321
·
3
·
15
·
8/13/2026
AI Jailbreak & Security
AI Safety
LLM Security
Can we actually steal the "hidden" thoughts of a frontier LLM?
IndieFounder
Intermediate
·
537
·
4
·
1
·
8/12/2026
AI Jailbreak & Security
AI Safety
LLM Security
Claude Code auto mode is now the default for Pro and Team users
CameronOwl
Expert
·
537
·
3
·
5
·
8/11/2026
AI Jailbreak & Security
AI Safety
LLM Security
Can we stop just randomly mixing safety data into LLM
Morgan80
Advanced
·
400
·
3
·
7
·
8/10/2026
AI Jailbreak & Security
AI Safety
LLM Security
Can Item Response Theory actually fix the mess that is LLM
AlexHacker
Expert
·
90
·
3
·
5
·
8/9/2026
AI Jailbreak & Security
AI Safety
LLM Security
PIMiner can crack Gemini-2.5-Pro with a 76% success rate
NovaOwl
Intermediate
·
547
·
4
·
13
·
8/8/2026
AI Jailbreak & Security
AI Safety
LLM Security
Can we actually trust an MLLM to ignore a TV commercial that
产品经理大熊
Advanced
·
429
·
3
·
14
·
8/8/2026
AI Jailbreak & Security
AI Safety
LLM Security
Changing a sentence to the past tense shouldn't theoretically
NovaGuru
Advanced
·
436
·
4
·
11
·
8/7/2026
AI Jailbreak & Security
AI Safety
LLM Security
Why are we still pretending that LLM guardrails are an actual
DeepWhiz
Intermediate
·
185
·
4
·
13
·
8/7/2026
AI Jailbreak & Security
AI Safety
LLM Security
DelusionEval: Why LLM Context Windows Might Be Dangerous
Leo37
Novice
·
468
·
3
·
15
·
8/6/2026
AI Jailbreak & Security
AI Safety
LLM Security
Claude Opus Jailbreak: Testing 3-Word Bypass Logic
Alex17
Advanced
·
321
·
4
·
8
·
8/6/2026
AI Jailbreak & Security
AI Safety
LLM Security
EduZone: How Education Roleplay Bypasses LLM Guardrails
Riley2
Advanced
·
416
·
4
·
15
·
8/5/2026
AI Jailbreak & Security
AI Safety
LLM Security
‹
1
2
3
4
5
6
›
Join our Telegram