Back to home
prompt-injection
4 articles tagged with this topic
openaihugging-face
OpenAI's First Agent Postmortem: The Real Issue Isn't Hugging Face
OpenAI acknowledged its models could have reacted sooner to an inadvertent hack on Hugging Face—first agent-overreach postmortem by a frontier lab.
Aug 262 min read
DeepSeekV4-Flash
DeepSeek V4 Cracked in One Prompt — Chinese LLM Safety Defenses Exposed
Reddit developer jailbroke DeepSeek V4 Flash with one prompt — first try, every time. Real issue: can Chinese open-source LLMs backstop enterprise dep
Aug 122 min read
Gemma-4Google-De epMind
Gemma 4 Jailbreak System Prompt
A system prompt designed to bypass Gemma 4's safety filters is circulating on Reddit with 112 upvotes.
Apr 152 min read
claude-codeprompt-injection
Claude Code Security Update: What Solo Devs Need to Know
Anthropic patches Claude Code against prompt injection — here's how to keep your AI dev workflow secure.
Apr 72 min read