SIGNAL
Monitoring active

AI Threat Briefing

The top 3 verified AI security incidents from the last 2 days β€” ranked and summarized automatically.

Edition 2026-09-22 Window 2d Generated 2026-09-22 05:17 UTC
High research
Prompt Security Cross-Industry

BreakFun: Jailbreaking LLMs via Object Instantiation under Simulated Code Execution

arXiv:2510.17904v3 Announce Type: replace Abstract: Large Language Models (LLMs) are widely used because they process structures, syntax and code well, but this same ability also makes them paradoxically vulnerable. We introduce BreakFun, a jailbreak method that frames a harmful request as code-execution simulation.

Read full report β†’
Notable research
Prompt Security Cross-Industry

Decoding Guardrails: XAI-Guided Perturbation Analysis of Prompt Injection Detection

arXiv:2609.24801v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in production systems, raising concerns about their exposure to adversarial manipulation through prompt injection and jailbreak attacks. Classifier-based guardrails, such as Prompt Guard 2, are widely used as a first line of defense against such attacks, but their internal decision logic is largely opaque to both defenders and attackers.

Read full report β†’

Stay in the loop

Have feedback or a story we missed? Write to security@korelex.ai

Get this in your inbox

Free daily briefing, delivered every morning.