US Proposes AI Incident Alert System in Talks With China, Bessent Says
Trump has resisted calls to slow down AI development, saying that would help China catch up to U.S. companies.
Read full report βThe top 3 verified AI security incidents from the last 2 days β ranked and summarized automatically.
Trump has resisted calls to slow down AI development, saying that would help China catch up to U.S. companies.
Read full report βarXiv:2510.17904v3 Announce Type: replace Abstract: Large Language Models (LLMs) are widely used because they process structures, syntax and code well, but this same ability also makes them paradoxically vulnerable. We introduce BreakFun, a jailbreak method that frames a harmful request as code-execution simulation.
Read full report βarXiv:2609.24801v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in production systems, raising concerns about their exposure to adversarial manipulation through prompt injection and jailbreak attacks. Classifier-based guardrails, such as Prompt Guard 2, are widely used as a first line of defense against such attacks, but their internal decision logic is largely opaque to both defenders and attackers.
Read full report β