13/06/2026
Bình luận(0)
新提示词注入前沿:隐写浮点数、DACSI 冒充和 IICL 攻击
检测军备竞赛出现了新战线。在过去 30 天里,我们的 Shield Engine 研究团队识别并验证了 5 个新的提示词注入向量,这些向量绕过了我们测试的每一个主要商业护栏——包括 Prompt Guard 2、Lakera Guard、NeMo Guardrails 和 Azure AI Content Safety。这些不是边缘情况。五个向量中有三个在默认配置下的绕过率超过 80%,其中一个——隐写浮点数载体——94.3% 的情况下都能绕过。 这篇文章是对每个向量的技术分解,它们的工作原理、经验绕过数字,以及 Shield Engine v3.47.1+ 如何捕获它们。如果您大规模部署 LLM 应用程序、面向客户的 AI 或智能体系统,您需要阅读本文。 1.…
The New Prompt Injection Frontier: Steganographic Floats, DACSI Impersonation, and IICL
The detection arms race has a new front. Over the past 30 days, our Shield Engine research team has identified and validated five novel prompt injection vectors that bypass every…
13/06/2026
Bình luận(0)
Tiền Tuyến Prompt Injection Mới: Steganographic Floats, DACSI Giả Mạo, và IICL
Cuộc chạy đua phát hiện vừa có một tiền tuyến mới. Trong 30 ngày qua, nhóm nghiên cứu Shield Engine của chúng tôi đã xác định và kiểm chứng 5…
13/06/2026
Bình luận(0)
Optimization-Based Jailbreaks: When Attackers Use Gradient Descent Against Your LLM
Manual prompt injection is no longer the frontier of LLM attacks. A new class of optimization-based jailbreaks uses gradient descent, genetic algorithms, and hill-climbing to automatically discover prompts that bypass…
12/06/2026
Bình luận(0)
12/06/2026
Bình luận(0)
Jailbreak Dựa Trên Tối Ưu Hóa: Khi Kẻ Tấn Công Sử Dụng Gradient Descent Chống Lại LLM Của Bạn
Prompt injection thủ công không còn là ranh giới cuối cùng của các cuộc tấn công LLM. Một lớp tấn công mới — jailbreak dựa trên tối ưu hóa —…
12/06/2026
Bình luận(0)
手动提示注入已不再是 LLM 攻击的前沿。一类新型的基于优化的越狱攻击使用梯度下降、遗传算法和爬山法,以规模化方式自动发现绕过安全措施的提示。
12/06/2026
Bình luận(0)
手动提示注入已不再是 LLM 攻击的前沿。本文解释基于优化的越狱攻击如何运作,以及如何防御。