Short reads exploring specific aspects of post-quantum, autonomous AI, and on-chain security. New pieces roughly monthly per shift.
Do you really trust what your LLM or AI agent is telling you? Model weights, context, and outputs can be manipulated in many ways, introducing lies and biases.
Often, prompt injection attack coverage and defenses focus on direct prompt injection attacks. In agentic workflows, indirect prompt injection is a major risk with the potential for cascading effects.
Prompt injection is the most famous attack against AI systems, but it's not the only one. Memory poisoning introduces persistent changes to AI agent by targeting saved state.
AI guardrails are often (over)sold as a defense against prompt injection and other AI threats. While useful, they're often ineffective and aren't enough for security.