ByteByteGo
Read post

LLM Security Basics: The Full Threat Model

A structured threat model for LLM security, organized around the core property that language models process instructions and data as a single token sequence with no boundary between them. Covers the full attack surface mapped to pipeline stages: prompt injection (direct and indirect), RAG poisoning, model-level attacks (weight theft, training-data extraction, poisoning), excessive agency (the 'lethal trifecta'), and supply chain risks. Real incidents are used as anchors — EchoLeak (CVE-2025-32711), PoisonedRAG, nullifAI on Hugging Face, and GitHub/GitLab MCP compromises. Concludes that model-interior attacks are largely bounded and mitigated by providers, while the highest real-world risk sits where an agent simultaneously holds private data, untrusted content, and an external action channel. Defense in depth across all pipeline stages is the recommended posture, with no single filter being sufficient.

    #security#llm#ai-agents#prompt-injection
Aug 03•13m read time•From blog.bytebytego.com
Post cover image
Table of contents
Matic: A new era of visually intelligent cleaning (Sponsored)Trust BoundariesAttack SurfaceModel AttacksExcessive AgencySupply ChainDefense in DepthConclusion
962 Impressions
ByteByteGo's image
ByteByteGo

ByteByteGo provides tutorials, articles, and resources for learning and mastering the Go programming...

7.4K Followers

•

30.2K Upvotes

Would you recommend this post?

Copy link
WhatsApp
Facebook
X
New Squad
  • © 2026 Daily Dev Ltd.
  • Guidelines
  • Explore
  • Tags
  • Sources
  • Squads
  • Leaderboard