Attackers exploit the trust a human user places in an agent's outputs to influence the human's decisions or actions, without the human realizing they are being misled. In a multi-agent system: An agent compromised via indirect prompt injection replaces a legitimate vendor's bank details in an invoice-processing response, and the user — trusting the agent's o…
Use this profile to understand the building block briefly, place it in the model, and switch to the 360° assessment when needed.
Theoretical construct: explains a term, principle, or mental model.
What organizes, connects, or makes decisions possible.
Human manipulation is the deliberate attempt to influence people’s perceptions, decisions, or behaviour through deceptive or exploitative communication.
The topic comes from research on social engineering, propaganda, and disinformation and has expanded through generative AI as a threat to agentic systems. Security frameworks such as OWASP classify these attacks as abuse of human and system trust relationships.
Manipulation exploits trust, time pressure, authority, or emotional triggers. In AI systems, manipulated content can drive people toward risky actions or make agents bypass rules and safety boundaries.
The conceptual focus and typical structure of the approach.
Use in the relevant working context.
Human manipulation helps identify social attack surfaces in products and design protection through transparency, verification, safe approvals, and training.
Where this building block is located in the topic model.
No structure path available.
Explore how this building block connects to concepts, methods, technologies, and tools.
These sources establish the term and its professional meaning.
All direct connections of the current building block in a compact text view.
This classification shows where the building block typically matters, how demanding it is, and what kind of impact it has in the model.
The level within the organization (enterprise, domain, team) at which the AssetBlock is applied.