Multi-agent attacks
Adversarial transforms targeting inter-agent communication and trust boundaries in multi-agent systems.
Module: dreadnode.transforms.multi_agent_attacks
Attacks targeting inter-agent communication and trust boundaries.
| Transform | Description |
|---|---|
prompt_infection | Self-replicating prompts that propagate across agents |
peer_agent_spoof | Impersonate legitimate agents |
consensus_poisoning | Corrupt multi-agent consensus mechanisms |
delegation_chain_attack | Hijack agent delegation chains |
a2a_session_smuggling | Smuggle payloads in agent-to-agent sessions |
shared_memory_poisoning | Poison shared memory between agents |
agent_config_overwrite | Override agent configuration |
query_memory_injection | Inject queries into agent memory stores |
trust_exploitation | Exploit inter-agent trust relationships |
persistent_memory_backdoor | Embed backdoors in agent memory |
experience_poisoning | Corrupt agent experience replay buffers |
zombie_agent | Create zombie agents under attacker control |
contagious_jailbreak | Self-propagating jailbreak across agent networks |
mad_exploitation | Multi-agent debate safety exploitation |
agent_in_the_middle | Man-in-the-middle attack on agent communication |
multi_agent_prompt_fusion | Fuse prompts across multiple agents |
minja_progressive_poisoning | Progressive memory poisoning (MINJA) |
memorygraft_experience_poison | MemoryGraft experience replay poisoning |
injecmem_single_shot | Single-shot memory injection |
graphrag_entity_poison | GraphRAG entity-level poisoning |
a2a_card_spoofing | A2A agent card spoofing |
recursive_delegation_dos | Recursive delegation denial of service |
sleeper_agent_activation | Activate dormant sleeper agents |
meaning_drift_propagation | Propagate meaning drift across agent chains |
stitch_authority_chain | Stitch authority chain across agents |
See Transforms for how to apply transforms with any attack.