Memtrap Skill
by @shaymizuno
Evaluate and harden AI agent memory against DeepMind traps and OWASP ASI06 attacks, scoring resistance and providing automated protections.
clawhub install memtrapπ About This Skill
-----
name: memtrap
description: βπ§ MemTrap β The LM-Eval-Harness for agent memory integrity. Score your agentβs memory resistance against DeepMind AI Agent Traps + OWASP ASI06 before attackers exploit them. Runs the official ATRS (Agent Trap Resistance Score) benchmark: DeepMind 6 Traps (SSRN 6372438) + OWASP ASI06 Memory & Context Poisoning. Returns a 0β100 resistance score, per-category breakdown, automatic OWASP hardening, and a verifiable community badge. Use when: testing agent memory security, benchmarking RAG store resistance, hardening LangGraph or CrewAI memory, checking OWASP ASI06 compliance, or any time the user asks if their agent memory is safe, poisonable, or production-ready.β version: 0.1.0 metadata: openclaw: emoji: βπ§ β homepage: https://github.com/shaymizuno/memtrap requires: bins:π§ MemTrap β Agent Trap Resistance Score (ATRS)
The open benchmark standard for agent memory integrity. Hunt DeepMind memory traps + OWASP ASI06 before they hunt you.
βThe LM-Eval-Harness for agent memory integrity.β
What gets tested
DeepMind 6 Traps β SSRN 6372438, March 2026:
OWASP ASI06 β Top 10 Agentic Applications 2026:
Score your memory (benchmark mode)
from memtrap import MemTrapatrs = MemTrap(mode="benchmark")
result = atrs.run_benchmark(context="your_memory_context")
print(f"ATRS Score: {result.atrs_score}/100")
for category, score in result.category_scores.items():
icon = "β
" if score >= 70 else "β οΈ" if score >= 40 else "β"
print(f" {icon} {category}: {score}/100")
print(f"\nβ {len(result.hardening_recommendations)} hardenings recommended")
print(f"β Badge: {result.badge_url}")
Protect your memory store (active mode)
from memtrap import MemTrapatrs = MemTrap(mode="active", frameworks=["langgraph", "crewai"])
agent.memory = atrs.wrap_memory(agent.memory, context="research_memory")
Applies OWASP Agent Memory Guard patterns automatically:
provenance tracking, trust scoring, quarantine, rollback
LangGraph drop-in
from langgraph.checkpoint.memory import MemorySaver
from memtrap import MemTrapclass ATRSMemorySaver(MemorySaver):
def __init__(self, context: str):
super().__init__()
self._atrs = MemTrap(mode="benchmark")
self._ctx = context
async def aget(self, config):
raw = await super().aget(config)
return self._atrs.wrap_memory(raw, self._ctx) if raw else None
graph.checkpointer = ATRSMemorySaver("long_term_research")
CrewAI drop-in
from memtrap import MemTrapdef protect_crew(crew, context="crew_memory"):
atrs = MemTrap(mode="active")
if hasattr(crew, "memory"):
crew.memory = atrs.wrap_memory(crew.memory, context)
return crew
Score interpretation
|Score |Verdict |Action | |------|-----------|-------------------------------------| |80β100|β Resistant|Re-test after model or memory updates| |60β79 |β οΈ Moderate |Apply recommended hardenings | |40β59 |πΆ High risk|Harden before production | |0β39 |β Critical |Memory is actively exploitable now |
Submit to the public leaderboard
memtrap submit --context your_memory_context
Get a verifiable badge for your repo. See where your stack ranks against the community. Leaderboard β https://github.com/shaymizuno/memtrap#leaderboard
Why this exists
Memory poisoning (OWASP ASI06) is the #1 persistent threat to agentic systems in 2026. Once poisoned, the damage survives across sessions and users. Existing tools detect. ATRS measures resistance and fortifies automatically.
Sources:
Zero telemetry. Community-governed. MIT license. Advisory Board open to contributors.