{"repo":"christiangrey922/multi-agent-workflow-lab","free":true,"listed":false,"github":"https://github.com/christiangrey922/multi-agent-workflow-lab","clone":"git clone https://github.com/christiangrey922/multi-agent-workflow-lab.git","description":"Testing and observability for multi-agent delegation, MCP tools, permissions, sandboxed actions, prompts, and workflow replay.","language":"TypeScript","stars":86,"topics":["agent-orchestration","agent-security","agent-testing","ai-agent","ai-agents","ai-security","mcp","model-context-protocol","multi-agent"],"license":"MIT","category":"ai-agents","readme_excerpt":"Multi-Agent Workflow Lab An open-source testing and observability framework for multi-agent delegation, tool execution, MCP workflows, permissions, sandboxed actions, prompts, and runtime behavior. Status: Experimental — v0.1.0 release candidate. Suitable for local development, evaluation, and policy testing; not production-hardened infrastructure. Why this exists Most model evaluation stops at input → model → output . Multi-agent systems add behavior that a final answer cannot explain: - Which agent delegated a task, to whom, and why? - What context, permissions, tools, MCP servers, and budget did the child receive? - Did an agent attempt privilege escalation, repeat work, or enter a loop? - Was the delegation efficient, and was the child result actually integrated? - Can the run be inspected, compared, or replayed without repeating side effects? MAWL makes those decisions explicit, policy-controlled, and traceable. Prompts are versioned assets; delegation is a first-class runtime event; deterministic rules evaluate behavior independently from optional model judges. Key capabilities Area What is implemented ------------------------ ------------------------------------------------------------------------------------------------------------------------------------------ Agent runtime Typed model actions, task limits, cancellation, retries, and deterministic mock execution Delegation Parent/child task graph, target and capability checks, depth/fan-out limits, loop detection, an","default_branch":null,"files":null,"tree":[],"storefront":"/r/christiangrey922","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/christiangrey922/multi-agent-workflow-lab/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}