{"repo":"vladkol/agent-evaluation-lab","free":true,"listed":false,"github":"https://github.com/vladkol/agent-evaluation-lab","clone":"git clone https://github.com/vladkol/agent-evaluation-lab.git","description":"Evaluation of Multi-Agent Systems on Cloud Run with the GenAI Client in Vertex AI SDK","language":"Python","stars":12,"topics":["adk","agent-tracing","ai-agent-evaluation","evaluation","google-cloud","observability"],"license":"Apache-2.0","category":"ai-agents","readme_excerpt":"Evaluation of Multi-Agent Systems This project implements a Continuous Evaluation Pipeline for a multi-agent system built with Google Agent Development Kit (ADK) and Agent2Agent (A2A) protocol on Cloud Run. It features a team of microservice agents that research, judge, and build content, orchestrated to deliver high-quality results. The goal of this project is to demonstrate Agentic Engineering practices for Continuous Evaluation : safely deploying agents to shadow revisions, running automated evaluation suites using Gemini Enterprise Agent Platform, and making data-driven decisions on agent deployments and improvements. It is a companion code repository to the codelab From \"vibe checks\" to data-driven Agent Evaluation . It uses Agent Platform GenAI Evaluation Service that provides enterprise-grade tools for objective, data-driven assessment of generative AI models and agents. Architecture The system uses a distributed microservices architecture where each agent runs in its own container and communicates via the A2A protocol: Orchestrator Service ( orchestrator ): The main entry point and \"brain\" of the operation. It manages the workflow using LoopAgent and SequentialAgent patterns, delegating tasks to other agents. Researcher Service ( researcher ): A standalone agent equipped with a Wikipedia Search tool. It gathers information based on queries. Judge Service ( judge ): A standalone agent that evaluates the quality and relevance of the research provided by the Researcher. ","default_branch":null,"files":null,"tree":[],"storefront":"/r/vladkol","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/vladkol/agent-evaluation-lab/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}