Opinionated end-to-end testing framework for AI agents: run the real agent, get a binary pass/fail verdict backed by tests, traces, and deliverables.
-
Updated
Sep 19, 2026 - Python
Opinionated end-to-end testing framework for AI agents: run the real agent, get a binary pass/fail verdict backed by tests, traces, and deliverables.
Official Seclai Go SDK
Official Seclai JavaScript SDK
Official Seclai C# SDK
Official Seclai Python SDK
Official Seclai Command Line Interface
Connect AI coding tools to Seclai via Model Context Protocol (MCP) — manage agents, knowledge bases, and content sources from Claude, Cursor, Windsurf, and more
Turn agent telemetry into eval jobs — outcome, quality, and spend — correlated with business outcomes per agent run.
Cognitive AutoGen (CoAG) extends Microsoft AutoGen with a CoALA-inspired cognitive layer — structured working memory, RAG long-term memory and a 6-phase deliberative cycle
To associate your repository with the agent-evaluations topic, visit your repo's landing page and select "manage topics."