Agent Testing & QA
Local CI
Agent-enablingSource AvailableSource-available local GitHub Actions runner designed for AI-agent development loops
Category
Tools that let agents plan, run, inspect, or maintain tests for software and agent-driven workflows.
Editorial scope version 1.0 ยท effective 3 September 2026
Category listing
Definition
This category covers testing and quality-assurance systems in which agents are active test authors, operators, evaluators, or debugging participants.
Includes systems where agents actively author, execute, adapt, evaluate, or diagnose tests and produce reviewable artifacts such as traces, screenshots, or reproducible steps.
Scope
Selection guide
Category evidence
LangSmith documents final-response, single-step, and trajectory evaluation as distinct, complementary ways to test agent behaviour.
accessed 2026-09-03Comparable facts
Unknown values are shown explicitly rather than inferred from marketing copy. Open a profile to inspect its claim-level sources.
| Tool | Classification | Pricing | Interfaces | Deployment | Verification |
|---|---|---|---|---|---|
| Local CI | Agent-enabling | Source Available | command-line interface, NDJSON event stream, agent skill | local Docker, remote Docker, local macOS virtual machines | documentation reviewed |
| Agent QA | Agent-native | Source Available | command-line interface, web dashboard, MCP server, agent skills | local, continuous integration | documentation reviewed |