
AgentGuard
- AI Safety
- Developer Tools
- Agent Evaluation
Built an open-source Python CLI for repeatable coding-agent evaluations with runtime guardrails, adversarial benchmarks, traces, dashboards, and CI-ready reports.
- Python
- Pytest
- Docker
- GitHub Actions
- YAML
- Static Reports
- Trace Replay