“quality-assurance”
Test-Driven Agent Development: Writing Evals Before You Build
Learn how to write evaluation-driven agent development tests that catch regressions, validate tool calls, and ensure reliable agent behavior before shipping.
LLM-as-a-Judge for Agent Outputs: Strengths and Pitfalls
How to use LLMs as evaluators for agent outputs — the strengths, systematic biases, and practical debiasing strategies for production evaluation.
Evals for Enterprise Agents: From Unit Tests to Regression Suites
Build comprehensive evaluation suites for enterprise agents. Learn to create unit tests, regression tests, and benchmark suites for multi-step agent reasoning and tool selection.
Building an Eval Test Suite for ADK Agents
Move from vibes-based testing to structured evaluation for ADK agents, tracking tool call accuracy, trajectory efficiency, and LLM-as-a-judge quality scoring.
Stop Writing Boilerplate: Generate Unit Tests Automatically with AI
Unit testing is essential but often tedious. Learn how to use AI to generate comprehensive test suites and edge cases in seconds.
Integrating Spec Kit and Gen AI for ISO 9001:2015 Documentation
A guide on how software development organizations can leverage Spec-Driven Development with Spec Kit and Generative AI to efficiently create ISO 9001:2015 compliant documentation.