Home › Glossary › Microsoft Foundry AI agent evaluation GitHub Action

Microsoft Foundry AI agent evaluation GitHub Action

Brings offline evaluation of Foundry agents into CI/CD on GitHub. A test dataset is sent to the agents, catalogue evaluators grade what comes back, versions are compared statistically, and the result can hold back a release.

Also called ai-agent-evals.

Read more: Microsoft Learn

In the Ultra Transcenders books

AI-300AI-103

Each book explains Microsoft Foundry AI agent evaluation GitHub Action in context, with comparison tables and the common traps.

Terms in this definition

See Microsoft Foundry AI agent evaluation GitHub Action in the full glossary