Session
Ship Responsibly: Evaluation and Safety in Azure AI Foundry
You have built an AI-powered application - now how do you know it actually works? Traditional software has unit tests and integration tests, but AI applications are probabilistic, context-dependent, and capable of producing harmful content. In this session, we tackle the hardest problem in AI engineering: measuring quality and ensuring safety before you ship. Using Azure AI Foundry's built-in evaluation framework, we will run quality evaluators (groundedness, relevance, coherence, fluency) and safety evaluators (violence, hate, self-harm, protected material, jailbreak detection) against a real application. You will see how to build test datasets, run evaluations locally and in the cloud, interpret scorecards, and integrate evaluation gates into your CI/CD pipeline so that a regression in quality or safety blocks deployment automatically. We will also cover Foundry's runtime content filters, the Risks and Safety monitoring dashboard, and how to build a responsible AI practice that goes beyond tooling - including human oversight, incident response, and transparency. Whether you are a developer, tech lead, or architect, you will leave with a practical playbook for shipping AI applications that your users can trust.
Target Audience
Tech leads, architects, and developers who are shipping or preparing to ship AI-powered applications to production.
Prerequisites for Attendees
- Basic understanding of LLM-based applications (chat, RAG, agents)
- Familiarity with CI/CD concepts
- An Azure subscription with an AI Foundry resource
Joseph Guadagno
Senior VP of Technology, Microsoft MVP, and international speaker specializing in engineering leadership, .NET, Azure, and AI‑assisted development. A long‑time community leader.
Chandler, Arizona, United States
Links
Please note that Sessionize is not responsible for the accuracy or validity of the data provided by speakers. If you suspect this profile to be fake or spam, please let us know.
Jump to top