People Nerds

Scaling Quality with AI Evaluation Models

July 22, 2026

overview

Leaders from The Cigna Group, Dscout, Datadog, and TD Bank Group explain how they evaluate AI-generated research output and govern trust in AI-assisted workflows.

Contributors

Design Executive Council

The Dscout Team

Author

Scaling Quality with AI Evaluation Models

July 22, 2026

Overview

Leaders from The Cigna Group, Dscout, Datadog, and TD Bank Group explain how they evaluate AI-generated research output and govern trust in AI-assisted workflows.

Contributors

Design Executive Council

The Dscout Team

Author

Join DXC and Dscout for the fourth and final session in our series exploring how AI is reshaping research and design.

Leaders from The Cigna Group, Dscout, Datadog, and TD Bank Group will discuss how their teams are approaching evaluation frameworks, last-mile review, and trust in AI-assisted workflows—without slowing the pace of delivery.

The group discussess...

Interpreting ‘good enough’ AI output

As AI-generated research artifacts spread across teams, "good output" means something different depending on who you ask (and that gap has real consequences).

The group will explore how orgs can align on shared quality standards and who should hold the authority to define and enforce them.

Last-mile reviews and evaluation at scale

Agreement on quality is only half the battle. The harder part is building review processes that hold up as AI output volumes grow.

Leaders will dig into how teams are structuring human-in-the-loop evaluation to stay rigorous without becoming a bottleneck.

Building trust and evolving governance

Trust in AI-assisted research isn't static, and the governance frameworks built for early adoption may not be fit for what's coming next.The group will explore how to read signals of shifting trust—internally and externally—and how they're adapting oversight as AI takes on greater autonomy.

Learn more about the speakers

HOT off the Press

More from People Nerds