Radio and PodcastRadio and PodcastLive Radio & Podcasts
An LLM Evaluation Framework for High-Stakes AI artwork
Technology

An LLM Evaluation Framework for High-Stakes AI

Software Engineering Institute (SEI) Podcast Series by Carnegie Mellon University Software Engineering Institute

Jun 11, 202616:33Technology

Experimentation and validation of LLM performance is critical when building LLM-driven systems that must reliably deliver a service, from customer service chat bots to intelligence analysis tools. To help teams meet the...