Operationalizing Evaluation For Voice & Language Related AI Use Cases

Description:

Building methods/tools to test productized (black-box) Conversational AI solutions for AI safety and Responsible AI compliance, since public testing tools mostly target underlying LLMs, not full products.