The Best Way to Test New AI Models
The episode outlines a repeatable framework for evaluating new AI models on your own tasks, weighing quality, speed, and cost to decide which models merit adoption. It also cites KPMG research showing that treating AI as a reasoning partner drives higher impact and that these skills can be taught at scale.
- Create a systematic testing process that measures model quality, latency, and cost for your specific use cases.
- Directly compare outputs across models to identify the best fit for each task.
- Treat AI as a reasoning partner; developing this mindset can be scaled across teams for greater impact.

