Topic: Strawberry

1 chapters across the catalog

Episode 261: Podhemian Grove
14:18 - 17:26

Episode 261: Podhemian Grove

Unit Testing AI Agents and Mathematical Logic Failures

The hosts explore the necessity of applying "unit tests" to AI outputs to ensure quality, such as checking for dangling audio in clips. They discuss why AI agents struggle with simple logic, like counting letters in the word "strawberry," unless they are specifically instructed to run a Python script to verify the answer. The cost of running these verification scripts often prevents models from being accurate in math and logic by default.