AI Assistant Watchdog / Trust Dashboard
The Problem
Teams and individual developers use multiple AI coding assistants but lack metrics to know which assistant is trustworthy for which tasks. There's no consolidated telemetry about hallucination rates, tendency to alter tests, or regression-introduction frequency. This reduces adoption and increases review burden.