promptdojo_

Data drift and concept drift — step 2 of 7

Input monitoring is green — every feature's serving distribution matches training. Yet rolling accuracy against late-arriving labels has slid from 0.86 to 0.71. What's happening?