Monitoring quality, cost and operations together
Technical availability says little about answers that are wrong in substance. We combine measurements of response time, errors and consumption with suitable quality feedback. Data drift is an indication of changed inputs, but not yet proof of worse results. Retraining or a model change is therefore evaluated against a stable test set. A staged rollout limits the impact; fallback versions and shutdown options are prepared. For external models, we take version changes and provider outages into account.


