Back to papers
March 19, 2026cs.AIIntermediate

Behavioral Fingerprints for LLM Endpoint Stability and Identity

AI-Generated Summary

This paper introduces Stability Monitor, a system that tracks whether AI language model endpoints behave consistently over time by periodically testing them with fixed prompts and analyzing how their outputs change. Rather than just checking if a service is running (uptime), it detects when a model's actual behavior shifts due to updates, hardware changes, or other modifications. The system successfully identified various types of changes in controlled tests and found that the same model can behave quite differently depending on which company is hosting it.

Difficulty
Intermediate
Categories

cs.AI

AI Tags
LLM monitoringmodel stabilitybehavioral consistencyfingerprintingreliabilitystatistical testinginfrastructuredeployment