Back to papers
March 19, 2026cs.AIIntermediate
Behavioral Fingerprints for LLM Endpoint Stability and Identity
AI-Generated Summary
This paper introduces Stability Monitor, a system that tracks whether AI language model endpoints behave consistently over time by periodically testing them with fixed prompts and analyzing how their outputs change. Rather than just checking if a service is running (uptime), it detects when a model's actual behavior shifts due to updates, hardware changes, or other modifications. The system successfully identified various types of changes in controlled tests and found that the same model can behave quite differently depending on which company is hosting it.
Difficulty
Intermediate
Categories
cs.AI
AI Tags
LLM monitoringmodel stabilitybehavioral consistencyfingerprintingreliabilitystatistical testinginfrastructuredeployment