Telnyx - Global Communications Platform ProviderHome
Voice AI AgentsText-to-SpeechSpeech-to-TextEmbeddingsSearch APIBrowser APIMeetingBotVoice DesignInference APIAgentSDKFunctionsStateful ActorsKVSQLDBStorageGlobal NumbersVoice APISIP TrunkingSMS APIEmail APIRCSWhatsAppWebRTCVerify APINumber ReputationNumber LookupDeepfake DetectionBranded CallingIoT SIMeSIMMobile VoicePrivate Wireless GatewaysVirtual Cross ConnectsCloud VPNGlobal IP200+ open-source buildsagent-signup.mdx402View all primitivesHealthcareFinanceTravel and HospitalityLogistics and TransportationContact CenterInsuranceRetail and E-CommerceSales and MarketingServices and DiningView all solutionsVoice AIVoice APIInferenceMobile VoiceSpeech-to-TextText-to-SpeechSIP TrunkingSMS APIEmail APIWhatsApp Business APIGlobal NumbersIoT SIM CardView all pricingOur NetworkGlobal communicationsEdge ComputeAgents PlatformPartnersCareersCustomer storiesResource centerMission Control PortalEventsSupport centerSETIDev DocsIntegrationsCode examples
Contact usLog in
Sign up
Contact usLog in
Sign up
Start building

Social

Compare

  • Twilio
  • Bandwidth
  • Plivo
  • Vonage
  • Wasabi
  • Amazon S3
  • ElevenLabs
  • Vapi
  • Baseten
  • Together.ai
  • Amazon Connect
  • Lumen
  • Cloudflare
  • Resend
  • SendGrid
  • Mailgun

Resources

  • Release Notes
  • Acceptable Use
  • Terms and conditions
  • Website Terms and Conditions
  • Data and Privacy
  • Report Abuse
  • Privacy Policy
  • Cookie Policy
  • Law Enforcement
  • Trust Center

Company

  • Why Telnyx
  • Our Network
  • Global Coverage
  • Customer Stories
  • Careers
  • Country Specific Requirements

Social

Company

  • Our Network
  • Global Coverage
  • Release Notes
  • Careers
  • Voice AI
  • AI Glossary
  • Shop

Legal

  • Data and Privacy
  • Report Abuse
  • Privacy Policy
  • Cookie Policy
  • Law Enforcement
  • Acceptable Use
  • Trust Center
  • Country Specific Requirements
  • Website Terms and Conditions
  • Terms and Conditions of Service

Compare

  • ElevenLabs
  • Vapi
  • Baseten
  • Together.ai
  • Twilio
  • Bandwidth
  • Vonage
  • Amazon Connect
  • Lumen
  • Cloudflare
  • Resend
  • SendGrid
  • Mailgun
Telnyx
© Telnyx LLC 2026
ISO • PCI • HIPAA • GDPR • SOC2 Type II

Ask AI

  • GPT
  • Claude
  • Perplexity
  • Gemini
  • Grok
Back to Glossary

Concept drift: why your model's accuracy is declining

Concept drift vs data drift in machine learning: data drift changes a model's inputs, concept drift changes what they mean. How to detect and fix each.

Andy Muns
Editor: Andy Muns

Updated September 2026

Data drift and concept drift are two of the main reasons a machine learning model loses accuracy after deployment. Data drift is a change in the distribution of the input data, written P(X), while the relationship between inputs and the target stays the same. Concept drift is a change in that relationship, P(y|X): the same inputs now call for a different output.

Neither data drift nor concept drift sets off an alarm. The model keeps running and returning answers, and nothing in its logs looks wrong, while the answers get less accurate. That is why Rabanser and colleagues describe machine learning systems as ones that "tend to fail silently." The two also surface in different places, data drift in the inputs and concept drift only in the outcomes, so each needs its own check.

This content was generated with the assistance of AI. Our AI prompt chain workflow is carefully grounded and preferences .gov and .edu citations when available. All content is reviewed by a Telnyx employee to ensure accuracy, relevance, and a high standard of quality.

Sign up and start building.

Sign UpContact Us

Understanding data drift

Data drift means the inputs a model receives in production no longer look like the inputs it trained on. A speech recognition model trained on clean headset audio that starts receiving noisy speakerphone calls has data drift: the audio changed, but the words still map to the same transcript.

The statistical term for data drift is covariate shift. Lu and colleagues call it virtual drift, because the decision boundary, the real dividing line between one answer and another, does not move. Data drift degrades model accuracy anyway. A model only learns where that line runs in the regions its training data covered, so when inputs move into regions it rarely saw, the model guesses.

Causes of data drift

Data drift starts with anything that changes who or what feeds the model:

  • Changes in user behavior, such as a shift from desktop to mobile checkouts.
  • Seasonal patterns and market trends, such as holiday shopping.
  • New data sources, such as a new region or customer segment.
  • New vocabulary, such as product names a text-to-speech voice has never seen.
  • Measurement changes, such as a new sensor, microphone, or audio codec.
  • Upstream pipeline changes, such as a field that switches units.

Understanding concept drift

Concept drift means the correct output for a given input has changed. The inputs can look exactly like the training data while the right answer moves, usually because of something the model never sees, such as the economy or a fraud ring's playbook. Lu's review calls these hidden variables and names the result actual drift: a change in P(y|X) that moves the decision boundary itself.

Concept drift degrades model accuracy even when every input looks familiar.

How drift breaks a model, shown in three scatter plots of the same two-feature space. At training, the model's boundary separates two classes where the training data sits, drawn solid where tested and dashed beyond. Under data drift, the inputs move into the untested region, where the dashed boundary cuts through one class and misclassifies part of it. Under concept drift, the inputs stay where they were but a new true boundary cuts through one class, so points the model still labels the old way are now misclassified.

Causes of concept drift

Concept drift comes from changes in the world that alter what an input means:

  • Economic shifts, such as a recession that changes which borrowers default at the same income.
  • Adversaries adapting, such as fraudsters changing tactics once a model learns the old ones.
  • Policy changes that redefine a label, such as a new rule for what counts as a compliant transaction.
  • Shifting preferences, such as customers who once responded to a discount now ignoring the same offer.
  • Seasonal cycles, in which the same inputs lead to different outcomes at different times of year.

Types of concept drift

Concept drift has four types: sudden, gradual, incremental, and recurring. Lu's review sorts them by how the new relationship between inputs and answers replaces the old one:

  • Sudden drift, where the relationship changes at once, such as the day a new regulation takes effect.
  • Gradual drift, where the old and new relationships alternate in the data until the new one takes over.
  • Incremental drift, where the relationship shifts a little at a time, such as delivery expectations tightening each month.
  • Recurring drift, where an old relationship returns, such as holiday buying each December.

Key differences between data drift and concept drift

The key difference between data drift and concept drift is distribution versus relationship change. Data drift refers to changes in the distribution of the input data, but the relationship between the input features and the target variable remains the same. Concept drift refers to changes in the relationship between the input features and the target variable, even if the input distribution remains the same.

In probability terms, data drift changes P(X) and concept drift changes P(y|X), and the two often arrive together.

Data driftConcept drift
What changesThe input distribution, P(X)The input-output relationship, P(y given X)
Also calledCovariate shift, virtual driftActual drift, concept shift
Visible in the inputsYesNot necessarily
Needs labels (true outcomes) to detect itNoUsually
Typical triggersNew users, devices, seasons, pipeline changesEconomic shifts, adversaries, policy changes
Usual responseRetrain on recent inputs, or fix the pipelineRelabel under the new concept, then retrain

The cause does not reveal the type: economic changes such as a recession can shift who applies for loans (data drift) and who defaults (concept drift) at once. Concept drift usually costs more to fix, because labels collected under the old concept teach the old answer: it takes new labels, not just new inputs.

What are examples of data drift and concept drift?

A common example of concept drift is fraud detection: once a model learns to flag a pattern, fraudsters shift to transactions that resemble approved ones, so the same features carry a different risk.

A common example of data drift, by contrast, is a medical imaging model moved to a new hospital. In the WILDS benchmark, a tumor classifier averaged 93.2% accuracy on tissue patches from the hospitals it trained on and 70.3% at a hospital it had never seen. Nearly one patch in three was misread. Tumor tissue still meant tumor; what changed was the slides, through differences in staining and image acquisition.

Some systems face both at once. Telnyx's guide to synthetic speech detection notes that new text-to-speech (TTS) models appear constantly, so a static detector starts drifting from day one. Each new voice generator changes both the audio a detector hears and the cues that mark a voice as synthetic.

How does model drift relate to data drift and concept drift?

In MLOps usage, model drift, also called model decay, is the decline in a model's performance in production. The difference between data drift, concept drift, and model drift is cause and effect: the first two are causes, and model drift is the accuracy lost. A third cause, label drift, changes how often each outcome occurs, such as flu becoming more common in winter while its symptoms stay the same.

Label drift can be corrected without new labels. Lipton and colleagues, who call it label shift, showed that the new outcome mix can be estimated from the model's own predictions and the model adjusted to match.

How to detect data drift and concept drift

Detect data drift by running a two-sample statistical test on each input feature, comparing a recent window of production data with the training data. Detect concept drift by scoring the model's predictions against the real outcomes as they come in, and watching for accuracy to fall.

For a numeric feature, the standard choice is the two-sample Kolmogorov-Smirnov (KS) test, which compares the distributions of the two samples and reports how far apart they are. On thousands of rows it flags even tiny, harmless shifts as significant, so alert on how big the shift is.

For images and other inputs with thousands of dimensions, test the model's outputs instead of the raw inputs. In Rabanser and colleagues' image experiments, running the same tests on a classifier's output probabilities caught shift best, so the model you already run can double as its drift detector.

Input and output checks share a blind spot: under pure concept drift the inputs do not change, so neither do the model's outputs. The only signal left is accuracy, measured against the ground truth. Track accuracy or F1 score per time window. Error-rate detectors, the largest family in Lu's review, automate the watch: the Drift Detection Method (DDM), one of the most cited, raises a warning when the error rate rises significantly and signals drift if it keeps rising.

This Python sketch runs a simple version of both checks, a KS test on one feature and an accuracy threshold on weekly labeled traffic:

import numpy as np
from scipy.stats import ks_2samp

rng = np.random.default_rng(0)
train_x = rng.normal(50, 10, 5000)  # a feature at training time
prod_x = rng.normal(56, 10, 5000)   # the same feature in production

# Data drift: two-sample KS test on the inputs, no labels needed
result = ks_2samp(train_x, prod_x)
print(f"KS statistic {result.statistic:.3f}, p-value {result.pvalue:.1e}")

# Concept drift: weekly accuracy on labeled traffic (example values)
weekly_accuracy = [0.94, 0.93, 0.94, 0.88, 0.85]
baseline = weekly_accuracy[0]
drifted = [week for week, acc in enumerate(weekly_accuracy) if acc < baseline - 0.05]
print("Accuracy dropped in weeks:", drifted)

The two checks run on different clocks. A data drift check needs only the inputs, so it can run as soon as production data arrives. A concept drift check needs labels, the true outcome for each prediction, such as whether a flagged transaction really was fraud, and a fraud label may arrive only when a chargeback lands. Concept drift shows up only as fast as those labels come back, which is why it is slower to catch.

In an MLOps pipeline, both checks run on a schedule and alert when they fire. Telnyx's guide to model deployment makes that monitoring a step of the deployment process itself, with automated alerts that catch drift before it reaches users.

Strategies for managing data drift and concept drift

Lu's review groups the strategies for managing data drift and concept drift into three families: retrain on recent data, keep past models for patterns that return, and adapt the model as data arrives.

  • Retrain on a detector's signal with a window of recent data, and keep a fixed schedule as a backstop.
  • Relabel before retraining for concept drift. Old labels teach the old answer, so data labeling has to capture today's right answer before the model sees it.
  • Keep past models for recurring drift, such as holiday buying, and switch back when the old pattern returns.
  • Update the model continuously when drift is local, changing only the part of the model the drift touched.
  • Fine-tune large models on recent data instead of retraining them: small adapter weights can train on a single GPU. Tuned weights go stale as the task changes too, so plan to repeat it.
  • Roll out each fix to a slice of traffic first, as canary deployments do, with a rollback ready.

If data drift traces to a pipeline bug, fix the pipeline. Retraining on broken data teaches the model the bug.

Four kinds of drift compared by what changes, labels needed, an example, and the usual response. Data drift changes P(X), is detected from inputs without labels, for example noisy speakerphone audio, and is handled by retraining on recent inputs or fixing the pipeline. Concept drift changes P(y|X), usually needs labels to detect, for example fraudsters changing tactics, and is handled by relabeling under the new concept, then retraining. Label drift changes P(y), can be handled from the model's predictions without new labels, for example flu becoming more common in winter, and is corrected for the new outcome mix. Model drift is the effect: weekly accuracy drops, measured against labels, and the response is to find which cause is behind it.

What is concept drift in LLMs?

Concept drift in a large language model is the growing gap between the world the model learned from and the world it answers questions about after its training cutoff. Lazaridou and colleagues found that language model performance "becomes increasingly worse with time" on text from beyond the training period, and that model size alone does not solve it. Grounding in AI narrows the gap at answer time by feeding a large language model current facts.

LLM applications also drift when the model itself changes. Chen, Zaharia, and Zou found that GPT-4's accuracy at identifying prime numbers fell from 84% in its March 2023 version to 51% in its June 2023 version. Pin model versions, and rerun your evaluations whenever a provider updates one.

Frequently asked questions

What does concept drift refer to in machine learning?

Concept drift in machine learning refers to a change over time in the relationship between a model's input features and the target it predicts. A model trained on past data then gives outdated answers to inputs that look familiar.

Is covariate shift the same as data drift?

Covariate shift is the research term for what practitioners usually call data drift: P(X) changes while P(y|X) stays fixed. Covariate shift differs from concept drift, which changes P(y|X) itself.

Is data drift a type of concept drift?

Data drift is a type of concept drift in much of the research literature, where many papers, Lu's review included, call any change in inputs or answers concept drift. In production monitoring, the two usually name separate problems: a change in the inputs, and a change in the right answer.

How often should you retrain a model to handle drift?

Retrain a model when a drift detector or a drop in labeled accuracy shows drift, with a fixed schedule as a backstop. A schedule alone can react weeks late to a sudden change.

How do you handle drift in a voice AI agent?

Handle drift in a voice AI agent by analyzing conversations for new failure patterns and shipping fixes to a slice of calls first. With Telnyx Voice AI agents, AI Assistants run analysis on every conversation through Insights, split calls between versions by percentage for gradual rollouts, and roll back to the main version. New product names and caller accents are data drift; a new refund policy, which changes the right answer, is concept drift.

Sources

  • Lu, Liu, Dong, Gu, Gama, and Zhang. Learning under Concept Drift: A Review, IEEE Transactions on Knowledge and Data Engineering, 2018.
  • Rabanser, Günnemann, and Lipton. Failing Loudly: An Empirical Study of Methods for Detecting Dataset Shift, 2018.
  • Lipton, Wang, and Smola. Detecting and Correcting for Label Shift with Black Box Predictors, 2018.
  • Lazaridou et al. Mind the Gap: Assessing Temporal Generalization in Neural Language Models, 2021.
  • Koh et al. WILDS: A Benchmark of in-the-Wild Distribution Shifts, 2020.
  • Chen, Zaharia, and Zou. How is ChatGPT's behavior changing over time?, 2023.
  • SciPy. scipy.stats.ks_2samp.
  • Telnyx. Testing and Traffic Distribution for AI Assistants.
  • Telnyx. Voice Assistant Quickstart.
Share on Social

Jump to:

Understanding data driftUnderstanding concept driftKey differences between data drift and concept driftWhat are examples of data drift and concept drift?How does model drift relate to data drift and concept drift?How to detect data drift and concept driftStrategies for managing data drift and concept driftWhat is concept drift in LLMs?Frequently asked questionsSources

Sign up for emails of our latest articles and news