Home / Companies / WhyLabs / Blog / Post Details
Content Deep Dive

Detecting Semantic Drift within Image Data: Monitoring Context-Full Data with whylogs

Blog post from WhyLabs

Post Details
Company
Date Published
Author
WhyLabs Admin
Word Count
2,726
Company Posts That Month
1
Language
English
Hacker News Points
2
Post removed?
No
Summary

This article discusses the use of whylogs for monitoring machine learning systems' data ingestion pipeline by enabling concept drift detection, specifically for image data. It presents two scenarios to demonstrate how to create more generalized semantic metrics and monitor specialized semantic information directly from datasets. The first scenario involves using metadata information and properties such as hue, saturation, or brightness (HSB) to detect data changes. In the second scenario, semantic drifts in data are detected by generating feature embeddings using transfer learning. Custom features like distances from cluster centers can be created to represent the distance from the logged images to the "ideal" representation of each class. The article concludes with a discussion on how whylogs can help detect data drift issues in images and provides examples of how these approaches can be integrated into different stages of the data pipeline for full observability of machine learning applications.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Vector Search 15 55 25 21 -27%
LLM 12 64 11 8 -7%
Data Pipeline 4 242 59 33 +2%
AI Guardrails 3 34 26 7 -47%
Observability 3 708 117 43 +73%
RAG 2 5 4 4 -17%
Serverless 2 681 103 49 +69%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.