Monitoring
Log Intelligence Center
AI Generated Summary
Error rate across all log sources is 0.4%, within normal range. A cluster of timeout errors on the feature_aggregation pipeline correlates with the schema change flagged in Data Pipelines — root cause identified.
Live Log Stream
StreamingError Analytics
Error count · last 24 hours
Event Correlation
AI Root Cause Analysis
The transform_events timeout cluster began 30 minutes after a schema change was pushed to customer_events. Confidence: 91%.
Recommended fix: add a backward-compatible schema version check to the transform_events pipeline before rollout.
AI Log Summary
312 error-level log entries in the last 24 hours, 84% of which trace back to a single root cause already identified above.
WARN-level volume is up 6% week over week, mostly from the recommendation-engine service.
INFO-level noise from the ETL scheduler accounts for 61% of total log volume — consider raising its default log level in non-debug environments.
No CRITICAL-level entries in the last 24 hours; the last was the timeout error cluster resolved 3 days ago.
Log Timeline
Timeout error cluster began
30 min ago
Schema change deployed to customer_events
1 hour ago
Nightly log rotation completed
4 hours ago
Archived Logs
Activity History