control system data analysis current issues and solution 20 november 2013

13
Control System Data Analysis Current Issues and Solution 20 November 2013 CERN - Openlab Workshop Author: Filippo Tilaro Supervised by: Axel Voitier

Upload: albina

Post on 23-Feb-2016

25 views

Category:

Documents


0 download

DESCRIPTION

Control System Data Analysis Current Issues and Solution 20 November 2013. CERN - Openlab Workshop. Author: Filippo Tilaro Supervised by: A xel Voitier. Summary. Current “production” state Use-cases under analysis GAS alarms breakdown Control System Health - PowerPoint PPT Presentation

TRANSCRIPT

Page 1: Control System Data Analysis Current Issues and  Solution 20 November 2013

Control System Data Analysis Current

Issues and Solution

20 November 2013

CERN - Openlab Workshop

Author: Filippo TilaroSupervised by: Axel Voitier

Page 2: Control System Data Analysis Current Issues and  Solution 20 November 2013

Summary

Current “production” state

Use-cases under analysis GAS alarms breakdown Control System Health Statistical Analysis of Alarms

Issues and Current limitations

Openlab Workshop – 20 November 2013 2

Page 3: Control System Data Analysis Current Issues and  Solution 20 November 2013

Current “production” StateServices

3

Fieldbus

TN

PLCs

Sensors &

Actuators

MOON(Monitoring)

High Voltage

DIM/CMW OPC

Control and Monitoring system Alerting and reporting system

Manually configured Based on threshold trespassing

pattern

Huge data volume: OS logs, performances metrics,

device status, Measurements, Alarms …

but not efficiently exploited yet

Field layer

Processlayer

Supervisionlayer

Openlab Workshop – 20 November 2013

Page 4: Control System Data Analysis Current Issues and  Solution 20 November 2013

4

Use cases under analysis

GAS system breakdown: system fault analysis and pattern extraction Events sequence pattern matching Post-mortem analysis so far Fault prediction based on recognizable trails of events

Control Systems Health Pattern matching and correlations of multivariate time series Structured (i.e. measurements) and unstructured (i.e. logs) data

Alarms statistical analysis Extract statistical indexes from the list of raised alarms Pragmatic approach: automatic threshold discovery and learning

Strategy: Use and extend the Siemens WatchCAT and other open-source analysis tools to

extract possible patterns and discover new insights hidden in the control data Take advantage of the huge amounts of control data produced by CERN facilities

Openlab Workshop – 20 November 2013

Page 5: Control System Data Analysis Current Issues and  Solution 20 November 2013

5

Gas System use-case

Openlab Workshop – 20 November 2013

28 Applications(Sub Detector)

7 Apps1 Data Server

9 Apps1 Data Server

6 Apps1 Data Server

6 Apps1 Data Server

Multi-wire chamber

Page 6: Control System Data Analysis Current Issues and  Solution 20 November 2013

6

Gas System Analysis

Openlab Workshop – 20 November 2013

Events ListExtraction

Simulation of Physical Control System: Complex System: more than 9000 equations to model all the system Validated against the real system Includes fault model!

Complex Diagnostic: Alarm flooding, “domino effect” A single fault can stop the whole process The 1st alarm is not necessarily the most

relevant for the diagnosis The alarm list depends on the system

status a knowledge-based model is not sufficient!

XML Conversion

SiemensWatchCAT

Pattern Extraction: Fault Signature Sequence

Alignment

Page 7: Control System Data Analysis Current Issues and  Solution 20 November 2013

7

Example: Distribution faultBubbler (safety device broken) line 2:

Initial impact on the Pump module, then on the Distribution

The Distribution seems to not have alarms yet

The Entire Control Process collapses

Openlab Workshop – 20 November 2013

Page 8: Control System Data Analysis Current Issues and  Solution 20 November 2013

Goal: control system faults/anomalies detection and diagnosis

8

Offline Control System Health

Openlab Workshop – 20 November 2013

Application WinCC OASystems

Parameters(Million dpes)

ALICE 100 3ATLAS 130 12 CMS 90 10LHCb 160 10

Accelerator Complex 120 10

System architecture under analysis: 16 Control Applications

QPS, nQPS, CRYO, CIET, CIS, PIC, WIC, LHC-CIRCUIT, PSEN … Linux control PCs : ~120 PLCs: ~300 FECs: ~100

Page 9: Control System Data Analysis Current Issues and  Solution 20 November 2013

9

Offline Control System Health Analysis

Openlab Workshop – 20 November 2013

Lemon

UNICOS

CMW FECs

LOGs

MOON long term storage diagnostic data, alarms,

devices status

Performances metrics Exceptions Status information

WinCC OA logs Sys logs

Unified Control SystemAlarms

FECs logs (from Splunk)

Pre-Data Analysis

I• Data Extraction

II• XML-Conversion

III• Data Cleaning / Completion

Repository

SiemensWatchCAT

Page 10: Control System Data Analysis Current Issues and  Solution 20 November 2013

10

Offline Control System Health: Status

Openlab Workshop – 20 November 2013

Issues: Huge amount of data [~130GB + LHC] Different data types:

Structured/Not Structured Numerical / Boolean / Plain-text Gaps, missing some metadata

Unsynchronized data sources Different relationships among the subsystems …

Initial conclusions no single framework out of the box to analyze numerical data and not

(next version of WatchCAT) Necessary a combination of tools for a complete data analysis (log

processing, statistical analysis, pattern recognition…) Split this use-case into smaller ones:

signal analysis use-case (next version of WatchCAT will provide predictive trending capabilities)

automatic extraction of statistical metrics and thresholds

Page 11: Control System Data Analysis Current Issues and  Solution 20 November 2013

11

Alarms Analysis Flow

Alarms List

Filtering & Aggregation

POJOsExtraction Conversion

Injection

Reporting

Openlab Workshop – 20 November 2013

MOON

CEP engine Open-source rules engine declarative paradigm

Page 12: Control System Data Analysis Current Issues and  Solution 20 November 2013

12

Typical Issues

Necessary actions: Access to the data (i.e. sensible or protected information) Deal with data heterogeneity: file formats, units of

measure, date formats, data structures Data synchronization Several different data sources Data enhancement: data classification, data

completeness, improve time resolution … Data selection / filtering Data input/output representations …

Openlab Workshop – 20 November 2013

Page 13: Control System Data Analysis Current Issues and  Solution 20 November 2013

13

Any Questions

Thank you for attending!

Openlab Workshop – 20 November 2013