WO2015059272A1 - Improved non-intrusive appliance load monitoring method and device - Google Patents
Improved non-intrusive appliance load monitoring method and device Download PDFInfo
- Publication number
- WO2015059272A1 WO2015059272A1 PCT/EP2014/072843 EP2014072843W WO2015059272A1 WO 2015059272 A1 WO2015059272 A1 WO 2015059272A1 EP 2014072843 W EP2014072843 W EP 2014072843W WO 2015059272 A1 WO2015059272 A1 WO 2015059272A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- events
- event
- clustering
- cluster
- clusters
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Classifications
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F17/00—Digital computing or data processing equipment or methods, specially adapted for specific functions
- G06F17/10—Complex mathematical operations
- G06F17/18—Complex mathematical operations for evaluating statistical data, e.g. average values, frequency distributions, probability functions, regression analysis
-
- G—PHYSICS
- G01—MEASURING; TESTING
- G01D—MEASURING NOT SPECIALLY ADAPTED FOR A SPECIFIC VARIABLE; ARRANGEMENTS FOR MEASURING TWO OR MORE VARIABLES NOT COVERED IN A SINGLE OTHER SUBCLASS; TARIFF METERING APPARATUS; MEASURING OR TESTING NOT OTHERWISE PROVIDED FOR
- G01D4/00—Tariff metering apparatus
- G01D4/002—Remote reading of utility meters
-
- G—PHYSICS
- G01—MEASURING; TESTING
- G01R—MEASURING ELECTRIC VARIABLES; MEASURING MAGNETIC VARIABLES
- G01R21/00—Arrangements for measuring electric power or power factor
- G01R21/133—Arrangements for measuring electric power or power factor by using digital technique
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06Q—INFORMATION AND COMMUNICATION TECHNOLOGY [ICT] SPECIALLY ADAPTED FOR ADMINISTRATIVE, COMMERCIAL, FINANCIAL, MANAGERIAL OR SUPERVISORY PURPOSES; SYSTEMS OR METHODS SPECIALLY ADAPTED FOR ADMINISTRATIVE, COMMERCIAL, FINANCIAL, MANAGERIAL OR SUPERVISORY PURPOSES, NOT OTHERWISE PROVIDED FOR
- G06Q10/00—Administration; Management
- G06Q10/04—Forecasting or optimisation specially adapted for administrative or management purposes, e.g. linear programming or "cutting stock problem"
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06Q—INFORMATION AND COMMUNICATION TECHNOLOGY [ICT] SPECIALLY ADAPTED FOR ADMINISTRATIVE, COMMERCIAL, FINANCIAL, MANAGERIAL OR SUPERVISORY PURPOSES; SYSTEMS OR METHODS SPECIALLY ADAPTED FOR ADMINISTRATIVE, COMMERCIAL, FINANCIAL, MANAGERIAL OR SUPERVISORY PURPOSES, NOT OTHERWISE PROVIDED FOR
- G06Q50/00—Information and communication technology [ICT] specially adapted for implementation of business processes of specific business sectors, e.g. utilities or tourism
- G06Q50/06—Energy or water supply
-
- G—PHYSICS
- G01—MEASURING; TESTING
- G01D—MEASURING NOT SPECIALLY ADAPTED FOR A SPECIFIC VARIABLE; ARRANGEMENTS FOR MEASURING TWO OR MORE VARIABLES NOT COVERED IN A SINGLE OTHER SUBCLASS; TARIFF METERING APPARATUS; MEASURING OR TESTING NOT OTHERWISE PROVIDED FOR
- G01D2204/00—Indexing scheme relating to details of tariff-metering apparatus
- G01D2204/20—Monitoring; Controlling
- G01D2204/24—Identification of individual loads, e.g. by analysing current/voltage waveforms
-
- Y—GENERAL TAGGING OF NEW TECHNOLOGICAL DEVELOPMENTS; GENERAL TAGGING OF CROSS-SECTIONAL TECHNOLOGIES SPANNING OVER SEVERAL SECTIONS OF THE IPC; TECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
- Y02—TECHNOLOGIES OR APPLICATIONS FOR MITIGATION OR ADAPTATION AGAINST CLIMATE CHANGE
- Y02B—CLIMATE CHANGE MITIGATION TECHNOLOGIES RELATED TO BUILDINGS, e.g. HOUSING, HOUSE APPLIANCES OR RELATED END-USER APPLICATIONS
- Y02B70/00—Technologies for an efficient end-user side electric power management and consumption
- Y02B70/30—Systems integrating technologies related to power network operation and communication or information technologies for improving the carbon footprint of the management of residential or tertiary loads, i.e. smart grids as climate change mitigation technology in the buildings sector, including also the last stages of power distribution and the control, monitoring or operating management systems at local level
-
- Y—GENERAL TAGGING OF NEW TECHNOLOGICAL DEVELOPMENTS; GENERAL TAGGING OF CROSS-SECTIONAL TECHNOLOGIES SPANNING OVER SEVERAL SECTIONS OF THE IPC; TECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
- Y04—INFORMATION OR COMMUNICATION TECHNOLOGIES HAVING AN IMPACT ON OTHER TECHNOLOGY AREAS
- Y04S—SYSTEMS INTEGRATING TECHNOLOGIES RELATED TO POWER NETWORK OPERATION, COMMUNICATION OR INFORMATION TECHNOLOGIES FOR IMPROVING THE ELECTRICAL POWER GENERATION, TRANSMISSION, DISTRIBUTION, MANAGEMENT OR USAGE, i.e. SMART GRIDS
- Y04S20/00—Management or operation of end-user stationary applications or the last stages of power distribution; Controlling, monitoring or operating thereof
- Y04S20/20—End-user application control systems
- Y04S20/242—Home appliances
-
- Y—GENERAL TAGGING OF NEW TECHNOLOGICAL DEVELOPMENTS; GENERAL TAGGING OF CROSS-SECTIONAL TECHNOLOGIES SPANNING OVER SEVERAL SECTIONS OF THE IPC; TECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
- Y04—INFORMATION OR COMMUNICATION TECHNOLOGIES HAVING AN IMPACT ON OTHER TECHNOLOGY AREAS
- Y04S—SYSTEMS INTEGRATING TECHNOLOGIES RELATED TO POWER NETWORK OPERATION, COMMUNICATION OR INFORMATION TECHNOLOGIES FOR IMPROVING THE ELECTRICAL POWER GENERATION, TRANSMISSION, DISTRIBUTION, MANAGEMENT OR USAGE, i.e. SMART GRIDS
- Y04S20/00—Management or operation of end-user stationary applications or the last stages of power distribution; Controlling, monitoring or operating thereof
- Y04S20/30—Smart metering, e.g. specially adapted for remote reading
Definitions
- the invention pertains to the technical field of automatic appliance detection, in particular Non-Intrusive Appliance Load Monitoring (NIALM), which refers to the automated detection of the state of appliances, e.g. household appliances, industrial appliances, a company's appliances, etc., from a total energy consumption signal. More in particular, the present invention relates to obtaining information at the component level of electrical appliances in a system, e.g. a household or a company, from a measurement of the total electricity consumption of said system. This information at the component level may include but is not limited to: the presence, the state and/or the energy consumption.
- the system may also be comprised of a single or a few appliances, whereby the methods, devices and/or systems of the present invention can be used for condition monitoring of one or more appliances. Background
- WO 2012/160062 Al discloses a method for detecting transitions in a measured signal e.g. a discrete time signal and current waveform, which is induced by elements i.e. components of appliances, of a physical system i.e. a house.
- the method involves generating a residual signal, i.e. transition likelihood signal, from a measured signal, i.e. a discrete time signal, where the residual signal is provided with high amplitude when transitions occur and with low amplitude in other cases.
- Rules concluding that transitions occur when the residual signal is larger than a threshold stable value are provided, where the stable value is automatically defined from local values of the residual signal and defined as a function of local background noise.
- the measured signal is filtered before generating residual signal.
- the method enables defining a time index corresponding to end of transient state to maximize a distance between transient states and a straight line passing through the transient state, thus improving separation between transient and steady states, and hence detecting transitions in the measured signal, and monitoring automatic-setup non-intrusive appliance load for identifying appliances energy consumption in a reliable manner.
- This document also discloses an automatic-setup non-intrusive appliance load monitoring method (50) for identifying appliances energy consumption and comprising the steps of:
- the step of identifying components is hereby performed by selecting spectral features or features describing a current waveform for the steady states, with the possibility of the chosen features being grouped together in well-separated clusters for different components.
- a first step of component identification consists in clustering and classifying the components according to their nature (motor, resistor, heater, television are examples of different natures).
- a second step classifies the components within their nature cluster; this is the component classification itself.
- Identifying a nature of a component is a classification problem with a predefined number of classes. Therefore, a K-means algorithm is preferably chosen.
- a K- means method is a clustering technique that partitions n observations into K clusters, where K is a predefined value. It classifies the components according to their nature; there will be as much classes as defined component natures.
- Three examples of different natures are resistors, motors and electronic devices. Not all features are relevant when looking only at a component nature.
- THD total harmonic distortion
- the second classification step is far different because the number of clusters to identify is generally large and unknown; no prior information is generally available about a number of appliances contained in a house, so no prior information is available concerning the identification of components within their nature.
- a DBScan (Density Based Spatial Clustering for Applications with Noise) classification algorithm is preferably used. Looking at the steady states, the components will differ for instance by a magnitude of their consumption from others within a component nature cluster. As far as the transient state features are concerned, two nearly identical but different components could have different transient shapes.
- the number of clusters does not have to be specified up front when using the DBSCAN algorithm, the number of clusters which is found by the DBSCAN algorithm has been observed not always to correspond with the actual number of electrical components in the system.
- the above methodology in particular regarding the second step above, may work in a controlled set-up, it has been found that problems (that are sometimes important) remain in real-life application of the above clustering and component identification methods. More in particular, the fact that the number of clusters to identify is unknown for any household, company, industry or other environment, poses severe problems for present state of the art clustering techniques within the field of NIALM.
- WO2009/103998 discloses a method of inference of appliance usage from a point measurement on a supply line, said supply line being common to multiple appliances and/or components of appliances comprising the steps of:
- This document does not seem to disclose how to separate clusters of events in an automated method, independent of the system is being monitored.
- This document discloses the possibility of using a look-up table in which ranges of clusters' properties are provided that correspond to particular appliances and appliance components. Obviously, such a look-up table technique will depend highly on the quality and continuous updating of the look-up table, and will allow identification of a limited number of electrical components only.
- this document discloses a step- wise cluster identification on the basis of the distance of events, whereby events within a pre-set distance are assumed to belong to the same cluster, after which the events belonging to the cluster with the greatest number of events are removed from the data, and the next cluster is identified.
- Prior art techniques for identifying electrical components within a system from a measured electrical power consumption signal have been noticed to require extra input with respect to the appliances comprised in the system, either by simply providing the method with characteristic signals of each single appliance up front, or by providing an extensive look-up table of characteristic signals of appliances.
- Such techniques could be deemed impractical for systems in which the number or type of appliances changes over time, such as a common household or a company, and could be deemed unworkable as it requires constant updating of a look-up table with characteristic signals from all possible appliances.
- Prior art techniques for identifying electrical components within a system from a measured electrical power consumption signal have been seen to be not completely reliable with respect to the grouping of events extracted from the measured signal into clusters. Furthermore, prior art techniques do not seem to allow fully computing or estimating the uncertainty or error which arises from the clustering of the events.
- the present invention provides a NIALM method which allows improved component detection, in particular in the case no prior information on the components or appliances in the system is known. In such a case, the number of components (N) is a priori unknown and an automated determination of this number of components is required. Hereby, each correctly detected event can be associated to one of the N components.
- the present invention hereto provides an improved clustering technique for events, which in cooperation with the other steps of the method, allows better identifying electrical components.
- the ideal case for the evaluation of the state sequence is as follows: first, events correspond to on/off switching of components and there are no fictitious events and, second, for a given component, signatures of both on and off transitions belong to one single cluster. Further, no events related to other components share that cluster.
- components are directly represented by clusters, their operations correspond to time intervals between on and off events and the evaluation of their state sequence simply consists to pair successive on and off events that are within identical clusters. The aggregate consumption is then modeled by combining the individual sequences.
- one cluster corresponds to one electric component and inversely.
- several clusters might correspond to a single component, e.g. if turn-on and turn-off features differ, and several components might correspond to a single cluster, e.g. a household or a company which own many similar electrical components. Therefore, prior art methods using clustering do not automatically lead to component detection.
- the present invention provides a NIALM method which allows an improved component detection, in particular in case a system comprises one or more components having different turn-on and turn-off features and thus giving rise to separate clusters, or in case of a system comprising many similar electrical components which gives rise to events which are attributed to one cluster, or in any of the above four cases SI to S4.
- the present invention hereto provides an improved component detection technique, which in cooperation with the other steps of the NIALM method, allows to better identify electrical components on the basis of clusters of events.
- the present invention provides a non-intrusive appliance load monitoring (NIALM) method for monitoring (preferably electrical) components of (preferably electrical) appliances in a system.
- NIALM non-intrusive appliance load monitoring
- the NIALM method comprises the steps of: detecting events in an electrical signal (that is preferably a measured electrical signal), said electrical signal comprising power consumption information of said system, said events representing state transitions of the appliances in the system;
- clustering events into a set of clusters on the basis of their event signatures; - identifying components on the basis of said set of clusters;
- said clustering comprises an initial clustering of said events into an initial set of clusters on a basis of a first clustering criterion, and a subsequent reclustering of at least one of said initial clusters on the basis of a second clustering criterion different from the first clustering criterion.
- said first clustering criterion is computed by taking into account event signatures from subtantially all events detected in said electrical signal
- said second clustering criterion is computed by taking into account event signatures of essentially only events in said one initial cluster which is to be reel uste red.
- said electrical signal is a measured electrical signal.
- the NIALM method can be used to identify (preferably electrical) components in a system on a basis of a (preferably measured) electrical signal comprising power consumption information of said system.
- the identification of the (preferably electrical) components allows a further identification of (preferably electrical) appliances, which is particularly useful in systems where the number and/or nature of appliances is not known a priori, or can change in time. Therefore, in a preferred embodiment, the method comprises a step of identifying appliances from the components on the basis of the identified (preferably electrical) components.
- the NIALM method can also be used for condition monitoring of an appliance. Typically, but not in a limitative way, condition monitoring of an appliance can be performed to check if or when the appliance is functioning normally or abnormally, and/or to give an alarm if a malfunction is observed.
- the appliance which is being monitored using the NIALM method of the present invention can be identified using a method as described in this document, or by another method.
- the method comprises providing energy consumption of appliances or of components.
- Energy consumption of appliances or of components can be given in a form of a report to a user of an appliance or of the system, or to the electricity provider to detect malfunctions or to optimize energy efficiency either for a single system or for a number of systems.
- the present invention further provides for a recursive clustering method for clustering events obtained from a (preferably measured) electrical signal comprising power consumption information of a system, said events being at least partially characterized by an event signature, said recursive clustering method comprising : an initial step of clustering events of an initial event set into a set of clusters using a clustering criterion for deciding whether two events belong to the same cluster and/or whether an event belongs to an existing cluster, said clustering criterion being computed on a basis of event signatures from substantially all events from said initial event set; and
- the present invention also provides for a method for clustering events detected in a (preferably measured) electrical signal comprising electrical power consumption information of a system, each of said events being at least partially characterized by an event signature comprising a set of event features, the method comprising the steps of:
- step (e) optionally, if two or more sub-clusters are obtained after step (d)(iii), performing step (d) taking at least one of said sub-clusters from step (d)(iii) as a cluster to be sub-clustered.
- step (d) is preferably performed recursively for each cluster of said set of clusters and/or for all sub-clusters obtained by performing step (d) until no further sub-clustering in two or more sub-clusters is obtained.
- the events relate to on-off switching of components of appliances in the system .
- Events can be classified by their signatures, which refer to a set of properties or features which characterize transitions in the measured signal. These properties or features can be computed from a current waveform extracted from the measured signal, and preferably from a delta waveform extracted from the measured signal. More preferably, features are computed on the basis of a frequency spectrum of the delta waveform, for example features could relate to harmonic magnitudes or phases.
- Preferred features which are computed to characterize events are : active power (P), reactive power (Q), fundamental frequency (Hi), odd harmonics preferably of order 3 to 13 and preferably normalized to the amplitude of the fundamental frequency com ponent of the current, continuous component and/or second order harmonics which a re preferably normalized to the amplitude of the fundamental frequency com ponent of the current, total harmonic distortion (THD), ratio
- a signature of an event comprises a time series of a quantity derivable from the electrica l signa l, preferably a quantity being the active power and/or the current.
- said time series of sa id quantity comprises a set of va lues for sa id number ordered chronologica lly.
- a decision criterion When clustering events, a decision criterion, or clustering criterion, needs to be applied to decide whether or not an event belongs to a cluster, or two events belong to the same cluster.
- a criterion preferably com prises computing an inter-event distance to compare events, said inter-event distance being compared to one or more delimiting va lues, which are parameters of the clustering a lgorithm .
- the clustering criterion is updated su bstantia lly only on the basis of the events which have been identified previously as being part of the pa rent cluster, i.e. events which could have been part of the initial event set, but were not identified as being part of the parent cluster, are not taken into account when updating the clustering criterion for reclustering or subclustering the parent cluster.
- events are compared by computing an inter-event distance.
- This distance could preferably, and in particular for comparing events characterized by steady-state features, be a Euclidean distance or a Mahalanobis distance, or any combination thereof.
- This distance could preferably, and in particular for comparing events characterized by transient-state features, be a Minkowski distance, a Dynamic Time Warping (DTW) distance, a Longest Common Subsequence (LCSS) distance, or any combination thereof.
- distances which are at least partly in a one-to-one correspondence with the above-mentioned distances or any combination thereof.
- events are characterized by steady-state features or transient-state events.
- the clustering criterion is computed on the basis of the principal components of events
- events which are characterized by steady- state features are clustered as these events have been noticed to comprise principal components which can be identified unambiguously in a principal component analysis, e.g. by means of looking at those linear functions of the event features which maximize the variance, e.g. as further explained in I. T. Jolliffe, "Principal Component Analysis", Vol. 30, Springer Series in Statistics, Springer, 2 nd ed., 2002.
- said clustering and/or reclustering and/or subclustering is performed using a density-based algorithm, such as a DBSCAN algorithm, an OPTICS algorithm and/or a DBCLASD algorithm, said algorithm based on a density defined by the amount of events which can be found in an e-neighborhood of a point of the cluster, said e-neighborhood of a point p defined as
- N £ (p) ⁇ q e X ⁇ d(p, q) ⁇ e ⁇ , where X is the signature space, preferably a feature space, d() represents a distance function between points and/or events, and e is an inter-event distance delimiting the neighborhood, whereby two kinds of points are defined : core points which have a density ⁇ N e (p) ⁇ MinPts, MinPts being a number higher than 1, and preferably lower than 20, and border points which belong to an e-neighborhood of a core point without being a core point itself, whereby two points are defined as density-reachable if there exists a sequence of points such that each one belongs to the an e-neighborhood of its predecessor, the latter being a core point, whereby two points are defined as density- connected if they are density-reachable from a common point, whereby said clustering criterion comprises evaluating if an event is density-connected to a point and/or an event in a cluster.
- said clustering criterion is computed and/or recomputed by computing and/or recomputing e and/or MinPts, preferably e.
- said events are characterized by steady-state features and/or transient- state features.
- prior art techniques seem to rely on external input, e.g. maximal inter-event distance or minimal number of events, in order to determine which event belongs to which cluster or which cluster can be deemed complete. Although such prior art techniques allow identification of relatively simple components at a reasonable level, these techniques seem to fail when applied to more intricate systems, such as households or companies, which comprise electrical appliances with rather complicated electrical components.
- the present invention also provides a component detection method comprising the steps of: obtaining clusters of events obtained from a (preferably measured) electrical signal comprising power consumption information of a system, said events comprising on-events and off-events, whereby each of said events comprises a cluster ID which allows to identify the cluster to which the event belongs, preferably whereby said clusters of events are obtained using a method according to the present invention;
- each paired event comprising paired cluster ID information representing the cluster ID of the on-event and the cluster ID of the off-event in said paired event;
- said step of pairing on-events and off-events comprises extremizing a cost function which depends on variations in energy, in power, in current, in energy fit errors, in power fit errors, in current fit errors or in any combination thereof, preferably in a linear combination of at least two of said variations. More preferably, this cost function depends on variations in power and/or in power fit errors.
- the present invention further concerns the use of any of the clustering methods and/or the component detection method in a non-intrusive appliance load monitoring method and/or for condition monitoring of an appliance.
- the component detection method of the present invention is particularly suited for a system wherein electrical appliances are on/off components, or consist of such on/off components, and, their consumption can be modeled by constant power segments during steady states.
- any on-event should be followed by an off- event with similar power step (with opposite sign).
- "Explained power steps” are the ones for which complementary power steps can be identified following the idea that: if a component has been turned on at time t on , it should be turned off at a time t off > t on . Further, constant power draw is assumed and we search for the pair (t on , t off ) minimizing fit-errors.
- the present invention also concerns a processing unit arranged for performing a NIALM method, any clustering method and/or a component dtetction method according to the present invention.
- processing unit is arranged for performing a recursive clustering method or a method for clustering according to the present invention, and a component detection method according to the present invention.
- the present invention further concerns a device, preferably a computer-mountable and/or meter-mountable device, comprising instructions for executing a NIALM method, any clustering method and/or a component detection method according to the present invention.
- 'meter-mountable' device refers to a device which can be linked to an electrical meter, such as an electrical meter which can be found in a residence, a business or any system comprising an electrically powered device, and which measures the consumption, and/or optionally the generation, of electrical energy.
- said processing unit is arranged for performing a recursive clustering method or a method for clustering according to the present invention, and a component detection method according to the present invention.
- said device is a separate device, i.e. which can be releasably linked to a computer or a meter or a system comprising electrical appliances.
- said device is integrated in a meter, such as a smart meter, or in a circuit breaker of said system.
- the present invention further concerns a NIALM system comprising a client device and a server device, whereby said client device and server device are linkable, and optionally are linked, and whereby said client device and server device are configured to together execute a NIALM method according to the present invention, a recursive clustering method according to the present invention, a method according to the present invention, and/or a component detection method according to the present invention.
- This system allows performing the separate steps of the methods according to the present invention on the client device or on the server device or on both. In particular certain steps of the methods of the present invention can be performed on the client device while the remaining steps of the methods can be performed on the server device, i.e.
- the client device can be configured to perform a first subset of the steps of the methods of the present invention, while the server device can be configure to perform a second subset of the remaining steps of the methods of the present invention.
- said client device is configured to obtain a measured electrical signal comprising power consumption information of a system.
- the client device is configured to perform the following steps: detecting events in a (preferably measured) electrical signal comprising power consumption information of said system, said events representing state transitions of the appliances in the system;
- the server device is configured to perform the following steps: obtaining said information representing said detected events from said client device;
- clustering events into a set of clusters on the basis of their event signatures; - identifying components on a basis of said set of clusters;
- the client device is configured to perform the following steps: detecting events in a (preferably measured) electrical signal comprising power consumption information of said system, said events representing state transitions of the appliances in the system;
- the server device is configured to perform the following steps: obtaining said information representing said event signatures from said client device;
- clustering events into a set of clusters on the basis of their signatures
- the client device is configured to perform the following steps: detecting events in a measured electrical signal comprising power consumption information of said system, said events representing state transitions of the appliances in the system;
- clustering events into a set of clusters on the basis of their signatures
- the server device is configured to perform the following steps: obtaining said information representing said set of clusters from said client device;
- said clustering comprises an initial clustering of said events into an initial set of clusters on a basis of a first clustering criterion, and a subsequent reclustering of at least one of said initial clusters on a basis of a second clustering criterion different from the first clustering criterion.
- a NIALM system with a client device and a server device provides a number of advantages, including, but not limited to: the possibility to optimize and update the algorithms in an easy way, in particular for the algorithms of the server device, and in particular the clustering algorithm,
- the present invention also concerns a database comprising information representing identified components in a system, preferably of a multitude of systems, said information obtained using a NIALM method according to a method disclosed in this document.
- the database also comprises information representing identified appliances and/or information representing energy consumption of the appliances and/or components in said system or said multitude of systems.
- Figure 1 shows three methods to evaluate similarities between time series: the Minkowski distance, the dynamic time warping (DTW) and the least common subsequence (LCSS).
- Minkowski distance the Minkowski distance
- DTW dynamic time warping
- LCSS least common subsequence
- Figure 2 shows how the amount of outliers mainly decreases with k, the amount of nearest neighbors in the k-dist criterion.
- Figure 3 shows that when increasing k, the decrease of outlier points is mainly located in the origin of the P-Q plane but small clusters are then considered as outliers.
- Figure 4 illustrates substructures which are better emphasized if principal components are evaluated for specific data subsets.
- Figure 5 shows a modeled aggregate power based on paired events. The evaluation is independent from the classification results.
- Figure 6 shows that, once pairs have been assigned to components, the considered power steps derive from the clusters.
- Figure 7 shows sequences of on-off cycles which are generated for nine appliances such that their aggregate consumption is not cyclic over the three-hour period.
- Figure 8 shows nine appliances which draw specific current waveforms.
- Figure 9 shows that the microwave is not a two state-device. Moreover, it exhibits cycling operations depending on the average power asked by the user.
- Figure 10 shows events and their features - events corresponding to: the kettle, the vacuum cleaner and the microwave oven.
- Figure 11 shows events and their features, zooming in at the origin of the P-Q plane.
- Figure 12 shows that some turn-on transients systematically lead to several detections.
- Figure 13 shows power variations of the ventilator which are significant compared to the power it consumes. Outliers risk to be associated with clusters of such a variable and low power device.
- Figure 14 shows that clusters which are obtained after a first clustering iteration do not correctly reveal the complete data structure.
- Figure 15 shows that an iterative clustering procedure according to the present invention efficiently captures the data substructure.
- Figure 16 shows a classification according to the highest membership value based on the Mahalanobis distance.
- Figure 17 shows that there is almost no confusion between clusters of transient patterns. However, no clusters are discovered for the kettle, the ventilator and the vacuum cleaner.
- Figure 18 shows that component ID are assigned to recurrent pairs of nonoutliers. As a result, equivalent component ID are assigned to on and off event clusters.
- Figure 19 shows first, the modeled aggregate power is evaluated with all paired events; and second, electrical components are discovered from recurrent pairs of nonoutliers. Finally, state sequences are evaluated based on classified event pairs.
- Figure 20 shows paired events are assigned to components according to their cluster identifiers.
- Figure 21 shows state sequences evaluated from the event pairs assigned to pure component. Results from the state sequence evaluation are given in the top plots and the measurement in the bottom ones.
- the drawings of the figures are neither drawn to scale nor proportioned. Generally, identical components are denoted by the same reference numerals in the figures.
- NIALM techniques are either pattern recognition-based or optimization-based approaches, also referred to as event-based or non event-based.
- samples correspond to state changes of appliances, referred to as 'events'.
- event-based NIALM signatures are used to associate events to appliances. State sequences of appliances then result from classification algorithms.
- supervised approaches changes of appliance states are matched one- by-one to known signatures. Where learning is unsupervised, the recurrence of signatures is exploited to recognize appliances.
- the present invention concerns an unsupervised, event-based NIALM method according to claim 1.
- the present invention also concerns methods which optimize the NIALM method, including a recursive clustering method for recursively clustering similar events according to claim 5, a method for clustering similar events taking into account principal components of the event's signature according to claim 9, a component detection method according to claim 13, as well as the use of any of these methods or their combination in a NIALM method or for condition monitoring of an appliance.
- the present invention also concerns a processing unit arranged for executing any of these methods and a device comprising instructions for carrying out these methods, as well as a NIALM system comprising a client device and a server device as specified in claims 19 and 20 and a database according to claim 21.
- a compartment refers to one or more than one compartment.
- the value to which the modifier "about” refers is itself also specifically disclosed.
- the DBSCAN algorithm refers to density-based spatial clustering of applications with noise.
- a distance parameter i.e. an inter-event distance parameter
- a zone is considered as dense if a sufficiently large number of points, i.e. events, is found within the given neighborhood.
- DBSCAN hereby requires only two parameters and no range of values must be specified for the number of clusters to investigate. Instead, the number of clusters is an output of the algorithm.
- M. Ester, H.-P. Kriegel, J. Sander and X. Xu "A Density-Based Algorithm for Discovering Clusters in Large Spatial Databases with Noise", Knowledge Discovery and Data Mining 96: 226-231, 1996 for more detailed information about this algorithm.
- the OPTICS algorithm refers to an algorithm for ordering points to identify the clustering structure.
- the DBCLASD algorithm refers to a distribution-based clustering algorithm for mining in large spatial databases.
- Such a criterion preferably comprises computing an inter-event distance to compare events. Therefore, in an embodiment, events are compared by computing an inter-event distance.
- This distance could preferably, and in particular for comparing events characterized by steady-state features, be a Euclidean distance or a Mahalanobis distance, or any combination thereof.
- This distance could preferably, and in particular for comparing events characterized by transient-state features, be a Minkowski distance, a Dynamic Time Warping (DTW) distance, a Longest Common Subsequence (LCSS) distance, or any combination thereof.
- distances which are at least partly in a one-to-one correspondence with the above-mentioned distances or any combination thereof.
- Euclidean distance As used herein, the term “Euclidean distance”, “Mahalanobis distance”, “Minkowski distance”, “Dynamic Time Warping (DTW) distance” and “Longest Common Subsequence (LCSS) distance” refer to specific types of distances between events, which are discussed in what follows.
- X and Y refer to event signatures characterized by a set of N features X, and y, resp.
- these features are normalized, e.g. with respect to the minimum and maximum values of the features from the events which are taken into account to define or compute the clustering criterion, according to wherein x, refers to the i'th feature in the signature, and x i;n refers to the normalized feature value.
- Mahalanobis distance e.g. with respect to the minimum and maximum values of the features from the events which are taken into account to define or compute the clustering criterion, according to wherein x, refers to the i'th feature in the signature, and x i;n refers to the normalized feature value.
- X c (n c x p) be a set of n c p-feature signatures grouped within a cluster.
- the variance-covariance matrix of cluster X c is
- x be a signature whose distance w.r.t. X c is to be evaluated and X c the vector of mean feature values of X c .
- the Mahalanobis distance is defined as follows:
- the Mahalanobis distance better deals with the dispersion of clusters along different directions than e.g. the Euclidean. Therefore, the Mahalanobis distance is better suited to identify the cluster to which outlier events belong.
- Outliers are isolated points in the feature space. They do not systematically exhibit extreme feature values. Instead, the distance to their nearest neighbors is high compared to the one of non outlier points. Consequently, the distance of a point to its k nearest neighbors is an image of its likelihood to be an outlier. We refer to such a distance as k-dist. Outliers are then points with the highest k-dist values.
- a threshold on k-dist values could be defined to decide whether or not points are outliers.
- a maximum nonoutlier (MNO) value of the k-dist distribution could be used : it is defined as the highest observed value below the upper quartile increased by 1.5 times the interquartile range.
- MNO maximum nonoutlier
- k can preferably take any value between 1 and 20, preferably between 4 and 15, more preferably 10.
- Dynamic time warping (DTW) distance and least common subsequence (LCSS) distance The principle of the dynamic time warping (DTW) and the least common subsequence (LCSS) are illustrated in figure 1, which shows three methods to evaluate distances, in particular distances between time series: the Minkowski distance, the dynamic time warping (DTW) and the least common subsequence (LCSS).
- DTW allows to deal with sequences with different lengths and different localization in the window.
- LCSS further allows to skip samples considered as outliers in the pattern.
- the globally optimal alignment obtained with DTW is a suitable way to measure the similarity.
- the DTW is chosen over the LCSS method.
- the DTW has the advantage that the entire sequences must be matched which avoids that spurious transient states be matched to subparts of true transients. In other words, DTW performs matching with time shift analysis robust to positive false detections.
- the average distortion is used instead of the total distortion in order to normalize for different segment lengths.
- a distance metric d(x,, y,) between two elements x, and y (e.g. Euclidean distance)
- the total cost of a warping path p is i
- the optimal warping path p* is the one minimizing the total cost among all possible paths.
- the distance DTW(X, Y) is defined as the total cost associated to the optimal path p* :
- the average DTW distance is DTW(X,Y)/L where L is the length of the optimal path p*.
- Example 1 Feature extraction on the basis of principa l components
- Electric signatures have been evaluated because they characterize the electric consumption of components, not because they allow to classify components.
- Features might be correlated a nd less features could then be used to represent the data structure.
- some features might be better than others to reveal similarities and differences between signatures. Considering non relevant features would add similarities between events.
- PCA principal component analysis
- PCA principal component analysis
- the first step is to look for a linear function ⁇ ⁇ x of the elements of x having maximum variance, where Oi is a vector of p constants an, a 12 , lp , and ' denotes transpose, so that
- the procedure is repeated so that at the k'th stage a linear function a x is found that has maximum variance subject to being uncorrelated with ⁇ ⁇ , ⁇ ' 2 ⁇ , ..., a' k -ix.
- the k'th derived varia ble, a' k x is the k'th PC.
- Detailed equations for the principal component analysis can be found in the work by Jolliffe.
- PC principa l components
- Outliers might impact the result of the principal component analysis. They are mainly related to errors from the underlying data generation, e.g . they can be related to errors insignal segmentation leading to outliers.
- the errors can be due to noise or non-linearities, or due to transients which can be very long and hence could be seen as two or more different events. They are also rare events resulting from simultaneous on/off switching of components. They should be removed from the data set before searching for the principa l components.
- Outliers a re isolated points in the feature space. They do not systematica lly exhibit extreme feature values. Instead, the distance to their nearest neighbors is high com pared to the one of non outlier points.
- Each feature x i;n ranges between 0 and 1 and the k-dist of events is the Euclidean distance between their normalized feature vectors and the one of their kth nearest neighbor.
- a threshold on k-dist values must be defined to decide whether or not points are outliers.
- FIG. 2 illustrates the relative evolution of the number of outliers according to the value of k for five different data subsets. We see that increasing the number of neighbors considered in the k-dist decreases the number of outliers. Variations are mainly located in the origin of the P - Q plane which corresponds to false positive detections and small loads whose signature evaluation is impacted by fluctuations of the mixture. However, increasing k also leads to considering points of small clusters as outliers. These two aspects are illustrated in figure 3.
- Non outlier PCA Detecting loads consuming low power compared to the aggregate consumption is the biggest challenge of NIALM. Consequently, it is desired that the false positive detections be correctly identified as outliers. This will facilitate the detection of points which cluster around the origin of the P - Q plane and are related to signatures of small loads. Taking this reasoning into account and the choice of DBSCAN parameters presented below, k is set to 10. 3. Non outlier PCA
- PCA selects features in order to maximize the variance of the data, variables with higher amplitudes will be given higher weight and will be systematically considered in the first PCs. For instance, the active power ranges from some dozens of watts up to some kilowatts whereas normalized harmonic amplitudes range between 0 and 1. Data should then be reduced before the PCA be applied in order to give the same weight to the different features.
- step c normalize with the transformation of step a
- PCA has been applied to the normalized dataset whose normalized active and reactive power are shown in the upper plot of figure 4.
- the two first PC are plotted in the leftmost below plot.
- the PCA is also applied to the same data set limited to events with active power first higher and then lower than 500W (respectively in the middle and right-most plots).
- the weights assigned to the features in the first PC are given in Table 1 and the percentage of total variance of the three first PCs in Table 2.
- Table 1 Weights of the features in the PCI for the data in figure 4
- the first PC is the combination of most features. But for both data subsets, only some features have relatively high weights: the PCI of the > 500W' subset involves the fundamental components whereas the one of the ⁇ 500W subset is mainly limited the harmonic content and the reactive power. This could have been expected.
- PC2 is needed in the case of the entire data set to satisfy the 85% criterion.
- the three first PCs must be used.
- Example 2 Component discovery
- the method to discover instances of electrical components from clusters and the event time sequence can be illustrated by the following example.
- cluster identifiers of paired events yield links between clusters which reflect electrical components.
- Pairing events by minimizing energy errors and error variations Considering two-state devices, if a component is turned off, it has previously been turned on. Following this, off-events should be paired to on-events to represent occurrences of unknown components. Constant power loads being assumed, the aggregate power is straightforwardly obtained for any sample k as the power of all pairs with k on ⁇ k ⁇ k off . If y k is the power measured at sample k and AP k is a vector whose nonzero entries correspond to the signed power step values of paired events, the modeled aggregate power at sample k, y k , can be written as y 0 stands to account for loads that would have been switched on before the first sample of the observation window.
- the ideal set of event pairs is the one for which the modeled aggregate power is the closest to the measured one. At each time, the power of the unknown components considered as active should sum up to the measured aggregate power. However, segmentation errors or simultaneous switch events lead to impossibility to truly fit the measured power simply by pairing events. Their impact on subsequent decisions should be minimized. This can be achieved by penalizing the bestfit solution with the total variation of the fit-error. We formalize this in the next paragraphs.
- y k is the median value of the power samples within steady state k.
- energy fit-errors and variations of energy fit-errors can be defined analogously for different time periods. This corresponds to modeling the observed aggregate power as constant segments.
- median values instead of mean values allows the method to be more robust to segmentation errors: if there is a non detected power step within a steady state (i.e. there are actually two steady states), the longer one determines the constant power segment.
- Pairing on-events and off-events can be done by extremizing a cost function which depends on variations in energy, in power, in current, in energy fit errors, in power fit errors, in current fit errors or in any combination thereof, preferably in a linear combination of at least two of said variations, more preferably variations in power and in power fit errors.
- the cost function l ⁇ k ⁇ is evaluated with summations limited to the local window.
- Step b allows to hide deviations from ideal fit due to previous erroneous pairing since is systematically initialized to the first sample in the local window.
- step c the minimum cost l ⁇ k ⁇ obtained with the different candidates in the window is compared to the costs of two other cases: • ⁇ pair with event k' ( ⁇ ⁇ ) : AP k is kept equal to zero, considering that event k is a fictitious transition
- the next step is conditioned by the case that minimized the cost function :
- the event is not paired. It probably results from a false positive detection or its partner has not been detected.
- the local window size is enlarged and new candidates are investigated in the new time interval.
- a component can be defined as a recurrent link between clusters required to appropriately model the aggregate power.
- Example 3 Results of laboratory testing
- Algorithms were developed considering both laboratory and field data. However, to evaluate performances at the different stages of the NIALM procedure, a reliable ground truth is required. The latter is only available with laboratory data. Algorithms have been developed with the aim of providing parameter settings which are not tuned to specific data. Consequently, a new data set was generated in the laboratory (ELDA) specifically for this performance evaluation.
- This appliance has more than two states: it consists of a light bulb, a motor and a magnetron. Furthermore, it can be noticed that, according to the average power asked by the user, the magnetron is turned on and off during its operations. Consequently events were manually added to the ground truth to account for these state changes.
- the rate of correct detections and the amount of missed detections can be evaluated for each individual load.
- the figures are reported in Table 3. We observe a global rate of 91%.
- the detected events are plotted in the ⁇ -AQ plan in figure 10 and figure 11.
- Each couple color-symbol represents one appliance, as indicated by the ground truth.
- Turn-on events with slow time decay lead to several detections. In the studied data set, this impacts the segmentation of turn-on transients corresponding to the mixer and the vacuum cleaner, as shown in figure 12. This leads to different steady state feature values for turn-on and turn-off events. This can be seen with the green point cluster in figure 10 for the vacuum cleaner and with the red point cluster in figure 11 for the mixer. The events corresponding to both these loads form substructures. Furthermore, fictitious second events limit the search for the knee in the power trace which prevent from properly capturing the transient dynamics.
- Second events of microwave turn-on transients are also divided in two clusters.
- Each column corresponds to one cluster: the first row gives the cluster identifier (ci D ) corresponding to figure 15 and events are distributed in rows according to the ground truth. Outliers are reported in the last column along with nondetected events.
- Rows correspond to appliances as defined by the ground truth. The last row corresponds to events for which no identifier has been assigned in the ground truth. They are mostly false positive detections, but second events of microwave turn-on transients also lie in this category.
- microwave oven events assigned to clusters 2 and 14 correspond to the light bulb belonging to this appliance.
- Ta ble 6 All a ppliances have a pure or quasi pure group of event pairs that can be used to learn component signatures.
- Table 7 Energy assigned to individual components (in Wh). The link between the components and the devices is manually defined. T (True) stands for energy consumed by the appliance and properly assigned to a component state sequence. F (False) stands for energy assigned to a state sequence whereas the corresponding appliance exhibits no consumption. 9 4 5 1 12 2 6 10 7 8 3 11 /
- Table 8 Paired events are assigned to components according to their cluster identifiers.
Landscapes
- Engineering & Computer Science (AREA)
- Physics & Mathematics (AREA)
- General Physics & Mathematics (AREA)
- Business, Economics & Management (AREA)
- Economics (AREA)
- Theoretical Computer Science (AREA)
- Data Mining & Analysis (AREA)
- Human Resources & Organizations (AREA)
- Strategic Management (AREA)
- Mathematical Physics (AREA)
- Mathematical Optimization (AREA)
- Tourism & Hospitality (AREA)
- Operations Research (AREA)
- General Business, Economics & Management (AREA)
- Marketing (AREA)
- Pure & Applied Mathematics (AREA)
- Mathematical Analysis (AREA)
- Computational Mathematics (AREA)
- Health & Medical Sciences (AREA)
- Bioinformatics & Cheminformatics (AREA)
- Bioinformatics & Computational Biology (AREA)
- Life Sciences & Earth Sciences (AREA)
- Evolutionary Biology (AREA)
- Power Engineering (AREA)
- Quality & Reliability (AREA)
- Probability & Statistics with Applications (AREA)
- Entrepreneurship & Innovation (AREA)
- Algebra (AREA)
- Game Theory and Decision Science (AREA)
- Databases & Information Systems (AREA)
- Software Systems (AREA)
- General Engineering & Computer Science (AREA)
- Development Economics (AREA)
- Public Health (AREA)
- Water Supply & Treatment (AREA)
- General Health & Medical Sciences (AREA)
- Primary Health Care (AREA)
- Remote Monitoring And Control Of Power-Distribution Networks (AREA)
Abstract
The current invention concerns a NIALM method for monitoring the electrical appliances in a system comprising the steps of: detecting events in a measured electrical signal comprising power consumption information of said system, said events representing state transitions of the appliances in the system; characterizing said events by an event signature, taking into account differences between steady states before and after the event and/or taking into account transient states located between steady states before and after the event; clustering events into a set of clusters on the basis of their signatures; identifying components on the basis of said set of clusters; identifying appliances from the components identified in the previous step;optionally providing appliances energy consumption, characterized in that said clustering comprises an initial clustering of said events into an initial set of clusters on the basis of a first clustering criterion, and a subsequent reclustering of at least one of said initial clusters on the basis of a second clustering criterion different from the first clustering criterion.
Description
Improved non-intrusive appliance load monitoring method and
device
Technical field
The invention pertains to the technical field of automatic appliance detection, in particular Non-Intrusive Appliance Load Monitoring (NIALM), which refers to the automated detection of the state of appliances, e.g. household appliances, industrial appliances, a company's appliances, etc., from a total energy consumption signal. More in particular, the present invention relates to obtaining information at the component level of electrical appliances in a system, e.g. a household or a company, from a measurement of the total electricity consumption of said system. This information at the component level may include but is not limited to: the presence, the state and/or the energy consumption. The system may also be comprised of a single or a few appliances, whereby the methods, devices and/or systems of the present invention can be used for condition monitoring of one or more appliances. Background
WO 2012/160062 Al discloses a method for detecting transitions in a measured signal e.g. a discrete time signal and current waveform, which is induced by elements i.e. components of appliances, of a physical system i.e. a house. The method involves generating a residual signal, i.e. transition likelihood signal, from a measured signal, i.e. a discrete time signal, where the residual signal is provided with high amplitude when transitions occur and with low amplitude in other cases. Rules concluding that transitions occur when the residual signal is larger than a threshold stable value, are provided, where the stable value is automatically defined from local values of the residual signal and defined as a function of local background noise. The measured signal is filtered before generating residual signal. The method enables defining a time index corresponding to end of transient state to maximize a distance between transient states and a straight line passing through the transient state, thus improving separation between transient and steady states, and hence detecting transitions in the measured signal, and monitoring automatic-setup non-intrusive appliance load for identifying appliances energy consumption in a reliable manner. This document also discloses an automatic-setup non-intrusive appliance load monitoring method (50) for identifying appliances energy consumption and comprising the steps of:
• detecting transitions in a measured signal;
• characterizing differences between steady states before and after these transitions;
• characterizing transient states located between steady states;
• identifying components;
• identifying appliances from the components identified in the previous step;
• providing appliances energy consumption. The step of identifying components is hereby performed by selecting spectral features or features describing a current waveform for the steady states, with the possibility of the chosen features being grouped together in well-separated clusters for different components.
Hereby, a first step of component identification consists in clustering and classifying the components according to their nature (motor, resistor, heater, television are examples of different natures). A second step classifies the components within their nature cluster; this is the component classification itself.
Identifying a nature of a component is a classification problem with a predefined number of classes. Therefore, a K-means algorithm is preferably chosen. A K- means method is a clustering technique that partitions n observations into K clusters, where K is a predefined value. It classifies the components according to their nature; there will be as much classes as defined component natures. Three examples of different natures are resistors, motors and electronic devices. Not all features are relevant when looking only at a component nature. Preferably, when searching a nature of a component, one can consider THD (total harmonic distortion), or ratio Q/P (Q = reactive power and P = active power).
The second classification step is far different because the number of clusters to identify is generally large and unknown; no prior information is generally available about a number of appliances contained in a house, so no prior information is available concerning the identification of components within their nature. Hence, a DBScan (Density Based Spatial Clustering for Applications with Noise) classification algorithm is preferably used. Looking at the steady states, the components will differ for instance by a magnitude of their consumption from others within a component nature cluster. As far as the transient state features are concerned, two nearly identical but different components could have different transient shapes. Note hereby, that although the number of clusters does not have to be specified up front when using the DBSCAN algorithm, the number of clusters which is found by the DBSCAN algorithm has been observed not always to correspond with the actual number of electrical components in the system.
Although the above methodology, in particular regarding the second step above, may work in a controlled set-up, it has been found that problems (that are sometimes important) remain in real-life application of the above clustering and component identification methods. More in particular, the fact that the number of clusters to identify is unknown for any household, company, industry or other environment, poses severe problems for present state of the art clustering techniques within the field of NIALM.
WO2009/103998 discloses a method of inference of appliance usage from a point measurement on a supply line, said supply line being common to multiple appliances and/or components of appliances comprising the steps of:
• obtaining data from said measurement point;
• sampling power and reactive power at intervals substantially throughout periods of operation of said appliances or components of appliances corresponding to appliances or components of appliances being in ON and/or OFF modes of use;
• identifying characteristics of events by assessing power and reactive power change during an event; and by
• assessing one or more additional characteristics derivable from said power and reactive power to characterize an appliance;
· grouping events and/or cycles of events into clusters of similar characteristics; and
• inferring appliance usage based on said grouping.
This document does not seem to disclose how to separate clusters of events in an automated method, independent of the system is being monitored. This document discloses the possibility of using a look-up table in which ranges of clusters' properties are provided that correspond to particular appliances and appliance components. Obviously, such a look-up table technique will depend highly on the quality and continuous updating of the look-up table, and will allow identification of a limited number of electrical components only. Furthermore, this document discloses a step- wise cluster identification on the basis of the distance of events, whereby events within a pre-set distance are assumed to belong to the same cluster, after which the events belonging to the cluster with the greatest number of events are removed from the data, and the next cluster is identified.
Note that methods such as the one described in WO2009/103998, use a pre-set maximal inter-event distance, on the basis of which events are assigned to a cluster, whereby clusters identification is made at least partially on the basis of a minimal
number of events assigned to that cluster. Both steps introduce a certain degree of arbitrariness into the method which results both in uncertainties with respect to the correctness, as well as in errors for the resulting identified clusters as well.
Prior art techniques for identifying electrical components within a system from a measured electrical power consumption signal, have been noticed to require extra input with respect to the appliances comprised in the system, either by simply providing the method with characteristic signals of each single appliance up front, or by providing an extensive look-up table of characteristic signals of appliances. Such techniques could be deemed impractical for systems in which the number or type of appliances changes over time, such as a common household or a company, and could be deemed unworkable as it requires constant updating of a look-up table with characteristic signals from all possible appliances.
Prior art techniques for identifying electrical components within a system from a measured electrical power consumption signal, have been seen to be not completely reliable with respect to the grouping of events extracted from the measured signal into clusters. Furthermore, prior art techniques do not seem to allow fully computing or estimating the uncertainty or error which arises from the clustering of the events.
The present invention provides a NIALM method which allows improved component detection, in particular in the case no prior information on the components or appliances in the system is known. In such a case, the number of components (N) is a priori unknown and an automated determination of this number of components is required. Hereby, each correctly detected event can be associated to one of the N components. The present invention hereto provides an improved clustering technique for events, which in cooperation with the other steps of the method, allows better identifying electrical components.
More in particular, techniques proposed in the literature use prior information about parametric finite state machine (FSM) models, parameters being automatically tuned to fit to the observations. As reported by Johnson and Willsky, "Bayesian Nonparametric Hidden Semi-Markov Models", arXiv preprint, 2012, this information should be learned from the data. Since information can be extracted from the aggregate consumption as e.g. detailed in this document, one could wonder if this information can be straightforwardly computed to obtain component state sequences. Given the assumption that appliances are either on/off loads or systems of such loads, the ideal case for the evaluation of the state sequence is as follows: first, events correspond to on/off switching of components and there are no fictitious events and, second, for a given component, signatures of both on and off transitions belong to one
single cluster. Further, no events related to other components share that cluster. In this ideal case, components are directly represented by clusters, their operations correspond to time intervals between on and off events and the evaluation of their state sequence simply consists to pair successive on and off events that are within identical clusters. The aggregate consumption is then modeled by combining the individual sequences.
Actual situations differ from this ideal case in several ways: SI :
Some on/off switching are missing from the event sequence and there are fictitious events.
Why: Because there are false (positive and negative) detections in the segmentation step, or because there is noise or non-linearities.
So what: Neither complementary events to missing ones nor fictitious detections should be paired with other events. S2 :
Events of a single component are classified into different clusters.
Why: On and off event signatures might differ because of nonideal segmentation of turn-on transients or variable consumption of loads during their operations, because of noise or because of non-linearities.
- So what: On and off events of different clusters should be paired to obtain component state sequences.
S3 :
Events corresponding to different components belong to a single cluster.
Why: There are components with similar properties w. r.t. the clustering parameters (or completely similar components).
So what: Events of that single cluster must be distributed into different component state sequences.
S4 :
There are events whose signatures are combinations of others
- Why: Some on/off switching are simultaneous switch events.
So what: These events should be paired to several events distributed in different clusters.
The problem consists to find the underlying electrical components responsible for the observed aggregate consumption. Only then could states sequences be correctly evaluated and cope with the deviations from the ideal case (SI to S4).
Ideally, one cluster corresponds to one electric component and inversely. In practice, several clusters might correspond to a single component, e.g. if turn-on and turn-off features differ, and several components might correspond to a single cluster, e.g. a household or a company which own many similar electrical components. Therefore, prior art methods using clustering do not automatically lead to component detection.
The present invention provides a NIALM method which allows an improved component detection, in particular in case a system comprises one or more components having different turn-on and turn-off features and thus giving rise to separate clusters, or in case of a system comprising many similar electrical components which gives rise to events which are attributed to one cluster, or in any of the above four cases SI to S4. The present invention hereto provides an improved component detection technique, which in cooperation with the other steps of the NIALM method, allows to better identify electrical components on the basis of clusters of events.
Summary of the invention
The present invention provides a non-intrusive appliance load monitoring (NIALM) method for monitoring (preferably electrical) components of (preferably electrical) appliances in a system. The NIALM method comprises the steps of: detecting events in an electrical signal (that is preferably a measured electrical signal), said electrical signal comprising power consumption information of said system, said events representing state transitions of the appliances in the system;
- characterizing each event by an event signature, by taking into account differences between steady states before and after each event and/or by taking into account transient states located between steady states before and after each event;
clustering events into a set of clusters on the basis of their event signatures; - identifying components on the basis of said set of clusters;
optionally identifying appliances from the components identified in the previous step;
optionally providing appliances' energy consumption or providing the components' energy consumption.
Hereby, said clustering comprises an initial clustering of said events into an initial set of clusters on a basis of a first clustering criterion, and a subsequent reclustering of at least one of said initial clusters on the basis of a second clustering criterion different from the first clustering criterion. Preferably said first clustering criterion is computed by taking into account event signatures from subtantially all events detected in said electrical signal, and said second clustering criterion is computed by taking into account event signatures of essentially only events in said one initial cluster which is to be reel uste red.
Preferably, said electrical signal is a measured electrical signal. The NIALM method can be used to identify (preferably electrical) components in a system on a basis of a (preferably measured) electrical signal comprising power consumption information of said system. The identification of the (preferably electrical) components allows a further identification of (preferably electrical) appliances, which is particularly useful in systems where the number and/or nature of appliances is not known a priori, or can change in time. Therefore, in a preferred embodiment, the method comprises a step of identifying appliances from the components on the basis of the identified (preferably electrical) components.
The NIALM method can also be used for condition monitoring of an appliance. Typically, but not in a limitative way, condition monitoring of an appliance can be performed to check if or when the appliance is functioning normally or abnormally, and/or to give an alarm if a malfunction is observed. The appliance which is being monitored using the NIALM method of the present invention can be identified using a method as described in this document, or by another method.
In a preferred embodiment, the method comprises providing energy consumption of appliances or of components. Energy consumption of appliances or of components can be given in a form of a report to a user of an appliance or of the system, or to the electricity provider to detect malfunctions or to optimize energy efficiency either for a single system or for a number of systems.
The present invention further provides for a recursive clustering method for clustering events obtained from a (preferably measured) electrical signal comprising power consumption information of a system, said events being at least partially characterized by an event signature, said recursive clustering method comprising : an initial step of clustering events of an initial event set into a set of clusters using a clustering criterion for deciding whether two events belong to the same cluster and/or whether an event belongs to an existing cluster, said clustering
criterion being computed on a basis of event signatures from substantially all events from said initial event set; and
a recursive step of reclustering events belonging to a cluster into a set of sub- clusters using an updated clustering criterion for deciding whether two events belong to the same cluster and/or whether an event belongs to an existing cluster, said updated clustering criterion being computed on a basis of event signatures substantially only of events from said cluster, thereby obtaining one, two or more sub-clusters, preferably whereby if two or more sub-clusters are obtained, these two or more sub-clusters are reclustered using the present recursive step.
The present invention also provides for a method for clustering events detected in a (preferably measured) electrical signal comprising electrical power consumption information of a system, each of said events being at least partially characterized by an event signature comprising a set of event features, the method comprising the steps of:
(a) defining principal components of said event features for said events;
(b) computing a clustering criterion on the basis of the principal components of substantially all of said events;
(c) clustering said events into a set of clusters; and being characterized in that it comprises the steps of:
(d) sub-clustering at least one cluster of said set of clusters by:
(i) redefining principal components of the event features for events belonging to said one cluster, thereby substantially excluding event features for events not belonging to said one cluster;
(ii) recomputing the clustering criterion for said one cluster on the basis of the redefined principal components of substantially all of the events belonging to said one cluster;
(iii) clustering said events belonging to said one cluster in one, two or more sub-clusters;
(e) optionally, if two or more sub-clusters are obtained after step (d)(iii), performing step (d) taking at least one of said sub-clusters from step (d)(iii) as a cluster to be sub-clustered.
Hereby, step (d) is preferably performed recursively for each cluster of said set of clusters and/or for all sub-clusters obtained by performing step (d) until no further sub-clustering in two or more sub-clusters is obtained.
Typically, the events relate to on-off switching of components of appliances in the system . Events can be classified by their signatures, which refer to a set of properties or features which characterize transitions in the measured signal. These properties or features can be computed from a current waveform extracted from the measured signal, and preferably from a delta waveform extracted from the measured signal. More preferably, features are computed on the basis of a frequency spectrum of the delta waveform, for example features could relate to harmonic magnitudes or phases. Preferred features which are computed to characterize events are : active power (P), reactive power (Q), fundamental frequency (Hi), odd harmonics preferably of order 3 to 13 and preferably normalized to the amplitude of the fundamental frequency com ponent of the current, continuous component and/or second order harmonics which a re preferably normalized to the amplitude of the fundamental frequency com ponent of the current, total harmonic distortion (THD), ratio | Q/P| , maximal IM and minima l Im values of the current. As a non-limiting, but preferred example, a signature of an event k, preferably a steady-state event, could be written as Xss(k) = [P, Q, Hi, H0/i, H2/i, Hodd(3→i3)/i, THD, | Q/P| , IM, Im], wherein Hn/i refers to the n'th Harmonic normalized to the magnitude of the fundamental frequency component of the current.
In a preferred embodiment, and most preferably for transient state events, a signature of an event comprises a time series of a quantity derivable from the electrica l signa l, preferably a quantity being the active power and/or the current. Typically said time series of sa id quantity comprises a set of va lues for sa id number ordered chronologica lly.
When clustering events, a decision criterion, or clustering criterion, needs to be applied to decide whether or not an event belongs to a cluster, or two events belong to the same cluster. Such a criterion preferably com prises computing an inter-event distance to compare events, said inter-event distance being compared to one or more delimiting va lues, which are parameters of the clustering a lgorithm .
In the present invention it should be noted that in the reclustering or recursive step, whereby a parent cluster is being reclustered or sub-clustered into child clusters or sub-clusters, the clustering criterion is updated su bstantia lly only on the basis of the events which have been identified previously as being part of the pa rent cluster, i.e. events which could have been part of the initial event set, but were not identified as being part of the parent cluster, are not taken into account when updating the clustering criterion for reclustering or subclustering the parent cluster. This allows a reclustering of the parent cluster which is specifica lly ada pted for that parent cluster,
which allows zooming in on the parent cluster and use a clustering criterion which is especially selective for events in that parent cluster. This concept makes the clustering method more robust, and more applicable for all types of measured signals coming from all types of systems, and in particular for systems comprising appliances which are not frequently or regularly used, or for systems in which new appliances are added or old ones are removed. Furthermore, it has been observed that methods of the present invention provide an improved clustering for systems comprising small and/or similar loads.
Therefore, in an embodiment, events are compared by computing an inter-event distance. This distance could preferably, and in particular for comparing events characterized by steady-state features, be a Euclidean distance or a Mahalanobis distance, or any combination thereof. This distance could preferably, and in particular for comparing events characterized by transient-state features, be a Minkowski distance, a Dynamic Time Warping (DTW) distance, a Longest Common Subsequence (LCSS) distance, or any combination thereof. Also preferred are distances which are at least partly in a one-to-one correspondence with the above-mentioned distances or any combination thereof.
In a preferred embodiment of the methods for clustering events as disclosed herein, events are characterized by steady-state features or transient-state events. In the method for clustering wherein the clustering criterion is computed on the basis of the principal components of events, preferably events which are characterized by steady- state features are clustered as these events have been noticed to comprise principal components which can be identified unambiguously in a principal component analysis, e.g. by means of looking at those linear functions of the event features which maximize the variance, e.g. as further explained in I. T. Jolliffe, "Principal Component Analysis", Vol. 30, Springer Series in Statistics, Springer, 2nd ed., 2002.
In a preferred embodiment, said clustering and/or reclustering and/or subclustering is performed using a density-based algorithm, such as a DBSCAN algorithm, an OPTICS algorithm and/or a DBCLASD algorithm, said algorithm based on a density defined by the amount of events which can be found in an e-neighborhood of a point of the cluster, said e-neighborhood of a point p defined as
N£(p) = {q e X\d(p, q)≤ e}, where X is the signature space, preferably a feature space, d() represents a distance function between points and/or events, and e is an inter-event distance delimiting the neighborhood, whereby two kinds of points are defined : core points which have a
density \Ne(p) \≥ MinPts, MinPts being a number higher than 1, and preferably lower than 20, and border points which belong to an e-neighborhood of a core point without being a core point itself, whereby two points are defined as density-reachable if there exists a sequence of points such that each one belongs to the an e-neighborhood of its predecessor, the latter being a core point, whereby two points are defined as density- connected if they are density-reachable from a common point, whereby said clustering criterion comprises evaluating if an event is density-connected to a point and/or an event in a cluster.
The applicant would like to note that descriptions of neighborhoods which are in correspondence, in particular in a bijectional correspondence, with the description given here above can also be used.
In a more preferred embodiment, said clustering criterion is computed and/or recomputed by computing and/or recomputing e and/or MinPts, preferably e.
Preferably, said events are characterized by steady-state features and/or transient- state features.
The inventors have found that prior art techniques seem to rely on external input, e.g. maximal inter-event distance or minimal number of events, in order to determine which event belongs to which cluster or which cluster can be deemed complete. Although such prior art techniques allow identification of relatively simple components at a reasonable level, these techniques seem to fail when applied to more intricate systems, such as households or companies, which comprise electrical appliances with rather complicated electrical components.
This is at least partly due to a degree of arbitrariness which is introduced in the method, in particular in a step related to clustering of events, prior to identification of components on the basis of the computed clusters. Although the results for the component identification can be improved on a case-by-case basis, i.e. largely depending on the type of system which is being monitored, it is clear that the prior art does not disclose a method which can be applied to any type of system, or the systems which change over time, e.g. a household or a factory. The inventors have also found that the current invention results in a clustering of events which is clearly better than prior art techniques, and thus leads to better component identification and appliance load monitoring. The present invention further allows to quantify or objectivate any error which is made during the clustering. The information on the error allows to improve the techniques further or to identify irregularities in the system.
In a further aspect, the present invention also provides a component detection method comprising the steps of: obtaining clusters of events obtained from a (preferably measured) electrical signal comprising power consumption information of a system, said events comprising on-events and off-events, whereby each of said events comprises a cluster ID which allows to identify the cluster to which the event belongs, preferably whereby said clusters of events are obtained using a method according to the present invention;
pairing on-events and off-events from said measured electrical signal, thereby obtaining a set of paired events, each paired event comprising paired cluster ID information representing the cluster ID of the on-event and the cluster ID of the off-event in said paired event;
detecting (preferably electrical) components by identifying recurrent paired cluster ID information within said set of paired events. In a preferred embodiment, said step of pairing on-events and off-events comprises extremizing a cost function which depends on variations in energy, in power, in current, in energy fit errors, in power fit errors, in current fit errors or in any combination thereof, preferably in a linear combination of at least two of said variations. More preferably, this cost function depends on variations in power and/or in power fit errors.
The present invention further concerns the use of any of the clustering methods and/or the component detection method in a non-intrusive appliance load monitoring method and/or for condition monitoring of an appliance.
The component detection method of the present invention is particularly suited for a system wherein electrical appliances are on/off components, or consist of such on/off components, and, their consumption can be modeled by constant power segments during steady states. For these systems, any on-event should be followed by an off- event with similar power step (with opposite sign). "Explained power steps" are the ones for which complementary power steps can be identified following the idea that: if a component has been turned on at time ton, it should be turned off at a time toff > ton. Further, constant power draw is assumed and we search for the pair (ton, toff) minimizing fit-errors.
Note that the component detection method of the present invention allows not relying on prior models of the system and/or its appliances and is robust to clustering errors.
The present invention also concerns a processing unit arranged for performing a NIALM method, any clustering method and/or a component dtetction method according to the present invention. Preferably said processing unit is arranged for performing a recursive clustering method or a method for clustering according to the present invention, and a component detection method according to the present invention.
The present invention further concerns a device, preferably a computer-mountable and/or meter-mountable device, comprising instructions for executing a NIALM method, any clustering method and/or a component detection method according to the present invention. Herein, 'meter-mountable' device refers to a device which can be linked to an electrical meter, such as an electrical meter which can be found in a residence, a business or any system comprising an electrically powered device, and which measures the consumption, and/or optionally the generation, of electrical energy. Preferably said processing unit is arranged for performing a recursive clustering method or a method for clustering according to the present invention, and a component detection method according to the present invention.
In a preferred method, said device is a separate device, i.e. which can be releasably linked to a computer or a meter or a system comprising electrical appliances. In another preferred method, said device is integrated in a meter, such as a smart meter, or in a circuit breaker of said system.
The present invention further concerns a NIALM system comprising a client device and a server device, whereby said client device and server device are linkable, and optionally are linked, and whereby said client device and server device are configured to together execute a NIALM method according to the present invention, a recursive clustering method according to the present invention, a method according to the present invention, and/or a component detection method according to the present invention. This system allows performing the separate steps of the methods according to the present invention on the client device or on the server device or on both. In particular certain steps of the methods of the present invention can be performed on the client device while the remaining steps of the methods can be performed on the server device, i.e. the client device can be configured to perform a first subset of the steps of the methods of the present invention, while the server device can be configure to perform a second subset of the remaining steps of the methods of the present invention. In a preferred embodiment, said client device is configured to obtain a measured electrical signal comprising power consumption information of a system.
In an embodiment, the client device is configured to perform the following steps: detecting events in a (preferably measured) electrical signal comprising power consumption information of said system, said events representing state transitions of the appliances in the system;
- transferring information representing said detected events to said server device; and the server device is configured to perform the following steps: obtaining said information representing said detected events from said client device;
- characterizing said events by an event signature, by taking into account differences between steady states before and after each event and/or by taking into account transient states located between steady states before and after each event;
clustering events into a set of clusters on the basis of their event signatures; - identifying components on a basis of said set of clusters;
optionally identifying appliances from the components identified in the previous step;
optionally providing appliances' energy consumption or providing the components' energy consumption. In an embodiment, the client device is configured to perform the following steps: detecting events in a (preferably measured) electrical signal comprising power consumption information of said system, said events representing state transitions of the appliances in the system;
characterizing said events by an event signature, by taking into account differences between steady states before and after the event and/or taking into account transient states located between steady states before and after the event;
transferring information representing said event signatures to said server device; and the server device is configured to perform the following steps: obtaining said information representing said event signatures from said client device;
clustering events into a set of clusters on the basis of their signatures;
identifying components on the basis of said set of clusters;
optionally identifying appliances from the components identified in the previous step;
optionally providing appliances' energy consumption or providing the components' energy consumption. In an embodiment, the client device is configured to perform the following steps: detecting events in a measured electrical signal comprising power consumption information of said system, said events representing state transitions of the appliances in the system;
characterizing said events by an event signature, taking into account differences between steady states before and after the event and/or taking into account transient states located between steady states before and after the event;
clustering events into a set of clusters on the basis of their signatures;
transferring information representing said set of clusters to said server device; and the server device is configured to perform the following steps: obtaining said information representing said set of clusters from said client device;
identifying components on the basis of said set of clusters;
optionally identifying appliances from the components identified in the previous step;
optionally providing appliances' energy consumption or providing the components' energy consumption.
In the NIALM system's embodiments given above, said clustering comprises an initial clustering of said events into an initial set of clusters on a basis of a first clustering criterion, and a subsequent reclustering of at least one of said initial clusters on a basis of a second clustering criterion different from the first clustering criterion.
The use of a NIALM system with a client device and a server device provides a number of advantages, including, but not limited to: the possibility to optimize and update the algorithms in an easy way, in particular for the algorithms of the server device, and in particular the clustering algorithm,
the possibility to collect data centrally in order to assess and improve the efficiency of the used algorithms,
the possibility to perform an analysis of the aggregate energy or power consumption of a number of systems.
Therefore, the present invention also concerns a database comprising information representing identified components in a system, preferably of a multitude of systems, said information obtained using a NIALM method according to a method disclosed in this document. Preferably the database also comprises information representing identified appliances and/or information representing energy consumption of the appliances and/or components in said system or said multitude of systems.
Description of figures Figure 1 shows three methods to evaluate similarities between time series: the Minkowski distance, the dynamic time warping (DTW) and the least common subsequence (LCSS).
Figure 2 shows how the amount of outliers mainly decreases with k, the amount of nearest neighbors in the k-dist criterion. Figure 3 shows that when increasing k, the decrease of outlier points is mainly located in the origin of the P-Q plane but small clusters are then considered as outliers.
Figure 4 illustrates substructures which are better emphasized if principal components are evaluated for specific data subsets. Figure 5 shows a modeled aggregate power based on paired events. The evaluation is independent from the classification results.
Figure 6 shows that, once pairs have been assigned to components, the considered power steps derive from the clusters.
Figure 7 shows sequences of on-off cycles which are generated for nine appliances such that their aggregate consumption is not cyclic over the three-hour period.
Figure 8 shows nine appliances which draw specific current waveforms.
Figure 9 shows that the microwave is not a two state-device. Moreover, it exhibits cycling operations depending on the average power asked by the user.
Figure 10 shows events and their features - events corresponding to: the kettle, the vacuum cleaner and the microwave oven.
Figure 11 shows events and their features, zooming in at the origin of the P-Q plane.
Figure 12 shows that some turn-on transients systematically lead to several detections.
Figure 13 shows power variations of the ventilator which are significant compared to the power it consumes. Outliers risk to be associated with clusters of such a variable and low power device.
Figure 14 shows that clusters which are obtained after a first clustering iteration do not correctly reveal the complete data structure.
Figure 15 shows that an iterative clustering procedure according to the present invention efficiently captures the data substructure. Figure 16 shows a classification according to the highest membership value based on the Mahalanobis distance.
Figure 17 shows that there is almost no confusion between clusters of transient patterns. However, no clusters are discovered for the kettle, the ventilator and the vacuum cleaner. Figure 18 shows that component ID are assigned to recurrent pairs of nonoutliers. As a result, equivalent component ID are assigned to on and off event clusters.
Figure 19 shows first, the modeled aggregate power is evaluated with all paired events; and second, electrical components are discovered from recurrent pairs of nonoutliers. Finally, state sequences are evaluated based on classified event pairs. Figure 20 shows paired events are assigned to components according to their cluster identifiers.
Figure 21 shows state sequences evaluated from the event pairs assigned to pure component. Results from the state sequence evaluation are given in the top plots and the measurement in the bottom ones. The drawings of the figures are neither drawn to scale nor proportioned. Generally, identical components are denoted by the same reference numerals in the figures.
Detailed description of the invention
NIALM techniques are either pattern recognition-based or optimization-based approaches, also referred to as event-based or non event-based. In the former case, samples correspond to state changes of appliances, referred to as 'events'. These methods are also referred to as event-based NIALM. Signatures are used to associate
events to appliances. State sequences of appliances then result from classification algorithms. In supervised approaches, changes of appliance states are matched one- by-one to known signatures. Where learning is unsupervised, the recurrence of signatures is exploited to recognize appliances. The present invention concerns an unsupervised, event-based NIALM method according to claim 1. The present invention also concerns methods which optimize the NIALM method, including a recursive clustering method for recursively clustering similar events according to claim 5, a method for clustering similar events taking into account principal components of the event's signature according to claim 9, a component detection method according to claim 13, as well as the use of any of these methods or their combination in a NIALM method or for condition monitoring of an appliance. The present invention also concerns a processing unit arranged for executing any of these methods and a device comprising instructions for carrying out these methods, as well as a NIALM system comprising a client device and a server device as specified in claims 19 and 20 and a database according to claim 21.
Unless otherwise defined, all terms used in disclosing the invention, including technical and scientific terms, have the meaning as commonly understood by one of ordinary skill in the art to which this invention belongs. By means of further guidance, term definitions are included to better appreciate the teaching of the present invention. As used herein, the following terms have the following meanings:
"A", "an", and "the" as used herein refers to both singular and plural referents unless the context clearly dictates otherwise. By way of example, "a compartment" refers to one or more than one compartment.
"About" as used herein referring to a measurable value such as a parameter, an amount, a temporal duration, and the like, is meant to encompass variations of +/- 20% or less, preferably +/-10% or less, more preferably +/-5% or less, even more preferably +/-1% or less, and still more preferably +/-0.1% or less of and from the specified value, in so far such variations are appropriate to perform in the disclosed invention. However, it is to be understood that the value to which the modifier "about" refers is itself also specifically disclosed.
"Comprise", "comprising", and "comprises" and "comprised of" as used herein are synonymous with "include", "including", "includes" or "contain", "containing", "contains" and are inclusive or open-ended terms that specifies the presence of what follows e.g. component and do not exclude or preclude the presence of additional,
non-recited components, features, element, members, steps, known in the art or disclosed therein.
The recitation of numerical ranges by endpoints includes all numbers and fractions subsumed within that range, as well as the recited endpoints. The expression "% by weight", "weight percent", "%wt" or "wt%", here and throughout the description unless otherwise defined, refers to the relative weight of the respective component based on the overall weight of the formulation.
As used herein, the DBSCAN algorithm refers to density-based spatial clustering of applications with noise. In DBSCAN, a distance parameter, i.e. an inter-event distance parameter, defines the neighborhood of a point, i.e. an event. A zone is considered as dense if a sufficiently large number of points, i.e. events, is found within the given neighborhood. DBSCAN hereby requires only two parameters and no range of values must be specified for the number of clusters to investigate. Instead, the number of clusters is an output of the algorithm. We refer to M. Ester, H.-P. Kriegel, J. Sander and X. Xu, "A Density-Based Algorithm for Discovering Clusters in Large Spatial Databases with Noise", Knowledge Discovery and Data Mining 96: 226-231, 1996 for more detailed information about this algorithm.
As used herein, the OPTICS algorithm refers to an algorithm for ordering points to identify the clustering structure. We refer to M. Ankerst, M. M. Breunig, h-P Kriegel and J. Sander, "OPTICS: ordering points to identify the clustering structure", ACM SIGMOD Record 28 :49-60, 1999, for more information.
As used herein, the DBCLASD algorithm refers to a distribution-based clustering algorithm for mining in large spatial databases. We refer to X. Xu, M. Ester, H. P. Kriegel and J. Sander, "A distribution-based clustering algorithm for mining in large spatial databases", International Conference on Data Engineering, Vol. 1, pp. 324-331, 1998, for further information.
When clustering events, a decision criterion needs to be applied on whether or not an event belongs to a cluster, or two events belong to the same cluster. Such a criterion preferably comprises computing an inter-event distance to compare events. Therefore, in an embodiment, events are compared by computing an inter-event distance. This distance could preferably, and in particular for comparing events characterized by steady-state features, be a Euclidean distance or a Mahalanobis distance, or any combination thereof. This distance could preferably, and in particular for comparing events characterized by transient-state features, be a Minkowski
distance, a Dynamic Time Warping (DTW) distance, a Longest Common Subsequence (LCSS) distance, or any combination thereof. Also preferred are distances which are at least partly in a one-to-one correspondence with the above-mentioned distances or any combination thereof.
As used herein, the term "Euclidean distance", "Mahalanobis distance", "Minkowski distance", "Dynamic Time Warping (DTW) distance" and "Longest Common Subsequence (LCSS) distance" refer to specific types of distances between events, which are discussed in what follows.
Euclidean and Minkowski distance
wherein p=2 for the Euclidean distance, and p= l, 3, 4 or more for a Minkowski distance. Herein, X and Y refer to event signatures characterized by a set of N features X, and y, resp.
Preferably, these features are normalized, e.g. with respect to the minimum and maximum values of the features from the events which are taken into account to define or compute the clustering criterion, according to
wherein x, refers to the i'th feature in the signature, and xi;n refers to the normalized feature value. Mahalanobis distance:
Let Xc (nc x p) be a set of nc p-feature signatures grouped within a cluster. The variance-covariance matrix of cluster Xc is
Let further x be a signature whose distance w.r.t. Xc is to be evaluated and Xc the vector of mean feature values of Xc. The Mahalanobis distance is defined as follows:
This distance is evaluated in the original feature space, i.e. in the features described in Chapter 6. The membership value is simply expressed as the inverse of the point-to- cluster distance. The Mahalanobis distance better deals with the dispersion of clusters along different directions than e.g. the Euclidean. Therefore, the Mahalanobis distance is better suited to identify the cluster to which outlier events belong.
Outliers are isolated points in the feature space. They do not systematically exhibit extreme feature values. Instead, the distance to their nearest neighbors is high compared to the one of non outlier points. Consequently, the distance of a point to its k nearest neighbors is an image of its likelihood to be an outlier. We refer to such a distance as k-dist. Outliers are then points with the highest k-dist values.
As an example for deciding which events are outlier events, a threshold on k-dist values could be defined to decide whether or not points are outliers. In an embodiment, a maximum nonoutlier (MNO) value of the k-dist distribution could be used : it is defined as the highest observed value below the upper quartile increased by 1.5 times the interquartile range. Following this definition, the threshold value on the k-dist is obtained with the following formula :
€ out Hers max [k ~™~ dist(p) < Q s + l.T5(Q¾— ¾s)
p where Q75 and Q25 are the upper and lower quartiles of the k-dist distribution. Herein, k can preferably take any value between 1 and 20, preferably between 4 and 15, more preferably 10.
Dynamic time warping (DTW) distance and least common subsequence (LCSS) distance: The principle of the dynamic time warping (DTW) and the least common subsequence (LCSS) are illustrated in figure 1, which shows three methods to evaluate distances, in particular distances between time series: the Minkowski distance, the dynamic time warping (DTW) and the least common subsequence (LCSS).
DTW allows to deal with sequences with different lengths and different localization in the window. LCSS further allows to skip samples considered as outliers in the pattern.
When sequences that we compare are isolated patterns, the globally optimal alignment obtained with DTW is a suitable way to measure the similarity. Considering that different transients have been separated by the segmentation and that turn-on transients follow a continuous process on the time axis (no step in time), the DTW is chosen over the LCSS method. Further, the DTW has the advantage that the entire sequences must be matched which avoids that spurious transient states be matched to subparts of true transients. In other words, DTW performs matching with time shift analysis robust to positive false detections. The average distortion is used instead of the total distortion in order to normalize for different segment lengths. The objective of DTW is to compare two sequences X = (xi, .., xN) and Y = (yi, .., yM) of length N and M. A warping path is a sequence p = (pi, .., pL) with pi = (πι,ιτΐι) satisfying the three conditions:
Boundary condition : Pi = (1, 1) and pL = (N,M)
Monotonicity condition : n! < n2 < ... < nL and mi < m2 < ... < mL
- Step size condition : p,+1 - p, e {(1, 0), (0, 0), (0, 1)} for I 6 [ 1,L - 1]
Provided a distance metric d(x,, y,) between two elements x, and y, (e.g. Euclidean distance), the total cost of a warping path p is i
eF(X, 1' ) = \^ t f\ .<·,„ . y. if )
1=1 and the optimal warping path p* is the one minimizing the total cost among all possible paths. The distance DTW(X, Y) is defined as the total cost associated to the optimal path p* :
DTW(X, Y) = cp*(X, Y)
= min {cp(X, Y) | p is a warping path}
Finally, the average DTW distance is DTW(X,Y)/L where L is the length of the optimal path p*.
The present invention will be now described in more details, referring to examples that are not limitative.
Examples
Example 1 : Feature extraction on the basis of principa l components
Electric signatures have been evaluated because they characterize the electric consumption of components, not because they allow to classify components. Features might be correlated a nd less features could then be used to represent the data structure. Also, some features might be better than others to reveal similarities and differences between signatures. Considering non relevant features would add similarities between events.
In this example, we present the principal component analysis (PCA) as a tool to extract relevant features for clustering purpose. Because PCA is based on data scattering analysis, outliers might severely impact the results. They should consequently be removed from the data before that PCA be performed. Their detection is the subject of Section 2. Then, in Section 3, we expose how the features are extracted from the steady state signatures in our NIALM procedure. 1. Principal component analysis
We propose a concise description of the principal component analysis based on the work by I. T. Jolliffe, "Principal Component Ana lysis", Vol. 30, Springer Series in Statistics, Springer, 2nd ed ., 2002.
"The central idea of principal component analysis (PCA) is to reduce the dimensionality of a data set consisting of a large number of interrelated variables, while retaining as much as possible of the variation present in the data set. This is achieved by transforming to a new set of varia bles, the principa l components (PCs), which are uncorrelated, and which are ordered so that the first few retain most of the variation present in all of the original variables". Suppose that x is a vector of p random varia bles, and that the variances of the p random varia bles. The first step is to look for a linear function α Ί x of the elements of x having maximum variance, where Oi is a vector of p constants an, a12, lp, and ' denotes transpose, so that
Next, look for a linear function a'2x, uncorrected with αΊχ and having maximum variance. The procedure is repeated so that at the k'th stage a linear function a x is found that has maximum variance subject to being uncorrelated with α Ίχ, α '2χ, ..., a'k-ix. The k'th derived varia ble, a'kx is the k'th PC. Detailed equations for the principal component analysis can be found in the work by Jolliffe.
The number of features will be smaller if the initial ones are not independent. Jolliffe suggested that a subset of the principal components be chosen such that their cumulative percentage of tota l variation be between 70% and 90%. Let o2 j be the variance of the data along principa l component PC,, the cumulative percentage of total variation corresponding to the PC k is :
We set that value to 85% in our implementation . The principa l components (PC) taken into account are then the set of first PCs covering 85% of the variance. In addition to adding relevant variations, considering more PCs will, on the one hand, add distinctions between points of a same cluster and, on the other hand, add similarities between points of different clusters.
2. Detection of outlier points
Outliers might impact the result of the principal component analysis. They are mainly related to errors from the underlying data generation, e.g . they can be related to errors insignal segmentation leading to outliers. Herein, the errors can be due to noise or non-linearities, or due to transients which can be very long and hence could be seen as two or more different events. They are also rare events resulting from simultaneous on/off switching of components. They should be removed from the data set before searching for the principa l components. Outliers a re isolated points in the feature space. They do not systematica lly exhibit extreme feature values. Instead, the distance to their nearest neighbors is high com pared to the one of non outlier points. Consequently, the distance of a point to its k nearest neighbors is an image of its likelihood to be an outlier. We refer to such a distance as k-dist. Outliers are then points with the highest k-dist values. Signatures consist of features, with very different magnitudes. In order to give similar impact to each feature in the k-dist evaluation, data are normalized . The min-max normalization can be appropriate to detect the outliers and the reduced data are then :
ΪΙΧ1ΙΧ 5¾
-Vt fc [l. jj
Each feature xi;n ranges between 0 and 1 and the k-dist of events is the Euclidean distance between their normalized feature vectors and the one of their kth nearest neighbor. A threshold on k-dist values must be defined to decide whether or not points are outliers. X. M. Lopez et al., "Clustering methods applied in the detection of Ki67 hot- spots in whole tumor slide images: an efficient way to characterize heterogeneous tissue-baked biomarkers", Cytometry, Part A: the journal of the International Society for Analytical Cytology, Vol. 81(9) : 765-775, September 2012, reported on the use of the maximum nonoutlier (MNO) value of the k-dist distribution : it is defined as the highest observed value below the upper quartile increased by 1.5 times the interquartile range. Following this definition, the threshold value on the k-dist is obtained with the following formula : toutliers = max [fc - Mst (p) < Q75 + 1.75(^75 ~ Q s)]
P where Q75 and Q25 are the upper and lower quartiles of the k-dist distribution. Another approach to detect outliers is proposed by Jolliffe. It consists to look at the extreme values of the last PCs.
A value must be chosen for k, the number of neighbors taken into account in the kdist. Figure 2 illustrates the relative evolution of the number of outliers according to the value of k for five different data subsets. We see that increasing the number of neighbors considered in the k-dist decreases the number of outliers. Variations are mainly located in the origin of the P - Q plane which corresponds to false positive detections and small loads whose signature evaluation is impacted by fluctuations of the mixture. However, increasing k also leads to considering points of small clusters as outliers. These two aspects are illustrated in figure 3.
Detecting loads consuming low power compared to the aggregate consumption is the biggest challenge of NIALM. Consequently, it is desired that the false positive detections be correctly identified as outliers. This will facilitate the detection of points which cluster around the origin of the P - Q plane and are related to signatures of small loads. Taking this reasoning into account and the choice of DBSCAN parameters presented below, k is set to 10.
3. Non outlier PCA
Outliers add non relevant scattering of the data. Obviously, that impacts the evaluation of the principal components which are based on variance-covariance evaluation. The principal component analysis should consequently be evaluated solely with the non outlier points.
Furthermore, because PCA selects features in order to maximize the variance of the data, variables with higher amplitudes will be given higher weight and will be systematically considered in the first PCs. For instance, the active power ranges from some dozens of watts up to some kilowatts whereas normalized harmonic amplitudes range between 0 and 1. Data should then be reduced before the PCA be applied in order to give the same weight to the different features.
Our non outlier PCA works as follows:
Ignore outliers and,
a. Normalize the data
Outliers have been identified. It is then possible to reduce their impact on data normalization. The min-max normalization defined above is applied to the data, but the outliers are ignored from the search for the minimum and maximum values. This avoids that data be compressed because of extreme feature values of outliers.
where the exponent NO stands for non outliers. X n = (x i,n , x NO p,n b. Evaluate the principal component transformation a = PCA(X O n) where PCA() refers to the calculations of the linear combinations O i to ap explained in Section 1.
Then, for the entire data set
With Xn = (x1; nr
d. project the data into the new features with the transformation matrix obtained in b.
XPC = ft %n
PCA has been applied to the normalized dataset whose normalized active and reactive power are shown in the upper plot of figure 4. The two first PC are plotted in the leftmost below plot. The PCA is also applied to the same data set limited to events with active power first higher and then lower than 500W (respectively in the middle and right-most plots). For each of these three cases, the weights assigned to the features in the first PC are given in Table 1 and the percentage of total variance of the three first PCs in Table 2.
Table 1 : Weights of the features in the PCI for the data in figure 4
Table 2: Cumulative percentage of total variance Different comment can be formulated based on this illustrative example.
In the three cases, three features are given quasi null weight | Q/P| , H0n, H2n. This means either that they are heavily correlated with other features ( | Q/P| ) or that there is no variations along these features (H0n and H2n).
When the entire data set is considered, the first PC is the combination of most features. But for both data subsets, only some features have relatively high weights: the PCI of the > 500W' subset involves the fundamental components
whereas the one of the < 500W subset is mainly limited the harmonic content and the reactive power. This could have been expected.
Whereas for the > 500W data subset the PCI is sufficient to explain the variance of the data, PC2 is needed in the case of the entire data set to satisfy the 85% criterion. For the < 500W subset, the three first PCs must be used.
It follows that adapting the extracted features to the considered data subset will improve the clustering.
Example 2 : Component discovery
The method to discover instances of electrical components from clusters and the event time sequence can be illustrated by the following example.
First, on and off events are paired such that the fit-error is minimized. Second, cluster identifiers of paired events yield links between clusters which reflect electrical components. Third, assign event pairs to components to derive state sequences.
1. Pairing events by minimizing energy errors and error variations Considering two-state devices, if a component is turned off, it has previously been turned on. Following this, off-events should be paired to on-events to represent occurrences of unknown components. Constant power loads being assumed, the aggregate power is straightforwardly obtained for any sample k as the power of all pairs with kon < k < koff . If yk is the power measured at sample k and APk is a vector whose nonzero entries correspond to the signed power step values of paired events, the modeled aggregate power at sample k, yk, can be written as
y0 stands to account for loads that would have been switched on before the first sample of the observation window.
The ideal set of event pairs is the one for which the modeled aggregate power is the closest to the measured one. At each time, the power of the unknown components considered as active should sum up to the measured aggregate power. However, segmentation errors or simultaneous switch events lead to impossibility to truly fit the measured power simply by pairing events. Their impact on subsequent decisions
should be minimized. This can be achieved by penalizing the bestfit solution with the total variation of the fit-error. We formalize this in the next paragraphs.
In the ideal case, all on/off switching have been detected and all events are correctly paired. In that case, differences between the modeled power (y) and the measured one (y) derive from non constant power consumptions. Among different pairing possibilities, the optimal one is the one which minimizes the difference between y and y. But to deal with missed events, a gap should be allowed and its variations should then be minimized. The fit-error and the variations of this error are respectively defined as:
where yk is the median value of the power samples within steady state k. Note that also energy fit-errors and variations of energy fit-errors can be defined analogously for different time periods. This corresponds to modeling the observed aggregate power as constant segments. The use of median values instead of mean values allows the method to be more robust to segmentation errors: if there is a non detected power step within a steady state (i.e. there are actually two steady states), the longer one determines the constant power segment.
Pairing on-events and off-events can be done by extremizing a cost function which depends on variations in energy, in power, in current, in energy fit errors, in power fit errors, in current fit errors or in any combination thereof, preferably in a linear combination of at least two of said variations, more preferably variations in power and in power fit errors.
where k varies on a considered window as explained in the next subsection. The optimal set of event pairs is the one minimizing this cost function. Summation limits are not explicited since no optimization window needs to be defined yet at this stage in the procedure.
The use of the energy error instead of the power error (first term of cost function equation) could be investigated. But variations of power fit errors (second term) should be kept since variations in modeled energy are not relevant. If the first term were to be replaced by energy errors, weights between both terms should be carefully studied.
1.1. Searching for locally optimal pairs
An approach to find the optimal solution consists to test all combinations of event pairs. Of course, this is neither possible nor judicious. We propose a windowed approach to find locally optimal solutions. The use of optimization programming techniques is discussed in the next section.
A local window is considered, the length of which is a parameter (we set the initial value to W = 10 minutes). If events candidate to be paired to an on-event are investigated, the window starts right before the studied on-event. Similarly, it ends right after studied off-events. In this local window, each candidate k' is tested as follows in three steps: a. the amplitude of the turn-off power step is assigned to APk and APk with appropriate signs,
b. the modeled aggregate power is evaluated with the equation given above where y0 is the power of the first sample in the window,
c. the cost function l~ k< is evaluated with summations limited to the local window.
Since off-event signatures are not impacted by turn-on transients, they are more reliable. This is why they are considered instead of on-event signatures in Step a. Step b allows to hide deviations from ideal fit due to previous erroneous pairing since is systematically initialized to the first sample in the local window. In step c, the minimum cost l~ k< obtained with the different candidates in the window is compared to the costs of two other cases:
• ηο pair with event k' (ΓΝΡ) : APk is kept equal to zero, considering that event k is a fictitious transition
• 'candidate for event k is out of window' (r0w) : APk is set to the observed power step at event k without assigning any complementary power step.
The next step is conditioned by the case that minimized the cost function :
• rk< is minimum
Events k and k' are paired. If the candidate k' was previously paired to another event, the previous pair is canceled. The entries of ΔΡ are updated accordingly.
• rNP is minimum
The event is not paired. It probably results from a false positive detection or its partner has not been detected.
• r0w is minimum
If the 'candidate out of window' solution minimizes the cost function, the local window size is enlarged and new candidates are investigated in the new time interval.
This process is repeated for all events in a global window. An example of modeled aggregate power is shown in figure 5.
2. From clusters to electrical components
Events have been paired and each event has been assigned to a cluster. We extract pairs which are constituted by non-outlier points, since these points are likely to be true events.
If events associated to different clusters are often paired, these two clusters are considered to contain events from a single component. In such a case, an instance of electrical component is recorded and its electrical properties derive from the electric signatures of the clusters. Pairs of nonoutlier events with identical cluster ID lead to a direct association of a cluster to a component.
To summarize, a component can be defined as a recurrent link between clusters required to appropriately model the aggregate power.
3. Evaluation of component state sequences Instances of electrical components have been learned from pairs of nonoutlier points. Since all events have been assigned to a cluster owing to the point-to-cluster Mahalanobis distances, a first approximation for state sequences of the discovered
components consists of pairs of events assigned to the corresponding clusters. An example is given in figure 6.
However, the discovered components could also be used as input for NIALM methods with prior information. Example 3 : Results of laboratory testing
Algorithms were developed considering both laboratory and field data. However, to evaluate performances at the different stages of the NIALM procedure, a reliable ground truth is required. The latter is only available with laboratory data. Algorithms have been developed with the aim of providing parameter settings which are not tuned to specific data. Consequently, a new data set was generated in the laboratory (ELDA) specifically for this performance evaluation.
In Section 1 below, the data set is presented. The results are given in Section 2 and discussed in Section 3.
1. Data set Measurements have been performed in the laboratory developed for this research (ELDA). In addition to measuring the voltage and the aggregate current, each appliance is individually monitored. The appliance level consumption is consequently known.
To obtain the ground truth about times of appliance state changes, the power consumed in each outlet is thresholded. However, this works only for two-state appliances since, with multi component appliances, once a device is in operation additional state changes cannot be detected by simply thresholding the power draw. For example, if the heating resistor of a washing machine is in operation, state changes of the drum cannot be detected by thresholding the power since the latter varies between e.g. 2000W and 2200W. Consequently, mainly two-state devices were considered.
Nine devices were cyclically turned on and off during a 3-hour period : a kitchen mixer, a hair dryer in low power mode, a kettle, a 50W incandescent light bulb, a ventilator, a vacuum cleaner, a 200W halogen light bulb, a 10W economic light and a microwave oven. All these devices but the microwave are two-state appliances. The power they consume during one on-off cycle is shown in figure 7 and their current waveforms are shown in figure 8.
Cycles have been generated such that repetitive simultaneous state changes are avoided. However, simultaneous on/off switching might still occur. Cyclic behavior of appliances has no impact on the results since times and durations of use are not taken into account. Furthermore, even if appliances are cyclically operated, the aggregate consumption is not cyclic and various load combinations are obtained.
Some occurrences of the microwave oven are shown in figure 9. This appliance has more than two states: it consists of a light bulb, a motor and a magnetron. Furthermore, it can be noticed that, according to the average power asked by the user, the magnetron is turned on and off during its operations. Consequently events were manually added to the ground truth to account for these state changes.
2. Performance evaluation
In this section, results obtained at the different stages of the consumption analysis procedure are assessed and illustrated. Results will also be analyzed with the aim to identify suggestions for future work. 2.1. Event detection
The power trace was median filtered with LM = 50 and the cusum detector parameters were set to vm = 20 and Nd = 50. The rate of correct detections and the amount of missed detections can be evaluated for each individual load. The figures are reported in Table 3. We observe a global rate of 91%.
Detection rate obtained with vm = 20 and Nd = 50. Most missed detections comes from low power loads.
2.2. Feature evaluation
The detected events are plotted in the ΔΡ-AQ plan in figure 10 and figure 11. Each couple color-symbol represents one appliance, as indicated by the ground truth. In the following paragraphs, we comment on what can be observed from these plots. Turn-on events with slow time decay lead to several detections. In the studied data set, this impacts the segmentation of turn-on transients corresponding to the mixer and the vacuum cleaner, as shown in figure 12. This leads to different steady state feature values for turn-on and turn-off events. This can be seen with the green point cluster in figure 10 for the vacuum cleaner and with the red point cluster in figure 11 for the mixer. The events corresponding to both these loads form substructures. Furthermore, fictitious second events limit the search for the knee in the power trace which prevent from properly capturing the transient dynamics.
Events corresponding to the microwave exhibit a more complex scattering. Based on the power cycles shown in figure 9, two observations can be made. First, events corresponding to its light lead to points in the cluster around the origin of the P-Q plane (see figure 10). Second, turn-on transients result systematically in two detected events, as illustrated in figure 12. The second ones form the cluster right under 1000W in figure 10. They are not labeled in the ground truth and are then considered as false positive detections. It can be observed in figure 11 that most fictitious detections lead to points located between clusters of low power loads. Some also lie in clusters formed by events of these loads (in the P-Q plan). Moreover, the relative impact of variations in the mixture on the feature evaluation is more important for these low power loads. In addition, the ventilator shows relatively important power variations with respect to the power it consumes. This is illustrated in figure 13. This will probably harden the identification task.
Note finally that, when transients states are defined such that they contain two switch events, the assigned ground truth label is the one of the firstly actuated appliance. This explains why red crosses indicating events of the economic light lie with events of the 200W light bulb in figure 11.
2.3. Clustering the signatures
Clusters obtained with a first run of DBSCAN are shown in figure 14 and the ones resulting from the iterative procedure are shown in figure 15. Note that these classification results are only based on steady-state features since no mean to obtain a single data partitioning from SS and TS clusters has been implemented. Clusters of
transient patterns are however shown at the end of this section. Three observations should be done:
• Low power loads are well identified thanks to the adaptive feature extraction and definition of the o-neighborhood in the iterative clustering procedure. This is observed by comparing figure 14 and figure 15.
• Second events of microwave turn-on transients are also divided in two clusters.
This is not desired but could be expected since these events clearly form two clusters in the P-Q plot.
• However, it can be observed that the cluster of on-events of the vacuum cleaner is not divided although distinct clusters appear in the P-Q plot. This means that they are close enough thanks to other features.
The link between clusters and monitored appliances is given in Table 4. All figures but the ones in the first row are numbers of events. The table is structured as follows:
• Each column corresponds to one cluster: the first row gives the cluster identifier (ciD) corresponding to figure 15 and events are distributed in rows according to the ground truth. Outliers are reported in the last column along with nondetected events.
• Rows correspond to appliances as defined by the ground truth. The last row corresponds to events for which no identifier has been assigned in the ground truth. They are mostly false positive detections, but second events of microwave turn-on transients also lie in this category.
Table 4: There is almost no confusion between clusters resulting from the iterative clustering procedure.
Differences in the total number of events in Table 4 (and following) compared to Table 3 are due to simultaneous switch events. Since, in these cases, one event appears in
the result instead of two, it is only counted once and is considered to belong to the appliance firstly turned on.
We analyze further the results given in Table 4.
• Single clusters are identified for the hair dryer (ciD = 3), the kettle (ciD = 1), the 50W light bulb (CiD = 14), the 200W light bulb (dD = 10) and the economic light (do = 11). The low confusion between events of the economic light and the ventilator observed at this stage is an interesting result. This would probably not be achievable if only P and Q were considered.
• Two clusters are identified for events corresponding to the mixer: one for on- events (CiD = 9) and one for off-events (CiD = 12). The situation is similar for the vacuum cleaner (ciD = 4 and CiD = 5). It is expected that the time sequence analysis reveal that these clusters belong to a single component.
• The power consumed by the ventilator is low and shows variation of the order of 20%. This is illustrated in figure 13. Consequently, outliers are associated to its cluster (ciD = 2).
• The situation is more complex for the microwave. Its turn-off events form a single cluster (ciD = 8). On the contrary, its turn-on transients are systematically segmented in two events: the first ones form one cluster (ciD = 6) while the second ones are grouped in two clusters (ciD = 7 and CiD = 15). · Except for clusters 2 and 13, all events appearing in entries others than the main one result from simultaneous switch events. In these cases, only the appliance firstly turned on was reported in the ground truth.
These results show that clusters are appropriate to learn electrical component signatures since there is little confusion between components: columns contain manly one single entry.
Membership values are evaluated with the Mahalanobis distance in the initial feature space. The classification corresponding to the minimum distance is shown in figure 16 and the corresponding figures are given in Table 5. In this table, figures in the rightmost column correspond to non detected events. The most important observation is that outliers are mostly assigned to one cluster, the one constituted by events of the ventilator (dD = 2). This suggest that a method to limit the impact of outliers be investigated in future research. For instance, it could be useful that a 'bin' cluster be defined or that a threshold on membership values be used not to associate spurious events to clusters.
§ 12 3 1 1 2 4 -5 10 6 8 11 13 15 '
Mixer 46 35 - - 2 17 - - - 1 - - 1 3 8
Hair dryer - 92 - - 5 2 - 2 - - - 3
Kettle 96 - - 5 2
SOW light bulb - - - 83 21 - 1 2 - 4
Ventilator - - 1 - 88 - - 2 2 2 - 1 15
Vacuum cleaner - - - - - - 49 51 1 - 1
20OW light bulb - - 14 1 - 127 2 1 2 4 microwave oven - - - 13 10 - 1 77 70 - 1 8 eco light - - 2 - 5.1 - 1 3 2 1 118 7 - 64
/ - - 257 2 1 1 1 3 21. 44
Table 5 : Outliers are associated to clusters with the Mahalanobis distance evaluated in the initial feature space. Spurious events are mostly assigned to the cluster
corresponding to the ventilator. Note that microwave oven events assigned to clusters 2 and 14 correspond to the light bulb belonging to this appliance.
Regarding transient states, patterns corresponding to the identified clusters are given in figure 17. No cluster are found for the kettle, the vacuum cleaner and the ventilator. This could be expected for the kettle and the ventilator since their turn-on transients are only step-ups. This is however surprising for the vacuum cleaner; this means its transient patterns are too scattered to be clustered. There are two clusters for the microwave oven corresponding to the two parts of its turn-on transients. Transients of the mixer are divided into two distinct clusters.
Again, the discovered clusters are appropriate to learn component signatures since there is almost no confusion between components.
2.4. Discovering the components
Events are paired following the local window approach of the best-fit technique exposed earlier in this document. The modeled aggregate power resulting from paired events is shown in figure 19. Components are defined as recurrent associations of cluster IDs within nonoutiier event pairs. The resulting component identifiers (compiD) are shown in figure 18 and given in Table 6.
9 4 5 1 12 £ 6 10 7 8 '1 11 /
Mix. - _ - - - _ - 1 - _ _ 94
H, J f ΘΓ - 44 10 _ - _ _ _ _ _ 50
Kettle _ - - 72 - - - _ _ - - _ *i 1
SOW light - - _ - IS - _ _ _ 2 _ - 91
Ventilator - - - 1 - 14 - _ 2 1 9 - 82
Vac, clean. - - _ - _ - _ _ - _ -
200W light - - 10 - _ - - 54 _ - _ - 87 in. oven - - - - - - - - 18 34 _ - 125 eco light. - - _ 1 - - 1 - - 1 1 20 214
/ - _ _ _ _ 2 _ - 1:9 36 5 _
Ta ble 6 : All a ppliances have a pure or quasi pure group of event pairs that can be used to learn component signatures.
The following observations can be formulated : · All appliances have a pure or quasi pure group of event pairs that can be used to learn component signatures.
• There is one component (compID = 5) composed of events from the hair dryer and the 200Wlight bulb, whereas these two components are not confused regarding the signature classification (Ta ble 5) . This suggests that the membership values be used in some ways in the pairing procedure. On the contrary, the confusion between the ventilator and the economic light (compID = 3) was already present in the classification.
• Since second events of microwave turn-on transients were divided in two clusters, pairs remain divided in two component identifier. Future works could address the analysis of their time sequences to detect that they belong to one single electrical component. Actually, they should not exhibit simulta neous on states and should have similar duration of use distribution .
• First events of microwave turn-on transients are successfully rejected . In future works, additiona l analysis could detect that such a rejected cluster derive from the segmentation of single transients. For instance, this could be based on the following considerations : events of D = 6 (figure 15) are always followed by events of D = 7/15 and the sum of their linea r features lies in cluster CiD = 8 which exhibits subsequent off-events.
2.5. Energy disaggregation
Although the proposed tool focuses on discovering the set of components, it is interesting to look at how component state sequences are evaluated and how
aggregated energy can be segmented into the one of individual components. Since a membership value has been assigned to all events. The method simply consists to assign component identifier to event pairs according to their cluster identifiers, the resulting classification is given in figure 20 and Table 8. The state sequence of each pure component is evaluated and shown in figure 21. The state sequence of the ventilator is based on compID = 2 although it is not a pure component. Furthermore, both compiD = 7 and compiD = 8 are manually associated to the microwave oven. Regarding the economic light, it can be seen in figure 21 that, although paired events correspond to events of this light, pairs do not properly link successive on and off events. This also occurs in a lesser way for the ventilator.
The energy corresponding to the state sequences is contrasted with the one obtained from the submetering in Table 7. The low percentages of energy recovery suggest that more robust techniques be implemented to exploit the discovered component signatures and better evaluate the state sequences.
Table 7: Energy assigned to individual components (in Wh). The link between the components and the devices is manually defined. T (True) stands for energy consumed by the appliance and properly assigned to a component state sequence. F (False) stands for energy assigned to a state sequence whereas the corresponding appliance exhibits no consumption.
9 4 5 1 12 2 6 10 7 8 3 11 /
Mix... 36 1 2 1 - 73
II. Dryer - 60 20 - - 3 - - - - - - 21
Kettle - - - 87 _ - _ - _ - _ _ 16
SOW light - _ _ - 26 4 _ - _ 2 1 - 78
Ventilator - - _ 1 - 42 - 2 2 1 14 - 47
Vac. clean. - - - _ _ 82 - _ - _ 20
200V light - - 19 - - _ 1 77 - 1 - - 53 in. oven - - 1 - 2 3 _ - 19 42 _ - 113 eco light - - _ 2 _ 16 1 1 - 1 50 37 143
/ - - _ - - 36 _ - 20 43 22 1
Table 8: Paired events are assigned to components according to their cluster identifiers.
3. Discussion The results presented in this example showed that it is possible to discover the electrical components responsible for the aggregate consumption, and this, without prior information. Component state sequences have been evaluated with the present algorithm, mainly for illustrative purpose.
Claims
1. Non-intrusive appliance load monitoring (NIALM) method for monitoring components of appliances in a system and comprising the steps of: detecting events in an electrical signal comprising power consumption information of said system, said events representing state transitions of the appliances in the system;
characterizing each event by an event signature, by taking into account differences between steady states before and after each event and/or by taking into account transient states located between steady states before and after each event;
clustering events into a set of clusters on the basis of their event signatures; identifying components on the basis of said set of clusters, characterized in that said clustering step comprises an initial clustering of said events into an initial set of clusters on a basis of a first clustering criterion, and a subsequent reclustering of at least one of said initial clusters on the basis of a second clustering criterion different from the first clustering criterion.
2. NIALM method according to claim 1, comprising the step of:
identifying appliances from the identified components.
3. NIALM method according to any of the claims 1 or 2, comprising the step of:
- providing energy consumption of identified appliances and/or of said identified components.
4. NIALM method according to any of the previous claims, characterized in that said first clustering criterion is computed by taking into account event signatures from substantially all events detected in said electrical signal, and in that said second clustering criterion is computed by taking into account event signatures of substantially only events in said one initial cluster which is to be reclustered.
5. Recursive clustering method for clustering events obtained from an electrical signal comprising power consumption information of a system, each event being at least partially characterized by an event signature, said recursive clustering method comprising : an initial step of clustering events of an initial event set into a set of clusters using a clustering criterion for deciding whether two events belong to the same cluster and/or whether an event belongs to an existing cluster, said clustering
criterion being computed on a basis of event signatures from substantially all events from said initial event set; and
a recursive step of reclustering events belonging to a cluster into a set of sub- clusters using an updated clustering criterion for deciding whether two events belong to the same cluster and/or whether an event belongs to an existing cluster, said updated clustering criterion being computed on a basis of event signatures substantially only of events from said cluster, thereby obtaining one, two or more sub-clusters, preferably whereby if two or more sub-clusters are obtained, these two or more sub-clusters are reclustered using the present recursive step.
6. Recursive clustering method according to claim 5, characterized in that clustering and/or reclustering is performed using a density-based algorithm, such as a DBSCAN algorithm, an OPTICS algorithm and/or a DBCLASD algorithm, said density- based algorithm being based on a density defined by the amount of events which can be found in an e-neighborhood of a point of the cluster, said e-neighborhood of a point p being defined as Ne(p) = {q e X\d(p, q)≤ e], where X is the signature or feature space, d() represents a distance function between points and/or events, and e is an inter- event distance delimiting the neighborhood, whereby two kinds of points are defined : core points which have a density \Ne(p) \≥MinPts, MinPts being a number higher than 1, and preferably lower than 20, and border points which belong to an e-neighborhood of a core point without being a core point itself, whereby two points are defined as density-reachable if there exists a sequence of points such that each one belongs to the an e-neighborhood of its predecessor, the latter being a core point, whereby two points are defined as density-connected if they are density-reachable from a common point, whereby said clustering criterion comprises evaluating if an event is density- connected to a point and/or an event in a cluster.
7. Recursive clustering method according to claim 6, characterized in that said clustering criterion is computed and/or recomputed by computing and/or recomputing e and/or MinPts, preferably e.
8. Recursive clustering method according to any of claims 5 to 7, characterized in that said events are characterized by steady-state features and/or transient-state features.
9. Method for clustering events detected in an electrical signal comprising power consumption information of a system, each of said events being at least partially characterized by an event signature comprising a set of event features, the method comprising the steps of:
(a) defining principal components of said event features for said events;
(b) com puting a clustering criterion on the basis of the principa l components of substantia lly a ll of sa id events;
(c) clustering sa id events into a set of clusters; characterized in that it comprises the steps of:
(d) sub-clustering at least one cluster of said set of clusters by:
(i) redefining principa l components of the event features for events belonging to said one cluster, thereby substantially excluding event features for events not belonging to sa id one cluster;
(ii) recomputing the clustering criterion for said one cluster on the basis of the redefined principa l components of substa ntially all of the events belonging to sa id one cluster;
(iii) clustering sa id events belonging to sa id one cluster in one, two or more sub-clusters;
(e) optiona lly, if two or more sub-clusters are obtained after step (d)(iii), performing step (d) by taking at least one of said sub-clusters from step (d)(iii) as a cluster to be sub-clustered.
10. Method according to claim 9 characterized in that step (d) is performed recursively for each cluster of said set of clusters and/or for all sub-clusters obtained by performing step (d) until no further sub-clustering in two or more sub-clusters is obta ined.
11. Method according to claims 9 or 10, characterized in that said events are characterized by steady-state features.
12. Method according to claims 9 to 11, whereby step (c) and/or step (d)(iii) is performed using a density-based algorithm, such as a DBSCAN algorithm, an OPTICS algorithm and/or a DBCLASD algorithm, said density-based a lgorithm being based on a density defined by the amount of events which can be found in an e-neighborhood of a point of the cluster, sa id e-neighborhood of a point p being defined as Ne(p) = {q e x\d(p, q)≤ e], where X is the signature or feature space, d() represents a distance function between points and/or events, and e is an inter-event distance delimiting the neighborhood, whereby two kinds of points a re defined : core points which have a density \Ne(p) \≥ MinPts, MinPts being a number higher than 1, and preferably lower than 20, and border points which belong to an e-neighborhood of a core point without being a core point itself, whereby two points are defined as density-reachable if there exists a sequence of points such that each one belongs to the an e-neighborhood of its
predecessor, the latter being a core point, whereby two points are defined as density- connected if they are density-reachable from a common point, whereby said clustering criterion comprises evaluating if an event is density-connected to a point and/or an event in a cluster, preferably whereby said clustering criterion is computed and/or recomputed by computing and/or recomputing e and/or MinPts, preferably e.
13. Component detection method comprising the steps of: obtaining clusters of events obtained from an electrical signal comprising power consumption information of a system, said events comprising on-events and off-events, whereby each of said events comprises a cluster ID which allows to identify the cluster to which the event belongs, preferably whereby said clusters of events are obtained using a method according to any of the claims 3 to 10;
pairing on-events and off-events from said electrical signal, thereby obtaining a set of paired events, each paired event comprising paired cluster ID information representing the cluster ID of the on-event and the cluster ID of the off-event in said paired event;
detecting components by identifying recurrent paired cluster ID information within said set of paired events.
14. Component detection method according to claim 13, characterized in that said step of pairing on-events and off-events comprises extremizing a cost function which depends on variations in energy, in power, in current, in energy fit errors, in power fit errors, in current fit errors or in any combination thereof, preferably in a linear combination of at least two of said variations, more preferably variations in power and in power fit errors.
15. Use of a method according to any of claims 5 to 14 in a non-intrusive appliance load monitoring method.
16. Use of a method according to any of claims 1 to 14 for condition monitoring of an appliance.
17. Processing unit arranged for performing a method according to any of the claims 1 to 14, preferably said processing unit being arranged for performing a recursive clustering method according to any of claims 5 to 8 or a method according to any of claims 9 to 12, or a component detection method according to claim 13 or 14.
18. Device, preferably a computer-mountable or meter-mountable device, comprising instructions for executing a method according to any of claims 1 to 14,
preferably said device being arranged for performing a recursive clustering method according to any of claims 5 to 8 or a method according to any of claims 9 to 12, or a component detection method according to claim 13 or 14.
19. NIALM system comprising a client device and a server device, whereby said 5 client device and server device are linkable, and optionally are linked, and whereby said client device and server device are configured to together execute a NIALM method according to any of the claims 1 to 4, a recursive clustering method according to any of the claims 5 to 8, a method according to any of the claims 9 to 12, and/or a component detection method according to claim 13 or 14.
10 20. NIALM system according to claim 19, whereby said client device is configured to obtain a measured electrical signal comprising power consumption information of a system.
21. Database comprising information representing identified components in a system, preferably of a multitude of systems, said information being obtained by using 15 a NIALM method according to any of the claims 1 to 4, preferably whereby the database comprises information representing identified appliances and/or information representing energy consumption of the appliances and/or of the components in said system or said multitude of systems.
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| EP13190061 | 2013-10-24 | ||
| EP13190061.5 | 2013-10-24 |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2015059272A1 true WO2015059272A1 (en) | 2015-04-30 |
Family
ID=51799087
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/EP2014/072843 Ceased WO2015059272A1 (en) | 2013-10-24 | 2014-10-24 | Improved non-intrusive appliance load monitoring method and device |
Country Status (1)
| Country | Link |
|---|---|
| WO (1) | WO2015059272A1 (en) |
Cited By (33)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| GB2538087A (en) * | 2015-05-06 | 2016-11-09 | Torro Ventures Ltd | Analysing a power circuit |
| WO2016177857A1 (en) * | 2015-05-06 | 2016-11-10 | Torro Ventures Limited | Analysing a power circuit |
| GB2543321A (en) * | 2015-10-14 | 2017-04-19 | British Gas Trading Ltd | Method and system for determining energy consumption of a property |
| GB2553514A (en) * | 2016-08-31 | 2018-03-14 | Green Running Ltd | A utility consumption signal processing system and a method of processing a utility consumption signal |
| EP3301586A1 (en) * | 2016-09-30 | 2018-04-04 | Hitachi Power Solutions Co., Ltd. | Pre-processor and diagnosis device |
| CN108460410A (en) * | 2018-02-08 | 2018-08-28 | 合肥工业大学 | Electricity consumption mode identification method and system, the storage medium of citizen requirement side |
| CN108923426A (en) * | 2018-08-13 | 2018-11-30 | 广东电网有限责任公司 | A kind of load recognition methods, device, equipment and computer readable storage medium |
| CN109145224A (en) * | 2018-08-20 | 2019-01-04 | 电子科技大学 | Social networks event-order serie relationship analysis method |
| CN109544399A (en) * | 2018-11-29 | 2019-03-29 | 广东电网有限责任公司 | Transmission facility method for evaluating state and device based on multi-source heterogeneous data |
| CN109543846A (en) * | 2018-10-25 | 2019-03-29 | 安徽理工大学 | One kind being based on the improved DBSCAN water bursting in mine spectral discrimination method of MVO |
| CN110222788A (en) * | 2019-06-14 | 2019-09-10 | 哈尔滨工业大学 | A kind of the load identifying system and method for intelligent socket |
| CN110244150A (en) * | 2019-07-05 | 2019-09-17 | 四川长虹电器股份有限公司 | A kind of non-intrusion type electrical appliance recognition based on root mean square and standard deviation |
| CN110244149A (en) * | 2019-07-05 | 2019-09-17 | 四川长虹电器股份有限公司 | Non-intrusion type electrical appliance recognition based on current amplitude standard deviation |
| CN110569876A (en) * | 2019-08-07 | 2019-12-13 | 武汉中原电子信息有限公司 | Non-invasive load identification method and device and computing equipment |
| CN110610192A (en) * | 2019-08-07 | 2019-12-24 | 电子科技大学 | A Spectrum Spatial Channel Clustering Method |
| US20200057098A1 (en) * | 2016-11-15 | 2020-02-20 | Electrolux Appliances Aktiebolag | Monitoring arrangement for domestic or commercial electrical appliances |
| CN111612650A (en) * | 2020-05-27 | 2020-09-01 | 福州大学 | A method and system for power user grouping based on DTW distance and neighbor propagation clustering algorithm |
| CN111665390A (en) * | 2020-06-15 | 2020-09-15 | 威胜集团有限公司 | Non-invasive load detection method, terminal device and readable storage medium |
| CN111766462A (en) * | 2020-05-14 | 2020-10-13 | 中国计量大学 | A Non-Intrusive Load Identification Method Based on V-I Trajectory |
| CN112560977A (en) * | 2020-12-23 | 2021-03-26 | 华北电力大学(保定) | Non-invasive load decomposition method based on sparse classifier hierarchical algorithm |
| CN112732748A (en) * | 2021-01-07 | 2021-04-30 | 西安理工大学 | Non-invasive household appliance load identification method based on adaptive feature selection |
| CN113238092A (en) * | 2021-05-25 | 2021-08-10 | 南京工程学院 | Non-invasive detection method based on machine learning |
| CN113326296A (en) * | 2021-02-25 | 2021-08-31 | 中国电力科学研究院有限公司 | Load decomposition method and system suitable for industrial and commercial users |
| CN113884782A (en) * | 2021-08-24 | 2022-01-04 | 国网天津市电力公司营销服务中心 | Method and system for identifying starting characteristic of microwave oven, computer equipment and storage medium |
| CN116205544A (en) * | 2023-05-06 | 2023-06-02 | 山东卓文信息科技有限公司 | Non-invasive load identification system based on deep neural network and transfer learning |
| CN116400244A (en) * | 2023-04-04 | 2023-07-07 | 华能澜沧江水电股份有限公司 | Abnormal detection method and device for energy storage battery |
| CN116595425A (en) * | 2023-07-13 | 2023-08-15 | 浙江大有实业有限公司杭州科技发展分公司 | A Defect Identification Method Based on Multi-source Data Fusion of Power Grid Dispatch |
| US11726117B2 (en) | 2020-04-30 | 2023-08-15 | Florida Power & Light Company | High frequency data transceiver and surge protection retrofit for a smart meter |
| US11755448B2 (en) | 2018-08-03 | 2023-09-12 | Nec Corporation | Event monitoring apparatus, method and program recording medium |
| CN116842413A (en) * | 2023-05-17 | 2023-10-03 | 国网宁夏电力有限公司电力科学研究院 | Non-invasive load identification method, medium and system based on unsupervised machine learning |
| CN117093821A (en) * | 2023-10-19 | 2023-11-21 | 中国标准化研究院 | Energy efficiency and water efficiency measuring system and method for washing machine |
| CN117239942A (en) * | 2023-11-16 | 2023-12-15 | 天津瑞芯源智能科技有限责任公司 | Ammeter with monitoring function |
| CN118013447A (en) * | 2024-04-10 | 2024-05-10 | 山东德源电力科技股份有限公司 | A method for processing electric energy meter monitoring data based on pattern recognition |
Citations (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO2009103998A2 (en) | 2008-02-21 | 2009-08-27 | Sentec Limited | A method of inference of appliance usage. data processing apparatus and/or computer software |
| WO2012082802A2 (en) * | 2010-12-13 | 2012-06-21 | Fraunhofer Usa, Inc. | Methods and system for nonintrusive load monitoring |
| WO2012160062A1 (en) | 2011-05-23 | 2012-11-29 | Universite Libre De Bruxelles | Transition detection method for automatic-setup non-intrusive appliance load monitoring |
-
2014
- 2014-10-24 WO PCT/EP2014/072843 patent/WO2015059272A1/en not_active Ceased
Patent Citations (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO2009103998A2 (en) | 2008-02-21 | 2009-08-27 | Sentec Limited | A method of inference of appliance usage. data processing apparatus and/or computer software |
| WO2012082802A2 (en) * | 2010-12-13 | 2012-06-21 | Fraunhofer Usa, Inc. | Methods and system for nonintrusive load monitoring |
| WO2012160062A1 (en) | 2011-05-23 | 2012-11-29 | Universite Libre De Bruxelles | Transition detection method for automatic-setup non-intrusive appliance load monitoring |
Non-Patent Citations (10)
| Title |
|---|
| ANONYMOUS: "Catalogue des thèses électroniques de l'ULB - Thèse ULBetd-11082013-074844", 25 February 2015 (2015-02-25), XP055172029, Retrieved from the Internet <URL:http://theses.ulb.ac.be/ETD-db/collection/available/ULBetd-11082013-074844/> [retrieved on 20150225] * |
| ANONYMOUS: "DI-fusion Unsupervised learning procedure for nonintrusive appliance load monitoring", 8 November 2013 (2013-11-08), XP055172033, Retrieved from the Internet <URL:http://difusion.ulb.ac.be/vufind/Record/ULB-ETD:oai:ulb.ac.be:ETDULB:ULBetd-11082013-074844/Details> [retrieved on 20150225] * |
| ANONYMOUS: "machine learning - K-Means Algorithm - Stack Overflow", ENTRY IN STACK OVERFLOW FORUM, 15 June 2011 (2011-06-15), XP055172126, Retrieved from the Internet <URL:http://stackoverflow.com/questions/6353537/k-means-algorithm/6354052#6354052> [retrieved on 20150225] * |
| ESTER M ET AL: "A density-based algorithm for discovering clusters in large spatial databases with noise", PROCEEDINGS. INTERNATIONAL CONFERENCE ON KNOWLEDGE DISCOVERY ANDDATA MINING, XX, XX, 8 February 1996 (1996-02-08), pages 226 - 231, XP002355949 * |
| I. T. JOLLIFFE: "Springer Series in Statistics", vol. 30, 2002, SPRINGER, article "Principal Component Analysis" |
| JOHNSON; WILLSKY, BAYESIAN NONPARAMETRIC HIDDEN SEMI-MARKOV MODELS, 2012 |
| M. ANKERST; M. M. BREUNIG; H-P KRIEGEL; J. SANDER: "OPTICS: ordering points to identify the clustering structure", ACM SIGMOD RECORD, vol. 28, 1999, pages 49 - 60, XP008144377 |
| M. ESTER; H.-P. KRIEGEL; J. SANDER; X. XU: "A Density-Based Algorithm for Discovering Clusters in Large Spatial Databases with Noise", KNOWLEDGE DISCOVERY AND DATA MINING, vol. 96, 1996, pages 226 - 231, XP002355949 |
| X. M. LOPEZ ET AL.: "Clustering methods applied in the detection of Ki67 hot-spots in whole tumor slide images: an efficient way to characterize heterogeneous tissue-baked biomarkers", CYTOMETRY, PART A: THE JOURNAL OF THE INTERNATIONAL SOCIETY FOR ANALYTICAL CYTOLOGY, vol. 81, no. 9, September 2012 (2012-09-01), pages 765 - 775, XP055249483, DOI: doi:10.1002/cyto.a.22085 |
| X. XU; M. ESTER; H. P. KRIEGEL; J. SANDER: "A distribution-based clustering algorithm for mining in large spatial databases", INTERNATIONAL CONFERENCE ON DATA ENGINEERING, vol. 1, 1998, pages 324 - 331, XP010268334, DOI: doi:10.1109/ICDE.1998.655795 |
Cited By (59)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| GB2538087A (en) * | 2015-05-06 | 2016-11-09 | Torro Ventures Ltd | Analysing a power circuit |
| WO2016177857A1 (en) * | 2015-05-06 | 2016-11-10 | Torro Ventures Limited | Analysing a power circuit |
| RU2687345C9 (en) * | 2015-05-06 | 2019-08-19 | Общество С Ограниченной Ответственностью "Вольтавер Сервисиз" | Method and apparatus for power register control, electric power supervisory power control system, and method of installation of device for controlling power circuit |
| RU2687345C2 (en) * | 2015-05-06 | 2019-05-13 | Общество С Ограниченной Ответственностью "Вольтавер Сервисиз" | Method and apparatus for power register control, electric power supervisory power control system, and method of installation of device for controlling power circuit |
| CN107980097B (en) * | 2015-05-06 | 2020-05-01 | 托尔控股有限公司 | Power circuit analysis |
| CN107980097A (en) * | 2015-05-06 | 2018-05-01 | 托尔控股有限公司 | Power circuit is analyzed |
| GB2538087B (en) * | 2015-05-06 | 2019-03-06 | Torro Ventures Ltd | Analysing a power circuit |
| US11309710B2 (en) | 2015-10-14 | 2022-04-19 | British Gas Trading Limited | Method and system for assessing usage of devices of a property |
| GB2543321B (en) * | 2015-10-14 | 2022-02-02 | British Gas Trading Ltd | Method and system for determining energy consumption of a property |
| US11183842B2 (en) | 2015-10-14 | 2021-11-23 | British Gas Trading Limited | Method and system for determining energy consumption of a property |
| GB2543321A (en) * | 2015-10-14 | 2017-04-19 | British Gas Trading Ltd | Method and system for determining energy consumption of a property |
| GB2553514A (en) * | 2016-08-31 | 2018-03-14 | Green Running Ltd | A utility consumption signal processing system and a method of processing a utility consumption signal |
| GB2553514B (en) * | 2016-08-31 | 2022-01-26 | Green Running Ltd | A utility consumption signal processing system and a method of processing a utility consumption signal |
| EP3301586A1 (en) * | 2016-09-30 | 2018-04-04 | Hitachi Power Solutions Co., Ltd. | Pre-processor and diagnosis device |
| US20200057098A1 (en) * | 2016-11-15 | 2020-02-20 | Electrolux Appliances Aktiebolag | Monitoring arrangement for domestic or commercial electrical appliances |
| CN108460410A (en) * | 2018-02-08 | 2018-08-28 | 合肥工业大学 | Electricity consumption mode identification method and system, the storage medium of citizen requirement side |
| US11755448B2 (en) | 2018-08-03 | 2023-09-12 | Nec Corporation | Event monitoring apparatus, method and program recording medium |
| CN108923426B (en) * | 2018-08-13 | 2021-09-03 | 广东电网有限责任公司 | Load identification method, device, equipment and computer readable storage medium |
| CN108923426A (en) * | 2018-08-13 | 2018-11-30 | 广东电网有限责任公司 | A kind of load recognition methods, device, equipment and computer readable storage medium |
| CN109145224B (en) * | 2018-08-20 | 2021-11-23 | 电子科技大学 | Social network event time sequence relation analysis method |
| CN109145224A (en) * | 2018-08-20 | 2019-01-04 | 电子科技大学 | Social networks event-order serie relationship analysis method |
| CN109543846A (en) * | 2018-10-25 | 2019-03-29 | 安徽理工大学 | One kind being based on the improved DBSCAN water bursting in mine spectral discrimination method of MVO |
| CN109543846B (en) * | 2018-10-25 | 2023-02-21 | 安徽理工大学 | An improved DBSCAN mine water inrush spectrum identification method based on MVO |
| CN109544399A (en) * | 2018-11-29 | 2019-03-29 | 广东电网有限责任公司 | Transmission facility method for evaluating state and device based on multi-source heterogeneous data |
| CN109544399B (en) * | 2018-11-29 | 2021-03-16 | 广东电网有限责任公司 | Power transmission equipment state evaluation method and device based on multi-source heterogeneous data |
| CN110222788A (en) * | 2019-06-14 | 2019-09-10 | 哈尔滨工业大学 | A kind of the load identifying system and method for intelligent socket |
| CN110244149A (en) * | 2019-07-05 | 2019-09-17 | 四川长虹电器股份有限公司 | Non-intrusion type electrical appliance recognition based on current amplitude standard deviation |
| CN110244149B (en) * | 2019-07-05 | 2021-03-16 | 四川长虹电器股份有限公司 | Non-invasive electrical appliance identification method based on current amplitude standard deviation |
| CN110244150A (en) * | 2019-07-05 | 2019-09-17 | 四川长虹电器股份有限公司 | A kind of non-intrusion type electrical appliance recognition based on root mean square and standard deviation |
| CN110610192B (en) * | 2019-08-07 | 2022-03-15 | 电子科技大学 | Spectrum space channel clustering method |
| CN110610192A (en) * | 2019-08-07 | 2019-12-24 | 电子科技大学 | A Spectrum Spatial Channel Clustering Method |
| CN110569876A (en) * | 2019-08-07 | 2019-12-13 | 武汉中原电子信息有限公司 | Non-invasive load identification method and device and computing equipment |
| US11726117B2 (en) | 2020-04-30 | 2023-08-15 | Florida Power & Light Company | High frequency data transceiver and surge protection retrofit for a smart meter |
| CN111766462B (en) * | 2020-05-14 | 2023-03-10 | 中国计量大学 | Non-invasive load identification method based on V-I track |
| CN111766462A (en) * | 2020-05-14 | 2020-10-13 | 中国计量大学 | A Non-Intrusive Load Identification Method Based on V-I Trajectory |
| CN111612650B (en) * | 2020-05-27 | 2022-06-17 | 福州大学 | DTW distance-based power consumer grouping method and system |
| CN111612650A (en) * | 2020-05-27 | 2020-09-01 | 福州大学 | A method and system for power user grouping based on DTW distance and neighbor propagation clustering algorithm |
| CN111665390A (en) * | 2020-06-15 | 2020-09-15 | 威胜集团有限公司 | Non-invasive load detection method, terminal device and readable storage medium |
| CN112560977B (en) * | 2020-12-23 | 2022-09-27 | 华北电力大学(保定) | Non-invasive load decomposition method based on sparse classifier hierarchical algorithm |
| CN112560977A (en) * | 2020-12-23 | 2021-03-26 | 华北电力大学(保定) | Non-invasive load decomposition method based on sparse classifier hierarchical algorithm |
| CN112732748A (en) * | 2021-01-07 | 2021-04-30 | 西安理工大学 | Non-invasive household appliance load identification method based on adaptive feature selection |
| CN112732748B (en) * | 2021-01-07 | 2024-03-15 | 西安理工大学 | Non-invasive household appliance load identification method based on self-adaptive feature selection |
| CN113326296B (en) * | 2021-02-25 | 2024-06-07 | 中国电力科学研究院有限公司 | Load decomposition method and system suitable for industrial and commercial users |
| CN113326296A (en) * | 2021-02-25 | 2021-08-31 | 中国电力科学研究院有限公司 | Load decomposition method and system suitable for industrial and commercial users |
| CN113238092B (en) * | 2021-05-25 | 2022-07-01 | 南京工程学院 | A non-invasive detection method based on machine learning |
| CN113238092A (en) * | 2021-05-25 | 2021-08-10 | 南京工程学院 | Non-invasive detection method based on machine learning |
| CN113884782A (en) * | 2021-08-24 | 2022-01-04 | 国网天津市电力公司营销服务中心 | Method and system for identifying starting characteristic of microwave oven, computer equipment and storage medium |
| CN116400244A (en) * | 2023-04-04 | 2023-07-07 | 华能澜沧江水电股份有限公司 | Abnormal detection method and device for energy storage battery |
| CN116400244B (en) * | 2023-04-04 | 2023-11-21 | 华能澜沧江水电股份有限公司 | Abnormality detection method and device for energy storage battery |
| CN116205544B (en) * | 2023-05-06 | 2023-10-20 | 山东卓文信息科技有限公司 | Non-invasive load identification system based on deep neural network and transfer learning |
| CN116205544A (en) * | 2023-05-06 | 2023-06-02 | 山东卓文信息科技有限公司 | Non-invasive load identification system based on deep neural network and transfer learning |
| CN116842413A (en) * | 2023-05-17 | 2023-10-03 | 国网宁夏电力有限公司电力科学研究院 | Non-invasive load identification method, medium and system based on unsupervised machine learning |
| CN116595425B (en) * | 2023-07-13 | 2023-11-10 | 浙江大有实业有限公司杭州科技发展分公司 | Defect identification method based on power grid dispatching multi-source data fusion |
| CN116595425A (en) * | 2023-07-13 | 2023-08-15 | 浙江大有实业有限公司杭州科技发展分公司 | A Defect Identification Method Based on Multi-source Data Fusion of Power Grid Dispatch |
| CN117093821A (en) * | 2023-10-19 | 2023-11-21 | 中国标准化研究院 | Energy efficiency and water efficiency measuring system and method for washing machine |
| CN117093821B (en) * | 2023-10-19 | 2024-01-09 | 中国标准化研究院 | A washing machine energy efficiency and water efficiency measurement system and method |
| CN117239942A (en) * | 2023-11-16 | 2023-12-15 | 天津瑞芯源智能科技有限责任公司 | Ammeter with monitoring function |
| CN117239942B (en) * | 2023-11-16 | 2024-01-19 | 天津瑞芯源智能科技有限责任公司 | Ammeter with monitoring function |
| CN118013447A (en) * | 2024-04-10 | 2024-05-10 | 山东德源电力科技股份有限公司 | A method for processing electric energy meter monitoring data based on pattern recognition |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| WO2015059272A1 (en) | Improved non-intrusive appliance load monitoring method and device | |
| WO2016079229A1 (en) | Improved non-intrusive appliance load monitoring method and device | |
| WO2016141978A1 (en) | Improved non-intrusive appliance load monitoring method and device | |
| Lai et al. | Multi-appliance recognition system with hybrid SVM/GMM classifier in ubiquitous smart home | |
| CN109387712B (en) | Non-intrusive load detection and decomposition method based on state matrix decision tree | |
| Aiad et al. | Unsupervised approach for load disaggregation with devices interactions | |
| He et al. | Non-intrusive load disaggregation using graph signal processing | |
| De Baets et al. | VI-based appliance classification using aggregated power consumption data | |
| US20150377935A1 (en) | Signal identification methods and systems | |
| US8918346B2 (en) | System and method employing a minimum distance and a load feature database to identify electric load types of different electric loads | |
| Altrabalsi et al. | A low-complexity energy disaggregation method: Performance and robustness | |
| Ghosh et al. | Remote appliance load monitoring and identification in a modern residential system with smart meter data | |
| WO2013081718A2 (en) | System and method employing a self-organizing map load feature database to identify electric load types of different electric loads | |
| Parson et al. | Using hidden markov models for iterative non-intrusive appliance monitoring | |
| US11573587B2 (en) | Apparatus and method for non-invasively analyzing behaviors of multiple power devices in circuit and monitoring power consumed by individual devices | |
| CN105324639A (en) | Method and system employing graphical electric load categorization to identify one of a plurality of different electric load types | |
| Kahl et al. | Appliance classification across multiple high frequency energy datasets | |
| Mittelsdorf et al. | Submeter based training of multi-class support vector machines for appliance recognition in home electricity consumption data | |
| Barker et al. | Powerplay: creating virtual power meters through online load tracking | |
| Nguyen et al. | Appliance classification method based on k-nearest neighbors for home energy management system | |
| Ghazali et al. | Extraction and selection of statistical harmonics features for electrical appliancesidentification using k-NN classifier combined with voting rules method | |
| Bagnall et al. | Finding motif sets in time series | |
| Zeifman et al. | Nonintrusive monitoring of miscellaneous and electronic loads | |
| Liu et al. | A review of nonintrusive load monitoring and its application in commercial building | |
| Bhattacharjee et al. | Appliance classification using energy disaggregation in smart homes |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 14789820 Country of ref document: EP Kind code of ref document: A1 |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 14789820 Country of ref document: EP Kind code of ref document: A1 |











