<?xml version="1.0" encoding="UTF-8"?><!DOCTYPE article  PUBLIC "-//NLM//DTD Journal Publishing DTD v3.0 20080202//EN" "http://dtd.nlm.nih.gov/publishing/3.0/journalpublishing3.dtd"><article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" dtd-version="3.0" xml:lang="en" article-type="research article"><front><journal-meta><journal-id journal-id-type="publisher-id">EPE</journal-id><journal-title-group><journal-title>Energy and Power Engineering</journal-title></journal-title-group><issn pub-type="epub">1949-243X</issn><publisher><publisher-name>Scientific Research Publishing</publisher-name></publisher></journal-meta><article-meta><article-id pub-id-type="doi">10.4236/epe.2017.94B044</article-id><article-id pub-id-type="publisher-id">EPE-75300</article-id><article-categories><subj-group subj-group-type="heading"><subject>Articles</subject></subj-group><subj-group subj-group-type="Discipline-v2"><subject>Engineering</subject></subj-group></article-categories><title-group><article-title>
 
 
  Residential Electricity Consumption Behavior Mining Based on System Cluster and Grey Relational Degree
 
</article-title></title-group><contrib-group><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Mengjia</surname><given-names>Xu</given-names></name><xref ref-type="aff" rid="aff1"><sup>1</sup></xref></contrib><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Yuhong</surname><given-names>Wang</given-names></name><xref ref-type="aff" rid="aff1"><sup>1</sup></xref></contrib></contrib-group><aff id="aff1"><addr-line>School of Electrical Engineering and Information, Sichuan University, Chengdu, China</addr-line></aff><pub-date pub-type="epub"><day>06</day><month>04</month><year>2017</year></pub-date><volume>09</volume><issue>04</issue><fpage>390</fpage><lpage>400</lpage><history><date date-type="received"><day>February</day>	<month>26,</month>	<year>2017</year></date><date date-type="rev-recd"><day>Accepted:</day>	<month>March</month>	<year>30,</year>	</date><date date-type="accepted"><day>April</day>	<month>6,</month>	<year>2017</year></date></history><permissions><copyright-statement>&#169; Copyright  2014 by authors and Scientific Research Publishing Inc. </copyright-statement><copyright-year>2014</copyright-year><license><license-p>This work is licensed under the Creative Commons Attribution International License (CC BY). http://creativecommons.org/licenses/by/4.0/</license-p></license></permissions><abstract><p>
 
 
   
   In order to improve the utilization of the residential electricity consumption data which contains the information on the user’s electricity consumption habits, a residential electricity consumption behaviors mining algorithm model is constructed. Firstly, according to the attribute, the collected data can be divided into the global data and the phase data, then the appropriate global variables are selected to mine the user’s electricity consumption patterns in the near future on the system clustering algorithm. Based on the theory of grey relational analysis, combing phase data with the power modes to analyze the potential characteristics of residential electricity consumption behaviors deeply that verify the ability of latest power mode to predict household electricity consumption situation in the coming few days and the effect of dominant phase variables on the peak load shifting. Finally, from the actual data of a certain family, the proposed data mining algorithm is testified that it can effectively explore the electricity consumption behavior habits and characteristics of the family. 
  
 
</p></abstract><kwd-group><kwd>Data Mining</kwd><kwd> Electricity Consumption Behavior</kwd><kwd> System Cluster</kwd><kwd> Grey  Relational Degree</kwd></kwd-group></article-meta></front><body><sec id="s1"><title>1. Introduction</title><p>In recent years, facing the growing problem of energy crisis and environment protection, as well as the increasingly tense relationship with the electric power supply and demand, smart electricity consumption which can improve the quality of electric power supply service and the electricity efficiency by means of intelligent measurement, high speed communication and high efficiency control, has become the focus of global attention [<xref ref-type="bibr" rid="scirp.75300-ref1">1</xref>] [<xref ref-type="bibr" rid="scirp.75300-ref2">2</xref>]. With the construction of smart grid, advanced measurement devices are more and more installed in the user [<xref ref-type="bibr" rid="scirp.75300-ref3">3</xref>], as of 2013, China has installed a total of about 150 million smart meters [<xref ref-type="bibr" rid="scirp.75300-ref4">4</xref>], so that the users’ power data is gathered to explosive growth now. These data imply the users’ electricity consumption habits and real needs [<xref ref-type="bibr" rid="scirp.75300-ref5">5</xref>], which is conducive for the power sector to master the user’s electricity consumption habits, and to draw up a detailed smart power strategy for the users to improve the passive response mode based on the time-of-use price before [<xref ref-type="bibr" rid="scirp.75300-ref6">6</xref>] [<xref ref-type="bibr" rid="scirp.75300-ref7">7</xref>], is advantageous to the power supply department to formulate economic power generation plan and improve the energy allocation efficiency. However, the data gathered is large and scattered, how to find the key variables in the messy data and extract the user’s power consumption characteristics is a difficult point.</p><p>At present, the research on the analysis of the user’s power using data has become a hot spot in the country that a variety of efficient data mining algorithms combined with the emerging computing model such as cloud computing platform has been studied and applied. References [<xref ref-type="bibr" rid="scirp.75300-ref8">8</xref>] [<xref ref-type="bibr" rid="scirp.75300-ref9">9</xref>] mine the massive power consumption data to get the user’s common power mode by the improved k-means clustering algorithm. Reference [<xref ref-type="bibr" rid="scirp.75300-ref10">10</xref>] use cluster average diameter and minimum distance to evaluate the clustering effect of K-means, FCM and other clustering algorithms, so as to determine the optimal number of clusters and realize the analysis and statistics of power consumption data of large power customers ultimately. References [<xref ref-type="bibr" rid="scirp.75300-ref11">11</xref>] [<xref ref-type="bibr" rid="scirp.75300-ref12">12</xref>] based on cloud computing platform, utilize parallel Apriori algorithm and parallel k-means algorithm respectively which have greatly improved the efficiency of data mining to achieve the relevance mining of the users electricity consumption behaviors as well as the load classification. But most of the existing researches are mainly aimed at the analysis of user groups’ electricity consumption behaviors, the individual user’s is less. Furthermore the future power supply mode will be transformed from the one- way to the two-way power supply service mode dominated by the users, which is more intelligent and targeted, so it is necessary to excavate the power behavior characteristics of individual user.</p><p>In this paper, based on the data of household electricity consumption, the definition and division are achieved firstly. Then the system clustering algorithm which overcomes the problem of the selection of the clusters’ number and the initial center in the traditional k-means algorithm is applied for the analysis and feature mining of the user’s total power consumption in the near future to get the user’s frequent electricity consumption mode. By the means of the gray correlation analysis, the household main power consumption behavior is gotten from each power mode, which is used to verify the user’s electricity behavior predictability and peak household load shifting by the main power consumption behavior.</p></sec><sec id="s2"><title>2. Analysis Process of Household Electricity Consumption Behavior</title><p>For a single family, considering the influencing factors of user’s electricity consumption behavior such as real time, climate change and human activity cycle, the family’s recent 10 - 15 days of electricity data is selected as the base of analysis. The collected information can be divided into two types: 1) Global data: including energy consumption, power, voltage and current, etc. 2) Phase data: the electricity data of the intelligent electrical appliances or the traditional electrical appliances with intelligent sockets, the electricity data of the home appliances with the similar attribute, and so on. Using the one day’s data as a unit sample, the characteristics of the recent household electricity consumption behavior is excavated by the means of the system clustering and gray relational degree. <xref ref-type="fig" rid="fig1">Figure 1</xref> is a block diagram of household electricity behavior mining process, which is divided into three modules: variable screening and sorting, household power pattern mining, potential power characteristics mining.</p><p>1) Variable screening and sorting completes the division and definition of the gathered data. The variables that have great contribution to the mining purpose are chosen as the input parameters of the data mining algorithm, what reduces the calculation process and increases the validity of the data mining result.</p><p>2) Household power mode mining selects the appropriate global variables to cluster the household electricity consumption using the system clustering algorithm and digs the family’s common electricity consumption modes to extract its characteristic curve.</p><p>3) Potential power characteristics mining. On the basis of the above-men- tioned electricity consumption modes, combined with the phase parameters, household electricity consumption behavior characteristics are analyzed by the gray relational degree, which provides the basis for smart power strategy.</p></sec><sec id="s3"><title>3. Household Power Modes Mining</title><sec id="s3_1"><title>3.1. Basic Theory of System Clustering</title><p>As an unsupervised learning process, clustering has a good effect on mining the</p><fig id="fig1"  position="float"><label><xref ref-type="fig" rid="fig1">Figure 1</xref></label><caption><title> Mining process framework of residential electricity consumption behaviors</title></caption><graphic mimetype="image"   position="float"  xlink:type="simple"  xlink:href="http://html.scirp.org/file/75300x2.png"/></fig><p>inner relations of unmarked data sets [<xref ref-type="bibr" rid="scirp.75300-ref13">13</xref>]. Compared with the k-means algorithm, the system clustering don’t need to set the number of clusters in advance, what is very suitable for processing such complex and unknown household electricity data and is one of the classical algorithms. This algorithm is divided into the cohesion method and the splitting method. Take the cohesion method as an example, each object is regarded as a cluster firstly. Then according to a similarity measure, two closest clusters are merged at every turn until all the clusters are merged into one and the affinity spectrum diagram is formed (specific principles are shown in <xref ref-type="fig" rid="fig2">Figure 2</xref> ).The split rule is just the opposite.</p></sec><sec id="s3_2"><title>3.2. Household Electricity Modes Mining Steps</title><p>Supposing the number of samples of household electricity consumption is s (10 ≤ s ≤ 15), global variables and phase variables are attribute parameters, and the kth sample can be described as: f<sub>k</sub> = [f<sub>kA</sub>,f<sub>kB</sub>] = {A<sub>k</sub><sub>1</sub>, A<sub>k</sub><sub>2</sub>, ・・・, A<sub>km</sub>, B<sub>k</sub><sub>1</sub>, B<sub>k</sub><sub>2</sub>, ・・・, B<sub>kn</sub>} (k = 1, 2, ・・・, s), where A<sub>ki</sub> (i = 1, 2, ・・・, m) is a global variable, B<sub>kj</sub> (j = 1, 2,・・・ n) is a phase variable, and both of which are time series parameters. Assuming that there are qacquisition points in the sample period, A<sub>ki</sub> = {<inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x3.png" xlink:type="simple"/></inline-formula>, <inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x4.png" xlink:type="simple"/></inline-formula>,・・・, <inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x5.png" xlink:type="simple"/></inline-formula>}, B<sub>kj</sub> = {<inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x6.png" xlink:type="simple"/></inline-formula>, <inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x7.png" xlink:type="simple"/></inline-formula>,・・・, <inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x8.png" xlink:type="simple"/></inline-formula>}. Because of the existence of different dimensions between different variables, standardized processing is carried out by Z-score method:</p><disp-formula id="scirp.75300-formula304"><label>(1)</label><graphic position="anchor" xlink:href="http://html.scirp.org/file/75300x9.png"  xlink:type="simple"/></disp-formula><p>where: μ is the sample mean of x, and τ is the sample standard deviation. For the convenience of writing, standardized variable symbols remain unchanged.</p><p>Choosing the global variables as the analysis parameters, s samples are clustered together to calculate the average distance between clusters by formula (2) and (3):</p><disp-formula id="scirp.75300-formula305"><label>(2)</label><graphic position="anchor" xlink:href="http://html.scirp.org/file/75300x10.png"  xlink:type="simple"/></disp-formula><disp-formula id="scirp.75300-formula306"><label>(3)</label><graphic position="anchor" xlink:href="http://html.scirp.org/file/75300x11.png"  xlink:type="simple"/></disp-formula><p>where: d(f<sub>i</sub>, f<sub>j</sub>) is the Euclidean distance between samples f<sub>i</sub> and f<sub>j</sub>, and I ≠ j, (i, j = 1, 2, ・・・, s), n<sub>i</sub> and n<sub>j</sub> are the number of objects of cluster C<sub>i</sub> and C<sub>j</sub>, respectively.</p><p>Merging the two clusters whose distance is the smallest, then updating the data,</p><fig id="fig2"  position="float"><label><xref ref-type="fig" rid="fig2">Figure 2</xref></label><caption><title> Hierarchical clustering process</title></caption><graphic mimetype="image"   position="float"  xlink:type="simple"  xlink:href="http://html.scirp.org/file/75300x12.png"/></fig><p>the calculation above is repeated until all the clusters are merged into one. According to the complicated relationship, the situation of the combination of two samples, and the combination of a sample and a cluster is classified as a class. Taking into account that the mean is susceptible to the discrete points and extreme values, the household energy modes M<sub>1</sub>, M<sub>2</sub>, ・・・, M<sub>p</sub> are obtained by utilizing the median eigenvalues to represent the general trend of a class of data sets:</p><disp-formula id="scirp.75300-formula307"><label>(4)</label><graphic position="anchor" xlink:href="http://html.scirp.org/file/75300x13.png"  xlink:type="simple"/></disp-formula><p>where:<inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x14.png" xlink:type="simple"/></inline-formula>, <inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x15.png" xlink:type="simple"/></inline-formula>(I = 1, 2, ・・・, m; j = 1, 2, ・・・, n) are the medians of the corresponding variables in the pth class.</p></sec></sec><sec id="s4"><title>4. Potential Electricity Consumption Features Mining</title><sec id="s4_1"><title>4.1. Basic Theory of Gray Relational Degree</title><p>Gray relational analysis is used to quantitatively analyze and compare the dynamic development process of the system, and to excavate the main factors influencing its change, which is mainly composed of three elements as reference sequence, comparison sequence, and gray relational degree. Assuming that the reference sequence at the ith time is x<sub>0</sub>(i), X<sub>0</sub> = (x<sub>0</sub>(1), x<sub>0</sub>(2), ・・・, x<sub>0</sub>(n)). Comparison sequences are generally more than one, recorded as X<sub>1</sub>, X<sub>2</sub>, ・・・, X<sub>k</sub>, where X<sub>k</sub> = (X<sub>k</sub>(1), X<sub>k</sub>(2), ・・・, X<sub>k</sub>(n)). The essence of gray relational analysis is to compare the similarity between the curves of X<sub>1</sub>, X<sub>2</sub>, ・・・, X<sub>k</sub> and X<sub>0</sub> with time respectively. The gray relational degree represents the relative order of the similarity of each comparison reference sequence to the reference sequence, and the more similar, the similarity degree is greater.</p></sec><sec id="s4_2"><title>4.2. Potential Electricity Consumption Features Mining Steps</title><p>Combined with phase data, the gray relational grade is applied to analyze the latent influencing factors of user’s power consumption modes, and the mode Mp is taken as an example:</p><p>1) The absolute difference between the phase and global variables is calculated by the formula (5):</p><disp-formula id="scirp.75300-formula308"><label>(5)</label><graphic position="anchor" xlink:href="http://html.scirp.org/file/75300x16.png"  xlink:type="simple"/></disp-formula><p>where: k = 1,2, ・・・, m.</p><p>2) The gray relational coefficient of corresponding elements between the phase variables and the global variable <inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x17.png" xlink:type="simple"/></inline-formula> is gotten by:</p><disp-formula id="scirp.75300-formula309"><label>(6)</label><graphic position="anchor" xlink:href="http://html.scirp.org/file/75300x18.png"  xlink:type="simple"/></disp-formula><p>where: <inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x19.png" xlink:type="simple"/></inline-formula>indicates the correlation between the ith variable <inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x19.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x20.png" xlink:type="simple"/></inline-formula> and the global variable <inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x19.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x20.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x21.png" xlink:type="simple"/></inline-formula> of mode M<sub>p</sub> at the jth time point.</p><p>3) The relational degree of each phase variable to the global variable <inline-formula><inline-graphic xlink:href="http://html.scirp.org/file/75300x22.png" xlink:type="simple"/></inline-formula> is computed by:</p><disp-formula id="scirp.75300-formula310"><label>(7)</label><graphic position="anchor" xlink:href="http://html.scirp.org/file/75300x23.png"  xlink:type="simple"/></disp-formula><p>4) Comprehensive evaluation is applied to analyze and compare the relational degree of each variable in each mode finally.</p></sec></sec><sec id="s5"><title>5. Experiment and Results Analysis</title><sec id="s5_1"><title>5.1. Data Preparation</title><p>This data is from the University of California, Irvine (UCI) database [<xref ref-type="bibr" rid="scirp.75300-ref17">17</xref>], mainly consists of four parts: 1) the total household electricity consumption, including global active power, global reactive power, average voltage, average current. 2) Kitchen power consumption, including a dishwasher, a microwave oven and oven. 3) Laundry power consumption, including a washing machine, a dryer, a refrigerator and a lamp. 4) Living area of electricity, including the water heater and air conditioner. The database contains the electricity consumption of the family from 2006 to 2010, measured once every minute. Because the analysis of household electricity consumption focuses on the information mining process, so although this paper uses foreign data that does not affect the final conclusion, and as China’s smart grid construction is maturing, household power split measurement is an inevitable trend.10 days ( 2010/8/2 to 2010/8/11 ) of electricity data were selected randomly as an analysis of samples to build data cube.</p></sec><sec id="s5_2"><title>5.2. Household Power Modes Mining</title><p>Considering the purpose of mining household electricity modes, and that household energy consumption is mainly based on active load, therefore, this paper chooses the daily load data as the analysis parameters, and uses the system clustering method in Section 2.2 to excavate the electricity consumption model to get the agglomeration schedule, as shown in <xref ref-type="table" rid="table1">Table 1</xref>.</p><table-wrap id="table1" ><label><xref ref-type="table" rid="table1">Table 1</xref></label><caption><title> Agglomeration schedule</title></caption><table><tbody><thead><tr><th align="center" valign="middle"  rowspan="2"  >Step</th><th align="center" valign="middle"  colspan="2"  >Aggregation cluster</th><th align="center" valign="middle"  rowspan="2"  >correlation coefficient</th><th align="center" valign="middle"  colspan="2"  >first clustering step</th></tr></thead><tr><td align="center" valign="middle" >Cluster 1</td><td align="center" valign="middle" >Cluster 2</td><td align="center" valign="middle" >Cluster 1</td><td align="center" valign="middle" >Cluster 2</td></tr><tr><td align="center" valign="middle" >1</td><td align="center" valign="middle" >9</td><td align="center" valign="middle" >10</td><td align="center" valign="middle" >131.640</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0</td></tr><tr><td align="center" valign="middle" >2</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >4</td><td align="center" valign="middle" >180.789</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0</td></tr><tr><td align="center" valign="middle" >3</td><td align="center" valign="middle" >5</td><td align="center" valign="middle" >6</td><td align="center" valign="middle" >189.296</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0</td></tr><tr><td align="center" valign="middle" >4</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >207.926</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0</td></tr><tr><td align="center" valign="middle" >5</td><td align="center" valign="middle" >8</td><td align="center" valign="middle" >9</td><td align="center" valign="middle" >220.040</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >1</td></tr><tr><td align="center" valign="middle" >6</td><td align="center" valign="middle" >5</td><td align="center" valign="middle" >7</td><td align="center" valign="middle" >269.404</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >0</td></tr><tr><td align="center" valign="middle" >7</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >289.126</td><td align="center" valign="middle" >4</td><td align="center" valign="middle" >2</td></tr><tr><td align="center" valign="middle" >8</td><td align="center" valign="middle" >5</td><td align="center" valign="middle" >8</td><td align="center" valign="middle" >329.282</td><td align="center" valign="middle" >6</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" >9</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >5</td><td align="center" valign="middle" >343.351</td><td align="center" valign="middle" >7</td><td align="center" valign="middle" >8</td></tr></tbody></table></table-wrap><p>The main purpose of clustering is to make the electricity consumption in the same mode as similar as possible, the electricity consumption of different modes as different as possible. Therefore, considering only the situation that the first agglomeration step is 0, agglomeration process of steps 7, 8, 9 are excluded. And the pedigree chart is shown in <xref ref-type="fig" rid="fig3">Figure 3</xref>.</p><p>From the figure above, we can see that, from 2010/8/2 to 8/11, the household electricity consumption can be roughly divided into four modes, and each mode is a clustering of days, which reflects the similarity of household electricity consumption in a short time. The median eigenvalue represents the corresponding modes, as shown in <xref ref-type="fig" rid="fig4">Figure 4</xref>. It can be seen that there are two peak powers in each household electricity consumption modes, where the peak period is gradually advanced from mode 4 to mode 1.</p><p>After analyzing a large number of sample data, it is found that the family has a basic load with amplitude of about 0.4 kW. The load fluctuation is small, which is the minimum basic energy loss of the household, while the user electricity</p><fig id="fig3"  position="float"><label><xref ref-type="fig" rid="fig3">Figure 3</xref></label><caption><title> Pedigree chart</title></caption><graphic mimetype="image"   position="float"  xlink:type="simple"  xlink:href="http://html.scirp.org/file/75300x24.png"/></fig><fig id="fig4"  position="float"><label><xref ref-type="fig" rid="fig4">Figure 4</xref></label><caption><title> Graph of modes’ feature</title></caption><graphic mimetype="image"   position="float"  xlink:type="simple"  xlink:href="http://html.scirp.org/file/75300x25.png"/></fig><p>consumption load which is mainly related to the use of electrical appliances is fluctuated obviously. In order to distinguish the user’s electricity consumption load and household basic load, the load fluctuation threshold is set to be 0.5. Then the basic characteristics of the power consumption modes are obtained as shown in <xref ref-type="table" rid="table2">Table 2</xref>.</p></sec><sec id="s5_3"><title>5.3. Household Electricity Feature Mining</title><p>1) Experiment 1</p><p>In order to verify the short-term similarity of the household electricity consumption behavior, the electricity data of 8/12 are adopted to compare the correlation between each power consumption modes and the power consumption in 8/12. Besides, based on the electricity data from 8/12 to 8/16, this paper compares the relationship between the past 5 days’ power consumption and the most recent power consumption mode. Detailed data are shown in <xref ref-type="table" rid="table3">Table 3</xref> and <xref ref-type="table" rid="table4">Table 4</xref>.</p><p>It can be seen from <xref ref-type="table" rid="table3">Table 3</xref> and <xref ref-type="table" rid="table4">Table 4</xref> that the gray relational degree between each mode with 8/12 household electricity consumption is basically reduced as the time interval larger, that is, the power characteristic of mode 1 is most similar to that of 8/12. In addition, the closer the date to mode 1, the greater the degree of correlation is, and the difference between the recent 3 days of</p><table-wrap id="table2" ><label><xref ref-type="table" rid="table2">Table 2</xref></label><caption><title> Basic characteristics of power mode</title></caption><table><tbody><thead><tr><th align="center" valign="middle" >mode</th><th align="center" valign="middle" >peak period</th><th align="center" valign="middle" >characteristics of the power consumption</th></tr></thead><tr><td align="center" valign="middle" >1</td><td align="center" valign="middle" >1:00-2:00 12:45-14:00</td><td align="center" valign="middle" >After the early peak, there is no obvious electricity consumption behavior and is mainly the basic load power consumption. After the latter peak, there is the same kind of electricity behavior or the same kind of electrical appliances is used.</td></tr><tr><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2:15-3:30 14:30-15:30</td><td align="center" valign="middle" >There is a similar power consumption behavior or the use of same electrical appliances after the power peak. There is no obvious power consumption behavior after latter peak.</td></tr><tr><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3:30-5:00 15:15-17:00</td><td align="center" valign="middle" >The electricity consumption is frequent and similar in the whole day.</td></tr><tr><td align="center" valign="middle" >4</td><td align="center" valign="middle" >4:15-5:30 16:00-17:30</td><td align="center" valign="middle" >The electricity consumption is frequent in the whole day. 2 hours after the early peak and 2.5 hours after the latter the peak exist active electricity consumption behaviors.</td></tr></tbody></table></table-wrap><table-wrap id="table3" ><label><xref ref-type="table" rid="table3">Table 3</xref></label><caption><title> Gray relational degree of household electricity in 8/12</title></caption><table><tbody><thead><tr><th align="center" valign="middle" ></th><th align="center" valign="middle" >Mode 1</th><th align="center" valign="middle" >Mode 2</th><th align="center" valign="middle" >Model 3</th><th align="center" valign="middle" >Model 4</th></tr></thead><tr><td align="center" valign="middle" >Grey relational degree</td><td align="center" valign="middle" >0.8776</td><td align="center" valign="middle" >0.8274</td><td align="center" valign="middle" >0.8196</td><td align="center" valign="middle" >0.8223</td></tr></tbody></table></table-wrap><table-wrap id="table4" ><label><xref ref-type="table" rid="table4">Table 4</xref></label><caption><title> Gray correlation degree of model 1</title></caption><table><tbody><thead><tr><th align="center" valign="middle" >Data</th><th align="center" valign="middle" >8/12</th><th align="center" valign="middle" >8/13</th><th align="center" valign="middle" >8/14</th><th align="center" valign="middle" >8/15</th><th align="center" valign="middle" >8/16</th></tr></thead><tr><td align="center" valign="middle" >Grey relational degree</td><td align="center" valign="middle" >0.8883</td><td align="center" valign="middle" >0.8850</td><td align="center" valign="middle" >0.8710</td><td align="center" valign="middle" >0.8165</td><td align="center" valign="middle" >0.7978</td></tr></tbody></table></table-wrap><p>gray relational degree is not significant. The above shows that in the short term, the household electricity behavior has certain continuity and similarity. Through the large number of fitting experiments on the random number of days’ load, it is found that the similar electricity behavior cycle is 2 - 3 days, so utilizing the recent household electricity consumption model to predict this family 2 - 3 days power consumption is feasible.</p><p>2) Experiment 2</p><p>In this experiment, the data of phase parameter in each mode are taken to analyze the correlation between the phase variables. As the kitchen power consumption is zero, we select the two phase variables as the laundry and living area to analyze, and ultimately get the potential impact of each power model, as shown in <xref ref-type="table" rid="table5">Table 5</xref>.</p><p>It can be seen that in the mode 1 to mode 4, the gray relational degree of living area is greater than the laundry’s. Compared with the energy consumption of the laundry, the load curve of the living area is more similar to the characteristic curve in the corresponding mode, indicating that the electricity consumption habits of living area is more dominant than the laundry’s in the whole household.</p><p>From <xref ref-type="table" rid="table2">Table 2</xref>, we can see that the power peak hours is about 1.5 h, so taking the mode 3 for example, the living area’s peak time is delayed 1.5 h to get the curve shown in <xref ref-type="fig" rid="fig5">Figure 5</xref>, which shows that compared with the original load curve, living area electricity of mode 3 is mainly concentrated in the peak hours, after delaying the living area, the household peak load is obviously reduced and the load curve is relatively gentle. In summary, appropriate adjustment on the dominant phase variables is beneficial to improve the overall electricity load</p><table-wrap id="table5" ><label><xref ref-type="table" rid="table5">Table 5</xref></label><caption><title> Gray relational degree of phase variables</title></caption><table><tbody><thead><tr><th align="center" valign="middle" ></th><th align="center" valign="middle" >Mode 1</th><th align="center" valign="middle" >Mode 2</th><th align="center" valign="middle" >Mode 3</th><th align="center" valign="middle" >Mode 4</th></tr></thead><tr><td align="center" valign="middle" >Laundry</td><td align="center" valign="middle" >0.7884</td><td align="center" valign="middle" >0.7852</td><td align="center" valign="middle" >0.7798</td><td align="center" valign="middle" >0.8445</td></tr><tr><td align="center" valign="middle" >Living area</td><td align="center" valign="middle" >0.9293</td><td align="center" valign="middle" >0.9276</td><td align="center" valign="middle" >0.9262</td><td align="center" valign="middle" >0.9432</td></tr><tr><td align="center" valign="middle" >Greater influencing factors</td><td align="center" valign="middle" >living area</td><td align="center" valign="middle" >living area</td><td align="center" valign="middle" >living area</td><td align="center" valign="middle" >living area</td></tr></tbody></table></table-wrap><fig id="fig5"  position="float"><label><xref ref-type="fig" rid="fig5">Figure 5</xref></label><caption><title> Comparison of the load curve after adjusting the power time of living area</title></caption><graphic mimetype="image"   position="float"  xlink:type="simple"  xlink:href="http://html.scirp.org/file/75300x26.png"/></fig><p>fluctuation, is helpful to peak shifting and valley filling, and finally obtain some economic significance</p></sec></sec><sec id="s6"><title>6. Conclusion</title><p>This paper presents a mining model for household electricity behavior based on system clustering and gray relational analysis. According to the analysis of a group of actual electricity data in a certain family, this model effectively excavates the household electricity consumption pattern of a certain period, as well as the electricity consumption behavior affects order of the corresponding model, and validates the predictive ability of the latest power consumption mode. This work will help to mine the users’ potential electricity consumption habits for the power companies to develop the appropriate smart power strategy and improve the quality of power service.</p></sec><sec id="s7"><title>Cite this paper</title><p>Xu, M.J. and Wang, Y.H. (2017) Residential Electricity Consumption Behavior Mining Based on System Cluster and Grey Relational Degree. Energy and Power Engineering, 9, 390-400. https://doi.org/10.4236/epe.2017.94B044</p></sec></body><back><ref-list><title>References</title><ref id="scirp.75300-ref1"><label>1</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Sun</surname><given-names> G.Q.</given-names></name>,<name name-style="western"><surname> Li</surname><given-names> Y.C.</given-names></name>,<name name-style="western"><surname> Wei</surname><given-names> Z.N.</given-names></name>,<name name-style="western"><surname> Yang Y.B.</surname><given-names> Zang</given-names></name>,<name name-style="western"><surname> H.X. andBian</surname><given-names> D. </given-names></name>,<etal>et al</etal>. (<year>2015</year>)<article-title>Discussion on Interactive Architecture of Smart Power Utilization</article-title><source> Automation of Electric Power System</source><volume> 39</volume>,<fpage> 68</fpage>-<lpage>74</lpage>.<pub-id pub-id-type="doi"></pub-id></mixed-citation></ref><ref id="scirp.75300-ref2"><label>2</label><mixed-citation publication-type="other" xlink:type="simple">Li, Y., Wang, B.B. andLi, F.X. (2015) Outlook and Thinking of Flexible and Interactive Utilization of Intelligent Power. Automation of Electric Power System, 39, 2-9.</mixed-citation></ref><ref id="scirp.75300-ref3"><label>3</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Wang</surname><given-names> G.H. </given-names></name>,<etal>et al</etal>. (<year>2012</year>)<article-title>Practice and Prospect of China Intelligent power Utilization</article-title><source> Electric Power</source><volume> 45</volume>,<fpage> 1</fpage>-<lpage>5</lpage>.<pub-id pub-id-type="doi"></pub-id></mixed-citation></ref><ref id="scirp.75300-ref4"><label>4</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Lin</surname><given-names> H.Y.</given-names></name>,<name name-style="western"><surname> Zhang</surname><given-names> J.</given-names></name>,<name name-style="western"><surname> Xu</surname><given-names> K.P.</given-names></name>,<name name-style="western"><surname> Pi</surname><given-names> X.J. </given-names></name>,<etal>et al</etal>. (<year>2012</year>)<article-title>Design of Interactive Service Platform for Smart Power Consumption</article-title><source> Power System Technology</source><volume> 36</volume>,<fpage> 255</fpage>-<lpage>259</lpage>.<pub-id pub-id-type="doi"></pub-id></mixed-citation></ref><ref id="scirp.75300-ref5"><label>5</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>He</surname><given-names> Y.X.</given-names></name>,<name name-style="western"><surname> Wang</surname><given-names> B.Xiong</given-names></name>,<name name-style="western"><surname> W.</surname><given-names> Zhang</given-names></name>,<name name-style="western"><surname> T. and Liu</surname><given-names> Y.Y. </given-names></name>,<etal>et al</etal>. (<year>2012</year>)<article-title>Analysis of Residents’ Smart Electricity Consumption Behavior Based on Fuzzy Synthetic Evaluation and the Design of Interactive Mechanism</article-title><source> Power System Technology</source><volume> 36</volume>,<fpage> 247</fpage>-<lpage>252</lpage>.<pub-id pub-id-type="doi"></pub-id></mixed-citation></ref><ref id="scirp.75300-ref6"><label>6</label><mixed-citation publication-type="other" xlink:type="simple">Sheng, W.X., Shi, C.K., Sun, J.P., Zhang, B. and Zhang, T.S. (2013)Characteristics and Research Framework of Automated Demand Response in Smart Utilization. Automation of Electric Power System, 37, 1-7.</mixed-citation></ref><ref id="scirp.75300-ref7"><label>7</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Yin</surname><given-names> S.G.</given-names></name>,<name name-style="western"><surname> Zhang</surname><given-names> Y.</given-names></name>,<name name-style="western"><surname> Bai</surname><given-names> K.M. </given-names></name>,<etal>et al</etal>. (<year>2009</year>)<article-title>A Smart Power Utilization System Based on Real-Time Electricity Prices</article-title><source> Power System Technology</source><volume> 33</volume>,<fpage> 11</fpage>-<lpage>16</lpage>.<pub-id pub-id-type="doi"></pub-id></mixed-citation></ref><ref id="scirp.75300-ref8"><label>8</label><mixed-citation publication-type="other" xlink:type="simple">Zhang, X., Li, D.H. andCheng, M. (2015) Study on Peak Load Shifting Management Based on the Big Data Technology. Modern Electric Power, 32, 66-70.</mixed-citation></ref><ref id="scirp.75300-ref9"><label>9</label><mixed-citation publication-type="other" xlink:type="simple">Zhao, L., Hou, X.Z., Hu, J., Bo, H. and Sun, H.L. (2014) Improved K-Means Algorithm Based Analysis on Massive Data of Intelligent Power Utilization. Power System Technology, 38, 2715-2720.</mixed-citation></ref><ref id="scirp.75300-ref10"><label>10</label><mixed-citation publication-type="other" xlink:type="simple">Peng, X.G., Lai, J.W. andChen, Y. (2014) Application of Clustering Analysis in Typical Power Consumption Profile Analysis. Power System Protection and Control, 42, 68-73.</mixed-citation></ref><ref id="scirp.75300-ref11"><label>11</label><mixed-citation publication-type="other" xlink:type="simple">Guo, X.L. andYu, Y. (2015) A Residential Smart Power Utilization Strategy Based on Cloud Computing. Automation of Electric Power System, 39, 114-119+133.</mixed-citation></ref><ref id="scirp.75300-ref12"><label>12</label><mixed-citation publication-type="other" xlink:type="simple">Zhang, S.X., Liu, J.M., Zhao, B.Z. and Cao, J.P. (2013) Cloud Computing-Based Analysis on Residential Electricity Consumption Behavior. Power System Technology, 37, 1542-1546.</mixed-citation></ref><ref id="scirp.75300-ref13"><label>13</label><mixed-citation publication-type="other" xlink:type="simple">Han, J.W. (2012) Data Mining Concepts and Techniques. China Machine Press, Beijing.</mixed-citation></ref><ref id="scirp.75300-ref14"><label>14</label><mixed-citation publication-type="other" xlink:type="simple">Zhang,X.L., Hao, S.P., Li, J. and Jiang, C.R. (2015) Grey Correlation Based Analysis on Impacting Factors of Maximum Power Point Tracking Control of Wind Power Generating Unit. Power System Technology, 39, 445-449.</mixed-citation></ref><ref id="scirp.75300-ref15"><label>15</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Zhang</surname><given-names> W.Y.</given-names></name>,<name name-style="western"><surname> Men</surname><given-names> D.Y.</given-names></name>,<name name-style="western"><surname> Liang</surname><given-names> J.F. and Wang W.Z. </given-names></name>,<etal>et al</etal>. (<year>2012</year>)<article-title>Monthly Load Forecasting Based on Grey Relational Degree and Least Squares Support Vector Machine</article-title><source> Power System Technology</source><volume> 36</volume>,<fpage> 228</fpage>-<lpage>232</lpage>.<pub-id pub-id-type="doi"></pub-id></mixed-citation></ref><ref id="scirp.75300-ref16"><label>16</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Kong L.H.</surname><given-names> Jiao</given-names></name>,<name name-style="western"><surname> Y.J. andDai</surname><given-names> Z.H. </given-names></name>,<etal>et al</etal>. (<year>2014</year>)<article-title>A New Substation Area Protection Principle Based on Gray Correlation Degree</article-title><source> Power System Technology</source><volume> 38</volume>,<fpage> 2274</fpage>-<lpage>2279</lpage>.<pub-id pub-id-type="doi"></pub-id></mixed-citation></ref><ref id="scirp.75300-ref17"><label>17</label><mixed-citation publication-type="other" xlink:type="simple">(2014) UCI Machine Learning Repository [EB/OL]. University of California, Irvine (UCI). http://archive.ics.uci.edu/ml/index.html.</mixed-citation></ref></ref-list></back></article>