{"id":1156,"date":"2019-05-14T09:13:13","date_gmt":"2019-05-14T09:13:13","guid":{"rendered":"https:\/\/controlloemisura.com\/?p=1156"},"modified":"2025-06-24T15:41:59","modified_gmt":"2025-06-24T15:41:59","slug":"exploring-big-data-to-make-strategic-decisions","status":"publish","type":"post","link":"https:\/\/www.publiteconline.it\/controlloemisura\/exploring-big-data-to-make-strategic-decisions\/","title":{"rendered":"Exploring Big Data to Make Strategic Decisions"},"content":{"rendered":"<p>Big Data may provide companies with very useful information. But a great amount of data is basically useless unless it is converted into something significant. Only once they are combined and normalized they can show their full potential<\/p>\n<p><em>di Valerio Alessandroni<\/em><\/p>\n<p>In 1999 the advent of blogs gave rise to what we would have learnt to know as the \u201cWeb 2.0\u201d: from users of contents generated by a limited number of sources, the users of Internet became the authors of these contents. The amount of information available exploded, and it did not take very long before companies and ICT experts understood that the amount of data generated online, although it was almost impossible to manage, had a huge potential in terms of business intelligence. In 2010, Google\u2019s former Managing Director, Erick Schmidt, recalled how five Exabyte (billions of billions of bytes) of information had been created in the period which goes from the dawning of civilization to 2003. Today these same five Exabytes are created every two days.<\/p>\n<p><strong>A seismic shift worth huge amounts of data<\/strong><br \/>\nJust a few examples are enough to understand this shift which made history. In 2016, the estimated volume of mobile traffic worldwide was equal to 6.2 billion Gbyte (6.2 Exabyte) a month. In 2020 almost 40.000 Exabytes (40 Zetabytes) of data are foreseen. Over 3.5 billion queries are carried out daily on Google, while FaceBook users increase by roughly 22% year after year. We might add the 187 million e-mails, 38 million WhatsApp messages and 18 million text messages exchanged every minute and so on.<br \/>\nIt is estimated that China alone, by 2020, will provide 20% of all data generated on the planet. It is difficult to imagine figures of this shirt. To provide you with an idea, here are a couple of benchmarks. Assuming all beaches on earth contain 700.5 bilion bilion grains of sand, the 40 Zb mentioned before would be equivalent to 57 times that amount. Or, if we could save all the 40 Zb of data on Blu-ray discs, the weight of those discs (excluding covers) would be equal to that of the aircraft carrier warship Nimitz. The expression \u201cBig Data\u201d refers to such large amounts of data.<\/p>\n<p><strong>Megadata will become the cornerstone of the future Internet 3.0<\/strong><br \/>\nOut of all these data, only one fourth could prove useful for companies, and consumers, if they were correctly classified and processed. But only 3% of these data get \u201ctagged\u201d and an even lower percentage, valued at about 0.5%, are actually examined. In their rough form, Big Data are indeed difficult to exploit, but this will not always be the case. We can indeed foresee that megadata will become the cornerstone of the future Internet 3.0 or \u201cSemantic web\u201d, which will imply the change to mass production and targeted consumption. To reach this result, however, we shall require tools capable of processing and structuring megadata, without excessive costs. Currently, only very few companies have the competence and means to exploit megadata, and this is heavily fettering the development of a market which, by 2027, should be worth 103 billion dollars.<br \/>\nIt is however evident that such a huge amount of data cannot be managed and analyzed using traditional methods to mine the information they \u201chide\u201d, and therefore carry out such operations as the predictive maintenance of plants, the analysis of consumers\u2019 behaviour, market projections and so on.<\/p>\n<p><strong>What are the five \u201cV\u201ds which characterize Big Data<\/strong><br \/>\nIn general, it may be said that Big Data are characterized by 5 \u201cV\u201ds: volume, velocity, variety, veracity and value. We mentioned volume already. Velocity refers to the amount of data accumulated in a unit of time, considering there is a huge and continuous flow of Big Data.<br \/>\nRegarding variety, this refers to the nature of data, which may be structured or not structured. Structured data are basically organized data, with a defined length and format, while not structured and generally not organized data do not fit into the traditional row ad column structure of relational data bases. Examples of these data are texts images, videos and so on.<br \/>\nVeracity indicated lack of congruence and uncertainty in data, because those available, coming from different sources, may be disorderly, making it difficult to check their quality and precision. Particularly, a mass of disorderly data may create confusion, while few data may provide incomplete information.<br \/>\nFinally, value. It must be remembered that a mass of data is worthless and basically useless unless it is turned into something significant, from where information may be derived. This is the most important \u201cV\u201d out of the five.<\/p>\n<p><strong>The importance of making decisions based on models and trends<\/strong><br \/>\nAfter combining and normalizing them, Big Data may display all their potential. For instance, analysis models and algorithms may be applied to identify the possible operative and energy savings, future malfunctioning of a company\u2019s production systems may be foreseen, and so on. Making decisions based on models and trends identified in the gathered data may make the difference in the management of the building. The problem is doing so using a BMS, which is typically designed to command and control the building\u2019s systems while data are gathered. A new platform must therefore be associated to the BMS which must take care of Data Mining, that is, of the software which can derive useful information from the large amount of available data.<\/p>\n<p>Data Mining: fully exploiting the data available in the company<br \/>\nData Mining is a process whereby knowledge is derived from large data banks by means of the application of algorithms which detect \u201chidden\u201d associations (patterns) in the data and make them visible.<br \/>\nStatistical data analysis may enhance abnormal behaviour and connections between environment and functioning parameters, which may be useful for instance, to optimize maintenance action.<br \/>\nIt may be discovered that a certain type of fault in an appliance always occurs when there is a drop in voltage superior to 5%, or when the appliance is used along with another device.<br \/>\nUnlike statistics, which allows to process such general information as unemployment or birth rates, data mining is used to seek connections between more variables relative to single subjects; by knowing the average behaviour of the clients of a telephone company, for instance, it is possible to try to foresee how much the average client will spend in the immediate future. Or else, by analyzing the vibrations of a mechanical line shaft, the moment of its breakage may be predicted in advance.<\/p>\n<p><strong>A concrete example in the inspection of hydraulic valves<\/strong><br \/>\nData Mining techniques are becoming very important considering the progress of Industry 4.0 and Internet of Things. The latter allows to collect great amounts of data using \u201cintelligent\u201d sensors positioned within the objects (tools, electric appliances, industrial devices and so on): These data, however, would be worthless if they did not provide us with useful information.<br \/>\nThe purpose of Data Mining is deriving \u201chidden\u201d information from large amounts of data which, observed as such, would appear random o meaningless.<br \/>\nBy evaluating production data, for instance, an important German company managed to reduce the time needed to inspect hydraulic valves by 17.4%.<br \/>\nWith 40,000 valves produces each year, the company can count on a saving of 14 days.<br \/>\nWhen we talk about millions of items, even a few seconds saved may build up rapidly, turning cents into millions of euros.<br \/>\nThe capability of generating new knowledge from Big Data is therefor proving to be one of the key competences of the future.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Big Data may provide companies with very useful information. But a great amount of data is basically useless unless it is converted into something significant. Only once they are combined and normalized they can show their full potential di Valerio Alessandroni In 1999 the advent of blogs gave rise to what we would have learnt &hellip;<\/p>\n","protected":false},"author":4,"featured_media":1154,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":"","_seopublitec_title":"","_seopublitec_meta_desc":"","_seopublitec_keyphrases":"","_seopublitec_canonical":"","_seopublitec_noindex":false,"_seopublitec_nofollow":false,"_seopublitec_og_title":"","_seopublitec_og_desc":"","_seopublitec_twitter_title":"","_seopublitec_twitter_desc":"","_seopublitec_twitter_image":"","_seopublitec_og_image":"","_seopublitec_twitter_card":"","_seopublitec_cornerstone":false,"_seopublitec_score":""},"categories":[27],"tags":[],"class_list":["post-1156","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-tecnologie"],"openstation_lock":null,"openstation_contributors":[],"openstation_attached_media":[1154],"_links":{"self":[{"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/posts\/1156","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/users\/4"}],"replies":[{"embeddable":true,"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/comments?post=1156"}],"version-history":[{"count":1,"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/posts\/1156\/revisions"}],"predecessor-version":[{"id":1157,"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/posts\/1156\/revisions\/1157"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/media\/1154"}],"wp:attachment":[{"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/media?parent=1156"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/categories?post=1156"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.publiteconline.it\/controlloemisura\/wp-json\/wp\/v2\/tags?post=1156"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}