Doing Big Data – Intelligently
14/12/2016

Last week we blogged about a term used by the Nobel Prize-winning author Daniel Kahneman in his book Thinking, Fast and Slow – WYSIATI (What you see is all there is). It refers to people wanting to see a complex world simplistically, whereas we would advocate that you have to understand (at some level at any rate) the complexity before simplifying – in other words “Simplicity on the far side of complexity”. Totally the opposite of WYSIATI – or “Simplistic on the near side of complexity”.
Many promises of Big Data fall into the area of simplistic on the near side of complexity. We hear Big Data is going to revolutionise our world from retail, financial services etc. through asset management and on into national security. This may be the case with unstructured data (i.e. free text), and possibly with structured data in some circumstances. It is obvious how a retailer might want to analyse all shoppers who purchased a certain washing powder for sensitive shin and run a promotion to also purchase sensitive skin cream and/or clothes made from materials for sensitive skin. Of course, retailers have been doing this for years – just not calling it Big Data.
However, there is a Big Question around Big Data revolutionising the Internet of Things or the Asset Management Industry. With sensor data coming from a myriad of equipment and machinery at the rate of every (typically) 100ms, we are being persuaded that it is necessary to capture every data point as if it has meaning – giving rise to vast so-called “data lakes” of petabytes and zeta bytes or whatever in size, with little hope of analysing it. It may revolutionise sales of Big Data solutions, but does it add value to understanding and analysis?
We came across a world-leading rubber moulding/sealing company in our travels over the past four years or so, and here we found this simplistic on the near side of complexity approach of “just put all the data into one place and run an algorithm to tell us what levers to pull” to get a perfect product. There was little interest in understanding, just “give me an answer”. They were generating large amounts of scrap – and still are.
More recently, we are working with a large transport organisation with many lifts in their buildings. We recognised that every data point did not hold value and that each data point was infected with electrical / mechanical noise that needed removal – in fact, through some intelligent analysis, we reduced the volume of rows by a factor of 400:1. Further analysis allowed us to enter the data into a proprietary time-series analysis platform and we were able to visualise normal behaviour hour by hour over several weeks, spot anomalies, see trends and patterns – all of which have never been seen before.
The Big Question – how would Big Data have provided such insight?
Categories & Tags:
Leave a comment on this post:
You might also like…
Using AI tools for your literature review
There are a proliferation of AI tools that can help you organise your life, work and study. This post focuses on academic or scholarly tools that have been developed to enhance the literature searching process, whether for independent research, an assignment or thesis. Bear in mind that these predominantly relate to finding journal/research papers, and not technical, business or trade sources such as standards, market research, industry reports or financial data. So ...
Finding successful past Cranfield theses
It’s always a good idea to look at examples of theses before you start work on your own. You may find them valuable for reading previous research, and for looking at structure, style and methodology. ...
On‑campus or off‑campus? How Cranfield students found their home away from home
Finding the right place to live is one of the biggest decisions you’ll make as you begin your student journey. Whether you’re looking for the convenience and community of living on-campus or the independence ...
Avoiding common referencing errors
As librarians, we get to see the full spectrum of reference lists in student work —from exemplary to … well, let’s just say, works still very much in progress! We are experts in spotting mistakes ...
Using your Mendeley library after you have left Cranfield
So you have spent the whole year (or more) lovingly collecting references around the topics that matter to you and now you have a large, personalised library in Mendeley Reference Manager containing all that information. ...
Referencing the use of generative AI in your work
We recognise that Artificial Intelligence (AI) has, and will increasingly, become a part of our everyday lives and that we need to adapt to it. Hopefully you will have already seen the guidance for staff ...
