Welcome!

Java IoT Authors: Elizabeth White, Pat Romanski, Carmen Gonzalez, Liz McMillan, Yeshim Deniz

Related Topics: @BigDataExpo, @CloudExpo, @ThingsExpo

@BigDataExpo: Blog Feed Post

Identifying Where and How to Start the Big Data Journey | @BigDataExpo #BigData #DataLake #Analytics

Organizations are eager to realize the business benefits of Big Data that they don’t take the time to do the little things first

Decisions Exercise: Identifying Where and How to Start the Big Data Journey

The recent deluge of rains in Northern California have flooded streets, brought down trees and plugged storm sewers.  As I was trying to make my way around the neighborhood, I thought of a classroom exercise to help my MBA students to identify the use cases upon which they could focus data and analytics.  In this exercise, I’m going to ask my students to pretend that they have been hired by the city to “Optimize Street Maintenance” after these rainstorms. In particular, the students need to address the following questions:

  • Where and how do you start to address this initiative?
  • What data might you need to support this initiative?

These are classic questions that I hear all the time when I meet with clients about their big data journeys.  Let’s walk through how I’ll teach my students to address this challenge.

Step 1:  Identify and Brainstorm the Decisions
“Where and how to start?” is such an open ended question.  How does one even begin to think about that question?  We recommend that organizations start by identifying the decisions that need to be made to support the targeted business initiative, which is “Optimize Street Maintenance” in this exercise.

I will break up the students into small groups (3 to 5 students) and ask them to brainstorm the decisions that need to be made with respect to the “Optimize Street Maintenance” initiative.  Those decisions could include:

  • What streets and intersections need maintenance?
  • What storm sewers are blocked?
  • What is blocking those storm sewers?
  • What sort of maintenance is needed?
  • What is the impact of street cleaning and debris removal on flooding?
  • What streets and intersections should we fix first?
  • How busy are the streets and intersections?
  • What worker skills are needed to fix the street?
  • What equipment and materials are needed to fix the street?
  • What time of the day / day of the week is ideal for doing that maintenance work?
  • How many workers are available?
  • Do I have access to temporary workers?
  • How much overtime can I afford?
  • How do I warn residents that a road is flooded?
  • What options do I give residents when the major arteries are flooded?

This brainstorming is much more effective when you have brought together the different business stakeholders who either impact or are impacted by the “Accelerate Street Maintenance” initiative (see Figure 1).

Figure 1: Brainstorm Decisions Across Different Stakeholders

Some key process points about Step 1:

  • Allow individuals to brainstorm on their own at first. When it is entirely a group exercise, some folks go quiet and we potentially lose some good ideas.
  • Be sure to capture each decision on a separate Post-It note for later usage.
  • Place the decisions/Post-it Notes on a flip chart (or two).
  • You don’t need to group decisions by business function. I just did it here to demonstrate the process.

Finally, “all ideas are worthy of consideration.”  This is the key to any brainstorming session; to create an environment where everyone feels comfortable to contribute without someone passing judgment about his or her thoughts or ideas.

Step 2:  Group Decisions Into Use Cases
Next, we want to group the decisions into common subject areas or use cases (which is much easier to do if each decision is captured on a separate Post-It note).  I will bring all the students together around the decisions on Post-it Notes, and have them look for logical groupings.

Looking over the decisions captured above, we can start to see some natural “Accelerate Street Maintenance” use cases emerging, such as:

Prioritize Streets and Intersections

  • What streets and intersections should we fix first?
  • What streets and intersections are busiest at what times of the day?
  • What are the alternative route options during maintenance?
  • What are the alternative transportation options during maintenance?
  • What business parks or malls will be disrupted by the maintenance work?
  • Which streets and intersections raise safety concerns for bikers and pedestrians?

Estimate Maintenance Effort

  • What streets and intersections need maintenance?
  • What storm sewers need maintenance?
  • How much maintenance is needed?
  • What type of maintenance is needed?
  • What worker maintenance skills are needed?
  • What types of equipment and materials are needed?

Optimize Maintenance Effort

  • What worker skills are needed to fix the street?
  • How many workers with those skills are available?
  • What equipment is available to fix the street?
  • What tools are needed to fix the street?
  • What materials (concrete, asphalt) are needed to fix the street?
  • How effective is street cleaning and debris removal in preventing flooding?

Minimize Traffic Disruptions

  • Which streets are bottlenecks for schools and at what times of the day?
  • Which streets are bottlenecks for shopping malls and at what times of the day?
  • Which streets are bottlenecks for business parks and at what times of the day?
  • What are the alternative route options?
  • What are the public transportation options?

Minimize Maintenance Costs

  • How many workers are available?
  • To what temporary workers do we have access?
  • How much overtime can I afford?
  • How much maintenance budget is available?

Improve Resident Communications

  • What streets need maintenance?
  • What streets and intersections are likely to need maintenance?
  • What are alternative travel routes?
  • What are alternative transportation options?

Increase Resident Satisfaction

  • How many residents did the flooding impact?
  • How long were those residents impacted?
  • What comments or feedback are most important and/or relevant?
  • What phone calls are most important and/or relevant?
  • What social media postings are important and/or relevant?

See Figure 2 for an example of how the end point of Step 2 might look.

A key process point about Step 2:

  • Ideally you will end up with 7 to 12 use cases. If you have fewer than 7, then look for ways to break up some of the groupings.  If you have more than 12, then look for ways to aggregate similar use cases.  Not sure why, but 7 to 12 use cases always seems to work out to the right level of granularity in the use cases.

Step 3:  Prioritize Use Cases
Not all use cases are equal, and some use cases are dependent upon other use cases.  The prioritization matrix takes the different business stakeholders through a facilitated process to prioritize each use case vis-à-vis its business value and implementation feasibility (see Figure 3).

Figure 3: Prioritization Matrix

For more details on the prioritization process, check out these blogs:

Summary
The news really surprised no one:  “MD Anderson Benches IBM Watson In Setback For Artificial Intelligence In Medicine.”  From the press release:

“The partnership between IBM and one of the world’s top cancer research institutions is falling apart. The project is on hold, MD Anderson confirms, and has been since late last year. MD Anderson is actively requesting bids from other contractors who might replace IBM in future efforts.  And a scathing report from auditors at the University of Texas says the project cost MD Anderson more than $62 million and yet did not meet its goals.”

If big data were only about buying and installing technology, then it would be easy.  Unfortunately, companies are learning the hard way that the “big bang” approach for implementing big data is fraught with misguided expectations and outright failures.

Organizations are so eager to realize the business benefits of big data, that they don’t take the time to do the little things first, like identifying and prioritizing those use cases that offer the optimal mix of business value and implementation feasibility. While I applaud all efforts to cure cancer (my mom died from cancer, so I have a vested interest like so many others), sometimes “curing cancer” might not be the best place to start.  Identifying and prioritizing those use cases that move the organization towards that “cure cancer” aspiration is the best way to achieve that goal.

The post Decisions Exercise: Identifying Where and How To Start the Big Data Journey appeared first on InFocus Blog | Dell EMC Services.

Read the original blog entry...

More Stories By William Schmarzo

Bill Schmarzo, author of “Big Data: Understanding How Data Powers Big Business”, is responsible for setting the strategy and defining the Big Data service line offerings and capabilities for the EMC Global Services organization. As part of Bill’s CTO charter, he is responsible for working with organizations to help them identify where and how to start their big data journeys. He’s written several white papers, avid blogger and is a frequent speaker on the use of Big Data and advanced analytics to power organization’s key business initiatives. He also teaches the “Big Data MBA” at the University of San Francisco School of Management.

Bill has nearly three decades of experience in data warehousing, BI and analytics. Bill authored EMC’s Vision Workshop methodology that links an organization’s strategic business initiatives with their supporting data and analytic requirements, and co-authored with Ralph Kimball a series of articles on analytic applications. Bill has served on The Data Warehouse Institute’s faculty as the head of the analytic applications curriculum.

Previously, Bill was the Vice President of Advertiser Analytics at Yahoo and the Vice President of Analytic Applications at Business Objects.

@ThingsExpo Stories
Five years ago development was seen as a dead-end career, now it’s anything but – with an explosion in mobile and IoT initiatives increasing the demand for skilled engineers. But apart from having a ready supply of great coders, what constitutes true ‘DevOps Royalty’? It’ll be the ability to craft resilient architectures, supportability, security everywhere across the software lifecycle. In his keynote at @DevOpsSummit at 20th Cloud Expo, Jeffrey Scheaffer, GM and SVP, Continuous Delivery Busine...
SYS-CON Events announced today that Outscale, a global pure play Infrastructure as a Service provider and strategic partner of Dassault Systèmes, will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Founded in 2010, Outscale simplifies infrastructure complexities and boosts the business agility of its customers. Outscale delivers a secure, reliable and industrial strength solution for its customers, which in...
SYS-CON Events announced today that CollabNet, a global leader in enterprise software development, release automation and DevOps solutions, will be a Bronze Sponsor of SYS-CON's 20th International Cloud Expo®, taking place from June 6-8, 2017, at the Javits Center in New York City, NY. CollabNet offers a broad range of solutions with the mission of helping modern organizations deliver quality software at speed. The company’s latest innovation, the DevOps Lifecycle Manager (DLM), supports Value S...
SYS-CON Events announced today that Peak 10, Inc., a national IT infrastructure and cloud services provider, will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Peak 10 provides reliable, tailored data center and network services, cloud and managed services. Its solutions are designed to scale and adapt to customers’ changing business needs, enabling them to lower costs, improve performance and focus intern...
A strange thing is happening along the way to the Internet of Things, namely far too many devices to work with and manage. It has become clear that we'll need much higher efficiency user experiences that can allow us to more easily and scalably work with the thousands of devices that will soon be in each of our lives. Enter the conversational interface revolution, combining bots we can literally talk with, gesture to, and even direct with our thoughts, with embedded artificial intelligence, whic...
In order to meet the rapidly changing demands of today’s customers, companies are continually forced to redefine their business strategies in order to meet these needs, stay relevant and continue to see profitable growth. IoT deployment and development is integral in this transformation, and today businesses are increasingly seeing the value of investing their resources into IoT deployments. These technologies are able increase ROI through projects such as connecting supply chains or enabling sm...
In his opening keynote at 20th Cloud Expo, Michael Maximilien, Research Scientist, Architect, and Engineer at IBM, will motivate why realizing the full potential of the cloud and social data requires artificial intelligence. By mixing Cloud Foundry and the rich set of Watson services, IBM's Bluemix is the best cloud operating system for enterprises today, providing rapid development and deployment of applications that can take advantage of the rich catalog of Watson services to help drive insigh...
SYS-CON Events announced today that SoftLayer, an IBM Company, has been named “Gold Sponsor” of SYS-CON's 18th Cloud Expo, which will take place on June 7-9, 2016, at the Javits Center in New York, New York. SoftLayer, an IBM Company, provides cloud infrastructure as a service from a growing number of data centers and network points of presence around the world. SoftLayer’s customers range from Web startups to global enterprises.
SYS-CON Events announced today that Super Micro Computer, Inc., a global leader in compute, storage and networking technologies, will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Supermicro (NASDAQ: SMCI), the leading innovator in high-performance, high-efficiency server technology, is a premier provider of advanced server Building Block Solutions® for Data Center, Cloud Computing, Enterprise IT, Hadoop/...
With major technology companies and startups seriously embracing Cloud strategies, now is the perfect time to attend @CloudExpo | @ThingsExpo, June 6-8, 2017, at the Javits Center in New York City, NY and October 31 - November 2, 2017, Santa Clara Convention Center, CA. Learn what is going on, contribute to the discussions, and ensure that your enterprise is on the right path to Digital Transformation.
SYS-CON Events announced today that EARP Integration will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. EARP Integration is a passionate software house. Since its inception in 2009 the company successfully delivers smart solutions for cities and factories that start their digital transformation. EARP provides bespoke solutions like, for example, advanced enterprise portals, business intelligence systems an...
Existing Big Data solutions are mainly focused on the discovery and analysis of data. The solutions are scalable and highly available but tedious when swapping in and swapping out occurs in disarray and thrashing takes place. The resolution for thrashing through machine learning algorithms and support nomenclature is through simple techniques. Organizations that have been collecting large customer data are increasingly seeing the need to use the data for swapping in and out and thrashing occurs ...
Amazon started as an online bookseller 20 years ago. Since then, it has evolved into a technology juggernaut that has disrupted multiple markets and industries and touches many aspects of our lives. It is a relentless technology and business model innovator driving disruption throughout numerous ecosystems. Amazon’s AWS revenues alone are approaching $16B a year making it one of the largest IT companies in the world. With dominant offerings in Cloud, IoT, eCommerce, Big Data, AI, Digital Assis...
SYS-CON Events announced today that Progress, a global leader in application development, has been named “Bronze Sponsor” of SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Enterprises today are rapidly adopting the cloud, while continuing to retain business-critical/sensitive data inside the firewall. This is creating two separate data silos – one inside the firewall and the other outside the firewall. Cloud ISVs oft...
The 21st International Cloud Expo has announced that its Call for Papers is open. Cloud Expo, to be held October 31 - November 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA, brings together Cloud Computing, Big Data, Internet of Things, DevOps, Digital Transformation, Machine Learning and WebRTC to one location. With cloud computing driving a higher percentage of enterprise IT budgets every year, it becomes increasingly important to plant your flag in this fast-expanding busin...
Internet of @ThingsExpo, taking place October 31 - November 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA, is co-located with the 21st International Cloud Expo and will feature technical sessions from a rock star conference faculty and the leading industry players in the world. @ThingsExpo Silicon Valley Call for Papers is now open.
As cloud adoption continues to transform business, today's global enterprises are challenged with managing a growing amount of information living outside of the data center. The rapid adoption of IoT and increasingly mobile workforce are exacerbating the problem. Ensuring secure data sharing and efficient backup poses capacity and bandwidth considerations as well as policy and regulatory compliance issues.
SYS-CON Events announced today that Interoute has been named “Bronze Sponsor” of SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Interoute is the owner operator of Europe's largest network and a global cloud services platform, which encompasses over 70,000 km of lit fiber, 15 data centers, 17 virtual data centers and 33 colocation centers, with connections to 195 additional partner data centers. Our full-service Unifie...
SYS-CON Events announced today that Progress, a global leader in application development, has been named “Bronze Sponsor” of SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Enterprises today are rapidly adopting the cloud, while continuing to retain business-critical/sensitive data inside the firewall. This is creating two separate data silos – one inside the firewall and the other outside the firewall. Cloud ISVs ofte...
DevOps is often described as a combination of technology and culture. Without both, DevOps isn't complete. However, applying the culture to outdated technology is a recipe for disaster; as response times grow and connections between teams are delayed by technology, the culture will die. A Nutanix Enterprise Cloud has many benefits that provide the needed base for a true DevOps paradigm.