Welcome!

Video Authors: Elizabeth White, Yakov Fain, Liz McMillan, Dan Ristic, Jnan Dash

Related Topics: @BigDataExpo, @CloudExpo, @ThingsExpo

@BigDataExpo: Blog Feed Post

Identifying Where and How to Start the Big Data Journey | @BigDataExpo #BigData #DataLake #Analytics

Organizations are eager to realize the business benefits of Big Data that they don’t take the time to do the little things first

Decisions Exercise: Identifying Where and How to Start the Big Data Journey

The recent deluge of rains in Northern California have flooded streets, brought down trees and plugged storm sewers.  As I was trying to make my way around the neighborhood, I thought of a classroom exercise to help my MBA students to identify the use cases upon which they could focus data and analytics.  In this exercise, I’m going to ask my students to pretend that they have been hired by the city to “Optimize Street Maintenance” after these rainstorms. In particular, the students need to address the following questions:

  • Where and how do you start to address this initiative?
  • What data might you need to support this initiative?

These are classic questions that I hear all the time when I meet with clients about their big data journeys.  Let’s walk through how I’ll teach my students to address this challenge.

Step 1:  Identify and Brainstorm the Decisions
“Where and how to start?” is such an open ended question.  How does one even begin to think about that question?  We recommend that organizations start by identifying the decisions that need to be made to support the targeted business initiative, which is “Optimize Street Maintenance” in this exercise.

I will break up the students into small groups (3 to 5 students) and ask them to brainstorm the decisions that need to be made with respect to the “Optimize Street Maintenance” initiative.  Those decisions could include:

  • What streets and intersections need maintenance?
  • What storm sewers are blocked?
  • What is blocking those storm sewers?
  • What sort of maintenance is needed?
  • What is the impact of street cleaning and debris removal on flooding?
  • What streets and intersections should we fix first?
  • How busy are the streets and intersections?
  • What worker skills are needed to fix the street?
  • What equipment and materials are needed to fix the street?
  • What time of the day / day of the week is ideal for doing that maintenance work?
  • How many workers are available?
  • Do I have access to temporary workers?
  • How much overtime can I afford?
  • How do I warn residents that a road is flooded?
  • What options do I give residents when the major arteries are flooded?

This brainstorming is much more effective when you have brought together the different business stakeholders who either impact or are impacted by the “Accelerate Street Maintenance” initiative (see Figure 1).

Figure 1: Brainstorm Decisions Across Different Stakeholders

Some key process points about Step 1:

  • Allow individuals to brainstorm on their own at first. When it is entirely a group exercise, some folks go quiet and we potentially lose some good ideas.
  • Be sure to capture each decision on a separate Post-It note for later usage.
  • Place the decisions/Post-it Notes on a flip chart (or two).
  • You don’t need to group decisions by business function. I just did it here to demonstrate the process.

Finally, “all ideas are worthy of consideration.”  This is the key to any brainstorming session; to create an environment where everyone feels comfortable to contribute without someone passing judgment about his or her thoughts or ideas.

Step 2:  Group Decisions Into Use Cases
Next, we want to group the decisions into common subject areas or use cases (which is much easier to do if each decision is captured on a separate Post-It note).  I will bring all the students together around the decisions on Post-it Notes, and have them look for logical groupings.

Looking over the decisions captured above, we can start to see some natural “Accelerate Street Maintenance” use cases emerging, such as:

Prioritize Streets and Intersections

  • What streets and intersections should we fix first?
  • What streets and intersections are busiest at what times of the day?
  • What are the alternative route options during maintenance?
  • What are the alternative transportation options during maintenance?
  • What business parks or malls will be disrupted by the maintenance work?
  • Which streets and intersections raise safety concerns for bikers and pedestrians?

Estimate Maintenance Effort

  • What streets and intersections need maintenance?
  • What storm sewers need maintenance?
  • How much maintenance is needed?
  • What type of maintenance is needed?
  • What worker maintenance skills are needed?
  • What types of equipment and materials are needed?

Optimize Maintenance Effort

  • What worker skills are needed to fix the street?
  • How many workers with those skills are available?
  • What equipment is available to fix the street?
  • What tools are needed to fix the street?
  • What materials (concrete, asphalt) are needed to fix the street?
  • How effective is street cleaning and debris removal in preventing flooding?

Minimize Traffic Disruptions

  • Which streets are bottlenecks for schools and at what times of the day?
  • Which streets are bottlenecks for shopping malls and at what times of the day?
  • Which streets are bottlenecks for business parks and at what times of the day?
  • What are the alternative route options?
  • What are the public transportation options?

Minimize Maintenance Costs

  • How many workers are available?
  • To what temporary workers do we have access?
  • How much overtime can I afford?
  • How much maintenance budget is available?

Improve Resident Communications

  • What streets need maintenance?
  • What streets and intersections are likely to need maintenance?
  • What are alternative travel routes?
  • What are alternative transportation options?

Increase Resident Satisfaction

  • How many residents did the flooding impact?
  • How long were those residents impacted?
  • What comments or feedback are most important and/or relevant?
  • What phone calls are most important and/or relevant?
  • What social media postings are important and/or relevant?

See Figure 2 for an example of how the end point of Step 2 might look.

A key process point about Step 2:

  • Ideally you will end up with 7 to 12 use cases. If you have fewer than 7, then look for ways to break up some of the groupings.  If you have more than 12, then look for ways to aggregate similar use cases.  Not sure why, but 7 to 12 use cases always seems to work out to the right level of granularity in the use cases.

Step 3:  Prioritize Use Cases
Not all use cases are equal, and some use cases are dependent upon other use cases.  The prioritization matrix takes the different business stakeholders through a facilitated process to prioritize each use case vis-à-vis its business value and implementation feasibility (see Figure 3).

Figure 3: Prioritization Matrix

For more details on the prioritization process, check out these blogs:

Summary
The news really surprised no one:  “MD Anderson Benches IBM Watson In Setback For Artificial Intelligence In Medicine.”  From the press release:

“The partnership between IBM and one of the world’s top cancer research institutions is falling apart. The project is on hold, MD Anderson confirms, and has been since late last year. MD Anderson is actively requesting bids from other contractors who might replace IBM in future efforts.  And a scathing report from auditors at the University of Texas says the project cost MD Anderson more than $62 million and yet did not meet its goals.”

If big data were only about buying and installing technology, then it would be easy.  Unfortunately, companies are learning the hard way that the “big bang” approach for implementing big data is fraught with misguided expectations and outright failures.

Organizations are so eager to realize the business benefits of big data, that they don’t take the time to do the little things first, like identifying and prioritizing those use cases that offer the optimal mix of business value and implementation feasibility. While I applaud all efforts to cure cancer (my mom died from cancer, so I have a vested interest like so many others), sometimes “curing cancer” might not be the best place to start.  Identifying and prioritizing those use cases that move the organization towards that “cure cancer” aspiration is the best way to achieve that goal.

The post Decisions Exercise: Identifying Where and How To Start the Big Data Journey appeared first on InFocus Blog | Dell EMC Services.

Read the original blog entry...

More Stories By William Schmarzo

Bill Schmarzo, author of “Big Data: Understanding How Data Powers Big Business”, is responsible for setting the strategy and defining the Big Data service line offerings and capabilities for the EMC Global Services organization. As part of Bill’s CTO charter, he is responsible for working with organizations to help them identify where and how to start their big data journeys. He’s written several white papers, avid blogger and is a frequent speaker on the use of Big Data and advanced analytics to power organization’s key business initiatives. He also teaches the “Big Data MBA” at the University of San Francisco School of Management.

Bill has nearly three decades of experience in data warehousing, BI and analytics. Bill authored EMC’s Vision Workshop methodology that links an organization’s strategic business initiatives with their supporting data and analytic requirements, and co-authored with Ralph Kimball a series of articles on analytic applications. Bill has served on The Data Warehouse Institute’s faculty as the head of the analytic applications curriculum.

Previously, Bill was the Vice President of Advertiser Analytics at Yahoo and the Vice President of Analytic Applications at Business Objects.

@ThingsExpo Stories
SYS-CON Events announced today that Calligo, an innovative cloud service provider offering mid-sized companies the highest levels of data privacy and security, has been named "Bronze Sponsor" of SYS-CON's 21st International Cloud Expo ®, which will take place on Oct 31 - Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Calligo offers unparalleled application performance guarantees, commercial flexibility and a personalised support service from its globally located cloud plat...
"We've been engaging with a lot of customers including Panasonic, we've been involved with Cisco and now we're working with the U.S. government - the Department of Homeland Security," explained Peter Jung, Chief Product Officer at Pulzze Systems, in this SYS-CON.tv interview at @ThingsExpo, held June 6-8, 2017, at the Javits Center in New York City, NY.
"We provide IoT solutions. We provide the most compatible solutions for many applications. Our solutions are industry agnostic and also protocol agnostic," explained Richard Han, Head of Sales and Marketing and Engineering at Systena America, in this SYS-CON.tv interview at @ThingsExpo, held June 6-8, 2017, at the Javits Center in New York City, NY.
Internet of @ThingsExpo, taking place October 31 - November 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA, is co-located with 21st Cloud Expo and will feature technical sessions from a rock star conference faculty and the leading industry players in the world. The Internet of Things (IoT) is the most profound change in personal and enterprise IT since the creation of the Worldwide Web more than 20 years ago. All major researchers estimate there will be tens of billions devic...
"The Striim platform is a full end-to-end streaming integration and analytics platform that is middleware that covers a lot of different use cases," explained Steve Wilkes, Founder and CTO at Striim, in this SYS-CON.tv interview at 20th Cloud Expo, held June 6-8, 2017, at the Javits Center in New York City, NY.
"We are focused on SAP running in the clouds, to make this super easy because we believe in the tremendous value of those powerful worlds - SAP and the cloud," explained Frank Stienhans, CTO of Ocean9, Inc., in this SYS-CON.tv interview at 20th Cloud Expo, held June 6-8, 2017, at the Javits Center in New York City, NY.
DX World EXPO, LLC., a Lighthouse Point, Florida-based startup trade show producer and the creator of "DXWorldEXPO® - Digital Transformation Conference & Expo" has announced its executive management team. The team is headed by Levent Selamoglu, who has been named CEO. "Now is the time for a truly global DX event, to bring together the leading minds from the technology world in a conversation about Digital Transformation," he said in making the announcement.
SYS-CON Events announced today that DXWorldExpo has been named “Global Sponsor” of SYS-CON's 21st International Cloud Expo, which will take place on Oct 31 – Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Digital Transformation is the key issue driving the global enterprise IT business. Digital Transformation is most prominent among Global 2000 enterprises and government institutions.
SYS-CON Events announced today that Datera, that offers a radically new data management architecture, has been named "Exhibitor" of SYS-CON's 21st International Cloud Expo ®, which will take place on Oct 31 - Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Datera is transforming the traditional datacenter model through modern cloud simplicity. The technology industry is at another major inflection point. The rise of mobile, the Internet of Things, data storage and Big...
"MobiDev is a Ukraine-based software development company. We do mobile development, and we're specialists in that. But we do full stack software development for entrepreneurs, for emerging companies, and for enterprise ventures," explained Alan Winters, U.S. Head of Business Development at MobiDev, in this SYS-CON.tv interview at 20th Cloud Expo, held June 6-8, 2017, at the Javits Center in New York City, NY.
While the focus and objectives of IoT initiatives are many and diverse, they all share a few common attributes, and one of those is the network. Commonly, that network includes the Internet, over which there isn't any real control for performance and availability. Or is there? The current state of the art for Big Data analytics, as applied to network telemetry, offers new opportunities for improving and assuring operational integrity. In his session at @ThingsExpo, Jim Frey, Vice President of S...
SYS-CON Events announced today that DXWorldExpo has been named “Global Sponsor” of SYS-CON's 21st International Cloud Expo, which will take place on Oct 31 – Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Digital Transformation is the key issue driving the global enterprise IT business. Digital Transformation is most prominent among Global 2000 enterprises and government institutions.
In his opening keynote at 20th Cloud Expo, Michael Maximilien, Research Scientist, Architect, and Engineer at IBM, discussed the full potential of the cloud and social data requires artificial intelligence. By mixing Cloud Foundry and the rich set of Watson services, IBM's Bluemix is the best cloud operating system for enterprises today, providing rapid development and deployment of applications that can take advantage of the rich catalog of Watson services to help drive insights from the vast t...
SYS-CON Events announced today that EnterpriseTech has been named “Media Sponsor” of SYS-CON's 21st International Cloud Expo, which will take place on Oct 31 – Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. EnterpriseTech is a professional resource for news and intelligence covering the migration of high-end technologies into the enterprise and business-IT industry, with a special focus on high-tech solutions in new product development, workload management, increased effic...
SYS-CON Events announced today that Massive Networks, that helps your business operate seamlessly with fast, reliable, and secure internet and network solutions, has been named "Exhibitor" of SYS-CON's 21st International Cloud Expo ®, which will take place on Oct 31 - Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. As a premier telecommunications provider, Massive Networks is headquartered out of Louisville, Colorado. With years of experience under their belt, their team of...
SYS-CON Events announced today that Cloud Academy named "Bronze Sponsor" of 21st International Cloud Expo which will take place October 31 - November 2, 2017 at the Santa Clara Convention Center in Santa Clara, CA. Cloud Academy is the industry’s most innovative, vendor-neutral cloud technology training platform. Cloud Academy provides continuous learning solutions for individuals and enterprise teams for Amazon Web Services, Microsoft Azure, Google Cloud Platform, and the most popular cloud com...
SYS-CON Events announced today that Cloudistics, an on-premises cloud computing company, has been named “Bronze Sponsor” of SYS-CON's 21st International Cloud Expo, which will take place on Oct 31 - Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Cloudistics delivers a complete public cloud experience with composable on-premises infrastructures to medium and large enterprises. Its software-defined technology natively converges network, storage, compute, virtualization, and ...
SYS-CON Events announced today that CHEETAH Training & Innovation will exhibit at SYS-CON's 21st International Cloud Expo®, which will take place on Oct. 31 – Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. CHEETAH Training & Innovation is a cloud consulting and IT training firm specializing in improving clients cloud strategies and infrastructures for medium to large companies.
SYS-CON Events announced today that Datanami has been named “Media Sponsor” of SYS-CON's 21st International Cloud Expo, which will take place on Oct 31 – Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Datanami is a communication channel dedicated to providing insight, analysis and up-to-the-minute information about emerging trends and solutions in Big Data. The publication sheds light on all cutting-edge technologies including networking, storage and applications, and thei...
The current age of digital transformation means that IT organizations must adapt their toolset to cover all digital experiences, beyond just the end users’. Today’s businesses can no longer focus solely on the digital interactions they manage with employees or customers; they must now contend with non-traditional factors. Whether it's the power of brand to make or break a company, the need to monitor across all locations 24/7, or the ability to proactively resolve issues, companies must adapt to...