Hadoop, the Data Lake, and a New World of Analytics

Size: px
Start display at page:

Download "Hadoop, the Data Lake, and a New World of Analytics"

Transcription

1 Hadoop, the Data Lake, and a New World of Analytics Hortonworks. We do Hadoop. Spring 2014 Version 1.0 Page 1 Hortonworks Inc. 2014

2 Traditional Data Architecture Pressured 2.8 ZB in % from New Data Types 15x Machine Data by ZB by 2020 Data source: IDC SOURCES OLTP, ERP, CRM Documents, s Web Logs, Click Streams Social Networks Machine Generated Sensor Data Geolocation Data Page 2 Hortonworks Inc. 2014

3 Modern Data Architecture with Hadoop DATA SYSTEMS APPLICATIONS ROOMS Statistical Analysis RDBMS EDW MPP Repositories BI / Reporting, Ad Hoc Analysis Interactive Web & Mobile Apps Governance & Integra.on Enterprise Applications ENTERPRISE HADOOP Data Access Data Management Security Opera.ons DEV & DATA TOOLS Build & Test OPERATIONS TOOLS Provision, Manage & Monitor SOURCES OLTP, ERP, CRM Documents, s Web Logs, Click Streams Social Networks Machine Generated Sensor Data Geolocation Data Page 3 Hortonworks Inc. 2014

4 YARN Transforms Hadoop s Architecture Mul.- Use Data Pla>orm Store all data in one place, process in many ways Batch Interac.ve Real-.me Streaming YARN : Data Opera.ng System 1 Store any/all raw data sources and processed data over extended periods of time. n Enables deep insight across a large, broad, diverse set of data at efficient scale Page 4 Hortonworks Inc. 2014

5 Unlock New Approach to Insight Current Approach Apply schema on write Dependent on IT Augment with Hadoop Apply schema on read Support range of access paierns to data stored in HDFS: polymorphic access SQL Single Query Engine Repeatable Linear Process Hadoop Mul(ple Query Engines Itera(ve Process: Explore, Transform, Analyze Determine list of ques.ons Design solu.ons Collect structured data Ask ques.ons from list Detect addi.onal ques.ons Batch Interac.ve Real-.me Streaming Page 5 Hortonworks Inc. 2014

6 Schema-on-Write and Schema-on-Read Standard Digital Camera Zoom & focus first Capture limited set of pixels Crop around the focused area Lytro Lightfield Camera Capture entire lightfield Infinite zoom & focus Crop any captured areas Page 6 Hortonworks Inc. 2014

7 Leverage Commodity Compute + Storage Cloud Storage HADOOP NAS Engineered System MPP Fully-loaded Cost Per Raw TB of Data (Min Max Cost) Hadoop Enables Scalable Compute & Storage at a Compelling Cost Structure SAN $0 $20,000 $40,000 $60,000 $80,000 $180,000 Source: Juergen Urbanski, Board Member Big Data & Analytics, BITKOM Page 7 Hortonworks Inc. 2014

8 The Common Journey with Hadoop Scale A Modern Data Architecture What s a Data Lake? RDBMS MPP EDW Governance & Integration Data Access Data Management Security Operations A modern data architecture that provides a shared service for broad insight across a large, diverse set of data at efficient scale New Analytic Apps New types of data LOB-driven Scope Page 8 Hortonworks Inc. 2014

9 Unlock Value in New Types of Data 1. Social Understand how people are feeling and interac(ng right now 2. Clickstream Capture and analyze website visitors data trails and op(mize your website 3. Sensor/Machine Discover paierns in data streaming from remote sensors and machines 4. Geographic Analyze loca(on- based data to manage opera(ons where they occur 5. Server Logs Diagnose process failures and prevent security breaches 6. Unstructured (txt, video, pictures, etc.) Understand paierns in files across millions of web pages, s, and documents Value + Online archive Data that was once purged or moved to tape can be stored in Hadoop to discover long term trends and previously hidden value Page 9 Hortonworks Inc. 2014

10 Business Applications of Hadoop Sensor Server Logs Text Social Geographi c Machine Clickstrea m Structured Unstructur ed Financial Services New Account Risk Screens Trading Risk Insurance Underwriting Telecom Call Detail Records (CDR) Infrastructure Investment Real-time Bandwidth Allocation Retail 360 View of the Customer Localized, Personalized Promotions Website Optimization Page 10 Hortonworks Inc. 2014

11 Business Applications of Hadoop Sensor Server Logs Text Social Geographi c Machine Clickstrea m Structured Unstructur ed Manufacturing Supply Chain and Logistics Assembly Line Quality Assurance Crowd-sourced Quality Assurance Healthcare Use Genomic Data in Medial Trials Monitor Patient Vitals in Real-Time Pharmaceuticals Recruit & Retain Patients for Drug Trials Improve Prescription Adherence Oil & Gas Unify Exploration & Production Data Government Monitor Rig Safety in Real-Time ETL Offload in Response to Budgetary Pressures Sentiment Analysis for Gov t Programs Page 11 Hortonworks Inc. 2014

12 UC Irvine Health does Hadoop UC Irvine Medical Center is ranked among the na(on's best hospitals by U.S. News & World Report for the 12th year More than 400 specialty and primary care physicians Opened in bed medical facility Migrated 22 years of patient data across admissions and clinical Predictive analytics to reduce patient re-admittance Real time monitoring for rapid response to changes in vital signs For healthcare, we have never had the ability to do this. We have always taken the approach that we think we know what data elements are important. Now with all the data, we let the data determine what is important for predictive analysis. Charles Boicey More details at: Page 12 Hortonworks Inc. 2014

13 Some of the Companies Telling Their Stories Page 13 Hortonworks Inc. 2014

14 Enabling Hadoop for the Enterprise Journey Scale A Modern Data Architecture RDBMS 1 Capabili.es Ensure enterprise capabili(es are delivered in 100% open source to benefit all MPP EDW Governance & Integration Data Access Data Management Security Operations 2 Integration Interoperable with existing data center investments New Analytic Apps New types of data LOB-driven 3 Skills Leverage your exis(ng skills: development, analy(cs, opera(ons Scope Page 14 Hortonworks Inc. 2014

15 Our Mission: Enable your Modern Data Architecture by delivering Enterprise Apache Hadoop Open Leadership Drive innovation in the open exclusively via the Apache community-driven open source process Enterprise Rigor Engineer, test and certify Apache Hadoop with the enterprise in mind Ecosystem Endorsement Focus on deep integration with existing data center technologies and skills Headquartered in Palo Alto, CA; 300+ employees and growing Reseller Partners: Page 15 Hortonworks Inc. 2014

16 A Traditional Approach Under Pressure APPLICATIONS Business Analy.cs Custom Applica.ons Packaged Applica.ons OLTP, ERP, CRM Systems Unstructured documents, s 2.8 ZB in 2012 Server logs DATA SYSTEM RDBMS EDW MPP REPOSITORIES 85% from New Data Types 15x Machine Data by 2020 Sen(ment, Web Data 40 ZB by 2020 Source: IDC Sensor. Machine Data SOURCES Exis.ng Sources (CRM, ERP, Clickstream, Logs) Clickstream Geoloca(on Page 16 Hortonworks Inc. 2014

17 An Emerging Modern Data Architecture APPLICATIONS Business Analy.cs Custom Applica.ons Packaged Applica.ons DEV & DATA TOOLS Build & Test DATA SYSTEM RDBMS EDW MPP REPOSITORIES Governance & Integration Data Access Data Management Security Operations OPERATIONS TOOLS Provision, Manage & Monitor SOURCES OLTP, ERP, Documents, Web Logs, CRM Systems s Click Streams Social Networks Machine Generated Sensor Data Geoloca(on Data Page 17 Hortonworks Inc. 2014

18 Drivers of Hadoop Adoption SCALE New Analytic Apps New types of data LOB-driven SCOPE Page 18 Hortonworks Inc. 2014

19 Hadoop Value: New types of data Sentiment Understand how your customers feel about your brand and products right now Clickstream Capture and analyze website visitors data trails and optimize your website Sensors Discover patterns in data streaming automatically from remote sensors and machines Geographic Analyze locationbased data to manage operations where they occur Server Logs Research logs to diagnose process failures and prevent security breaches Unstructured Understand patterns in files across millions of web pages, s, and documents Page 19 Hortonworks Inc. 2014

20 Net New Analytic Applications Are Everywhere $ Financial Services Retail Telecom Manufacturing New Account Risk Screens Fraud Prevention Trading Risk Maximize Deposit Spread Insurance Underwriting Accelerate Loan Processing 360 View of the Customer Analyze Brand Sentiment Localized, Personalized Promotions Website Optimization Optimal Store Layout Call Detail Records (CDRs) Infrastructure Investment Next Product to Buy (NPTB) Real-time Bandwidth Allocation New Product Development Supplier Consolidation Supply Chain and Logistics Assembly Line Quality Assurance Proactive Maintenance Crowdsourced Quality Assurance Healthcare Utilities, Oil & Gas Public Sector Genomic data for medical trials Monitor patient vitals Reduce re-admittance rates Smart meter stream analysis Slow oil well decline curves Optimize lease bidding Analyze public sentiment Protect critical networks Prevent fraud and waste Store medical research data Recruit cohorts for pharmaceutical trials Compliance reporting Proactive equipment repair Seismic image processing Crowdsource reporting for repairs to infrastructure Fulfill open records requests Page 20 Hortonworks Inc. 2014

21 Drivers of Hadoop Adoption SCALE A Modern Data Architecture/Data Lake RDBMS MPP EDW Governance & Integration Data Access Data Management Security Operations New Analytic Apps New types of data LOB-driven SCOPE Page 21 Hortonworks Inc. 2014

22 Unlock A New Approach to Insight Current Reality Apply schema on write Dependent on IT Augment w/ Hadoop Apply schema on read Support range of access patterns to data stored in HDFS: polymorphic access SQL Single Query Engine Repeatable Linear Process Hadoop Mul(ple Query Engines Itera(ve Process: Explore, Transform, Analyze Determine list of ques.ons Design solu.ons Collect structured data Ask ques.ons from list Detect addi.onal ques.ons Batch Interac.ve Real-.me Streaming Page 22 Hortonworks Inc. 2014

23 Architectural Drivers of the MDA EDW Optimization Commodity Compute & Storage OPERATIONS 50% ANALYTICS 20% ETL PROCESS 30% OPERATIONS 50% ANALYTICS 50% Cloud Storage HADOOP NAS Fully-loaded Cost Per Raw TB of Data (Min Max Cost) Current Reality EDW at capacity: some usage from low value workloads Older data archived, unavailable for ongoing exploration Source data often discarded Augment w/ Hadoop Free up EDW resources from low value tasks Keep 100% of source data and historical data for ongoing exploration Mine data for value after loading it because of schema-on-read Engineered System MPP SAN $0 $20,000 $40,000 $60,000 $80,000 $180,000 Hadoop Enables Scalable Compute & Storage at a Compelling Cost Structure Page 23 Hortonworks Inc. 2014

24 Delivering the Core Capabilities of Enterprise Hadoop Page 24 Hortonworks Inc. 2014

25 Core Capabilities of Enterprise Hadoop PRESENTATION & APPLICATION Enable both existing and new application to provide value to the organization ENTERPRISE MGMT & SECURITY Empower existing operations and security tools to manage Hadoop GOVERNANCE & INTEGRATION DATA ACCESS SECURITY OPERATIONS Load data and manage according to policy Access your data simultaneously in multiple ways (batch, interactive, real-time) Store and process all of your Corporate Data Assets Provide layered approach to security through Authentication, Authorization, Accounting, and Data Protection Deploy and effectively manage the platform DATA MANAGEMENT Provide deployment choice across physical, virtual, cloud DEPLOYMENT OPTIONS Page 25 Hortonworks Inc. 2014

26 Core Capabilities of Enterprise Hadoop GOVERNANCE & INTEGRATION DATA ACCESS SECURITY OPERATIONS Data Workflow, Lifecycle & Governance Falcon Sqoop Flume NFS WebHDFS Batch Map Reduce Script Pig SQL Hive/Tez, HCatalog NoSQL HBase Accumulo Stream Storm YARN : Data Opera.ng System Search Solr 1 HDFS (Hadoop Distributed File System) Others In- Memory Analy(cs, ISV engines Authen.ca.on Authoriza.on Accoun.ng Data Protec.on Storage: HDFS Resources: YARN Access: Hive, Pipeline: Falcon Cluster: Knox Provision, Manage & Monitor Ambari Zookeeper Scheduling Oozie N DATA MANAGEMENT Page 26 Hortonworks Inc. 2014

27 HDP Delivers Enterprise Hadoop HDP 2.1 Hortonworks Data Platform GOVERNANCE & INTEGRATION DATA ACCESS SECURITY OPERATIONS Data Workflow, Lifecycle & Governance Falcon Sqoop Flume NFS WebHDFS Batch Map Reduce Script Pig SQL Hive/Tez, HCatalog NoSQL HBase Accumulo Stream Storm YARN : Data Opera.ng System Search Solr 1 HDFS (Hadoop Distributed File System) Others In- Memory Analy(cs, ISV engines Authen.ca.on Authoriza.on Accoun.ng Data Protec.on Storage: HDFS Resources: YARN Access: Hive, Pipeline: Falcon Cluster: Knox Provision, Manage & Monitor Ambari Zookeeper Scheduling Oozie N DATA MANAGEMENT Deployment Choice Linux Windows On-Premise Cloud Page 27 Hortonworks Inc. 2014

28 Driving Innovation Through the Apache Community Apache Project Committers PMC Members Hadoop Tez 10 4 Hive 12 3 HBase 8 3 Pig 6 5 Sqoop 1 0 Ambari Knox 6 2 Falcon 4 2 Oozie 2 2 Zookeeper 2 1 Flume 1 0 Accumulo 2 2 Storm 1 0 TOTAL Total Net Lines Contributed to Apache Hadoop 449, ,933 End Users 614,041 3 LinkedIn 3 IBM 5 Facebook 10 Others 7 Cloudera Total Number of Committers to Apache Hadoop 10 Yahoo 21 Page 28 Hortonworks Inc. 2014

29 The Partners You Rely On, Rely On Hortonworks for Hadoop Page 29 Hortonworks Inc. 2014

30 Broad Ecosystem Integration APPLICATIONS DATA SYSTEM SOURCES BusinessObjects BI RDBMS EDW MPP HANA Exis.ng Sources (CRM, ERP, Clickstream, Logs) Governance & Integration HDP 2.1 Data Access Data Management Security Operations Emerging Sources (Sensor, Sen.ment, Geo, Unstructured) DEV & DATA TOOLS OPERATIONAL TOOLS INFRASTRUCTURE DEPTH Hortonworks engages in deep engineered relationships with the leaders in the data center, such as Microsoft, Teradata, Redhat & SAP BREADTH Hundreds of partners work with us to certify their applications to work with Hadoop so they can extend big data to their users Page 30 Hortonworks Inc. 2014

31 The Partners You Rely on, Rely on Hortonworks HDInsight & HDP for Windows Only Hadoop Distribution for Windows Azure & Windows Server Native integration with SQL Server, Excel, and System Center Extends Hadoop to.net community Teradata Portfolio for Hadoop Seamless data access between Teradata and Hadoop (SQL-H) Simple management & monitoring with Viewpoint integration Flexible deployment options Instant Access + Infinite Scale SAP can assure their customers they are deploying an SAP HANA + Hadoop architecture fully supported by SAP Enables analytics apps (BOBJ) to interact with Hadoop Complete Portfolio for Hadoop UDA Diagram Appliances Page 31 Hortonworks Inc. 2014

32 Hortonworks Data Platform Community Driven Enterprise Apache Hadoop Page 32 Hortonworks Inc. 2014

33 HDP 2.0: Enterprise Hadoop Platform HDP 2.1 Hortonworks Data Platform Hortonworks Data Platform (HDP) GOVERNANCE & INTEGRATION Data Workflow, Lifecycle & Governance Falcon Sqoop Flume NFS WebHDFS Batch Map Reduce Script Pig SQL Hive/Tez, HCatalog DATA ACCESS NoSQL HBase Accumulo Stream Storm YARN : Data Opera.ng System DATA MANAGEMENT Search Solr 1 HDFS (Hadoop Distributed File System) Others In- Memory Analy(cs, ISV engines N SECURITY Authen.ca.on Authoriza.on Accoun.ng Data Protec.on Storage: HDFS Resources: YARN Access: Hive, Pipeline: Falcon Cluster: Knox OPERATIONS Provision, Manage & Monitor Ambari Zookeeper Scheduling Oozie The Only Completely Open Distribution for Apache Hadoop Fundamentally Versatile and Comprehensive enterprise capabilities Wholly Integrated for deep ecosystem interoperability Deployment Choice Linux Windows On-Premise Cloud Page 33 Hortonworks Inc. 2014

34 HDP 2.1: Reliable, Consistent & Current HDP certifies most recent & stable community innovation HDP April HDP 2.0 October 2013 HDP 1.3 May * Hadoop &YARN Tez Pig Hive & HCatalog HBase Storm Mahout Solr Sqoop Flume Falcon Ambari Oozie Zookeeper Knox Data Management Data Access Governance & Integration Operations Security Hortonworks Data Platform Page 34 Hortonworks Inc. 2014

35 Hortonworks Process for Enterprise Hadoop Upstream Community Projects Apache HBase Apache Hive Downstream Enterprise Product Certified at scale using the most advanced Hadoop test bed on the planet 1000 s of production nodes at Yahoo! Over 1500 unit & system tests Apache Pig Apache Falcon Test & Patch Apache Hadoop Design & Develop Release Fixed Issues Design & Develop Integrate & Test HDP 2.1 Package & Certify Apache Knox Apache Storm Stable Project Releases Distribute Virtuous cycle when development & fixed issues done upstream & stable project releases flow downstream Page 35 Hortonworks Inc. 2014

36 Working in the Community for the Enterprise Hortonworks Reseller, System Integrator, Technology, and Training Partners Support Organization Implementation Team Engineering Team Hortonworks De-risks your investment All support and implementation backed by the largest and most experienced Hadoop team on the planet A Two Way Street Gain access to the expertise and knowledge of the Hadoop community All issues or feature requests represented in the community Apache Community Leadership Page 36 Hortonworks Inc. 2014

37 Flexible Support Subscription Programs Our support subscriptions provide unlimited support across development, pilot, staging, & production Resources Available Web /Phone Response Time 5 contacts 24 hours x 7 days Yes Severity One: 1 hour Severity Two: 4 hours All Other: 1 Bus Day Services Provided Customer Support Portal Advanced Knowledgebase Diagnosis of Install, Config & Cluster Mgmt Issues Access to Upgrades, Updates and Patches Diagnosis of Performance Issues Diagnosis of Data Loading, Processing &Query Issues Application Development Support Remote Troubleshooting What s Supported table for specific components Page 37 Hortonworks Inc. 2014

38 Transferring Hadoop Expertise Apache Hadoop Training & Certification World class training programs Designed to help you learn fast Role-based hands on classes with 50% lab time Hadoop Certification demonstrates expertise in Development & Administration Programs designed to transfer knowledge Industry leading Hadoop Sandbox Free download Fastest way to learn Apache Hadoop Personal, portable Hadoop environment Page 38 Hortonworks Inc. 2014

39 Who is Using Hadoop & Hortonworks? Page 39 Hortonworks Inc HORTONWORKS CONFIDENTIAL & PROPRIETARY INFORMATION

40 Hortonworks: A Leader The Forrester Wave : Big Data Hadoop Solutions, Q Hortonworks loves and lives open source innovation Vision & Execution for Enterprise Hadoop. Hortonworks leads with a strong strategy and roadmap for open source innovation with Hadoop and a strong delivery of that innovation in Hortonworks Data Platform. World Class Support and Services. Hortonworks' Customer Support received a maximum score and was significantly higher than both Cloudera and MapR. Key Strategic Partnerships. Hortonworks unique strategic partnerships with Microsoft, SAP, Teradata and others are a key strength as part of its overall strategy of ecosystem partnership to accelerate Hadoop adoption in the enterprise. Page 40 Hortonworks Inc. 2014

41 The Value of Open Connect With the Community Avoid Vendor Lock In Partnerships that Matter Certified for the Enterprise Support from the Experts We employ a large number of Apache project committers & innovators so that you are represented in the open source community Hortonworks Data Platform is close to the open source trunk as possible and is developed 100% in the open so you are never locked in We work with partners to deeply integrate Hadoop with data center technologies so you can leverage existing skills and investments We engineer, test and certify the Hortonworks Data Platform at scale to ensure reliability and stability you require for enterprise use We provide the highest quality of support for deploying at scale. You are supported by hundreds of years of Hadoop experience Page 41 Hortonworks Inc. 2014

HDP Enabling the Modern Data Architecture

HDP Enabling the Modern Data Architecture HDP Enabling the Modern Data Architecture Herb Cunitz President, Hortonworks Page 1 Hortonworks enables adoption of Apache Hadoop through HDP (Hortonworks Data Platform) Founded in 2011 Original 24 architects,

More information

HDP Hadoop From concept to deployment.

HDP Hadoop From concept to deployment. HDP Hadoop From concept to deployment. Ankur Gupta Senior Solutions Engineer Rackspace: Page 41 27 th Jan 2015 Where are you in your Hadoop Journey? A. Researching our options B. Currently evaluating some

More information

Upcoming Announcements

Upcoming Announcements Enterprise Hadoop Enterprise Hadoop Jeff Markham Technical Director, APAC jmarkham@hortonworks.com Page 1 Upcoming Announcements April 2 Hortonworks Platform 2.1 A continued focus on innovation within

More information

Big Data Realities Hadoop in the Enterprise Architecture

Big Data Realities Hadoop in the Enterprise Architecture Big Data Realities Hadoop in the Enterprise Architecture Paul Phillips Director, EMEA, Hortonworks pphillips@hortonworks.com +44 (0)777 444 3857 Hortonworks Inc. 2012 Page 1 Agenda The Growth of Enterprise

More information

A Modern Data Architecture with Apache Hadoop

A Modern Data Architecture with Apache Hadoop Modern Data Architecture with Apache Hadoop Talend Big Data Presented by Hortonworks and Talend Executive Summary Apache Hadoop didn t disrupt the datacenter, the data did. Shortly after Corporate IT functions

More information

Apache Hadoop's Role in Your Big Data Architecture

Apache Hadoop's Role in Your Big Data Architecture Apache Hadoop's Role in Your Big Data Architecture Chris Harris EMEA, Hortonworks charris@hortonworks.com Twi

More information

Hortonworks and ODP: Realizing the Future of Big Data, Now Manila, May 13, 2015

Hortonworks and ODP: Realizing the Future of Big Data, Now Manila, May 13, 2015 Hortonworks and ODP: Realizing the Future of Big Data, Now Manila, May 13, 2015 We Do Hadoop Fall 2014 Page 1 HDP delivers a comprehensive data management platform GOVERNANCE Hortonworks Data Platform

More information

Session 0202: Big Data in action with SAP HANA and Hadoop Platforms Prasad Illapani Product Management & Strategy (SAP HANA & Big Data) SAP Labs LLC,

Session 0202: Big Data in action with SAP HANA and Hadoop Platforms Prasad Illapani Product Management & Strategy (SAP HANA & Big Data) SAP Labs LLC, Session 0202: Big Data in action with SAP HANA and Hadoop Platforms Prasad Illapani Product Management & Strategy (SAP HANA & Big Data) SAP Labs LLC, Bellevue, WA Legal disclaimer The information in this

More information

Hortonworks & SAS. Analytics everywhere. Page 1. Hortonworks Inc. 2011 2014. All Rights Reserved

Hortonworks & SAS. Analytics everywhere. Page 1. Hortonworks Inc. 2011 2014. All Rights Reserved Hortonworks & SAS Analytics everywhere. Page 1 A change in focus. A shift in Advertising From mass branding A shift in Financial Services From Educated Investing A shift in Healthcare From mass treatment

More information

Modern Data Architecture for Predictive Analytics

Modern Data Architecture for Predictive Analytics Modern Data Architecture for Predictive Analytics David Smith VP Marketing and Community - Revolution Analytics John Kreisa VP Strategic Marketing- Hortonworks Hortonworks Inc. 2013 Page 1 Your Presenters

More information

Comprehensive Analytics on the Hortonworks Data Platform

Comprehensive Analytics on the Hortonworks Data Platform Comprehensive Analytics on the Hortonworks Data Platform We do Hadoop. Page 1 Page 2 Back to 2005 Page 3 Vertical Scaling Page 4 Vertical Scaling Page 5 Vertical Scaling Page 6 Horizontal Scaling Page

More information

The Future of Data Management

The Future of Data Management The Future of Data Management with Hadoop and the Enterprise Data Hub Amr Awadallah (@awadallah) Cofounder and CTO Cloudera Snapshot Founded 2008, by former employees of Employees Today ~ 800 World Class

More information

Modern Data Architecture for Retail with Apache Hadoop on Windows

Modern Data Architecture for Retail with Apache Hadoop on Windows 1 Modern Data Architecture for Retail with Apache Hadoop on Windows A Hortonworks and Microsoft White Paper JUNE 2014 2 Executive Summary Retailers have a long history of investing in data and analytics

More information

Modern Data Architecture for Financial Services with Apache Hadoop on Windows

Modern Data Architecture for Financial Services with Apache Hadoop on Windows 1 Modern Data Architecture for Financial Services with Apache Hadoop on Windows A Hortonworks and Microsoft White Paper JUNE 2014 2 Executive Summary Financial services firms have long been dependent on

More information

The Future of Data Management with Hadoop and the Enterprise Data Hub

The Future of Data Management with Hadoop and the Enterprise Data Hub The Future of Data Management with Hadoop and the Enterprise Data Hub Amr Awadallah Cofounder & CTO, Cloudera, Inc. Twitter: @awadallah 1 2 Cloudera Snapshot Founded 2008, by former employees of Employees

More information

Data Security in Hadoop

Data Security in Hadoop Data Security in Hadoop Eric Mizell Director, Solution Engineering Page 1 What is Data Security? Data Security for Hadoop allows you to administer a singular policy for authentication of users, authorize

More information

SAP and Hortonworks Reference Architecture

SAP and Hortonworks Reference Architecture SAP and Hortonworks Reference Architecture Hortonworks. We Do Hadoop. June Page 1 2014 Hortonworks Inc. 2011 2014. All Rights Reserved A Modern Data Architecture With SAP DATA SYSTEMS APPLICATIO NS Statistical

More information

Big Data: Making Sense of it all!

Big Data: Making Sense of it all! Big Data: Making Sense of it all! Jamie Engesser E-mail : jamie@hortonworks.com Page 1 Data Driven Business? Facts not Intuition! Data driven decisions are better decisions its as simple as that. Using

More information

Modernizing Your Data Warehouse for Hadoop

Modernizing Your Data Warehouse for Hadoop Modernizing Your Data Warehouse for Hadoop Big data. Small data. All data. Audie Wright, DW & Big Data Specialist Audie.Wright@Microsoft.com O 425-538-0044, C 303-324-2860 Unlock Insights on Any Data Taking

More information

Hortonworks Data Platform for Hadoop and SAP HANA

Hortonworks Data Platform for Hadoop and SAP HANA Hortonworks Data Platform for Hadoop and SAP HANA Prasad illapani, Big Data & SAP HANA- Product Management & Strategy SAP Labs LLC., Bellevue, WA Bob Page, VP Partner Products, Hortonworks Inc. Palo Alto,

More information

A Modern Data Architecture with Apache Hadoop

A Modern Data Architecture with Apache Hadoop A Modern Data Architecture with Apache Hadoop A Hortonworks White Paper March 2014 Executive Summary Apache Hadoop didn t disrupt the datacenter, the data did. Shortly after Corporate IT functions within

More information

#TalendSandbox for Big Data

#TalendSandbox for Big Data Evalua&on von Apache Hadoop mit der #TalendSandbox for Big Data Julien Clarysse @whatdoesdatado @talend 2015 Talend Inc. 1 Connecting the Data-Driven Enterprise 2 Talend Overview Founded in 2006 BRAND

More information

Bringing Big Data to People

Bringing Big Data to People Bringing Big Data to People Microsoft s modern data platform SQL Server 2014 Analytics Platform System Microsoft Azure HDInsight Data Platform Everyone should have access to the data they need. Process

More information

Data Services Advisory

Data Services Advisory Data Services Advisory Modern Datastores An Introduction Created by: Strategy and Transformation Services Modified Date: 8/27/2014 Classification: DRAFT SAFE HARBOR STATEMENT This presentation contains

More information

Hadoop s Advantages for! Machine! Learning and. Predictive! Analytics. Webinar will begin shortly. Presented by Hortonworks & Zementis

Hadoop s Advantages for! Machine! Learning and. Predictive! Analytics. Webinar will begin shortly. Presented by Hortonworks & Zementis Webinar will begin shortly Hadoop s Advantages for Machine Learning and Predictive Analytics Presented by Hortonworks & Zementis September 10, 2014 Copyright 2014 Zementis, Inc. All rights reserved. 2

More information

Evolution from Big Data to Smart Data

Evolution from Big Data to Smart Data Evolution from Big Data to Smart Data Information is Exploding 120 HOURS VIDEO UPLOADED TO YOUTUBE 50,000 APPS DOWNLOADED 204 MILLION E-MAILS EVERY MINUTE EVERY DAY Intel Corporation 2015 The Data is Changing

More information

GAIN BETTER INSIGHT FROM BIG DATA USING JBOSS DATA VIRTUALIZATION

GAIN BETTER INSIGHT FROM BIG DATA USING JBOSS DATA VIRTUALIZATION GAIN BETTER INSIGHT FROM BIG DATA USING JBOSS DATA VIRTUALIZATION Syed Rasheed Solution Manager Red Hat Corp. Kenny Peeples Technical Manager Red Hat Corp. Kimberly Palko Product Manager Red Hat Corp.

More information

Ganzheitliches Datenmanagement

Ganzheitliches Datenmanagement Ganzheitliches Datenmanagement für Hadoop Michael Kohs, Senior Sales Consultant @mikchaos The Problem with Big Data Projects in 2016 Relational, Mainframe Documents and Emails Data Modeler Data Scientist

More information

Big Data Analytics. with EMC Greenplum and Hadoop. Big Data Analytics. Ofir Manor Pre Sales Technical Architect EMC Greenplum

Big Data Analytics. with EMC Greenplum and Hadoop. Big Data Analytics. Ofir Manor Pre Sales Technical Architect EMC Greenplum Big Data Analytics with EMC Greenplum and Hadoop Big Data Analytics with EMC Greenplum and Hadoop Ofir Manor Pre Sales Technical Architect EMC Greenplum 1 Big Data and the Data Warehouse Potential All

More information

Harnessing big data with Hortonworks Data Platform and Red Hat JBoss Data Virtualization

Harnessing big data with Hortonworks Data Platform and Red Hat JBoss Data Virtualization Harnessing big data with Hortonworks Data Platform and Red Hat JBoss Data Virtualization Kimberly Palko, Product Manager Red Hat JBoss Doug Reid, Director Partner Product Management Hortonworks Cojan van

More information

The Digital Enterprise Demands a Modern Integration Approach. Nada daveiga, Sr. Dir. of Technical Sales Tony LaVasseur, Territory Leader

The Digital Enterprise Demands a Modern Integration Approach. Nada daveiga, Sr. Dir. of Technical Sales Tony LaVasseur, Territory Leader The Digital Enterprise Demands a Modern Integration Approach Nada daveiga, Sr. Dir. of Technical Sales Tony LaVasseur, Territory Leader Yesterday s approach to data and application integration is a barrier

More information

Talend Big Data. Delivering instant value from all your data. Talend 2014 1

Talend Big Data. Delivering instant value from all your data. Talend 2014 1 Talend Big Data Delivering instant value from all your data Talend 2014 1 I may say that this is the greatest factor: the way in which the expedition is equipped. Roald Amundsen race to the south pole,

More information

SOLVING REAL AND BIG (DATA) PROBLEMS USING HADOOP. Eva Andreasson Cloudera

SOLVING REAL AND BIG (DATA) PROBLEMS USING HADOOP. Eva Andreasson Cloudera SOLVING REAL AND BIG (DATA) PROBLEMS USING HADOOP Eva Andreasson Cloudera Most FAQ: Super-Quick Overview! The Apache Hadoop Ecosystem a Zoo! Oozie ZooKeeper Hue Impala Solr Hive Pig Mahout HBase MapReduce

More information

Microsoft Big Data. Solution Brief

Microsoft Big Data. Solution Brief Microsoft Big Data Solution Brief Contents Introduction... 2 The Microsoft Big Data Solution... 3 Key Benefits... 3 Immersive Insight, Wherever You Are... 3 Connecting with the World s Data... 3 Any Data,

More information

Hortonworks CISC Innovation day

Hortonworks CISC Innovation day Hortonworks CISC Innovation day Simon gregory sgregory@hortonworks.com Here was the ask Hortonworks' data reposition - how this works and the types of data you work with. 1: Data Types & Value. What have

More information

Why Spark on Hadoop Matters

Why Spark on Hadoop Matters Why Spark on Hadoop Matters MC Srivas, CTO and Founder, MapR Technologies Apache Spark Summit - July 1, 2014 1 MapR Overview Top Ranked Exponential Growth 500+ Customers Cloud Leaders 3X bookings Q1 13

More information

BIG DATA: FROM HYPE TO REALITY. Leandro Ruiz Presales Partner for C&LA Teradata

BIG DATA: FROM HYPE TO REALITY. Leandro Ruiz Presales Partner for C&LA Teradata BIG DATA: FROM HYPE TO REALITY Leandro Ruiz Presales Partner for C&LA Teradata Evolution in The Use of Information Action s ACTIVATING MAKE it happen! Insights OPERATIONALIZING WHAT IS happening now? PREDICTING

More information

Hortonworks Data Platform. Buyer s Guide

Hortonworks Data Platform. Buyer s Guide Hortonworks Data Platform Buyer s Guide Hortonworks Data Platform (HDP Completely Open and Versatile Hadoop Data Platform 2 2014 Hortonworks, Inc. All rights reserved. Hadoop and the Hadoop elephant logo

More information

Extend your analytic capabilities with SAP Predictive Analysis

Extend your analytic capabilities with SAP Predictive Analysis September 9 11, 2013 Anaheim, California Extend your analytic capabilities with SAP Predictive Analysis Charles Gadalla Learning Points Advanced analytics strategy at SAP Simplifying predictive analytics

More information

Apache Hadoop: The Big Data Refinery

Apache Hadoop: The Big Data Refinery Architecting the Future of Big Data Whitepaper Apache Hadoop: The Big Data Refinery Introduction Big data has become an extremely popular term, due to the well-documented explosion in the amount of data

More information

Big Data & QlikView. Democratizing Big Data Analytics. David Freriks Principal Solution Architect

Big Data & QlikView. Democratizing Big Data Analytics. David Freriks Principal Solution Architect Big Data & QlikView Democratizing Big Data Analytics David Freriks Principal Solution Architect TDWI Vancouver Agenda What really is Big Data? How do we separate hype from reality? How does that relate

More information

Apache Hadoop Patterns of Use

Apache Hadoop Patterns of Use Community Driven Apache Hadoop Apache Hadoop Patterns of Use April 2013 2013 Hortonworks Inc. http://www.hortonworks.com Big Data: Apache Hadoop Use Distilled There certainly is no shortage of hype when

More information

Data Lake In Action: Real-time, Closed Looped Analytics On Hadoop

Data Lake In Action: Real-time, Closed Looped Analytics On Hadoop 1 Data Lake In Action: Real-time, Closed Looped Analytics On Hadoop 2 Pivotal s Full Approach It s More Than Just Hadoop Pivotal Data Labs 3 Why Pivotal Exists First Movers Solve the Big Data Utility Gap

More information

Cloudera Enterprise Data Hub in Telecom:

Cloudera Enterprise Data Hub in Telecom: Cloudera Enterprise Data Hub in Telecom: Three Customer Case Studies Version: 103 Table of Contents Introduction 3 Cloudera Enterprise Data Hub for Telcos 4 Cloudera Enterprise Data Hub in Telecom: Customer

More information

Datenverwaltung im Wandel - Building an Enterprise Data Hub with

Datenverwaltung im Wandel - Building an Enterprise Data Hub with Datenverwaltung im Wandel - Building an Enterprise Data Hub with Cloudera Bernard Doering Regional Director, Central EMEA, Cloudera Cloudera Your Hadoop Experts Founded 2008, by former employees of Employees

More information

The Enterprise Data Hub and The Modern Information Architecture

The Enterprise Data Hub and The Modern Information Architecture The Enterprise Data Hub and The Modern Information Architecture Dr. Amr Awadallah CTO & Co-Founder, Cloudera Twitter: @awadallah 1 2013 Cloudera, Inc. All rights reserved. Cloudera Overview The Leader

More information

Data Governance in the Hadoop Data Lake. Michael Lang May 2015

Data Governance in the Hadoop Data Lake. Michael Lang May 2015 Data Governance in the Hadoop Data Lake Michael Lang May 2015 Introduction Product Manager for Teradata Loom Joined Teradata as part of acquisition of Revelytix, original developer of Loom VP of Sales

More information

Hadoop Introduction. Olivier Renault Solution Engineer - Hortonworks

Hadoop Introduction. Olivier Renault Solution Engineer - Hortonworks Hadoop Introduction Olivier Renault Solution Engineer - Hortonworks Hortonworks A Brief History of Apache Hadoop Apache Project Established Yahoo! begins to Operate at scale Hortonworks Data Platform 2013

More information

THE JOURNEY TO A DATA LAKE

THE JOURNEY TO A DATA LAKE THE JOURNEY TO A DATA LAKE 1 THE JOURNEY TO A DATA LAKE 85% OF DATA GROWTH BY 2020 WILL COME FROM NEW TYPES OF DATA ACCORDING TO IDC, AS MUCH AS 85% OF DATA GROWTH BY 2020 WILL COME FROM NEW TYPES OF DATA,

More information

Investor Presentation. Second Quarter 2015

Investor Presentation. Second Quarter 2015 Investor Presentation Second Quarter 2015 Note to Investors Certain non-gaap financial information regarding operating results may be discussed during this presentation. Reconciliations of the differences

More information

Cisco IT Hadoop Journey

Cisco IT Hadoop Journey Cisco IT Hadoop Journey Srini Desikan, Program Manager IT 2015 MapR Technologies 1 Agenda Hadoop Platform Timeline Key Decisions / Lessons Learnt Data Lake Hadoop s place in IT Data Platforms Use Cases

More information

Please give me your feedback

Please give me your feedback Please give me your feedback Session BB4089 Speaker Claude Lorenson, Ph. D and Wendy Harms Use the mobile app to complete a session survey 1. Access My schedule 2. Click on this session 3. Go to Rate &

More information

Hadoop in the Hybrid Cloud

Hadoop in the Hybrid Cloud Presented by Hortonworks and Microsoft Introduction An increasing number of enterprises are either currently using or are planning to use cloud deployment models to expand their IT infrastructure. Big

More information

BIG DATA TRENDS AND TECHNOLOGIES

BIG DATA TRENDS AND TECHNOLOGIES BIG DATA TRENDS AND TECHNOLOGIES THE WORLD OF DATA IS CHANGING Cloud WHAT IS BIG DATA? Big data are datasets that grow so large that they become awkward to work with using onhand database management tools.

More information

VIEWPOINT. High Performance Analytics. Industry Context and Trends

VIEWPOINT. High Performance Analytics. Industry Context and Trends VIEWPOINT High Performance Analytics Industry Context and Trends In the digital age of social media and connected devices, enterprises have a plethora of data that they can mine, to discover hidden correlations

More information

HADOOP. Revised 10/19/2015

HADOOP. Revised 10/19/2015 HADOOP Revised 10/19/2015 This Page Intentionally Left Blank Table of Contents Hortonworks HDP Developer: Java... 1 Hortonworks HDP Developer: Apache Pig and Hive... 2 Hortonworks HDP Developer: Windows...

More information

Intel HPC Distribution for Apache Hadoop* Software including Intel Enterprise Edition for Lustre* Software. SC13, November, 2013

Intel HPC Distribution for Apache Hadoop* Software including Intel Enterprise Edition for Lustre* Software. SC13, November, 2013 Intel HPC Distribution for Apache Hadoop* Software including Intel Enterprise Edition for Lustre* Software SC13, November, 2013 Agenda Abstract Opportunity: HPC Adoption of Big Data Analytics on Apache

More information

Roadmap Talend : découvrez les futures fonctionnalités de Talend

Roadmap Talend : découvrez les futures fonctionnalités de Talend Roadmap Talend : découvrez les futures fonctionnalités de Talend Cédric Carbone Talend Connect 9 octobre 2014 Talend 2014 1 Connecting the Data-Driven Enterprise Talend 2014 2 Agenda Agenda Why a Unified

More information

Information Builders Mission & Value Proposition

Information Builders Mission & Value Proposition Value 10/06/2015 2015 MapR Technologies 2015 MapR Technologies 1 Information Builders Mission & Value Proposition Economies of Scale & Increasing Returns (Note: Not to be confused with diminishing returns

More information

Apache Hadoop in the Enterprise. Dr. Amr Awadallah, CTO/Founder @awadallah, aaa@cloudera.com

Apache Hadoop in the Enterprise. Dr. Amr Awadallah, CTO/Founder @awadallah, aaa@cloudera.com Apache Hadoop in the Enterprise Dr. Amr Awadallah, CTO/Founder @awadallah, aaa@cloudera.com Cloudera The Leader in Big Data Management Powered by Apache Hadoop The Leading Open Source Distribution of Apache

More information

Community Driven Apache Hadoop. Apache Hadoop Basics. May 2013. 2013 Hortonworks Inc. http://www.hortonworks.com

Community Driven Apache Hadoop. Apache Hadoop Basics. May 2013. 2013 Hortonworks Inc. http://www.hortonworks.com Community Driven Apache Hadoop Apache Hadoop Basics May 2013 2013 Hortonworks Inc. http://www.hortonworks.com Big Data A big shift is occurring. Today, the enterprise collects more data than ever before,

More information

Cisco IT Hadoop Journey

Cisco IT Hadoop Journey Cisco IT Hadoop Journey Alex Garbarini, IT Engineer, Cisco 2015 MapR Technologies 1 Agenda Hadoop Platform Timeline Key Decisions / Lessons Learnt Data Lake Hadoop s place in IT Data Platforms Use Cases

More information

INVESTOR PRESENTATION. First Quarter 2014

INVESTOR PRESENTATION. First Quarter 2014 INVESTOR PRESENTATION First Quarter 2014 Note to Investors Certain non-gaap financial information regarding operating results may be discussed during this presentation. Reconciliations of the differences

More information

Artur Borycki. Director International Solutions Marketing

Artur Borycki. Director International Solutions Marketing Artur Borycki Director International Solutions Agenda! Evolution of Teradata s Unified Architecture Analytical and Workloads! Teradata s Reference Information Architecture Evolution of Teradata s" Unified

More information

SQL Server 2012 PDW. Ryan Simpson Technical Solution Professional PDW Microsoft. Microsoft SQL Server 2012 Parallel Data Warehouse

SQL Server 2012 PDW. Ryan Simpson Technical Solution Professional PDW Microsoft. Microsoft SQL Server 2012 Parallel Data Warehouse SQL Server 2012 PDW Ryan Simpson Technical Solution Professional PDW Microsoft Microsoft SQL Server 2012 Parallel Data Warehouse Massively Parallel Processing Platform Delivers Big Data HDFS Delivers Scale

More information

Collaborative Big Data Analytics. Copyright 2012 EMC Corporation. All rights reserved.

Collaborative Big Data Analytics. Copyright 2012 EMC Corporation. All rights reserved. Collaborative Big Data Analytics 1 Big Data Is Less About Size, And More About Freedom TechCrunch!!!!!!!!! Total data: bigger than big data 451 Group Findings: Big Data Is More Extreme Than Volume Gartner!!!!!!!!!!!!!!!

More information

Stinger Initiative: Introduction

Stinger Initiative: Introduction Stinger Initiative: Introduction Interactive Query on Hadoop Chris Harris E-Mail : charris@hortonworks.com Twitter : cj_harris5 Page 1 The World of Data is Changing Data Explosion 1 Zettabyte (ZB) = 1

More information

Aligning Your Strategic Initiatives with a Realistic Big Data Analytics Roadmap

Aligning Your Strategic Initiatives with a Realistic Big Data Analytics Roadmap Aligning Your Strategic Initiatives with a Realistic Big Data Analytics Roadmap 3 key strategic advantages, and a realistic roadmap for what you really need, and when 2012, Cognizant Topics to be discussed

More information

TE's Analytics on Hadoop and SAP HANA Using SAP Vora

TE's Analytics on Hadoop and SAP HANA Using SAP Vora TE's Analytics on Hadoop and SAP HANA Using SAP Vora Naveen Narra Senior Manager TE Connectivity Santha Kumar Rajendran Enterprise Data Architect TE Balaji Krishna - Director, SAP HANA Product Mgmt. -

More information

BIG DATA AND MICROSOFT. Susie Adams CTO Microsoft Federal

BIG DATA AND MICROSOFT. Susie Adams CTO Microsoft Federal BIG DATA AND MICROSOFT Susie Adams CTO Microsoft Federal THE WORLD OF DATA IS CHANGING Cloud What s making this possible? Electrical efficiency of computers doubles every year and ½. Laptops and mobile

More information

Big Data, Why All the Buzz? (Abridged) Anita Luthra, February 20, 2014

Big Data, Why All the Buzz? (Abridged) Anita Luthra, February 20, 2014 Big Data, Why All the Buzz? (Abridged) Anita Luthra, February 20, 2014 Defining Big Not Just Massive Data Big data refers to data sets whose size is beyond the ability of typical database software tools

More information

Hadoop Beyond Hype: Complex Adaptive Systems Conference Nov 16, 2012. Viswa Sharma Solutions Architect Tata Consultancy Services

Hadoop Beyond Hype: Complex Adaptive Systems Conference Nov 16, 2012. Viswa Sharma Solutions Architect Tata Consultancy Services Hadoop Beyond Hype: Complex Adaptive Systems Conference Nov 16, 2012 Viswa Sharma Solutions Architect Tata Consultancy Services 1 Agenda What is Hadoop Why Hadoop? The Net Generation is here Sizing the

More information

Dominik Wagenknecht Accenture

Dominik Wagenknecht Accenture Dominik Wagenknecht Accenture Improving Mainframe Performance with Hadoop October 17, 2014 Organizers General Partner Top Media Partner Media Partner Supporters About me Dominik Wagenknecht Accenture Vienna

More information

OPEN MODERN DATA ARCHITECTURE FOR FINANCIAL SERVICES RISK MANAGEMENT

OPEN MODERN DATA ARCHITECTURE FOR FINANCIAL SERVICES RISK MANAGEMENT WHITEPAPER OPEN MODERN DATA ARCHITECTURE FOR FINANCIAL SERVICES RISK MANAGEMENT A top-tier global bank s end-of-day risk analysis jobs didn t complete in time for the next start of trading day. To solve

More information

Hadoop and Data Warehouse Friends, Enemies or Profiteers? What about Real Time?

Hadoop and Data Warehouse Friends, Enemies or Profiteers? What about Real Time? Hadoop and Data Warehouse Friends, Enemies or Profiteers? What about Real Time? Kai Wähner kwaehner@tibco.com @KaiWaehner www.kai-waehner.de Disclaimer! These opinions are my own and do not necessarily

More information

INVESTOR PRESENTATION. Third Quarter 2014

INVESTOR PRESENTATION. Third Quarter 2014 INVESTOR PRESENTATION Third Quarter 2014 Note to Investors Certain non-gaap financial information regarding operating results may be discussed during this presentation. Reconciliations of the differences

More information

WWW.WIPRO.COM HADOOP VENDOR DISTRIBUTIONS THE WHY, THE WHO AND THE HOW? Guruprasad K.N. Enterprise Architect Wipro BOTWORKS

WWW.WIPRO.COM HADOOP VENDOR DISTRIBUTIONS THE WHY, THE WHO AND THE HOW? Guruprasad K.N. Enterprise Architect Wipro BOTWORKS WWW.WIPRO.COM HADOOP VENDOR DISTRIBUTIONS THE WHY, THE WHO AND THE HOW? Guruprasad K.N. Enterprise Architect Wipro BOTWORKS Table of contents 01 Abstract 01 02 03 04 The Why - Need for The Who - Prominent

More information

Next Gen Hadoop Gather around the campfire and I will tell you a good YARN

Next Gen Hadoop Gather around the campfire and I will tell you a good YARN Next Gen Hadoop Gather around the campfire and I will tell you a good YARN Akmal B. Chaudhri* Hortonworks *about.me/akmalchaudhri My background ~25 years experience in IT Developer (Reuters) Academic (City

More information

BIG DATA TECHNOLOGY. Hadoop Ecosystem

BIG DATA TECHNOLOGY. Hadoop Ecosystem BIG DATA TECHNOLOGY Hadoop Ecosystem Agenda Background What is Big Data Solution Objective Introduction to Hadoop Hadoop Ecosystem Hybrid EDW Model Predictive Analysis using Hadoop Conclusion What is Big

More information

A Tour of the Zoo the Hadoop Ecosystem Prafulla Wani

A Tour of the Zoo the Hadoop Ecosystem Prafulla Wani A Tour of the Zoo the Hadoop Ecosystem Prafulla Wani Technical Architect - Big Data Syntel Agenda Welcome to the Zoo! Evolution Timeline Traditional BI/DW Architecture Where Hadoop Fits In 2 Welcome to

More information

End to End Solution to Accelerate Data Warehouse Optimization. Franco Flore Alliance Sales Director - APJ

End to End Solution to Accelerate Data Warehouse Optimization. Franco Flore Alliance Sales Director - APJ End to End Solution to Accelerate Data Warehouse Optimization Franco Flore Alliance Sales Director - APJ Big Data Is Driving Key Business Initiatives Increase profitability, innovation, customer satisfaction,

More information

Getting Started Practical Input For Your Roadmap

Getting Started Practical Input For Your Roadmap Getting Started Practical Input For Your Roadmap Mike Ferguson Managing Director, Intelligent Business Strategies BA4ALL Big Data & Analytics Insight Conference Stockholm, May 2015 About Mike Ferguson

More information

Self-service BI for big data applications using Apache Drill

Self-service BI for big data applications using Apache Drill Self-service BI for big data applications using Apache Drill 2015 MapR Technologies 2015 MapR Technologies 1 Management - MCS MapR Data Platform for Hadoop and NoSQL APACHE HADOOP AND OSS ECOSYSTEM Batch

More information

How to Hadoop Without the Worry: Protecting Big Data at Scale

How to Hadoop Without the Worry: Protecting Big Data at Scale How to Hadoop Without the Worry: Protecting Big Data at Scale SESSION ID: CDS-W06 Davi Ottenheimer Senior Director of Trust EMC Corporation @daviottenheimer Big Data Trust. Redefined Transparency Relevance

More information

Driving Growth in Insurance With a Big Data Architecture

Driving Growth in Insurance With a Big Data Architecture Driving Growth in Insurance With a Big Data Architecture The SAS and Cloudera Advantage Version: 103 Table of Contents Overview 3 Current Data Challenges for Insurers 3 Unlocking the Power of Big Data

More information

Using Tableau Software with Hortonworks Data Platform

Using Tableau Software with Hortonworks Data Platform Using Tableau Software with Hortonworks Data Platform September 2013 2013 Hortonworks Inc. http:// Modern businesses need to manage vast amounts of data, and in many cases they have accumulated this data

More information

QUEST meeting Big Data Analytics

QUEST meeting Big Data Analytics QUEST meeting Big Data Analytics Peter Hughes Business Solutions Consultant SAS Australia/New Zealand Copyright 2015, SAS Institute Inc. All rights reserved. Big Data Analytics WHERE WE ARE NOW 2005 2007

More information

Open Source in Financial Services: Meet the challenges of new business models and disruption

Open Source in Financial Services: Meet the challenges of new business models and disruption Open Source in Financial Services: Meet the challenges of new business models and disruption Speakers Vamsi Chemitiganti, General Manager Financial Services, Hortonworks Josh West, Senior Solutions Architect,

More information

Hadoop Ecosystem Overview. CMSC 491 Hadoop-Based Distributed Computing Spring 2015 Adam Shook

Hadoop Ecosystem Overview. CMSC 491 Hadoop-Based Distributed Computing Spring 2015 Adam Shook Hadoop Ecosystem Overview CMSC 491 Hadoop-Based Distributed Computing Spring 2015 Adam Shook Agenda Introduce Hadoop projects to prepare you for your group work Intimate detail will be provided in future

More information

Self-service BI for big data applications using Apache Drill

Self-service BI for big data applications using Apache Drill Self-service BI for big data applications using Apache Drill 2015 MapR Technologies 2015 MapR Technologies 1 Data Is Doubling Every Two Years Unstructured data will account for more than 80% of the data

More information

Build Your Competitive Edge in Big Data with Cisco. Rick Speyer Senior Global Marketing Manager Big Data Cisco Systems 6/25/2015

Build Your Competitive Edge in Big Data with Cisco. Rick Speyer Senior Global Marketing Manager Big Data Cisco Systems 6/25/2015 Build Your Competitive Edge in Big Data with Cisco Rick Speyer Senior Global Marketing Manager Big Data Cisco Systems 6/25/2015 Big Data Trends Increasingly Everything will be Connected to Everything Massive

More information

Impact of Big Data in Oil & Gas Industry. Pranaya Sangvai Reliance Industries Limited 04 Feb 15, DEJ, Mumbai, India.

Impact of Big Data in Oil & Gas Industry. Pranaya Sangvai Reliance Industries Limited 04 Feb 15, DEJ, Mumbai, India. Impact of Big Data in Oil & Gas Industry Pranaya Sangvai Reliance Industries Limited 04 Feb 15, DEJ, Mumbai, India. New Age Information 2.92 billions Internet Users in 2014 Twitter processes 7 terabytes

More information

Apache Hadoop: The Pla/orm for Big Data. Amr Awadallah CTO, Founder, Cloudera, Inc. aaa@cloudera.com, twicer: @awadallah

Apache Hadoop: The Pla/orm for Big Data. Amr Awadallah CTO, Founder, Cloudera, Inc. aaa@cloudera.com, twicer: @awadallah Apache Hadoop: The Pla/orm for Big Data Amr Awadallah CTO, Founder, Cloudera, Inc. aaa@cloudera.com, twicer: @awadallah 1 The Problems with Current Data Systems BI Reports + Interac7ve Apps RDBMS (aggregated

More information

Bringing the Power of SAS to Hadoop. White Paper

Bringing the Power of SAS to Hadoop. White Paper White Paper Bringing the Power of SAS to Hadoop Combine SAS World-Class Analytic Strength with Hadoop s Low-Cost, Distributed Data Storage to Uncover Hidden Opportunities Contents Introduction... 1 What

More information

Transactions & Interactions

Transactions & Interactions Transactions & Interactions The Correlation of Structured and Unstructured Data Shaun Connolly, Hortonworks December 15, 2011 Big Data Has Reached Every Market Digital data is personal, everywhere, increasingly

More information

Big Data and Industrial Internet

Big Data and Industrial Internet Big Data and Industrial Internet Keijo Heljanko Department of Computer Science and Helsinki Institute for Information Technology HIIT School of Science, Aalto University keijo.heljanko@aalto.fi 16.6-2015

More information

Architecting for Big Data Analytics and Beyond: A New Framework for Business Intelligence and Data Warehousing

Architecting for Big Data Analytics and Beyond: A New Framework for Business Intelligence and Data Warehousing Architecting for Big Data Analytics and Beyond: A New Framework for Business Intelligence and Data Warehousing Wayne W. Eckerson Director of Research, TechTarget Founder, BI Leadership Forum Business Analytics

More information

Il mondo dei DB Cambia : Tecnologie e opportunita`

Il mondo dei DB Cambia : Tecnologie e opportunita` Il mondo dei DB Cambia : Tecnologie e opportunita` Giorgio Raico Pre-Sales Consultant Hewlett-Packard Italiana 2011 Hewlett-Packard Development Company, L.P. The information contained herein is subject

More information

BIG DATA AND THE ENTERPRISE DATA WAREHOUSE WORKSHOP

BIG DATA AND THE ENTERPRISE DATA WAREHOUSE WORKSHOP BIG DATA AND THE ENTERPRISE DATA WAREHOUSE WORKSHOP Business Analytics for All Amsterdam - 2015 Value of Big Data is Being Recognized Executives beginning to see the path from data insights to revenue

More information

ENABLING GLOBAL HADOOP WITH EMC ELASTIC CLOUD STORAGE

ENABLING GLOBAL HADOOP WITH EMC ELASTIC CLOUD STORAGE ENABLING GLOBAL HADOOP WITH EMC ELASTIC CLOUD STORAGE Hadoop Storage-as-a-Service ABSTRACT This White Paper illustrates how EMC Elastic Cloud Storage (ECS ) can be used to streamline the Hadoop data analytics

More information