Welcome!

Eclipse Authors: Pat Romanski, Elizabeth White, Liz McMillan, David H Deans, JP Morgenthal

Blog Feed Post

The Enterprise Data Hub: A place to store all your data with enterprise grade data management, integrations

By

Hadoop World/Strata has been full of activity and announcements. We will be providing more through the week. One of the more important announcements was Cloudera’s articulation of an Enterprise Data hub. The significance of this is huge for enterprise data. Imagine a place where you can store all data, structured and unstructured, for a very economical cost. This alone is a fantastic, highly desired capability. The enterprise data hub construct has far more capability and features you would expect from a well engineered solution. This includes enhancements to simplify storage, processing, analyzing and managing data. It also includes enhanced security and auditing. And very tight integration into your existing legacy infrastructure and applications.  This is going to be big.

The press release from Cloudera is below:

Cloudera Enterprise 5 Sets New Standard for Data Management; Lays Foundation for The Enterprise Data Hub

Oct/29/2013

Company Extends Category Leadership With Public Beta Release of CDH 5 and Cloudera Enterprise 5; Unveils Industry’s First Enterprise Data Hub and Analysis Platform

PALO ALTO, CA and NEW YORK, NY–(Marketwired – Oct 29, 2013) - From Strata + Hadoop World: Cloudera, the leader in enterprise analytic data management powered by Apache Hadoop™, today unveiled the fifth generation of its Platform for Big Data, Cloudera Enterprise, which is now available for public beta. The new product release, powered by Apache Hadoop 2, offers unique features and advancements that simplify storing, processing, analyzing and managing large structured and unstructured datasets, while offering increased security, robust data management and tight integration with third-party applications. The combination of innovative updates to CDH (Cloudera’s Distribution Including Apache Hadoop) at the core — plus enhancements to Cloudera Manager for Hadoop system administration and Cloudera Navigator for Hadoop audit and access control, data discovery and lineage analysis — together deliver the industry’s first Enterprise Data Hub.

“With Cloudera Enterprise 5, Cloudera has taken several important steps toward realizing its vision to transform Hadoop into an enterprise data hub for analytics,” said Tony Baer, Principal Analyst for Ovum. “Adding support for in-memory data tiering and user-defined functions are essential for delivering the kind of performance that enterprises expect from their analytic data platforms.”

Rethink Data: Introducing the Enterprise Data Hub
Organizations currently employ a variety of systems to support their diverse data hub goals: data warehouses for operational reporting; storage systems to keep data available and safe; specialized massively-parallel databases for large-scale analytics; and search systems for finding and exploring documents. While these systems are suitable for traditional data and workloads, they are not equipped to handle today’s exponential growth in data volume and variety, or the range of users who seek insights from that data. And because each system is purpose-built for a particular class of data and workload, no single system can provide unified access to all relevant information to diverse business users. A new hybrid approach is required, which pragmatically extends the value of existing investments while enabling fundamentally new ways of delivering value from data.

The objective is simple: Acquire and combine any amount or type of data in its original fidelity, in one place, for as long as is necessary, and deliver insights to all kinds of users, as fast as possible. And do so with maximum efficiency of capital and resources.

The solution? The Enterprise Data Hub. One place to store and work with all data, with the flexibility to run a variety of enterprise workloads — including batch processing, interactive SQL, enterprise search and advanced analytics — together with the integrations to existing systems, robust security, governance, data protection, and management that enterprises require. The Enterprise Data Hub is the emerging and necessary center of enterprise data management, complementing existing infrastructure.

Cloudera Enterprise 5: Next Generation Platform for Big Data powered by Apache Hadoop
Built for the demanding requirements of enterprise customers, Cloudera Enterprise enables companies to store, process and analyze unlimited amounts of data and applications from a single system. The newest innovations in Cloudera Enterprise 5 offer customers a significant leap forward in the evolution of the platform, which can now be used to efficiently address an even wider range of business problems. Customers can now use Cloudera to easily handle the rapidly increasing data volume and variety they face, absorbing a growing share of data and workloads from legacy infrastructure while optimizing the efficiency of those existing systems.

Cloudera Enterprise 5 offers a single platform from which organizations can tackle diverse critical business problems:

  • Automatically archiving the complete set of enterprise data to meet compliance requirements while retaining queryable access;
  • Complementing data warehouses to offload data and workloads to help customers increase efficiency and manage costs, while delivering faster ETL/ELT data processing at scale;
  • Supporting business intelligence, through familiar tools, on more data and more kinds of data than ever before possible;
  • Enabling and consolidating enterprise search on data and documents in-place within the single environment; and
  • Accelerating a diverse array of advanced analytics solutions, like recommendation engines, fraud detection or image processing.

Increasingly, strategic partners like Informatica are certifying reference architectures to bring these benefits to joint customers. For example, Informatica and Cloudera together provide a “Data Warehouse Optimization” solution to address the challenges facing traditional data warehouse infrastructures, where capacity is too quickly consumed by increasing data volumes, leading to performance bottlenecks and costly upgrades.

Key advances in Cloudera Enterprise 5 include:

Accelerated Time-to-Value

  • In-Memory HDFS Caching: Datasets from HDFS can now be cached in-memory, boosting MapReduce data processing performance and Cloudera Impala’s analytic query response times for even faster time to insight.
  • User-Defined Functions (UDFs): Customers can now use the custom query functions they depend on in conjunction with Cloudera Impala to deliver the business insights they require. They can also take advantage of the popular open source MADlib library of pre-built statistical and analytic functions to enable scalable in-database analytics.

Improved Efficiency

  • Resource Management: Cloudera Enterprise now delivers advanced resource management for running multiple frameworks for data processing and analysis on a single cluster through the powerful combination of Hadoop YARN (Yet Another Resource Negotiator) and Cloudera Manager. For the first time, administrators can allocate resources not only by workload, but by workgroup, ensuring the best combination of performance and utilization. For example, customers can dedicate 50% of capacity for IT to run mission critical data processing jobs, 30% to the marketing team for ad-hoc BI queries, and so on.
  • Unified Management of Third Party Applications. Cloudera Manager now provides extensibility to enable customers to deploy, manage and monitor products from Cloudera partners such as SAS, Revolution Analytics, Syncsort and many more. Now, customers can manage complex clustered environments from within a single, intuitive interface.

Comprehensive Data Management

  • Manage and Explore Big Data. In addition to enabling centralized data auditing for Hadoop, Cloudera Navigator now provides:
    • Data Discovery: Analysts and data modelers can search, explore, define and tag datasets through the Cloudera Navigator interface, to help identify relevant information for downstream analysis or processing.
    • Data Lineage: As the amount of data in Cloudera Enterprise grows, so does the importance of understanding how that data is used across the organization. Cloudera Navigator delivers the industry’s first data lineage solution for Hadoop, enabling customers to meet regulatory requirements, find associated datasets, and satisfy data governance and retention policies.
  • Data Protection: HDFS and HBase now support snapshots to help prevent data loss.
  • NFS-based Data and Application Access: Easily integrate Cloudera Enterprise with data in and applications running on existing filesystems with native support for NFSv3.

“Over the last five years, we have worked closely with enterprises around the world to help them capture the value in the data they have. Resoundingly, they have asked for a more secure, more reliable real-time data platform that streamlines their existing architectures and speeds up time to insight,” said Mike Olson, chairman and chief strategy officer, Cloudera. “The market has spoken and we are listening. The new capabilities introduced in Cloudera Enterprise 5 deliver the industry’s first Enterprise Data Hub.”

Product Availability and Documentation
Public beta releases of Cloudera Enterprise 5 and CDH 5 are now available. To learn more about Cloudera Enterprise 5, visit http://cloudera.com/CE5. To learn more about CDH 5, or to download it for free, visithttp://www.cloudera.com/content/cloudera/en/products/cdh.html.

The Cloudera Enterprise Data Hub is available today on Cloudera Enterprise 4, for more information contact Cloudera on [email protected].

This information is not a commitment, promise or legal obligation to deliver any material, code, or functionality. Cloudera does not guarantee that the beta software will be made generally available or that any individual feature in the beta version will be made generally available. Cloudera may make the beta software generally available, or not, in its sole discretion and without obligation to make any communication of any kind with regard to such availability.

About Cloudera
Cloudera is revolutionizing enterprise data management by offering the first unified Platform for Big Data: The Enterprise Data Hub. Cloudera offers enterprises one place to store, process and analyze all their data, empowering them to extend the value of existing investments, while enabling fundamental new ways to derive value from their data. Founded in 2008, Cloudera was the first, and is still today, the leading provider and supporter of Hadoop for the enterprise. Cloudera also offers software for business critical data challenges, including storage, access, management, analysis, security and search. With over 15,000 individuals trained, Cloudera is a leading educator of data professionals, offering the industry’s broadest array of Hadoop training and certification programs. Cloudera works with over 700 hardware, software and services partners to meet customers’ big data goals. Leading organizations in every industry run Cloudera in production, including finance, telecommunications, retail, internet, utilities, oil and gas, healthcare, biopharmaceuticals, networking and media, plus top public sector organizations globally. www.cloudera.com

Connect with Cloudera
Read our blog: http://blog.cloudera.com/blog/
Follow us on Twitter: https://twitter.com/cloudera
Visit us on Facebook: https://www.facebook.com/cloudera

Cloudera, Cloudera Manager, Cloudera Navigator, CDH, Cloudera Enterprise, Cloudera Standard and Cloudera Enterprise Data Hub are trademarks or registered trademarks of Cloudera in the United States and in jurisdictions throughout the world. All other company and product names may be trade names or trademarks of their respective owners.

Read the original blog entry...

More Stories By Bob Gourley

Bob Gourley writes on enterprise IT. He is a founder of Crucial Point and publisher of CTOvision.com

IoT & Smart Cities Stories
According to Forrester Research, every business will become either a digital predator or digital prey by 2020. To avoid demise, organizations must rapidly create new sources of value in their end-to-end customer experiences. True digital predators also must break down information and process silos and extend digital transformation initiatives to empower employees with the digital resources needed to win, serve, and retain customers.
Early Bird Registration Discount Expires on August 31, 2018 Conference Registration Link ▸ HERE. Pick from all 200 sessions in all 10 tracks, plus 22 Keynotes & General Sessions! Lunch is served two days. EXPIRES AUGUST 31, 2018. Ticket prices: ($1,295-Aug 31) ($1,495-Oct 31) ($1,995-Nov 12) ($2,500-Walk-in)
Business professionals no longer wonder if they'll migrate to the cloud; it's now a matter of when. The cloud environment has proved to be a major force in transitioning to an agile business model that enables quick decisions and fast implementation that solidify customer relationships. And when the cloud is combined with the power of cognitive computing, it drives innovation and transformation that achieves astounding competitive advantage.
Machine learning has taken residence at our cities' cores and now we can finally have "smart cities." Cities are a collection of buildings made to provide the structure and safety necessary for people to function, create and survive. Buildings are a pool of ever-changing performance data from large automated systems such as heating and cooling to the people that live and work within them. Through machine learning, buildings can optimize performance, reduce costs, and improve occupant comfort by ...
René Bostic is the Technical VP of the IBM Cloud Unit in North America. Enjoying her career with IBM during the modern millennial technological era, she is an expert in cloud computing, DevOps and emerging cloud technologies such as Blockchain. Her strengths and core competencies include a proven record of accomplishments in consensus building at all levels to assess, plan, and implement enterprise and cloud computing solutions. René is a member of the Society of Women Engineers (SWE) and a m...
IoT is rapidly becoming mainstream as more and more investments are made into the platforms and technology. As this movement continues to expand and gain momentum it creates a massive wall of noise that can be difficult to sift through. Unfortunately, this inevitably makes IoT less approachable for people to get started with and can hamper efforts to integrate this key technology into your own portfolio. There are so many connected products already in place today with many hundreds more on the h...
Digital Transformation: Preparing Cloud & IoT Security for the Age of Artificial Intelligence. As automation and artificial intelligence (AI) power solution development and delivery, many businesses need to build backend cloud capabilities. Well-poised organizations, marketing smart devices with AI and BlockChain capabilities prepare to refine compliance and regulatory capabilities in 2018. Volumes of health, financial, technical and privacy data, along with tightening compliance requirements by...
Charles Araujo is an industry analyst, internationally recognized authority on the Digital Enterprise and author of The Quantum Age of IT: Why Everything You Know About IT is About to Change. As Principal Analyst with Intellyx, he writes, speaks and advises organizations on how to navigate through this time of disruption. He is also the founder of The Institute for Digital Transformation and a sought after keynote speaker. He has been a regular contributor to both InformationWeek and CIO Insight...
Digital Transformation is much more than a buzzword. The radical shift to digital mechanisms for almost every process is evident across all industries and verticals. This is often especially true in financial services, where the legacy environment is many times unable to keep up with the rapidly shifting demands of the consumer. The constant pressure to provide complete, omnichannel delivery of customer-facing solutions to meet both regulatory and customer demands is putting enormous pressure on...
Andrew Keys is Co-Founder of ConsenSys Enterprise. He comes to ConsenSys Enterprise with capital markets, technology and entrepreneurial experience. Previously, he worked for UBS investment bank in equities analysis. Later, he was responsible for the creation and distribution of life settlement products to hedge funds and investment banks. After, he co-founded a revenue cycle management company where he learned about Bitcoin and eventually Ethereal. Andrew's role at ConsenSys Enterprise is a mul...