The Hortonworks Blog

Posts categorized by : Data Lake

On April 30, learn from experts at Hortonworks, Cisco, and Red Hat about accelerating the implementation of a scalable, cost-efficient and robust Big Data solution. Here is a sneak preview of what you’ll hear from our speakers:

  • Ali Bajawa, Senior Partner Solution Engineer, Hortonworks
  • Ron Graham, System Engineer for Big Data Analytics, Cisco
  • Irshad Raihan, Senior Principal, Big Data Product Marketing, Red Hat

Register Now

1. What should a company consider when looking for a big data solution?…

Waterline Data is a Hortonworks Technology Partner and recently earned HDP Certification and YARN Ready with their solution that automates the inventory of data assets in the data lake, enables data governance, and provides self-service to data engineers and data scientists to find and understand their data. Learn more by joining the upcoming webinar on May 6, download the Sandbox tutorial or joint whitepaper. Our guest blogger is Oliver Claude, CMO at Waterline Data.…

Can you identify the unused data in your data warehouse? Are you using your “big data” efficiently? Are your data migration projects cost effective? Is your data in compliance with industry regulations? If you answered “no” to any or all of these questions, then you may want to learn more about how to optimize your data warehouse.

On April 23rd at 11:00 am PST, Adis Cesir, Big Data Solution Engineer at Hortonworks, Ramu Kalvakuntla, Principal at RCG Global Services Big Data Practice, and Santosh Chitakki, Director of Product Management at Attunity, will be telling us more about rebalancing data warehouses and integrating your current enterprise data warehouse with a Modern Data Architecture.…

Enterprises across all major industries adopt Apache Hadoop for its ability to store and process an abundance of new types of data in a modern data architecture. This “Any Data” capability has always been a hallmark feature of Hadoop, opening insight from new data sources such as clickstream, web and social, geo-location, IoT, server logs, or traditional data sets from ERP, CRM, SCM or other existing data systems.…

Hortonworks is pleased to announce the general availability of Apache Spark in Hortonworks Data Platform (HDP)— now available on our downloads page. With HDP 2.2.4 Hortonworks now offers support for your developers and data scientists using Apache Spark 1.2.1.

HDP’s YARN-based architecture enables multiple applications to share a common cluster and dataset while ensuring consistent levels of service and response. Now Spark is one of the many data access engines that works with YARN and that is supported in an HDP enterprise data lake.…

Today EMC is launching their EMC® Business Data Lake solution, the first fully-engineered, enterprise-grade solution for a Data Lake running on EMC infrastructure. At Hortonworks, we’ve been assisting customers on their journey to a data lake via a Modern Data Architecture (MDA) and our vision and EMC’s vision are highly complementary and so we’re delighted to be part of the EMC Business Data Lake.

The Data Lake enabled by a Modern Data Architecture allows enterprises to be a Data-First Enterprise.…

Forrester recently called Apache Hadoop adoption “mandatory” for the enterprise. For most organizations, moving forward with Hadoop is no longer a question of if, but when. Hadoop-powered insight into big data is enabling market disruption in every industry and the market winners are those who handle that data most effectively and at the lowest cost.

As with any new platform, making decisions on how best to implement and for what purpose can be challenging.…

Cisco and Hortonworks established their official alliance back in 2013. Together, they have been bringing to life the vision of a single big data platform for the enterprise. As every industry is witnessing unprecedented quantities of data and a variety of new data types e.g. clickstream and behavior, machine and sensor, geographic data, server logs, sentiment and web…, Cisco and Hortonworks have been collaborating to empower companies with their data. Oftentimes, organizations need to optimize their IT infrastructure and free up their Enterprise Data Warehouse (EDW) to make the most of all of their data, building new analytic applications and moving towards the vision of the Data Lake.

In this guest blog, Kumar Srivastava, senior director of product management at ClearStory Data, shares his thoughts on ClearStory’s integration with Hortonworks Data Platform (HDP)

We are excited to be working with and announcing ClearStory Data’s integration with Hortonworks Data Platform (HDP) during Strata + Hadoop World 2015. This partnership with Hortonworks is significant as it brings ClearStory’s business-ready, fast-cycle, scalable analysis on Hadoop Data Lakes and specifically on the Hortonworks Data Platform (HDP).…

There are lots of ways to interact with Hortonworks at this weeks Strata +Hadoop World event.

Exhibitor Booth 1321

While at our booth you can talk with our experts and get the latest on Hortonworks, get an overview of Apache Hadoop or hear more about how we are helping organizations drive success with Hadoop. You can also get one of the popular Hortonworks elephants!

Passport Program

While at our booth you can pick up a Passport Card to that you can enter for a chance to win some great prizes from one of the 24!…

By now, we have all heard about Big Data. However, approaches to derive value from the phenomenon vary greatly from one organization to another. While companies like Facebook, Google or Yahoo! were birthplaces of game-changing innovations, most corporations are still trying to figure out how to unlock the power of Big Data.

In this video series created in collaboration between Informatica and Hortonworks, two pioneers and leaders in the data space, you will hear about a wide range of topics addressed in simple business terms.…

Have you ever wondered how to share content infrastructure that transparently synchronizes information with your existing systems? Are you looking for ways to build an open standards-based platform for deep analysis and data monetization? If so, you will want to join our webinar on Wednesday, January 21st, at 10 AM PT.

Our Big Data experts will teach you how to:

  • Leverage 100% of your data, including text, images, audio, video, and many more data types to be automatically consumed and enriched using HP Haven and Hortonworks Data Platform (HDP).
  • As we approach the opening bell on Nasdaq and another milestone for open source Apache Hadoop, we at Hortonworks want to thank those who have contributed deeply to this journey. We owe you – our customers – a huge thank you. Your active collaboration with us in the Apache Hadoop community has greatly impacted the trajectory of this platform for data management and has established a path for how thousands of other enterprises can successfully build a new open data architecture that brings all data under management.…

    Many types of industries are finding new opportunities from an abundance of new types of data stored at scale in Hadoop, combined with Hadoop’s ability to process that data at lower costs than traditional platforms. Apache Hadoop and the Hortonworks Data Platform (HDP) can help enterprises turn what used to be data fumes into high-octane fuel that propels their businesses.

    Sign up for the Hadoop industry solutions email series to find out how Hortonworks customers use Hadoop to solve real-world business challenges.…

    The successful Hadoop journey typically starts with new analytic applications, which lead to a Data Lake. As more and more applications are created that derive value from the new types of data, an architectural shift happens in the data center: companies gain deeper insight across a large, broad, diverse set of data at efficient scale. They create a Data Lake.

    Cisco and Hortonworks have partnered to build a highly efficient, highly scalable way to manage all your enterprise data in a data lake.…