Feed abonnieren
“When elephants cross the world's hottest desert…” “When elephants cross the world's hottest desert…”

Introduction
Anyone who is serious about big data, scale out applications and cloud infrastructure should want to intimately understand the benefits of scale out architecture and the resource elasticity of cloud services. As we continue our evolution into a deeper understanding of data, we see a need agile access to an elastic big data platform. Such a platform can allow us to capture, synthesize and quantify data into business value.

Enter OpenStack Sahara - the intersection of Hadoop and OpenStack.

As an OpenStack project started by Red Hat, Mirantis and Hortonworks during the OpenStack Havana summit in Portland, Sahara was incubated for the OpenStack Icehouse release and is expected to be integrated for OpenStack Juno by the end of 2014.

Sahara’s mission is to provide a scalable data processing stack and associated management interfaces. Sahara delivers on that mission by providing the ability to rapidly create and manage Apache Hadoop™ clusters and easily run workloads across them. All on OpenStack managed infrastructure, without having to deal with the details of cluster management.

With full cluster lifecycle management, provisioning, scaling and termination, Sahara allows the user to select different Hadoop versions, cluster topology and node hardware details.

Sahara key features and use cases:

  • Fast and agile Hadoop cluster deployment
  • An extensible framework for management and provisioning components
  • Run Hadoop workloads in few clicks without expertise in Hadoop operations
  • “Analytics as a Service” utilization of unused compute capacity for ad-hoc or bursty analytic workloads
  • Sahara supports different types of jobs: MapReduce, Hive, Pig and Oozie workflows. The data could be taken from various sources: Swift, HDFS, NoSQL and SQL databases. It also  supports various provisioning plugins.
  • The intersection of two of the largest open source movements
  • OpenStack provides  the foundation and hub of innovation for cleanly managing infrastructure resources. While Apache Hadoop™ serves as the core and innovation driver for storing and processing data.

Sahara graph

Bringing these two technologies together not only strengthens and catalyzes their ecosystems, but offers an increasing wealth of value to their users.

The OpenStack Sahara project aims to facilitate this combination and enable customers and partners alike to take advantage of a growing big data processing platform on OpenStack.

hadoop openstack

Our vision is to bring Big Data and OpenStack together, with a broad ecosystem of partner interoperability, reliability & choice.

You can use Sahara now in RDO and as technology preview in RHEL OSP 5

Over the next few months, we’ll bring you examples of how to use Sahara in RDO and RHEL OSP, how to get involved as a customer or partner, and tell you about the value provided by merging the infrastructure and data processing universes. Look for post by Keith Basil and Matthew Farrelle.

To learn more and get involved with the Sahara project, please visit the Sahara OpenStack Wiki at: https://wiki.openstack.org/wiki/Sahara

 


Über den Autor

Sean Cohen is the Director of Product Management in the Hybrid Platforms organization at Red Hat, a leading provider of open-source technologies and hybrid cloud solutions. He oversees the infrastructure and observability business, including strategy and delivery for Red Hat OpenShift and OpenStack Platforms. With over 15 years of experience, Sean has a robust background in senior product management and delivery across enterprise and telco markets, driving cloud infrastructure strategy, product lifecycle management, and leading high-performance cross-functional teams.

Read full bio
UI_Icon-Red_Hat-Close-A-Black-RGB

Nach Thema durchsuchen

automation icon

Automatisierung

Das Neueste zum Thema IT-Automatisierung für Technologien, Teams und Umgebungen

AI icon

Künstliche Intelligenz

Erfahren Sie das Neueste von den Plattformen, die es Kunden ermöglichen, KI-Workloads beliebig auszuführen

open hybrid cloud icon

Open Hybrid Cloud

Erfahren Sie, wie wir eine flexiblere Zukunft mit Hybrid Clouds schaffen.

security icon

Sicherheit

Erfahren Sie, wie wir Risiken in verschiedenen Umgebungen und Technologien reduzieren

edge icon

Edge Computing

Erfahren Sie das Neueste von den Plattformen, die die Operations am Edge vereinfachen

Infrastructure icon

Infrastruktur

Erfahren Sie das Neueste von der weltweit führenden Linux-Plattform für Unternehmen

application development icon

Anwendungen

Entdecken Sie unsere Lösungen für komplexe Herausforderungen bei Anwendungen

Original series icon

Original Shows

Interessantes von den Experten, die die Technologien in Unternehmen mitgestalten