OpenStack Sahara Essentials

Packt Publishing Limited
  • 1. Auflage
  • |
  • erschienen am 25. April 2016
  • |
  • 178 Seiten
E-Book | ePUB mit Adobe DRM | Systemvoraussetzungen
978-1-78588-014-8 (ISBN)
Integrate, deploy, rapidly configure, and successfully manage your own big data-intensive clusters in the cloud using OpenStack SaharaAbout This BookA fast paced guide to help you utilize the benefits of Sahara in OpenStack to meet the Big Data world of Hadoop.A step by step approach to simplify the complexity of Hadoop configuration, deployment and maintenance.Who This Book Is ForThis book targets data scientists, cloud developers and Devops Engineers who would like to become proficient with OpenStack Sahara. Ideally, this book is well suitable for readers who are familiars with databases, Hadoop and Spark solutions. Additionally, a basic prior knowledge of OpenStack is expected. The readers should also be familiar with different Linux boxes, distributions and virtualization technology.What You Will LearnIntegrate and Install Sahara with OpenStack environmentLearn Sahara architecture under the hoodRapidly configure and scale Hadoop clusters on top of OpenStackExplore the Sahara REST API to create, deploy and manage a Hadoop clusterLearn the Elastic Processing Data (EDP) facility to execute jobs in clusters from SaharaCover other Hadoop stable plugins existing supported by SaharaDiscover different features provided by Sahara for Hadoop provisioning and deploymentLearn how to troubleshoot OpenStack Sahara issuesIn DetailThe Sahara project is a module that aims to simplify the building of data processing capabilities on OpenStack.The goal of this book is to provide a focused, fast paced guide to installing, configuring, and getting started with integrating Hadoop with OpenStack, using Sahara.The book should explain to users how to deploy their data-intensive Hadoop and Spark clusters on top of OpenStack. It will also cover how to use the Sahara REST API, how to develop applications for Elastic Data Processing on Openstack, and setting up hadoop or spark clusters on Openstack.Style and approachThis book takes a step by step approach teaching how to integrate, deploy and manage data using OpenStack Sahara. It will teach how the OpenStack Sahara is beneficial by simplifying the complexity of Hadoop configuration, deployment and maintenance.
  • Englisch
  • Birmingham
  • |
  • Großbritannien
978-1-78588-014-8 (9781785880148)
1785880144 (1785880144)
weitere Ausgaben werden ermittelt
Omar Khedher is a systems and network engineer. He worked for a few years in cloud computing environments and was involved in several private cloud projects based on OpenStack. Leveraging his skills as a system administrator in virtualisation, storage, and networking, he works as cloud system engineer for a leading advertising technology company, Fyber, based in Berlin. Currently, together with several highly skilled professional teams in the market, they collaborate to build a high scalable infrastructure based on the cloud platform.
Omar is also the author of another OpenStack book, Mastering OpenStack, Packt Publishing. He has authored also few academic publications based on new researches for the cloud performance improvement.
  • Cover
  • Copyright
  • Credits
  • About the Author
  • About the Reviewer
  • Table of Contents
  • Preface
  • Chapter 1: The Essence of Big Data in the Cloud
  • It is all about data
  • The dimensions of big data
  • The big challenge of big data
  • The revolution of big data
  • A key of big data success
  • Use case: Elastic MapReduce
  • OpenStack crossing big data
  • Sahara: bringing big data to the cloud
  • Sahara in OpenStack
  • The Sahara OpenStack mission
  • Summary
  • Chapter 2: Integrating OpenStack Sahara
  • Preparing the test infrastructure environment
  • OpenStack test topology environment
  • OpenStack test networking layout
  • OpenStack test environment design
  • Installing OpenStack
  • Network requirements
  • System requirements
  • Running the RDO installation
  • Integrating Sahara
  • Installing and configuring OpenStack Sahara
  • Installing the Sahara user interface
  • Summary
  • Chapter 3: Using OpenStack Sahara
  • Planning a Hadoop deployment
  • Assigning Hadoop nodes
  • Sahara provisioning plugins
  • Creating a Hadoop cluster
  • Preparing the image from Horizon
  • Preparing the image using CLI
  • Creating the Node Group Template
  • Creating the Node Group Template in Horizon
  • Creating a Node Group Template using CLI
  • Creating the Node Cluster Template
  • Creating the Node Cluster Template with Horizon
  • Creating the Node Cluster Template using CLI
  • Launching the Hadoop cluster
  • Launching the Hadoop cluster with Horizon
  • Launching the Hadoop cluster using the CLI
  • Summary
  • Chapter 4: Executing Jobs with Sahara
  • Job glossary in Sahara
  • Job binaries in Sahara
  • Jobs in Sahara
  • Running jobs in Sahara
  • Executing jobs via Horizon
  • Executing jobs using the Sahara RESTful API
  • API authentication
  • Launching an EDP job
  • Registering a Spark image using REST API
  • Creating Spark node group templates
  • Creating a Spark cluster template
  • Launching the Spark cluster
  • Creating a job binary
  • Creating a Spark job template
  • Executing the Spark job
  • Extending the Spark job
  • Summary
  • Chapter 5: Discovering Advanced Features with Sahara
  • Sahara plugins
  • Vanilla Apache Hadoop
  • Building an image for the Apache Vanilla plugin
  • Vanilla Apache requirements and limitations
  • Hortonworks Data Platform plugin
  • Building an image for the HDP plugin
  • HDP requirements and limitations
  • Cloudera Distribution Hadoop plugin
  • Building an image for the CDH plugin
  • CDH requirements and limitations
  • Apache Spark plugin
  • Building an image for the Spark plugin
  • Spark requirements and limitations
  • Affinity and anti-affinity
  • Anti-affinity in action
  • Boosting Elastic Data Processing performance
  • Defining the network
  • Increasing data reliability
  • Summary
  • Chapter 6: Hadoop High Availability Using Sahara
  • HDP high-availability support
  • Minimum requirements for the HA Hadoop cluster in Sahara
  • HA Hadoop cluster templates
  • CDH high-availability support
  • Summary
  • Chapter 7: Troubleshooting
  • Troubleshooting OpenStack
  • OpenStack debug tool
  • Troubleshooting SELinux
  • Troubleshooting identity
  • Troubleshooting networking
  • Troubleshooting data processing
  • Debugging Sahara
  • Logging Sahara
  • Troubleshooting missing services
  • Troubleshooting cluster creation
  • Troubleshooting user quota
  • Troubleshooting cluster scaling
  • Troubleshooting cluster access
  • Summary
  • Index

Dateiformat: EPUB
Kopierschutz: Adobe-DRM (Digital Rights Management)


Computer (Windows; MacOS X; Linux): Installieren Sie bereits vor dem Download die kostenlose Software Adobe Digital Editions (siehe E-Book Hilfe).

Tablet/Smartphone (Android; iOS): Installieren Sie bereits vor dem Download die kostenlose App Adobe Digital Editions (siehe E-Book Hilfe).

E-Book-Reader: Bookeen, Kobo, Pocketbook, Sony, Tolino u.v.a.m. (nicht Kindle)

Das Dateiformat EPUB ist sehr gut für Romane und Sachbücher geeignet - also für "fließenden" Text ohne komplexes Layout. Bei E-Readern oder Smartphones passt sich der Zeilen- und Seitenumbruch automatisch den kleinen Displays an. Mit Adobe-DRM wird hier ein "harter" Kopierschutz verwendet. Wenn die notwendigen Voraussetzungen nicht vorliegen, können Sie das E-Book leider nicht öffnen. Daher müssen Sie bereits vor dem Download Ihre Lese-Hardware vorbereiten.

Weitere Informationen finden Sie in unserer E-Book Hilfe.

Download (sofort verfügbar)

32,73 €
inkl. 19% MwSt.
Download / Einzel-Lizenz
ePUB mit Adobe DRM
siehe Systemvoraussetzungen
E-Book bestellen

Unsere Web-Seiten verwenden Cookies. Mit der Nutzung dieser Web-Seiten erklären Sie sich damit einverstanden. Mehr Informationen finden Sie in unserem Datenschutzhinweis. Ok