Virtualizing Hadoop : How to Install, Deploy, and Optimize Hadoop in a Virtualized Architecture

Λεπτομέρειες βιβλιογραφικής εγγραφής
Τίτλος: Virtualizing Hadoop : How to Install, Deploy, and Optimize Hadoop in a Virtualized Architecture
Περιγραφή: Plan and Implement Hadoop Virtualization for Maximum Performance, Scalability, and Business Agility Enterprises running Hadoop must absorb rapid changes in big data ecosystems, frameworks, products, and workloads. Virtualized approaches can offer important advantages in speed, flexibility, and elasticity. Now, a world-class team of enterprise virtualization and big data experts guide you through the choices, considerations, and tradeoffs surrounding Hadoop virtualization. The authors help you decide whether to virtualize Hadoop, deploy Hadoop in the cloud, or integrate conventional and virtualized approaches in a blended solution. First, Virtualizing Hadoop reviews big data and Hadoop from the standpoint of the virtualization specialist. The authors demystify MapReduce, YARN, and HDFS and guide you through each stage of Hadoop data management. Next, they turn the tables, introducing big data experts to modern virtualization concepts and best practices. Finally, they bring Hadoop and virtualization together, guiding you through the decisions you'll face in planning, deploying, provisioning, and managing virtualized Hadoop. From security to multitenancy to day-to-day management, you'll find reliable answers for choosing your best Hadoop strategy and executing it. Coverage includes the following: • Reviewing the frameworks, products, distributions, use cases, and roles associated with Hadoop • Understanding YARN resource management, HDFS storage, and I/O • Designing data ingestion, movement, and organization for modern enterprise data platforms • Defining SQL engine strategies to meet strict SLAs • Considering security, data isolation, and scheduling for multitenant environments • Deploying Hadoop as a service in the cloud • Reviewing the essential concepts, capabilities, and terminology of virtualization • Applying current best practices, guidelines, and key metrics for Hadoop virtualization • Managing multiple Hadoop frameworks and products as one unified system • Virtualizing master and worker nodes to maximize availability and performance • Installing and configuring Linux for a Hadoop environment

Συγγραφείς: George Trujillo, Charles Kim, Steve Jones, Rommel Garcia, Justin Murray
Resource Type: eBook.
Θέματα: Apache Hadoop, Virtual computer systems, Electronic data processing--Distributed processi, File organization (Computer science), File processing (Computer science)
Categories: COMPUTERS / Data Science / Data Analytics, COMPUTERS / Certification Guides / General
Βάση Δεδομένων: eBook Index
FullText Text:
  Availability: 0
Header DbId: edsebk
DbLabel: eBook Index
An: 1601470
RelevancyScore: 918
AccessLevel: 6
PubType: eBook
PubTypeId: ebook
PreciseRelevancyScore: 918.373718261719
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: Virtualizing Hadoop : How to Install, Deploy, and Optimize Hadoop in a Virtualized Architecture
– Name: Abstract
  Label: Description
  Group: Ab
  Data: Plan and Implement Hadoop Virtualization for Maximum Performance, Scalability, and Business Agility Enterprises running Hadoop must absorb rapid changes in big data ecosystems, frameworks, products, and workloads. Virtualized approaches can offer important advantages in speed, flexibility, and elasticity. Now, a world-class team of enterprise virtualization and big data experts guide you through the choices, considerations, and tradeoffs surrounding Hadoop virtualization. The authors help you decide whether to virtualize Hadoop, deploy Hadoop in the cloud, or integrate conventional and virtualized approaches in a blended solution. First, Virtualizing Hadoop reviews big data and Hadoop from the standpoint of the virtualization specialist. The authors demystify MapReduce, YARN, and HDFS and guide you through each stage of Hadoop data management. Next, they turn the tables, introducing big data experts to modern virtualization concepts and best practices. Finally, they bring Hadoop and virtualization together, guiding you through the decisions you'll face in planning, deploying, provisioning, and managing virtualized Hadoop. From security to multitenancy to day-to-day management, you'll find reliable answers for choosing your best Hadoop strategy and executing it. Coverage includes the following: • Reviewing the frameworks, products, distributions, use cases, and roles associated with Hadoop • Understanding YARN resource management, HDFS storage, and I/O • Designing data ingestion, movement, and organization for modern enterprise data platforms • Defining SQL engine strategies to meet strict SLAs • Considering security, data isolation, and scheduling for multitenant environments • Deploying Hadoop as a service in the cloud • Reviewing the essential concepts, capabilities, and terminology of virtualization • Applying current best practices, guidelines, and key metrics for Hadoop virtualization • Managing multiple Hadoop frameworks and products as one unified system • Virtualizing master and worker nodes to maximize availability and performance • Installing and configuring Linux for a Hadoop environment <p style=
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22George+Trujillo%22">George Trujillo</searchLink><br /><searchLink fieldCode="AR" term="%22Charles+Kim%22">Charles Kim</searchLink><br /><searchLink fieldCode="AR" term="%22Steve+Jones%22">Steve Jones</searchLink><br /><searchLink fieldCode="AR" term="%22Rommel+Garcia%22">Rommel Garcia</searchLink><br /><searchLink fieldCode="AR" term="%22Justin+Murray%22">Justin Murray</searchLink>
– Name: TypePub
  Label: Resource Type
  Group: TypPub
  Data: eBook.
– Name: Subject
  Label: Subjects
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Apache+Hadoop%22">Apache Hadoop</searchLink><br /><searchLink fieldCode="DE" term="%22Virtual+computer+systems%22">Virtual computer systems</searchLink><br /><searchLink fieldCode="DE" term="%22Electronic+data+processing--Distributed+processi%22">Electronic data processing--Distributed processi</searchLink><br /><searchLink fieldCode="DE" term="%22File+organization+%28Computer+science%29%22">File organization (Computer science)</searchLink><br /><searchLink fieldCode="DE" term="%22File+processing+%28Computer+science%29%22">File processing (Computer science)</searchLink>
– Name: SubjectBISAC
  Label: Categories
  Group: Su
  Data: <searchLink fieldCode="ZK" term="%22COMPUTERS+%2F+Data+Science+%2F+Data+Analytics%22">COMPUTERS / Data Science / Data Analytics</searchLink><br /><searchLink fieldCode="ZK" term="%22COMPUTERS+%2F+Certification+Guides+%2F+General%22">COMPUTERS / Certification Guides / General</searchLink>
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=edsebk&AN=1601470
RecordInfo BibRecord:
  BibEntity:
    Classifications:
      – Code: 005.4
        Scheme: ddc
        Type: prePub
    Languages:
      – Code: eng
        Text: English
    Subjects:
      – SubjectFull: Apache Hadoop
        Type: general
      – SubjectFull: Virtual computer systems
        Type: general
      – SubjectFull: Electronic data processing--Distributed processi
        Type: general
      – SubjectFull: File organization (Computer science)
        Type: general
      – SubjectFull: File processing (Computer science)
        Type: general
    Titles:
      – TitleFull: Virtualizing Hadoop : How to Install, Deploy, and Optimize Hadoop in a Virtualized Architecture
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: George Trujillo
      – PersonEntity:
          Name:
            NameFull: Charles Kim
      – PersonEntity:
          Name:
            NameFull: Steve Jones
      – PersonEntity:
          Name:
            NameFull: Rommel Garcia
      – PersonEntity:
          Name:
            NameFull: Justin Murray
      – PersonEntity:
          Name:
            NameFull: George Trujillo
      – PersonEntity:
          Name:
            NameFull: Charles Kim
      – PersonEntity:
          Name:
            NameFull: Steve Jones
      – PersonEntity:
          Name:
            NameFull: Rommel Garcia
      – PersonEntity:
          Name:
            NameFull: Justin Murray
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 01
              M: 01
              Type: published
              Y: 2015
            – D: 23
              M: 09
              Type: profile
              Y: 2024
          Identifiers:
            – Type: isbn-print
              Value: 9780133811025
            – Type: isbn-electronic
              Value: 9780133811117
            – Type: isbn-electronic
              Value: 9780133811131
          Titles:
            – TitleFull: Virtualizing Hadoop : How to Install, Deploy, and Optimize Hadoop in a Virtualized Architecture
              Type: main
ResultId 1