{"@id":"https://credentialengineregistry.org/resources/ce-5cd5ca4f-0d58-45db-83f6-9a503d01247a","@type":"ceterms:Certification","@context":"https://credreg.net/ctdl/schema/context/json","ceterms:ctid":"ce-5cd5ca4f-0d58-45db-83f6-9a503d01247a","ceterms:name":{"en-US":"Big Data Hadoop Developer or Administration"},"ceterms:ownedBy":["https://credentialengineregistry.org/resources/ce-b8aeaa6c-1e7b-4e30-a71d-58ca884d324f"],"ceterms:requires":[{"@type":"ceterms:ConditionProfile","ceterms:description":{"en-US":"Big Data is a term used to define data sets that have the potential to rapidly grow so large that they become unmanageable. The Big Data movement includes new tools and ways of storing information that allow efficient processing and analysis for informed business decision-making The availability of large data sets presents new opportunities and challenges to organizations of all sizes. This course provides the hands-on programming skills to leverage the Apache Hadoop platform to efficiently process a variety of Big Data. Additionally, you learn to test and deploy Big Data solutions on commodity clusters. This course also covers Pig, Hive, HBase and other components of the Hadoop ecosystem. Further, this course teaches testing, deployment and best practices to architect and develop a complete Big Data solution. This course provides the hands-on experience to install, configure and manage the Apache Hadoop platform and its associated ecosystem. Attendees learn to monitor Hadoop using built-in functionality and associated tools like Ganglia. Additionally, they learn to optimize resource allocation related to the file system and MapReduce. This course also covers techniques of ensuring robustness, efficiency and high availability using approaches such as redundancy and NameNode Federation. In addition to coverage of core Hadoop administration, attendees are exposed to the management of Pig, Hive, ZooKeeper, Oozie and HBase as well as the challenges related to backup, recovery and security. Sqoop and Flume will be employed to demonstrate the migration of data into and out of Hadoop. Students in this program can opt to take either Big Data Hadoop Developer or Big Data Hadoop Administration class."},"ceterms:targetLearningOpportunity":["https://credentialengineregistry.org/resources/ce-2b87a567-2b58-4ac9-a6ba-b7db64dd619d"]}],"ceterms:offeredBy":["https://credentialengineregistry.org/resources/ce-b8aeaa6c-1e7b-4e30-a71d-58ca884d324f"],"ceterms:inLanguage":["en-US"],"ceterms:description":{"en-US":"Big Data is a term used to define data sets that have the potential to rapidly grow so large that they become unmanageable. The Big Data movement includes new tools and ways of storing information that allow efficient processing and analysis for informed business decision-making The availability of large data sets presents new opportunities and challenges to organizations of all sizes. This course provides the hands-on programming skills to leverage the Apache Hadoop platform to efficiently process a variety of Big Data. Additionally, you learn to test and deploy Big Data solutions on commodity clusters. This course also covers Pig, Hive, HBase and other components of the Hadoop ecosystem. Further, this course teaches testing, deployment and best practices to architect and develop a complete Big Data solution. This course provides the hands-on experience to install, configure and manage the Apache Hadoop platform and its associated ecosystem. Attendees learn to monitor Hadoop using built-in functionality and associated tools like Ganglia. Additionally, they learn to optimize resource allocation related to the file system and MapReduce. This course also covers techniques of ensuring robustness, efficiency and high availability using approaches such as redundancy and NameNode Federation. In addition to coverage of core Hadoop administration, attendees are exposed to the management of Pig, Hive, ZooKeeper, Oozie and HBase as well as the challenges related to backup, recovery and security. Sqoop and Flume will be employed to demonstrate the migration of data into and out of Hadoop. Students in this program can opt to take either Big Data Hadoop Developer or Big Data Hadoop Administration class."},"ceterms:subjectWebpage":"http://www.comnetgroup.com","ceterms:credentialStatusType":{"@type":"ceterms:CredentialAlignmentObject","ceterms:framework":"https://credreg.net/ctdl/terms/CredentialStatus","ceterms:targetNode":"credentialStat:Active","ceterms:frameworkName":{"en-US":"Credential Status"},"ceterms:targetNodeName":{"en-US":"Active"},"ceterms:targetNodeDescription":{"en-US":"Awards of the credential are ongoing."}},"ceterms:instructionalProgramType":[{"@type":"ceterms:CredentialAlignmentObject","ceterms:framework":"https://nces.ed.gov/ipeds/cipcode/Default.aspx?y=56","ceterms:targetNode":"https://nces.ed.gov/ipeds/cipcode/cipdetail.aspx?y=56\u0026cip=11.0802","ceterms:codedNotation":"11.0802","ceterms:frameworkName":{"en-US":"Classification of Instructional Programs"},"ceterms:targetNodeName":{"en-US":"Data Modeling/Warehousing and Database Administration."},"ceterms:targetNodeDescription":{"en-US":"A program that prepares individuals to design and manage the construction of databases and related software programs and applications, including the linking of individual data sets to create complex searchable databases (warehousing) and the use of analytical search tools (mining). Includes instruction in database theory, logic, and semantics; operational and warehouse modeling; dimensionality; attributes and hierarchies; data definition; technical architecture; access and security design; integration; formatting and extraction; data delivery; index design; implementation problems; planning and budgeting; and client and networking issues."}}]}