Generated by All in One SEO v4.9.4.1, this is an llms.txt file, used by LLMs to index the site. # datacognate.com ## Sitemaps - [XML Sitemap](https://datacognate.com/sitemap.xml): Contains all public & indexable URLs for this website. ## Posts - [Developing Java Map-Reduce on local machine to run on Hadoop Cluster](https://datacognate.com/developing-java-map-reduce-on-local-machine-to-run-on-hadoop-cluster/) - Introduction In this post, I have explained how to develop hadoop jobs in Java and export JAR to run on Hadoop clusters. Most of the articles on internet, talk about installing eclipse-plugin and using maven or ANT to build JAR. To install eclipse-plugin for hadoop, one needs to install eclipse on the same Linux machine - [Implementing Security in Hadoop Cluster](https://datacognate.com/implementing-security-in-hadoop-cluster/) - When we talk about security in Hadoop, we need to explore all the aspect of cluster networking and understand how the Nodes and Client communicate to each other. Let’s list down possible communication in a simple cluster. Master – Slave communication => Namenode - Datanode / Jobtracker - Tasktracker communicationSlave to slave communication => Datanode - [Cloudera vs AWS vs AZURE vs Google Cloud: How to decide on the right big data platform?](https://datacognate.com/cloudera-vs-aws-vs-azure-vs-google-cloud-how-to-decide-on-the-right-big-data-platform/) - UPDATED:28 Sep 2024: This article was published many years ago. Most of the facts described in this article may not be valid in today's scenario. The updated version of this article will be published soon. Background Big data concepts evolved to solve a specific problem of processing data of diversified nature, high volume and streaming - [Hadoop Streaming with Perl Script](https://datacognate.com/hadoop-streaming-with-perl-script/) - In this article, I am going to explain how to use Hadoop streaming with Perl scripts. First, let’s understand some theory behind Hadoop streaming. Hadoop has been written in Java. Therefore, the native language to write MapReduce program is Java. But, Hadoop also provide an API to MapReduce that allows you to write your map - [Tool & ToolRunner – Simplifying the concept](https://datacognate.com/tool-toolrunner-simplifying-the-concept/) - Writing a mapper & reducer Program definition is easy. Just extend your class by org.apache.hadoop.mapreduce.Mapper and org.apache.hadoop.mapreduce.Reducer respectively and override the map and reduce methods to implement your logics. But, when it comes to write driver program (contain main method of program) for the MapReduce Job, it’s always preferable to use ToolRunner class & Tool ## Pages - [Terms Of Use](https://datacognate.com/terms-of-use/) - Last updated: February 14, 2019 Please read these Terms and Conditions carefully before using the https://www.datacognate.com website operated by DataCognate. Your access to and use of the content is conditioned on your acceptance of and compliance with these Terms. These Terms apply to all visitors, users and others who access or use the Website. By - [About](https://datacognate.com/about/) - About DataCognate DataCognate is managed by Paaword digital media. This blog site is primarily focused on Information technology and data transformation. The objective of this blog site is to share the knowledge and provide consulting expertise by the industry experts on the subject related to big data, artificial intelligence, machine learning and cloud computing. It - [Home](https://datacognate.com/home/) - [Privacy Policy](https://datacognate.com/privacy-policy/) - Effective date: February 14, 2019 DataCognate operates the https://datacognate.com website (the "Service"). This page informs you of our policies regarding the collection, use, and disclosure of personal data when you use our Service and the choices you have associated with that data. Information Collection And Use We do not collect any type of personal information - [Contact Us](https://datacognate.com/contact/) ## Categories - [Big Data & Hadoop](https://datacognate.com/category/big-data-hadoop/) ## Tags - [Apache Hadoop](https://datacognate.com/tag/apache-hadoop/) - [AWS](https://datacognate.com/tag/aws/) - [Azure](https://datacognate.com/tag/azure/) - [Big Data](https://datacognate.com/tag/big-data/) - [Cloud Computing](https://datacognate.com/tag/cloud-computing/) - [Cloudera](https://datacognate.com/tag/cloudera/) - [Hortonworks](https://datacognate.com/tag/hortonworks/) - [Kerberos](https://datacognate.com/tag/kerberos/) - [MapReduce](https://datacognate.com/tag/mapreduce/) - [Perl](https://datacognate.com/tag/perl/)