{"id":13071,"date":"2014-04-21T12:34:16","date_gmt":"2014-04-21T07:04:16","guid":{"rendered":"https:\/\/2thenew.today\/blog\/?p=13071"},"modified":"2015-07-09T15:32:15","modified_gmt":"2015-07-09T10:02:15","slug":"introduction-to-hadoop","status":"publish","type":"post","link":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/","title":{"rendered":"Introduction To Hadoop"},"content":{"rendered":"<p align=\"JUSTIFY\"><span style=\"font-size: large;\"><span style=\"color: #000000;\"><span style=\"font-family: Century Schoolbook L,serif;\">A Brief History of Hadoop: <\/span><\/span><span style=\"color: #000000;\"><span style=\"font-family: Century Schoolbook L,serif;\">Hadoop was created by Doug Cutting, the creator of Apache Lucene, the widely used text search library. Hadoop has its origins in Apache Nutch, an open source web search engine, itself a part of the Lucene project.<\/span><\/span><\/span><\/p>\n<p align=\"JUSTIFY\"><strong><span style=\"font-family: Century Schoolbook L,serif;\"><span style=\"font-size: large;\">The Origin Of The Name \u201chadoop\u201d<\/span><\/span><\/strong><span style=\"font-family: Century Schoolbook L,serif;\"><span style=\"font-size: large;\">. Hadoop is not an acronym; it\u2019s a made-up name. The project\u2019s creator, Doug Cutting, explains how the name came about:<\/span><\/span> \u201c<span style=\"font-family: Century Schoolbook L,serif;\"><span style=\"font-size: large;\">The name my kid gave a stuffed yellow elephant. Short, relatively easy to spell and pronounce, meaningless, and not used elsewhere: those are my naming criteria. Kids are good at generating such. Googol is a kid\u2019s term.\u201d<\/span><\/span><\/p>\n<p align=\"JUSTIFY\"><strong><span style=\"font-family: Century Schoolbook L,serif;\"><span style=\"font-size: large;\">Hadoop Core Component <\/span><\/span><\/strong><\/p>\n<p align=\"JUSTIFY\"><strong><span style=\"font-family: Century Schoolbook L,serif;\"><span style=\"font-size: large;\">HDFS:(Hadoop Distributed File System):<\/span><\/span><\/strong><span style=\"color: #000000;\"><span style=\"font-family: Century Schoolbook L,serif;\"><span style=\"font-size: large;\">a distributed file system designed to run on commodity hardware. It has many similarities with existing distributed file systems. However, the differences from other distributed file systems are significant. HDFS is highly fault-tolerant and is designed to be deployed on low-cost hardware. HDFS provides high throughput access to application data and is suitable for applications that have large data sets. HDFS relaxes a few POSIX requirements to enable streaming access to file system<\/span><\/span><\/span><\/p>\n<p align=\"JUSTIFY\"><strong><span style=\"font-family: Century Schoolbook L,serif;\"><span style=\"font-size: large;\">Mapreduce:<\/span><\/span><\/strong><\/p>\n<p align=\"JUSTIFY\"><span style=\"font-family: Century Schoolbook L,serif;\"><span style=\"font-size: large;\"><span style=\"color: #000000;\">MapReduce is a software framework that <a title=\"Hadoop Developers\" href=\"https:\/\/2thenew.today\/analytics\">allows developers to write programs<\/a> that process massive amounts of unstructured data in parallel across a distributed cluster of processors or stand-alone computers. It was developed at Google for indexing Web pages and replaced their original indexing algorithms and heuristics in 2004.<\/span><\/span><\/span><\/p>\n","protected":false},"excerpt":{"rendered":"<p>A Brief History of Hadoop: Hadoop was created by Doug Cutting, the creator of Apache Lucene, the widely used text search library. Hadoop has its origins in Apache Nutch, an open source web search engine, itself a part of the Lucene project. The Origin Of The Name \u201chadoop\u201d. Hadoop is not an acronym; it\u2019s a [&hellip;]<\/p>\n","protected":false},"author":114,"featured_media":0,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"iawp_total_views":22,"footnotes":""},"categories":[1395],"tags":[1396,1398],"class_list":["post-13071","post","type-post","status-publish","format-standard","hentry","category-big-data","tag-big-data-2","tag-hadoop"],"aioseo_notices":[],"aioseo_head":"\n\t\t<!-- All in One SEO 5.0.0.1 - aioseo.com -->\n\t<meta name=\"description\" content=\"A Brief History of Hadoop: Hadoop was created by Doug Cutting, the creator of Apache Lucene, the widely used text search library. Hadoop has its origins in Apache Nutch, an open source web search engine, itself a part of the Lucene project. The Origin Of The Name \u201chadoop\u201d. Hadoop is not an acronym; it\u2019s a\" \/>\n\t<meta name=\"robots\" content=\"max-image-preview:large\" \/>\n\t<meta name=\"author\" content=\"Pavan\"\/>\n\t<link rel=\"canonical\" href=\"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/\" \/>\n\t<meta name=\"generator\" content=\"All in One SEO (AIOSEO) 5.0.0.1\" \/>\n\t\t<meta property=\"og:locale\" content=\"en_US\" \/>\n\t\t<meta property=\"og:site_name\" content=\"TO THE NEW BLOG\" \/>\n\t\t<meta property=\"og:type\" content=\"article\" \/>\n\t\t<meta property=\"og:title\" content=\"Introduction To Hadoop | TO THE NEW Blog\" \/>\n\t\t<meta property=\"og:description\" content=\"A Brief History of Hadoop: Hadoop was created by Doug Cutting, the creator of Apache Lucene, the widely used text search library. Hadoop has its origins in Apache Nutch, an open source web search engine, itself a part of the Lucene project. The Origin Of The Name \u201chadoop\u201d. Hadoop is not an acronym; it\u2019s a\" \/>\n\t\t<meta property=\"og:url\" content=\"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/\" \/>\n\t\t<meta property=\"og:image\" content=\"https:\/\/2thenew.today\/blog\/wp-content\/themes\/ttn\/images\/social-logo.png\" \/>\n\t\t<meta property=\"og:image:secure_url\" content=\"https:\/\/2thenew.today\/blog\/wp-content\/themes\/ttn\/images\/social-logo.png\" \/>\n\t\t<meta property=\"article:tag\" content=\"big data\" \/>\n\t\t<meta property=\"article:tag\" content=\"hadoop\" \/>\n\t\t<meta property=\"article:published_time\" content=\"2014-04-21T07:04:16+00:00\" \/>\n\t\t<meta property=\"article:modified_time\" content=\"2015-07-09T10:02:15+00:00\" \/>\n\t\t<meta name=\"twitter:card\" content=\"summary\" \/>\n\t\t<meta name=\"twitter:site\" content=\"@tothenew\" \/>\n\t\t<meta name=\"twitter:title\" content=\"Introduction To Hadoop | TO THE NEW Blog\" \/>\n\t\t<meta name=\"twitter:description\" content=\"A Brief History of Hadoop: Hadoop was created by Doug Cutting, the creator of Apache Lucene, the widely used text search library. Hadoop has its origins in Apache Nutch, an open source web search engine, itself a part of the Lucene project. The Origin Of The Name \u201chadoop\u201d. Hadoop is not an acronym; it\u2019s a\" \/>\n\t\t<meta name=\"twitter:image\" content=\"https:\/\/2thenew.today\/blog\/wp-content\/themes\/ttn\/images\/social-logo.png\" \/>\n\t\t<script type=\"application\/ld+json\" class=\"aioseo-schema\">\n\t\t\t{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/introduction-to-hadoop\\\/#article\",\"name\":\"Introduction To Hadoop | TO THE NEW Blog\",\"headline\":\"Introduction To Hadoop\",\"author\":{\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/author\\\/pavan\\\/#author\"},\"publisher\":{\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/#organization\"},\"datePublished\":\"2014-04-21T12:34:16+05:30\",\"dateModified\":\"2015-07-09T15:32:15+05:30\",\"inLanguage\":\"en-US\",\"mainEntityOfPage\":{\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/introduction-to-hadoop\\\/#webpage\"},\"isPartOf\":{\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/introduction-to-hadoop\\\/#webpage\"},\"articleSection\":\"Big Data, big data, hadoop\"},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/introduction-to-hadoop\\\/#breadcrumblist\",\"itemListElement\":[{\"@type\":\"ListItem\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog#listItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\\\/\\\/www.tothenew.com\\\/blog\",\"nextItem\":{\"@type\":\"ListItem\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/category\\\/big-data\\\/#listItem\",\"name\":\"Big Data\"}},{\"@type\":\"ListItem\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/category\\\/big-data\\\/#listItem\",\"position\":2,\"name\":\"Big Data\",\"item\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/category\\\/big-data\\\/\",\"nextItem\":{\"@type\":\"ListItem\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/introduction-to-hadoop\\\/#listItem\",\"name\":\"Introduction To Hadoop\"},\"previousItem\":{\"@type\":\"ListItem\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog#listItem\",\"name\":\"Home\"}},{\"@type\":\"ListItem\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/introduction-to-hadoop\\\/#listItem\",\"position\":3,\"name\":\"Introduction To Hadoop\",\"previousItem\":{\"@type\":\"ListItem\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/category\\\/big-data\\\/#listItem\",\"name\":\"Big Data\"}}]},{\"@type\":\"Organization\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/#organization\",\"name\":\"TO THE NEW Blog\",\"url\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/\"},{\"@type\":\"Person\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/author\\\/pavan\\\/#author\",\"url\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/author\\\/pavan\\\/\",\"name\":\"Pavan\",\"image\":{\"@type\":\"ImageObject\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/introduction-to-hadoop\\\/#authorImage\",\"url\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/5b433937b104b79072ed364af9c379663e6072b6aa5bed9ced6da4052c3aece9?s=96&d=mm&r=g\",\"width\":96,\"height\":96,\"caption\":\"Pavan\"}},{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/introduction-to-hadoop\\\/#webpage\",\"url\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/introduction-to-hadoop\\\/\",\"name\":\"Introduction To Hadoop | TO THE NEW Blog\",\"description\":\"A Brief History of Hadoop: Hadoop was created by Doug Cutting, the creator of Apache Lucene, the widely used text search library. Hadoop has its origins in Apache Nutch, an open source web search engine, itself a part of the Lucene project. The Origin Of The Name \\u201chadoop\\u201d. Hadoop is not an acronym; it\\u2019s a\",\"inLanguage\":\"en-US\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/#website\"},\"breadcrumb\":{\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/introduction-to-hadoop\\\/#breadcrumblist\"},\"author\":{\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/author\\\/pavan\\\/#author\"},\"creator\":{\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/author\\\/pavan\\\/#author\"},\"datePublished\":\"2014-04-21T12:34:16+05:30\",\"dateModified\":\"2015-07-09T15:32:15+05:30\"},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/#website\",\"url\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/\",\"name\":\"TO THE NEW Blog\",\"inLanguage\":\"en-US\",\"publisher\":{\"@id\":\"https:\\\/\\\/www.tothenew.com\\\/blog\\\/#organization\"}}]}\n\t\t<\/script>\n\t\t<!-- All in One SEO -->\n\n","aioseo_head_json":{"title":"Introduction To Hadoop | TO THE NEW Blog","description":"A Brief History of Hadoop: Hadoop was created by Doug Cutting, the creator of Apache Lucene, the widely used text search library. Hadoop has its origins in Apache Nutch, an open source web search engine, itself a part of the Lucene project. The Origin Of The Name \u201chadoop\u201d. Hadoop is not an acronym; it\u2019s a","canonical_url":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/","robots":"max-image-preview:large","keywords":"","webmasterTools":{"miscellaneous":""},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/#article","name":"Introduction To Hadoop | TO THE NEW Blog","headline":"Introduction To Hadoop","author":{"@id":"https:\/\/2thenew.today\/blog\/author\/pavan\/#author"},"publisher":{"@id":"https:\/\/2thenew.today\/blog\/#organization"},"datePublished":"2014-04-21T12:34:16+05:30","dateModified":"2015-07-09T15:32:15+05:30","inLanguage":"en-US","mainEntityOfPage":{"@id":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/#webpage"},"isPartOf":{"@id":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/#webpage"},"articleSection":"Big Data, big data, hadoop"},{"@type":"BreadcrumbList","@id":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/#breadcrumblist","itemListElement":[{"@type":"ListItem","@id":"https:\/\/2thenew.today\/blog#listItem","position":1,"name":"Home","item":"https:\/\/2thenew.today\/blog","nextItem":{"@type":"ListItem","@id":"https:\/\/2thenew.today\/blog\/category\/big-data\/#listItem","name":"Big Data"}},{"@type":"ListItem","@id":"https:\/\/2thenew.today\/blog\/category\/big-data\/#listItem","position":2,"name":"Big Data","item":"https:\/\/2thenew.today\/blog\/category\/big-data\/","nextItem":{"@type":"ListItem","@id":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/#listItem","name":"Introduction To Hadoop"},"previousItem":{"@type":"ListItem","@id":"https:\/\/2thenew.today\/blog#listItem","name":"Home"}},{"@type":"ListItem","@id":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/#listItem","position":3,"name":"Introduction To Hadoop","previousItem":{"@type":"ListItem","@id":"https:\/\/2thenew.today\/blog\/category\/big-data\/#listItem","name":"Big Data"}}]},{"@type":"Organization","@id":"https:\/\/2thenew.today\/blog\/#organization","name":"TO THE NEW Blog","url":"https:\/\/2thenew.today\/blog\/"},{"@type":"Person","@id":"https:\/\/2thenew.today\/blog\/author\/pavan\/#author","url":"https:\/\/2thenew.today\/blog\/author\/pavan\/","name":"Pavan","image":{"@type":"ImageObject","@id":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/#authorImage","url":"https:\/\/secure.gravatar.com\/avatar\/5b433937b104b79072ed364af9c379663e6072b6aa5bed9ced6da4052c3aece9?s=96&d=mm&r=g","width":96,"height":96,"caption":"Pavan"}},{"@type":"WebPage","@id":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/#webpage","url":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/","name":"Introduction To Hadoop | TO THE NEW Blog","description":"A Brief History of Hadoop: Hadoop was created by Doug Cutting, the creator of Apache Lucene, the widely used text search library. Hadoop has its origins in Apache Nutch, an open source web search engine, itself a part of the Lucene project. The Origin Of The Name \u201chadoop\u201d. Hadoop is not an acronym; it\u2019s a","inLanguage":"en-US","isPartOf":{"@id":"https:\/\/2thenew.today\/blog\/#website"},"breadcrumb":{"@id":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/#breadcrumblist"},"author":{"@id":"https:\/\/2thenew.today\/blog\/author\/pavan\/#author"},"creator":{"@id":"https:\/\/2thenew.today\/blog\/author\/pavan\/#author"},"datePublished":"2014-04-21T12:34:16+05:30","dateModified":"2015-07-09T15:32:15+05:30"},{"@type":"WebSite","@id":"https:\/\/2thenew.today\/blog\/#website","url":"https:\/\/2thenew.today\/blog\/","name":"TO THE NEW Blog","inLanguage":"en-US","publisher":{"@id":"https:\/\/2thenew.today\/blog\/#organization"}}]},"og:locale":"en_US","og:site_name":"TO THE NEW BLOG","og:type":"article","og:title":"Introduction To Hadoop | TO THE NEW Blog","og:description":"A Brief History of Hadoop: Hadoop was created by Doug Cutting, the creator of Apache Lucene, the widely used text search library. Hadoop has its origins in Apache Nutch, an open source web search engine, itself a part of the Lucene project. The Origin Of The Name \u201chadoop\u201d. Hadoop is not an acronym; it\u2019s a","og:url":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/","og:image":"https:\/\/2thenew.today\/blog\/wp-content\/themes\/ttn\/images\/social-logo.png","og:image:secure_url":"https:\/\/2thenew.today\/blog\/wp-content\/themes\/ttn\/images\/social-logo.png","article:tag":{"0":"big data","2":"hadoop"},"article:published_time":"2014-04-21T07:04:16+00:00","article:modified_time":"2015-07-09T10:02:15+00:00","twitter:card":"summary","twitter:site":"@tothenew","twitter:title":"Introduction To Hadoop | TO THE NEW Blog","twitter:description":"A Brief History of Hadoop: Hadoop was created by Doug Cutting, the creator of Apache Lucene, the widely used text search library. Hadoop has its origins in Apache Nutch, an open source web search engine, itself a part of the Lucene project. The Origin Of The Name \u201chadoop\u201d. Hadoop is not an acronym; it\u2019s a","twitter:image":"https:\/\/2thenew.today\/blog\/wp-content\/themes\/ttn\/images\/social-logo.png"},"aioseo_meta_data":{"post_id":"13071","title":null,"description":null,"keywords":null,"keyphrases":null,"primary_term":null,"canonical_url":null,"og_title":"","og_description":"","og_object_type":"article","og_image_type":"default","og_image_url":null,"og_image_width":null,"og_image_height":null,"og_image_custom_url":null,"og_image_custom_fields":null,"og_video":"","og_custom_url":null,"og_article_section":"","og_article_tags":"","twitter_use_og":false,"twitter_card":"summary","twitter_image_type":"default","twitter_image_url":null,"twitter_image_custom_url":null,"twitter_image_custom_fields":null,"twitter_title":null,"twitter_description":null,"schema":{"blockGraphs":[],"customGraphs":[],"default":{"data":{"Article":[],"Course":[],"Dataset":[],"FAQPage":[],"Movie":[],"Person":[],"Product":[],"ProductReview":[],"Car":[],"Recipe":[],"Service":[],"SoftwareApplication":[],"WebPage":[]},"graphName":"Article","isEnabled":true},"graphs":[]},"schema_type":null,"schema_type_options":null,"pillar_content":false,"robots_default":true,"robots_noindex":false,"robots_noarchive":false,"robots_nosnippet":false,"robots_nofollow":false,"robots_noimageindex":false,"robots_noodp":false,"robots_notranslate":false,"robots_max_snippet":null,"robots_max_videopreview":null,"robots_max_imagepreview":"large","priority":null,"frequency":null,"local_seo":null,"limit_modified_date":false,"created":"2021-04-30 08:02:18","updated":"2024-02-29 08:55:45","focus_keyword":null,"additional_keywords":null,"truseo_locale":null,"ai":null,"breadcrumb_settings":null,"seo_analyzer_scan_date":null},"aioseo_breadcrumb":"<div class=\"aioseo-breadcrumbs\"><span class=\"aioseo-breadcrumb\">\n\t\t\t<a href=\"https:\/\/2thenew.today\/blog\" title=\"Home\">Home<\/a>\n\t\t<\/span><span class=\"aioseo-breadcrumb-separator\">&raquo;<\/span><span class=\"aioseo-breadcrumb\">\n\t\t\t<a href=\"https:\/\/2thenew.today\/blog\/category\/big-data\/\" title=\"Big Data\">Big Data<\/a>\n\t\t<\/span><span class=\"aioseo-breadcrumb-separator\">&raquo;<\/span><span class=\"aioseo-breadcrumb\">\n\t\t\tIntroduction To Hadoop\n\t\t<\/span><\/div>","aioseo_breadcrumb_json":[{"label":"Home","link":"https:\/\/2thenew.today\/blog"},{"label":"Big Data","link":"https:\/\/2thenew.today\/blog\/category\/big-data\/"},{"label":"Introduction To Hadoop","link":"https:\/\/2thenew.today\/blog\/introduction-to-hadoop\/"}],"_links":{"self":[{"href":"https:\/\/2thenew.today\/blog\/wp-json\/wp\/v2\/posts\/13071","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/2thenew.today\/blog\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/2thenew.today\/blog\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/2thenew.today\/blog\/wp-json\/wp\/v2\/users\/114"}],"replies":[{"embeddable":true,"href":"https:\/\/2thenew.today\/blog\/wp-json\/wp\/v2\/comments?post=13071"}],"version-history":[{"count":0,"href":"https:\/\/2thenew.today\/blog\/wp-json\/wp\/v2\/posts\/13071\/revisions"}],"wp:attachment":[{"href":"https:\/\/2thenew.today\/blog\/wp-json\/wp\/v2\/media?parent=13071"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/2thenew.today\/blog\/wp-json\/wp\/v2\/categories?post=13071"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/2thenew.today\/blog\/wp-json\/wp\/v2\/tags?post=13071"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}