{"id":5922,"date":"2015-12-21T07:00:19","date_gmt":"2015-12-21T15:00:19","guid":{"rendered":"http:\/\/formtek.com\/blog\/?p=5922"},"modified":"2015-11-06T08:24:04","modified_gmt":"2015-11-06T16:24:04","slug":"big-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets","status":"publish","type":"post","link":"https:\/\/formtek.com\/blog\/big-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets\/","title":{"rendered":"Big Data and Nomadic Computing: Speeding up Analytics of Complex Data Sets"},"content":{"rendered":"<p>Researchers at The University of Texas at Austin and the Texas Advanced Computing Center (TACC) are developing ways to speed up analytics processing of massive data sets. \u00a0One of the techniques developed by the university team is called nomadic computing.<\/p>\n<p>Nomadic computing has nothing to do with\u00a0bedouins or gypsies. \u00a0Actually, the tool developed by the researchers is\u00a0called NOMAD, which stands for &#8220;non-locking, stochastic multi-machine algorithm for asynchronous and decentralized matrix completion.&#8221; \u00a0Data analysis using NOMAD is significantly faster than by using other techniques that are considered to be\u00a0state-of-the-art. \u00a0The tool enables analysis on data sets so large that most other analytic tools are just not able to handle them.<\/p>\n<p>The NOMAD technique involves simplifying the complexity of data sets. \u00a0The idea is that if data comes with many parameters or dimensions that it can be simplified by identifying\u00a0just the subset of the dimensions that are the most meaningful.<\/p>\n<p><a href=\"https:\/\/www.cs.utexas.edu\/~inderjit\/\" target=\"_blank\">Inderjit Dhillon<\/a>, a professor of computer science at The University of Texas at Austin,<a href=\"http:\/\/phys.org\/news\/2015-11-nomadic-big-analytics.html#jCp\" target=\"_blank\"> said that<\/a> &#8220;suppose you have a massive computational problem and you need to run it on datasets that do not fit in a computer&#8217;s memory. \u00a0If you want the answer in a reasonable amount of time, the logical thing to do would be to distribute the computations over different machines. We are trying to develop an asynchronous method where each parameter is, in a sense, a nomad. \u00a0The parameters go to different processors, but instead of synchronizing this computation followed by communication, the nomadic framework does its work whenever a variable is available at a particular processor.&#8221;<\/p>\n<p><a href=\"http:\/\/people.cs.clemson.edu\/~aapon\/\" target=\"_blank\">Amy Apon<\/a>, a program director at NSF, <a href=\"http:\/\/www.infozine.com\/news\/stories\/op\/storiesView\/sid\/63143\/\" target=\"_blank\">said that<\/a>\u00a0&#8220;traditionally, machine learning inference algorithms run on a single large&#8211;and sometimes expensive&#8211;server, and this limits the size of the problem that can be addressed. \u00a0This team has noticed a property of some machine learning algorithms that if a few slowly changing variables can be only occasionally synchronized, then the work can be more easily distributed across different computers. \u00a0Their clever mathematical approach is opening doors to running machine algorithms on the kind of massive-scale, distributed, commodity computers that we find in today&#8217;s cloud computing environment.&#8221;<\/p>\n<p>&nbsp;<\/p>\n<div class=\"lightsocial_container\"><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/digg.com\/submit?url=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F&amp;title=\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/digg.png\" alt=\"Digg This\" title=\"Digg This\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/www.reddit.com\/submit?url=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F&amp;title=\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/reddit.png\" alt=\"Reddit This\" title=\"Reddit This\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/www.stumbleupon.com\/submit?url=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F&amp;title=\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/stumbleupon.png\" alt=\"Stumble Now!\" title=\"Stumble Now!\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/buzz.yahoo.com\/buzz?targetUrl=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F&amp;headline=\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/yahoo_buzz.png\" alt=\"Buzz This\" title=\"Buzz This\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/www.dzone.com\/links\/add.html?title=&amp;url=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/dzone.png\" alt=\"Vote on DZone\" title=\"Vote on DZone\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/www.facebook.com\/sharer.php?t=&amp;u=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/facebook.png\" alt=\"Share on Facebook\" title=\"Share on Facebook\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/delicious.com\/save?title=&amp;url=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/delicious.png\" alt=\"Bookmark this on Delicious\" title=\"Bookmark this on Delicious\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/www.dotnetkicks.com\/kick\/?title=&amp;url=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/dotnetkicks.png\" alt=\"Kick It on DotNetKicks.com\" title=\"Kick It on DotNetKicks.com\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/dotnetshoutout.com\/Submit?title=&amp;url=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/dotnetshoutout.png\" alt=\"Shout it\" title=\"Shout it\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/www.linkedin.com\/shareArticle?mini=true&amp;url=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F&amp;title=&amp;summary=&amp;source=\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/linkedin.png\" alt=\"Share on LinkedIn\" title=\"Share on LinkedIn\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/www.technorati.com\/faves?add=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/technorati.png\" alt=\"Bookmark this on Technorati\" title=\"Bookmark this on Technorati\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/twitter.com\/home?status=Reading+https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/twitter.png\" alt=\"Post on Twitter\" title=\"Post on Twitter\" \/><\/a><\/div><div class=\"lightsocial_element\"><a class=\"lightsocial_a\" href=\"http:\/\/www.google.com\/buzz\/post?url=https%3A%2F%2Fformtek.com%2Fblog%2Fbig-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets%2F\" target=\"_blank\"><img decoding=\"async\" class=\"lightsocial_img\" src=\"https:\/\/formtek.com\/blog\/wp-content\/plugins\/light-social\/google_buzz.png\" alt=\"Google Buzz (aka. Google Reader)\" title=\"Google Buzz (aka. Google Reader)\" \/><\/a><\/div><\/div>","protected":false},"excerpt":{"rendered":"<p>Researchers at The University of Texas at Austin and the Texas Advanced Computing Center (TACC) are developing ways to speed up analytics processing of massive data sets. \u00a0One of the techniques developed by the university team is called nomadic computing.<span class=\"ellipsis\">&hellip;<\/span><\/p>\n<div class=\"read-more\"><a href=\"https:\/\/formtek.com\/blog\/big-data-and-nomadic-computing-speeding-up-analytics-of-complex-data-sets\/\">Read more &#8250;<\/a><\/div>\n<p><!-- end of .read-more --><\/p>\n","protected":false},"author":2,"featured_media":0,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[3,58],"tags":[],"class_list":["post-5922","post","type-post","status-publish","format-standard","hentry","category-big-data","category-data-analytics"],"_links":{"self":[{"href":"https:\/\/formtek.com\/blog\/wp-json\/wp\/v2\/posts\/5922","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/formtek.com\/blog\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/formtek.com\/blog\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/formtek.com\/blog\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/formtek.com\/blog\/wp-json\/wp\/v2\/comments?post=5922"}],"version-history":[{"count":0,"href":"https:\/\/formtek.com\/blog\/wp-json\/wp\/v2\/posts\/5922\/revisions"}],"wp:attachment":[{"href":"https:\/\/formtek.com\/blog\/wp-json\/wp\/v2\/media?parent=5922"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/formtek.com\/blog\/wp-json\/wp\/v2\/categories?post=5922"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/formtek.com\/blog\/wp-json\/wp\/v2\/tags?post=5922"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}