Welcome!

Open Source Authors: RealWire News Distribution, Bernard Golden, Harald Zeitlhofer, Yeshim Deniz, Pat Romanski

Related Topics: Cloud Expo, Open Source

Cloud Expo: Blog Feed Post

Big Data Trees with Hadoop HDFS

I've embedded the slides below, and you can also watch the webinar recording on YouTube

Last month's release of Revolution R Enterprise 6.1 added the capability to fit decision and regression trees on large data sets (using a new parallel external memory algorithm included in the RevoScaleR package).

It also introduced the possibility of applying this and the other big-data statistical methods of RevoScaleR to data files distributed in in Hadoop's HDFS file system*, using the Hadoop nodes themselves as the compute engine (with Revolution R Enterprise installed).


Revolution Analytics' VP of Development Sue Ranney explained how this works in a recent webinar. I've embedded the slides below, and you can also watch the webinar recording on YouTube.

 

[*] Or to use the department of redundancy department-approved acronym, HHFDSFS

Read the original blog entry...

More Stories By David Smith

David Smith is Vice President of Marketing and Community at Revolution Analytics. He has a long history with the R and statistics communities. After graduating with a degree in Statistics from the University of Adelaide, South Australia, he spent four years researching statistical methodology at Lancaster University in the United Kingdom, where he also developed a number of packages for the S-PLUS statistical modeling environment. He continued his association with S-PLUS at Insightful (now TIBCO Spotfire) overseeing the product management of S-PLUS and other statistical and data mining products.<

David smith is the co-author (with Bill Venables) of the popular tutorial manual, An Introduction to R, and one of the originating developers of the ESS: Emacs Speaks Statistics project. Today, he leads marketing for REvolution R, supports R communities worldwide, and is responsible for the Revolutions blog. Prior to joining Revolution Analytics, he served as vice president of product management at Zynchros, Inc. Follow him on twitter at @RevoDavid

Comments (0)

Share your thoughts on this story.

Add your comment
You must be signed in to add a comment. Sign-in | Register

In accordance with our Comment Policy, we encourage comments that are on topic, relevant and to-the-point. We will remove comments that include profanity, personal attacks, racial slurs, threats of violence, or other inappropriate material that violates our Terms and Conditions, and will block users who make repeated violations. We ask all readers to expect diversity of opinion and to treat one another with dignity and respect.