Posts Tagged ‘ Big Data ’

Where to find data to use with R

October 11, 2011
By

(Contributing blogger Joe Rickert has put together a fantastic list of data sources suitable for use with R. If you're looking for data to use in the Applications of R Contest -- entries close October 31 -- this is a great resource for you -- Ed.) Hardly a day goes by without someone or something reminding me that we...

Read more »

Slides and replay for "Backtesting FINRA’s Limit Up/Down Rules" available

October 5, 2011
By

If you missed last week's webinar on using Revolution R and IBM Netezza to analyze the effectiveness of new rules intended to prevent another financial "Flash Crash", you can watch a replay by filling in this form. Once the replay begins, you can download the slides by clicking the "Download" button that appears below the media player. Revolution Analytics...

Read more »

Oracle’s Big Data Appliance to include R

October 3, 2011
By

At the Oracle OpenWorld conference in San Francisco today, Oracle announced the new Oracle Big Data Appliance, "a new engineered system that includes an open source distribution of Apache™ Hadoop™, Oracle NoSQL Database, Oracle Data Integrator Application Adapter for Hadoop, Oracle Loader for Hadoop, and an open source distribution of R." Oracle's foray into the Hadoop and NoSQL spaces...

Read more »

Revolution Analytics partners with Cloudera

September 26, 2011
By

Revolution Analytics today announced that it has partnered with Cloudera, the leader in Apache Hadoop-based software and services, to make big-data analytics with Hadoop and R available to Revolution R Enterprise users. As we announced earlier this month, we have created three open-source R packages which make it possible for R users to write map-reduce programs in the R...

Read more »

Are new SEC rules enough to prevent another Flash Crash?

September 22, 2011
By
Are new SEC rules enough to prevent another Flash Crash?

At 2:42PM on March 10 2010, without warning, the Dow Jones Industrial Index plunged more than 1000 points in just 5 minutes. It remains the biggest one-day decline in this stock market index in history. On an intra-day basis, anyway: by the end of the day, the market had regained 600 points of the drop. At the time, the...

Read more »

Slides and replay from "R and Hadoop" webinar

September 21, 2011
By
Slides and replay from "R and Hadoop" webinar

So ... there's clearly a lot of interest in integrating R and Hadoop. Today's webinar was a record-setter for Revolution Analytics, with more than 1000 people signing up to learn how to access Hadoop data from R with the packages from the open-source RHadoop project. If you didn't catch the live webinar, don't fret: the slides and replay are...

Read more »

How to extract time series from large timestamped logs with R

September 16, 2011
By

Revolution Analytics' Joe Rickert has a new post on inside-R.org, demonstrating how you can use R and the RevoScaleR package to extract time series data from time-stamped logs (in this case, the "US Domestic Flights From 1990 to 2009" dataset on Infochimps): Analyzing time series data of all sorts is a fundamental business analytics task to which the R...

Read more »

How to program MapReduce jobs in Hadoop with R

September 13, 2011
By

MapReduce is a powerful programming framework for efficiently processing very large amounts of data stored in the Hadoop distributed filesystem. But while several programming frameworks for Hadoop exist, few are tuned to the needs of data analysts who typically work in the R environment as opposed to general-purpose languages like Java. That's why the dev team at Revolution Analytics...

Read more »

Unlocking Big Data with R

September 9, 2011
By

I have an article out this week on ReadWriteHack: Unlocking Big Data with R. My thanks to the folks at ReadWriteWeb for giving us the opportunity to showcase some of the many real-world Big Data applications of R. Here are some additional links about the applications mentioned in the article: New York Times: Destruction of the Haiti earthquake; 2010...

Read more »

Analyzing big data in R: two presentations from useR! 2011

September 7, 2011
By

At last month's useR! 2011 conference at Warwick University, there were two talks on the RevoScaleR package for big data statistics in R. The first was a keynote presentation from Revolution Analytics' Chief Scientist, Lee Edlefsen. Here is the overview of his talk, Scalable Data Analysis in R: For the past several decades the rising tide of technology --...

Read more »