## Bayesian and Frequentist Approaches: Ask the Right Question

May 6, 2013
It occurred to us recently that we don’t have any articles about Bayesian approaches to statistics here. I’m not going to get into the “Bayesian versus Frequentist” war; in my opinion, which style of approach to use is less about philosophy, and more about figuring out the best way to answer a question. Once you

## Incomplete Data by Design: Bringing Machine Learning to Marketing Research

May 6, 2013
Survey research deals with the problem of question wording by always asking the same question.  Thus, the Gallup Daily Tracking is filled with examples of moving averages for the exact same question asked precisely the same way every day. &nb...

## Creating a QGIS-Style (qml-file) with an R-Script

May 6, 2013
How to get from a txt-file with short names and labels to a QGIS-Style (qml-file)?
I used the below R-script to create a style for this legend table where I copy-pasted the parts I needed to a txt-file, like for the WRB-FULL (WRB-FULL: Full soil code of the STU from the World Reference Base for Soil Resources). The...

## The half variance approximation for mean returns

May 6, 2013
What’s that thing about arithmetic and geometric returns and the variance? Previously An introduction to the difference between simple and log returns is: A tale of two returns Issue Suppose you are predicting the mean annual return of an asset for some number of years.  To simplify the discussion, let’s buy into the fantasy that … Continue reading...

## analyze the social security administration public use microdata files (ssapumf) with r

May 5, 2013
the social security administration (ssa) must be overflowing with quiet heroes, because their public-use microdata files are as inconspicuous as they are thorough.  sure, ssa publishes enough great statistical research of their own that outside re...

## Google Analytics + R = FUN!

May 5, 2013
The scope of this post it to show how simple it is to get data out of the Google Analytics and create your own reports (that you hope that they can be semi-automated at least) and you favourite statistical graphs (those that GA is currently missing). As you already know R is a favourite tool ...read more

## R/Finance 2013 Is Coming Quickly…

May 5, 2013
There is about two weeks remaining until R/Finance 2013 - being held on May 17th and 18th at UIC in Chicago.  Make sure you register beforehand to ensure you have a spot, and – yes - you do want to come to the conference dinner on Friday.   I am particularly excited about the lineup of keynotes

## Simulation shows gain of clmm over ANOVA is small

May 5, 2013
After last post's setting up for a simulation, it is now time to look how the models compare. To my disappointment with my simple simulations of assessors behavior the gain is minimal. Unfortunately, the simulation took much more time than I ...

## Quandl Package – 5,000,000 free datasets at the tip of your fingers!

May 5, 2013
# Yes, you read that correctly and no Quandl (http://www.quandl.com/) did not pay me anything.# Quandl is a new database management tool which seeks to become the place to find datasets.  They boast of having over 5x10^6 data sets available t...

## Backporting R 3.0.0 to Quantal, Precise, and Lucid

May 4, 2013
Today (May 4, 2013) I will begin the process of backporting R 3.0.0 to Quantal, Precise, and Lucid. This will include all the recommended packages and the packages for R found in the universe repository for Ubuntu. Things to keep in mind: If you do...

## TV shows rated by episode as a Shiny App

May 3, 2013
A few days ago there was an interesting R based article by diffuseprior on the decline and fall in the quality of The Simpsons The author scraped results from GEOS, an online survey of TV programs, and applied the R package changepoint to offer an analysis of the show over time This seemed a candidate aaaa

## LaTeX in R graphs

May 3, 2013
A nice post was recently published on the rsnippets blog, about the tikzDevice R package. This package is – indeed – awesome. Even if it has been removed from the CRAN website. Of course, it can be download from the archive folder, on http://cran.r-project.org/…, but also (for a more recent version)  on http://download.r-forge.r-project.org/…. But first, it is necessary to install...

## Animation, from R to LaTeX

May 3, 2013
$X_{i,j}\sim\mathcal{B}(1/2)$

Just a short post, to share some codes used to generate animated graphs, with R. Assume that we would like to illustrate the law of large number, and the convergence of the average value from binomial sample. We can generate samples  using > n=200 > k=1000 > set.seed(1) > X=matrix(sample(0:1,size=n*k,replace=TRUE),n,k) Each row  will be a trajectory of heads and...

## Old Post with New d3 Life–GARCH and MA Performance

May 3, 2013
Parallel coordinates become much more useful when they are interactive, so I recreated one of my favorite blog posts "Trend is Not Your Friend" Applied to 48 Industries and convert the chart to a living breathing d3 parallel coordinates chart courtesy ...

## Extending RevoScaleR for Mining Big Data – Naive Bayes

May 3, 2013
by Derek McCrae Norton, Senior Sales Engineer In this third installment (following part 1 and part 2) of Extending RevoScaleR for Mining Big Data we look at how to use the building blocks provided by RevoScaleR to create a Naive Bayes model. Motivation: Fit a Naive Bayes model to big data. Naive Bayes is a simple probabilistic classifier based...

## All About Spherically Distributed Regression Errors

May 2, 2013
(This article was first published on Econometrics Beat: Dave Giles' Blog, and kindly contributed to R-bloggers) This post is based on a handout that I use for one of my courses, and it relates to the usual linear regression model,                                  y = Xβ...

## Improved R Profiling Summaries

May 2, 2013
In my last post I mentioned that I had improved on R’s `summaryRprof()` function with a custom function called `proftable()`. I’ve updated `proftable()` to take advantage of R 3.0.0’s ability to record line numbers while profiling. I’ve put it on github – you can get it there or below.

`proftable` reads in a file generated by...

## How R Grows

May 2, 2013
by Joseph Rickert Saturday morning I was drinking my coffee wondering how much effort goes into R worldwide. (It’s my job.) I noticed that there were 4469 packages on CRAN, and it occurred to me that tabulating the packages by publication date would give some indication of how much effort is being expended to improve packags and keep them...

## Changing The Presidential Election with R in the Browser

May 2, 2013
After I finished with the tutorial post d3 <- R with rCharts and slidify and then saw R creates d3/javascript charts in Ipython Style Notebook, a light clicked.  I could finally answer the lingering question I have had ever since I saw the NYT ...

## Concerto Simple Demonstration Test v4.b

May 2, 2013
# This is my first attempt to post a example of a test using Concerto 4 beta.

# See Simple Test

# This is an extremely simple test.  Too simple to be useful to most people yet I believe it is didactically helpful.

The Following is the HTML code for Template 9. It is the form that the user first inputs...

## Do Torontonians Want a New Casino? Survey Analysis Part 1

May 2, 2013
Toronto City Council is in the midst of a very lengthy process of considering whether or not to allow the OLG to build of a new casino in Toronto, and where.  The process started in November of 2012, and set … Continue reading

## Cut Dates Into Quarters

May 2, 2013
Frequently I need to recode a date column to quarters. For example, at Excelsior College we have continuous enrollment so we report new enrollments per quarter. To complicate things a bit, our fiscal year starts in July so that July, August, and September represent the first quarter, January, February, and March are actually the third quarter. But sometimes...

## Writing from R to Excel with xlsx

May 1, 2013
Paul Teetor, who is doing yeoman’s duty as one of the organizers of the Chicago R User Group (CRUG), asked recently if I would do a short presentation about a “favorite package”.  I picked xlsx, one of the many packages that provides a bridge between spreadsheets and R.  Here are the slides from my presentation

## NYT uses R to investigate NFL draft picks

May 1, 2013
Last week, the New York Times published online an interactive tool to explore NFL draft picks, revealing the fact that there's not much relationship between an early pick and the star performers in the season: Kevin Quealy, graphics editor at the NYT, detailed the process behind creating this graphic on his chartsnthings blog. He and others on the graphics...

## R for dummies

May 1, 2013
I already mentioned R for dummies a while ago on the ‘Og and never got around to read it from cover to back. Now that I am reduced to a dummy state with too much free time!, I can produce a full review of the book. R for dummies was written by two Belgian statistics

## R creates d3/javascript charts in Ipython Style Notebook

May 1, 2013
I am not sure I have ever done a post like this, but I was so blown away I had to do this post simply to embed this amazing Youtube video from the author of the R packages rCharts and slidify.  Watch this screencast as he creates d3/raphael charts...

## When the “reorder” function just isn’t good enough…

May 1, 2013
(This article was first published on Data and Analysis with R, at Work, and kindly contributed to R-bloggers) The reorder function, in R 3.0.0, is behaving strangely (or I’m really not understanding something).  Take the following simple data frame: df = data.frame(a1 = c(4,1,1,3,2,4,2), a2 = c(“h”,”j”,”j”,”e”,”c”,”h”,”c”)) I expect that if I call the reorder function on the a2...

## Book Review: The R Book, Second Edition (2013)

May 1, 2013
The first edition of The R Book by Michael J. Crawley was an ambitious work, but managed to be slightly rubbish due to the atrocious typographical layout of the original book. The good news is that the new 2nd edition, released in 2013, has a substanti...