1712 search results for "regression"

Data Science Labs: Predictive Models to Improve Vaccine Quality and Production

June 20, 2013
By
Data Science Labs: Predictive Models to Improve Vaccine Quality and Production

The age of "blockbuster drugs" is coming to an end, as personalized medicine becomes a reality. Data science will be a major driver of innovation in these and other areas of the pharmaceutical industry. This was demonstrated during a project the Data Science Labs team executed on with a major pharmaceuticals company.

Read more »

A Toy Instrumental Variable Application

June 19, 2013
By
A Toy Instrumental Variable Application

An R package for Smith-Wilson yield curves

June 19, 2013
By
An R package for Smith-Wilson yield curves

Yield Curve fitting - the Smith-Wilson method Yield Curve fitting - the Smith-Wilson method This article illustrates the R package SmithWilsonYieldCurve, and provides some additional background on yield curve fitting. The method implemented in the package fits a curve to interest rate market...

Read more »

Oracle R Connector for Hadoop 2.1.0 released

June 17, 2013
By

(This article was first published on Oracle R Enterprise, and kindly contributed to R-bloggers) Oracle R Connector for Hadoop (ORCH), a collection of R packages that enables Big Data analytics using HDFS, Hive, and Oracle Database from a local R environment, continues to make advancements. ORCH 2.1.0 is now available, providing a flexible framework while remarkably improving performance and...

Read more »

Top 100 R packages for 2013 (Jan-May)!

June 13, 2013
By
Top 100 R packages for 2013 (Jan-May)!

(This article was first published on R-statistics blog » RR-statistics blog, and kindly contributed to R-bloggers) What are the top 100 (most downloaded) R packages in 2013? Thanks to the recent release of RStudio of their “0-cloud” CRAN log files (but without including downloads from the primary CRAN mirror or any of the 88 other CRAN mirrors), we can now answer this question...

Read more »

Data imputation I

June 12, 2013
By

I recently entered kaggle titanic learning competition for fun and to see where my out of the box utilization of random forest would rank me (303 out of 5,882). It was interesting to see that much of the scoring differentiation came from score imputation, that is filling missing values based on other data. For example, we might have

Read more »

Using Quandl in R

June 12, 2013
By
Using Quandl in R

Image by Jan Zander Our mantra here at Quandl is making data easy to find and easy to use. Following that goal we (and subsequently the community) have created packages that integrate Quandl’s API into a number of software platforms. Today we’ll take a look at R. R is a free statistical computing language created

Read more »

Sobol Sensitivity Analysis

June 10, 2013
By
Sobol Sensitivity Analysis

Sensitivity analysis is the task of evaluating the sensitivity of a model output Y to input variables (X1,…,Xp). Quite often, it is assumed that this output is related to the input through a known function f :Y= f(X1,…,Xp). Sobol indices are generalizing the coefficient of the coefficient of determination in regression. The ith first order indice is the proportion of...

Read more »

At what sample size do correlations stabilize?

June 6, 2013
By
At what sample size do correlations stabilize?

Maybe you have encountered this situation: you run a large-scale study over the internet, and out of curiosity, you frequently check the correlation between two variables. My experience with this practice is usually frustrating, as in small sample sizes (and we will see what “small” means in this context) correlations go up and down, change sign,

Read more »

The Frisch–Waugh–Lovell Theorem for Both OLS and 2SLS

June 5, 2013
By
The Frisch–Waugh–Lovell Theorem for Both OLS and 2SLS

The Frisch–Waugh–Lovell (FWL) theorem is of great practical importance for econometrics. FWL establishes that it is possible to re-specify a linear regression model in terms of orthogonal complements. In other words, it permits econometricians to partial out right-hand-side, or control, variables. This is useful in a variety of settings. For example, there may be cases

Read more »