Spark update brings R support and machine learning chops

One of the most popular big data processing platforms, Spark, now supports one of the premier statistical programming languages, R, which could pave the way for easier big data statistical analysis.

“R is the lingua franca of data scientists and its adoption has exploded in the last two years,” wrote Patrick Wendell, one of the chief contributors to Spark, in an email. Wendell is also a cofounder and software engineer at Databricks, which offers a commercial cloud-based version of Spark for enterprises.

The new version “will let R users work directly on large datasets, scaling to hundreds or thousands of machines, well beyond the limits of a stand-alone R program,” Wendell wrote.

To read this article in full or to leave a comment, please click here

Read more 0 Comments

Spark update brings R support and machine learning chops

One of the most popular big data processing platforms, Spark, now supports one of the premier statistical programming languages, R, which could pave the way for easier big data statistical analysis.

“R is the lingua franca of data scientists and its adoption has exploded in the last two years,” wrote Patrick Wendell, one of the chief contributors to Spark, in an email. Wendell is also a cofounder and software engineer at Databricks, which offers a commercial cloud-based version of Spark for enterprises.

The new version “will let R users work directly on large datasets, scaling to hundreds or thousands of machines, well beyond the limits of a stand-alone R program,” Wendell wrote.

To read this article in full or to leave a comment, please click here

Read more 0 Comments