Data Analytics Using Open-Source Tools

Data Analytics Using Open-Source Tools

VonJeffrey Strickland

Normalerweise in 3-5 Werktagen gedruckt
This book is about Data Analytics. In that respect, it is like others. What distinguishes it from the rest is the variety of open-source tool applications. This book incorporates the use of R Studio, Python, SAS Studio (University Edition), and KNIME. This book is also about manipulating Big Data. Apache Hadoop on Hortonworks Sandbox is introduced and we manage, move, handle, and transform data using Apache Hive, Apache Spark, MapReduce and TEZ, with terminal shell commands and Ambari. We show you how to set up a virtual machine in Microsoft Azure. We then use the data in later chapters for modeling. We cover Descriptive Modeling and Predictive. The content includes Support Vector Machines, Decision Tree learning, Random Forests, Naïve and Empirical Bayes, Gradient Boosting, Cluster Modeling, Generalized Linear Models, Logistic Regression, and Artificial Neural Networks. Every chapter includes completely worked examples using one or more open-source tools.

Details

Veröffentlicht am
Jul 20, 2016
Sprache
English
ISBN
9781365270413
Kategorie
Business & Wirtschaft
Copyright
Alle Rechte vorbehalten - Standard-Urheberrechtslizenz
Autoren/Mitwirkende
Von (Autor): Jeffrey Strickland

Spezifikationen

Seiten
706
Bindung
Paperback Paperback
Farbe für den Innenteil des Buches
schwarz & weiß
Abmessungen
US Trade (6 x 9 Zoll / 152 x 229 mm)

Bewertungen & Rezensionen