Janssens, Jeroen

Data Science at the - Command Line: Obtain, - Scrub, Explore, and Model - Data with Unix Power Tools

(No reviews yet) Write a Review
ISBN 13:
9781492087915
author:
Janssens, Jeroen
format:
Paperback
publisher:
O'Reilly Media
language:
English
Publication Year:
2021
Pages:
280
Dimensions:
23.3 x 17.8 x 1.5 centimetres (0
Genre:
Computers, Operating Systems, Linux,
Condition:
New
Availability:
Item usually sent within 10 working days
£40.35

Description

Efficient data science starts at the command line. This comprehensive guide shows you how to harness the power of Unix tools to quickly obtain, scrub, explore, and model your data. With over 100 included tools, you'll be able to process data from various sources, including websites, APIs, databases, and spreadsheets. You'll learn how to perform essential operations such as text scrubbing, CSV and JSON file management, and data visualization. The book also covers managing your data science workflow, creating custom tools, and parallelizing data-intensive pipelines. Whether you're a data scientist, analyst, engineer, or system administrator, this guide will help you improve your productivity and efficiency. By combining command-line tools with Python, R, Jupyter, RStudio, and Apache Spark, you'll be able to model your data using dimensionality reduction, regression, and classification algorithms. With this book as your guide, you'll discover the agility, scalability, and extensibility of the command line technology.

View AllClose