Data Science at the Command Line /

This hands-on guide demonstrates how the flexibility of the command line can help you become a more efficient and productive data scientist. You'll learn how to combine small, yet powerful, command-line tools to quickly obtain, scrub, explore, and model your data. To get you started-whether you...

Full description

Bibliographic Details
Main Author: Janssens, Jeroen (Author)
Corporate Author: Safari, an O'Reilly Media Company
Format: eBook
Language:English
Published: O'Reilly Media, Inc., 2014.
Edition:1st edition.
Subjects:
Online Access:Connect to this electronic resource

MARC

Tag First Indicator Second Indicator Subfields
LEADER 00000uam a2200000 a 4500
001 in00004099331
005 20260128224805.2
006 m o d
007 cr cn
008 240614s2014 xx o eng
020 |z 9781491947852 
035 |a (CaSebORM)9781491947845 
040 |d UtOrBLW 
041 0 |a eng 
100 1 |a Janssens, Jeroen,  |e author.  |0 http://id.loc.gov/authorities/names/no2002085513 
245 1 0 |a Data Science at the Command Line /  |c Janssens, Jeroen. 
250 |a 1st edition. 
264 1 |b O'Reilly Media, Inc.,  |c 2014. 
300 |a 1 online resource (210 pages) 
336 |a text  |b txt  |2 rdacontent 
337 |a computer  |b c  |2 rdamedia 
338 |a online resource  |b cr  |2 rdacarrier 
347 |a text file 
520 |a This hands-on guide demonstrates how the flexibility of the command line can help you become a more efficient and productive data scientist. You'll learn how to combine small, yet powerful, command-line tools to quickly obtain, scrub, explore, and model your data. To get you started-whether you're on Windows, OS X, or Linux-author Jeroen Janssens introduces the Data Science Toolbox, an easy-to-install virtual environment packed with over 80 command-line tools. Discover why the command line is an agile, scalable, and extensible technology. Even if you're already comfortable processing data with, say, Python or R, you'll greatly improve your data science workflow by also leveraging the power of the command line. Obtain data from websites, APIs, databases, and spreadsheets Perform scrub operations on plain text, CSV, HTML/XML, and JSON Explore data, compute descriptive statistics, and create visualizations Manage your data science workflow using Drake Create reusable tools from one-liners and existing Python or R code Parallelize and distribute data-intensive pipelines using GNU Parallel Model data with dimensionality reduction, clustering, regression, and classification algorithms 
533 |a Electronic reproduction.  |b Boston, MA :  |c Safari,  |n Available via World Wide Web. 
538 |a Mode of access: World Wide Web. 
542 |f Copyright © 2014 Jeroen H.M. Janssens 
588 |a Online resource; Title from title page (viewed October 2, 2014) 
500 |a Electronic resource. 
655 7 |a Electronic books.  |2 local 
710 2 |a Safari, an O'Reilly Media Company. 
856 4 0 |u https://proxy.library.tamu.edu/login?url=https://go.oreilly.com/TAMU/library/view/-/9781491947845/?ar  |z Connect to this electronic resource  |t 0 
999 f f |s 3275b8e6-4080-3c86-b0a6-aa9faa514145  |i e56f6ac8-2405-311b-870a-9bbc394bf2f2  |t 0 
952 f f |a Texas A&M University  |b College Station  |c Electronic Resources  |s www_evans  |d Available Online  |t 0  |h No information provided 
998 f f |t 0  |l Available Online