This hands-on course covers tools and methods used by data scientists, from researching solutions to scaling prototypes on Spark clusters. Students engage with the full data engineering and data science pipeline, from data acquisition to extracting insight ...
Biology is becoming more and more a data science, as illustrated by the explosion of available genome sequences. This course aims to show how we can make sense of such data and harness it in order to understand biological processes in a quantitative way. ...
This class is an introduction to Machine Learning and High Dimensional Statistics in Finance. We start with purely empirical approach, focusing first on high dimensional regressions then moving to kernel methods and deep learning, and then study equilibriu ...
This course teaches the basic techniques, methodologies, and practical skills required to draw meaningful insights from a variety of data, with the help of the most acclaimed software tools in the data science world (pandas, scikit-learn, Spark, etc.) ...
Give students a feel for how single-cell genomics datasets are analyzed from raw data to data interpretation. Different steps of the analysis will be demonstrated and the most common statistical and bioinformatic techniques applied by the students. Data an ...
This course is intended for students who want to understand modern large-scale data analysis systems and database systems. It covers a wide range of topics and technologies, and will prepare students to be able to build such systems as well as read and und ...