Data Science Bootcamp (openHPI)

Data Science Bootcamp (openHPI)

The ultimate goal of the bootcamp is to cultivate strong data science skills with an emphasis on machine learning techniques to satisfactorily meet and exceed the requests of the Data science world. In the process, we will develop good habits for operating independently as data scientists and for operating as members of productive data science teams.

Why is the topic relevant? Why is it on everyone's mind?
Having data science and machine learning skills nowadays can potentially increase your success chances, whether that be as an individual or a business. Many industries offer their employees the opportunity to enroll in upskilling programs. In that way, domain experts can leverage the knowledge in their given field and seek higher roles in their company. As the demand for data science skills rises higher and higher, having a rounded understanding of data science and applying that knowledge practically can help widen your scope of knowledge.

Who should take this course?

  • People with basic python knowledge. That includes variables, conditional statements, while loops, and data structures.
  • People with domain knowledge that need to apply modern data analysis in their daily workload.

What will be taught in the course?

  • What is data science? why is it relevant?
  • Data Analysis and making sense of the data you have.
  • The use of libraries such as NumPy, Pandas, and Matplotlib for interesting data visualizations
  • Leveraging the power of ML

How will it be taught?

  • Videos
  • h5p
  • Quizzes
  • Two live streams

What needs to be accomplished?

  • Jupyter notebook Exercises
  • Weekly challenges
  • Exposure to real-life scenarios and datasets (no easy data)

How much time is expected to be spent?
The workload for the course is approximately 5 - 7 hours per week, depending on prior knowledge.

What you'll learn

  • What is Jupyter Notebooks and how to use it for Data Science
  • Work with real-life datasets and apply Numpy, Pandas and Matplotlib
  • Use scikit-learn to create powerful ML models

Who this course is for
High School and College Students
Domain Experts

Go to Class
MOOC List is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

Related Courses

AI for Efficient Programming: Harnessing the Power of LLMs (Coursera) Coursera
Fred Hutchinson Cancer Center

AI for Efficient Programming: Harnessing the Power of LLMs (Coursera)

This course on Artificial Intelligence (AI) for software development explores the use of AI large language models such as ChatGPT, Bard, and others and their potential benefits and challenges. Through examples and hands-on activities, you will develop an understanding of the ways in which AI can speed up software development tasks and free up time for more creative and strategic work.

Oct 5th 2026
4 Weeks
Data Engineering und Data Science – Klarheit in den Schlagwort-Dschungel (openHPI) OpenHPI
Hasso-Plattner-Institut

Data Engineering und Data Science – Klarheit in den Schlagwort-Dschungel (openHPI)

Die Schlagwörter Künstliche Intelligenz, Data Science, Data Engineering, und Big Data dominieren seit einigen Jahren nicht nur die IT-Schlagzeilen. In unserem Kurs wollen wir diese Wörter mit grundlegendem Inhalt füllen und die typischen Arbeitsschritte eines Data Scientists nachvollziehen. Insbesondere schauen wir hinter die Kulissen und betrachten den oft mühsamen Weg der Daten bis sie endlich genutzt werden können um z.B. mittels maschinellem Lernen Modelle trainieren zu können. Dazu gehören die Datenbeschaffung, die Datenreinigung, und die Datenintegration. Anschließend lernen wir, wie man aus diesen Daten und auch aus Texten neue Erkenntnisse mittels Data Mining und maschinellem Lernen gewinnt. Der Abschluss bildet eine Diskussion über Ethik und Fairness bei der automatisierten Datenanalyse.

Self Paced
Self-Paced
KI und Datenqualität - Perspektiven aus Data Science, Ethik, Normung und Recht (openHPI) OpenHPI
Hasso-Plattner-Institut

KI und Datenqualität - Perspektiven aus Data Science, Ethik, Normung und Recht (openHPI)

Ohne Daten gibt es keine Künstliche Intelligenz. Maschinelles Lernen benutzt große Datenmengen, um KI-Modelle zu trainieren. Eine der größten Herausforderungen beim Einsatz von gesellschaftlich verträglicher KI ist die Bereitstellung ausreichender, besonders aber qualitativ hochwertiger Trainingsdaten. In dem Kurs “KI und Datenqualität” berichten Expertinnen und Experten aus den Bereichen Informatik, Recht, Ethik und Normung über diese vielfältigen Aspekte der Daten für die Künstliche Intelligenz.

Apr 19th 2023
2 Weeks
Foundations of Data Science: K-Means Clustering in Python (Coursera) Coursera
University of London,Goldsmiths, University of London

Foundations of Data Science: K-Means Clustering in Python (Coursera)

This MOOC, designed by an academic team from Goldsmiths, University of London, will quickly introduce you to the core concepts of Data Science to prepare you for intermediate and advanced Data Science courses. It focuses on the basic mathematics, statistics and programming skills that are necessary for typical data analysis tasks.

Oct 5th 2026
5-12 Weeks
Functional Programming in Scala Capstone (Coursera) Coursera
École Polytechnique Fédérale de Lausanne

Functional Programming in Scala Capstone (Coursera)

In the final capstone project you will apply the skills you learned by building a large data-intensive application using real-world data. You will implement a complete application processing several gigabytes of data. This application will show interactive visualizations of the evolution of temperatures over time all over the world.

Oct 5th 2026
5-12 Weeks
Population Health: Responsible Data Analysis (Coursera) Coursera
Leiden University

Population Health: Responsible Data Analysis (Coursera)

In most areas of health, data is being used to make important decisions. As a health population manager, you will have the opportunity to use data to answer interesting questions. In this course, we will discuss data analysis from a responsible perspective, which will help you to extract useful information from data and enlarge your knowledge about specific aspects of interest of the population.

Oct 5th 2026
4 Weeks
Practical Time Series Analysis (Coursera) Coursera
The State University of New York

Practical Time Series Analysis (Coursera)

Many of us are "accidental" data analysts. We trained in the sciences, business, or engineering and then found ourselves confronted with data for which we have no formal analytic training. This course is designed for people with some technical competencies who would like more than a "cookbook" approach, but who still need to concentrate on the routine sorts of presentation and analysis that deepen the understanding of our professional topics.

Oct 5th 2026
5-12 Weeks
Preparing for the Google Cloud Professional Data Engineer Exam en Español (Coursera) Coursera
Google Cloud

Preparing for the Google Cloud Professional Data Engineer Exam en Español (Coursera)

En este curso, se emplea un enfoque descendente a fin de identificar las habilidades y los conocimientos adquiridos, así como poner en evidencia la información y las áreas de habilidades que requieren una preparación adicional. Puede aprovechar este curso para crear su propio plan de preparación personalizado. Lo ayudará a distinguir lo que sabe de lo que no. Además, le permitirá desarrollar y practicar las habilidades que se les exigen a los profesionales que realizan este trabajo.

Oct 5th 2026
1 Week
Data intelligence for businesses and managers (Coursera) Coursera
Institut Mines-Telecom

Data intelligence for businesses and managers (Coursera)

With the proliferation of connected objects (computers, tablets, watches, etc.), huge masses of data are generated every second. This Big Data has led to the emergence of a data economy, where data is the main source of competitive advantage for companies. In this sense, data and its processing tools have become a strategic priority for companies, and the main gas pedal of their digital transformation.

Oct 5th 2026
5-12 Weeks
Introduction to Bayesian Data Analysis (openHPI) OpenHPI
Hasso-Plattner-Institut

Introduction to Bayesian Data Analysis (openHPI)

Bayesian data analysis is increasingly becoming the tool of choice for many data-analysis problems. This free course on Bayesian data analysis will teach you basic ideas about random variables and probability distributions, Bayes' rule, and its application in simple data analysis problems. You will learn to use the R package brms (which is a front-end for the probabilistic programming language Stan). The focus will be on regression modeling, culminating in a brief introduction to hierarchical models (otherwise known as mixed or multilevel models). This course is appropriate for anyone familiar with the programming language R and for anyone who has done some frequentist data analysis (e.g., linear modeling and/or linear mixed modeling) in the past.

Jan 25th 2023
5-12 Weeks
A Step-by-Step Introduction to Process Mining (openHPI) OpenHPI
Hasso-Plattner-Institut

A Step-by-Step Introduction to Process Mining (openHPI)

Process mining is widely used in organizations to improve the understanding of business processes, based on data. Therefore, process mining is also called “data science for business processes”. While process mining has gone mainstream, there are many underlying concepts and techniques, and these are complex. The goal of this online course is to provide a general understanding of the concepts and techniques behind process mining. The course will be most valuable for domain experts, whose business processes are investigated, and for professionals in IT and in business consulting. We aim at providing a common understanding and a common language that facilitates communication between all stakeholders involved in process mining projects.

May 5th 2021
2 Weeks