FUN

Exploratory Multivariate Data Analysis (FUN)

Offered by Agrocampus Ouest,
Exploratory Multivariate Data Analysis (FUN)

Exploratory multivariate data analysis is studied and teached in a French-way since a long time in France. This course focuses on four essential and basic methods, those with the largest potential in terms of applications: principal component analysis (PCA) when variables are quantitative, correspondence analysis (CA) and multiple correspondence analysis (MCA) when variables are categorical and clustering. This course has been designed for scientists whose aim is not to become statisticians but who feel the need to analyze the data themselves. It is therefore addressed to practitioners who are confronted with the analysis of data in marketing, surveys, ecology, biology, geography, etc.

This course is application-oriented; formalism and mathematics writing have been reduced as much as possible while examples and intuition have been emphasized and the numerous exercises done with FactoMineR (a package of the free R software) will make the participant efficient and reliable face to data analysis.

We hope that with this course, the participant will be fully equipped (theory, examples, software) to confront multivariate real-life data.
What you will learn
At the end of this course, you will be able to:

  • résumer et synthétiser des tableaux de données par des graphes simples ;
  • utiliser des méthodes de visualisation adaptées à l'analyse exploratoire multidimensionnelle ;
  • interpréter les résultats d'une analyse factorielle et d'une classification ;
  • reconnaître, par rapport à la problématique et aux données, la méthode adaptée à l'exploration d'un jeu de données selon la nature et la structure des variables ;
  • analyser les réponses à une enquête ;
  • mettre en oeuvre une méthode d'analyse de données textuelles
  • mettre en oeuvre les méthodes factorielles et de classification sur le logiciel gratuit R

En résumé, vous serez autonome sur la mise en œuvre et l'interprétation d'analyses exploratoires multidimensionnelles.

To whom is this course addressed?
This course will be held in English. It has been designed for scientists whose aim is not to become statisticians but who feel the need to analyze the data themselves. It is therefore addressed to practitioners who are confronted with the analysis of data in marketing, surveys, ecology, biology, geography, etc.
Suggested Readings: Exploratory Multivariate Analysis by Example Using R (Chapman & Hall/CRC Computer Science & Data Analysis)

Course Schedule

Week 1. Principal Component Analysis

  • Data - Practicalities
  • Studying individuals and variables
  • Aids for interpretation
  • PCA in practice using FactoMineR

Week 2. Correspondence Analysis

  • Data - introduction and independence model
  • Visualizing the row and column clouds
  • Inertia and percentage of inertia
  • Simultaneous representation
  • Interpretation aids
  • Correspondance Analysis in practice using FactoMineR

Week 3. Multiple Correspondence Analysis

  • Data - issues
  • Visualizing the point cloud of individuals
  • Visualizing the point cloud of categories - simultaneous representation
  • Interpretation aids
  • Multiple Correspondance Analysis in practice using FactoMineR

Week 4. Clustering

  • Hierarchical clustering
  • An example, and choosing the number of classes
  • Partitioning methods and other details
  • Characterizing the classes
  • Clustering in practice using FactoMineR

Week 5 : Multiple Factor Analysis

  • Data - issues
  • Balancing groups and choosing a weighting for the variables
  • Studying and visualizing the groups of variables
  • Visualizing the partial points
  • Visualizing the separate analyses
  • Taking into account groups of categorical variables
  • Taking into account contingency tables
  • Interpretation aids
  • Multiple Factor Analysis in practice using FactoMineR
Go to Class
MOOC List is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

Related Courses

Microsoft Azure Machine Learning (Coursera) Coursera
Microsoft

Microsoft Azure Machine Learning (Coursera)

Machine learning is at the core of artificial intelligence, and many modern applications and services depend on predictive machine learning models. Training a machine learning model is an iterative process that requires time and compute resources. Automated machine learning can help make it easier. In this course, you will learn how to use Azure Machine Learning to create and publish models without writing code.

Sep 21st 2026
4 Weeks
Statistical Inference and Modeling for High-throughput Experiments (edX) EdX
HarvardX,Harvard University

Statistical Inference and Modeling for High-throughput Experiments (edX)

A focus on the techniques commonly used to perform statistical inference on high throughput data. In this course you’ll learn various statistics topics including multiple testing problem, error rates, error rate controlling procedures, false discovery rates, q-values and exploratory data analysis. We then introduce statistical modeling and how it is applied to high-throughput data. In particular, we will discuss parametric distributions, including binomial, exponential, and gamma, and describe maximum likelihood estimation.

Self Paced
Self-Paced
High-Dimensional Data Analysis (edX) EdX
HarvardX,Harvard University

High-Dimensional Data Analysis (edX)

A focus on several techniques that are widely used in the analysis of high-dimensional data. If you’re interested in data analysis and interpretation, then this is the data science course for you. We start by learning the mathematical definition of distance and use this to motivate the use of the singular value decomposition (SVD) for dimension reduction and multi-dimensional scaling and its connection to principle component analysis.

Self Paced
Self-Paced
Introduction to Probability and Data with R (Coursera) Coursera
Duke University

Introduction to Probability and Data with R (Coursera)

This course introduces you to sampling and exploring data, as well as basic probability theory and Bayes' rule. You will examine various types of sampling methods, and discuss how such methods can impact the scope of inference. A variety of exploratory data analysis techniques will be covered, including numeric summary statistics and basic data visualization.

Sep 14th 2026
5-12 Weeks
Statistical Thinking for Industrial Problem Solving, presented by JMP (Coursera) Coursera
SAS

Statistical Thinking for Industrial Problem Solving, presented by JMP (Coursera)

Statistical Thinking for Industrial Problem Solving is an applied statistics course for scientists and engineers offered by JMP, a division of SAS. By completing this course, students will understand the importance of statistical thinking, and will be able to use data and basic statistical methods to solve many real-world problems.

Sep 21st 2026
5-12 Weeks
Exploratory Data Analysis with MATLAB (Coursera) Coursera
MathWorks

Exploratory Data Analysis with MATLAB (Coursera)

In this course, you will learn to think like a data scientist and ask questions of your data. You will use interactive features in MATLAB to extract subsets of data and to compute statistics on groups of related data. You will learn to use MATLAB to automatically generate code so you can learn syntax as you explore. You will also use interactive documents, called live scripts, to capture the steps of your analysis, communicate the results, and provide interactive controls allowing others to experiment by selecting groups of data.

Sep 21st 2026
5-12 Weeks
Exploratory Data Analysis for Machine Learning (Coursera) Coursera
IBM

Exploratory Data Analysis for Machine Learning (Coursera)

This first course in the IBM Machine Learning Professional Certificate introduces you to Machine Learning and the content of the professional certificate. In this course you will realize the importance of good, quality data. You will learn common techniques to retrieve your data, clean it, apply feature engineering, and have it ready for preliminary analysis and hypothesis testing.

Sep 21st 2026
2 Weeks
Healthcare Information Design and Visualizations (Coursera) Coursera
Northeastern University

Healthcare Information Design and Visualizations (Coursera)

Introduces processes and design principles for creating meaningful displays of information that support effective business decision-making. Studies how to collect and process data; create visualizations (both static and interactive); and use them to provide insight into a problem, situation, or opportunity. Introduces methods to critique visualizations along with ways to answer the elusive question: “What makes a visualization effective?”

Oct 5th 2026
4 Weeks
Create Machine Learning Models in Microsoft Azure (Coursera) Coursera
Microsoft

Create Machine Learning Models in Microsoft Azure (Coursera)

Machine learning is the foundation for predictive modeling and artificial intelligence. If you want to learn about both the underlying concepts and how to get into building models with the most common machine learning tools this path is for you. In this course, you will learn the core principles of machine learning and how to use common tools and frameworks to train, evaluate, and use machine learning models. This course is designed to prepare you for roles that include planning and creating a suitable working environment for data science workloads on Azure.

Sep 21st 2026
3 Weeks
AI and Public Health (Coursera) Coursera
DeepLearning.AI

AI and Public Health (Coursera)

The AI and Public Health course is designed to introduce learners to the concept of using artificial intelligence to address social and environmental issues. We'll start by defining AI for Good, and explore various examples of AI for Good projects. You will get hands-on with a real-world concern, air quality in Bogotá, and use the AI for Good framework to define the problem, identify stakeholders, and determine where AI could fit.

Sep 21st 2026
3 Weeks