Bioinformatic Methods I (Coursera)

Offered by University of Toronto,
Bioinformatic Methods I (Coursera)

Large-scale biology projects such as the sequencing of the human genome and gene expression surveys using RNA-seq, microarrays and other technologies have created a wealth of data for biologists. However, the challenge facing scientists is analyzing and even accessing these data to extract useful information pertaining to the system being studied. This course focuses on employing existing bioinformatic resources – mainly web-based programs and databases – to access the wealth of data to answer questions relevant to the average biologist, and is highly hands-on.

Class Deals by MOOC List - Click here and see Coursera's Active Discounts, Deals, and Promo Codes.

Topics covered include multiple sequence alignments, phylogenetics, gene expression data analysis, and protein interaction networks, in two separate parts.
The first part, Bioinformatic Methods I (this one), deals with databases, Blast, multiple sequence alignments, phylogenetics, selection analysis and metagenomics.
The second part, Bioinformatic Methods II, covers motif searching, protein-protein interactions, structural bioinformatics, gene expression data analysis, and cis-element predictions.
This pair of courses is useful to any student considering graduate school in the biological sciences, as well as students considering molecular medicine. Both provide an overview of the many different bioinformatic tools that are out there.
These courses are based on one taught at the University of Toronto to upper-level undergraduates who have some understanding of basic molecular biology. No programming is required for this course.
Course 1 of 4 in the Plant Bioinformatic Methods Specialization.

Syllabus

WEEK 1
NCBI/Blast I
In this module we'll be exploring the amazing resources available at NCBI, the National Centre for Biotechnology Information, run by the National Library of Medicine in the USA. We'll also be doing a Blast search to find similar sequences in the enormous NR sequence database. We can use similar sequences to infer homology, which is the primary predictor of gene or protein function.

WEEK 2
Blast II/Comparative Genomics
In this module we'll continue exploring the incredible resources available at NCBI, the National Centre for Biotechnology Information. We will be performing several different kinds of Blast searches: BlastP, PSI-Blast, and Translated Blast. We can use similar sequences identified by such methods to infer homology, which is the primary predictor of gene or protein function. We'll also be comparing parts of the genomes of a couple of different species, to see how similar they are.

WEEK 3
Multiple Sequence Alignments
In this module we'll be doing multiple sequence alignments with Clustal (as implemented in MEGA), DiAlign, and MAFFT. Multiple sequences alignments can tell you where in a sequence the conserved and variable regions are, which is important for understanding the biology of the sequences under investigation. It also has practical applications, such as being able to design PCR primers that will amplify sequences from a number of different species, for example.

WEEK 4
Review: NCBI/Blast I, Blast II/Comparative Genetics, and Multiple Sequence Alignments

WEEK 5
Phylogenetics
In this module we'll be using the multiple sequence alignments we generated last lab to do some phylogenetic analyses with both neighbour-joining and maximum likelihood methods. The tree-like structure generated by such analyses tells us how closely sequences are related one to another, and suggests when in evolutionary time a speciation or gene duplication event occurred.

WEEK 6
Selection Analysis
In this module we'll take a set of orthologous sequences from bacteria and use DataMonkey to analyze them for the presence of certain sites under positive, negative or neutral selection. Such an analysis can help understand the biology of a set of protein coding sequences by identifying residues that might be important for biological function (those residues under negative selection) or those that might be involved in response to external influences, such as drugs, pathogens or other factors (residues under positive selection).

WEEK 7
'Next Gen' Sequence Analysis (RNA-Seq) / Metagenomics
In this module we'll explore some of the data that have been generated as a result of the rapid decrease in the cost of sequencing DNA. We'll be exploring a couple of RNA-Seq data sets that can tell us where any given gene is expressed, and also how that gene might be alternatively spliced. We'll also be looking at a couple of metagenome data sets that can tell us about the kinds of species (especially microbial species that might otherwise be hard to culture) that are in a given environmental niche.

WEEK 8
Review: Phylogenetics, Selection Analysis, and 'Next Gen' Sequence Analysis (RNA-seq)/Metagenomics + Final Assignment

Go to Class
MOOC List is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

Related Courses

Data Science Ethics (Coursera) Coursera
University of Michigan

Data Science Ethics (Coursera)

What are the ethical considerations regarding the privacy and control of consumer information and big data, especially in the aftermath of recent large-scale data breaches? This course provides a framework to analyze these concerns as you examine the ethical and privacy implications of collecting and managing big data. Explore the broader impact of the data science field on modern society and the principles of fairness, accountability and transparency as you gain a deeper understanding of the importance of a shared set of ethical values.

Sep 21st 2026
4 Weeks
SQL: A Practical Introduction for Querying Databases (Coursera) Coursera
IBM

SQL: A Practical Introduction for Querying Databases (Coursera)

Much of the world's data lives in databases. SQL (or Structured Query Language) is a powerful programming language that is used for communicating with and manipulating data in databases. A working knowledge of databases and SQL is a must for anyone who wants to start a career in Data Engineering, Data Warehousing, Data Analytics, Data Science or Business Intelligence. The purpose of this course is to help you learn and apply foundational and intermediate knowledge of the SQL language, and become familiar with many relational database (RDBMS) concepts along the way.

Sep 21st 2026
5-12 Weeks
Введение в биоинформатику (Introduction to Bioinformatics) (Coursera) Coursera
Saint Petersburg State University

Введение в биоинформатику (Introduction to Bioinformatics) (Coursera)

Курс «Введение в биоинформатику» адресован тем, кто хочет получить расширенное представление о том, что такое биоинформатика и как она помогает биологам и медикам в их работе. The course is aimed at those who would like to have a better idea of what bioinformatics is and how it helps biologists and medical scientists in research and clinical work.

Sep 21st 2026
5-12 Weeks
Big Data Science with the BD2K-LINCS Data Coordination and Integration Center (Coursera) Coursera
Icahn School of Medicine at Mount Sinai

Big Data Science with the BD2K-LINCS Data Coordination and Integration Center (Coursera)

In this course we briefly introduce the DCIC and the various Centers that collect data for LINCS. We then cover metadata and how metadata is linked to ontologies. We then present data processing and normalization methods to clean and harmonize LINCS data. This follow discussions about how data is served as RESTful APIs. Most importantly, the course covers computational methods including: data clustering, gene-set enrichment analysis, interactive data visualization, and supervised learning. Finally, we introduce crowdsourcing/citizen-science projects where students can work together in teams to extract expression signatures from public databases and then query such collections of signatures against LINCS data for predicting small molecules as potential therapeutics.

Sep 21st 2026
5-12 Weeks
Data Science for Business Innovation (Coursera) Coursera
Politecnico di Milano,EIT Digital

Data Science for Business Innovation (Coursera)

The course is a compendium of the must-have expertise in data science for executive and middle-management to foster data-driven innovation. It consists of introductory lectures spanning big data, machine learning, data valorization and communication. Topics cover the essential concepts and intuitions on data needs, data analysis, machine learning methods, respective pros and cons, and practical applicability issues.

Sep 21st 2026
4 Weeks
Lactation Biology (Coursera) Coursera
University of Illinois at Urbana-Champaign

Lactation Biology (Coursera)

Lactation and especially milk, which is the product of that unique mammalian process, are routinely encountered within our daily lives. Nevertheless, they often are poorly understood by many, even including many who are engaged in the business of producing milk. The overall course goal is to introduce fundamental concepts that form the basis for understanding the biology of lactation, the biology of the mammary gland, and the products of that important physiological process.

Sep 21st 2026
5-12 Weeks
Hacking COVID-19 — Course 2: Decoding SARS-CoV-2's Secrets (Coursera) Coursera
University of California, San Diego

Hacking COVID-19 — Course 2: Decoding SARS-CoV-2's Secrets (Coursera)

In this course, you will follow in the footsteps of the bioinformaticians investigating the COVID-19 outbreak by annotating the SARS-CoV-2 genome and using the annotation to design a COVID-19 diagnostic test. Whether you’re new to the world of computational biology, or you’re a bioinformatics expert seeking to learn about its applications in the COVID-19 pandemic, or somewhere in between, this course is for you!

Sep 21st 2026
2 Weeks
Combining and Analyzing Complex Data (Coursera) Coursera
University of Maryland, College Park

Combining and Analyzing Complex Data (Coursera)

In this course you will learn how to use survey weights to estimate descriptive statistics, like means and totals, and more complicated quantities like model parameters for linear and logistic regressions. Software capabilities will be covered with R® receiving particular emphasis. The course will also cover the basics of record linkage and statistical matching—both of which are becoming more important as ways of combining data from different sources. Combining of datasets raises ethical issues which the course reviews. Informed consent may have to be obtained from persons to allow their data to be linked. You will learn about differences in the legal requirements in different countries.

Sep 21st 2026
4 Weeks
Case studies in business analytics with ACCENTURE (Coursera) Coursera
ESSEC Business School

Case studies in business analytics with ACCENTURE (Coursera)

This course is RESTRICTED TO LEARNERS ENROLLED IN Strategic Business Analytics SPECIALIZATION as a preparation to the capstone project. During the first two MOOCs, we focused on specific techniques for specific applications. Instead, with this third MOOC, we provide you with different examples to open your mind to different applications from different industries and sectors. The objective is to give you an helicopter overview on what's happening in this field. You will see how the tools presented in the two previous courses of the Specialization are used in real life projects.

Sep 21st 2026
3 Weeks
Plant Bioinformatics (Coursera) Coursera
University of Toronto

Plant Bioinformatics (Coursera)

The past 15 years have been exciting ones in plant biology. Hundreds of plant genomes have been sequenced, RNA-seq has enabled transcriptome-wide expression profiling, and a proliferation of "-seq"-based methods has permitted protein-protein and protein-DNA interactions to be determined cheaply and in a high-throughput manner. These data sets in turn allow us to generate hypotheses at the click of a mouse.

Sep 21st 2026
5-12 Weeks