Robotics: Perception (Coursera)

Robotics: Perception (Coursera)

How can robots perceive the world and their own movements so that they accomplish navigation and manipulation tasks? In this module, we will study how images and videos acquired by cameras mounted on robots are transformed into representations like features and optical flow. Such 2D representations allow us then to extract 3D information about where the camera is and in which direction the robot moves. You will come to understand how grasping objects is facilitated by the computation of 3D posing of objects and navigation can be accomplished by visual odometry and landmark-based localization.

Class Deals by MOOC List - Click here and see Coursera's Active Discounts, Deals, and Promo Codes.

Course 4 of 6 in the Robotics Specialization.

Syllabus

WEEK 1
Geometry of Image Formation
Welcome to Robotics: Perception! We will begin this course with a tutorial on the standard camera models used in computer vision. These models allow us to understand, in a geometric fashion, how light from a scene enters a camera and projects onto a 2D image. By defining these models mathematically, we will be able understand exactly how a point in 3D corresponds to a point in the image and how an image will change as we move a camera in a 3D environment. In the later modules, we will be able to use this information to perform complex perception tasks such as reconstructing 3D scenes from video.

WEEK 2
Projective Transformations
Now that we have a good camera model, we will explore the geometry of perspective projections in depth. We will find that this projection is the cause of the main challenge in perception, as we lose a dimension that we can no longer directly observe. In this module, we will learn about several properties of projective transformations in depth, such as vanishing points, which allow us to infer complex information beyond our basic camera model.

WEEK 3
Pose Estimation
In this module we will be learning about feature extraction and pose estimation from two images. We will learn how to find the most salient parts of an image and track them across multiple frames (i.e. in a video sequence). We will then learn how to use features to find the position of the camera with respect to another reference frame on a plane using Homographies. We will also learn about how to make these techniques more robust, using least squares to hand noisy feature points or RANSAC to remove completely erroneous feature points.

WEEK 4
Multi-View Geometry
Now we will use what we learned from two view geometry and extend it to sequences of images, such as a video. We will explain the fundamental geometric constraints between point features in images, the Epipolar constraint, and learn how to use it to extract the relative poses between multiple frames. We will finish by combining all this information together for the application of Structure from Motion, where we will compute the trajectory of a camera and a map throughout many frames and refine our estimates using Bundle adjustment.

Go to Class
MOOC List is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

Related Courses

Landing.AI for Beginners: Build Data Visualization AI Models (Coursera) Coursera
Coursera Project Network

Landing.AI for Beginners: Build Data Visualization AI Models (Coursera)

In this 1-hour long project-based course, you'll step into the exciting field of Computer Vision and Generative AI using the LandingLens platform. We'll start by exploring the concept of visual prompting, and initiating a visual prompting project. LandingLens simplifies the model creation, training, and deployment process, making it a user-friendly platform for this endeavor.

Oct 19th 2026
1 Week
AI and Disaster Management (Coursera) Coursera
DeepLearning.AI

AI and Disaster Management (Coursera)

In Course 3, AI and Disaster Management, you will begin by learning how natural disasters create both short and long-term impacts on the economy, environment, and community. You will learn how the disaster management cycle can be used to reduce the impacts of a disaster and guidelines for how you can engage in disaster-related projects.

Oct 19th 2026
3 Weeks
Digitalisation in the Aerospace Industry (Coursera) Coursera
Technische Universität München - TUM

Digitalisation in the Aerospace Industry (Coursera)

The online course Digitalisation in Aerospace aims at making you aware of special production requirements connected with digitalisation. You will learn about the role of robotics and automation in manufacturing and gain a better understanding of differing perspectives on research and manufacturing as well as the points where these intersect.

Oct 19th 2026
3 Weeks
Computer Vision in Microsoft Azure (Coursera) Coursera
Microsoft

Computer Vision in Microsoft Azure (Coursera)

In Microsoft Azure, the Computer Vision cognitive service uses pre-trained models to analyze images, enabling software developers to easily build applications"see" the world and make sense of it. This ability to process images is the key to creating software that can emulate human visual perception. In this course, you'll explore some of these capabilities as you learn how to use the Computer Vision service to analyze images.

Oct 19th 2026
3 Weeks
Synapses, Neurons and Brains (Coursera) Coursera
Hebrew University of Jerusalem

Synapses, Neurons and Brains (Coursera)

These are very unique times for brain research. The aperitif for the course will thus highlight the present “brain-excitements” worldwide. You will then become intimately acquainted with the operational principles of neuronal “life-ware” (synapses, neurons and the networks that they form) and consequently, on how neurons behave as computational microchips and how they plastically and constantly change - a process that underlies learning and memory.

Oct 19th 2026
5-12 Weeks
Geometría Analítica Preuniversitaria (Coursera) Coursera
Universidad Autónoma Metropolitana

Geometría Analítica Preuniversitaria (Coursera)

Líneas rectas, círculos, parábolas, elipses e hipérbolas son figuras geométricas que encontramos en nuestro derredor. Por ejemplo, mucha gente sabe que los planetas en nuestro sistema solar se mueven en órbitas elípticas teniendo al astro rey en un foco de esta figura. Sin embargo, pocos saben que la plaza de San Pedro en el Vaticano está construída sobre elipses donde sus focos se encuentran sobre las fuentes donde mucha gente se toma fotos. Estos son dos ejemplos que muestran la importancia de las figuras geométricas en nuestra vida.

Nov 2nd 2026
5-12 Weeks
Advanced Neurobiology II (Coursera) Coursera
Peking University

Advanced Neurobiology II (Coursera)

Hello everyone! Welcome to advanced neurobiology! Neuroscience is a wonderful branch of science on how our brain perceives the external world, how our brain thinks, how our brain responds to the outside of the world, and how during disease or aging the neuronal connections deteriorate. We’re trying to understand the molecular, cellular nature and the circuitry arrangement of how nervous system works.

Nov 2nd 2026
5-12 Weeks
Input and Interaction (Coursera) Coursera
University of California, San Diego

Input and Interaction (Coursera)

In this course, you will learn relevant fundamentals of human motor performance, perception, and cognition that inform effective interaction design. You will use these models of how people work to design more effective input and interaction techniques. You’ll apply these to both traditional graphic and gestural interfaces.

Oct 19th 2026
3 Weeks
Modern Robotics, Course 3: Robot Dynamics (Coursera) Coursera
Northwestern University

Modern Robotics, Course 3: Robot Dynamics (Coursera)

Do you want to know how robots work? Are you interested in robotics as a career? Are you willing to invest the effort to learn fundamental mathematical modeling techniques that are used in all subfields of robotics? If so, then the "Modern Robotics: Mechanics, Planning, and Control" specialization may be for you. This specialization, consisting of six short courses, is serious preparation for serious students who hope to work in the field of robotics or to undertake advanced study. It is not a sampler.

Oct 19th 2026
4 Weeks
Computer Vision with Embedded Machine Learning (Coursera) Coursera
Edge Impulse

Computer Vision with Embedded Machine Learning (Coursera)

Computer vision (CV) is a fascinating field of study that attempts to automate the process of assigning meaning to digital images or videos. In other words, we are helping computers see and understand the world around us! A number of machine learning (ML) algorithms and techniques can be used to accomplish CV tasks, and as ML becomes faster and more efficient, we can deploy these techniques to embedded systems.

Oct 26th 2026
3 Weeks
Features and Boundaries (Coursera) Coursera
Columbia University

Features and Boundaries (Coursera)

This course focuses on the detection of features and boundaries in images. Feature and boundary detection is a critical preprocessing step for a variety of vision tasks including object detection, object recognition and metrology – the measurement of the physical dimensions and other properties of objects. The course presents a variety of methods for detecting features and boundaries and shows how features extracted from an image can be used to solve important vision tasks.

Oct 19th 2026
5-12 Weeks
Visual Perception for Self-Driving Cars (Coursera) Coursera
University of Toronto

Visual Perception for Self-Driving Cars (Coursera)

Welcome to Visual Perception for Self-Driving Cars, the third course in University of Toronto’s Self-Driving Cars Specialization. This course will introduce you to the main perception tasks in autonomous driving, static and dynamic object detection, and will survey common computer vision methods for robotic perception. By the end of this course, you will be able to work with the pinhole camera model, perform intrinsic and extrinsic camera calibration, detect, describe and match image features and design your own convolutional neural networks.

Oct 19th 2026
5-12 Weeks