Skip to main content

Supervised Classification

In the context of Remote Sensing (RS) and Digital Image Processing (DIP), supervised classification is the process where an analyst defines "training sites" (Areas of Interest or ROIs) representing known land cover classes (e.g., Water, Forest, Urban). The computer then uses these training samples to teach an algorithm how to classify the rest of the image pixels.

The algorithms used to classify these pixels are generally divided into two broad categories: Parametric and Nonparametric decision rules.


Parametric Decision Rules

These algorithms assume that the pixel values in the training data follow a specific statistical distribution—almost always the Gaussian (Normal) distribution (the "Bell Curve").

  • Key Concept: They model the data using statistical parameters: the Mean vector ($\mu$) and the Covariance matrix ($\Sigma$).

  • Analogy: Imagine trying to fit a smooth hill over your data points. If a new point lands high up on the hill, it belongs to that class.

Nonparametric Decision Rules

These algorithms make no assumptions about the statistical distribution of the data. They do not care if the data fits a bell curve.

  • Key Concept: They classify based on discrete geometric shapes (polygons, boxes) or the relative position of the data points themselves.

  • Analogy: Imagine drawing a literal box or fence around your data points. If a new point falls inside the fence, it belongs to that class.


A. Minimum-Distance-to-Means (MDM)

  • Classification: Generally considered a simple Parametric classifier (as it relies on the mean parameter), though it operates geometrically.

  • How it works:

    1. The algorithm calculates the spectral mean vector (the center point or centroid) for each training class.

    2. For every unclassified pixel in the image, it calculates the Euclidean distance to the mean of every class.

    3. The pixel is assigned to the class with the shortest distance.


  • Pros: Very fast computationally; mathematically simple.

  • Cons: It is insensitive to the variance (spread) of the data.

    • Example: If "Urban" data is very scattered (high variance) and "Water" is very tight (low variance), a pixel far from the Urban center might actually belong to Urban, but MDM might classify it as Water just because the Water mean is slightly closer geometrically.

B. Parallelepiped Classification

  • Classification: Nonparametric.

  • How it works:

    1. The algorithm looks at the training data and finds the minimum and maximum brightness values for each band.

    2. It creates a rectangular box (a parallelepiped in multi-dimensional space) defined by these limits.

    3. If a pixel's value falls within the box, it is assigned to that class.

  • Pros: Extremely fast; easy to understand conceptually.

  • Cons:

    • The Correlation Problem: Real remote sensing data (like vegetation in Red vs. NIR bands) is often correlated (diagonal distribution). A rectangular box cannot fit a diagonal data cloud efficiently, leading to large "empty corners" in the box that capture noise/wrong pixels.

    • Overlapping: Pixels often fall into the overlapping area of two boxes, leaving the computer unable to decide.

C. Gaussian Maximum Likelihood (GML/MLC)

  • Classification: Parametric (The standard industry workhorse).

  • How it works:

    1. It assumes the data for each class is normally distributed.

    2. It uses both the Mean vector AND the Covariance matrix to calculate the probability density function.

    3. It calculates the statistical probability of a pixel belonging to each class.

    4. It constructs ellipsoidal equiprobability contours (rather than circles or boxes).

  • Pros: Highly accurate because it accounts for the variance (spread) and covariance (correlation/direction) of the data. It handles "diagonal" data clouds perfectly.

  • Cons: Computationally expensive (slow on massive images); requires a large number of training pixels per class to compute a stable covariance matrix (usually $10N$ to $100N$ pixels, where $N$ is the number of bands).


FeatureParallelepipedMinimum DistanceMaximum Likelihood
TypeNonparametricParametric (Simple)Parametric (Advanced)
GeometryRectangular BoxesCircles/SpheresEllipsoids
AssumptionsNone (Min/Max thresholds)Mean Center PointGaussian Distribution
SpeedVery FastFastSlow / Intensive
AccuracyLow to ModerateModerateHigh
Best Used ForQuick looks; Uncorrelated dataWell-separated classesComplex, correlated data


Comments

Popular posts from this blog

Regional Geography, Systematic Geography, Idiographic, Nomothetic, Inductive and Deductive Approaches

T wo major ways of studying Geography : the Regional Approach and the Systematic Approach . It also explains the related ideas of idiographic vs. nomothetic and inductive vs. deductive reasoning , especially in the context of the Hartshorne–Schaefer debate . 1. Regional Geography: “All About One” Regional Geography studies one particular region in detail . A region is an area that has some degree of homogeneity (sameness) within its boundary but is also unique or different from other regions . For example, if we study Palakkad District , we may study: Relief and drainage Climate Soil Vegetation Agriculture Population Occupation Economy Culture Political characteristics The purpose is to understand the complete geographical personality of Palakkad and the relationships among its different features. Key concepts Region: A geographical area with identifiable characteristics and boundaries. Homogeneity: Si...

Kuhn’s Paradigms

The given content explains Thomas S. Kuhn’s model of scientific development , its application to geography, and criticisms by Karl Popper, Paul Feyerabend, Michel Foucault , and others. 1. Basic Idea Kuhn argued that science does not develop continuously in a straight line . Instead, scientific development occurs through: Preparadigm → Paradigm → Normal Science → Crisis → Scientific Revolution → New Paradigm A new paradigm may replace an older one, producing a major change in the way scientists understand and study a subject. Concepts and Terminologies Concept / Term Simple Meaning Paradigm A commonly accepted framework/model that guides scientific research Exemplar A successful concrete problem-solution used as a model for future research Disciplinary Matrix Shared beliefs, values, concepts, methods and techniques of a scientific community Preparadig...

Multispectral and Hyperspectral Imaging Systems

The main idea is how a remote-sensing sensor collects information about an area . A sensor does not simply take an ordinary photograph. It measures the electromagnetic energy reflected or emitted by objects in different wavelength bands . Depending on how many bands are measured and how the sensor collects them, different imaging systems are used. 1. Multispectral vs. Hyperspectral Multispectral imaging (MSI) records information in a limited number of relatively broad, separate spectral bands , such as blue, green, red, near-infrared and shortwave infrared. Hyperspectral imaging (HSI) records information in many narrow and usually contiguous spectral bands . Therefore, it provides a much more detailed spectral signature of each pixel. The resulting dataset is commonly called a hyperspectral data cube (hypercube) because it contains: X-axis → spatial information Y-axis → spatial information Z-axis → wavelength/spectral information Thus, hype...

Discrete Detectors and Scanning mirrors Across the track scanner Whisk broom scanner.

Multispectral Imaging Using Discrete Detectors and Scanning Mirrors (Across-Track Scanner or Whisk Broom Scanner) Multispectral Imaging:  This technique involves capturing images of the Earth's surface using multiple sensors that are sensitive to different wavelengths of electromagnetic radiation.  This allows for the identification of various features and materials based on their spectral signatures. Discrete Detectors:  These are individual sensors that are arranged in a linear or array configuration.  Each detector is responsible for measuring the radiation within a specific wavelength band. Scanning Mirrors:  These are optical components that are used to deflect the incoming radiation onto the discrete detectors.  By moving the mirrors,  the sensor can scan across the scene,  capturing data from different points. Across-Track Scanner or Whisk Broom Scanner:  This refers to the scanning mechanism where the mirror moves perpendicular to the direction of flight.  This allows for t...

Satalite

Landsat → Land resources SPOT → High-resolution mapping IRS → Indian natural-resource mapping ASTER → Geology + thermal + DEM QuickBird → Very high spatial resolution MODIS → Daily global monitoring GOES → Weather monitoring AVHRR → Weather + vegetation + ocean AVIRIS → Hyperspectral imaging Highest spectral resolution: AVIRIS (224 narrow bands) Highest spatial resolution in this list: QuickBird (~0.61 m PAN) Highest temporal frequency: GOES (minutes) Best broad global monitoring: MODIS Indian sensors: IRS-LISS III and LISS IV Hyperspectral: AVIRIS Thermal + multispectral + DEM: ASTER abbreviations MSS – Multispectral Scanner System TM – Thematic Mapper ETM+ – Enhanced Thematic Mapper Plus GOES – Geostationary Operational Environmental Satellite AVHRR – Advanced Very High Resolution Radiometer HRV – High Resolution Visible HRVIR – High Resolution Visible and Infrared HRG – High Resolution Geometric ...