Skip to main content

Image Classification → Steps


Assembling the Training Data

Training data (also called training samples or signature sets) are the foundation of supervised image classification in remote sensing.
This is where the analyst selects representative examples of each land-cover class—such as water, vegetation, urban, soil, etc.—from the satellite image.

To prepare training data properly, several analytical and interactive steps are used. These help ensure that the classes are well separated and that the classifier receives the correct spectral information.

1. Graphical Representation of Spectral Response Patterns

✔ What it means

For each class (e.g., water, forest, built-up), the training pixels have a spectral signature—a pattern of reflectance values across the image's spectral bands.

This pattern is visualized using:

  • Spectral reflectance curves

  • Band-by-band scatter plots

  • Histograms for each band

✔ Purpose

  • To understand how different classes behave in different bands

  • To check if the selected training pixels are spectrally consistent

  • To identify overlaps between classes (e.g., dark soil and turbid water)

✔ Key terminology

  • Spectral profile / spectral signature

  • Spectral separability

  • Spectral scatterplot

  • Feature space

2. Quantitative Expressions of Category Separation

This step uses mathematical measures to check if classes are well-separated in spectral space.

✔ Why it matters

Classification accuracy depends on how distinct one class is from another.
If training classes overlap too much, classification errors will occur.

✔ Common quantitative measures

  • Transformed Divergence (TD)

  • Jeffries–Matusita Distance (JM)

  • Bhattacharyya Distance (BD)

✔ What these values indicate

  • Values close to 2.0 (JM scale) → excellent class separability

  • Values close to 0.0 → poor separability; classes overlap

  • Helps decide whether to:

    • combine classes

    • redefine training samples

    • collect more samples

    • split a mixed class

✔ Key terminology

  • Separability index

  • Statistical distance

  • Cluster separation

  • Spectral overlap

3. Self-Classification of the Training Data Set

✔ Concept

Before performing classification on the full image, the classifier is run only on the training pixels themselves.

✔ Purpose

  • To check if the algorithm correctly "recognizes" the classes it was trained on.

  • If the classifier mislabels the training samples, the training data need to be corrected.

✔ What it reveals

  • Misclassified pixels → inaccurate training sets

  • Mixed or overlapping classes

  • Inconsistencies in attribute statistics (means, variances)

  • Too much variability within a class

✔ Key terminology

  • Internal accuracy check

  • Confusion among training classes

  • Spectral homogeneity

4. Interactive Preliminary Classification

✔ What it is

A rough or temporary classification is generated on the image using preliminary training samples.

✔ Purpose

  • To visually inspect how the training data behave when applied to the entire image

  • To refine training sites

  • To identify new sub-classes or remove misidentified ones

✔ What the analyst checks

  • Are water bodies correctly classified?

  • Are vegetation areas split properly (forest vs cropland)?

  • Are built-up areas being confused with dry soil?

✔ Why "interactive"?

The analyst reviews the output and actively adjusts:

  • training polygons

  • class definitions

  • band combinations

  • class separability

✔ Key terminology

  • Pre-classification map

  • Trial classification

  • Interactive refinement

5. Representative Subscene Classification

✔ Concept

Instead of classifying the whole image, a small but representative subscene is used.

A subscene:

  • contains all major land-cover types

  • captures geographic and spectral variability

  • is easier to evaluate and test

✔ Purpose

  • To test classifier performance on a manageable area

  • To refine spectral signatures before final classification

  • To avoid wasting processing time on the full image if training data are weak

✔ What it helps detect

  • Class confusion in specific regions

  • Spectral variability across the scene

  • Need for more training samples

  • Problems with similar classes (e.g., shallow water vs wet soil)

✔ Key terminology

  • Subscene

  • Training refinement

  • Pilot classification

  • Signature validation


Assembling training data for supervised image classification involves:

  1. Graphical representation of spectral response patterns – using spectral curves, histograms, and scatter plots to visualize class behavior.

  2. Quantitative expressions of category separation – using statistical measures (JM, TD, BD) to evaluate how distinct classes are.

  3. Self-classification of training data – testing if the classifier correctly labels its own training samples.

  4. Interactive preliminary classification – producing a trial classification to visually refine training sites.

  5. Representative subscene classification – testing the classifier on a smaller, diverse image subset to check accuracy and refine signatures.


Comments

Popular posts from this blog

CREATION OF SPATIAL DATA

Spatial data creation is the process of generating, organizing, and managing geographically referenced information in a Geographic Information System (GIS). It involves converting maps, satellite images, GPS observations, and field survey data into digital datasets that can be stored, analyzed, and visualized. The quality of GIS analysis depends largely on the accuracy of spatial data creation. 1. Creation of Shapefile and Geodatabase A. Shapefile A Shapefile is one of the most widely used vector data formats developed by Esri for storing geographic features. Definition A shapefile stores the geometry and attributes of geographic features such as points, lines, and polygons. Components of a Shapefile A shapefile consists of several files: .shp – Stores geometry (shape) .shx – Shape index .dbf – Attribute table .prj – Coordinate Reference System (CRS) .sbn/.sbx – Spatial index (optional) Geometry Types Point – W...

Nature and Scope of Geography

Geography is the scientific study of the Earth's surface, its physical features, human populations, and the interactions between people and their environment. The word Geography is derived from the Greek words Geo (Earth) and Graphien (to describe or write), meaning "description of the Earth." Modern geography goes far beyond description; it seeks to explain where phenomena occur, why they occur there, how they are spatially distributed, and how they change over time. Geography is regarded as a spatial science , an environmental science , and an integrative discipline because it bridges natural sciences, social sciences, and geospatial technologies. Nature The nature of geography refers to the characteristics and fundamental features that define the discipline. 1. Geography as a Spatial Science Terminology: Spatial Science A discipline concerned with the location, distribution, arrangement, organization, and interaction of phenomena in ...

Geography of Health or Medical Geography

Health Geography (also known as Medical Geography ) is a sub-discipline of Human Geography that studies the relationships between place, environment, society, and health . It examines how spatial location, environmental conditions, and social and economic factors influence human health, disease patterns, and access to healthcare services. Health geography integrates concepts from geography, epidemiology, medicine, public health, environmental science, sociology, and Geographic Information Systems (GIS) to understand and improve population health. Major Components of Health Geography Health geography is generally divided into two major branches : The Geography of Disease and Ill Health The Geography of Health Care 1. The Geography of Disease and Ill Health This branch studies the spatial distribution, determinants, and diffusion of diseases across different geographical scales, from neighborhoods to global regions. It seeks t...

Historical Development of Geography in the Ancient Period

The Ancient Period marks the earliest stage in the evolution of geographical thought, extending from approximately 3000 BCE to the 5th century CE . During this period, geography evolved from simple descriptions of the Earth's surface to systematic scientific inquiry. Early civilizations developed geographical knowledge to meet practical needs such as navigation, trade, agriculture, military expansion, taxation, and administration . The greatest contributions came from the Mesopotamian, Egyptian, Indian, Chinese, Greek, and Roman civilizations , with the Greeks laying the foundations of scientific geography . Meaning Terminology: Historical Development Historical development refers to the gradual evolution of geographical knowledge, concepts, methods, and theories over time. Concept Geographical knowledge evolved through: Observation of the natural environment Exploration and travel Cartography (map-making) Astronomical observations ...

Development of Health Geography

Health Geography (formerly Medical Geography ) is the branch of geography that studies the relationship between health, disease, environment, place, and healthcare systems . The discipline has evolved over more than 2,500 years through contributions from physicians, geographers, epidemiologists, microbiologists, and public health experts. The development of Health Geography can be divided into the following periods: Ancient Period Medieval Period Renaissance and Pre-Modern Period Nineteenth Century (Pre-World War Era) World War Period Post-World War Period Modern Health Geography 1. Ancient Period (5th Century BC – 500 AD) Characteristics Health closely linked with the natural environment. Diseases explained through climate, water, air, and seasons. No knowledge of microorganisms. Medical observations were descriptive. Major Concepts Environmental Determinism Disease Ecology Climate and Healt...