Unit 4 Notes
Unit 4 Notes
[Link] spatial data model will be advantages for urban growth studies?
For urban growth studies in remote sensing and GIS, a advantageous spatial data model is the
vector data model.
Here are two key reasons why the vector data model is advantageous for urban growth
studies:
1. Precision and Detail: Vector data models represent geographic features as discrete
objects with precise boundaries, such as points, lines, and polygons. This level of
detail is well-suited for studying urban growth, where the delineation of city
boundaries, roads, buildings, and other infrastructure elements is crucial for accurate
analysis.
2. Topology: The vector data model can capture topological relationships between
geographic features, making it ideal for modeling the connectivity and adjacency of
urban features. This is important for understanding how urban areas expand, where
new developments occur, and how they relate to existing infrastructure.
Overall, the vector data model provides the necessary level of detail, precision, and
topological information required for in-depth urban growth analysis in remote sensing
and GIS.
7. What are the levels of measurement used in GIS?
Nominal (categories)
Ordinal (ordered categories)
Interval (equal intervals)
Ratio (true zero point)
8. What do you mean by data compression ?
Large amounts of data can create enormous problems in storage space and
transmission time.
The main reasons of data compression could be summarized as:
•Multimedia data have large data volume
•Difficulty sending real-time uncompressed data over current
network
The design goal of image compression is to represent images with as few bits as
possible, to save storage and transmission channel capacity.
Compression data principle: “Eliminate data redundancy and try to find a code
with less data volume.”
9. What are the advantages & disadvantages of raster and vector data ?
Advantages of raster data structures
Simple data structure
Easy to generate (e.g. from remote sensing or scandigitizing)
Easy workflows and analysis
Technology is cheap and is being energetically developed
Simulation is very easy as each spatial unit has the same shape and size
Same set of grid cells are used for several variables
Disadvantages of raster data structures
Non-adaptive data structure
Definition: realization of the discrete model of real world using an object centered approach
in which an object has both physical (attribute) and geometric characteristics. ′
Different types of objects can interact because they are not confined to separate layers
SPATIAL DATA MODEL
There are presently three types of representations for geographic data: raster vector,
and objects.
raster - set of cells on a grid that represents an entity (entity --> symbol/color -->
cells).
vector -an entity is represented by nodes and their connecting arc or line segment
(entity --> points, lines or areas --> connectivity) ′
object – an entity is represented by an object which has as one of its attributes spatial
information
[Link] about the data elevation model in GIS
Digital Elevation Model (DEM) is the digital representation of the land surface
elevation with respect to any reference datum.
DEM is frequently used to refer to any digital representation of a topographic
surface.
DEM is the simplest form of digital representation of topography
DEMs are used to determine terrain attributes such as elevation at any point, slope
and aspect.
Terrain features like drainage basins and channel networks can also be identified
from the DEMs.
DEMs are widely used in hydrologic and geologic analyses, hazard monitoring,
natural resources exploration, agricultural management etc
Hydrologic applications of the DEM include groundwater modelling, estimation of
the volume of proposed reservoirs, determining landslide probability, flood prone
area mapping etc.
Three main type of structures used are the following.
a) Regular square grids
b) Triangulated irregular networks (TIN)
c) Contours
A digital elevation model (DEM) is a digital model or 3D representation of a
terrain's surface, created from terrain elevation data.
There are three similar names as digital elevation model (DEM), digital terrain
model (DTM) and digital surface model (DSM).
DEM is a subset of the DTM, which also represents other morphological
elements
′Digital Elevation Model: It is a bare-earth raster grid referenced to a vertical
datum. The built (power lines, buildings and towers) and natural (trees and other
types of vegetation) aren’t included in a DEM.
Digital Terrain Model (DTM): DTM is a mathematical representation (model) of
the ground surface, most often in the form of a regular grid, in which a unique
elevation value is assigned to each pixel. ′
Digital Surface Model (DSM): Represents the MSL elevations of the reflective
surfaces of trees, buildings, and other features elevated above the “Bare Earth”
Types of DEM:
Triangular irregular network (TIN) (primary DEM ): It is a vector-based. It uses
irregular sampling points connected through non-overlapping triangles.
Raster DEM (secondary DEM): Also known as a heightmap when representing
elevation
Gridded DEM (GDEM): Consists of regularly placed, uniform grids with the
elevation information of each grid
Data Sources: Data retrieval in remote sensing and GIS begins with identifying and
selecting appropriate data sources. These sources may include:
1. Remote Sensing Satellites: Obtaining satellite imagery and data from Earth-
observing satellites such as Landsat, Sentinel, and MODIS for various applications
like land cover mapping, environmental monitoring, and disaster management.
2. Aerial Photography: Collecting aerial imagery through aircraft or drones for high-
resolution mapping, land use planning, and infrastructure monitoring.
3. Ground Surveys: Acquiring ground-based data through field surveys, GPS
measurements, and sensor deployments to validate and supplement remote sensing
data.
4. GIS Databases: Accessing existing GIS databases and repositories containing
geographic data layers, such as land parcels, transportation networks, and
demographic information.
5. Open Data Portals: Utilizing open data portals and government agencies' resources
to access publicly available geospatial datasets and maps.
Data retrieval involves ensuring compatibility with specific data formats and standards:
1. Raster Formats: Remote sensing data, such as satellite imagery and elevation
models, are often stored in raster formats like GeoTIFF or HDF.
2. Vector Formats: GIS data, such as shapefiles, GeoJSON, or KML, use vector
formats to represent geographic features and attributes.
Data Integration: Once retrieved, data is integrated into GIS software or remote sensing
platforms for analysis, visualization, and modeling. This involves:
1. Data Preprocessing: Cleaning and preparing the data, including georeferencing,
mosaicking, and removing artifacts or noise.
2. Geospatial Analysis: Performing spatial analysis tasks like spatial querying, overlay
analysis, and geoprocessing to extract meaningful information.
3. Visualization: Creating maps, charts, and visual representations to convey spatial
information effectively.
4. Modelling: Using retrieved data as inputs for spatial models and simulations, such as
hydrological modelling or land use change prediction.
Data Updates: Data retrieval is an ongoing process as new data becomes available, or
existing data is updated. Regular updates and maintenance are essential for ensuring the
accuracy and relevance of GIS and remote sensing datasets.
In conclusion, data retrieval in remote sensing and GIS involves the identification,
acquisition, and integration of spatial data from various sources to support decision-making,
analysis, and modelling in a wide range of applications, from environmental monitoring to
urban planning. The choice of data sources and methods depends on the specific project
requirements and objectives.
Overlay analysis
Overlay analysis is a fundamental spatial analysis technique used in both remote sensing and
Geographic Information Systems (GIS). It involves the layering or combination of multiple
spatial datasets to derive new information, make informed decisions, and gain insights into
the relationships between different geographic features. Here's a short note on overlay
analysis in remote sensing and GIS:
Overlay analysis serves several important purposes in remote sensing and GIS:
1. Spatial Querying: It allows users to answer complex spatial questions by identifying
features that meet specific criteria or conditions.
2. Integration of Information: Overlay combines diverse data sources, such as land
use, population, and environmental variables, to create composite datasets for
comprehensive analysis.
3. Decision Support: It aids in decision-making by visualizing spatial relationships and
patterns, making it valuable for urban planning, environmental management, and
resource allocation.
Applications:
Challenges:
Overlay analysis can be complex and computationally intensive, particularly when dealing
with large datasets or numerous layers. Ensuring the quality and accuracy of input data is
crucial, as errors or inaccuracies can propagate into the results. Additionally, defining
appropriate criteria and rules for overlay operations requires careful consideration.
In conclusion, overlay analysis is a powerful spatial analysis technique in remote sensing and
GIS that allows users to combine and analyze spatial data layers to extract valuable insights,
make informed decisions, and address a wide range of spatial problems in diverse fields.
[Link] in detail about the various type of data &data formats used in GIS
File formats
File formats are intended to store particular kinds of digital information.
E.g. JPEG format is designed only to store still images, while the GIF format
supports storage of both still images and simple animations.
The most well-known formats have file specifications that describe exactly how the
data is to be encoded.
Digital image formats
The standard greyscale images use 256 shades of grey from 0 (black) to 255
(white).
For a given number of pixels, considerably more data is required to represent a
color image.
The important data formats used for GIS are:
1. GIF
2. JPEG
3. TIFF
4. BMP
5. PNG
1. GIF (Graphic Interchange Format):
GIF files offer optimum compression (smallest files) for solid color graphics.
The GIF format can only show 256 colors but it can choose any of the 16 million colors
in a 24- bit image.
⚫ 8 bit = 256 colors
⚫ 7 bit = 128 colors
⚫ 6 bit = 64 colors
⚫ 5 bit = 32 colors
⚫ 4 bit = 16 colors
⚫ 3 bit = 8 colors
⚫ 2 bit = 4 colors
⚫ 1 bit = 2 colors (usually B/W)
[Link] (Tagged Image File Format):
TIFF files have many formats: Black and white, greyscale, 4- and 8-bit color, full
color (24-bit) images.
TIFF files support the use of data compression using many compression standards.
[Link] (Bitmap):
B/W Bitmap is monochrome and the color table contains two entries.
Each bit in the bitmap array represents a pixel. ′
4-bit Bitmap has a maximum of 16 colors.
8-bit Bitmap has a maximum of 256 colors.
[Link] (Joint Photographic Expert Group Format):
A JPEG image provides very good compression, but does not uncompress exactly as
it was; JPEG is a lossy compression techniques.
Compressions up to 50:1 are easily obtainable.
[Link] (Portable Network Graphics):
The PNG format was designed to replace the antiquated GIF format, and to some
extent, the TIFF format. ′
It utilizes lossless compression.
It is a universal format that is recognized by the World Wide Web consortium, and
supported by modern web browsers.
File formats and supported data types
[Link] detailed notes on vector & raster data structures .
1. Raster data structures:
⚫ represents geography via grid cells
⚫ single values associated with each cell
⚫ typically 8 bits assigned to values therefore 256 possible values (0-255)
[Link] data structures:
⚫ represents geography via coordinates
⚫ point (node): 0-dimension
⚫ line (arc): 1-dimension
⚫ polygon : 2-dimensions
Usage Scenarios of: Raster Data Structures
Photos
Photogrammetry and remote sensing
Scanned images of maps
Terrain modelling
Landcover analysis ′
Hydrologic modelling and analysis
General GIS surface modelling and analysis for continuous surface
Usage scenarios of:Vector Data Structures
CAD, technical drawings
Street or river networks,
cadastral map
Network analysis Cartography
Advantages of Raster Data Structures
Simple data structure ′ Easy to generate (e.g. from remote sensing or
scandigitizing) ′
Easy workflows and analysis
Technology is cheap and is being energetically developed
Simulation is very easy as each spatial unit has the same shape and size
Same set of grid cells are used for several variable
Disadvantages of Raster Data Structures
Non-adaptive data structure
Tends to generate huge files, depending on resolution ′
Cell arrangement is usually random and does not respect natural borders
Errors in evaluating perimeter of shape ′
Topology or network linkages are difficultto establish
Geometric transformations are difficult to handle
Use of large cells to reduce data volumes result into loss of information
Advantages of Vector Data Structures
Small amount of data
Logical data structure
Attributes are combined with objects
Preserves quality after interactivity (e.g. scaling)
Topology described with network linkages ′
Retrieval, updating and generalization of graphics and attributes are possible
Widely used to described administrative zones
Disadvantages of Vector Data Structures
Complex data structure
Continuous data is not represented effectively
Spatial variability is not implicitly represented
Spatial analysis and filtering within polygons is impossible
Needs a lot of manual editing to get good quality
It always introduces hard boundaries
Simulation is difficult as each unit has different topological form
Display and plotting can be expensive
[Link] in detailed about spatial analysis & overlay analysis.
Spatial analysis and overlay analysis are essential techniques in the fields of remote sensing
and Geographic Information Systems (GIS). They involve the manipulation, integration, and
interpretation of spatial data to extract meaningful information and gain insights into the
relationships between geographic features. Let's discuss these concepts in detail:
Spatial Analysis:
Spatial analysis is a broad term that encompasses a wide range of operations and techniques
used to analyze and interpret spatial data. It involves processing and examining the spatial
relationships, patterns, and attributes of geographic features. Here are key aspects of spatial
analysis in remote sensing and GIS:
1. Geoprocessing Operations: Spatial analysis includes various geoprocessing
operations such as buffering, clipping, merging, and dissolving. These operations
modify or extract information from spatial datasets.
2. Proximity Analysis: Proximity analysis involves measuring distances between
features, finding nearest neighbors, and identifying areas within a certain distance of
specific features. It is crucial for applications like site selection and wildlife habitat
analysis.
3. Spatial Queries: Spatial queries involve selecting or filtering features based on their
spatial relationships with other features. For example, finding all houses within a
certain distance from a park.
4. Spatial Statistics: Spatial statistics techniques analyze patterns and distributions of
features, including hotspot analysis, spatial autocorrelation, and spatial interpolation
methods like kriging.
5. Network Analysis: Network analysis deals with route optimization, finding the
shortest path, and solving transportation problems. It is valuable in urban planning,
logistics, and emergency response.
6. Overlay Analysis: Overlay analysis, which we'll discuss in detail in the next section,
is a critical component of spatial analysis.
7. Terrain Analysis:Analyzing elevation and slope data to understand terrain
characteristics, suitable locations for infrastructure, and flood modeling.
8. Temporal Analysis: Examining changes and trends in spatial data over time,
especially important in monitoring environmental changes and land use/land cover
changes.
Overlay Analysis:
Overlay analysis is a specific type of spatial analysis that involves combining multiple spatial
datasets, often in raster or vector format, to derive new information. The primary objective is
to explore the intersection of various datasets to answer complex spatial questions and
generate new data layers. Here's a detailed look at overlay analysis:
1. Data Layers: Overlay analysis requires multiple layers of spatial data, each
representing different geographic features or attributes. These layers can include land
cover, land use, elevation, administrative boundaries, and more.
2. Overlay Operations: Overlay operations combine these layers based on spatial
relationships. Common overlay operations include union (combining features),
intersection (finding common areas), difference (highlighting differences), and
buffering (creating zones around features).
3. Attribute Data: In addition to the spatial relationships, overlay analysis considers
attribute data associated with features. Attributes may include population counts, land
use categories, or environmental characteristics.
4. Resultant Maps: The output of overlay analysis is often visualized as thematic maps,
where areas with specific combinations of features or conditions are symbolized
differently. These maps help visualize patterns and relationships.
5. Applications: Overlay analysis is widely used in urban planning, environmental
management, land suitability analysis, epidemiology, disaster management, and many
other fields. Examples include land use planning, habitat suitability modeling, and
flood risk assessment.
6. Challenges: Overlay analysis can be computationally intensive, especially with large
datasets or numerous layers. Care must be taken to ensure data quality and accuracy,
as errors can propagate into the results.
In summary, spatial analysis and overlay analysis are crucial components of remote sensing
and GIS. They enable professionals to extract valuable information from spatial data, make
informed decisions, and gain insights into complex spatial relationships, ultimately
supporting a wide range of applications in fields such as environmental science, urban
planning, public health, and more.
[Link] about modelling using GIS with a case study
Modelling using Geographic Information Systems (GIS) is the process of using spatial data
and analytical tools to create representations of real-world phenomena, make predictions, and
support decision-making. GIS models help users understand complex spatial relationships,
simulate scenarios, and analyze the potential impacts of different factors on geographic areas.
Below, I'll explain the concept of modelling using GIS and provide a case study to illustrate
its application.
Background:
A municipality in a coastal region wants to assess and mitigate flood risks in the face of
increasing sea-level rise due to climate change. They decide to use GIS modelling to identify
vulnerable areas and plan for future flood events.
In this case study, GIS modelling is used to assess flood risks, taking into account various
spatial factors and scenarios. Such modelling enables proactive planning and helps the
municipality make informed decisions to protect its community and infrastructure from future
flood events.
[Link] the spatial data models of GIS & their merits and demerits .
Spatial data models are fundamental components of Geographic Information Systems (GIS)
that help organize and represent geographical information. There are primarily two types of
spatial data models: vector and raster. Each has its own merits and demerits, making them
suitable for different types of geospatial data and applications.
1. Vector Data Model:
• Merits:
• Precision and Accuracy: Vector data models are highly precise and
accurate for representing point, line, and polygon features. This makes
them suitable for applications where precise geographic information is
essential, such as urban planning and navigation.
• Topology: Vector data models support topological relationships,
which can help in tasks like network analysis and spatial query
optimization.
• Efficient Storage: They are efficient in terms of storage for data with
discrete features.
• Easy to Edit and Update: Vector data can be easily edited, updated,
and maintained.
• Demerits:
• Complexity: Representing continuous data (e.g., elevation) can be
challenging in vector format as it requires discretization, potentially
leading to data loss.
• Data Volume: Vector data can lead to large data volumes for complex
geometries or high-resolution data.
• Performance: Complex spatial operations on vector data can be
computationally intensive.
2. Raster Data Model:
• Merits:
• Efficiency for Continuous Data: Raster data models are well-suited
for representing continuous data, such as elevation models, satellite
imagery, and temperature maps.
• Ease of Analysis: They are efficient for grid-based spatial analysis,
including terrain analysis, spatial interpolation, and remote sensing.
• Compact Storage: Raster data can be more compact in storage when
dealing with continuous data over large areas.
• Speed: Raster data operations can be faster for certain types of
analyses, especially when using parallel processing.
• Demerits:
• Loss of Detail: Raster data may lose detail and precision when
representing discrete features, such as roads, buildings, or
administrative boundaries.
• Data Volume: High-resolution raster data can lead to large file sizes.
• Complex Topology: Topological relationships between features are
not as easily represented in raster data.
3. Hybrid Models:
• Some GIS systems use hybrid models that combine vector and raster data
models to take advantage of the strengths of both. For example, vector data
may be used for representing discrete features and raster data for continuous
attributes.
4. TIN (Triangulated Irregular Network) Model:
• TIN is another spatial data model that represents terrain as a set of non-
overlapping triangles. It is particularly useful for representing complex terrain
surfaces.
In summary, the choice between vector and raster data models in GIS depends on the specific
application and the type of data being dealt with. Vector data excels in representing discrete
features with high precision, while raster data is better suited for continuous data and grid-
based analysis. Hybrid models and TIN models may be employed to combine the advantages
of both in some situations.
Clip: A cookie cutter approach where one layer is subsetted based on the extent of
another. The procedure is often used to clip data to some study area boundary.
Erase: A procedure that is opposite to a clip procedure; keeping all features outside of
an erase area.
Intersect: A procedure that combines layers and retains information common to the
intersected areas.
Union: Retains information from both layers entirely
Buffering
A process that defines/identifies areas within a specified distance of a coverage (e.g.,
within 300m of a stream),Process is applicable to points, line and polygons.
A buffer in GIS is a zone around a map feature measured in units of distance or time.
A buffer is an area defined by the bounding region determined by a set of points at a
specified maximum distance from all nodes along segments of an object
Ways to perform geoprocessing:
1. Arctoolbox: Simple way to integrate the tools and explore procedures.
2. Command line: more advanced way and may lead to script languages, for those computer
programmers in the course.
3. Model builder: Graphical way of stitching together commands.
4. Scripts: Customized text based sequence of commands creating a small program. Similar
to the command line or a command-line version of the Model builder
[Link] discuss the data input devices of vector & raster data
METHODSOFGEOSPATIAL DATA INPUT
Methods of Raster Data Input
Aerial photographs and satellite imagery are the examples of raster data in
digital format.
Scanner
Scanners are used to convert images from analog maps or photographs into
digital image data in raster format, which is then converted to vector format
through digital tracing or digitisation.
This process of tracing is also called vectorisation.
Scanning converts the map into binary file in raster format, with pixel having a
value of either ‘1’ representing the map feature or ‘0’ representing the
background (Chang, 2010)
A scanner has a light source, a background for source document and a lens.
There are three different types of scanners:
• Flat-bed scanners
• Drum scanners
• Large-formatfeed scanners
After scanning, the raster image needs to be first corrected for errors caused by scanning.
Image distortions are corrected by:
Despeckling: It removes speckles or stray pixels that appear in an image when
we scan a dirty or wrinkled image.
Greyscaling: It converts a colour image into greyscale
Brightness and Contrast: It can be adjusted in a colour or greyscale image.
Increasing the contrast enhances he distinction between dark and light areas
whereas increasing the brightness lightens the image so as to enhance even the
shadow areas.
Thresholding: It segregates the image grey values into two distinct values (that
is 0 for black and 255 for white) by a threshold value.
Methods of Vector Data Input
Vector data is usually captured or digitised from a hardcopy print out or a digital
raster image. ′
A number of input devices are in use for map digitisation.
Some of the commonly used devices for vector data input are:
o Manual (via typing on keyboard or importing text files
o Digitizing
o Scanning
Manual (via typing on keyboard or importing text files)
Manual data entry can bring into GIS either collected or measured data.
These data exist as simple text files (have at least two columns with X and
Y coordinates ) or binary files (for example files from Global Positioning
System data collection).
Keyboard is the simplest device to manually input data into a computer .
The process is also known as key coding.
Mouse is the simplest and the most accurate type of digitiser device
Cell arrangement is usually random and does not respect natural borders
Errors in evaluating perimeter of shape
Use of large cells to reduce data volumes result into loss of information
Advantages of vector data structures
Small amount of data
Logical data structure
Attributes are combined with objects
Preserves quality after interactivity (e.g. scaling)
Topology described with network linkages
Retrieval, updating and generalization of graphics and attributes are possible
Widely used to described administrative zone
Disadvantages of vector data structures
Complex data structure
Continuous data is not represented effectively
Spatial variability is not implicitly represented
Spatial analysis and filtering within polygons is impossible
Needs a lot of manual editing to get good quality
It always introduces hard boundaries
Simulation is difficult as each unit has different topological form
Display and plotting can be expensive
[Link] the GIS data input using digitiser ,scanner ,automatic &onscreen
digitisation.
GIS (Geographic Information Systems) data input is a crucial step in the creation and
maintenance of spatial databases. Various methods are used to input geographic data into
GIS, including digitization, scanning, automatic digitization, and on-screen digitization.
Here's an overview of each of these methods in the context of remote sensing and GIS:
1. Digitization:
• Manual Digitization: This method involves the manual tracing of features
from paper maps, aerial photographs, or other hard-copy sources onto a digital
map using a digitizing tablet or digitizing board. An operator uses a cursor or
puck to capture the coordinates of specific points or lines.
• Merits: Provides high precision when handled by skilled operators. Suitable
for converting existing paper maps and historical data into digital format.
• Demerits: Labor-intensive, time-consuming, and may introduce errors if not
done carefully. Limited to the features visible on source documents.
2. Scanning:
• Scanning and Rasterization: In this method, hard-copy maps, aerial
photographs, or satellite images are scanned using specialized scanners. The
scanned image is then converted into a raster format, where each pixel
represents a specific location on the Earth's surface.
• Merits: Rapid data capture, especially for large-scale maps and imagery.
Suitable for handling large volumes of data. Preserves the original appearance
of source documents.
• Demerits: Results in raster data, which may have limited accuracy for precise
spatial analysis. It does not capture vector data or attribute information
directly.
3. Automatic Digitization:
• Feature Extraction Algorithms: Automated software tools use feature
extraction algorithms to identify and digitize geographic features from
scanned or raster data. These algorithms detect edges, lines, and shapes to
create vector representations.
• Merits: Faster than manual digitization, especially for large datasets. Reduces
the risk of human error.
• Demerits: May require post-processing and manual editing to correct errors.
Accuracy depends on the quality of the algorithms and the input data.
4. On-Screen Digitization:
• Interactive Digitization: In this method, GIS software is used to display a
scanned image, aerial photograph, or satellite image on a computer screen. An
operator interacts with the software to trace and create vector features directly
on the digital image using a mouse or stylus.
• Merits: Allows for real-time editing and quality control. Suitable for
digitizing features that require interpretation or are not clearly defined on
source documents.
• Demerits: Operator-dependent accuracy. May be time-consuming for
complex datasets.
In remote sensing and GIS, the choice of data input method depends on factors such as the
type of data source, the level of accuracy required, the scale of the project, and available
resources. While manual digitization and on-screen digitization are suitable for small-scale
projects and feature interpretation, scanning and automatic digitization are preferred for
large-scale projects and data extraction from imagery. A combination of these methods is
often used to achieve the best results, ensuring both accuracy and efficiency in GIS data
input.
[Link] do you understand by spatial data model ? describe conceptual &logical data
model .
Spatial data model:
There are presently three types of representations for geographic data:
raster vector, and objects.
raster - set of cells on a grid that represents an entity (entity -->
symbol/color --> cells).
vector -an entity is represented by nodes and their connecting arc or line
segment (entity --> points, lines or areas --> connectivity)
object – an entity is represented by an object which has as one of its
attributes spatial information.
Conceptual & logical data model:
Conceptual and logical data models are fundamental components of Geographic Information
Systems (GIS) and are used to represent the structure and relationships of data in the context
of remote sensing and GIS applications.
1. Conceptual Data Model:
• Purpose: The conceptual data model provides a high-level, abstract
representation of the data and its relationships within a GIS or remote sensing
application. It is primarily concerned with defining what data should be stored
and managed and how it relates to real-world entities and phenomena.
• Characteristics:
• High-Level Abstraction: It focuses on the concepts and entities
relevant to the application domain rather than specific technical details.
• No Implementation Specifics: It is implementation-agnostic and does
not concern itself with how the data will be stored or accessed.
• User-Centric: The conceptual data model is designed to be easily
understood by non-technical stakeholders and domain experts.
• Components:
• Entities: These are the main objects or things of interest in the
application domain, such as rivers, buildings, or land parcels.
• Attributes: Attributes describe properties or characteristics of entities.
For example, the attributes of a building entity might include height,
area, and construction material.
• Relationships: Relationships define how entities are related or
connected to each other. For instance, a river entity may be related to a
city entity through the "flows through" relationship.
• Example: In a land management GIS, the conceptual model might include
entities like "Parcel," "Owner," and "Land Use," with attributes such as parcel
area, owner name, and land use type. Relationships might include "owned by"
between Parcel and Owner entities.
2. Logical Data Model:
• Purpose: The logical data model takes the conceptual model a step further by
defining how the data will be structured, organized, and stored within a
database or GIS system. It serves as a bridge between the high-level concepts
of the conceptual model and the physical implementation of the database.
• Characteristics:
• Implementation-Friendly: It includes details on data tables, keys,
relationships, and data types, making it suitable for database design
and development.
• Abstracted from Technology: While it defines the data structure, it
remains independent of specific database management systems or
software.
• Balances Complexity: The logical model balances the complexity of
the real-world domain (as represented in the conceptual model) with
the efficiency and simplicity required for data storage and retrieval.
• Components:
• Tables: Tables represent data entities and contain rows and columns
for data storage. Each table corresponds to an entity from the
conceptual model.
• Keys: Keys are used to uniquely identify records in tables. Primary
keys ensure each row is unique, and foreign keys establish
relationships between tables.
• Data Types: Data types specify the format of attribute values, such as
text, numbers, dates, or spatial data types for GIS applications.
• Relationships: Logical data models define how tables are related
through primary and foreign keys, establishing relationships consistent
with the conceptual model.
• Example: In a land management GIS, the logical model might include
database tables like "Parcel," "Owner," and "LandUse," with primary keys,
foreign keys, and data types defined. Relationships between tables would be
established, specifying how parcels are linked to owners and land use types.
In summary, the conceptual data model defines what data should be captured and how it
relates to real-world concepts, while the logical data model takes these concepts and
structures them in a way that can be implemented in a database or GIS system. Together,
these models provide a foundation for the design and development of effective GIS databases
and applications in the field of remote sensing and GIS.
[Link] the concept of integrated data analysis.
Integrated data analysis in remote sensing and GIS (Geographic Information Systems) refers
to the process of combining and analyzing multiple sources of spatial and non-spatial data to
gain a deeper understanding of a specific geographic area, phenomenon, or problem. This
approach allows users to leverage diverse datasets and information types to derive
meaningful insights and make informed decisions. Here's an explanation of the concept:
1. Multiple Data Sources: Integrated data analysis involves the use of various data
sources, which can include but are not limited to:
• Remote Sensing Data: Satellite imagery, aerial photographs, LiDAR (Light
Detection and Ranging) data, and other sensor data are commonly used for
capturing information about the Earth's surface and atmosphere.
• GIS Data: Geospatial data layers, such as land use/land cover, topography,
infrastructure, and administrative boundaries, provide foundational geographic
information.
• Field Data: Ground-truth data collected through surveys, measurements, and
sampling to validate and supplement remote sensing and GIS data.
• Demographic Data: Population statistics, socioeconomic data, and other
demographic information.
• Environmental Data: Climate data, environmental measurements (e.g., air
quality, water quality), and ecological data.
2. Data Integration: The key to integrated data analysis is merging and harmonizing
these diverse datasets into a single, coherent framework. This can involve data
preprocessing, such as georeferencing, data conversion, and data transformation, to
ensure that all data are compatible in terms of spatial reference, resolution, and
format.
3. Spatial Analysis: Once integrated, spatial analysis techniques are applied to explore
relationships, patterns, and trends within the data. This can include operations such as
overlay analysis, proximity analysis, spatial interpolation, and spatial statistics.
4. Multiscale Analysis: Integrated data analysis often involves examining data at
multiple scales, from local to regional to global. This allows for a comprehensive
understanding of the studied area and its context within larger geographic and
environmental systems.
5. Decision Support: Integrated data analysis provides decision-makers with valuable
information for various applications, including urban planning, environmental
management, disaster response, agriculture, and natural resource management. It can
aid in identifying trends, hotspots, anomalies, and areas of interest.
6. Example Applications:
• Environmental Monitoring: Combining remote sensing data with ground-
based measurements and GIS layers to monitor changes in land cover, assess
vegetation health, and study environmental impacts.
• Emergency Response: Using real-time satellite imagery and GIS data to
assess the extent of natural disasters (e.g., wildfires, floods) and plan rescue
and relief efforts.
• Urban Planning: Integrating demographic data, land use data, transportation
networks, and environmental data to support urban planning and infrastructure
development.
• Precision Agriculture: Analyzing soil data, weather data, and remote sensing
imagery to optimize crop management practices.
Compression ratio: The ratio of the two file sizes. E.g., original image is 100MB, after
compression, the new file is 10MB, then the compression ratio is 10:1.
Lossless compression: Lossless data compression is compression without any loss of
data quality.
Lossy compression: Compressing data and then decompressing it Eliminates
irrelevant information as well, and permits only an approximate reconstruction
of the original file.
[Link] briefly about attribute & spatial data .
In remote sensing and GIS (Geographic Information Systems), data can be broadly
categorized into two main types: attribute data and spatial data. These data types play distinct
but complementary roles in the analysis and representation of geographic information.
1. Attribute Data:
• Definition: Attribute data, also known as non-spatial data or tabular data,
consists of information that describes the characteristics, attributes, or
properties of geographic features. These attributes are typically stored in tables
or databases.
• Examples: Attribute data can include a wide range of information, such as
names, IDs, population counts, temperature values, land use categories,
ownership information, and more.
• Role: Attribute data provide context and additional information about spatial
features. They enable users to answer questions about what, when, who, and
how much. Attribute data are critical for attribute queries, statistical analysis,
and decision-making.
• Representation: Attribute data are typically represented in tabular form, with
rows corresponding to individual features (e.g., points, lines, polygons) and
columns representing specific attributes or properties associated with those
features.
• Analysis: Attribute data are used in various GIS operations, including
attribute queries, data classification, thematic mapping, and statistical analysis.
They are essential for creating data-driven maps and reports.
2. Spatial Data:
• Definition: Spatial data, also known as geospatial data, consists of
information that defines the location, shape, and spatial relationships of
geographic features. This data type represents the geometry or geography of
objects in the real world.
• Examples: Spatial data can include point coordinates, line segments,
polygons, boundary shapes, and any other geometric representations of
features on the Earth's surface.
• Role: Spatial data provide the spatial context necessary for visualizing and
analyzing the geographic world. They answer questions about where and how
features are distributed in space.
• Representation: Spatial data are typically represented as vector data or raster
data. Vector data represent features as discrete points, lines, and polygons,
while raster data represent features as grids of cells with pixel values.
• Analysis: Spatial data are fundamental to GIS operations such as spatial
queries, spatial analysis, cartographic mapping, overlay analysis, proximity
analysis, and network analysis. They are used for creating maps, conducting
spatial analysis, and making spatial decisions.
In summary, attribute data and spatial data are two core components of remote sensing and
GIS. Attribute data provide descriptive information about geographic features, while spatial
data define the location and geometry of those features. Combining both types of data enables
comprehensive analysis and visualization of geographic information, facilitating decision-
making, planning, and problem-solving in various fields, including environmental science,
urban planning, natural resource management, and more.
[Link] notes on following
I. Attribute data analysis.
II. Integrated data analysis.
Attribute data analysis:
Attribute data analysis in remote sensing and GIS (Geographic Information Systems)
involves examining and drawing insights from the non-spatial information associated with
geographic features. Attribute data analysis is a crucial component of geospatial analysis,
allowing users to answer questions related to the characteristics, attributes, and properties of
geographic features. Here are some key aspects of attribute data analysis in remote sensing
and GIS:
1. Data Exploration:
• Before conducting in-depth analysis, it's essential to explore the attribute data
to understand its content, structure, and quality. This includes checking for
missing values, outliers, and data distribution.
2. Data Summary and Descriptive Statistics:
• Calculate summary statistics, such as mean, median, mode, standard deviation,
and range, to gain a basic understanding of the data's central tendencies and
variability.
3. Data Visualization:
• Visualizing attribute data is essential for gaining insights. Common
visualization techniques include histograms, bar charts, scatter plots, box
plots, and thematic maps.
• Thematic mapping involves representing attribute data on a map, allowing
users to visualize spatial patterns and relationships among geographic features.
4. Spatial Queries:
• Attribute data can be used in combination with spatial data to perform spatial
queries. For example, you can query for all land parcels with an area greater
than a specified threshold or identify all buildings owned by a particular
individual.
5. Data Classification:
• Attribute data can be classified into categories or classes to simplify analysis
and interpretation. Common classification methods include equal intervals,
quantiles, natural breaks, and manual classification based on domain
knowledge.
6. Correlation Analysis:
• Investigate relationships between different attributes. For example, you can
examine the correlation between land use type and property values or analyze
how temperature relates to elevation in a region.
7. Spatial Joins and Relational Analysis:
• Combine attribute data from different datasets through spatial joins or
relational analysis. This can be used to associate attributes from one dataset
with features in another, facilitating data enrichment and analysis.
8. Geostatistical Analysis:
• Geostatistical methods, such as kriging or spatial autocorrelation analysis,
incorporate both attribute and spatial data to model and predict attribute values
across space.
9. Decision Support:
• Attribute data analysis plays a critical role in decision-making processes. For
example, in urban planning, it can inform decisions related to zoning,
infrastructure development, and resource allocation.
10. Quality Assessment:
• Evaluate the quality of attribute data, including data accuracy, completeness,
and consistency. Identify and rectify data errors and discrepancies.
11. Temporal Analysis:
• Analyze changes in attribute data over time, which is particularly important
for monitoring environmental variables, land use dynamics, and demographic
trends.
12. Statistical Modeling:
• Attribute data can be used in statistical models, such as regression analysis, to
understand the relationships between different attributes and predict outcomes.
13. Reporting and Visualization:
• Communicate the results of attribute data analysis through reports,
dashboards, and interactive maps to facilitate data-driven decision-making.
Attribute data analysis is an integral part of GIS and remote sensing applications, as it
enhances the understanding of spatial phenomena and supports a wide range of applications
in fields like environmental science, urban planning, agriculture, public health, and more.
Integrated data analysis in remote sensing and GIS (Geographic Information Systems) refers
to the process of combining and analyzing various types of data from multiple sources to gain
a comprehensive understanding of geographic phenomena and make informed decisions. This
approach recognizes that valuable insights can be derived by integrating data from diverse
sources, such as remote sensing imagery, geographic databases, field surveys, and other
relevant information. Here's a more detailed explanation of integrated data analysis in remote
sensing and GIS:
1. Data Integration:
• Integrated data analysis begins with the integration of spatial and non-spatial
data from various sources. This involves harmonizing data formats, spatial
references, and scales to ensure compatibility.
2. Remote Sensing Data:
• Remote sensing data sources can include satellite imagery, aerial photographs,
LiDAR data, thermal imagery, and radar data. These sources provide valuable
information about the Earth's surface and atmosphere.
3. GIS Data:
• Geographic Information Systems store and manage geospatial data layers,
such as land cover, land use, elevation, transportation networks, administrative
boundaries, and infrastructure data.
4. Field Data:
• Ground-truth data collected through field surveys, GPS measurements, and
sample observations are essential for validating and supplementing remote
sensing and GIS data.
5. Attribute Data:
• Attribute data, which describe the characteristics and properties of geographic
features, are often integrated with spatial data. These attributes may include
population counts, land values, environmental parameters, and more.
6. Temporal Data:
• Time-series data provide insights into temporal changes and trends. This can
include historical satellite imagery, climate data, and data collected over
different time intervals.
7. Analysis Techniques:
• Integrated data analysis employs various spatial analysis techniques to explore
relationships, patterns, and trends in the integrated dataset. Common
operations include overlay analysis, proximity analysis, change detection,
interpolation, and spatial statistics.
8. Multiscale Analysis:
• The analysis may occur at multiple scales, ranging from local to regional to
global, to capture variations in geographic phenomena and consider different
spatial contexts.
9. Decision Support:
• The insights gained from integrated data analysis serve as the foundation for
informed decision-making in a wide range of applications, including
environmental monitoring, urban planning, disaster management, agriculture,
and natural resource management.
10. Visualization and Reporting:
• Effective visualization techniques, such as maps, charts, and graphs, are used
to present the results of integrated data analysis. These visual representations
aid in communicating findings to stakeholders and decision-makers.
11. Modeling:
• Integrated data analysis often involves the use of spatial models to simulate
and predict future scenarios or to understand the impacts of different decisions
and policies.
12. Quality Assurance:
• Ensuring data quality and accuracy is a critical component of integrated data
analysis. Data validation and quality control procedures are essential to
maintain the integrity of the analysis results.
Integrated data analysis is a powerful approach that leverages the strengths of multiple data
sources to address complex spatial and environmental challenges. It plays a central role in the
fields of environmental science, land management, urban planning, disaster response, and
many other domains where a comprehensive understanding of geographic phenomena is
required for effective decision-making and problem-solving.