0% found this document useful (0 votes)
4 views48 pages

Unit 4 Notes

The document discusses various aspects of GIS, including the importance of hierarchical data models, buffering, overlay analysis, and data models such as raster and vector. It also covers the advantages and disadvantages of raster and vector data structures, the concept of spaghetti models, and the significance of Digital Elevation Models (DEMs) in terrain analysis. Additionally, it explains raster data compression techniques, highlighting lossless and lossy methods with examples.

Uploaded by

Sinduja N
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
4 views48 pages

Unit 4 Notes

The document discusses various aspects of GIS, including the importance of hierarchical data models, buffering, overlay analysis, and data models such as raster and vector. It also covers the advantages and disadvantages of raster and vector data structures, the concept of spaghetti models, and the significance of Digital Elevation Models (DEMs) in terrain analysis. Additionally, it explains raster data compression techniques, highlighting lossless and lossy methods with examples.

Uploaded by

Sinduja N
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

UNIT-4

DATA INPUT AND ANALYSIS


PART-A
[Link] the importance of Hierarchical Data Model.
 One to many relationship.
 Easy to update and expand.
 Easy data access for keys.
 Ideal for data that is inherently hierarchical.
 Poor access for associated attributes.
 Restrictive paths.
2. What is buffering in GIS analysis?
A process that defines/identifies areas within a specified distance of a coverage (e.g., within
300m of a stream), Process is applicable to points, line and polygons
3. Write the purpose of overlaying analysis.
 Data Integration:
Overlapping analysis allows for the integration of multiple layers of spatial data, such
as satellite imagery, topographic maps, and land use information. By overlaying these
layers, it becomes possible to combine and analyse different types of data to gain a
more comprehensive understanding of a particular geographic area or phenomenon.
 Spatial Relationships:
Overlapping analysis helps in identifying and analysing spatial relationships between
different features or datasets. This can be crucial for various applications, such as land use
planning, environmental monitoring, and disaster management, where understanding how
different elements on the Earth's surface interact with each other is essential for decision-
making.
4. What is map overlay ?
 Overlay is a GIS operation that superimposes multiple data sets (representing
different themes) together for the purpose of identifying relationships between them.
 An overlay creates a composite map by combining the geometry and attributes of the
input data sets
5. What is object oriented data base model ?
 Definition: realization of the discrete model of real world using an object centered
approach in which an object has both physical (attribute) and geometric
characteristics.
 Different types of objects can interact because they are not confined to separate
layer

[Link] spatial data model will be advantages for urban growth studies?
For urban growth studies in remote sensing and GIS, a advantageous spatial data model is the
vector data model.
Here are two key reasons why the vector data model is advantageous for urban growth
studies:
1. Precision and Detail: Vector data models represent geographic features as discrete
objects with precise boundaries, such as points, lines, and polygons. This level of
detail is well-suited for studying urban growth, where the delineation of city
boundaries, roads, buildings, and other infrastructure elements is crucial for accurate
analysis.
2. Topology: The vector data model can capture topological relationships between
geographic features, making it ideal for modeling the connectivity and adjacency of
urban features. This is important for understanding how urban areas expand, where
new developments occur, and how they relate to existing infrastructure.
Overall, the vector data model provides the necessary level of detail, precision, and
topological information required for in-depth urban growth analysis in remote sensing
and GIS.
7. What are the levels of measurement used in GIS?
 Nominal (categories)
 Ordinal (ordered categories)
 Interval (equal intervals)
 Ratio (true zero point)
8. What do you mean by data compression ?
 Large amounts of data can create enormous problems in storage space and
transmission time.
 The main reasons of data compression could be summarized as:
•Multimedia data have large data volume
•Difficulty sending real-time uncompressed data over current
network
 The design goal of image compression is to represent images with as few bits as
possible, to save storage and transmission channel capacity.
 Compression data principle: “Eliminate data redundancy and try to find a code
with less data volume.”
9. What are the advantages & disadvantages of raster and vector data ?
Advantages of raster data structures
 Simple data structure
 Easy to generate (e.g. from remote sensing or scandigitizing)
 Easy workflows and analysis
 Technology is cheap and is being energetically developed
 Simulation is very easy as each spatial unit has the same shape and size
 Same set of grid cells are used for several variables
Disadvantages of raster data structures
 Non-adaptive data structure

 Tends to generate huge files, depending on resolution


 Cell arrangement is usually random and does not respect natural borders

 Errors in evaluating perimeter of shape


 Topology or network linkages are difficultto establish

 Geometric transformations are difficult to handle


 Use of large cells to reduce data volumes result into loss of information
Advantages of vector data structures
 Small amount of data
 Logical data structure
 Attributes are combined with objects
 Preserves quality after interactivity (e.g. scaling)
 Topology described with network linkages
 Retrieval, updating and generalization of graphics and attributes are possible
 Widely used to described administrative zone
Disadvantages of vector data structures
 Complex data structure
 Continuous data is not represented effectively
 Spatial variability is not implicitly represented
 Spatial analysis and filtering within polygons is impossible
 Needs a lot of manual editing to get good quality
 It always introduces hard boundaries
 Simulation is difficult as each unit has different topological form
 Display and plotting can be expensive
10. what is meant by spaghetti model ?
The term "spaghetti model" in the context of remote sensing and GIS (Geographic
Information Systems) typically refers to a situation where there is a complex and unorganized
network of spatial data layers or vector features in a GIS project. In a spaghetti model:
1. Complexity: There are numerous overlapping and interconnected layers of data, often
making it difficult to discern individual features or understand the relationships
between them.
2. Lack of Organization: The data layers may not be properly organized or structured,
leading to confusion and inefficiency when trying to analyze or visualize spatial
information.
A spaghetti model is generally undesirable in GIS and remote sensing because it can
hinder the ability to effectively manage and utilize spatial data. GIS professionals strive to
create organized and logically structured data models to ensure that data can be efficiently
analyzed, visualized, and interpreted for various applications, such as land use planning,
environmental assessment, or disaster management.
[Link] is buffering ?
 A process that defines/identifies areas within a specified distance of a coverage (e.g.,
within 300m of a stream),Process is applicable to points, line and polygons.
 A buffer in GIS is a zone around a map feature measured in units of distance or time.
 A buffer is an area defined by the bounding region determined by a set of points at a
specified maximum distance from all nodes along segments of an object CONTD
12. What is data model ?
 Data Model : an abstraction of real world entities and their relationships into
structures that can be implemented with a computer language.
 Types of Data Base Models in GIS:
[Link]-Relational Data Mode
2. Hierarchical Data Models
3. Network Data Models
4. Relational Data Models
5. Object Oriented Data Models
6. Spatial data models
a) raster
b) vector
[Link] between spatial and non spatial data

[Link] Spatial data Non spatial data


1 It includes location, shape, size and  It is also known attribute or
orientation information of features or characteristic data.
objects.  It consists of the characteristics of
spatial features which are
independent of all geometric
considerations.
2 A particular square in which its center The non-spatial data of town comprise of
(the intersection of its diagonals) name of the town, its population, settlement
specifies its location; its shape is a type, means of transportation and
square; length of one of its sides communication, administration set-up,
specifies its size and angle its diagonals education institutions, occupations and
e.g., the xaxis specifies its orientation. facilities.
3 Spatial data includes spatial  It is important to note that all the
relationships, for example, the above mentioned data of town are
arrangement of three stumps in a cricket not dependent on their location
ground identities.
 Hence, non-spatial data is
independentfrom location
information.
PART-B

[Link] are the GIS data models ? Explain any three


Data Model : an abstraction of real world entities and their relationships into structures that
can be implemented with a computer language.
Types of Data Base Models in GIS:
1. Entity-Relational Data Model
2. Hierarchical Data Models
3. Network Data Models
4. Relational Data Models
5. Object Oriented Data Models
6. Spatial data models
1. raster
2. Vector

 A conceptual data model in which information is represented by entities and


relationships between entities
 one to many relationship. ′ easy to update and expand. ′
 one to many relationship.
 easy to update and expand.
 easy data access for keys.
 ideal for data that is inherently hierarchical.
 poor access for associated attributes.
 Restrictive paths.
 one to many relationships
 many to many relationships.
 reduces redundancy.
 more flexible paths to data.
 very fast
 Pointers* expensive and difficult to update when inserting and deleting.
 *Pointer – A memory address that references the start of data in the RAM
RELATIONAL DATA MODEL

 Data stored as records known as tuples grouped together in two-dimensional tables


known as relations.
 Whereas hierarchical structures rely on the hierarchy and networks depend on
pointers to associate entities, the relational model uses data redundancy in the form of
unique keys that identify records in each file.
 Simplifies data maintenance because data for an entity type is stored in simple tables.
 Relational joins are used to cross reference entities using a primary key in one table
and a foreign key in another table.
 Thus, in order to perform relational joins there needs to be at least one column in
common between tables being related.
 The relational model is design to reduce redundancy of data whenever possible.
• structures very flexible.
• insert and delete easy.
• often use sequential search unless previously sorted

Definition: realization of the discrete model of real world using an object centered approach
in which an object has both physical (attribute) and geometric characteristics. ′
 Different types of objects can interact because they are not confined to separate layers
SPATIAL DATA MODEL
 There are presently three types of representations for geographic data: raster vector,
and objects.
 raster - set of cells on a grid that represents an entity (entity --> symbol/color -->
cells).
 vector -an entity is represented by nodes and their connecting arc or line segment
(entity --> points, lines or areas --> connectivity) ′
 object – an entity is represented by an object which has as one of its attributes spatial
information
[Link] about the data elevation model in GIS
 Digital Elevation Model (DEM) is the digital representation of the land surface
elevation with respect to any reference datum.
 DEM is frequently used to refer to any digital representation of a topographic
surface.
 DEM is the simplest form of digital representation of topography
 DEMs are used to determine terrain attributes such as elevation at any point, slope
and aspect.
 Terrain features like drainage basins and channel networks can also be identified
from the DEMs.
 DEMs are widely used in hydrologic and geologic analyses, hazard monitoring,
natural resources exploration, agricultural management etc
 Hydrologic applications of the DEM include groundwater modelling, estimation of
the volume of proposed reservoirs, determining landslide probability, flood prone
area mapping etc.
 Three main type of structures used are the following.
a) Regular square grids
b) Triangulated irregular networks (TIN)
c) Contours
 A digital elevation model (DEM) is a digital model or 3D representation of a
terrain's surface, created from terrain elevation data.
 There are three similar names as digital elevation model (DEM), digital terrain
model (DTM) and digital surface model (DSM).
 DEM is a subset of the DTM, which also represents other morphological
elements
 ′Digital Elevation Model: It is a bare-earth raster grid referenced to a vertical
datum. The built (power lines, buildings and towers) and natural (trees and other
types of vegetation) aren’t included in a DEM.
 Digital Terrain Model (DTM): DTM is a mathematical representation (model) of
the ground surface, most often in the form of a regular grid, in which a unique
elevation value is assigned to each pixel. ′
 Digital Surface Model (DSM): Represents the MSL elevations of the reflective
surfaces of trees, buildings, and other features elevated above the “Bare Earth”
Types of DEM:
 Triangular irregular network (TIN) (primary DEM ): It is a vector-based. It uses
irregular sampling points connected through non-overlapping triangles.
 Raster DEM (secondary DEM): Also known as a heightmap when representing
elevation
 Gridded DEM (GDEM): Consists of regularly placed, uniform grids with the
elevation information of each grid

Methods for obtaining elevation data used to create DEMs:


Lidar-Light Detection And Ranging (sometimes Light Imaging, Detection, And Ranging)
 Stereo photogrammetry from aerial surveys
 Interferometry from radar data
 Real Time Kinematic GPS
 Topographic maps
 Theodolite or totalstation
 Doppler radar
 Surveying and mapping drones
 Range imaging
LiDAR and DEM:
 Light Detection and Ranging (LiDAR) sensors operate on the same principle as
that of laser equipment.
 Pulses are sent from a laser onboard an aircraft and the scattered pulses are
recorded.
 The time lapse for the returning pulses is used to determine the two-way distance
to the object
DEM applications:
1. Estimating elevation
2. Estimating slope and aspect
3. Determining drainage networks
4. Determining the watershed
5. Terrain stability – Areas prone to avalanches are high slope areas with sparse
vegetation, which is useful when planning a highway or residential subdivision.
6. Soil mapping – DEMs assist in mapping soils which is a function of elevation (as
well as geology, time and climate)
7. To create a profile graph from digitized features of a surface
BHUVAN data sources-CARTOSAT in India:
 Indian Space Research Organisation provides all DEM data of India through it portal
Bhuvan.
 ′Catrosat-1 and 2 provide all these source
3. Discuss about raster data compression with suitable examples
Raster compression plays a crucial role in remote sensing and Geographic Information
Systems (GIS) to reduce the storage and transmission requirements of large spatial datasets.
Introduction to Raster Data: Raster data represents geographic information as a grid of cells,
where each cell stores a value. This format is commonly used in remote sensing and GIS to
represent data such as satellite imagery, elevation models, and land cover maps.
1. Need for Compression: Raster datasets can be massive, often requiring significant
storage space and time to transmit over networks. Compression is essential to mitigate
these challenges.
2. Lossless vs. Lossy Compression: There are two main types of raster compression:
lossless and lossy.
Lossless compression retains all original data, ensuring data accuracy but may not
achieve high compression ratios.
Lossy compression sacrifices some data accuracy for greater compression, suitable for
applications where minor data loss is acceptable.
3. Lossless Compression Examples:
Run-Length Encoding (RLE): A simple method where consecutive identical values
are encoded as a count and a single value. Example: Compression of a binary land
cover map with large contiguous areas of the same land type.
Lempel-Ziv-Welch (LZW): A dictionary-based method used in formats like TIFF.
Suitable for compressing multispectral imagery.
4. Lossy Compression Examples:
JPEG (Joint Photographic Experts Group): Widely used for compressing satellite
imagery. JPEG reduces file size by encoding frequencies and quantizing color values.
Example: Compressing a high-resolution satellite image for web viewing.
Discrete Cosine Transform (DCT): Used in JPEG compression, it converts spatial
data into frequency domain data for compression.
Wavelet Transform: Suitable for multispectral and hyperspectral imagery
compression, it decomposes data into various scales and frequencies.
5. Compression Factors: The choice of compression method and the degree of
compression depend on the application's requirements. For critical applications like
scientific analysis, lossless compression is preferred, while lossy compression is
acceptable for visual applications.
6. Compression in GIS: In GIS, raster data compression is used to reduce the storage and
processing requirements of large datasets. For example, a high-resolution digital
elevation model (DEM) can be compressed to save disk space while preserving
essential topographic information.
7. Remote Sensing Applications: In remote sensing, compression is vital for satellite
data transmission and archiving. For instance, hyperspectral data from remote sensing
satellites like Sentinel-2 are compressed before transmission to Earth, reducing the
data volume while maintaining key spectral information.
8. Data Transmission: Compressed raster data is easier to transmit over limited
bandwidth connections, making it more feasible to share remote sensing data between
researchers and organizations worldwide.
9. Limitations:
Lossy compression may result in the loss of fine details in imagery.
Compression artifacts can affect the quality of visual interpretation.
Compression and decompression processes require computational resources.
10. Conclusion: Raster data compression is essential in remote sensing and GIS to
manage the storage, transmission, and processing of large datasets efficiently. The
choice between lossless and lossy compression should be made based on the specific
requirements of the application.
11. Future Trends: Ongoing research in remote sensing and GIS focuses on developing
more advanced compression techniques, such as machine learning-based approaches,
to balance data size reduction with preservation of information content.
12. Overall Impact: Effective data raster compression significantly enhances the usability
and accessibility of remote sensing and GIS data, making it feasible to work with
large datasets for various scientific, environmental, and commercial applications.
4. Distinguish the difference between raster &vector data
Raster and vector data are two fundamental data models used in remote sensing and
Geographic Information Systems (GIS) to represent and analyze spatial information.
Here, I'll distinguish the differences between raster and vector data in the context of
remote sensing and GIS:
1. Data Structure:
Raster Data (Grid-Based): Raster data is organized as a regular grid or matrix of cells,
where each cell represents a discrete unit of space (pixel). Each cell stores a single
value or multiple values, such as pixel values in an image or elevation values in a
Digital Elevation Model (DEM).
Vector Data (Geometry-Based): Vector data represents spatial features using
geometric objects such as points, lines, and polygons. Points represent discrete
locations, lines represent linear features, and polygons represent enclosed areas.
2. Data Representation:
Raster Data: Raster data is suitable for representing continuous phenomena and
regularly sampled data, making it well-suited for imagery, elevation data, and
continuous environmental variables. It is pixel-based and provides a continuous field
representation.
Vector Data: Vector data is best for representing discrete, distinct features with
precise boundaries. It excels in representing objects like buildings, roads,
administrative boundaries, and individual trees.
3. Data Volume:
Raster Data: Raster datasets can be large, especially for high-resolution imagery. The
volume of data increases proportionally with spatial resolution, making it storage-
intensive for detailed imagery.
Vector Data: Vector data typically consumes less storage space compared to raster data
for the same area because it only records the geometric coordinates of feature boundaries
and attributes.
4. Data Analysis:
Raster Data: Raster data is suitable for various analytical operations involving continuous
data, such as terrain analysis, image classification, and spatial modeling using techniques
like interpolation and proximity analysis.
Vector Data: Vector data is well-suited for operations involving discrete features, such as
overlay analysis, network analysis, and attribute queries. It allows for precise spatial
analysis and topological relationships.
5. Data Precision:
Raster Data: Raster data may suffer from pixelation, where fine details can be lost when
represented at lower resolutions. However, it's suitable for representing data with
continuous variation, even if at the expense of precision.
Vector Data: Vector data maintains precise geometric representation, making it ideal for
capturing fine details and accurately representing feature boundaries.
6. Data Visualization:
Raster Data: Raster data is typically visualized as images or thematic maps, making it
suitable for visual interpretation and cartographic display.
Vector Data: Vector data is used to create maps with precise, well-defined features and
labels. It's also suitable for symbolizing different feature types.
7. Common Applications:
Raster Data: Remote sensing imagery, satellite data, elevation models, climate grids, and
continuous environmental variables.
Vector Data: Land parcels, transportation networks, administrative boundaries, utilities,
and point-of-interest data.
In summary, the choice between raster and vector data models in remote sensing and GIS
depends on the nature of the data, the specific analytical tasks, and the required precision.
Raster data is well-suited for continuous data and imagery, while vector data excels at
representing discrete features and maintaining geometric precision. Often, GIS projects
involve a combination of both data types to address different aspects of spatial analysis
and representation.
[Link] a short note on [Link] retrieval &[Link] analysis
Data retrieval
Data retrieval in remote sensing and Geographic Information Systems (GIS) involves the
process of obtaining, accessing, and using spatial data and information for various
applications. Here is a short note on data retrieval in these fields:

Data Sources: Data retrieval in remote sensing and GIS begins with identifying and
selecting appropriate data sources. These sources may include:
1. Remote Sensing Satellites: Obtaining satellite imagery and data from Earth-
observing satellites such as Landsat, Sentinel, and MODIS for various applications
like land cover mapping, environmental monitoring, and disaster management.
2. Aerial Photography: Collecting aerial imagery through aircraft or drones for high-
resolution mapping, land use planning, and infrastructure monitoring.
3. Ground Surveys: Acquiring ground-based data through field surveys, GPS
measurements, and sensor deployments to validate and supplement remote sensing
data.
4. GIS Databases: Accessing existing GIS databases and repositories containing
geographic data layers, such as land parcels, transportation networks, and
demographic information.
5. Open Data Portals: Utilizing open data portals and government agencies' resources
to access publicly available geospatial datasets and maps.

Data Retrieval Methods:


1. Online Sources: Many remote sensing and GIS data sources are available online
through data portals, government agencies, and commercial providers. Users can
search, browse, and download data directly from these sources.
2. Data Purchase: In some cases, users may need to purchase data from commercial
providers or agencies that offer specialized remote sensing or GIS datasets. This is
common for high-resolution satellite imagery.
3. Data Interpolation: In GIS, users can interpolate data to fill gaps or create
continuous surfaces from sparse data points. Interpolation techniques like kriging or
inverse distance weighting are commonly used.
4. Geospatial APIs: Application Programming Interfaces (APIs) provide programmatic
access to geospatial data. Users can retrieve data dynamically and integrate it into
their applications or workflows.

Data Formats and Standards:

Data retrieval involves ensuring compatibility with specific data formats and standards:
1. Raster Formats: Remote sensing data, such as satellite imagery and elevation
models, are often stored in raster formats like GeoTIFF or HDF.
2. Vector Formats: GIS data, such as shapefiles, GeoJSON, or KML, use vector
formats to represent geographic features and attributes.

Data Integration: Once retrieved, data is integrated into GIS software or remote sensing
platforms for analysis, visualization, and modeling. This involves:
1. Data Preprocessing: Cleaning and preparing the data, including georeferencing,
mosaicking, and removing artifacts or noise.
2. Geospatial Analysis: Performing spatial analysis tasks like spatial querying, overlay
analysis, and geoprocessing to extract meaningful information.
3. Visualization: Creating maps, charts, and visual representations to convey spatial
information effectively.
4. Modelling: Using retrieved data as inputs for spatial models and simulations, such as
hydrological modelling or land use change prediction.

Data Updates: Data retrieval is an ongoing process as new data becomes available, or
existing data is updated. Regular updates and maintenance are essential for ensuring the
accuracy and relevance of GIS and remote sensing datasets.

In conclusion, data retrieval in remote sensing and GIS involves the identification,
acquisition, and integration of spatial data from various sources to support decision-making,
analysis, and modelling in a wide range of applications, from environmental monitoring to
urban planning. The choice of data sources and methods depends on the specific project
requirements and objectives.

Overlay analysis
Overlay analysis is a fundamental spatial analysis technique used in both remote sensing and
Geographic Information Systems (GIS). It involves the layering or combination of multiple
spatial datasets to derive new information, make informed decisions, and gain insights into
the relationships between different geographic features. Here's a short note on overlay
analysis in remote sensing and GIS:

Purpose of Overlay Analysis:

Overlay analysis serves several important purposes in remote sensing and GIS:
1. Spatial Querying: It allows users to answer complex spatial questions by identifying
features that meet specific criteria or conditions.
2. Integration of Information: Overlay combines diverse data sources, such as land
use, population, and environmental variables, to create composite datasets for
comprehensive analysis.
3. Decision Support: It aids in decision-making by visualizing spatial relationships and
patterns, making it valuable for urban planning, environmental management, and
resource allocation.

Key Concepts and Steps:


1. Spatial Layers: Overlay analysis requires multiple spatial layers, each representing
different geographic features or attributes. These layers can include vector data (e.g.,
points, lines, polygons) and raster data (e.g., satellite imagery, elevation models).
2. Overlay Operations: Overlay operations involve combining layers based on spatial
relationships, such as intersection, union, difference, and proximity. Common
operations include:
• Union: Combines features from multiple layers into a single layer, preserving
all attributes.
• Intersection: Identifies the common area or features shared by multiple
layers.
• Difference: Highlights the areas that differ between two layers.
• Buffering: Creates a zone or buffer around features within a specified
distance.
3. Attribute Data: In addition to spatial relationships, overlay analysis considers
attribute data associated with features. Attributes may include population counts, land
use categories, or environmental characteristics.
4. Resultant Maps: The output of overlay analysis is often displayed as thematic maps,
where areas with specific combinations of features or conditions are symbolized
differently. These maps help visualize patterns and relationships.

Applications:

Overlay analysis is widely used in various fields, including:


1. Urban Planning: Identifying suitable locations for new developments, assessing land
use changes, and analyzing transportation networks.
2. Environmental Management: Assessing environmental impact, habitat suitability,
and conservation planning.
3. Epidemiology: Identifying disease clusters and understanding the spatial distribution
of health-related factors.
4. Natural Resource Management:Analyzing the availability and suitability of
resources like water, minerals, and timber.
5. Emergency Management: Identifying vulnerable areas during disasters and planning
evacuation routes.
6. Site Selection: Choosing optimal locations for facilities such as schools, hospitals,
and retail outlets.

Challenges:

Overlay analysis can be complex and computationally intensive, particularly when dealing
with large datasets or numerous layers. Ensuring the quality and accuracy of input data is
crucial, as errors or inaccuracies can propagate into the results. Additionally, defining
appropriate criteria and rules for overlay operations requires careful consideration.

In conclusion, overlay analysis is a powerful spatial analysis technique in remote sensing and
GIS that allows users to combine and analyze spatial data layers to extract valuable insights,
make informed decisions, and address a wide range of spatial problems in diverse fields.

[Link] various application of object oriented data base models


Object-oriented database models have found numerous applications in the fields of remote
sensing and Geographic Information Systems (GIS) due to their ability to efficiently manage
and manipulate complex spatial data. Here are various applications of object-oriented
database models in remote sensing and GIS:
1. Spatial Data Representation: Object-oriented database models provide an effective
way to represent spatial data as objects. Geographic features like points, lines, and
polygons can be stored as objects with associated attributes, allowing for more
comprehensive and structured data representation.
2. Data Integration: Object-oriented databases enable the integration of diverse data
types and formats. Remote sensing imagery, vector layers, attribute data, and
metadata can all be stored as objects within the same database, simplifying data
management and retrieval.
3. Multisource Data Fusion: In remote sensing, data fusion involves combining
information from multiple sources, such as satellite imagery, aerial photography, and
ground-based sensors. Object-oriented databases facilitate the fusion of heterogeneous
data, allowing users to analyze and visualize information from different sensors and
platforms in a unified manner.
4. Temporal Analysis: Temporal data, such as time-series satellite imagery, is critical
in both remote sensing and GIS applications. Object-oriented databases support the
efficient storage and retrieval of temporal data, enabling the analysis of changes over
time, such as land cover change detection and monitoring vegetation growth.
5. 3D and LiDAR Data: Object-oriented databases are well-suited for handling 3D
spatial data, including LiDAR point clouds. They can store and manage 3D objects,
terrain models, and building information models (BIM), making them valuable for
urban planning, infrastructure management, and 3D visualization.
6. Topological Relationships: Object-oriented databases support the modeling of
topological relationships between spatial objects. This is essential for GIS operations
like connectivity analysis, network modeling, and spatial queries.
7. Complex Spatial Analysis: Object-oriented databases offer capabilities for complex
spatial analysis and modeling. Users can define custom spatial operations and queries,
making them suitable for advanced geospatial analytics.
8. Web-Based GIS: Object-oriented databases can serve as the backend for web-based
GIS applications. They enable efficient data retrieval and management for online
mapping and spatial data services.
9. Geospatial Metadata Management: Metadata, including information about data
sources, accuracy, and metadata standards, can be efficiently managed within object-
oriented databases. This is crucial for maintaining data quality and traceability.
10. Natural Resource Management: Object-oriented databases are valuable for
managing and analyzing natural resource data, such as forestry inventory data,
wildlife habitat modeling, and hydrological modeling.
11. Disaster Management: In disaster management and emergency response, object-
oriented databases facilitate the integration of real-time sensor data, satellite imagery,
and geospatial information to assess and mitigate the impact of disasters.
12. Land Use Planning: Object-oriented databases support land use planning by storing
cadastral data, zoning regulations, and land parcel information. This aids in decision-
making related to land allocation and urban development.

In summary, object-oriented database models have a wide range of applications in remote


sensing and GIS, enhancing data management, analysis, and integration capabilities. Their
ability to represent complex spatial data in a structured manner makes them particularly
valuable for handling the diverse and intricate datasets common in these fields.

[Link] in detail about the various type of data &data formats used in GIS
File formats
 File formats are intended to store particular kinds of digital information.
 E.g. JPEG format is designed only to store still images, while the GIF format
supports storage of both still images and simple animations.
 The most well-known formats have file specifications that describe exactly how the
data is to be encoded.
Digital image formats
 The standard greyscale images use 256 shades of grey from 0 (black) to 255
(white).
 For a given number of pixels, considerably more data is required to represent a
color image.
 The important data formats used for GIS are:
1. GIF
2. JPEG
3. TIFF
4. BMP
5. PNG
1. GIF (Graphic Interchange Format):
 GIF files offer optimum compression (smallest files) for solid color graphics.
 The GIF format can only show 256 colors but it can choose any of the 16 million colors
in a 24- bit image.
⚫ 8 bit = 256 colors
⚫ 7 bit = 128 colors
⚫ 6 bit = 64 colors
⚫ 5 bit = 32 colors
⚫ 4 bit = 16 colors
⚫ 3 bit = 8 colors
⚫ 2 bit = 4 colors
⚫ 1 bit = 2 colors (usually B/W)
[Link] (Tagged Image File Format):
 TIFF files have many formats: Black and white, greyscale, 4- and 8-bit color, full
color (24-bit) images.
 TIFF files support the use of data compression using many compression standards.
[Link] (Bitmap):
 B/W Bitmap is monochrome and the color table contains two entries.
 Each bit in the bitmap array represents a pixel. ′
 4-bit Bitmap has a maximum of 16 colors.
 8-bit Bitmap has a maximum of 256 colors.
[Link] (Joint Photographic Expert Group Format):
 A JPEG image provides very good compression, but does not uncompress exactly as
it was; JPEG is a lossy compression techniques.
 Compressions up to 50:1 are easily obtainable.
[Link] (Portable Network Graphics):
 The PNG format was designed to replace the antiquated GIF format, and to some
extent, the TIFF format. ′
 It utilizes lossless compression.
 It is a universal format that is recognized by the World Wide Web consortium, and
supported by modern web browsers.
File formats and supported data types
[Link] detailed notes on vector & raster data structures .
1. Raster data structures:
⚫ represents geography via grid cells
⚫ single values associated with each cell
⚫ typically 8 bits assigned to values therefore 256 possible values (0-255)
[Link] data structures:
⚫ represents geography via coordinates
⚫ point (node): 0-dimension
⚫ line (arc): 1-dimension
⚫ polygon : 2-dimensions
Usage Scenarios of: Raster Data Structures
 Photos
 Photogrammetry and remote sensing
 Scanned images of maps
 Terrain modelling
 Landcover analysis ′
 Hydrologic modelling and analysis
 General GIS surface modelling and analysis for continuous surface
Usage scenarios of:Vector Data Structures
 CAD, technical drawings
 Street or river networks,
 cadastral map
 Network analysis Cartography
Advantages of Raster Data Structures
 Simple data structure ′ Easy to generate (e.g. from remote sensing or
scandigitizing) ′
 Easy workflows and analysis
 Technology is cheap and is being energetically developed
 Simulation is very easy as each spatial unit has the same shape and size
 Same set of grid cells are used for several variable
Disadvantages of Raster Data Structures
 Non-adaptive data structure
 Tends to generate huge files, depending on resolution ′
 Cell arrangement is usually random and does not respect natural borders
 Errors in evaluating perimeter of shape ′
 Topology or network linkages are difficultto establish
 Geometric transformations are difficult to handle
 Use of large cells to reduce data volumes result into loss of information
Advantages of Vector Data Structures
 Small amount of data
 Logical data structure
 Attributes are combined with objects
 Preserves quality after interactivity (e.g. scaling)
 Topology described with network linkages ′
 Retrieval, updating and generalization of graphics and attributes are possible
 Widely used to described administrative zones
Disadvantages of Vector Data Structures
 Complex data structure
 Continuous data is not represented effectively
 Spatial variability is not implicitly represented
 Spatial analysis and filtering within polygons is impossible
 Needs a lot of manual editing to get good quality
 It always introduces hard boundaries
 Simulation is difficult as each unit has different topological form
 Display and plotting can be expensive
[Link] in detailed about spatial analysis & overlay analysis.

Spatial analysis and overlay analysis are essential techniques in the fields of remote sensing
and Geographic Information Systems (GIS). They involve the manipulation, integration, and
interpretation of spatial data to extract meaningful information and gain insights into the
relationships between geographic features. Let's discuss these concepts in detail:

Spatial Analysis:

Spatial analysis is a broad term that encompasses a wide range of operations and techniques
used to analyze and interpret spatial data. It involves processing and examining the spatial
relationships, patterns, and attributes of geographic features. Here are key aspects of spatial
analysis in remote sensing and GIS:
1. Geoprocessing Operations: Spatial analysis includes various geoprocessing
operations such as buffering, clipping, merging, and dissolving. These operations
modify or extract information from spatial datasets.
2. Proximity Analysis: Proximity analysis involves measuring distances between
features, finding nearest neighbors, and identifying areas within a certain distance of
specific features. It is crucial for applications like site selection and wildlife habitat
analysis.
3. Spatial Queries: Spatial queries involve selecting or filtering features based on their
spatial relationships with other features. For example, finding all houses within a
certain distance from a park.
4. Spatial Statistics: Spatial statistics techniques analyze patterns and distributions of
features, including hotspot analysis, spatial autocorrelation, and spatial interpolation
methods like kriging.
5. Network Analysis: Network analysis deals with route optimization, finding the
shortest path, and solving transportation problems. It is valuable in urban planning,
logistics, and emergency response.
6. Overlay Analysis: Overlay analysis, which we'll discuss in detail in the next section,
is a critical component of spatial analysis.
7. Terrain Analysis:Analyzing elevation and slope data to understand terrain
characteristics, suitable locations for infrastructure, and flood modeling.
8. Temporal Analysis: Examining changes and trends in spatial data over time,
especially important in monitoring environmental changes and land use/land cover
changes.

Overlay Analysis:

Overlay analysis is a specific type of spatial analysis that involves combining multiple spatial
datasets, often in raster or vector format, to derive new information. The primary objective is
to explore the intersection of various datasets to answer complex spatial questions and
generate new data layers. Here's a detailed look at overlay analysis:
1. Data Layers: Overlay analysis requires multiple layers of spatial data, each
representing different geographic features or attributes. These layers can include land
cover, land use, elevation, administrative boundaries, and more.
2. Overlay Operations: Overlay operations combine these layers based on spatial
relationships. Common overlay operations include union (combining features),
intersection (finding common areas), difference (highlighting differences), and
buffering (creating zones around features).
3. Attribute Data: In addition to the spatial relationships, overlay analysis considers
attribute data associated with features. Attributes may include population counts, land
use categories, or environmental characteristics.
4. Resultant Maps: The output of overlay analysis is often visualized as thematic maps,
where areas with specific combinations of features or conditions are symbolized
differently. These maps help visualize patterns and relationships.
5. Applications: Overlay analysis is widely used in urban planning, environmental
management, land suitability analysis, epidemiology, disaster management, and many
other fields. Examples include land use planning, habitat suitability modeling, and
flood risk assessment.
6. Challenges: Overlay analysis can be computationally intensive, especially with large
datasets or numerous layers. Care must be taken to ensure data quality and accuracy,
as errors can propagate into the results.

In summary, spatial analysis and overlay analysis are crucial components of remote sensing
and GIS. They enable professionals to extract valuable information from spatial data, make
informed decisions, and gain insights into complex spatial relationships, ultimately
supporting a wide range of applications in fields such as environmental science, urban
planning, public health, and more.
[Link] about modelling using GIS with a case study

Modelling using Geographic Information Systems (GIS) is the process of using spatial data
and analytical tools to create representations of real-world phenomena, make predictions, and
support decision-making. GIS models help users understand complex spatial relationships,
simulate scenarios, and analyze the potential impacts of different factors on geographic areas.
Below, I'll explain the concept of modelling using GIS and provide a case study to illustrate
its application.

Concept of Modelling Using GIS:

Modelling using GIS involves the following key steps:


1. Data Acquisition: The first step is to gather relevant spatial and attribute data. This
data can come from various sources, including remote sensing, field surveys,
government databases, and historical records. It should accurately represent the
geographic area of interest.
2. Data Preparation: Data preparation involves cleaning, organizing, and converting
data into a format compatible with GIS software. This step also includes
georeferencing to ensure that all data layers align correctly in space.
3. Model Development: In this stage, GIS analysts use spatial analysis tools to create
models that represent the real-world processes or relationships they want to study.
This can involve developing mathematical equations, rule-based models, or
simulation models.
4. Parameterization: Many GIS models require parameterization, which involves
assigning values to variables and parameters within the model. These values are often
based on empirical data, expert knowledge, or statistical analysis.
5. Model Calibration and Validation: Before applying the model to real-world
scenarios, it's essential to calibrate and validate it. Calibration fine-tunes the model to
match observed data, while validation tests its accuracy and reliability against
independent data.
6. Scenario Testing: Once validated, the model can be used to simulate various
scenarios by inputting different datasets or changing parameters. This allows
decision-makers to explore potential outcomes under different conditions.
7. Results Analysis: GIS models generate output data that can be analyzed and
visualized to gain insights into the modeled phenomena. Users can create maps,
charts, and reports to interpret and communicate the results effectively.

Case Study: Flood Risk Assessment Using GIS

Background:

A municipality in a coastal region wants to assess and mitigate flood risks in the face of
increasing sea-level rise due to climate change. They decide to use GIS modelling to identify
vulnerable areas and plan for future flood events.

Steps in the Modelling Process:


1. Data Collection: The municipality collects various spatial data, including
topographic maps, elevation data, rainfall records, land use/land cover maps, and
historical flood data. They also acquire sea-level rise projections for different time
horizons.
2. Data Preparation: The collected data is converted into a common GIS format,
georeferenced, and quality-checked. The elevation data is used to create a digital
elevation model (DEM) for the area.
3. Model Development: The GIS team develops a flood risk model using hydrological
and hydraulic modelling techniques. This model considers factors such as rainfall,
terrain, land use, sea-level rise, and river discharge.
4. Parameterization: Parameters, such as rainfall intensities and flood discharge values,
are assigned based on historical data and expert input. Sea-level rise projections are
incorporated into the model as time-variable parameters.
5. Model Calibration and Validation: The model is calibrated using historical flood
data, ensuring that it accurately reproduces past flood events. Validation is done by
comparing model predictions with independent flood events not used for calibration.
6. Scenario Testing: The municipality runs various scenarios, including different sea-
level rise scenarios, rainfall patterns, and land use changes. They assess the potential
flood extent, depth, and impact on critical infrastructure.
7. Results Analysis: The GIS team produces flood risk maps, showing areas susceptible
to flooding under different scenarios. They identify high-risk zones and prioritize
mitigation efforts. These results are shared with local authorities and residents to
inform decision-making and land use planning.

In this case study, GIS modelling is used to assess flood risks, taking into account various
spatial factors and scenarios. Such modelling enables proactive planning and helps the
municipality make informed decisions to protect its community and infrastructure from future
flood events.

[Link] the spatial data models of GIS & their merits and demerits .
Spatial data models are fundamental components of Geographic Information Systems (GIS)
that help organize and represent geographical information. There are primarily two types of
spatial data models: vector and raster. Each has its own merits and demerits, making them
suitable for different types of geospatial data and applications.
1. Vector Data Model:
• Merits:
• Precision and Accuracy: Vector data models are highly precise and
accurate for representing point, line, and polygon features. This makes
them suitable for applications where precise geographic information is
essential, such as urban planning and navigation.
• Topology: Vector data models support topological relationships,
which can help in tasks like network analysis and spatial query
optimization.
• Efficient Storage: They are efficient in terms of storage for data with
discrete features.
• Easy to Edit and Update: Vector data can be easily edited, updated,
and maintained.
• Demerits:
• Complexity: Representing continuous data (e.g., elevation) can be
challenging in vector format as it requires discretization, potentially
leading to data loss.
• Data Volume: Vector data can lead to large data volumes for complex
geometries or high-resolution data.
• Performance: Complex spatial operations on vector data can be
computationally intensive.
2. Raster Data Model:
• Merits:
• Efficiency for Continuous Data: Raster data models are well-suited
for representing continuous data, such as elevation models, satellite
imagery, and temperature maps.
• Ease of Analysis: They are efficient for grid-based spatial analysis,
including terrain analysis, spatial interpolation, and remote sensing.
• Compact Storage: Raster data can be more compact in storage when
dealing with continuous data over large areas.
• Speed: Raster data operations can be faster for certain types of
analyses, especially when using parallel processing.
• Demerits:
• Loss of Detail: Raster data may lose detail and precision when
representing discrete features, such as roads, buildings, or
administrative boundaries.
• Data Volume: High-resolution raster data can lead to large file sizes.
• Complex Topology: Topological relationships between features are
not as easily represented in raster data.
3. Hybrid Models:
• Some GIS systems use hybrid models that combine vector and raster data
models to take advantage of the strengths of both. For example, vector data
may be used for representing discrete features and raster data for continuous
attributes.
4. TIN (Triangulated Irregular Network) Model:
• TIN is another spatial data model that represents terrain as a set of non-
overlapping triangles. It is particularly useful for representing complex terrain
surfaces.

In summary, the choice between vector and raster data models in GIS depends on the specific
application and the type of data being dealt with. Vector data excels in representing discrete
features with high precision, while raster data is better suited for continuous data and grid-
based analysis. Hybrid models and TIN models may be employed to combine the advantages
of both in some situations.

[Link] the map overlay analysis.


Map overlay
 Overlay is a GIS operation that superimposes multiple data sets (representing
different themes) together for the purpose of identifying relationships between them
 An overlay creates a composite map by combining the geometry and attributes of the
input data sets.

Map overlay procedure:


 Extraction approaches those that extract one layer based on the properties of another.
 Combination approaches those that combine two or more layers together

 Clip: A cookie cutter approach where one layer is subsetted based on the extent of
another. The procedure is often used to clip data to some study area boundary.
 Erase: A procedure that is opposite to a clip procedure; keeping all features outside of
an erase area.

 Intersect: A procedure that combines layers and retains information common to the
intersected areas.
 Union: Retains information from both layers entirely

 Dissolve-Used when you want to eliminate complexity in a spatial database. The


process groups information together when they are neighbors and share the same
attribute.
 Append: Stitches together layers of the same type and coordinate system resulting
in a single mega layer
 Merge: stitches data together as a way reducing file numbers and creating a large
(mega-area) dataset based on two or more files. Offers greater flexibility than append
as the attribute tables are not required to be the same

Buffering
A process that defines/identifies areas within a specified distance of a coverage (e.g.,
within 300m of a stream),Process is applicable to points, line and polygons.
 A buffer in GIS is a zone around a map feature measured in units of distance or time.
 A buffer is an area defined by the bounding region determined by a set of points at a
specified maximum distance from all nodes along segments of an object
Ways to perform geoprocessing:
1. Arctoolbox: Simple way to integrate the tools and explore procedures.
2. Command line: more advanced way and may lead to script languages, for those computer
programmers in the course.
3. Model builder: Graphical way of stitching together commands.
4. Scripts: Customized text based sequence of commands creating a small program. Similar
to the command line or a command-line version of the Model builder
[Link] discuss the data input devices of vector & raster data
METHODSOFGEOSPATIAL DATA INPUT
Methods of Raster Data Input
 Aerial photographs and satellite imagery are the examples of raster data in
digital format.
 Scanner
Scanners are used to convert images from analog maps or photographs into
digital image data in raster format, which is then converted to vector format
through digital tracing or digitisation.
 This process of tracing is also called vectorisation.
 Scanning converts the map into binary file in raster format, with pixel having a
value of either ‘1’ representing the map feature or ‘0’ representing the
background (Chang, 2010)
 A scanner has a light source, a background for source document and a lens.
 There are three different types of scanners:
• Flat-bed scanners
• Drum scanners
• Large-formatfeed scanners

After scanning, the raster image needs to be first corrected for errors caused by scanning.
Image distortions are corrected by:
 Despeckling: It removes speckles or stray pixels that appear in an image when
we scan a dirty or wrinkled image.
 Greyscaling: It converts a colour image into greyscale
 Brightness and Contrast: It can be adjusted in a colour or greyscale image.
Increasing the contrast enhances he distinction between dark and light areas
whereas increasing the brightness lightens the image so as to enhance even the
shadow areas.
 Thresholding: It segregates the image grey values into two distinct values (that
is 0 for black and 255 for white) by a threshold value.
Methods of Vector Data Input
 Vector data is usually captured or digitised from a hardcopy print out or a digital
raster image. ′
 A number of input devices are in use for map digitisation.
 Some of the commonly used devices for vector data input are:
o Manual (via typing on keyboard or importing text files
o Digitizing
o Scanning
 Manual (via typing on keyboard or importing text files)
Manual data entry can bring into GIS either collected or measured data.
 These data exist as simple text files (have at least two columns with X and
Y coordinates ) or binary files (for example files from Global Positioning
System data collection).
 Keyboard is the simplest device to manually input data into a computer .
The process is also known as key coding.
 Mouse is the simplest and the most accurate type of digitiser device

i. Digitizing and iii. Scanning


 Digitizing is a process of entering digital codes of analyzed data into
computer.
 Digitizing can be manual (using digitizing table) or automatic (using
scanner).
 The difference between two methods is that digitizing tablet allows to do
georeferencing during the digitizing process, while scanning require
georeferencing later, after digital file (usually TIFF, GIF or JPEG image)
has been created.
 Scanning allows automatic layer separation while digitizing of the map on a
tablet requires manual creation of separate themes.
[Link] about advantages & disadvantages for raster & vector analysis
Advantages of raster data structures
 Simple data structure
 Easy to generate (e.g. from remote sensing or scandigitizing)
 Easy workflows and analysis
 Technology is cheap and is being energetically developed
 Simulation is very easy as each spatial unit has the same shape and size
 Same set of grid cells are used for several variables
Disadvantages of raster data structures
 Non-adaptive data structure
 Tends to generate huge files, depending on resolution

 Cell arrangement is usually random and does not respect natural borders
 Errors in evaluating perimeter of shape

 Topology or network linkages are difficultto establish


 Geometric transformations are difficult to handle

 Use of large cells to reduce data volumes result into loss of information
Advantages of vector data structures
 Small amount of data
 Logical data structure
 Attributes are combined with objects
 Preserves quality after interactivity (e.g. scaling)
 Topology described with network linkages
 Retrieval, updating and generalization of graphics and attributes are possible
 Widely used to described administrative zone
Disadvantages of vector data structures
 Complex data structure
 Continuous data is not represented effectively
 Spatial variability is not implicitly represented
 Spatial analysis and filtering within polygons is impossible
 Needs a lot of manual editing to get good quality
 It always introduces hard boundaries
 Simulation is difficult as each unit has different topological form
 Display and plotting can be expensive
[Link] the GIS data input using digitiser ,scanner ,automatic &onscreen
digitisation.
GIS (Geographic Information Systems) data input is a crucial step in the creation and
maintenance of spatial databases. Various methods are used to input geographic data into
GIS, including digitization, scanning, automatic digitization, and on-screen digitization.
Here's an overview of each of these methods in the context of remote sensing and GIS:
1. Digitization:
• Manual Digitization: This method involves the manual tracing of features
from paper maps, aerial photographs, or other hard-copy sources onto a digital
map using a digitizing tablet or digitizing board. An operator uses a cursor or
puck to capture the coordinates of specific points or lines.
• Merits: Provides high precision when handled by skilled operators. Suitable
for converting existing paper maps and historical data into digital format.
• Demerits: Labor-intensive, time-consuming, and may introduce errors if not
done carefully. Limited to the features visible on source documents.
2. Scanning:
• Scanning and Rasterization: In this method, hard-copy maps, aerial
photographs, or satellite images are scanned using specialized scanners. The
scanned image is then converted into a raster format, where each pixel
represents a specific location on the Earth's surface.
• Merits: Rapid data capture, especially for large-scale maps and imagery.
Suitable for handling large volumes of data. Preserves the original appearance
of source documents.
• Demerits: Results in raster data, which may have limited accuracy for precise
spatial analysis. It does not capture vector data or attribute information
directly.
3. Automatic Digitization:
• Feature Extraction Algorithms: Automated software tools use feature
extraction algorithms to identify and digitize geographic features from
scanned or raster data. These algorithms detect edges, lines, and shapes to
create vector representations.
• Merits: Faster than manual digitization, especially for large datasets. Reduces
the risk of human error.
• Demerits: May require post-processing and manual editing to correct errors.
Accuracy depends on the quality of the algorithms and the input data.
4. On-Screen Digitization:
• Interactive Digitization: In this method, GIS software is used to display a
scanned image, aerial photograph, or satellite image on a computer screen. An
operator interacts with the software to trace and create vector features directly
on the digital image using a mouse or stylus.
• Merits: Allows for real-time editing and quality control. Suitable for
digitizing features that require interpretation or are not clearly defined on
source documents.
• Demerits: Operator-dependent accuracy. May be time-consuming for
complex datasets.

In remote sensing and GIS, the choice of data input method depends on factors such as the
type of data source, the level of accuracy required, the scale of the project, and available
resources. While manual digitization and on-screen digitization are suitable for small-scale
projects and feature interpretation, scanning and automatic digitization are preferred for
large-scale projects and data extraction from imagery. A combination of these methods is
often used to achieve the best results, ensuring both accuracy and efficiency in GIS data
input.

[Link] do you understand by spatial data model ? describe conceptual &logical data
model .
Spatial data model:
 There are presently three types of representations for geographic data:
raster vector, and objects.
 raster - set of cells on a grid that represents an entity (entity -->
symbol/color --> cells).
 vector -an entity is represented by nodes and their connecting arc or line
segment (entity --> points, lines or areas --> connectivity)
 object – an entity is represented by an object which has as one of its
attributes spatial information.
Conceptual & logical data model:
Conceptual and logical data models are fundamental components of Geographic Information
Systems (GIS) and are used to represent the structure and relationships of data in the context
of remote sensing and GIS applications.
1. Conceptual Data Model:
• Purpose: The conceptual data model provides a high-level, abstract
representation of the data and its relationships within a GIS or remote sensing
application. It is primarily concerned with defining what data should be stored
and managed and how it relates to real-world entities and phenomena.
• Characteristics:
• High-Level Abstraction: It focuses on the concepts and entities
relevant to the application domain rather than specific technical details.
• No Implementation Specifics: It is implementation-agnostic and does
not concern itself with how the data will be stored or accessed.
• User-Centric: The conceptual data model is designed to be easily
understood by non-technical stakeholders and domain experts.
• Components:
• Entities: These are the main objects or things of interest in the
application domain, such as rivers, buildings, or land parcels.
• Attributes: Attributes describe properties or characteristics of entities.
For example, the attributes of a building entity might include height,
area, and construction material.
• Relationships: Relationships define how entities are related or
connected to each other. For instance, a river entity may be related to a
city entity through the "flows through" relationship.
• Example: In a land management GIS, the conceptual model might include
entities like "Parcel," "Owner," and "Land Use," with attributes such as parcel
area, owner name, and land use type. Relationships might include "owned by"
between Parcel and Owner entities.
2. Logical Data Model:
• Purpose: The logical data model takes the conceptual model a step further by
defining how the data will be structured, organized, and stored within a
database or GIS system. It serves as a bridge between the high-level concepts
of the conceptual model and the physical implementation of the database.
• Characteristics:
• Implementation-Friendly: It includes details on data tables, keys,
relationships, and data types, making it suitable for database design
and development.
• Abstracted from Technology: While it defines the data structure, it
remains independent of specific database management systems or
software.
• Balances Complexity: The logical model balances the complexity of
the real-world domain (as represented in the conceptual model) with
the efficiency and simplicity required for data storage and retrieval.
• Components:
• Tables: Tables represent data entities and contain rows and columns
for data storage. Each table corresponds to an entity from the
conceptual model.
• Keys: Keys are used to uniquely identify records in tables. Primary
keys ensure each row is unique, and foreign keys establish
relationships between tables.
• Data Types: Data types specify the format of attribute values, such as
text, numbers, dates, or spatial data types for GIS applications.
• Relationships: Logical data models define how tables are related
through primary and foreign keys, establishing relationships consistent
with the conceptual model.
• Example: In a land management GIS, the logical model might include
database tables like "Parcel," "Owner," and "LandUse," with primary keys,
foreign keys, and data types defined. Relationships between tables would be
established, specifying how parcels are linked to owners and land use types.

In summary, the conceptual data model defines what data should be captured and how it
relates to real-world concepts, while the logical data model takes these concepts and
structures them in a way that can be implemented in a database or GIS system. Together,
these models provide a foundation for the design and development of effective GIS databases
and applications in the field of remote sensing and GIS.
[Link] the concept of integrated data analysis.
Integrated data analysis in remote sensing and GIS (Geographic Information Systems) refers
to the process of combining and analyzing multiple sources of spatial and non-spatial data to
gain a deeper understanding of a specific geographic area, phenomenon, or problem. This
approach allows users to leverage diverse datasets and information types to derive
meaningful insights and make informed decisions. Here's an explanation of the concept:
1. Multiple Data Sources: Integrated data analysis involves the use of various data
sources, which can include but are not limited to:
• Remote Sensing Data: Satellite imagery, aerial photographs, LiDAR (Light
Detection and Ranging) data, and other sensor data are commonly used for
capturing information about the Earth's surface and atmosphere.
• GIS Data: Geospatial data layers, such as land use/land cover, topography,
infrastructure, and administrative boundaries, provide foundational geographic
information.
• Field Data: Ground-truth data collected through surveys, measurements, and
sampling to validate and supplement remote sensing and GIS data.
• Demographic Data: Population statistics, socioeconomic data, and other
demographic information.
• Environmental Data: Climate data, environmental measurements (e.g., air
quality, water quality), and ecological data.
2. Data Integration: The key to integrated data analysis is merging and harmonizing
these diverse datasets into a single, coherent framework. This can involve data
preprocessing, such as georeferencing, data conversion, and data transformation, to
ensure that all data are compatible in terms of spatial reference, resolution, and
format.
3. Spatial Analysis: Once integrated, spatial analysis techniques are applied to explore
relationships, patterns, and trends within the data. This can include operations such as
overlay analysis, proximity analysis, spatial interpolation, and spatial statistics.
4. Multiscale Analysis: Integrated data analysis often involves examining data at
multiple scales, from local to regional to global. This allows for a comprehensive
understanding of the studied area and its context within larger geographic and
environmental systems.
5. Decision Support: Integrated data analysis provides decision-makers with valuable
information for various applications, including urban planning, environmental
management, disaster response, agriculture, and natural resource management. It can
aid in identifying trends, hotspots, anomalies, and areas of interest.
6. Example Applications:
• Environmental Monitoring: Combining remote sensing data with ground-
based measurements and GIS layers to monitor changes in land cover, assess
vegetation health, and study environmental impacts.
• Emergency Response: Using real-time satellite imagery and GIS data to
assess the extent of natural disasters (e.g., wildfires, floods) and plan rescue
and relief efforts.
• Urban Planning: Integrating demographic data, land use data, transportation
networks, and environmental data to support urban planning and infrastructure
development.
• Precision Agriculture: Analyzing soil data, weather data, and remote sensing
imagery to optimize crop management practices.

Integrated data analysis is essential in today's data-driven world, where information is


generated from various sources and sensors. It enables organizations and researchers to
harness the power of multidimensional data for more informed decision-making, better
resource management, and a deeper understanding of complex geographic phenomena.

18. write a detailed note on data models &data compression.


Data models : Refers Qn 1
Data compression :
 Large amounts of data can create enormous problems in storage space and
transmission time.
 The main reasons of data compression could be summarized as:
• Multimedia data have large data volume
• Difficulty sending real-time uncompressed data over current network
 The design goal of image compression is to represent images with as few bits as
possible, to save storage and transmission channel capacity.
 Compression data principle: “Eliminate data redundancy and try to find a code with
less data volume.”
Data compression principles:
 Is the substitution of frequently occurring data items, or symbols, with short codes
that require fewer bits of storage than the original symbol.
 Saves space, but requires time to save and extract.
 Success varies with type of data.
 Works best on data with low spatial variability and limited possible values.
 Works poorly with high spatial variability data or continuous surfaces.
 Exploits inherent redundancy and irrelevancy by transforming a data file into a
smaller one.

 Compression ratio: The ratio of the two file sizes. E.g., original image is 100MB, after
compression, the new file is 10MB, then the compression ratio is 10:1.
 Lossless compression: Lossless data compression is compression without any loss of
data quality.
 Lossy compression: Compressing data and then decompressing it Eliminates
irrelevant information as well, and permits only an approximate reconstruction
of the original file.
[Link] briefly about attribute & spatial data .
In remote sensing and GIS (Geographic Information Systems), data can be broadly
categorized into two main types: attribute data and spatial data. These data types play distinct
but complementary roles in the analysis and representation of geographic information.
1. Attribute Data:
• Definition: Attribute data, also known as non-spatial data or tabular data,
consists of information that describes the characteristics, attributes, or
properties of geographic features. These attributes are typically stored in tables
or databases.
• Examples: Attribute data can include a wide range of information, such as
names, IDs, population counts, temperature values, land use categories,
ownership information, and more.
• Role: Attribute data provide context and additional information about spatial
features. They enable users to answer questions about what, when, who, and
how much. Attribute data are critical for attribute queries, statistical analysis,
and decision-making.
• Representation: Attribute data are typically represented in tabular form, with
rows corresponding to individual features (e.g., points, lines, polygons) and
columns representing specific attributes or properties associated with those
features.
• Analysis: Attribute data are used in various GIS operations, including
attribute queries, data classification, thematic mapping, and statistical analysis.
They are essential for creating data-driven maps and reports.
2. Spatial Data:
• Definition: Spatial data, also known as geospatial data, consists of
information that defines the location, shape, and spatial relationships of
geographic features. This data type represents the geometry or geography of
objects in the real world.
• Examples: Spatial data can include point coordinates, line segments,
polygons, boundary shapes, and any other geometric representations of
features on the Earth's surface.
• Role: Spatial data provide the spatial context necessary for visualizing and
analyzing the geographic world. They answer questions about where and how
features are distributed in space.
• Representation: Spatial data are typically represented as vector data or raster
data. Vector data represent features as discrete points, lines, and polygons,
while raster data represent features as grids of cells with pixel values.
• Analysis: Spatial data are fundamental to GIS operations such as spatial
queries, spatial analysis, cartographic mapping, overlay analysis, proximity
analysis, and network analysis. They are used for creating maps, conducting
spatial analysis, and making spatial decisions.

In summary, attribute data and spatial data are two core components of remote sensing and
GIS. Attribute data provide descriptive information about geographic features, while spatial
data define the location and geometry of those features. Combining both types of data enables
comprehensive analysis and visualization of geographic information, facilitating decision-
making, planning, and problem-solving in various fields, including environmental science,
urban planning, natural resource management, and more.
[Link] notes on following
I. Attribute data analysis.
II. Integrated data analysis.
Attribute data analysis:

Attribute data analysis in remote sensing and GIS (Geographic Information Systems)
involves examining and drawing insights from the non-spatial information associated with
geographic features. Attribute data analysis is a crucial component of geospatial analysis,
allowing users to answer questions related to the characteristics, attributes, and properties of
geographic features. Here are some key aspects of attribute data analysis in remote sensing
and GIS:
1. Data Exploration:
• Before conducting in-depth analysis, it's essential to explore the attribute data
to understand its content, structure, and quality. This includes checking for
missing values, outliers, and data distribution.
2. Data Summary and Descriptive Statistics:
• Calculate summary statistics, such as mean, median, mode, standard deviation,
and range, to gain a basic understanding of the data's central tendencies and
variability.
3. Data Visualization:
• Visualizing attribute data is essential for gaining insights. Common
visualization techniques include histograms, bar charts, scatter plots, box
plots, and thematic maps.
• Thematic mapping involves representing attribute data on a map, allowing
users to visualize spatial patterns and relationships among geographic features.
4. Spatial Queries:
• Attribute data can be used in combination with spatial data to perform spatial
queries. For example, you can query for all land parcels with an area greater
than a specified threshold or identify all buildings owned by a particular
individual.
5. Data Classification:
• Attribute data can be classified into categories or classes to simplify analysis
and interpretation. Common classification methods include equal intervals,
quantiles, natural breaks, and manual classification based on domain
knowledge.
6. Correlation Analysis:
• Investigate relationships between different attributes. For example, you can
examine the correlation between land use type and property values or analyze
how temperature relates to elevation in a region.
7. Spatial Joins and Relational Analysis:
• Combine attribute data from different datasets through spatial joins or
relational analysis. This can be used to associate attributes from one dataset
with features in another, facilitating data enrichment and analysis.
8. Geostatistical Analysis:
• Geostatistical methods, such as kriging or spatial autocorrelation analysis,
incorporate both attribute and spatial data to model and predict attribute values
across space.
9. Decision Support:
• Attribute data analysis plays a critical role in decision-making processes. For
example, in urban planning, it can inform decisions related to zoning,
infrastructure development, and resource allocation.
10. Quality Assessment:
• Evaluate the quality of attribute data, including data accuracy, completeness,
and consistency. Identify and rectify data errors and discrepancies.
11. Temporal Analysis:
• Analyze changes in attribute data over time, which is particularly important
for monitoring environmental variables, land use dynamics, and demographic
trends.
12. Statistical Modeling:
• Attribute data can be used in statistical models, such as regression analysis, to
understand the relationships between different attributes and predict outcomes.
13. Reporting and Visualization:
• Communicate the results of attribute data analysis through reports,
dashboards, and interactive maps to facilitate data-driven decision-making.
Attribute data analysis is an integral part of GIS and remote sensing applications, as it
enhances the understanding of spatial phenomena and supports a wide range of applications
in fields like environmental science, urban planning, agriculture, public health, and more.

Integrated data analysis :

Integrated data analysis in remote sensing and GIS (Geographic Information Systems) refers
to the process of combining and analyzing various types of data from multiple sources to gain
a comprehensive understanding of geographic phenomena and make informed decisions. This
approach recognizes that valuable insights can be derived by integrating data from diverse
sources, such as remote sensing imagery, geographic databases, field surveys, and other
relevant information. Here's a more detailed explanation of integrated data analysis in remote
sensing and GIS:
1. Data Integration:
• Integrated data analysis begins with the integration of spatial and non-spatial
data from various sources. This involves harmonizing data formats, spatial
references, and scales to ensure compatibility.
2. Remote Sensing Data:
• Remote sensing data sources can include satellite imagery, aerial photographs,
LiDAR data, thermal imagery, and radar data. These sources provide valuable
information about the Earth's surface and atmosphere.
3. GIS Data:
• Geographic Information Systems store and manage geospatial data layers,
such as land cover, land use, elevation, transportation networks, administrative
boundaries, and infrastructure data.
4. Field Data:
• Ground-truth data collected through field surveys, GPS measurements, and
sample observations are essential for validating and supplementing remote
sensing and GIS data.
5. Attribute Data:
• Attribute data, which describe the characteristics and properties of geographic
features, are often integrated with spatial data. These attributes may include
population counts, land values, environmental parameters, and more.
6. Temporal Data:
• Time-series data provide insights into temporal changes and trends. This can
include historical satellite imagery, climate data, and data collected over
different time intervals.
7. Analysis Techniques:
• Integrated data analysis employs various spatial analysis techniques to explore
relationships, patterns, and trends in the integrated dataset. Common
operations include overlay analysis, proximity analysis, change detection,
interpolation, and spatial statistics.
8. Multiscale Analysis:
• The analysis may occur at multiple scales, ranging from local to regional to
global, to capture variations in geographic phenomena and consider different
spatial contexts.
9. Decision Support:
• The insights gained from integrated data analysis serve as the foundation for
informed decision-making in a wide range of applications, including
environmental monitoring, urban planning, disaster management, agriculture,
and natural resource management.
10. Visualization and Reporting:
• Effective visualization techniques, such as maps, charts, and graphs, are used
to present the results of integrated data analysis. These visual representations
aid in communicating findings to stakeholders and decision-makers.
11. Modeling:
• Integrated data analysis often involves the use of spatial models to simulate
and predict future scenarios or to understand the impacts of different decisions
and policies.
12. Quality Assurance:
• Ensuring data quality and accuracy is a critical component of integrated data
analysis. Data validation and quality control procedures are essential to
maintain the integrity of the analysis results.

Integrated data analysis is a powerful approach that leverages the strengths of multiple data
sources to address complex spatial and environmental challenges. It plays a central role in the
fields of environmental science, land management, urban planning, disaster response, and
many other domains where a comprehensive understanding of geographic phenomena is
required for effective decision-making and problem-solving.

You might also like