Module 01
Module 01
1980-2000: Portable computers and pervasive devices emerged in both wired and
wireless applications.
Since 1990, HPC and HTCsystems have grown in use, hidden in clusters, grids, or
Internet clouds, serving consumers and web-scale services.
The trendis toleverage shared web resources andvast data overthe Internet.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Supercomputers are being replaced by clusters of cooperative, homogeneous compute
nodes.
HTC systems like P2P networks are used fordistributed file sharing and content delivery,
withgloballydistributed machines.
P2P, cloud computing, and web services focus more on HTC than HPC, leading to the
development of computational and data grids.
[Link] High-PerformanceComputing
HPC systems have focused on raw speed, increasing from Gflops in the 1990s to Pflops
by2010.
This growth was driven by demands from scientific, engineering, and manufacturing
sectors.
The Top 500 most powerful supercomputers are ranked by floating-point speed in the
Linpackbenchmark.
However, supercomputers are used byless than10% of all computerusers.
Most users now rely on desktop computers or large servers for Internet searches and
market-driven tasks.
[Link] High-ThroughputComputing
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
High-end computing is shifting from the HPC to HTC paradigm, focusing on high-flux
computing for tasks like Internet searches and web services.
The performance goal now measures throughput—the number of tasks completed per
unit oftime.
HTC needs improvements in batch processing speed, cost, energy efficiency, security,
and reliability at datacenters.
[Link] ThreeNewComputingParadigms
In 1969, Leonard Kleinrock predicted that computer networks would evolve into utilities
servicinghomes and offices.
Overtime, the definitionof "computer" has evolved:
Cloud computing overlaps with all three: distributed, centralized, and parallel
computing.
Centralized Computing:
All resources (processors, memory, storage) are centralizedinone physical system.
Parallel Computing
Processors are either tightly coupled with centralized shared memory or loosely coupled
withdistributedmemory.
Interprocessorcommunicationoccurs through sharedmemory ormessage passing.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Programs running on parallel computers are called parallel programs, and the process is
knownas parallel programming.
Distributed Computing
Involves multiple autonomous computers, each with its own private memory,
communicatingvia anetwork.
Informationexchange occurs throughmessage passing.
Programs running in a distributed system are called distributed programs, and writing
themis knownas distributed programming.
CloudComputing
AnInternet cloud canbe centralized ordistributed.
[Link] DistributedSystemFamilies
Since the mid-1990s, P2P networks and clusters have formed computational grids and
data gridsforwide-area computing.
Interest in Internet cloud resources for data-intensive applications has increased, with
clouds shifting desktop computing to service-oriented computing using server clusters
and databases.
Grids and clouds focusonresource sharing in hardware, software, and data.
These systems aim toexploit parallelism and concurrencyacross manymachines.
P2P networks can involve millions of client machines, and cloud computing clusters may
have thousands of nodes.
Future HPC and HTC systems will require multicore or many-core processors to handle
large numbers of computing threads, emphasizing parallelism, scalability, efficiency, and
reliability.
Efficiency depends on speed, programming, and energy factors (throughput per watt),
withdesign goals centered onthroughput, scalability, andenergyefficiency:
o Efficiency: Resource utilizationinHPC and job throughput, data access, and
power efficiencyinHTC.
o Dependability: ReliabilityandQoS under failure conditions.
o Adaptation: Supportingbillions of job requests across large datasetsandcloud
resources.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
o Flexibility: Running efficiently in both HPC (science) and HTC (business)
applications.
1.1.2 ScalableComputing Trendsand NewParadigms
Predictable technology trends drive computing applications, with designers aiming to
forecast future system capabilities.
Moore’ s law (processor speed doubles every 18 months) and Gilder’ s law (network
bandwidthdoubles annually) have shaped technology, thoughfuture validityis uncertain.
The price/performance ratio of commodity hardware, driven by desktop, notebook, and
tablet markets, hasboosted large-scale computing adoption.
Distributed systems focus on resource distribution, concurrency, and high degrees of
parallelism(DoP).
[Link] DegreesofParallelism
Bit-Level Parallelism (BLP): Transition from bit-serial to word-level processing over
time (e.g., 4-bit to 64-bit CPUs).
Instruction-LevelParallelism (ILP): Executes multiple instructions simultaneouslyusing
pipelining, superscalar computing, and multithreading. ILP requires branch prediction,
dynamic scheduling, and compilersupport.
Data-Level Parallelism (DLP): Uses SIMD and vector machines for parallel data
processing, requiring more hardware andcompiler support.
Task-Level Parallelism (TLP): Explores parallelismwithmulticore processors, but faces
challenges inprogramming and efficient execution.
Job-Level Parallelism (JLP): As we move to distributed processing, parallelism
increases to job-level granularity, built on fine-grainparallelism.
[Link] InnovativeApplications
HPC and HTC systems aim for transparency in data access, resource allocation, process
location, concurrency, job replication, and failure recovery.
Key applications drive the development of parallel and distributed systems across
domains like science, engineering, business, healthcare, education, military, and
government.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
All applications require computing economics, web-scale data, reliability, and scalable
performance.
Example: Distributed transaction processing in banking demands consistency in
replicated transaction records, despite challenges like lack of software support, network
saturation, and security threats.
1.1.2.3TheTrendtowardUtilityComputing
1. Computing Paradigms
Ubiquitous, reliable, scalable, autonomic, and composable withQoS/SLAs.
Alignwiththe computerutilityvision.
2. Utility Computing
Business model where customers pay forcomputing resources.
4. Technological Challenges
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Advancesneededinprocessors, storage, OS, virtualization, programmingmodels,
and resource management.
1.1.2.4TheHypeCycleofNewTechnologies
HypeCyclePhases:
1. Innovation Trigger: Initial breakthroughorinnovation.
2. Peak of Inflated Expectations: Hype and exaggerated expectations.
3. Trough of Disillusionment: Disappointment as technology fails to meet
expectations.
4. Slope of Enlightenment: Understandingandacceptance grow.
5. Plateau of Productivity: Mainstream adoption and maturity.
Symbols forAdoptionTimeline:
o Hollow Circles: Mainstream adoptionin<2 years.
o Gray Circles: Mainstream adoption in2-5 years.
o Solid Circles: Mainstreamadoptionin5-10 years.
o Triangles: Adoption>10 years.
o Crossed Circles: Technologies becomingobsolete before plateau.
2010 Example:
o Consumer-generated media: In disillusionment, expected to reach adoption in<2
years.
o Internet micropayments: Moving fromenlightenment tomaturity in 2-5 years.
o 3Dprinting: 5-10 years from rising expectations to mainstream adoptionJG.
o Meshnetworksensors:>10 years toplateau.
Cloud Technology: Expected to move from peak expectations to productivity in 2-5
years.
Obsolete Technologies:
o Broadband over power line: Expected to become obsolete before leaving the
valley.
Promising Technologies (2010):
o Cloud, biometric authentication, interactive TV, speech recognition, predictive
analytics, and media tablets.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
IP Addresses: With IPv6, 2128 IP addresses are available, enabling identification of all
objects on Earth, including computers and devices.
Scale of IoT:
o Estimatedthat each personwill be surrounded by1,000 to 5,000 objects.
o The IoT needs totrack 100 trillionobjects (static ormoving)simultaneously.
Addressability: Universal addressability is required for all IoT objects to ensure they are
identifiable, searchable, and manageable.
Geographical Development: The IoT is more heavilydeveloped inAsia and Europe.
CommunicationPatterns:
1. H2H(Human-to-Human)
2. H2T(Human-to-Thing)
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
3. T2T(Thing-to-Thing)
Dynamic Connectivity: Connections are made at any time (day, night, indoors, outdoors,
onthe move)and anyplace (PC, mobile phones, etc.)withlow cost.
IoTGrowth: The IoT is still initsearlystages, with many prototypes in limited areas.
Cloud Computing: Cloud technologies are expected to support fast, efficient, and
intelligent interactions amonghumans, machines, and objects.
Smart Earth Vision: The goal is to create intelligent cities, efficient resources (water,
power, transport), green IT, better schools, healthcare, and more, though achieving this
globallywill take time.
1.1.3.2Cyber-Physical Systems(CPS)-
Definition: A CPS integrates computational processes withthe physical world, combining
cyber (computational)and physical (real-world)components.
Technologies: Merges computation, communication, and control into a feedback system
betweenthe physical and informationworlds.
Focus: IoTconnects physical objects, while CPS explores virtual realityapplicationsinthe
physical world.
Impact: CPS may transform interactions with the physical world, similar to how the
Internet transformed virtual interactions.
1.2TechnologiesforNetwork-BasedSystems-
Focus on hardware, software, and network technologies for distributed computing
systems.
Emphasis on designing distributed operating systems to manage massive parallelism in
distributed environments.
1.2.1Multicore CPUs andMultithreadingTechnologies
1.2.1.1Advances inCPU Processors – infig:1.4
Multicore Architecture: Modern CPUs feature multiple cores (dual, quad, etc.) to
exploit ILP andTLP.
Performance Growth: CPU speed has grown from 1 MIPS in 1978 (VAX 780) to 22,000
MIPS in2008(Sun Niagara 2), followingMoore's law.
Clock Rate Limitation: The clock rate, once increasingto 4 GHz, nowfaceslimitsdue to
power and heat constraints.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
ILP Exploitation: Techniques like superscalar architecture, branch prediction, and
speculative executionare usedto enhance ILP.
GPU Advances: GPUs utilize many-core architecture to exploit DLP and TLP, with
hundreds tothousands of simple cores.
1.2.1.2MulticoreCPUandMany-CoreGPUArchitectures-
Multi-Core andMany-CoreProcessors –
Architecture: Multi-core CPUs have multiple cores, eachwithits ownL1cache, and
share anL2 cache.
Future chipsmayalso include anL3 cache.
High-End Processors: Examples include Intel i7, Xeon, AMDOpteron, and Sun Niagara.
Multithreading: Cores canbe multithreaded, suchas NiagaraII with8coresand8
threads per core (64 threads total).
Performance: The Intel Core i7990x (2011)reported 159,000 MIPS executionrate.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Future of Multicore CPUs: CPUs may scale to hundreds of cores, but are limited in
exploiting massive DLP due to memorywall issues.
Many-Core GPUs: GPUs with hundreds of thin cores are developed to handle massive
parallelism.
Processor Trends: x-86 processors are replacing RISC processors in high-performance
systems, including data centers andsupercomputers.
Heterogeneous Chips: Future processors may combine fat CPU cores and thin GPU
cores onthe same chip.
1.2.1.3MultithreadingTechnology-
ProcessorTypes:
1. SuperscalarProcessor: Single-threaded withmultiple functional units.
2. Fine-Grain MultithreadedProcessor: Switches threads everycycle.
3. Coarse-Grain Multithreaded Processor: Executes multiple instructions from the
same thread forseveral cyclesbefore switching.
4. Dual-Core CMP: Twocores, eachsingle-threaded andtwo-way superscalar.
5. SMT Processor: Simultaneously schedules instructions from multiple threads in
one cycle.
Execution Patterns: These processors handle instructions from different threads in
distinct ways, withsome more efficient thanothers inutilizingfunctional units.
1.2.2GPUComputingtoExascaleandBeyond-
GPU Definition: A graphics coprocessor or accelerator used to offload graphics tasks
fromthe CPU, first introduced byNVIDIAin1999 (GeForce 256).
GPU vs CPU: GPUs have hundreds of cores, while CPUs like Xeon X5670 have only six
cores.
Parallelism: GPUs utilize a throughput architecture to execute many concurrent threads
slowly, unlike CPUs that focus onfast executionofa single thread.
GPGPU: General-purpose computing on GPUs (GPGPUs) has gained traction in HPC,
withNVIDIA’ sCUDA model supporting thisshift.
1.2.2.1HowGPUsWork-
GPU Evolution: Early GPUs were CPU coprocessors; modern NVIDIA GPUs have 128
cores, eachhandling8threads, allowing 1,024 concurrent threads.
Parallelism: GPUs excel at massive parallelism, executing many threads compared to
CPUs withfewerthreads.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Optimization: CPUs focus on latency, while GPUs prioritize high throughput and manage
on-chip memory.
Applications: GPUs power HPC systems and supercomputers for parallel processing
beyond graphics, handlingdata-intensive calculations.
Usage: GPUs are common in mobile phones, game consoles, PCs, and servers, with
NVIDIA CUDA Tesla/Fermi used inHPC systems.
1.2.2.2GPUProgrammingModel -
CPU-GPU Interaction: The CPU offloads floating-point computations to the GPU, which
has manysimple processing cores (multiprocessors)for massive parallelism.
CUDA Programming: NVIDIA’ s CUDA is used for large-scale data processing in GPUs
like the GeForce 8800 or Tesla Fermi.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Fermi GPU Architecture: The Fermi GPU (2011) has 16 streaming multiprocessors (SMs),
eachwith512CUDA cores, totaling 3 billiontransistors.
GPU Capabilities: Fermi GPUs can handle up to 82.4 Tflops, and future GPUs may reach
Exascale systems (1Eflop).
Challenges for Exascale Computing: Key challenges include energy, memory,
concurrency, and system resiliency.
1.2.2.3PowerEfficiency of theGPU
Bill Dally's View: Power efficiency and parallelism are keyadvantagesof GPUs overCPUs.
Exaflops System Requirement: 60 Gflops/watt per core is needed to run an exaflops
system.
Power Consumption:
CPU: 2 nJ/instruction.
GPU: 200 pJ/instruction (10 timesmore efficient thanCPU).
Optimization Differences:
CPU: Optimizedfor latencyin caches and memory.
GPU: Optimized forthroughput withexplicit memorymanagement.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
2010 Performance/PowerRatio:
GPU: 5 Gflops/watt per core.
CPU: Less than1Gflop/watt percore.
Challenges forFuture Supercomputers:
Data movement dominates powerconsumption.
Need foroptimized storage hierarchiesand memory tailored to applications.
KeyFocus Areas:
Development of self-aware OS and runtimes.
Creationof locality-aware compilers and auto-tunersfor GPU-based systems.
Conclusion: Both power efficiencyand software development are crucial for the future of
parallel and distributedcomputing systems.
1.2.3Memory,Storage,andWide-AreaNetworking
1.2.3.1Memory Technology(Fig 1.10)
DRAM Growth: DRAM capacity grew from 16KB in 1976 to 64GB in 2011, increasing 4x
everythree years.
Memory Access: Memory access times have improved little, and the memory wall issue
worsensas processors get faster.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Hard Drive Growth: Hard drive capacitygrewfrom260MB in 1981to 250GB in2004, and
Seagate's Barracuda XT reached 3TB in 2011(10x increase everyeight years).
Future Trends: Disk array capacity will continue to grow. The gap between processor
speed and memory capacity may worsen, further limiting CPU performance due to the
memory wall.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Data-intensive science, cloud computing, and multicore computing are converging,
transforming the next generation of computing in architecture, design, and programming
challenges. These advancements turndata intoknowledge, drivingmachine wisdominSOA.
1.3SystemModelsforDistributedandCloudComputing (Table 1.2)
Distributed and cloudcomputing systems consist of many autonomous nodes connectedby
SANs, LANs, or WANs. LAN switchescan link hundreds of machines, while WANs connect
multiple local clusters toformlarge-scale systems withmillions of computers.
These systems are highlyscalable and canachieve web-scale connectivity. Theyare classified
into fourtypes: clusters, P2P networks, computing grids, and cloud datacenters. These
systems mayinvolve hundreds tomillions of nodes working cooperativelyat various levels,
each serving different technical and applicationneeds.
Clusters dominate supercomputing, with 417of the Top 500 in 2009 using thisarchitecture,
formingthe base forgrids andclouds. P2P networks are preferred forbusiness but face
copyright issues in content industries. National gridswere underused due to unreliable
middleware. Cloud computing offers low cost and simplicityfor bothusers and providers.
1.3.1ClustersofCooperativeComputers
A computing clusteris a group of interconnected stand-alone computersworking together as a
single resource, achievingimpressive resultsinhandling heavy workloads and large data sets.
1.3.1.1ClusterArchitecture (Fig: 1.15)
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
A typical serverclusteruses alow-latency, high-bandwidthinterconnectionnetworklike SAN
(Myrinet) orLAN (Ethernet). For largerclusters, networks canbe built with multiple levels of
Gigabit Ethernet, Myrinet, or InfiniBand switches. Clustersare connected tothe Internet via a
VPN gateway, whichlocates the cluster. Clusters typically have looselycoupled nodes, each
withits ownOS, leadingto multiple system imagesacross autonomous nodes.
1.3.1.2Single-SystemImage
Greg Pfistersuggests an ideal clustershould merge multiple systemimages intoa
single-systemimage (SSI).
SSI is supportedbya cluster OS ormiddleware toshare CPUs, memory, andI/O across
all nodes.
SSI creates the illusionofa unified, powerful resource, making the cluster appearasa
single machine to the user.
Without SSI, a clusteris just a collection of independent computers.
1.3.1.3Hardware,Software,andMiddlewareSupport
Cluster designprinciples forsmall andlarge clusters will be discussedinChapter2.
Massive parallel processing (MPP)clusters dominate the Top 500 HPC list.
Building blocks: computer nodes (PCs, workstations, servers, SMP), communication
software (PVM, MPI), and network interface cards.
Most clusters use Linux OS and are interconnected via high-bandwidthnetworks
(Gigabit Ethernet, Myrinet, InfiniBand).
Special middleware supports featureslike SSI and highavailability (HA).
Both sequential and parallel applicationscan run; special parallel environments are
neededforcluster resource use.
Distributed shared memory(DSM)maybe used toshare memoryacross servers.
Achieving full SSI can be expensive ordifficult; manyclusters remainloosely coupled.
Virtual clusters canbe dynamicallycreated via virtualization.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
1.3.1.4MajorClusterDesignIssues (Table 1.3)
Cluster-wide OS for full resource sharing is not yet available.
Middleware or OS extensionsare used to achieve SSI at selected levels.
Without middleware, clusternodes cannot collaborate effectively.
Middleware is essential forhigh performance insoftware environments and applications.
Keybenefitsof clusters: scalable performance, efficient message passing, high
availability, fault tolerance, andjobmanagement.
1.3.2GridComputingInfrastructures
Overthe past 30 years, usershave experienced growth fromInternet toweb andgrid
computingservices.
Internet services like Telnet allowlocal-to-remote computerconnections.
Web services like HTTP provide remote access to web pages.
Gridcomputing enablesclose interaction betweenapplications ondistant computers.
Forbes projectsIT-based economy growth from$1 trillionin2001 to $20 trillion by2015,
drivenbythese services.
1.3.2.1Computational Grids (Fig 1.16)
A computing grid combines computers, software, middleware, instruments, people, and
sensors, similar to an electric utilitygrid.
Grids are built overLAN, WAN, or Internet backbones at regional, national, or global
scales.
Theyserve as integratedcomputing resources and virtual platforms for virtual
organizations.
Grids use workstations, servers, clusters, and supercomputers, withPCs, laptops, and
PDAs asaccess devices.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
1.3.3Peer-to-PeerNetworkFamilies
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
A client-serverarchitecture is a common distributed system where client machines
connect to a central serverfor tasks like computing, email, file access, and databases.
The P2P architecture offers a client-oriented distributed model, as opposedto
server-based systems.
Thissectioncovers P2P systemsat the physical and overlaynetwork(logical)levels.
1.3.3.1P2PSystems (Fig 1.17)
InP2P systems, eachnode acts as both aclient anda server, contributing resources.
Peer machines autonomouslyjoin or leave the system, withno central coordinationor
database.
The systemis self-organizing withdistributed control, meaning no peer has a global view
of the entire network.
The P2P network is ad hoc, formed using the TCP/IP and NAI protocols, and its size and
topologychange dynamically.
1.3.3.2OverlayNetworks (Fig1.17)
Data items/filesare distributedacrosspeers in the network, forminganoverlay network
at the logical level.
New peers joinbyadding their peerIDto the overlay, while leaving peershave theirID
removedautomatically.
There are two typesof overlaynetworks:
o Unstructured: Randomgraph, where messages are sent via flooding, leading to
hightraffic andunpredictable searchresults.
o Structured: Follows a set topology, with rules for adding/removing nodes and
optimized routing mechanisms.
1.3.3.3P2PApplicationFamilies (Table 1.5)
P2P networks are classified into fourapplication-basedgroups:
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
1. Distributed File Sharing: Includes networks like Gnutella, Napster, andBitTorrent for
sharing digital content.
2. CollaborationNetworks: Forinstant messaging and collaborative tools like MSN, Skype.
3. Distributed P2P Computing: Used inapplications like SETI@home for collective
computing power.
4. Platform-Specific Networks: Platforms like [Link] fortasks suchas naming,
communication, andresource aggregation.
1.3.4CloudComputingovertheInternet
Computational science is becoming data-intensive, requiring balancedsystems with
petascale I/O and networking.
Future computingwill involve sending programs tolarge data sets instead of moving
data to workstations.
Thistrend leadsto the rise of cloud computing, where software, hardware, anddata are
provided on-demand.
IBMdefines cloudcomputing asa pool of virtualizedresources that canhost various
workloads.
Cloud computing supports rapid scaling, self-recoveryfromfailures, and real-time
monitoring forresource management.
“A cloud is apool ofvirtualized computer resources. Acloudcan host a variety of
differentworkloads, including batch-style backend jobs andinteractive and
user-facing applications.”
1.3.4.1Internet Clouds (Fig 1.18)
Cloud computing uses virtualizedplatforms withelastic resources, provisioning
hardware, software, and data dynamically.
It shifts desktop computing to service-oriented platforms using serverclusters and large
databases.
Cloud computing offerslowcost and simplicity forusersand providers, enabledby
machine virtualization.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Portals like OGFCE and HUBzero provide access to users, integratingwebservice and
Web 2.0 technologies.
1.4.1.4GridsversusClouds
The boundarybetween grids andcloudsis blurring.
Workflowtechnologies inweb services coordinate services forbusiness processeslike
two-phase transactions.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Workflowapproaches include Pegasus, Taverna, Kepler, Trident, and Swift.
Grids use static resources, while clouds emphasize elastic resources.
Some researchers see the maindifference betweengridsand clouds as dynamic
resource allocationvia virtualizationandautonomic computing.
A grid canbe built from multiple clouds, supportingnegotiatedresource allocationbetter
thanpure clouds.
Thisleads toa systemof systems like cloud of clouds, grid of clouds, orinter-clouds in
SOA architecture.
1.4.2TrendstowardDistributedOperatingSystems
Distributed systems have looselycoupled computers, eachrunningindependent
operating systems.
To enable resource sharing and fast communication, adistributedOS is ideal for
managingresources efficiently.
Thissystemis typicallyclosed and uses message passing and RPCsforcommunication
betweennodes.
A distributed OS is crucial for improving performance, efficiency, and flexibilityin
distributed applications.
1.4.2.1DistributedOperatingSystems
Tanenbaumidentifiesthree approaches for distributing resource management ina
system:
1. NetworkOS: Built over heterogeneous OS platforms, offeringlowtransparency,
mainlyrelyingonfile sharing for communication.
2. Middleware: Provides limited resource sharing, like MOSIX/OS for clustered
systems.
3. Distributed OS: Achieveshighersystemtransparency forbetterresource sharing.
Table 1.6compares the functionalitiesof these approaches.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
1.4.2.2AmoebaversusDCE
DCE: A middleware-based systemfor distributed computing, developed bythe Open
Software Foundation(OSF).
Amoeba: Developed at Free University, the Amoeba, DCE, and MOSIX2are research
prototypesprimarily used in academia, withno commercial success.
Need forWeb-based OS: New operating systems are needed to support resource
virtualization indistributed environments, still an area of active research.
Distributed OS Design: Should use a lightweight microkernel approach (like Amoeba)or
extend existing OS (like DCE withUNIX)tobalance resource management across
servers, freeingusers from most management tasks.
1.4.2.3MOSIX2forLinuxClusters
MOSIX2 is a distributed OS witha virtualizationlayer inthe Linux environment.
Providesa partial single-systemimage touserapplications.
Supports bothsequential and parallel applications.
Discovers resources andmigrates software processes amongLinux nodes.
Canmanage Linux clusters orgrids of multiple clusters.
Allows resource sharingamong multiple cluster owners.
Exploredfor managing resourcesinLinux clusters, GPU clusters, grids, andcloudsusing
VMs.
1.4.2.4Transparencyin ProgrammingEnvironments(Fig1.22)
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
Transparent computing infrastructure separates user data, applications, OS, and
hardware into fourlevels.
User data is independent of applications.
OS provides clearinterfaces and systemcalls forapplicationprogrammers.
Infuture cloud infrastructure, hardware will be separatedfromthe OS by standard
interfaces, allowingusers to choose preferredOSes.
Cloud applications (SaaS)allowusersto switchservices without binding datato specific
applications.
1.4.3Parallel andDistributedProgrammingModels
Thissectioncovers fourdistributedcomputing programming models with scalable
performance and flexibility.
Table 1.7summarizesthree models and associatedsoftware tools.
MPI is the most popularmodel for message-passingsystems.
Google’ s MapReduce and BigTable optimize resource use inclouds and datacenters.
Service cloudsextend Hadoop, EC2, and S3for distributed computing overstorage
systems.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
1.4.3.1Message-PassingInterface(MPI)
Message-Passing Interface (MPI) is a primary standard for developing parallel and
concurrent programs on distributed systems.
MPI is a libraryfor C orFORTRAN usedto write parallel programs.
It supports clusters, grid systems, and P2P with upgraded web services and utility
computing applications.
Parallel Virtual Machine (PVM) is another low-level primitive for distributed
programming, alongside MPI.
1.4.3.2MapReduce
MapReduce is a web programming model forscalable data processing onlarge clusters.
It is mainlyusedinweb-scale search and cloud computing applications.
Users specify a Map function to generate key/value pairs and a Reduce function to
merge intermediate values withthe same key.
MapReduce scales highly, handling terabytes of data across tens of thousands of
machines.
Google executesthousands of MapReduce jobsdailyacross itsclusters.
1.4.3.3HadoopLibrary
Hadoop is a software platform developed by Yahoo! for running applications on
distributed data.
It scales easilytostore and processpetabytes of data.
Hadoopis cost-effective, with anopen-source MapReduce versionminimizingoverhead.
It processes data with high parallelism across many nodes and ensures reliability by
keeping multiple data copies forfailover.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
1.4.3.4Open GridServicesArchitecture(OGSA)
Gridinfrastructure development isdriven bylarge-scale distributed computing
applications requiring highresource anddata sharing.
OGSA is astandard forpublic grid services, withGenesis II being akeyimplementation.
Features include adistributedexecutionenvironment, PKI services, trust management,
and securitypolicies.
1.4.3.5GlobusToolkitsandExtensions
Globus is amiddleware librarydevelopedbyU.S. Argonne National Laboratoryand USC
InformationScience Institute.
Implements OGSA standards forresource discovery, allocation, and securityingrid
environments.
Supports PKI certificates formultisite mutual authentication.
GT4, the current version, hasbeeninuse since 2008, withIBMextendingit for business
applications.
1.5.1PerformanceMetricsandScalabilityAnalysis
[Link] Performance Metrics
Performance Factors:
o MIPS (Million Instructions PerSecond)
o Tflops (TeraFloating-Point Operations per Second)
o TPS (Transactions PerSecond)
o Job response time
o Networklatency
Preferred Network:
o Low-latency, high-bandwidthinterconnectionnetwork
System Overhead:
o OS boot time
o Compile time
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
o I/O data rate
o Runtime support system
OtherPerformance Metrics:
o QoS forInternet and web services
o System availabilityanddependability
o Security resilience against network attacks
[Link] Dimensions ofScalability
ScalabilityinDistributed Systems:
o Size Scalability: Achieving higher performance by increasingmachine size
(processors, memory, I/O). Not all systems are equally scalable (e.g., IBM
BlueGene/L scaledto 65,000 processors).
o Software Scalability: Upgrading OS, compilers, libraries, and environments. Some
software may not work with larger systems and require testing and fine-tuning.
o ApplicationScalability: Matchingproblem size withmachine size. Increasing
problemsize (data or workload)can improve systemefficiency without
expanding machine size.
o Technology Scalability: Adapting to changes in technology(hardware and
software). Includes:
Time: Generation scalability, consideringthe impactof upgrading
components.
Space: Packaging, energyconcerns, and portability.
Heterogeneity: Use of components fromdifferent vendors, which maylimit
scalability.
[Link] Scalability versus OSImage Count
Scalable Performance: Refers toachieving higherspeeds byadding more processors,
memory, disk capacity, or I/O channels.
1.5.1.6Gustafson’ sLaw
To achieve higherefficiency whenusing a large cluster, we must considerscaling the problem
size to matchthe cluster capability. This leads to the following speedup lawproposed byJohn
Gustafson, referred as scaled-workload speedup
Let W be the workloadina givenprogram. Whenusing ann-processorsystem, the user scales
the workload toW′ = αW+(1− α)nW. Note that onlythe parallelizable portionof the workload
is scaled ntimes in the second term. This scaledworkloadW′ is essentially the sequential
executiontime ona single processor.
The parallel executiontime of ascaledworkload W′ onn processors is defined bya
scaled-workload speedup as follows:
---------------------------------- (1.3)
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
1.5.4.3Application Layer
Traditional apps: Focus onincreasing speedorquality.
Energy-aware apps: Challenge isto manage energywithout compromisingperformance.
Keyfactors: Energyconsumptiondependsonthe number of instructionsand memory
transactions.
Correlation: Compute and storage impact completion time and energy use.
[Link] Middleware Layer
Middleware layer: Bridges the applicationand resource layers.
Functions: Provides resource broker, communication, task analysis, scheduling, security,
reliability, andinformationservices.
Energy-efficiency: Focusesonenergy-efficient taskscheduling.
New approach: Combines makespan (executiontime) and energyconsumptionin
scheduling.
[Link] Resource Layer
Resource layer: Manages computingnodes, storage, and hardware resources in
distributed systems.
Powermanagement: Includes Dynamic Power Management (DPM)and Dynamic
Voltage-FrequencyScaling (DVFS)techniques.
DPM: Allows hardware, like CPUs, to switch to lowerpowermodes.
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
DVFS: Controls powerconsumptionbyadjusting frequencyandvoltage, optimizing
executiontime andenergyuse.
1.5.4.6Network Layer
Networklayer: Responsible forrouting packets and enabling network services to the
resource layerindistributed systems.
Challenges:
1. Models shouldfullyunderstand the interactions of time, space, and energy.
2. Need fornew energy-efficient routingalgorithms andprotocols tocombat
network attacks.
Data centers: Crucial forinformationstorage and processing but suffer fromhighcosts,
complexmanagement, lowsecurity, andexcessive energyconsumption. Next-gen
technologies are needed for improvement.
[Link] DVFS Methodfor Energy Efficiency
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25
CloudComputingandSecurity-BIS613D
BGSCET/AI&ML/2022-SCHEME/MODULE-01/2024-25