Telecommunication Standardization Sector: Question(s) : Input Document Source: Title
Telecommunication Standardization Sector: Question(s) : Input Document Source: Title
References
[ITU-T AI Challenge] ITU AI/ML in 5G Challenge website
[Link]
[ITU AI/ML Primer] ITU AI/ML 5G Challenge: Participation Guidelines (17th
April, 2020)
[ITU AI/ML Summary] ITU AI/ML 5G Challenge: Summary Slides (23rd April, 2020)
1. Introduction
[ITU AI/ML Participation Guidelines] described the proposal for ITU Global Challenge on
AI/ML in 5G networks.
Problem statements which are relevant to ITU and IMT-2020 networks are the backbone of
the challenge. They should be aligned with the theme/tracks of the challenge and should
provide enough intellectual challenge while being practical within the time period of the
challenge. They should address short term pain points for industry while pointing to long
term research directions for academia. In addition, many of them may need quality data to
solve them. This contribution collates the problem statements from our partners in a standard
format. Future steps for these problem statements are:
analyse the submitted problem statements from our partners and colleagues,
present them for selection by the challenge management team
host the selected problem statements on the challenge website.
While discussing and disseminating the challenge with our partners, an important and
frequent question posed to us is about the relevant resources. This document contains a
collection of resources pointed to us by our members and partners in the context of ITU
ML5G global challenge. This is an attempt to compile and classify them so that it is useful to
all our partners. We invite our members and partners to add pointers to private as well as
public resources which may be of relevance to the Challenge.
2. Problem statements
NOTE 1- the structure of the list below is derived from the many discussions that we had
with partners across the globe.
NOTE 2- this list is in no specific order.
NOTE- some problem statements are “restricted problem statements”. These are available
in this document with red Title but the registration to the regional host’s website to such
problem statements and data are subject to conditions set forth by the Regional host. E.g.
currently the problem statements offered by AIIA-ITU challenge are restricted problem
statements and are available only to Chinese citizens with authorized Chinese identification.
NOTE- some problem statements use “restricted data” which is available only under a
certain conditions set forth by the Regional host.
Id ITU-ML5G-PS-TEMPLATE
Title Do not modify this particular table, this serves as a template, use
the one below.
Description NOTE 3- include a brief overview followed by a description about
the problem, its importance to IMT-2020 networks and ITU,
highlight any specific research or industry problem under
consideration.
Challenge Track NOTE 4- include a brief note on why it belongs in this track
Evaluation criteria NOTE 5- this should include the expected submission format e.g.
video, comma separated value (CSV) file, etc.
NOTE 6- this should include any currently available benchmarks.
e.g. accuracy.
Data source NOTE 7- e.g. description of private data which may be available
only under certain conditions to certain participants, pointers to
open data, pointers to simulated data.
Resources NOTE 7- e.g. simulators, APIs, lab setups, tools, algorithms, add a
link in clause 2.
Any controls or NOTE 8- e.g. this problem statement is open only to students or
restrictions academia, data is under export control, employees of XYZ
corporation cannot participate in this problem statement, any other
rules applicable for this problem, specific IPR conditions, etc.
Specification/Paper NOTE 9- e.g. arxiv link, ITU-T link to specifications, etc.
reference
Contact NOTE 10- email id or social media contact of the person who can
answer questions about this problem statement.
Id ITU-ML5G-PS-001
Title 5G+AI+AR (Zhejiang Division)
Description Background:
Augmented Reality, which enriches the real world experience through
digital means. Its realization depends on a variety of technical means
such as multimedia, three-dimensional modeling, real-time tracking and
registration, intelligent interaction, and sensing. It simulates computer-
generated virtual information such as text, images, three-dimensional
models, music, and videos, and then applies it to the real world. The
two kinds of information complement each other to achieve "augment"
of the real world.
The final breakthrough of AI technology comes from the rapid
development of big data and computing power. The combination of AI
and AR is based on data and hardware to improve perception
recognition, knowledge calculation, sameness and interaction fidelity,
so that virtual objects and real environment can have natural,
continuous and in-depth interaction with users. The deep integration of
AI and AR will enable the virtual world to be seamlessly connected to
the real world, and ultimately enable digital applications in various
industries.
Problems:
Focusing on the intelligence application demand of industry, the
artificial intelligence technology and augmented reality technology are
applied to the digital upgrade of the industrial Internet. It can be
expanded around the following two topics:
Direction 1: AI+AR entertainment application
"AI + AR Entertainment" combines 5G, AI, and AR technologies with
the consumption, entertainment, and business fields. It empowers the
entertainment market through technological means, changes existing
communication methods, strengthens the participation and interaction of
audiences, and brings people an immersive sensory experience.
"AI + AR Entertainment" includes rich industrial scenes such as city
landmarks, business district interaction, games, and digital venues.
Participants can choose any scene to play their creativity and
imagination and combine science and technology to achieve the purpose
of improving the audience experience, innovating the communication
and marketing methods, and enhancing the cultural and entertainment
content. This helps to ensure that the solution is innovative and
accessible and uses technology to help the development of the
entertainment industry.
[Link]+AR city landmark interaction:
The tourism supply side reform is shifting from relying heavily on large
resources, large capital and large commercial district to focusing on
differentiation, innovation, experience and operation. As the showcase
project of the city, the city landmark is not only the name card of the
city, but also the display window of city multiculturalism. In the city
landmark scene, the technologies of combination of virtual and reality
are introduced to provide rich and diverse interactive experience for
different groups of people, strengthen the digital operation value of
urban landmarks, rebuild the relationship between people and city, and
make the city identity more full and dynamic.
[Link]+AR commercial district interaction:
With the deepening of urbanization, the single shopping mall with a
large serving range has gradually disappeared. More and more shopping
zones and the impact of e-commerce makes it a new challenge for the
business complex to attract more young customers with strong
consumption ability and high consumption desire. In the era of 5G,
digital empowerment enables the effective connection between online
and offline. "Smart commercial district" will become a visible trend.
AI+AR technology is likely to break the space limitation of shopping
malls and create unprecedented experience upgrade and consumption
upgrade by using new interaction and communication methods.
[Link]+AR games:
Gaming is the most widely used area of AR technology at present.
Since Pokémon Go, the phenomenal-level AR interactive game, became
popular all over the world, AR games have become popular among
more and more players due to the high sense of immersion brought by
the combination of virtual and reality. When compared to the high
degree of homogeneity and repetitive patterns in traditional games,
AI+AR has great potential to bring fresh gameplay, visual expression
and new experience to games, realizing more creativity and
imagination.
[Link] digital venues:
In the era led by digital technology, more and more digital interactive
exhibition items are being used in the design of exhibition halls and
pavilions, which has also become the new vane of the industry. The
introduction of 5G and AR technologies further breaks the physical
space constraints of indoor pavilions, bringing possibilities for the
enhanced memory, experience and cognition of viewers, as well as new
market benefits.
Direction 2: AI+AR industrial Internet application
Driven by 5G technology, the Industrial Internet will develop rapidly,
and at the same time, it will bring opportunities for AI + AR
applications that are involved in multiple parts of the Industrial Internet,
and digital applications for vertical industries will emerge in succession.
This "AI + AR industry application" competition theme is closely
related to the theme of empowering the industry's digital upgrade and
improving production efficiency. It calls for solutions and products that
are innovative, useful and of practical value to industry needs.
Submitting:
Submission of works Our competition schedule is divided into two
stages: preliminary and final. The two stages need to submit different
competition works.
Challenge Track Vertical-track (invite participant to make solutions for 5G, AI and AR
application in vertical industries)
Evaluation criteria Evaluation Standard of preliminary:
Project ( full mark: 100) Evaluation Standard
Benefit evaluation The economic and social benefits of the project to the
industry
(10 marks)
Team (10 marks) Team members have relevant education and work
background; Reasonable division of work; Rigorous
organization; Proper division of property rights and equity
rights; The team has a strong ability to work under pressure,
and it is fully prepared for possible difficulties in starting a
business. The team has a strong interest in the industry
Evaluation Standard of
Project ( full mark: final:
Evaluation Standard
100)
Id ITU-ML5G-PS-003
Title Configuration Knowledge Graph Construction of Loop Network
Devices based on MEC Architecture (Guangdong Division)
Description Background:
If knowledge is the ladder of human progress, knowledge graph is the
ladder of AI. In the past few years, Google, Microsoft, Facebook,
Alibaba, Baidu and other major companies have announced their own
knowledge graph products. Knowledge graph is the premise of
intelligence. The knowledge graph is trying to make the computer think
like human brain, which provides a new perspective and opportunity
for the interpretable AI.
2. Data example
In this contest, A and B data are provided. These two data are
generated by network devices of different manufacturers, and
the data structure will be slightly different.
Resources No
Any controls or restricted data
restrictions Data is under export control and employees of partners cannot
participate in this problem
Specification/Pape
No
r reference
Contact liutf24@[Link]; Tel +86 15652955883; wechat:
yudajiangshan
wangw200@[Link]; weijx29@[Link];
Id ITU-ML5G-PS-004
Title Alarm and prevention for public health emergency based on telecom
data (Beijing Division)
Description Background:
In recent years, the worldwide outbreak of Covid-19, Ebola, MERS and
SARS posed grievous and global affects on human beings and seriously
challenged WHO as well as the health department of many countries.
Apart from the effort of health department, modern informational
technologies and data can help in health emergencies. In this problem
statement, competitors should use the tracking data of telecom users’
geographical movements and DPI information, technologies including
machine learning and big data, to propose comprehensive solutions,
product developing or advises on infrastructure for serious public
health emergencies. All these works can be considered on aspects of
epidemic surveillance, spread monitoring, precise prevention, resource
allocation, effect evaluation for health incidents.
Problems:
This topic focuses on epidemic surveillance, spread monitoring, precise
prevention, resource allocation, effect evaluation by telecom users’
tracking data and DPI information while the outbreak of Covid-19.
Participants should propose related products or solutions by using the
data, resources and developing environment provided by the
competition organizer. If participants use the data from anywhere else,
it should be taken in account that the accessibility and scalability of the
data.
Submitting:
Participants do mining and modeling based on the data provided by the
organizer and yield corresponding solutions or products. The final
submission should cover the following aspects:
Detailed introduction of the solutions or products.
The source code of mining and modeling, as well as the completed zip
file of applications; The model and explanations.
The product prototype, website or APP (optional, plus).
Challenge Track Vertical-track
Evaluation criteria Full marks 100
Id ITU-ML5G-PS-005
Title Energy-Saving Prediction of Base Station Cells in Mobile
Communication Network (Shanghai Division)
Description Background:
With the arrival of the era of mobile Internet + artificial intelligence,
Internet giants have occupied the forefront of AI in the era of AI and
IoT. Operators need to think deeply about how to give play to their
professional advantages, accelerate cross-industry integration and
enhance industry value.
Problems:
The service load of the base station is unevenly distributed in time and
space, and the power supply of the base station cannot follow the
service load of the base station, resulting in energy consumption waste.
Base station AI energy saving project is aimed at the accumulated
operation and maintenance data of operators. Taking AI as the starting
point, the base station is modeled and analyzed based on the historical
data of base station and base station cell, and the energy saving
optimization strategy is generated on the premise of ensuring the
service carrying capacity and coverage.
Submitting:
Contestants need to submit two parts of content in the preliminary
competition: one is to submit the algorithm model and the analysis
results (submitted in. CSV format); The second is the annotated core
code and documentation (a separate attached file submitted as [Link]
file). Finally, all the files are packaged and compressed into a zip file
for submission.
Challenge Track Network-track
Evaluation criteria TP (True Positive): 1 for True and 1 for prediction; FN (False
Negative): true 0, predicted 1; FP (False Positive): true is 1, prediction
is 0; TN (True Negative): 0 for True and 0 for prediction.
According to the following formula, the scores of the contestants are
calculated. According to the accuracy rate (formula 1) and recall rate
(formula 2), F1-score (formula 3) is calculated. Finally, all the
contestants are ranked according to F1-score.
P = TP/(TP+FP) (1)
R = TP/(TP+FN) (2)
F1-score = 2*P*R/(P+R) (3)
Data source This contest provides the resource data of the base station (eci, enodeb,
antenna, carrier frequency, etc.), the resource data of the base station
cell (flow, coverage, PRB, etc.), the cell phone bill information of the
base station cell, the perception data, etc.
In order to protect users' privacy and data security, the data has been
sampled and desensitized. There are null values or junk data in the data
table, and the participants need to handle it by themselves.
Resources None
Id ITU-ML5G-PS-006
Title Core network KPI index anomaly detection (Shanghai Division)
Description Background:
The core network occupies a pivotal position in the entire mobile
operator network. Once the fault occurs, the service quality of the
whole network will be greatly affected. Therefore, it is necessary to
quickly discover the risk of the core network and timely eliminate the
fault before the influence scope is expanded.
Problems:
Key performance indicators (KPIs) reflect network performance and
quality. Analysis and mining of KPI can timely find the risk of network
quality deterioration. The organizer will provide the real data of a
certain operator's core network KPI during the competition, with
sampling interval of 1 hour. Contestants are required to train the model
and detect anomalies in the following 11 days (test data set) according
to the KPI data (training data set) with a history of two and a half
months, including normal labels and abnormal labels.
Submitting:
Contestants need to submit two parts of content in the preliminary
competition: one is to submit the algorithm model and the analysis
results (submitted in. CSV format); The second is the annotated core
code and documentation (a separate attached file submitted as [Link]
file). Finally, all the files are packaged and compressed into a zip file
for submission.
Challenge Track Network-track
Evaluation criteria TP (True Positive): 1 for True and 1 for prediction; FN (False
Negative): true 0, predicted 1; FP (False Positive): true is 1, prediction
is 0; TN (True Negative): 0 for True and 0 for prediction.
According to the following formula, the scores of the contestants are
calculated. According to the accuracy rate (formula 1) and recall rate
(formula 2), F1-score (formula 3) is calculated. Finally, all the
contestants are ranked according to F1-score.
P = TP/(TP+FP) (1)
R = TP/(TP+FN) (2)
F1-score = 2*P*R/(P+R) (3)
Data source [Link] of core network KPI and its meaning.
[Link] data set: data list file of 23 KPIs under different scenarios,
label 1 at abnormal moments.
[Link] data set: data list file of 23 KPIs in subsequent 11 days.
In order to protect users' privacy and data security, the data has been
sampled and desensitized. There are null values or junk data in the data
table, and the participants need to handle it by themselves.
Resources None
Id ITU-ML5G-PS-008
Title Out of Service(OOS) Alarm Prediction of 4/5G Network Base
Station
Description At present, the operation and maintenance of 4/5G BS(base
station) follow a passive pattern, repairing orders will not be
generated until the out of service(OOS) fault occurs. Once the BS
is out of service, users will not be able to connect to the wireless
network, and their regular communication will be affected. In
general, there are some secondary alarms before the major alarm
(OOS alarm). Therefore, in this challenge, the participants are
expected to train an AI model using historical alarm data with
labels of major ones. By excavating the relationship between
alarms, one may use the secondary alarms to predict the
probability of the important alarm happening in a future period, so
that the operation and maintenance personnel can solve the fault in
advance and avoid network deterioration. Due to the similar
operation and maintenance mode of 4G/5G network, after the
large scale commercial use of 5G network, the AI model can be
smoothly transferred as a pre-trained model.
Challenge Track Network-track
Evaluation criteria Submit a comma separated value (CSV) file. The content includes
whether the given base station will have an out of service alarm in
the next 24 hours (or other period). The accuracy of the current
prediction model has reached 78%
Data source 4/5G network fault alarm data from China Mobile.
The data is fault alarm data of several months, including alarm
start time, alarm name, base station name, base station ID, vendor
name, city, etc.
Resources None
Id ITU-ML5G-PS-009
Title Radio signal coverage analysis and prediction based on UE
measurement report
Description Multiple frequency bands are usually deployed in the commercial
network to increase the network coverage and capacity. With the
increasing number of bands, inter-frequency measurements by UEs
may cause amount of signalling overhead and cost huge UE power
consumption and severely impact on running service by the data
interruption for inter-frequency measurement gap. It takes too long
time for UE to choose the proper cell to reside in. This will degrade
the network performance and UE experience. So quick inter-
frequency measurement is desired. One way to obtain the coverage
information of UEs' radio signal quickly is to divide the cell into the
grids by serving cell’s and neighbouring cell’s radio signal levels,
then locate the UE’s grid and perceive UE’s coverage information
based on statistical analysis or directly predict the inter-frequency
measurement based on the intra-frequency measurement, which can
largely reduce the numbers of UE inter-frequency measurement and
benefit for mobility based handover, load balancing, dual
connection and carrier aggregation.
Challenge Track Secure-track
Evaluation criteria Solution, criteria hasn’t been determined
Data source Training data from commercial LTE network with feedback on UE
MR data including RSRP,RSRQ,Earfcn,PCI of serving cell and
neighboring cells.
Resources No
Specification/Paper No
reference
Contact xieyuxuan@[Link]
Id ITU-ML5G-PS-010
Title UE Moblity Analytics in 5G network
Description Background: In 3GPP, the NWDAF is the AI related network
function (NF), which collects data from NFs, OAM and to feedback
around 9 categories analytics to requested NFs (Please refer to
TS23.288). Within the category “UE related analytics”, the UE
mobility analytics or predications could be utilized by NFs, e.g.
AMF, SMF, EIR for some purposes, such as mobility management
parameter adjustment, detect UE been stolen, and etc.
Id ITU-ML5G-PS-011
Title Intelligent spectrum management for future networks
Description Background: Future networks are heterogeneous, e,g, Multi-RAT
(5G, 4G, licensed, unlicensed, fixed, mobile), Multiple platforms
(edge cloud vs. centralized cloud, VNF vs. PNF, Multiple
levels/domains (Access Network vs. Core, network slices with varied
KPI demands, various management and orchestration layers). Also
there several potential data sources e.g. (Peer-to-peer networks, NF,
applications, UEs.
Any controls or Data privacy: No data should be moved from the region.
restrictions Private data from VF (available only to VF approved candidates)
Specification/Paper ITU-T Y.3172 and Y.3174
reference
Contact
[Link]-Eissa@[Link]
Id ITU-ML5G-PS-012
Title ML5G-PHY: Machine Learning Applied to the Physical Layer of
Millimeter-Wave MIMO Systems
Description The increasing complexity of configuring cellular networks
suggests that machine learning (ML) can effectively improve 5G
and future networks. One of the technologies for applications such
as vehicular systems is millimeter (mmWave) MIMO, which
enables fast exchange of data. A main challenge is that mmWave,
as initially envisioned for this application, requires the pointing of
narrow beams at both the transmitter and receiver. Taking into
account extra information such as out-of-band measurements and
vehicles positions can reduce the time needed to find the best beam
pair. Beam training is part of standards such as IEEE 802.11ad and
5G, and has also been extensively studied in the context of wireless
personal and local area networks. Hence, one of the tasks focuses
on beam-selection. Another task is channel estimation, which is
challenging due to mobility, strong attenuation in mmWave and
other issues. This challenge uses datasets obtained with the
Raymobtime methodology. The data consists of millimeter wave
(mmWave) multiple-input multiple-output (MIMO) channels,
paired with data from sensors such as LIDAR.
Challenge Track Network-track, as the challenge consists of use cases related to
signalling or management.
Evaluation criteria Top-K classification for beam selection and normalized mean
squared error for channel estimation
Data source Raymobtime datasets - [Link]
Resources None
Any controls or This Challenge is open to all participants.
restrictions
Specification/Paper [7] 5G MIMO Data for Machine Learning: Application to Beam-
reference Selection using Deep Learning, 2018 -
[Link]
[8] MmWave Vehicular Beam Training with Situational
Awareness by Machine Learning, 2018 -
[Link]
[9] LIDAR Data for Deep Learning-Based mmWave Beam-
Selection, 2019 - [Link]
[10] MIMO Channel Estimation with Non-Ideal ADCS: Deep
Learning Versus GAMP, 2019 -
[Link]
Contact Aldebaro Klautau – aldebaro@[Link]. Tel: +55 91 3201-7181
Id ITU-ML5G-PS-013
Title Improving the capacity of IEEE 802.11 WLANs through Machine
Learning
Description The usage of Machine Learning (ML) is foreseen to be a key
enabler to address the challenges podes by future wireless
networks. In IEEE 802.11 Wireless Local Area Networks
(WLANs), the major challenges will be the user’s density and lack
of coordination, which, given the current channel allocation
mechanisms, lead to sub-optimal performance. One potential
solution is the application of Dynamic Channel Bonding (DCB),
whereby an Overlapping Basic Service Set (OBSS) adapts the
spectrum to be used so that their performance is maximized.
Nevertheless, due to the complexity of massively crowded
deployments, choosing the appropriate channel width is not trivial.
Moreover, increasing the channel width entails a trade-off between
the link capacity and the quality of the link (using more bandwidth
entails a lower received signal strength and leads to a higher
contention). To address the abovementioned challenges, we
propose using Deep Learning (DL) to predict the performance that
will be obtained in an OBSS by using different channel bonding
strategies.
Challenge Track Network-track (students)
Evaluation criteria Participants should provide a .csv file containing the predicted
performance of each BSS (columns) in the different test
deployments (rows).
The evaluation of the proposed algorithms will be based on the
average squared-root error obtained from all the predictions
compared to the actual result in each type of deployment.
Data source To be provided
Resources The IEEE 802.11ax-oriented Komondor simulator [3] has been
used to generate both training and test datasets.
Any controls or This Challenge is open to all student participants.
restrictions
Specification/Paper [11] Barrachina-Muñoz, S., Wilhelmi, F., & Bellalta, B. (2019).
reference Dynamic channel bonding in spatially distributed high-density
WLANs. IEEE Transactions on Mobile Computing.
[12] Barrachina-Muñoz, S., Wilhelmi, F., & Bellalta, B. (2019). To
overlap or not to overlap: Enabling channel bonding in high-density
WLANs. Computer Networks, 152, 40-53.
[13] Barrachina-Muñoz, S., Wilhelmi, F., Selinis, I., & Bellalta, B.
(2019, April). Komondor: a wireless network simulator for next-
generation high-density WLANs. In 2019 Wireless Days (WD) (pp.
1-8). IEEE.
Contact Francesc Wilhelmi, [Link]@[Link] (+34 93 5422906)
Id ITU-ML5G-PS-014
Title Graph Neural Networking Challenge 2020
Description Network modelling is essential to construct optimization tools for
networking. For instance, an accurate network model enables to
predict the resulting performance (e.g., delay, jitter, loss) and
helps finding the configuration maximizes the network
performance according to a target policy.
Currently, network models are either based on packet-level
simulators or analytic models. The former are very costly
computationally while the latter are fast but not accurate. In this
context, Machine Learning (ML) arises as a promising solution to
build accurate network models able to operate in real time.
Recently, Graph Neural Networks (GNN) have shown a strong
potential to be integrated into commercial products for network
control and management. Early works using GNN have
demonstrated an unprecedented capability to learn from different
network characteristics that are fundamentally represented as
graphs, such as the topology, the routing configuration, or the
traffic that flows along a series of nodes in the network. In
contrast to previous ML-based solutions, GNN enables to produce
accurate predictions even in networks unseen during the training
phase. Nowadays, GNN is a hot topic in the ML field and, as such,
we are witnessing significant efforts to leverage its potential in
many different fields (e.g., chemistry, physics, social networks). In
the networking field, the application of GNN is gaining increasing
attention and, as it becomes more mature, is expected to have a
major impact in the networking industry.
Problem statement:
The goal of this challenge is to create a neural network model that
estimates performance metrics given a network snapshot.
Specifically, this model must predict the resulting per-source-
destination performance (delay, jitter, loss) given a network
topology, a routing configuration, and a source-destination traffic
matrix.
Id ITU-ML5G-PS-016
Description Background: In the 5G era, multiple new services are emerging, and
various Internet applications are constantly being enriched, which
has doubled Internet traffic. The rapid growth of traffic has brought a
lot of pressure to network bandwidth, computing, and storage. DPI
data records and presents key traffic information (data statistics start,
end time, and upstream and downstream traffic) in the application
dimension. The analysis of current network traffic models and traffic
service development trends through DPI data is the basis for solving
network congestion, improving user experience, and rationally
allocating and utilizing network resources to improve network
bandwidth utilization.
Problem: Based on the DPI traffic data collected by the big data
platform and the distance between base stations, artificial intelligence
technology can be used to analyse and predict base station traffic, in
order to provide guidance to subsequent network planning, operation
and maintenance. In this problem, we will provide a unified data set
for the participating teams. Each participating team can split the data
set into a training set, a test set, and a verification set, and use it for
training and testing of the AI algorithm model. The purpose of the
algorithm is to predict the traffic trend of base station in the future
through the historical DPI traffic data in the target area and the traffic
information in the surrounding area.
Submitting:
Competitors need to submit two parts in the preliminary competition:
one is to submit the algorithm model and analysis results (submitted
in .csv format); the other is the annotated complete code and
explanatory documents (separately attached files, submitted in .pdf
file format). Finally, all the files are packaged and compressed into a
zip file for submission.
Challenge Track Network-track
Evaluation
criteria
Evaluation criteria: (Mean Absolute Percentage Error,
MAPE),
Data source DPI traffic data collected from the current network and desensitized.
Resources No
Specification/Pap No
er reference
Contact xudan6@[Link]
Id ITU-ML5G-PS-017
Title User-Specific Demand Prediction
Description Background:
In recent years, more and more research has pointed out that by
proactively caching content items, for which users may request, to
the edge of the network, the wireless network can reduce the
download time when users request the data. However, the benefits
of this approach relay heavily on the accuracy of user’s demand
prediction. The more accurate the user's demand prediction, the
greater the benefits of this approach.
Problem:
This topic focuses on user-specific mobile traffic demand
prediction. Competitors need to build mathematical models or
design algorithms to predict the time-varying requesting
probability of each user requesting each content item in the next 24
hours. The time-varying requesting probability can be modelled by
probability density function for continuous random variables and
probability mass function for discrete random variables. This
problem covers four sub-problems as follows.
1. Competitors need to collect datasets by themselves to solve the
problem. They can collect any dataset according to their needs,
e.g., the time spent by each user on TikTok.
Submitting:
Competitors need to solve the problem based on the data collected
by themselves. The final submission should cover the following
aspects:
1. The dataset. In order to facilitate the verification and repeat of
the experiment results, if the competitors solve the problem
based on a public dataset, they need to indicate the source and
download link for the public dataset; if the competitors solve
the problem based on the dataset collected by themselves, they
need to upload their dataset and a detailed report to explain
how they collect the data. (If the dataset is too large, a
download link for the dataset is acceptable.)
2. An annotated source code. In order to facilitate the verification
and repeat of the experiment results, competitors need to
submit all source code and corresponding explanatory
documents.
3. A detailed report. Competitors need to submit a detailed report
to explain how they process the data, build models, design
algorithms, and verify algorithm performance.
(All the files are packaged and compressed into a zip file for
submission.)
Challenge Track Network-track
Evaluation criteria 1. Competitors need upload a detailed report in PDF format to
explain how they process the data, build models, design
algorithms, and verify algorithm performance. The report will
be rated based on the innovation of solutions, the completeness
of implementation, the accuracy of results, and the writing
quality.
2. Competitors need upload a detailed file in CSV format to
record the prediction results and the caching policy.
3. Competitors can use the hit ratio, i.e., the amount of data the
user reads from the cache, to evaluate their caching policy.
Data source Competitors need to collect the data by themselves.
Resources None.
Any controls or None.
restrictions
Specification/Paper [1] M. Lee, A. F. Molisch, N. Sastry and A. Raman, "Individual
reference Preference Probability Modeling and Parameterization for Video
Content in Wireless Caching Networks," in IEEE/ACM
Transactions on Networking, vol. 27, no. 2, pp. 676-690, April
2019.
[2] B. Wu, W. Cheng, Y. Zhang, Q. Huang, J. Li, and T. Mei,
“Sequential prediction of social media popularity with deep
temporal context networks,” in Proceedings of the 26th
International Joint Conference on Artificial Intelligence
(IJCAI’17). AAAI Press, 3062–3068, 2017.
[3] S. D. Roy, T. Mei, W. Zeng and S. Li, "Towards Cross-Domain
Learning for Social Video Popularity Prediction," in IEEE
Transactions on Multimedia, vol. 15, no. 6, pp. 1255-1267, Oct.
2013.
Contact guoxin9@[Link]
2. Resources
NOTE 1- the structure of the list below is intentionally kept simple for our partners to easily
add or change it. The structure is as below:
<<type of resource: 1-line description, link, contact>>
NOTE 2- this list is in no specific order.
[RayMobTime] Data set: Raymobtime is a collection of ray-tracing datasets for wireless
communications. [Link] aldebaro@[Link]
[CUBE-AI] ML marketplace: It is an open source network AI platform developed by China
Unicom Network Technology Research Institute, which integrates AI model development,
model sharing. [Link] , liutf24@[Link]
[Adlik] Toolkit: an end-to-end optimizing framework for deep learning
models. [Link] , [Link]@[Link]
[KNOW] Challenge platform: a data challenge platform which lists several challenges and
competitions. [Link]
[SE-CAID] Data sets: An open AI research and innovation platform for networks and digital
infrastructures for industries, SMEs and academia to share a broad range of telecom data and
AI models. [Link]
[AIIA] Challenge: past competition, led by AIIA in China
[Link]
[Link]
[Link]
__biz=MzU0MTEwNjg1OA==&mid=2247487451&idx=1&sn=cb4370e9fa9d7f827dc632c7
9fe41d2d&chksm=fb2fb81ecc583108221592c69fdea3eb226da933859514dbd9fb8c15288c6f
cb392c65399ddc&mpshare=1&scene=1&srcid=&sharer_sharetime=1575542631509&sharer
_shareid=75fb4d5f665341fa1dafcbc554417e75&key=67a2c7aa29623c33d72ba777f7853d10
2e6f4db8ac8b23733613e267ce0dae54ca817de36bde651b3cf32c3a0daf055c432e46c3b8f43b
088f60edcdef801a54201eea05d0de9051201391ee19fd326f&ascene=1&uin=MjEzNjY3NDQ
5Mw%3D%3D&devicetype=Windows+7&version=62070141&lang=en&exportkey=AoB
%2BIuWyreUPRCOzxdLg0q0%3D&pass_ticket=fCmC%2FiTFfXlmGxvOLq
%2BdVPRElGBj59sZO2eVMyeABxg07Ve7tOfmRWTtKc1rmCRV
[DuReader] Challenge: past competition, includes data sets, including the largest Chinese
public domain reading comprehension dataset, DuReader
[Link]
[IUDX] Data and challenge: a research project for an open source data exchange software
platform, [Link]
[PUDX] Past challenge, Datathon to develop innovative solutions based on India Urban Data
Exchange (IUDX), [Link]
[TI-bigdata] Data: a large dataset of 30+ kinds of data (mobile, weather, energy, etc. from
Telcom Italia big data challenge. [Link]
[TI-phone] Data: The Mobile phone activity dataset is a part of the Telecom Italia Big Data
Challenge 2014. [Link]
[MDC] Data: Mobile Data Challenge (MDC) Dataset, restricted to non-profit organizations,
[Link] need to make a request to get a copy)
[MIRAGE] Data: MIRAGE-2019 is a human-generated dataset for mobile traffic analysis
with associated ground-truth, [Link]
[Urban-Air] Data: An air quality dataset that could be useful for
verticals [Link]
[UCR] Data: UCR STAR is built to serve the geospatial community and facilitate the finding
of public geospatial datasets to use in research and development. [Link]
[NYU] Data: NYU Metropolitan Mobile Bandwidth Trace, a.k.a. NYU-METS, is a LTE
mobile bandwidth dataset that were measured in New York City metropolitian area;
[Link]
[Omnet] Data: Challenge and dataset from comes from Omnet++ network simulator, contains
several topologies and thousands of labeled routings, traffic matrices with the corresponding
per-flow performance (delay, jitter and losses). [Link]
[GNN] Data: data sets for Unveiling the potential of GNN for network modeling and
optimization in SDN. This data set can be divided in two components: (i) the data sets used to
train the delay/jitter RoutNet models and (ii) the delay/jitter RouteNet models already trained
[Link]
network-modeling-and-optimization-in-SDN/tree/master/datasets
[Unity] [Link]
[ETSI ARF] ETSI GS ARF 003 V1.1.1 (2020-03) Augmented Reality Framework (ARF);
AR framework architecture
[Link]
df
[TH_COVID] COVID-19 Live Updates of Tencent Health is developed to track the live
updates of COVID-19, including the global pandemic trends, domestic live updates,
and overseas live updates. [Link]
[HW_NAIE] NAIE Learning Service Telecommunication scenario AI training solutions,
providing pre-consultation from now on. [Link]
[IBM_COVID] IBM has resources to share — like supercomputing power, virus tracking and
an AI assistant to answer citizens’ questions [Link]
[FB-COVID] public data sets from Facebook Data for Good [Link]
[GOOG_COVID] Google Cloud COVID-19 public dataset program: Making data freely
accessible for better public outcomes [Link]
analytics/free-public-datasets-for-covid19
Appendix I: Academic papers of interest
[1] ` "Very Long Term Field of View Prediction for 360-degree Video Streaming", Chenge
Li, Weixi Zhang, Yong Liu, and Yao Wang, 2019 IEEE Conference on Multimedia
Information Processing and Retrieval.
[2] "A Two-Tier System for On-Demand Streaming of 360 Degree Video Over Dynamic
Networks", Liyang Sun, Fanyi Duanmu, Yong Liu, Yao Wang, Hang Shi, Yinghua Ye, and
David Dai, IEEE Journal on Emerging and Selected Topics in Circuits and Systems (March
2019 )
[3] “Multi-path Multi-tier 360-degree Video Streaming in 5G Networks”, Liyang Sun,
Fanyi Duanmu, Yong Liu, Yao Wang, Hang Shi, Yinghua Ye, and David Dai, in the
Proceedings of ACM Multimedia Systems 2018 Conference (MMSys 2018),
[4] “Prioritized Buffer Control in Two-tier 360 Video Streaming”, Fanyi Duanmu,
Eymen Kurdoglu, S. Amir Hosseini, Yong Liu and Yao Wang, in the Proceedings of ACM
SIGCOMM Workshop on Virtual Reality and Augmented Reality Network, August 2017;
[5] Rusek, K., Suárez-Varela, J., Mestres, A., Barlet-Ros, P., & Cabellos-Aparicio, A, “Unveiling the
potential of Graph Neural Networks for network modeling and optimization in SDN,” In Proceedings
of ACM SOSR, pp. 140-151, 2019. [ACM SOSR] [arXiv]
[6] Source code and tutorial of RouteNet. (URL:
[Link]
[7] 5G MIMO Data for Machine Learning: Application to Beam-Selection using Deep
Learning, 2018 - [Link]
[8] MmWave Vehicular Beam Training with Situational Awareness by Machine Learning,
2018 - [Link]
[9] LIDAR Data for Deep Learning-Based mmWave Beam-Selection, 2019 -
[Link]
[10] MIMO Channel Estimation with Non-Ideal ADCS: Deep Learning Versus GAMP, 2019
- [Link]
[11] Barrachina-Muñoz, S., Wilhelmi, F., & Bellalta, B. (2019). Dynamic channel bonding in
spatially distributed high-density WLANs. IEEE Transactions on Mobile Computing.
[12] Barrachina-Muñoz, S., Wilhelmi, F., & Bellalta, B. (2019). To overlap or not to overlap:
Enabling channel bonding in high-density WLANs. Computer Networks, 152, 40-53.
[13] Barrachina-Muñoz, S., Wilhelmi, F., Selinis, I., & Bellalta, B. (2019, April). Komondor:
a wireless network simulator for next-generation high-density WLANs. In 2019 Wireless
Days (WD) (pp. 1-8). IEEE.
_____________