Athletic Performance Scoring Model
Athletic Performance Scoring Model
1 Université
Paris-Saclay & Université de Paris, CNRS/IN2P3, IJCLab, 91405 Orsay, France, 2 Barry University, 11300 NE 2nd Ave,
Miami Shores, FL 33161, USA, 3 Florida Atlantic University, 777 Glades Rd, Boca Raton, FL 33431, USA
* Corresponding
Author E-mail:
basigram@[Link]
Abstract
We present an application of our recently proposed scoring method to running performances. To this end we use the performances cor-
responding to the elite level, as given by the World Athletics scoring tables, in order to calibrate the high-end (max-score) of our scoring
formula. For the lower end (null-score) we start from the prescription put forward by one of the authors (G.P.), stated epigrammatically
as “walking is not running” and obtain an estimate of the corresponding performances. The result is a set of parameters, calculated once
and for all, which allow, given the distance and the registered time, to obtain the scoring for the performance for distances ranging from
sprints to ultramarathons (for both sexes and adjusted for the age of the runner).
Keywords: scoring, running performances, athletics, age factors
1 Introduction
Humans are naturally attracted to competition. The domain of sports is special in the sense that the contest is not motivated by a scarcity
of resources but rather pursues the attaining of a goal (usually in the form of a personal best, and, for the elite, a record). Competition is
a central dynamic in most (if not all) sports. According to R. Mertes [Martens 1976] competition is a form of social evaluation. Viewed
under this angle, the term competition encompasses not only the comparison of individuals against one another but also against some
objective standard of excellence. Organized competition brings order to the process since it fixes the rules defining the means which are
allowed in order to achieve a goal.
Looking back at the origins of organized sport in Ancient Greece [Golden 2004] one encounters the first facet of competition, that of
the comparison to others. In fact, in the Ancient Greek tradition [Miller 2004] competitions were governed by the quest for excellence,
where the most important thing was victory [Leonard 2016]. This had as consequence that no precise measurements were required, and
neither the distances nor the implement weights were standardized. In fact, very few data on the performances of the ancient Greek
athletes have been preserved. However, over the centuries, the situation evolved and the technological progress helped make sports more
quantitative. Athletics profited from the increasing accuracy in time and distance measurements.
Having precise measurements facilitates also the second facet of competition, where the aim is not the comparison to others but
rather the accomplishment of a self-fixed goal. Thanks to the measurements one has a way to assess one’s level of ability and appreciate
the progress towards the fixed goal. Athletics is par excellence a quantitative sport where the performance is expressed in time or length
allowing thus straightforward comparisons within a specific event. But once one wishes to compare performances obtained in two
different events, the mere result of a precise measurement is not sufficient. This is where scoring tables can be valuable. They made their
appearance in Athletics at the end of the 19th century in order to allow the classification in combined events [Zarnowski 1989]. However,
the question of performance scoring can be asked in a more general setting. How does one attribute a score for a performance in a sport
where precise measurements are not really helpful (like mountain climbing or open-sea sailing)? D. Harder addressed this question
[Harder 2001] and his answer, to put it in a nutshell, is that “two performances are equivalent when the fraction of the practitioners who
obtain them is comparable". The mathematical foundation for the construction of scoring tables has been set by one of the co-authors in
[Grammaticos 2007] and served as a basis for the recent proposal [Grammaticos-Meloun-Purdy 2022] of a scoring system extending the
one proposed by another co-author [Purdy 1974-77]. The idea behind our approach is that the score of a performance should be linked
to the inverse of the probability of exceeding this performance. The best, practical, way to assess this probability is through the statistics
of the distribution of performances establishing, thus, a link between our scoring approach and that of Harder.
As already explained (and, in fact, announced in the title of the paper) this work will focus on scoring performances in athletics, and
more specifically on running. Scoring the athletic performance is far from trivial. Several questions must be answered before setting
up a scoring system and, what is worse, the various choices exert influence upon each other. In order to construct a well-balanced and
fair scoring one must first determine what is the performance warranting an arbitrarily fixed high mark (typically we are talking here
about 1000 points), the max-score, and what is the performance below which one will obtain zero points, the null-score. Once this is
done for one event it has to be done for every other one taking particular care that the performances corresponding to 1000 and 0 points
are indeed equivalent. Even when this is ensured, there is no guarantee that the intermediate performances will be equivalent all along
the scoring table. This depends on how the performances are distributed for the various events and the simplifying, tacit, assumption of
scoring tables specialists is that these distributions follow indeed the same law for all events.
An essential requirement for fair scoring is that it be progressive [Trkal 2003]. In practice ‘progressive’ means that the same increase
in performance garners a larger increase in points when the performance increases. For instance, a gain of 1 second in 400 m will bring
more points when it corresponds to going from 46 s to 45 s compared to the case when the athlete goes from 58 s to 57 s. The progressive
character of scoring is a sine qua non for fair scoring. It was already a main feature of the tables proposed by one of us [Purdy 1972] and
will be present in the tables that will be presented in the next section.
One caveat is necessary at this point. Everybody has their preferred events. Nobody will score equally well in short-, middle- and
long-distance events. Having a good mark in some range of distances and so-so marks outside this range is natural and in no way an
indication of unfair scoring. To use a striking example, a sprinter may be excellent in short-distance running but will score poorly over
longer distances and even may well be unable to finish a marathon. Thus, it’s important to realise that the scoring function represents a
statistical summary of the entire population and not a representation of any one member of the population. With that said, the scoring
function is useful by individual members of the population for measuring the value of their performance.
An important remark concerns the differences between men and women. When it comes to running events women are roughly 10 %
slower than men. And this difference is more or less constant over the whole spectrum of running events, with distances covering several
orders of magnitude. A fair scoring must account for this difference and attribute the adequate number of points.
Finally a fair scoring method should take into account the variations of the performance with age. In order to account for this one
must introduce appropriate age factors which correct the performance converting it, in some sense, to the performance the athlete would
have registered in his prime.
The present paper will focus on scoring running performances covering the whole spectrum, from very short sprints to multi-
day ultramarathons. We shall start with a short presentation of our scoring system explaining how one can use data in order to fix the
parameters of the scoring formula. We shall then explain how we have extracted those data from the World Athletics tables [Spiriev 2022]
and the best performances on ultra-running [Wikipedia] allowing to fix the upper part of the tables. The approach presented in the work
of one of us, G.P., is used as a basis for the determination of the null-score, the latter being another necessary ingredient for the
determination of the parameters. The result is a closed-form mathematical formula which allows to compute the scoring for any distance
between 50 m and over 1000 km from the simple datum of the time to cover the distance (as well as the sex and the age of the runner).
One of the authors, G.P., developed the brand name “TraxScore” a number of years ago as a vehicle to help road race organizers, coaches
and runners have a handle by which they could represent their performance from any distance and time. It will be used throughout this
paper.
where a, b, c, d, f are parameters. The first term of (1), as explained in [Grammaticos-Meloun-Purdy 2022], corresponds to the Gompertz
part of the distribution of performances, with b controlling the rate of growth of scoring with the velocity. The factor a adjusts the
contribution of this term to the total score (up to a global scaling). The second term, corresponds to the skewed logistic part of the
performance distribution. Again, the parameter d controls the rate of growth of scoring with the velocity (and its effect is felt mainly for
low velocities). Quite expectedly, f calibrates the contribution of this term to the score.
Obviously, we could have constructed the scoring formula in terms of the time t registered during the running event over a distance
s, but we have opted for the velocity since the latter is the natural physical quantity through which to assess the athlete’s effort. The
velocity has been normalized so as to make the parameters b and d dimensionless. Starting from the mean velocity v = s/t, obtained
during the running event, we divide it by the value vm corresponding to the value that obtains the maximum number of points. Thus the
normalized velocity u = v/vm can be understood as u = tm /t where tm is the time that obtains the maximum number of points. Speaking
about the latter, it is clear that an overall scaling of the scoring points can be applied depending on whether we wish to score over 1000
points or any other value. For instance, if we have the coefficients fixed so as to attribute 1000 points to a performance u = 1 and decide
to attribute 1400 points to this performance, it suffices to multiply a and f by a factor 1.4. (It goes without saying that it is possible to
apply the scoring formula to values of u exceeding 1, i.e. corresponding to velocities higher than vm ).
In Figure 1 we present a graphic of the scoring curve where both the scoring points and the performance (normalized mean velocity)
have their maximum fixed to 1.
Figure 1. An example of normalized scoring. The solid line represent the total number of points, the dashed line corresponds to the
part involving the logarithm and the dot-dashed line comes form the exponential part
Using this curve as a guide we can now explain the procedure for the fitting of the parameters. We see that the exponential part
is negligible for the very low performances. This is a choice of ours, so as to have the fast increasing part entering only at higher
performances thus allowing us to fix separately the parameters of the two parts of the scoring equation.
In [Grammaticos-Meloun-Purdy 2022] it was decided that the parameter c can be fixed once and for all to a value in the 0.01-0.0001
range (corresponding roughly to the fraction of the population unable to register any performance). For simplicity, we are going to work
here with a value c = 0.001. Second, we remark that the curve obtained from (1) starts from the origin of the coordinates, i.e. zero points
correspond to zero performance. Thus the null-score normalized performance z = v0 /vm , being finite, will inevitably score a non-zero
number of points q. However we can make this number of points as small as we like, for instance something between 1 and 10 when the
maximum score is 1000 (or between 0.001 and 0.01 if we normalize the maximum score). This leads to a first relation
f log 1 + c(edz − 1) = q.
(2)
The effect of the term involving the logarithm on higher performances can be seen in Figure 1: the number of points grows practically
linearly with the performance.
Next we turn to the exponential term. As explained in [Grammaticos-Meloun-Purdy 2022], the parameters a and b can be fixed in
a very simple way if one decides what is the contribution of this term at a performance half of the maximum one, i.e. u = 1/2 and at
the maximum u = 1. We can assume that at u = 1/2 we have a very small number r and that at the maximum u = 1 the contribution
of the exponential term is w. (The choice of the value of w is arbitrary. In fact a value of w in the 0.20-0.25 range, as in Figure 1, is
guaranteed to provide a, perfectly acceptable, moderately progressive scoring). The parameters a and b are then given by the expressions
a = r2 /(w − 2r) and b = 2 log((w − r)/r). Finally we use the fact that the maximum number of points m (in principle normalized to 1),
obtained for u = 1 is the sum of the contributions of the exponential term, which as we saw is equal to w, and the logarithmic term. This
allows us to write a second equation involving f and d:
w + f log 1 + c(ed − 1) = m.
(3)
Solving (2) and (3) allows us to obtain f and d. However given the form of the equations the solution can only be obtained numerically
(or graphically). In fact, eliminating between (2) and (3) we obtain an equation for d
d m−w
log 1 + c(edz − 1) ,
log 1 + c(e − 1) = (4)
q
MEN WOMEN
distance s (m) velocity vm (m/s) velocity vm (m/s)
50 8.83 8.24
55 9.02 8.38
60 9.20 8.50
100 9.98 9.09
200 9.94 8.95
300 9.52 8.43
400 8.95 7.98
500 8.52 7.63
600 8.14 7.17
800 7.68 6.78
1000 7.46 6.55
1500 7.03 6.23
2000 6.82 6.08
3000 6.58 5.85
5000 6.39 5.67
10000 6.11 5.40
The second table, also obtained from the World Athletics tables, gives the velocities for road events. Note that for the distances of
5000 and 10000 m there exist separate scorings for track and road events. However, given the increasing popularity of road events and
the high level of performances registered there, the difference in scoring (in the most recent, 2022, tables) between track and road events
is negligible.
MEN WOMEN
distance s (m) velocity vm (m/s) velocity vm (m/s)
15000 5.97 5.36
20000 5.92 5.29
21097 5.89 5.28
25000 5.80 5.18
30000 5.69 5.08
42195 5.51 4.90
100000 4.44 4.14
The third table gives the velocities as a function of distance for ultramarathon events. Since some events are based on time duration
rather than distance, we have separated accordingly the performances of men and women.
MEN WOMEN
distance s (m) velocity vm (m/s) velocity vm (m/s)
50000 5.14 4.63
80450 4.62 3.94
85492 (6 hr) - 3.96
97200 (6 hr) 4.50 -
100000 4.51 4.24
149130 (12 hr) - 3.45
160900 4.12 3.52
177410 (12 hr) 4.11 -
270116 (24 hr) - 3.13
309400 (24 hr) 3.58 -
397103 (48 hr) - 2.32
473496 (48 hr) 2.74 -
883631 (6 days) - 1.71
1000000 2.04 1.51
1036800 (6 days) 2.00 -
1609000 1.78 1.48
A first remark concerns the ratio of women to men mean velocities. This is a question that was addressed in a previous publication
of one of the authors [Grammaticos-Charon 2014]. The conclusion there was that for distances up to the marathon, the ratio of women
to men velocities is roughly 0.9. Perusing the results of the three tables we see that this ratio holds for distances up to 100 km.
It diminishes beyond this point but we believe that this is not a physiological effect but rather one due to the fact that there are fewer
women participating in ultramarathon events. In order to simplify our approach we shall assume that the ratio 0.9 holds over all distances
and thus posit that the max-score velocity for women is 9/10 that of men for the same distance.
Having settled the question of women versus men we return now to the initial question, namely of how to obtain the variation of
velocity as a function of the distance. Clearly this is a delicate matter given the fact that we pretend to cover a range where the ratio of
distances (or of durations) of the longest to the shortest events is of the order of 104 -105 . It is clear that the physiological mechanisms
entering the athlete’s effort vary substantially along the whole spectrum of events and the function v = f (s) can be quite complicated. In
[Purdy 1974] one of the authors had presented a graphical representation of that function together with a best fit, albeit on distances not
exceeding 100 km. In what follows we are going to distinguish five domains over which we shall obtain a relation between v and s. The
first domain covers the distances from 50 to 300 m. This domain is governed by the fact that the athlete must accelerate from velocity
zero up to his maximal velocity. This explains the left branch of the curve in Figure 3.
Figure 3. Velocity as a function of distance for races spanning the interval 50-300 m.
However the maximal velocity, once reached, cannot be maintained for long, the athlete depleting fast his alactic anaerobic energy
reserves. The lactic anaerobic mechanism enters into play and is crucial in races spanning the distances from 300 to 1000 m, Figure 4.
Figure 4. Velocity as a function of distance for races spanning the interval 300-1000 m.
Beyond these distances, and roughly up to the marathon, the main energy production mechanism is the aerobic one leading to the
curve represented in Figure 5.
Figure 5. Velocity as a function of distance for races spanning the interval 1000 m to the marathon.
Once the duration of the effort exceeds two to two and a half hours the glycogen reserves of the organism are depleted and the
aerobic mechanism relies now on the use of lipids. This results into a sharp decline of the velocity, as can be seen in Figure 6, covering
the distances up to 24 hours.
Figure 6. Velocity as a function of distance for races from the marathon to 300 km.
Finally for events that last for more than 24 hours, the athlete’s effort is interspersed by unavoidable pauses corresponding to sleep,
ingesting food and taking care of other bodily functions, resulting in a substantial decrease of the velocity, Figure 7.
such would have made impossible the simplifications we are aiming at, for the construction of our scoring formula. Still we are going
to use this prescription as a basis for the determination of the null-score velocity. To this end we start from the null-score velocities
given by the expression v0 = 2 − s/105 for races from 5000 to 40000 m and we fit the points obtained with an expression v0 = B/sγ
where γ = 0.0721, i.e. the very same value of the exponent obtained from the fit of the max-score velocities in the range 1000 m to the
marathon.
the number of corresponding points. A word of caution is necessary at this point concerning very short distances. Although the scoring
formula is valid from a purely mathematical point of view, we believe that the domain of sprints would necessitate a special treatment
(and we shall deal with this in some future work of ours), while a blind application of some mathematical recipe may lead to unwarranted
conclusions. Thus, to be on the safe side, we recommend some caution in the application of our scoring approach to distances of 50-150
m.
The procedure we just described is elementary and its practical implementation would necessitate just a few lines of code.
6 Conclusion
The main objective of this paper was to show that the scoring tables we proposed in [Grammaticos-Meloun-Purdy 2022] can be applied
to a real-life situation. To this end, we decided to provide scoring for running events, since running is an ubiquitous athletic activity,
practiced by a substantial fraction of the active population and has an extreme variety as far as the distances involved are concerned.
Our aim was not to introduce scoring tables for running aiming at replacing existing ones but rather to provide a proof-of-concept
constructing a scoring formula starting from reliable, already statistically treated, data. To this end we chose the World Athletics tables
(the ones covering all athletics events and not just the combined events ones) and decided to obtain the max-score performance of our
scoring from the one corresponding to 1200 points in the World Athletics tables. The slight inconvenience of this approach is that no
scoring of performances is available beyond the 100 km race. In order to palliate this we complemented the max-score performances
obtained from the tables with the best performances for ultramarathon events.
Having the max-score velocity as a function of the distance we proceeded to fit it with a simple mathematical expression that
accurately models human performance in running. We did not attempt an overall fit with a unique formula since we believe that fitting
per range of distances is better adapted to the various physiological regimes which play the major role in the races in question. The null-
score performance was obtained based on estimates by one of us (G.P.) leading in the end to a constant ratio of the null-score velocity
to that of the max-score. In order to provide scoring tailored to the performance differences between men and women we decided to fix
the max-score velocity for women to 90 % of that of men.
While the mathematical expression of our scoring formula may look a little bit complicated it has the advantage of being straight-
forward to implement. In fact the scoring is provided by the sum of two terms only one of which plays a role in low-to-moderate
performances and which becomes very simple when the performances grow. The second term plays a role only at the high end of the
spectrum and thus the two terms can be adjusted almost independently. This, combined with the fact that we have fixed the ratio of null-
to max-score, allows to calculate the parameters of the scoring once and for all, making them in some sense universal. We have thus a
scoring formula which is valid for any distance, say from 20 m to 2000 km, and which can be practically implemented with minimal
effort.
This work is the fruit of a collaboration which aimed, first, at correcting mistakes in performance scoring and, second, at providing a
scientific basis thereof, constituting thus a significant contribution to the science of modelling human performance. The initial motivation
for such an enterprise came from the domain of combined events. Several mathematical models [Grammaticos-Meloun-Purdy 2022]
have been proposed for their scoring over the years, none being devoid of drawbacks. Although we believe that our approach could lead
to a definite improvement of combined events scoring, we decided to focus on running, since this is an exercise enjoyed by millions of
people, rather than a handpicked elite.
Most recreational runners have great difficulty when comparing performances from one event to another. This becomes even more
arduous when it comes to different individuals participating in different races. Our model allows to answer the question of performance
comparison in a precise, quantitative, way. It can provide a universal scoring, expressed as a number of points, for any running distance.
In order to give runners a friendly ’handle’ we have introduced the term TraxScore. This allows comparisons between performances of
a given runner in different events but also comparisons between different runners.
The usefulness of our approach is multiple. Knowing his TraxScore a runner can adjust his objectives over various distances in a
more precise way. A universal score allows also to monitor in a precise way the runner’s progress. Concerning race organizers, the use
of the TraxScore would allow to group the participants in a race into corrals of runners of roughly the same value and make possible
well-ordered staggered starts. Moreover, knowing one’s level of ability should allow the runner to plan his training in a rational way,
helping him to improve his performances while minimizing the injury risk.
The development of the TraxScore is part of an ambitious project of ours, the Human Performance Modelling. A first objective
consists in making the TraxScore available to road races all over the world. Using race results we intend to refine the parametrization of
our scoring formula. We will thus be able to provide a more accurate scoring for the various age groups for both sexes. Accounting for
the altitude of the course, could be included in a future version of TraxScore. A more ambitious extension would be to provide scoring
taking into account environmental factors, like temperature, humidity and weather in general. These improvements would make possible
handicap races, both in real time and virtually. The existence of a top-quality TraxScore would assist organizers in their task and be a
most useful guide to millions of runners worldwide.
Several directions of further research appear at this point. First, it would be easy, provided organized data exist, to apply our scoring
approach to other locomotion-based disciplines like cycling or swimming. Second, it would be interesting by analyzing real-world data
from various popular races to provide a feedback to the scoring proposed by World Athletics. Of course, we are aware that this is a tall
order necessitating a substantial investment (not only) in time. Finally, since scoring tables are traditionaly related to combined events, it
would be interesting to extend the treatment presented here to field events, reconnecting thus with the initial program of one of us, who
in [Purdy 1972] proposed scoring tables with the decathlon in mind.
References
[Martens 1976] R. Martens, Competition: In need of a theory. In D. M. Landers (Ed.), Social problems in athletics. Urbana: University
of Illinois Press (1976) 9.
[Golden 2004] M. Golden, Sport in the Ancient World, Routledge (London), 2004.
[Miller 2004] S.G. Miller, Ancient Greek Athletics, Yale Univ. Press (New Haven), 2004.
[Leonard 2016] J. Leonard, The Value of Athletic Glory in Ancient Greece, online at [Link] 2016.
[Zarnowski 1989] F. Zarnowski, The Decathlon, Leisure Press (New York), 1989.
[Harder 2001] D. Harder, Sports Comparisons, Education Plus (Castro Valley CA) 2001.
[Grammaticos 2007] B. Grammaticos, The physical basis of scoring the athletic performance, New Stud. Athl. 22:3 (2007) 47.
[Grammaticos-Meloun-Purdy 2022] B. Grammaticos, J. Meloun and J.G. Purdy, Distribution of performances and scoring in athletics,
Math. and Sports 3 (2022) 1.
[Purdy 1974-77] J.G. Purdy, Computer generated track and field scoring tables, in three parts: I. Historical development, II. Theoretical
foundation and development of a model, III. Model evaluation and analysis, Med Sci. Sports 6 (1974) 287, 7 (1975) 111,
9 (1977) 212.
[Trkal 2003] V. Trkal, The development of combined events scoring tables and implications for the training of decathletes, New Stud.
Athl. 18:4 (2003) 7.
[Purdy 1972] J.G. Purdy, The application of computers to model physiological effort in scoring tables for track and field, PhD thesis,
Stanford Univ. 1972.
[Spiriev 2022] B. Spiriev, IAAF Scoring Tables of Athletics, World Athletics (Monaco), 2022.
[Wikipedia] Wikipedia, Ultramarathon, online at [Link]
[Grammaticos-Charon 2014] B. Grammaticos and Y. Charon, Comparing the best athletic performances of the two sexes, New Stud.
Athl. 29:4 (2014) 37.
[Purdy 1974] J.G. Purdy, Least square model for the running curve, Res. Q. Exerc. Sport 45 (1974) 224.
[Grammaticos 2009] B. Grammaticos, Scoring the athletic performance for age groups, New Stud. Athl. 24:3 (2009) 63.
[Grammaticos 2020] B. Grammaticos, On the Rise and Fall of athletic performances, online at
[Link]