Introduction
Order statistics is a very useful concept in statistical science. They have a wide range of applications including auction modeling, auto racing and insurance policies, optimization of production processes, estimation of parametersThe "parameters" are variables or criteria that are used to define, measure or evaluate a phenomenon or system. In various fields such as statistics, Computer Science and Scientific Research, Parameters are critical to establishing norms and standards that guide data analysis and interpretation. Their proper selection and handling are crucial to obtain accurate and relevant results in any study or project.... of distributions, et al. Through this article, we will understand the idea of order statistics. We will first understand its meaning and gradually proceed to its distribution., eventually covering more advanced concepts.
Suppose we have a set of random variables X1, X2, …, XNorth, that are independent and identically distributed (iid). For independence, we mean that the value taken by a variableIn statistics and mathematics, a "variable" is a symbol that represents a value that can change or vary. There are different types of variables, and qualitative, that describe non-numerical characteristics, and quantitative, representing numerical quantities. Variables are fundamental in experiments and studies, since they allow the analysis of relationships and patterns between different elements, facilitating the understanding of complex phenomena.... random variable is not influenced by the values taken by other random variables. By identical distribution, we mean that the probability density function (PDF) (or equivalently, the cumulative distribution function, CDF) for random variables it is the same. The Kth The order statistic for this set of random variables is defined as kth smallest sample value.
To better understand this concept, we will take 5 random variables X1, X2, X3, X4, X5. We will observe a realization / random result of the distribution of each of these random variables. Suppose we obtain the following values:

The Kth the order statistic for this experiment is kth smallest value of the set {4, 2, 7, 11, 5}. Then, the 1S t the order statistic is 2 (smallest value), the 2North Dakota the order statistic is 4 (the next smallest), and so on. The 5th the order statistic is the fifth smallest value (the greatest value), What is it 11. We repeat this process many times, namely, we extract samples from the distribution of each of these iid random variables and find the kth smallest value for each set of observations. The probability distribution of these values gives the distribution of kth order statistics.
In general, if we order random variables X1, X2, …, XNorth in ascending order, then the kth the order statistic is displayed as:

The general notation of the kth the order statistic is X(k). Note X(k) is different from Xk. Xk is the kth random variable of our set, while X(k) is the kth statistical order of our set. X(k) takes the value of Xk yes Xk is the kth Random variable when realizations are sorted in ascending order.
The 1S t X-order statistic(1) is the set of minimum values of the realization of the set of 'n’ random variables. Laterth X-order statistic(North) is the set of maximum values (nth minimum values) of the realization of the set of 'n’ random variables. They can be expressed as:

Order statistics distribution
Now we will try to find out the distribution of the order statistics. We will first describe the distribution of the nth order statistics, then he 1S t statistical order and finally the kth general order statistics.
A) Distribution of nth Order statistics:
Let the probability density function (PDF) and the cumulative distribution function (CDF) our random variables let fX(x) y FX(x) respectively. By definition of CDF,

Since our random variables are identically distributed, have the same PDF fX(x) y CDF FX(X). Now we will calculate the CDF of nth order statistics (FNorth(x)) as follows:

Random variables X1, X2, …, XNorth they are also independent. Therefore, by property of independence,

The PDF of the nth statistical order (fNorth(x)) is calculated as follows:

Therefore, the expression for PDF and CDF of nth The order statistic has been obtained.
B) Distribution of 1S t Order statistics:
The CDF of a random variable can also be calculated as the one minus the probability that the random variable X takes a value greater than or equal to x. Mathematically,

We will determine the CDF of 1S t order statistics (F1(x)) as follows:

One more time, using the independence property of random variables,

The PDF of the 1S t statistical order (f1(x)) is calculated as follows:

Therefore, the expression for PDF and CDF of 1S t The order statistic has been obtained.
C) Distribution of the kth Order statistics:
Forkth order statistics, in general, the following equation describes your CDF (Fk(X)):

The PDF of kth statistical order (fk(x)) is expressed as:

To prevent confusions, we will use geometric proofs to understand the equation. As discussed above, the set of random variables has the same PDF (fX(X)). The graphic below shows a sample PDF with the kth Order statistic obtained from random sampling:

Then, the PDF of the random variables fX(x) is defined between the interval [a,b]. The k-th order statistic for a random sample is shown with the red line. The other variable realizations (for the random sample) shown by the small black lines on the x-axis.
There are exactly (k – 1) observations of random variables that fall in the yellow region of the graph (the region between & kth order statistics). The probability that a particular observation falls into this region is given by the CDF of the random variables (FX(X)). But we are aware that (k – 1) observations fell in the region, what gives us the term (for independence) (FX(X))(k – 1).
There are exactly (n – k) observations of random variables that fall in the blue region of the graph (the region between kth statistical order & b). The probability that a particular observation will fall in this region is given by the 1 – CDF of the random variables (1– FX(X)). But we are aware that (n – k) observations fell in the region, what gives us the term (for independence) (1-FX(X))(n – k).
Finally, exactly 1 observation falls exactly on the k-th order statistic with probability fX(X). Therefore, the product of 3 terms gives us an idea of the geometric meaning of the equation for PDF of the k-th order statistic. But, Where does the factorial term come from? The above scenario only showed one of the many orderings. There can be many of these combinations. The total number of such combinations is shown below:

Therefore, the product of all these terms gives us the general distribution of kth order statistics.
Useful functions of order statistics
Order statistics lead to several useful functions. Among them, notable ones include the sample range and the medianThe median is a statistical measure that represents the central value of a set of ordered data. To calculate it, the data is organized from lowest to highest and the number in the middle is identified. If there are an even number of observations, the two core values are averaged. This indicator is especially useful in asymmetric distributions, since it is not affected by extreme values.... of the sample.
1) Sample range: It is defined as the difference between the largest and smallest value. It is expressed as follows:

2) Sample median: The sample median divides the random sample (realizations of the set of random variables) in two halves, one containing the lowest value samples and the other containing the highest value samples. It's like the middle order statistic / central. It is mathematically defined as:

Order statistics set PDF
A joint probability density function can help us better understand the relationship between two random variables (two-order statistics
in our case). The joint PDF for any statistics of 2 X orders(a) & X(B), such that 1 ≤ a ≤ b ≤ n is given by the following equation:

Example
We will use a very simple example to illustrate the distribution of the order statistics: the standard uniform distribution (U[0, 1] distribution). We will take 5 random variables X1, X2, X3, X4, X5, everyone has the U[0, 1] distribution. For this set of random variables, we will calculate and plot the 1S t, 3rd (the sample median) Y 5th (Northth) order statistics. The following figure"Figure" is a term that is used in various contexts, From art to anatomy. In the artistic field, refers to the representation of human or animal forms in sculptures and paintings. In anatomy, designates the shape and structure of the body. What's more, in mathematics, "figure" it is related to geometric shapes. Its versatility makes it a fundamental concept in multiple disciplines.... shows the U[0, 1] distribution:

We will draw random samples as follows and find the 1S t, 3rd & 5th order statistic for each sample. Below are two of the samples:

The standard uniform distribution PDF and CDF are given as:

We will use this information and calculate X(1), X(3) & X(5) using the formulas we derived. We will take the case only when x is between 0 Y 1 (for other cases, the order statistic is zero since PDF is zero).
A) To 1S t order statistics:

Plot for f1(X):

B) To 3rd order statistics:

Plot for f5(X):

C) To 5th order statistics:

Plot for f5(X):

Conclution
Therefore, we have explored the concepts of order statistics in depth. A wide range of physical processes can be modeled through order statistics, exploiting its properties, particularly their distributions.
The media shown in this article is not the property of DataPeaker and is used at the author's discretion.



