scieee AI-readable full text Open interactive document viewer

On the Statistical Characterization of Sprays

Panão, Miguel O.,Moita, Ana S.,Moreira, António L.

Abstract

The statistical characterization of sprays is an essential way of organizing data on drop size and velocity to provide reliable information on the spray dynamics. A clear presentation of data using statistical tools provides evidence of a clear research question underlying the spray characterization. In this article, a review of the best practices to build histograms is presented, as well as three relevant details on spray characterization: (i) the application of information theory to assess if we have enough information (not data); (ii) the link between mathematical probability distributions and the physical interpretation of spray data; (iii) and introducing, for the first time, the concept of drop size diversity, with the quantification of the polydispersion and heterogeneity degrees. Finally, the view presented is applied to the characterization of nanofluid sprays for thermal management.

Full text

applied sciences Review On the Statistical Characterization of Sprays Miguel O. Panão 1,* , Ana S. Moita 2,3 and António L. Moreira 2 1ADAI, LAETA, Mechanical Engineering Department, University of Coimbra, Rua Luis Reis Santos, 3030-788 Coimbra, Portugal 2 IN+, Mechanical Engineering Department, Instituto Superior Técnico, University of Lisbon, Av. Rovisco Pais, 1049-001 Lisboa, Portugal; [email protected] (A.S.M.); [email protected] (A.L.M.) 3CINAMIL, Department of Exact Sciences and Engineering, Portuguese Military Academy, Rua Gomes Freire, 203, 1169-203 Lisboa, Portugal *Correspondence: [email protected] Received: 10 August 2020; Accepted: 29 August 2020; Published: 3 September 2020   Featured Application: The work establishes the grounds for the spray characterization using statistical analysis, how to explore it to improve the physical interpretation of spray processes, and advance the methods to report it in a way that further advances spray science. Abstract: The statistical characterization of sprays is an essential way of organizing data on drop size and velocity to provide reliable information on the spray dynamics. A clear presentation of data using statistical tools provides evidence of a clear research question underlying the spray characterization. In this article, a review of the best practices to build histograms is presented, as well as three relevant details on spray characterization: (i) the application of information theory to assess if we have enough information (not data); (ii) the link between mathematical probability distributions and the physical interpretation of spray data; (iii) and introducing, for the first time, the concept of drop size diversity, with the quantification of the polydispersion and heterogeneity degrees. Finally, the view presented is applied to the characterization of nanofluid sprays for thermal management. Keywords: drop size distributions; drop size diversity; information theory; nanofluid sprays 1. Introducing the Statistical Organization of Spray Data A spray is a two-phase flow of droplets interacting with a gaseous continuous phase. The physical process of liquid atomization depends on the atomizer type and breakup process, and once completed, the droplets formed have multiple sizes and velocities, and statistical histograms are the most common way of organizing the large amount of data on their characteristics. In the sense of data organization, the histograms organizing the sizes and velocities of droplets by classes do not represent a probability of occurrence, as in conventional statistical analysis, but a probability of presence of droplets in a spray, since the atomization mechanisms already occurred. This small language shift allows considering each probability value as representing the degree of relevance of a certain class in the spray—a notion which will be essential to understand drop size diversity. In practice, after sorting data by classes, the way histograms represent the probability of presence is dividing the counts in each class ( nk ) by the total sample size ( N ), pk=nk/N . However, if there is a need to increase the detail of the distribution, one can build the discrete distribution in terms of density of probability of presence by considering the bin width ( δDk ) in the probability value as pdk=pk/δDk . The bin width is constant if size classes are regularly spaced within the spectrum, or can vary the size if irregularly spaced. Appl. Sci. 2020,10, 6122; doi:10.3390/app10176122 www.mdpi.com/journal/applsci Appl. Sci. 2020,10, 6122 2 of 18 In the case of discrete probability distributions, the sum of all probability values associated with each class k is equal to one, ∑pk= 1, which means that each probability value pk corresponds to the weight a number of drops within a characteristic class k has in the entire spray. It is why one designates this way of presenting spray data as a number-weighted probability distribution. There are other ways as shown later. The interpretation of the probability as a number-weighted value of class k , for example, applied to drop sizes dk , allows for calculating the moments of the size distribution, such as the average size of droplets, d10 =∑kdkpk=∑kdkpdkδDk , whether using a probability discrete distribution or a probability discrete density distribution, respectively. Although this is basic statistical knowledge, in several research works, it is unclear which is the distribution reported, considering that each approach (probability or probability density) reacts differently when we increase the detail of a distribution by changing the bin width δDk, as illustrated in the example of Figure 1. While a smaller bin size implies a higher number of classes, in probability distributions, it leads, ultimately, to a uniform probability distribution with one class per sample (Figure 1a). However, in probability density distributions, it increases the detail to allow identifying eventual multimodalities dues to different drop cluster with similar characteristics, or increase the noise in pdk values, as illustrated in the examples depicted in Figure 1b. Figure 1. Example of the effect of changing the bin size δ in discrete probability ( a ) and probability density (b) size distributions. The simulated data follow a lognormal distribution function. The purpose of remarking the difference between probability and probability density distribution is particularly relevant when comparing experimental with simulated drop characteristics. In the case of using a probability distribution ( pk ), the number of classes must be the same, or the representative values of each class ( dk for drop size and uk for one velocity component). Otherwise, if the presentation of spray data opts for the probability density distribution, which is dimensional ( pdk [ µ m −1 ]), it is not necessary to use the same number of classes. Ultimately, to avoid erroneous comparisons between experimental and simulated data, it is essential to be clear about the approach followed when presenting spray data in statistical format. The question is why one should change or tune the number of classes when describing the spray characteristics. The underlying idea of increasing the number of classes is to obtain a greater detail of the probability distributions and detect eventual multimodality associated with clusters of data with dissimilar characteristics. In the case of drop sizes, an example generating such multimodality would be the presence of multiple atomization mechanisms (e.g., aerodynamic and/or hydrodynamic). In addition, in the case of the velocity, different two-phase flow events like the impact of droplets on solid surfaces contain information about the axial velocity component with positive values from those impinging on the surface, but the secondary droplets resulting after impact have a negative velocity component. Therefore, what is the criterion for choosing a given number of classes, k ? In addition, should the spacing of these classes be regular or irregular? Considering regularly spaced classes, one of the basic principles introduced by Sturges [1] stated that the number of classes k>log2(N) , with N as the total sample size. Doane [2] further elaborated on Sturge’s rule, but the problem is the over-smoothing of histrograms produced and its applicability Appl. Sci. 2020,10, 6122 3 of 18 limited to a number of data samples below 200, as analyzed by Hyndman [3] . The alternatives for the Sturge’s and Doane’s rules are the: •Scott’s rule for the bin width as δD=3.49sN−1/3, where sis the standard deviation [4]; • Freedman–Diaconis’ rule, also for the bin width as δD= 2 (IQR)N−1/3 , with IQR as the interquartile range [5]; •Rice’s rule for the number of classes is k=2N1/3 [6]; • and a rule based on J = 6 interlaced Fibonacci series with a number of classes defined as k=Jln(N)/ ln(1.618)[7]. Considering a simulated example of two clusters of droplets, each following a lognormal distribution function fLN(dg,γ) = 1 dγ√2πexp −ln(d/dg)2 2γ2!(1) with dg as the geometric diameter and γ as the geometric standard deviation, the final distribution function is a mixture between the two as f(d) = w1fLN( 40, 0.5 /√6)+( 1 −w1)fLN( 70, 0.5 /√6) with w1= 0.3. Figure 2shows the effect of using different criteria to organize drop size data with (a) N = 10 4 and (b) N= 105measurements in the form of probability density distributions of drop size. Figure 2. Example of the effect of changing the number of classes in details obtained on the probability density distribution of a mixture between two clusters described by distinct lognormal distributions, considering different sample sizes of (a)N=104and (b)N=105droplets. For the two sample sizes tested, the Sturges’ rule clearly over-smooths the distribution’s bimodality. The Scott’s rule generates less classes but is enough to capture the multimodality of the drop size distribution while producing the minimum number of empty classes. The Freedman and Diaconis proposal, and the interlaced Fibonacci series, can provide greater detail, but the results with the lower number of samples (Figure 2a) prove the cost of increasing the number of classes by a larger noise Appl. Sci. 2020,10, 6122 4 of 18 observed in probability density values. It is noteworthy that the best approach to represent drop size distributions for comparison purposes is a probability density ( pd fk [ µ m −1] ), since the amplitude of distributions with a different number of classes does not change as the amplitude of probability distributions (pk). The second approach to define classes in drop size statistics (less so in velocity) is the use of irregular bin widths. For example, in laser diffraction measurement systems, like Malvern’s Spraytec, the ratio (dupperBoundary −dlowerBoundary)/dk is constant, leading to broader classes for larger diameters representing each class. However, a systematic method for using irregular bin width to improve the description of drop size and velocity distributions is still open for further research. These details seldom appear reported in the literature when authors present the results of spray characterization and compare sprays obtained in different operating conditions. In this sense, the approach less prone to error would be to present the statistical data of the spray characteristics in terms of cumulative probability distributions, as explored later in this introduction. A final remark on the presentation and analysis of spray data in the form of statistical distributions is to consider the weight given to each class. In the case of the velocity of spray droplets, a number-weighted probability distribution is the most adequate. However, for histograms of drop size, other weighting factors, such as the •area-weighted pa,k=sk S with sk=πnkd2 kand S=∑sk; •and volume-weighted pv,k=vk V with vk= (π/6)nkd3 kand V=∑vk; applied to probability distributions improve the interpretation of its moments, as reported in the work of Sowa [8] . The reason for organizing spray data with other weight values besides the number-weighted case relates to the physics associated with the research question. Namely, if the investigation involves phenomena occurring at the droplet surface area, such as heat and mass transfer events, an area-weighted drop size distribution is more adequate. However, if the spray liquid volume is more important, such as spray cooling applications, the most adequate is the volume-weighted drop size distribution. The characterization of a spray often involves moments of the measured discrete probability distribution using single-point diagnostic techniques, like the Phase-Doppler Interferometry, or field diagnostic techniques using imaging. In the case of drop size, the calculation of each moment from drop size raw data corresponds to dab = ∑N i=1da i ∑N i=1db i!1 a−b ,∀a>b,{a,b}∈Z+(2) where di is a measurement of drop size in the sample acquired. If these moments use, instead, the number-weighted probability distribution values, the expression is dab =∑knk(dk)a ∑knk(dk)b1 a−b,∀a>b,{a,b}∈Z+(3) with nk as the number of measurements counted within the size class k where dk represents the mid-point in the interval between a lower and an upper bound. The characterization of the spray droplets using the average size obtained from a number-weighted probability distribution— d10 —means considering all droplets have, on average, the same size of d10 . For example, when comparing d10 for different locations in the spray, or the same location for different sprays, any increase implies more droplets of larger sizes. When analyzing the physics of droplet transport while interacting with the surrounding environment, the arithmetic Appl. Sci. 2020,10, 6122 5 of 18 mean diameter might be a valuable characteristic measure to consider. However, when the research aims at combustion applications, with evaporation and mass diffusion phenomena occurring at the surface area of each droplet, the best characteristic size is d32 , since it is the average of an area-weighted probability distribution—or else, if the research points to spray cooling applications, where the mass deposited on the surface is the relevant parameter due to its contribution to the formation of liquid films and their dynamic behavior, the best characteristic size for analyzing the heat and mass transfer involved would be d43, the average of a volume-weighted probability distribution. The main point when choosing the best way to present spray data is the awareness that each characteristic parameter, obtained statistically, has an underlying physical meaning, depending on the thermofluid phenomena involved. Finally, it is noteworthy that providing the moments of drop size and velocity distributions may limit the use of spray data in future works and spray simulations because of the inability to reconstruct the original distributions from the moments reported. Therefore, the next section discusses the implications for spray science of the attempt to fit a probability distribution functions to the histograms of drop size. 2. Drop Size Distribution Functions and Spray Science Spray characterization provides relevant information of droplets dynamics for the application of its mean quantities in empirical correlations related with heat transfer and fluid flow processes. The several optical diagnostic techniques acquire large amounts of data and the challenge is often how to process it. The criterion for the sample size ( N ) of spray data recurs often to the notion of statistical uncertainty. When its value is below a pre-defined threshold, the measurement stops because the experimentalist has enough data. However, this criterion often interprets spray data as a random probabilistic event, while spray statistics is more a method for organizing data. Therefore, the right question is not whether there is enough data to post-process and characterize a spray, but whether there is enough information. Section 2.1 reviews an approach based on information theory to make this assessment. Secondly, the meaning of using an average quantity to describe a spray, where droplets have multiple sizes, is to assume that whatever physical phenomenon affects an average size represents what occurs to the entire spray. The limitation of presenting the spray data based solely on mean quantities is the loss of information of the local or global drop polydispersed sizes or velocities in the spray and its potential usefulness in the development of numerical models that simulate sprays. However, while it is difficult to retrieve information of the original statistical distributions from their mean quantities, the ability to reconstruct drop size or velocity distributions from characteristic parameters of the mathematical probability distribution functions ( pd f ), instead of its moments, allows for obtaining the mean quantities without losing the information of the original distributions. Therefore, approaching spray characterization from the point of view of reconstructing probability distributions has significant advantages over the approach that uses solely moments retrieved from discrete probability distributions that describe a spray (e.g., see Panão and Radu [9] ). Section 2.2 reviews the fitting of probability distribution functions to histograms of drop data to retrieve the characteristic parameters allowing the reconstruction of spray data distributions, but introduces the argument of whether or not such fitting can provide some insight into the physics of the atomization process. The final Section 2.3 introduces, for the first time, the notion of drop size diversity, distinguishing the polydispersion degree from the heterogeneity degree and presenting the best parameters for their characterization. An accurate characterization of drop size diversity is relevant for the design of sprays. 2.1. Defining Enoughness in Large Data Samples An experimentalist stops a measurement based of criteria related to a statistical uncertainty. The definition of statistical uncertainty contains three conditions: 1. it is maximum for a uniform distribution where all classes have the same probability; 2. a small variation in the probability of a class generates a small variation in the uncertainty; Appl. Sci. 2020,10, 6122 6 of 18 3. and, finally, it depends on the distribution itself. The common measure used for the statistical uncertainty considers the standard deviation ( sx ) and the sample size (N) as ε=Zcsx √N(4) with Zc as the coefficient associated with the confidence interval considered (e.g., Zc= 1.96 for a 95% Confidence Interval). When the mean value ( x ) is different from zero, dividing ε by the mean provides the uncertainty in percentage. Since the standard deviation depends on the spray characteristics, reducing the statistical uncertainty implies adding more data to increase N . However, as argued by Panão [10] , if the spray begins to operate in a different way and the distribution changes, the statistical uncertainty as defined in Equation (4) continues to decrease, without providing any evidence to the experimentalist about the changes occurring in the spray. In a certain sense, it fails to comply to the third condition defining a statistical uncertainty because it is more sensitive to the sample size than the distribution itself. For this reason, an approach based on information theory is a better option. In information theory, the Shannon entropy complies to all the aforementioned characteristics of a statistical uncertainty. Considering the probability values of any discrete distribution ( pi ), the expression for calculating the Shannon entropy His H=−∑ i (piln(pi))(5) As an example, if we consider drop size, the minimum value corresponds to a monosize spray with all droplets having the same size, thus, p= 1, and H= 0. The maximum value occurs for a hypothetical spray where all droplets have the same probability of being present, corresponding to a uniform distribution with k classes, each for a different drop size. Therefore, the maximum Shannon entropy is max(H) = 1 /k . In sprays, while measuring size and velocity, considering the Shannon entropy normalized by its maximum value, Hn=H/max(H) , it tends to stabilize (see Panão [10] for details). The meaning is that adding more data does not mean adding more information because the shape and scale of the distribution stabilized. However, as argued in Panão [10] , the enoughness requires a criterion with interpretative value and proposes the excess entropy ( EE ) because, as defined in Feldman et al. [ 11 ], it is what best captures the nature of convergence of the entropy rate as the amount of memory gained, or the cost of amnesia if all data would suddenly be lost. To evaluate the evolution of EE while measuring, • one calculates the entropy rate that quantifies the difference between the normalized Shannon entropy with adding one value the N samples and the Shannon entropy with N samples- ˙ Hn= |Hn(N+1)−Hn(N)|; •considers the limit when N→∞, which is ˙ h; •and the excess entropy EE is formulated as EE =∑∞ N=1˙ Hn(N)−˙ h In the case of spray data, the stabilization implies ˙ h = 0, thus, it simplifies EE . The method proposed is setting a convergence criterion– εEE and stop measuring when |EE(N)−median(EE)|<εEE , considering the number of samples corresponding to the median as the minimum required (see Panão [10] for more details). Once the experimentalist has enough information and organizes the spray data with histograms of drop velocity and weighted distributions of drop sizes, depending on the physical process under analysis, one of the methods to avoid losing information of the measured distributions is fitting data to known mathematical or empirical distribution functions. However, the question is whether such fitting can provide further insight into liquid atomization mechanisms. This is the topic explored in the next sub-section. Appl. Sci. 2020,10, 6122 7 of 18 2.2. Underlying Physics of Probability Distribution Functions Applied to Sprays An important consideration when characterizing sprays is the effect of the interaction between droplets and the continuous phase on the local (or even overall) size and velocity distributions. This interaction involves different transport phenomena, with momentum and energy exchanges between the dispersed phase (spray) and the carrier phase (surrounding environment), generating eventual secondary breakup of the spray droplets leading to changes in the shape and scale of drop size distributions, or acceleration (positive or negative) captured by changes in local velocity probability distributions. However, the physical reason why a certain distribution function might fit better than another is still open for further research. This work advances an argument in favor of distinguishing between modeling and characterizing drop size distributions. The purpose of modeling drop size distributions is to predict them from the information of the atomizer geometry and operating conditions. The simulation of sprays using a numerical approach [ 12 ], or a statistical or stochastic approach [ 13 ] can produce data on the droplets characteristics, but it is different from statistically processing such data. Therefore, according to Déchelette et al. [14], there are four methods for modeling drop size distributions: •the empirical; •the Maximum Entropy Formalism (MEF); •the Discrete Probability Function (DPF); •and the Stochastic. However, the purpose of characterizing a spray is to describe, as accurately as possible, the polydispersion of sizes and velocities of its droplets. This description aims at obtaining mean quantities for analyzing heat transfer and flow processes—or its aim is to improve our understanding of the nature underlying the atomization mechanisms. There are two categories of probability distribution functions used to describe droplets’ characteristics: mathematical and empirical. Lefebvre and McDonell [15] provide a synthesis of the main probability distributions in each category. Except for the Rosin–Rammler or Weibull, the Nukiyama–Tanasawa and Upper-Limit empirical distribution functions are complex and problems arise when determining the best-fit values for their parameters. As to the mathematical distribution functions, the simplest is the Log-Normal, while the Log-Hyperbolic is also complex and problems arise with finding the best fitting parameters. One distribution absent from Lefebvre and McDonell [15] and other review works is the Gamma distribution function which Villermaux et al. [ 16 ] associated with the distribution of droplets resulting from the fragmentation of ligaments, generating a spray, reviewed later in this section. While most research on spray characterization focuses on the fitting process, few works such as Villermaux [17] and Villermaux et al. [ 16 ] address the meaning of the mathematical distribution function used and the physical background for such fitting. As mentioned in the Introduction, instead of focusing our attention on probability or probability density functions, we propose a greater focus on cumulative distribution functions ( F(d) ). Therefore, any comparison between different F(d) becomes universal and if such distribution properly describes the local or global experimental results, its digitization to simulate a spray is relatively accessible. In the case of the Log-Normal distribution, its cumulative form given by FLN(d,dg,γ) = 1 21+erf ln(d/dg) √2γ (6) includes dg as the scale parameter, the geometric mean diameter, and γ as the shape parameter of the distribution. Applied to spray characterization, the reason for using a Log-Normal distribution function is related to the multiplicative effect of subsequent stages of droplets breaking up during atomization, as a cascade process, where one drop breaks into two or more and so on. However, if we consider the Appl. Sci. 2020,10, 6122 8 of 18 interaction between droplets and the continuous gaseous phase, in time, the dragging of droplets leads to secondary flows which eventually produce a vortical effect on the transport of subsequent droplets. Namely, smaller droplets, with lower response times, dragged by secondary flows, may return to upward locations, redistributing the counts of certain drop size classes in locations further downstream of the spray trajectory. Therefore, even if it is not the result of a multiplicative breakup process, the presence of these smaller drops affects the probability distributions describing the spray, having an effect similar to that of a cascade of multiple breakup stages. One could even speculate whether or not the reason for a Log-Normal distribution best fitting experimental results expresses the way the multiphase flow organizes the transport of droplets according to their size in a cascade pattern. Besides several breakup stages and transport phenomena, some atomization processes result from the disintegration of liquid sheets or jets into ligaments, and those ligaments further fragmenting into droplets. In this case, Villermaux et al. [ 16 ] argues each ligament constituted of several blobs, and when it fragments into several droplets, the size distribution that reasonably fits is a Gamma probability distribution function, which cumulative form is expressed by FΓ(d,a,b) = 1 baΓ(a)Zd 0xa−1exp(−x/b)dx (7) with a and b as the shape and scale parameter, respectively. In this case, a characteristic size corresponds the product of both: dc=a·b. Considering empirical distribution functions, this work focuses on the Weibull distribution, first applied to describe the distribution of drop sizes by Rosin and Rammler, FWB(d,dc,q) = 1−exp −d dcq(8) where dc is a scale parameter related to a characteristic drop size and q is the shape parameter and considered a measure of the spreading in drop sizes. The accuracy of this empirical probability distribution is best related to drop size distributions with fewer smaller droplets or narrower size distributions [15]. A final note considers the Nukiyama–Tanasawa empirical distribution that is often used to fit experimental data. Here, particular attention is given to the work of Li and Tankin [18] that derived the expression using an information-theory approach, and for the spray liquid volume, its cumulative form results in FNT(d,dc,q) = 1−(1+qd3)exp −qd3(9) where qis also the shape parameter. The average quantities referred to so far are related to moments in probability distributions, but considering the cumulative distribution, any quantity corresponds to a representative diameter, generally expressed as Dxw , where x is the type of distribution, and w is the percent cumulative value related to the representative diameter. Therefore, if the cumulative distribution is •number-based, Dnw represents the size containing w% of the droplets in the spray; •area-based, Daw represents the size containing w% of the spray surface area; •volume-based, Dvw represents the size containing w% of the liquid spray volume. One of the most relevant representative diameters corresponds to 50% ( Dn0.5 , Da0.5 , or Dv0.5 ) because dividing the classes by that number allows a better comparison between cumulative distributions, useful for analyzing the effect of parametric variations in the spray. 2.3. Introducing Drop Size Diversity in a Spray Drop size distributions are a way to organize the data acquired to characterize a spray. The characterization of the diversity of drop sizes answers two distinct questions: (1) how many Appl. Sci. 2020,10, 6122 9 of 18 different sizes are relevant in a spray; (2) and how different are the relevant sizes in a spray. The word relevant links to the probability of presence of certain drop size classes relative to others. In the known textbook on Atomization and Sprays, Lefebvre and McDonell [15] refer to this diversity as drop spray dispersion associated with the size range of droplets. On the other hand, several research articles address the different sizes of the spray droplets as a polydispersed spray. In the authors’ opinion, these are two different things, which is why we introduce the concept of Drop Size Diversity (DSD) measured by two different degrees: •the polydispersion degree to quantify the multitude of different sizes that are relevant in a spray. Thus, the maximum for the case where all different sizes have the same probability of being present in the spray, and •the heterogeneity degree to quantify how different are the relevant sizes in the spray. Therefore, it is related with the size range or size dispersion. The challenge is to devise the right indicators to measure both degrees. Among the several indicators available and synthesized in Lefebvre and McDonell [15] , the most known and used indicator is the Relative Span obtained from the representative diameters of a volume-based cumulative size distribution (DvX with 0 <X<1 as the fraction of the spray liquid volume) as ∆v=Dv0.9 −Dv0.1 Dv0.5 (10) Considering the normalization of a representative diameter by the value representing half of the liquid volumeDv0.5 -as D∗ vX =DvX/Dv0.5 , the interpretation of Equation (10) relative to the range of drop sizes corresponds to a difference- ∆v=D∗ v0.9 −D∗ v0.1 which is equal to zero when all droplets in the spray have the same size, and maximum when all droplets have the same probability of occurrence (a limit unrealistic case). Panão [19] proposed a different approach based on information theory, through the concept of the normalized Shannon entropy, already defined in Section 2.1 as Hn=H ln(Nbins)(11) In the information theory terminology applied to spray characterization, a spray where all droplets have the same size, p= 1, resulting in a null normalized Shannon entropy, Hn= 0, while an unrealistic spray with all classes having the same probability of presence (uniform distribution), Hn=1, because the Shannon entropy–numerator in Equation (11)–is maximum. Finally, García et al. [ 20 ] proposed a third approach based on the standard deviation of the volume-weighted drop size distribution expressed as SDv=qd2 53 −d2 43 (12) where d53 and d43 are the secondand first-order moments of the volume-weighted drop size distribution, respectively. The authors compared this standard deviation with Shannon entropy and found inconsistencies. They state that the Shannon entropy has a main drawback, since the information of the drop sizes representing each class is not explicitly included; thus, if these probability values would be randomly rearranged, the H value would be the same. This is an important insight because it allows to understand the difference between polydispersion and size dispersion in a spray. To compare the three approaches, consider the simulation of a spray mixing two monosize droplet streams of 10 µ m and 20 µ m. A weight parameter w , varying between 0 and 1, sets the percentage of drops present in the mixed spray from each of the monosize sources. Therefore, if w=0, all droplets have the size of 20 µ m, and if w= 1, all droplets have 10 µ m. Figure 3shows the result for the Relative Span ( ∆v ), normalized Shannon entropy based on the volume-weight drop size distribution ( Hn,v ), Appl. Sci. 2020,10, 6122 16 of 18 Finally, in this work, we introduced the concept of drop size diversity to better understand the many different sizes relevant in a spray, as well as how different the relevant sizes in a spray are. The polydispersion degree given by the normalized Shannon entropy, Hn , and spray heterogeneity degree given by the volume-weighted standard deviation, SDv [ µ m], can describe this diversity, as depicted in Figure 10. Figure 10. Polydispersion degree given by the normalized Shannon entropy, Hn , and spray heterogeneity degree given by the volume-weighted standard deviation, SDv [ µ m], for the planes of Z=10 mm (top) and Z=20 mm (bottom). In the case of distilled water, the results show a higher polydispersion degree around r=8 mm , where the spray heterogeneity degree also becomes higher. However, despite local changes in Hn , the behavior does not seem significantly affected by adding surfactant to produce the base fluid, and further adding nanoparticles. However, in terms of heterogeneity, the results are different. The addition of surfactant exerts a major effect on the heterogeneity of the spray characteristics. However, the addition of nanoparticles produces a negligibly effect. Consequently, one may also expect a minor effect of the nanoparticles during spray impact. This can actually be a beneficial feature, for instance, in spray cooling applications, as it suggests that the nanoparticles can be used to alter the thermal properties of the working fluids, without significantly affecting the main spray characteristics. 4. Conclusions The characterization of a spray is not merely acquiring information on the size and velocity of its droplets in several locations from single-point measurement techniques, or in several planes from 2Dor 3D-measurement techniques. Although it is essential to acquire enough information, as explored here through an information theory approach, the clarity of the research question associated with the spray characterization is very much like the clarity in the display and analysis of data. Therefore, a well-defined research question is what guides the kind of spray characterization performed. In this article, we review the statistical language used in spray characterization, and: • the differences between organizing spray data using probability histograms or histograms of probability density; •how to choose the number of classes in histrograms; • the different kinds of probability histograms considering the number of droplets, their area, or the liquid volume, and corresponding moments; • a method based on the excess entropy from information theory to assess if there is enough information for post-processing; Appl. Sci. 2020,10, 6122 17 of 18 • the physical meaning of fitting mathematical probability distribution functions, namely the Log-Normal, Gamma, and Weibull, to spray data; • and introduce, for the first time, the notion of Drop Size Diversity with its polydispersion and heterogeneity degrees quantified by the normalized Shannon entropy and volume-weighted standard deviation, respectively. Finally, the topics explored on spray characterization are applied to nanofluid sprays to explore the effect of introducing a surfactant and nanoparticles on the spray structure and dynamic characteristics. Author Contributions: Conceptualization, investigation, M.O.P.; resources, A.S.M. and A.L.M.; writing–original draft preparation, M.O.P; writing–review and editing, M.O.P and A.S.M.; project administration, A.L.M.; funding acquisition, A.S.M. and A.L.M. All authors have read and agreed to the published version of the manuscript. Funding: Ana S. Moita would like to acknowledge project No. 030171 funded by LISBOA-01-0145-FEDER030171/PTDC/EME-SIS/30171/2017, and project UTAP-EXPL/CTE/0064/2017. Conflicts of Interest: The authors declare no conflict of interest. Abbreviations The following abbreviations are used in this manuscript: CTAB CetylTrimethylAmmonium Bromide DSD Drop Size Diversity PDI Phase-Doppler Interferometer References 1. Sturges, H.A. The choice of a class interval. J. Am. Stat. Assoc. 1926,21, 65–66. [CrossRef] 2. Doane, D.P. Aesthetic frequency classifications. Am. Stat. 1976,30, 181–183. 3. Hyndman, R.J. The Problem with Sturges Rule for Constructing Histograms; Monash University: Melbourne, Australia, 1995. 4. Scott, D.W. On optimal and data-based histograms. Biometrika 1979,66, 605–610. [CrossRef] 5. Freedman, D.; Diaconis, P. On the histogram as a density estimator: L2 theory. Z. Wahrscheinlichkeit 1981 ,57, 453–476. [CrossRef] 6. Terrell, G.R.; Scott, D.W. Oversmoothed nonparametric density estimates. J. Am. Stat. Assoc. 1985 ,80, 209–214. [CrossRef] 7. Panão, M.; Moreira, A. A real-time assessment of measurement uncertainty in the experimental characterization of sprays. Meas. Sci. Technol. 2008,19, 095402. [CrossRef] 8. Sowa, W. Interpreting mean drop diameters using distribution moments. At. Sprays 1992 ,2, 1–15. [CrossRef] 9. Panão, M.R.; Radu, L. Advanced statistics to improve the physical interpretation of atomization processes. Int. J. Heat Fluid Flow 2013,40, 151–164. [CrossRef] 10. Panão, M.R.O. Assessment of measurement efficiency in laser-and phase-Doppler techniques: An information theory approach. Meas. Sci. Technol. 2012,23, 125304. [CrossRef] 11. Feldman, D.P.; McTague, C.S.; Crutchfield, J.P. The organization of intrinsic computation: Complexity-entropy diagrams and the diversity of natural information processing. Chaos: Interdiscip. J. Nonlinear Sci. 2008,18, 043106. [CrossRef] [PubMed] 12. Naruemon, I.; Liu, L.; Liu, D.; Ma, X.; Nishida, K. An Analysis on the Effects of the Fuel Injection Rate Shape of the Diesel Spray Mixing Process Using a Numerical Simulation. Appl. Sci. 2020,10, 4983. [CrossRef] 13. Archambault, M.R. A discussion on the stochastic modeling of spray flows. At. Sprays 2009 ,19, 1171–1191. [CrossRef] 14. Déchelette, A.; Babinsky, E.; Sojka, P. Drop size distributions. In Handbook of Atomization and Sprays; Springer: Berlin, Germany, 2011; pp. 479–495. 15. Lefebvre, A.H.; McDonell, V.G. Atomization and Sprays; CRC Press: Boca Raton, FL, USA, 2017. 16. Villermaux, E.; Marmottant, P.; Duplat, J. Ligament-mediated spray formation. Phys. Rev. Lett. 2004,92, 074501. [CrossRef] [PubMed] 17. Villermaux, E. Fragmentation. Annu. Rev. Fluid Mech. 2007,39, 419–446. [CrossRef] Appl. Sci. 2020,10, 6122 18 of 18 18. Li, X.; Tankin, R.S. Droplet size distribution: A derivation of a Nukiyama-Tanasawa type distribution function. Combust. Sci. Technol. 1987,56, 65–76. 19. Panão, M.R.O. Redefining spray uniformity through an information theory approach. At. Sprays 2016 ,26, 1069–1081. [CrossRef] 20. García, J.; Lozano, A.; Alconchel, J.; Calvo, E.; Barreras, F.; Santolaya, J. Atomization of glycerin with a twin-fluid swirl nozzle. Int. J. Multiph. Flow 2017,92, 150–160. [CrossRef] 21. Mal ` y, M.; Moita, A.; Jedelsky, J.; Ribeiro, A.; Moreira, A. Effect of nanoparticles concentration on the characteristics of nanofluid sprays for cooling applications. J. Therm. Anal. Calorim. 2019 ,135, 3375–3386. [CrossRef] c 2020 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (http://creativecommons.org/licenses/by/4.0/).