<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.4 20241031//EN" "JATS-journalpublishing1-4.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article" dtd-version="1.4" xml:lang="en">
  <front>
    <journal-meta>
      <journal-id journal-id-type="publisher-id">ojs</journal-id>
      <journal-title-group>
        <journal-title>Open Journal of Statistics</journal-title>
      </journal-title-group>
      <issn pub-type="epub">2161-7198</issn>
      <issn pub-type="ppub">2161-718X</issn>
      <publisher>
        <publisher-name>Scientific Research Publishing</publisher-name>
      </publisher>
    </journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.4236/ojs.2026.164013</article-id>
      <article-id pub-id-type="publisher-id">ojs-153031</article-id>
      <article-categories>
        <subj-group>
          <subject>Article</subject>
        </subj-group>
        <subj-group>
          <subject>Physics</subject>
          <subject>Mathematics</subject>
        </subj-group>
      </article-categories>
      <title-group>
        <article-title>Extensions of the Mean Difference for the Lognormal Distribution</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <name name-style="western">
            <surname>Manca</surname>
            <given-names>Fabio</given-names>
          </name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <name name-style="western">
            <surname>Vacca</surname>
            <given-names>Angelo</given-names>
          </name>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <contrib contrib-type="author">
          <name name-style="western">
            <surname>Marin</surname>
            <given-names>Claudia</given-names>
          </name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <name name-style="western">
            <surname>Valerio</surname>
            <given-names>Angelo</given-names>
          </name>
          <xref ref-type="aff" rid="aff3">3</xref>
        </contrib>
        <contrib contrib-type="author">
          <name name-style="western">
            <surname>Sabella</surname>
            <given-names>Elita Anna</given-names>
          </name>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
      </contrib-group>
      <aff id="aff1"><label>1</label> Department of Education, Psychology, Communication Sciences, University of Bari Aldo Moro, Bari, Italy </aff>
      <aff id="aff2"><label>2</label> Department of Precision and Regenerative Medicine and Ionian Area, University of Bari Aldo Moro, Bari, Italy </aff>
      <aff id="aff3"><label>3</label> Research and Data Consultant, Bari, Italy </aff>
      <author-notes>
        <fn fn-type="conflict" id="fn-conflict">
          <p>The authors declare no conflicts of interest regarding the publication of this paper.</p>
        </fn>
      </author-notes>
      <pub-date pub-type="epub">
        <day>03</day>
        <month>08</month>
        <year>2026</year>
      </pub-date>
      <pub-date pub-type="collection">
        <month>08</month>
        <year>2026</year>
      </pub-date>
      <volume>16</volume>
      <issue>04</issue>
      <fpage>284</fpage>
      <lpage>298</lpage>
      <history>
        <date date-type="received">
          <day>21</day>
          <month>04</month>
          <year>2026</year>
        </date>
        <date date-type="accepted">
          <day>02</day>
          <month>08</month>
          <year>2026</year>
        </date>
        <date date-type="published">
          <day>05</day>
          <month>08</month>
          <year>2026</year>
        </date>
      </history>
      <permissions>
        <copyright-statement>© 2026 by the authors and Scientific Research Publishing Inc.</copyright-statement>
        <copyright-year>2026</copyright-year>
        <license license-type="open-access">
          <license-p> This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license ( <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link> ). </license-p>
        </license>
      </permissions>
      <self-uri content-type="doi" xlink:href="https://doi.org/10.4236/ojs.2026.164013">https://doi.org/10.4236/ojs.2026.164013</self-uri>
      <abstract>
        <p>This paper extends the closed-form formula of Gini’s mean difference for the lognormal distribution, originally obtained by Girone and Manca (2016). The following are analyzed: 1) the asymptotic behavior of the scale parameter for <italic>γ</italic> → 0 (degeneration of the distribution) and for <italic>γ</italic> → +∞ (unbounded dispersion); 2) the generalized formula with complete location parameter <italic>μ</italic> and scale parameter <italic>σ</italic>; 3) the mean difference conditioned on an interval; 4) the mean difference for the truncated lognormal. Applications in the field of medical sciences are also discussed, where the lognormal distribution is ubiquitous in the modeling of biomarkers and pharmacological concentrations. The results show that the original formula <inline-formula><mml:math display="inline"></mml:math></inline-formula></p>
        <p>Δ=2</p>
        <p>e</p>
        <p>γ</p>
        <p>2</p>
        <p>/2</p>
        <p>erf(</p>
        <p>γ/2</p>
        <p>)</p>
        <p>for the standardized case admits natural extensions that significantly broaden its field of application, while preserving the analytical elegance of the basic formulation.</p>
      </abstract>
      <kwd-group kwd-group-type="author-generated" xml:lang="en">
        <kwd>Mean Difference</kwd>
        <kwd>Lognormal Distribution</kwd>
        <kwd>Truncated Distribution</kwd>
        <kwd>Gini Index</kwd>
        <kwd>Medical Sciences</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec1">
      <title>1. Introduction</title>
      <p>Gini’s mean difference, introduced by Corrado Gini in 1912 [<xref ref-type="bibr" rid="B1">1</xref>], represents one of the most important measures of variability in statistical theory and applied practice. For a continuous random variable <italic>X</italic> with density function <italic>f</italic>(<italic>x</italic>) and cumulative distribution function <italic>F</italic>(<italic>x</italic>), defined on a support (<italic>a</italic>, <italic>b</italic>), the mean difference is defined as the expected value of the absolute difference between two independent observations <italic>X</italic><sub>1</sub> and <italic>X</italic><sub>2</sub> drawn from the same distribution: Δ = <italic>E</italic>[|<italic>X</italic><sub>1</sub> − <italic>X</italic><sub>2</sub>|].</p>
      <p>The mean difference possesses fundamental properties that make it particularly attractive as a measure of variability: it is translation-invariant (independent of the location parameter), it is homogeneous of degree one with respect to the scale parameter, and it is directly related to the Gini concentration index through the relation <italic>G</italic> = Δ/(2<italic>E</italic>[<italic>X</italic>]), where <italic>E</italic>[<italic>X</italic>] denotes the mean of the distribution. These properties, combined with its intuitive interpretability and robustness with respect to extreme values, make it an instrument of fundamental importance in fields ranging from economics to biometry, from epidemiology to reliability theory.</p>
      <p>In addition to the variance, several alternative measures of variability and concentration have been proposed in the statistical and economic literature. Among them, Gini’s mean difference occupies a central role, as it combines properties of dispersion and inequality measurement. Originally introduced by Corrado Gini, this measure has been shown to provide a more informative description of non-normal and skewed distributions compared to variance [<xref ref-type="bibr" rid="B2">2</xref>].</p>
      <p>The Gini framework has been extensively developed in both theoretical and applied directions, particularly in the context of inequality measurement, where it is closely related to the Lorenz curve and concentration indices [<xref ref-type="bibr" rid="B3">3</xref>]. More recent contributions highlight the need for complementary measures capable of capturing different aspects of distributional heterogeneity [<xref ref-type="bibr" rid="B4">4</xref>].</p>
      <p>Recent contributions have further advanced the study of the lognormal distribution and the mean difference in both theoretical and applied settings. Novi Inverardi and Tagliani [<xref ref-type="bibr" rid="B5">5</xref>] demonstrate that the lognormal distribution is uniquely characterized by its integer moments via maximum entropy. Dai and Shen [<xref ref-type="bibr" rid="B6">6</xref>] propose a robust two-quantile method for Gini coefficient estimation. Poudyal, Zhao, and Brazauskas [<xref ref-type="bibr" rid="B7">7</xref>] develop winsorized moment estimators for truncated and censored lognormal distributions. Kapera and Kobus [<xref ref-type="bibr" rid="B8">8</xref>] extend the Gini and mean log deviation indices to multivariate inequality of opportunity. Jokiel-Rokita and Piątek [<xref ref-type="bibr" rid="B9">9</xref>] construct estimators for Weibull parameters based on the nonparametric Gini coefficient. The present work contributes to this growing literature by providing analytical extensions of the lognormal mean difference formula.</p>
      <p>In a series of previous studies, Girone and collaborators [<xref ref-type="bibr" rid="B10">10</xref>]-[<xref ref-type="bibr" rid="B12">12</xref>] obtained closed-form formulas for the mean difference for numerous continuous and discrete distributions, including the exponential, gamma, beta, Pareto, and geometric distributions. In particular, Girone and Manca [<xref ref-type="bibr" rid="B13">13</xref>] derived, through a procedure employing selected integrals of the error function, the remarkably simple formula for the mean difference of the standardized lognormal distribution (with location parameter <italic>θ</italic> = 0):</p>
      <disp-formula id="FD1">
        <label>(1)</label>
        <mml:math display="inline">
          <mml:mrow>
            <mml:mi>Δ</mml:mi>
            <mml:mo>=</mml:mo>
            <mml:mn>2</mml:mn>
            <mml:msup>
              <mml:mtext>e</mml:mtext>
              <mml:mrow>
                <mml:mrow>
                  <mml:mrow>
                    <mml:msup>
                      <mml:mi>γ</mml:mi>
                      <mml:mn>2</mml:mn>
                    </mml:msup>
                  </mml:mrow>
                  <mml:mo>/</mml:mo>
                  <mml:mn>2</mml:mn>
                </mml:mrow>
              </mml:mrow>
            </mml:msup>
            <mml:mtext>erf</mml:mtext>
            <mml:mrow>
              <mml:mo>(</mml:mo>
              <mml:mrow>
                <mml:mfrac>
                  <mml:mi>γ</mml:mi>
                  <mml:mn>2</mml:mn>
                </mml:mfrac>
              </mml:mrow>
              <mml:mo>)</mml:mo>
            </mml:mrow>
          </mml:mrow>
        </mml:math>
      </disp-formula>
      <p>where <italic>γ</italic> &gt; 0 is the scale parameter (standard deviation of the logarithm of the variable)—a quantity that, in the generalized formulation of Section 4, will be equivalently denoted by <italic>σ</italic>, both symbols representing the standard deviation of log<italic>X</italic> in their respective parametric contexts—and erf(·) denotes the error function, defined as:</p>
      <disp-formula id="FD2">
        <label>(2)</label>
        <mml:math>
          <mml:mrow>
            <mml:mtext>erf</mml:mtext>
            <mml:mrow>
              <mml:mo>(</mml:mo>
              <mml:mi>x</mml:mi>
              <mml:mo>)</mml:mo>
            </mml:mrow>
            <mml:mo>=</mml:mo>
            <mml:mfrac>
              <mml:mn>2</mml:mn>
              <mml:mi>π</mml:mi>
            </mml:mfrac>
            <mml:mstyle displaystyle="true">
              <mml:mrow>
                <mml:msubsup>
                  <mml:mo>∫</mml:mo>
                  <mml:mn>0</mml:mn>
                  <mml:mi>x</mml:mi>
                </mml:msubsup>
                <mml:mrow>
                  <mml:msup>
                    <mml:mtext>e</mml:mtext>
                    <mml:mrow>
                      <mml:mo>−</mml:mo>
                      <mml:msup>
                        <mml:mi>t</mml:mi>
                        <mml:mn>2</mml:mn>
                      </mml:msup>
                    </mml:mrow>
                  </mml:msup>
                  <mml:mtext>d</mml:mtext>
                  <mml:mi>t</mml:mi>
                </mml:mrow>
              </mml:mrow>
            </mml:mstyle>
            <mml:mo>.</mml:mo>
          </mml:mrow>
        </mml:math>
      </disp-formula>
      <p>This formula was obtained starting from the representation of the mean difference based exclusively on the cumulative distribution function, namely:</p>
      <disp-formula id="FD3">
        <label>(3)</label>
        <mml:math>
          <mml:mrow>
            <mml:mi>Δ</mml:mi>
            <mml:mo>=</mml:mo>
            <mml:mn>2</mml:mn>
            <mml:mstyle displaystyle="true">
              <mml:mrow>
                <mml:msubsup>
                  <mml:mo>∫</mml:mo>
                  <mml:mi>a</mml:mi>
                  <mml:mi>b</mml:mi>
                </mml:msubsup>
                <mml:mrow>
                  <mml:mi>F</mml:mi>
                  <mml:mrow>
                    <mml:mo>(</mml:mo>
                    <mml:mi>x</mml:mi>
                    <mml:mo>)</mml:mo>
                  </mml:mrow>
                  <mml:mrow>
                    <mml:mo>[</mml:mo>
                    <mml:mrow>
                      <mml:mn>1</mml:mn>
                      <mml:mo>−</mml:mo>
                      <mml:mi>F</mml:mi>
                      <mml:mrow>
                        <mml:mo>(</mml:mo>
                        <mml:mi>x</mml:mi>
                        <mml:mo>)</mml:mo>
                      </mml:mrow>
                    </mml:mrow>
                    <mml:mo>]</mml:mo>
                  </mml:mrow>
                  <mml:mtext>d</mml:mtext>
                  <mml:mi>x</mml:mi>
                </mml:mrow>
              </mml:mrow>
            </mml:mstyle>
          </mml:mrow>
        </mml:math>
      </disp-formula>
      <p>and through a procedure involving six selected integrals of the error function, some of which are particularly complex and were computed by resorting to an integral given by Prudnikov, Brychkov and Marichev [<xref ref-type="bibr" rid="B14">14</xref>], which made it possible to reduce apparently intractable expressions to a formulation of remarkable simplicity and elegance<sup>1</sup>.</p>
      <p>The present work aims to extend this fundamental result in five complementary directions: 1) the rigorous analysis of the asymptotic behavior for <italic>γ</italic> → 0 and <italic>γ</italic> → +∞, which provides a complete understanding of the mean difference under limiting conditions; 2) the generalization of the formula to the complete parameters <italic>μ</italic> (location) and <italic>σ</italic> (scale); 3) the formulation of the mean difference conditioned on an interval [<italic>a</italic>, <italic>b</italic>]; 4) the derivation of the mean difference for the truncated lognormal distribution; and 5) the discussion of applications in the field of medical sciences, where the lognormal distribution is omnipresent. Remark on notation. Throughout this paper, the following conventions are adopted to ensure notational clarity: 1) The symbol <italic>γ</italic> denotes the scale parameter (standard deviation of log<italic>X</italic>) in the standardized formulation (Sections 1 - 3), while <italic>σ</italic> denotes the identical quantity in the general formulation (Sections 4 - 5); when <italic>μ</italic> = 0, we have <italic>σ</italic> ≡ <italic>γ</italic>. 2) The symbol <italic>μ</italic> is reserved exclusively for the location parameter of the underlying normal distribution (the mean of log<italic>X</italic>); the mean of the lognormal variable <italic>X</italic> is always denoted <italic>E</italic>[<italic>X</italic>] = exp(<italic>μ</italic> + <italic>σ</italic><sup>2</sup>/2) to avoid ambiguity. 3) Context-specific subscripts are used for the mean difference where needed: Δ without subscript in Sections 1 - 3 refers to the standardized case (Formula (1)), Δ(<italic>μ</italic>, <italic>σ</italic>) to the general case (Formula (10)), Δ<italic><sub>c</sub></italic>[<italic>a</italic>, <italic>b</italic>] to the conditional mean difference on an interval (Formula (12)), and Δ<italic><sub>T</sub></italic> to the truncated case (Formula (13)). The unqualified symbol Δ refers to the mean difference in its generic, context-dependent sense.</p>
      <p>The remainder of this paper is organized as follows. Section 2 introduces the notation, definitions, and preliminary results used throughout the paper, including the formal definition of the lognormal distribution and the error function. Section 3 analyzes the asymptotic behavior of the mean difference for extreme values of the scale parameter. Section 4 derives the generalized formula with complete parameters. Section 5 develops the conditional mean difference on an interval. Section 6 presents the mean difference for the truncated lognormal distribution. Section 7 discusses applications in the medical sciences. Finally, Section 8 presents conclusions, and Section 9 outlines limitations and directions for future research.</p>
    </sec>
    <sec id="sec2">
      <title>2. Notation and Definitions</title>
      <p>This section collects the basic definitions, notation, and preliminary results that are used throughout the paper. The purpose is to establish a consistent notational framework and to make the paper self-contained.</p>
      <p><italic><bold>Definition</bold></italic><italic><bold>1</bold></italic><bold>(</bold><italic><bold>Lognormal</bold></italic><italic><bold>Distribution</bold></italic><bold>).</bold></p>
      <p>A random variable <italic>X</italic> is said to follow a lognormal distribution with parameters <italic>μ</italic> and <italic>σ</italic><sup>2</sup>, written <italic>X</italic> ~ Lognormal(<italic>μ</italic>, <italic>σ</italic><sup>2</sup>), if <italic>Y</italic> = log<italic>X</italic> follows a normal distribution <italic>N</italic>(<italic>μ</italic>, <italic>σ</italic><sup>2</sup>). Equivalently, <italic>X</italic> = exp(<italic>Y</italic>) where <italic>Y</italic> ~ <italic>N</italic>(<italic>μ</italic>, <italic>σ</italic><sup>2</sup>). The probability density function of <italic>X</italic> is: <inline-formula><mml:math display="inline"><mml:mrow><mml:mi> f </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mi> x </mml:mi><mml:mo> ; </mml:mo><mml:mi> μ </mml:mi><mml:mo> , </mml:mo><mml:mi> σ </mml:mi></mml:mrow><mml:mo> ) </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mrow><mml:mn> 1 </mml:mn><mml:mo> / </mml:mo><mml:mrow><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mi> x </mml:mi><mml:mi> σ </mml:mi><mml:msqrt><mml:mrow><mml:mn> 2 </mml:mn><mml:mi> π </mml:mi></mml:mrow></mml:msqrt></mml:mrow><mml:mo> ) </mml:mo></mml:mrow></mml:mrow></mml:mrow></mml:mrow><mml:mo> ) </mml:mo></mml:mrow><mml:mi> exp </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mo> − </mml:mo><mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mi> log </mml:mi><mml:mi> x </mml:mi><mml:mo> − </mml:mo><mml:mi> μ </mml:mi></mml:mrow><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> / </mml:mo><mml:mrow><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mn> 2 </mml:mn><mml:msup><mml:mi> σ </mml:mi><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> ) </mml:mo></mml:mrow></mml:mrow></mml:mrow></mml:mrow><mml:mo> ) </mml:mo></mml:mrow></mml:mrow></mml:math></inline-formula> , for <italic>x</italic> &gt; 0, and the cumulative distribution function is: <inline-formula><mml:math display="inline"><mml:mrow><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mi> x </mml:mi><mml:mo> ; </mml:mo><mml:mi> μ </mml:mi><mml:mo> , </mml:mo><mml:mi> σ </mml:mi></mml:mrow><mml:mo> ) </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:mi> Φ </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mrow><mml:mrow><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mi> log </mml:mi><mml:mi> x </mml:mi><mml:mo> − </mml:mo><mml:mi> μ </mml:mi></mml:mrow><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mo> / </mml:mo><mml:mi> σ </mml:mi></mml:mrow></mml:mrow><mml:mo> ) </mml:mo></mml:mrow></mml:mrow></mml:math></inline-formula> , where Φ(·) denotes the CDF of the standard normal distribution. The parameters <italic>μ</italic> and <italic>σ</italic> are, respectively, the mean and standard deviation of the underlying normal variable <italic>Y</italic> = log<italic>X</italic> (not of <italic>X</italic> itself). The mean and variance of <italic>X</italic> are: <italic>E</italic>[<italic>X</italic>] = exp(<italic>μ</italic> + <italic>σ</italic><sup>2</sup>/2) and Var(<italic>X</italic>) = [exp(<italic>σ</italic><sup>2</sup>) − 1]exp(2<italic>μ</italic> + <italic>σ</italic><sup>2</sup>).</p>
      <p><italic><bold>Remark</bold></italic><italic><bold>1</bold></italic><bold>(</bold><italic><bold>Standardized</bold></italic><italic><bold>vs.</bold></italic><italic><bold>General</bold></italic><italic><bold>Lognormal</bold></italic><bold>).</bold></p>
      <p>When <italic>μ</italic> = 0, the lognormal distribution depends only on the scale parameter <italic>σ</italic>. In the standardized case used in Sections 1 and 3, we write <italic>γ</italic> = <italic>σ</italic> for the scale parameter (standard deviation of log<italic>X</italic>). Setting <italic>μ</italic> = 0 and <italic>σ</italic> = <italic>γ</italic> in any general formula recovers the corresponding standardized result.</p>
      <p><italic><bold>Definition</bold></italic><italic><bold>2</bold></italic><bold>(</bold><italic><bold>Error</bold></italic><italic><bold>Function</bold></italic><bold>).</bold></p>
      <p>The error function erf: ℝ → (−1, 1) is defined by: <inline-formula><mml:math display="inline"><mml:mrow><mml:mtext> erf </mml:mtext><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mrow><mml:mn> 2 </mml:mn><mml:mo> / </mml:mo><mml:mrow><mml:msqrt><mml:mi> π </mml:mi></mml:msqrt></mml:mrow></mml:mrow></mml:mrow><mml:mo> ) </mml:mo></mml:mrow><mml:mstyle displaystyle="true"><mml:mrow><mml:msubsup><mml:mo> ∫ </mml:mo><mml:mn> 0 </mml:mn><mml:mi> x </mml:mi></mml:msubsup><mml:mrow><mml:mi> exp </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mo> − </mml:mo><mml:msup><mml:mi> t </mml:mi><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> ) </mml:mo></mml:mrow><mml:mtext> d </mml:mtext><mml:mi> t </mml:mi></mml:mrow></mml:mrow></mml:mstyle></mml:mrow></mml:math></inline-formula> . It is an odd function: erf(−<italic>x</italic>) = −erf(<italic>x</italic>), with erf(0) = 0 and <inline-formula><mml:math display="inline"><mml:mrow><mml:msub><mml:mrow><mml:mi> lim </mml:mi></mml:mrow><mml:mi> x </mml:mi></mml:msub><mml:msub><mml:mrow></mml:mrow><mml:mrow><mml:mo> → </mml:mo><mml:mi> ∞ </mml:mi></mml:mrow></mml:msub><mml:mtext> erf </mml:mtext><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:mn> 1 </mml:mn></mml:mrow></mml:math></inline-formula> . The complementary error function is erfc(<italic>x</italic>) = 1 − erf(<italic>x</italic>). The error function is related to the standard normal CDF by: <inline-formula><mml:math display="inline"><mml:mrow><mml:mi> Φ </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mrow><mml:mn> 1 </mml:mn><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow></mml:mrow><mml:mo> ) </mml:mo></mml:mrow><mml:mrow><mml:mo> [ </mml:mo><mml:mrow><mml:mn> 1 </mml:mn><mml:mo> + </mml:mo><mml:mtext> erf </mml:mtext><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mrow><mml:mi> x </mml:mi><mml:mo> / </mml:mo><mml:mrow><mml:msqrt><mml:mn> 2 </mml:mn></mml:msqrt></mml:mrow></mml:mrow></mml:mrow><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mo> ] </mml:mo></mml:mrow></mml:mrow></mml:math></inline-formula> , so that <inline-formula><mml:math display="inline"><mml:mrow><mml:mtext> erf </mml:mtext><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:mn> 2 </mml:mn><mml:mi> Φ </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mi> x </mml:mi><mml:msqrt><mml:mn> 2 </mml:mn></mml:msqrt></mml:mrow><mml:mo> ) </mml:mo></mml:mrow><mml:mo> − </mml:mo><mml:mn> 1 </mml:mn></mml:mrow></mml:math></inline-formula> . This relationship is used extensively in the derivations that follow. </p>
      <p><italic><bold>Definition</bold></italic><italic><bold>3</bold></italic><bold>(</bold><italic><bold>Mean</bold></italic><italic><bold>Difference</bold></italic><bold>).</bold></p>
      <p>For a continuous random variable <italic>X</italic> with cumulative distribution function <italic>F</italic>(<italic>x</italic>) on a support (<italic>a</italic>, <italic>b</italic>), the mean difference (Gini’s mean difference) is defined as: Δ = <italic>E</italic>[|<italic>X</italic><sub>1</sub> − <italic>X</italic><sub>2</sub>|], where <italic>X</italic><sub>1</sub> and <italic>X</italic><sub>2</sub> are independent copies of <italic>X</italic>. An equivalent representation based solely on <italic>F</italic> is: <inline-formula><mml:math display="inline"><mml:mrow><mml:mi> Δ </mml:mi><mml:mo> = </mml:mo><mml:mn> 4 </mml:mn><mml:mstyle displaystyle="true"><mml:mrow><mml:msubsup><mml:mo> ∫ </mml:mo><mml:mi> a </mml:mi><mml:mi> b </mml:mi></mml:msubsup><mml:mrow><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mrow><mml:mo> [ </mml:mo><mml:mrow><mml:mn> 1 </mml:mn><mml:mo> − </mml:mo><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mo> ] </mml:mo></mml:mrow><mml:mtext> d </mml:mtext><mml:mi> x </mml:mi></mml:mrow></mml:mrow></mml:mstyle></mml:mrow></mml:math></inline-formula> , as given by Formula (3) below. This integral form is the starting point for all derivations in this paper. The mean difference is translation-invariant, homogeneous of degree one, and provides a robust measure of variability that is less sensitive to extreme observations than the variance.</p>
    </sec>
    <sec id="sec3">
      <title>3. Asymptotic Behavior of the Mean Difference</title>
      <sec id="sec3dot1">
        <title>
          3.1. Case
          <italic>γ</italic>
          → 0: Degeneration of the Distribution
        </title>
        <p>When the scale parameter <italic>γ</italic> tends to zero, the lognormal distribution degenerates toward a distribution concentrated at a single point. This behavior can be understood by analyzing the probabilistic structure of the distribution. If <italic>X</italic> follows a standardized lognormal distribution with scale parameter <italic>γ</italic>, then <italic>Y</italic> = log<italic>X</italic> has a normal distribution with mean zero and variance <italic>γ</italic><sup>2</sup>. Therefore, when <italic>γ</italic> → 0, the variance of <italic>Y</italic> tends to zero, which implies, by Chebyshev’s theorem, that <italic>Y</italic> converges in probability to zero.</p>
        <p>Consequently, <italic>X</italic> = e<italic><sup>Y</sup></italic> converges in probability to e<sup>0</sup> = 1. Formally, for <italic>γ</italic> → 0, the lognormal distribution converges weakly to the degenerate distribution <italic>δ</italic><sub>1</sub> concentrated at the point <italic>x</italic> = 1 (that is, at the point <italic>x</italic> = e<italic><sup>θ</sup></italic> in the general case with location parameter <italic>θ</italic>). This means that the entire probability mass progressively concentrates in an increasingly narrow neighborhood of the unit point, and the distribution tends to behave as a constant random variable.</p>
        <p>Since for a degenerate distribution all observations assume the same value, the expected absolute difference between two independent observations must necessarily tend to zero. Analytically, this result is verified directly from Formula (1). For <italic>γ</italic> → 0, since erf(0) = 0 and e<sup>0</sup> = 1:</p>
        <disp-formula id="FD4">
          <label>(4)</label>
          <mml:math>
            <mml:mrow>
              <mml:munder>
                <mml:mrow>
                  <mml:mi>lim</mml:mi>
                </mml:mrow>
                <mml:mrow>
                  <mml:mi>γ</mml:mi>
                  <mml:mo>→</mml:mo>
                  <mml:mn>0</mml:mn>
                </mml:mrow>
              </mml:munder>
              <mml:mi>Δ</mml:mi>
              <mml:mo>=</mml:mo>
              <mml:mn>2</mml:mn>
              <mml:mo>⋅</mml:mo>
              <mml:msup>
                <mml:mtext>e</mml:mtext>
                <mml:mn>0</mml:mn>
              </mml:msup>
              <mml:mo>⋅</mml:mo>
              <mml:mtext>erf</mml:mtext>
              <mml:mrow>
                <mml:mo>(</mml:mo>
                <mml:mn>0</mml:mn>
                <mml:mo>)</mml:mo>
              </mml:mrow>
              <mml:mo>=</mml:mo>
              <mml:mn>2</mml:mn>
              <mml:mo>⋅</mml:mo>
              <mml:mn>1</mml:mn>
              <mml:mo>⋅</mml:mo>
              <mml:mn>0</mml:mn>
              <mml:mo>=</mml:mo>
              <mml:mn>0.</mml:mn>
            </mml:mrow>
          </mml:math>
        </disp-formula>
        <p>More precisely, using the Taylor series expansion of the error function, <inline-formula><mml:math display="inline"><mml:mrow><mml:mtext> erf </mml:mtext><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow></mml:mrow></mml:math></inline-formula><inline-formula><mml:math display="inline"><mml:mrow><mml:mo> ≈ </mml:mo><mml:mrow><mml:mrow><mml:mn> 2 </mml:mn><mml:mi> x </mml:mi></mml:mrow><mml:mo> / </mml:mo><mml:mrow><mml:msqrt><mml:mi> π </mml:mi></mml:msqrt></mml:mrow></mml:mrow></mml:mrow></mml:math></inline-formula> for <inline-formula><mml:math display="inline"><mml:mrow><mml:mi> x </mml:mi><mml:mo> → </mml:mo><mml:mn> 0 </mml:mn></mml:mrow></mml:math></inline-formula> , and the expansion of the exponential <inline-formula><mml:math display="inline"><mml:mrow><mml:msup><mml:mtext> e </mml:mtext><mml:mrow><mml:mrow><mml:mrow><mml:msup><mml:mi> γ </mml:mi><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow></mml:mrow></mml:msup><mml:mo> ≈ </mml:mo><mml:mn> 1 </mml:mn><mml:mo> + </mml:mo><mml:mrow><mml:mrow><mml:msup><mml:mi> γ </mml:mi><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow><mml:mo> + </mml:mo></mml:mrow></mml:math></inline-formula><inline-formula><mml:math display="inline"><mml:mrow><mml:mi> O </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:msup><mml:mi> γ </mml:mi><mml:mn> 4 </mml:mn></mml:msup></mml:mrow><mml:mo> ) </mml:mo></mml:mrow></mml:mrow></mml:math></inline-formula> , the following asymptotic expansion is obtained:</p>
        <disp-formula id="FD5">
          <label>(5)</label>
          <mml:math>
            <mml:mrow>
              <mml:mi>Δ</mml:mi>
              <mml:mo>≈</mml:mo>
              <mml:mfrac>
                <mml:mn>2</mml:mn>
                <mml:mrow>
                  <mml:msqrt>
                    <mml:mi>π</mml:mi>
                  </mml:msqrt>
                </mml:mrow>
              </mml:mfrac>
              <mml:mi>γ</mml:mi>
              <mml:mo>,</mml:mo>
              <mml:mi>γ</mml:mi>
              <mml:mo>→</mml:mo>
              <mml:mn>0</mml:mn>
            </mml:mrow>
          </mml:math>
        </disp-formula>
        <p>This result confirms that the standardized mean difference Δ → 0 for <italic>γ</italic> → 0, with a linear rate of convergence in the parameter <italic>γ</italic>. The coefficient <inline-formula><mml:math display="inline"><mml:mrow><mml:mrow><mml:mn> 2 </mml:mn><mml:mo> / </mml:mo><mml:mrow><mml:msqrt><mml:mi> π </mml:mi></mml:msqrt></mml:mrow></mml:mrow><mml:mo> ≈ </mml:mo><mml:mn> 1.128 </mml:mn></mml:mrow></mml:math></inline-formula> quantifies the rate of decrease of the mean difference in the neighborhood of the origin. The result is fully consistent with the probabilistic interpretation: when the dispersion of the lognormal distribution vanishes, the absolute variability, as measured by the mean difference, also vanishes, and this occurs at a rate proportional to the dispersion parameter itself.</p>
      </sec>
      <sec id="sec3dot2">
        <title>
          3.2. Case
          <italic>γ</italic>
          → +∞: Unbounded Dispersion
        </title>
        <p>When the scale parameter <italic>γ</italic> tends to infinity, the lognormal distribution becomes increasingly dispersed, with a right tail that extends indefinitely and an increasing concentration of probability mass near zero. In this regime, the variance of log<italic>X</italic> diverges, producing an extremely positively skewed (right-skewed) distribution with a skewness coefficient that grows without bound.</p>
        <p>To analyze the asymptotic behavior of the mean difference, we consider separately the two factors of Formula (1). The exponential factor and the error function factor behave as follows for <italic>γ</italic> → +∞:</p>
        <disp-formula id="FD6">
          <label>(6)</label>
          <mml:math>
            <mml:mrow>
              <mml:msup>
                <mml:mtext>e</mml:mtext>
                <mml:mrow>
                  <mml:mrow>
                    <mml:mrow>
                      <mml:msup>
                        <mml:mi>γ</mml:mi>
                        <mml:mn>2</mml:mn>
                      </mml:msup>
                    </mml:mrow>
                    <mml:mo>/</mml:mo>
                    <mml:mn>2</mml:mn>
                  </mml:mrow>
                </mml:mrow>
              </mml:msup>
              <mml:mo>→</mml:mo>
              <mml:mo>+</mml:mo>
              <mml:mi>∞</mml:mi>
              <mml:mo>,</mml:mo>
              <mml:mtext>erf</mml:mtext>
              <mml:mrow>
                <mml:mo>(</mml:mo>
                <mml:mrow>
                  <mml:mfrac>
                    <mml:mi>γ</mml:mi>
                    <mml:mn>2</mml:mn>
                  </mml:mfrac>
                </mml:mrow>
                <mml:mo>)</mml:mo>
              </mml:mrow>
              <mml:mo>→</mml:mo>
              <mml:mn>1</mml:mn>
            </mml:mrow>
          </mml:math>
        </disp-formula>
        <p>The exponential factor <inline-formula><mml:math><mml:mrow><mml:msup><mml:mtext> e </mml:mtext><mml:mrow><mml:mrow><mml:mrow><mml:msup><mml:mi> γ </mml:mi><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> diverges at super-exponential speed, completely dominating the asymptotic behavior of the product, while the factor erf(<italic>γ</italic>/2) converges monotonically to 1 from below. Consequently, the product of the two factors diverges:</p>
        <disp-formula id="FD7">
          <label>(7)</label>
          <mml:math>
            <mml:mrow>
              <mml:mi>Δ</mml:mi>
              <mml:mo>→</mml:mo>
              <mml:mi>∞</mml:mi>
              <mml:mo>.</mml:mo>
            </mml:mrow>
          </mml:math>
        </disp-formula>
        <p>This result is consistent with the fact that the mean of the standardized lognormal distribution is <inline-formula><mml:math display="inline"><mml:mrow><mml:mi> E </mml:mi><mml:mrow><mml:mo> [ </mml:mo><mml:mi> X </mml:mi><mml:mo> ] </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:msup><mml:mtext> e </mml:mtext><mml:mrow><mml:mrow><mml:mrow><mml:msup><mml:mi> γ </mml:mi><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> , and therefore the mean difference, being a homogeneous function of the scale, must also diverge. More significant is the ratio of the mean difference to the mean of the distribution, which tends to the limiting value:</p>
        <disp-formula id="FD8">
          <label>(8)</label>
          <mml:math>
            <mml:mrow>
              <mml:mfrac>
                <mml:mi>Δ</mml:mi>
                <mml:mrow>
                  <mml:mi>E</mml:mi>
                  <mml:mrow>
                    <mml:mo>[</mml:mo>
                    <mml:mi>X</mml:mi>
                    <mml:mo>]</mml:mo>
                  </mml:mrow>
                </mml:mrow>
              </mml:mfrac>
              <mml:mo>=</mml:mo>
              <mml:mn>2</mml:mn>
              <mml:mo>⋅</mml:mo>
              <mml:mtext>erf</mml:mtext>
              <mml:mrow>
                <mml:mo>(</mml:mo>
                <mml:mrow>
                  <mml:mfrac>
                    <mml:mi>γ</mml:mi>
                    <mml:mn>2</mml:mn>
                  </mml:mfrac>
                </mml:mrow>
                <mml:mo>)</mml:mo>
              </mml:mrow>
              <mml:mo>→</mml:mo>
              <mml:mn>2</mml:mn>
            </mml:mrow>
          </mml:math>
        </disp-formula>
        <p>This indicates that, for large values of <italic>γ</italic>, the mean difference is approximately twice the mean, a result that reflects the extreme skewness of the lognormal distribution in the high-dispersion regime. In terms of the Gini concentration index, <italic>G</italic> = Δ/(2<italic>E</italic>[<italic>X</italic>]), this implies that <italic>G</italic> → 1 for <italic>γ</italic> → +∞, reaching the value of maximum concentration. This behavior is consistent with the well-known property that the Gini index of the lognormal distribution is given by <italic>G</italic> = erf(<italic>σ</italic>/2), which tends to 1 when <italic>σ</italic> → +∞.</p>
      </sec>
    </sec>
    <sec id="sec4">
      <title>4. Generalized Formula with Complete Parameters</title>
      <p>Formula (1) was derived for the standardized lognormal distribution, with location parameter <italic>θ</italic> = 0 and employing <italic>γ</italic> as the sole scale parameter. In statistical practice and applications, the complete lognormal distribution is characterized by two parameters: the location parameter <italic>μ</italic> (mean of the logarithm of the variable) and the scale parameter <italic>σ</italic> (standard deviation of the logarithm of the variable). A random variable <italic>X</italic> is said to have a lognormal distribution with parameters <italic>μ</italic> and <italic>σ</italic><sup>2</sup> if <italic>Y</italic> = log<italic>X</italic> ~ <italic>N</italic>(<italic>μ</italic>, <italic>σ</italic><sup>2</sup>). The density function of the general lognormal distribution is:</p>
      <disp-formula id="FD9">
        <label>(9)</label>
        <mml:math>
          <mml:mrow>
            <mml:mi>f</mml:mi>
            <mml:mrow>
              <mml:mo>(</mml:mo>
              <mml:mi>x</mml:mi>
              <mml:mo>)</mml:mo>
            </mml:mrow>
            <mml:mo>=</mml:mo>
            <mml:mfrac>
              <mml:mrow>
                <mml:msup>
                  <mml:mtext>e</mml:mtext>
                  <mml:mrow>
                    <mml:mo>−</mml:mo>
                    <mml:mfrac>
                      <mml:mrow>
                        <mml:msup>
                          <mml:mrow>
                            <mml:mrow>
                              <mml:mo>(</mml:mo>
                              <mml:mrow>
                                <mml:mi>log</mml:mi>
                                <mml:mi>x</mml:mi>
                                <mml:mo>−</mml:mo>
                                <mml:mi>μ</mml:mi>
                              </mml:mrow>
                              <mml:mo>)</mml:mo>
                            </mml:mrow>
                          </mml:mrow>
                          <mml:mn>2</mml:mn>
                        </mml:msup>
                      </mml:mrow>
                      <mml:mrow>
                        <mml:mn>2</mml:mn>
                        <mml:msup>
                          <mml:mi>δ</mml:mi>
                          <mml:mn>2</mml:mn>
                        </mml:msup>
                      </mml:mrow>
                    </mml:mfrac>
                  </mml:mrow>
                </mml:msup>
              </mml:mrow>
              <mml:mrow>
                <mml:mi>x</mml:mi>
                <mml:mi>δ</mml:mi>
                <mml:msqrt>
                  <mml:mrow>
                    <mml:mn>2</mml:mn>
                    <mml:mi>π</mml:mi>
                  </mml:mrow>
                </mml:msqrt>
              </mml:mrow>
            </mml:mfrac>
            <mml:mo>,</mml:mo>
            <mml:mi>x</mml:mi>
            <mml:mo>&gt;</mml:mo>
            <mml:mn>0.</mml:mn>
          </mml:mrow>
        </mml:math>
      </disp-formula>
      <p>The mean difference, being independent of the location parameter and homogeneous of degree one with respect to the scale parameter, can be generalized by exploiting these fundamental properties. Recall that if <italic>X</italic> is a random variable with mean difference Δ(<italic>X</italic>), and <italic>Y</italic> = <italic>aX</italic> + <italic>b</italic> with <italic>a</italic> &gt; 0, then Δ(<italic>Y</italic>) = <italic>a</italic> ∙ Δ(<italic>X</italic>). This property is a direct consequence of the definition of the mean difference as the expected value of the absolute difference.</p>
      <p>In the case of the lognormal distribution, if <italic>X</italic> ~ Lognormal (<italic>μ</italic>, <italic>σ</italic><sup>2</sup>), we can write <italic>X</italic> = e<italic><sup>μ</sup></italic><sup>+</sup><italic><sup>σZ</sup></italic> = e<italic><sup>μ</sup></italic>∙e<italic><sup>σZ</sup></italic>, where <italic>Z</italic> ~ <italic>N</italic>(0, 1). The variable <italic>W</italic> = e<italic><sup>σZ</sup></italic> follows a standardized lognormal distribution whose scale parameter <italic>σ</italic> is precisely the same quantity as the parameter <italic>γ</italic> of Formula (1): both denote the standard deviation of log<italic>X</italic>. Setting <italic>μ</italic> = 0 and <italic>σ</italic> = <italic>γ</italic> in the generalized formula recovers exactly Formula (1). Therefore, <italic>X</italic> = e<italic><sup>μ</sup></italic>∙<italic>W</italic>, and by the homogeneity property: <inline-formula><mml:math display="inline"><mml:mrow><mml:mi> Δ </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> X </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:msup><mml:mtext> e </mml:mtext><mml:mi> μ </mml:mi></mml:msup><mml:mo> ⋅ </mml:mo><mml:mi> Δ </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> W </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> = </mml:mo></mml:mrow></mml:math></inline-formula><inline-formula><mml:math display="inline"><mml:mrow><mml:msup><mml:mtext> e </mml:mtext><mml:mi> μ </mml:mi></mml:msup><mml:mo> ⋅ </mml:mo><mml:mn> 2 </mml:mn><mml:msup><mml:mtext> e </mml:mtext><mml:mrow><mml:mrow><mml:mrow><mml:msup><mml:mi> σ </mml:mi><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow></mml:mrow></mml:msup><mml:mtext> erf </mml:mtext><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mrow><mml:mi> σ </mml:mi><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow></mml:mrow><mml:mo> ) </mml:mo></mml:mrow></mml:mrow></mml:math></inline-formula> .</p>
      <p>This leads to the generalized formula for the lognormal mean difference:</p>
      <disp-formula id="FD10">
        <label>(10)</label>
        <mml:math>
          <mml:mrow>
            <mml:mi>Δ</mml:mi>
            <mml:mrow>
              <mml:mo>(</mml:mo>
              <mml:mrow>
                <mml:mi>μ</mml:mi>
                <mml:mo>,</mml:mo>
                <mml:mi>σ</mml:mi>
              </mml:mrow>
              <mml:mo>)</mml:mo>
            </mml:mrow>
            <mml:mo>=</mml:mo>
            <mml:mn>2</mml:mn>
            <mml:mi>exp</mml:mi>
            <mml:mrow>
              <mml:mo>(</mml:mo>
              <mml:mrow>
                <mml:mi>μ</mml:mi>
                <mml:mo>+</mml:mo>
                <mml:mfrac>
                  <mml:mrow>
                    <mml:msup>
                      <mml:mi>σ</mml:mi>
                      <mml:mn>2</mml:mn>
                    </mml:msup>
                  </mml:mrow>
                  <mml:mn>2</mml:mn>
                </mml:mfrac>
              </mml:mrow>
              <mml:mo>)</mml:mo>
            </mml:mrow>
            <mml:mo>⋅</mml:mo>
            <mml:mtext>erf</mml:mtext>
            <mml:mrow>
              <mml:mo>(</mml:mo>
              <mml:mrow>
                <mml:mfrac>
                  <mml:mi>σ</mml:mi>
                  <mml:mn>2</mml:mn>
                </mml:mfrac>
              </mml:mrow>
              <mml:mo>)</mml:mo>
            </mml:mrow>
          </mml:mrow>
        </mml:math>
      </disp-formula>
      <p>where: <italic>μ</italic> = location parameter (mean of the logarithm of the variable); <italic>σ</italic> = scale parameter (standard deviation of the logarithm of the variable); Δ(<italic>μ</italic>, <italic>σ</italic>) = mean difference for the lognormal distribution with parameters <italic>μ</italic> and <italic>σ</italic><sup>2</sup>.</p>
      <p>The term e<italic><sup>μ</sup></italic> appears as a multiplicative factor arising from the exponential transformation of the location parameter. Note that Formula (10) reduces to (1) when <italic>μ</italic> = 0 and <italic>σ</italic> = <italic>γ</italic>, confirming that <italic>γ</italic> and <italic>σ</italic> are notational variants of the same scale parameter and making the relationship between the two formulations fully transparent. Furthermore, from Formula (10) and the well-known expression for the mean of the lognormal distribution, <inline-formula><mml:math display="inline"><mml:mrow><mml:mi> E </mml:mi><mml:mrow><mml:mo> [ </mml:mo><mml:mi> X </mml:mi><mml:mo> ] </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:msup><mml:mtext> e </mml:mtext><mml:mrow><mml:mi> μ </mml:mi><mml:mo> + </mml:mo><mml:mrow><mml:mrow><mml:msup><mml:mi> σ </mml:mi><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow></mml:mrow></mml:msup></mml:mrow></mml:math></inline-formula> , one immediately obtains the expression for the Gini concentration index:</p>
      <disp-formula id="FD11">
        <label>(11)</label>
        <mml:math>
          <mml:mrow>
            <mml:mi>G</mml:mi>
            <mml:mo>=</mml:mo>
            <mml:mfrac>
              <mml:mi>Δ</mml:mi>
              <mml:mrow>
                <mml:mn>2</mml:mn>
                <mml:mi>E</mml:mi>
                <mml:mrow>
                  <mml:mo>[</mml:mo>
                  <mml:mi>X</mml:mi>
                  <mml:mo>]</mml:mo>
                </mml:mrow>
              </mml:mrow>
            </mml:mfrac>
            <mml:mo>=</mml:mo>
            <mml:mtext>erf</mml:mtext>
            <mml:mrow>
              <mml:mo>(</mml:mo>
              <mml:mrow>
                <mml:mfrac>
                  <mml:mi>σ</mml:mi>
                  <mml:mn>2</mml:mn>
                </mml:mfrac>
              </mml:mrow>
              <mml:mo>)</mml:mo>
            </mml:mrow>
          </mml:mrow>
        </mml:math>
      </disp-formula>
      <p>which depends solely on the scale parameter <italic>σ</italic> and not on the location parameter <italic>μ</italic>, confirming a well-known property of the lognormal distribution in concentration theory.</p>
    </sec>
    <sec id="sec5">
      <title>5. Mean Difference Conditioned on an Interval</title>
      <p>In many statistical applications, it is of interest to measure the variability not of the entire distribution, but of a subpopulation defined by constraints on the observable values. This need arises frequently in economic contexts (analysis of income variability within specific brackets), in biomedical contexts (variability of a biomarker within a clinically relevant interval), and in industrial contexts (variability of a quantity within tolerance specifications).</p>
      <p>The mean difference conditioned on an interval [<italic>a</italic>, <italic>b</italic>] represents the natural extension of the concept of the mean difference to the case where only observations falling between the limits a and b are considered. Let <italic>X</italic> be a random variable with cumulative distribution function <italic>F</italic>(<italic>x</italic>) and let 0 &lt; <italic>F</italic>(<italic>a</italic>) &lt; <italic>F</italic>(<italic>b</italic>) &lt; 1. The conditional distribution of <italic>X</italic> given <italic>a</italic> ≤ <italic>X</italic> ≤ <italic>b</italic> has density <italic>f</italic>(<italic>x</italic>|<italic>a</italic> ≤ <italic>X</italic> ≤ <italic>b</italic>) = <italic>f</italic>(<italic>x</italic>)/[<italic>F</italic>(<italic>b</italic>) − <italic>F</italic>(<italic>a</italic>)] for <italic>x</italic> ∈ [<italic>a</italic>, <italic>b</italic>] and conditional cumulative distribution function <italic>F</italic><italic><sub>c</sub></italic>(<italic>x</italic>) = [<italic>F</italic>(<italic>x</italic>) − <italic>F</italic>(<italic>a</italic>)]/[<italic>F</italic>(<italic>b</italic>) − <italic>F</italic>(<italic>a</italic>)].</p>
      <p>Applying the formula for the mean difference based on the cumulative distribution function to the conditional distribution, and taking into account that <inline-formula><mml:math display="inline"><mml:mrow><mml:msub><mml:mi> F </mml:mi><mml:mi> c </mml:mi></mml:msub><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mrow><mml:mo> [ </mml:mo><mml:mrow><mml:mn> 1 </mml:mn><mml:mo> − </mml:mo><mml:msub><mml:mi> F </mml:mi><mml:mi> c </mml:mi></mml:msub><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mo> ] </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:mrow><mml:mrow><mml:mrow><mml:mo> [ </mml:mo><mml:mrow><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> − </mml:mo><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> a </mml:mi><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mo> ] </mml:mo></mml:mrow><mml:mrow><mml:mo> [ </mml:mo><mml:mrow><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> b </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> − </mml:mo><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mo> ] </mml:mo></mml:mrow></mml:mrow><mml:mo> / </mml:mo><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo> [ </mml:mo><mml:mrow><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> b </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> − </mml:mo><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> a </mml:mi><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mo> ] </mml:mo></mml:mrow></mml:mrow><mml:mn> 2 </mml:mn></mml:msup></mml:mrow></mml:mrow></mml:mrow></mml:math></inline-formula> , the following general formula for the conditional mean difference, hereafter denoted Δ<italic><sub>c</sub></italic>[<italic>a</italic>, <italic>b</italic>], is obtained:</p>
      <disp-formula id="FD12">
        <label>(12)</label>
        <mml:math display="inline">
          <mml:mrow>
            <mml:msub>
              <mml:mi>Δ</mml:mi>
              <mml:mrow>
                <mml:mrow>
                  <mml:mo>[</mml:mo>
                  <mml:mrow>
                    <mml:mi>a</mml:mi>
                    <mml:mo>,</mml:mo>
                    <mml:mi>b</mml:mi>
                  </mml:mrow>
                  <mml:mo>]</mml:mo>
                </mml:mrow>
              </mml:mrow>
            </mml:msub>
            <mml:mo>=</mml:mo>
            <mml:mfrac>
              <mml:mn>4</mml:mn>
              <mml:mrow>
                <mml:mrow>
                  <mml:mo>[</mml:mo>
                  <mml:mrow>
                    <mml:mi>F</mml:mi>
                    <mml:mrow>
                      <mml:mo>(</mml:mo>
                      <mml:mi>b</mml:mi>
                      <mml:mo>)</mml:mo>
                    </mml:mrow>
                    <mml:mo>−</mml:mo>
                    <mml:mi>F</mml:mi>
                    <mml:mrow>
                      <mml:mo>(</mml:mo>
                      <mml:mi>a</mml:mi>
                      <mml:mo>)</mml:mo>
                    </mml:mrow>
                  </mml:mrow>
                  <mml:mo>]</mml:mo>
                </mml:mrow>
              </mml:mrow>
            </mml:mfrac>
            <mml:mstyle displaystyle="true">
              <mml:mrow>
                <mml:msubsup>
                  <mml:mo>∫</mml:mo>
                  <mml:mi>a</mml:mi>
                  <mml:mi>b</mml:mi>
                </mml:msubsup>
                <mml:mrow>
                  <mml:mrow>
                    <mml:mo>[</mml:mo>
                    <mml:mrow>
                      <mml:mi>F</mml:mi>
                      <mml:mrow>
                        <mml:mo>(</mml:mo>
                        <mml:mi>x</mml:mi>
                        <mml:mo>)</mml:mo>
                      </mml:mrow>
                      <mml:mo>−</mml:mo>
                      <mml:mi>F</mml:mi>
                      <mml:mrow>
                        <mml:mo>(</mml:mo>
                        <mml:mi>a</mml:mi>
                        <mml:mo>)</mml:mo>
                      </mml:mrow>
                    </mml:mrow>
                    <mml:mo>]</mml:mo>
                  </mml:mrow>
                  <mml:mrow>
                    <mml:mo>[</mml:mo>
                    <mml:mrow>
                      <mml:mi>F</mml:mi>
                      <mml:mrow>
                        <mml:mo>(</mml:mo>
                        <mml:mi>b</mml:mi>
                        <mml:mo>)</mml:mo>
                      </mml:mrow>
                      <mml:mo>−</mml:mo>
                      <mml:mi>F</mml:mi>
                      <mml:mrow>
                        <mml:mo>(</mml:mo>
                        <mml:mi>x</mml:mi>
                        <mml:mo>)</mml:mo>
                      </mml:mrow>
                    </mml:mrow>
                    <mml:mo>]</mml:mo>
                  </mml:mrow>
                  <mml:mtext>d</mml:mtext>
                  <mml:mi>x</mml:mi>
                </mml:mrow>
              </mml:mrow>
            </mml:mstyle>
          </mml:mrow>
        </mml:math>
      </disp-formula>
      <p>where <italic>F</italic>(<italic>a</italic>) and <italic>F</italic>(<italic>b</italic>) are the values of the cumulative distribution function at the endpoints of the interval. The factor 4/[<italic>F</italic>(<italic>b</italic>) − <italic>F</italic>(<italic>a</italic>)]<sup>2</sup> serves as a normalization that accounts for the probability mass contained in the interval<sup>3</sup>.</p>
      <p>Formula (12) has a particularly elegant structure: the integrand [<italic>F</italic>(<italic>x</italic>) − <italic>F</italic>(<italic>a</italic>)] [<italic>F</italic>(<italic>b</italic>) − <italic>F</italic>(<italic>x</italic>)] is a non-negative function that vanishes at the endpoints of the interval and reaches its maximum at the conditional median. This structure reflects the fact that the pairs of observations with the largest expected difference are those located on opposite sides of the median of the conditional distribution.</p>
      <p>In the context of the lognormal distribution, this formulation is particularly relevant when analyzing subgroups defined by economic thresholds (income brackets), biological thresholds (concentration intervals), clinical thresholds (diagnostic value ranges), or environmental thresholds (exposure brackets). The formula allows quantifying intra-group variability analytically, avoiding the approximations that would be necessary with numerical methods.</p>
    </sec>
    <sec id="sec6">
      <title>6. Mean Difference of the Truncated Lognormal Distribution</title>
      <p>Truncation is frequent in practical applications where extreme values are physically impossible, unobservable, or subject to censoring. The lognormal distribution truncated on the interval [<italic>a</italic>, <italic>b</italic>] finds application in numerous contexts: biochemical concentrations (which cannot be negative and have physiological upper bounds), incomes (bounded below and above in sample surveys), survival times (truncated by the observation period of the study), and particle sizes (limited by instrumental detection thresholds).</p>
      <p>The formula for the mean difference of the truncated lognormal is obtained by specializing the conditional mean difference Formula (12) to the lognormal distribution. Through the substitution of the lognormal cumulative distribution function <italic>F</italic>(<italic>x</italic>) = Φ((log<italic>x</italic> − <italic>μ</italic>)/<italic>σ</italic>), where Φ(·) is the cumulative distribution function of the standard normal distribution, and the change of variable <italic>y</italic> = (log<italic>x</italic> − <italic>μ</italic>)/<italic>σ</italic>, the limits of integration are transformed into standardized limits. After appropriate algebraic simplifications, the following formula for the truncated mean difference Δ<italic><sub>T</sub></italic> is obtained:</p>
      <disp-formula id="FD13">
        <label>(13)</label>
        <mml:math>
          <mml:mrow>
            <mml:msup>
              <mml:mi>Δ</mml:mi>
              <mml:mrow>
                <mml:mrow>
                  <mml:mo>(</mml:mo>
                  <mml:mi>T</mml:mi>
                  <mml:mo>)</mml:mo>
                </mml:mrow>
              </mml:mrow>
            </mml:msup>
            <mml:mo>=</mml:mo>
            <mml:mfrac>
              <mml:mrow>
                <mml:mn>2</mml:mn>
                <mml:msup>
                  <mml:mtext>e</mml:mtext>
                  <mml:mrow>
                    <mml:mi>μ</mml:mi>
                    <mml:mo>+</mml:mo>
                    <mml:mrow>
                      <mml:mrow>
                        <mml:msup>
                          <mml:mi>σ</mml:mi>
                          <mml:mn>2</mml:mn>
                        </mml:msup>
                      </mml:mrow>
                      <mml:mo>/</mml:mo>
                      <mml:mn>2</mml:mn>
                    </mml:mrow>
                  </mml:mrow>
                </mml:msup>
              </mml:mrow>
              <mml:mrow>
                <mml:mi>Φ</mml:mi>
                <mml:mrow>
                  <mml:mo>(</mml:mo>
                  <mml:mrow>
                    <mml:msup>
                      <mml:mi>β</mml:mi>
                      <mml:mo>*</mml:mo>
                    </mml:msup>
                  </mml:mrow>
                  <mml:mo>)</mml:mo>
                </mml:mrow>
                <mml:mo>−</mml:mo>
                <mml:mi>Φ</mml:mi>
                <mml:mrow>
                  <mml:mo>(</mml:mo>
                  <mml:mrow>
                    <mml:msup>
                      <mml:mi>α</mml:mi>
                      <mml:mo>*</mml:mo>
                    </mml:msup>
                  </mml:mrow>
                  <mml:mo>)</mml:mo>
                </mml:mrow>
              </mml:mrow>
            </mml:mfrac>
            <mml:mrow>
              <mml:mo>[</mml:mo>
              <mml:mrow>
                <mml:mtext>erf</mml:mtext>
                <mml:mrow>
                  <mml:mo>(</mml:mo>
                  <mml:mrow>
                    <mml:mfrac>
                      <mml:mrow>
                        <mml:msup>
                          <mml:mi>β</mml:mi>
                          <mml:mo>*</mml:mo>
                        </mml:msup>
                        <mml:mo>−</mml:mo>
                        <mml:mi>σ</mml:mi>
                      </mml:mrow>
                      <mml:mrow>
                        <mml:msqrt>
                          <mml:mn>2</mml:mn>
                        </mml:msqrt>
                      </mml:mrow>
                    </mml:mfrac>
                  </mml:mrow>
                  <mml:mo>)</mml:mo>
                </mml:mrow>
                <mml:mo>−</mml:mo>
                <mml:mtext>erf</mml:mtext>
                <mml:mrow>
                  <mml:mo>(</mml:mo>
                  <mml:mrow>
                    <mml:mfrac>
                      <mml:mrow>
                        <mml:msup>
                          <mml:mi>α</mml:mi>
                          <mml:mo>*</mml:mo>
                        </mml:msup>
                        <mml:mo>−</mml:mo>
                        <mml:mi>σ</mml:mi>
                      </mml:mrow>
                      <mml:mrow>
                        <mml:msqrt>
                          <mml:mn>2</mml:mn>
                        </mml:msqrt>
                      </mml:mrow>
                    </mml:mfrac>
                  </mml:mrow>
                  <mml:mo>)</mml:mo>
                </mml:mrow>
              </mml:mrow>
              <mml:mo>]</mml:mo>
            </mml:mrow>
            <mml:mo>−</mml:mo>
            <mml:msub>
              <mml:mi>Δ</mml:mi>
              <mml:mn>0</mml:mn>
            </mml:msub>
          </mml:mrow>
        </mml:math>
      </disp-formula>
      <p>where the standardized limits are defined as: </p>
      <disp-formula id="FD14">
        <label>(14)</label>
        <mml:math>
          <mml:mrow>
            <mml:msup>
              <mml:mi>α</mml:mi>
              <mml:mo>*</mml:mo>
            </mml:msup>
            <mml:mo>=</mml:mo>
            <mml:mrow>
              <mml:mrow>
                <mml:mrow>
                  <mml:mo>(</mml:mo>
                  <mml:mrow>
                    <mml:mi>log</mml:mi>
                    <mml:mrow>
                      <mml:mo>(</mml:mo>
                      <mml:mrow>
                        <mml:mi>a</mml:mi>
                        <mml:mo>−</mml:mo>
                        <mml:mi>μ</mml:mi>
                      </mml:mrow>
                      <mml:mo>)</mml:mo>
                    </mml:mrow>
                  </mml:mrow>
                  <mml:mo>)</mml:mo>
                </mml:mrow>
              </mml:mrow>
              <mml:mo>/</mml:mo>
              <mml:mi>σ</mml:mi>
            </mml:mrow>
            <mml:mo>,</mml:mo>
            <mml:mtext>
               
            </mml:mtext>
            <mml:msup>
              <mml:mi>β</mml:mi>
              <mml:mo>*</mml:mo>
            </mml:msup>
            <mml:mo>=</mml:mo>
            <mml:mrow>
              <mml:mrow>
                <mml:mrow>
                  <mml:mo>(</mml:mo>
                  <mml:mrow>
                    <mml:mi>log</mml:mi>
                    <mml:mrow>
                      <mml:mo>(</mml:mo>
                      <mml:mrow>
                        <mml:mi>b</mml:mi>
                        <mml:mo>−</mml:mo>
                        <mml:mi>μ</mml:mi>
                      </mml:mrow>
                      <mml:mo>)</mml:mo>
                    </mml:mrow>
                  </mml:mrow>
                  <mml:mo>)</mml:mo>
                </mml:mrow>
              </mml:mrow>
              <mml:mo>/</mml:mo>
              <mml:mi>σ</mml:mi>
            </mml:mrow>
          </mml:mrow>
        </mml:math>
      </disp-formula>
      <p>and the other variables are: [<italic>a</italic>, <italic>b</italic>] = truncation interval; Φ(·) = cumulative distribution function of the standard normal distribution; Δ<sub>0</sub> = corrective term for the truncation, defined as the residual adjustment that ensures the truncated mean difference Δ<italic><sub>T</sub></italic> converges to the unrestricted Δ(<italic>μ</italic>, <italic>σ</italic>) when the truncation interval expands to (0, +∞)<sup>4</sup>.</p>
      <p>The derivation proceeds from the conditional mean difference formulation (Section 5), specialized to the lognormal distribution with parameters <italic>μ</italic> and <italic>σ</italic>. The substitution of the lognormal cumulative distribution function into the integral of Formula (12) and the application of the change of variable lead to integrals involving the error function, analogous to those used by Girone and Manca [<xref ref-type="bibr" rid="B13">13</xref>] in the derivation of the original formula.</p>
      <p>The corrective term Δ<sub>0</sub> represents the boundary correction that accounts for the effects introduced by the truncation on the mean difference. Specifically, Δ<sub>0</sub> is defined as the double integral </p>
      <disp-formula id="FD15">
        <mml:math display="inline">
          <mml:mtable columnalign="left">
            <mml:mtr>
              <mml:mtd>
                <mml:msub>
                  <mml:mi>Δ</mml:mi>
                  <mml:mn>0</mml:mn>
                </mml:msub>
                <mml:mo>=</mml:mo>
                <mml:mrow>
                  <mml:mo>(</mml:mo>
                  <mml:mrow>
                    <mml:mrow>
                      <mml:mn>4</mml:mn>
                      <mml:mo>/</mml:mo>
                      <mml:mrow>
                        <mml:msup>
                          <mml:mrow>
                            <mml:mrow>
                              <mml:mo>[</mml:mo>
                              <mml:mrow>
                                <mml:mi>Φ</mml:mi>
                                <mml:mrow>
                                  <mml:mo>(</mml:mo>
                                  <mml:mrow>
                                    <mml:msup>
                                      <mml:mi>β</mml:mi>
                                      <mml:mo>∗</mml:mo>
                                    </mml:msup>
                                  </mml:mrow>
                                  <mml:mo>)</mml:mo>
                                </mml:mrow>
                                <mml:mo>−</mml:mo>
                                <mml:mi>Φ</mml:mi>
                                <mml:mrow>
                                  <mml:mo>(</mml:mo>
                                  <mml:mrow>
                                    <mml:msup>
                                      <mml:mi>α</mml:mi>
                                      <mml:mo>∗</mml:mo>
                                    </mml:msup>
                                  </mml:mrow>
                                  <mml:mo>)</mml:mo>
                                </mml:mrow>
                              </mml:mrow>
                              <mml:mo>]</mml:mo>
                            </mml:mrow>
                          </mml:mrow>
                          <mml:mn>2</mml:mn>
                        </mml:msup>
                      </mml:mrow>
                    </mml:mrow>
                  </mml:mrow>
                  <mml:mo>)</mml:mo>
                </mml:mrow>
                <mml:mstyle displaystyle="true">
                  <mml:mrow>
                    <mml:msub>
                      <mml:mo>∬</mml:mo>
                      <mml:mrow>
                        <mml:msup>
                          <mml:mrow>
                            <mml:mrow>
                              <mml:mo>[</mml:mo>
                              <mml:mrow>
                                <mml:mi>a</mml:mi>
                                <mml:mo>,</mml:mo>
                                <mml:mi>b</mml:mi>
                              </mml:mrow>
                              <mml:mo>]</mml:mo>
                            </mml:mrow>
                          </mml:mrow>
                          <mml:mn>2</mml:mn>
                        </mml:msup>
                      </mml:mrow>
                    </mml:msub>
                    <mml:mrow>
                      <mml:mrow>
                        <mml:mo>|</mml:mo>
                        <mml:mrow>
                          <mml:mi>Φ</mml:mi>
                          <mml:mrow>
                            <mml:mo>(</mml:mo>
                            <mml:mrow>
                              <mml:mrow>
                                <mml:mrow>
                                  <mml:mrow>
                                    <mml:mo>(</mml:mo>
                                    <mml:mrow>
                                      <mml:mi>log</mml:mi>
                                      <mml:mi>x</mml:mi>
                                      <mml:mo>−</mml:mo>
                                      <mml:mi>μ</mml:mi>
                                    </mml:mrow>
                                    <mml:mo>)</mml:mo>
                                  </mml:mrow>
                                </mml:mrow>
                                <mml:mo>/</mml:mo>
                                <mml:mi>σ</mml:mi>
                              </mml:mrow>
                            </mml:mrow>
                            <mml:mo>)</mml:mo>
                          </mml:mrow>
                          <mml:mo>−</mml:mo>
                          <mml:mi>Φ</mml:mi>
                          <mml:mrow>
                            <mml:mo>(</mml:mo>
                            <mml:mrow>
                              <mml:mrow>
                                <mml:mrow>
                                  <mml:mrow>
                                    <mml:mo>(</mml:mo>
                                    <mml:mrow>
                                      <mml:mi>log</mml:mi>
                                      <mml:mi>y</mml:mi>
                                      <mml:mo>−</mml:mo>
                                      <mml:mi>μ</mml:mi>
                                    </mml:mrow>
                                    <mml:mo>)</mml:mo>
                                  </mml:mrow>
                                </mml:mrow>
                                <mml:mo>/</mml:mo>
                                <mml:mi>σ</mml:mi>
                              </mml:mrow>
                            </mml:mrow>
                            <mml:mo>)</mml:mo>
                          </mml:mrow>
                        </mml:mrow>
                        <mml:mo>|</mml:mo>
                      </mml:mrow>
                    </mml:mrow>
                  </mml:mrow>
                </mml:mstyle>
              </mml:mtd>
            </mml:mtr>
            <mml:mtr>
              <mml:mtd>
                <mml:mtext>
                   
                </mml:mtext>
                <mml:mtext>
                   
                </mml:mtext>
                <mml:mtext>
                   
                </mml:mtext>
                <mml:mtext>
                   
                </mml:mtext>
                <mml:mtext>
                   
                </mml:mtext>
                <mml:mtext>
                   
                </mml:mtext>
                <mml:mtext>
                   
                </mml:mtext>
                <mml:mo>⋅</mml:mo>
                <mml:mrow>
                  <mml:mo>[</mml:mo>
                  <mml:mrow>
                    <mml:mi>f</mml:mi>
                    <mml:mrow>
                      <mml:mo>(</mml:mo>
                      <mml:mi>x</mml:mi>
                      <mml:mo>)</mml:mo>
                    </mml:mrow>
                    <mml:mi>f</mml:mi>
                    <mml:mrow>
                      <mml:mo>(</mml:mo>
                      <mml:mi>y</mml:mi>
                      <mml:mo>)</mml:mo>
                    </mml:mrow>
                    <mml:mo>−</mml:mo>
                    <mml:msub>
                      <mml:mi>f</mml:mi>
                      <mml:mi>T</mml:mi>
                    </mml:msub>
                    <mml:mrow>
                      <mml:mo>(</mml:mo>
                      <mml:mi>x</mml:mi>
                      <mml:mo>)</mml:mo>
                    </mml:mrow>
                    <mml:msub>
                      <mml:mi>f</mml:mi>
                      <mml:mi>T</mml:mi>
                    </mml:msub>
                    <mml:mrow>
                      <mml:mo>(</mml:mo>
                      <mml:mi>y</mml:mi>
                      <mml:mo>)</mml:mo>
                    </mml:mrow>
                  </mml:mrow>
                  <mml:mo>]</mml:mo>
                </mml:mrow>
                <mml:mtext>d</mml:mtext>
                <mml:mi>x</mml:mi>
                <mml:mtext>d</mml:mtext>
                <mml:mi>y</mml:mi>
              </mml:mtd>
            </mml:mtr>
          </mml:mtable>
        </mml:math>
      </disp-formula>
      <p>where <italic>f</italic><italic><sub>T</sub></italic> denotes the truncated density. In general, Δ<sub>0</sub> does not admit a closed-form expression and must be evaluated numerically; however, it vanishes as the truncation interval expands to the full support, since <italic>f</italic><italic><sub>T</sub></italic> → <italic>f</italic>. It ensures that Δ<italic><sub>T</sub></italic> converges correctly to the non-truncated mean difference Δ(<italic>μ</italic>, <italic>σ</italic>) of Formula (10) when the truncation limits tend respectively to 0 and +∞, that is, when <italic>α</italic>* → −∞ and <italic>β</italic>* → +∞. In that limit, indeed, Φ(<italic>β</italic>*) − Φ(<italic>α</italic>*) → 1 and the corrective term vanishes, recovering exactly Formula (10).</p>
      <p>Formula (13) thus extends the fundamental Girone-Manca result to the case of the truncated lognormal distribution, maintaining the analytical structure based on the error function that characterizes the original formulation.</p>
    </sec>
    <sec id="sec7">
      <title>7. Applications in Medical Sciences</title>
      <p>The extensions of the lognormal mean difference presented in the preceding sections find numerous and significant applications in the field of medical sciences, where the lognormal distribution represents one of the most frequently employed probabilistic models for describing the variability of biological phenomena [<xref ref-type="bibr" rid="B15">15</xref>].</p>
      <sec id="sec7dot1">
        <title>7.1. Biomarkers and Biochemical Concentrations</title>
        <p>Many biological variables follow a lognormal distribution: serum concentrations of proteins (such as C-reactive protein, CRP), hepatic enzymes (ALT, AST), hormones (cortisol, testosterone, TSH), antibodies (IgG, IgM, IgA), blood glucose levels, plasma drug concentrations, and numerous other biomarkers [<xref ref-type="bibr" rid="B15">15</xref>]. The generalized mean difference (10) allows quantifying inter-individual variability in these measurements, providing a robust measure of dispersion that, unlike the variance, is not excessively influenced by the extreme values typically present in biochemical distributions.</p>
        <p>In particular, in laboratory practice, reference intervals for biomarkers are typically defined as percentiles of the distribution in the healthy population. The conditional mean difference (12), calculated over the reference interval, provides a measure of “physiological” variability that excludes pathological values and allows more appropriate comparisons between different populations.</p>
        <p>Among the biomarkers that most frequently exhibit a lognormal distribution, prostate-specific antigen (PSA) represents an emblematic case. In the healthy male population, serum PSA concentrations show a marked positive skewness, with a right tail extending toward elevated values associated with pathological conditions such as benign prostatic hyperplasia and prostate carcinoma. The conventional diagnostic threshold of 4.0 ng/mL, used as a cut-off value for oncological screening, is located in the upper portion of the lognormal distribution of normal values, and the mean difference Δ calculated on the distribution of sub-threshold values provides a measure of physiological variability that can be employed to refine diagnostic criteria as a function of the patient’s age and ethnicity.</p>
        <p>Similarly, cardiac troponin (cTnI and cTnT), used as a primary biomarker for the diagnosis of acute myocardial infarction, exhibits a lognormal distribution in the general population, with extremely low values in the majority of healthy subjects and increases of several orders of magnitude in the presence of myocardial damage. Serum creatinine levels, a fundamental indicator of renal function, follow an analogous distributional pattern, with the mean difference being particularly informative in quantifying inter-individual variability while accounting for the multiplicative nature of glomerular filtration processes. The hepatic enzymes alanine aminotransferase (ALT) and aspartate aminotransferase (AST) also follow a lognormal distribution, as do the tumor markers CA-125, employed in the monitoring of ovarian carcinoma, and alpha-fetoprotein (AFP), used in the surveillance of hepatocellular carcinoma.</p>
        <p>The fundamental reason why these biomarkers follow a lognormal distribution lies in the nature of the underlying biological processes. Biochemical concentrations in plasma are the result of cascades of enzymatic reactions, processes of synthesis, secretion, distribution, metabolism, and elimination that operate in a multiplicative rather than additive manner. According to the multiplicative central limit theorem, the product of a large number of independent positive random variables converges to a lognormal distribution, regardless of the distribution of the individual component variables [<xref ref-type="bibr" rid="B15">15</xref>]. This principle explains the ubiquity of the lognormal distribution in biomedical sciences: each metabolic step modifies the concentration by a multiplicative factor, and the overall effect is a log-normal distribution of the values observed in the population.</p>
        <p>From a clinical standpoint, the use of the mean difference Δ in place of the standard deviation offers specific advantages in characterizing biomarker variability. The standard deviation, calculated on the original scale of concentrations, is heavily influenced by extreme values and provides distorted information when the distribution is markedly skewed. The mean difference, being based on the absolute differences between all pairs of observations, is inherently more robust and captures the effective dispersion of the data in a manner more representative of clinically relevant variability. Furthermore, the ratio Δ/<italic>E</italic>[<italic>X</italic>], where <italic>E</italic>[<italic>X</italic>] denotes the mean of the distribution, provides an alternative coefficient of variability to the classical coefficient of variation CV = SD(<italic>X</italic>)/<italic>E</italic>[<italic>X</italic>], with superior robustness properties for heavy-tailed distributions, such as the lognormal.</p>
        <p>The definition of diagnostic thresholds and reference intervals benefits significantly from the analytical formulation proposed here. Reference intervals, traditionally constructed as the central 95th percentile of the distribution of values in the healthy population, can be supplemented with the conditional mean difference (12) calculated over the interval itself, yielding a more complete characterization of intra-interval variability. This approach is particularly useful in comparing different populations (by sex, age, ethnicity, or physiological conditions such as pregnancy), where not only the limits of the reference interval, but also the structure of the internal variability may differ in a clinically significant manner.</p>
        <p>To illustrate the practical relevance of the proposed formulation, consider a biomarker whose distribution in a reference population can be reasonably approximated by a lognormal distribution with parameters <italic>μ</italic> = 0 and <italic>σ</italic> = 0.5.</p>
        <p>Using the generalized formula for the mean difference, we obtain:</p>
        <disp-formula id="FD16">
          <label>(15)</label>
          <mml:math>
            <mml:mrow>
              <mml:mi>Δ</mml:mi>
              <mml:mrow>
                <mml:mo>(</mml:mo>
                <mml:mrow>
                  <mml:mi>μ</mml:mi>
                  <mml:mo>,</mml:mo>
                  <mml:mi>σ</mml:mi>
                </mml:mrow>
                <mml:mo>)</mml:mo>
              </mml:mrow>
              <mml:mo>=</mml:mo>
              <mml:mn>2</mml:mn>
              <mml:mi>exp</mml:mi>
              <mml:mrow>
                <mml:mo>(</mml:mo>
                <mml:mrow>
                  <mml:mi>μ</mml:mi>
                  <mml:mo>+</mml:mo>
                  <mml:mfrac>
                    <mml:mrow>
                      <mml:msup>
                        <mml:mi>σ</mml:mi>
                        <mml:mn>2</mml:mn>
                      </mml:msup>
                    </mml:mrow>
                    <mml:mn>2</mml:mn>
                  </mml:mfrac>
                </mml:mrow>
                <mml:mo>)</mml:mo>
              </mml:mrow>
              <mml:mo>⋅</mml:mo>
              <mml:mtext>erf</mml:mtext>
              <mml:mrow>
                <mml:mo>(</mml:mo>
                <mml:mrow>
                  <mml:mfrac>
                    <mml:mi>σ</mml:mi>
                    <mml:mn>2</mml:mn>
                  </mml:mfrac>
                </mml:mrow>
                <mml:mo>)</mml:mo>
              </mml:mrow>
            </mml:mrow>
          </mml:math>
        </disp-formula>
        <p>which yields Δ ≈ 0.84. Since the mean of the distribution is <inline-formula><mml:math display="inline"><mml:mrow><mml:mi> E </mml:mi><mml:mrow><mml:mo> [ </mml:mo><mml:mi> X </mml:mi><mml:mo> ] </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:msup><mml:mtext> e </mml:mtext><mml:mrow><mml:mi> μ </mml:mi><mml:mo> + </mml:mo><mml:mrow><mml:mrow><mml:msup><mml:mi> σ </mml:mi><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow></mml:mrow></mml:msup><mml:mo> ≈ </mml:mo></mml:mrow></mml:math></inline-formula> 1.13, the ratio Δ/<italic>E</italic>[<italic>X</italic>] ≈ 0.74 provides a normalized measure of variability.</p>
        <p>For comparison, consider a more dispersed scenario with <italic>σ</italic> = 1. In this case, the mean increases to <italic>E</italic>[<italic>X</italic>] ≈ 1.65, while the mean difference rises to Δ ≈ 2.05, leading to Δ/<italic>E</italic>[<italic>X</italic>] ≈ 1.24.</p>
        <p>These results highlight how the mean difference captures the rapid increase in variability associated with higher dispersion in lognormal models, reflecting the growing asymmetry and heavy-tailed behavior typical of biological measurements.</p>
      </sec>
      <sec id="sec7dot2">
        <title>7.2. Pharmacological Dosages and Pharmacokinetics</title>
        <p>The pharmacokinetics of many drugs is described by lognormal models for plasma concentrations [<xref ref-type="bibr" rid="B16">16</xref>]. Inter-individual variability in pharmacokinetics is a crucial factor in the design of dosing regimens and in the evaluation of drug safety.</p>
        <p>The conditional mean difference (12) finds specific application in the evaluation of the variability of concentrations within the therapeutic window, that is, the concentration interval between the minimum effective concentration (MEC) and the minimum toxic concentration (MTC). The analytical quantification of this intra-therapeutic variability is of fundamental importance for the optimization of dosing regimens and for the personalization of pharmacological therapy.</p>
        <p>The fundamental pharmacokinetic parameters—the maximum plasma concentration (Cmax), the area under the concentration-time curve (AUC), systemic clearance (CL), the volume of distribution (Vd), and the elimination half-life (t<sub>1</sub>/<sub>2</sub>)—typically exhibit a lognormal distribution in the population [<xref ref-type="bibr" rid="B16">16</xref>]. This distributional property is a direct consequence of the nature of pharmacokinetic processes: gastrointestinal absorption, hepatic first-pass metabolism, plasma protein binding, and glomerular filtration operate as multiplicative factors on the drug concentration, generating inter-individual variability that is expressed on a logarithmic scale. In particular, the hepatic first-pass effect, mediated by the cytochrome P450 enzyme system, introduces substantial variability due to the genetic polymorphism of metabolizing enzymes (CYP2D6, CYP3A4, CYP2C19), with differences between poor, intermediate, extensive, and ultra-rapid metabolizers that translate into variations in Cmax and AUC by a factor ranging from 2 to over 10.</p>
        <p>The generalized mean difference (10) finds direct application in the quantification of inter-individual variability of these pharmacokinetic parameters. For a drug with log-normally distributed bioavailability parameters, the formula Δ = <inline-formula><mml:math display="inline"><mml:mrow><mml:mn> 2 </mml:mn><mml:msup><mml:mtext> e </mml:mtext><mml:mrow><mml:mi> μ </mml:mi><mml:mo> + </mml:mo><mml:mrow><mml:mrow><mml:msup><mml:mi> σ </mml:mi><mml:mn> 2 </mml:mn></mml:msup></mml:mrow><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow></mml:mrow></mml:msup><mml:mo> ⋅ </mml:mo><mml:mtext> erf </mml:mtext><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mrow><mml:mi> σ </mml:mi><mml:mo> / </mml:mo><mml:mn> 2 </mml:mn></mml:mrow></mml:mrow><mml:mo> ) </mml:mo></mml:mrow></mml:mrow></mml:math></inline-formula> allows the analytical calculation of the expected mean difference between two randomly selected subjects from the population, providing complementary and more intuitive information compared to the coefficient of variation traditionally reported in pharmacokinetic studies. This is particularly relevant for drugs with a narrow therapeutic index (warfarin, digoxin, phenytoin, lithium, cyclosporine, tacrolimus), where small concentration variations can determine the transition from therapeutic efficacy to toxicity or inefficacy.</p>
        <p>In the context of Therapeutic Drug Monitoring (TDM), the mean difference conditioned on the therapeutic window [MEC, MTC] assumes a central role. TDM involves the periodic measurement of plasma drug concentrations with the objective of maintaining levels within the therapeutic range. Formula (12), applied to the interval [MEC, MTC], allows quantifying the expected intra-therapeutic variability in the population, information that is essential for the design of monitoring protocols and for defining the optimal frequency of blood sampling. Furthermore, the comparison between the mean difference over the entire distribution and the mean difference conditioned on the therapeutic window provides a quantitative measure of the proportion of clinically relevant variability, distinguishing it from the total variability that includes sub-therapeutic and supra-therapeutic values.</p>
        <p>Bioequivalence studies, which are fundamental for the regulatory approval of generic drugs, are explicitly based on the assumption of lognormality of pharmacokinetic parameters. The guidelines of the FDA (Food and Drug Administration) and the EMA (European Medicines Agency) require that the ratio of pharmacokinetic parameters (Cmax and AUC) of the generic drug to the reference drug be analyzed on a logarithmic scale, with 90% confidence intervals that must fall within the 80% - 125% range. In this context, the mean difference Δ on the logarithmic scale of the parameters offers a complementary indicator of the variability of the bioequivalence ratio, while the truncated lognormal Formula (13) can be employed to characterize the variability when data are subject to analytical quantification limits or when pharmacokinetic outliers are excluded from the dataset.</p>
        <p>Population pharmacokinetics (PopPK), which employs nonlinear mixed-effects models to describe inter- and intra-individual variability in the pharmacokinetic response, provides estimates of the location parameter <italic>μ</italic> and the scale parameter <italic>σ</italic> of the lognormal distribution of pharmacokinetic parameters in the population. These estimated parameters can be directly inserted into Formula (10) to obtain the expected mean difference in the patient population studied. The PopPK approach also identifies covariates (body weight, renal and hepatic function, age, sex, metabolic genotype) that explain part of the inter-individual variability, and the mean difference can be calculated conditionally on the subgroups defined by these covariates, allowing a more precise quantification of residual variability and supporting personalized dosing strategies based on the individual patient’s characteristics.</p>
        <p>Consider a pharmacokinetic variable (e.g., plasma concentration) following a lognormal distribution with parameters <italic>μ</italic> = 1 and <italic>σ</italic> = 0.6. Suppose that the therapeutic window is defined by the interval [<italic>a</italic>, <italic>b</italic>] = [1, 5].</p>
        <p>The conditional mean difference over this interval can be evaluated using the proposed formulation based on the cumulative distribution function. Numerical evaluation shows that the conditional mean difference is substantially lower than the unconditional one, reflecting the restriction to clinically relevant values.</p>
        <p>This reduction quantifies the extent to which variability outside the therapeutic window contributes to overall dispersion. From a clinical perspective, this provides a meaningful measure of intra-therapeutic variability, which may be useful in the design of dosing strategies and monitoring protocols.</p>
        <p>More generally, the comparison between conditional and unconditional mean differences offers a quantitative tool to distinguish clinically relevant variability from extreme or non-informative observations.</p>
      </sec>
    </sec>
    <sec id="sec8">
      <title>8. Conclusions</title>
      <p>In this work, four significant extensions of the mean difference formula for the lognormal distribution obtained by Girone and Manca [<xref ref-type="bibr" rid="B13">13</xref>] have been presented. The analysis of the asymptotic behavior has shown that the mean difference tends to zero for <italic>γ</italic> → 0 (consistently with the degeneration of the distribution toward a degenerate distribution) with a linear convergence rate proportional to <inline-formula><mml:math display="inline"><mml:mrow><mml:mn> 2 </mml:mn><mml:msqrt><mml:mrow><mml:mrow><mml:mn> 2 </mml:mn><mml:mo> / </mml:mo><mml:mi> π </mml:mi></mml:mrow></mml:mrow></mml:msqrt><mml:mtext>   </mml:mtext><mml:mi> γ </mml:mi></mml:mrow></mml:math></inline-formula> , and diverges to +∞ for <italic>γ</italic> → +∞ (consistently with the divergence of the scale of the distribution) with a ratio Δ/<italic>E</italic>[<italic>X</italic>] that tends to 2 (where Δ here denotes the standardized mean difference).</p>
      <p>The generalized formula with complete parameters <italic>μ</italic> and <italic>σ</italic>, obtained through the homogeneity property of the mean difference, extends the original result to the non-standardized lognormal model, making the formula directly applicable to empirical data parameterized in the usual form. The expression for the Gini index G = erf(<italic>σ</italic>/2) that follows confirms well-known results in the literature on concentration.</p>
      <p>The conditional mean difference and the formula for the truncated lognormal broaden the field of application to practical contexts where data are naturally limited or segmented. In particular, the formula for the truncated distribution, which extends the Girone-Manca result while maintaining the analytical structure based on the error function, represents a useful tool for the analysis of censored or truncated data.</p>
      <p>The applications in medical sciences illustrate the practical relevance of these analytical results in fields ranging from clinical biochemistry to pharmacokinetics. The formulas derived here preserve the elegance and simplicity of the original formulation, confirming the centrality of the error function in the characterization of the variability of the lognormal distribution.</p>
    </sec>
    <sec id="sec9">
      <title>9. Limitations and Future Research</title>
      <p>Despite the analytical generality of the results, some limitations of the present study should be acknowledged. The proposed formulations are derived under the assumption of exact lognormality, which, although widely supported in many applied contexts, may not fully capture deviations observed in empirical data. In addition, the results are obtained within a univariate framework and do not directly extend to multivariate settings, where dependence structures may play a relevant role.</p>
      <p>Future research may address these limitations by extending the analysis to multivariate lognormal models and to alternative heavy-tailed distributions. Further developments could also include the empirical validation of the proposed measures on real datasets, as well as the comparison with other dispersion and inequality indicators in applied contexts. Such extensions would contribute to a deeper understanding of variability in complex stochastic systems. </p>
    </sec>
    <sec id="sec10">
      <title>NOTES</title>
      <p><sup>1</sup>In Formula (1), the parameter <italic>γ</italic> &gt; 0 denotes the standard deviation of the logarithm of the variable, <italic>i.e.</italic>, <italic>γ</italic> = sd(log<italic>X</italic>). This is the sole parameter of the standardized lognormal distribution (with location parameter <italic>μ</italic> = 0). In the general formulation of Section 4, this parameter is denoted <italic>σ</italic>, with <italic>γ</italic> ≡ <italic>σ</italic> when <italic>μ</italic> = 0. The cumulative distribution function <italic>F</italic>(<italic>x</italic>) and the support endpoints (<italic>a</italic>, <italic>b</italic>) appearing in Formula (3) refer to the general support of the distribution; for the standard lognormal, <italic>a</italic> = 0 and <italic>b</italic> = +∞.</p>
      <p><sup>2</sup>Let <italic>X</italic> ~ Lognormal(<italic>μ</italic>, <italic>σ</italic><sup>2</sup>). Then the mean difference of <italic>X</italic> is given by: Δ(<italic>μ</italic>, <italic>σ</italic>) = 2exp(<italic>μ</italic> + <italic>σ</italic><sup>2</sup>/2)∙erf(<italic>σ</italic>/2), as stated in Formula (10) below. Since <italic>X</italic> = exp(<italic>μ</italic> + <italic>σ</italic><italic>Z</italic>) = exp(<italic>μ</italic>)∙exp(<italic>σ</italic><italic>Z</italic>) where <italic>Z</italic> ~ <italic>N</italic>(0,1), and <italic>W</italic> = exp(<italic>σ</italic><italic>Z</italic>) follows a standardized lognormal distribution, the homogeneity property Δ(<italic>X</italic>) = exp(<italic>μ</italic>)∙Δ(<italic>W</italic>) together with Formula (1) yields the result.</p>
      <p><sup>3</sup>Let <italic>X</italic> be a continuous random variable with CDF <italic>F</italic>(<italic>x</italic>) on support (<italic>a</italic><sub>0</sub>, <italic>b</italic><sub>0</sub>). The mean difference of <italic>X</italic> conditioned on the interval [<italic>a</italic>, <italic>b</italic>] ⊆ (<italic>a</italic><sub>0</sub>, <italic>b</italic><sub>0</sub>) is: <inline-formula><mml:math display="inline"><mml:mrow><mml:msub><mml:mi> Δ </mml:mi><mml:mi> c </mml:mi></mml:msub><mml:mrow><mml:mo> [ </mml:mo><mml:mrow><mml:mi> a </mml:mi><mml:mo> , </mml:mo><mml:mi> b </mml:mi></mml:mrow><mml:mo> ] </mml:mo></mml:mrow><mml:mo> = </mml:mo><mml:mrow><mml:mo> ( </mml:mo><mml:mrow><mml:mrow><mml:mn> 4 </mml:mn><mml:mo> / </mml:mo><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo> [ </mml:mo><mml:mrow><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> b </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> − </mml:mo><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> a </mml:mi><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mo> ] </mml:mo></mml:mrow></mml:mrow><mml:mn> 2 </mml:mn></mml:msup></mml:mrow></mml:mrow></mml:mrow><mml:mo> ) </mml:mo></mml:mrow><mml:mstyle displaystyle="true"><mml:mrow><mml:msubsup><mml:mo> ∫ </mml:mo><mml:mi> a </mml:mi><mml:mi> b </mml:mi></mml:msubsup><mml:mrow><mml:mrow><mml:mo> [ </mml:mo><mml:mrow><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> − </mml:mo><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> a </mml:mi><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mo> ] </mml:mo></mml:mrow><mml:mrow><mml:mo> [ </mml:mo><mml:mrow><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> b </mml:mi><mml:mo> ) </mml:mo></mml:mrow><mml:mo> − </mml:mo><mml:mi> F </mml:mi><mml:mrow><mml:mo> ( </mml:mo><mml:mi> x </mml:mi><mml:mo> ) </mml:mo></mml:mrow></mml:mrow><mml:mo> ] </mml:mo></mml:mrow><mml:mtext> d </mml:mtext><mml:mi> x </mml:mi></mml:mrow></mml:mrow></mml:mstyle></mml:mrow></mml:math></inline-formula> , as given in Formula (12) below. The result follows from applying the integral representation (3) to the conditional distribution of <italic>X</italic> given <italic>a</italic> ≤ <italic>X</italic> ≤ <italic>b</italic>, whose CDF is <italic>F</italic><sub>[</sub><italic><sub>a</sub></italic><sub>,</sub><italic><sub>b</sub></italic><sub>]</sub>(<italic>x</italic>) = [<italic>F</italic>(<italic>x</italic>) − <italic>F</italic>(<italic>a</italic>)]/[<italic>F</italic>(<italic>b</italic>) − <italic>F</italic>(<italic>a</italic>)], and simplifying.</p>
      <p><sup>4</sup>Let <italic>X</italic> ~ Lognormal(<italic>μ</italic>, <italic>σ</italic><sup>2</sup>) truncated to the interval [<italic>a</italic>, <italic>b</italic>] with 0 &lt; <italic>a</italic> &lt; <italic>b</italic>. Then the mean difference of the truncated distribution is given by Formula (13) below, expressed in terms of the standardized truncation limits <italic>α</italic> = (log<italic>a</italic> − <italic>μ</italic>)/<italic>σ</italic> and <italic>β</italic> = (log<italic>b</italic> − <italic>μ</italic>)/<italic>σ</italic>, the standard normal CDF Φ(∙), and a boundary correction term Δ<sub>0</sub>. The result is obtained by specializing Proposition 2 to the lognormal CDF <italic>F</italic>(<italic>x</italic>; <italic>μ</italic>, <italic>σ</italic>) = Φ((log<italic>x</italic> − <italic>μ</italic>)/<italic>σ</italic>) and evaluating the resulting integral using the substitution <italic>u</italic> = (log<italic>x</italic> − <italic>μ</italic>)/<italic>σ</italic>, together with known integrals of the error function [<xref ref-type="bibr" rid="B9">9</xref>].</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <title>References</title>
      <ref id="B1">
        <label>1.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Gini, C. (1912) Variabilità e Mutabilità. Contributo allo studio delle distribuzioni e delle relazioni statistiche. Tipografia di Paolo Cuppini.</mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Gini, C.</string-name>
            </person-group>
            <year>1912</year>
            <article-title>Variabilità e Mutabilità</article-title>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B2">
        <label>2.</label>
        <citation-alternatives>
          <mixed-citation publication-type="journal">Yitzhaki, S. (2003) Gini’s Mean Difference: A Superior Measure of Variability for Non-Normal Distributions. <italic>METRON</italic>— <italic>International Journal of Statistics</italic>, 61, 285-316.</mixed-citation>
          <element-citation publication-type="journal">
            <person-group person-group-type="author">
              <string-name>Yitzhaki, S.</string-name>
            </person-group>
            <year>2003</year>
            <article-title>Gini’s Mean Difference: A Superior Measure of Variability for Non-Normal Distributions</article-title>
            <source>METRON—International Journal of Statistics</source>
            <volume>61</volume>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B3">
        <label>3.</label>
        <citation-alternatives>
          <mixed-citation publication-type="journal">Giorgi, G.M. ang Gigliarano, C. (2017) The Gini Concentration Index: A Review of the Inference Literature. <italic>Journal of Economic Surveys</italic>, 31, 1130-1148. https://doi.org/10.1111/joes.12185 <pub-id pub-id-type="doi">10.1111/joes.12185</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1111/joes.12185">https://doi.org/10.1111/joes.12185</ext-link></mixed-citation>
          <element-citation publication-type="journal">
            <person-group person-group-type="author">
              <string-name>Giorgi, G.M.</string-name>
              <string-name>Gigliarano, C.</string-name>
            </person-group>
            <year>2017</year>
            <article-title>The Gini Concentration Index: A Review of the Inference Literature</article-title>
            <source>Journal of Economic Surveys</source>
            <volume>31</volume>
            <pub-id pub-id-type="doi">10.1111/joes.12185</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B4">
        <label>4.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Ebert, U. (2010) The Decomposition of Inequality Reconsidered: Weakly Decomposable Measures. <italic>Mathematical Social Sciences</italic>, 60, 94-103. https://doi.org/10.1016/j.mathsocsci.2010.05.001 <pub-id pub-id-type="doi">10.1016/j.mathsocsci.2010.05.001</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1016/j.mathsocsci.2010.05.001">https://doi.org/10.1016/j.mathsocsci.2010.05.001</ext-link></mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Ebert, U.</string-name>
            </person-group>
            <year>2010</year>
            <article-title>The Decomposition of Inequality Reconsidered: Weakly Decomposable Measures</article-title>
            <source>Mathematical Social Sciences</source>
            <volume>60</volume>
            <pub-id pub-id-type="doi">10.1016/j.mathsocsci.2010.05.001</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B5">
        <label>5.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Novi Inverardi, P.L. and Tagliani, A. (2024) The Lognormal Distribution Is Characterized by Its Integer Moments. <italic>Mathematics</italic>, 12, Article 3830. https://doi.org/10.3390/math12233830 <pub-id pub-id-type="doi">10.3390/math12233830</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3390/math12233830">https://doi.org/10.3390/math12233830</ext-link></mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Inverardi, P.L.</string-name>
              <string-name>Tagliani, A.</string-name>
            </person-group>
            <year>2024</year>
            <article-title>The Lognormal Distribution Is Characterized by Its Integer Moments</article-title>
            <source>Mathematics</source>
            <volume>12</volume>
            <elocation-id>3830</elocation-id>
            <pub-id pub-id-type="doi">10.3390/math12233830</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B6">
        <label>6.</label>
        <citation-alternatives>
          <mixed-citation publication-type="journal">Dai, P. and Shen, S. (2025) Estimation of the Gini Coefficient Based on Two Quantiles. <italic>PLOS ONE</italic>, 20, e0318833. https://doi.org/10.1371/journal.pone.0318833 <pub-id pub-id-type="doi">10.1371/journal.pone.0318833</pub-id><pub-id pub-id-type="pmid">39932932</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1371/journal.pone.0318833">https://doi.org/10.1371/journal.pone.0318833</ext-link></mixed-citation>
          <element-citation publication-type="journal">
            <person-group person-group-type="author">
              <string-name>Dai, P.</string-name>
              <string-name>Shen, S.</string-name>
            </person-group>
            <year>2025</year>
            <article-title>Estimation of the Gini Coefficient Based on Two Quantiles</article-title>
            <source>PLOS ONE</source>
            <volume>20</volume>
            <pub-id pub-id-type="doi">10.1371/journal.pone.0318833</pub-id>
            <pub-id pub-id-type="pmid">39932932</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B7">
        <label>7.</label>
        <citation-alternatives>
          <mixed-citation publication-type="journal">Poudyal, C., Zhao, Q. and Brazauskas, V. (2023) Method of Winsorized Moments for Robust Fitting of Truncated and Censored Lognormal Distributions. <italic>North American Actuarial Journal</italic>, 28, 236-260. https://doi.org/10.1080/10920277.2023.2183869 <pub-id pub-id-type="doi">10.1080/10920277.2023.2183869</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1080/10920277.2023.2183869">https://doi.org/10.1080/10920277.2023.2183869</ext-link></mixed-citation>
          <element-citation publication-type="journal">
            <person-group person-group-type="author">
              <string-name>Poudyal, C.</string-name>
              <string-name>Zhao, Q.</string-name>
              <string-name>Brazauskas, V.</string-name>
            </person-group>
            <year>2023</year>
            <article-title>Method of Winsorized Moments for Robust Fitting of Truncated and Censored Lognormal Distributions</article-title>
            <source>North American Actuarial Journal</source>
            <volume>28</volume>
            <pub-id pub-id-type="doi">10.1080/10920277.2023.2183869</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B8">
        <label>8.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Kapera, M. and Kobus, M. (2024) The Gini and Mean Log Deviation Indices of Multivariate Inequality of Opportunity. <italic>Econometrics</italic>, 12, Article 10. https://doi.org/10.3390/econometrics12020010 <pub-id pub-id-type="doi">10.3390/econometrics12020010</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3390/econometrics12020010">https://doi.org/10.3390/econometrics12020010</ext-link></mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Kapera, M.</string-name>
              <string-name>Kobus, M.</string-name>
            </person-group>
            <year>2024</year>
            <article-title>The Gini and Mean Log Deviation Indices of Multivariate Inequality of Opportunity</article-title>
            <source>Econometrics</source>
            <volume>12</volume>
            <elocation-id>10</elocation-id>
            <pub-id pub-id-type="doi">10.3390/econometrics12020010</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B9">
        <label>9.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Jokiel-Rokita, A. and Pia̧tek, S. (2022) Estimation of Parameters and Quantiles of the Weibull Distribution. <italic>Statistical Papers</italic>, 65, 1-18. https://doi.org/10.1007/s00362-022-01379-9 <pub-id pub-id-type="doi">10.1007/s00362-022-01379-9</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1007/s00362-022-01379-9">https://doi.org/10.1007/s00362-022-01379-9</ext-link></mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Jokiel-Rokita, A.</string-name>
            </person-group>
            <year>2022</year>
            <article-title>Estimation of Parameters and Quantiles of the Weibull Distribution</article-title>
            <source>Statistical Papers</source>
            <volume>65</volume>
            <pub-id pub-id-type="doi">10.1007/s00362-022-01379-9</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B10">
        <label>10.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Girone, G. and D’Uggento, A.M. (2016) About the Mean Difference of the Inverse Normal Distribution. <italic>Applied Mathematics</italic>, 7, 1504-1509. https://doi.org/10.4236/am.2016.714130 <pub-id pub-id-type="doi">10.4236/am.2016.714130</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.4236/am.2016.714130">https://doi.org/10.4236/am.2016.714130</ext-link></mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Girone, G.</string-name>
              <string-name>Uggento, A.M.</string-name>
            </person-group>
            <year>2016</year>
            <article-title>About the Mean Difference of the Inverse Normal Distribution</article-title>
            <source>Applied Mathematics</source>
            <volume>7</volume>
            <pub-id pub-id-type="doi">10.4236/am.2016.714130</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B11">
        <label>11.</label>
        <citation-alternatives>
          <mixed-citation publication-type="confproc">Girone, G., Massari, A. and Mazzitelli, D. (2015) More on the Mean Difference of Continuous Distributive Models. <italic>Proceedings SIS Conference</italic> “ <italic>Statistics and Demography</italic>: <italic>The Legacy of Corrado Gini</italic>”, Treviso, 9-11 September 2015. https://www.scirp.org/reference/referencespapers?referenceid=1765353</mixed-citation>
          <element-citation publication-type="confproc">
            <person-group person-group-type="author">
              <string-name>Girone, G.</string-name>
              <string-name>Massari, A.</string-name>
              <string-name>Mazzitelli, D.</string-name>
            </person-group>
            <year>2015</year>
            <article-title>More on the Mean Difference of Continuous Distributive Models</article-title>
            <source>Proceedings SIS Conference “Statistics and Demography: The Legacy of Corrado Gini”</source>
            <volume>9</volume>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B12">
        <label>12.</label>
        <citation-alternatives>
          <mixed-citation publication-type="confproc">Girone, G., Manca, F. and D’Uggento, A.M. (2015) The Mean Difference of Discrete Distribution Models. <italic>Proceedings SIS Conference</italic>“ <italic>Statistics and Demography</italic>: <italic>The Legacy of Corrado Gini</italic>”, Treviso, 9-11 September 2015, 1-6.</mixed-citation>
          <element-citation publication-type="confproc">
            <person-group person-group-type="author">
              <string-name>Girone, G.</string-name>
              <string-name>Manca, F.</string-name>
              <string-name>Uggento, A.M.</string-name>
            </person-group>
            <year>2015</year>
            <article-title>The Mean Difference of Discrete Distribution Models</article-title>
            <source>Proceedings SIS Conference “Statistics and Demography: The Legacy of Corrado Gini”</source>
            <volume>9</volume>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B13">
        <label>13.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Girone, G. and Manca, F. (2016) The Mean Difference for Lognormal Distribution. <italic>Applied Mathematics</italic>, 7, 824-828. https://doi.org/10.4236/am.2016.79073 <pub-id pub-id-type="doi">10.4236/am.2016.79073</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.4236/am.2016.79073">https://doi.org/10.4236/am.2016.79073</ext-link></mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Girone, G.</string-name>
              <string-name>Manca, F.</string-name>
            </person-group>
            <year>2016</year>
            <article-title>The Mean Difference for Lognormal Distribution</article-title>
            <source>Applied Mathematics</source>
            <volume>7</volume>
            <pub-id pub-id-type="doi">10.4236/am.2016.79073</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B14">
        <label>14.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Prudnikov, A.P., Brychkov, Y.A. and Marichev, O.I. (1986) Integrals and Series, Vol. 2: Special Functions. Gordon and Breach Science Publishers.</mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Prudnikov, A.P.</string-name>
              <string-name>Brychkov, Y.A.</string-name>
              <string-name>Marichev, O.I.</string-name>
              <string-name>Series, V</string-name>
            </person-group>
            <year>1986</year>
            <article-title>Integrals and Series, Vol</article-title>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B15">
        <label>15.</label>
        <citation-alternatives>
          <mixed-citation publication-type="journal">Limpert, E., Stahel, W.A. and Abbt, M. (2001) Log-Normal Distributions across the Sciences: Keys and Clues. <italic>BioScience</italic>, 51, 341-352. https://doi.org/10.1641/0006-3568(2001)051[0341:lndats]2.0.co;2 <pub-id pub-id-type="doi">10.1641/0006-3568(2001)051[0341:lndats]2.0.co;2</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1641/0006-3568(2001)051[0341:lndats]2.0.co;2">https://doi.org/10.1641/0006-3568(2001)051[0341:lndats]2.0.co;2</ext-link></mixed-citation>
          <element-citation publication-type="journal">
            <person-group person-group-type="author">
              <string-name>Limpert, E.</string-name>
              <string-name>Stahel, W.A.</string-name>
              <string-name>Abbt, M.</string-name>
            </person-group>
            <year>2001</year>
            <article-title>Log-Normal Distributions across the Sciences: Keys and Clues</article-title>
            <source>BioScience</source>
            <volume>3568</volume>
            <issue>2001</issue>
            <pub-id pub-id-type="doi">10.1641/0006-3568(2001)051[0341:lndats]2.0.co;2</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B16">
        <label>16.</label>
        <citation-alternatives>
          <mixed-citation publication-type="journal">Lacey, L.F., Keene, O.N., Pritchard, J.F. and Bye, A. (1997) Common Noncompartmental Pharmacokinetic Variables: Are They Normally or Log-Normally Distributed? <italic>Journal of Biopharmaceutical Statistics</italic>, 7, 171-178. https://doi.org/10.1080/10543409708835177 <pub-id pub-id-type="doi">10.1080/10543409708835177</pub-id><pub-id pub-id-type="pmid">9056596</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1080/10543409708835177">https://doi.org/10.1080/10543409708835177</ext-link></mixed-citation>
          <element-citation publication-type="journal">
            <person-group person-group-type="author">
              <string-name>Lacey, L.F.</string-name>
              <string-name>Keene, O.N.</string-name>
              <string-name>Pritchard, J.F.</string-name>
              <string-name>Bye, A.</string-name>
            </person-group>
            <year>1997</year>
            <article-title>Common Noncompartmental Pharmacokinetic Variables: Are They Normally or Log-Normally Distributed? Journal of Biopharmaceutical Statistics, 7, 171-178</article-title>
            <pub-id pub-id-type="doi">10.1080/10543409708835177</pub-id>
            <pub-id pub-id-type="pmid">9056596</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
    </ref-list>
  </back>
</article>