<?xml version="1.0" encoding="UTF-8"?><!DOCTYPE article  PUBLIC "-//NLM//DTD Journal Publishing DTD v3.0 20080202//EN" "http://dtd.nlm.nih.gov/publishing/3.0/journalpublishing3.dtd"><article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" dtd-version="3.0" xml:lang="en" article-type="research article"><front><journal-meta><journal-id journal-id-type="publisher-id">OJS</journal-id><journal-title-group><journal-title>Open Journal of Statistics</journal-title></journal-title-group><issn pub-type="epub">2161-718X</issn><publisher><publisher-name>Scientific Research Publishing</publisher-name></publisher></journal-meta><article-meta><article-id pub-id-type="doi">10.4236/ojs.2021.116062</article-id><article-id pub-id-type="publisher-id">OJS-114207</article-id><article-categories><subj-group subj-group-type="heading"><subject>Articles</subject></subj-group><subj-group subj-group-type="Discipline-v2"><subject>Physics&amp;Mathematics</subject></subj-group></article-categories><title-group><article-title>
 
 
  Computational Identification of Confirmatory Factor Analysis Model with Simplimax Procedures
 
</article-title></title-group><contrib-group><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Jingyu</surname><given-names>Cai</given-names></name><xref ref-type="aff" rid="aff1"><sup>1</sup></xref><xref ref-type="corresp" rid="cor1"><sup>*</sup></xref></contrib><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Henk</surname><given-names>A. L. Kiers</given-names></name><xref ref-type="aff" rid="aff2"><sup>2</sup></xref></contrib><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Kohei</surname><given-names>Adachi</given-names></name><xref ref-type="aff" rid="aff1"><sup>1</sup></xref></contrib></contrib-group><aff id="aff1"><addr-line>Graduate School of Human Sciences, Osaka University, Osaka, Japan</addr-line></aff><aff id="aff2"><addr-line>Faculty of Behavioural and Social Sciences, University of Groningen, Groningen, The Netherlands</addr-line></aff><pub-date pub-type="epub"><day>29</day><month>11</month><year>2021</year></pub-date><volume>11</volume><issue>06</issue><fpage>1044</fpage><lpage>1061</lpage><history><date date-type="received"><day>15,</day>	<month>November</month>	<year>2021</year></date><date date-type="rev-recd"><day>25,</day>	<month>December</month>	<year>2021</year>	</date><date date-type="accepted"><day>28,</day>	<month>December</month>	<year>2021</year></date></history><permissions><copyright-statement>&#169; Copyright  2014 by authors and Scientific Research Publishing Inc. </copyright-statement><copyright-year>2014</copyright-year><license><license-p>This work is licensed under the Creative Commons Attribution International License (CC BY). http://creativecommons.org/licenses/by/4.0/</license-p></license></permissions><abstract><p>
 
 
  Confirmatory factor analysis (CFA) refers to the FA procedure with some loadings constrained to be zeros. A difficulty in CFA is that the constraint must be specified by users in a subjective manner. For dealing with this difficulty, we propose a computational method, in which the best CFA solution is obtained optimally without relying on users’ judgements. The method consists of the procedures at lower (L) and higher (H) levels: at the L level, for a fixed number of zero loadings, it is determined both which loadings are to be zeros and what values are to be given to the remaining nonzero parameters; at the H level, the procedure at the L level is performed over the different numbers of zero loadings, to provide the best solution. In the L level procedure, Kiers’ (1994) simplimax rotation fulfills a key role: the CFA solution under the constraint computationally specified by that rotation is used for initializing the parameters of a new FA procedure called simplimax FA. The task at the H level can be easily performed using information criteria. The usefulness of the proposed method is demonstrated numerically.
 
</p></abstract><kwd-group><kwd>Confirmatory Factor Analysis</kwd><kwd> Cardinality of Loadings</kwd><kwd> Simplimax Rotation</kwd></kwd-group></article-meta></front><body><sec id="s1"><title>1. Introduction</title><p>In factor analysis (FA), the variation of p observed variables is assumed to be explained by m common factors and p unique factors, with m &lt; p and the two types of factors mutually uncorrelated. The m common factors serve to explain the variations of all p variables. On the other hand, each of the p unique factors has a one-to-one correspondence to each variable: a unique factor explains specifically the variation of the corresponding variable that remains unaccounted for by the common factors [<xref ref-type="bibr" rid="scirp.114207-ref1">1</xref>] [<xref ref-type="bibr" rid="scirp.114207-ref2">2</xref>] [<xref ref-type="bibr" rid="scirp.114207-ref3">3</xref>].</p><p>The parameters to be estimated in FA are a factor loading matrix L = (λ<sub>ij</sub>) (p &#215; m), a unique variance matrix Y = (ψ<sub>ii</sub><sub>&#162;</sub>) (p &#215; p), and a factor correlation matrix F = (f<sub>jk</sub>) (m &#215; m). Here, the loadings in L stand for how the variables load on the common factors, Y is the diagonal matrix whose diagonal element ψ<sub>ii</sub> expresses the variance of the ith unique factor, and F contains the correlation coefficients among the m common factors. Upon making certain distributional assumptions for the factors, the covariance matrix S among p observed variables is modeled as</p><p>Σ = Λ Φ Λ ′ + Ψ . (1)</p><p>(e.g., Adachi, 2019 [<xref ref-type="bibr" rid="scirp.114207-ref1">1</xref>]; Mulaik, 2010 [<xref ref-type="bibr" rid="scirp.114207-ref2">2</xref>] ). Thus, a loss function f ( Λ , Ψ , Φ | S ) can be defined, which stands for the discrepancy between (1) and its sample counterpart S (p &#215; p): FA can be formulated as minimizing f ( Λ , Ψ , Φ | S ) over L, Y, and F, for a given S.</p><p>FA can be classified into two types; confirmatory (CFA) and exploratory (EFA) [<xref ref-type="bibr" rid="scirp.114207-ref2">2</xref>]. In CFA, some loadings in L are constrained to take zero values, while no constraints are imposed on the loadings of EFA. In this paper, we focus on CFA. An example of the CFA model with a particular constraint is illustrated in <xref ref-type="fig" rid="fig1">Figure 1</xref>. There, a constrained loading matrix L is shown on the left, and the corresponding CFA model is depicted on the right as a diagram, in which only pairs of variables and factors with unconstrained loadings are linked by paths. In <xref ref-type="fig" rid="fig1">Figure 1</xref>, the binary matrix B is also presented which specifies the constraints in L and the links in the right diagram. The elements in B = (b<sub>ij</sub>) (p &#215; m) are defined generally as</p><p>b i j = { 0 ,         iff   λ i j = 0 1 ,   otherwise , (2)</p><p>in other words, b<sub>ij</sub> = 1 if variable i is linked to factor j; otherwise, b<sub>ij</sub> = 0. In this sense, we call B a link matrix. Any CFA constraint can be expressed as Λ = B • Λ . Here, &#183; denotes the element-wise Hadamard product with B • Λ = ( b i j λ i j ) . It should be kept in mind that specifying a CFA model amounts to selecting a particular link matrix B. Thus, CFA can be formally expressed as</p><p>min Λ , Ψ , Φ f ( Λ , Ψ , Φ | S ) s.t. Λ = B • Λ for a specified B ∈ S B (3)</p><p>with “s.t.” the abbreviation for “subject to” and S B denoting a set of considered matrices B.</p><p>A problem in CFA (3) is that the link matrix B = (b<sub>ij</sub>) defined as (2) must be selected by users. That is, it must be decided in a subjective manner, which elements in B are to be zeros/ones, in other words, which pairs of variables and factors are linked as in <xref ref-type="fig" rid="fig1">Figure 1</xref> whether the CFA model with a particular constraint is accepted or not is checked afterwards by referring to the goodness-of-fit of the solution [<xref ref-type="bibr" rid="scirp.114207-ref4">4</xref>] [<xref ref-type="bibr" rid="scirp.114207-ref5">5</xref>] [<xref ref-type="bibr" rid="scirp.114207-ref6">6</xref>]. However, even if the model is found acceptable, better CFA models with other link matrices B may exist. It implies that all possible B in S B must be considered for finding the best B. However, this is unfeasible, unless pm is very small, as the number of all possible B is enormous. That number can be calculated as 2<sup>pm</sup>, since each of the pm elements in B takes zero or one as in (2): for example, for p = 12 and m = 3 one finds 2<sup>pm</sup> @ 6.87 &#215; 10<sup>10</sup>. In short, the problems in CFA can be summarized next:</p><p>[P1] The link matrix B (specifying a CFA model) must be selected subjectively by users.</p><p>[P2] An enormous number of possible matrices B in S B must be considered for finding the best B.</p><p>To the best of our knowledge, the problems [P1] and [P2] in CFA have not been considered in the existing papers. In order to deal with those problems, we propose an FA procedure for computationally and optimally identifying a suitable CFA model. Here, the model identification includes estimating the model parameter values. The outline of our proposed procedure is described in the next section. Then, we detail the procedure in Section 3, report its assessment in a simulation study in Section 4, give numerical examples in Section 5, and conclude this paper in Section 6.</p></sec><sec id="s2"><title>2. Outline of the Proposed Procedure</title><p>First, we outline our approach to the CFA model identification under the condition that the number of zero loadings is fixed to a particular integer in Section 2.1. Then, in Section 2.2, the approach is extended to the cases with the number of zero loadings not being fixed. We then summarize the prospects for the following sections.</p><sec id="s2_1"><title>2.1. Model Identification for a Specified Number of Nonzero Loadings (Cardinality)</title><p>For dealing with the difficulties [P1] and [P2] in CFA, we can consider the two procedures introduced in the next paragraphs, on the condition that Card ( Λ ) = Card ( B ) , that is, the number of nonzero values in L and B the matrix between parentheses equals a specified integer c hence</p><p>Card ( Λ ) = c , or equivalently, Card ( B ) = c . (4)</p><p>Clearly, pm - c equals the number of zeros in B or L.</p><p>The first procedure considered can be formulated as</p><p>min Λ , Ψ , Φ f ( Λ , Ψ , Φ | S ) s.t. Λ = B • Λ , after B is estimated optimally s.t. (4). (5)</p><p>This minimization is rewritten as performing CFA under the constraint indicated by link matrix B, with B estimated by another method in advance. Thus, the problem [P1] is overcome. Further, we can also consider that [P2] is dealt with, supposed that the value c in (4) and B are suitable.</p><p>The above procedure consists of two stages: first B is estimated, then CFA is performed, see (5). In contrast, the second procedure considered is formulated with a single stage as follows:</p><p>min B , Λ , Ψ , Φ f ( B , Λ , Ψ , Φ | S ) s.t. Λ = B • Λ and (4). (6)</p><p>Here, B has been added to the subscripts of “min”: the link matrix B, which indicates the pairs of variables and factors to be linked, is estimated jointly with the other parameters L, Y, and F.</p><p>Between (5) and (6) or (6&#162;), we can find the following difference: in minimizing loss function f ( Λ , Ψ , Φ | S ) , the B value is kept fixed in (5), but allowed to change in (6). The difference implies that the resulting loss function value in (6) cannot exceed that value in (5):</p><p>min B , Λ , Ψ , Φ f ( B , Λ , Ψ , Φ | S ) ≤ min Λ , Ψ , Φ f ( Λ , Ψ , Φ | S ) . (7)</p><p>This inequality shows that (6) can provide a better solution than (5). We will propose an iterative algorithm for (6), which will be started by the optimum for (5). As empirically shown later, in almost all cases, the solutions of (5) and (6) are equivalent: we can obtain the final solutions only by (5), without performing the iterative algorithm for (6). However, it is worth to perform the latter step, as (6) can provide a better solution in a few cases.</p><p>In this paper, we use the maximum likelihood (ML) method for estimating the parameters in (5) and (6). This implies that the loss function to be minimized is given as the negative of the log likelihood. It is explicitly expressed as</p><p>f ( Λ , Ψ , Φ | S ) = log | Σ | + tr   S Σ − 1 = log | Λ Φ Λ ′ + Ψ | + tr   S ( Λ Φ Λ ′ + Ψ ) − 1 , (8)</p><p>following from the normality assumptions for factors [<xref ref-type="bibr" rid="scirp.114207-ref7">7</xref>]. For minimizing (8), we use Rubin &amp; Thayer’s [<xref ref-type="bibr" rid="scirp.114207-ref8">8</xref>] EM algorithm for FA, whose properties are discussed in Adachi [<xref ref-type="bibr" rid="scirp.114207-ref9">9</xref>]. This algorithm is detailed in Appendix A1.</p></sec><sec id="s2_2"><title>2.2. Selection of the Best Cardinality</title><p>We should notice that the above approach is conditional upon c in (4). Thus, it remains to select the best value for c. This can be attained by the following procedure:</p><p>Select the value c with the lowest IC(c) among c = c min , ⋯ , c max . (9)</p><p>Here, c<sub>min</sub>/c<sub>max</sub> expresses the reasonable minimum/maximum of c, and IC(c) denotes the value of an information criterion statistic [<xref ref-type="bibr" rid="scirp.114207-ref10">10</xref>], with the statistic being a function of c. The information criteria consist of some statistics which can be defined in ML-based procedures, and they show smaller values for better solutions. Explicit expressions of IC(c) are introduced later. We can rewrite (9) as follows: the procedure of (6) following (5) (in the last subsection) is performed for c = c min , ⋯ , c max , so that we have multiple solutions, among which the solution with the least IC(c) is selected as the best one. The largest interval of [c<sub>min</sub>, c<sub>max</sub>] is obviously [1, pm]. It can be reasonably reduced as</p><p>c min = p and c max = p m − m ( m − 1 ) 2 (10)</p><p>This is because, solutions with c smaller than p would surely have one or more zero rows of loadings. On the other hand, it is considered in the c<sub>max</sub> value that m(m - 1)/2 elements in L can be set to zeros without a change in the values of loss function (8).</p><p>Now, let us discuss how our proposed procedure is related to the set S B containing all possible B, using S B ( c ) for the subset of S B that contains the matrices B with c nonzeros. The number of B in S B ( c ) is given by</p><p>N B ( c ) = C p m c , (11)</p><p>i.e., the number of the combinations of the c elements being ones among all pm elements in B. Thus, the number of all B contained in S B is given by</p><p>N B = ∑ c = c min c max N B ( c ) (12)</p><p>with (10) and (11). For example, when p = 12 and m = 3, the value in (12) is approximately 1.25 &#215; 10<sup>9</sup>, which is enormous (though less than 2<sup>pm</sup> @ 6.87 &#215; 10<sup>10</sup> presented in the last section where (10) was not yet considered). However, it is not required to assess all N<sub>B</sub> matrices B in S B in our proposed procedure, as the performance of (6) following (5) allows us to find the optimal B in S B ( c ) for a particular c, with this performance made for c = c min , ⋯ , c max as in (9). That is, our proposed procedure can find the optimal B among all N<sub>B</sub> matrices B in S B by the c max − c min + 1 runs of (6) following (5), but not by assessing all B. The example of c max − c min + 1 = 22 for p = 12 and m = 3 demonstrates how easily we can arrive at the optimal B.</p><p>The remaining parts of this paper are organized as follows: in Sect 3.1 and Sect 3.2, we detail (5) and (6) in turn. There, it is described that Kiers’ [<xref ref-type="bibr" rid="scirp.114207-ref11">11</xref>] simplimax rotation is used for the optimal estimation of B subject to (4) in (5) and the idea of this rotation also fulfills a key role in (6). Thus, the procedures for (5) and (6) are particularly called simplimax-based CFA (SbCFA) and simplimax FA (SimpFA), respectively. In the section after treating SbCFA (5) and SimpFA (6), we detail the whole process of the proposed procedure, i.e., SimpFA (6) following SbCFA (5) with the cardinality selection by (9). The proposed procedure is studied in a simulation study and illustrated with real data examples, as reported in the sections before concluding this paper.</p></sec></sec><sec id="s3"><title>3. Simplimax-Based Confirmatory Factor Analysis</title><p>In this section, we detail the procedure formulated as (5) with f ( Λ , Ψ , Φ | S ) defined as (8). A key point is that we will use exploratory FA (EFA) followed by Kiers’ [<xref ref-type="bibr" rid="scirp.114207-ref11">11</xref>] simplimax rotation for estimating B optimally subject to (4) in (5). That is, in the procedure, firstly EFA is performed for a data set, secondly simplimax is applied to the EFA solution for estimating the link matrix B, and, finally, the resulting B is used for minimizing (8) subject to Λ = B • Λ . These three stages are described respectively in the following subsections. The last one can be regarded as confirmatory FA (CFA) on the basis of the B given by simplimax. We thus call the procedure in this section simplimax-based CFA (SbCFA).</p><sec id="s3_1"><title>3.1. Exploratory Factor Analysis</title><p>Let T denote any m &#215; m nonsingular matrix satisfying</p><p>diag ( T T ′ ) = I m , (13)</p><p>where diag(TT&#162;) is the diagonal matrix whose diagonal elements are those of TT&#162;. EFA has rotational indeterminacy as shown next:</p><p>Σ = Λ T − 1 T T ′ T − 1 ′ Λ ′ + Ψ = Λ T Φ Λ ′ T + Ψ (14)</p><p>with F = TT&#162; and Λ<sub>T</sub> = ΛT<sup>−1</sup>. Here, the latter matrix can also be regarded as the loading matrix. Thus, even if F = TT&#162; is fixed to I<sub>m</sub> by choosing T that meets TT&#162; = I<sub>m</sub>, the (8) value remains unchanged when Λ is replaced by ΛT<sup>−1</sup>. By taking account of this property, f ( Λ , Ψ , Φ = Ι m | S ) , i.e., (8) with F = I<sub>m</sub>, is minimized over Λ and Y in EFA. This is attained with the EM algorithm described in Appendix A1. We use Λ<sub>E</sub> for Λ resulting from EFA.</p><p>The indeterminacy (14) implies that Λ<sub>T</sub> = Λ<sub>E</sub>T<sup>−1</sup> and F = TT&#162; can also be the EFA solutions of factor loading and correlation matrices, respectively, with f ( Λ E , Ψ , Φ = Ι m | S ) = f ( Λ T , Ψ , Φ | S ) . This fact is often exploited by obtaining T that allows Λ<sub>T</sub> = Λ<sub>E</sub>T to be interpretable. The procedures for obtaining T are generally referred to as factor rotation.</p></sec><sec id="s3_2"><title>3.2. Simplimax Rotation</title><p>This subsection concerns the factor rotation intermediating between the first EFA and the final CFA stages. Here, the loading matrix is unconstrained in EFA, but constrained to meet Λ = B • Λ in CFA as in (3). The latter constrainedness leads to</p><p>min Λ , Ψ , Φ f ( B • Λ , Ψ , Φ | S ) ≥ f ( Λ E , Ψ , Φ = Ι | S ) = f ( Λ T , Ψ , Φ | S ) , (15)</p><p>where f ( B • Λ , Ψ , Φ | S ) stands for the loss function in (3) in which the CFA constraint Λ = B • Λ substituted. Inequality (15) implies that the value of the EFA loss function f ( Λ E , Ψ , Φ = Ι | S ) = f ( Λ T , Ψ , Φ | S ) gives the lower limit of f ( B • Λ , Ψ , Φ | S ) to be minimized in CFA (3). This suggests that the best attainable CFA solution would be one which provides the min Λ , Ψ , Φ f ( B • Λ , Ψ , Φ | S ) value close to the EFA counterpart the f ( Λ E , Ψ , Φ = Ι | S ) = f ( Λ T , Ψ , Φ | S ) value.</p><p>Now if T can be chosen in the factor rotation such that Λ<sub>T</sub> = Λ<sub>E</sub>T<sup>−1</sup> is similar to a matrix that can be written as B • Λ for particular matrices B and L, then we can expect that such B and L will be good candidates for giving a CFA solution with a fit close to that of EFA. This can be attained by the simplimax rotation [<xref ref-type="bibr" rid="scirp.114207-ref11">11</xref>], which is formulated as minimizing the simplimax function</p><p>s p x ( B , Λ , T ) = ‖ B • Λ − Λ E T − 1 ‖ 2 (16)</p><p>over B, L, and T subject to (4) and (13). Thus, this rotation serves as a suitable bridge between the first and final stages.</p><p>For the constrained minimization of (16), two steps are iterated alternately. In one of them, the simplimax function (16) is minimized over T under (13) for a given B • Λ , using Browne’s [<xref ref-type="bibr" rid="scirp.114207-ref12">12</xref>] algorithm. In another step, (16) is minimized over B = (b<sub>ij</sub>) and Lunder (8) for a given Λ<sub>E</sub>T<sup>−1</sup>. As this problem is also related to the procedure in the next section, we express the minimization in a generalized form as</p><p>min B , Λ s p x ( B , Λ | H ) = ‖ B • Λ − H ‖ 2 for a given p&#215;m matrix H = ( h j k ) , (17)</p><p>with H = Λ<sub>E</sub>T<sup>−1</sup> in the present context. As explained in Appendix A3, (17) can be attained for</p><p>b i j = { 1 ,     if   h i j 2 ≥ h [ c ] 2 0 ,     otherwise and Λ = H , (18)</p><p>with h [ c ] 2 denoting the cth largest value of all elements in H • H = ( h i j 2 ) . The steps in the simplimax rotation algorithm is listed in Appendix A2.</p></sec><sec id="s3_3"><title>3.3. Confirmatory Factor Analysis Following Simplimax Rotation</title><p>The final stage is simply to perform CFA using B obtained by the simplimax rotation. Thus, SbCFA is formulated by making (5) concrete as</p><p>min Λ , Ψ , Φ f ( Λ , Ψ , Φ | S ) s.t. Λ = B • Λ</p><p>after B is estimated s.t. Card ( B ) = c by simplimax (5&#162;)</p><p>with f ( Λ , Ψ , Φ | S ) defined as (8). The minimization in (5) or (5&#162;) is attained with the EM algorithm, which is described in Appendix A1.</p></sec><sec id="s3_4"><title>3.4. Simplimax Factor Analysis</title><p>We will now describe simplimax FA (SimpFA). This is the procedure minimizing (6), using the solution from SbCFA as a starting configuration. Specifically, we minimize the negative of the log likelihood (see (8)), with B • Λ is substituted into L, over B, L, Y, and F subject to card(&#229;) = c (see (4)). We call this procedure simplimax FA (SimpFA), as this is a new FA procedure, and the minimization of the simplimax function (17) fulfills a key role as will be seen in the next paragraph.</p><p>The EM algorithm described in Appendix A1 can also be used for SimpFA. The algorithm for SimpFA differs from that for the other FA procedures, in that the former includes the step for minimizing (8) over both B and L with Y and F fixed. An innovative feature in the algorithm is to use Kiers’ [<xref ref-type="bibr" rid="scirp.114207-ref13">13</xref>] [<xref ref-type="bibr" rid="scirp.114207-ref14">14</xref>] majorization method for dealing with difficulties in treating (8) as a function of B and L. In that method, an auxilary function that majorizes (8) is minimized. For this minimization,</p><p>s p x ( B , Λ | W ) = ‖ B • Λ − W ‖ 2 , (19)</p><p>i.e., the simplimax function in (17) with H = W, fulfills a key role. Here, W = Λ C + Q ′ Λ ′ C Ψ − 1 − tr   C ′ Ψ − 1 Λ C with Q and C defined in A.1 and L<sub>C</sub> is the current L value (before update), see Appendix A4 for a derivation.</p></sec><sec id="s3_5"><title>3.5. Multiple Runs of SimpFA Following SbCFA</title><p>Here, we describe the procedures in SimpFA following SbCFA, in which SbCFA provides the initial values of the parameters in SimpFA providing the final solution. We take a multiple run approach for SimpFA following SbCFA, in order to reduce the possibility of selecting a local minimizer as the optimal solution: the algorithm of SbCFA followed by SimpFA is run multiple times by starting SbCFA with mutually different initial values, and the best solution is selected among the multiple SimpFA solutions. The procedure is listed as follows:</p><p>Stage 1. For S, perform EFA to provide L<sub>E</sub> and set l = 1.</p><p>Stage 2. For each of l = 1 , ⋯ , 100 , perform the following sub-stages:</p><p>Stage 2.1. Initialize T to T<sub>l</sub> and perform simplimax. Express the resulting B as B<sub>l</sub>.</p><p>Stage 2.2. For S, perform CFA using B<sub>l</sub> as B. Express the resulting L, Y, and F as Λ ˜ l , Ψ ˜ l , and Φ ˜ l , respectively, with their set Θ ˜ l = { Λ ˜ l , Ψ ˜ l , Φ ˜ l } .</p><p>Stage 2.3. For S, perform SimpFA with L, Y, and F initialized at Λ ˜ l , Ψ ˜ l , and Φ ˜ l , respectively. Express the resulting L, Y, and F as L<sub>l</sub>, Y<sub>l</sub>, and F<sub>l</sub>, respectively, with their set Θ l = { Λ l , Ψ l , Φ l } .</p><p>Stage 3. Select { Λ ^ , Ψ ^ , Φ ^ } with f ( Λ ^ , Ψ ^ , Φ ^ | S ) = min l f ( Λ l , Ψ l , Φ l | S ) as the optimal solution.</p><p>Here, the initial value T<sub>l</sub> for T in Stage 2.1 is chosen as follows: T<sub>l</sub> is obtained by the varimax rotation for L<sub>E</sub> if l = 1; otherwise, T<sub>l</sub> is set to diag ( T 0 T ′ 0 ) − 1 / 2 T 0 with the elements in T<sub>0</sub> chosen randomly: the resulting T<sub>l</sub> can be substituted into T. In <xref ref-type="fig" rid="fig2">Figure 2</xref>, the above stages are graphically illustrated, so that they can be captured visually.</p></sec><sec id="s3_6"><title>3.6. Selection of Cardinality by Information Criteria</title><p>The final solution { Λ ^ , Ψ ^ , Φ ^ } resulting in Stage 3 (the last subsection) depends on the cardinality c value in (8). We thus use { Λ ^ c , Ψ ^ c , Φ ^ c } for the solution { Λ ^ , Ψ ^ , Φ ^ } for a particular value of c. The best value for c can be selected with the procedure (9). This can be rewritten using { Λ ^ c , Ψ ^ c , Φ ^ c } as</p><p>Choose the { Λ ^ c ∗ , Ψ ^ c ∗ , Φ ^ c ∗ } for the value c ∗ = arg min c min ≤ c ≤ max I C ( c ) (20)</p><p>with c<sub>min</sub> and c<sub>max</sub> defined as (10).</p><p>As IC(c) in (20), we consider using either of Akaike’s [<xref ref-type="bibr" rid="scirp.114207-ref15">15</xref>] information criterion (AIC) or Schwarz’s [<xref ref-type="bibr" rid="scirp.114207-ref16">16</xref>] Bayesian information criterion (BIC), which are particularly popular in the information criteria. AIC and BIC can be calculated as</p><p>A I C ( c ) = 2 f ( Λ ^ c , Ψ ^ c , Φ ^ c | S ) + 2 κ ( c ) , (21)</p><p>B I C ( c ) = 2 f ( Λ ^ c , Ψ ^ c , Φ ^ c | S ) + log n κ ( c ) , (22)</p><p>respectively, for a particular value of c. Here, n is the number of observations, and k(c) = c + α with α the number of unique variances plus the number of inter-factor correlations. Thus, (21) or (22) is substituted into IC(c) in (20). Whether we should use (21) or (22) is assessed in the next section.</p></sec></sec><sec id="s4"><title>4. Simulation Study</title><p>We performed a simulation study to assess the proposed procedure in the last section, with respect to [<xref ref-type="bibr" rid="scirp.114207-ref1">1</xref>] how often the c value in (8) is selected correctly by AIC and BIC; [<xref ref-type="bibr" rid="scirp.114207-ref2">2</xref>] how similar the SbCFA and SimpFA solutions are; [<xref ref-type="bibr" rid="scirp.114207-ref3">3</xref>] how well parameter values are recovered. The procedures for synthesizing and analyzing data in the study are described in the first subsection, then the results for [<xref ref-type="bibr" rid="scirp.114207-ref1">1</xref>] [<xref ref-type="bibr" rid="scirp.114207-ref2">2</xref>] and [<xref ref-type="bibr" rid="scirp.114207-ref3">3</xref>] are reported in the following ones.</p><sec id="s4_1"><title>4.1. Data Synthesis and Analysis</title><p>With p = 12, and m =3, we set the true {L, Y, F} as in <xref ref-type="table" rid="table1">Table 1</xref> and sampled n = 300 rows of an n &#215; p data matrix X from the p-variate normal distribution whose average is 0<sub>p</sub> and covariance matrix is defined as Σ = Λ Φ Λ ′ + Ψ , see (1). The resulting X provided the sample covariance matrix S = n<sup>−1</sup>X&#162;X. We replicated this procedure 200 times to have 200 matrices S.</p><p>For each S, we carried out the procedure of SimpFA following SbCFA with the selection of the best c by (20), using both AIC and BIC. It gives the SimpFA solutions { Λ ^ c ∗ , Ψ ^ c ∗ , Φ ^ c ∗ } , which are classified into two types according to whether c was chosen by AIC or BIC.</p></sec><sec id="s4_2"><title>4.2. Correctness of Selected Cardinality</title><p>Let c<sub>true</sub> denote the true Card(L). As found in <xref ref-type="table" rid="table1">Table 1</xref>, c<sub>true</sub> = 15. On the other</p><table-wrap id="table1" ><label><xref ref-type="table" rid="table1">Table 1</xref></label><caption><title> True parameters in L, Y1<sub>p</sub>, and F with blank cells indicating zero elements</title></caption><table><tbody><thead><tr><th align="center" valign="middle"  colspan="3"  >L</th><th align="center" valign="middle" >Y1<sub>p </sub></th><th align="center" valign="middle"  colspan="3"  >F</th></tr></thead><tr><td align="center" valign="middle" >−0.9</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.2</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >0.2</td><td align="center" valign="middle" >−0.3</td></tr><tr><td align="center" valign="middle" >0.8</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.3</td><td align="center" valign="middle" >0.2</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >0.1</td></tr><tr><td align="center" valign="middle" >0.7</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.5</td><td align="center" valign="middle" >−0.3</td><td align="center" valign="middle" >0.1</td><td align="center" valign="middle" >1</td></tr><tr><td align="center" valign="middle" >0.6</td><td align="center" valign="middle" >−0.6</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.4</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.9</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.2</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" ></td><td align="center" valign="middle" >−0.8</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.4</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.7</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.5</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.6</td><td align="center" valign="middle" >−0.6</td><td align="center" valign="middle" >0.3</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.9</td><td align="center" valign="middle" >0.2</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.8</td><td align="center" valign="middle" >0.3</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >−0.7</td><td align="center" valign="middle" >0.5</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" >−0.6</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.6</td><td align="center" valign="middle" >0.4</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td></tr></tbody></table></table-wrap><p>hand, c<sup>*</sup> expresses Card(L) selected by the proposed procedure (20). We obtained the deviation c<sup>*</sup> − c<sub>true</sub> for each of the 200 data set. The averages (standard deviation) of these deviations over all data sets are 2.43 (1.40) when using AIC and 0.34 (0.62) when using BIC. This result shows that AIC tends to overestimate Card(L). Moreover, the mean absolute bias |c<sup>*</sup> − c<sub>true</sub>| for BIC was smaller than that for IC for 184 data sets among the 200 ones. This result shows that BIC is substantially better than AIC. Accordingly, we take only the BIC-based solution into consideration from here.</p></sec><sec id="s4_3"><title>4.3. Equivalence of SbCFA and SimpFA Solutions</title><p>It was found that 98 percent of the 200 solutions of SimpFA were equivalent to those of the SbCFA ones used for initializing the SimpFA parameters: no iteration in the SimpFA algorithm was required. Also in the remaining two percent of the runs, SbCFA and SimFA solutions were almost equivalent, with the average of the differences between the two solutions over the 200 data sets being 0.001. Here, the difference is defined as ( ‖ Λ ^ c ∗ − Λ ˜ c ∗ ‖ 1 + ‖ Ψ ^ c ∗ − Ψ ˜ c ∗ ‖ 1 + ‖ Φ ^ c ∗ − Φ ˜ c ∗ ‖ 1 ) / { p m + p − m ( m − 1 ) } with ‖ Γ ‖ 1 = Σ i Σ j | γ i j | denoting the L<sub>1</sub> norm of a matrix G = (γ<sub>ij</sub>) and { Λ ˜ c ∗ , Ψ ˜ c ∗ , Φ ˜ c ∗ } being the SbCFA counterpart of { Λ ^ c ∗ , Ψ ^ c ∗ , Φ ^ c ∗ } .</p><p>The above results do not show that the SimpFA (following SbCFA) is useless, as SimpFA serves for showing the equivalence of its solution to the SbCFA one, which allows us to find the optimality of the SbCFA solution.</p></sec><sec id="s4_4"><title>4.4. Recovery of Parameters</title><p>We assess how well the true parameter values are recovered by the SimpFA solution Θ ^ c ∗ = { Λ ^ c ∗ , Ψ ^ c ∗ , Φ ^ c ∗ } , which is provided by (20) with IC(c) = BIC(c). As the indices standing for the badness in the recovery for L, Y, and F,we obtained ‖ Λ ^ c ∗ − Λ true ‖ 1 / ( p m ) , ‖ Ψ ^ c ∗ − Ψ true ‖ 1 / p , and ‖ Φ ^ c ∗ − Φ true ‖ 1 / M , respectively. Further, for assessing the incorrectness in identifying the true zero and nonzero loadings, we recorded MIR<sub>0</sub> = the number of the true zero loadings estimated as nonzero/(pm − c), and MIR<sub>#</sub> = the number of the true nonzero loadings estimated as zero/c, for each data set, with M = m(m − 1). Here MIR stands for misidentification rate.</p><p>The statistics of the resulting five index values over the 200 data sets are presented in <xref ref-type="table" rid="table2">Table 2</xref>. It shows that the proposed procedure recovered { Λ ^ c ∗ , Ψ ^ c ∗ , Φ ^ c ∗ } , i.e., the links of variables to factors and parameter values, fairly well, except that the 95 percentile for factor correlations can be considered large. These results allow us to conclude that the true CFA model can be recovered satisfactorily, though the estimates of factor correlations might deviate from the true counterparts in a few cases.</p></sec></sec><sec id="s5"><title>5. Real Data Demonstration</title><p>In order to demonstrate how useful our proposed procedure is, we apply it to two data sets that have already been analyzed by CFA with zero constraints selected by users. The solutions of the latter user-based CFA are compared with those for our proposed procedure.</p><p>The first data set is Carlson and Mulaik’s [<xref ref-type="bibr" rid="scirp.114207-ref17">17</xref>] personality trait data matrix of n = 280 (participants) &#215; p = 15 (personality traits). Its correlation matrix (Mulaik, 2010, p. 198) has been analyzed by Mulaik [<xref ref-type="bibr" rid="scirp.114207-ref2">2</xref>] using CFA with his selected constraints, and its solution is shown in Mulaik (2010, p. 437) and also in <xref ref-type="table" rid="table3">Table 3</xref> together with the solution of the proposed procedure. Here, the latter solution is the same as the one of SbCFA before SimpFA: iteration was not required in the SimpFA algorithm. Comparing the BIC values in <xref ref-type="table" rid="table3">Table 3</xref>, we can find that the solution of the proposed procedure was better than Mulaik’s [<xref ref-type="bibr" rid="scirp.114207-ref2">2</xref>] one.</p><p>Another data set is Kojima’s housing preference one with n = 1120 (participants) and p = 13 (features). It describes to what degree the participants wish to</p><table-wrap id="table2" ><label><xref ref-type="table" rid="table2">Table 2</xref></label><caption><title> Percentiles, average (Avg) and standard deviation (SD) of each index over200 data sets</title></caption><table><tbody><thead><tr><th align="center" valign="middle"  rowspan="2"  >Index</th><th align="center" valign="middle"  colspan="5"  >Percentiles</th><th align="center" valign="middle"  rowspan="2"  >Avg.</th><th align="center" valign="middle"  rowspan="2"  >SD</th></tr></thead><tr><td align="center" valign="middle" >5</td><td align="center" valign="middle" >25</td><td align="center" valign="middle" >50</td><td align="center" valign="middle" >75</td><td align="center" valign="middle" >95</td></tr><tr><td align="center" valign="middle" >‖ Λ ^ c ∗ − Λ true ‖ 1 / ( p m )</td><td align="center" valign="middle" >0.011</td><td align="center" valign="middle" >0.013</td><td align="center" valign="middle" >0.015</td><td align="center" valign="middle" >0.017</td><td align="center" valign="middle" >0.088</td><td align="center" valign="middle" >0.023</td><td align="center" valign="middle" >0.035</td></tr><tr><td align="center" valign="middle" >‖ Ψ ^ c ∗ − Ψ true ‖ 1 / p</td><td align="center" valign="middle" >0.029</td><td align="center" valign="middle" >0.035</td><td align="center" valign="middle" >0.039</td><td align="center" valign="middle" >0.045</td><td align="center" valign="middle" >0.053</td><td align="center" valign="middle" >0.040</td><td align="center" valign="middle" >0.008</td></tr><tr><td align="center" valign="middle" >‖ Φ ^ c ∗ − Φ true ‖ 1 / M</td><td align="center" valign="middle" >0.020</td><td align="center" valign="middle" >0.035</td><td align="center" valign="middle" >0.049</td><td align="center" valign="middle" >0.069</td><td align="center" valign="middle" >0.227</td><td align="center" valign="middle" >0.071</td><td align="center" valign="middle" >0.089</td></tr><tr><td align="center" valign="middle" >MIR<sub>0</sub></td><td align="center" valign="middle" >0.000</td><td align="center" valign="middle" >0.000</td><td align="center" valign="middle" >0.000</td><td align="center" valign="middle" >0.000</td><td align="center" valign="middle" >0.095</td><td align="center" valign="middle" >0.006</td><td align="center" valign="middle" >0.026</td></tr><tr><td align="center" valign="middle" >MIR<sub>#</sub></td><td align="center" valign="middle" >0.000</td><td align="center" valign="middle" >0.000</td><td align="center" valign="middle" >0.000</td><td align="center" valign="middle" >0.067</td><td align="center" valign="middle" >0.200</td><td align="center" valign="middle" >0.032</td><td align="center" valign="middle" >0.067</td></tr></tbody></table></table-wrap><table-wrap id="table3" ><label><xref ref-type="table" rid="table3">Table 3</xref></label><caption><title> Solution for personality trait data, with blank cells indicating zero elements</title></caption><table><tbody><thead><tr><th align="center" valign="middle"  rowspan="2"  >Variable</th><th align="center" valign="middle"  colspan="4"  >User-based CFA:BIC= 924.1</th><th align="center" valign="middle"  colspan="4"  >Proposed: BIC = 12.3</th></tr></thead><tr><td align="center" valign="middle"  colspan="3"  >L</td><td align="center" valign="middle" >Y1<sub>p</sub></td><td align="center" valign="middle"  colspan="3"  >L</td><td align="center" valign="middle" >Y1<sub>p</sub></td></tr><tr><td align="center" valign="middle" >Friendly</td><td align="center" valign="middle" >0.85</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.53</td><td align="center" valign="middle" >0.85</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.29</td></tr><tr><td align="center" valign="middle" >Sympathetic</td><td align="center" valign="middle" >0.92</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.39</td><td align="center" valign="middle" >1.03</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >-0.19</td><td align="center" valign="middle" >0.15</td></tr><tr><td align="center" valign="middle" >Kind</td><td align="center" valign="middle" >0.93</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.38</td><td align="center" valign="middle" >1.10</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >-0.28</td><td align="center" valign="middle" >0.11</td></tr><tr><td align="center" valign="middle" >Affectionate</td><td align="center" valign="middle" >0.90</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.43</td><td align="center" valign="middle" >0.90</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.19</td></tr><tr><td align="center" valign="middle" >Intelligent</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.88</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.48</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.88</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.23</td></tr><tr><td align="center" valign="middle" >Capable</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.93</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.38</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.93</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.15</td></tr><tr><td align="center" valign="middle" >Competent</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.93</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.38</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.93</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.15</td></tr><tr><td align="center" valign="middle" >Smart</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.90</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.43</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.90</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.19</td></tr><tr><td align="center" valign="middle" >Talkative</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.80</td><td align="center" valign="middle" >0.60</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.80</td><td align="center" valign="middle" >0.35</td></tr><tr><td align="center" valign="middle" >Outgoing</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.95</td><td align="center" valign="middle" >0.32</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.95</td><td align="center" valign="middle" >0.10</td></tr><tr><td align="center" valign="middle" >Gregarious</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.89</td><td align="center" valign="middle" >0.45</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.89</td><td align="center" valign="middle" >0.21</td></tr><tr><td align="center" valign="middle" >Extrovert</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.90</td><td align="center" valign="middle" >0.44</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.90</td><td align="center" valign="middle" >0.19</td></tr><tr><td align="center" valign="middle" >Helpful</td><td align="center" valign="middle" >0.73</td><td align="center" valign="middle" >0.22</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.59</td><td align="center" valign="middle" >0.74</td><td align="center" valign="middle" >0.20</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.35</td></tr><tr><td align="center" valign="middle" >Cooperative</td><td align="center" valign="middle" >0.72</td><td align="center" valign="middle" >0.23</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.60</td><td align="center" valign="middle" >0.73</td><td align="center" valign="middle" >0.21</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.36</td></tr><tr><td align="center" valign="middle" >Sociable</td><td align="center" valign="middle" >0.17</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.83</td><td align="center" valign="middle" >0.34</td><td align="center" valign="middle" >0.18</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.81</td><td align="center" valign="middle" >0.12</td></tr><tr><td align="center" valign="middle" >Factor</td><td align="center" valign="middle"  colspan="3"  >F</td><td align="center" valign="middle" ></td><td align="center" valign="middle"  colspan="3"  >F</td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" >1</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" >0.22</td><td align="center" valign="middle" >0.56</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" >0.25</td><td align="center" valign="middle" >0.64</td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" >2</td><td align="center" valign="middle" >0.22</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" >0.30</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.25</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" >0.30</td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" >3</td><td align="center" valign="middle" >0.56</td><td align="center" valign="middle" >0.30</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.64</td><td align="center" valign="middle" >0.30</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" ></td></tr></tbody></table></table-wrap><p>live in the houses featured by the variables. The correlation matrix for this data set is presented in <xref ref-type="table" rid="table4">Table 4</xref>, as it is described in Japanese and not easily available. This matrix has been analyzed by Kojima using CFA with his selected constraints, and its solution was shown in Kojima and also in <xref ref-type="table" rid="table5">Table 5</xref> with the solution of the proposed procedure. Here, this solution differs from the SbCFA counterpart: iteration was required in the SimpFA algorithm. In <xref ref-type="table" rid="table5">Table 5</xref>, BIC shows that the solution of the proposed procedure was better.</p><p>In both of the above examples, the proposed procedure outperformed in the BIC values, but its solutions are similar to the counterparts of CFA with users’ selected constraints (<xref ref-type="table" rid="table3">Table 3</xref> and <xref ref-type="table" rid="table5">Table 5</xref>). This similarity shows that the selected constraints were rational. However, the proposed procedure is fully computational and does not require any users’ effort for considering constraints. This property would be helpful as soon as there is some uncertainty as to which loadings to constrain to 0.</p></sec><sec id="s6"><title>6. Conclusions</title><p>A problem in the confirmatory factor analysis (CFA) is that users must select</p><table-wrap id="table4" ><label><xref ref-type="table" rid="table4">Table 4</xref></label><caption><title> Correlation matrix for Kojima’s (2013) housing preference data</title></caption><table><tbody><thead><tr><th align="center" valign="middle"  colspan="2"  >Variable</th><th align="center" valign="middle" >1</th><th align="center" valign="middle" >2</th><th align="center" valign="middle" >3</th><th align="center" valign="middle" >4</th><th align="center" valign="middle" >5</th><th align="center" valign="middle" >6</th><th align="center" valign="middle" >7</th><th align="center" valign="middle" >8</th><th align="center" valign="middle" >9</th><th align="center" valign="middle" >10</th><th align="center" valign="middle" >11</th><th align="center" valign="middle" >12</th><th align="center" valign="middle" >13</th></tr></thead><tr><td align="center" valign="middle" >1</td><td align="center" valign="middle" >Food services</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.423</td><td align="center" valign="middle" >0.404</td><td align="center" valign="middle" >0.127</td><td align="center" valign="middle" >0.179</td><td align="center" valign="middle" >0.132</td><td align="center" valign="middle" >0.148</td><td align="center" valign="middle" >0.214</td><td align="center" valign="middle" >0.225</td><td align="center" valign="middle" >0.252</td><td align="center" valign="middle" >0.125</td><td align="center" valign="middle" >0.122</td><td align="center" valign="middle" >0.192</td></tr><tr><td align="center" valign="middle" >2</td><td align="center" valign="middle" >Tea services</td><td align="center" valign="middle" >0.423</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.751</td><td align="center" valign="middle" >0.187</td><td align="center" valign="middle" >0.267</td><td align="center" valign="middle" >0.115</td><td align="center" valign="middle" >0.188</td><td align="center" valign="middle" >0.223</td><td align="center" valign="middle" >0.265</td><td align="center" valign="middle" >0.300</td><td align="center" valign="middle" >0.144</td><td align="center" valign="middle" >0.097</td><td align="center" valign="middle" >0.193</td></tr><tr><td align="center" valign="middle" >3</td><td align="center" valign="middle" >Home for the old</td><td align="center" valign="middle" >0.404</td><td align="center" valign="middle" >0.751</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.238</td><td align="center" valign="middle" >0.276</td><td align="center" valign="middle" >0.178</td><td align="center" valign="middle" >0.239</td><td align="center" valign="middle" >0.249</td><td align="center" valign="middle" >0.281</td><td align="center" valign="middle" >0.325</td><td align="center" valign="middle" >0.174</td><td align="center" valign="middle" >0.147</td><td align="center" valign="middle" >0.220</td></tr><tr><td align="center" valign="middle" >4</td><td align="center" valign="middle" >Lush greenery</td><td align="center" valign="middle" >0.127</td><td align="center" valign="middle" >0.187</td><td align="center" valign="middle" >0.238</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.411</td><td align="center" valign="middle" >0.399</td><td align="center" valign="middle" >0.350</td><td align="center" valign="middle" >0.286</td><td align="center" valign="middle" >0.274</td><td align="center" valign="middle" >0.273</td><td align="center" valign="middle" >0.150</td><td align="center" valign="middle" >0.192</td><td align="center" valign="middle" >0.176</td></tr><tr><td align="center" valign="middle" >5</td><td align="center" valign="middle" >Walking and joking</td><td align="center" valign="middle" >0.179</td><td align="center" valign="middle" >0.267</td><td align="center" valign="middle" >0.276</td><td align="center" valign="middle" >0.411</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.545</td><td align="center" valign="middle" >0.330</td><td align="center" valign="middle" >0.302</td><td align="center" valign="middle" >0.296</td><td align="center" valign="middle" >0.333</td><td align="center" valign="middle" >0.235</td><td align="center" valign="middle" >0.248</td><td align="center" valign="middle" >0.224</td></tr><tr><td align="center" valign="middle" >6</td><td align="center" valign="middle" >Large park</td><td align="center" valign="middle" >0.132</td><td align="center" valign="middle" >0.115</td><td align="center" valign="middle" >0.178</td><td align="center" valign="middle" >0.399</td><td align="center" valign="middle" >0.545</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.298</td><td align="center" valign="middle" >0.327</td><td align="center" valign="middle" >0.229</td><td align="center" valign="middle" >0.362</td><td align="center" valign="middle" >0.173</td><td align="center" valign="middle" >0.242</td><td align="center" valign="middle" >0.143</td></tr><tr><td align="center" valign="middle" >7</td><td align="center" valign="middle" >River and lake</td><td align="center" valign="middle" >0.148</td><td align="center" valign="middle" >0.188</td><td align="center" valign="middle" >0.239</td><td align="center" valign="middle" >0.350</td><td align="center" valign="middle" >0.330</td><td align="center" valign="middle" >0.298</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.266</td><td align="center" valign="middle" >0.205</td><td align="center" valign="middle" >0.240</td><td align="center" valign="middle" >0.140</td><td align="center" valign="middle" >0.153</td><td align="center" valign="middle" >0.161</td></tr><tr><td align="center" valign="middle" >8</td><td align="center" valign="middle" >Communal events</td><td align="center" valign="middle" >0.214</td><td align="center" valign="middle" >0.223</td><td align="center" valign="middle" >0.249</td><td align="center" valign="middle" >0.286</td><td align="center" valign="middle" >0.302</td><td align="center" valign="middle" >0.327</td><td align="center" valign="middle" >0.266</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.417</td><td align="center" valign="middle" >0.554</td><td align="center" valign="middle" >0.338</td><td align="center" valign="middle" >0.293</td><td align="center" valign="middle" >0.324</td></tr><tr><td align="center" valign="middle" >9</td><td align="center" valign="middle" >Multigenerational living</td><td align="center" valign="middle" >0.225</td><td align="center" valign="middle" >0.265</td><td align="center" valign="middle" >0.281</td><td align="center" valign="middle" >0.274</td><td align="center" valign="middle" >0.296</td><td align="center" valign="middle" >0.229</td><td align="center" valign="middle" >0.205</td><td align="center" valign="middle" >0.417</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.371</td><td align="center" valign="middle" >0.223</td><td align="center" valign="middle" >0.171</td><td align="center" valign="middle" >0.222</td></tr><tr><td align="center" valign="middle" >10</td><td align="center" valign="middle" >Active community</td><td align="center" valign="middle" >0.252</td><td align="center" valign="middle" >0.300</td><td align="center" valign="middle" >0.325</td><td align="center" valign="middle" >0.273</td><td align="center" valign="middle" >0.333</td><td align="center" valign="middle" >0.362</td><td align="center" valign="middle" >0.240</td><td align="center" valign="middle" >0.554</td><td align="center" valign="middle" >0.371</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.286</td><td align="center" valign="middle" >0.291</td><td align="center" valign="middle" >0.283</td></tr><tr><td align="center" valign="middle" >11</td><td align="center" valign="middle" >Interactions for hobbies</td><td align="center" valign="middle" >0.125</td><td align="center" valign="middle" >0.144</td><td align="center" valign="middle" >0.174</td><td align="center" valign="middle" >0.150</td><td align="center" valign="middle" >0.235</td><td align="center" valign="middle" >0.173</td><td align="center" valign="middle" >0.140</td><td align="center" valign="middle" >0.338</td><td align="center" valign="middle" >0.223</td><td align="center" valign="middle" >0.286</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.335</td><td align="center" valign="middle" >0.452</td></tr><tr><td align="center" valign="middle" >12</td><td align="center" valign="middle" >House party</td><td align="center" valign="middle" >0.122</td><td align="center" valign="middle" >0.097</td><td align="center" valign="middle" >0.147</td><td align="center" valign="middle" >0.192</td><td align="center" valign="middle" >0.248</td><td align="center" valign="middle" >0.242</td><td align="center" valign="middle" >0.153</td><td align="center" valign="middle" >0.293</td><td align="center" valign="middle" >0.171</td><td align="center" valign="middle" >0.291</td><td align="center" valign="middle" >0.335</td><td align="center" valign="middle" >1.000</td><td align="center" valign="middle" >0.333</td></tr><tr><td align="center" valign="middle" >13</td><td align="center" valign="middle" >Utilizing own’s careers</td><td align="center" valign="middle" >0.192</td><td align="center" valign="middle" >0.193</td><td align="center" valign="middle" >0.220</td><td align="center" valign="middle" >0.176</td><td align="center" valign="middle" >0.224</td><td align="center" valign="middle" >0.143</td><td align="center" valign="middle" >0.161</td><td align="center" valign="middle" >0.324</td><td align="center" valign="middle" >0.222</td><td align="center" valign="middle" >0.283</td><td align="center" valign="middle" >0.452</td><td align="center" valign="middle" >0.333</td><td align="center" valign="middle" >1.000</td></tr></tbody></table></table-wrap><table-wrap id="table5" ><label><xref ref-type="table" rid="table5">Table 5</xref></label><caption><title> Solution for housing preference data, with blank cells indicating zero elements</title></caption><table><tbody><thead><tr><th align="center" valign="middle"  rowspan="2"  >Variable</th><th align="center" valign="middle"  colspan="5"  >(A) User-based CFA: BIC = 10915.5</th><th align="center" valign="middle"  colspan="5"  >(B) Proposed: BIC = 10864.2</th></tr></thead><tr><td align="center" valign="middle"  colspan="4"  >L</td><td align="center" valign="middle" >Y1<sub>p</sub></td><td align="center" valign="middle"  colspan="4"  >L</td><td align="center" valign="middle" >Y1<sub>p</sub></td></tr><tr><td align="center" valign="middle" >Food services</td><td align="center" valign="middle" >0.48</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.77</td><td align="center" valign="middle" >0.39</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.16</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.76</td></tr><tr><td align="center" valign="middle" >Tea services</td><td align="center" valign="middle" >0.86</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.27</td><td align="center" valign="middle" >0.89</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.21</td></tr><tr><td align="center" valign="middle" >Home for the old</td><td align="center" valign="middle" >0.88</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.23</td><td align="center" valign="middle" >0.81</td><td align="center" valign="middle" >0.09</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.28</td></tr><tr><td align="center" valign="middle" >Lush greenery</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.59</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.65</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.58</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.67</td></tr><tr><td align="center" valign="middle" >Walking and joking</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.74</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.46</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.82</td><td align="center" valign="middle" >−0.11</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.45</td></tr><tr><td align="center" valign="middle" >Large park</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.69</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.52</td><td align="center" valign="middle" >−0.18</td><td align="center" valign="middle" >0.78</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.48</td></tr><tr><td align="center" valign="middle" >River and lake</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.49</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.77</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.48</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.77</td></tr><tr><td align="center" valign="middle" >Communal events</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.74</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.45</td><td align="center" valign="middle" >−0.17</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.84</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.40</td></tr><tr><td align="center" valign="middle" >Multigenerational living</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.55</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.70</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.55</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.70</td></tr><tr><td align="center" valign="middle" >Active community</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.73</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.47</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.72</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.49</td></tr><tr><td align="center" valign="middle" >Interactions for hobbies</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.66</td><td align="center" valign="middle" >0.57</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.67</td><td align="center" valign="middle" >0.55</td></tr><tr><td align="center" valign="middle" >House party</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.53</td><td align="center" valign="middle" >0.72</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.15</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.43</td><td align="center" valign="middle" >0.73</td></tr><tr><td align="center" valign="middle" >Utilizing own’s careers</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.66</td><td align="center" valign="middle" >0.57</td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.67</td><td align="center" valign="middle" >0.55</td></tr><tr><td align="center" valign="middle" >Factor</td><td align="center" valign="middle"  colspan="4"  >F</td><td align="center" valign="middle" ></td><td align="center" valign="middle"  colspan="4"  >F</td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" >1</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" >0.38</td><td align="center" valign="middle" >0.46</td><td align="center" valign="middle" >0.32</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" >0.41</td><td align="center" valign="middle" >0.50</td><td align="center" valign="middle" >0.28</td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" >2</td><td align="center" valign="middle" >0.38</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" >0.65</td><td align="center" valign="middle" >0.47</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.41</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" >0.69</td><td align="center" valign="middle" >0.44</td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" >3</td><td align="center" valign="middle" >0.46</td><td align="center" valign="middle" >0.65</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" >0.65</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.50</td><td align="center" valign="middle" >0.69</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" >0.62</td><td align="center" valign="middle" ></td></tr><tr><td align="center" valign="middle" >4</td><td align="center" valign="middle" >0.32</td><td align="center" valign="middle" >0.47</td><td align="center" valign="middle" >0.65</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0.28</td><td align="center" valign="middle" >0.44</td><td align="center" valign="middle" >0.62</td><td align="center" valign="middle" >1.00</td><td align="center" valign="middle" ></td></tr></tbody></table></table-wrap><p>what pairs of variables and factors are linked, in other words, what loadings are to be zero or nonzero, in a subjective manner. To deal with this problem, we proposed the procedure of SimpFA following SbCFA for computating an optimally suitable CFA model and its solution without relying on any user’s judgment. The simulation study showed that the true CFA model and parameter values can be recovered fairly well by the proposed procedure. Real data examples demonstrated that it can outperform CFA with users’ selected constraints in terms of the BIC statistic.</p><p>In Section 4.3, we found the SbCFA solutions to be equivalent to SimpFA in almost all cases. In particular, the good performance of SbCFA was somewhat surprising. To theoretically study, reasons for such a SbCFA performance are considered as a subject for a future study.</p></sec><sec id="s7"><title>Conflicts of Interest</title><p>The authors declare no conflicts of interest regarding the publication of this paper.</p></sec><sec id="s8"><title>Cite this paper</title><p>Cai, J.Y., Kiers, H.A.L. and Adachi, K. (2021) Computational Identification of Confirmatory Factor Analysis Model with Simplimax Procedures. Open Journal of Statistics, 11, 1044-1061. https://doi.org/10.4236/ojs.2021.116062</p></sec><sec id="s9"><title>Appendix</title>A1. EM Algorithm for Factor Analysis<p>In Rubin and Thayer’s (1982) EM algorithm for FA, anauxiliary function of (8) is defined as</p><p>ϕ ( Λ , Ψ , Φ | S ) = log | Ψ | + tr   Ψ − 1 ( S − 2 C Λ ′ + Λ Q Λ ′ ) + log | Φ | + tr   Q Φ − 1 . (A1)</p><p>Here, C = SA and Q = A&#162;SA + U, with A and U being computed as</p><p>A = ( Λ Φ Λ ′ + Ψ ) − 1 Λ Φ and U = Φ 1 / 2 [ I m + ( Φ 1 / 2 Λ ′ ) Ψ − 1 ( Λ Φ 1 / 2 ) ] − 1 Φ 1 / 2 (A2)</p><p>using the current L,Y, and F values. In the EM algorithm, (A1) rather than (8) is minimized iteratively. This minimization is known to allow (8) to be minimized (Rubin &amp; Thayer, 1982). This fact also holds true, when L is constrained as Λ = B • Λ . Thus, the minimization of (8) in EFA, SbCFA (5) or (5&#162;), and SimpFA (6) or (6&#162;) is attained with the EM algorithm.</p><p>The EM algorithm for EFA, SbCFA, and SimpFA can be summarized as follows:</p><p>Step 1. Initialize L,Y, and F, with B also initialized in SimpFA.</p><p>Step 2. Obtain C and Q in (A.1) through (A.2).</p><p>Step 3. Minimize (A.1) over Y with L and F fixed.</p><p>Step 4. Minimize (A.1) over L with Y and F fixed, with “over L“ replaced by “over B and L under (4)” only in SimpFA.</p><p>Step 5. Minimize (A.1) over F with Y and L fixed.</p><p>Step 6. Transform F and L so that F is a correlation matrix and finish if convergence is reached; otherwise. Go back to Step 2.</p><p>Here, the minimization in Step 3 is attained when the ith diagonal elements in Y are updated as ψ i i = s i i − 2 λ ′ i c i + λ ′ i Q λ i ( i = 1 , ⋯ , p ), with λ ′ i the ith row of L, c ′ i that of C, and s<sub>ii</sub> the (i,i) element of S. The task of Step 5 is attained simply by updating F as F = Q. This is because F may be regarded simply as a covariance matrix during the iteration and finally be transromed into a correlation matrix. Thus, ifconvergenceisreachedin Step 6, we must transform F and L as Φ R = diag ( Φ ) − 1 / 2 Φ diag ( Φ ) − 1 / 2 and Λ R = Λ diag ( Φ ) 1 / 2 , then regard the resulting F<sub>R</sub> and L<sub>R</sub> as the solutions of F and L, respectively. Here, we should notice that the transformation does not change the value of (8), which is shown by Λ Φ Λ ′ = Λ D ( D − 1 Φ D − 1 ) ( Λ D ) ′ with D an m &#215; m diagonal matrix. Convergence in Step 6 is defined here as the decrease in (6) from the previous round being less than 10<sup>−6</sup>. The details in Steps 1 and 4 differ among EFA, SbCFA, and SimpFA, as described in the next paragraphs.</p><p>The initialization in Step 1 is detailed here. In EFA, the eigenvalue decomposition of S defined as S = LD<sup>2</sup>L&#162; is used, with D<sup>2</sup> the p &#215; p diagonal matrix whose diagonal elements are arranged in descending order, and LL&#162; = I<sub>p</sub>. That is, the initial L and Y are set to L<sub>m</sub>D<sub>m</sub> and diag ( S − L m Δ m 2 L ′ m ) , respectively, in EFA, with Δ m 2 the first m &#215; m diagonal block of D<sup>2</sup> and L<sub>m</sub> the p &#215; m matrix containing the first m columns of L. In SbCFA, the initial Y is set to the one obtained in the preceding EFA, while the initial L and F are set to the matrices B • Λ and TT&#162;, respectively, that are obtained by the preceding simplimax rotation. In SimpFA, L, Y, and F are intialized at their SbCFA solution and B at its simplimax rotation solution.</p><p>The minimization in Step 4 is attained by the update of L with L = CQ<sup>−</sup><sup>1</sup> in EFA and the row-wise λ i = B ′ i ( B i Q B ′ i ) − 1 B i c i ( i = 1 , ⋯ , p ) in SbCFA. Here, B<sub>i</sub> is the m<sub>i</sub> &#215; m binary matrix satisfying 1 ′ m i B i = b ′ i with b ′ i the ith row of the link matrix B = (b<sub>ij</sub>) defined as (2) and m i = b ′ i 1 m : for example, if b ′ i = [ 1 , 0 , 1 ] ,</p><p>then B i = [ 1 0 0 0 0 1 ] . The Step 4 in SimpFA is detailed in Appendix A4.</p>A2. Algorithm for Simplimax Rotation<p>The algorithm for the simplimax rotation can be summarized as follows:</p><p>Step 1. Initialize T</p><p>Step 2. Update B and L as (18) with H = L<sub>E</sub>T<sup>−1 </sup></p><p>Step 3. Update T with Browne’s (1972) algorithm</p><p>Step 4. Finish if convergence is reached; otherwise, go back to Step 2.</p><p>How T is initialized is described in the section for the whole process of the proposed procedure. The convergence is defined that the decrease in (16) from the previous round is less than 10<sup>−6</sup> in this paper.</p>A3. Derivation of (18)<p>The simplimax function s p x ( B , Λ | H ) = ‖ B • Λ − H ‖ 2 in (17) can be rewritten as</p><p>s p x ( B , Λ | H ) = ∑ ( i , j ) ∈ ℵ ( b i j λ i j − h i j ) 2 + ∑ ( i , j ) ∈ ℵ ⊥ h i j 2 ≥ ∑ ( i , j ) ∈ ℵ ⊥ h i j 2 . (A3)</p><p>with H = (h<sub>ij</sub>). Here, ℵ denotes the set of the index pairs (i, j) for b<sub>ij</sub> = 1, while ℵ ⊥ is the set of (i, j) for b<sub>ij</sub> = 0, and we have used ∑ ( i , j ) ∈ ℵ ⊥ ( b i j λ i j − h i j ) 2 = ∑ ( i , j ) ∈ ℵ ⊥ h i j 2 . The inequality in (A.3) shows that the lower limit of s p x ( B , Λ | H ) is ∑ ( i , j ) ∈ ℵ ⊥ h i j 2 which is attained if b<sub>ij</sub>λ<sub>ij</sub> = h<sub>ij</sub>. Furthermore, the limit ∑ ( i , j ) ∈ ℵ ⊥ h i j 2 is minimal when ℵ ⊥ contains the (i, j) for h i j 2 ≤ [ h i j 2 ] q , with q = p m − c and [ h i j 2 ] q the qth smallest h i j 2 value among h i j 2 , i = 1 , ⋯ , p ; j = 1 , ⋯ , m . This can be rewritten as (18). That is, (17) is attained for (18).</p>A4. Majorization Algorithm for SimpFA Loadings<p>Let us rewrite (A.1) as ϕ * ( Λ ) + c o n s t , where const is a part independent of Λ = B • Λ and</p><p>ϕ * ( Λ ) = − 2tr Ψ − 1 C Λ ′ + tr Ψ − 1 Λ Q Λ ′ = tr Ψ − 1 Λ Q Λ ′ − 2tr C ′ Ψ − 1 Λ . (A4)</p><p>This minimization over B • Λ is found to be the task of Step 4 in Appendix A1 for SimpFA. Using Δ = Λ − Λ C or Λ = Λ C + Δ with L<sub>C</sub> the current L value, (A4) can be rewritten as</p><p>ϕ * ( Λ ) = tr Ψ − 1 ( Λ C + Δ ) Q ( Λ C + Δ ) ′ − 2tr C ′ Ψ − 1 ( Λ C + Δ ) = ϕ * ( Λ C ) + tr Ψ − 1 Δ Q Δ ′ + tr 2 V Δ (A4&#162;)</p><p>with V = Q ′ Λ ′ C Ψ − 1 − C ′ Ψ − 1 Λ C . Kiers (1990, Theorem 1) shows that tr Ψ − 1 Δ Q Δ in (A4&#162;) satisfies the inequality tr Ψ − 1 Δ Q Δ ′ / ‖ Δ ‖ 2 ≤ α , i.e., tr Ψ − 1 Δ Q Δ ′ ≤ α ‖ Δ ‖ 2 , where α is the greatest eigenvalue of Ψ − 1 ⊗ Q , with ⊗ denoting the Kronecker product. Using that inequality, a function majorizing (A4&#162;) can be defined as</p><p>η ( Λ ) = ϕ * ( Λ C ) + β tr ‖ Δ ‖ 2 + 2tr V Δ = g ( Λ C ) + β ‖ Δ − V ‖ 2 − β ‖ V ‖ 2 , (A5)</p><p>with β ≥ α and β ≥ 0: (14) satisfies ϕ * ( Λ C ) = η ( Λ C ) ≥ η ( Λ ) ≥ ϕ * ( Λ ) if L is the minimizer of ϕ * ( Λ C ) . Since only β ‖ Δ − V ‖ 2 is a function of L with β &#179; 0 on the right side of (A5), the task of Step 4 is attained by minimizing η * ( Λ ) = ‖ Δ − V ‖ 2 . From Δ = Λ − Λ C , we can rewrite η * ( Λ ) as η * ( Λ ) = ‖ Λ − Λ C − V ‖ 2 = ‖ Λ − W ‖ 2 , which is the simplimax function (19), with W = Λ C + V . Thus, B and L to be obtained in Step 4 are given by (18) with H = (h<sub>ij</sub>) and h [ c ] 2 replaced by W = (w<sub>ij</sub>) and w [ c ] 2 , respectively. Here, w [ c ] 2 is the cth largest value of all elements in W • W = ( w i j 2 ) .</p></sec></body><back><ref-list><title>References</title><ref id="scirp.114207-ref1"><label>1</label><mixed-citation publication-type="other" xlink:type="simple">Adachi, K. (2019) Factor Analysis: Latent Variable, Matrix Decomposition, and Constrained Uniqueness Formulations. WIREs Computational Statistics, 11, e1458.  
https://onlinelibrary.wiley.com/doi/abs/10.1002/wics.1458  
https://doi.org/10.1002/wics.1458</mixed-citation></ref><ref id="scirp.114207-ref2"><label>2</label><mixed-citation publication-type="other" xlink:type="simple">Mulaik, S.A. (2010) Foundations of Factor Analysis. 2nd Edition, CRC Press, Boca Raton.</mixed-citation></ref><ref id="scirp.114207-ref3"><label>3</label><mixed-citation publication-type="book" xlink:type="simple">Yanai, H. and Ichikawa, M. (2007) Factor Analysis. In: Rao, C.R. and Sinharay, S., Eds., Handbook of Statistics, Vol. 26, Elsevier, Amsterdam, 257-296.  
https://doi.org/10.1016/S0169-7161(06)26009-7</mixed-citation></ref><ref id="scirp.114207-ref4"><label>4</label><mixed-citation publication-type="other" xlink:type="simple">Arbuckle, J.L. (2008) AMOS 17 User’s Guide. SPSS Inc., Chicago.</mixed-citation></ref><ref id="scirp.114207-ref5"><label>5</label><mixed-citation publication-type="other" xlink:type="simple">Bollen, K.A. (1989) Structural Equations with Latent Variables. Wiley, New York.  
https://doi.org/10.1002/9781118619179</mixed-citation></ref><ref id="scirp.114207-ref6"><label>6</label><mixed-citation publication-type="other" xlink:type="simple">J&amp;ouml;reskog, K.G. (1969) A General Approach to Confirmatory Maximum Likelihood Factor Analysis. Psychometrika, 34, 183-202.  
https://doi.org/10.1002/9781118619179</mixed-citation></ref><ref id="scirp.114207-ref7"><label>7</label><mixed-citation publication-type="other" xlink:type="simple">Lawley, D.N. (1941) Further Investigation of Factor Estimation. Proceedings of the Royal Society of Edinburgh A, 61, 176-185.  
https://doi.org/10.1017/S0080454100006178</mixed-citation></ref><ref id="scirp.114207-ref8"><label>8</label><mixed-citation publication-type="other" xlink:type="simple">Rubin, D.B. and Thayer, D.T. (1982) EM Algorithms for ML Factor Analysis. Psychometrika, 47, 69-76. https://doi.org/10.1007/BF02293851</mixed-citation></ref><ref id="scirp.114207-ref9"><label>9</label><mixed-citation publication-type="other" xlink:type="simple">Adachi, K. (2013) Factor Analysis with EM Algorithm Never Gives Improper Solutions When Sampleand Initial Parameter Matrices Are Proper. Psychometrika, 78, 380-394. https://doi.org/10.1007/s11336-012-9299-8</mixed-citation></ref><ref id="scirp.114207-ref10"><label>10</label><mixed-citation publication-type="other" xlink:type="simple">Konishi, S. and Kitagawa, G. (2008) Information Criteria and Statistical Modeling. Springer, New York. https://doi.org/10.1007/978-0-387-71887-3</mixed-citation></ref><ref id="scirp.114207-ref11"><label>11</label><mixed-citation publication-type="other" xlink:type="simple">Kiers, H.A.L. (1994) Simplimax: Oblique Rotation to an Optimal Target with Simple Structure. Psychometrika, 59, 567-579. https://doi.org/10.1007/BF02294392</mixed-citation></ref><ref id="scirp.114207-ref12"><label>12</label><mixed-citation publication-type="other" xlink:type="simple">Browne, M.W. (1972) Oblique Rotation to a Partially Specified Target. British Journal of Mathematical and Statistical Psychology, 25, 207-212.  
https://doi.org/10.1007/BF02294392</mixed-citation></ref><ref id="scirp.114207-ref13"><label>13</label><mixed-citation publication-type="other" xlink:type="simple">Kiers, H.A.L. (1990) Majorization as a Tool for Optimizing a Class of Matrix Functions. Psychometrika, 55, 417-428. https://doi.org/10.1007/BF02294758</mixed-citation></ref><ref id="scirp.114207-ref14"><label>14</label><mixed-citation publication-type="other" xlink:type="simple">Kiers, H.A.L. (2002) Setting Up Alternating Least Squares and Iterative Majorization Algorithms for Solving Various Matrix Optimization Problems. Computational Statistics and Data Analysis, 41, 157-170.  
https://doi.org/10.1016/S0167-9473(02)00142-1</mixed-citation></ref><ref id="scirp.114207-ref15"><label>15</label><mixed-citation publication-type="other" xlink:type="simple">Akaike, H. (1974) A New Look at the Statistical Model Identification. IEEE Transactions on Automatic Control, 19, 716-723.  
https://doi.org/10.1109/TAC.1974.1100705</mixed-citation></ref><ref id="scirp.114207-ref16"><label>16</label><mixed-citation publication-type="other" xlink:type="simple">Schwarz, G. (1978) Estimating the Dimension of a Model. Annals of Statistics, 6, 461-464. https://doi.org/10.1214/aos/1176344136</mixed-citation></ref><ref id="scirp.114207-ref17"><label>17</label><mixed-citation publication-type="other" xlink:type="simple">Carlson, M. and Mulaik, S.A. (1993) Trait Ratings from Descriptions of Behavior as Mediated by Components of Meaning. Multivariate Behavioral Research, 28, 111-159.  
https://doi.org/10.1207/s15327906mbr2801_7</mixed-citation></ref></ref-list></back></article>