<?xml version="1.0" encoding="UTF-8"?><!DOCTYPE article  PUBLIC "-//NLM//DTD Journal Publishing DTD v3.0 20080202//EN" "http://dtd.nlm.nih.gov/publishing/3.0/journalpublishing3.dtd"><article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" dtd-version="3.0" xml:lang="en" article-type="research article"><front><journal-meta><journal-id journal-id-type="publisher-id">JMF</journal-id><journal-title-group><journal-title>Journal of Mathematical Finance</journal-title></journal-title-group><issn pub-type="epub">2162-2434</issn><publisher><publisher-name>Scientific Research Publishing</publisher-name></publisher></journal-meta><article-meta><article-id pub-id-type="doi">10.4236/jmf.2020.104038</article-id><article-id pub-id-type="publisher-id">JMF-103953</article-id><article-categories><subj-group subj-group-type="heading"><subject>Articles</subject></subj-group><subj-group subj-group-type="Discipline-v2"><subject>Business&amp;Economics</subject><subject> Physics&amp;Mathematics</subject></subj-group></article-categories><title-group><article-title>
 
 
  Topology Data Analysis Using Mean Persistence Landscapes in Financial Crashes
 
</article-title></title-group><contrib-group><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Alejandro</surname><given-names>Aguilar</given-names></name><xref ref-type="aff" rid="aff1"><sup>1</sup></xref><xref ref-type="corresp" rid="cor1"><sup>*</sup></xref></contrib><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Katherine</surname><given-names>Ensor</given-names></name><xref ref-type="aff" rid="aff1"><sup>1</sup></xref></contrib></contrib-group><aff id="aff1"><addr-line>Department of Statistics, Rice University, Houston, USA</addr-line></aff><pub-date pub-type="epub"><day>10</day><month>10</month><year>2020</year></pub-date><volume>10</volume><issue>04</issue><fpage>648</fpage><lpage>678</lpage><history><date date-type="received"><day>26,</day>	<month>September</month>	<year>2020</year></date><date date-type="rev-recd"><day>3,</day>	<month>November</month>	<year>2020</year>	</date><date date-type="accepted"><day>6,</day>	<month>November</month>	<year>2020</year></date></history><permissions><copyright-statement>&#169; Copyright  2014 by authors and Scientific Research Publishing Inc. </copyright-statement><copyright-year>2014</copyright-year><license><license-p>This work is licensed under the Creative Commons Attribution International License (CC BY). http://creativecommons.org/licenses/by/4.0/</license-p></license></permissions><abstract><p>
 
 
  Topological features in high dimensional time series are used to characterize changes in stock market dynamics over time. We explored the daily log returns of four major US stock market indices and 10 
  ETF
   sectors between January 2010-June 2020. Topological data analysis and persistence homology were used on two sequences of point cloud data sets the stock indices and the 
  ETF
   sectors, respectively. Using these sequences, the daily log returns, persistence diagrams, persistence landscapes, and mean landscapes were used to quantify topological patterns in the multidimensional time series. For example, norms of the persistence landscapes were generated to detect critical transitions in the daily log returns. To measure statistical significance, we implemented three permutation tests with a significance level α = 0.05 to determine if topological features change within a particular time frame by comparing sliding windows in the sequence of point cloud data sets. We found that between July 1, 2019 and July 1, 2020, there is evidence of changing structure in the US stock market. Critical transitions are identified by the statistical properties of the norms of the persistence landscape between contiguous daily sliding windows of the stock indices and 
  ETF
   sector series.
 
</p></abstract><kwd-group><kwd>Topological Data Analysis</kwd><kwd> Topological Time Series</kwd><kwd> Persistent Homology</kwd><kwd> Mathematical Finance</kwd></kwd-group></article-meta></front><body><sec id="s1"><title>1. Introduction</title><p>Topological data analysis (TDA) extracts topological features by examining the shape of the data through persistent homology to produce topological summaries. Two topological summaries, the persistent barcode [<xref ref-type="bibr" rid="scirp.103953-ref1">1</xref>] [<xref ref-type="bibr" rid="scirp.103953-ref2">2</xref>], and the persistent diagram [<xref ref-type="bibr" rid="scirp.103953-ref3">3</xref>], provide visual representation of persistent topological features. However, these topological summaries lack geometric properties and do not have a unique (Fr&#233;chet) mean [<xref ref-type="bibr" rid="scirp.103953-ref4">4</xref>], which makes it difficult to conduct statistical analysis and machine learning. In fact, Bubenik [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>] states effective algorithms do not exist for computing means for a wide variety of examples, but he notes that [<xref ref-type="bibr" rid="scirp.103953-ref6">6</xref>] and [<xref ref-type="bibr" rid="scirp.103953-ref7">7</xref>] have made noteworthy advancement in this direction.</p><p>In Bubenik [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>], the persistence landscape is provided as an alternative topological summary. The computation time is less for the persistence landscape than the persistence barcode and persistence diagram, because the persistence landscape is a sequence of piece-wise linear functions. Yet, the main advantage of using a persistence landscape is that they are situated in a separable Banach space, which means that we may use probability theory and random variables. After defining the persistence landscape and the norms of persistence landscapes, Bubenik [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>] further develops his work by introducing the mean or average persistence landscape, which may be used for statistical analysis and again is something the persistence diagram and persistence barcode do not have. Furthermore, he proves many statistical properties for using persistence landscapes, such as convergence, stability, the Central Limit Theorem, and the Strong Law of Large Numbers (SLLN), which is important, so that one may conduct statistical inference. In particular, Bubenik [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>] conducts a permutation test using multiple persistence landscapes to obtain a p-value (see Section 2.5, Section 2.6, Section 3.4), which is something that also cannot be done with persistence diagrams and persistence barcodes. This permutation test and the persistence landscape can be found in [<xref ref-type="bibr" rid="scirp.103953-ref8">8</xref>]. Other notable topological data analysis applications include the discovery of a subgroup of breast cancers [<xref ref-type="bibr" rid="scirp.103953-ref9">9</xref>], an understanding of the topology of the space of natural images [<xref ref-type="bibr" rid="scirp.103953-ref10">10</xref>], brain signals [<xref ref-type="bibr" rid="scirp.103953-ref11">11</xref>], and pre-clinical spinal cord injury [<xref ref-type="bibr" rid="scirp.103953-ref12">12</xref>].</p><p>With this alternative topological summary and the ability to conduct statistical inference, we bring our focus to critical transitions in complex dynamical systems, in particular, the financial market. Scheffer et al. [<xref ref-type="bibr" rid="scirp.103953-ref13">13</xref>] asserted that predicting critical transitions in a complex dynamical system prior to occurring is unreliable and very challenging, because the state of the complex system may not fluctuate substantially before reaching a critical threshold. However, in their work, they found for vast classes of systems, early warning signals may exist to indicate when a critical transition is imminent. Even though Scheffer et al. [<xref ref-type="bibr" rid="scirp.103953-ref13">13</xref>] presented examples of early warning signals in ecosystems, time series, climate dynamics, and epileptic seizures, there is not an example of an early warning signal for a financial crash. While Scheffer et al. [<xref ref-type="bibr" rid="scirp.103953-ref13">13</xref>] clarified that some predictability may be employed by experts, in general financial markets are complicated to predict. Although Scheffer et al. [<xref ref-type="bibr" rid="scirp.103953-ref13">13</xref>] referenced excellent works with financial early warning indicators, such as the VIX (volatility based index), systematic relationships in the variance and first order auto-correlation, and correlation increases across returns in falling markets, it was not relevant to our study, which leads us to research more about financial market dynamics.</p><p>Ensor and Koev [<xref ref-type="bibr" rid="scirp.103953-ref14">14</xref>] focused on the multivariate GARCH (MGARCH) and the hierarchical regime switching dynamic covariance (HRSDC) models, in which both models examined the co-variance structure within and between market sectors for the time period January 2, 1998 to December 2001. The HRSDC model provided early detection of several anomalous behaviors, such as the decline of Enron, the unusual returns for Silicon Graphics, and the fall of Lehmann Brothers. The detection of these anomalous behaviors is based on price movement of individual securities when viewed as a system of securities with the correlation within and between sectors. Therefore, Ensor and Koev [<xref ref-type="bibr" rid="scirp.103953-ref14">14</xref>] demonstrated a nested model is useful for understanding the correlation structure between different market sectors and how these sectors interacted as the market changes between regimes.</p><p>While Ensor and Koev [<xref ref-type="bibr" rid="scirp.103953-ref14">14</xref>] study prompted us to use ETF sectors in our study and is effective in identifying anomalous behaviors, such as the decline of Enron and the fall of Lehnman Brothers, our interest is to use TDA to detect early warning signals for financial crashes and examine topological features changing within time with statistical significance. A number of recent studies have explored the use of TDA on financial time series data to detect early warning signals of financial crashes. Gidea [<xref ref-type="bibr" rid="scirp.103953-ref2">2</xref>] analyzed the cross correlation network of the daily returns (adjusted closing prices) of the Dow Jones Industrial Average (DJIA) stocks listed as February 19, 2008 from January 2004 to September 2008. They tracked the topological changes when approaching a critical transition and showed some presence of early signs of a critical transition. On the other hand, Gidea et al. [<xref ref-type="bibr" rid="scirp.103953-ref15">15</xref>] analyzed four major cryptocurrencies (Bitcoin, Ethereum, Litecoin, and Ripple) before the beginning of 2018 and showed these cryptocurriences exhibiting highly erratic behavior. The paper introduced a method that combines TDA with machine learning to understand what happens before a critical transition. Moreover, they use Takens’ theorem, the time delay embedding theorem, and C<sup>1</sup>-norms of persistence landscapes. While the paper has valid analysis, our interest is in stocks and ETF sectors rather than cyprtocurrencies.</p><p>Alternatively, Gidea and Katz [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>] investigated the daily log-returns of four stock indices (DJIA, S&amp;P500, NASDAQ, and Russell 2000) from December 23, 1987 and December 08, 2016, where the topological properties of these stock indices were examined. This paper uses a sequence of a point cloud data set with a sliding window. Gidea and Katz [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>] provided an excellent framework using persistence diagrams, persistence landscapes, and norms for persistence landscapes and we were able to replicate all of their results for 2000 and 2008 crashes. They demonstrated that the variance as defined in [<xref ref-type="bibr" rid="scirp.103953-ref17">17</xref>] shows rising trends, we are not convinced about the average spectral densities and auto-correlation function (ACF) with their associated Kendall-Mau tests demonstrated trends.</p><p>While these papers provide insightful groundwork for TDA in financial markets and cryptocurrencies, such as showing how to use cross correlation networks to track topological changes, using TDA with machine learning to understand what happens before a critical transition, and using the norms of persistence landscapes to indicate an approaching critical transition, these financial papers lack statistical inference. We are motivated to explore how the topological features change within a given time period for stocks and ETF sectors and find any statistical significant using a permutation test [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>] [<xref ref-type="bibr" rid="scirp.103953-ref8">8</xref>], which we discuss in detail in Section 2.5, Section 2.6, and Section 3.4.</p><p>While we acknowledge the previous cited authors, we deem our contributions as an empirical framework that adapts their analytical models to new data sets and expand by conducting statistical inference. Similar to Gidea and Katz [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>], we investigate the same four major indices (DJIA, S&amp;P500, NASDAQ, and Russell 2000), but we extend our data set to include 10 ETF sectors (Consumer Discretionary, Consumer Staples, Energy, Financials, Health Care, Industrials, Materials, Information Technology, Utilities, and Index) for January 4, 2010-July 1, 2020 to examine their topological features to detect a critical transition or transitions. Moreover, we generate several topological summaries, norms for persistence landscapes p = 1 and p = 2 , and conduct statistical inference on how these topological features change over time. In particular, we want to compare only sliding windows within a sliding step of one day from each other, which will be done separately for all the stock indices and for all the ETF sectors. We also compare all the stock indices against ETF sectors within the same sliding window. Our hypotheses tests will distinguish for two groups at a time if the means of topological features are the same either within a sliding step of one day in their respective sliding windows or within the same sliding window. The statistical tests of interest have not been seen before in any financial papers and will be our main contribution. The remainder of this paper is organized as follows.</p><p>In Section 2, we provide background information on algebraic topology, homology, constructing the Vietoris-Rips complex, persistent homology, topological summaries, norms of persistent landscapes, and statistical inference. In Section 3, we outline our methods for obtaining the data, constructing a sequence of a point cloud data, using persistent homology on a sequence of a point cloud data set, generating topological summaries, and performing statistical inference. In Section 4, we present our findings from our data. In Section 5, we discuss and provide an interpretation of our results. In Section 6, we conclude the paper.</p></sec><sec id="s2"><title>2. Background</title><p>This study presents a topological data analysis of financial time series data. Here we provide background material about four relevant areas: algebraic topology, homology, topological summaries, and norms for persistent landscapes. We apply topological data analysis to a sequence of point cloud data sets to examine their topological properties within a point cloud matrix of d 1-dimensional time series. For our analysis, a sequence of point cloud data sets denoted X n is shown below:</p><p>X 1 = [ x ( t 1 ) x ( t 2 ) ⋮ x ( t w ) ] = [ x 1 1 ⋯ x 1 d x 2 1 ⋯ x 2 d ⋮ ⋱ ⋮ x w 1 ⋯ x w d ]               ⋮ X q = [ x ( t q ) x ( t q + 1 ) ⋮ x ( t q + w − 1 ) ] = [ x q 1 ⋯ x q d x q + 1 1 ⋯ x q + 1 d ⋮ ⋱ ⋮ x q + w − 1 1 ⋯ x q + w − 1 d ] , (1)</p><p>where each point in the sequence is expressed as x ( t n ) = ( x q 1 , x q 2 , ⋯ , x q d ) ∈ ℝ d , d is the column number from a 1-dimensional time series, w is the sliding window size for a certain number of trading days ( n t d ) with a sliding step of one day, and n = 1,2, ⋯ , q . To obtain q, the difference is taken between the total number of days of the daily log returns ( n d l r ) and one less than the sliding window size w − 1 , so that q becomes q = n d l r − ( w − 1 ) or q = n d l r − w + 1 . The total number of days of the daily log returns ( n d l r ) is the total number of trading days ( n t d ) minus 1 or n d l r = n t d − 1 . To approximate the daily log returns, the formula is discussed in Section 3.1. So, every point cloud is compromised of a d &#215; w matrix, where w &gt; d [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>]. Note our method uses a sliding window w as seen in [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>] and it does not apply the sliding window embedding theorem or Takens’ theorem. In the next two subsections, we provide background information on algebraic topology and persistent homology, so that for every point cloud, we generate topological summaries and compute their L p norms based on their corresponding persistence landscapes to conduct statistical inference. For a more in depth background, we refer readers to [<xref ref-type="bibr" rid="scirp.103953-ref3">3</xref>] [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>] [<xref ref-type="bibr" rid="scirp.103953-ref18">18</xref>] [<xref ref-type="bibr" rid="scirp.103953-ref19">19</xref>] [<xref ref-type="bibr" rid="scirp.103953-ref20">20</xref>] [<xref ref-type="bibr" rid="scirp.103953-ref21">21</xref>].</p><sec id="s2_1"><title>2.1. Algebraic Topology</title><p>To produce topological summaries, we must first construct a Vietoris-Rips filtration for each point cloud in a sequence of point cloud data sets, which requires understanding simplices and simplicial complexes and are defined below [<xref ref-type="bibr" rid="scirp.103953-ref19">19</xref>]:</p><p>Definition 2.1 Let { a 0 , ⋯ , a n } be a geometrically independent set in ℝ N . We define the n-simplex σ spanned by a 0 , ⋯ , a n to be the set of all points x of ℝ N such that:</p><p>x = ∑ i = 0 n     t i a i ,   where   t = ∑ i = 0 n     t i = 1 , (2)</p><p>and t i ≥ 0 for all i.</p><p>Definition 2.2 A simplicial complex K in ℝ N in a collection of simplices in ℝ N such that:</p><p>&#183; Every face of a simplex of K is in K.</p><p>&#183; The intersection of any two simplexes of K is a face.</p></sec><sec id="s2_2"><title>2.2. Homology</title><p>In homology, we are interested in a vector space H i ( X ) to a space X for each natural number i ∈ { 0,1,2 , ⋯ } , because H i ( X ) counts the number of k-dimensional holes in X. For example, H 0 ( X ) counts the number of 0-dimensional holes or the number of connected components in X, while H 1 ( X ) counts number of 1-dimensional holes or the number of loops in X. Furthermore, the algebraic structures must be homotopy invariant, meaning they must not change through deformations. Yet, it is very challenging to determine the homology of arbitrary topological spaces, because it is computationally inefficient, so instead we approximate using simplicial complexes.</p><p>Now that simplicial complexes have been defined, we are introducing the p<sup>th</sup> homology of a simplicial complex K. First, we denote the field with two elements as F 2 . Second, for a given simplicial complex K, we let C p ( K ) denote the F 2 -vector space with basis given by the p-simplices of K. Third, for any p ∈ { 1,2, ⋯ } , we define the linear map:</p><p>∂ p : C p ( K ) → C p − 1 ( K ) : σ ↦ ∑ τ ⊂ σ , τ ∈ K p − 1 τ , (3)</p><p>The kernel of ∂ p : C p ( K ) → C p − 1 ( K ) is the subgroup ∂ p − 1 ( 0 ) of C p ( K ) and is called the group of p-cycles. The image of ∂ p + 1 : C p + 1 ( K ) → C p ( K ) is the image ∂ p + 1 is the subgroup of ∂ p + 1 ( C p + 1 ( K ) ) of C p ( K ) and is called the group of p-boundaries [<xref ref-type="bibr" rid="scirp.103953-ref19">19</xref>].</p><p>Definition 2.3 For any p ∈ { 0,1,2, ⋯ } , the p<sup>th</sup> homology of a simplicial complex K is the quotient vector space is defined as:</p><p>H p ( K ) = kernel ( ∂ p ) / image ( ∂ p + 1 ) . (4)</p><p>Its dimension is defined by:</p><p>β p ( K ) : = dim H p ( K ) = dimkernel ( ∂ p ) − dimimage ( ∂ p + 1 ) , (5)</p><p>which is called the p<sup>th</sup> Betti number of K [<xref ref-type="bibr" rid="scirp.103953-ref22">22</xref>].</p><p>The p-cycles that are not boundaries represent p-dimensional holes, which the p<sup>th</sup> Betti number counts. For the p<sup>th</sup> homology of a filtered simplicial complex K, we apply definition 2.3 and define as:</p><p>Definition 2.4 Let K be a finite simplicial complex, and let K 1 ⊂ K 2 ⊂ K 3 … ⊂ K l = K be a finite sequence of nested subcomplexes of K. The simplicial complex K with such a sequence of subcomplexes is called a filtered simplicial complex. The p<sup>th</sup> persistent homology of K is the pair</p><p>( { H p ( K i ) } 1 ≤ i ≤ l , { f p i , j } 1 ≤ i ≤ j ≤ l ) ,</p><p>where i , j ∈ { 1, ⋯ , l } for all i ≤ j , f p i , j : H p ( K i ) → H p ( K j ) are the linear maps induced by the inclusion maps K i → K j [<xref ref-type="bibr" rid="scirp.103953-ref22">22</xref>].</p><p>The p<sup>th</sup> persistent homology of a filtered simplicial complex provides more information about the maps between each subcomplex than the homologies of single subcomplexes, which is explained further in Section 2.2.2. While there are several filtered simplicial complexes, such as the Cech, Alpha, and Delaunay, we chose the Vietoris-Rips complex, because it is computationally efficient [<xref ref-type="bibr" rid="scirp.103953-ref22">22</xref>].</p><sec id="s2_2_1"><title>2.2.1. Vietoris-Rips Construction</title><p>Definition 2.5 Let X = { x 1 , ⋯ , x n } be a collection of points in ℝ d . Given a distance ϵ &gt; 0 , R ( X , ϵ ) denotes the simplicial complex on n vertices x 1 , ⋯ , x n , where an edge between the vertices x i and x j with i ≠ j is included if and only if d ( x i , x j ) ≤ ϵ or generally the k-simplex are included with vertices x i 0 , ⋯ , x i k if and only if all of the pairwise distances are at most ϵ . This type of simplicial complex is called a Vietoris-Rips complex [<xref ref-type="bibr" rid="scirp.103953-ref8">8</xref>] [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>].</p><p>When ϵ &lt; ϵ ′ , the Vietoris-Rips complex forms a filtration, R ( X , ϵ ) ⊆ R ( X , ϵ ′ ) , which by definition 2.4 is a filtered simplicial complex. While there is no clear criteria for select ϵ ′ , [<xref ref-type="bibr" rid="scirp.103953-ref23">23</xref>] used ϵ ′ = 0.05 in their study. In this study, the Vietoris-Rips complex of X n is denoted as R ( X n , ϵ ) and follows definition 2.5, where X n is a sequence of point cloud data sets as given by Equation (1). Moreover, the filtration of R ( X n , ϵ ) ⊆ R ( X , ϵ ′ ) is shown below:</p><p>R ( X 1 , ϵ ) ⊆ R ( X 1 , ϵ ′ )                           ⋮ R ( X q , ϵ ) ⊆ R ( X q , ϵ ′ ) , (6)</p><p>where q is the difference between the number of the daily log returns and the sliding window ( w + 1 ) or q = n d l r − w + 1 . By definition 2.4, R ( X n , ϵ ) is a filtered simplicial complex.</p></sec><sec id="s2_2_2"><title>2.2.2. Persistent Homology</title><p>Using definition 2.4 and definition 2.5, it is possible to find the p-dimensional homology of the Vietoris-Rips complex of X n labelled as H p ( R ( X n , ϵ ) ) with coefficients in field ℤ / 2 ℤ for small values of p and for different values of ϵ [<xref ref-type="bibr" rid="scirp.103953-ref8">8</xref>]. Recall from section 2.2, H p ( K i ) is a vector space and β p ( K i ) counts the number of p-dimensional holes. When ϵ &lt; ϵ ′ , we apply definition 2.4 to the filtration R ( X n , ϵ ) ⊆ R ( X n , ϵ ′ ) , which induces linear maps f p i , j : H p ( R ( X n , ϵ ) ) → H p ( R ( X n , ϵ ′ ) ) as seen below:</p><p>H p ( R ( X 1 , ϵ ) ) → H p ( R ( X 1 , ϵ ′ ) )                                       ⋮ H p ( R ( X q , ϵ ) ) → H p ( R ( X q , ϵ ′ ) ) , (7)</p><p>where q = n d l r − w + 1 . Each H p ( R ( X n , ϵ ) ) is a vector space whose generators correspond to holes in R ( X n , ϵ ) , and the linear maps f p i , j allow us to track the generators from H p ( R ( X n , ϵ ) ) → H p ( R ( X n , ϵ ′ ) ) . A suitable basis is selected by applying the Fundamental Theorem of Persistence Homology.</p><p>Theorem 2.1 (Fundamental Theorem of Persistent Homology) The Fundamental Theorem of Persistent Homology states there is a choice of basis vectors H p ( K i ) for each i ∈ { 1, ⋯ , l } such that each map is determined by a bipartite matching of basis vectors [<xref ref-type="bibr" rid="scirp.103953-ref22">22</xref>].</p><p>Given Theorem 2.1, there is a choice of basis vectors of H p ( R ( X n , ϵ ) ) , such that one may construct a well-defined and unique collection of disjoint half-open intervals, where a generator x ∈ H p ( R ( X n , ϵ ) ) corresponds to a half-open interval [ b i , d i ) , which represents the lifetime of x. The endpoints b i and d i refer to x first appearing and finally disappearing respectively in R ( X n , ϵ ) . Specifically, if x ≠ 0 is not in the image of f p b i − 1 , b i , then x is born in H p ( R ( X n , ϵ ) ) . Conversely, if d i &gt; b i is the smallest index for which f p b i , d i ( x ) = 0 , then x dies in H p ( R ( X n , ϵ ) ) . Persistence is determined by a generator’s lifetime in the half-open interval, where a generator is considered more persistent the longer it appears in the half-open interval. If f p b i , d i ( x ) = 0 for all b i &gt; d i in I j , then x lives forever, and its lifetime is represented by the interval [ b i , ∞ ) [<xref ref-type="bibr" rid="scirp.103953-ref22">22</xref>]. Then, the set of vector spaces H p ( R ( X n , ϵ ) ) together with the corresponding linear maps is referred to as a persistence module, which is the foundation for constructing topological summaries.</p></sec></sec><sec id="s2_3"><title>2.3. Topological Summaries</title><p>To visualize, construct, and produce topological summaries, Theorem 2.1 is used to select the choice of basis vectors from H p ( R ( X n , ϵ ) ) and the corresponding linear maps f p b i , d i , in which all topological summaries are derived from the persistent modules.</p><sec id="s2_3_1"><title>2.3.1. Persistence Module, Persistence Barcode, Persistence Diagram</title><p>Definition 2.6 A persistence module is defined as a vector space M α for all a ∈ ℝ and linear maps M ( a ≤ b ) : M a → M b for all a ≤ b such that:</p><p>1) M ( a ≤ a ) is the identity map;</p><p>2) For all a ≤ b ≤ c , M ( b ≤ c ) ∘ M ( a ≤ b ) = M ( a ≤ c ) .</p><p>For additional information about the construction of a persistence module, see [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>]. There are three main types of topological summaries associated with a persistence module. The first type of topological summary is a called a barcode. It represents a finite collection of disjoint half-open intervals I j , in which each interval’s endpoints are a birth-death pairs, (b) and (d) respectively. In particular, an interval starts with the time of birth (b) and ends with the time of death (d) of a topological feature. The p<sup>th</sup> barcode is denoted by B p = { I j } . A topological feature’s survival or persistence is represented by the interval’s length. The second type of topological summary is the p<sup>th</sup> persistence diagram, which is denoted as D p = { ( b i , d i ) } i ∈ I j , where b i and d i are the bar codes intervals’ end points and − ∞ &lt; b i &lt; d i &lt; ∞ .</p><p>Unfortunately, the geometric properties of the barcodes and persistence diagrams present a difficult challenge for the calculation of means and variances, since two barcodes or two persistence diagrams may not have the same unique Friechet mean, which means statistical inference cannot be done. While the barcode and the persistence diagram are conventional topological summaries, Bubenik [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>] showed how the persistence landscape is a better alternative.</p></sec><sec id="s2_3_2"><title>2.3.2. Persistence Landscapes and Mean Landscape</title><p>Bubenik and Dlotko [<xref ref-type="bibr" rid="scirp.103953-ref18">18</xref>] proved numerous statistical properties of persistence landscapes that we may use for statistical inference, such as stability, convergence, central limit theorem, and strong law of large numbers. The persistent landscape and mean landscape are also used as topological summaries to indicate how persistence changes by examining the number of peaks. First, given a pair of numbers ( b , d ) with b &lt; d , the piecewise linear (PL) function f ( b , d ) : ℝ → [ 0, ∞ ] is defined by [<xref ref-type="bibr" rid="scirp.103953-ref18">18</xref>]:</p><p>f ( b , d ) = ( 0 if   x ∉ ( b , d ) x − b if   x ∈ ( b , b + d 2 ] − x + d if   x ∈ ( b + d 2 , d ) (8)</p><p>Second, given a persistence module, M, the persistence landscape may be defined as the function λ : ℕ &#215; ℝ → R given by:</p><p>λ ( k , t ) = sup ( h &gt; 0 | rank M ( t − h ≤ t + h ) &gt; k ) . (9)</p><p>Third, given a persistence diagram D p = { ( b i , d i ) } i ∈ I for b &lt; d , f ( b , d ) ( t ) = max ( 0, min ( b + t , d − t ) ) , and the persistence landscape is defined as follows:</p><p>λ ( k , t ) = kmax { f p b i , d i ( t ) | ( b i , d i ) ∈ D p ( t ) } i ∈ I , (10)</p><p>where kmax denotes the k<sup>th</sup> largest element. Using Equation (10) for X n , the persistence landscape of X n denoted by λ ( X n ) is the following:</p><p>λ ( X 1 ) = k-max { f p b i , d i ( X 1 ) | ( b i , d i ) ∈ D p ( X 1 ) } i ∈ I                                                                 ⋮ λ ( X q ) = k-max { f p b i , d i ( X q ) | ( b i , d i ) ∈ D p ( X q ) } i ∈ I , (11)</p><p>where q = n d l r − w + 1 . This results in the following lemma from [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>]:</p><p>Lemma 2.2</p><p>The persistence landscape has the following properties:</p><p>1) λ k ( t ) ≥ 0 ,</p><p>2) λ k ( t ) ≥ λ k + 1 ( t ) , and</p><p>3) λ k ( t ) 0 is 1-Lipschitz.</p><p>From Equation (10), the persistence landscape is obtained and used to calculate the mean landscape, which is defined below:</p><p>Definition 2.7 Let Y 1 , ⋯ , Y n be independent and identically distributed copies of Y, and let Λ 1 , ⋯ , Λ n be corresponding persistence landscapes. The mean landscape Λ &#175; n is given by the point wise mean, in particular, Λ &#175; n ( ω ) = Λ &#175; n , where</p><p>λ &#175; n ( k , t ) = 1 n ∑ i = 1 n     λ i ( k , t ) . (12)</p><p>Using Equation (12) for X n , we have the following:</p><p>λ &#175; n ( X 1 ) = 1 n ∑ i = 1 n     λ i ( X 1 )                           ⋮ λ &#175; n ( X q ) = 1 n ∑ i = 1 n     λ i ( X q ) , (13)</p><p>where q = n d l r − w + 1 . The mean landscape is used in section 2.5 and section 2.6.</p></sec></sec><sec id="s2_4"><title>2.4. Norms for Persistence Landscapes</title><p>Gidea and Katz [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>] applied L p norms of the persistence landscapes to identify the signs of a financial crash, which usually occurs within a time of high variance and cross-correlations among stocks or ETFs, and demonstrated that L 1 and L 2 norms of the persistence landscapes of four stock indices exhibited significant rising trends before the financial crashes. We adopt their approach in our study.</p><p>Therefore, for real valued functions on ℝ &#215; ℝ , for 1 ≤ p &lt; ∞ , p-norms of persistence landscapes are defined as:</p><p>‖ λ ‖ p = ∑ i = 1 ∞ [ ∫ − ∞ ∞     λ k ( t ) p d t ] 1 p , (14)</p><p>and for p = ∞ ,</p><p>‖ λ ‖ ∞ = sup k , t λ k ( t ) . (15)</p><p>Applying Equation (14) to our sequence of point cloud data sets X n results in:</p><p>‖ λ ( X 1 ) ‖ p = ∑ i = 1 ∞ [ ∫ − ∞ ∞     λ ( X 1 ) p d t ] 1 p ‖ λ ( X q ) ‖ p = ∑ i = 1 ∞ [ ∫ − ∞ ∞     λ ( X q ) p d t ] 1 p , (16)</p><p>where q = n d l r − w + 1 .</p></sec><sec id="s2_5"><title>2.5. Statistical Inference: Part I</title><p>To compare the topological features between two groups, the persistence landscape is used to conduct a hypothesis test and statistical inference, which require several assumptions provided by [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>]. First, the persistence landscapes lie in a separable Banach space L p ( S ) for 1 ≤ p ≤ ∞ , where S = ℕ &#215; ℝ . Second, Y is to be a random variable on some underlying probability space ( Ω , F , P ) with a corresponding landscape Λ . Third, if we have ω ∈ Ω , then Y ( ω ) is the random variable and Λ ( ω ) = λ ( Y ( ω ) ) : = λ is the corresponding topological summary statistic. To avoid confusion, we use Y instead of X as a random variable, because our sequence of point cloud data sets uses the variable X n . In addition, Bubenik [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>] proved the convergence of persistence landscapes using the Strong Law of Large Numbers and the Central Limit Theorem, which is extremely important for setting up our random variables and hypothesis test. Our random variable Y is defined as:</p><p>Y = f ( λ ( k , t ) ) = ∑ k ∫ ℝ   λ k ( t ) d t , (17)</p><p>where f ∈ L b ( S ) is a continuous linear functional, 1 a + 1 b = 1 , and Y satisfies the (SLLN) and (CLT) as seen in [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>], which implies Y has an adequate sample size and follows an approximately normal distribution.</p><p>The statistical properties and definitions above are utilized to a conduct hypothesis tests with corresponding p-value based on a permutation test. To compare the topological features of two groups, Y 1 and Y 2 , where k 1 and k 2 are samples taken from these groups respectively, and Λ 1 and Λ 2 are the corresponding landscapes respectively. The associate sample values of Y 1 and Y 2 are denoted as y 1 1 , ⋯ , y 1 k 1 and y 2 1 , ⋯ , y 2 k 2 and the corresponding landscapes of these sample values are labelled as λ 1 1 , ⋯ , λ 1 k 1 and λ 2 1 , ⋯ , λ 2 k 2 . We apply Equation (17) to Y 1 and Y 2 , so the functional of Y 1 and Y 2 are as follows:</p><p>Y 1 = f ( y 1 1 ) , ⋯ , f ( y 1 k 1 ) = f ( λ 1 1 ( k , t ) ) , ⋯ , f ( λ 1 k 1 ( k , t ) ) = ∑ i = 1 k 1 ∫ ℝ     λ 1 i ( k , t ) d t Y 2 = f ( y 2 1 ) , ⋯ , f ( y 2 k 2 ) = f ( λ 2 1 ( k , t ) ) , ⋯ , f ( λ 2 k 2 ( k , t ) ) = ∑ i = 1 k 2 ∫ ℝ     λ 2 i ( k , t ) d t . (18)</p><p>Recall the sample mean is Y &#175; = 1 n ∑ i = 1 n     Y i , so the sample means of the Y 1 and Y 2 are the following:</p><p>Y &#175; 1 = 1 k 1 ∑ i = 1 k 1     f ( y 1 i ) = 1 k 1 ∑ i = 1 k 1     f ( λ 1 i ( k , t ) ) Y &#175; 2 = 1 k 2 ∑ i = 1 k 2     f ( y 2 i ) = 1 k 2 ∑ i = 1 k 2     f ( λ 2 i ( k , t ) ) , (19)</p><p>where again k 1 and k 2 are the samples taken from Y 1 and Y 2 . We assume that μ 1 and μ 2 are the expectations of Y 1 and Y 2 . So, μ 1 and μ 2 are assumed to be the population means of Y 1 and Y 2 . Therefore, the statistical hypothesis is:</p><p>H 0 : μ 1 = μ 2   H a : μ 1 ≠ μ 2 . (20)</p><p>To test the null-hypothesis, we use a two sample permutation test. Let</p><p>t = | Y &#175; 1 − Y &#175; 2 | V a r ( Y 1 ) k 1 + V a r ( Y 2 ) k 2 . (21)</p><p>Using Equation (21), t 1 , ⋯ , t m of the test statistic are calculated for permutations s = 1 , ⋯ , m . The observed value of the test statistic is expressed as t observed . The p-value is calculated by comparing t observed with t s and averaging the number of times t observed ≤ t s . Thus, Equation (21) becomes:</p><p>t { 1, Y 1 , Y 2 } = | Y &#175; 1 − Y &#175; 2 | V a r ( Y 1 ) k 1 + V a r ( Y 2 ) k 2                                 ⋮ t { m , Y 1 , Y 2 } = | Y &#175; 1 − Y &#175; 2 | V a r ( Y 1 ) k 1 + V a r ( Y 2 ) k 2 . (22)</p><p>A general form of Equation (22) is:</p><p>t { s , Y 1 , Y 2 } = | Y &#175; 1 − Y &#175; 2 | V a r ( Y 1 ) k 1 + V a r ( Y 2 ) k 2 . (23)</p><p>Hence, using Equation (22) and every instance where t observed ≤ t s , the p-value is obtained as:</p><p>p -value { Y 1 , Y 2 } = 1 m ∑ i = 1 m     t { i , Y 1 , Y 2 } . (24)</p><p>To measure the statistical significance, [<xref ref-type="bibr" rid="scirp.103953-ref8">8</xref>] used a significance level α = 0.05 in their study, which we incorporate in our study. We may apply the above assumptions, equations, and definitions to compare the topological features of more groups.</p></sec><sec id="s2_6"><title>2.6. Statistical Inference: Part II</title><p>Instead of conducting one hypothesis test, multiple hypotheses tests are conducted to determine how the topological features in our sequence of point cloud data sets X n change within a particular time frame. The hypotheses tests are done on all the sliding window matrices within X n . In particular, two adjacent sliding window matrices are compared, where adjacent means the sliding window matrices differ by a sliding step of one day. For example, the sliding window matrices X 1 and X 2 would be compared, while the sliding window matrices X 1 and X 3 would not be compared. Therefore, the assumptions, equations, and definitions from section 2.5 are applied to X n . When hypotheses tests are performed, there are q = n d l r − w + 1 random variables (see Equation (1)), which is also the size of the sequence of the point cloud data set X n .</p><p>So, we let Y 1 , Y 2 , ⋯ , Y q be random variables, where k 1 , k 2 , ⋯ , k q are taken as samples from these groups respectively, and Λ 1 , Λ 2 , ⋯ , Λ q are the corresponding landscapes respectively. The associate sample values of Y 1 , Y 2 , ⋯ , Y q are denoted as y 1 1 , ⋯ , y 1 k 1 , y 2 1 , ⋯ , y 2 k 2 , ⋯ , y q 1 , ⋯ , y q k q , and the corresponding landscapes of these sample values are labelled as λ 1 1 , ⋯ , λ 1 k 1 , λ 2 1 , ⋯ , λ 2 k 2 , ⋯ , λ q 1 , ⋯ , λ q k q . The functional in Equation (17) is used to define the following for Y 1 , Y 2 , ⋯ , Y q :</p><p>Y 1 = ∑ i = 1 k 1 ∫ ℝ     λ 1 i ( X 1 ) d t                     ⋮ Y q = ∑ i = 1 k q ∫ ℝ     λ j i ( X q ) d t , (25)</p><p>where q = n d l r − w + 1 . Recall the sample mean is Y &#175; = 1 n ∑ i = 1 n     Y i , so the sample means of the Y 1 , Y 2 , ⋯ , Y q as follows:</p><p>Y &#175; 1 = 1 k 1 ∑ i = 1 k 1     f ( λ i ( X 1 ) ) Y &#175; q = 1 k q ∑ i = 1 k q     f ( λ i ( X q ) ) , (26)</p><p>where q = n d l r − w + 1 . We assume that μ 1 , μ 2 , ⋯ , μ q are the expectations of Y 1 , Y 2 , ⋯ , Y q . So, μ 1 , μ 2 , ⋯ , μ q are assumed to be population means of Y 1 , Y 2 , ⋯ , Y q , and the statistical hypotheses are:</p><p>H 0 : μ 1 = μ 2   H a : μ 1 ≠ μ 2                                     ⋮ H 0 : μ q − 1 = μ q   H a : μ q − 1 ≠ μ q , (27)</p><p>where q = n d l r − w + 1 . To test the null-hypothesis, we use a two sample permutation test with statistics,</p><p>t { Y 1 , Y 2 } = | Y &#175; 1 − Y &#175; 2 | V a r ( Y 1 ) k 1 + V a r ( Y 2 ) k 2                                     ⋮ t { Y q − 1 , Y q } = | Y &#175; q − 1 − Y &#175; q | V a r ( Y q − 1 ) k q − 1 + V a r ( Y q ) k q . (28)</p><p>where q = n d l r − w + 1 . Using Equation (28), t 1 , ⋯ , t m of the test statistic are calculated for permutations s = 1 , ⋯ , m . The observed value of the test statistic is expressed as t observed . The p-value is calculated by comparing t observed with t s and averaging the number of times t observed ≤ t s . Using Equation (23), Equation (28) becomes:</p><p>t { s , Y 1 , Y 2 } = | Y &#175; 1 − Y &#175; 2 | V a r ( Y 1 ) k 1 + V a r ( Y 2 ) k 2                                     ⋮ t { s , Y q − 1 , Y q } = | Y &#175; q − 1 − Y &#175; q | V a r ( Y q − 1 ) k q − 1 + V a r ( Y q ) k q , (29)</p><p>where q = n d l r − w + 1 . Hence, using Equation (29) and every instance where t observed ≤ t s , the p-value is obtained as:</p><p>p -value { Y 1 , Y 2 } = 1 m ∑ i = 1 m     t { i , Y 1 , Y 2 } p -value { Y q − 1 , Y q } = 1 m ∑ i = 1 m     t { i , Y q − 1 , Y q } , (30)</p><p>where q = n d l r − w + 1 . In our study, we also conduct hypotheses tests between two sequences of point cloud data sets, X n 1 and X n 2 , within the same sliding window, so using the same assumptions, definitions, and results from this section. The only difference is a change in subscripts and superscripts. This case is presented in Section 3.4.</p></sec></sec><sec id="s3"><title>3. Methods</title><p>In this section, we describe the methods to obtain the data and analyze the financial time series using topological data analysis, statistical inference, and RStudio [<xref ref-type="bibr" rid="scirp.103953-ref23">23</xref>]. The data, which were obtained from Yahoo Finance, consisted of daily adjusted closing prices (amended for corporate actions such as stocks and dividends) for four major US stock indices: S&amp;P 500, DJIA, NASDAQ, and Russell 2000 and 10 ETF sectors between January 4, 2010 and July 1, 2020 (2641 trading days). During this time period, a decline in the daily log returns happened on March 16, 2020. In order to examine this date of interest, we limited our data sets to 1001 trading days ( n t d ) before March 16, 2020 to observe any patterns in the L p norms and determine any critical thresholds. To analyze the data, we first approximated the daily log returns of the adjusted closing prices. A return is defined as r t = ( x t − x t − 1 x t − 1 ) , where x t is the actual value (adjusted closing price) of the desired stock index or ETF sector. The daily log returns are defined as:</p><p>log 10 ( x t x t − 1 ) = log 10 ( x t ) − log 10 ( x t − 1 ) ≈ r t ,</p><p>which is an approximation of a return [<xref ref-type="bibr" rid="scirp.103953-ref24">24</xref>]. Since the daily log returns are forward daily changes, then the time frame of the daily log returns is from January 5, 2010 to June 30, 2020.</p><sec id="s3_1"><title>3.1. Point Cloud Data</title><p>After approximating the daily log returns, we designed two sequences of point cloud data sets, each with a sliding window of w = 50 and a sliding step of one day, which is based on the same method found in [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>]. The first sequence of point cloud data set denoted by X n S I examined the four major US stock indices ( d = 4 ), which resulted in a 4 &#215; 50 matrix for each individual point cloud for a total of q = n d l r − w + 1 = ( 1001 − 1 ) − ( 50 − 1 ) = 951 point clouds as seen below from using Equation (1):</p><p>X 1 S I = [ x ( t 1 ) x ( t 2 ) ⋮ x ( t 50 ) ] = [ x 1 1 ⋯ x 1 4 x 2 1 ⋯ x 2 4 ⋮ ⋱ ⋮ x 50 1 ⋯ x 50 4 ]               ⋮ X 951 S I = [ x ( t 951 ) x ( t 952 ) ⋮ x ( t 1000 ) ] = [ x 951 1 ⋯ x 951 4 x 952 1 ⋯ x 952 4 ⋮ ⋱ ⋮ x 1000 1 ⋯ x 1000 4 ] . (31)</p><p>The second sequence of point cloud data set denoted by X n E T F examined the 10 ETF sectors ( d = 10 ), which yielded a 10 &#215; 50 matrix for each single point cloud for a total of q = n d l r − w + 1 = 951 point clouds as seen below from using Equation (1):</p><p>X 1 E T F = [ x ( t 1 ) x ( t 2 ) ⋮ x ( t 50 ) ] = [ x 1 1 ⋯ x 1 10 x 2 1 ⋯ x 2 10 ⋮ ⋱ ⋮ x 50 1 ⋯ x 50 10 ]                     ⋮ X 951 E T F = [ x ( t 951 ) x ( t 952 ) ⋮ x ( t 1000 ) ] = [ x 951 1 ⋯ x 951 10 x 952 1 ⋯ x 952 10 ⋮ ⋱ ⋮ x 1000 1 ⋯ x 1000 10 ] . (32)</p></sec><sec id="s3_2"><title>3.2. Vietoris-Rips Complex and Persistent Homology</title><p>Next, we constructed Vietoris-Rips complexes and filtration for each point cloud in X n S I and X n E T F from definition 2.4, definition 2.5, and Equation (6) and R-package “TDA” [<xref ref-type="bibr" rid="scirp.103953-ref25">25</xref>]. The Rips filtration for all the stock indices and all the ETFs are denoted by R ( X n S I , ϵ ) and R ( X n E T F , ϵ ) respectively for ϵ &gt; 0 . For the maximum filtration, we used ϵ ′ S I = 0.055 and ϵ ′ E T F = 0.08 , which are based on similar methods found in [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>]. Therefore, we obtained the following Rips filtration:</p><p>R ( X n S I , ϵ ) = R ( X n S I , 0 ) ⊂ ⋯ ⊂ R ( X n S I , 0.055 ) , (33)</p><p>R ( X n E T F , ϵ ) = R ( X n E T F , 0 ) ⊂ ⋯ ⊂ R ( X n E T F , 0.08 ) , (34)</p><p>where n = 1 , ⋯ , 951 . Based on the Equations (6), (33), and (34), we computed only the p = 1 dimensional homology H 1 ( R ( X n , ϵ ) ) with coefficients in the field ℤ / 2 ℤ from Equation (7) as follows:</p><p>H 1 ( R ( X n S I ,0 ) ) → H 1 ( R ( X n S I ,0.055 ) ) , (35)</p><p>H 1 ( R ( X n E T F ,0 ) ) → H 1 ( R ( X n E T F ,0.08 ) ) , (36)</p><p>where n = 1 , ⋯ , 951 . Also, we are only interested in the persistence of loops in as they appear in each point cloud during the transition states of the market, which is why we did the first dimensional homology. From definition 2.4, the filtration from Equations (33) and (34) induced a sequence of linear maps f 1 b i , d i , S I : H 1 ( R ( X n S I , 0 ) ) → H 1 ( R ( X n S I , 0.055 ) ) and f 1 b i , d i , E T F : H 1 ( R ( X n S I ,0 ) ) → H 1 ( R ( X n S I ,0.08 ) ) . The images of these maps are the persistent homology groups. The collection of vector spaces H 1 ( R ( X n S I ) ) and H 1 ( R ( X n E T F ) ) along with the corresponding linear maps is a persistent module, which leads us to the topological summaries.</p></sec><sec id="s3_3"><title>3.3. Topological Summaries</title><p>By modifying the R script in [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>], the first dimensional persistence diagrams denoted by D 1 ( X n S I ) = { b i , d i } i ∈ I and D 1 ( X n E T F ) = { b i , d i } i ∈ I for each point cloud data set were used along with Equations (10) and (11) to produce the analogous first dimensional persistent landscapes λ ( X n S I ) and λ ( X n E T F ) as seen below:</p><disp-formula id="scirp.103953-formula1"><label>(37)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x308.png"  xlink:type="simple"/></disp-formula><disp-formula id="scirp.103953-formula2"><label>(38)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x309.png"  xlink:type="simple"/></disp-formula><p>where<inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x310.png" xlink:type="simple"/></inline-formula>. Next, the norms of the persistence landscapes <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x311.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x312.png" xlink:type="simple"/></inline-formula> were computed for <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x313.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x314.png" xlink:type="simple"/></inline-formula> using Equations (14)-(16). The norms of the persistence landscapes and the daily log returns were plotted in juxtaposition, where it is important to remember that a point in the norms of persistence landscapes refers to a sliding window of 50 trading days in the daily log returns. After generating the topological summaries, the mean landscape is constructed using definition 2.7 and Equations (12)-(13) for the time period between July 1, 2019 and July 1, 2020 (253 trading days) or the time frame between July 2, 2019 to June 30, 2019 (252 days for the daily log returns). For this reason, the sequences of point cloud data sets will go from <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x315.png" xlink:type="simple"/></inline-formula> to<inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x316.png" xlink:type="simple"/></inline-formula>. Also, recall that the daily log returns are forward daily changes, so the time frame of the daily log returns are from July 2, 2019 to June 30, 2020. We assigned <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x317.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x318.png" xlink:type="simple"/></inline-formula> to be the corresponding landscapes for all the point clouds in <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x319.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x320.png" xlink:type="simple"/></inline-formula> to obtain the mean landscapes as seen below:</p><disp-formula id="scirp.103953-formula3"><label>(39)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x321.png"  xlink:type="simple"/></disp-formula><disp-formula id="scirp.103953-formula4"><label>(40)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x322.png"  xlink:type="simple"/></disp-formula><p>where <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x323.png" xlink:type="simple"/></inline-formula> for <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x324.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x325.png" xlink:type="simple"/></inline-formula> samples [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>]. We are interested in time period between July 1, 2019 and July 1, 2020, which has 253 trading days, because we wanted to observe market conditions prior to our market decline of interest and see if we are able to detect any critical transitions. Therefore, we provide summary statistics for this time period for all the stock indices and all the ETF sectors. The daily log returns, persistent diagrams, persistent landscapes, and the mean landscapes for sliding windows of 50 trading days were generated and plotted together for July 2, 2019 and June 30, 2020, but we only highlighted specific date ranges near our market fall of interest and peaks in the norms in the persistence landscape for all the stock indices and ETF sectors, which is discussed in Section 4.</p></sec><sec id="s3_4"><title>3.4. Statistical Inference</title><p>While the topological summaries were useful for examining topological features, we were also interested in finding statistical significant for any changes of these topological features within time. The time period of interest is July 1, 2019 to July 1, 2020, which has <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x326.png" xlink:type="simple"/></inline-formula> trading days.</p><p>We make the same assumptions from Section 2.5 and Section 2.6. Our random variables will derive from our two sequences of point cloud data sets, <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x327.png" xlink:type="simple"/></inline-formula>and<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x327.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x328.png" xlink:type="simple"/></inline-formula>. Since our time period of interest has 253 trading days, our sequence of point cloud data sets are size<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x327.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x328.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x329.png" xlink:type="simple"/></inline-formula>. For this reason, we have <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x327.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x328.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x329.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x330.png" xlink:type="simple"/></inline-formula> random variables in each sequence of point cloud data sets.</p><p>For all the stock indices and all the ETF sectors, we have <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula> be random variables respectively for <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x334.png" xlink:type="simple"/></inline-formula> samples for these groups respectively and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x334.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x335.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x334.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x335.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x336.png" xlink:type="simple"/></inline-formula> are the corresponding landscapes respectively for<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x334.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x335.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x336.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x337.png" xlink:type="simple"/></inline-formula>. The associate sample values of <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x334.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x335.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x336.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x337.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x338.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x334.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x335.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x336.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x337.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x338.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x339.png" xlink:type="simple"/></inline-formula> are denoted as <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x334.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x335.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x336.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x337.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x338.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x339.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x340.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x334.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x335.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x336.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x337.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x338.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x339.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x340.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x341.png" xlink:type="simple"/></inline-formula> respectively and the corresponding landscapes of these sample values are labelled as <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x334.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x335.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x336.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x337.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x338.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x339.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x340.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x341.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x342.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x331.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x332.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x333.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x334.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x335.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x336.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x337.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x338.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x339.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x340.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x341.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x342.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x343.png" xlink:type="simple"/></inline-formula> respectively.</p><p>The functional in Equation (25) is used to define the random variables for all the stock indices and all the ETFs as follows:</p><disp-formula id="scirp.103953-formula5"><label>(41)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x344.png"  xlink:type="simple"/></disp-formula><disp-formula id="scirp.103953-formula6"><label>(42)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x345.png"  xlink:type="simple"/></disp-formula><p>where <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x346.png" xlink:type="simple"/></inline-formula> for <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x346.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x347.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x346.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x347.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x348.png" xlink:type="simple"/></inline-formula> samples. We recall the sample mean<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x346.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x347.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x348.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x349.png" xlink:type="simple"/></inline-formula>, so using Equation (26) for the sample means for the random variables of all the stock indices and ETF sectors, we have the following:</p><disp-formula id="scirp.103953-formula7"><label>(43)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x350.png"  xlink:type="simple"/></disp-formula><disp-formula id="scirp.103953-formula8"><label>(44)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x351.png"  xlink:type="simple"/></disp-formula><p>where <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x352.png" xlink:type="simple"/></inline-formula> for <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x352.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x353.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x352.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x353.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x354.png" xlink:type="simple"/></inline-formula> samples. We assume that <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x352.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x353.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x354.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x355.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x352.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x353.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x354.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x355.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x356.png" xlink:type="simple"/></inline-formula> are the expectations and population means of <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x352.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x353.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x354.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x355.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x356.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x357.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x352.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x353.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x354.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x355.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x356.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x357.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x358.png" xlink:type="simple"/></inline-formula> respectively for<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x352.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x353.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x354.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x355.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x356.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x357.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x358.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x359.png" xlink:type="simple"/></inline-formula>. We set up three sets of hypotheses test and an analogous p-value based on a permutation test. For our first two sets of statistical hypotheses, we desire separate hypotheses tests for all the stock indices and for all the ETF sectors within a one day lag in their respective sliding windows. Our statistical hypotheses will distinguish for two groups at a time if the means of topological features are the same within a one day lag in their respective sliding windows and point cloud data sets as seen below:</p><disp-formula id="scirp.103953-formula9"><label>(45)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x360.png"  xlink:type="simple"/></disp-formula><disp-formula id="scirp.103953-formula10"><label>(46)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x361.png"  xlink:type="simple"/></disp-formula><p>For our third set of statistical hypotheses, we also wish to compare all the stock indices against all the ETF sectors within the same sliding windows. Our statistical hypotheses will determine for two groups at a time if the means of topological features are the same within the same sliding window as shown below:</p><disp-formula id="scirp.103953-formula11"><label>(47)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x362.png"  xlink:type="simple"/></disp-formula><p>where in Equations (45)-(47),<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x363.png" xlink:type="simple"/></inline-formula>. To test the null hypotheses found in Equations (45) and (46), we used a two-sample permutation test from Equation (28) to obtain:</p><disp-formula id="scirp.103953-formula12"><label>(48)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x364.png"  xlink:type="simple"/></disp-formula><disp-formula id="scirp.103953-formula13"><label>(49)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x365.png"  xlink:type="simple"/></disp-formula><p>where <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x366.png" xlink:type="simple"/></inline-formula> for <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x366.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x367.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x366.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x367.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x368.png" xlink:type="simple"/></inline-formula> samples. To test the null hypotheses found in Equation (47), we used a two-sample permutation test from Equations (28) to obtain:</p><disp-formula id="scirp.103953-formula14"><label>(50)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x369.png"  xlink:type="simple"/></disp-formula><p>where <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x370.png" xlink:type="simple"/></inline-formula> for <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x370.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x371.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x370.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x371.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x372.png" xlink:type="simple"/></inline-formula> samples. Using Equations (48), (49), and (50), <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x370.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x371.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x372.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x373.png" xlink:type="simple"/></inline-formula>of the test statistic were calculated for permutations<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x370.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x371.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x372.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x373.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x374.png" xlink:type="simple"/></inline-formula>. The observed value of the test statistic is expressed as<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x370.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x371.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x372.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x373.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x374.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x375.png" xlink:type="simple"/></inline-formula>. The p-value is calculated by comparing <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x370.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x371.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x372.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x373.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x374.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x375.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x376.png" xlink:type="simple"/></inline-formula> with <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x370.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x371.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x372.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x373.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x374.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x375.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x376.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x377.png" xlink:type="simple"/></inline-formula> and averaging the number of times<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x370.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x371.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x372.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x373.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x374.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x375.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x376.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x377.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x378.png" xlink:type="simple"/></inline-formula>. Using Equation (23), Equations (48) and (49) become:</p><disp-formula id="scirp.103953-formula15"><label>(51)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x379.png"  xlink:type="simple"/></disp-formula><disp-formula id="scirp.103953-formula16"><label>(52)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x380.png"  xlink:type="simple"/></disp-formula><p>Similarly, Equation (50) becomes:</p><disp-formula id="scirp.103953-formula17"><label>(53)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x381.png"  xlink:type="simple"/></disp-formula><p>where in Equations (51)-(53), where <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x382.png" xlink:type="simple"/></inline-formula> for <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x382.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x383.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x382.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x383.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x384.png" xlink:type="simple"/></inline-formula> samples. Hence, using Equations (48) and (52) and every instance where<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x382.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x383.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x384.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x385.png" xlink:type="simple"/></inline-formula>, the p-values were obtained as:</p><disp-formula id="scirp.103953-formula18"><label>(54)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x386.png"  xlink:type="simple"/></disp-formula><disp-formula id="scirp.103953-formula19"><label>(55)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x387.png"  xlink:type="simple"/></disp-formula><p>Similarly, using Equation (53) and every instance where<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x388.png" xlink:type="simple"/></inline-formula>, the p-value was obtained as:</p><disp-formula id="scirp.103953-formula20"><label>(56)</label><graphic position="anchor" xlink:href="//html.scirp.org/file/9-1490875x389.png"  xlink:type="simple"/></disp-formula><p>where in Equations (54)-(56),<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x390.png" xlink:type="simple"/></inline-formula>. To evaluate statistical significance, using Equations (41)-(56), a permutation is completed at a significance level of <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x390.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x391.png" xlink:type="simple"/></inline-formula> for homology in degree 1 for all our hypothesis tests. Since we are only interested in the number of loops, we will look at homology in degree 1. All these hypothesis testing methods were modified from the R script in [<xref ref-type="bibr" rid="scirp.103953-ref5">5</xref>]. After finding the p-values, we plotted the daily log returns with the p-values that were less than or greater than or equal to our significant level <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x390.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x391.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x392.png" xlink:type="simple"/></inline-formula> for either all the stock indices or all the ETF sectors along a sliding window of 50 trading days.</p></sec></sec><sec id="s4"><title>4. Results</title><p>The goal of this study is to detect a statistically relevant critical transition and characterize any changes in topological features over time. To assess the statistical significance of observed differences in the topological features that change over time, we used a permutation test. For degree 1, we obtained ten sample values of the random variables <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x393.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x393.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x394.png" xlink:type="simple"/></inline-formula> as in Equation (41). Using Equations (43), (45), (48), (51), and (54), the permutation test is implemented with a significance level <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x393.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x394.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x395.png" xlink:type="simple"/></inline-formula> when comparing all stock indices in different sliding windows between July 1, 2019 and July 1, 2020. The permutation test yields 164 p-values of 0.0000, 2 p-values of 0.001, and 36 p-values of 1 for homology in degree 1.</p><p>Using Equations (44), (46), (49), (52), and (55), the permutation test is conducted with a significance level <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x396.png" xlink:type="simple"/></inline-formula> when comparing all the ETF sectors in different sliding windows between July 1, 2019 and July 1, 2020. The permutation test returns 164 p-values of 0.0000, 4 p-values of 0.001, and 33 p-values of 1 for homology in degree 1. Using Equations (43), (44), (47), (50), (53), and (56), the permutation is performed with a significance level <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x396.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x397.png" xlink:type="simple"/></inline-formula> when comparing all the stock indices and all the ETF sectors in the same sliding windows between July 1, 2019 and July 1, 2020, which results in 199 p-values of 0.0000 and 2 p-values of 0.001 for homology in degree 1.</p><p>In order to understand these results, we will review the daily log returns, the norms of the persistence landscapes, and the topological summaries of all the stock indices and all the ETF sectors. When reviewing the daily log returns for DJIA, the S&amp;P 500, NASDAQ, and Russell 2000 between January 5, 2010 and June 30, 2020 (see <xref ref-type="fig" rid="fig1">Figure 1</xref>), the stock indices range from −0.05 and 0.05 from 2010 to mid 2011, with some positive and negative spikes that appear leading up to 2012. From 2012 to March 2020, the daily log returns once again fall between −0.05 and 0.05. However, from March 2020 to June 2020, the market is highly volatile. Similar patterns are observed for the ETF sectors, but there is a notable spike around 2017 and from March 2020 to June 2020, the ETF sectors are more volatile than the stock indices as shown in <xref ref-type="fig" rid="fig2">Figure 2</xref>.</p><p>When we examine the daily log returns of all of the stock indices between January 5, 2010-June 1, 2020, the minimum daily log return occur on March 16, 2020, where Russell 2000 had a return of −0.154, the S&amp;P 500 had a return at −0.1277, and the other stock indices were in between these values. When reviewing the daily log returns for all of the ETFs sectors for the same time period, the minimum daily log return also occurs on March 16, 2020, where Information Technology (XLK) had a return of −0.1487, Consumer Staples (XLP) had a return of −0.0702, and the other ETF sectors were in between these values. While March 16, 2020 is not recognized as an official financial crash or meltdown, this</p><p>date is noteworthy, and warrants closer examination for potential critical transitions prior to this date. Focusing on when the peaks occur, we include summary statistics for July 1, 2019 to July 1, 2020 for all the stock indices and all the ETF sectors in <xref ref-type="table" rid="table1">Table 1</xref> and <xref ref-type="table" rid="table2">Table 2</xref>, respectively.</p><p>The norms of the persistence landscapes in homology degree 1 presented in <xref ref-type="fig" rid="fig3">Figure 3</xref> and <xref ref-type="fig" rid="fig4">Figure 4</xref> display all of the stock indices and all of the ETFs respectively for <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x400.png" xlink:type="simple"/></inline-formula> and <inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x400.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x401.png" xlink:type="simple"/></inline-formula> for 1001 trading days prior to March 16, 2020. For the stock indices, the L<sup>1</sup> distances are less than 0.01 between 2017 and 2018, less than 0.02 between 2018 and 2020, but the greatest L<sup>1</sup> distance occurs in 2020 at approximately 0.08 as seen in <xref ref-type="fig" rid="fig3">Figure 3</xref>. The L<sup>1</sup> distance for all of the ETFs, have more spikes than the L<sup>1</sup> distances of the stock indices, especially between 2018 and 2020, but the greatest L<sup>1</sup> distance occurs in March 2020 at approximately 0.14 as seen in <xref ref-type="fig" rid="fig3">Figure 3</xref>. While the L<sup>2</sup> norms for all of the stock indices and all of the ETFs have similar distances, there is a noticeable spike in 2020. However, the distances in L<sup>2</sup> are not as great as in L<sup>1</sup>, as shown in <xref ref-type="fig" rid="fig3">Figure 3</xref> and <xref ref-type="fig" rid="fig4">Figure 4</xref>.</p><p><xref ref-type="fig" rid="fig3">Figure 3</xref> and <xref ref-type="fig" rid="fig4">Figure 4</xref> highlight the peak in more detail for the time period between January 3, 2020 to June 30, 2020. While critical points are discernible in the month of February 2020, the peaks occurred on February 21, 2020 and March 3, 2020 for all of the stock indices and for all of the ETF sectors respectively as seen in <xref ref-type="fig" rid="fig4">Figure 4</xref>. Recall that a point on the norms of the persistence landscapes coincides with a sliding window of 50 trading days in the daily log returns, which means the peaks are from February 21, 2020 to May 1, 2020 and March 3, 2020 to May 12, 2020 for all of the stock indices and for all of the ETF</p><table-wrap id="table1" ><label><xref ref-type="table" rid="table1">Table 1</xref></label><caption><title> Summary statistics for stock indices</title></caption><table><tbody><thead><tr><th align="center" valign="middle" >Stock Name</th><th align="center" valign="middle" ><inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x402.png" xlink:type="simple"/></inline-formula></th><th align="center" valign="middle" ><inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x403.png" xlink:type="simple"/></inline-formula></th><th align="center" valign="middle" ><inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x404.png" xlink:type="simple"/></inline-formula></th><th align="center" valign="middle" ><inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x405.png" xlink:type="simple"/></inline-formula></th><th align="center" valign="middle" ><inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x406.png" xlink:type="simple"/></inline-formula></th></tr></thead><tr><td align="center" valign="middle" >Dow Jones</td><td align="center" valign="middle" >−1e−04</td><td align="center" valign="middle" >5e−04</td><td align="center" valign="middle" >0.0228</td><td align="center" valign="middle" >−0.8479</td><td align="center" valign="middle" >13.1333</td></tr><tr><td align="center" valign="middle" >S&amp;P 500</td><td align="center" valign="middle" >2e−04</td><td align="center" valign="middle" >5e−04</td><td align="center" valign="middle" >0.0213</td><td align="center" valign="middle" >−0.8691</td><td align="center" valign="middle" >12.6087</td></tr><tr><td align="center" valign="middle" >NASDAQ</td><td align="center" valign="middle" >9e−04</td><td align="center" valign="middle" >5e−04</td><td align="center" valign="middle" >0.0213</td><td align="center" valign="middle" >−1.0494</td><td align="center" valign="middle" >12.6029</td></tr><tr><td align="center" valign="middle" >Russell 2000</td><td align="center" valign="middle" >−3e−04</td><td align="center" valign="middle" >7e−04</td><td align="center" valign="middle" >0.0263</td><td align="center" valign="middle" >−1.3226</td><td align="center" valign="middle" >11.2358</td></tr></tbody></table></table-wrap><table-wrap id="table2" ><label><xref ref-type="table" rid="table2">Table 2</xref></label><caption><title> Summary statistics for ETF sectors</title></caption><table><tbody><thead><tr><th align="center" valign="middle" >Stock Symbol</th><th align="center" valign="middle" ><inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x407.png" xlink:type="simple"/></inline-formula></th><th align="center" valign="middle" ><inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x408.png" xlink:type="simple"/></inline-formula></th><th align="center" valign="middle" ><inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x409.png" xlink:type="simple"/></inline-formula></th><th align="center" valign="middle" ><inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x410.png" xlink:type="simple"/></inline-formula></th><th align="center" valign="middle" ><inline-formula><inline-graphic xlink:href="/html.scirp.org/file/9-1490875x411.png" xlink:type="simple"/></inline-formula></th></tr></thead><tr><td align="center" valign="middle" >XLY</td><td align="center" valign="middle" >0.0003</td><td align="center" valign="middle" >0.0004</td><td align="center" valign="middle" >0.0210</td><td align="center" valign="middle" >−1.3141</td><td align="center" valign="middle" >14.1179</td></tr><tr><td align="center" valign="middle" >XLP</td><td align="center" valign="middle" >0.0001</td><td align="center" valign="middle" >0.0003</td><td align="center" valign="middle" >0.0173</td><td align="center" valign="middle" >−0.2487</td><td align="center" valign="middle" >12.6118</td></tr><tr><td align="center" valign="middle" >XLE</td><td align="center" valign="middle" >−0.0017</td><td align="center" valign="middle" >0.0012</td><td align="center" valign="middle" >0.0348</td><td align="center" valign="middle" >−1.3392</td><td align="center" valign="middle" >13.4664</td></tr><tr><td align="center" valign="middle" >XLF</td><td align="center" valign="middle" >−0.0006</td><td align="center" valign="middle" >0.0008</td><td align="center" valign="middle" >0.0276</td><td align="center" valign="middle" >−0.6256</td><td align="center" valign="middle" >10.4510</td></tr><tr><td align="center" valign="middle" >XLV</td><td align="center" valign="middle" >0.0004</td><td align="center" valign="middle" >0.0004</td><td align="center" valign="middle" >0.0187</td><td align="center" valign="middle" >−0.4299</td><td align="center" valign="middle" >10.1362</td></tr><tr><td align="center" valign="middle" >XLI</td><td align="center" valign="middle" >−0.0004</td><td align="center" valign="middle" >0.0006</td><td align="center" valign="middle" >0.0243</td><td align="center" valign="middle" >−0.5575</td><td align="center" valign="middle" >10.0397</td></tr><tr><td align="center" valign="middle" >XLB</td><td align="center" valign="middle" >−0.0001</td><td align="center" valign="middle" >0.0005</td><td align="center" valign="middle" >0.0232</td><td align="center" valign="middle" >−0.7113</td><td align="center" valign="middle" >10.0980</td></tr><tr><td align="center" valign="middle" >XLK</td><td align="center" valign="middle" >0.0012</td><td align="center" valign="middle" >0.0006</td><td align="center" valign="middle" >0.0242</td><td align="center" valign="middle" >−0.6869</td><td align="center" valign="middle" >12.5942</td></tr><tr><td align="center" valign="middle" >XLU</td><td align="center" valign="middle" >−0.0001</td><td align="center" valign="middle" >0.0006</td><td align="center" valign="middle" >0.0235</td><td align="center" valign="middle" >−0.0751</td><td align="center" valign="middle" >11.9571</td></tr><tr><td align="center" valign="middle" >SPY</td><td align="center" valign="middle" >0.0002</td><td align="center" valign="middle" >0.0004</td><td align="center" valign="middle" >0.0207</td><td align="center" valign="middle" >−0.8911</td><td align="center" valign="middle" >11.7189</td></tr></tbody></table></table-wrap><p>sectors respectively. In particular, <xref ref-type="fig" rid="fig5">Figure 5</xref> and <xref ref-type="fig" rid="fig6">Figure 6</xref> emphasize this point, where the norms of the persistence landscapes and the daily log returns of either all of the stock indices or all of the ETF sectors are next to each other. The sliding windows of 50 trading days of the daily log returns synchronize to the first point in the norms of the persistence landscape and to the maximum values of the norms of the persistence landscapes as indicated by <xref ref-type="fig" rid="fig5">Figure 5</xref> and <xref ref-type="fig" rid="fig6">Figure 6</xref>.</p><p>Aside from the norms of the persistence landscapes, we produce topological summaries to represent the persistence of topological features for all the stock indices and for all the ETF sectors between January 3, 2020 and June 30, 2020. Along with these topological summaries (the persistence diagram, the persistence landscape, the mean landscape), we plotted the daily log returns for the corresponding sliding window of 50 trading days shown in Figures 7-12.</p><p><xref ref-type="fig" rid="fig7">Figure 7</xref> and <xref ref-type="fig" rid="fig8">Figure 8</xref> indicate that the daily log returns are centered around zero from January 3, 2020 to February 21, 2020 for all of the stock indices and from January 3, 2020 to February 26, 2020 for all of the ETF sectors. Not much persistence is evident in the persistence diagram and few spikes appear in persistence landscape and mean landscape. <xref ref-type="fig" rid="fig9">Figure 9</xref> and <xref ref-type="fig" rid="fig1">Figure 1</xref>0 illustrate more variability in the daily log returns from March 1, 2020 to April 16, 2020 for all of</p><p>the stock indices and from March 1, 2020 to May 1, 2020 for all of the ETF sectors. Significant persistence is apparent in the persistence diagram and more spikes appear in the persistence landscape and mean landscape.</p></sec><sec id="s5"><title>5. Discussion</title><p>From reviewing the norms of the persistence landscape, the daily log returns, persistence diagrams, persistence landscapes, and mean landscapes for all of the selected dates, it is clear that the number of the loops in the relevant point clouds are more pronounced resulting in more persistence, which signifies that the stock market is transitioning from a stable state to a more unpredictable, volatile state. Moreover, the ETF sectors demonstrate more volatility than the stock indices. These stock indices’ findings coincide with the 2000 and 2008 market crashes findings found in [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>]. Similar to Gidea and Katz [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>], we observe L<sup>1</sup> distances that confirm the critical thresholds prior to the 2020 peak and exhibit more than the L<sup>2</sup> norm. In other words, the L<sup>p</sup>-norms exhibit strong growth around the emergence of the primary peak.</p><p>While the highest peak occurred on February 21, 2020 for all of the stock indices and March 3, 2020 for all of the ETF sectors in the L<sup>p</sup> norms, the Coronavirus (COVID-19) broke out in 2019 in Wuhan, China, but on January 21, 2020, the first US case was confirmed. The most important dates are March 13, 2020 when President Trump declares national emergency, March 15, 2020 when the Center of Disease Control and Prevention warns against large gatherings, and March 17, 2020 when COVID is present in all 50 states. The daily log returns for all of the stock indices and for all of the ETF sectors do not include negative values. Yet, there are other dates that could have lead to a market decline in March 16, 2020. For example, on January 30, 2020 when World Health Organization (WHO) declares a global health emergency or between February 5, 2020 and February 29, 2020 when the outbreak becomes an epidemic. While we acknowledge that it is quite difficult to predict a market crash, the norms of the persistence landscape performed really well as indicator in detecting critical transitions and the topological summaries authenticated volatility by of the number of loops increasing.</p><p>Our hypotheses tests aimed to find how topological features change within time, notably between July 1, 2019 and July 1, 2020. Our hypotheses tests for all of the stock indices found evidence of difference in topological features when comparing adjacent sliding windows of a sliding step of one day. In particular, we found for the chosen time frame that the daily log returns of all the stock indices significantly differ in the number of loops. Equivalently, our hypotheses tests for all of the ETF sectors found evidence of difference in topological features when comparing adjacent sliding windows of a sliding step of one day. Specifically, we found for the selected time frame that the daily log returns of all the ETF sectors significantly differ in the number of loops. Our last hypotheses tests between all of the stock indices and all of the ETF sectors within the same sliding window found inconclusive evidence of difference in topological features for the entire time frame.</p></sec><sec id="s6"><title>6. Conclusions</title><p>In this paper, we investigated the topological features of four major indices and 10 ETF sectors for January 4, 2010-July 1, 2020. We used two sequences of point cloud data sets, one for all the stock indices and the other for all the ETFs with a sliding window<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x430.png" xlink:type="simple"/></inline-formula>. Both sequences were used to perform TDA through algebraic topology and persistent homology. From there, topological summaries are generated to determine persistence and the norms for persistence landscapes are used to detect a critical transition by adapting methods found in [<xref ref-type="bibr" rid="scirp.103953-ref16">16</xref>]. Our goal is to determine how the statistical significance of topological features of stock indices and ETF sectors change for a specific time frame. We found that between July 1, 2019 and July 1, 2020, there is evidence of difference of topological features for all the stock indices and all the ETFs. As a result, critical transitions are determined using the norms of the persistence landscape and topological features of stock indices and ETF sectors change within time when comparing two sliding windows of a sliding step of one day.</p><p>We conclude with possible future research goals. Further work could be done analyzing persistence landscapes for homology in degree two. It would be interesting to study topological features based on higher degree persistence. Furthermore, it would be fascinating to expand to commodities, futures, and other financial time series. Moreover, it would be more resourceful to expand topological data analysis to statistics beyond statistical inference and use for predictive modeling with machine learning.</p><p>This table presents summary statistics for all the stock indices. We estimated the mean (<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x431.png" xlink:type="simple"/></inline-formula>), standard deviation (<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x431.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x432.png" xlink:type="simple"/></inline-formula>), variance (<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x431.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x432.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x433.png" xlink:type="simple"/></inline-formula>), skewness (<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x431.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x432.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x433.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x434.png" xlink:type="simple"/></inline-formula>), and kurtosis (<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x431.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x432.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x433.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x434.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x435.png" xlink:type="simple"/></inline-formula>) of the daily log returns from July 2, 2019 to June 30, 2020. The reporting period of this table contains 253 trading days from July 1, 2019 to July 1, 2020.</p><p>This table presents summary statistics for all the ETF sectors. We estimated the mean (<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x436.png" xlink:type="simple"/></inline-formula>), standard deviation (<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x436.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x437.png" xlink:type="simple"/></inline-formula>), variance (<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x436.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x437.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x438.png" xlink:type="simple"/></inline-formula>), skewness (<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x436.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x437.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x438.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x439.png" xlink:type="simple"/></inline-formula>), and kurtosis (<inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x436.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x437.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x438.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x439.png" xlink:type="simple"/></inline-formula><inline-formula><inline-graphic xlink:href="//html.scirp.org/file/9-1490875x440.png" xlink:type="simple"/></inline-formula>) of the daily log returns from July 2, 2019 to June 30, 2020. The reporting period of this table contains 253 trading days from July 1, 2019 to July 1, 2020.</p></sec><sec id="s7"><title>Acknowledgements</title><p>The authors would like to thank Tracy Volz for helpful discussions in the editing process. The authors would also like to thank the Center of Computational Finance and Economic Systems (https://cofes.rice.edu).</p></sec><sec id="s8"><title>Conflicts of Interest</title><p>The authors declare no conflicts of interest regarding the publication of this paper.</p></sec><sec id="s9"><title>Cite this paper</title><p>Aguilar, A. and Ensor, K. (2020) Topology Data Analysis Using Mean Persistence Landscapes in Financial Crashes. Journal of Mathematical Finance, 10, 648-678. https://doi.org/10.4236/jmf.2020.104038</p></sec></body><back><ref-list><title>References</title><ref id="scirp.103953-ref1"><label>1</label><mixed-citation publication-type="other" xlink:type="simple">Collins, A., Zomorodian, A., Carlsson, G. and Guibas, L.J. (2004) A Barcode Shape Descriptor for Curve Point Cloud Data. Computers &amp; Graphics, 28, 881-894. https://doi.org/10.1016/j.cag.2004.08.015</mixed-citation></ref><ref id="scirp.103953-ref2"><label>2</label><mixed-citation publication-type="other" xlink:type="simple">Mileyko, Y., Mukherjee, S. and Harer, J. (2011) Probability Measures on the Space of Persistence Diagrams. Inverse Problems, 27, Article ID: 124007. https://doi.org/10.1088/0266-5611/27/12/124007</mixed-citation></ref><ref id="scirp.103953-ref3"><label>3</label><mixed-citation publication-type="other" xlink:type="simple">Turner, K., Mileyko, Y., Mukherjee, S. and Harer, J. (2014) Fréchet Means for Distributions of Persistence Diagrams. Discrete &amp; Computational Geometry, 52, 44-70. https://doi.org/10.1007/s00454-014-9604-7</mixed-citation></ref><ref id="scirp.103953-ref4"><label>4</label><mixed-citation publication-type="other" xlink:type="simple">Munch, E., Bendich, P., Turner, K., Mukherjee, S., Mattingly, J. and Harer, J. (2013) Probabilistic Fréchet Means and Statistics on Vineyards.</mixed-citation></ref><ref id="scirp.103953-ref5"><label>5</label><mixed-citation publication-type="other" xlink:type="simple">Nicolau, M., Levine, A.J. and Carlsson, G. (2011) Topology Based Data Analysis Identifies a Subgroup of Breast Cancers with a Unique Mutational Profile and Excellent Survival. Proceedings of the National Academy of Sciences, 108, 7265-7270. https://doi.org/10.1073/pnas.1102826108</mixed-citation></ref><ref id="scirp.103953-ref6"><label>6</label><mixed-citation publication-type="other" xlink:type="simple">Carlsson, G., Ishkhanov, T., De Silva, V. and Zomorodian, A. (2008) On the Local Behavior of Spaces of Natural Images. International Journal of Computer Vision, 76, 1-12. https://doi.org/10.1007/s11263-007-0056-x</mixed-citation></ref><ref id="scirp.103953-ref7"><label>7</label><mixed-citation publication-type="other" xlink:type="simple">Wang, Y., Ombao, H. and Chung, M.K. (2019) Statistical Persistent Homology of Brain Signals. IEEE International Conference on Acoustics, Speech and Signal Processing, Brighton, 12-17 May 2019, 1125-1129. https://doi.org/10.1109/ICASSP.2019.8682978</mixed-citation></ref><ref id="scirp.103953-ref8"><label>8</label><mixed-citation publication-type="other" xlink:type="simple">Nielson, J.L., Paquette, J., Liu, A.W., Guandique, C.F., Tovar, C.A., Inoue, T., Irvine, K.-A., Gensel, J.C., Kloke, J., Petrossian, T.C., et al. (2015) Topological Data Analysis for Discovery in Preclinical Spinal Cord Injury and Traumatic Brain Injury. Nature Communications, 6, 8581. https://doi.org/10.1038/ncomms9581</mixed-citation></ref><ref id="scirp.103953-ref9"><label>9</label><mixed-citation publication-type="other" xlink:type="simple">Scheffer, M., Bascompte, J., Brock, W.A., Brovkin, V., Carpenter, S.R., Dakos, V., Held, H., Van Nes, E.H., Rietkerk, M. and Sugihara, G. (2009) Early-Warning Signals for Critical Transitions. Nature, 461, 53-59. https://doi.org/10.1038/nature08227</mixed-citation></ref><ref id="scirp.103953-ref10"><label>10</label><mixed-citation publication-type="other" xlink:type="simple">Ensor, K.B. and Koev, G.M. (2014) Computational Finance: Correlation, Volatility, and Markets. Wiley Interdisciplinary Reviews: Computational Statistics, 6, 326-340. https://doi.org/10.1002/wics.1323</mixed-citation></ref><ref id="scirp.103953-ref11"><label>11</label><mixed-citation publication-type="other" xlink:type="simple">Gidea, M. (2017) Topological Data Analysis of Critical Transitions in Financial Networks. In: International Conference and School on Network Science, Springer, Berlin, 47-59. https://doi.org/10.1007/978-3-319-55471-6_5</mixed-citation></ref><ref id="scirp.103953-ref12"><label>12</label><mixed-citation publication-type="other" xlink:type="simple">Gidea, M., Goldsmith, D., Katz, Y., Roldan, P., Shmalo, Y., et al. (2020) Topological Recognition of Critical Transitions in Time Series of Cryptocurrencies. Physica A: Statistical Mechanics and Its Applications, 548, Article ID: 123843. https://doi.org/10.1016/j.physa.2019.123843</mixed-citation></ref><ref id="scirp.103953-ref13"><label>13</label><mixed-citation publication-type="other" xlink:type="simple">Guttal, V., Raghavendra, S., Goel, N. and Hoarau, Q. (2016) Lack of Critical Slowing Down Suggests That Financial Meltdowns Are Not Critical Transitions, Yet Rising Variability Could Signal Systemic Risk. PLoS ONE, 11, e0144198. https://doi.org/10.1371/journal.pone.0144198</mixed-citation></ref><ref id="scirp.103953-ref14"><label>14</label><mixed-citation publication-type="other" xlink:type="simple">Edelsbrunner, H., Letscher, D. and Zomorodian, A. (2002) Topological Persistence and Simplification. Discrete &amp; Computational Geometry, 28, 511-533. https://doi.org/10.1007/s00454-002-2885-2</mixed-citation></ref><ref id="scirp.103953-ref15"><label>15</label><mixed-citation publication-type="other" xlink:type="simple">Hatcher, A. (2002) Algebraic Topology. Cambridge University Press, Cambridge.</mixed-citation></ref><ref id="scirp.103953-ref16"><label>16</label><mixed-citation publication-type="other" xlink:type="simple">Munkres, J.R. (1984) Elements of Algebraic Topology. The Benjamin/Cummings Publishing Company, Meno Park, CA.</mixed-citation></ref><ref id="scirp.103953-ref17"><label>17</label><mixed-citation publication-type="other" xlink:type="simple">Dundas, B.I. (2013) Differential Topology.</mixed-citation></ref><ref id="scirp.103953-ref18"><label>18</label><mixed-citation publication-type="other" xlink:type="simple">Bubenik, P. and D&amp;#322;otko, P. (2017) A Persistence Landscapes Toolbox for Topological Statistics. Journal of Symbolic Computation, 78, 91-114. https://doi.org/10.1016/j.jsc.2016.03.009</mixed-citation></ref><ref id="scirp.103953-ref19"><label>19</label><mixed-citation publication-type="other" xlink:type="simple">Otter, N., Porter, M.A., Tillmann, U., Grindrod, P. and Harrington, H.A. (2017) A Roadmap for the Computation of Persistent Homology. EPJ Data Science, 6, 17. https://doi.org/10.1140/epjds/s13688-017-0109-5</mixed-citation></ref><ref id="scirp.103953-ref20"><label>20</label><mixed-citation publication-type="other" xlink:type="simple">Kovacev-Nikolic, V., Bubenik, P., Nikoli&amp;#263;, D. and Heo, G. (2016) Using Persistent Homology and Dynamical Distances to Analyze Protein Binding. Statistical Applications in Genetics and Molecular Biology, 15, 19-38. https://doi.org/10.1515/sagmb-2015-0057</mixed-citation></ref><ref id="scirp.103953-ref21"><label>21</label><mixed-citation publication-type="other" xlink:type="simple">RStudio Team (2020) RStudio: Integrated Development Environment for R. RStudio, PBC, Boston.</mixed-citation></ref><ref id="scirp.103953-ref22"><label>22</label><mixed-citation publication-type="other" xlink:type="simple">Shumway, R.H. and Stoffer, D.S. (2016) Time Series Analysis and Its Applications. Springer, Berlin. https://doi.org/10.1007/978-3-319-52452-8</mixed-citation></ref><ref id="scirp.103953-ref23"><label>23</label><mixed-citation publication-type="other" xlink:type="simple">Gidea, M. and Katz, Y. (2018) Topological Data Analysis of Financial Time Series: Landscapes of Crashes. Physica A: Statistical Mechanics and Its Applications, 491, 820-834. https://doi.org/10.1016/j.physa.2017.09.028</mixed-citation></ref><ref id="scirp.103953-ref24"><label>24</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Bubenik</surname><given-names> P. </given-names></name>,<etal>et al</etal>. (<year>2015</year>)<article-title>Statistical Topological Data Analysis Using Persistence Landscapes</article-title><source> The Journal of Machine Learning Research</source><volume> 16</volume>,<fpage> 77</fpage>-<lpage>102</lpage>.<pub-id pub-id-type="doi"></pub-id></mixed-citation></ref><ref id="scirp.103953-ref25"><label>25</label><mixed-citation publication-type="other" xlink:type="simple">Fasy, B.T., Kim, J., Lecci, F., Maria, C. and Rouvreau, V. (2019) TDA: Statistical Tools for Topological Data Analysis. RStudio, PBC, Boston. https://cran.r-project.org/web/packages/TDA/index.html</mixed-citation></ref></ref-list></back></article>