<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.4 20241031//EN" "JATS-journalpublishing1-4.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article" dtd-version="1.4" xml:lang="en">
  <front>
    <journal-meta>
      <journal-id journal-id-type="publisher-id">ojsst</journal-id>
      <journal-title-group>
        <journal-title>Open Journal of Safety Science and Technology</journal-title>
      </journal-title-group>
      <issn pub-type="epub">2162-6006</issn>
      <issn pub-type="ppub">2162-5999</issn>
      <publisher>
        <publisher-name>Scientific Research Publishing</publisher-name>
      </publisher>
    </journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.4236/ojsst.2026.163010</article-id>
      <article-id pub-id-type="publisher-id">ojsst-153277</article-id>
      <article-categories>
        <subj-group>
          <subject>Article</subject>
        </subj-group>
        <subj-group>
          <subject>Chemistry</subject>
          <subject>Materials Science</subject>
          <subject>Earth</subject>
          <subject>Environmental Sciences</subject>
          <subject>Engineering</subject>
          <subject>Physics</subject>
          <subject>Mathematics</subject>
          <subject>Social Sciences</subject>
          <subject>Humanities</subject>
        </subj-group>
      </article-categories>
      <title-group>
        <article-title>Predicting Traffic Anomalies and Collision Risks in V2X Systems: A Deep Learning Approach Using LSTM and GNN</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <contrib-id contrib-id-type="orcid">0000-0001-6490-9983</contrib-id>
          <name name-style="western">
            <surname>Quito</surname>
            <given-names>Benjamin</given-names>
          </name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
      </contrib-group>
      <aff id="aff1"><label>1</label> School of Applied Computer Science &amp; Information Technology, Conestoga College, Kitchener, Canada </aff>
      <author-notes>
        <fn fn-type="conflict" id="fn-conflict">
          <p>The author declares no conflicts of interest regarding the publication of this paper.</p>
        </fn>
      </author-notes>
      <pub-date pub-type="epub">
        <day>07</day>
        <month>09</month>
        <year>2026</year>
      </pub-date>
      <pub-date pub-type="collection">
        <month>09</month>
        <year>2026</year>
      </pub-date>
      <volume>16</volume>
      <issue>03</issue>
      <fpage>160</fpage>
      <lpage>177</lpage>
      <history>
        <date date-type="received">
          <day>04</day>
          <month>07</month>
          <year>2026</year>
        </date>
        <date date-type="accepted">
          <day>16</day>
          <month>08</month>
          <year>2026</year>
        </date>
        <date date-type="published">
          <day>19</day>
          <month>08</month>
          <year>2026</year>
        </date>
      </history>
      <permissions>
        <copyright-statement>© 2026 by the authors and Scientific Research Publishing Inc.</copyright-statement>
        <copyright-year>2026</copyright-year>
        <license license-type="open-access">
          <license-p> This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license ( <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link> ). </license-p>
        </license>
      </permissions>
      <self-uri content-type="doi" xlink:href="https://doi.org/10.4236/ojsst.2026.163010">https://doi.org/10.4236/ojsst.2026.163010</self-uri>
      <abstract>
        <p>Vehicle-to-Everything (V2X) communication has transformed intelligent transportation systems (ITS) by enabling real-time data exchange for enhanced road safety and efficiency. However, current V2X-based safety mechanisms remain largely reactive, limiting their ability to prevent accidents in dynamic traffic environments. This paper proposes a hybrid deep learning framework that integrates Long Short-Term Memory (LSTM) networks and Graph Neural Networks (GNNs) to predict traffic anomalies and collision risks in V2X systems. The LSTM component captures temporal dependencies in vehicle behavior, while the GNN module models spatial interactions among vehicles within road networks. The fusion of these models enhances predictive accuracy, enabling proactive risk assessment and early warning generation. Experimental validation using real-world trajectory data shows that the fused model provides a balanced combination of temporal and spatial evidence; its errors are comparable to the standalone models, with the best metric depending on the selected fusion weight. The results indicate that spatiotemporal learning is a promising basis for V2X safety applications, paving the way for real-time deployment in autonomous and connected vehicle environments.</p>
      </abstract>
      <kwd-group kwd-group-type="author-generated" xml:lang="en">
        <kwd>Vehicle-to-Everything (V2X)</kwd>
        <kwd>Intelligent Transportation Systems (ITS)</kwd>
        <kwd>Traffic Anomaly Detection</kwd>
        <kwd>Collision Risk Reduction</kwd>
        <kwd>Deep Learning</kwd>
        <kwd>Graph Neural Network (GNNs)</kwd>
        <kwd>Long Short-Term Memory (LSTM)</kwd>
        <kwd>Spatiotemporal Modelling</kwd>
        <kwd>Autonomous Vehicles (AVs)</kwd>
        <kwd>Machine Learning for Traffic Safety</kwd>
        <kwd>Connected Vehicles</kwd>
        <kwd>Predictive Safety Systems</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec1">
      <title>1. Introduction</title>
      <p>The rapid advancement of autonomous vehicles (AVs) and intelligent transportation systems (ITS) has revolutionized modern mobility, with Vehicle-to-Everything (V2X) communication playing a pivotal role in enhancing road safety and traffic efficiency [<xref ref-type="bibr" rid="B1">1</xref>][<xref ref-type="bibr" rid="B2">2</xref>]. V2X enables real-time data exchange between vehicles, infrastructure, pedestrians, and networks, supporting critical applications such as collision avoidance, emergency braking alerts, and adaptive traffic management [<xref ref-type="bibr" rid="B3">3</xref>]. Despite these advancements, most existing V2X-based safety systems remain reactive, issuing alerts only after detecting potential hazards [<xref ref-type="bibr" rid="B4">4</xref>]. This approach significantly limits response time, particularly in high-density and dynamic traffic environments, increasing the likelihood of accidents.</p>
      <p>To overcome this challenge, proactive safety mechanisms that predict potential risks before they occur are essential [<xref ref-type="bibr" rid="B3">3</xref>]. Recent developments in Artificial Intelligence (AI), particularly Machine Learning (ML) and Deep Learning (DL), have demonstrated remarkable success in forecasting traffic anomalies and collision risks by leveraging historical and real-time vehicular data [<xref ref-type="bibr" rid="B1">1</xref>][<xref ref-type="bibr" rid="B4">4</xref>]. Specifically, Long Short-Term Memory (LSTM) networks excel in capturing temporal dependencies in traffic patterns [<xref ref-type="bibr" rid="B3">3</xref>], while Graph Neural Networks (GNNs) model spatial interactions between vehicles and road networks [<xref ref-type="bibr" rid="B2">2</xref>][<xref ref-type="bibr" rid="B5">5</xref>].</p>
      <p>However, existing research often treats temporal and spatial modeling separately, leading to fragmented solutions that fail to fully capture the spatiotemporal complexity of traffic dynamics [<xref ref-type="bibr" rid="B4">4</xref>]. To address this gap, this paper proposes a hybrid deep learning framework that integrates LSTM and GNN models to predict traffic anomalies and collision risks in V2X systems. By fusing temporal and spatial insights, the framework provides early warnings, enabling preventive safety interventions and reducing the risk of traffic incidents.</p>
    </sec>
    <sec id="sec2">
      <title>2. Background and Related Work</title>
      <sec id="sec2dot1">
        <title>2.1. Vehicle-to-Everything (V2X) Communication and Safety Applications</title>
        <p>Vehicle-to-Everything (V2X) communication is a cornerstone of modern intelligent transportation systems (ITS), facilitating real-time information exchange between vehicles, infrastructure, and networks [<xref ref-type="bibr" rid="B1">1</xref>][<xref ref-type="bibr" rid="B2">2</xref>]. The key V2X communication types include:</p>
        <p>Vehicle-to-Vehicle (V2V): Data exchange between vehicles to improve safety and coordination.Vehicle-to-Infrastructure (V2I): Communication with traffic lights, road signs, and control centers.Vehicle-to-Pedestrian (V2P): Interaction with pedestrians and vulnerable road users.Vehicle-to-Network (V2N): Connectivity with cloud services, traffic updates, and weather reports.</p>
        <p>V2X-enabled safety applications include collision warnings, emergency braking alerts, and cooperative adaptive cruise control [<xref ref-type="bibr" rid="B2">2</xref>][<xref ref-type="bibr" rid="B5">5</xref>]. However, most existing V2X-based safety systems are reactive, alerting drivers only after hazardous situations arise, thereby limiting the time available for intervention in complex urban environments [<xref ref-type="bibr" rid="B2">2</xref>][<xref ref-type="bibr" rid="B4">4</xref>].</p>
      </sec>
      <sec id="sec2dot2">
        <title>2.2. Traffic Anomaly and Collision Risk Prediction</title>
        <p>Predicting traffic anomalies and collision risks is crucial for proactive road safety management. Traditional rule-based and statistical models often fail to capture the complex, non-linear relationships in traffic dynamics [<xref ref-type="bibr" rid="B3">3</xref>]. Machine Learning (ML) and Deep Learning (DL) techniques have improved predictive performance by analyzing large-scale traffic data, but many models do not effectively incorporate spatiotemporal dependencies, reducing real-world accuracy [<xref ref-type="bibr" rid="B5">5</xref>].</p>
      </sec>
      <sec id="sec2dot3">
        <title>2.3. Long Short-Term Memory (LSTM) Networks for Temporal Prediction</title>
        <p>Long Short-Term Memory (LSTM) networks, a variant of Recurrent Neural Networks (RNNs), are widely used for time-series forecasting due to their ability to learn long-term dependencies [<xref ref-type="bibr" rid="B3">3</xref>]. LSTMs have been successfully applied to traffic flow forecasting, congestion detection, and anomaly prediction by analyzing historical vehicle trajectory data [<xref ref-type="bibr" rid="B1">1</xref>]. Ji <italic>et al</italic>. [<xref ref-type="bibr" rid="B1">1</xref>] demonstrated that LSTM-based models improved predictive accuracy in V2X communication resource allocation. However, LSTMs alone lack spatial awareness, making them insufficient for modeling vehicle interactions across road networks.</p>
      </sec>
      <sec id="sec2dot4">
        <title>2.4. Graph Neural Networks (GNNs) for Spatial Modeling</title>
        <p>Graph Neural Networks (GNNs) effectively capture spatial dependencies in road networks and vehicle interactions [<xref ref-type="bibr" rid="B2">2</xref>]. By representing vehicles as nodes and their interactions as edges, GNNs model complex spatial relationships in traffic environments. Li <italic>et al</italic>. [<xref ref-type="bibr" rid="B2">2</xref>] applied GNNs to analyze Cellular V2X (C-V2X) data, significantly improving multi-vehicle collision prediction. However, GNNs alone fail to capture temporal dependencies, thereby limiting their predictive accuracy in rapidly changing traffic conditions.</p>
      </sec>
      <sec id="sec2dot5">
        <title>2.5. Integration of LSTM and GNN Models</title>
        <p>While LSTMs and GNNs are highly effective for modeling temporal and spatial dependencies, their isolated applications are insufficient for fully capturing the dynamic nature of V2X systems [<xref ref-type="bibr" rid="B4">4</xref>]. Recent research has explored hybrid deep learning architectures, but studies focusing on LSTM-GNN integration for predictive V2X safety applications remain limited [<xref ref-type="bibr" rid="B5">5</xref>]. Zoghlami <italic>et al</italic>. [<xref ref-type="bibr" rid="B5">5</xref>] examined dynamic data collection in V2X systems but did not incorporate predictive spatiotemporal modeling. This paper addresses this research gap by proposing a hybrid LSTM-GNN framework that enables proactive detection of traffic anomalies and collision risks with high accuracy.</p>
      </sec>
    </sec>
    <sec id="sec3">
      <title>3. Proposed Conceptual Framework</title>
      <p>The proposed framework integrates LSTM and GNN models to predict traffic anomalies and collision risks in V2X systems. By leveraging both temporal and spatial learning, the framework provides proactive safety alerts, improving response times for drivers and autonomous systems.</p>
      <p>Framework Components</p>
      <p>1) Data Collection &amp; Preprocessing Module</p>
      <p>Aggregates data from V2X communications and external sources.</p>
      <p>2) Modeling Module</p>
      <p>Implements LSTM for temporal anomaly prediction.Implements GNN for spatial collision risk estimation.Combines predictions through a spatiotemporal fusion mechanism.</p>
      <p>3) Risk Assessment &amp; Alert Generation Module</p>
      <p>Computes final risk scores.Issues real-time safety alerts.</p>
      <sec id="sec3dot1">
        <title>3.1. Abbreviations and Acronyms</title>
        <p>The following abbreviations and acronyms are used throughout this paper: AI (Artificial Intelligence), AV (Autonomous Vehicle), CAM (Cooperative Awareness Message), C-V2X (Cellular Vehicle-to-Everything), DL (Deep Learning), DSRC (Dedicated Short-Range Communications), GNN (Graph Neural Network), ITS (Intelligent Transportation System), LSTM (Long Short-Term Memory), ML (Machine Learning), QoS (Quality of Service), RNN (Recurrent Neural Network), V2I (Vehicle-to-Infrastructure), V2N (Vehicle-to-Network), V2P (Vehicle-to-Pedestrian), V2V (Vehicle-to-Vehicle), V2X (Vehicle-to-Everything), and URLLC (Ultra-Reliable Low-Latency Communication).</p>
      </sec>
      <sec id="sec3dot2">
        <title>3.2. Data Collection and Preprocessing</title>
        <p>The conceptual framework can accept multi-source V2X data; in the NGSIM experiment, only trajectory-derived equivalents were available:</p>
        <p>V2V: Speed, acceleration, lane changes, and braking events.V2I-equivalent: Lane identifiers and roadway geometry annotations.V2N-equivalent: Global_Time and Frame_ID synchronization metadata; weather, traffic-update, and accident-report feeds were not available.</p>
        <p>Preprocessing Steps:</p>
        <p>Normalization: Standardizes speed, acceleration, and positional data.Missing Data Handling: Uses interpolation for temporal consistency.Temporal Segmentation: Splits data into 50-time step sequences for LSTM input.Graph Construction: Represents vehicles as nodes and proximity-based interactions as edges.</p>
      </sec>
      <sec id="sec3dot3">
        <title>3.3. Temporal Prediction Using LSTM</title>
        <p>The LSTM module predicts traffic anomalies based on historical vehicle trajectories.</p>
        <p>Input: Speed, acceleration, lane positions (50-time steps).Architecture: Two stacked LSTM layers (128, 64 hidden units).Output: Normalized anomaly score (0 - 1).</p>
      </sec>
      <sec id="sec3dot4">
        <title>3.4. Spatial Collision Risk Prediction Using GNN</title>
        <p>The GNN module models vehicle interactions and collision risks.</p>
        <p>Graph Structure:Nodes: Vehicles.Edges: Vehicles within 20m proximity.Features: Speed, acceleration, X-Y coordinates.Model Architecture:Two GCN layers (32, 16 hidden units).Global graph pooling for frame-level risk estimation.</p>
      </sec>
      <sec id="sec3dot5">
        <title>3.5. Spatiotemporal Fusion Module</title>
        <p>Outputs from LSTM and GNN are combined:</p>
        <disp-formula id="FD1">
          <mml:math>
            <mml:mrow>
              <mml:mi>F</mml:mi>
              <mml:mi>i</mml:mi>
              <mml:mi>n</mml:mi>
              <mml:mi>a</mml:mi>
              <mml:mi>l</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>R</mml:mi>
              <mml:mi>i</mml:mi>
              <mml:mi>s</mml:mi>
              <mml:mi>k</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>S</mml:mi>
              <mml:mi>c</mml:mi>
              <mml:mi>o</mml:mi>
              <mml:mi>r</mml:mi>
              <mml:mi>e</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>=</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>α</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>·</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>L</mml:mi>
              <mml:mi>S</mml:mi>
              <mml:mi>T</mml:mi>
              <mml:mi>M</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>O</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mi>p</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>+</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>β</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>·</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>G</mml:mi>
              <mml:mi>N</mml:mi>
              <mml:mi>N</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>O</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mi>p</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
            </mml:mrow>
          </mml:math>
        </disp-formula>
        <p>Threshold-based alerts are issued for low, medium, and high-risk events.</p>
        <p><xref ref-type="fig" rid="fig1">Figure 1</xref> illustrates the complete processing sequence from trajectory-derived V2X-equivalent inputs through temporal and spatial modeling, score fusion, and alert generation.</p>
        <fig id="fig1">
          <label>Figure 1</label>
          <graphic xlink:href="https://html.scirp.org/file/1480499-rId17.jpeg?20260819024725" />
        </fig>
        <p><bold>Figure 1.</bold>Predictive safety framework architecture diagram.</p>
      </sec>
      <sec id="sec3dot6">
        <title>3.6. Framework Architecture Diagram</title>
        <p>1) V2X Data Sources</p>
        <p>V2V: Speed, acceleration, braking, and lane changes.V2I-equivalent: Lane identifiers and roadway geometry annotations.V2N-equivalent: Global_Time and Frame_ID synchronization metadata; no external network feeds were used.</p>
        <p>2) Data Collection &amp; Preprocessing </p>
        <p>Aggregates raw data from V2X sources.Performs cleaning, normalization, and interpolation to handle missing values.Segments data for LSTM and GNN model compatibility.</p>
        <p>3) Feature Extraction &amp; Embedding </p>
        <p>Convert raw traffic data into structured feature vectors.Embeds relevant attributes for temporal and spatial modeling.</p>
        <p>4) Model Modules</p>
        <p>Temporal Module (LSTM): Captures historical trends and sequential dependencies to predict traffic anomalies.Spatial Module (GNN): Models Road network interactions and vehicle relationships to assess collision risk.</p>
        <p>5) Spatiotemporal Fusion Module</p>
        <p>Combines outputs from LSTM (anomaly scores) and GNN (collision risk scores).Use a weighted fusion strategy to optimize predictive accuracy.</p>
        <p>6) Risk Assessment &amp; Alert Generation</p>
        <p>Compute a final risk score based on spatiotemporal predictions.Categorizes risk into low, medium, or high levels based on predefined thresholds.</p>
        <p>7) Safety Alerts</p>
        <p>Transmits real-time safety alerts to vehicles, infrastructure, and network entities.Supports automated interventions (e.g., adaptive braking, traffic signal adjustments) to prevent collisions.</p>
      </sec>
    </sec>
    <sec id="sec4">
      <title>4. Experimental Setup and Methodology</title>
      <p>This section details the experimental setup and methodology designed to implement and validate the proposed spatiotemporal predictive safety framework for V2X systems. The framework integrates Long Short-Term Memory (LSTM) networks for temporal prediction and Graph Neural Networks (GNNs) for spatial modeling, effectively capturing both the temporal evolution of traffic patterns and spatial interactions between vehicles.</p>
      <sec id="sec4dot1">
        <title>4.1. Data Collection and Preprocessing</title>
        <p><bold>1)</bold><bold>Dataset Collecti</bold><bold>on</bold></p>
        <p>The framework relies on real-world vehicle trajectory data from the Next Generation Simulation (NGSIM) dataset, with observed fields mapped into V2X-equivalent analytical groups as follows:</p>
        <p>NGSIM is not a native V2X communications dataset; therefore, the V2V, V2I, and V2N terms in this experiment denote analytical input groups rather than recorded wireless messages. Observed NGSIM fields were mapped as follows: per-vehicle speed, acceleration/deceleration, lane changes, and relative positions were treated as V2V-equivalent state information; lane identifiers and roadway geometry annotations were treated as V2I-equivalent context; and Global_Time and Frame_ID were treated as V2N-equivalent synchronization metadata. Weather, packet-level network measurements, infrastructure messages, and accident reports were not observed and were not synthesized. Accordingly, the experiment evaluates trajectory-derived V2X safety proxies rather than communication-channel performance.</p>
        <p>V2V (Vehicle-to-Vehicle):Vehicle speed profilesAcceleration and braking patternsLane change eventsV2I (Vehicle-to-Infrastructure):Lane identifiers and positioningRoad geometry information (from dataset annotations)V2N (Vehicle-to-Network):Temporal information (e.g., Global Time, Frame ID)</p>
        <p>(V2P data was not included due to the absence of pedestrian information in the NGSIM dataset.)</p>
        <p><bold>2)</bold><bold>Preprocessing Ste</bold><bold>ps</bold></p>
        <p>To ensure the data’s suitability for model training, the following preprocessing steps were performed:</p>
        <p>Normalization: Speed, acceleration, and positional coordinates were normalized to maintain consistency across sequences.Handling Missing Data: Linear interpolation was applied to preserve temporal consistency and fill missing values.Temporal Segmentation: The dataset was segmented into fixed-length sequences (50 time steps) to serve as input to the LSTM model.Graph Construction:Nodes: Represented individual vehicles in each frame.Edges: Created between vehicles within a 20-meter proximity, capturing vehicle-to-vehicle interactions.Node Features: Included speed, acceleration, and positional coordinates.Edge Weights: Based on inverse distance to represent interaction strength.Target construction: Because NGSIM does not provide explicit anomaly or collision-risk labels, both outcomes were operationalized as continuous trajectory-derived proxy scores in the range [0, 1]. For a sequence ending at frame t, the anomaly target summarizes the normalized abrupt longitudinal change (speed change and acceleration/deceleration magnitude) along with a lane-change indicator; larger departures from the vehicle’s recent trajectory yield higher scores. For each frame graph, the collision-risk target summarizes the most critical connected-vehicle pair based on inter-vehicle separation and closing motion; smaller separation and faster closing produce higher scores. Each component was scaled using training-set statistics only, then combined on a normalized scale and clipped to [0, 1]. These are surrogate safety targets rather than observed crashes, and MAE/RMSE therefore measure error in normalized score units.Prediction unit: The LSTM produces one anomaly score for each 50-step vehicle sequence ending at a given frame. The GNN applies graph convolution at the vehicle-node level, followed by global pooling, to produce a single collision-risk score for the corresponding frame graph. Fusion is therefore performed at the frame level after aligning the LSTM sequence-end scores with the GNN frame; when several vehicle sequences end in the same frame, their temporal scores are aggregated by the maximum so that the most safety-critical vehicle determines the frame-level temporal input. The final fused score and alert category are frame-level outputs.</p>
        <p>The final dataset consisted of:</p>
        <p>2,399,992 temporal sequences for LSTM modeling.11,207 spatial graphs for GNN modeling.</p>
      </sec>
      <sec id="sec4dot2">
        <title>4.2. Temporal Prediction Using LSTM</title>
        <p>The LSTM module was designed to predict traffic anomalies based on historical vehicle trajectories.</p>
        <p><bold>1) Input Featur</bold><bold>es</bold></p>
        <p>Vehicle speed (normalize)Acceleration patternsLane positions</p>
        <p><bold>2) Model Architectu</bold><bold>re</bold></p>
        <p>The LSTM network consists of two stacked layers to learn complex temporal dependencies:</p>
        <p>Input Layer: Accepts sequences of 50 time steps with 3 features each (speed, acceleration, position).LSTM Layers:First LSTM layer with 128 hidden unitsSecond LSTM layer with 64 hidden unitsDense Layer: Maps LSTM outputs to normalized anomaly scores.Output Layer: Produces a probability between 0 and 1, where larger values indicate a greater trajectory deviation.</p>
        <p><bold>3</bold><bold>)</bold><bold>Training and Validati</bold><bold>on</bold></p>
        <p>Dataset Split:Training Set: 1,679,934 sequences (70%)Validation Set: 360,059 sequences (15%)Test Set: 359,999 sequences (15%)Split protocol: Partitioning was performed before sliding-window extraction. All sequences from a given vehicle were assigned to a single subset, and contiguous scene/time blocks were kept intact so that overlapping windows from the same vehicle or local traffic episode would not cross the training, validation, and test boundaries. The frame graphs used by the GNN followed the same scene-level assignment. Normalization parameters and fusion-weight selection were computed from the training and validation subsets only; the test subset was held out for final reporting.Training Performance (3 Epochs):</p>
        <p>Due to computational constraints and rapid convergence, training was limited to 3 epochs instead of the initially planned 10. Despite this, the model reached optimal performance with near-zero loss.</p>
        <p>Epoch 1 Loss: 0.000633Epoch 2 Loss: 0.000000Epoch 3 Loss: 0.000000</p>
      </sec>
      <sec id="sec4dot3">
        <title>4.3. Spatial Collision Risk Prediction Using GNN</title>
        <p>The GNN module focuses on modeling spatial interactions between vehicles to predict collision risks.</p>
        <p>Graph Construction:</p>
        <p>Nodes: Represented vehicles in each time frame.Edges: Formed between vehicles within 20 meters to capture proximity interactions.Node Features:SpeedAccelerationPositional coordinates (X, Y)Edge Weights: Inverse of the Euclidean distance between connected vehicles.</p>
        <p>Model Architecture:</p>
        <p>Input Layer: Processes node features (speed, acceleration, X, Y).Graph Convolution Layers:First GCN layer with 32 hidden unitsSecond GCN layer with 16 hidden units</p>
        <p>Design choices were fixed before test evaluation. A 50-step window retained short-term maneuver history while limiting latency and memory demand. The 20 m edge threshold focuses message-equivalent interactions on nearby vehicles likely to influence immediate safety. Two stacked LSTM layers provide hierarchical temporal features without the training cost and overfitting risk of a deeper recurrent model, while two GCN layers permit information exchange through immediate and two-hop neighborhoods without excessive smoothing. The 128/64 LSTM and 32/16 GCN widths progressively compress each representation before scoring. Fusion weights were compared on the validation set; 0.6/0.4 was retained as a balanced operating point rather than as the minimum for every individual error metric.</p>
        <p>Global Pooling Layer: Aggregates node-level information into a graph-level representation.Output Layer: Generates one collision-risk score for each frame graph.</p>
        <p>Training and Evaluation (3 epochs):</p>
        <p>Epoch 1 Avg Loss: 0.000639Epoch 2 Avg Loss: 0.000006Epoch 3 Avg Loss: 0.000001Dataset Split:Training Graphs: 7844 (70%)Validation Graphs: 1681 (15%)Test Graphs: 1682 (15%)</p>
      </sec>
      <sec id="sec4dot4">
        <title>4.4. Spatiotemporal Fusion Module</title>
        <p>To enhance prediction accuracy, outputs from the LSTM (temporal anomaly scores) and GNN (spatial collision risk scores) modules were combined using a weighted fusion approach:</p>
        <disp-formula id="FD2">
          <mml:math>
            <mml:mrow>
              <mml:mi>F</mml:mi>
              <mml:mi>i</mml:mi>
              <mml:mi>n</mml:mi>
              <mml:mi>a</mml:mi>
              <mml:mi>l</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>R</mml:mi>
              <mml:mi>i</mml:mi>
              <mml:mi>s</mml:mi>
              <mml:mi>k</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>S</mml:mi>
              <mml:mi>c</mml:mi>
              <mml:mi>o</mml:mi>
              <mml:mi>r</mml:mi>
              <mml:mi>e</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>=</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>α</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>·</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>L</mml:mi>
              <mml:mi>S</mml:mi>
              <mml:mi>T</mml:mi>
              <mml:mi>M</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>O</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mi>p</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>+</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>β</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>·</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>G</mml:mi>
              <mml:mi>N</mml:mi>
              <mml:mi>N</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>O</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mi>p</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
            </mml:mrow>
          </mml:math>
        </disp-formula>
        <p>where:</p>
        <p><italic>α</italic> (Alpha) = 0.6 (weight for LSTM outputs)<italic>β</italic> (Beta) = 0.4 (weight for GNN outputs)</p>
        <p>Weights were determined through cross-validation to balance the influence of temporal and spatial features.</p>
      </sec>
      <sec id="sec4dot5">
        <title>4.5. Risk Assessment and Alert Generation</title>
        <p>The final risk score was compared against predefined thresholds to trigger proactive safety alerts; the tiered presentation follows established crash-warning interface principles [<xref ref-type="bibr" rid="B6">6</xref>]:</p>
        <p>Low Risk (Score &lt; 0.3): No alert issued.Medium Risk (0.3 ≤ Score &lt; 0.6): Advisory warning provided.High Risk (Score ≥ 0.6): Immediate safety intervention (e.g., automatic braking for AVs).</p>
        <p>These alerts were designed to be communicated via V2X channels, enabling nearby vehicles and infrastructure systems to respond promptly and prevent potential collisions.</p>
      </sec>
      <sec id="sec4dot6">
        <title>4.6. Methodology</title>
        <p>The following methodology was implemented to develop and evaluate the proposed spatiotemporal predictive safety framework for V2X systems. The framework integrates LSTM for temporal anomaly prediction and GNN for spatial collision risk modeling.</p>
        <p><bold>1) Data Collection and Preprocessi</bold><bold>ng</bold></p>
        <p>The dataset was collected from the Next Generation Simulation (NGSIM) dataset, containing real-world vehicle trajectory data. The data was preprocessed to ensure consistency and reliability for model training.</p>
        <p>Preprocessing Steps:Feature Normalization: Speed, acceleration, and positional coordinates were normalized for consistency.Handling Missing Data: Linear interpolation was applied to preserve temporal consistency.Temporal Segmentation: The dataset was segmented into fixed-length sequences (50-time steps) for LSTM input.Graph Construction for GNN: Nodes: Individual vehicles per frame.Edges: Vehicles within 20 meters were connected to capture V2V interactions.Node Features: Speed, acceleration, positional coordinates.Edge Weights: Based on inverse Euclidean distance.Final Dataset Structure:2,399,992 temporal sequences for LSTM modeling.11,207 spatial graphs for GNN modeling.</p>
        <p><bold>2) LSTM Temporal Predicti</bold><bold>on</bold></p>
        <p>The LSTM model was designed to predict traffic anomalies using historical vehicle trajectory data.</p>
        <p>Training Process:LSTM trained on speed, acceleration, and lane positions over time.Evaluated using regression metrics (MAE, RMSE).Model Architecture:Input Layer: Sequences of 50 time steps with 3 features (speed, acceleration, position).LSTM Layers: First LSTM layer: 128 hidden unitsSecond LSTM layer: 64 hidden unitsDense Layer: Maps LSTM outputs to a normalized anomaly score.Output Layer: Produces a continuous anomaly score (0 - 1).Training Performance (3 Epochs) (Due to rapid convergence and near-zero loss, training was stopped at 3 epochs instead of the initially planned 10 epochs):Epoch 1 Loss: 0.000633Epoch 2 Loss: 0.000000Epoch 3 Loss: 0.000000Evaluation Metrics:Mean Absolute Error (MAE): 0.018697Root Mean Squared Error (RMSE): 0.074871</p>
        <p><bold>3) GNN Spatial Collision Risk Predicti</bold><bold>on</bold></p>
        <p>The GNN module was trained to model spatial interactions among vehicles and predict collision risk.</p>
        <p>Graph Construction:Nodes: Represented individual vehicles.Edges: Formed between vehicles within 20 meters.Node Features: Speed, acceleration, positional coordinates (X, Y).Edge Weights: Inverse Euclidean distance.Model Architecture:Input Layer: Processes 4-dimensional node features (speed, acceleration, X, Y).Graph Convolution Layers: First GCN layer: 32 hidden unitsSecond GCN layer: 16 hidden unitsGlobal Pooling Layer: Aggregates node embeddings into a frame-graph representation.Output Layer: Generates one collision-risk score for each frame graph.Training Performance (3 Epochs). Like LSTM, the GNN training was limited to 3 epochs due to rapid convergence:Epoch 1 Avg Loss: 0.000639Epoch 2 Avg Loss: 0.000006Epoch 3 Avg Loss: 0.000001Evaluation Metrics:Mean Absolute Error (MAE): 0.019446Root Mean Squared Error (RMSE): 0.074733</p>
        <p><bold>4) Spatiotemporal Fusi</bold><bold>on</bold></p>
        <p>To enhance prediction accuracy, outputs from the LSTM (temporal anomaly scores) and GNN (spatial collision risk scores) were combined using weighted fusion:</p>
        <disp-formula id="FD3">
          <mml:math>
            <mml:mrow>
              <mml:mi>F</mml:mi>
              <mml:mi>i</mml:mi>
              <mml:mi>n</mml:mi>
              <mml:mi>a</mml:mi>
              <mml:mi>l</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>R</mml:mi>
              <mml:mi>i</mml:mi>
              <mml:mi>s</mml:mi>
              <mml:mi>k</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>S</mml:mi>
              <mml:mi>c</mml:mi>
              <mml:mi>o</mml:mi>
              <mml:mi>r</mml:mi>
              <mml:mi>e</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>=</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>α</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>·</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>L</mml:mi>
              <mml:mi>S</mml:mi>
              <mml:mi>T</mml:mi>
              <mml:mi>M</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>O</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mi>p</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>+</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>β</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>·</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>G</mml:mi>
              <mml:mi>N</mml:mi>
              <mml:mi>N</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>O</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mi>p</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
            </mml:mrow>
          </mml:math>
        </disp-formula>
        <p>Fusion Weights:<italic>α</italic> (LSTM Weight): 0.6<italic>β</italic> (GNN Weight): 0.4</p>
        <p>Fusion weights were determined through cross-validation to balance the impact of temporal and spatial features.</p>
        <p>Fusion Model Performance:MSE: 0.005585RMSE: 0.074733MAE: 0.018697</p>
        <p><bold>5) Risk Assessment and Alert Generati</bold><bold>on</bold></p>
        <p>The final risk score was used to trigger proactive safety alerts.</p>
        <p>Risk Thresholds:Low Risk (Score &lt; 0.3): No alert issued.Medium Risk (0.3 ≤ Score &lt; 0.6): Advisory warning.High Risk (Score ≥ 0.6): Immediate safety intervention (e.g., automatic braking).</p>
      </sec>
    </sec>
    <sec id="sec5">
      <title>5. Results</title>
      <p>This section presents the results obtained from training, evaluating, and deploying the proposed spatiotemporal predictive safety framework for V2X systems. The results are categorized into temporal anomaly prediction (LSTM), spatial collision risk prediction (GNN), and spatiotemporal fusion performance.</p>
      <sec id="sec5dot1">
        <title>5.1. Training Performance</title>
        <p>The LSTM model was trained to detect traffic anomalies based on historical vehicle trajectories, focusing on speed, acceleration, and lane position.</p>
        <p><bold>1) Training Performan</bold><bold>ce</bold></p>
        <p>The model was trained for 3 epochs, instead of the initially planned 10, due to rapid convergence and diminishing loss values. The final training and validation loss approached near-zero values, indicating highly optimized performance.</p>
        <p>LSTM Training Performance (3 Epochs):Epoch 1 Loss: 0.000633Epoch 2 Loss: 0.000000Epoch 3 Loss: 0.000000</p>
        <p><bold>2) Evaluation Metri</bold><bold>cs</bold></p>
        <p>The LSTM model was evaluated on the test dataset (359,999 sequences) using standard regression metrics.</p>
        <p>LSTM Test Performance:Mean Absolute Error (MAE): 0.018697Root Mean Squared Error (RMSE): 0.074871Mean Squared Error (MSE): 0.005606</p>
        <p>These results indicate that the LSTM model effectively captures temporal traffic patterns and provides reliable anomaly predictions.</p>
      </sec>
      <sec id="sec5dot2">
        <title>5.2. Spatial Collision Risk Prediction Using GNN</title>
        <p>The GNN model was trained to predict collision risks by analyzing spatial interactions between vehicles.</p>
        <p><bold>1) Training Performan</bold><bold>ce</bold></p>
        <p>As with LSTM, GNN training was limited to 3 epochs due to rapid convergence and minimal loss reduction beyond epoch 3.</p>
        <p>GNN Training Performance (3 Epochs):Epoch 1 Avg Loss: 0.000639Epoch 2 Avg Loss: 0.000006Epoch 3 Avg Loss: 0.000001</p>
        <p><bold>2) Evaluation Metri</bold><bold>cs</bold></p>
        <p>The GNN model was evaluated on the test dataset (1682 graphs) to assess its ability to predict vehicle collision risks.</p>
        <p>GNN Test Performance:Mean Absolute Error (MAE): 0.019446Root Mean Squared Error (RMSE): 0.074733Mean Squared Error (MSE): 0.005567</p>
        <p>The results confirm that the GNN model successfully captures spatial dependencies between vehicles and predicts collision risk with high accuracy. </p>
      </sec>
      <sec id="sec5dot3">
        <title>5.3. Spatiotemporal Fusion Performance</title>
        <p>To enhance prediction accuracy, the outputs from the LTSM (temporal anomaly scores) and the GNN (spatial collision risk scores) were fused via weighted summation.</p>
        <disp-formula id="FD4">
          <mml:math>
            <mml:mrow>
              <mml:mi>F</mml:mi>
              <mml:mi>i</mml:mi>
              <mml:mi>n</mml:mi>
              <mml:mi>a</mml:mi>
              <mml:mi>l</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>R</mml:mi>
              <mml:mi>i</mml:mi>
              <mml:mi>s</mml:mi>
              <mml:mi>k</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>S</mml:mi>
              <mml:mi>c</mml:mi>
              <mml:mi>o</mml:mi>
              <mml:mi>r</mml:mi>
              <mml:mi>e</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>=</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>α</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>·</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>L</mml:mi>
              <mml:mi>S</mml:mi>
              <mml:mi>T</mml:mi>
              <mml:mi>M</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>O</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mi>p</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>+</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>β</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mo>·</mml:mo>
              <mml:mo>
              </mml:mo>
              <mml:mi>G</mml:mi>
              <mml:mi>N</mml:mi>
              <mml:mi>N</mml:mi>
              <mml:mo>
              </mml:mo>
              <mml:mi>O</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
              <mml:mi>p</mml:mi>
              <mml:mi>u</mml:mi>
              <mml:mi>t</mml:mi>
            </mml:mrow>
          </mml:math>
        </disp-formula>
        <p>The retained fusion weights were selected through validation-set comparison as a balanced operating point:</p>
        <p><italic>α</italic> (LSTM Weight) = 0.6, <italic>β</italic> (GNN Weight) = 0.4</p>
        <p><bold>Fusion Model Evaluati</bold><bold>on</bold></p>
        <p>The fused model was tested to determine its effectiveness in predicting traffic risks.</p>
        <p>Fusion Model Test Performance:Mean Absolute Error (MAE): 0.018697Root Mean Squared Error (RMSE): 0.074733Mean Squared Error (MSE): 0.005585</p>
        <p>The fusion model achieved a balanced prediction performance by leveraging both temporal (LSTM) and spatial (GNN) insights.</p>
      </sec>
      <sec id="sec5dot4">
        <title>5.4. Sensitivity Analysis: Fusion Weight Optimization</title>
        <p>To validate the effectiveness of spatiotemporal fusion, different weighting combinations were tested:</p>
        <p>Fusion Weights (LSTM: 0.3, GNN: 0.7)MSE: 0.005578RMSE: 0.074683MAE: 0.018995Fusion Weights (LSTM: 0.5, GNN: 0.5)MSE: 0.005585RMSE: 0.074733MAE: 0.018697Fusion Weights (LSTM: 0.7, GNN: 0.3)MSE: 0.005593RMSE: 0.074786MAE: 0.018404</p>
        <p>No single fusion ratio minimized all three metrics. The 0.3/0.7 setting produced the lowest MSE and RMSE (0.005578 and 0.074683), whereas 0.7/0.3 produced the lowest MAE (0.018404). The 0.6/0.4 setting was retained as a preselected balance between temporal and spatial contributions, not as an across-metric optimum.</p>
      </sec>
      <sec id="sec5dot5">
        <title>5.5. Risk Assessment and Alert Generation</title>
        <p><bold>Table 1</bold> summarizes the three final-risk categories and their associated alert responses.</p>
        <p><bold>Table 1.</bold>The final risk score was categorized into three safety levels.</p>
        <table-wrap id="tbl1">
          <label>Table 1</label>
          <table>
            <tbody>
              <tr>
                <td rowspan="2">
                  <bold>Risk Score</bold>
                </td>
                <td colspan="2">
                  <bold>Alert Generation</bold>
                </td>
              </tr>
              <tr>
                <td>
                  <italic>
                    <bold>Alert</bold>
                  </italic>
                </td>
                <td>
                  <bold>System Response</bold>
                </td>
              </tr>
              <tr>
                <td>&lt;0.3</td>
                <td>Low Risk</td>
                <td>No alert issued</td>
              </tr>
              <tr>
                <td>0.3 - 0.6</td>
                <td>Medium Risk</td>
                <td>Advisory warning provided</td>
              </tr>
              <tr>
                <td>≥0.6</td>
                <td>High Risk</td>
                <td>Immediate safety intervention (e.g., automatic braking)</td>
              </tr>
            </tbody>
          </table>
        </table-wrap>
      </sec>
    </sec>
    <sec id="sec6">
      <title>6. Conceptual Analysis and Discussion</title>
      <p>This section presents a conceptual analysis of the proposed spatiotemporal predictive safety framework, discussing its effectiveness, limitations, and potential for real-world deployment. The discussion is structured into key aspects: model performance analysis, fusion strategy evaluation, system robustness, computational efficiency, and practical implementation considerations.</p>
      <sec id="sec6dot1">
        <title>6.1. Model Performance Analysis</title>
        <p>The LSTM-based temporal anomaly prediction and GNN-based spatial collision risk modeling both demonstrated high accuracy and low error rates, validating their effectiveness in predictive V2X safety systems.</p>
        <p><bold>Table 2</bold> compares the test-set errors of the temporal, spatial, and fused models.</p>
        <p><bold>Table 2.</bold>Model performance summary.</p>
        <table-wrap id="tbl2">
          <label>Table 2</label>
          <table>
            <tbody>
              <tr>
                <td rowspan="2">
                  <bold>Model</bold>
                </td>
                <td colspan="3">
                  <bold>Model Performance</bold>
                </td>
              </tr>
              <tr>
                <td>
                  <bold>MSE</bold>
                </td>
                <td>
                  <bold>RMSE</bold>
                </td>
                <td>
                  <bold>MAE</bold>
                </td>
              </tr>
              <tr>
                <td>LSTM (Temporal Prediction)</td>
                <td>0.005606</td>
                <td>0.074871</td>
                <td>0.018697</td>
              </tr>
              <tr>
                <td>GNN (Spatial Risk Modeling)</td>
                <td>0.005567</td>
                <td>0.074733</td>
                <td>0.019446</td>
              </tr>
              <tr>
                <td>Fusion Model</td>
                <td>0.005585</td>
                <td>0.074733</td>
                <td>0.018697</td>
              </tr>
            </tbody>
          </table>
        </table-wrap>
        <p>The LSTM model efficiently captured temporal trends in vehicle behavior, identifying anomalies in speed and acceleration. The GNN model successfully modeled spatial dependencies, detecting high-risk vehicle interactions. The fusion model did not outperform both standalone models on every metric. At 0.6/0.4, its MAE equaled the LSTM result (0.018697), its RMSE equaled the GNN result (0.074733), and its MSE (0.005585) lay between the LSTM (0.005606) and GNN (0.005567) values. The experiment therefore supports complementary spatiotemporal modeling and an explicit metric-dependent trade-off, rather than uniform error reduction.</p>
      </sec>
      <sec id="sec6dot2">
        <title>6.2. Effectiveness of Spatiotemporal Fusion</title>
        <p>The validation-set comparison retained LSTM: 0.6 and GNN: 0.4 as a balanced operating point between temporal and spatial contributions; it was not the numerical optimum for all metrics.</p>
        <p><bold>Table 3</bold> reports the metric-dependent sensitivity of the fusion model to alternative LSTM/GNN weights.</p>
        <p><bold>Table 3.</bold>Fusion weight sensitivity analysis.</p>
        <table-wrap id="tbl3">
          <label>Table 3</label>
          <table>
            <tbody>
              <tr>
                <td colspan="5">
                  <bold>Fusion Weight Sensitivity</bold>
                </td>
              </tr>
              <tr>
                <td>
                  <bold>LSTM Weight (</bold>
                  <italic>
                    <bold>α</bold>
                  </italic>
                  <bold>)</bold>
                </td>
                <td>
                  <bold>GNN Weight (</bold>
                  <italic>
                    <bold>β</bold>
                  </italic>
                  <bold>)</bold>
                </td>
                <td>
                  <bold>MSE</bold>
                </td>
                <td>
                  <bold>RMSE</bold>
                </td>
                <td>
                  <bold>MAE</bold>
                </td>
              </tr>
              <tr>
                <td>0.3</td>
                <td>0.7</td>
                <td>0.005578</td>
                <td>0.074683</td>
                <td>0.018995</td>
              </tr>
              <tr>
                <td>0.5</td>
                <td>0.5</td>
                <td>0.005585</td>
                <td>0.074733</td>
                <td>0.018697</td>
              </tr>
              <tr>
                <td>0.6</td>
                <td>0.4</td>
                <td>0.005585</td>
                <td>0.074733</td>
                <td>0.018697</td>
              </tr>
              <tr>
                <td>0.7</td>
                <td>0.3</td>
                <td>0.005593</td>
                <td>0.074786</td>
                <td>0.018404</td>
              </tr>
            </tbody>
          </table>
        </table-wrap>
        <p>The sensitivity analysis reveals a metric-dependent trade-off. Increasing the GNN weight to 0.7 yielded the lowest MSE and RMSE but increased MAE, while increasing the LSTM weight to 0.7 yielded the lowest MAE but higher squared-error measures. The selected 0.6/0.4 setting preserves contributions from both modalities, but these results do not establish that it is more accurate than each standalone model under every metric.</p>
      </sec>
      <sec id="sec6dot3">
        <title>6.3. Robustness of the Framework</title>
        <p>The proposed predictive safety framework was tested on a large-scale real-world dataset (NGSIM) and demonstrated high generalizability. However, several factors influence model robustness, including:</p>
        <p>Scalability: The LSTM-GNN hybrid can be extended to larger datasets without significant performance degradation.Adaptability: The framework can incorporate additional V2X data sources, such as weather conditions and pedestrian activity, to improve accuracy.Fault Tolerance: The system was designed to handle missing or noisy data through interpolation and normalization techniques.</p>
        <p>The flexibility of the model architecture makes it suitable for real-time deployment in intelligent transportation systems (ITS).</p>
      </sec>
      <sec id="sec6dot4">
        <title>6.4. Computational Efficiency</title>
        <p>Given the high-volume nature of V2X data, computational efficiency is critical for real-time deployment.</p>
        <p>Optimizations Applied:Batch Processing: Improved memory management for large datasets.Early Stopping: Limited training to 3 epochs after rapid convergence.Graph Batching: Processed spatial data in mini-batches, reducing GPU memory usage.</p>
        <p><bold>Table 4</bold> summarizes the hardware, memory consumption, and observed training time.</p>
        <p><bold>Table 4.</bold>GPU utilization summary.</p>
        <table-wrap id="tbl4">
          <label>Table 4</label>
          <table>
            <tbody>
              <tr>
                <td colspan="2">
                  <bold>GPU Utilization</bold>
                </td>
              </tr>
              <tr>
                <td>
                  <bold>Metric</bold>
                </td>
                <td>
                  <bold>Value</bold>
                </td>
              </tr>
              <tr>
                <td>GPU Used</td>
                <td>Tesla T4</td>
              </tr>
              <tr>
                <td>Memory Consumption</td>
                <td>~10 GB (LSTM + GNN)</td>
              </tr>
              <tr>
                <td>Processing Time</td>
                <td>~15 Minutes (Training)</td>
              </tr>
            </tbody>
          </table>
        </table-wrap>
        <p>The efficient training process ensures that the model can be integrated into real-time V2X safety systems.</p>
      </sec>
      <sec id="sec6dot5">
        <title>6.5. Real-World Deployment Considerations</title>
        <p>To deploy this predictive safety framework in real-world V2X environments, the following factors must be addressed:</p>
        <p><bold>1) Latency Constrain</bold><bold>ts</bold></p>
        <p>Real-time risk assessment requires low-latency processing.Solution: Deploy models on edge computing devices in vehicles or roadside units.</p>
        <p><bold>2) Data Privacy and Securi</bold><bold>ty</bold></p>
        <p>V2X data contains sensitive vehicle and location information.Solution: Implement secure federated learning techniques to process data without sharing raw information.</p>
        <p><bold>3) V2X Communication Infrastructu</bold><bold>re</bold></p>
        <p>Model effectiveness depends on fast and reliable V2X networks (e.g., 5G, DSRC, C-V2X).Solution: Optimize the model to function even with intermittent connectivity.</p>
      </sec>
      <sec id="sec6dot6">
        <title>6.6. Limitations and Future Work</title>
        <p>Despite strong performance and efficiency, certain limitations remain:</p>
        <p><bold>1) Limited Data Sourc</bold><bold>es</bold></p>
        <p>The study did not include V2P (Vehicle-to-Pedestrian) interactions, as NGSIM lacks pedestrian data.Future Work: Integrate multimodal V2X data, including weather, road conditions, and pedestrian activity.</p>
        <p><bold>2) Dependency on Graph Structu</bold><bold>re</bold></p>
        <p>The GNN model relies on predefined edge connections between vehicles.Future Work: Use dynamic graph learning techniques to adjust edges in real-time based on evolving traffic flow.</p>
        <p><bold>3) Model Calibration for Different Traffic Scenari</bold><bold>os</bold></p>
        <p>The model was trained on freeway scenarios (NGSIM) but may require retraining for urban traffic.Future Work: Train on diverse datasets, including city intersections, roundabouts, and mixed traffic conditions.</p>
      </sec>
    </sec>
    <sec id="sec7">
      <title>7. Conclusion</title>
      <p>The proposed LSTM-GNN hybrid model integrates temporal anomaly prediction and spatial collision-risk modeling using trajectory-derived proxy targets. Spatiotemporal fusion at 0.6/0.4 produced balanced errors comparable to the standalone models, while alternative weights favored different metrics; the present results therefore demonstrate feasibility rather than uniform superiority. Because NGSIM contains trajectories rather than native V2X messages or observed crash labels, communication performance and real-world alert effectiveness remain to be validated. Further work should add native multimodal V2X data, observed safety outcomes, external-site evaluation, and real-time deployment tests. Future research will focus on enhancing multimodal V2X integration and optimizing model deployment for autonomous vehicle safety systems.</p>
    </sec>
    <sec id="sec8">
      <title>Acknowledgements</title>
      <p>I would like to thank my family for supporting me in writing this paper. To Brother Eduardo V. Manalo for his spiritual guidance. And above all else, to our Almighty God, who made this journey possible and attainable.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <title>References</title>
      <ref id="B1">
        <label>1.</label>
        <citation-alternatives>
          <mixed-citation publication-type="journal">Ji, M., Wu, Q., Fan, P., Cheng, N., Chen, W., Wang, J., <italic>et al</italic>. (2025) Graph Neural Networks and Deep Reinforcement Learning-Based Resource Allocation for V2X Communications. <italic>IEEE</italic><italic>Internet</italic><italic>of</italic><italic>Things</italic><italic>Journal</italic>, 12, 3613-3628. https://doi.org/10.1109/jiot.2024.3469547 <pub-id pub-id-type="doi">10.1109/jiot.2024.3469547</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1109/jiot.2024.3469547">https://doi.org/10.1109/jiot.2024.3469547</ext-link></mixed-citation>
          <element-citation publication-type="journal">
            <person-group person-group-type="author">
              <string-name>Ji, M.</string-name>
              <string-name>Wu, Q.</string-name>
              <string-name>Fan, P.</string-name>
              <string-name>Cheng, N.</string-name>
              <string-name>Chen, W.</string-name>
              <string-name>Wang, J.</string-name>
            </person-group>
            <year>2025</year>
            <article-title>Graph Neural Networks and Deep Reinforcement Learning-Based Resource Allocation for V2X Communications</article-title>
            <source>IEEE Internet of Things Journal</source>
            <volume>12</volume>
            <pub-id pub-id-type="doi">10.1109/jiot.2024.3469547</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B2">
        <label>2.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Li, P., Wu, K., Cheng, Y., Parker, S.T. and Noyce, D.A. (2024) How Does C-V2X Perform in Urban Environments? Results from Real-World Experiments on Urban Arterials. <italic>IEEE</italic><italic>Transactions</italic><italic>on</italic><italic>Intelligent</italic><italic>Vehicles</italic>, 9, 2520-2530. https://doi.org/10.1109/tiv.2023.3326735 <pub-id pub-id-type="doi">10.1109/tiv.2023.3326735</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1109/tiv.2023.3326735">https://doi.org/10.1109/tiv.2023.3326735</ext-link></mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Li, P.</string-name>
              <string-name>Wu, K.</string-name>
              <string-name>Cheng, Y.</string-name>
              <string-name>Parker, S.T.</string-name>
              <string-name>Noyce, D.A.</string-name>
            </person-group>
            <year>2024</year>
            <article-title>How Does C-V2X Perform in Urban Environments? Results from Real-World Experiments on Urban Arterials</article-title>
            <source>IEEE Transactions on Intelligent Vehicles</source>
            <volume>9</volume>
            <pub-id pub-id-type="doi">10.1109/tiv.2023.3326735</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B3">
        <label>3.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Nazzal, M., Khreishah, A., Lee, J., Angizi, S., Al-Fuqaha, A. and Guizani, M. (2024) Semi-Decentralized Inference in Heterogeneous Graph Neural Networks for Traffic Demand Forecasting: An Edge-Computing Approach. <italic>IEEE</italic><italic>Transactions</italic><italic>on</italic><italic>Vehicular</italic><italic>Technology</italic>, 73, 19400-19416. https://doi.org/10.1109/tvt.2024.3355971 <pub-id pub-id-type="doi">10.1109/tvt.2024.3355971</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1109/tvt.2024.3355971">https://doi.org/10.1109/tvt.2024.3355971</ext-link></mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Nazzal, M.</string-name>
              <string-name>Khreishah, A.</string-name>
              <string-name>Lee, J.</string-name>
              <string-name>Angizi, S.</string-name>
              <string-name>Al-Fuqaha, A.</string-name>
              <string-name>Guizani, M.</string-name>
            </person-group>
            <year>2024</year>
            <article-title>Semi-Decentralized Inference in Heterogeneous Graph Neural Networks for Traffic Demand Forecasting: An Edge-Computing Approach</article-title>
            <source>IEEE Transactions on Vehicular Technology</source>
            <volume>73</volume>
            <pub-id pub-id-type="doi">10.1109/tvt.2024.3355971</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B4">
        <label>4.</label>
        <citation-alternatives>
          <mixed-citation publication-type="other">Malik, S., Khan, M.J., Khan, M.A. and El-Sayed, H. (2023) Collaborative Perception—The Missing Piece in Realizing Fully Autonomous Driving. <italic>Sensors</italic>, 23, Article 7854. https://doi.org/10.3390/s23187854 <pub-id pub-id-type="doi">10.3390/s23187854</pub-id><pub-id pub-id-type="pmid">37765911</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3390/s23187854">https://doi.org/10.3390/s23187854</ext-link></mixed-citation>
          <element-citation publication-type="other">
            <person-group person-group-type="author">
              <string-name>Malik, S.</string-name>
              <string-name>Khan, M.J.</string-name>
              <string-name>Khan, M.A.</string-name>
              <string-name>El-Sayed, H.</string-name>
            </person-group>
            <year>2023</year>
            <article-title>Collaborative Perception—The Missing Piece in Realizing Fully Autonomous Driving</article-title>
            <source>Sensors</source>
            <volume>23</volume>
            <elocation-id>7854</elocation-id>
            <pub-id pub-id-type="doi">10.3390/s23187854</pub-id>
            <pub-id pub-id-type="pmid">37765911</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B5">
        <label>5.</label>
        <citation-alternatives>
          <mixed-citation publication-type="confproc">Zoghlami, C., Kacimi, R. and Dhaou, R. (2022) A Study on Dynamic Collection of Cooperative Awareness Messages in V2X Safety Applications. 2022 <italic>IEEE</italic>19 <italic>th Annual Consumer Communications &amp; Networking Conference</italic>( <italic>CCNC</italic>), Las Vegas, 8-11 January 2022, 723-724. https://doi.org/10.1109/ccnc49033.2022.9700720 <pub-id pub-id-type="doi">10.1109/ccnc49033.2022.9700720</pub-id><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1109/ccnc49033.2022.9700720">https://doi.org/10.1109/ccnc49033.2022.9700720</ext-link></mixed-citation>
          <element-citation publication-type="confproc">
            <person-group person-group-type="author">
              <string-name>Zoghlami, C.</string-name>
              <string-name>Kacimi, R.</string-name>
              <string-name>Dhaou, R.</string-name>
            </person-group>
            <year>2022</year>
            <article-title>A Study on Dynamic Collection of Cooperative Awareness Messages in V2X Safety Applications</article-title>
            <source>2022 IEEE 19th Annual Consumer Communications &amp; Networking Conference (CCNC)</source>
            <volume>8</volume>
            <pub-id pub-id-type="doi">10.1109/ccnc49033.2022.9700720</pub-id>
          </element-citation>
        </citation-alternatives>
      </ref>
      <ref id="B6">
        <label>6.</label>
        <citation-alternatives>
          <mixed-citation publication-type="web">National Highway Traffic Safety Administration (2007) Crash Warning System Interfaces: Human Factors Insights and Lessons Learned. https://www.nhtsa.gov/document/crash-warning-system-interfaces-human-factors-insights-and-lessons-learned-0</mixed-citation>
          <element-citation publication-type="web">
            <year>2007</year>
            <article-title>Crash Warning System Interfaces: Human Factors Insights and Lessons Learned</article-title>
          </element-citation>
        </citation-alternatives>
      </ref>
    </ref-list>
  </back>
</article>