Hierarchical Federated Transfer Learning in Digital Twin-Based Vehicular Networks
Summary
This paper proposes Hierarchical Federated Transfer Learning (HFTL) for Digital Twin-based Vehicular Ad hoc Networks, addressing data heterogeneity and sparsity via vehicle clustering and a data quality score mechanism to defend against malicious vehicles.
View Cached Full Text
Cached at: 08/13/26, 03:36 PM
# Hierarchical Federated Transfer Learning in Digital Twin-Based Vehicular Networks
Source: [https://arxiv.org/html/2608.11532](https://arxiv.org/html/2608.11532)
## Hierarchical Federated Transfer Learning in Digital Twin\-Based Vehicular NetworksJournal:High\-Confidence Computing
Qasim Ziaa, Saide Zhub, Haoxin Wanga, Zafar Iqbala, Yingshu LiaAffiliation:Department of Computer Science, Georgia State University, 7th Floor, 25 Park Place Bldg, Atlanta, 30303, GA, USAAffiliation:College of Information Sciences and Technology, Penn State Berks, 288 N\. Burrowes Rd, Reading, 19610, PA, USA
###### Abstract
In recent research on the Digital Twin\-based Vehicular Ad hoc Network\(DT\-VANET\), Federated Learning \(FL\) has shown its ability to provide data privacy\. However, Federated learning struggles to adequately train a global model when confronted with data heterogeneity and data sparsity among vehicles, which ensure suboptimal accuracy in making precise predictions for different vehicle types\. To address these challenges, this paper combines Federated Transfer Learning \(FTL\) to conduct vehicle clustering related to types of vehicles and proposes a novel Hierarchical Federated Transfer Learning \(HFTL\)\. We construct a framework for DT\-VANET, along with two algorithms designed for cloud server model updates and intra\-cluster federated transfer learning, to improve the accuracy of the global model\. In addition, we developed a data quality score\-based mechanism to prevent the global model from being affected by malicious vehicles\. Lastly, detailed experiments on real\-world datasets are conducted, considering different performance metrics that verify the effectiveness and efficiency of our algorithm\.
###### Keywords:
Vehicular ad\-hoc network, Hierarchical federated transfer learning, Vehicular Digital Twin, autonomous vehicle, Digital Twin\-based Vehicular Networks\.
††graphicalabstract:![[Uncaptioned image]](https://arxiv.org/html/2608.11532v1/Graphic_Abstract_HFTL.png)††highlights:Research highlight 1 Research highlight 2## 1Introduction
Vehicular networks enable numerous applications for Intelligent Transportation Systems \(ITSs\), including traffic control, infotainment, and accident reporting[Noor\-A\-Rahim et al\. 2020](https://arxiv.org/html/2608.11532#bib.bib1)\. These applications rely on various performance criteria, such as reliability, latency, and user\-defined features like the quality of physical experience\. Thus, the effective handling of vehicle heterogeneity, data privacy, and limited computational resources has become an urgent challenge to discuss, while current wireless technologies struggle to assist it thoroughly[Zia et al\. 2016](https://arxiv.org/html/2608.11532#bib.bib2)\. Despite advanced onboard sensors and processing capabilities of autonomous vehicles, they still face challenges such as limited sensing range, limited local processing capacity, and communication barriers[He et al\. 2022](https://arxiv.org/html/2608.11532#bib.bib3)\. To meet these diverse requirements, the development of vehicular networks must incorporate two key advancements: proactive, intelligent analytics and self\-sustaining wireless systems\.
Table 1:NomenclatureProactive Intelligent Analytics is a data analytics technique that analyzes incoming data while anticipating potential challenges, trends, and opportunities in advance\. It involves real\-time or nearly real\-time data analysis using the most advanced machine learning, artificial intelligence, and prediction models\. Proactive online learning wireless systems will make it possible to optimize wireless resources to ensure the quality of service \(QoS\) of diverse ITS applications with varying requirements\.
Self\-sustaining wireless systems can function without constant external power sources\. These systems are helpful for Internet of Things\(IoT\) devices, remote locations, and scenarios where battery replacement or maintenance is challenging\. Self\-sustaining wireless networks will allow ITS to run with the least operator or user interaction\.
The combination of digital twins and vehicular networks has become a promising paradigm in recent years, transforming our understanding of and approach to managing transportation systems\. The development and growth of the digital twin provided insight into how to address the problems associated with autonomous vehicles\. Virtual representation of physical systems serves as the foundation for digital twins\. A digital twin is an accurate virtual representation of a physical thing on a cloud that shows its lifetime and status\. The term “Digital Twin\-based Vehicular Ad\-Hoc Network” refers to the cloud\-based digital twin of a physical vehicle that synchronizes real\-time sensing data from actual vehicles through wireless communication\. Although digital twins can enhance the understanding of vehicle systems through virtual representations, they still require efficient data processing and analysis technologies to enable decision\-making\. Distributed or centralized training can serve as a foundation for Machine Learning \(ML\) because vehicles are reluctant to share their private information, and centralized training causes privacy leaks by moving data from dispersed devices to a centralized location\. At the same time, traditional centralized machine learning methods have proven insufficient when dealing with nonindependent and identically distributed \(non\-IID\) data\.
Artificial Intelligence \(AI\) and Machine Learning \(ML\) have revolutionized biomedical research, enabling breakthroughs in areas such as antibody design[Ahmed et al\. 2026b](https://arxiv.org/html/2608.11532#bib.bib21), medical imaging[Ahmed et al\. 2023](https://arxiv.org/html/2608.11532#bib.bib13), and protein\-protein interaction prediction[Ahmed et al\. 2026a](https://arxiv.org/html/2608.11532#bib.bib17)\. This demonstrates the power of AI in understanding complex molecular interactions[Ahmed et al\. 2025](https://arxiv.org/html/2608.11532#bib.bib25)\. However, the data sharing remains a bottleneck for research due to privacy concerns\. Federated Learning\(FL\) is the most recent technique in digital twin\-based vehicular ad hoc networks that addresses centralized learning \(CL\) issues[Khan et al\. 2023](https://arxiv.org/html/2608.11532#bib.bib4)\. FL is well known for its ability to promote collaboration and protect data privacy in various domains, including industrial and medical cyber\-physical systems[AbdulRahman et al\. 2020](https://arxiv.org/html/2608.11532#bib.bib5)[Li et al\. 2020](https://arxiv.org/html/2608.11532#bib.bib6)\. FL, which employs dispersed vehicles to learn a global FL model without transporting data from devices to a centralized site for training, is presented as a solution to this centralized ML constraint[Khan et al\. 2021](https://arxiv.org/html/2608.11532#bib.bib7)\. Compared to centralized ML, FL maintains better privacy\. However, traditional federated learning faces issues such as slow convergence and insufficient accuracy when dealing with the diversity of vehicle types and limited vehicle data\. Therefore, we use the Hierarchical Federated Transfer Learning \(HFTL\) technique in digital twin\-based vehicular networks \(DT\-VANET\)\. HFTL groups vehicles into different clusters based on their type, and pre\-trained models are used to perform personalized model fine\-tuning that fits each vehicle type’s specific needs\. It also improves both prediction accuracy and convergence speed in a meaningful way\.
Figure 1:High\-level architecture of Hierarchical Federated Transfer Learning for digital twin\-based vehicular networkWe also resolved another issue related to optimizing the route associated with the specific vehicle type and managing the aggregation of communication overhead between various vehicle types that makes precise and accurate predictions for each vehicle type\. Our HFTL for DT\-VANET has also successfully resolved this issue\. To ensure our architecture is trustworthy and reliable, each vehicle will assess its data quality, vehicle health, and safe driving performance score, and these scores will be stored in the blockchain of the cloud, which can be used to give rewards later\. We use blockchain’s tamper\-proof feature to ensure that recorded data cannot be maliciously altered\. In that way, they are stopping untrustworthy vehicle nodes from uploading low\-quality or fake data\. Additionally, the distributed consensus mechanism of blockchain allows data quality scores to be shared across multiple nodes, making the global model updates more fair and reliable\. Furthermore, the reward mechanism for vehicles can be automatically carried out through smart contracts\. So, our research paper addresses issues and challenges related to privacy, efficiency, accuracy, and scalability of machine learning models in digital twin vehicle networks, enhancing model performance through implementing an HFTL approach\.
To the best of our knowledge and understanding, our work is the first to review HFTL and Trustworthiness metrics in a single architecture of DT\-VANET\. Also, we illustrate a scenario to enable vehicle networks based on digital twins through the diagram\. We also discussed the related work, algorithms of HFTL, and performance metrics in twin\-based vehicular networks\. Therefore, the following is a summary of our contributions:
1. 1\.We build a general thorough architecture for vehicular networks based on digital twins\. A scenario for a digital twin\-based vehicular network is shown, which enables efficient data synchronization between physical and virtual nodes of vehicles, enhancing the model’s scalability and response speed\. Additionally, we offer diagrams for twin\-based vehicle networks\.
2. 2\.We outline the algorithms of federated transfer learning in digital twin\-based vehicular networks\. It allows for model customization based on different vehicle types, which improves prediction accuracy and training efficiency\.
3. 3\.Lastly, we conducted experiments on real\-time datasets to evaluate the performance metrics\. By comparing the performance of HFTL, Federated Learning \(FL\), and Centralized Learning \(CL\) in terms of model accuracy, training time, communication overhead, and resource consumption, we validated the advantages of HFTL in different scenarios\.
The rest of the paper is structured as follows: The related works are reviewed in Section[2](https://arxiv.org/html/2608.11532#S2)\. We suggest a Weighted Cloud Server Cycling Model Update Algorithm and an Inner Cluster Federated Transfer Learning Algorithm in Section[3](https://arxiv.org/html/2608.11532#S3)\. Subsequently, in Section[4](https://arxiv.org/html/2608.11532#S4), comprehensive experiments are conducted and evaluated to determine the efficiency and effectiveness of our suggested architecture\. Finally, Section[6](https://arxiv.org/html/2608.11532#S6)discusses our conclusions\.
## 2Related Works
This section overviews recent research on Hierarchical Federated Learning, Federated Transfer Learning, and Clustering Techniques for VANET\.
### 2\.1Hierarchical Federated Learning in VANET
Numerous research has presented the use of Federated Learning to address data privacy\. However, Google suggested the first work on Federated Learning \(FL\) in 2016[Konečnỳ et al\. 2016](https://arxiv.org/html/2608.11532#bib.bib8)for conventional FL\. In 2022, Li et al\.[Li et al\. 2022](https://arxiv.org/html/2608.11532#bib.bib9)proposed a FEEL framework using Federated Learning for non\-IID data\. To predict the high accuracy of network traffic with a high volume of data, Sepasgozar et al\.[Sepasgozar and Pierre 2022](https://arxiv.org/html/2608.11532#bib.bib10)proposed a new Network Traffic Prediction \(Fed\-NTP\) based on federated learning\. Hierarchical Learning in Machine Learning was first introduced by Zhang et al\.[Zhang and Zhang 2006](https://arxiv.org/html/2608.11532#bib.bib11)\. The human approach to problem\-solving in hierarchies inspires their work\. In 2021, Goncalves et al\.[Gonçalves et al\. 2021](https://arxiv.org/html/2608.11532#bib.bib12)discuss the Hierarchical approach for Security Framework in VANET\. In 2024, HaghighiFard et al\.[HaghighiFard and Coleri 2024](https://arxiv.org/html/2608.11532#bib.bib14)discuss the use of Hierarchical Federated Learning in VANET\. It uses the cosine similarity of FL model parameters and average relative speed as a metric for making clusters\. Although Li et al\.[Li et al\. 2022](https://arxiv.org/html/2608.11532#bib.bib9), the FEEL framework successfully applied federated learning to non\-IID data environments\. It did not fully address the data communication challenges in highly dynamic VANET environments\. In contrast, HaghighiFard et al\.[HaghighiFard and Coleri 2024](https://arxiv.org/html/2608.11532#bib.bib14)introduced hierarchical federated learning, which reduced communication overhead to some extent\. However, their clustering method still relies on fixed vehicle attributes, making it less adaptable to the rapid changes in vehicle environments\.
### 2\.2Federated Transfer Learning in VANET
Federated Transfer Learning\(FTL\) facilitates knowledge transfer without affecting user privacy\. The FTL was first discussed by Liu et al\. in 2020[Liu et al\. 2020](https://arxiv.org/html/2608.11532#bib.bib15), where they enabled the target\-domain party to use the source domain’s rich level to create adaptable and efficient models\. In 2022, Otoum et al\.[Otoum et al\. 2022](https://arxiv.org/html/2608.11532#bib.bib16)proposed an intrusion detection system in VANET\. On that account, they have compared Split Learning, Federated Learning, and Transfer Learning\. However, Liu et al\.[Liu et al\. 2020](https://arxiv.org/html/2608.11532#bib.bib15)paper discusses Federated Transfer Learning\. Still, the paper does not consider DT\-VANET for their findings\. Otoum et al\.[Otoum et al\. 2022](https://arxiv.org/html/2608.11532#bib.bib16)consider federated transfer learning only for intrusion detection scenarios, not for digital twin\-based communication\.
### 2\.3Clustering Techniques for VANET
The primary purpose of clustering is to find discrete groups of vehicles in the environment\. It will reduce the communication with the roadside units\.[Zia 2015](https://arxiv.org/html/2608.11532#bib.bib18)discusses various data\-centric protocols related to clustering\. The PasCar routing protocol was presented by Wang et al\.[Wang and Lin 2013](https://arxiv.org/html/2608.11532#bib.bib19)in 2013 to make a reliable cluster structure based on a single hop\. It chooses suitable participants from candidate nodes\. However, the single\-hop mechanism limits the system’s coverage and stability\. Because of this, there has been a lot of research on multi\-hop clustering techniques in the years that followed\. In 2022, Rashid et al\.[Rashid et al\. 2020](https://arxiv.org/html/2608.11532#bib.bib20)discussed efficient multi\-hop clustering based on prediction\. This paper discusses increasing the cluster coverage area increase, link stability, energy efficiency, and mobility characteristics increase\. Temurnikar et al\.[Temurnikar et al\. 2022](https://arxiv.org/html/2608.11532#bib.bib22)in 2022 discussed a new Particle Swarm Optimization \(PSO\) based multi\-hop method because of which you can use the best route\. We can also select the Cluster head, which is much more stable\. It also identified false messages from malicious vehicles\.[Zia et al\. 2024](https://arxiv.org/html/2608.11532#bib.bib23)in 2024, discussed the possibility of using priorities to achieve faster response time\. Compared to the techniques already mentioned, this paper introduces Hierarchical Federated Transfer Learning \(HFTL\), which offers more flexible clustering to address vehicle heterogeneity\. Additionally, integrating digital twin technology enables real\-time data synchronization, improving prediction accuracy and reducing communication overhead\.
In this section, we proposed a high\-level secure and robust federated transfer learning architecture for vehicular networks, as shown in Figure[1](https://arxiv.org/html/2608.11532#S1.F1)\. There is a digital twin layer, and the physical layer forms the two main layers of the proposed architecture\. The physical layer consists of objects like base stations \(BS\), which help end users and distribute autonomous vehicles\. The Digital Twin Layer consists of four essential parts: data storage, twin management, virtual model mapping, and blockchain[Dai and Zhang 2022](https://arxiv.org/html/2608.11532#bib.bib24)\. It uses digital twins to effectively map the virtual twin system objects and real\-vehicle edge computing\. The cloud server implements the digital twin layer, and the idea of digital twin objects is used[Khan et al\. 2022a](https://arxiv.org/html/2608.11532#bib.bib26)\. A virtual representation of the physical system is known as a digital twin object, and it is doable to simulate numerous vehicle network operations and applications using this virtual representation type of the physical vehicular network\. Mathematical and experimental methodologies can be used to model such network applications/functions[Khan et al\. 2022b](https://arxiv.org/html/2608.11532#bib.bib27)\. However, as we know, mathematical modeling is highly dependent on assumptions; it might not be able to denote the actual network accurately\. Meanwhile, mistakes may prevent experimental modeling from being accurate during experimentation\. In digital twin\-based vehicular networks, a data\-driven modeling approach based on Federated Transfer Learning may be considered to overcome these problems\. It also provides enhanced network performance and efficiency, improved safety and reliability, and allows advanced applications to run\. Hierarchical Federated Transfer Learning\(HFTL\) in digital twins is used for precise and accurate prediction according to the vehicle type\. Blockchain is also used to secure the trustworthiness metrics of the individual vehicle\. We will now discuss the features of this architecture one by one\.
Table 2:Categorization of Vehicles### 3\.1Hierarchical Federated Transfer Learning \(HFTL\)
HFTL’s clustering considers vehicle size, purpose, driving patterns, and computational resources\. It will help to build a customized model\. This model aligns with the learning goals and data properties specific to each kind of vehicle\. As shown in Figure[1](https://arxiv.org/html/2608.11532#S1.F1), vehicles are grouped according to the type of vehicles\. The types of vehicles are classified according to vehicle size, purpose, and computational resources in Table[2](https://arxiv.org/html/2608.11532#S3.T2)\. Generally, the vehicle with the most advanced features is selected as the cluster head\.
DT=∑i=1n∑j=1nTvj∈CiDT=\\sum\_\{i=1\}^\{n\}\\sum\_\{j=1\}^\{n\}Tv\_\{j\}\\in C\_\{i\}\(1\)
WhereCiC\_\{i\}is a cluster that contains a type of vehicle j\.
Equation[1](https://arxiv.org/html/2608.11532#S3.E1)discusses that in the Digital Twin, there are many clusters fromCiC\_\{i\}toCnC\_\{n\}where each clusterCiC\_\{i\}contains only one type of vehicleTvjTv\_\{j\}\.
Secondly, HFTL improves the convergence and heterogeneity in Dispersed Federated Learning\. In HFTL, a pre\-trained model is shared depending on the type of vehicle and target application, such as traffic flow prediction, collision risk prediction, intelligent parking prediction, etc\. This pre\-trained model will be fine\-tuned using the vehicle’s private local data\. This enables the model to be customized for the environment and based on the vehicle’s driving patterns\. For example, public vehicles that require higher precision will have their model fine\-tuning focus more on safety and route optimization\. However, private vehicles may prioritize reducing idle times and optimizing the driving experience\. After training, share the model update to the cloud server for the global model update\. Therefore, model updates are combined from every vehicle’s digital twin to the cloud to make it a more robust and heterogeneous global model\. Several digital twins of vehicles use these model updates to make decisions and predictions\. The model updates from every vehicle digital twin are given weight according to the trustworthiness metrics\. These metrics are recorded and verified through blockchain\. Vehicles with higher trustworthiness have greater weight in global model updates, ensuring the security and accuracy of the model updates\. These weights will be considered for the Global model update using the Federated Averaging Method\.
Figure 2:The Flowchart and Vehicle type data flow for the Inner Cluster of Hierarchical Federated Transfer Learning
### 3\.2Digital Twin Based Vehicular Networks
As we discussed briefly above, the digital twin architecture in VANET allows us to have many benefits, including improved decision\-making, increased safety, predictive maintenance, optimized network performance, enhanced scalability, enhanced simulation and testing, real\-time monitoring and maintenance, and cost reduction\.
### 3\.3Blockchain Integration in Architecture
In this architecture, blockchain can be defined as a distributed ledger integrated into DT\-VANET\-based HFTL architecture to manage and store weights\. The trustworthiness metrics calculate every weight, as discussed in the next section\. Low\-weight model updates are not considered to update the global model to maintain integrity\. It will also help us detect anomalies\. We ensure the secure DT\-VANET\-based HFTL architecture through the weights and tamper\-proof blockchain\.
### 3\.4Trustworthiness Metrics
DTSz=w1⋅DTCz\+w2⋅SRCz\+w3⋅Ez\+w4⋅CSFz\+w5⋅GRHLz\+w6⋅SFRz\\begin\{split\}DTS\_\{z\}=&\\,w\_\{1\}\\cdot DTC\_\{z\}\+w\_\{2\}\\cdot SRC\_\{z\}\+w\_\{3\}\\cdot E\_\{z\}\\\\ &\+w\_\{4\}\\cdot CSF\_\{z\}\+w\_\{5\}\\cdot GRHL\_\{z\}\+w\_\{6\}\\cdot SFR\_\{z\}\\end\{split\}\(2\)where:
- 1\.DTSzDTS\_\{z\}= The Data Quality Score of vehicle z
- 2\.DTCzDTC\_\{z\}= The Data Completeness of vehicle z
- 3\.SRCzSRC\_\{z\}= The sensor collaboration of vehicle z
- 4\.EzE\_\{z\}= The events \(traffic incidents, road conditions, etc\.\) recorded by vehicle z
- 5\.CSFzCSF\_\{z\}= The constant format \(According to Standard Data Formats\) for vehicle z
- 6\.GRHLzGRHL\_\{z\}= The General Health of vehicle z \(e\.g\., looking after maintenance issues\)
- 7\.SFRzSFR\_\{z\}= The reliable and safe driving patterns for vehicle z\.
- 8\.w1w\_\{1\},w2w\_\{2\},w3w\_\{3\}…w6w\_\{6\}are the weights that are assigned with each factor
Every vehicle will calculate its data quality score depending on data completeness, sensor collaboration, events, constant format, the vehicle’s general health, and safe, reliable driving patterns\.
RPSz\(tm\+1\)=RPSz\(tm\)\+α⋅DQSz−β⋅MzRPS\_\{z\}\(tm\+1\)=RPS\_\{z\}\(tm\)\+\\alpha\\cdot DQS\_\{z\}\-\\beta\\cdot M\_\{z\}\(3\)where:
- 1\.RPSzRPS\_\{z\}\(tm\) = The Reputation Score of vehicle z
- 2\.α\\alpha= The Weight for data quality score
- 3\.β\\beta= The malicious behaviour Penalty factor
- 4\.MzM\_\{z\}= The malicious activity indicator for vehicle z \(0 or 1\)
- 5\.RPSzRPS\_\{z\}\(tm \+ 1\) = At tm \+ 1, updated reputation score of vehicle z
Wtz=RPSz∑m=1nRPSmWt\_\{z\}=\\frac\{RPS\_\{z\}\}\{\\sum\_\{m=1\}^\{n\}RPS\_\{m\}\}\(4\)where:
- 1\.WTzWT\_\{z\}= Weight of vehicle z’s model update\.
- 2\.nn= Total number of vehicles
These scores are input to the reputation system maintained at the cloud end\. This reputation system helps DT\-VANET to assign weight to the model update as described in the previous section\. It will also help to keep the malicious vehicle from interfering\.
ΔRPSz=RPSz\(tm\+1\)−RSz\(tm\)\\Delta RPS\_\{z\}=RPS\_\{z\}\(tm\+1\)\-RS\_\{z\}\(tm\)\(5\)BCz=\{ΔRPSz\(tm1\),ΔRPSz\(tm2\),…,ΔRPSz\(tmn\)\}BC\_\{z\}=\\\{\\Delta RPS\_\{z\}\(tm\_\{1\}\),\\Delta RPS\_\{z\}\(tm\_\{2\}\),\.\.\.,\\Delta RPS\_\{z\}\(tm\_\{n\}\)\\\}where:
- 1\.BCzBC\_\{z\}is the set of reputation changes over the timet1,t2,…tmt\_\{1\},t\_\{2\},\.\.\.t\_\{m\}for vehicle z
The reputation of a vehicle changes over time, and digital twins on the cloud are in charge of figuring out and recording these value changes on the blockchain\.
RWz=f\(RPSz\)wheref′\(RPSz\)\>0RW\_\{z\}=f\(RPS\_\{z\}\)\\quad\\text\{where\}\\quad f^\{\\prime\}\(RPS\_\{z\}\)\>0\(6\)Where:
- 1\.RWzRW\_\{z\}for vehicle z be a function of reputation scoreRPSzRPS\_\{z\}
RWz=γ⋅RPSzRW\_\{z\}=\\gamma\\cdot RPS\_\{z\}Where:
- 1\.γ\\gammais used for changing reputation scoreRPSzRPS\_\{z\}into a specific reward with a scaling factor\.
Vehicles with good scores and positive contributions may be rewarded with tax or toll reductions or urgent access to services\.
RPSz<θ⟹Vehiclezis excluded from participation\.RPS\_\{z\}<\\theta\\implies\\text\{Vehicle \}z\\text\{ is excluded from participation\.\}\(7\)Where:
- 1\.Malicious Vehicle reputation scoreRPSzRPS\_\{z\}for vehicle z is considered dangerous and stopped from participation
A blockchain ensures tamper\-proof records of participation, high\-quality data, and protocol enforcement\.
### 3\.5Cloud Server Hierarchical Federated Transfer Learning
Algorithm 1Weighted Cloud Server Cycling Model Update1:The central cloud server has existing model parameters
WexistW\_\{exist\}and digital twin of the entire system for simulation;
2:Initialize
CnC\_\{n\};
3:foreach update
j=0,1,2,⋯,j−1j=0,1,2,\\cdots,j\-1do
4:
Wexist−W\_\{exist\}^\{\-\}=
WexistW\_\{exist\};
5:Select cluster i from
CnC\_\{n\};
6:Send
Wexist−W\_\{exist\}^\{\-\}to Digital Twin of cluster i ;
7:c =Inner\-Cluster\(
Wexist−W\_\{exist\}^\{\-\}\);
8:Calculate
FweightF\_\{weight\}\(
Wexist\+W\_\{exist\}^\{\+\}\) ;
9:Eliminate i from
CnC\_\{n\};
10:Update
WexistW\_\{exist\};
11:endfor
12:Acquire the final model parameters
WexistW\_\{exist\};
#### 3\.5\.1Initialization
First, the central cloud server has existing model parameters for the Federated Transfer Learning Process and the digital twins for the entire VANET\. It initializes the collection of clusters that have not yet participated in Federated Transfer Learning\.
#### 3\.5\.2Selection of Training Clusters
We have a collection of clustersCnC\_\{n\}that participates in FTL Learning\. If we want to take updates from all the clusters in the environment, thenCnC\_\{n\}= n\. There are j model updates in total\. Each cluster contains a group of specific types of vehicles\. The central cloud server sends the latest model parameters to a base station that forwards it to the particular cluster inCnC\_\{n\}that needs the prediction to be done\. The cluster may be chosen based on the requirement to make some predictions\. Pre\-trained model parameters at the start of the jth update are sent to the base station of a selected cluster’s digital twin inCnC\_\{n\}\. This allows the vehicle in that cluster to carry out inner\-cluster training to derive the model parametersWexist\-W\_\{\\textbf\{exist\}\}^\{\\textbf\{\-\}\}\. The details of the inner\-cluster training will be covered in the following subsection\. It is not observable to the central cloud server\.
#### 3\.5\.3Updates to the Cloud Central Model
We denote the model parameters that the central cloud server sends to and receives from the base station for cluster i in the jth update asWexist\-W\_\{\\textbf\{exist\}\}^\{\\textbf\{\-\}\}andWexist\+W\_\{\\textbf\{exist\}\}^\{\\textbf\{\+\}\}\. We calculate the weight of the received model parameters using the scores we received for each vehicle in the given cluster\.
Fweight\(Wexist\+\)=∑Sr\|v\|F\_\{\\text\{weight\}\}\(W\_\{\\text\{exist\}\}^\{\+\}\)=\\frac\{\\sum Sr\}\{\|v\|\}\(8\)In the above equation[8](https://arxiv.org/html/2608.11532#S3.E8),∑Sr\\sum Sris the sum of all the scores that each vehicle DT calculated in the given cluster, and\|v\|\|v\|is the number representing all the vehicles in each cluster\. If there is only one vehicle in the cluster due to a lack of homogeneous vehicle type in the surroundings, we will consider a cluster of only one vehicle\. This overall provides us with the average score of vehicles in a cluster\. However, If the output weight is below a certain threshold, regarded as the data quality it is trained on is not good enough to update the global model ModelUpdate\(\.\) at the central cloud server\. Cluster i will be removed from theCnC\_\{n\}after deciding whether to update the global model parameters based on the average score of clusters\. After j updates, a complete network model with parametersWexistW\_\{exist\}is determined\. Algorithm[1](https://arxiv.org/html/2608.11532#alg1)indicates that our suggested architecture framework needs j\-cluster iterations of the inner\-cluster algorithm for the central cloud server updates\.
### 3\.6Inter Cluster Hierarchical Federated Transfer Learning
Algorithm 2Inner Cluster Hierarchical Federated Transfer LearningInput:The model parameters of the cluster’s digital twin that need to be fine\-tunedWexist\-W\_\{\\textbf\{exist\}\}^\{\\textbf\{\-\}\}; Output:The updated fine\-tuned trained model parametersWexist\+W\_\{\\textbf\{exist\}\}^\{\\textbf\{\+\}\};
1:Inner\-Cluster\(\):
2:\(I\)\. For base station:
3:Receive
Wexist−W\_\{exist\}^\{\-\}from the central cloud server;
4:Send
Wexist−W\_\{exist\}^\{\-\}to
VhV\_\{h\};
5:Receive
Wexist\+W\_\{exist\}^\{\+\}and Sr from
VhV\_\{h\};
6:Send
Wexist\+W\_\{exist\}^\{\+\}and Sr to the central cloud server;
7:
8:\(II\)\. For head vehicle DTVhV\_\{h\}:
9:Receive
wh−w\_\{h\}^\{\-\}=
Wexist−W\_\{exist\}^\{\-\}from the base station;
10:Note its post arrange vehicles into set
PohPo\_\{h\};
11:forall vehicles
v∈Pohv\\in Po\_\{h\}do
12:Send
wh−w\_\{h\}^\{\-\}to vehicle DT ;
13:Receive
wvw\_\{v\},
DQvDQ\_\{v\}and Sr from each vehicle DT;
14:Update
wh−w\_\{h\}^\{\-\}= combine\(
wh−w\_\{h\}^\{\-\},
wvw\_\{v\},
DtQvDtQ\_\{v\}, Sr\) ;
15:endfor
16:
wh−w\_\{h\}^\{\-\}=Fine\_Tun\(
whw\_\{h\},
DtQvDtQ\_\{v\}\) ;
17:Send
wh\+w\_\{h\}^\{\+\}and Sr to the base station;
18:
19:\(III\)\. For route vehicle DTVrV\_\{r\}:
20:Receive
wr−w\_\{r\}^\{\-\}from its previous vehicle DT ;
21:Note its post arrange vehicles into set
PorPo\_\{r\};
22:forall vehicles
v∈Porv\\in Po\_\{r\}do
23:Send
wr−w\_\{r\}^\{\-\}to vehicle DT ;
24:Receive
wvw\_\{v\},
DQvDQ\_\{v\}and Sr from each vehicle DT;
25:Update
wr−w\_\{r\}^\{\-\}= combine\(
wr−w\_\{r\}^\{\-\},
wvw\_\{v\},
DtQvDtQ\_\{v\}, Sr\) ;
26:endfor
27:
wr−w\_\{r\}^\{\-\}=Fine\_Tun\(
wrw\_\{r\},
DtQvDtQ\_\{v\}\) ;
28:Update
DtQrDtQ\_\{r\}and Sr;
29:Send
wrw\_\{r\},
DtQrDtQ\_\{r\}and Sr to its previous vehicle DT;
30:
31:\(IV\)\. For edge vehicle DTVeV\_\{e\}:
32:Receive
we−w\_\{e\}^\{\-\}from its previous vehicle DT;
33:
we−w\_\{e\}^\{\-\}=Fine\_Tun\(
wew\_\{e\},
DtQeDtQ\_\{e\}\) ;
34:
DtQeDtQ\_\{e\}=
\|Dataev\|\|Data\_\{ev\}\|;
35:Send
wew\_\{e\},
DtQeDtQ\_\{e\}and Sr to its previous vehicle DT;
The Algorithm[2](https://arxiv.org/html/2608.11532#alg2)sequence of collaboration and flowchart across vehicles in a cluster is depicted in Figure[2](https://arxiv.org/html/2608.11532#S3.F2)\.In our proposed framework, vehicles are further divided into three groups in a cluster named edge vehicle DTVeV\_\{e\}, route vehicle DTVrV\_\{r\}, and head vehicle DTVhV\_\{h\}if there is more than one vehicle in a cluster which draws inspiration from earlier work[HaghighiFard and Coleri 2024](https://arxiv.org/html/2608.11532#bib.bib14)that separates vehicles in each cluster into cluster head and cluster members\.
Assumption 1: The Inner Cluster training occurs on any cluster i digital twin that belongs to theCnCncluster’s collection\. We assume that our cluster consists of head, route, and edge vehicles\. However, if any cluster is not formed of the same vehicle collection, the relevant portion of the cluster will be discarded\.
#### 3\.6\.1Base Station
After the head vehicle, DTvhv\_\{h\}, completes its training, it shares the updated model parameters and data quality score with the base station\. The Base Station then shares its model parameters and data quality score with the global model that resides on the cloud server\. The relevant global model will be updated only after the cluster has a good aggregated data quality score\.
#### 3\.6\.2Head Vehicle
When there is more than one vehicle in the cluster, there is always a Head Vehiclevhv\_\{h\}\. No vehicle exists before the head vehiclevhv\_\{h\}in the cluster\. It is directly connected to the base station if there is more than one vehicle\. For HFTL to execute, the base station sends global model parameters tovhv\_\{h\}vehicle DT, which forwards global parameters to the vehicle DT that exists next to this vehicle DT in the cluster\. To represent the collection of all the vehicles that are next tovhv\_\{h\}vehicle, we usePohPo\_\{h\}\. The updated fine\-tuned global parameterswvw\_\{v\}, for constructing model parameters Data QuantityDQvDQ\_\{v\}and Sr is the Data Quality scores are received by thevhv\_\{h\}vehicle DT for integration when the next vehicle DT is finished with their fine\-tuning\.
While integrating, we give weightage to the parameters according to the amount of data on which it is fine\-tuned and their data quality scores\. After this, thevhv\_\{h\}vehicle DT forwards the updated integrated model parameters and scores to the next vehicle’s DT, which remains in the cluster for fine\-tuning\. The above procedure continues until all the remaining vehicle DTs are completed with the fine\-tuning\. Ultimately, the model is fine\-tuned using the local data ofvhv\_\{h\}vehicle DT, and the Data quality score is calculated\. The finalized fine\-tuned model parameters and Data Quality Score are then sent to the base station\.
#### 3\.6\.3Route Vehicle
A vehicle that functions as a routing vehicle DT is represented by the symbolvrv\_\{r\}, which is a type of vehicle that has at least one predecessor vehicle and one successor vehicle\. Thevrv\_\{r\}vehicle DT gets the model parameterswr−w\_\{r\}^\{\-\}fromvhv\_\{h\}the predecessor vehicle DT and sends them to all the successor vehicles DT one at a time\. Likevhv\_\{h\}DT,vrv\_\{r\}DT integrateswr−w\_\{r\}^\{\-\}with model parameters received after fine\-tuning from the successor vehicle and continues with the same process until the successor vehicles in the collection are done with the same\. After that,vrv\_\{r\}DT, using its local data, fine\-tunes the model parameters and calculates the Data Quality Score\. After this,vrv\_\{r\}sent it to the predecessor vehicle\.
#### 3\.6\.4Edge Vehicle
The termvev\_\{e\}describes a vehicle DT that is either the only vehicle in the cluster or has only predecessor vehicles but no successor vehicles\. If it is the only vehicle, it will forward model parameters directly to the base station; otherwise, it will send them to the predecessor vehicles\. Whenvev\_\{e\}DT gets model parameters from its predecessor vehicle’s DTs, it performs fine\-tuning and sends back the model parameters and data quality score\(Sr\) to the predecessor vehicle DT\.
By making clusters of similar vehicle types, we also make a low communication overhead with fewer links to the base station\. We also make sure that our trained model is customized to each type of vehicle separately\.
Figure 3:Twelve random topologies Network X generated
#### 3\.6\.5Example Scenario
We present an example of specific Intelligent Transportation System\(ITS\) applications where Hierarchical Federated Transfer Learning \(HFTL\) in Digital Twin Vehicular Networks could be particularly beneficial:
1. 1\.Traffic Flow Management HFTL in DT\-VANET can enhance traffic flow by aggregating data from sensors of multiple vehicles to accurately predict real\-time traffic congestion patterns\. The digital twin also adjusts the signal timings in real\-time according to traffic conditions, helping to maintain traffic flow\.
2. 2\.Predictive Maintenance By analyzing the patterns of component failures, the system can predict the maintenance needs of the vehicle’s components\(e\.g\., Brake issues, engine performance\)\.
3. 3\.Real\-time Route Optimization The system can provide route recommendations based on the weather, traffic, and driver patterns\. The digital twin can also optimize routes for emergency vehicles\.
4. 4\.Safety Applications The system can also alert drivers of a potential collision by acquiring data from multiple vehicles\. The architecture can also be beneficial for getting weather updates from vehicles and for providing predictions and alerts to upcoming vehicles about the weather\.
5. 5\.Smart Parking, Charging and Environmental Solutions Digital Twins helps vehicles find parking and charging spots in crowded areas\. It also helps identify the polluted regions of the city\.
## 4PERFORMANCE EVALUATION
This section includes a thorough evaluation of our suggested framework using Hierarchical Federated Transfer Learning \(HFTL\) with Clustered Federated Learning \(CFL\), Federated Learning \(FL\), and Centralized Learning \(CL\) by different performance evaluations\. We chose Centralized Learning \(CL\), Federated Learning \(FL\), and Clustered Federated Learning \(CFL\) as our baseline comparisons because CL represents the simplest centralized training method, while FL represents distributed learning\. Conversely, CFL extends federated learning with clusters that resonate with our algorithm\. The above techniques conflict with HFTL regarding data heterogeneity and resource utilization, effectively highlighting the advantages of HFTL in heterogeneous vehicle environments\. The generation of the cluster topology in our experiment is based on the Network X network structure, using a random number of vehicles distributed according to different vehicle types\. There are twelve clusters, each having more than a random ten number of vehicles\. It helps us to analyze results in high\-density traffic scenarios, as shown in Figure[3](https://arxiv.org/html/2608.11532#S3.F3)\. It also helps to evaluate the system’s ability to handle high\-traffic data, which is very helpful for performance metrics results\. Although the number of vehicles in each cluster is random, the clustering is based on vehicle type\. To ensure the realism of the experiment in each generated cluster, only one head vehicle connects with the base station, and the rest of the vehicles are either routing vehicles or edge vehicles, simulating the actual distribution of different vehicle types\. In Figure[3](https://arxiv.org/html/2608.11532#S3.F3), node A represents the head vehicle, and the rest of the node represents vehicles that are either route or edge vehicles\. Below are the specific experimental settings\.
### 4\.1Experiment Settings
#### 4\.1\.1Experimental Setup
Network X111A Network structure in Python \(https://networkx\.org/\)creates the random Inner Cluster topologies while Torch222Deep Learning Library for Python\(https://pytorch\.org/\)implements our suggested federated transfer learning framework\. The Google Colab[Google Colab 2022](https://arxiv.org/html/2608.11532#bib.bib28)is used for implementation\.
#### 4\.1\.2Setting for Cluster
We form twelve random clusters, each with a variable number of vehicles but the same vehicle type as described in table[2](https://arxiv.org/html/2608.11532#S3.T2)\. As mentioned in Figure[2](https://arxiv.org/html/2608.11532#S3.F2), if each cluster has more than one vehicle, there would be a single cluster head; otherwise, in the case of the single vehicle, it would act as an edge vehicle\. In our case for the evaluation of performance metrics, we keep the number of vehicles to more than a random ten number of vehicles because it will help us to evaluate the system’s ability to handle high traffic data as shown in Figure[3](https://arxiv.org/html/2608.11532#S3.F3)
#### 4\.1\.3Datasets
For our performance metrics, we have conducted our experiments on the real\-world dataset named vehicle mobility trace333https://vehicular\-mobilitytrace\.github\.io/index\.html\#data\. The data sets comprise real\-time values, such as vehicle position, angle, vehicle ID, coordinates, speed, etc\. As part of preprocessing, we first perform data cleaning, which handles outliers by replacing them with extreme values and using interpolation techniques to fill in missing data\. After that, we use a standard scalar to normalize the data\. It makes sure that all features are on the same scale\. We split the data set in such a way that 80 % is used as training while 20% as testing\. The main reason behind this data split is to make sure there is enough data for training to retain some portion for evaluation\. However, we also used cross\-validation to verify the model performance across several data subsets to guarantee the reliability and stability of the results\. We assigned vehicles a numerical ID according to their category, as mentioned in Table[2](https://arxiv.org/html/2608.11532#S3.T2)\. This will help us implement Hierarchical Federated Learning by randomly clustering vehicles according to their similar category\.
#### 4\.1\.4Baseline Studies
We will compare centralized learning, federated learning, clustered federated learning, and federated transfer learning\. For comparison, we have used centralized, federated learning, and clustered federated learning as they are most closely related to federated transfer learning\. Furthermore, while implementing the search for the most efficient hyperparameters, we utilized a checkpoint that led us to set the number of epochs to 200, the learning rate\(α\)\(\\alpha\)to 0\.1, and the batch size to 128 for both the proposed Federated transfer learning and the baseline algorithms\.
#### 4\.1\.5Performance Metrics
We computed average Model Accuracy, Training Time, Resource Consumption \(Computational\), Communication Overhead, Latency, Convergence Time, and Throughput\. Model accuracy tells us our model’s accuracy in correctly predicting output\. Training time refers to the duration required to complete the training process\. Resource Consumption refers to the utilization of computational resources during the training process\. Communication Overhead describes the amount of data exchanged during transmission\. The amount of time required with an input and producing output \(prediction\) by the model is known as Latency\. Convergence Time is the time needed for the training procedure to attain the required accuracy or stable condition \(in our case, 90% accuracy\)\. The number of predictions a model can make in an amount of time given is known as Throughput\.
Figure 4:Performance metrics graph results of our proposed Federated Transfer Learning algorithm with Clustered Federated Learning, Federated Learning, and Centralized Learning
#### 4\.1\.6Experimental Results and Performance Analysis
We have compared existing centralized learning, federated learning, and clustered federated learning with federated transfer learning\. We used a real\-time vehicular mobility trace CSV file dataset to make our findings realistic\. Figure[4](https://arxiv.org/html/2608.11532#S4.F4)shows the graph for the performance metrics with 100 simulations to eliminate any bias scenario\. Table[3](https://arxiv.org/html/2608.11532#S4.T3)shows the average performance of our HFTL algorithm over CL and FL\.
Table 3:Experimental results regarding the average performance metrics across learning algorithms with Federated Transfer Learning Algorithm444Model Accur = Model Accuracy \(Percentage %\), Train Time = Training Time \(Sec\), Resour Consu = Resource Consumption \(Percentage %\), Comm Ohead = Communication Overhead \(Messages\), Latncy = Latency \(Sec\), Conver Time = Convergence Time \(sec\), Thrput = Throughput \(Messages\)The Federated transfer learning\-based algorithm has the highest model accuracy average because it uses a pre\-trained model according to the specific vehicle type\. So, its prediction accuracy is better on average than the other two models\.
For the performance metric Training Time, on average, of a hundred different simulations, our algorithm takes less time because fine\-tuning a model requires less time in seconds compared to training a model from scratch\.
The federated transfer learning model requires fewer computational resources for resource consumption than the other models\. That is because fine\-tuning requires, on average, fewer computational resources than training from scratch, especially when we have customized models for each vehicle type\.
The federated transfer learning algorithm has slightly more communication overhead when compared to Centralized learning\(CL\), Federated learning\(FL\), and Clustered Federated Learning \(CFL\)\. First, a pre\-trained model must be fine\-tuned and optimized based on each vehicle’s personalized requirements, which will take up more communication overhead\. Centralized learning\(CL\), Federated learning \(FL\), and Clustered Federated Learning \(CFL\) don’t require pre\-trained models for training, so their communication overhead is lower than our algorithm’s\. Having a slight difference in communication overhead does not affect our algorithm efficiency\. However, the increased communication overhead is offset by significant improvements in accuracy and latency, giving HFTL a clear advantage in real\-time decision\-making and prediction precision, particularly in scenarios with high vehicle heterogeneity\.
The Federated Transfer Learning algorithm has less latency than Centralized learning\(CL\), Federated learning\(FL\), and Clustered Federated Learning \(CFL\)\. This is due to its use of pre\-trained models that require less time to make predictions\. Because of pre\-trained model fine\-tuning, this algorithm also has more throughput, faster convergence, a hierarchical structure, and low resource usage compared to Federated and Centralized learning\.
Figure 5:Graphical Comparison of Different Algorithm’s Model Accuracy, Resource Consumption, Convergence Time, Training Time, Communication Overhead, Latency, ThroughputFigure[5](https://arxiv.org/html/2608.11532#S4.F5)shows the bar graph comparison of HFTL over CFL, FL, and CL for Model Accuracy, Training Time, and Resource Consumption\. HFTL outperforms the other algorithms\. However, Figure[5](https://arxiv.org/html/2608.11532#S4.F5)shows the bar graph comparison of HFTL over CFL, FL, and CL for Communication Overhead, where the other algorithm’s communication overhead is slightly lower than HFTL\.
Figure 6:Scalability of our Federated Transfer Algorithm with Clustered Federated, Federated, and Centralized Learning
#### 4\.1\.7Scalability Experimental Results and Performance Analysis
The above results are tailored to a few vehicles and specific network conditions\. We want to see how our algorithm performs if we gradually increase the number of vehicles in our network to check whether it remains effective\. Figure[6](https://arxiv.org/html/2608.11532#S4.F6)shows the performance of learning algorithms for scalability metrics\. In the Dynamic Environment, as compared to Clustered Federated, Federated, and Centralized Learning, our proposed algorithm performs better except for some underperformance in communication overhead\. Table[4](https://arxiv.org/html/2608.11532#S4.T4)shows the average scalability metrics values for the algorithms\.
Table 4:Experimental results regarding the average scalability metrics across the learning algorithms with Federated Transfer Learning AlgorithmFigure[7](https://arxiv.org/html/2608.11532#S4.F7)shows these algorithms’ differences in scalability performance metrics output as a bar graph\.555Model Accur = Model Accuracy \(Percentage %\), Train Time = Training Time \(Sec\), Resour Consu = Resource Consumption \(Percentage %\), Comm Ohead = Communication Overhead \(Messages\), Latncy = Latency \(Sec\), Conver Time = Convergence Time \(sec\), Thrput = Throughput \(Messages\)
Figure 7:Scalability of our Federated Transfer Algorithm with Federated and Centralized LearningScalability testing in large\-scale VANETs shows that HFTL maintains high model accuracy and fast convergence as the number of nodes increases\. However, as the network grows continuously, communication overhead and resource consumption may become limiting factors according to the results\. Future work could explore communication compression techniques or asynchronous update mechanisms to further optimize HFTL’s performance in large\-scale networks\. As we increase the number of vehicles, the communication overhead increases based on the following:
1. 1\.The model updates that need to be communicated also increase\.
2. 2\.Aggregation complexity also increases communication overhead as vehicle updates are integrated with the head vehicle DT and sent to the central server\.
3. 3\.More Communication Rounds are needed to achieve convergence because the number of vehicles grows, increasing communication overhead\.
As the number of vehicles increases, the communication overhead increases because of more data traffic and aggregation complexity\. A slightly higher communication overhead than the other algorithms will not affect efficiency, as other performance metrics have greatly improved\.
## 5Challenges and Limitations
Although the above results show the significance of our proposed research, some challenges and limitations are associated with it\.
1. 1\.Data heterogeneity: Significant differences in the data distribution between digital twins could result in less accuracy, more training time, and inefficient model convergence\.
2. 2\.Network limitations: Real\-time synchronization and data transfer between vehicles and their respective digital twins can be affected by environments with erratic, unstable, or low\-bandwidth communication links\.
3. 3\.Model Bias: The global model may inherit bias if the local training data is biased
4. 4\.Dynamic Contexts: The hierarchical model may take time to adjust in real\-time where the vehicle patterns change quickly
5. 5\.Heterogeneous VANET Types: Our HFTL works well in urban and highways\. It may struggle in sparely populated regions where data is less available\.
## 6Conclusion
This paper proposes the Hierarchical Federated Transfer Learning architecture and algorithm for the Digital Twin\-based Vehicular Ad Hoc Network that provides faster and more accurate predictions\. We have also formed clusters based on the specific vehicle type to make precise decision\-making\. We first designed the high\-level architecture of Hierarchical Federated Learning\. Then, we describe algorithms for our proposed architecture that are time\-efficient and more accurate\. Our experimental results demonstrate that HFTL outperforms Clustered Federated Learning \(CFL\), Federated Learning \(FL\), and Centralized Learning \(CL\) across several key metrics, including model convergence speed, throughput, and resource consumption\. We also scale our network to see how our technique performs in the dynamic environment\. In the future, we plan to conduct further research in the following areas: First, we aim to explore efficient communication compression techniques to reduce communication overhead\. Second, we investigate asynchronous model update strategies to improve system real\-time performance\. Lastly, we intend to optimize lightweight blockchain implementations to ensure data privacy and security in large\-scale vehicular networks\.
## References
- Noor\-A\-Rahim et al\. \(2020\)M\. Noor\-A\-Rahim, Z\. Liu, H\. Lee, G\. M\. N\. Ali, D\. Pesch, and P\. Xiao, “A survey on resource allocation in vehicular networks,”*IEEE Transactions on Intelligent Transportation Systems*, vol\. 23, no\. 2, pp\. 701–721, 2020\.
- Zia et al\. \(2016\)Q\. Zia, M\. S\. Farooq, and A\. Abid, “Improving response time of vehicular ad hoc networks \(VANET\),” 2016\.
- He et al\. \(2022\)C\. He, T\. H\. Luan, R\. Lu, Z\. Su, and M\. Dong, “Security and privacy in vehicular digital twin networks: Challenges and solutions,”*IEEE Wireless Communications*, 2022\.
- Khan et al\. \(2023\)L\. U\. Khan, E\. Mustafa, J\. Shuja, F\. Rehman, K\. Bilal, Z\. Han, and C\. S\. Hong, “Federated learning for digital twin\-based vehicular networks: Architecture and challenges,”*IEEE Wireless Communications*, 2023\.
- AbdulRahman et al\. \(2020\)S\. AbdulRahman, H\. Tout, H\. Ould\-Slimane, A\. Mourad, C\. Talhi, and M\. Guizani, “A survey on federated learning: The journey from centralized to distributed on\-site learning and beyond,”*IEEE Internet of Things Journal*, vol\. 8, no\. 7, pp\. 5476–5497, 2020\.
- Li et al\. \(2020\)B\. Li, Y\. Wu, J\. Song, R\. Lu, T\. Li, and L\. Zhao, “DeepFed: Federated deep learning for intrusion detection in industrial cyber–physical systems,”*IEEE Transactions on Industrial Informatics*, vol\. 17, no\. 8, pp\. 5615–5624, 2020\.
- Khan et al\. \(2021\)L\. U\. Khan, W\. Saad, Z\. Han, and C\. S\. Hong, “Dispersed federated learning: Vision, taxonomy, and future directions,”*IEEE Wireless Communications*, vol\. 28, no\. 5, pp\. 192–198, 2021\.
- Konečnỳ et al\. \(2016\)J\. Konečnỳ, H\. B\. McMahan, F\. X\. Yu, P\. Richtárik, A\. T\. Suresh, and D\. Bacon, “Federated learning: Strategies for improving communication efficiency,”*arXiv preprint arXiv:1610\.05492*, 2016\.
- Li et al\. \(2022\)B\. Li, Y\. Jiang, Q\. Pei, T\. Li, L\. Liu, and R\. Lu, “FEEL: Federated end\-to\-end learning with non\-IID data for vehicular ad hoc networks,”*IEEE Transactions on Intelligent Transportation Systems*, vol\. 23, no\. 9, pp\. 16 728–16 740, 2022\.
- Sepasgozar and Pierre \(2022\)S\. S\. Sepasgozar and S\. Pierre, “Fed\-NTP: A federated learning algorithm for network traffic prediction in VANET,”*IEEE Access*, vol\. 10, pp\. 119 607–119 616, 2022\.
- Zhang and Zhang \(2006\)L\. Zhang and B\. Zhang, “Hierarchical machine learning—a learning methodology inspired by human intelligence,” in*International Conference on Rough Sets and Knowledge Technology*\. Springer, 2006, pp\. 28–30\.
- Gonçalves et al\. \(2021\)F\. Gonçalves, J\. Macedo, and A\. Santos, “An intelligent hierarchical security framework for VANETs,”*Information*, vol\. 12, no\. 11, p\. 455, 2021\.
- Ahmed et al\. \(2023\)M\. Ahmed, U\. Sardar, S\. Ali, S\. Alam, M\. Patterson, and I\. U\. Khan, “Robust brain age estimation via regression models and MRI\-derived features,”*International Conference on Computational Collective Intelligence*, pp\. 661–674, 2023\.
- HaghighiFard and Coleri \(2024\)M\. S\. HaghighiFard and S\. Coleri, “Hierarchical federated learning in multi\-hop cluster\-based VANETs,”*arXiv preprint arXiv:2401\.10361*, 2024\.
- Liu et al\. \(2020\)Y\. Liu, Y\. Kang, C\. Xing, T\. Chen, and Q\. Yang, “A secure federated transfer learning framework,”*IEEE Intelligent Systems*, vol\. 35, no\. 4, pp\. 70–82, 2020\.
- Otoum et al\. \(2022\)S\. Otoum, N\. Guizani, and H\. Mouftah, “On the feasibility of split learning, transfer learning and federated learning for preserving security in ITS systems,”*IEEE Transactions on Intelligent Transportation Systems*, 2022\.
- Ahmed et al\. \(2026a\)M\. Ahmed, H\. Chai, H\. Wang, H\. Venkateswara, and M\. Patterson, “EpiFormer: Learning antigen–antibody interactions for epitope prediction via geometric deep learning,”*arXiv preprint arXiv:2606\.04154*, 2026\.
- Zia \(2015\)Q\. Zia, “A survey of data\-centric protocols for wireless sensor networks,”*Computer Science Systems Biology, OMICS Publishing Group*, vol\. 8, no\. 3, pp\. 127–131, 2015\.
- Wang and Lin \(2013\)S\.\-S\. Wang and Y\.\-S\. Lin, “PassCAR: A passive clustering aided routing protocol for vehicular ad hoc networks,”*Computer Communications*, vol\. 36, no\. 2, pp\. 170–179, 2013\.
- Rashid et al\. \(2020\)S\. A\. Rashid, L\. Audah, M\. M\. Hamdi, and S\. Alani, “Prediction based efficient multi\-hop clustering approach with adaptive relay node selection for VANET,”*J\. Commun\.*, vol\. 15, no\. 4, pp\. 332–344, 2020\.
- Ahmed et al\. \(2026b\)M\. Ahmed, N\. Taj, I\. U\. Khan, H\. Venkateswara, and M\. Patterson, “ChiMERa\-Bench: A benchmark dataset for epitope\-specific antibody design,”*ICLR 2026 Workshop on Generative and Experimental Perspectives for Biomolecular Design*, 2026\.
- Temurnikar et al\. \(2022\)A\. Temurnikar, P\. Verma, and G\. Dhiman, “A PSO enable multi\-hop clustering algorithm for VANET,”*International Journal of Swarm Intelligence Research \(IJSIR\)*, vol\. 13, no\. 2, pp\. 1–14, 2022\.
- Zia et al\. \(2024\)Q\. Zia, C\. Wang, S\. Zhu, and Y\. Li, “Priority based inter\-twin communication in vehicular digital twin networks,”*International Journal of Parallel, Emergent and Distributed Systems*, pp\. 1–16, 2024\.
- Dai and Zhang \(2022\)Y\. Dai and Y\. Zhang, “Adaptive digital twin for vehicular edge computing and networks,”*Journal of Communications and Information Networks*, vol\. 7, no\. 1, pp\. 48–59, 2022\.
- Ahmed et al\. \(2025\)M\. Ahmed, S\. Ali, A\. Jan, I\. U\. Khan, and M\. Patterson, “Improved graph\-based antibody\-aware epitope prediction with protein language model\-based embeddings,”*International Conference on Computational Advances in Bio and Medical Sciences*, pp\. 290–302, 2025\.
- Khan et al\. \(2022a\)L\. U\. Khan, W\. Saad, D\. Niyato, Z\. Han, and C\. S\. Hong, “Digital\-twin\-enabled 6G: Vision, architectural trends, and future directions,”*IEEE Communications Magazine*, vol\. 60, no\. 1, pp\. 74–80, 2022\.
- Khan et al\. \(2022b\)L\. U\. Khan, Z\. Han, W\. Saad, E\. Hossain, M\. Guizani, and C\. S\. Hong, “Digital twin of wireless systems: Overview, taxonomy, challenges, and opportunities,”*IEEE Communications Surveys & Tutorials*, vol\. 24, no\. 4, pp\. 2230–2254, 2022\.
- Google Colab \(2022\)Google Colab, Jun 2022\. \[Online\]\. Available:[https://colab\.research\.google\.com/](https://colab.research.google.com/)Similar Articles
Federated Foundation Models over Vehicular Networks
This paper presents a vision for integrating multi-modal multi-task federated foundation models (M3T FedFMs) into vehicular networks, discussing training principles, use cases, challenges, and a case study on the Waymo Open Dataset.
FoggyTrust: Robust Federated Learning with Hierarchical Trust Networks
FoggyTrust is a hierarchical extension of FLTrust that localizes trust computation to fog nodes, improving robustness against Byzantine attacks in heterogeneous federated learning settings, achieving over 50% improvement on challenging attacks like Krum and Trim on CIFAR-10.
Towards Effective Federated Multimodal Graph Learning via Navigating Multifaceted Heterogeneity
Proposes FedTCR, the first systematic federated multimodal graph learning algorithm that handles task, modality, and topology heterogeneity via topology-aware cross-modal routing and tri-level contrastive learning, outperforming baselines across 7 domains.
Decentralised Federated Learning over Temporal Networks: The Role of Heterogeneities
This paper analyzes the effect of structural and temporal heterogeneities in decentralized federated learning over temporal networks, showing that ignoring these heterogeneities leads to unrealistically rapid convergence and that real-world networks slow down diffusion.
HADT: A Heterogeneous Multi-Agent Differential Transformer for Autonomous Earth Observation Satellite Cluster
This paper proposes HADT, a transformer-based architecture for autonomous resource management in heterogeneous satellite clusters for Earth observation, using differential attention and relational tokenization. Experiments show significant improvements over baselines and strong adaptability to varying cluster sizes.