
==== Front
BioData Min
BioData Min
BioData Mining
1756-0381
BioMed Central London

39232802
386
10.1186/s13040-024-00386-w
Research
QIGTD: identifying critical genes in the evolution of lung adenocarcinoma with tensor decomposition
Chen Bolin blchen@nwpu.edu.cn

12
Zhang Jinlei 1
Shao Ci 1
Bian Jun 3
Kang Ruiming 4
Shang Xuequn 12
1 https://ror.org/01y0j0j86 grid.440588.5 0000 0001 0307 1240 School of Computer Science, Northwestern Polytechnical University, Xi’an, 710012 China
2 grid.424018.b 0000 0004 0605 0826 Key Laboratory of Big Data Storage and Management, Northwestern Polytechnical University, Ministry of Industry and Information Technology, Xi’an, 710012 China
3 grid.43169.39 0000 0001 0599 1243 Department of General Surgery, Xi’an Children’s Hosptial, Xi’an Jiaotong University Affiliated Children’s Hosptial, Xi’an, 710003 China
4 Rewise (Hangzhou) Information Technology Co., LTD, Hangzhou, 310000 China
4 9 2024
4 9 2024
2024
17 3025 2 2024
28 8 2024
© The Author(s) 2024
2024
https://creativecommons.org/licenses/by-nc-nd/4.0/ Open Access This article is licensed under a Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License, which permits any non-commercial use, sharing, distribution and reproduction in any medium or format, as long as you give appropriate credit to the original author(s) and the source, provide a link to the Creative Commons licence, and indicate if you modified the licensed material. You do not have permission under this licence to share adapted material derived from this article or parts of it. The images or other third party material in this article are included in the article’s Creative Commons licence, unless indicated otherwise in a credit line to the material. If material is not included in the article’s Creative Commons licence and your intended use is not permitted by statutory regulation or exceeds the permitted use, you will need to obtain permission directly from the copyright holder. To view a copy of this licence, visit http://creativecommons.org/licenses/by-nc-nd/4.0/.
Background

Identifying critical genes is important for understanding the pathogenesis of complex diseases. Traditional studies typically comparing the change of biomecules between normal and disease samples or detecting important vertices from a single static biomolecular network, which often overlook the dynamic changes that occur between different disease stages. However, investigating temporal changes in biomolecular networks and identifying critical genes is critical for understanding the occurrence and development of diseases.

Methods

A novel method called Quantifying Importance of Genes with Tensor Decomposition (QIGTD) was proposed in this study. It first constructs a time series network by integrating both the intra and inter temporal network information, which preserving connections between networks at adjacent stages according to the local similarities. A tensor is employed to describe the connections of this time series network, and a 3-order tensor decomposition method was proposed to capture both the topological information of each network snapshot and the time series characteristics of the whole network. QIGTD is also a learning-free and efficient method that can be applied to datasets with a small number of samples.

Results

The effectiveness of QIGTD was evaluated using lung adenocarcinoma (LUAD) datasets and three state-of-the-art methods: T-degree, T-closeness, and T-betweenness were employed as benchmark methods. Numerical experimental results demonstrate that QIGTD outperforms these methods in terms of the indices of both precision and mAP. Notably, out of the top 50 genes, 29 have been verified to be highly related to LUAD according to the DisGeNET Database, and 36 are significantly enriched in LUAD related Gene Ontology (GO) terms, including nuclear division, mitotic nuclear division, chromosome segregation, organelle fission, and mitotic sister chromatid segregation.

Conclusion

In conclusion, QIGTD effectively captures the temporal changes in gene networks and identifies critical genes. It provides a valuable tool for studying temporal dynamics in biological networks and can aid in understanding the underlying mechanisms of diseases such as LUAD.

Keywords

Temporal network
Tensor decomposition
Critical genes
LUAD
Evolution
National Key R&D Program of China2021YFA1000402 National Natural Science Foundation of China61972320 Xi’an municipal bureau of science and technology22YXYJ0057 22YXYJ0057 issue-copyright-statement© BioMed Central Ltd., part of Springer Nature 2024
==== Body
pmcIntroduction

Complex networks are common in real life and the research has shifted from discovering the macro laws of structure and dynamics to uncovering the role of macro elements as nodes in real systems [1–5]. In the past several decades, designing effective centrality methods to identify critical nodes in complex networks has become a significant topic. Many methods have been designed to measure the importance of nodes in static networks, such as degree [6], closeness [7] and betweenness [8, 9]. Those centrality measures have been used in predicting essential proteins [10], identifying influential nodes from social networks [11, 12], finding critical links [13]. With the development and evolution of organisms, the structure of biological networks changes dynamically over time [14–17], by increasing or decreasing the number of nodes or edges [18, 19]. The research on dynamic biological networks and the identification of critical nodes are helpful to better understand biological processes [20].

At present, several methods have paying attention to identify critical genes in biological networks [21–23]. Liu et al. [24] identified potential critical genes related to the pathogenesis and prognosis of gastric cancer by protein-protein interaction (PPI) network and Cox proportional hazards. Li et al. [25] identified critical miRNAs, genes and transcription factors of lung adenocarcinoma by analyzing Gene Ontology terms, pathways, and PPI networks. Liu et al. [26] used the robust rank aggregation method, re-constructed the PPI network and performed modules analysis to identify critical genes. However, most of those methods are based on static network and ignoring the stage heterogeneity of complex diseases. He et al. [20] investigated miRNAs in serum exosome-like microvesicles to identify stage-common and stage-specific miRNAs, but ignored the connections between stages. Kim et al. [27] defined the temporal version of degree, closeness and betweenness on temporal networks, which reduced a dynamic network to a static one with directed flows. Nevertheless, those methods simply calculated the degree, closness and betweenness centrality of nodes in different time snapshots and obtained a mean value. The information of nodes changing with time would be lost. In our previous studies, we have proven that the studies of cancer stages is important for understanding the evolution of cancers [28, 29].Fig. 1 The module of forming the tensor and decomposing tensor. a The way of constructing temporal network into tensor. The edges, consist of inter-stage edges and intra-stage edges are both taken into consideration. The yellow one is the network of Stage I, the green one is Stage II, the blue one is Stage III and the red one is Stage IV. The black represents inter-stage network. b The decomposition of tensor

In this study, a lightweight and effective method that quantify the importance of genes with tensor decomposition (QIGTD) was proposed to identify the critical genes along with the progression of lung adenocarcinoma (LUAD). To start with, a time-series network was constructed to represent the molecular connections of individual pathological stages of LUAD, and a third-order tensor was employed to capture topological information of both intra-stage and inter-stage. The intra-stage topological connections were obtained from gene co-expression relationships, while the inter-stage topological connections were calculated by combining both local similarities and a pre-defined parameters. Then, a tensor decomposition method was proposed to identify critical gene from the temporal network, which considers not only the intra-stage topological information, but also the inter-stage temporal characteristics. It is also a learning free method, which can work well with a small amount of samples. The precision and mAP are presented to evaluate the performance of QIGTD, and the other three state-of-the-art methods: temporal versions of degree, betweenness and closeness [30, 31] were employed as benchmark methods. The overall framework of the proposed method was show in Fig. 1.

Materials and methods

Data collection and processing

The critical genes are identified in the Stage I - Stage IV temporal networks of LUAD. The LUAD related gene expression dataset were downloaded from Xena (https://xenabrowser.net/), where there are 206 samples in Stage I, 93 samples in Stage II, 59 samples in Stage III, and 20 samples in Stage IV.

The networks of the four stages are constructed separately according to the PCC and the obtained p-value. In this study, the selection criteria were p-value<0.01 and |PCC|>0.8 according to the characteristics of biological networks. As a result, there are 17,830 edges in Stage I, 21,951 in Stage II, 11,170 in Stage III and 611 edges in Stage IV.

The known critical genes can be obtained from DisGeNET (https://www.disgenet.org/), where there are 3,899 genes appearing in the temporal network, and 566 of them have been verified to be associated with LUAD.

The temporal network construction

The network representing each stage of LUAD was constructed with pearson correlation coefficient (PCC) calculated with gene expression. Besides the connections within stage, there was also a fixed set of genes connected between different stages of LUAD.

Currently, there are two typical ways to construct connections between networks of adjacent stages. One is to use a fixed constant to represent the interlayer relationship, and the value of the parameter can indicate the strength of the interlayer relationship. The other method is to use similarity metrics to measure inter layer relationships. It was stated that the features in temporal networks are studied by converting time into a snapshot sequence of the network, so the similarity measurement method between nodes in static networks can be extended to the node relationships between adjacent layers.

In this study, a novel method for measuring inter layer node similarity in temporal networks was proposed by combining the calculation of node local similarity index with fixed parameters, which is1 TLSi(t,t+1)=C+∑jwijt+∑jwijt+12N+|SNi(t,t+1)|N

where C is a constant parameter that indicate the strength of the interlayer relationship, N is the number of vertices in the network, if ∑jwijt=1, then the vertex i and vertex j has connection in the network of Gt, while if ∑jwijt=0, then the vertex i and vertex j does not have connection. SNi(t,t+1) represents the number of common vertices in two adjacent network Gt and Gt+1.

The first part of Eq. 1 is a constant parameter, which can be setup according to the experimental requirement. If a relatively small parameter was used, then it enhances the importance of vertices with high inter layer similarity in temporal networks, while selecting larger fixed parameters strengthens the importance of isolated vertices. In this study, the value of C was set to 0.5 based on the characteristics of the biological network. The second part of Eq. 1 is network local similarity, which represents the proportion of local neighbors of adjacent snapshot nodes in the entire network at different times. The third part of Eq. 1 represents the proportion of shared neighbors of adjacent snapshot vertices in the entire network. Hence, the overall value of TLS can characterize the degree of vertices in a temporal network and the inter layer relationship of node adjacency in different time snapshots. The larger the TLS value, the higher the probability of the node continuously appearing on two snapshot layers, and the more stable the node adjacency relationship.

The tensor description of the temporal network

The temporal network was represented as X=Gt,C. The Gt is the network of different stages of LUAD and C is the set of interconnections between different networks. The elements in C were concerned as ‘cross network’. The temporal network could be represented in tensor as follows. Let X∈RI×J×K. The elements can be defined according to Formula 2.2 Xijk=wijkifXijk∈GtcijkifXijk∈C0otherwise

where 0≤i<I,0≤j<J,0≤k<K, the wijk is the element in Gt and cijk is the element in C.

The process to form the tensor from temporal network is shown in Fig. 1a, where different colors represent different stages. The edges between the different stages compose the cross network. There are four kinds of networks in Gt and three kinds of cross networks in C.

The tensor could be transformed into matrix by unfolding or flattening. In this study, we expanded the n-order tensor X along mode-n into a matrix Xn. The mode-1 corresponds to the 1-order of tensor, mode-2 corresponds to the 2-order of tensor and mode-3 corresponds to the 3-order of tensor. After the matricization of the tensor, the Kronecker, Khatri-Rao, and Hadamard products can be calculated respectively as follows.3 A⊙B=a1⊗b1a1⊗b1a2⊗b3⋯aN⊗bN

4 A⊗B=a11B⋯a1NB⋮⋱⋮aN1B⋯aNNB

5 A∗B=a11B11⋯a1NB1N⋮⋱⋮aN1BN1⋯aNNBNN

The Canonical Polyadic (CP) decomposition of tensor

A tensor can be expressed as the sum of finite rank tensors. In this study, the 3-order could be decomposed as follows.6 X≈[A,B,C]=∑r=1Rar∘br∘cr

whereX∈RI×J×KA=(a1,a2,a3,...,aR)∈RI×RB=(b1,b2,b3,...,bR)∈RJ×RC=(c1,c2,c3,...,cR)∈RK×R

The symbol “∘” is the outer product, the vector ar∈RI is column r of factor matrix A∈RI×R, the vector br∈RJ is column r of factor matrix B∈RJ×R, and the vector cr∈RK is column r of factor matrix C∈RK×R.

The outer product of these vectors is a rank one tensor, so the R rank-one tensors was used to approximate the original data, which is shown in Fig. 1b. By utilizing the factor matrix, the 3-order tensor can be decomposed as follows.7 minA∑i,j,kxijk-∑r=1Rairbjrckr2=minAX(1)-A(C⊙B)⊤F2minB∑i,j,kxijk-∑r=1Rairbjrckr2=minBX(2)-B(C⊙A)⊤F2minC∑i,j,kxijk-∑r=1Rairbjrckr2=minBX(2)-B(C⊙A)⊤F2

The formulas can be approximately as8 X(1)≈[A(C⊙B)⊤]X(2)≈[B(C⊙A)⊤]X(3)≈[C(B⊙A)⊤]

Consequently, the A[n] could be calculated with back propagation and gradient descent. Since the goal is to make the tensor X^ estimated by A, B, C as close as possible to the original tensor X, the loss function is set as follows.9 Loss1=12[X(1)-A(C⊙B)⊤]Loss2=12[X(2)-B(C⊙A)⊤]Loss3=12[X(3)-C(B⊙A)⊤]

The partial derivative of A, B and C could be quantified and the parameters could be updated by the following formulas.10 A=A-α∗∂Loss1∂AB=B-α∗∂Loss2∂BC=C-α∗∂Loss3∂C

where α is the learning rate. The vertex centrality is now can be calculated as11 si=1T∑t=1T((a1)i(c1)t+(b1)i(c1)t)

In this study, where I=J represents the number of genes, and K=4 represents the number of layers in the network. The importance score of every gene could be determined with either I=A⊙C or I=B⊙C. Additionally, if R is set to 1, so I=A or I=B.

Results

The evaluation indices

The performance of QIGTD is evaluated by the precision, mean average precision (mAP) and fold enrichment.

The precision show the true positive ratio by giving a list of predictions, which is12 precision=TPTP+FP

The primary objective revolves around the task of ranking, where precision alone may not insufficiently reflect the algorithm’s performance. The mAP does not only consider the accuracy of identifying the critical genes, but also considers the differences in genes order. More robustly, the mAP is utilized to reflect the model’s performance, which can be defined as Formula 13.13 APqi=∑i∈i1,i2,...,iMPi×LiM

where Pi=1∑j=1ij, and the Li is the label of the i-th gene. In this study, the label comprises 0 and 1. Since there is only one query in the problem, so the mAP is equal to the AP.

Comparing with the benchmark methods

The important score of every gene in temporal network were calculated with three benchmark methods: T-degree, T-closeness, T-betweenness and the proposed QIGTD method.

The T-degree is defined as14 T-deg(v)=∑t=1TDt(v)T

where Dt(v) represents the degree of vertex v at the tth network snapshot, and T is the total number of network snapshot.

The T-closeness is defined as15 T-clo(v)=∑1≤t≤T∑u∈V\v1Δt,T(u,v)

where Δt,T(u,v) represents the shortest path length between vertex u and vertex v.

The T-betweenness is defined as16 T-bet(v)=∑1≤t≤T∑s≠v≠d∈Vσ(s,d,v)σ(s,v)

where σ(s,d,v) represents the number of shortest path between vertex s and vertex v that through vertex d, whileσ(s,v) represents the number of all shortest path between vertex s and vertex v.

The precision of different methods are summarized in Table 1. The QIGTD consistently performs better than the other three state-of-the-art methods from the top 10 to top 500 predictions. In top 10, the precision of QIGTD is 0.50, while the best in the other three methods is T-betweenness with the precision of only 0.20. This situation also hold for predictions from the top 50 to top 500. The top three important genes calculated by QIGTD are highly related to LUAD. Table 1 The precision of QIGTD and the three SOTA methods

Rank	T-deg.	T-clo.	T-bet.	QIGTD	
Top 10	0.10	0.00	0.20	0.50	
Top 50	0.22	0.08	0.16	0.58	
Top 100	0.28	0.11	0.19	0.52	
Top 150	0.27	0.15	0.17	0.46	
Top 500	0.21	0.20	0.20	0.23	
The rank is calculated with the four methods. And the precision is obtained by comparing with DisGeNET. For example, the precision of QIGTD is 0.50 in top 10 and it means there are 50% genes appear in DisGeNET among top 10 genes. The precision of QIGTD consistently surpasses that of the other three methods

The results of mAP@M, presented in Table 2, indicate that QIGTD outperforms the other three methods. QIGTD exhibits superior performance in accurately identifying the LUAD related genes without learning. Table 2 The mAP of the four methods

Rank	T-deg.	T-clo.	T-bet.	QIGTD	
mAP@5	0.000	0.000	0.250	0.458	
mAP@10	0.017	0.000	0.120	0.214	
mAP@20	0.008	0.003	0.063	0.125	
mAP@50	0.009	0.003	0.028	0.061	
mAP@100	0.007	0.002	0.016	0.034	
The mAP of QIGTD demonstrates higher performance compared to the other methods

The fold enrichment is carried out to measure the performance of the model, which indicates how precisely the method can locate disease-related genes. QIGTD consistently exhibits significantly higher values compared to the other three methods as shown in Fig. 2.Fig. 2 The curve of fold enrichment in the top 500 genes. The x axis is the rank of genes in every method. The y axis is the score of fold enrichment. The fold enrichment could be calculated with precision and the correlation rate between all genes and LUAD

Biological evidences of the predictions

Table 3 illustrated top 10 genes identified by the four methods as well as their rank. Among the top 10 genes identified by QIGTD, 5 genes are associated with LUAD in DisGenNET, indicating a higher level of association with the disease compared to the other methods.

Additionally, the rest 5 genes have the potential to become biomarkers of LUAD. The NCAPH was verified to be negatively associated with Mcl-1 in non-small cell lung cancer [32]. Nguyen et al. [33] found that CDCA5 (cell division cycle associated 5) upregulated in the majority of lung cancers. The study of Wei et al. [34] found that the knockdown of HJURP inhibits non-small cell lung cancer cell proliferation, migration, and invasion. Coincidentally, there are many researchers found that BUB1 may hopefully become a novel marker and therapeutic target for LUAD [35–37]. The BUB1B was also identified to be a significant biomarker for a poor prognosis and poor clinicopathological outcomes in patients with LUAD [38]. Table 3 The top 10 genes identified by the four methods and verified by DisGeNET

T-deg.	Dis.	T-clo.	Dis.	T-bet.	Dis.	QIGTD	Dis.	
SASH3	0	CD4	0	ZEB2	1	TPX2	1	
NCKAP1L	0	IKZF1	0	GIMAP8	0	NCAPG	1	
CD53	0	ARHGEF6	0	CD53	0	KIF23	1	
LCP2	0	PLEK	0	OLFML1	0	NCAPH	0	
BTK	0	ACAP1	0	TCF4	1	KIF2C	1	
PTPRC	1	C17orf87	0	WAS	0	CDCA5	0	
IL10RA	0	MPEG1	0	STARD8	0	HJURP	0	
CD4	0	FGD2	0	IL16	0	BUB1	0	
PLEK	0	SELPLG	0	FLI1	0	KIF4A	1	
ARHGAP25	0	MS4A4A	0	PLXDC2	0	BUB1B	0	
There is the rank of top 10 genes calculated with the four methods. In the gene columns, genes are arranged in descending order based on the scores obtained in each method. In the disease columns, 1 represents the gene has been verified to be associated with LUAD in DisGeNET. The top 3 ranked with QIGTD have all verified in dataset. And 5 in top 10 genes have verified to be highly related to LUAD according to DisGeNET

The differentially expressed analysis is also performed on the top 10 genes, which is shown in Fig. 3. The blue box is the gene expression in control and the rest 4 boxes are that in four stages. The top 10 genes are obviously differentially expressed in stages compared to control.Fig. 3 The boxplot of the expression of top 10 genes. The different color in the plots is different stages. The blue box represents the expression in control. It shows that the genes identified with QIGTD are differentially expressed genes

Moreover, in top 50 genes identified by QIGTD, 29 genes are verified have a strong association with the disease, which is demonstrated in Table 4. Table 4 The top 50 genes identified by QIGTD and verified in DisGeNET

Gene	Score	Disease	Gene	Score	Disease	
TPX2	0.13189	1	PLK1	0.11190	1	
NCAPG	0.12907	1	CKAP2L	0.11183	1	
KIF23	0.12859	1	TOP2A	0.10984	1	
NCAPH	0.12847	0	NEK2	0.10982	1	
KIF2C	0.12699	1	SPAG5	0.10966	1	
CDCA5	0.12596	0	NDC80	0.10933	0	
HJURP	0.12580	0	KIFC1	0.10894	0	
BUB1	0.12566	0	GTSE1	0.10841	0	
KIF4A	0.12427	1	ZWINT	0.10797	0	
BUB1B	0.12413	0	CDC6	0.10767	0	
CENPA	0.12363	1	KIF18B	0.10750	1	
SGOL1	0.12327	0	RRM2	0.10728	1	
MCM10	0.12241	1	SKA3	0.10700	0	
TTK	0.12233	1	NUF2	0.10657	0	
EXO1	0.12208	0	RAD54L	0.10550	0	
CCNB2	0.12115	0	SKA1	0.10485	0	
CCNA2	0.12071	1	KIF15	0.10482	1	
KIF11	0.11880	0	SPC25	0.10462	1	
DEPDC1	0.11584	1	KIF14	0.10388	1	
DLGAP5	0.11541	0	FAM72B	0.10358	0	
NUSAP1	0.11408	0	CDK1	0.10251	1	
CDCA8	0.11397	0	FOXM1	0.10137	1	
PRC1	0.11379	1	ESPL1	0.10133	1	
CEP55	0.11346	1	ASPM	0.10064	1	
CDC20	0.11205	1	PLK4	0.10004	1	
The scores in the table represent the significance of each gene under the QIGTD method, where higher scores indicate a higher ranking. Among the top 50 genes, 29 genes are associated with LUAD

The sub-networks of top 10 genes are extracted as Fig. 4a. The figure shows the subgraphs of top 10 genes of Stage I to Stage IV respectively. The thickness of the edge in the figure indicates the weight of the edge. The thick edges gradually decrease from Stage I to Stage IV, and some edges also disappear at Stage IV, thus the subgraphs of top 10 genes exhibit the evolution of LUAD.

The sub-networks of top 50 genes are extracted as Fig. 4b. The red nodes are top 10 genes, the purple are top 20 genes and the green are top 50. The thickness of the edges is not obviously as there are too many edges in the networks, but the sub-networks gradually become sparse with the stages, which shows a signal of the evolution of LUAD.Fig. 4 The sub networks of Top 10 genes and Top 50 genes

Among top 50 genes, 36 genes are enriched to 5 GO terms in Fig. 5. The GO terms are nuclear division, mitotic nuclear division, chromosome segregation, organelle fission and mitotic sister chromatid segregation, all of which have been verified to be associated with LUAD [39–41]. The different color of the ribbon represents the different GO terms. The numbers of ribbon means the number of GO terms that genes enrich. For example, the DLGAP5 has five ribbons, which means it enriches all 5 GO terms. The SPC25 has a green ribbon, which means it only enriches the chromosome segregation.Fig. 5 The GO enrichment of top 50 genes. 5 Go terms chosen from the result of the enrichment are exhibited. The ribbons in different colors represents different GO terms. The number of the ribbons in gene is the number of GO terms it enriches. The boxed genes are LUAD-related genes. There are 32 genes enriched on nuclear division, 29 enriched on mitotic nuclear division, 30 on chromosome segregation, 32 on organelle fission and 25 enriched on mitotic sister chromatid segregation

Discussion

The investigation of critical genes in temporal networks has become increasingly prevalent. Most of previous studies concentrate on the the structure of the network itself, but ignore the connections and changes between network at adjacent stages. Inspired by tensor decomposition, QIGTD is proposed in this research. Both the connections of genes inter and intra are taken into consideration.

The experimental results show that QIGTD outperforms the other three SOTA methods, especially in identifying the most critical genes. In the result, 5 genes in top 10 identified by QIGTD have been verified to be critical. At the same time, the other five genes may also be critical according to recent researches. The top 10 genes also differentially expression in stages compared to control. Furthermore, 29 genes are highly related to LUAD in top 50. The GO terms show indicate the top 50 genes ranked by QIGTD is associated with LUAD. The sub network of top 10 to top 50 undergoes changes across stages, which means the genes identified are potential to be biomarkers of the evolution of LUAD.

Additionally, QIGTD is a learning free and effective method, which does not require too many samples. The QIGTD has a low computational complexity and can be utilized in large-scale networks, which also could be easily embedded into the research of other complex problems.

Acknowledgements

Thanks to all members of the laboratory for their valuable discussion and comments.

Authors' contributions

B.C. initialized this study. J.Z. conducted the numerical experiments and drafted the manuscript. C.S. gave suggestions many times, also gave idea of some part of model to J.Z.. Everyone read the manuscript and revised it, and agreed with the final version.

Funding

This work was supported by the National Key R&D Program of China under Grant No. 2021YFA1000402, the National Natural Science Foundation of China under Grant No. 61972320, and Xi’an municipal bureau of science and technology under Grant No. 22YXYJ0057.

Availability of data and materials

The publicly dataset of LUAD could be downloaded from Xena: https://xenabrowser.net/. The gene-disease association information could be collected via DisGeNET: https://www.disgenet.org/.

Declarations

Competing interests

The authors declare no competing interests.

Publisher's Note

Springer Nature remains neutral with regard to jurisdictional claims in published maps and institutional affiliations.
==== Refs
References

1. Boccaletti S Bianconi G Criado R Del Genio CI Gómez-Gardenes J Romance M The structure and dynamics of multilayer networks Phys Rep. 2014 544 1 1 122 10.1016/j.physrep.2014.07.001 32834429
Boccaletti S, Bianconi G, Criado R, Del Genio CI, Gómez-Gardenes J, Romance M, et al. The structure and dynamics of multilayer networks. Phys Rep. 2014;544(1):1–122.32834429 10.1016/j.physrep.2014.07.001
2. Jin S Li Y Pan R Zou X Characterizing and controlling the inflammatory network during influenza A virus infection Sci Rep. 2014 4 1 1 14 10.1038/srep03799
Jin S, Li Y, Pan R, Zou X. Characterizing and controlling the inflammatory network during influenza A virus infection. Sci Rep. 2014;4(1):1–14.10.1038/srep03799
3. Li Y Jin S Lei L Pan Z Zou X Deciphering deterioration mechanisms of complex diseases based on the construction of dynamic networks and systems analysis Sci Rep. 2015 5 1 1 11
Li Y, Jin S, Lei L, Pan Z, Zou X. Deciphering deterioration mechanisms of complex diseases based on the construction of dynamic networks and systems analysis. Sci Rep. 2015;5(1):1–11.
4. Morone F Makse HA Influence maximization in complex networks through optimal percolation Nature. 2015 524 7563 65 68 10.1038/nature14604 26131931
Morone F, Makse HA. Influence maximization in complex networks through optimal percolation. Nature. 2015;524(7563):65–8.26131931 10.1038/nature14604
5. Lü L Chen D Ren XL Zhang QM Zhang YC Zhou T Vital nodes identification in complex networks Phys Rep. 2016 650 1 63 10.1016/j.physrep.2016.06.007
Lü L, Chen D, Ren XL, Zhang QM, Zhang YC, Zhou T. Vital nodes identification in complex networks. Phys Rep. 2016;650:1–63.10.1016/j.physrep.2016.06.007
6. Bonacich P Factoring and weighting approaches to status scores and clique identification J Math Sociol. 1972 2 1 113 120 10.1080/0022250X.1972.9989806
Bonacich P. Factoring and weighting approaches to status scores and clique identification. J Math Sociol. 1972;2(1):113–20.10.1080/0022250X.1972.9989806
7. Freeman LC Centrality in social networks conceptual clarification Soc Networks. 1978 1 3 215 239 10.1016/0378-8733(78)90021-7
Freeman LC. Centrality in social networks conceptual clarification. Soc Networks. 1978;1(3):215–39.10.1016/0378-8733(78)90021-7
8. Zhang J, Luo Y. Degree centrality, betweenness centrality, and closeness centrality in social network. In: Proceedings of the 2017 2nd International Conference on Modelling, Simulation and Applied Mathematics (MSAM2017), vol. 132. Atlantis press. 2017. pp. 300–303.
9. Freeman LC. A set of measures of centrality based on betweenness. Sociometry. 1977;40:35–41.
10. Tang X Wang J Zhong J Pan Y Predicting essential proteins based on weighted degree centrality IEEE/ACM Trans Comput Biol Bioinforma. 2013 11 2 407 418 10.1109/TCBB.2013.2295318
Tang X, Wang J, Zhong J, Pan Y. Predicting essential proteins based on weighted degree centrality. IEEE/ACM Trans Comput Biol Bioinforma. 2013;11(2):407–18.10.1109/TCBB.2013.2295318
11. Srinivas A, Velusamy RL, Identification of influential nodes from social networks based on Enhanced Degree Centrality Measure. In: 2015 IEEE international advance computing conference (IACC). IEEE; 2015. pp. 1179–84.
12. Okamoto K, Chen W, Li XY. Ranking of closeness centrality for large-scale social networks. In: International workshop on frontiers in algorithmics. Springer; 2008. pp. 186–195.
13. Veremyev A Prokopyev OA Pasiliao EL Finding critical links for closeness centrality INFORMS J Comput. 2019 31 2 367 389 10.1287/ijoc.2018.0829
Veremyev A, Prokopyev OA, Pasiliao EL. Finding critical links for closeness centrality. INFORMS J Comput. 2019;31(2):367–89.10.1287/ijoc.2018.0829
14. Hintze A Adami C Evolution of complex modular biological networks PLoS Comput Biol. 2008 4 2 e23 10.1371/journal.pcbi.0040023 18266463
Hintze A, Adami C. Evolution of complex modular biological networks. PLoS Comput Biol. 2008;4(2):e23.18266463 10.1371/journal.pcbi.0040023
15. Tenazinha N Vinga S A survey on methods for modeling and analyzing integrated biological networks IEEE/ACM Trans Comput Biol Bioinforma. 2010 8 4 943 958 10.1109/TCBB.2010.117
Tenazinha N, Vinga S. A survey on methods for modeling and analyzing integrated biological networks. IEEE/ACM Trans Comput Biol Bioinforma. 2010;8(4):943–58.10.1109/TCBB.2010.117
16. Pavlopoulos GA Secrier M Moschopoulos CN Soldatos TG Kossida S Aerts J Using graph theory to analyze biological networks BioData Min. 2011 4 1 1 27 10.1186/1756-0381-4-10 21232136
Pavlopoulos GA, Secrier M, Moschopoulos CN, Soldatos TG, Kossida S, Aerts J, et al. Using graph theory to analyze biological networks. BioData Min. 2011;4(1):1–27.21232136 10.1186/1756-0381-4-10
17. Alon U Network motifs: theory and experimental approaches Nat Rev Genet. 2007 8 6 450 461 10.1038/nrg2102 17510665
Alon U. Network motifs: theory and experimental approaches. Nat Rev Genet. 2007;8(6):450–61.17510665 10.1038/nrg2102
18. Kramer MA Kolaczyk ED Kirsch HE Emergent network topology at seizure onset in humans Epilepsy Res. 2008 79 2–3 173 186 10.1016/j.eplepsyres.2008.02.002 18359200
Kramer MA, Kolaczyk ED, Kirsch HE. Emergent network topology at seizure onset in humans. Epilepsy Res. 2008;79(2–3):173–86.18359200 10.1016/j.eplepsyres.2008.02.002
19. Gao W Gilmore JH Giovanello KS Smith JK Shen D Zhu H Temporal and spatial evolution of brain network topology during the first two years of life PLoS ONE. 2011 6 9 e25278 10.1371/journal.pone.0025278 21966479
Gao W, Gilmore JH, Giovanello KS, Smith JK, Shen D, Zhu H, et al. Temporal and spatial evolution of brain network topology during the first two years of life. PLoS ONE. 2011;6(9):e25278.21966479 10.1371/journal.pone.0025278
20. Yang Q, He S, Huang L, Shao C, Nie T, Xia L, et al. Serum Exosomal miRNAs as Biomarkers of Early Diagnosis and Progression in Parkinson’s Disease. Transl Neurodegener. 2021;10:25.
21. Liu X Hong Z Liu J Lin Y Rodríguez-Patón A Zou Q Computational methods for identifying the critical nodes in biological networks Brief Bioinform. 2020 21 2 486 497 10.1093/bib/bbz011 30753282
Liu X, Hong Z, Liu J, Lin Y, Rodríguez-Patón A, Zou Q, et al. Computational methods for identifying the critical nodes in biological networks. Brief Bioinform. 2020;21(2):486–97.30753282 10.1093/bib/bbz011
22. Abedi M Gheisari Y Nodes with high centrality in protein interaction networks are responsible for driving signaling pathways in diabetic nephropathy PeerJ. 2015 3 e1284 10.7717/peerj.1284 26557424
Abedi M, Gheisari Y. Nodes with high centrality in protein interaction networks are responsible for driving signaling pathways in diabetic nephropathy. PeerJ. 2015;3:e1284.26557424 10.7717/peerj.1284
23. Rezaei J Zare Mirakabad F Marashi SA MirHassani SA The assessment of essential genes in the stability of PPI networks using critical node detection problem AUT J Math Comput. 2022 3 1 59 76
Rezaei J, Zare Mirakabad F, Marashi SA, MirHassani SA. The assessment of essential genes in the stability of PPI networks using critical node detection problem. AUT J Math Comput. 2022;3(1):59–76.
24. Liu X Wu J Zhang D Bing Z Tian J Ni M Identification of potential key genes associated with the pathogenesis and prognosis of gastric cancer based on integrated bioinformatics analysis Front Genet. 2018 9 265 10.3389/fgene.2018.00265 30065754
Liu X, Wu J, Zhang D, Bing Z, Tian J, Ni M, et al. Identification of potential key genes associated with the pathogenesis and prognosis of gastric cancer based on integrated bioinformatics analysis. Front Genet. 2018;9:265.30065754 10.3389/fgene.2018.00265
25. Li J Li Z Zhao S Song Y Si L Wang X Identification key genes, key miRNAs and key transcription factors of lung adenocarcinoma J Thorac Dis. 2020 12 5 1917 10.21037/jtd-19-4168 32642095
Li J, Li Z, Zhao S, Song Y, Si L, Wang X. Identification key genes, key miRNAs and key transcription factors of lung adenocarcinoma. J Thorac Dis. 2020;12(5):1917.32642095 10.21037/jtd-19-4168
26. Liu L He C Zhou Q Wang G Lv Z Liu J Identification of key genes and pathways of thyroid cancer by integrated bioinformatics analysis J Cell Physiol. 2019 234 12 23647 23657 10.1002/jcp.28932 31169306
Liu L, He C, Zhou Q, Wang G, Lv Z, Liu J. Identification of key genes and pathways of thyroid cancer by integrated bioinformatics analysis. J Cell Physiol. 2019;234(12):23647–57.31169306 10.1002/jcp.28932
27. Kim H Anderson R Temporal node centrality in complex networks Phys Rev E. 2012 85 2 026107 10.1103/PhysRevE.85.026107
Kim H, Anderson R. Temporal node centrality in complex networks. Phys Rev E. 2012;85(2):026107.10.1103/PhysRevE.85.026107
28. Chen B, Wang Y, Zhang J, Han Y, Benhammouda H, Bian J, et al. Specific feature recognition on group specific networks (SFR-GSN): a biomarker identification model for cancer stages. Front Genet. 2024;15:1407072.
29. Chen B, Chakrobortty N, Saha AK, Shang X. Identifying colon cancer stage related genes and their cellular pathways. Front Gen. 2023;14:1120185.
30. Tang J, Musolesi M, Mascolo C, Latora V, Nicosia V. Analysing information flows and key mediators through temporal centrality metrics. In: Proceedings of the 3rd Workshop on Social Network Systems. New York, Paris: Association for Computing Machinery; 2010. pp. 1–6. 10.1145/1852658.1852661.
31. Tsalouchidou I Baeza-Yates R Bonchi F Liao K Sellis T Temporal betweenness centrality in dynamic graphs Int J Data Sci Anal. 2020 9 3 257 272 10.1007/s41060-019-00189-x
Tsalouchidou I, Baeza-Yates R, Bonchi F, Liao K, Sellis T. Temporal betweenness centrality in dynamic graphs. Int J Data Sci Anal. 2020;9(3):257–72.10.1007/s41060-019-00189-x
32. Xiong YC Wang J Cheng Y Zhang XY Ye XQ Overexpression of MYBL2 promotes proliferation and migration of non-small-cell lung cancer via upregulating NCAPH Mol Cell Biochem. 2020 468 1 185 193 10.1007/s11010-020-03721-x 32200471
Xiong YC, Wang J, Cheng Y, Zhang XY, Ye XQ. Overexpression of MYBL2 promotes proliferation and migration of non-small-cell lung cancer via upregulating NCAPH. Mol Cell Biochem. 2020;468(1):185–93.32200471 10.1007/s11010-020-03721-x
33. Nguyen MH Koinuma J Ueda K Ito T Tsuchiya E Nakamura Y Phosphorylation and activation of cell division cycle associated 5 by mitogen-activated protein kinase play a crucial role in human lung carcinogenesis Cancer Res. 2010 70 13 5337 5347 10.1158/0008-5472.CAN-09-4372 20551060
Nguyen MH, Koinuma J, Ueda K, Ito T, Tsuchiya E, Nakamura Y, et al. Phosphorylation and activation of cell division cycle associated 5 by mitogen-activated protein kinase play a crucial role in human lung carcinogenesis. Cancer Res. 2010;70(13):5337–47.20551060 10.1158/0008-5472.CAN-09-4372
34. Wei Y, Ouyang G, Yao W, Zhu Y, Li X, Huang L, et al. Knockdown of HJURP inhibits non-small cell lung cancer cell proliferation, migration, and invasion by repressing Wnt/-catenin signaling. Eur Rev Med Pharmacol Sci. 2019;23(9):3847–56.
35. Wang L Li S Wang Y Tang Z Liu C Jiao W Identification of differentially expressed protein-coding genes in lung adenocarcinomas Exp Ther Med. 2020 19 2 1103 1111 32010276
Wang L, Li S, Wang Y, Tang Z, Liu C, Jiao W, et al. Identification of differentially expressed protein-coding genes in lung adenocarcinomas. Exp Ther Med. 2020;19(2):1103–11.32010276
36. Jeganathan K Malureanu L Baker DJ Abraham SC Van Deursen JM Bub1 mediates cell death in response to chromosome missegregation and acts to suppress spontaneous tumorigenesis J Cell Biol. 2007 179 2 255 267 10.1083/jcb.200706015 17938250
Jeganathan K, Malureanu L, Baker DJ, Abraham SC, Van Deursen JM. Bub1 mediates cell death in response to chromosome missegregation and acts to suppress spontaneous tumorigenesis. J Cell Biol. 2007;179(2):255–67.17938250 10.1083/jcb.200706015
37. Zhou X, Yuan Y, Kuang H, Tang B, Zhang H, Zhang M. BUB1B (BUB1 Mitotic Checkpoint Serine/Threonine Kinase B) Promotes Lung Adenocarcinoma by Interacting with Zinc Finger Protein ZNF143 and Regulating Glycolysis. Bioengineered. 2022;13(2):2471–85.
38. Chen J Liao Y Fan X Prognostic and clinicopathological value of BUB1B expression in patients with lung adenocarcinoma: a meta-analysis Expert Rev Anticancer Ther. 2021 21 7 795 803 10.1080/14737140.2021.1908132 33764838
Chen J, Liao Y, Fan X. Prognostic and clinicopathological value of BUB1B expression in patients with lung adenocarcinoma: a meta-analysis. Expert Rev Anticancer Ther. 2021;21(7):795–803.33764838 10.1080/14737140.2021.1908132
39. Sun ZY Wang W Gao H Chen QF Potential therapeutic targets of the nuclear division cycle 80 (NDC80) complexes genes in lung adenocarcinoma J Cancer. 2020 11 10 2921 10.7150/jca.41834 32226507
Sun ZY, Wang W, Gao H, Chen QF. Potential therapeutic targets of the nuclear division cycle 80 (NDC80) complexes genes in lung adenocarcinoma. J Cancer. 2020;11(10):2921.32226507 10.7150/jca.41834
40. Rao CV Yamada HY Yao Y Dai W Enhanced genomic instabilities caused by deregulated microtubule dynamics and chromosome segregation: a perspective from genetic studies in mice Carcinogenesis. 2009 30 9 1469 1474 10.1093/carcin/bgp081 19372138
Rao CV, Yamada HY, Yao Y, Dai W. Enhanced genomic instabilities caused by deregulated microtubule dynamics and chromosome segregation: a perspective from genetic studies in mice. Carcinogenesis. 2009;30(9):1469–74.19372138 10.1093/carcin/bgp081
41. Chiang YY Chen SL Hsiao YT Huang CH Lin TY Chiang IP Nuclear expression of dynamin-related protein 1 in lung adenocarcinomas Mod Pathol. 2009 22 9 1139 1150 10.1038/modpathol.2009.83 19525928
Chiang YY, Chen SL, Hsiao YT, Huang CH, Lin TY, Chiang IP, et al. Nuclear expression of dynamin-related protein 1 in lung adenocarcinomas. Mod Pathol. 2009;22(9):1139–50.19525928 10.1038/modpathol.2009.83
