
==== Front
Eur J Hum Genet
Eur J Hum Genet
European Journal of Human Genetics
1018-4813
1476-5438
Springer International Publishing Cham

35953519
1161
10.1038/s41431-022-01161-3
Article
Identification of novel genes whose expression in adipose tissue affects body fat mass and distribution: an RNA-Seq and Mendelian Randomization study
http://orcid.org/0000-0002-9966-6819
Konigorski Stefan stefan.konigorski@hpi.de

123
Janke Jürgen 2
Patone Giannino 4
Bergmann Manuela M. 5
http://orcid.org/0000-0001-6363-2556
Lippert Christoph 13
Hübner Norbert 467
Kaaks Rudolf 8
Boeing Heiner 5
http://orcid.org/0000-0003-1568-767X
Pischon Tobias tobias.pischon@mdc-berlin.de

267910
1 grid.11348.3f 0000 0001 0942 1117 Digital Health & Machine Learning Research Group, Hasso Plattner Institute for Digital Engineering, University of Potsdam, Potsdam, Germany
2 https://ror.org/04p5ggc03 grid.419491.0 0000 0001 1014 0849 Molecular Epidemiology Research Group, Max Delbrück Center (MDC) for Molecular Medicine in the Helmholtz Association, Berlin, Germany
3 https://ror.org/04a9tmd77 grid.59734.3c 0000 0001 0670 2351 Hasso Plattner Institute for Digital Health at Mount Sinai, Icahn School of Medicine at Mount Sinai, New York, NY USA
4 https://ror.org/04p5ggc03 grid.419491.0 0000 0001 1014 0849 Genetics and Genomics of Cardiovascular Diseases Research Group, Max Delbrück Center (MDC) for Molecular Medicine in the Helmholtz Association, Berlin, Germany
5 https://ror.org/05xdczy51 grid.418213.d 0000 0004 0390 0098 German Institute of Human Nutrition Potsdam–Rehbrücke (DIfE), Nuthetal, Germany
6 grid.452396.f 0000 0004 5937 5237 DZHK (German Center for Cardiovascular Research) partner site Berlin, Berlin, Germany
7 https://ror.org/001w7jn25 grid.6363.0 0000 0001 2218 4662 Charité Universitätsmedizin Berlin, Berlin, Germany
8 https://ror.org/04cdgtt98 grid.7497.d 0000 0004 0492 0584 Department of Cancer Epidemiology, German Cancer Research Center (DKFZ), Heidelberg, Germany
9 https://ror.org/04p5ggc03 grid.419491.0 0000 0001 1014 0849 MDC Biobank, Max Delbrück Center for Molecular Medicine (MDC) in the Helmholtz Association, Berlin, Germany
10 grid.484013.a 0000 0004 6879 971X BIH Biobank, Berlin Institute of Health, Berlin, Germany
11 8 2022
11 8 2022
9 2024
32 9 11271135
9 2 2022
26 6 2022
19 7 2022
© The Author(s) 2022
2022
https://creativecommons.org/licenses/by/4.0/ Open Access This article is licensed under a Creative Commons Attribution 4.0 International License, which permits use, sharing, adaptation, distribution and reproduction in any medium or format, as long as you give appropriate credit to the original author(s) and the source, provide a link to the Creative Commons license, and indicate if changes were made. The images or other third party material in this article are included in the article’s Creative Commons license, unless indicated otherwise in a credit line to the material. If material is not included in the article’s Creative Commons license and your intended use is not permitted by statutory regulation or exceeds the permitted use, you will need to obtain permission directly from the copyright holder. To view a copy of this license, visit http://creativecommons.org/licenses/by/4.0/.
Many studies have shown that abdominal adiposity is more strongly related to health risks than peripheral adiposity. However, the underlying pathways are still poorly understood. In this cross-sectional study using data from RNA-sequencing experiments and whole-body MRI scans of 200 participants in the EPIC-Potsdam cohort, our aim was to identify novel genes whose gene expression in subcutaneous adipose tissue has an effect on body fat mass (BFM) and body fat distribution (BFD). The analysis identified 625 genes associated with adiposity, of which 531 encode a known protein and 487 are novel candidate genes for obesity. Enrichment analyses indicated that BFM-associated genes were characterized by their higher than expected involvement in cellular, regulatory and immune system processes, and BFD-associated genes by their involvement in cellular, metabolic, and regulatory processes. Mendelian Randomization analyses suggested that the gene expression of 69 genes was causally related to BFM and BFD. Six genes were replicated in UK Biobank. In this study, we identified novel genes for BFM and BFD that are BFM- and BFD-specific, involved in different molecular processes, and whose up-/downregulated gene expression may causally contribute to obesity.

Subject terms

Gene expression
High-throughput screening
Genetic association study
issue-copyright-statement© European Society of Human Genetics 2024
==== Body
pmcIntroduction

Obesity is an established risk factor for many chronic diseases and for premature death. Numerous studies have shown that abdominal adiposity is more strongly related to health risks than peripheral adiposity [1]. In line with this observation is evidence that visceral adipose tissue (VAT, which is the major compartment that determines abdominal adiposity) is metabolically more active than subcutaneous adipose tissue (SAT, which is the major determinant of peripheral adiposity). SAT might even have protective effects [2].

Previous authors estimated a heritability between 40% and 90% for obesity, expressed as body mass index (BMI) or absolute fat mass [3, 4], and slightly lower heritability between 15% and 60% for body fat distribution, expressed as waist to hip ratio (WHR) or various other ratios of fat mass in different body compartments [5–7]. These estimates have partly come out of large genetic association studies with obesity traits, which have identified many associated genetic loci, again with a higher number of identified loci associated with fat mass compared to fat distribution [7–10]. However, for many of those loci and genes, it is unclear how they affect obesity and particularly body fat distribution. In addition, their functional attributes are poorly understood. There are a few studies in humans that have investigated the association of genes with obesity on the transcriptomic or proteomic level [11–20]. Campbell et al. [11], Armenise et al. [15], Day et al. [16] and Kerr et al. [17] investigated the association between weight loss and gene expression in adipose tissue as well as in whole blood. The most often implicated pathways were inflammatory pathways [12, 17–20] and lipid as well as glucose pathways [12, 17, 19], with evidence that they are upregulated and changed in obesity. These changes might in turn be part of the role of obesity in cancer [11, 12]. Other studies based on RNA-sequencing of subcutaneous adipose tissue focused on identifying eQTLs which was followed-up by colocalization analyses with GWAS hits for cardiometabolic traits [14] or investigated cellular heterogeneity of gene expression in adipose tissue [13]. Little is known about whether gene expression in SAT affects body fat distribution, and if different genes are implicated in body fat mass and body fat distribution. Such knowledge may help to gain information about biological processes contributing to adiposity and body fat distribution as major determinants of health risks.

In this study, our aim was therefore to identify and to characterize novel genes whose gene expression in SAT is associated with obesity traits of body fat mass and body fat distribution. We performed cross-sectional analyses of ribonucleic acid (RNA)-sequencing experiments from abdominal SAT biopsies and whole-body magnetic resonance imaging (MRI) scans on 200 participants in the EPIC Potsdam study. SAT mass and the ratio of SAT and total adipose tissue (TAT) were obtained from whole-body MRI scans as measures of fat mass and fat distribution. In the analysis, we first investigated the association of gene expression with SAT and SAT/TAT. For the association tests, we used a recently developed method based on joint copula models [21] to improve power of association tests with multiple phenotypes. We followed-up the results with a gene ontology term enrichment analysis which indicated that SAT-associated genes were characterized by their higher than expected involvement in cellular, regulatory and immune system processes, and SAT/TAT-associated genes by their involvement in cellular, metabolic, and regulatory processes. Mendelian Randomization (MR) analyses confirmed that these novel genes are specific for body fat mass or body fat distribution, i.e. implicating different molecular processes, and suggested that the up-regulation or downregulation of the gene expression may causally contribute to obesity. Finally, we replicated the results using UK Biobank data, where we imputed AT gene expression based on exome sequencing data and weights learned in the analysis of the EPIC Potsdam study.

Methods

Study population

This study was conducted in a sub–cohort of EPIC Potsdam within the large European Prospective Investigation into Cancer and Nutrition (EPIC) study [22]. EPIC Potsdam is an ongoing cohort study among 27,548 persons aged 35–65 at recruitment between 1994–1998 from the general population of the city of Potsdam and surrounding area in Germany [22]. From 2010 to 2013, a random sample of 1472 participants was re–invited to the study center of whom 816 agreed to participate [23, 24].

MRI scans were obtained to assess body compartments from 594 participants on a separate visit [25]. Based on automated segmentation algorithms of the MRI scans [26, 27], for the analysis in this manuscript, SAT mass (fat mass in subcutaneous adipose tissue) was extracted as a measure of absolute fat mass, and the ratio of SAT and total adipose tissue (TAT) mass, SAT/TAT, as a measure of body fat distribution [28].

Subcutaneous adipose tissue biopsies were taken from 278 participants with sufficient material extracted from 200 participants [28]. The total RNA was extracted for RNA-sequencing (RNA-Seq). Single nucleotide variants (SNVs) were called from the RNA-Seq data for the 200 participants. For 160 of the participants, MRI measurements were available, which therefore constituted the sample for this study. In comparison to the full EPIC–Potsdam cohort as well as to the 816 participants of the substudy, these 160 probands were very similar regarding their age and sex distribution, disease prevalence, and anthropometric measures (data not shown). Sex was set to equal the assessed gender.

For a replication analysis of the associated genes, we used UK Biobank data (www.ukbiobank.ac.uk). The UK Biobank is a prospective cohort study encompassing data of about 500,000 participants (40–69 years of age at baseline) from Great Britain [29], including whole-exome sequencing data of about 49,960 participants at the time of analysis [30] and MRI scans of about 10,000 participants [31]. In the replication, we analyzed the subset of unrelated (i.e. excluding one person of each pair with greater than 3rd-degree relatedness) white British participants (based on self-report and their genetic principal components), which yielded a sample size of n = 4904.

Assessment of gene expression in subcutaneous adipose tissue

The multiplexed probes were sequenced on the Illumina HiSeq 2000 platform. After the sequencing, the reads were aligned to hg38 (GRCh38.78) using TopHat2 version 2.0.12 [32] and Bowtie 2 version 2.0.6.0 [33] and quality-controlled. In order to obtain gene expression measures, the aligned reads were counted using htseq-count [34] and trimmed mean of M values (TMM)-normalized transcripts per million (TPM) counts were obtained. [35, 36] Finally, the normalized read counts were quality-controlled and low-expressed genes (expressed in less than 25% or the participants) were filtered, which yielded 30,917 genes for the main analysis.

Assessment of genetic variation

In order to investigate genetic variants, SNVs (in coding regions) were called from the RNA-Seq reads using the mpileup tool of bcftools version 1.9 [37] and further quality-controlled, trimmed and imputed. For the complete-case analysis in the sample of n = 160, 4,776,233 autosomal biallelic non-monomorphic quality-controlled SNVs were available.

See Supplementary Text for more details regarding the study population, pre-processing of the RNA-seq data, quality control steps, and details on genetic variation processing.

Statistical analysis

All analyses were performed in R 3.6.3 [38]. SAT mass was log-transformed for all analyses to yield a normally-distributed measure. The Yeo-Johnson transformation [39] was used to remove skewness and yield normally-distributed gene expression measures (based on the TMM-normalized TPM counts) for all analyses described in the following. SAT/TAT was not transformed.

In the first part of the analysis, the SAT gene expression of each of the 30,917 genes was tested for its association with SAT and SAT/TAT separately, in copula models [21, 40, 41] of the joint distribution of SAT and SAT/TAT conditional on the respective gene expression and the covariates age, sex, smoking status, physical activity and education. Copula functions can be used as a flexible tool to model the joint distribution of multiple outcomes, here SAT and SAT/TAT. By modeling the dependence of SAT and SAT/TAT, which had dependence Kendall’s τ = 0.36, the power of the association tests can be increased. In more detail, the joint distribution F of SAT and SAT/TAT was constructed using the copula function Cψ, FSAT,SAT/TAT∣x=CψF1SAT∣x,F2SAT/TAT∣x, with marginal models1 SAT=γ0+γ1x1+γ2x2+γ3x3+γ4x4+γ5x5+βjgj+ε

2 SAT/TAT=γ0′+γ1′x1+γ2′x2+γ3′x3+γ4′x4+γ5′x5+βj′gj+ε′

and the 2-parameter copula function Cψ(u1,u2,…up,φ,θ)={[∑l=1p(ul−φ−1)θ]1/θ+1}−1/φ with 0 ≤ u1,u2 ≤ 1 and ψ=(φ,θ)T,φ>0,θ≥1. Here, F1 and F2 are the marginal distributions of SAT and SAT/TAT and x=(x1,x2,x3,x4,x5,gj)T includes the gene expression gj and covariates x1…x5 sex, age, smoking, physical activity, education. Hence, in the marginal models, parameter estimates of βj and βj′ can be interpreted analogously to linear regression models and quantify the change in SAT or SAT/TAT for a 1 unit increase in gene expression, given the covariates. The copula models were fitted sequentially for the gene expression of each gene gj, j=1,…,30,917, and the large-sample Wald test statistics were computed to test the null hypotheses H0j:βj=0 (vs. HAj:βj≠0) and H0j:βj′=0 (vs. HAj:βj′≠0) using the cjamp function in the CJAMP (copula-based joint analysis of multiple phenotypes) R package [42].

Next, the obesity-associated genes identified in the above analyses were characterized regarding their functional properties, performing a gene ontology (GO)-term enrichment analysis in order to identify which gene ontology terms are enriched (under-/overrepresented) in the obesity-associated genes compared to all 30,917 analyzed genes. For details see the Supplementary Text.

In the subsequent analyses, the focus was restricted to the autosomal obesity-associated genes, and it was investigated how many of the associated genes have been found to be associated with obesity or body fat distribution in previous studies. For this, the NCBI gene database (accessed at https://www.ncbi.nlm.nih.gov/gene on July 25, 2020) was searched for genes associated with “obesity”, and all entries were extracted filtering for humans. Furthermore, the GWAS Catalog [10] (accessed at https://www.ebi.ac.uk/gwas/ on July 25, 2020) was searched for Experimental Factor Ontology (EFO) traits “obesity”, “fat body mass”, “body mass index”, “body composition measurement“, “body fat distribution”, “BMI-adjusted waist-hip ratio”, “visceral:subcutaneous adipose tissue ratio” and “visceral:total adipose tissue ratio”, and all SNVs (associations) with a p value <10−5 for all relevant reported traits and child traits were extracted by restricting the results to main overall effects (i.e. ignoring interaction effects, subgroup analyses, and proxy traits such as protein levels of obesity as traits), and restricting to body fat distribution traits of the trunk/abdomen (i.e. ignoring e.g. leg fat distribution). Finally, we queried the AstraZeneca PheWAS Portal (accessed at https://azphewas.com on June 25, 2022) which is based on a recent phenome-wide association study [43] of 18,762 genes and 2,108,983 SNVs using exome-sequencing data in UK Biobank. We extracted genes that were associated (i.e. had a p value <0.05/18,762 in any of the 11 performed collapsing gene-level tests, or contained a SNV with p value <0.05/2,108,983 in genotypic variant-level association tests) with BMI, waist circumference, whole body fat mass, abdominal SAT mass or VAT mass. For all lists, gene symbols were extracted and in order to match this list with the list of the genes associated with SAT and SAT/TAT in our study, gene symbols were extracted from the Ensembl identifiers (ID) using the biological DataBase network (accessed at https://biodbnet-abcc.ncifcrf.gov/db/db2db.php on July 25, 2020). Next, we investigated for each of the identified genes whether they encode a known protein, also by using the biological DataBase network––in more detail, by inputting the Ensembl ID of the genes and outputting the encoded UniProt protein name.

Next, we investigated the causal role of the identified genes in obesity. To this aim, we performed a Mendelian randomization (MR) study to investigate the association of genetically-determined gene expression with SAT mass and SAT/TAT. For this, all SNVs were included in the MR analysis that (i) have been identified as single-tissue cis expression quantitative trait loci (eQTLs) for SAT gene expression in Genotype-Tissue Expression (GTEx) version 8 for that respective gene (obtained from https://www.gtexportal.org/home/datasets), (ii) were not associated with any confounder of the gene expression–SAT, gene expression–SAT/TAT association, i.e. not associated with covariates sex, age, smoking, physical activity, education (tested using Wald tests of the regression coefficients in linear regression models or Fisher’s exact tests, with statistical significance threshold 0.001), (iii) were not associated with SAT and SAT/TAT, respectively, conditional on the expression of the respective gene and confounders sex, age, smoking, physical activity, education (tested in Wald tests of the regression coefficients in linear regression models, with statistical significance threshold 0.001), and (iv) with further filtering by excluding one SNV of each SNV pair with Spearman correlation greater than 0.9 (or smaller than −0.9). The analysis was performed using the mr_ivw function in the Mendelian Randomization R package [44], with the “weights =’delta’” option, psi being set to the sample correlation between gene expression and obesity measure, otherwise default settings and using the “correl=TRUE” option, which computes the inverse-variance weighted method (IVW) and allows to incorporate multiple correlated SNVs.

For a replication of the causally associated genes, we used UK Biobank data. Based on the filtered SNVs that were used in the Mendelian Randomization analysis, and their weights from a multiple linear regression model predicting the respective gene expression of the gene in the EPIC-Potsdam data, we imputed SAT gene expression based on the whole-exome sequencing data in the UK Biobank data. Then, we tested the association of this imputed genetically-determined SAT gene expression with abdominal SAT (aSAT) mass and aSAT/(aSAT+VAT) from MRI scans, in a linear regression model adjusting for age, sex, smoking and education.

Results

Description of the participants’ characteristics

The characteristics of the study population from the EPIC Potsdam study are shown in Table 1. There were slightly more women than men, and participants constituted an older and predominantly healthy sample from the general population.Table 1 Sex-stratified characteristics of the study population from the EPIC-Potsdam substudy (n=160).

Characteristics	Women	Men	
Sample size	88	72	
Age, years	62.7 (8.4)	66.9 (8.3)	
Smoking, %	
  Never	54.5	29.2	
  Former	31.8	55.6	
  Current	13.6	15.3	
CPAI, %	
  Inactive	8.0	8.3	
  Moderately inactive	26.1	29.1	
  Moderately active	35.2	33.3	
  Active	30.7	29.2	
Vocational training, %	
  No vocational training/ vocational training	40.9	34.7	
  Technical college	26.1	11.1	
  University	31.8	54.2	
 Myocardial infarction, %	0	1.4	
 Stroke, %	0	0	
 Heart failure, %	0	0	
 Diabetes, %	2.3	11.1	
 BMI, kg/m2	27.8 (4.3)	28.1 (3.9)	
 SAT, kg*	20.1 (5.2)	14.8 (4.3)	
 TAT, kg*	23.7 (5.4)	21.0 (5.9)	
Values are relative frequencies, mean and SD, or *median and median absolute deviation.

CPAI Cambridge physical activity index, BMI body mass index, AT adipose tissue, SAT subcutaneous AT, TAT total AT.

Screen of associations between gene expression and obesity

In the first part of the analysis, the SAT gene expression of each of the quality-controlled 30,917 genes was tested for its association with SAT and SAT/TAT separately. The transcriptome-wide association analysis using C-JAMP identified 441 genes associated with SAT mass and 225 associated with SAT/TAT after a respective Bonferroni-correction for multiple testing of the 30,917 genes, i.e. with a respective p value cutoff of 0.05/30,917. Of these genes, 41 overlapped so that in total, 625 genes were identified to be associated with adiposity (see Table 2). For sensitivity checks, standard univariate regression models were computed of the respective marginal models. The results showed that there was a large overlap of the associated genes identified in the copula analysis and the regression analysis, also with similar ranking (see Figure S1). The copula-based analysis identified more associated genes as compared to standard linear regression of the marginal models (410 with respect to SAT and 121 genes with respect to SAT/TAT).Table 2 Overview of the number of genes associated with only SAT, only SAT/TAT, with both SAT and SAT/TAT, and their total sum, in the analysis of 30,917 genes in the study population from the EPIC-Potsdam substudy (n=160).

Description	SAT only	SAT/TAT only	SAT and SAT/TAT	Total	
Associated genesa	400	184	41	625	
Associated autosomal genesa	392	177	38	607	
Associated autosomal genes that are knownb	69	38	13	120	
Associated autosomal genes that are novelb	323	139	25	487	
Associated autosomal genes which encode known proteinc	338	158	35	531	
aafter Bonferroni-correction for multiple testing of the 30,917 genes.

bOverlap with genes that have been associated with obesity and body fat distribution in previous studies on the molecular or genetic level. These lists have extracted on July 25, 2020 from the (i) NCBI gene database (https://www.ncbi.nlm.nih.gov/gene) by searching and extracting all associated with “obesity” in humans, and (ii) GWAS Catalog (https://www.ebi.ac.uk/gwas/), where all SNVs (associations) with a p value <10−5 for all relevant reported traits and child traits were extracted when searching for EFO (Experimental Factor Ontology) traits “obesity”, “fat body mass”, “body mass index”, “body composition measurement“, “body fat distribution”, “BMI-adjusted waist-hip ratio”, “visceral:subcutaneous adipose tissue ratio” and “visceral:total adipose tissue ratio”, and restricting the results to main overall effects (i.e. ignoring interaction effects, subgroup analyses, and proxy traits such as protein levels of obesity as traits), and restricting to body fat distribution traits of the trunk/abdomen (i.e. ignoring e.g. leg fat distribution). (iii) Further, the AstraZeneca PheWAS Portal was accessed at https://azphewas.com on June 25, 2022 and genes were extracted that were associated with BMI, waist circumference, whole body fat mass, abdominal SAT mass or VAT mass in gene-level or variant level association tests. (i) yielded 1962 genes, (ii) SNVs in 2509 genes, and (iii) 23 genes and SNPs in 562 genes, which yields in total 4460 known candidate genes for obesity and body composition.

cWhether each gene encodes a known protein was checked using the UniProt annotation from https://biodbnet-abcc.ncifcrf.gov/db/db2db.php, queried on July 25, 2020.

Characterization of the identified adiposity-associated genes

Of the 625 identified adiposity-associated genes (400 for SAT only, 184 for SAT/TAT only, 41 for both), 607 are autosomal genes and 18 are sex-chromosomal genes. Of the 607 autosomal genes, only 38 were associated with both SAT mass as well as SAT/TAT. In all further analyses described below, we focused on the 607 autosomal genes. In order to identify known and novel genes, the NCBI gene database, GWAS Catalog and AstraZeneca PheWAS Portal were searched which yielded 1962 genes, SNPs in 2509 genes, and as well as 23 genes and SNPs in 562 genes, respectively, for in total 4460 known genes associated with obesity and body composition. Of the 607 obesity-associated genes in our study, 120 have been found to be associated with adiposity in previous studies, such as the LEP gene encoding the adipokine leptin and several cytokines of the interleukin and tumor-necrosis-factor alpha families [45, 46]. Regarding a first functional characterization of the 607 genes, 531 encode a known protein. An overview of these numbers is given in Table 2.

Gene ontology (GO) term enrichment analyses indicated that the identified adiposity-associated genes are overrepresented in metabolic, cellular, regulatory and immune system processes, and that there are differences between those genes associated with body fat mass and those associated with body fat distribution. In more detail, there were 15 GO terms that were overrepresented in the 441 genes associated with SAT compared to the full pool of 30,917 genes, and 36 GO terms that were overrepresented in the 225 genes associated with SAT/TAT compared to the full pool of 30,917 genes. While the genes associated with body fat mass are mainly overrepresented in cellular, regulatory and immune system processes, those genes associated with body fat distribution are mainly overrepresented with cellular, metabolic, and regulatory processes (see Table 3 and Tables S1–S4). For example, there were 35 GO terms related to metabolic processes overrepresented in the genes associated with SAT/TAT, but no GO term related to metabolic processes overrepresented or underrepresented in the genes associated with SAT.Table 3 Summary of the results of the GO term enrichment analysis in the study population from the EPIC-Potsdam substudy (n=160).

GO terms	SAT	SAT/TAT	
Metabolic process	0	35	
Cellular process	7	27	
Immune system process	4	1	
Localization	7	1	
Locomotion	0	1	
Response to stimulus	2	0	
Biological regulation	1	0	
Cell proliferation	1	0	
There were 15 Gene Ontology (GO) terms that were (statistically significantly after multiple testing correction for all 6287 analyzed GO terms) overrepresented in the 441 genes associated with SAT compared to the full pool of 30,917 genes, and 36 GO terms that were (statistically significantly after multiple testing correction for all 6287 analyzed GO terms) overrepresented in the 225 genes associated with SAT/TAT compared to the full pool of 30,917 genes, using Fisher’s exact test. In this summary table, the number of highest-level parent GO terms of these overrepresented 15 (left panel) and 36 (right panel) GO terms is shown. For the full results of all 15 and 36 GO terms, see Tables S3, S4.

Causal gene expression effects on obesity

Next, we investigated the causal role of the identified genes in obesity in more detail. To this aim, we performed a MR study to investigate the association of genetically-determined gene expression with SAT mass and SAT/TAT. The stringent filtering steps as described in the Methods section allowed to perform a MR analysis of 261 (of the 430) genes for SAT and of 122 (of the 215) genes for SAT/TAT, which each contained at least one single nucleotide variant (SNV) after the filtering steps. In the analysis, on average 12 and 9 SNVs were included per gene for SAT mass and SAT/TAT, respectively, (min=1, max=97 for SAT and min=1, max=38 for SAT/TAT). They explained on average 10% variance of the respective gene expression for SAT and 7% variance of the respective gene expression for SAT/TAT.

In the MR analyses, the genetically-determined gene expression of 53 genes was associated with SAT mass and of 16 genes with SAT/TAT, supporting a causal effect of gene expression on adiposity for these genes. Both sets of genes were non-overlapping. They explained on average 20% variance of the respective gene expression for SAT and 15% variance of the respective gene expression for SAT/TAT. Of these 69 genes, 57 are novel genes for obesity (i.e. have not been reported to be associated with adiposity in the NCBI database and GWAS Catalog), and 46 are novel and encode a known protein. An overview of these numbers is given in Table 4 and an overview of the p-values and results for all genes in Tables S5, S6.Table 4 Overview of the number of genes associated with only SAT, only SAT/TAT, with both SAT and SAT/TAT, and their total sum, in Mendelian Randomization and replication analyses, of the 607 autosomal genes associated with SAT and/or with SAT/TAT (see Table 2) in the study population from the EPIC-Potsdam substudy (n = 160).

Description	SAT only	SAT/TAT only	SAT and SAT/TAT	Total	
Associated autosomal genesa	392	177	38	607	
Associated autosomal genes that were investigated in MR analysisb	234	95	27	356	
Associated autosomal genes identified to be causalc	53	16	0	69	
Associated autosomal genes identified to be causal and novelc,d	46	11	0	57	
Associated autosomal genes identified to be causal, novel, and to encode a known proteinc,d,e	36	10	0	46	
Associated autosomal genes identified to be causal that were investigated in the UK Biobank replication analysis	38	10	0	48	
Associated autosomal genes identified to be causal that were confirmed in the UK Biobank replication analysis	5	1	0	6	
AT adipose tissue, MR Mendelian Randomization, SAT subcutaneous AT, TAT total AT.

aafter Bonferroni-correction for multiple testing of the 30,917 genes.

bi.e. number of genes for which valid instruments were available to perform a MR analysis.

cTo test whether the identified associations are causal, MR analysis were performed predicting SAT and SAT/TAT by genetically determined gene expression.

dOverlap with genes that have been associated with obesity and body fat distribution in previous studies on the molecular or genetic level. These lists have extracted on July 25, 2020 from the (i) NCBI gene database (https://www.ncbi.nlm.nih.gov/gene) by searching and extracting all associated with “obesity” in humans, and (ii) GWAS Catalog (https://www.ebi.ac.uk/gwas/), where all single nucleotide variant (SNV) associations with a p value <10−5 for all relevant reported traits and child traits were extracted when searching for EFO (Experimental Factor Ontology) traits “obesity”, “fat body mass”, “body mass index”, “body composition measurement“, “body fat distribution”, “BMI-adjusted waist-hip ratio”, “visceral:subcutaneous adipose tissue ratio” and “visceral:total adipose tissue ratio”, and restricting the results to main overall effects (i.e. ignoring interaction effects, subgroup analyses, and proxy traits such as protein levels of obesity as traits), and restricting to body fat distribution traits of the trunk/abdomen (i.e. ignoring e.g. leg fat distribution). (iii) Further, the AstraZeneca PheWAS Portal was accessed at https://azphewas.com on June 25, 2022 and genes were extracted that were associated with BMI, waist circumference, whole body fat mass, abdominal SAT mass or VAT mass in gene-level or variant level association tests. (i) yielded 1962 genes, (ii) SNVs in 2509 genes, and (iii) 23 genes and SNPs in 562 genes, which yields in total 4460 known candidate genes for obesity and body composition.

eWhether each gene encodes a known protein was checked using the UniProt annotation from https://biodbnet-abcc.ncifcrf.gov/db/db2db.php, queried on July 25, 2020.

Replication of causally associated genes in UK Biobank

Finally, we used data generated from the participants of the UK Biobank for replication of the results. See Table S7 for characteristics of the study population containing n = 4904 participants. We were able to investigate 38 of the 53 genes for SAT mass and 10 of the 16 genes for SAT/TAT (see Table 4), with at least one SNV being available in the quality-filtered whole-exome sequencing data in the UK Biobank. For these genes, the SNVs explained on average 7% variance for SAT mass and 4% variance for SAT/TAT. Using the weights from the analysis of the EPIC-Potsdam dataset, the computed genetically-determined SAT gene expression score in UK Biobank of 5 genes was associated with SAT mass and of 1 gene was associated with aSAT/(aSAT+VAT). These 6 genes are DBNDD1, PTPRU, ERAP1, ANKDD1A and LINC02798 for SAT and MCC1 for aSAT/(aSAT+VAT), see Table 5 for an overview of these final genes with their functional annotation.Table 5 Results and annotations for the 5 autosomal genes associated with SAT mass and 1 autosomal gene associated with SAT/TAT in the Mendelian Randomization analysis that were replicated in the UK Biobank analysis.

Gene information	Gene expression analysis	Mendelian Randomization analysis	Replication analysis	
Trait	Ensembl	Gene	Chr	Start	End	Novel	Protein	beta	SE	p value	no_var	R2	p value	no_var	R2	p value	
SAT	ENSG00000003249	DBNDD1	16	90004865	90020128	YES	Dysbindin domain-containing protein 1	0,30	0,06	2,79 × 10−7	333	0,87	5,31 × 10−4	11	0,151	3,13 × 10−3	
SAT	ENSG00000060656	PTPRU	1	29236516	29326813	YES	Receptor-type tyrosine-protein phosphatase U	0,34	0,06	4,97 × 10−8	39	0,03	5,05 × 10−3	1	0,012	2,63 × 10−4	
SAT	ENSG00000164307	ERAP1	5	96760810	96808100	NO	Endoplasmic reticulum aminopeptidase 1	0,36	0,07	8,37 × 10−7	368	0,88	6,26 × 10−11	15	0,349	3,70 × 10−2	
SAT	ENSG00000166839	ANKDD1A	15	64911902	64958700	YES	Ankyrin repeat and death domain-containing protein 1 A	0,37	0,06	8,39 × 10−11	128	0,29	1,09 × 10−2	6	0,077	2,37 × 10−2	
SAT	ENSG00000227082	LINC02798	1	121396754	121463129	YES	--	0,32	0,06	5,82 × 10−7	250	0,57	4,38 × 10−6	1	0,001	2,63 × 10−4	
SAT/TAT	ENSG00000078070	MCCC1	3	183015218	183116075	YES	Methylcrotonoyl-CoA carboxylase subunit alpha, mitochondrial	0,27	0,05	1,02 × 10−7	164	0,30	2,85 × 10−2	2	0,037	4,16 × 10−2	
Shown are the Ensembl ID of the gene (“Ensembl”), the gene symbol (“Gene”), the genomic position in form of the chromosome number (“Chr”) as well as the start (“Start”) and end (“End”) position in base pairs of the gene; whether the gene is known to be associated with obesity (based on the NCBI gene,GWAS Catalog and AstraZeneca PheWAS Portal databases), the corresponding protein encoded by the gene if it is known (from UniProt); and results of the statistical analyses: the estimates of the effect size (“beta”; estimate of beta_j in the marginal model of C-JAMP in Eq. (1), or (2) respectively, in the main text) its standard error estimates (“SE”) and p value (“p value”) from the C-JAMP association analysis of SAT and SAT/TAT conditional on the gene expression and covariates; the results from the Mendelian randomization (MR) analysis in form of the number of SNVs in the gene that were incorporated in the MR analysis (“no_var”), the explained variance of the gene expression based on a multiple linear regression model containing these SNVs (“R2”), and the p value from the inverse-variance weighted MR method (“p value”); as well as the results of the replication analysis in form of the number of SNVs in the gene that were incorporated in the MR analysis (“no_var”), the explained variance of the gene expression based on a multiple linear regression model containing these SNVs (“R2”), and the p value from the linear regression analysis (“p value”).

The following gene ontology terms of the Biological Process (BP) ontology are associated with each of the genes, obtained using the annFUN.org() function in the topGO R package:

DBNDD1: --

PTPRU: protein dephosphorylation; cell adhesion; transmembrane receptor protein tyrosine phosphatase signaling pathway; negative regulation of cell proliferation; cell differentiation; negative regulation of cell migration; animal organ regeneration; homotypic cell-cell adhesion; protein localization to cell surface; peptidyl-tyrosine dephosphorylation; response to glucocorticoid; negative regulation of canonical Wnt signaling pathway; positive regulation of cell-cell adhesion mediated by cadherin

ERAP1: angiogenesis; adaptive immune response; antigen processing and presentation of peptide antigen via MHC class I; membrane protein ectodomain proteolysis; regulation of blood pressure; response to bacterium; antigen processing and presentation of endogenous peptide antigen via MHC class I; regulation of innate immune response; fat cell differentiation; positive regulation of angiogenesis

ANKDD1A: signal transduction

LINC02798: --

MCCC1: leucine catabolic process; biotin metabolic process; branched-chain amino acid catabolic process; protein heterooligomerization

Discussion

In this study, we identified 487 novel candidate genes for adiposity and 120 genes that have previously been found to be related to adiposity. The MR analysis indicates that for 69 genes, there is evidence for a causal role of their gene expression in adiposity. Importantly, 57 of these 69 genes––46 genes for body fat mass and 11 genes for body fat distribution––have not been established as adiposity genes in previous studies, and are interesting novel candidate genes whose gene expression may causally affect adiposity. Six genes were confirmed in stringent replication analysis using UK Biobank data.

Investigating these genes in follow-up studies can provide novel evidence of the molecular correlates and pathways underlying both abdominal adiposity as well as peripheral adiposity, and provide a fine-grained view on the different obesity traits that altogether constitute an established risk factor for many chronic diseases and for premature death. Interestingly, the results of our study provide ample view that body fat mass and body fat distribution are distinct phenotypes with distinct molecular correlates and underlying pathways. This was observed in the results of the transcriptomic association analysis, where 441 genes were associated with SAT, 225 genes were associated with SAT/TAT, and only 41 of these genes were overlapping. All subsequent follow-up analyses including the MR analysis fortified this separation further and revealed non-overlapping sets of genes. In addition to these separate sets of associated genes, the GO-term enrichment analyses provided further evidence. Their results indicated that SAT-associated genes were characterized by their higher than expected involvement in cellular, regulatory and immune system processes, and SAT/TAT-associated genes by their involvement in cellular, metabolic, and regulatory processes.

We investigated the causal role of the identified genes in obesity regarding the question whether they affect obesity causally, e.g. through an upregulation or downregulation of their gene expression which contributes to a metabolic imbalance which in turn contributes to obesity. The results of the Mendelian Randomization analysis suggest a causal effect of gene expression of 53 genes on SAT mass and of 16 genes on SAT/TAT. The replication analyses of these results in the UK Biobank provide support for a causal role of the gene expression of six genes on adiposity: DBNDD1, PTPRU, ERAP1, ANKDD1A and LINC02798 for body fat mass and MCC1 for body fat distribution. All these genes have not been listed in neither the NCBI database nor the GWAS Catalog as being associated with adiposity and are novel candidate genes. Only the ERAP1 gene has recently been associated in rare-variant association studies with BMI and body fat mass [43]. Further, all genes except for the Long Intergenic Non-Protein Coding RNA 2798 (LINC02798) have known proteins that could be further candidate biomarkers of interest. Regarding DBNDD1 (Dysbindin Domain Containing 1), there is evidence of its involvement in gluco-metabolic pathways [47] and through its function of binding dystrobrevin, a protein involved in intracellular processes in muscle tissue, has also some evidence for an involvement in type 2 diabetes [48]. PTPRU (Protein Tyrosine Phosphatase Receptor Type U) encodes a protein of the protein tyrosine phosphatase (PTP) family, and is a key regulator of cell communication through regulating cellular phosphotyrosine levels. The PTP family is involved, among others, in different metabolic pathways [49] and regulatory processes of cancer and diabetes [50]. Another candidate of interest for follow-up investigations is ERAP1, which encodes the endoplasmic reticulum aminopeptidase 1 and is involved in MHC class I antigen processing, hence immune processes [51], as well as in peptide catabolic processes and type 1 diabetes [52]. ANKDD1A (Ankyrin Repeat And Death Domain Containing 1 A) is a functional tumor suppressor gene and involved in signal transduction [53]. Similarly to MCC1 (Methylcrotonoyl-CoA Carboxylase 1), the involvement of ANKDD1A in metabolic and catabolic processes is still unclear. In addition to these genes and proteins, further interesting genes identified by the MR analyses but that could not be investigated in the UK Biobank replication are, for example, CD44, PLCXD3, ANG, GPR39 and GALNT10 [54].

Our study has some limitations. First, the analyses of the EPIC-Potsdam and UK Biobank cohorts were based on comparably small sample sizes, and the MR analyses as well as replication analysis were based on weak instruments. Nevertheless, we found a high number of genes to be associated with our outcomes in the gene expression analysis. This was made possible by high-quality phenotyping, analysis of quantitative phenotypes, deep RNA-sequencing data, extensive quality-control to reduce measurement error and imprecise measures, and the use of powerful statistical modeling using copula models. The association tests based on copula tests yielded more associations compared to linear regression models, and their validity to keep the nominal type I error is supported by detailed evaluations of the copula models and Wald tests in previous studies [21, 41]. These strengths of our study counterbalanced the smaller sample size compared to published studies on gene expression in subcutaneous adipose tissue that didn’t have MRI data available [14] or another focus in the analysis on cell-type composition [13]. In the choice between using more liberal instruments in the MR analysis and using more stringent filtering criteria on the SNVs, we opted for the latter to minimize risk of bias. As such, we believe that our results provide a lower bound on the genes whose gene expression in SAT is causally linked to adiposity. Similarly, only very few SNVs could be used in the replication analysis in the UK Biobank to impute SAT gene expression based on the whole-exome sequencing data. Still, the imputed gene expression of 6 genes was associated with body fat mass and body fat distribution, supporting the robustness of our analyses and results. As further limitation, the genotype calls in our study were not available from genotyping microarrays or DNA-sequencing and were called from the RNA-sequencing data. Due to again stringent filtering steps, few SNVs remained and were used in imputation. In our opinion, this might have rather decreased the power of the subsequent MR analysis, instead of increasing bias and false positive findings. As another point of discussion, about 10% of participants in our sample took antidiabetic drugs, which might affect gene expression of some genes. However, different drugs might have different effects and detailed information on drugs was not available. Therefore, we opted not to add antidiabetic medication use as a covariate in the copula model, which was supported by sensitivity checks we performed in an earlier study where results of association studies did not change [28]. Even if the association tests of some genes might have been affected by this, our choice of following up the identified genes in a Mendelian Randomization analysis ensured that these results would not be affected. Finally, in our study, we only investigated gene expression in subcutaneous adipose tissue. Without the assessment of VAT gene expression and secretion rates from SAT and VAT, however, parts of the overall molecular picture remain unclear. For example, in terms of molecular mechanisms, it cannot be ruled out that the gene expression in VAT is upregulated or downregulated in parallel to the gene expression in SAT. However, assessing VAT in a population–based study is rarely possible, and VAT gene expression measured after bariatric or other surgeries might not allow for a valid approximation of the metabolic activity in the general population.

In summary, we identified novel adiposity genes that are fat mass specific and fat distribution-specific, involved in different molecular processes, and whose upregulated or downregulated gene expression may causally contribute to obesity. These findings can provide guidance for future work in finding pieces in the puzzle of molecular mechanisms contributing to adiposity.

Supplementary information

Supplementary Material

Supplementary Table 5

Supplementary Table 6

Supplementary information

The online version contains supplementary material available at 10.1038/s41431-022-01161-3.

Acknowledgements

The authors thank the participants of the EPIC Potsdam substudy and the teams at the Human Study Center of the German Institute of Human Nutrition Potsdam-Rehbrücke and the German Cancer Center Heidelberg for the handling of the data and images, and the transformation of the MRI scans into body composition information. We thank Johannes Hierholzer for his contributions and supervision of the MRI scans at the Ernst-von-Bergmann Clinic and to Martin Küper for the accurate performance of the MRI scans. We are grateful to Dagmar Drogan for help in the setup of the analysis of the EPIC data and data preparation, to Sarah Moreno Garcia and Henning Damm for the sample handling and for performing the ELISA and PCR experiments. The Genotype-Tissue Expression (GTEx) Project was supported by the Common Fund of the Office of the Director of the National Institutes of Health, and by NCI, NHGRI, NHLBI, NIDA, NIMH, and NINDS. We thank the participants of the UK Biobank study. The analyses of the UK Biobank have been conducted using the UK Biobank Resource under application number 40502.

Author contributions

TP, NH, HB, SK conceived and designed the gene expression experiments; MMB, HB, conceived, designed and maintained the EPIC Potsdam study, and MMB, HB, TP conceived and designed the EPIC Potsdam substudy including the assessment of MRI scans together with RK. JJ extracted the RNA from the SAT biopsies; TP, JJ, NH, SK, GP supervised the further preparation of samples and sequencing. GP performed initial bioinformatic processing of the RNA sequencing (demultiplexing, read alignment, obtaining read counts). SK performed all further bioinformatics processing steps including the SNP calling of the EPIC data, all data analyses including the gene expression, Mendelian Randomization, and replication analysis, and interpretation of all analyses. SK wrote the manuscript with intellectual input from all authors. TP supervised the project. All authors reviewed and approved the final manuscript.

Funding

Open Access funding enabled and organized by Projekt DEAL.

Data availability

For the analyses described in this manuscript, the file Adipose_Subcutaneous.v8.signif_variant_gene_pairs.txt.gz from GTEx_Analysis_v8_eQTL.tar was obtained from the GTEx Portal https://www.gtexportal.org/home/datasets on 05/18/2020. The UK Biobank data can be requested upon application at https://www.ukbiobank.ac.uk/register-apply/. The analyzed datasets from the still ongoing EPIC-Potsdam study have ethical and legal restrictions for public deposition due to the data protection rules applicable to its informed consent. The gene expression data and phenotypic data will therefore be available upon application with an appropriate research proposal which goes through a process of review and evaluation. To request the data, please contact Prof. Matthias Schulze at mschulze@dife.de. All bioinformatic processing and statistical analyses were conducted using publicly available software as referenced in the Methods section and Supplementary Text.

Competing interests

The authors declare no competing interests.

Ethical approval and consent to participate

The study was approved by the Ethics Committee of the medical association of the state of Brandenburg (Germany) and all participants gave written informed consent.

Publisher’s note Springer Nature remains neutral with regard to jurisdictional claims in published maps and institutional affiliations.

Change history

8/29/2022

A Correction to this paper has been published: 10.1038/s41431-022-01183-x
==== Refs
References

1. Nimptsch K Konigorski S Pischon T Diagnosis of obesity and use of obesity biomarkers in science and clinical medicine Metabolism 2019 92 61 70 10.1016/j.metabol.2018.12.006 30586573
Nimptsch K, Konigorski S, Pischon T. Diagnosis of obesity and use of obesity biomarkers in science and clinical medicine. Metabolism. 2019;92:61–70.30586573 10.1016/j.metabol.2018.12.006
2. Porter SA Massaro JM Hoffmann U Vasan RS O’Donnel CJ Fox CS Abdominal subcutaneous adipose tissue: a protective fat depot? Diabetes Care 2009 32 1068 75 10.2337/dc08-2280 19244087
Porter SA, Massaro JM, Hoffmann U, Vasan RS, O’Donnel CJ, Fox CS. Abdominal subcutaneous adipose tissue: a protective fat depot? Diabetes Care. 2009;32:1068–75.19244087 10.2337/dc08-2280
3. Herrera BM Lindgren CM The genetics of obesity Curr Diab Rep. 2010 10 498 505 10.1007/s11892-010-0153-z 20931363
Herrera BM, Lindgren CM. The genetics of obesity. Curr Diab Rep. 2010;10:498–505.20931363 10.1007/s11892-010-0153-z
4. Yang W Kelly T He J Genetic epidemiology of obesity Epidemiol Rev 2007 29 49 61 10.1093/epirev/mxm004 17566051
Yang W, Kelly T, He J. Genetic epidemiology of obesity. Epidemiol Rev 2007;29:49–61.17566051 10.1093/epirev/mxm004
5. Rask-Andersen M Karlsson T Ek WE Åsa J Genome-wide association study of body fat distribution identifies adiposity loci and sex-specific genetic effects Nat Commun 2019 10 339 10.1038/s41467-018-08000-4 30664634
Rask-Andersen M, Karlsson T, Ek WE, Åsa J. Genome-wide association study of body fat distribution identifies adiposity loci and sex-specific genetic effects. Nat Commun 2019;10:339.30664634 10.1038/s41467-018-08000-4
6. Schleinitz D Böttcher Y Blüher M Kovacs P The genetics of fat distribution Diabetologia 2014 57 1276 86 10.1007/s00125-014-3214-z 24632736
Schleinitz D, Böttcher Y, Blüher M, Kovacs P. The genetics of fat distribution. Diabetologia. 2014;57:1276–86.24632736 10.1007/s00125-014-3214-z
7. Shungin D Winkler TW Croteau-Chonka DC Ferreira T Locke AE Mägi R New genetic loci link adipose and insulin biology to body fat distribution Nature 2015 518 187 96 10.1038/nature14132 25673412
Shungin D, Winkler TW, Croteau-Chonka DC, Ferreira T, Locke AE, Mägi R, et al. New genetic loci link adipose and insulin biology to body fat distribution. Nature. 2015;518:187–96.25673412 10.1038/nature14132
8. Locke AE Kahali B Berndt SI Justice AE Pers TH Day FR Genetic studies of body mass index yield new insights for obesity biology Nature 2015 518 197 206 10.1038/nature14177 25673413
Locke AE, Kahali B, Berndt SI, Justice AE, Pers TH, Day FR, et al. Genetic studies of body mass index yield new insights for obesity biology. Nature. 2015;518:197–206.25673413 10.1038/nature14177
9. Lu Y Day FR Gustafsson S Buchkovich ML Na J Bataille V New loci for body fat percentage reveal link between adiposity and cardiometabolic disease risk Nat Commun 2016 7 10495 10.1038/ncomms10495 26833246
Lu Y, Day FR, Gustafsson S, Buchkovich ML, Na J, Bataille V, et al. New loci for body fat percentage reveal link between adiposity and cardiometabolic disease risk. Nat Commun 2016;7:10495.26833246 10.1038/ncomms10495
10. MacArthur J Bowler E Cerezo M Gil L Hall P Hastings E The new NHGRI-EBI Catalog of published genome-wide association studies (GWAS catalog) Nucleic Acids Res 2017 45 D896 D901. 10.1093/nar/gkw1133 27899670
MacArthur J, Bowler E, Cerezo M, Gil L, Hall P, Hastings E, et al. The new NHGRI-EBI Catalog of published genome-wide association studies (GWAS catalog). Nucleic Acids Res. 2017;45:D896–D901.27899670 10.1093/nar/gkw1133
11. Campbell KL Foster-Schubert KE Makar KW Kratz M Hagman D Schur EA Gene expression changes in adipose tissue with diet–and/or exercise–induced weight loss Cancer Prev Res 2013 6 217 31 10.1158/1940-6207.CAPR-12-0212
Campbell KL, Foster-Schubert KE, Makar KW, Kratz M, Hagman D, Schur EA, et al. Gene expression changes in adipose tissue with diet–and/or exercise–induced weight loss. Cancer Prev Res. 2013;6:217–31.10.1158/1940-6207.CAPR-12-0212
12. Del Cornò M Baldassarre A Calura E Conti L Martini P Romualdi C Transcriptome profiles of human visceral adipocytes in obesity and colorectal cancer unravel the effects of body mass index and polyunsaturated fatty acids on genes and biological processes related to tumorigenesis Front Immunol 2019 10 265 10.3389/fimmu.2019.00265 30838002
Del Cornò M, Baldassarre A, Calura E, Conti L, Martini P, Romualdi C, et al. Transcriptome profiles of human visceral adipocytes in obesity and colorectal cancer unravel the effects of body mass index and polyunsaturated fatty acids on genes and biological processes related to tumorigenesis. Front Immunol 2019;10:265.30838002 10.3389/fimmu.2019.00265
13. Glastonbury CA Couto Alves A El-Sayed Moustafa JS Small KS Cell-type heterogeneity in adipose tissue is associated with complex traits and reveals disease-relevant cell-specific eQTLs Am J Hum Genet 2019 104 1013 24 10.1016/j.ajhg.2019.03.025 31130283
Glastonbury CA, Couto Alves A, El-Sayed Moustafa JS, Small KS. Cell-type heterogeneity in adipose tissue is associated with complex traits and reveals disease-relevant cell-specific eQTLs. Am J Hum Genet 2019;104:1013–24.31130283 10.1016/j.ajhg.2019.03.025
14. Raulerson CK Ko A Kidd JC Currin KW Brotman SM Cannon ME Adipose tissue gene expression associations reveal hundreds of candidate genes for cardiometabolic traits Am J Hum Genet 2019 105 773 87 10.1016/j.ajhg.2019.09.001 31564431
Raulerson CK, Ko A, Kidd JC, Currin KW, Brotman SM, Cannon ME, et al. Adipose tissue gene expression associations reveal hundreds of candidate genes for cardiometabolic traits. Am J Hum Genet 2019;105:773–87.31564431 10.1016/j.ajhg.2019.09.001
15. Armenise C Lefebvre G Carayol J Bonnel S Bolton J Di Cara A Transcriptome profiling from adipose tissue during a low-calorie diet reveals predictors of weight and glycemic outcomes in obese, nondiabetic subjects Am J Clin Nutr 2017 106 736 46. 10.3945/ajcn.117.156216 28793995
Armenise C, Lefebvre G, Carayol J, Bonnel S, Bolton J, Di Cara A, et al. Transcriptome profiling from adipose tissue during a low-calorie diet reveals predictors of weight and glycemic outcomes in obese, nondiabetic subjects. Am J Clin Nutr 2017;106:736–46.28793995 10.3945/ajcn.117.156216
16. Day K Dordevic AL Truby H Southey MC Coort S Murgia C Transcriptomic changes in peripheral blood mononuclear cells with weight loss: systematic literature review and primary data synthesis Genes Nutr 2021 16 12 10.1186/s12263-021-00692-6 34281497
Day K, Dordevic AL, Truby H, Southey MC, Coort S, Murgia C. Transcriptomic changes in peripheral blood mononuclear cells with weight loss: systematic literature review and primary data synthesis. Genes Nutr. 2021;16:12.34281497 10.1186/s12263-021-00692-6
17. Kerr AG Andersson DP Rydén M Arner P Dahlman I Long-term changes in adipose tissue gene expression following bariatric surgery J Intern Med 2020 288 219 33 10.1111/joim.13066 32406570
Kerr AG, Andersson DP, Rydén M, Arner P, Dahlman I. Long-term changes in adipose tissue gene expression following bariatric surgery. J Intern Med 2020;288:219–33.32406570 10.1111/joim.13066
18. Paczkowska-Abdulsalam M Niemira M Bielska A Szałkowska A Raczkowska BA Junttila S Evaluation of Transcriptomic Regulations behind Metabolic Syndrome in Obese and Lean Subjects Int J Mol Sci 2020 21 1455 10.3390/ijms21041455 32093387
Paczkowska-Abdulsalam M, Niemira M, Bielska A, Szałkowska A, Raczkowska BA, Junttila S, et al. Evaluation of Transcriptomic Regulations behind Metabolic Syndrome in Obese and Lean Subjects. Int J Mol Sci. 2020;21:1455.32093387 10.3390/ijms21041455
19. Rodriguez-Ayala E Gallegos-Cabrales EC Gonzalez-Lopez L Laviada-Molina HA Salinas-Osornio RA Nava-Gonzalez EJ Towards precision medicine: defining and characterizing adipose tissue dysfunction to identify early immunometabolic risk in symptom-free adults from the GEMM family study Adipocyte 2020 9 153 69 10.1080/21623945.2020.1743116 32272872
Rodriguez-Ayala E, Gallegos-Cabrales EC, Gonzalez-Lopez L, Laviada-Molina HA, Salinas-Osornio RA, Nava-Gonzalez EJ, et al. Towards precision medicine: defining and characterizing adipose tissue dysfunction to identify early immunometabolic risk in symptom-free adults from the GEMM family study. Adipocyte. 2020;9:153–69.32272872 10.1080/21623945.2020.1743116
20. Zhou Q Fu Z Gong Y Seshachalam VP Li J Ma Y Metabolic health status contributes to transcriptome alternation in human visceral adipose tissue during obesity Obes (Silver Spring) 2020 28 2153 62 10.1002/oby.22950
Zhou Q, Fu Z, Gong Y, Seshachalam VP, Li J, Ma Y, et al. Metabolic health status contributes to transcriptome alternation in human visceral adipose tissue during obesity. Obes (Silver Spring). 2020;28:2153–62.10.1002/oby.22950
21. Konigorski S Yilmaz YE Janke J Bergmann MM Boeing H Pischon T Powerful rare variant association testing in a copula-based joint analysis of multiple traits Genet Epidemiol 2020 44 26 40 10.1002/gepi.22265 31732979
Konigorski S, Yilmaz YE, Janke J, Bergmann MM, Boeing H, Pischon T. Powerful rare variant association testing in a copula-based joint analysis of multiple traits. Genet Epidemiol 2020;44:26–40.31732979 10.1002/gepi.22265
22. Boeing H Wahrendorf J Becker N EPIC-Germany – A source for studies into diet and risk of chronic diseases. European Investigation into Cancer and Nutrition Ann Nutr Metab 1999 43 195 204 10.1159/000012786 10592368
Boeing H, Wahrendorf J, Becker N. EPIC-Germany – A source for studies into diet and risk of chronic diseases. European Investigation into Cancer and Nutrition. Ann Nutr Metab 1999;43:195–204.10592368 10.1159/000012786
23. Gottschald M Knüppel S Boeing H Buijsse B The influence of adjustment for energy misreporting on relations of cake and cookie intake with cardiometabolic disease risk factors Eur J Clin Nutr 2016 70 1318 24 10.1038/ejcn.2016.131 27460264
Gottschald M, Knüppel S, Boeing H, Buijsse B. The influence of adjustment for energy misreporting on relations of cake and cookie intake with cardiometabolic disease risk factors. Eur J Clin Nutr 2016;70:1318–24.27460264 10.1038/ejcn.2016.131
24. Wientzek A Vigl M Steindorf K Brühmann B Bergmann MM Harttig U The improved physical activity index for measuring physical activity in EPIC Germany PLoS One 2014 9 e92005 10.1371/journal.pone.0092005 24642812
Wientzek A, Vigl M, Steindorf K, Brühmann B, Bergmann MM, Harttig U, et al. The improved physical activity index for measuring physical activity in EPIC Germany. PLoS One. 2014;9:e92005.24642812 10.1371/journal.pone.0092005
25. Neamat-Allah J Wald D Hüsing A Teucher B Wendt A Delorme S Validation of anthropometric indices of adiposity against whole-body magnetic resonance imaging - a study within the German European Prospective Investigation into Cancer and Nutrition (EPIC) cohorts PLoS One 2014 9 e91586 10.1371/journal.pone.0091586 24626110
Neamat-Allah J, Wald D, Hüsing A, Teucher B, Wendt A, Delorme S, et al. Validation of anthropometric indices of adiposity against whole-body magnetic resonance imaging - a study within the German European Prospective Investigation into Cancer and Nutrition (EPIC) cohorts. PLoS One. 2014;9:e91586.24626110 10.1371/journal.pone.0091586
26. Wald D Teucher B Dinkel J Kaaks R Delorme S Boeing H Automatic quantification of subcutaneous and visceral adipose tissue from whole-body magnetic resonance images suitable for large cohort studies J Magn Reson Imaging 2012 36 1421 34 10.1002/jmri.23775 22911921
Wald D, Teucher B, Dinkel J, Kaaks R, Delorme S, Boeing H, et al. Automatic quantification of subcutaneous and visceral adipose tissue from whole-body magnetic resonance images suitable for large cohort studies. J Magn Reson Imaging. 2012;36:1421–34.22911921 10.1002/jmri.23775
27. Wald D, Teucher B, Dinkel J, Kaaks R, Delorme S, Meinzer HP, et al. (2012). Automated quantification of adipose and skeletal muscle tissue in whole-body MRI data for epidemiological studies. Medical Imaging 2012: Computer-Aided Diagnosis. Edited by van Ginneken B, Novak CL Proceedings of the SPIE 8315, 831519.
28. Konigorski S Janke J Drogan D Bergmann MM Hierholzer J Kaaks R Prediction of circulating adipokine levels based on body fat compartments and adipose tissue gene expression Obes Facts 2019 12 590 605 10.1159/000502117 31698359
Konigorski S, Janke J, Drogan D, Bergmann MM, Hierholzer J, Kaaks R, et al. Prediction of circulating adipokine levels based on body fat compartments and adipose tissue gene expression. Obes Facts. 2019;12:590–605.31698359 10.1159/000502117
29. Bycroft C Freeman C Petkova D Band G Elliott LT Sharp K The UK Biobank resource with deep phenotyping and genomic data Nature 2018 562 203 9 10.1038/s41586-018-0579-z 30305743
Bycroft C, Freeman C, Petkova D, Band G, Elliott LT, Sharp K, et al. The UK Biobank resource with deep phenotyping and genomic data. Nature. 2018;562:203–9.30305743 10.1038/s41586-018-0579-z
30. Van Hout CV, Tachmazidou I, Backman JD, Hoffman JX, Ye B, Pandey AK, et al. Whole exome sequencing and characterization of coding variation in 49,960 individuals in the UK Biobank. 2019. bioRxiv. 10.1101/572347.
31. Linge J Borga M West J Tuthill T Miller MR Dumitriu A Body composition profiling in the UK Biobank imaging study Obesity 2018 26 1785 1795 10.1002/oby.22210 29785727
Linge J, Borga M, West J, Tuthill T, Miller MR, Dumitriu A, et al. Body composition profiling in the UK Biobank imaging study. Obesity. 2018;26:1785–1795.29785727 10.1002/oby.22210
32. Kim D Pertea G Trapnell C Pimentel H Kelley R Salzberg SL TopHat2: accurate alignment of transcriptomes in the presence of insertions, deletions and gene fusions Genome Biol 2013 14 R36 10.1186/gb-2013-14-4-r36 23618408
Kim D, Pertea G, Trapnell C, Pimentel H, Kelley R, Salzberg SL. TopHat2: accurate alignment of transcriptomes in the presence of insertions, deletions and gene fusions. Genome Biol. 2013;14:R36.23618408 10.1186/gb-2013-14-4-r36
33. Langmead B Salzberg SL Fast gapped-read alignment with Bowtie 2 Nat Methods 2012 14 357 9 10.1038/nmeth.1923
Langmead B, Salzberg SL. Fast gapped-read alignment with Bowtie 2. Nat Methods. 2012;14:357–9.10.1038/nmeth.1923
34. Anders S Pyl PT Huber W HTSeq - a Python framework to work with highthroughput sequencing data Bioinformatics 2015 31 166 9 10.1093/bioinformatics/btu638 25260700
Anders S, Pyl PT, Huber W. HTSeq - a Python framework to work with highthroughput sequencing data. Bioinformatics. 2015;31:166–9.25260700 10.1093/bioinformatics/btu638
35. Li B Dewey CN RSEM: accurate transcript quantification from RNA-Seq data with or without a reference genome BMC Bioinforma 2011 12 323 10.1186/1471-2105-12-323
Li B, Dewey CN. RSEM: accurate transcript quantification from RNA-Seq data with or without a reference genome. BMC Bioinforma. 2011;12:323.10.1186/1471-2105-12-323
36. Robinson MD Oshlack A A scaling normalization method for differential expression analysis of RNA-seq data Genome Biol 2010 11 R25 10.1186/gb-2010-11-3-r25 20196867
Robinson MD, Oshlack A. A scaling normalization method for differential expression analysis of RNA-seq data. Genome Biol. 2010;11:R25.20196867 10.1186/gb-2010-11-3-r25
37. Li H Handsaker B Wysoker A Fennell T Ruan J Homer N The Sequence alignment/map (SAM) format and SAMtools Bioinformatics 2009 25 2078 9 10.1093/bioinformatics/btp352 19505943
Li H, Handsaker B, Wysoker A, Fennell T, Ruan J, Homer N, et al. The Sequence alignment/map (SAM) format and SAMtools. Bioinformatics. 2009;25:2078–9.19505943 10.1093/bioinformatics/btp352
38. R Core Team (2017). R: a language and environment for statistical computing. R Foundation for Statistical Computing, Vienna, Austria. https://www.R-project.org/.
39. Yeo IK Johnson R A new family of power transformations to improve normality or symmetry Biometrika 2000 87 954 9 10.1093/biomet/87.4.954
Yeo IK, Johnson R. A new family of power transformations to improve normality or symmetry. Biometrika. 2000;87:954–9.10.1093/biomet/87.4.954
40. Konigorski S Yilmaz YE Bull SB Bivariate genetic association analysis of systolic and diastolic blood pressure by copula models BMC Proc 2014 8 S72 S77. 10.1186/1753-6561-8-S1-S72 25519342
Konigorski S, Yilmaz YE, Bull SB. Bivariate genetic association analysis of systolic and diastolic blood pressure by copula models. BMC Proc. 2014;8:S72–S77.25519342 10.1186/1753-6561-8-S1-S72
41. Konigorski S Yilmaz YE Pischon T Genetic association analysis based on a joint model of gene expression and blood pressure BMC Proc 2016 10 57 10.1186/s12919-016-0045-6
Konigorski S, Yilmaz YE, Pischon T. Genetic association analysis based on a joint model of gene expression and blood pressure. BMC Proc. 2016;10:57.10.1186/s12919-016-0045-6
42. Konigorski S, Yilmaz YE (2018). CJAMP: Copula-based joint analysis of multiple phenotypes. R package version 0.1.0.
43. Wang Q Dhindsa RS Carss K Harper AR Nag A Tachmazidou I Rare variant contribution to human disease in 281,104 UK Biobank exomes Nature 2021 597 527 32 10.1038/s41586-021-03855-y 34375979
Wang Q, Dhindsa RS, Carss K, Harper AR, Nag A, Tachmazidou I, et al. Rare variant contribution to human disease in 281,104 UK Biobank exomes. Nature 2021;597:527–32.34375979 10.1038/s41586-021-03855-y
44. Yavorska O (2018). MendelianRandomization: Mendelian Randomization package. R package version 0.3.0.
45. Aleksandrova K Boeing H Jenab M Bueno-de-Mesquita HB Jansen E van Duijnhoven FJB Leptin and soluble leptin receptor in risk of colorectal cancer in the European prospective investigation into cancer and nutrition cohort Cancer Res 2012 72 5328 37 10.1158/0008-5472.CAN-12-0465 22926557
Aleksandrova K, Boeing H, Jenab M, Bueno-de-Mesquita HB, Jansen E, van Duijnhoven FJB, et al. Leptin and soluble leptin receptor in risk of colorectal cancer in the European prospective investigation into cancer and nutrition cohort. Cancer Res. 2012;72:5328–37.22926557 10.1158/0008-5472.CAN-12-0465
46. Tilg H Moschen AR Adipocytokines: mediators linking adipose tissue, inflammation and immunity Nat Rev Immunol 2006 6 772 83 10.1038/nri1937 16998510
Tilg H, Moschen AR. Adipocytokines: mediators linking adipose tissue, inflammation and immunity. Nat Rev Immunol 2006;6:772–83.16998510 10.1038/nri1937
47. Langefeld CD Comeau ME Sharma NK Bowden DW Freedman BI Das SK Transcriptional regulatory mechanisms in adipose and muscle tissue associated with composite glucometabolic phenotypes Obesity 2018 26 559 69 10.1002/oby.22113 29377571
Langefeld CD, Comeau ME, Sharma NK, Bowden DW, Freedman BI, Das SK. Transcriptional regulatory mechanisms in adipose and muscle tissue associated with composite glucometabolic phenotypes. Obesity. 2018;26:559–69.29377571 10.1002/oby.22113
48. Keildson S Fadista J Ladenvall C Hedman ÅK Elgzyri T Small KS Expression of phosphofructokinase in skeletal muscle is influenced by genetic variation and associated with insulin sensitivity Diabetes 2014 63 1154 65 10.2337/db13-1301 24306210
Keildson S, Fadista J, Ladenvall C, Hedman ÅK, Elgzyri T, Small KS, et al. Expression of phosphofructokinase in skeletal muscle is influenced by genetic variation and associated with insulin sensitivity. Diabetes. 2014;63:1154–65.24306210 10.2337/db13-1301
49. Elchebly M Payette P Michaliszyn E Cromlish W Collins S Loy AL Increased insulin sensitivity and obesity resistance in mice lacking the protein tyrosine phosphatase-1B gene Science 1999 283 1544 8 10.1126/science.283.5407.1544 10066179
Elchebly M, Payette P, Michaliszyn E, Cromlish W, Collins S, Loy AL, et al. Increased insulin sensitivity and obesity resistance in mice lacking the protein tyrosine phosphatase-1B gene. Science. 1999;283:1544–8.10066179 10.1126/science.283.5407.1544
50. Hay IM Fearnley GW Rios P Köhn M Sharpe HJ Deane JE The receptor PTPRU is a redox sensitive pseudophosphatase Nat Commun 2020 11 3219 10.1038/s41467-020-17076-w 32591542
Hay IM, Fearnley GW, Rios P, Köhn M, Sharpe HJ, Deane JE. The receptor PTPRU is a redox sensitive pseudophosphatase. Nat Commun 2020;11:3219.32591542 10.1038/s41467-020-17076-w
51. Pepelyayeva Y Amalfitano A The role of ERAP1 in autoinflammation and autoimmunity Hum Immunol 2019 80 302 9 10.1016/j.humimm.2019.02.013 30817945
Pepelyayeva Y, Amalfitano A. The role of ERAP1 in autoinflammation and autoimmunity. Hum Immunol 2019;80:302–9.30817945 10.1016/j.humimm.2019.02.013
52. Kronenberg-Versteeg D Eichmann M Russell MA de Ru A Hehn B Yusuf N Molecular pathways for immune recognition of preproinsulin signal peptide in type 1 diabetes Diabetes 2018 67 687 96 10.2337/db17-0021 29343547
Kronenberg-Versteeg D, Eichmann M, Russell MA, de Ru A, Hehn B, Yusuf N, et al. Molecular pathways for immune recognition of preproinsulin signal peptide in type 1 diabetes. Diabetes. 2018;67:687–96.29343547 10.2337/db17-0021
53. Feng J Zhang Y She X Sun Y Fan L Ren X Hypermethylated gene ANKDD1A is a candidate tumor suppressor that interacts with FIH1 and decreases HIF1a stability to inhibit cell autophagy in the glioblastoma multiforme hypoxia microenvironment Oncogene 2019 38 103 19 10.1038/s41388-018-0423-9 30082910
Feng J, Zhang Y, She X, Sun Y, Fan L, Ren X, et al. Hypermethylated gene ANKDD1A is a candidate tumor suppressor that interacts with FIH1 and decreases HIF1a stability to inhibit cell autophagy in the glioblastoma multiforme hypoxia microenvironment. Oncogene. 2019;38:103–19.30082910 10.1038/s41388-018-0423-9
54. Monda KL Chen GK Taylor KC Palmer C Edwards TL Lange LA A meta-analysis identifies new loci associated with body mass index in individuals of African ancestry Nat Genet 2013 45 690 6 10.1038/ng.2608 23583978
Monda KL, Chen GK, Taylor KC, Palmer C, Edwards TL, Lange LA, et al. A meta-analysis identifies new loci associated with body mass index in individuals of African ancestry. Nat Genet. 2013;45:690–6.23583978 10.1038/ng.2608
