TM content and topology in viral and microbial proteins The distributions of membrane proteins by the number of TM segments in representative sets of prokaryotic and viral genomes show striking, highly significant differences ( p -value < 1e-300 by Chi-Squared test and <2e-16 by Mann-Whitney-Wilcoxon test) whereas the archaeal and bacterial distributions are indistinguishable (Fig. 3 ).
Excerpts
All of these degrees of overlap are statistically highly significant ( p -value <1 × 10 –300 ).
Association analysis of normalized methylation data (beta-values) from 391,885 CpGs revealed 11,010 sex-methylation associations (SMAs) ( p < 1.26E−07) with some of them being highly significant ( p < 1E−300).
cuniculus , which is extremely significant (r 2 = 0.305, p -value = 6.24 × 10 −295 ); furthermore, H. sapiens shows significant correlations even with M. musculus (r 2 = 0.177, p -value = 2.7 × 10 −113 ) and C. lupus familiaris (r2 = 0.260, p -value = 4.15 × 10 −226 ).
The rejection of unimodality was supported by a highly significant dip test 21 (dip = 0.086275, p < 10 –293 ) using the Python “diptest” package available at https://pypi.org/project/diptest/ .
While due to its relative simplicity only a weak positive correlation was expected for the Zhang model, for reasons unclear, a highly significant ( p < 10 −293 ) negative correlation is observed ( Fig 5 , left).
(2018) , which also tested for ASE in crosses between PAXB and RABS in the VTP, we found a highly significant overlap (Fisher's exact test, P = 8.83 × 10 −292 ).
The overlap between the 67 genes with those 1008 is highly significant ( p = 1 − 290 , hypergeometric test), suggesting that our 126 genes are particularly enriched in DE genes.
We compared the overlap of MPeM interactome with the MPM interactome [ 29 ], revealing 989 shared genes, a highly significant overlap ( P -value = 3.18E−289, odds ratio = 2.92).
As expected, these DEGs clustered samples by karyotype, but also featured highly significant overlap and concordance in direction ( Fig. 4 A ), including 936 concordant DEGs out of 1,283 common to all three comparisons (73% concordant, sign test P = 8.2 × 10 −285 ).
This intersection is highly significant (p value = 9,38E −283 ), indicating that H2A.Z isoforms regulate overlapping sets of genes.
In agreement with the elevated transcription in the HEID, genome wide, RAD50 showed a strongly positive (cor = 0.565) and highly significant ( P = 2.4 × 10 −281 ) correlation with transcriptional strength.
Consistent with this idea, plotting eQTL effect size vs MAF in the baboons results in a very strong, highly significant negative correlation (r = −0.723, p < 10 −280 ; Figure 3C ), with no large effect eQTL detected at higher MAFs.
First, a modified t-test reduced the effective sample size to 11,010.13 (df = 11,008.13), yet the corrected test statistic remained highly significant (t = 36.84, p = 3.06 × 10 −280 ).
Moreover, when considering all the genes, we observed a positive correlation between expression change and the change in H3.3 enrichment that is modest, but highly significant (Spearman rank correlation of 0.28, p-value<1e-275) ( Figure 3C ).
Despite all this, 9414 of our predicted interactions are found in the relatively small set of eCLIP experiments, a highly significant overlap considering the 82 proteins and 7381 transcripts present in both eCLIP and cat RAPID datasets ( P -value < 2.2e-271, OR = 1.85, two-tailed Fisher's exact test), therefore increasing our confidence in the predicted network.
Such an overlap is highly significant considering the size of the yeast proteome ( p < 10 -269 by Fisher's Exact Test), indicating that our approach is robust with regard to the gold standard used.
The multivariable robust linear regression model that tested smoking versus methylation at AHRR cg05575921 was highly significant (smoking regression coefficient p = 7.27E-269, r 2 = 0.58; adjusted for age, sex, race, lipids, cell types and batch).
At enhancer sites, CTBP2 peaks, as well as HDAC3 peaks, demonstrated a highly significant overlap with the TRPS1 peaks that we identified in our previous experiments in MCF7 cells ( P = 4.28 × 10 −268 for CTBP2; P = 6.08 × 10 −53 for HDAC3).
Similarly, for sites detected in standard alignment and PAC (but not WASP-filtered data), PAC quantification is highly significant (496 sites, R 2 = 0.956, P = 2.6 × 10 −266 ) ( Fig. 2 A), showing that PAC performs well at sites with lower coverage that may be missed by other approaches.
In total, 568 predicted promoters were found in proximity (≤40 nucleotides) of the putative TSSs, showing a highly significant co-occurrence of predicted promoter and TSS (p-value<10 −263 ).
Statistical analysis using the Kruskal–Wallis test confirmed highly significant differences in expression distributions across regulatory categories ( P = 3.28e−261).
Furthermore, rs2796265 genotypes correlate with a highly significant CD46 mRNA splicing quantitative locus in several tissues: fibroblast ( p = 1.4 × 10 −259 ), oesophagus ( p = 1.7 × 10 −188 ), immortalized B-lymphocytes ( p = 2.1 × 10 −57 ), or brain-cortex ( p = 1.3 × 10 −25 ) ( Figure 4 ). 4.
We found highly significant overlap between the two replicates (p = 7.5 x 10 −257 ) and there was no correlation between the editing frequency of a nucleotide and the number of times that nucleotide was read ( Fig 5B ), suggesting that editing frequency did not simply reflect mRNA abundance.
This overlap is highly significant, as the probability of such an overlap to occur in random is very low (p<10 −256 ). 45% of the 402 overlapped CpG islands are located in intronic and intergenic areas, 6% in promoters and 49% in exons.