CroCoNet: a framework for the quantitative comparison of gene regulatory networks across species.
The 40 matches · 5 of them tie a paragraph to a whole file, not to given lines: weak matches, whose lines are not tinted
- [1] § Methods › Cell lines ↔ 4.paper_figures_and_tables/suppl.tables.R, lines 32–83 · score 0.98 · peripheral blood mononuclear, urine derived stem, dermal fibroblasts, integration free Sendai, OSKM vectors, RNA seq
- [2] § Methods › Quantification of cross-species divergence ↔ R/data.R, lines 482–563 · score 0.94 · weighted linear model, upper boundary, inversely proportional, lower boundary, prediction interval, identify modules
- [3] § Methods › Quantification of cross-species divergence ↔ R/findConservedDivergedModules.R, lines 1–72 · score 0.94 · upper boundary, inversely proportional, lower boundary, dependent variable, prediction interval, linear model
- [4] § Results › Consensus network and module assignment ↔ R/pruneModules.R, lines 162–294 · score 0.90 · cumulative sum curves, median module, knee point, regulator target adjacency, pruning step, intramodular connectivity
- [5] § Methods › Quantification of cross-species divergence ↔ R/plotConservedDivergedModules.R, lines 1–60 · score 0.87 · lower boundary, prediction interval, linear model, jackknife versions, subtree length, module tree
- [6] § Methods › Target gene contribution ↔ R/findConservedDivergedTargets.R, lines 1–61 · score 0.87 · jackknifed module version, diverged target genes, target gene contribution, linear model, subtree length, original module
- [7] § Methods › Quantification of cross-species divergence ↔ R/findConservedDivergedModules.R, lines 1–72 · score 0.87 · externally studentized residuals, Benjamini Hochberg, discovery rate, robust subset, diverged modules, FDR
- [8] § Methods › POU5F1 perturbations using CRISPRi ↔ 2.validations/2.8.POU5F1_CRISPRi/2.8.9.stemness_scores_and_knockdown_efficiencies.R, lines 1–40 · score 0.84 · knockdown efficiencies, gRNAs, POU5F1 perturbed, control cells, validation, species
- [9] § Results › Lineage-specific divergence ↔ vignettes/CroCoNet.Rmd, lines 1032–1052 · score 0.83 · recent common ancestor, human MRCA, minimal requirement, branch leading, assess module divergence, human replicates form
- [10] § Methods › Quantification of cross-species divergence ↔ R/data.R, lines 482–563 · score 0.83 · externally studentized residuals, Benjamini Hochberg, discovery rate, FDR, POU5F1, cutoff
- [11] § Methods › Module assignment ↔ R/pruneModules.R, lines 162–294 · score 0.83 · cumulative sum curve, knee point, Pruning steps, pruned module, initial module, iterative
- [12] § Results › Applying CroCoNet to primate scRNA-seq data ↔ R/pruneModules.R, lines 64–159 · score 0.82 · cumulative sum curve, knee point, regulator target adjacency, intramodular connectivity, edge weights, pruned modules
- [13] § Results › Quantifying module divergence ↔ 4.paper_figures_and_tables/figure4.R, lines 83–162 · score 0.81 · negatively correlated POU5F1, stemness score, positively correlated, POU5F1 module member, network genes, module genes
- [14] § Results › Quantifying module divergence ↔ R/plotTreeStats.R, lines 112–257 · score 0.81 · linear regression model, neighbor joining trees, branch lengths, trees represent, module topology, preservation scores
- [15] § Methods › Network inference ↔ vignettes/CroCoNet.Rmd, lines 113–145 · score 0.80 · network inference, GRNBoost2, inferred networks, network versions, importance scores, potential regulators
- [16] § Methods › Inferring binding potential and binding site divergence of the regulators ↔ 2.validations/2.3.binding_site_enrichment_and_divergence/2.3.4.ATAC_seq_mapping.sh, the whole file · a weak match · score 0.79 · NPC samples, gorGor6, human NPC, ATAC seq, BWA, MEM2
- [17] § Methods › Inferring binding potential and binding site divergence of the regulators ↔ 2.validations/2.3.binding_site_enrichment_and_divergence/2.3.18.associate_peaks_to_gene.R, lines 82–118 · score 0.79 · gorGor6, cynomolgus NPC, gorilla NPCs, human NPC, LiftOver, ATAC seq
- [18] § Results › Applying CroCoNet to primate scRNA-seq data ↔ R/data.R, lines 41–93 · score 0.76 · neural progenitor cells, GRNBoost2, scRNA, cynomolgus macaque, day, trajectory
- [19] § Results › Preservation statistics to compare module topologies ↔ R/comparePresStats.R, lines 1–39 · score 0.75 · fine grained topology, intramodular connectivities, module topologies, cor adj, CroCoNet, adjacencies
- [20] § Results › Preservation statistics to compare module topologies ↔ R/calculatePresStats.R, lines 1–44 · score 0.74 · joint module assignment, preserved module, intramodular connectivities, cor adj, consensus network, WGCNA
- [21] § Results › Rewiring of the POU5F1 module ↔ 2.validations/2.7.POU5F1_LTR7_enrichment/2.7.3.calculate_enrichment_near_POU5F1_module_members.R, lines 50–96 · score 0.72 · near POU5F1 module, POU5F1 module member, LTR7 elements, iPSCs, genes expressed, bound
- [22] § Methods › Module preservation, tree reconstruction and module filtering ↔ R/filterModuleTrees.R, the whole file · a weak match · score 0.71 · probability density functions, bivariate normal distributions, random modules, tree, filtering
- [23] § Methods › Experimental procedures and data processing ↔ R/data.R, lines 96–137 · score 0.71 · gorGor6, macFas6, cynomolgus macaque, transferred, Liftoff, GENCODE
- [24] § Methods › Module preservation, tree reconstruction and module filtering ↔ R/plotTreeStats.R, lines 112–257 · score 0.71 · neighbor joining tree, chosen statistic, subtree length, preservation scores, species diversity, reconstruction
- [25] § Results › Quantifying module divergence ↔ vignettes/CroCoNet.Rmd, lines 927–950 · score 0.69 · positive score indicates, target gene weakens, negative score, divergence score, Target gene contribution, conserved modules
- [26] § Methods › Inferring binding potential and binding site divergence of the regulators ↔ 2.validations/2.3.binding_site_enrichment_and_divergence/2.3.21.summarize_motif_scores_per_gene.Rmd, lines 1157–1196 · score 0.69 · peaks associated, motif scores, gene associations, ATAC seq, Cluster, random modules
- [27] § Methods › Inferring binding potential and binding site divergence of the regulators ↔ 4.paper_figures_and_tables/suppl.tables.R, lines 128–169 · score 0.68 · binding site divergence, human cynomolgus, human gorilla, species pair, median, conserved
- [28] § Methods › Network inference ↔ 1.neural_differentiation_dataset/1.2.network_inference/1.2.1.prepare_data.R, the whole file · a weak match · score 0.67 · randomized quantile residual, network inference, transformations, matrix, regulators, genes
- [29] § Methods › Module preservation, tree reconstruction and module filtering ↔ R/data.R, lines 344–387 · score 0.66 · neighbor joining tree, distance matrix, subtree length, preservation scores, species diversity, infer
- [30] § Methods › Inferring binding potential and binding site divergence of the regulators ↔ 2.validations/2.7.POU5F1_LTR7_enrichment/2.7.3.calculate_enrichment_near_POU5F1_module_members.R, lines 1–48 · score 0.66 · ATAC seq peaks, iPSCs, gene associations, kb, Active, validation
- [31] § Methods › Network inference ↔ R/addDirectionality.R, the whole file · a weak match · score 0.66 · correlation coefficients, expression matrices, gene pairs, Spearman, inference, cells
- [32] § Methods › Module preservation, tree reconstruction and module filtering ↔ R/data.R, lines 390–435 · score 0.66 · cor.adj, intramodular connectivities, cor.kIM, module preservation, correlation, quantified
- [33] § Methods › Experimental procedures and data processing ↔ 3.brain_dataset/3.1.data_preparation/3.1.2.filtering.R, lines 45–85 · score 0.66 · H18.30.001, atypical cell, donor, composition, metadata, brain
- [34] § Results › Rewiring of the POU5F1 module ↔ 4.paper_figures_and_tables/suppl.figureS12_S13.R, lines 161–229 · score 0.65 · POU5F1 binding, LTR7 elements, binding sites, S12, S13, HERVH
- [35] § Results › Preservation statistics to compare module topologies ↔ R/comparePresStats.R, lines 1–39 · score 0.65 · fine grained, phylogenetic information, cor adj, CroCoNet, signal, topologies
- [36] § Methods › Cell lines ↔ 1.neural_differentiation_dataset/1.3.CroCoNet_analysis/1.3.1.CroCoNet_analysis.R, lines 43–81 · score 0.62 · Homo sapiens, Macaca fascicularis, cynomolgus, gorilla, human
- [37] § Results › Preservation statistics to compare module topologies ↔ 4.paper_figures_and_tables/suppl.figureS6.R, the whole file · a weak match · score 0.60 · cor.adj, phylogenetic information, cor.kIM, S6, inference, neural
- [38] § Methods › POU5F1 perturbations using CRISPRi ↔ 2.validations/2.8.POU5F1_CRISPRi/2.8.8.QC_and_filtering.R, lines 258–334 · score 0.60 · dCas9, gRNAs, POU5F1, cynomolgus, cell, human
- [39] § Results › Overview of the CroCoNet pipeline ↔ R/reconstructTrees.R, lines 101–184 · score 0.59 · neighbor joining tree, branch lengths, distance matrix, preservation score, CroCoNet, pipeline
- [40] § Results › Consensus network and module assignment ↔ 4.paper_figures_and_tables/figure2.R, lines 322–365 · score 0.59 · PAX6 ChIP seq, ChIP seq enrichment, NANOG, fraction, motifs, peak
Paper
Loaded from Europe PMC by your browser, not stored by OSCR: doi.org · Europe PMC
The paper is loaded when this pane is shown.
The authors' code
R · 563 lines · 49 KB · GPL-3.0 · 6 matches
- #' Transcriptional regulators
- #'
- #' 7 important transcriptional regulators during early neural differentiation that are used as cores to assemble modules around them.
- #'
- #' @format A character vector of 7 elements.
- "regulators"
- #' All genes
- #'
- #' The 300 genes that are used for the network inference.
- #'
- #' @format A character vector of 300 elements.
- "genes"
- #' JASPAR 2024 vertebrate core transcriptional regulators
- #'
- #' Transcriptional regulators that have at least 1 annotated motif in the JASPAR 2024 vertebrate core collection.
- #'
- #' @format A character vector of 784 elements.
- "jaspar_core_TRs"
- #' JASPAR 2024 unvalidated transcriptional regulators
- #'
- #' Transcriptional regulators that have at least 1 annotated motif in the JASPAR 2024 unvalidated collection.
- #'
- #' @format A character vector of 590 elements.
- "jaspar_unvalidated_TRs"
- #' IMAGE transcriptional regulators
- #'
- #' Transcriptional regulators that have at least 1 annotated motif in the IMAGE database (Madsen et al. 2018).
- #'
- #' @format A character vector of 1351 elements.
- "image_TRs"
- #' Replicate-species conversion
- #'
- #' A data frame that specifies which replicate belongs to which species.
- #'
- #' @format A data frame with 9 rows and 2 columns:
- #' \describe{
- #' \item{replicate}{Name of the replicate/cell line.}
- #' \item{species}{Name of the species.}
- #' }
- "replicate2species"
- #' SCE object of the primate neural differentiation dataset
- #'
- #' A subset of the primate neural differentiation scRNA-seq dataset in an SCE format. The data was collected during the early neural differentiation of human, gorilla and cynomolgus macaque iPS cells with 3 human, 2 gorilla and 4 cynomolgus cell lines (replicates). Using a directed differentiation protocol, cells were differentiated into neural progenitor cells (NPCs) over the course of 9 days, and scRNA-seq data was obtained at six time points (days 0, 1, 3, 5, 7 and 9) during this process. The SCE object contains the raw and log-normalized counts as well as the metadata for 300 genes and 900 cells (100 cells per replicate).
- #'
- #' @format An SCE object with 300 rows, 900 columns, 9 metadata columns and 2 assays.
- #'
- #' Metadata columns:
- #' \describe{
- #' \item{species}{Name of the species.}
- #' \item{replicate}{Name of the replicate/cell line.}
- #' \item{day}{The day when the cell was collected.}
- #' \item{n_UMIs}{Number of UMIs detected.}
- #' \item{n_genes}{Number of genes detected.}
- #' \item{perc_mito}{Percent of mitochondrial reads.}
- #' \item{sizeFactor}{Size factor for scaling normalization, calculated first per replicate using [scran::computeSumFactors] and [scran::quickCluster], then adjusted by [batchelor::multiBatchNorm] to remove systematic differences in covergae across replicates.}
- #' \item{pseudotime}{Pseudotime inferred by [SCORPIUS::infer_trajectory].}
- #' \item{cell_type}{Cell type labels predicted by [SingleR::classifySingleR] using the embryoid body dataset from Rhodes et al. 2022 as reference.}
- #' }
- #' Assays:
- #' \describe{
- #' \item{counts}{Raw counts.}
- #' \item{logcounts}{Log-normalized counts created by first calculating size factors per replicate using [scran::computeSumFactors] and [scran::quickCluster], then adjusted them to remove systematic differences in covergae across replicates and log-normalizing by [batchelor::multiBatchNorm].}
- #' }
- "sce"
- #' List of raw networks
- #'
- #' List of networks per replicate with raw edge weights. The networks were inferred using GRNBoost2 based on a subset of the primate neural differentiation scRNA-seq dataset. All 300 genes in the subsetted data were used as potential regulators. To circumvent the stochastic nature of the algorithm, GRNBoost2 was run 10 times on the same count matrices, then the results were averaged across runs, and rarely occurring edges were removed altogether. In addition, edges inferred between the same gene pair but in opposite directions were also averaged.
- #'
- #' @format A named list of 9 [igraph] objects. Each network contains 300 nodes, and has 1 node and 2 edge attributes:
- #'
- #' Node attributes:
- #' \describe{
- #' \item{name}{Name of the node (gene).}
- #' }
- #' Edge attributes:
- #' \describe{
- #' \item{weight}{Edge weight, the importance score calculate by GRNBoost2.}
- #' }
- "network_list_raw"
- #' List of genomic annotations
- #'
- #' List of genomic annotations per species as GRanges objects. The human annotation is the GTF of Hg38 GENCODE release 32 primary assembly, while the gorilla and cynomolgus macaque annotations were created by transferring the human annotation onto the gorGor6 and macFas6 genomes via the tool Liftoff (https://github.com/agshumate/Liftoff). Each annotation was subsetted for the 300 genes that feature in this example dataset.
- #'
- #' @format A named list of 3 GRanges objects. Each object has 8 columns.
- #'
- #' Attributes:
- #' \describe{
- #' \item{seqnames}{Chromosome or contig name.}
- #' \item{start}{Genomic start location.}
- #' \item{end}{Genomic end location.}
- #' \item{width}{Width of the feature in base pairs.}
- #' \item{strand}{Genomic strand ("+" or "-").}
- #' \item{source}{The prediction program or public database where the annotations came from.}
- #' \item{type}{Feature type (gene, transcript, exon, CDS, UTR, start_codon, stop_codon or Selenocysteine).}
- #' \item{score}{The degree of confidence in the feature's existence and coordinates.}
- #' \item{phase}{One of '0', '1' or '2'. '0' means that the first base of the feature is the first base of a codon, '1' that the second base is the first base of a codon, and so on.}
- #' \item{gene_id}{Unique identifier of the gene.}
- #' \item{gene_name}{Name of the gene.}
- #' \item{transcript_id}{Unique identifier of the transcript.}
- #' \item{transcript_name}{Name of the transcript.}
- #' }
- #' @source <ftp.ebi.ac.uk/pub/databases/gencode/Gencode_human/release_32/gencode.v32.primary_assembly.annotation.gtf.gz>, <https://hgdownload.soe.ucsc.edu/goldenPath/gorGor6/bigZips/gorGor6.fa.gz>, <https://ftp.ensembl.org/pub/release-109/fasta/macaca_fascicularis/dna/Macaca_fascicularis.Macaca_fascicularis_6.0.dna_sm.toplevel.fa.gz>
- "gtf_list"
- #' List of networks
- #'
- #' List of networks per replicate with edge weights re-scaled between 0 and 1, and gene pairs (edges) with overlapping annotations removed. The networks were inferred using GRNBoost2 based on a subset of the primate neural differentiation scRNA-seq dataset. All 300 genes in the subsetted data were used as potential regulators. To circumvent the stochastic nature of the algorithm, GRNBoost2 was run 10 times on the same count matrices, then the results were averaged across runs, and rarely occurring edges were removed altogether. In addition, edges inferred between the same gene pair but in opposite directions were also averaged. Edge weights were scaled by the maximum edge weight across all replicates. Gene pairs that have overlapping annotations in any of the species' genomes were removed from all networks.
- #'
- #' @format A named list of 9 [igraph] objects. Each network contains 300 nodes, and has 1 node attribute and 3 edge attributes:
- #'
- #' Node attributes:
- #' \describe{
- #' \item{name}{Name of the node (gene).}
- #' }
- #' Edge attributes:
- #' \describe{
- #' \item{weight}{Edge weight, the importance score calculate by GRNBoost2 rescaled between 0 and 1.}
- #' \item{genomic_dist}{Numeric, the genomic distance of the 2 genes that form the edge (Inf if the 2 genes are annotated on different chromosomes/contigs).}
- #' }
- "network_list"
- #' Phylogenetic tree
- #'
- #' Rooted phylogenetic tree of 3 primate species: human (Homo sapiens), gorilla (Gorilla gorilla) and cynomolgus macaque (Macaca Fascicularis). The tree was created by subsetting the mammalian tree with the best estimates of branch lengths from Bininda-Edmons et al. 2007.
- #'
- #' @format A [phylo] object with 4 edges and 3 nodes.
- #' @source Olaf R. P. Bininda-Emonds, Marcel Cardillo, Kate E. Jones, Ross D. E. MacPhee, Robin M. D. Beck, Richard Grenyer, Samantha A. Price, Rutger A. Vos, John L. Gittleman & Andy Purvis. "The delayed rise of present-day mammals" Nature 446, 507-512(29 March 2007). doi:10.1038/nature05634
- "tree"
- #' Consensus network
- #'
- #' Consensus network of the 9 primate replicates in the example dataset. For each edge, the consensus adjacency was calculated as the weighted average of replicate-wise adjacencies using weights that correct for 1) the phylogenetic distances between species and 2) the different numbers of replicates per species. If an edge was not detected in certain replicate, the adjacency of that replicate was regarded as 0 for the calculation of the consensus. The directionality of each edge was determined based on a modified Spearman's correlation between the corresponding 2 genes' expression profiles (positive expression correlation - activating interaction, negative expression correlation - repressing interaction). The correlations were calculated per replicate, then the mean correlation was taken across all replicates.
- #'
- #' @format An [igraph] object with 300 nodes, and has 1 node attribute and 3 edge attributes:
- #'
- #' Node attributes:
- #' \describe{
- #' \item{name}{Name of the node (gene).}
- #' }
- #' Edge attributes:
- #' \describe{
- #' \item{weight}{Consensus edge weight/adjacency, the weighted average of replicate-wise adjacencies.}
- #' \item{rho}{Approximate Spearman's correlation coefficient of the 2 genes' expression profiles that form the edge.}
- #' \item{p.adj}{BH-corrected approximate p-value of rho.}
- #' \item{direction}{Direction of the interaction between the 2 genes that form the edge ("+" or "-").}
- #' }
- "consensus_network"
- #' Initial modules
- #'
- #' Initial modules created by assigning the top 250 targets to each of 7 transcriptional regulators involved in the early neuronal differentiation of primates (\code{regulators}). For each regulator, the top 250 targets were selected based edge weights in the consensus network: all targets of the regulator were ranked based on their edge weight to the regulator (regulator-taregt adjacency) and the 250 targets with the highest regulator-target adjacencies were kept.
- #'
- #' @format A data frame with 1750 rows and 8 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{target}{Target gene of the transcriptional regulator (member of the regulator's initial module).}
- #' \item{weight}{Consensus edge weight/adjacency, the weighted average of replicate-wise adjacencies.}
- #' \item{rho}{Approximate Spearman's correlation coefficient of the 2 genes' expression profiles that form the edge.}
- #' \item{p.adj}{BH-corrected approximate p-value of rho.}
- #' \item{direction}{Direction of the interaction between the 2 genes that form the edge ("+" or "-").}
- #' }
- "initial_modules"
- #' Pruned modules
- #'
- #' Pruned modules created by the dynamic pruning of the initial modules via the "UIK_adj_kIM" method. The initial module members were filtered in successive steps based on their adjacency to the regulator and their intramodular connectivitiy alternately. In each step, the cumulative sum curve based on one of these two characteristics was calculated per module, then the targets below the knee point of the curve were kept. This process was continued until the module sizes became as small as possible without the median falling below a pre-defined minimum of 20 genes.
- #'
- #' @format A data frame with 225 rows and 9 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{target}{Character, target gene of the transcriptional regulator (member of the regulator's pruned module).}
- #' \item{weight}{Numeric, consensus edge weight/adjacency, the weighted average of replicate-wise adjacencies.}
- #' \item{rho}{Approximate Spearman's correlation coefficient of the 2 genes' expression profiles that form the edge.}
- #' \item{p.adj}{BH-corrected approximate p-value of rho.}
- #' \item{direction}{Character specifying the direction of regulation between the regulator and the target, either "+" or "-".}
- #' \item{module_size}{Module size, the numer of target genes assigned to a regulator.}
- #' }
- "pruned_modules"
- #' Random modules
- #'
- #' Random modules that match the actual (pruned) modules in size. These random modules have the same regulators and contain the same number of target genes as the actual modules, but the target genes were randomly drawn from all genes in the network.
- #'
- #' @format A data frame with 225 rows and 3 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{target}{Member of the regulator's random module.}
- #' \item{module_size}{Module size, the numer of target genes assigned to a regulator.}
- #' }
- "random_modules"
- #' Eigengenes
- #'
- #' Eigengenes of the pruned modules calculated using the activated targets in each module. An eigengene summarizes the expression profile of the module as a whole, mathematically it is the first principal component of the module expression data (i.e. the scaled and centered logcounts subsetted for the activated targets and the regulators of the given module). In this case, the principal component was taken across all cells irrespective of species.
- #'
- #' @format A data frame with 6300 rows and 8 columns:
- #'\describe{
- #' \item{cell}{Character, the cell barcode.}
- #' \item{species}{Character, the name of the species.}
- #' \item{pseudotime}{Numeric, inferred pseudotime.}
- #' \item{cell_type}{Character, cell type annotation.}
- #' \item{module}{Character, transcriptional regulator and direction of regulation (in this case always nameOfRegulator(+)).}
- #' \item{eigengene}{Numeric, the eigengene (i.e. the first principal component of the scaled and centered logcounts) of the module.}
- #' \item{mean_expr}{Numeric, the mean of the scaled and centered logcounts across all genes in the module.}
- #' \item{regulator_expr}{Numeric, the scaled and centered logcounts of the regulator.}
- #' }
- "eigengenes"
- #' Eigengenes per species
- #'
- #' Eigengenes of the pruned modules calculated per species using the activated targets in each module. An eigengene summarizes the expression profile of the module as a whole, mathematically it is the first principal component of the module expression data (i.e. the scaled and centered logcounts subsetted for the activated targets and the regulator of the given module). In this case, the principal component was calculated for each species separately.
- #'
- #' @format A data frame with 6300 rows and 8 columns:
- #'\describe{
- #' \item{cell}{Character, the cell barcode.}
- #' \item{species}{Character, the name of the species.}
- #' \item{pseudotime}{Numeric, inferred pseudotime.}
- #' \item{cell_type}{Character, cell type annotation.}
- #' \item{module}{Character, transcriptional regulator and direction of regulation (in this case always nameOfRegulator(+)).}
- #' \item{eigengene}{Numeric, the eigengene (i.e. the first principal component of the scaled and centered logcounts) of the module.}
- #' \item{mean_expr}{Numeric, the mean of the scaled and centered logcounts across all genes in the module.}
- #' \item{regulator_expr}{Numeric, the scaled and centered logcounts of the regulator.}
- #' }
- "eigengenes_per_species"
- #' Preservation statistics of the original and jackknifed pruned modules
- #'
- #' Correlation of intramodular connectivities (cor_kIM) per replicate pair for the original and all jackknifed versions of the pruned modules. The jackknifed versions of the modules were created by removing each target gene assigned to a module (the regulators were never excluded). The preservation statistic cor.kIM was then calculated for the original module as well as each jackknife module version by comparing each replicate to all others, both within and across species. cor.kIM quantifies how well the connectivity patterns are preserved between the networks of two replicates, mathematically it is the correlation of the intramodular connectivities per module member gene in the network of the 1st replicate VS the intramodular connectivities per module member gene in the network of the 2nd replicate.
- #'
- #' @format A data frame with 8352 rows and 8 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Integer, the numer of target genes assigned to a regulator.}
- #' \item{type}{Character, module type (orig = original or jk = jackknifed).}
- #' \item{id}{Character, the unique ID of the module version (format: nameOfRegulator_jk_nameOfGeneRemoved in case of module type 'jk' and nameOfRegulator_orig in case of module type 'orig').}
- #' \item{gene_removed}{Character, the name of the gene removed by jackknifing (NA in case of module type 'orig').}
- #' \item{replicate1, replicate2}{Character, the names of the replicates compared.}
- #' \item{species1, species2}{Character, the names of the species \code{replicate1} and \code{replicate2} belong to, respectively.}
- #' \item{cor_adj}{Numeric, correlation of adjacencies.}
- #' \item{cor_kIM}{Numeric, correlation of intramodular connectivities.}
- #' }
- "pres_stats_jk"
- #' Preservation statistics of the the original and jackknifed random modules
- #'
- #' Correlation of intramodular connectivities (cor_kIM) per replicate pair for the original and all jackknifed versions of the random modules. The jackknifed versions of the modules were created by removing each target gene assigned to a module (the regulators were never excluded). The preservation statistic cor.kIM was then calculated for the original module as well as each jackknife module version by comparing each replicate to all others, both within and across species. cor.kIM quantifies how well the connectivity patterns are preserved between the networks of two replicates, mathematically it is the correlation of the intramodular connectivities per module member gene in the network of the 1st replicate VS the intramodular connectivities per module member gene in the network of the 2nd replicate.
- #'
- #' @format A data frame with 8352 rows and 8 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Integer, the numer of target genes assigned to a regulator.}
- #' \item{type}{Character, module type (orig = original or jk = jackknifed).}
- #' \item{id}{Character, the unique ID of the module version (format: nameOfRegulator_jk_nameOfGeneRemoved in case of module type 'jk' and nameOfRegulator_orig in case of module type 'orig').}
- #' \item{gene_removed}{Character, the name of the gene removed by jackknifing (NA in case of module type 'orig').}
- #' \item{replicate1, replicate2}{Character the names of the replicates compared.}
- #' \item{species1, species2}{Character, the names of the species \code{replicate1} and \code{replicate2} belong to, respectively.}
- #' \item{cor_adj}{Numeric, correlation of adjacencies.}
- #' \item{cor_kIM}{Numeric, correlation of intramodular connectivities.}
- #' }
- "random_pres_stats_jk"
- #' Distance measures of the original and jackknifed pruned modules
- #'
- #' Distance measures per replicate pair for the original and all jackknifed versions of the pruned modules. The jackknifed versions of the modules were created by removing each target gene assigned to a module (the regulators were never excluded). The distance measures were calculated based on the correlation of intramodular connectivities: \eqn{dist = \frac{1 - cor.kIM}{2}} (a correlation of 1 corresponds to a distance of 0, whereas a correlation of -1 corresponds to a distance of 1). Each element of the list corresponds a jackknifed/original module version and contains the distance measures between all possible pairs of replicates for this module version.
- #'
- #' @format A named list with 232 elements containing the distance measures per (original or jackknifed) module version. Each element is a data frame with 72 rows and 10 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Integer, the numer of target genes assigned to a regulator.}
- #' \item{type}{Character, module type (orig = original or jk = jackknifed).}
- #' \item{id}{Character, the unique ID of the module version (format: nameOfRegulator_jk_nameOfGeneRemoved in case of module type 'jk' and nameOfRegulator_orig in case of module type 'orig').}
- #' \item{gene_removed}{Character, the name of the gene removed by jackknifing (NA in case of module type 'orig').}
- #' \item{replicate1, replicate2}{Character the names of the replicates compared.}
- #' \item{species1, species2}{Character, the names of the species \code{replicate1} and \code{replicate2} belong to, respectively.}
- #' \item{dist}{Numeric, distance measure ranging from 0 to 1, calculated based on the correlation of intramodular connectivities.}
- #' }
- "dist_jk"
- #' Distance measures of the pruned modules
- #'
- #' Distance measures per replicate pair for the original (i.e. not jackknifed) pruned modules.
- #'
- #' The distance measures were calculated based on the correlation of intramodular connectivities: \deqn{dist = \frac{1 - cor.kIM}{2}} (a correlation of 1 corresponds to a distance of 0, whereas a correlation of -1 corresponds to a distance of 1). Each element of the list corresponds to a module and contains the distance measures between all possible pairs of replicates for this module.
- #'
- #' @format A named list with 12 elements containing the distance measures per module. Each element is a data frame with 72 rows and 7 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Integer, the numer of target genes assigned to a regulator.}
- #' \item{replicate1, replicate2}{Character the names of the replicates compared.}
- #' \item{species1, species2}{Character, the names of the species \code{replicate1} and \code{replicate2} belong to, respectively.}
- #' \item{dist}{Numeric, distance measure ranging from 0 to 1, calculated based on the correlation of intramodular connectivities.}
- #' }
- "dist"
- #' Distance measures of the the original and jackknifed random modules
- #'
- #' Distance measures per replicate pair for the original and all jackknifed versions of the random modules.
- #'
- #' The jackknifed versions of the modules were created by removing each target gene assigned to a module (the regulators were never excluded). The distance measures were calculated based on the correlation of intramodular connectivities: \deqn{dist = \frac{1 - cor.kIM}{2}} (a correlation of 1 corresponds to a distance of 0, whereas a correlation of -1 corresponds to a distance of 1). Each element of the list corresponds a jackknife module version and contains the distance measures between all possible pairs of replicates for this module version.
- #'
- #' @format A named list with 232 elements containing the distance measures per (original or jackknifed) module version. Each element is a data frame with 72 rows and 10 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Integer, the numer of target genes assigned to a regulator.}
- #' \item{type}{Character, module type (orig = original or jk = jackknifed).}
- #' \item{id}{Character, the unique ID of the module version (format: nameOfRegulator_jk_nameOfGeneRemoved in case of module type 'jk' and nameOfRegulator_orig in case of module type 'orig').}
- #' \item{gene_removed}{Character, the name of the gene removed by jackknifing (NA in case of module type 'orig').}
- #' \item{replicate1, replicate2}{Character the names of the replicates compared.}
- #' \item{species1, species2}{Character, the names of the species \code{replicate1} and \code{replicate2} belong to, respectively.}
- #' \item{dist}{Numeric, distance measure ranging from 0 to 1, calculated based on the correlation of intramodular connectivities.}
- #' }
- "random_dist_jk"
- #' Trees of the original and jackknifed pruned modules
- #'
- #' Neighbor-joining trees representing the similarities of connectivity patterns across the 9 primate replicates for the original and all jackknifed versions of the pruned modules. The jackknifed versions of the modules were created by removing each target gene assigned to a module (the regulators were never excluded). For each of these module versions, the trees were inferred based on the preservation statistic cor.kIM (correlation of intramodular connectivities): first, the preservation scores were calculated between all possible replicate pairs, then they were converted into a distance matrix of replicates, and finally trees were reconstructed based on this distance matrix using the neighbor-joining algorithm. The result is a single tree for each original or jackknife module where the tips represent the replicates and the branch lengths represent the dissimilarity of connectivity patterns between these replicates.
- #' @format A named list with 232 elements containing the neighbor-joining trees as [phylo] objects.
- "trees_jk"
- #' Trees of the pruned modules
- #'
- #' Neighbor-joining trees representing the similarities of connectivity patterns across the 9 primate replicates for the original (i.e. not jackknifed) pruned modules. For each of the modules, the trees were inferred based on the preservation statistic cor.kIM (correlation of intramodular connectivities): first, the preservation scores were calculated between all possible replicate pairs, then they were converted into a distance matrix of replicates, and finally trees were reconstructed based on this distance matrix using the neighbor-joining algorithm. The result is a single tree per module where the tips represent the replicates and the branch lengths represent the dissimilarity of connectivity patterns between these replicates.
- #' @format A named list with 12 elements containing the neighbor-joining trees as [phylo] objects.
- "trees"
- #' Trees of the original and jackknifed random modules
- #'
- #' Neighbor-joining trees representing the similarities of connectivity patterns across the 9 primate replicates for the original and all jackknifed versions of the random modules. The jackknifed versions of the modules were created by removing each target gene assigned to a module (the regulators were never excluded). For each of these module versions, the trees were inferred based on the preservation statistic cor.kIM (correlation of intramodular connectivities): first, the preservation scores were calculated between all possible replicate pairs, then they were converted into a distance matrix of replicates, and finally trees were reconstructed based on this distance matrix using the neighbor-joining algorithm. The result is a single tree for each original or jackknife module where the tips represent the replicates and the branch lengths represent the dissimilarity of connectivity patterns between these replicates.
- #' @format A named list with 232 elements containing the neighbor-joining trees as [phylo] objects.
- "random_trees_jk"
- #' Tree-based statistics of the original and jackknifed pruned modules
- #'
- #' Tree-based statistics calculated for the original and all jackknifed versions of the pruned modules. For each module version, a tree was reconstructed based on the module preservation scores within and across species. The tips of resulting tree represent the replicates and the branch lengths represent the dissimilarity of connectivity patterns between the replicates. Branches between replicates of different species carry information about cross-species differences, while the branches between replicates of the same species carry information about the within-species diversity.
- #' @format A data frame with 232 rows and 16 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Module size, the numer of target genes assigned to a regulator.}
- #' \item{type}{Character, module type (orig = original or jk = jackknifed).}
- #' \item{id}{Character, the unique ID of the module version (format: nameOfRegulator_jk_nameOfGeneRemoved in case of module type 'jk' and nameOfRegulator_orig in case of module type 'orig').}
- #' \item{gene_removed}{The name of the gene removed by jackknifing (NA in case of module type 'orig').}
- #' \item{within_species_diversity}{Numeric, the sum of the human, gorilla and cynomolgus diversity.}
- #' \item{total_tree_length}{Numeric, the sum of all branch lengths in the tree.}
- #' \item{human_diversity}{Numeric, the sum of the branch lengths in the subtree that contains only the human tips.}
- #' \item{gorilla_diversity}{Numeric, the sum of the branch lengths in the subtree that contains only the gorilla tips.}
- #' \item{cynomolgus_diversity}{Numeric, the sum of the branch lengths in the subtree that contains only the cynomolgus tips.}
- #' \item{human_monophyl}{Logical indicating whether the tree is monophyletic for the human replicates.}
- #' \item{gorilla_monophyl}{Logical indicating whether the tree is monophyletic for the gorilla replicates.}
- #' \item{cynomolgus_monophyl}{Logical indicating whether the tree is monophyletic for the cynomolgus replicates.}
- #' \item{human_subtree_length}{Numeric, the sum of the branch lengths in the subtree that is defined by the human replicates and includes the internal branch connecting the human replicates to the rest of the tree. NA if the tree is not monophyletic for the human replicates.}
- #' \item{gorilla_subtree_length}{Numeric, the sum of the branch lengths in the subtree that is defined by the gorilla replicates and includes the internal branch connecting the gorilla replicates to the rest of the tree. NA if the tree is not monophyletic for the gorilla replicates.}
- #' \item{cynomolgus_subtree_length}{Numeric, the sum of the branch lengths in the subtree that is defined by the cynomolgus replicates and includes the internal branch connecting the cynomolgus replicates to the rest of the tree. NA if the tree is not monophyletic for the cynomolgus replicates.}
- #' }
- "tree_stats_jk"
- #' Tree-based statistics of the original and jackknifed random modules
- #'
- #' Tree-based statistics calculated for the original and all jackknifed versions of the random modules. For each module version, a tree was reconstructed based on the module preservation scores within and across species. The tips of resulting tree represent the replicates and the branch lengths represent the dissimilarity of connectivity patterns between the replicates. Branches between replicates of different species carry information about cross-species differences, while the branches between replicates of the same species carry information about the within-species diversity.
- #' @format A data frame with 232 rows and 16 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Module size, the numer of target genes assigned to a regulator.}
- #' \item{type}{Character, module type (orig = original or jk = jackknifed).}
- #' \item{id}{Character, the unique ID of the module version (format: nameOfRegulator_jk_nameOfGeneRemoved in case of module type 'jk' and nameOfRegulator_orig in case of module type 'orig').}
- #' \item{gene_removed}{The name of the gene removed by jackknifing (NA in case of module type 'orig').}
- #' \item{within_species_diversity}{Numeric, the sum of the human, gorilla and cynomolgus diversity.}
- #' \item{total_tree_length}{Numeric, the sum of all branch lengths in the tree.}
- #' \item{human_diversity}{Numeric, the sum of the branch lengths in the subtree that contains only the human tips.}
- #' \item{gorilla_diversity}{Numeric, the sum of the branch lengths in the subtree that contains only the gorilla tips.}
- #' \item{cynomolgus_diversity}{Numeric, the sum of the branch lengths in the subtree that contains only the cynomolgus tips.}
- #' \item{human_monophyl}{Logical indicating whether the tree is monophyletic for the human replicates.}
- #' \item{gorilla_monophyl}{Logical indicating whether the tree is monophyletic for the gorilla replicates.}
- #' \item{cynomolgus_monophyl}{Logical indicating whether the tree is monophyletic for the cynomolgus replicates.}
- #' \item{human_subtree_length}{Numeric, the sum of the branch lengths in the subtree that is defined by the human replicates and includes the internal branch connecting the human replicates to the rest of the tree. NA if the tree is not monophyletic for the human replicates.}
- #' \item{gorilla_subtree_length}{Numeric, the sum of the branch lengths in the subtree that is defined by the gorilla replicates and includes the internal branch connecting the gorilla replicates to the rest of the tree. NA if the tree is not monophyletic for the gorilla replicates.}
- #' \item{cynomolgus_subtree_length}{Numeric, the sum of the branch lengths in the subtree that is defined by the cynomolgus replicates and includes the internal branch connecting the cynomolgus replicates to the rest of the tree. NA if the tree is not monophyletic for the cynomolgus replicates.}
- #' }
- "random_tree_stats_jk"
- #' Preservation statistics of the pruned modules
- #'
- #' Correlation of intramodular connectivities (cor_kIM) per replicate pair and module. The preservation statistic cor.kIM quantifies how well the connectivity patterns are preserved between the networks of two replicates, mathematically it is the correlation of the intramodular connectivities per module member gene in the network of the 1st replicate VS the intramodular connectivities per module member gene in the network of the 2nd replicate.
- #' This statistic was calculated for all possible jackknifed versions of the modules, each of which was created by removing a target gene assigned to the given module (the regulators were never excluded). Each jackknifed module was compared between all posible pairs of replicates, both within and across species, resulting in a cor.kIM value per jackknifed module version and replicate pair. Finally, the cor.kIM values were summarized per module and replicate pair by taking the median and its 95\% confidence interval across all jackknifed module versions.
- #'
- #' @format A data frame with 252 rows and 10 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Module size, the numer of target genes assigned to a regulator.}
- #' \item{replicate1, replicate2}{The names of the replicates compared.}
- #' \item{species1, species2}{The names of the species 'replicate1' and 'replicate2' belongs to, respectively.}
- #' \item{cor_kIM}{The median of cor.kIM across all jackknifed versions of the module.}
- #' \item{var_cor_kIM}{The variance of cor.kIM across all jackknifed versions of the module.}
- #' \item{lwr_cor_kIM}{The lower bound of the 95\% confidence interval of cor.kIM calculated by jackknifing.}
- #' \item{upr_cor_kIM}{The upper bound of the 95\% confidence interval of cor.kIM calculated by jackknifing.}
- #' \item{cor_adj}{The median of cor.adj across all jackknifed versions of the module.}
- #' \item{var_cor_adj}{The variance of cor.adj across all jackknifed versions of the module.}
- #' \item{lwr_cor_adj}{The lower bound of the 95\% confidence interval of cor.adj calculated by jackknifing.}
- #' \item{upr_cor_adj}{The upper bound of the 95\% confidence interval of cor.adj calculated by jackknifing.}
- #' }
- "pres_stats"
- #' Preservation statistics of the random modules
- #'
- #' Correlation of intramodular connectivities (cor_kIM) per replicate pair and random module. The preservation statistic cor.kIM quantifies how well the connectivity patterns are preserved between the networks of two replicates, mathematically it is the correlation of the intramodular connectivities per module member gene in the network of the 1st replicate VS the intramodular connectivities per module member gene in the network of the 2nd replicate.
- #' This statistic was calculated for all possible jackknifed versions of the modules, each of which was created by removing a target gene assigned to the given module (the regulators were never excluded). Each jackknifed module was compared between all posible pairs of replicates, both within and across species, resulting in a cor.kIM value per jackknifed module version and replicate pair. Finally, the cor.kIM values were summarized per module and replicate pair by taking the median and its 95\% confidence interval across all jackknifed module versions.
- #'
- #' @format A data frame with 252 rows and 10 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Module size, the numer of target genes assigned to a regulator.}
- #' \item{replicate1, replicate2}{The names of the replicates compared.}
- #' \item{species1, species2}{The names of the species 'replicate1' and 'replicate2' belongs to, respectively.}
- #' \item{cor_kIM}{The median of cor.kIM across all jackknifed versions of the module.}
- #' \item{var_cor_kIM}{The variance of cor.kIM across all jackknifed versions of the module.}
- #' \item{lwr_cor_kIM}{The lower bound of the 95\% confidence interval of cor.kIM calculated by jackknifing.}
- #' \item{upr_cor_kIM}{The upper bound of the 95\% confidence interval of cor.kIM calculated by jackknifing.}
- #' \item{cor_adj}{The median of cor.adj across all jackknifed versions of the module.}
- #' \item{var_cor_adj}{The variance of cor.adj across all jackknifed versions of the module.}
- #' \item{lwr_cor_adj}{The lower bound of the 95\% confidence interval of cor.adj calculated by jackknifing.}
- #' \item{upr_cor_adj}{The upper bound of the 95\% confidence interval of cor.adj calculated by jackknifing.}
- #' }
- "random_pres_stats"
- #' Tree-based statistics of the pruned modules
- #'
- #' Tree-based statistics per module. The trees were reconstructed based on module preservation scores (cor.kIM) within and across species. The tips of resulting tree represent the replicates and the branch lengths represent the dissimilarity of connectivity patterns between the replicates. The total tree length is the sum of all branch lengths in the tree and it measures module variability both within and across species, while the within-species_diversity is the sum of the human, gorilla and cynomolgus within-species branch lengths and it measures module variability within species only.
- #' These statistics were calculated for all possible jackknifed versions of the modules, each of which was created by removing a target gene assigned to the given module (the regulators were never excluded). Then the statistics were summarized per module by taking the median and its 95\% confidence interval across all jackknifed module versions.
- #'
- #' @format A data frame with 7 rows and 10 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Module size, the numer of target genes assigned to a regulator.}
- #' \item{total_tree_length}{The median of total tree lengths across all jackknifed versions of the module.}
- #' \item{var_total_tree_length}{The variance of total tree lengths across all jackknifed versions of the module.}
- #' \item{lwr_total_tree_length}{The lower bound of the 95\% confidence interval of the total tree length calculated by jackknifing.}
- #' \item{upr_total_tree_length}{The upper bound of the 95\% confidence interval of the total tree length calculated by jackknifing.}
- #' \item{within_species_diversity}{The median of within-species_diversity values across all jackknifed versions of the module.}
- #' \item{var_within_species_diversity}{The variance of within-species_diversity values across all jackknifed versions of the module.}
- #' \item{lwr_within_species_diversity}{The lower bound of the 95\% confidence interval of the within-species_diversity calculated by jackknifing.}
- #' \item{upr_within_species_diversity}{The upper bound of the 95\% confidence interval of the within-species_diversity calculated by jackknifing.}
- #' }
- "tree_stats"
- #' Tree-based statistics of the random modules
- #'
- #' Tree-based statistics per random module. The trees were reconstructed based on module preservation scores (cor.kIM) within and across species. The tips of resulting tree represent the replicates and the branch lengths represent the dissimilarity of connectivity patterns between the replicates. The total tree length is the sum of all branch lengths in the tree and it measures module variability both within and across species, while the within-species_diversity is the sum of the human, gorilla and cynomolgus within-species branch lengths and it measures module variability within species only.
- #' These statistics were calculated for all possible jackknifed versions of the modules, each of which was created by removing a target gene assigned to the given module (the regulators were never excluded). Then the statistics were summarized per module by taking the median and its 95\% confidence interval across all jackknifed module versions.
- #'
- #' @format A data frame with 7 rows and 10 columns:
- #' \describe{
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Module size, the numer of target genes assigned to a regulator.}
- #' \item{total_tree_length}{The median of total tree lengths across all jackknifed versions of the module.}
- #' \item{var_total_tree_length}{The variance of total tree lengths across all jackknifed versions of the module.}
- #' \item{lwr_total_tree_length}{The lower bound of the 95\% confidence interval of the total tree length calculated by jackknifing.}
- #' \item{upr_total_tree_length}{The upper bound of the 95\% confidence interval of the total tree length calculated by jackknifing.}
- #' \item{within_species_diversity}{The median of within-species_diversity values across all jackknifed versions of the module.}
- #' \item{var_within_species_diversity}{The variance of within-species_diversity values across all jackknifed versions of the module.}
- #' \item{lwr_within_species_diversity}{The lower bound of the 95\% confidence interval of the within-species_diversity calculated by jackknifing.}
- #' \item{upr_within_species_diversity}{The upper bound of the 95\% confidence interval of the within-species_diversity calculated by jackknifing.}
- #' }
- "random_tree_stats"
- #' Linear model for the characterization of overall module conservation
- #'
- #' A weighted linear model between the total tree length and within-species diversity of the 12 modules in the subsetted early neuronal differentiation dataset. It captures the general relationship between the two tree-based statistics in this dataset and thus also describes the expected total tree length for any given within-species diversity. By identifying modules that differ the most from this expectation, i.e. identifying the outlier data points, it can be used to pinpoint conserved and overall diverged modules. In combination with jackknifing, it can also be used to identify target genes within these modules that contribute the most to the conservation/divergence. The linear model was fit by calling \code{\link{lm}} with weights inversely proportional to the variance of the total tree lengths.
- #'
- #' @format An object of class \code{\link{lm}} fitted on the object \code{tree_stats} with the formula "total_tree_length ~ within_species_diversity".
- "lm_overall"
- #' Cross-species conservation measures per module
- #'
- #' Cross-species conservation measures of the 12 modules in the subsetted early neuronal differentiation dataset. Modules that were found to be conserved or diverged overall (across all species) are labelled as "conserved" or "diverged" in the column \code{conservation}.
- #'
- #' To determine whether a module as whole is conserved or diverged overall, module trees were reconstructed from pairwise preservation scores between replicates and based on these trees 2 useful statistics were calculated for each module: the total tree length and the within-species diversity (for details please see \code{\link{calculatePresStats}}, \code{\link{reconstructTrees}}, \code{\link{calculateTreeStats}}). After fitting a weighted linear model between the total tree length and within-species diversity values of all modules, a module was considered diverged if it fell above the prediction interval of the regression line, while a module was considered conserved if it fell below the prediction interval of the regression line (for details please see \code{\link{fitTreeStatsLm}} and \code{\link{findConservedDivergedModules}}. The degree of conservation/divergence can be further compared between the modules categorized as conserved/diverged using 2 measures, the residual and the t-score.
- #'
- #' @format A data frame with 12 rows and 16 columns:
- #' \describe{
- #' \item{focus}{Character, the focus of interest in terms of cross-species conservation ("overall").}
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Integer, the numer of target genes assigned to a regulator.}
- #' \item{total_tree_length}{Numeric, total tree length per module (the median across all jackknife versions of the module).}
- #' \item{lwr_total_tree_length}{Numeric, the lower bound of the confidence interval of the total tree length calculated based on the jackknifed versions of the module.}
- #' \item{upr_total_tree_length}{Numeric, the upper bound of the confidence interval of the total tree length calculated based on the jackknifed versions of the module.}
- #' \item{within_species_diversity}{Numeric, within-species diveristy per module (the median across all jackknife versions of the module).}
- #' \item{lwr_within_species_diversity}{Numeric, the lower bound of the confidence interval of the within-species diversity calculated based on the jackknifed versions of the module.}
- #' \item{upr_within_species_diversity}{Numeric, the upper bound of the confidence interval of the within-species diversity calculated based on the jackknifed versions of the module.}
- #' \item{fit}{Numeric, the fitted total tree length at the within-species diversity value of the module.}
- #' \item{lwr_fit}{Numeric, the lower bound of the prediction interval of the fit.}
- #' \item{upr_fit}{Numeric, the upper bound of the prediction interval of the fit.}
- #' \item{residual}{Numeric, the residual of the module in the linear model. It is calculated as the difference between the observed and expected (fitted) total tree lengths.}
- #' \item{studentized_residual}{Numeric, the externally studentized residual of the module. This statistic measures how strongly a module deviates from the fitted regression line relative to the expected variability.}
- #' \item{p_value}{Numeric, two-sided p-value associated with the externally studentized residual.}
- #' \item{fdr}{Numeric, false discovery rate obtained by adjusting the p-values across all modules using the Benjamini–Hochberg method.}
- #' \item{category}{Character, classification of the module based on the prediction interval: "diverged" if a module has a higher total tree length than the upper boundary of the prediction interval, "conserved" if a module has a lower total tree length than the lower boundary of the prediction interval, and "within_expectation" otherwise.}
- #' \item{robust}{Logical, indicates whether the module passes an additional robustness filter based on the false discovery rate (FDR). Modules with an FDR below \code{fdr_cutoff} are marked as \code{TRUE}.}
- #' }
- "module_conservation_overall"
- #' Cross-species conservation measures of the target genes in the POU5F1 module
- #'
- #' Cross-species conservation measures of the 21 target genes in the POU5F1 module in the subsetted early neuronal differentiation dataset.
- #'
- #' To determine which target genes are the most responsible for the divergence of the POU5F1 module, the same statistics were be used as for the identification of conserved/diverged modules but in combination with jackknifing. This means that each target gene in the module was removed and all statistics, including the tree-based statistics that carry information about cross-species conservation, were re-calculated (please see \code{\link{calculatePresStats}}, \code{\link{reconstructTrees}} and \code{\link{calculateTreeStats}}). Our working hypothesis is that if removing a target gene from the diverged POU5F1 module makes the module more conserved, then this gene contributed to the divergence of the original module in the first place. This can be quantified using the residuals compared to the regression line between the total tree lengths and within-species diversities of the tree representations across all modules (please see \code{\link{fitTreeStatsLm}} and \code{\link{findConservedDivergedTargets}}. The jackknifed versions of the POU5F1 module that have the lowest residuals are the most conserved, therefore the corresponding target genes are considered the most diverged.
- #'
- #' @format A data frame with 23 rows and 12 columns:
- #' \describe{
- #' \item{focus}{Character, the focus of interest in terms of cross-species conservation ("overall").}
- #' \item{regulator}{Character, transcriptional regulator.}
- #' \item{module_size}{Integer, the numer of target genes assigned to a regulator.}
- #' \item{type}{Character, module type (orig = original, jk = jackknifed or summary = the summary across all jackknife module versions).}
- #' \item{id}{Character, the unique ID of the module version (format: nameOfRegulator_jk_nameOfGeneRemoved in case of module type 'jk', nameOfRegulator_orig in case of module type 'orig' and nameOfRegulator_summary in case of module type 'summary').}
- #' \item{gene_removed}{Character, the name of the gene removed by jackknifing (NA in case of module types 'orig' and 'summary').}
- #' \item{total_tree_length}{Numeric, total tree length per jackknife module version.}
- #' \item{within_species_diversity}{Numeric, within-species diversity per jackknife module version.}
- #' \item{fit}{Numeric, the fitted total tree length at the within-species diversity value of a jackknife module version.}
- #' \item{lwr_fit}{Numeric, the lower bound of the prediction interval of the fit.}
- #' \item{upr_fit}{Numeric, the upper bound of the prediction interval of the fit.}
- #' \item{residual}{Numeric, the residual of the jackknife module version in the linear model. It is calculated as the difference between the observed and expected (fitted) total tree lengths.}
- #' \item{contribution}{Numeric, the contribution score of the removed target gene.}
- #' }
- "POU5F1_target_conservation"
data.R at commit c0bbc8a, under GPL-3.0 · at the source
Overview
- Anthropology and Human Genomics, Ludwig-Maximilians-Universität München,82152, Planegg, Germany
- Research Unit Brain Epigenomics, Helmholtz Center Munich,81377, Munich, Germany
Abstract
To understand phenotypic evolution, it is essential to investigate the underlying gene regulatory networks (GRNs). However, most comparative GRN analyzes remain descriptive due to the low signal-to-noise ratio inherent in single-cell transcriptomics data. To address this, we introduce CroCoNet (Cross-species Comparison of Networks), an R-package for quantitative GRN comparison across species. CroCoNet builds comparable network modules centered on putative regulators and compares module topologies within and between species, distinguishing true evolutionary divergence from technical and biological confounders. We demonstrate its utility by comparing early neural differentiation across primates and validating results with a CRISPRi analysis of the diverged POU5F1 module.
Reproduced under the paper's license (CC BY), from the paper cited above.
Repositories
Its files are read in the Code ↔ Paper reader above, with 40 matches between paragraphs and lines of code.
anita-termeg/gRNA-design
Availability: 1 check, the latest on 27 September 2026: the link is dead
- 27 September 2026: the link is dead
Hellmann-Lab/CroCoNet
c0bbc8a98747f33d2dfdd4251129c277df2ea502, 17 March 2026Availability: 1 check, the latest on 27 September 2026: the link answers
- 27 September 2026: the link answers
41 files
- R/
CroCoNet-package.R , R, 23 lines - R/
addDirectionality.R , R, 89 lines, 1 match - R/
assignInitialModules.R , R, 70 lines - R/
calculateEigengenes.R , R, 266 lines - R/
calculatePresStats.R , R, 373 lines, 1 match - R/
calculateTreeStats.R , R, 238 lines - R/
comparePresStats.R , R, 168 lines, 2 matches - R/
createConsensus.R , R, 171 lines - R/
createRandomModules.R , R, 58 lines - R/
data.R , R, 563 lines, 6 matches - R/
filterModuleTrees.R , R, 70 lines, 1 match - R/
findConservedDivergedMod , R, 156 lines, 2 matchesules.R - R/
findConservedDivergedTar , R, 184 lines, 1 matchgets.R - R/
fitTreeStatsLm.R , R, 124 lines - R/
getRegulators.R , R, 84 lines - R/
loadNetworks.R , R, 147 lines - R/
normalizeEdgeWeights.R , R, 104 lines - R/
plotConservedDivergedMod , R, 218 lines, 1 matchules.R - R/
plotConservedDivergedTar , R, 181 linesgets.R - R/
plotDistMats.R , R, 130 lines - R/
plotExpr.R , R, 1,423 lines - R/
plotModuleSizeDistributi , R, 36 lineson.R - R/
plotNetworks.R , R, 424 lines - R/
plotPresStats.R , R, 366 lines - R/
plotTreeStats.R , R, 257 lines, 2 matches - R/
plotTrees.R , R, 240 lines - R/
pruneModules.R , R, 505 lines, 3 matches - R/
reconstructTrees.R , R, 286 lines, 1 match - R/
removeOverlappingGenePai , R, 151 linesrs.R - R/
scale_free_transformatio , R, 182 linesn.R - R/
summarizeJackknifeStats. , R, 138 linesR - R/
utils.R , R, 123 lines - README.Rmd, R, 57 lines
- data-raw/
arboreto_with_multiproce , Python, 160 linesssing.py - data-raw/
grnboost2_per_clone.sh , Shell, 22 lines - data-raw/
internal_data.R , R, 52 lines - data-raw/
logo.R , R, 3 lines - data-raw/
toy_dataset.R , R, 344 lines - vignettes/
CroCoNet.Rmd , R, 1,189 lines, 3 matches - LICENSE.md, License, 595 lines
- README.md, Text, 56 lines
Hellmann-Lab/CroCoNet-analyses
4fd700f49c35ad41c161b80e6f021682c1826507, 1 June 2026Availability: 1 check, the latest on 27 September 2026: the link answers
- 27 September 2026: the link answers
115 files
- 1.neural_differentiation
_dataset/ , Shell, 26 lines1.1.mapping_and_QC/ 1.1.1.download_ref_genom es.sh - 1.neural_differentiation
_dataset/ , R, 926 lines1.1.mapping_and_QC/ 1.1.10.QC_and_filtering. R - 1.neural_differentiation
_dataset/ , Shell, 21 lines1.1.mapping_and_QC/ 1.1.2.run_liftoff.sh - 1.neural_differentiation
_dataset/ , R, 56 lines1.1.mapping_and_QC/ 1.1.3.filter_liftoff_gtf s.R - 1.neural_differentiation
_dataset/ , R, 49 lines1.1.mapping_and_QC/ 1.1.4.remove_small_conti gs.R - 1.neural_differentiation
_dataset/ , Shell, 20 lines1.1.mapping_and_QC/ 1.1.5.create_STAR_indice s.sh - 1.neural_differentiation
_dataset/ , Shell, 34 lines1.1.mapping_and_QC/ 1.1.6.download_FASTQ.sh - 1.neural_differentiation
_dataset/ , Shell, 15 lines1.1.mapping_and_QC/ 1.1.7.trimming.sh - 1.neural_differentiation
_dataset/ , Shell, 24 lines1.1.mapping_and_QC/ 1.1.8.mapping.sh - 1.neural_differentiation
_dataset/ , Shell, 23 lines1.1.mapping_and_QC/ 1.1.9.download_cell_type _annotation_ref.sh - 1.neural_differentiation
_dataset/ , R, 63 lines, 1 match1.2.network_inference/ 1.2.1.prepare_data.R - 1.neural_differentiation
_dataset/ , Shell, 15 lines1.2.network_inference/ 1.2.2.create_GRNBoost2_e nv.sh - 1.neural_differentiation
_dataset/ , Shell, 27 lines1.2.network_inference/ 1.2.3.run_GRNBoost2.sh - 1.neural_differentiation
_dataset/ , R, 34 lines1.2.network_inference/ 1.2.4.prepare_data_for_S pearman.R - 1.neural_differentiation
_dataset/ , Shell, 21 lines1.2.network_inference/ 1.2.5.run_correlatePairs .sh - 1.neural_differentiation
_dataset/ , Python, 160 lines1.2.network_inference/ arboreto_with_multiproce ssing.py - 1.neural_differentiation
_dataset/ , R, 23 lines1.2.network_inference/ correlatePairs.R - 1.neural_differentiation
_dataset/ , R, 516 lines, 1 match1.3.CroCoNet_analysis/ 1.3.1.CroCoNet_analysis. R - 1.neural_differentiation
_dataset/ , R, 109 lines1.3.CroCoNet_analysis/ 1.3.2.CroCoNet_analysis_ with_top50_pruning.R - 1.neural_differentiation
_dataset/ , R, 314 lines1.3.CroCoNet_analysis/ 1.3.3.dynamic_VS_top50_p runing.R - 1.neural_differentiation
_dataset/ , R, 40 lines1.3.CroCoNet_analysis/ 1.3.4.CroCoNet_analysis_ with_cor_adj.R - 1.neural_differentiation
_dataset/ , R, 246 lines1.3.CroCoNet_analysis/ 1.3.5.cor_kIM_VS_cor_adj .R - 1.neural_differentiation
_dataset/ , R, 105 lines1.3.CroCoNet_analysis/ 1.3.6.CroCoNet_analysis_ on_Spearman_networks.R - 2.validations/
2.1.regulator_interactio , R, 40 linesns_and_module_overlaps/ 2.1.1.get_regulator_inte ractions.R - 2.validations/
2.1.regulator_interactio , R, 54 linesns_and_module_overlaps/ 2.1.2.calculate_module_o verlaps.R - 2.validations/
2.1.regulator_interactio , R, 90 linesns_and_module_overlaps/ 2.1.3.test_association.R - 2.validations/
2.2.pathway_enrichment/ , R, 132 lines2.2.1.run_Reactome.R - 2.validations/
2.2.pathway_enrichment/ , R, 118 lines2.2.2.summarize_Reactome _results.R - 2.validations/
2.3.binding_site_enrichm , Shell, 83 linesent_and_divergence/ 2.3.1.download_ATAC_seq_ FASTQ.sh - 2.validations/
2.3.binding_site_enrichm , R, 158 linesent_and_divergence/ 2.3.10.liftOver_human_NP C_peaks_to_gorGor6.R - 2.validations/
2.3.binding_site_enrichm , Shell, 17 linesent_and_divergence/ 2.3.11.download_nanopore _FASTQ.sh - 2.validations/
2.3.binding_site_enrichm , Shell, 13 linesent_and_divergence/ 2.3.12.create_minimap2_i ndex.sh - 2.validations/
2.3.binding_site_enrichm , Shell, 59 linesent_and_divergence/ 2.3.13.nanopore_mapping. sh - 2.validations/
2.3.binding_site_enrichm , Shell, 38 linesent_and_divergence/ 2.3.14.merge_BAMs.sh - 2.validations/
2.3.binding_site_enrichm , Shell, 21 linesent_and_divergence/ 2.3.15.reconstruct_nanop ore_transcripts.sh - 2.validations/
2.3.binding_site_enrichm , R, 175 linesent_and_divergence/ 2.3.16.annotate_nanopore _transcripts.R - 2.validations/
2.3.binding_site_enrichm , R, 518 linesent_and_divergence/ 2.3.17.identify_TSS.R - 2.validations/
2.3.binding_site_enrichm , R, 397 lines, 1 matchent_and_divergence/ 2.3.18.associate_peaks_t o_gene.R - 2.validations/
2.3.binding_site_enrichm , R, 74 linesent_and_divergence/ 2.3.19.get_binding_motif s.R - 2.validations/
2.3.binding_site_enrichm , Shell, 35 linesent_and_divergence/ 2.3.2.ATAC_seq_trimming. sh - 2.validations/
2.3.binding_site_enrichm , Shell, 38 linesent_and_divergence/ 2.3.20.run_cbust.sh - 2.validations/
2.3.binding_site_enrichm , R, 1,613 lines, 1 matchent_and_divergence/ 2.3.21.summarize_motif_s cores_per_gene.Rmd - 2.validations/
2.3.binding_site_enrichm , R, 99 linesent_and_divergence/ 2.3.22.binding_site_enri chment_across_module_typ es.R - 2.validations/
2.3.binding_site_enrichm , R, 126 linesent_and_divergence/ 2.3.23.binding_site_dive rgence_across_species.R - 2.validations/
2.3.binding_site_enrichm , Shell, 15 linesent_and_divergence/ 2.3.3.create_bwa_mem2_in dices.sh - 2.validations/
2.3.binding_site_enrichm , Shell, 63 lines, 1 matchent_and_divergence/ 2.3.4.ATAC_seq_mapping.s h - 2.validations/
2.3.binding_site_enrichm , Shell, 59 linesent_and_divergence/ 2.3.5.calculate_coverage .sh - 2.validations/
2.3.binding_site_enrichm , Shell, 17 linesent_and_divergence/ 2.3.6.name_sorting.sh - 2.validations/
2.3.binding_site_enrichm , Shell, 64 linesent_and_divergence/ 2.3.7.peak_calling_witho ut_blacklist.sh - 2.validations/
2.3.binding_site_enrichm , R, 24 linesent_and_divergence/ 2.3.8.create_blacklists. R - 2.validations/
2.3.binding_site_enrichm , Shell, 70 linesent_and_divergence/ 2.3.9.peak_calling_with_ blacklist.sh - 2.validations/
2.4.sequence_divergence/ , Shell, 25 lines2.4.1.download_protein_a lignments_and_phastCons_ scores.sh - 2.validations/
2.4.sequence_divergence/ , R, 152 lines2.4.2.calculate_sequence _divergence.R - 2.validations/
2.4.sequence_divergence/ , R, 66 lines2.4.3.calculate_phastCon s_scores.R - 2.validations/
2.4.sequence_divergence/ , R, 73 lines2.4.4.sequence_divergenc e_VS_network_divergence. R - 2.validations/
2.4.sequence_divergence/ , R, 87 linesalignment_scoring_functi on.R - 2.validations/
2.5.expression_pattern_d , R, 156 linesivergence/ 2.5.1.calculate_expressi on_pattern_divergence.R - 2.validations/
2.5.expression_pattern_d , R, 56 linesivergence/ 2.5.2.expression_pattern _divergence_VS_network_d ivergence.R - 2.validations/
2.5.expression_pattern_d , R, 111 linesivergence/ subsampling_helper_funct ion.R - 2.validations/
2.6.ChIP_seq_enrichment/ , R, 346 linesNANOG_ChIP_seq.Rmd - 2.validations/
2.6.ChIP_seq_enrichment/ , R, 342 linesPAX6_ChIP_seq.Rmd - 2.validations/
2.6.ChIP_seq_enrichment/ , R, 321 linesPOU5F1_ChIP_seq.Rmd - 2.validations/
2.7.POU5F1_LTR7_enrichme , Shell, 9 linesnt/ 2.7.1.run_RepeatMasker_f or_macFas6.sh - 2.validations/
2.7.POU5F1_LTR7_enrichme , R, 181 linesnt/ 2.7.2.get_LTR7_positions _and_info.R - 2.validations/
2.7.POU5F1_LTR7_enrichme , R, 146 lines, 2 matchesnt/ 2.7.3.calculate_enrichme nt_near_POU5F1_module_me mbers.R - 2.validations/
2.7.POU5F1_LTR7_enrichme , R, 31 linesnt/ 2.7.4.create_Ggorilla_BS genome_package.R - 2.validations/
2.7.POU5F1_LTR7_enrichme , R, 136 linesnt/ 2.7.5.get_CREs_near_SPP1 _and_SCGB3A2.R - 2.validations/
2.7.POU5F1_LTR7_enrichme , R, 55 linesnt/ 2.7.6.get_POU5F1_motifs. R - 2.validations/
2.7.POU5F1_LTR7_enrichme , Shell, 30 linesnt/ 2.7.7.indentify_POU5F1_m otifs_near_SPP1_and_SCGB 3A2.sh - 2.validations/
2.7.POU5F1_LTR7_enrichme , R, 334 linesnt/ 2.7.8.plot_regions_near_ SPP1_and_SCGB3A2.R - 2.validations/
2.7.POU5F1_LTR7_enrichme , R, 421 linesnt/ helper_functions.R - 2.validations/
2.8.POU5F1_CRISPRi/ , R, 70 lines2.8.1.annotate_dCas9_cas sette.R - 2.validations/
2.8.POU5F1_CRISPRi/ , R, 143 lines2.8.10.DE_analysis.R - 2.validations/
2.8.POU5F1_CRISPRi/ , Shell, 78 lines2.8.2.create_ref_genomes _with_dCas9_casette.sh - 2.validations/
2.8.POU5F1_CRISPRi/ , Shell, 17 lines2.8.3.download_FASTQ.sh - 2.validations/
2.8.POU5F1_CRISPRi/ , Shell, 5 lines2.8.4.mapping.sh - 2.validations/
2.8.POU5F1_CRISPRi/ , Shell, 7 lines2.8.5.species_demultiple xing.sh - 2.validations/
2.8.POU5F1_CRISPRi/ , R, 59 lines2.8.6.get_cells_per_spec ies.R - 2.validations/
2.8.POU5F1_CRISPRi/ , Shell, 9 lines2.8.7.individual_demulti plexing.sh - 2.validations/
2.8.POU5F1_CRISPRi/ , R, 616 lines, 1 match2.8.8.QC_and_filtering.R - 2.validations/
2.8.POU5F1_CRISPRi/ , R, 388 lines, 1 match2.8.9.stemness_scores_an d_knockdown_efficiencies .R - 2.validations/
2.8.POU5F1_CRISPRi/ , R, 711 lineshelper_functions.R - 2.validations/
2.8.POU5F1_CRISPRi/ , Shell, 74 linesindiv_demultiplexing_per _lane_and_genome.sh - 2.validations/
2.8.POU5F1_CRISPRi/ , R, 10 linesload_genbank_seq.R - 2.validations/
2.8.POU5F1_CRISPRi/ , Shell, 43 linesmapping_per_lane_and_gen ome.sh - 2.validations/
2.8.POU5F1_CRISPRi/ , Shell, 77 linesspecies_demultiplexing_p er_lane_and_genome.sh - 3.brain_dataset/
3.1.data_preparation/ , Shell, 20 lines3.1.1.download_processed _data.sh - 3.brain_dataset/
3.1.data_preparation/ , R, 209 lines, 1 match3.1.2.filtering.R - 3.brain_dataset/
3.1.data_preparation/ , R, 101 lines3.1.3.subsampling.R - 3.brain_dataset/
3.2.network_inference/ , Shell, 19 lines3.2.1.run_correlatePairs .sh - 3.brain_dataset/
3.2.network_inference/ , R, 23 linescorrelatePairs.R - 3.brain_dataset/
3.3.CroCoNet_analysis/ , R, 508 lines3.3.1.CroCoNet_analysis. R - 3.brain_dataset/
3.3.CroCoNet_analysis/ , R, 41 lines3.3.2.CroCoNet_analysis_ with_cor_kIM.R - 3.brain_dataset/
3.3.CroCoNet_analysis/ , R, 107 lines3.3.3.cor_kIM_VS_cor_adj .R - 4.paper_figures_and_tabl
es/ , R, 576 lines, 1 matchfigure2.R - 4.paper_figures_and_tabl
es/ , R, 426 linesfigure3.R - 4.paper_figures_and_tabl
es/ , R, 162 lines, 1 matchfigure4.R - 4.paper_figures_and_tabl
es/ , R, 1,168 lineshelper_functions.R - 4.paper_figures_and_tabl
es/ , R, 386 linessuppl.figureS1.R - 4.paper_figures_and_tabl
es/ , R, 201 linessuppl.figureS10.R - 4.paper_figures_and_tabl
es/ , R, 280 linessuppl.figureS11.R - 4.paper_figures_and_tabl
es/ , R, 300 lines, 1 matchsuppl.figureS12_S13.R - 4.paper_figures_and_tabl
es/ , R, 144 linessuppl.figureS14.R - 4.paper_figures_and_tabl
es/ , R, 228 linessuppl.figureS2.R - 4.paper_figures_and_tabl
es/ , R, 258 linessuppl.figureS3.R - 4.paper_figures_and_tabl
es/ , R, 333 linessuppl.figureS4.R - 4.paper_figures_and_tabl
es/ , R, 240 linessuppl.figureS5.R - 4.paper_figures_and_tabl
es/ , R, 86 lines, 1 matchsuppl.figureS6.R - 4.paper_figures_and_tabl
es/ , R, 148 linessuppl.figureS7.R - 4.paper_figures_and_tabl
es/ , R, 43 linessuppl.figureS8.R - 4.paper_figures_and_tabl
es/ , R, 195 linessuppl.figureS9.R - 4.paper_figures_and_tabl
es/ , R, 419 lines, 2 matchessuppl.tables.R - README.Rmd, R, 386 lines
- LICENSE, License, 674 lines
- README.md, Text, 333 lines
Zenodo 20343560
Availability: 1 check, the latest on 27 September 2026: the link answers (HTTP 200)
- 27 September 2026: the link answers (HTTP 200)
115 files
- CroCoNet_analysis_script
s.zip/ , Shell, 26 lines1.neural_differentiation _dataset/ 1.1.mapping_and_QC/ 1.1.1.download_ref_genom es.sh - CroCoNet_analysis_script
s.zip/ , R, 926 lines1.neural_differentiation _dataset/ 1.1.mapping_and_QC/ 1.1.10.QC_and_filtering. R - CroCoNet_analysis_script
s.zip/ , Shell, 21 lines1.neural_differentiation _dataset/ 1.1.mapping_and_QC/ 1.1.2.run_liftoff.sh - CroCoNet_analysis_script
s.zip/ , R, 56 lines1.neural_differentiation _dataset/ 1.1.mapping_and_QC/ 1.1.3.filter_liftoff_gtf s.R - CroCoNet_analysis_script
s.zip/ , R, 49 lines1.neural_differentiation _dataset/ 1.1.mapping_and_QC/ 1.1.4.remove_small_conti gs.R - CroCoNet_analysis_script
s.zip/ , Shell, 20 lines1.neural_differentiation _dataset/ 1.1.mapping_and_QC/ 1.1.5.create_STAR_indice s.sh - CroCoNet_analysis_script
s.zip/ , Shell, 34 lines1.neural_differentiation _dataset/ 1.1.mapping_and_QC/ 1.1.6.download_FASTQ.sh - CroCoNet_analysis_script
s.zip/ , Shell, 15 lines1.neural_differentiation _dataset/ 1.1.mapping_and_QC/ 1.1.7.trimming.sh - CroCoNet_analysis_script
s.zip/ , Shell, 24 lines1.neural_differentiation _dataset/ 1.1.mapping_and_QC/ 1.1.8.mapping.sh - CroCoNet_analysis_script
s.zip/ , Shell, 23 lines1.neural_differentiation _dataset/ 1.1.mapping_and_QC/ 1.1.9.download_cell_type _annotation_ref.sh - CroCoNet_analysis_script
s.zip/ , R, 63 lines1.neural_differentiation _dataset/ 1.2.network_inference/ 1.2.1.prepare_data.R - CroCoNet_analysis_script
s.zip/ , Shell, 15 lines1.neural_differentiation _dataset/ 1.2.network_inference/ 1.2.2.create_GRNBoost2_e nv.sh - CroCoNet_analysis_script
s.zip/ , Shell, 27 lines1.neural_differentiation _dataset/ 1.2.network_inference/ 1.2.3.run_GRNBoost2.sh - CroCoNet_analysis_script
s.zip/ , R, 34 lines1.neural_differentiation _dataset/ 1.2.network_inference/ 1.2.4.prepare_data_for_S pearman.R - CroCoNet_analysis_script
s.zip/ , Shell, 21 lines1.neural_differentiation _dataset/ 1.2.network_inference/ 1.2.5.run_correlatePairs .sh - CroCoNet_analysis_script
s.zip/ , Python, 160 lines1.neural_differentiation _dataset/ 1.2.network_inference/ arboreto_with_multiproce ssing.py - CroCoNet_analysis_script
s.zip/ , R, 23 lines1.neural_differentiation _dataset/ 1.2.network_inference/ correlatePairs.R - CroCoNet_analysis_script
s.zip/ , R, 516 lines1.neural_differentiation _dataset/ 1.3.CroCoNet_analysis/ 1.3.1.CroCoNet_analysis. R - CroCoNet_analysis_script
s.zip/ , R, 109 lines1.neural_differentiation _dataset/ 1.3.CroCoNet_analysis/ 1.3.2.CroCoNet_analysis_ with_top50_pruning.R - CroCoNet_analysis_script
s.zip/ , R, 314 lines1.neural_differentiation _dataset/ 1.3.CroCoNet_analysis/ 1.3.3.dynamic_VS_top50_p runing.R - CroCoNet_analysis_script
s.zip/ , R, 40 lines1.neural_differentiation _dataset/ 1.3.CroCoNet_analysis/ 1.3.4.CroCoNet_analysis_ with_cor_adj.R - CroCoNet_analysis_script
s.zip/ , R, 246 lines1.neural_differentiation _dataset/ 1.3.CroCoNet_analysis/ 1.3.5.cor_kIM_VS_cor_adj .R - CroCoNet_analysis_script
s.zip/ , R, 105 lines1.neural_differentiation _dataset/ 1.3.CroCoNet_analysis/ 1.3.6.CroCoNet_analysis_ on_Spearman_networks.R - CroCoNet_analysis_script
s.zip/ , R, 40 lines2.validations/ 2.1.regulator_interactio ns_and_module_overlaps/ 2.1.1.get_regulator_inte ractions.R - CroCoNet_analysis_script
s.zip/ , R, 54 lines2.validations/ 2.1.regulator_interactio ns_and_module_overlaps/ 2.1.2.calculate_module_o verlaps.R - CroCoNet_analysis_script
s.zip/ , R, 90 lines2.validations/ 2.1.regulator_interactio ns_and_module_overlaps/ 2.1.3.test_association.R - CroCoNet_analysis_script
s.zip/ , R, 132 lines2.validations/ 2.2.pathway_enrichment/ 2.2.1.run_Reactome.R - CroCoNet_analysis_script
s.zip/ , R, 118 lines2.validations/ 2.2.pathway_enrichment/ 2.2.2.summarize_Reactome _results.R - CroCoNet_analysis_script
s.zip/ , Shell, 83 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.1.download_ATAC_seq_ FASTQ.sh - CroCoNet_analysis_script
s.zip/ , R, 158 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.10.liftOver_human_NP C_peaks_to_gorGor6.R - CroCoNet_analysis_script
s.zip/ , Shell, 17 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.11.download_nanopore _FASTQ.sh - CroCoNet_analysis_script
s.zip/ , Shell, 13 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.12.create_minimap2_i ndex.sh - CroCoNet_analysis_script
s.zip/ , Shell, 59 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.13.nanopore_mapping. sh - CroCoNet_analysis_script
s.zip/ , Shell, 38 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.14.merge_BAMs.sh - CroCoNet_analysis_script
s.zip/ , Shell, 21 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.15.reconstruct_nanop ore_transcripts.sh - CroCoNet_analysis_script
s.zip/ , R, 175 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.16.annotate_nanopore _transcripts.R - CroCoNet_analysis_script
s.zip/ , R, 518 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.17.identify_TSS.R - CroCoNet_analysis_script
s.zip/ , R, 397 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.18.associate_peaks_t o_gene.R - CroCoNet_analysis_script
s.zip/ , R, 74 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.19.get_binding_motif s.R - CroCoNet_analysis_script
s.zip/ , Shell, 35 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.2.ATAC_seq_trimming. sh - CroCoNet_analysis_script
s.zip/ , Shell, 38 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.20.run_cbust.sh - CroCoNet_analysis_script
s.zip/ , R, 1,613 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.21.summarize_motif_s cores_per_gene.Rmd - CroCoNet_analysis_script
s.zip/ , R, 99 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.22.binding_site_enri chment_across_module_typ es.R - CroCoNet_analysis_script
s.zip/ , R, 126 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.23.binding_site_dive rgence_across_species.R - CroCoNet_analysis_script
s.zip/ , Shell, 15 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.3.create_bwa_mem2_in dices.sh - CroCoNet_analysis_script
s.zip/ , Shell, 63 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.4.ATAC_seq_mapping.s h - CroCoNet_analysis_script
s.zip/ , Shell, 59 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.5.calculate_coverage .sh - CroCoNet_analysis_script
s.zip/ , Shell, 17 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.6.name_sorting.sh - CroCoNet_analysis_script
s.zip/ , Shell, 64 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.7.peak_calling_witho ut_blacklist.sh - CroCoNet_analysis_script
s.zip/ , R, 24 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.8.create_blacklists. R - CroCoNet_analysis_script
s.zip/ , Shell, 70 lines2.validations/ 2.3.binding_site_enrichm ent_and_divergence/ 2.3.9.peak_calling_with_ blacklist.sh - CroCoNet_analysis_script
s.zip/ , Shell, 25 lines2.validations/ 2.4.sequence_divergence/ 2.4.1.download_protein_a lignments_and_phastCons_ scores.sh - CroCoNet_analysis_script
s.zip/ , R, 152 lines2.validations/ 2.4.sequence_divergence/ 2.4.2.calculate_sequence _divergence.R - CroCoNet_analysis_script
s.zip/ , R, 66 lines2.validations/ 2.4.sequence_divergence/ 2.4.3.calculate_phastCon s_scores.R - CroCoNet_analysis_script
s.zip/ , R, 73 lines2.validations/ 2.4.sequence_divergence/ 2.4.4.sequence_divergenc e_VS_network_divergence. R - CroCoNet_analysis_script
s.zip/ , R, 87 lines2.validations/ 2.4.sequence_divergence/ alignment_scoring_functi on.R - CroCoNet_analysis_script
s.zip/ , R, 156 lines2.validations/ 2.5.expression_pattern_d ivergence/ 2.5.1.calculate_expressi on_pattern_divergence.R - CroCoNet_analysis_script
s.zip/ , R, 56 lines2.validations/ 2.5.expression_pattern_d ivergence/ 2.5.2.expression_pattern _divergence_VS_network_d ivergence.R - CroCoNet_analysis_script
s.zip/ , R, 111 lines2.validations/ 2.5.expression_pattern_d ivergence/ subsampling_helper_funct ion.R - CroCoNet_analysis_script
s.zip/ , R, 346 lines2.validations/ 2.6.ChIP_seq_enrichment/ NANOG_ChIP_seq.Rmd - CroCoNet_analysis_script
s.zip/ , R, 342 lines2.validations/ 2.6.ChIP_seq_enrichment/ PAX6_ChIP_seq.Rmd - CroCoNet_analysis_script
s.zip/ , R, 321 lines2.validations/ 2.6.ChIP_seq_enrichment/ POU5F1_ChIP_seq.Rmd - CroCoNet_analysis_script
s.zip/ , Shell, 9 lines2.validations/ 2.7.POU5F1_LTR7_enrichme nt/ 2.7.1.run_RepeatMasker_f or_macFas6.sh - CroCoNet_analysis_script
s.zip/ , R, 181 lines2.validations/ 2.7.POU5F1_LTR7_enrichme nt/ 2.7.2.get_LTR7_positions _and_info.R - CroCoNet_analysis_script
s.zip/ , R, 146 lines2.validations/ 2.7.POU5F1_LTR7_enrichme nt/ 2.7.3.calculate_enrichme nt_near_POU5F1_module_me mbers.R - CroCoNet_analysis_script
s.zip/ , R, 31 lines2.validations/ 2.7.POU5F1_LTR7_enrichme nt/ 2.7.4.create_Ggorilla_BS genome_package.R - CroCoNet_analysis_script
s.zip/ , R, 136 lines2.validations/ 2.7.POU5F1_LTR7_enrichme nt/ 2.7.5.get_CREs_near_SPP1 _and_SCGB3A2.R - CroCoNet_analysis_script
s.zip/ , R, 55 lines2.validations/ 2.7.POU5F1_LTR7_enrichme nt/ 2.7.6.get_POU5F1_motifs. R - CroCoNet_analysis_script
s.zip/ , Shell, 30 lines2.validations/ 2.7.POU5F1_LTR7_enrichme nt/ 2.7.7.indentify_POU5F1_m otifs_near_SPP1_and_SCGB 3A2.sh - CroCoNet_analysis_script
s.zip/ , R, 334 lines2.validations/ 2.7.POU5F1_LTR7_enrichme nt/ 2.7.8.plot_regions_near_ SPP1_and_SCGB3A2.R - CroCoNet_analysis_script
s.zip/ , R, 421 lines2.validations/ 2.7.POU5F1_LTR7_enrichme nt/ helper_functions.R - CroCoNet_analysis_script
s.zip/ , R, 70 lines2.validations/ 2.8.POU5F1_CRISPRi/ 2.8.1.annotate_dCas9_cas sette.R - CroCoNet_analysis_script
s.zip/ , R, 143 lines2.validations/ 2.8.POU5F1_CRISPRi/ 2.8.10.DE_analysis.R - CroCoNet_analysis_script
s.zip/ , Shell, 78 lines2.validations/ 2.8.POU5F1_CRISPRi/ 2.8.2.create_ref_genomes _with_dCas9_casette.sh - CroCoNet_analysis_script
s.zip/ , Shell, 17 lines2.validations/ 2.8.POU5F1_CRISPRi/ 2.8.3.download_FASTQ.sh - CroCoNet_analysis_script
s.zip/ , Shell, 5 lines2.validations/ 2.8.POU5F1_CRISPRi/ 2.8.4.mapping.sh - CroCoNet_analysis_script
s.zip/ , Shell, 7 lines2.validations/ 2.8.POU5F1_CRISPRi/ 2.8.5.species_demultiple xing.sh - CroCoNet_analysis_script
s.zip/ , R, 59 lines2.validations/ 2.8.POU5F1_CRISPRi/ 2.8.6.get_cells_per_spec ies.R - CroCoNet_analysis_script
s.zip/ , Shell, 9 lines2.validations/ 2.8.POU5F1_CRISPRi/ 2.8.7.individual_demulti plexing.sh - CroCoNet_analysis_script
s.zip/ , R, 616 lines2.validations/ 2.8.POU5F1_CRISPRi/ 2.8.8.QC_and_filtering.R - CroCoNet_analysis_script
s.zip/ , R, 388 lines2.validations/ 2.8.POU5F1_CRISPRi/ 2.8.9.stemness_scores_an d_knockdown_efficiencies .R - CroCoNet_analysis_script
s.zip/ , R, 711 lines2.validations/ 2.8.POU5F1_CRISPRi/ helper_functions.R - CroCoNet_analysis_script
s.zip/ , Shell, 74 lines2.validations/ 2.8.POU5F1_CRISPRi/ indiv_demultiplexing_per _lane_and_genome.sh - CroCoNet_analysis_script
s.zip/ , R, 10 lines2.validations/ 2.8.POU5F1_CRISPRi/ load_genbank_seq.R - CroCoNet_analysis_script
s.zip/ , Shell, 43 lines2.validations/ 2.8.POU5F1_CRISPRi/ mapping_per_lane_and_gen ome.sh - CroCoNet_analysis_script
s.zip/ , Shell, 77 lines2.validations/ 2.8.POU5F1_CRISPRi/ species_demultiplexing_p er_lane_and_genome.sh - CroCoNet_analysis_script
s.zip/ , Shell, 20 lines3.brain_dataset/ 3.1.data_preparation/ 3.1.1.download_processed _data.sh - CroCoNet_analysis_script
s.zip/ , R, 209 lines3.brain_dataset/ 3.1.data_preparation/ 3.1.2.filtering.R - CroCoNet_analysis_script
s.zip/ , R, 101 lines3.brain_dataset/ 3.1.data_preparation/ 3.1.3.subsampling.R - CroCoNet_analysis_script
s.zip/ , Shell, 19 lines3.brain_dataset/ 3.2.network_inference/ 3.2.1.run_correlatePairs .sh - CroCoNet_analysis_script
s.zip/ , R, 23 lines3.brain_dataset/ 3.2.network_inference/ correlatePairs.R - CroCoNet_analysis_script
s.zip/ , R, 508 lines3.brain_dataset/ 3.3.CroCoNet_analysis/ 3.3.1.CroCoNet_analysis. R - CroCoNet_analysis_script
s.zip/ , R, 41 lines3.brain_dataset/ 3.3.CroCoNet_analysis/ 3.3.2.CroCoNet_analysis_ with_cor_kIM.R - CroCoNet_analysis_script
s.zip/ , R, 107 lines3.brain_dataset/ 3.3.CroCoNet_analysis/ 3.3.3.cor_kIM_VS_cor_adj .R - CroCoNet_analysis_script
s.zip/ , R, 576 lines4.paper_figures_and_tabl es/ figure2.R - CroCoNet_analysis_script
s.zip/ , R, 426 lines4.paper_figures_and_tabl es/ figure3.R - CroCoNet_analysis_script
s.zip/ , R, 162 lines4.paper_figures_and_tabl es/ figure4.R - CroCoNet_analysis_script
s.zip/ , R, 1,168 lines4.paper_figures_and_tabl es/ helper_functions.R - CroCoNet_analysis_script
s.zip/ , R, 386 lines4.paper_figures_and_tabl es/ suppl.figureS1.R - CroCoNet_analysis_script
s.zip/ , R, 201 lines4.paper_figures_and_tabl es/ suppl.figureS10.R - CroCoNet_analysis_script
s.zip/ , R, 280 lines4.paper_figures_and_tabl es/ suppl.figureS11.R - CroCoNet_analysis_script
s.zip/ , R, 300 lines4.paper_figures_and_tabl es/ suppl.figureS12_S13.R - CroCoNet_analysis_script
s.zip/ , R, 144 lines4.paper_figures_and_tabl es/ suppl.figureS14.R - CroCoNet_analysis_script
s.zip/ , R, 228 lines4.paper_figures_and_tabl es/ suppl.figureS2.R - CroCoNet_analysis_script
s.zip/ , R, 258 lines4.paper_figures_and_tabl es/ suppl.figureS3.R - CroCoNet_analysis_script
s.zip/ , R, 333 lines4.paper_figures_and_tabl es/ suppl.figureS4.R - CroCoNet_analysis_script
s.zip/ , R, 240 lines4.paper_figures_and_tabl es/ suppl.figureS5.R - CroCoNet_analysis_script
s.zip/ , R, 86 lines4.paper_figures_and_tabl es/ suppl.figureS6.R - CroCoNet_analysis_script
s.zip/ , R, 148 lines4.paper_figures_and_tabl es/ suppl.figureS7.R - CroCoNet_analysis_script
s.zip/ , R, 43 lines4.paper_figures_and_tabl es/ suppl.figureS8.R - CroCoNet_analysis_script
s.zip/ , R, 195 lines4.paper_figures_and_tabl es/ suppl.figureS9.R - CroCoNet_analysis_script
s.zip/ , R, 419 lines4.paper_figures_and_tabl es/ suppl.tables.R - CroCoNet_analysis_script
s.zip/ , R, 386 linesREADME.Rmd - CroCoNet_analysis_script
s.zip/ , License, 674 linesLICENSE - CroCoNet_analysis_script
s.zip/ , Text, 333 linesREADME.md
The paper's code and data availability statement is in the Data section.
Tracing map
Proposed by the machine: these links were found in the paper and verified at the source, without human review. The map will receive a Zenodo DOI once one of the paper's authors has validated it with their ORCID.
What the map holds:
- 4 repositories of the authors' code, each at its verified commit, with its license and how the link was found in the paper;
- 265 scripts, each with its path and the digest of its content;
- 40 matches between paragraphs of the paper and lines of the code (method lexical-v1);
- neither the text of the paper nor the code itself.
Its JSON (tracing-map.json) is deposited on Zenodo with its DOI once the map is validated.
Data
Datasets cited
- arrayexpress:E-MTAB-1569
5 , at ArrayExpress; found in the references - figshare:33000921, at figshare; found in DataCite
- geo:GSE326106, at NCBI GEO; found in the references
Data availability
The CroCoNet R-package is available at https://
Datasets generated for this study have been deposited in public repositories. The neural differentiation scRNA-seq data have been deposited in ArrayExpress under the accession number E-MTAB-15695 [111]. The ATAC-seq data of the cell lines G1c2, G1c3 and C1c3 have been deposited in ArrayExpress under accession number E-MTAB-15654 [112]. The long-read RNA-seq data have been deposited in GEO under the accession number GSE326106 [113]. The single-cell CRISPRi data, profiling the effects of the POU5F1 perturbation, have been deposited in GEO under the accession number GSE325993 [114].
All other datasets were previously published and reused for this study. The raw data for the primate brain snRNA-seq dataset are available at the Neuroscience Multi-omic Data Archive (NeMO) under the ID nemo:dat-net1412 [115] and the corresponding processed data can be downloaded from the NeMO Archive repository ”Great Ape MTG Analysis”. The ATAC-seq data of the cell lines H1c2, H2c1, C1c1 and C2c1 are available at ArrayExpress under the accession number E-MTAB-13373 [116]. The POU5F1 and NANOG ChIP-seq data are available at GEO under the accession numbers GSE69646 [117] and GSE99627 [118], respectively. The PAX6 ChIP-seq data are available at the Genome Sequence Archive (GSA) of the China National Center for Bioinformation (CNCB-NGDC) under the accession number CRA002482 [119]. The Hi-C data were published under the DOI 10.1101/
Reproduced under the paper's license (CC BY), from the paper cited above.
Versions
The history of this record: each version stored by the harvester or made by a correction of its authors or of the maintainers of its code, and what changed in its facts. The texts of the paper (its abstract, its availability statements) are not part of it; versions that changed only those are not listed.
Version 1, 27 September 2026: the first record
Recorded: type, language, journal, volume, issue, pages, dates, 11 authors, 5 MeSH terms, 4 funders, 107 references.
Cite
This paper
Térmeg, A., Storozhuk, V., Kliesmete, Z., Edenhofer, F. C., Geuder, J., Dietl, T., Vieth, B., Janssen, P., Richter, D., Bonev, B., & Hellmann, I. (2026). CroCoNet: a framework for the quantitative comparison of gene regulatory networks across species. Genome biology, 27(1), 228. https://
BibTeX
@article{termeg2026croco
author = {Térmeg, Anita and Storozhuk, Vladyslav and Kliesmete, Zane and Edenhofer, Fiona C. and Geuder, Johanna and Dietl, Tamina and Vieth, Beate and Janssen, Philipp and Richter, Daniel and Bonev, Boyan and Hellmann, Ines},
title = {{CroCoNet: a framework for the quantitative comparison of gene regulatory networks across species}},
journal = {Genome biology},
year = {2026},
month = jul,
volume = {27},
number = {1},
pages = {228},
publisher = {BMC},
issn = {1474-7596},
doi = {10.1186/
url = {https://
pmid = {42458527},
pmcid = {PMC13371563}
}
RIS
TY - JOUR
AU - Térmeg, Anita
AU - Storozhuk, Vladyslav
AU - Kliesmete, Zane
AU - Edenhofer, Fiona C.
AU - Geuder, Johanna
AU - Dietl, Tamina
AU - Vieth, Beate
AU - Janssen, Philipp
AU - Richter, Daniel
AU - Bonev, Boyan
AU - Hellmann, Ines
TI - CroCoNet: a framework for the quantitative comparison of gene regulatory networks across species
T2 - Genome biology
J2 - Genome Biol
PY - 2026
DA - 2026/
VL - 27
IS - 1
SP - 228
SN - 1474-7596
PB - BMC
DO - 10.1186/
UR - https://
LA - en
ER -
CSL-JSON
{
"id": "10.1186/
"type": "article-journal",
"title": "CroCoNet: a framework for the quantitative comparison of gene regulatory networks across species",
"container-title": "Genome biology",
"author": [
{
"family": "Térmeg",
"given": "Anita"
},
{
"family": "Storozhuk",
"given": "Vladyslav"
},
{
"family": "Kliesmete",
"given": "Zane"
},
{
"family": "Edenhofer",
"given": "Fiona C."
},
{
"family": "Geuder",
"given": "Johanna"
},
{
"family": "Dietl",
"given": "Tamina"
},
{
"family": "Vieth",
"given": "Beate"
},
{
"family": "Janssen",
"given": "Philipp"
},
{
"family": "Richter",
"given": "Daniel"
},
{
"family": "Bonev",
"given": "Boyan"
},
{
"family": "Hellmann",
"given": "Ines"
}
],
"container-title-short":
"volume": "27",
"issue": "1",
"page": "228",
"DOI": "10.1186/
"PMID": "42458527",
"PMCID": "PMC13371563",
"ISSN": "1474-7596",
"publisher": "BMC",
"URL": "https://
"language": "en",
"issued": {
"date-parts": [
[
2026,
7,
15
]
]
}
}
The tracing map gets a citation of its own once an author has validated it and it has a DOI.
Similar papers
The papers with a page that share the most with this one: the tools found in their code, their categories, datasets, cited references and authors, the rarest counting most.
- [1] doi:10.1016/j.xcrm.2026.102766 [code]
- A longitudinal single-cell and spatial multiomic atlas of pediatric high-grade glioma.Journal: Cell reports. MedicineIn common: Harmony, SingleCellExperiment, limma, 13 other tools, 3 references
- [2] doi:10.1038/s41467-026-73305-8 [code]
- Comparative analysis of the cellular landscape in mammalian striatum.Journal: Nature communicationsIn common: Cell Ranger, Harmony, SingleCellExperiment, 11 other tools, 2 references
- [3] doi:10.1038/s41586-026-10214-2 [code]
- Multidimensional profiling of heterogeneity in supratentorial ependymomas.Journal: NatureIn common: Cell Ranger, Harmony, SingleCellExperiment, 12 other tools
- [4] doi:10.1038/s41586-026-10629-x [code]
- Whole-genome duplication shaped cell-type evolution in the vertebrate brain.Journal: NatureIn common: BEDTools, Harmony, igraph, 11 other tools, 2 references
- [5] doi:10.1038/s44318-026-00818-9 [code]
- FAM134B-mediated ER-phagy degrades APP and suppresses Alzheimer's disease pathology.Journal: The EMBO journalIn common: BEDTools, Harmony, SingleCellExperiment, 12 other tools
- [6] doi:10.1038/s41398-026-04200-5 [code]
- Postmortem brain single-nucleus and bulk gene expression analyses identify shared and distinct abnormalities in bipolar disorder and major depressive disorder.Journal: Translational psychiatryIn common: Harmony, SingleCellExperiment, limma, 9 other tools, 3 references
- [7] doi:10.1038/s42003-026-10957-8 [code]
- Brain defence by the extracellular matrix protein Cochlin.Journal: Communications biologyIn common: SAMtools, limma, igraph, 12 other tools
- [8] doi:10.1016/j.xcrm.2026.102651 [code]
- Integrative CSF profiling identifies disease-specific immune responses in leptomeningeal disease.Journal: Cell reports. MedicineIn common: Harmony, SingleCellExperiment, igraph, 12 other tools
- [9] doi:10.1002/imt2.70163 [code]
- Spatial multi-omics unveils sphingolipid metabolic reprogramming within the retinal pathological niche.Journal: iMetaIn common: Harmony, SingleCellExperiment, limma, 12 other tools
- [10] doi:10.1038/s41586-026-10612-6 [code]
- Acquired genetic and cell-state changes in IDH-mutant glioma progression.Journal: NatureIn common: BEDTools, Harmony, SingleCellExperiment, 11 other tools
Contribute
The authors of this paper can claim it, correct its record and validate its tracing map, and the maintainers of its code (its owner, or a public member of its organization) correct what it says of their repository; anyone signed in can ask for its removal. Every request goes to OSCR's own machine, which answers it; your account page follows them.
Sign in with ORCID to claim this paper as one of its authors, correct its record or validate its tracing map: when the paper's metadata lists your ORCID iD, you are recognized at once. Maintainers of its code: sign in with GitHub, then claim the repository on your account page.
Claim this paper
Correct its record
Say what each link of this record is, remove the ones that are not the paper's, add the ones that are missing. The correction becomes a new version of the record, in its Versions section.
Validate its tracing map
You validate the map as this page shows it: 4 repositories of the authors' code, each at its verified commit and with its license, 265 scripts, and 40 matches between paragraphs and code (see the Code and Map sections). It then receives a DOI on Zenodo, with you (your ORCID iD) and OSCR as its creators; the code itself is not deposited.
The map's fingerprint: sha256:8734da304f92a9a0…
Add the badge to its README
The badge links the code to this page. Copy one of these into the README of the paper's code: only you decide where it goes, and nothing is changed for you.
Markdown
[, paste the snippet at the top, then “Commit changes…” and, to review it first, “Create a new branch and start a pull request”. You open the pull request; OSCR asks for no permission.
Request its removal
To ask OSCR to remove this record, the copies of its authors' scripts or its tracing map, use the removal request page: signed in, you say who you are, what to remove and why, then review and confirm the request. Published rules decide every request (how).
Discussion, reproductions, activity
Discussion: questions and error reports about this paper and its code, from signed-in readers and its authors. It opens with sign-in.
Reproductions: reports from readers who ran the authors' code: what they reproduced, with which environment, commit and data. It opens with sign-in.
Activity: what happens around this paper: new versions of its record, its map's validation, discussions and reproductions. It opens with sign-in.
