Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Software tool to organize, retrieve, and share genome analysis resources. Reference genome assembly asset manager. In addition to genome indexes, can manage any files related to reference genomes, including sequences and annotation files. Includes command line interface and server application that provides RESTful API, so it is useful for both tool development and analysis.
Proper citation: refgenie (RRID:SCR_017574) Copy
https://bioconductor.org/packages/TCGAbiolinks/
Software R Bioconductor package for integrative analysis with TCGA data.TCGAbiolinks is able to access National Cancer Institute Genomic Data Commons thorough its GDC Application Programming Interface to search, download and prepare relevant data for analysis in R.
Proper citation: TCGAbiolinks (RRID:SCR_017683) Copy
https://crispy.secondarymetabolites.org
Web tool to design sgRNAs for CRISPR applications. Web tool based on CRISPy to design sgRNAs for any user-provided microbial genome. Implemented as standalone web application for Cas9 target prediction.
Proper citation: CRISPy-web (RRID:SCR_017970) Copy
https://github.com/slimsuite/pafscaff
Software as Pairwise mApping Format reference based Scaffold anchoring and super scaffolding tool. Dsigned for mapping genome assembly scaffolds to closely related chromosome level reference genome assembly.
Proper citation: PAFScaff (RRID:SCR_017976) Copy
Visualization and analysis software for interactive visual exploration and mining of fiber-tracts and brain networks with their genetic determinants and functional outcomes. BECA includes an fMRI and Diseases Analysis version as well as a Genome Explorer version.
Proper citation: BECA (RRID:SCR_015846) Copy
http://www.alliancegenome.org/
Organization that aims to develop and maintain sustainable genome information resources to promote understanding of the genetic and genomic basis of human biology, health, and disease. The Alliance is composed of FlyBase, Mouse Genome Database (MGD), the Gene Ontology Consortium (GOC), Saccharomyces Genome Database (SGD), Rat Genome Database (RGD), WormBase, and the Zebrafish Information Network (ZFIN).
Proper citation: Alliance of Genome Resources (RRID:SCR_015850) Copy
https://github.com/thackl/cross-species-scaffolding
Software that generates in silico mate-pair reads from single-/paired-end reads of your organism of interest, and a closely related reference genome. It can improve draft genomes by using preferred scaffolding software with the newly created read data. Software that generates in silico mate-pair reads from single-/paired-end reads of your organism of interest, and a closely related reference genome. It can improve draft genomes by using preferred scaffolding software with the newly created read data. Super-scaffolding of draft genome assemblies with in silico mate-pair libraries derived from (closely) related references.
Proper citation: Cross-species scaffolding (RRID:SCR_015932) Copy
https://github.com/harry-thorpe/piggy
Pipeline for analyzing intergenic regions in bacteria. It is designed to be used in conjunction with Roary (https://github.com/sanger-pathogens/Roary).
Proper citation: Piggy (RRID:SCR_015941) Copy
Web application to perform automated model construction and genome annotation for large-scale metabolic networks. Platform for accessing, analyzing and manipulating genome-scale metabolic networks (GSM) as well as biochemical pathways.
Proper citation: MetaNetX (RRID:SCR_015882) Copy
http://www.vicbioinformatics.com/software.barrnap.shtml
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software to predict the location of ribosomal RNA genes in genomes. It supports bacteria, archaea, mitochondria, and eukaryotes. It takes FASTA DNA sequence as input, writes GFF3 as output, and supports multithreading., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Barrnap (RRID:SCR_015995) Copy
http://standage.github.io/AEGeAn
Software toolkit for the analysis and evaluation of genome annotations. The toolkit includes a variety of analysis programs, e.g. for comparing distinct sets of gene structure annotations (ParsEval), computation of gene loci (LocusPocus) and more.
Proper citation: Aegean (RRID:SCR_015965) Copy
https://github.com/EvolBioInf/andi
Software tool for rapidly computing and estimating evolutionary distance between closely related genomes. Because andi does not compute full alignments it scales even up to thousands of bacterial genomes.
Proper citation: andi (RRID:SCR_015971) Copy
https://github.com/PacificBiosciences/FALCON
Software package for aligning long sequencing reads as a diploid-aware genome assembler. Used for assembling non-inbred or rearranged heterozygous genomes.
Proper citation: Falcon (RRID:SCR_016089) Copy
http://baderlab.org/Software/EnrichmentMap
Source code of a Cytoscape plugin for functional enrichment visualization. It organizes gene-sets, such as pathways and Gene Ontology terms, into a network to reveal which mutually overlapping gene-sets cluster together.
Proper citation: EnrichmentMap (RRID:SCR_016052) Copy
http://www.xavierdidelot.xtreemhost.com/clonalframe.htm
Software package for the inference of bacterial microevolution using multilocus sequence data. It is used to identify the clonal relationships between the members of a sample, while also estimating the chromosomal position of homologous recombination events that have disrupted the clonal inheritance.
Proper citation: Clonalframe (RRID:SCR_016060) Copy
https://sanger-pathogens.github.io/gubbins/
Software application as an algorithm that iteratively identifies loci containing elevated densities of base substitutions while concurrently constructing a phylogeny based on the putative point mutations outside of these regions. It is used for phylogenetic analysis of genome sequences and generating highly accurate reconstructions under realistic models of short-term bacterial evolution., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Gubbins (RRID:SCR_016131) Copy
http://projects.tcag.ca/humandup/
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 17, 2013. It contains information about segmental duplications in the human genome. The criteria used to identify regions of segmental duplication are: Sequence identity of at least 90, Sequence length of at least 5 kb, Not be entirely composed of repetitive elements. Background Previous studies have suggested that recent segmental duplications, which are often involved in chromosome rearrangements underlying genomic disease, account for some 5 of the human genome. We have developed rapid computational heuristics based on BLAST analysis to detect segmental duplications, as well as regions containing potential sequence misassignments in the human genome assemblies. Results Our analysis of the June 2002 public human genome assembly revealed that 107.4 of 3,043.1 megabases (Mb) (3.53) of sequence contained segmental duplications, each with size equal or more than 5 kb and 90 identity. We have also detected that 38.9 Mb (1.28) of sequence within this assembly is likely to be involved in sequence misassignment errors. Furthermore, we have identified a significant subset (199,965 of 2,327,473 or 8.6) of single-nucleotide polymorphisms (SNPs) in the public databases that are not true SNPs but are potential paralogous sequence variants. Conclusion Using two distinct computational approaches, we have identified most of the sequences in the human genome that have undergone recent segmental duplications. Near-identical segmental duplications present a major challenge to the completion of the human genome sequence. Potential sequence misassignments detected in this study would require additional efforts to resolve. The segmental duplication data and summary statistics are available for download. Data for Human Genome (based on the May 2004 Human Genome Assembly (hg17)) Visualize duplication relationships in GBrowse (GBrowse) Duplicon Pair relationships (GFF) Genes within duplication regions (HTML) Genome duplication content (MS Excel) The segmental duplication data can be visualized in a genome browser in the GBrowse section. Selected human genome annotation tracks (except the segmental duplication track) have also been obtained from UCSC and loaded into the genome browser. Detailed information (e.g. overlapping genes, overlapping clones, detailed alignment) can be obtained by clicking on a duplication cluster in GBrowse. Both keyword search and BLAT search are available. Analyses based on previous human genome assemblies can be found in the Previous Analyses section. Acknowledgments We thank The Centre for Applied Genomics at the Hospital for Sick Children (HSC) as well as collaborators worldwide. Supported by Genome Canada the Howard Hughes Medical Institute International Scholar Program (to S.W.S.) and the HSC Foundation.
Proper citation: Human Genome Segmental Duplication Database (RRID:SCR_007728) Copy
Collection of male germ cell transcriptiome information derived from Serial Analysis of Gene Expression (SAGE). It includes the three key germ cell stages in spermatogenesis, including mouse type A spermatogonia (Spga), pachytene spermatocytes (Spcy), and round spermatids (Sptd). A total of 452,095 SAGE tags are represented in all the libraries and is by far the most comprehensive resource available. Users can choose a global view of germ cell transcriptome data in the UCSC Genome browser. They can also search genes or specify searching criteria based on tag sequence, chromosomal location or tag counts.
Proper citation: GermSAGE (RRID:SCR_007689) Copy
Database of compiled, public, deep sequencing miRNA data and several novel tools to facilitate exploration of massive data. The miR-seq browser supports users to examine short read alignment with the secondary structure and read count information available in concurrent windows. Features such as sequence editing, sorting, ordering, import and export of user data are of great utility for studying iso-miRs, miRNA editing and modifications. miRNA����??target relation is essential for understanding miRNA function. Coexpression analysis of miRNA and target mRNAs, based on miRNA-seq and RNA-seq data from the same sample, is visualized in the heat-map and network views where users can investigate the inverse correlation of gene expression and target relations, compiled from various databases of predicted and validated targets.
Proper citation: miRGator (RRID:SCR_007793) Copy
SYSTERS is a database of protein sequences grouped into homologous families and superfamilies. The SYSTERS project aims to provide a meaningful partitioning of the whole protein sequence space by a fully automatic procedure. A refined two-step algorithm assigns each protein to a family and a superfamily. The sequence data underlying SYSTERS release 4 now comprise several protein sequence databases derived from completely sequenced genomes (ENSEMBL, TAIR, SGD and GeneDB), in addition to the comprehensive Swiss-Prot/TrEMBL databases. To augment the automatically derived results, information from external databases like Pfam and Gene Ontology are added to the web server. Furthermore, users can retrieve pre-processed analyses of families like multiple alignments and phylogenetic trees. New query options comprise a batch retrieval tool for functional inference about families based on automatic keyword extraction from sequence annotations. A new access point, PhyloMatrix, allows the retrieval of phylogenetic profiles of SYSTERS families across organisms with completely sequenced genomes. Gene, Human, Vertebrate, Genome, Human ORFs
Proper citation: SYSTERS (RRID:SCR_007955) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.