Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
https://web.uri.edu/riinbre/mic/
Core provides sequencing and bioinformatics support for INBRE and non-INBRE researchers. Provides data science services adjacent to traditional bioinformatics; access to computational and software resources for INBRE network institutions, particularly primarily undergraduate institutions; training for students and faculty in data science methods. Maintains professional network with other core and user facilities in Rhode Island and beyond to maximize resources available to our users.Utilizes novel technologies such as virtual/augmented reality for use in teaching and research.
Proper citation: Rhode Island INBRE Molecular Informatics Core Facility (RRID:SCR_017685) Copy
http://www.sanger.ac.uk/science/tools/ssaha2-0
A program designed for the efficient mapping of sequence reads onto genomic references. The software is capable of reading most sequencing platforms and giving a range of outputs are supported.
Proper citation: Sequence Search and Alignment by Hashing Algorithm (RRID:SCR_000544) Copy
The EBI genomes pages give access to a large number of complete genomes including bacteria, archaea, viruses, phages, plasmids, viroids and eukaryotes. Methods using whole genome shotgun data are used to gain a large amount of genome coverage for an organism. WGS data for a growing number of organisms are being submitted to DDBJ/EMBL/GenBank. Genome entries have been listed in their appropriate category which may be browsed using the website navigation tool bar on the left. While organelles are all listed in a separate category, any from Eukaryota with chromosome entries are also listed in the Eukaryota page. Within each page, entries are grouped and sorted at the species level with links to the taxonomy page for that species separating each group. Within each species, entries whose source organism has been categorized further are grouped and numbered accordingly. Links are made to: * taxonomy * complete EMBL flatfile * CON files * lists of CON segments * Project * Proteomes pages * FASTA file of Proteins * list of Proteins
Proper citation: EBI Genomes (RRID:SCR_002426) Copy
http://www.broad.mit.edu/annotation/fungi/fgi/
Produces and analyzes sequence data from fungal organisms that are important to medicine, agriculture and industry. The FGI is a partnership between the Broad Institute and the wider fungal research community, with the selection of target genomes governed by a steering committee of fungal scientists. Organisms are selected for sequencing as part of a cohesive strategy that considers the value of data from each organism, given their role in basic research, health, agriculture and industry, as well as their value in comparative genomics.
Proper citation: Fungal Genome Initiative (RRID:SCR_003169) Copy
Research oriented service laboratory providing informatics support to research community. Services include data analysis and mining in proteomics, genomics and chemistry, systems biology approaches such as pathway, network and interaction analyses, large scale statistical and machine learning studies, protein structure, function and stability prediction, sequence and domain analyses,d esign and implementation of relational databases and software programs, consultation on experimental design involving data acquisition, management and analysis, report, grant, and manuscript preparation.
Proper citation: Kansas University at Lawrence Applied Bioinformatics Laboratory Core Facility (RRID:SCR_017751) Copy
http://genetics.group.shef.ac.uk/index.html
Core provides DNA sequencing services including DNA extraction, cell line identification, microsatellite analysis, and antibody sequencing,DNA Sequencing, Monoclonal Antibody Sequencing,Nucleic Acid Quantification,PCR Machine Hire,Real-Time PCR Robotic Liquid Handling,Taqman SNP Analysis.
Proper citation: University of Sheffield Genomic Core Facility (RRID:SCR_017912) Copy
https://github.com/BackofenLab/HVSeeker/tree/main
Software tool for distinguishing between bacterial and phage sequences. Consists of two separate models: one analyzing DNA sequences and the other focusing on proteins.
Proper citation: HVSeeker (RRID:SCR_026120) Copy
https://github.com/DerrickWood/kraken2
Software tool as second version of Kraken taxonomic sequence classification system.
Proper citation: kraken2 (RRID:SCR_026838) Copy
https://github.com/SCANDAN-Team/SCANDAN-DICOM-labelling
Software tool for rules for DICOM tag based labelling. Regular expression used during the SCANDAN project to label MRI scans based on DICOM tag.
Proper citation: SCANDAN-DICOM-labelling (RRID:SCR_028365) Copy
http://virome.diagcomputing.org/#view=home
A web-application designed for scientific exploration of metagenome sequence data collected from viral assemblages occurring within a number of different environmental contexts. The VIROME informatics pipeline focuses on the classification of predicted open-reading frames (ORFs) from viral metagenomes. The portal allows you to submit your viral metagenome to be processed through the VIROME analysis pipeline, and enable you to investigate your data via the VIROME user interface., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: VIROME (RRID:SCR_004362) Copy
A collaborative ontology for the definition of sequence features used in biological sequence annotation. SO was initially developed by the Gene Ontology Consortium. Contributors to SO include the GMOD community, model organism database groups such as WormBase, FlyBase, Mouse Genome Informatics group, and institutes such as the Sanger Institute and the EBI. Input to SO is welcomed from the sequence annotation community. The OBO revision is available here: http://sourceforge.net/p/song/svn/HEAD/tree/ SO includes different kinds of features which can be located on the sequence. Biological features are those which are defined by their disposition to be involved in a biological process. Biomaterial features are those which are intended for use in an experiment such as aptamer and PCR_product. There are also experimental features which are the result of an experiment. SO also provides a rich set of attributes to describe these features such as polycistronic and maternally imprinted. The Sequence Ontologies use the OBO flat file format specification version 1.2, developed by the Gene Ontology Consortium. The ontology is also available in OWL from Open Biomedical Ontologies. This is updated nightly and may be slightly out of sync with the current obo file. An OWL version of the ontology is also available. The resolvable URI for the current version of SO is http://purl.obolibrary.org/obo/so.owl.
Proper citation: SO (RRID:SCR_004374) Copy
http://gmod.org/wiki/Main_Page
A collection of open source software tools for creating and managing genome-scale biological databases. GMOD is made up databases, applications, and adaptor software that connects these components together. You can use it to create a small laboratory database of genome annotations, or a large web-accessible community database. At first GMOD just featured model organisms but now any organism with any kind of sequence associated with it is a good candidate as a subject for a GMOD database. There are GMOD databases with just protein sequence in them, with EST sequence only, those that are concerned primarily with gene expression, and even those dedicated to collections of RNA sequence. They have also heard of GMOD databases for oligonucleotides and plasmids.
Proper citation: Generic Model Organism Database Project (RRID:SCR_001731) Copy
Web application to search protein databases using a translated nucleotide query. Translated BLAST services are useful when trying to find homologous proteins to a nucleotide coding region. Blastx compares translational products of the nucleotide query sequence to a protein database. Because blastx translates the query sequence in all six reading frames and provides combined significance statistics for hits to different frames, it is particularly useful when the reading frame of the query sequence is unknown or it contains errors that may lead to frame shifts or other coding errors. Thus blastx is often the first analysis performed with a newly determined nucleotide sequence and is used extensively in analyzing EST sequences. This search is more sensitive than nucleotide blast since the comparison is performed at the protein level.
Proper citation: BLASTX (RRID:SCR_001653) Copy
http://csg.sph.umich.edu//abecasis/MACH/index.html
A Markov Chain based software tool for haplotyping, genotype imputation and disease association analysis that can resolve long haplotypes or infer missing genotypes in samples of unrelated individuals.
Proper citation: MACH 1.0 (RRID:SCR_001759) Copy
https://github.com/benedictpaten/pecan
A Java consistency based multiple sequence alignment software program.
Proper citation: Pecan (RRID:SCR_001909) Copy
http://www.aspergillus-genomes.org.uk/
A resource for viewing annotated genes arising from various Aspergillus sequencing and annotation projects, resulting from the merging of Central Aspergillus Data REpository (CADRE) and The Aspergillus Website, which took place in June 2008. The principal role of CADRE is to aid the Aspergillus research community by managing Aspergillus genome data and by providing visualization tools, ranging from relatively simple annotation displays to more complex data integration displays. In contrast, The Aspergillus Website provides a range of information to the medical community (i.e., clinicians, patients and scientists) regarding the genus Aspergillus and the diseases, such as Aspergillosis, that it can cause. CADRE has been implemented using the Ensembl v22 suite. This suite comprises: * a database schema, which has been devised for storing annotated eukaryotic genomes. The schema is implemented with the MySQL relational database management system. * several specialized programming modules for building interfaces (i.e., BioPerl and Ensembl API modules). * a series of programs (i.e., Perl CGI scripts using the API modules) for viewing genomic data within a web browser., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Aspergillus Genomes (RRID:SCR_001880) Copy
http://rgp.dna.affrc.go.jp/E/index.html
Rice Genome Research Program (RGP) is an integral part of the Japanese Ministry of Agriculture, Forestry and Fisheries (MAFF) Genome Research Project. RGP now aims to completely sequence the entire rice genome and subsequently to pursue integrated goals in functional genomics, genome informatics and applied genomics. It is jointly coordinated by the National Institute of Agrobiological Sciences (NIAS), a government research institute under MAFF and the Society for Techno-innovation of Agriculture, Forestry and Fisheries (STAFF), a semi-private research organization managed and supported by MAFF and a consortium of some twenty Japanese companies. The research is funded with yearly grants from MAFF and additional funds from the Japan Racing Association (JRA). It is now the leading member of the International Rice Genome Sequencing Project (IRGSP), a consortium of ten countries sharing the sequencing of the 12 rice chromosomes. The IRGSP adopts the clone-by-clone shotgun sequencing strategy so that each sequenced clone can be associated with a specific position on the genetic map and adheres to the policy of immediate release of the sequence data to the public domain. In December 2004, the IRGSP completed the sequencing of the rice genome. The high-quality and map-based sequence of the entire genome is now available in public databases.
Proper citation: Rice Genome Research Project (RRID:SCR_002268) Copy
http://www.structuralgenomics.org/
The Structural Genomics Project aims at determination of the 3D structure of all proteins. It also aims to reduce the cost and time required to determine three-dimensional protein structures. It supports selection, registration, and tracking of protein families and representative targets. This aim can be achieved in four steps : -Organize known protein sequences into families. -Select family representatives as targets. -Solve the 3D structure of targets by X-ray crystallography or NMR spectroscopy. -Build models for other proteins by homology to solved 3D structures. PSI has established a high-throughput structure determination pipeline focused on eukaryotic proteins. NMR spectroscopy is an integral part of this pipeline, both as a method for structure determinations and as a means for screening proteins for stable structure. Because computational approaches have estimated that many eukaryotic proteins are highly disordered, about 1 year into the project, CESG began to use an algorithm. The project has been organized into two separate phases. The first phase was dedicated to demonstrating the feasibility of high-throughput structure determination, solving unique protein structures, and preparing for a subsequent production phase. The second phase, PSI-2, has focused on implementing the high-throughput structure determination methods developed in PSI-1, as well as homology modeling and addressing bottlenecks like modeling membrane proteins. The first phase of the Protein Structure Initiative (PSI-1) saw the establishment of nine pilot centers focusing on structural genomics studies of a range of organisms, including Arabidopsis thaliana, Caenorhabditis elegans and Mycobacterium tuberculosis. During this five-year period over 1,100 protein structures were determined, over 700 of which were classified as unique due to their < 30% sequence similarity with other known protein structures. The primary goal of PSI-1 was to develop methods to streamline the structure determination process, resulted in an array of technical advances. Several methods developed during PSI-1 enhanced expression of recombinant proteins in systems like Escherichia coli, Pichia pastoris and insect cell lines. New streamlined approaches to cell cloning, expression and protein purification were also introduced, in which robotics and software platforms were integrated into the protein production pipeline to minimize required manpower, increase speed, and lower costs. The goal of the second phase of the Protein Structure Initiative (PSI-2) is to use methods introduced in PSI-1 to determine a large number of proteins and continue development in streamlining the structural genomics pipeline. Currently, the third phase of the PSI is being developed and will be called PSI: Biology. The consortia will propose work on substantial biological problems that can benefit from the determination of many protein structures Sponsors: PSI is funded by the U.S. National Institute of General Medical Sciences (NIGMS),
Proper citation: Protein Structure Initiative (RRID:SCR_002161) Copy
Collection of genome databases for vertebrates and other eukaryotic species with DNA and protein sequence search capabilities. Used to automatically annotate genome, integrate this annotation with other available biological data and make data publicly available via web. Ensembl tools include BLAST, BLAT, BioMart and the Variant Effect Predictor (VEP) for all supported species.
Proper citation: Ensembl (RRID:SCR_002344) Copy
http://www.ncbi.nlm.nih.gov/SNP/
General database of genetic variations maintained by the NCBI. Database as central repository for both single base nucleotide substitutions and short deletion and insertion polymorphisms. Distinguishes report of how to assay SNP from use of that SNP with individuals and populations. This separation simplifies some issues of data representation. However, these initial reports describing how to assay SNP will often be accompanied by SNP experiments measuring allele occurrence in individuals and populations. Community can contribute to this resource.
Proper citation: dbSNP (RRID:SCR_002338) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.