Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://www.ebi.ac.uk/thornton-srv/databases/WSsas/
SAS is a tool for applying structural information to a given protein sequence. It uses FASTA to scan a given protein sequence against all the proteins of known 3D structure in the Protein Data Bank and provides functional residue annotation based on data from the Catalytic Site Atlas and PDBsum. The web service is aimed to facilitate the use of the SAS tool when having a huge number of queries. Currently, the web service provides annotation for binding sites (to ligand, metal or nucleic acid), catalytic residues and amino acids related to protein-protein interactions.
Proper citation: WSsas - Web Service for the SAS tool (RRID:SCR_007051) Copy
http://weizhong-lab.ucsd.edu/cd-hit/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software program for clustering biological sequences with many applications in various fields such as making non-redundant databases, finding duplicates, identifying protein families, filtering sequence errors and improving sequence assembly etc. It is very fast and can handle extremely large databases. CD-HIT helps to significantly reduce the computational and manual efforts in many sequence analysis tasks and aids in understanding the data structure and correct the bias within a dataset. The CD-HIT package has CD-HIT, CD-HIT-2D, CD-HIT-EST, CD-HIT-EST-2D, CD-HIT-454, CD-HIT-PARA, PSI-CD-HIT, CD-HIT-OTU and over a dozen scripts. * CD-HIT (CD-HIT-EST) clusters similar proteins (DNAs) into clusters that meet a user-defined similarity threshold. * CD-HIT-2D (CD-HIT-EST-2D) compares 2 datasets and identifies the sequences in db2 that are similar to db1 above a threshold. * CD-HIT-454 identifies natural and artificial duplicates from pyrosequencing reads. * CD-HIT-OTU cluster rRNA tags into OTUs The usage of other programs and scripts can be found in CD-HIT user''s guide. CD-HIT was originally developed by Dr. Weizhong Li at Dr. Adam Godzik''s Lab at the Burnham Institute (now Sanford-Burnham Medical Research Institute)., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: CD-HIT (RRID:SCR_007105) Copy
This database provides a platform to query and compare gene expression data during the development of the major model animals (zebrafish, drosophila, medaka, mouse). The name 4DXpress stands for expression database in 4D. The 4D (four dimensions) of 4DXpress can be interpreted either as: 3 spatial dimensions plus time, or as 1. species 2. gene 3. developmental stage 4. anatomical structure. The major focus of this database lies in cross species comparison. The high resolution expression data was acquired through whole mount in situ hybridsation-, antibody- or transgenic experiments. Data was integrated from several species specific expression pattern databases, such as ZFIN, BDGP, GXD, MEPD as well as directly submitted by researchers of the participating groups at EMBL. The 4DXpress database is a project within the Centre for Computational Biology at EMBL. It is developed by Yannick Haudry, Thorsten Henrich and Ivica Letunic and coordinated by Thorsten Henrich. Hugo Berube is developing the 4D ArrayExpress Data Warehouse at EBI for integrating in situ data with microarray data.
Proper citation: Expression Database in 4D (RRID:SCR_007066) Copy
Database containing the DNA sequence and annotation of the entire human chromosome 7, encompassing nearly 158 million nucleotides of DNA and 1917 gene structures, are presented; the most up to date collation of sequence, gene, and other annotations from all databases (eg. Celera published, NCBI, Ensembl, RIKEN, UCSC) as well as unpublished data. To generate a higher order description, additional structural features such as imprinted genes, fragile sites, and segmental duplications were integrated at the level of the DNA sequence with medical genetic data, including 440 chromosome rearrangement breakpoints associated with disease. The objective of this project is to generate a comprehensive description of human chromosome 7 to facilitate biological discovery, disease gene research and medical genetic applications. There are over 360 disease-associated genes or loci on chromosome 7. A major challenge ahead will be to represent chromosome alterations, variants, and polymorphisms and their related phenotypes (or lack thereof), in an accessible way. In addition to being a primary data source, this site serves as a weighing station for testing community ideas and information to produce highly curated data to be submitted to other databases such as NCBI, Ensembl, and UCSC. Therefore, any useful data submitted will be curated and shown in this database. All Chromosome 7 genomic clones (cosmids, BACs, YACs) listed in GBrowser and in other data tables are freely distributed.
Proper citation: Chromosome 7 Annotation Project (RRID:SCR_007134) Copy
http://zhanglab.ccmb.med.umich.edu/I-TASSER/
Web server as integrated platform for automated protein structure and function prediction. Used for protein 3D structure prediction. Resource for automated protein structure prediction and structure-based function annotation.
Proper citation: I-TASSER (RRID:SCR_014627) Copy
http://cole-trapnell-lab.github.io/cufflinks/cuffmerge/
Software tool for transcriptome assembly and differential expression analysis for RNA-Seq. Includes script called cuffmerge that can be used to merge together several Cufflinks assemblies. It also handles running Cuffcompare as well as automatically filtering a number of transfrags that are likely to be artifacts. If the researcher has a reference GTF file, the researcher can provide it to the script to more effectively merge novel isoforms and maximize overall assembly quality.
Proper citation: Cufflinks (RRID:SCR_014597) Copy
A SEED-quality automated service that annotates complete or nearly complete bacterial and archaeal genomes across the entire phylogenetic tree. RAST can also be used to analyze draft genomes.
Proper citation: RAST Server (RRID:SCR_014606) Copy
http://bix.ucsd.edu/repeatscout/
Algorithm used to identify de novo repeat families in newly sequenced genomes. Repeat libraries for C. briggsae, M. muscles (X chromosome), R. novegicus (X chromosome), armadillo, H. sapiens (X chromosome), and various other mammals created using RepeatScout are available on the main site., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: RepeatScout (RRID:SCR_014653) Copy
http://www.vicbioinformatics.com/software.prokka.shtml
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software tool for the rapid annotation of prokaryotic genomes. It produces GFF3, GBK and SQN files that are ready for editing in Sequin and ultimately submitted to Genbank/DDJB/ENA. A typical 4 Mbp genome can be fully annotated in less than 10 minutes on a quad-core computer, and scales well to 32 core SMP systems., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Prokka (RRID:SCR_014732) Copy
http://www.sailing.cs.cmu.edu/main/?page_id=511
Automatic software program for profiling spatial gene expression patterns from Fly embryo ISH images. It utilizes image-based genome-scale profiling of whole-body mRNA patterns.
Proper citation: SPEX2 (RRID:SCR_014923) Copy
https://dogma.ccbb.utexas.edu/
Web-based annotation tool for plant chloroplasts and animal mitochondrial genomes. DOGMA allows the use of BLAST searches against a custom database, and conservation of basepairing in the secondary structure of animal mitochondrial tRNAs to identify and annotate genes.
Proper citation: DOGMA (RRID:SCR_015060) Copy
http://bioinformatics.psb.ugent.be/orcae/
Online genome annotation tool for validating and correcting gene annotations. OrcAE is community-driven and can be edited by account-holders in the research community.
Proper citation: Online Resource for Community Annotation of Eukaryotes (RRID:SCR_014989) Copy
http://www.mybiosoftware.com/seaview-4-2-12-sequence-alignment-phylogenetic-tree-building.html
Graphical user interface for multiple sequence alignment and molecular phylogeny. SeaView also generates phylogenetic trees.
Proper citation: SeaView (RRID:SCR_015059) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on June 29,2023. Software tool for the analysis of cross-linking/mass spectrometry datasets using MS-cleavable cross-linkers. MeroX is specialized for MS/MS-cleavable cross linking reagents and identifies the specific fragmentation products of the cleavable cross links.
Proper citation: MeroX (RRID:SCR_014956) Copy
Software tool to quantitatively measure genome assembly and annotation completeness based on evolutionarily informed expectations of gene content.
Proper citation: BUSCO (RRID:SCR_015008) Copy
http://floresta.eead.csic.es/primers4clades
Web application for the design of PCR primers for cross-species amplification of novel sequences from metagenomic DNA or from uncharacterized organisms belonging to user-specified phylogenetic lineages. It implements an extended CODEHOP strategy and evaluates thermodynamic properties of the oligonucleotide pairs.
Proper citation: primers4clades (RRID:SCR_015714) Copy
http://amp.pharm.mssm.edu/clustergrammer/
Clustergrammer is a web-based tool for visualizing and analyzing high-dimensional data as interactive and shareable hierarchically clustered heatmaps. Clustergrammer enables intuitive exploration of high-dimensional data and has several optional biology-specific features.
Proper citation: clustergrammer (RRID:SCR_015681) Copy
https://bioconductor.org/packages/release/bioc/html/oligo.html
Software package to analyze oligonucleotide arrays (expression/SNP/tiling/exon) at probe-level. It currently supports Affymetrix (CEL files) and NimbleGen arrays (XYS files).
Proper citation: oligo (RRID:SCR_015729) Copy
https://github.com/BGI-SZ/BSVF
Software code for bisulfite sequencing virus integration. This finder is for directional libraries only and does not support PBAT and indirectional libraries.
Proper citation: BSVF (RRID:SCR_015727) Copy
http://www.genepattern-notebook.org/
Interactive analysis notebook environment that streamlines genomics research by interleaving text, multimedia, and executable code into unified, sharable, reproducible “research narratives.” It integrates the dynamic capabilities of notebook systems with an investigator-focused, simple interface that provides access to hundreds of genomic tools without the need to write code.
Proper citation: GenePattern Notebook (RRID:SCR_015699) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.