Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://bix.ucsd.edu/projects/singlecell/
Software package for short read data from single cells that improves assembly through use of progressively increasing coverage cutoff. Used for single cell Illumina sequences, allows variable coverage datasets to be utilized with assembly of E. coli and S. aureus single cell reads. Assembles single cell genome of uncultivated SAR324 clade of Deltaproteobacteria.
Proper citation: Velvet-SC (RRID:SCR_004377) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on December 17, 2021. Database to store, annotate, view, analyze and share microarray data. It provides registered users access to their own data, provides users access to public data, and tools with which to analyze those data, to any public user anywhere in the world. The GenePattern software package has been incorporated directly into SMD, providing access to many new analysis tools, as well as a plug-in architecture that allows users to directly integrate and share additional tools through SMD. This extension is available with the SMD source code that is fully and freely available to others under an Open Source license, enabling other groups to create a local installation of SMD with an enriched data analysis capability. SMD search options allow the user to Search By Experiments, Search By Datasets, or Search By Gene Names. Web services are provided using common standards, such as Simple Object Access Protocol (SOAP). This enables both local and remote researchers to connect to an installation of the database and retrieve data using pre-defined methods, without needing to resort to use of a web browser.
Proper citation: SMD (RRID:SCR_004987) Copy
Only worldwide authority that provides standardized nomenclature, i.e. gene names and symbols (short form abbreviations), for all known human genes, and stores all approved symbols in the HGNC database. Approved human gene nomenclature. Database of gene symbols and names. Manually curated genes into groups based on shared characteristics such as homology, function or phenotype. Data for protein-coding genes, pseudogenes and non-coding RNAs.
Proper citation: HGNC (RRID:SCR_002827) Copy
Central data repository for nematode biology including complete genomic sequence, gene predictions and orthology assignments from range of related nematodes.Data concerning genetics, genomics and biology of C. elegans and related nematodes. Derived from initial ACeDB database of C. elegans genetic and sequence information, WormBase includes genomic, anatomical and functional information of C. elegans, other Caenorhabditis species and other nematodes. Maintains public FTP site where researchers can find many commonly requested files and datasets, WormBase software and prepackaged databases.
Proper citation: WormBase (RRID:SCR_003098) Copy
BioPerl is a community effort to produce Perl code which is useful in biology. This toolkit of perl modules is useful in building bioinformatics solutions in Perl. It is built in an object-oriented manner so that many modules depend on each other to achieve a task. The collection of modules in the bioperl-live repository consist of the core of the functionality of bioperl. Additionally auxiliary modules for creating graphical interfaces (bioperl-gui), persistent storage in RDMBS (bioperl-db), running and parsing the results from hundreds of bioinformatics applications (Run package), software to automate bioinformatic analyses (bioperl-pipeline) are all available as Git modules in our repository. The BioPerl toolkit provides a library of hundreds of routines for processing sequence, annotation, alignment, and sequence analysis reports. It often serves as a bridge between different computational biology applications assisting the user to construct analysis pipelines. This chapter illustrates how BioPerl facilitates tasks such as writing scripts summarizing information from BLAST reports or extracting key annotation details from a GenBank sequence record. BioPerl includes modules written by Sohel Merchant of the GO Consortium for parsing and manipulating OBO ontologies. Platform: Windows compatible, Mac OS X compatible, Linux compatible, Unix compatible
Proper citation: BioPerl (RRID:SCR_002989) Copy
http://compgen.bscb.cornell.edu/phast/
A freely available software package for comparative and evolutionary genomics that consists of about half a dozen major programs, plus more than a dozen utilities for manipulating sequence alignments, phylogenetic trees, and genomic annotations. For the most part, PHAST focuses on two kinds of applications: the identification of novel functional elements, including protein-coding exons and evolutionarily conserved sequences; and statistical phylogenetic modeling, including estimation of model parameters, detection of signatures of selection, and reconstruction of ancestral sequences. It consists of over 60,000 lines of C code.
Proper citation: PHAST (RRID:SCR_003204) Copy
Organization that provides biomedical researchers with online tools and a web portal enabling them to access, review, and integrate disparate ontological resources in all aspects of biomedical investigation and clinical practice. A major focus of the work involves the use of biomedical ontologies to aid in the management and analysis of data derived from complex experiments.
Proper citation: National Center for Biomedical Ontology (RRID:SCR_003304) Copy
http://ccb.jhu.edu/software/FLASH/
Open source software tool to merge paired-end reads from next-generation sequencing experiments. Designed to merge pairs of reads when original DNA fragments are shorter than twice length of reads. Can improve genome assemblies and transcriptome assembly by merging RNA-seq data.
Proper citation: FLASH (RRID:SCR_005531) Copy
http://www.phrap.org/consed/consed.html
A graphical tool for sequence finishing (BAM File Viewer, Assembly Editor, Autofinish, Autoreport, Autoedit, and Align Reads To Reference Sequence)
Proper citation: Consed (RRID:SCR_005650) Copy
Web based integrative platform for transcriptional regulation studies.
Proper citation: Cistrome (RRID:SCR_000242) Copy
http://www.repeatmasker.org/RepeatModeler/
Sequence analysis software that performs repeat family identification and creates models for sequence data. RepeatModeler utilizes RepeatScout and RECON to identify repeat element boundaries and family relationships., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: RepeatModeler (RRID:SCR_015027) Copy
https://github.com/BioDepot/BioDepot-workflow-builder
Software tool to create and execute reproducible bioinformatics workflows using drag and drop interface. Graphical widgets represent Docker containers executing modular task. Widgets are linked graphically to build bioinformatics workflows that can be reproducibly deployed across different local and cloud platforms. Each widget contains form-based user interface to facilitate parameter entry and console to display intermediate results.
Proper citation: BioDepot-workflow-builder (RRID:SCR_017402) Copy
https://ccb.jhu.edu/software/stringtie/
Software application for assembling of RNA-Seq alignments into potential transcripts. It enables improved reconstruction of a transcriptome from RNA-seq reads. This transcript assembling and quantification program is implemented in C++ .
Proper citation: StringTie (RRID:SCR_016323) Copy
https://bitbucket.org/charade/svengine
Software for analysis and simulation of gene sequences and structural variants. This software works with FASTA, FASTQ, BAM, VAR, META, and NEWICK file formats.
Proper citation: SVEngine (RRID:SCR_016235) Copy
http://www.oreganno.org/oregano/
Open source, open access database and literature curation system for community based annotation of experimentally identified DNA regulatory regions, transcription factor binding sites and regulatory variants. Automatically cross referenced against PubMED, Entrez Gene, EnsEMBL, dbSNP, eVOC: Cell type ontology, and Taxonomy database. Community driven resource for curated regulatory annotation.
Proper citation: Open Regulatory Annotation Database (RRID:SCR_007835) Copy
http://www.uniprot.org/help/uniref
Databases which provide clustered sets of sequences from UniProt Knowledgebase and selected UniParc records, in order to obtain complete coverage of sequence space at several resolutions while hiding redundant sequences from view. The UniRef100 database combines identical sequences and sub-fragments with 11 or more residues (from any organism) into a single UniRef entry. The sequence of a representative protein, the accession numbers of all the merged entries, and links to the corresponding UniProtKB and UniParc records are all displayed in the entry. UniRef90 and UniRef50 are built by clustering UniRef100 sequences with 11 or more residues such that each cluster is composed of sequences that have at least 90% (UniRef90) or 50% (UniRef50) sequence identity to the longest sequence (UniRef seed sequence). All the sequences in each cluster are ranked to facilitate the selection of a representative sequence for the cluster.
Proper citation: UniRef (RRID:SCR_010646) Copy
http://ccr.coriell.org/Sections/Collections/NHGRI/?SsId=11
DNA samples and cell lines from fifteen populations, including the samples used for the International HapMap Project, the HapMap 3 Project and the 1000 Genomes Project (except for the CEPH samples). All of the samples were contributed with consent to broad data release and to their use in many future studies, including for extensive genotyping and sequencing, gene expression and proteomics studies, and all other types of genetic variation research. NHGRI led the contribution of the NIH to the International HapMap Project, which developed a haplotype map of the human genome. This haplotype map, called the HapMap is a publicly available tool that allows researchers to find genes and genetic variations that affect health and disease. The samples from four populations used to develop the HapMap were initially housed in the Human Genetic Cell Repository of the National Institute of General Medical Sciences (NIGMS). Except for the Utah CEPH samples that were in the NIGMS Repository before the initiation of the HapMap Project and remain there, the NHGRI Repository now houses all of the HapMap samples. The NHGRI repository also houses the extended set of HapMap samples, which includes additional samples from the HapMap populations and samples from seven additional populations. All of the samples were collected with extensive community engagement, including discussions with members of the donor communities about the ethical and social implications of human genetic variation research. These samples were studied as part of the HapMap 3 Project. The NHGRI repository also houses the samples for the International 1000 Genomes Project. This Project is lightly sequencing genome-wide 2500 samples from 27 populations. This project aims to provide a detailed map of human genetic variation, including common and rare SNPs and structural variants. This map will allow more precise localization of genomic regions that contribute to health and disease. The 1000 Genomes Project includes many of the samples from the HapMap and extended set of HapMap samples, as well as samples being collected from additional populations. Currently, samples from five additional populations are available; the others will become available during 2011 and 2012. No identifying or phenotypic information is available for the samples. Donors gave broad consent for use of the samples, including for genotyping, sequencing, and cellular phenotype studies. Samples collected from other populations for the study of human genetic variation may be added to the collection in the future. The NHGRI Repository distributes high quality lymphoblastoid cell lines and DNA from the samples to researchers. DNA is provided in plates or panels of 70 to 100 samples or as individual samples. Cell cultures and DNA samples are distributed only to qualified professional persons who are associated with recognized research, medical, educational, or industrial organizations engaged in health-related research or health delivery.
Proper citation: NHGRI Sample Repository for Human Genetic Research (RRID:SCR_004528) Copy
Software tool to enable biologists without training in computer vision or programming to quantitatively measure phenotypes from thousands of images automatically. It counts cells and also measures the size, shape, intensity and texture of every cell (and every labeled subcellular compartment) in every image. It was designed for high throughput screening but can perform automated image analysis for images from time-lapse movies and low-throughput experiments. CellProfiler has an increasing number of algorithms to identify and measure properties of neuronal cell types.
Proper citation: CellProfiler Image Analysis Software (RRID:SCR_007358) Copy
Model organism database that serves as central repository and web-based resource for zebrafish genetic, genomic, phenotypic and developmental data. Data represented are derived from three primary sources: curation of zebrafish publications, individual research laboratories and collaborations with bioinformatics organizations. Data formats include text, images and graphical representations.Serves as primary community database resource for laboratory use of zebrafish. Developed and supports integrated zebrafish genetic, genomic, developmental and physiological information and link this information extensively to corresponding data in other model organism and human databases.
Proper citation: Zebrafish Information Network (ZFIN) (RRID:SCR_002560) Copy
http://sonorus.princeton.edu/hefalmp/
HEFalMp (Human Experimental/FunctionAL MaPper) is a tool developed by Curtis Huttenhower in Olga Troyanskaya's lab at Princeton University. It was created to allow interactive exploration of functional maps. Functional mapping analyzes portions of these networks related to user-specified groups of genes and biological processes and displays the results as probabilities (for individual genes), functional association p-values (for groups of genes), or graphically (as an interaction network). HEFalMp contains information from roughly 15,000 microarray conditions, over 15,000 publications on genetic and physical protein interactions, and several types of DNA and protein sequence analyses and allows the exploration of over 200 H. sapiens process-specific functional relationship networks, including a global, process-independent network capturing the most general functional relationships. Looking to download functional maps? Keep an eye on the bottom of each page of results: every functional map of any kind is generated with a Download link at the bottom right. Most functional maps are provided as tab-delimited text to simplify downstream processing; graphical interaction networks are provided as Support Vector Graphics files, which can be viewed using the Adobe Viewer, any recent version of Firefox, or the excellent open source Inkscape tool.
Proper citation: Human Experimental/FunctionAL MaPper: Providing Functional Maps of the Human Genome (RRID:SCR_003506) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.