Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://bibiserv.techfak.uni-bielefeld.de/dialign/
Tool for multiple sequence alignment using various sources of external information that is particularly useful to detect local homologies in sequences with low overall similarity. While standard alignment methods rely on comparing single residues and imposing gap penalties, DIALIGN constructs pairwise and multiple alignments by comparing entire segments of the sequences. No gap penalty is used. This approach can be used for both global and local alignment, but it is particularly successful in situations where sequences share only local homologies. Several versions of DIALIGN are available online at GOBICS, http://dialign.gobics.de/
Proper citation: DIALIGN (RRID:SCR_003041) Copy
http://resexomedb.bioinf-dz.org/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on October 28,2025. An online catalog for whole-exome sequencing (WES) results including mutations and gene-disease associations identified by WES. It is browsable and searchable by mutation, gene, study or publication. In addition, it centralizes all publications, software, platforms related to exome / whole genome sequencing.
Proper citation: resExomeDB (RRID:SCR_003224) Copy
http://compbio.cs.sfu.ca/software-novelseq
Software pipeline to detect novel sequence insertions using high throughput paired-end whole genome sequencing data.
Proper citation: NovelSeq (RRID:SCR_003136) Copy
http://mrcanavar.sourceforge.net/
Copy number caller that analyzes the whole-genome next-generation sequence mapping read depth to discover large segmental duplications and deletions. It also has the capability of predicting absolute copy numbers of genomic intervals.
Proper citation: mrCaNaVaR (RRID:SCR_003135) Copy
Nematode & Neglected Genomics (at) The Blaxter Lab is a nematode related portal including databases and services. Resources include genomic and transcriptomic databases for nematodes and other metazoan phyla and freely downloadable software tools for expressed sequence tag analysis, DNA barcode analysis and phylogenomics. Major categories include: * GenePool * 959 Nematode Genomes * Teaching * Research Projects * Bioinformatics Software Tools * Lab Personnel * Lab Wiki * Genomics Databases * NEMBASE4 * Tardigrada: Hypsibius dujardini * Earthworm: Lumbricus rubellus * MolluscDB * ArthropodDB * other Neglected Genomes
Proper citation: nematodes.org (RRID:SCR_003267) Copy
A biopharmaceutical company applying its discoveries in human genetics to develop drugs and diagnostics for common diseases. They specialize in gene discovery - their population approach and resources have enabled them to isolate key genes contributing to major public health challenges from cardiovascular disease to cancer. The company's genotyping capacity is now one of the highest in the world. They have a large population-based biobank containing whole blood and DNA samples with extensive relevant phenotypic information from around 120.000 Icelanders. In the company's work in more than 50 disease projects, their statistical and informatics departments have established themselves in data processing and analysis. deCODE genetics is widely recognized as a center of excellence in genetic research.
Proper citation: deCODE genetics (RRID:SCR_003334) Copy
http://www.lgm.upmc.fr/parseq/
Statistical software for transcription landscape reconstruction at a basepair resolution from RNA Seq read counts. It is based on a state-space model which describes, in terms of abrupt shifts and more progressive drifts, the transcription level dynamics along the genome. Alongside variations of transcription level, it incorporates a component of short-range variation to pull apart local artifacts causing correlated dispersion. Reconstruction of the transcription level relies on a conditional sequential Monte Carlo approach that is combined with parameter estimation in a Markov chain Monte Carlo algorithm known as particle Gibbs. The method allows to estimate the local transcription level, to call transcribed regions, and to identify the transcript borders.
Proper citation: Parseq (RRID:SCR_003464) Copy
http://bejerano.stanford.edu/phenotree/
Web server to search for genes involved in given phenotypic difference between mammalian species. The mouse-referenced multiple alignment data files used to perform the forward genomics screen is also available. The webserver implements one strategy of a Forward Genomics approach aiming at matching phenotype to genotype. Forward genomics matches a given pattern of phenotypic differences between species to genomic differences using a genome-wide screen. In the implementation, the divergence of the coding region of genes in mammals is measured. Given an ancestral phenotypic trait that is lost in independent mammalian lineages, it is shown that searching for genes that are more diverged in all trait-loss species can discover genes that are involved in the given phenotype.
Proper citation: Phenotree (RRID:SCR_003591) Copy
Maintains and provides archival, retrieval and analytical resources for biological information. Central DDBJ resource consists of public, open-access nucleotide sequence databases including raw sequence reads, assembly information and functional annotation. Database content is exchanged with EBI and NCBI within the framework of the International Nucleotide Sequence Database Collaboration (INSDC). In 2011, DDBJ launched two new resources: DDBJ Omics Archive and BioProject. DOR is archival database of functional genomics data generated by microarray and highly parallel new generation sequencers. Data are exchanged between the ArrayExpress at EBI and DOR in the common MAGE-TAB format. BioProject provides organizational framework to access metadata about research projects and data from projects that are deposited into different databases.
Proper citation: DNA DataBank of Japan (DDBJ) (RRID:SCR_002359) Copy
http://www.ncbi.nlm.nih.gov/genome
Database that organizes information on genomes including sequences, maps, chromosomes, assemblies, and annotations in six major organism groups: Archaea, Bacteria, Eukaryotes, Viruses, Viroids, and Plasmids. Genomes of over 1,200 organisms can be found in this database, representing both completely sequenced organisms and those for which sequencing is in progress. Users can browse by organism, and view genome maps and protein clusters. Links to other prokaryotic and archaeal genome projects, as well as BLAST tools and access to the rest of the NCBI online resources are available.
Proper citation: NCBI Genome (RRID:SCR_002474) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on March 17, 2022. A secure repository for storing, cataloging, and accessing cancer genome sequences, alignments, and mutation information from the Cancer Genome Atlas (TCGA) consortium and related projects. CGHub gives scientific researchers the statistical power of large cancer genome datasets to attack the molecular complexity of cancer.
Proper citation: Cancer Genomics Hub (RRID:SCR_002657) Copy
http://bioweb.ensam.inra.fr/esther
Database and tools for analysis of protein and nucleic acid sequences belonging to superfamily of alpha/beta hydrolases homologous to cholinesterases. Covers multiple species, including human, mouse caenorhabditis and drosophila., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: ESTHER (RRID:SCR_002621) Copy
http://www.nitrc.org/projects/penncnv
A free software tool for Copy Number Variation (CNV) detection from SNP genotyping arrays. Currently it can handle signal intensity data from Illumina and Affymetrix arrays. With appropriate preparation of file format, it can also handle other types of SNP arrays and oligonucleotide arrays. PennCNV implements a hidden Markov model (HMM) that integrates multiple sources of information to infer CNV calls for individual genotyped samples. It differs form segmentation-based algorithm in that it considered SNP allelic ratio distribution as well as other factors, in addition to signal intensity alone. In addition, PennCNV can optionally utilize family information to generate family-based CNV calls by several different algorithms. Furthermore, PennCNV can generate CNV calls given a specific set of candidate CNV regions, through a validation-calling algorithm.
Proper citation: PennCNV (RRID:SCR_002518) Copy
http://genespeed.ccf.org/home/
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 16, 2013. Database and customized tools to study the PFAM protein domain content of the transcriptome for all expressed genes of Homo sapiens, Mus musculus, Drosophila melanogaster, and Caenorhabditis elegans tethered to both a genomics array repository database and a range of external information resources. GeneSpeed has merged information from several existing data sets including the Gene Ontology Consortium, InterPro, Pfam, Unigene, as well as micro-array datasets. GeneSpeed is a database of PFAM domain homology contained within Unigene. Because Unigene is a non-redundant dbEST database, this provides a wide encompassing overview of the domain content of the expressed transcriptome. We have structured the GeneSpeed Database to include a rich toolset allowing the investigator to study all domain homology, no matter how remote. As a result, homology cutoff score decisions are determined by the scientist, not by a computer algorithm. This quality is one of the novel defining features of the GeneSpeed database giving the user complete control of database content. In addition to a domain content toolset, GeneSpeed provides an assortment of links to external databases, a unique and manually curated Transcription Factor Classification list, as well as links to our newly evolving GeneSpeed BetaCell Database. GeneSpeed BetaCell is a micro-array depository combined with custom array analysis tools created with an emphasis around the meta analysis of developmental time series micro-array datasets and their significance in pancreatic beta cells.
Proper citation: GeneSpeed- A Database of Unigene Domain Organization (RRID:SCR_002779) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 14,2026. Integrated database of genomic, expression and protein data for Drosophila, Anopheles, C. elegans and other organisms. You can run flexible queries, export results and analyze lists of data. FlyMine presents data in categories, with each providing information on a particular type of data (for example Gene Expression or Protein Interactions). Template queries, as well as the QueryBuilder itself, allow you to perform searches that span data from more than one category. Advanced users can use a flexible query interface to construct their own data mining queries across the multiple integrated data sources, to modify existing template queries or to create your own template queries. Access our FlyMine data via our Application Programming Interface (API). We provide client libraries in the following languages: Perl, Python, Ruby and & Java API
Proper citation: FlyMine (RRID:SCR_002694) Copy
Database that collects, integrates and links all relevant primary information from the GABI plant genome research projects and makes them accessible via internet. Its purpose is to support plant genome research in Germany, to yield information about commercial important plant genomes, and to establish a scientific network within plant genomic research.
GreenCards is the main interface for text based retrieval of sequence, SNP, mapping data etc. Sharing and interchange of data among collaborating research groups, industry and the patent- and licensing agency are facilitated.
* GreenCards: Text based search for sequence, mapping, SNP data etc. * Maps: Visualization of genetic or physical maps. * BLAST: Secure BLAST search against different public databases or non-public sequence data stored in GabiPD. * Proteomics: View interactive 2D-gels and view or download information for identified protein spots. Registered users can submit data via secure file upload.
Proper citation: Gabi Primary Database (RRID:SCR_002755) Copy
Database of information regarding genome and metagenome sequencing projects, and their associated metadata, around the world. It also provides information related to organism properties such as phenotype, ecotype and disease. Both complete and ongoing projects, along with their associated metadata, can be accessed. Users can also register, annotate and publish genome and metagenome data.
Proper citation: Genomes Online Database (RRID:SCR_002817) Copy
A database designed for plant comparative and functional genomics based on complete genomes. It comprises complete proteome sequences from the major phylum of plant evolution. The clustering of these proteomes was performed to define a consistent and extensive set of homeomorphic plant families. Based on this, lists of gene families such as plant or species specific families and several tools are provided to facilitate comparative genomics within plant genomes. The analyses follow two main steps: gene family clustering and phylogenomic analysis of the generated families. Once a group of sequences (cluster) is validated, phylogenetic analyses are performed to predict homolog relationships such as orthologs and ultraparalogs.
Proper citation: GreenPhylDB (RRID:SCR_002834) Copy
Portal that supports Ambystoma-related research and educational efforts. It is composed of several resources: Salamander Genome Project, Ambystoma EST Database, Ambystoma Gene Collection, Ambystoma Map and Marker Collection, Ambystoma Genetic Stock Center, and Ambystoma Research Coordination Network.
Proper citation: Sal-Site (RRID:SCR_002850) Copy
Computational biology research at Memorial Sloan-Kettering Cancer Center (MSKCC) pursues computational biology research projects and the development of bioinformatics resources in the areas of: sequence-structure analysis; gene regulation; molecular pathways and networks, and diagnostic and prognostic indicators. The mission of cBio is to move the theoretical methods and genome-scale data resources of computational biology into everyday laboratory practice and use, and is reflected in the organization of cBio into research and service components ~ the intention being that new computational methods created through the process of scientific inquiry should be generalized and supported as open-source and shared community resources. Faculty from cBio participate in graduate training provided through the following graduate programs: * Gerstner Sloan-Kettering Graduate School of Biomedical Sciences * Graduate Training Program in Computational Biology and Medicine Integral to much of the research and service work performed by cBio is the creation and use of software tools and data resources. The tools that we have created and utilize provide evidence of our involvement in the following areas: * Cancer Genomics * Data Repositories * iPhone & iPod Touch * microRNAs * Pathways * Protein Function * Text Analysis * Transcription Profiling
Proper citation: Computational Biology Center (RRID:SCR_002877) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.