Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://compbio.cs.sfu.ca/software-variation-hunter
A software tool for discovery of structural variation in one or more individuals simultaneously using high throughput technologies.
Proper citation: VariationHunter (RRID:SCR_004865) Copy
http://www.cbcb.umd.edu/software/phymm/
Software for Phylogenetic Classification of Metagenomic Data with Interpolated Markov Models to taxonomically classify DNA sequences and accurately classify reads as short as 100 bp. PhymmBL, the hybrid classifier included in this distribution which combines analysis from both Phymm and BLAST, produces even higher accuracy.
Proper citation: Phymm and PhymmBL (RRID:SCR_004751) Copy
https://code.google.com/p/destruct/
A software tool for identifying structural variation in tumour genomes from whole genome illumina sequencing.
Proper citation: deStruct (RRID:SCR_004747) Copy
http://www.genedb.org/Homepage/Tbruceibrucei927
Database of the most recent sequence updates and annotations for the T. brucei genome. New annotations are constantly being added to keep up with published manuscripts and feedback from the Trypanosomatid research community. You may search by Protein Length, Molecular Mass, Gene Type, Date, Location, Protein Targeting, Transmembrane Helices, Product, GO, EC, Pfam ID, Curation and Comments, and Dbxrefs. BLAST and other tools are available. T. brucei possesses a two-unit genome, a nuclear genome and a mitochondrial (kinetoplast) genome with a total estimated size of 35Mb/haploid genome. The nuclear genome is split into three classes of chromosomes according to their size on pulsed-field gel electrophoresis, 11 pairs of megabase chromosomes (0.9-5.7 Mb), intermediate (300-900 kb) and minichromosomes (50-100 kb). The T. brucei genome contains a ~0.5Mb segmental duplication affecting chromosomes 4 and 8, which is responsible for some 75 gene duplicates unique to this species. A comparative chromosome map of the duplicons can be accessed here (PubmedID 18036214). Protozoan parasites within the species Trypanosoma brucei are the etiological agent of human sleeping sickness and Nagana in animals. Infections are limited to patches of sub-Saharan Africa where insects vectors of the Glossina genus are endemic. The most recent estimates indicate between 50,000 - 70,000 human cases currently exist, with 17 000 new cases each year (WHO Factsheet, 2006). In collaboration with GeneDB, the EuPathDB genomic sequence data and annotations are regularly deposited on TriTrypDB where they can be integrated with other datasets and queried using customized queries.
Proper citation: GeneDB Tbrucei (RRID:SCR_004786) Copy
http://www.baseclear.com/landingpages/basetools-a-wide-range-of-bioinformatics-solutions/sspacev12/
A stand-alone software program for scaffolding pre-assembled contigs using paired-read data. Main features are: a short runtime, multiple library input of paired-end and/or mate pair datasets and possible contig extension with unmapped sequence reads.
Proper citation: SSPACE (RRID:SCR_005056) Copy
http://www.biomedcentral.com/1471-2105/13/189
An algorithm to use optical map information directly within the de Bruijn graph framework to help produce an accurate assembly of a genome that is consistent with the optical map information provided. AGORA takes as input two data structures: OpMap ? an ordered list of fragment sizes representing the optical map; and Edges ? a list of de Bruijn graph edges with their corresponding sequences.
Proper citation: AGORA (RRID:SCR_005070) Copy
https://github.com/tk2/RetroSeq
A tool for discovery and genotyping of transposable element variants (TEVs) (also known as mobile element insertions) from next-gen sequencing reads aligned to a reference genome in BAM format. The goal is to call TEVs that are not present in the reference genome but present in the sample that has been sequenced. It should be noted that RetroSeq can be used to locate any class of viral insertion in any species where whole-genome sequencing data with a suitable reference genome is available. RetroSeq is a two phase process, the first being the read pair discovery phase where discorandant mate pairs are detected and assigned to a TE class (Alu, SINE, LINE, etc.) by using either the annotated TE elements in the reference and/or aligned with Exonerate to the supplied library of viral sequences.
Proper citation: RetroSeq (RRID:SCR_005133) Copy
http://bioinfo.mc.vanderbilt.edu/VirusFinder/
Software tool for efficient and accurate detection of viruses and their integration sites in host genomes through next generation sequencing data. Specifically, it detects virus infection, co-infection with multiple viruses, virus integration sites in host genomes, as well as mutations in the virus genomes. It also facilitates virus discovery by reporting novel contigs, long sequences assembled from short reads that map neither to the host genome nor to the genomes of known viruses. VirusFinder 2 works with both paired-end and single-end data, unlike the previous 1.x versions that accepted only paired-end reads. The types of NGS data that VirusFinder 2 can deal with include whole genome sequencing (WGS), whole transcriptome sequencing (RNA-Seq), targeted sequencing data such as whole exome sequencing (WES) and ultra-deep amplicon sequencing.
Proper citation: VirusFinder (RRID:SCR_005205) Copy
NIH established expectations for sharing data obtained through NIH-funded genome-wide association studies (GWAS) with the implementation of the GWAS Policy. Information and resources related to the GWAS Policy can be found on this website.
Proper citation: Genomic Datasharing (RRID:SCR_005233) Copy
http://seqant.genetics.emory.edu/
A free web service and open source software package that performs rapid, automated annotation of DNA sequence variants (single base mutations, insertions, deletions) discovered with any sequencing platform. Variant sites are characterized with respect to their functional type (Silent, Replacement, 5' UTR, 3' UTR, Intronic, Intergenic), whether they have been previously submitted to dbSNP, and their evolutionary conservation. Annotated variants can be viewed directly on the web browser, downloaded in a tab delimited text file, or directly uploaded in a Browser Extended Data (BED) format to the UCSC genome browser. SeqAnt further identifies all loci harboring two or more coding sequence variants that help investigators identify potential compound heterozygous loci within exome sequencing experiments. In total, SeqAnt resolves a significant bottleneck by allowing an investigator to rapidly prioritize the functional analysis of those variants of interest.
Proper citation: SeqAnt (RRID:SCR_005186) Copy
http://stothard.afns.ualberta.ca/downloads/NGS-SNP/
A collection of command-line scripts for providing rich annotations for SNPs identified by the sequencing of transcripts or whole genomes from organisms with reference sequences in Ensembl. Included among the annotations, several of which are not available from any existing SNP annotation tools, are the results of detailed comparisons with orthologous sequences. These comparisons allow, for example, SNPs to be sorted or filtered based on how drastically the SNP changes the score of a protein alignment. Other fields indicate the names of overlapping protein domains or features, and the conservation of both the SNP site and flanking regions. NCBI, Ensembl, and Uniprot IDs are provided for genes, transcripts, and proteins when applicable, along with Gene Ontology terms, a gene description, phenotypes linked to the gene, and an indication of whether the SNP is novel or known. A ?Model_Annotations? field provides several annotations obtained by transferring in silico the SNP to an orthologous gene, typically in a well-characterized species.
Proper citation: NGS-SNP (RRID:SCR_005182) Copy
http://www.animalgenome.org/pig/genome/db/
Database facilitating information integration and mining within the pig and across species of all genomics / genetics research results accumulated over the years including pig gene expression, quantitative trait loci (QTL), candidate gene, and whole genome association study (WGAS) results. The key functions developed so far include pig gene pages (a centralized gene search tool), a local copy of Biomart (for customizable genome information queries), genome feature alignment tools (Pig QTLdb and Gbrowse), integrated gene expression information (ANEXDB and ESTdb), a dedicated pig genome and gene set BLAST server, and virtual comparative map database and tools (VCmap). By developing the PGD, it is our aim to collaboratively utilize existing databases and tools via networked functions, such as web services, database API, etc., to maximize the potential of all related databases through the PGD implementation.
Proper citation: Pig Genome Database (RRID:SCR_006367) Copy
A comparative platform for green plant genomics. Families of orthologous and paralogous genes that represent the modern descendents of ancestral gene sets are constructed at key phylogenetic nodes. These families allow easy access to clade specific orthology / paralogy relationships as well as clade specific genes and gene expansions. As of release v9.1, Phytozome provides access to forty-one sequenced and annotated green plant genomes which have been clustered into gene families at 20 evolutionarily significant nodes. Where possible, each gene has been annotated with PFAM, KOG, KEGG, and PANTHER assignments, and publicly available annotations from RefSeq, UniProt, TAIR, JGI are hyper-linked and searchable., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Phytozome (RRID:SCR_006507) Copy
We at NRSP-8 bioinformatics coordination program strive to serve the animal genomics research community to better use computer tools and methods, to best utilize available resources, and in working with researchers in the community, to effectively share, combine, manage, manipulate, and analyze information from genomics/genetics studies. This site is designed as an information center to serve the national animal genome research projects of cattle, chicken, pigs, sheep, horse, and aquaculture species. This is home to databases and web sites (being) built for structural, functional and application oriented studies of the animal genomics, to serve the purpose of research, education and related activities in the scientific, industrial and educational communities in the states and world wide. The challenges in bioinformatics support/research for animal genomics may involve * Effective data collection, organization and management * Rapid development of most needed bioinformatics tools and resources * Efficient use of these tools for innovative data analysis Projects: * Animal Trait Ontology (ATO) Project * Virtual Comparative Genomics * The Past, the Current, and the Potentials * Collaborative and Hosted Works
Proper citation: NAGRP Bioinformatics Coordination Program (RRID:SCR_006564) Copy
http://www.ncbi.nlm.nih.gov/projects/genome/assembly/grc/
Consortium that puts sequences into a chromosome context and provides the best possible reference assembly for human, mouse, and zebrafish via FTP. Tools to facilitate the curation of genome assemblies based on the sequence overlaps of long, high quality sequences.
Proper citation: Genome Reference Consortium (RRID:SCR_006553) Copy
http://www.geenivaramu.ee/en/tools/gwama
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software tool for meta analysis of whole genome association data.
Proper citation: GWAMA (RRID:SCR_006624) Copy
Model organism database for the social amoeba Dictyostelium discoideum that provides the biomedical research community with integrated, high quality data and tools for Dictyostelium discoideum and related species. dictyBase houses the complete genome sequence, ESTs, and the entire body of literature relevant to Dictyostelium. This information is curated to provide accurate gene models and functional annotations, with the goal of fully annotating the genome to provide a ''''reference genome'''' in the Amoebozoa clade. They highlight several new features in the present update: (i) new annotations; (ii) improved interface with web 2.0 functionality; (iii) the initial steps towards a genome portal for the Amoebozoa; (iv) ortholog display; and (v) the complete integration of the Dicty Stock Center with dictyBase. The Dicty Stock Center currently holds over 1500 strains targeting over 930 different genes. There are over 100 different distinct amoebozoan species. In addition, the collection contains nearly 600 plasmids and other materials such as antibodies and cDNA libraries. The strain collection includes: * strain catalog * natural isolates * MNNG chemical mutants * tester strains for parasexual genetics * auxotroph strains * null mutants * GFP-labeled strains for cell biology * plasmid catalog The Dicty Stock Center can accept Dictyostelium strains, plasmids, and other materials relevant for research using Dictyostelium such as antibodies and cDNA or genomic libraries.
Proper citation: Dictyostelium discoideum genome database (RRID:SCR_006643) Copy
http://rice.plantbiology.msu.edu/
Database and resource that provides sequence and annotation data for the rice genome. This website provides genome sequence from the Nipponbare subspecies of rice and annotation of the 12 rice chromosomes. All structural and functional annotation is viewable through our Rice Genome Browser which currently supports 75 tracks of annotation. Enhanced data access is available through web interfaces, FTP downloads and a Data Extractor tool developed in order to support discrete dataset downloads. Rice is a model species for the monocotyledonous plants and the cereals which are the greatest source of food for the world''s population. While rice genome sequence is available through multiple sequencing projects, high quality, uniform annotation is required in order for genome sequence data to be fully utilized by researchers. The existence of a common gene set and uniform annotation allows researchers within the rice community to work from a common resource so that their results can be more easily interpreted by other scientists. The objective of this project has always been to provide high quality annotation for the rice genome. They generated, refined and updated gene models for the estimated 40,000-60,000 total rice genes, provided standardized annotation for each model, linked each model to functional annotation including expression data, gene ontologies, and tagged lines. They have provided a resource to extend the annotation of the rice genome to other plant species by providing comparative alignments to other plant species. Analysis/Tools are available including: BLAST, Locus Name Search, Functional Term Search, Protein Domain Search, Anatomy Expression Viewer, Highly Expressed Genes
Proper citation: Rice Genome Annotation (RRID:SCR_006663) Copy
DPVweb provides a central source of information about viruses, viroids and satellites of plants, fungi and protozoa. Comprehensive taxonomic information, including brief descriptions of each family and genus, and classified lists of virus sequences are provided. The database also holds detailed, curated, information for all sequences of viruses, viroids and satellites of plants, fungi and protozoa that are complete or that contain at least one complete gene. For comparative purposes, it also contains a single representative sequence of all other fully sequenced virus species with an RNA or single-stranded DNA genome. The start and end positions of each feature (gene, non-translated region and the like) have been recorded and checked for accuracy. As far as possible, nomenclature for genes and proteins are standardized within genera and families. Sequences of features (either as DNA or amino acid sequences) can be directly downloaded from the website in FASTA format. The sequence information can also be accessed via client software for PC computers (freely downloadable from the website) that enable users to make an easy selection of sequences and features of a chosen virus for further analyses. The public sequence databases contain vast amounts of data on virus genomes but accessing and comparing the data, except for relatively small sets of related viruses can be very time consuming. The procedure is made difficult because some of the sequences on these databases are incorrectly named, poorly annotated or redundant. The NCBI Reference Sequence project (1) provides a comprehensive, integrated, non-redundant set of sequences, including genomic DNA, transcript (RNA) and protein products, for major research organisms. This now includes curated information for a single sequence of each fully sequenced virus species. While this is a welcome development, it can only deal with complete sequences. An important feature of DPV is the opportunity to access genes (and other features) of multiple sequences quickly and accurately. Thus, for example, it is easy to obtain the nucleotide or amino acid sequences of all the available accessions of the coat protein gene of a given virus species or for a group of viruses. To increase its usefulness further, DPVweb also contains a single representative sequence of all other fully sequenced virus species with an RNA or single-stranded DNA (ssDNA) genome. Sponsors: This site is supported by the Association of Applied Biologists and the Zhejiang Academy of Agricultural Sciences, Hangzhou, People''s Republic of China.
Proper citation: Descriptions of Plant Viruses (RRID:SCR_006656) Copy
http://inparanoid.sbc.su.se/cgi-bin/index.cgi
Collection of pairwise comparisons between 100 whole genomes generated by a fully automatic method for finding orthologs and in-paralogs between TWO species. Ortholog clusters in the InParanoid are seeded with a two-way best pairwise match, after which an algorithm for adding in-paralogs is applied. The method bypasses multiple alignments and phylogenetic trees, which can be slow and error-prone steps in classical ortholog detection. Still, it robustly detects complex orthologous relationships and assigns confidence values for in-paralogs. The original data sets can be downloaded.
Proper citation: InParanoid: Eukaryotic Ortholog Groups (RRID:SCR_006801) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.