Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Web application to search protein databases using a translated nucleotide query. Translated BLAST services are useful when trying to find homologous proteins to a nucleotide coding region. Blastx compares translational products of the nucleotide query sequence to a protein database. Because blastx translates the query sequence in all six reading frames and provides combined significance statistics for hits to different frames, it is particularly useful when the reading frame of the query sequence is unknown or it contains errors that may lead to frame shifts or other coding errors. Thus blastx is often the first analysis performed with a newly determined nucleotide sequence and is used extensively in analyzing EST sequences. This search is more sensitive than nucleotide blast since the comparison is performed at the protein level.
Proper citation: BLASTX (RRID:SCR_001653) Copy
International collaboration producing an extensive public catalog of human genetic variation, including SNPs and structural variants, and their haplotype contexts, in an effort to provide a foundation for investigating the relationship between genotype and phenotype. The genomes of about 2500 unidentified people from about 25 populations around the world were sequenced using next-generation sequencing technologies. Redundant sequencing on various platforms and by different groups of scientists of the same samples can be compared. The results of the study are freely and publicly accessible to researchers worldwide. The consortium identified the following populations whose DNA will be sequenced: Yoruba in Ibadan, Nigeria; Japanese in Tokyo; Chinese in Beijing; Utah residents with ancestry from northern and western Europe; Luhya in Webuye, Kenya; Maasai in Kinyawa, Kenya; Toscani in Italy; Gujarati Indians in Houston; Chinese in metropolitan Denver; people of Mexican ancestry in Los Angeles; and people of African ancestry in the southwestern United States. The goal Project is to find most genetic variants that have frequencies of at least 1% in the populations studied. Sequencing is still too expensive to deeply sequence the many samples being studied for this project. However, any particular region of the genome generally contains a limited number of haplotypes. Data can be combined across many samples to allow efficient detection of most of the variants in a region. The Project currently plans to sequence each sample to about 4X coverage; at this depth sequencing cannot provide the complete genotype of each sample, but should allow the detection of most variants with frequencies as low as 1%. Combining the data from 2500 samples should allow highly accurate estimation (imputation) of the variants and genotypes for each sample that were not seen directly by the light sequencing. All samples from the 1000 genomes are available as lymphoblastoid cell lines (LCLs) and LCL derived DNA from the Coriell Cell Repository as part of the NHGRI Catalog. The sequence and alignment data generated by the 1000genomes project is made available as quickly as possible via their mirrored ftp sites. ftp://ftp.1000genomes.ebi.ac.uk ftp://ftp-trace.ncbi.nlm.nih.gov/1000genomes
Proper citation: 1000 Genomes: A Deep Catalog of Human Genetic Variation (RRID:SCR_006828) Copy
http://www.ncbi.nlm.nih.gov/SNP/
General database of genetic variations maintained by the NCBI. Database as central repository for both single base nucleotide substitutions and short deletion and insertion polymorphisms. Distinguishes report of how to assay SNP from use of that SNP with individuals and populations. This separation simplifies some issues of data representation. However, these initial reports describing how to assay SNP will often be accompanied by SNP experiments measuring allele occurrence in individuals and populations. Community can contribute to this resource.
Proper citation: dbSNP (RRID:SCR_002338) Copy
http://hapmap.ncbi.nlm.nih.gov/
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 22, 2016. A multi-country collaboration among scientists and funding agencies to develop a public resource where genetic similarities and differences in human beings are identified and catalogued. Using this information, researchers will be able to find genes that affect health, disease, and individual responses to medications and environmental factors. All of the information generated by the Project will be released into the public domain. Their goal is to compare the genetic sequences of different individuals to identify chromosomal regions where genetic variants are shared. Public and private organizations in six countries are participating in the International HapMap Project. Data generated by the Project can be downloaded with minimal constraints. HapMap project related data, software, and documentation include: bulk data on genotypes, frequencies, LD data, phasing data, allocated SNPs, recombination rates and hotspots, SNP assays, Perlegen amplicons, raw data, inferred genotypes, and mitochondrial and chrY haplogroups; Generic Genome Browser software; protocols and information on assay design, genotyping and other protocols used in the project; and documentation of samples/individuals and the XML format used in the project.
Proper citation: International HapMap Project (RRID:SCR_002846) Copy
http://blast.ncbi.nlm.nih.gov/Blast.cgi
Web search tool to find regions of similarity between biological sequences. Program compares nucleotide or protein sequences to sequence databases and calculates statistical significance. Used for identifying homologous sequences.
Proper citation: NCBI BLAST (RRID:SCR_004870) Copy
http://biologylabs.utah.edu/jorgensen/wayned/ape/
Software tool for plasmid and sequence editing, annotating and drawing plasmid sequences. Used to view circular or linear maps of DNA sequences. Users can perform virtual digests whereby they select predefined DNA ladder, or specify their own, and visualize theoretical DNA fragments. Used to highlight restriction sites in editing window, accurately reflect Dam/Dcm blocking of enzyme sites, highlighting and drawing graphic maps using feature annotations from genbank and embl files, highlighting text using pre-defined and custom feature libraries, and directly BLASTing selected sequence at NCBI or Wormbase. Runs across Windows, OS X, and Linux/Unix.
Proper citation: A plasmid Editor (RRID:SCR_014266) Copy
http://www.ncbi.nlm.nih.gov/protein
Databases of protein sequences and 3D structures of proteins. Collection of sequences from several sources, including translations from annotated coding regions in GenBank, RefSeq and TPA, as well as records from SwissProt, PIR, PRF, and PDB.
Proper citation: NCBI Protein Database (RRID:SCR_003257) Copy
http://www.ncbi.nlm.nih.gov/igblast/
THIS RESOURCE IS NO LONGER IN SERVICE.Documented on January 4,2023. IgBLAST was developed at NCBI to facilitate analysis of immunoglobulin V region sequences in GenBank. In addition to performing a regular BLAST search, IgBLAST has several additional functions: - Reports the germline V, D and J gene matches to the query sequence. - Annotates the immunoglobulin domains (FWR1 through FWR3). - Matches the returned hits (for databases other than germline genes) to the closest germline V genes, making it easier to identify related sequences. - Reveals the V(D)J junction details such as nucleotide homology between the ends of V(D)J segments and N nucleotide insertions. D and J gene reporting is only for nucleotide sequence search and requires a stretch of five or more nucleotide identity between the query and D or J genes. Sponsors: This resource is supported by the National Center for Biotechnology Information, a division of the U.S. National Library of Medicine.
Proper citation: IgBLAST (RRID:SCR_002873) Copy
https://www.ncbi.nlm.nih.gov/genbank/tbl2asn2/
Software tool as a command-line program that automates the creation of sequence records for submission to GenBank. Records need no additional manual editing before submission.
Proper citation: tbl2asn (RRID:SCR_016636) Copy
http://www.ncbi.nlm.nih.gov/genbank/tpa/
Database designed to capture experimental or inferential results that support submitter-provided annotation for sequence data that the submitter did not directly determine but derived from GenBank primary data. Records are divided into two categories: * TPA:experimental: Annotation of sequence data is supported by peer-reviewed wet-lab experimental evidence. * TPA:inferential: Annotation of sequence data by inference (where the source molecule or its product(s) have not been the subject of direct experimentation) TPA records are retrieved through the Nucleotide Database and feature information on the sequence, how it was cataloged, and proper way to cite the sequence information.
Proper citation: TPA (RRID:SCR_003593) Copy
https://www.ncbi.nlm.nih.gov/Web/Search/entrezfs.html
Web portal for global query cross database search and retrieval system that provides access to all databases simultaneously with a single query string and user interface. Retrieves nucleotide and protein sequence data, gene centered and genomic mapping information, 3D structures, and references. Covers databases including protein sequence data from PIR-International, PRF, Swiss-Prot, and PDB and nucleotide sequence data from GenBank that includes information from EMBL and DDBJ.
Proper citation: Entrez (RRID:SCR_016640) Copy
https://www.ncbi.nlm.nih.gov/sites/batchentrez
Software program for loading numbers of genome records. Allows the retrieval of a large number of nucleotide sequences or protein sequences, in a batch mode, by importing a file containing a list of the desired GI or accession numbers.
Proper citation: Batch Entrez (RRID:SCR_016634) Copy
https://www.ncbi.nlm.nih.gov/projects/Sequin/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on November 12,2024. Software tool for DNA sequence submission. Used for submitting and updating entries to the GenBank or EMBL sequence databases.
Proper citation: Sequin (RRID:SCR_016581) Copy
https://ftp.ncbi.nlm.nih.gov/pub/mhc/mhc/Final%20Archive/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on August 23, 2019 Database was open, publicly accessible platform for DNA and clinical data related to human Major Histocompatibility Complex (MHC). Data from IHWG workshops were provided as well., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: dbMHC (RRID:SCR_002302) Copy
http://www.ncbi.nlm.nih.gov/genome
Database that organizes information on genomes including sequences, maps, chromosomes, assemblies, and annotations in six major organism groups: Archaea, Bacteria, Eukaryotes, Viruses, Viroids, and Plasmids. Genomes of over 1,200 organisms can be found in this database, representing both completely sequenced organisms and those for which sequencing is in progress. Users can browse by organism, and view genome maps and protein clusters. Links to other prokaryotic and archaeal genome projects, as well as BLAST tools and access to the rest of the NCBI online resources are available.
Proper citation: NCBI Genome (RRID:SCR_002474) Copy
Web application to search nucleotide databases using a nucleotide query. Algorithms: blastn, megablast, discontiguous megablast.
Proper citation: BLASTN (RRID:SCR_001598) Copy
http://www.ncbi.nlm.nih.gov/HTGS/
Database of high-throughput genome sequences from large-scale genome sequencing centers, including unfinished and finished sequences. It was created to accommodate a growing need to make unfinished genomic sequence data rapidly available to the scientific community in a coordinated effort among the International Nucleotide Sequence databases, DDBJ, EMBL, and GenBank. Sequences are prepared for submission by using NCBI's software tools Sequin or tbl2asn. Each center has an FTP directory into which new or updated sequence files are placed. Sequence data in this division are available for BLAST homology searches against either the htgs database or the month database, which includes all new submissions for the prior month. Unfinished HTG sequences containing contigs greater than 2 kb are assigned an accession number and deposited in the HTG division. A typical HTG record might consist of all the first-pass sequence data generated from a single cosmid, BAC, YAC, or P1 clone, which together make up more than 2 kb and contain one or more gaps. A single accession number is assigned to this collection of sequences, and each record includes a clear indication of the status (phase 1 or 2) plus a prominent warning that the sequence data are unfinished and may contain errors. The accession number does not change as sequence records are updated; only the most recent version of a HTG record remains in GenBank.
Proper citation: High Throughput Genomic Sequences Division (RRID:SCR_002150) Copy
http://www.ncbi.nlm.nih.gov/projects/genome/assembly/grc/
Consortium that puts sequences into a chromosome context and provides the best possible reference assembly for human, mouse, and zebrafish via FTP. Tools to facilitate the curation of genome assemblies based on the sequence overlaps of long, high quality sequences.
Proper citation: Genome Reference Consortium (RRID:SCR_006553) Copy
https://www.ncbi.nlm.nih.gov/genbank/dbest/
Database as a division of GenBank that contains sequence data and other information on single-pass cDNA sequences, or Expressed Sequence Tags, from a number of organisms.
Proper citation: dbEST (RRID:SCR_008132) Copy
https://www.ncbi.nlm.nih.gov/Web/Newsltr/Spring04/blastlab.html
Software tool as a program within the standalone BLAST package used to cluster either protein or nucleotide sequences. Used to make non redundant sequence sets.
Proper citation: BLASTClust (RRID:SCR_016641) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.