Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Public archive providing a comprehensive record of the world''''s nucleotide sequencing information, covering raw sequencing data, sequence assembly information and functional annotation. All submitted data, once public, will be exchanged with the NCBI and DDBJ as part of the INSDC data exchange agreement. The European Nucleotide Archive (ENA) captures and presents information relating to experimental workflows that are based around nucleotide sequencing. A typical workflow includes the isolation and preparation of material for sequencing, a run of a sequencing machine in which sequencing data are produced and a subsequent bioinformatic analysis pipeline. ENA records this information in a data model that covers input information (sample, experimental setup, machine configuration), output machine data (sequence traces, reads and quality scores) and interpreted information (assembly, mapping, functional annotation). Data arrive at ENA from a variety of sources including submissions of raw data, assembled sequences and annotation from small-scale sequencing efforts, data provision from the major European sequencing centers and routine and comprehensive exchange with their partners in the International Nucleotide Sequence Database Collaboration (INSDC). Provision of nucleotide sequence data to ENA or its INSDC partners has become a central and mandatory step in the dissemination of research findings to the scientific community. ENA works with publishers of scientific literature and funding bodies to ensure compliance with these principles and to provide optimal submission systems and data access tools that work seamlessly with the published literature. ENA is made up of a number of distinct databases that includes the EMBL Nucleotide Sequence Database (Embl-Bank), the newly established Sequence Read Archive (SRA) and the Trace Archive. The main tool for downloading ENA data is the ENA Browser, which is available through REST URLs for easy programmatic use. All ENA data are available through the ENA Browser. Note: EMBL Nucleotide Sequence Database (EMBL-Bank) is entirely included within this resource.
Proper citation: European Nucleotide Archive (ENA) (RRID:SCR_006515) Copy
http://amigo.geneontology.org/
Web tool to search, sort, analyze, visualize and download data of interest. Along with providing details of the ontologies, gene products and annotations, features a BLAST search, Term Enrichment and GO Slimmer tools, the GO Online SQL Environment and a user help guide.Used at the Gene Ontology (GO) website to access the data provided by the GO Consortium. Developed and maintained by the GO Consortium.
Proper citation: AmiGO (RRID:SCR_002143) Copy
https://github.com/blackrim/phyutility
Command line program that performs analyses or modifications on both trees and data matrices. Software phyloinformatics tool for trees, alignments and molecular data. Used for summarizing and manipulating phylogenetic trees, manipulating molecular data and retrieving data from NCBI.
Proper citation: Phyutility (RRID:SCR_018545) Copy
Web application to search protein databases using a translated nucleotide query. Translated BLAST services are useful when trying to find homologous proteins to a nucleotide coding region. Blastx compares translational products of the nucleotide query sequence to a protein database. Because blastx translates the query sequence in all six reading frames and provides combined significance statistics for hits to different frames, it is particularly useful when the reading frame of the query sequence is unknown or it contains errors that may lead to frame shifts or other coding errors. Thus blastx is often the first analysis performed with a newly determined nucleotide sequence and is used extensively in analyzing EST sequences. This search is more sensitive than nucleotide blast since the comparison is performed at the protein level.
Proper citation: BLASTX (RRID:SCR_001653) Copy
http://www.ncbi.nlm.nih.gov/pmc/
Collection of full text archive of biomedical and life sciences journal literature at U.S. National Institutes of Health National Library of Medicine (NIH/NLM). With PubMed Central, NCBI is taking lead in preserving and maintaining open access to electronic literature. Value of PubMed Central, in addition to its role as an archive, lies in what can be done when data from diverse sources is stored in common format in single repository. All articles in PMC are free (sometimes on a delayed basis). Some journals go beyond free, to Open Access.
Proper citation: PubMed Central (RRID:SCR_004166) Copy
http://pubchem.ncbi.nlm.nih.gov/
Collection of information about chemical structures and biological properties of small molecules and siRNA reagents hosted by the National Center for Biotechnology Information (NCBI).
Proper citation: PubChem (RRID:SCR_004284) Copy
http://blast.ncbi.nlm.nih.gov/Blast.cgi
Web search tool to find regions of similarity between biological sequences. Program compares nucleotide or protein sequences to sequence databases and calculates statistical significance. Used for identifying homologous sequences.
Proper citation: NCBI BLAST (RRID:SCR_004870) Copy
http://www.ncbi.nlm.nih.gov/SNP/
General database of genetic variations maintained by the NCBI. Database as central repository for both single base nucleotide substitutions and short deletion and insertion polymorphisms. Distinguishes report of how to assay SNP from use of that SNP with individuals and populations. This separation simplifies some issues of data representation. However, these initial reports describing how to assay SNP will often be accompanied by SNP experiments measuring allele occurrence in individuals and populations. Community can contribute to this resource.
Proper citation: dbSNP (RRID:SCR_002338) Copy
http://hapmap.ncbi.nlm.nih.gov/
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 22, 2016. A multi-country collaboration among scientists and funding agencies to develop a public resource where genetic similarities and differences in human beings are identified and catalogued. Using this information, researchers will be able to find genes that affect health, disease, and individual responses to medications and environmental factors. All of the information generated by the Project will be released into the public domain. Their goal is to compare the genetic sequences of different individuals to identify chromosomal regions where genetic variants are shared. Public and private organizations in six countries are participating in the International HapMap Project. Data generated by the Project can be downloaded with minimal constraints. HapMap project related data, software, and documentation include: bulk data on genotypes, frequencies, LD data, phasing data, allocated SNPs, recombination rates and hotspots, SNP assays, Perlegen amplicons, raw data, inferred genotypes, and mitochondrial and chrY haplogroups; Generic Genome Browser software; protocols and information on assay design, genotyping and other protocols used in the project; and documentation of samples/individuals and the XML format used in the project.
Proper citation: International HapMap Project (RRID:SCR_002846) Copy
International collaboration producing an extensive public catalog of human genetic variation, including SNPs and structural variants, and their haplotype contexts, in an effort to provide a foundation for investigating the relationship between genotype and phenotype. The genomes of about 2500 unidentified people from about 25 populations around the world were sequenced using next-generation sequencing technologies. Redundant sequencing on various platforms and by different groups of scientists of the same samples can be compared. The results of the study are freely and publicly accessible to researchers worldwide. The consortium identified the following populations whose DNA will be sequenced: Yoruba in Ibadan, Nigeria; Japanese in Tokyo; Chinese in Beijing; Utah residents with ancestry from northern and western Europe; Luhya in Webuye, Kenya; Maasai in Kinyawa, Kenya; Toscani in Italy; Gujarati Indians in Houston; Chinese in metropolitan Denver; people of Mexican ancestry in Los Angeles; and people of African ancestry in the southwestern United States. The goal Project is to find most genetic variants that have frequencies of at least 1% in the populations studied. Sequencing is still too expensive to deeply sequence the many samples being studied for this project. However, any particular region of the genome generally contains a limited number of haplotypes. Data can be combined across many samples to allow efficient detection of most of the variants in a region. The Project currently plans to sequence each sample to about 4X coverage; at this depth sequencing cannot provide the complete genotype of each sample, but should allow the detection of most variants with frequencies as low as 1%. Combining the data from 2500 samples should allow highly accurate estimation (imputation) of the variants and genotypes for each sample that were not seen directly by the light sequencing. All samples from the 1000 genomes are available as lymphoblastoid cell lines (LCLs) and LCL derived DNA from the Coriell Cell Repository as part of the NHGRI Catalog. The sequence and alignment data generated by the 1000genomes project is made available as quickly as possible via their mirrored ftp sites. ftp://ftp.1000genomes.ebi.ac.uk ftp://ftp-trace.ncbi.nlm.nih.gov/1000genomes
Proper citation: 1000 Genomes: A Deep Catalog of Human Genetic Variation (RRID:SCR_006828) Copy
http://www.ncbi.nlm.nih.gov/blast/html/megablast.html
Software that uses the greedy algorithm for nucleotide sequence alignment search.
Proper citation: Mega BLAST (RRID:SCR_011920) Copy
http://blast.ncbi.nlm.nih.gov/Blast.cgi?PROGRAM=tblastn&PAGE_TYPE=BlastSearch&LINK_LOC=blasthome
Tool to search translated nucleotide databases using a protein query.
Proper citation: TBLASTN (RRID:SCR_011822) Copy
http://www.ncbi.nlm.nih.gov/igblast/
THIS RESOURCE IS NO LONGER IN SERVICE.Documented on January 4,2023. IgBLAST was developed at NCBI to facilitate analysis of immunoglobulin V region sequences in GenBank. In addition to performing a regular BLAST search, IgBLAST has several additional functions: - Reports the germline V, D and J gene matches to the query sequence. - Annotates the immunoglobulin domains (FWR1 through FWR3). - Matches the returned hits (for databases other than germline genes) to the closest germline V genes, making it easier to identify related sequences. - Reveals the V(D)J junction details such as nucleotide homology between the ends of V(D)J segments and N nucleotide insertions. D and J gene reporting is only for nucleotide sequence search and requires a stretch of five or more nucleotide identity between the query and D or J genes. Sponsors: This resource is supported by the National Center for Biotechnology Information, a division of the U.S. National Library of Medicine.
Proper citation: IgBLAST (RRID:SCR_002873) Copy
http://www.ncbi.nlm.nih.gov/books/NBK1116/
Provides clinically relevant and medically actionable information for inherited conditions in standardized journal-style format, covering diagnosis, management, and genetic counseling for patients and their families. Searchable book of expert-authored, peer-reviewed disease descriptions presented in standardized format and focused on clinically relevant and medically actionable information on diagnosis, management, and genetic counseling of patients and families with specific inherited conditions.
Proper citation: GeneReviews (RRID:SCR_006560) Copy
http://www.ncbi.nlm.nih.gov/unigene
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 11, 2023. Web tool for an organized view of the transcriptome. Collection of the computationally identified transcripts from the same locus. Information on protein similarities, gene expression, cDNA clones, and genomic location. System for automatically partitioning GenBank sequences into a non redundant set of gene oriented clusters.
Proper citation: UniGene (RRID:SCR_004405) Copy
http://www.ncbi.nlm.nih.gov/CBBresearch/Wilbur/IRET/PIE/
A web service to extract Protein-protein interaction (PPI)-relevant articles from MEDLINE that provides protein interaction information (PPI) articles for biologists, baseline system performance for bio-text mining researchers and a compact PubMed-search environment for PubMed users. It accepts PubMed input formats including All Fields, Author, Journal, MeSH Terms, Publication Date, Title, and Title/Abstract with Boolean operations (AND, OR, and NOT). However, the output is the list of articles prioritized by PPI confidence rates. Some words (mostly gene/protein names) which contributed for PPI prediction are underlined and linked to Entrez or Entrez Gene. Even though our system focuses on a PubMed search environment, it also provides a CGI access for bio-text mining researchers. Using the CGI program, a list of PubMed IDs can be obtained as a query result, thus it can be utilized as a baseline system performance. PIE the search is based on a winning approach in the BioCreative III ACT competition (BC3)1. For input queries, MEDLINE articles are first retrieved through the PubMed service. PPI scores are calculated for the retrieved articles, and the articles are re-ranked based on scores. To effectively capture PPI patterns from biomedical literature, their approach utilizes both word and syntactic features for machine learning classifiers. Dependency parsing, gene mention tagging, and term-based features are utilized along with a Huber classifier.
Proper citation: PIE the search (RRID:SCR_005296) Copy
http://www.ncbi.nlm.nih.gov/clinvar/
Archive of aggregated information about sequence variation and its relationship to human health. Provides reports of relationships among human variations and phenotypes along with supporting evidence. Submissions from clinical testing labs, research labs, locus-specific databases, expert panels and professional societies are welcome. Collects reports of variants found in patient samples, assertions made regarding their clinical significance, information about submitter, and other supporting data. Alleles described in submissions are mapped to reference sequences, and reported according to HGVS standard.
Proper citation: ClinVar (RRID:SCR_006169) Copy
http://www.ncbi.nlm.nih.gov/gene
Database for genomes that have been completely sequenced, have active research community to contribute gene-specific information, or that are scheduled for intense sequence analysis. Includes nomenclature, map location, gene products and their attributes, markers, phenotypes, and links to citations, sequences, variation details, maps, expression, homologs, protein domains and external databases. All entries follow NCBI's format for data collections. Content of Entrez Gene represents result of curation and automated integration of data from NCBI's Reference Sequence project (RefSeq), from collaborating model organism databases, and from many other databases available from NCBI. Records are assigned unique, stable and tracked integers as identifiers. Content is updated as new information becomes available.
Proper citation: Entrez Gene (RRID:SCR_002473) Copy
http://www.ncbi.nlm.nih.gov/homologene
Automated system for constructing putative homology groups from complete gene sets of wide range of eukaryotic species. Databse that provides system for automatic detection of homologs, including paralogs and orthologs, among annotated genes of sequenced eukaryotic genomes. HomoloGene processing uses proteins from input organisms to compare and sequence homologs, mapping back to corresponding DNA sequences. Reports include homology and phenotype information drawn from Online Mendelian Inheritance in Man, Mouse Genome Informatics, Zebrafish Information Network, Saccharomyces Genome Database and FlyBase.
Proper citation: HomoloGene (RRID:SCR_002924) Copy
http://www.ncbi.nlm.nih.gov/mapview/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 4, 2023. Database that provides special browsing capabilities for a subset of organisms in Entrez Genomes. Map Viewer allows users to view and search an organism's complete genome, display chromosome maps, and zoom into progressively greater levels of detail, down to the sequence data for a region of interest. If multiple maps are available for a chromosome, it displays them aligned to each other based on shared marker and gene names, and, for the sequence maps, based on a common sequence coordinate system.
Proper citation: MapViewer (RRID:SCR_003092) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.