Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Public archive providing a comprehensive record of the world''''s nucleotide sequencing information, covering raw sequencing data, sequence assembly information and functional annotation. All submitted data, once public, will be exchanged with the NCBI and DDBJ as part of the INSDC data exchange agreement. The European Nucleotide Archive (ENA) captures and presents information relating to experimental workflows that are based around nucleotide sequencing. A typical workflow includes the isolation and preparation of material for sequencing, a run of a sequencing machine in which sequencing data are produced and a subsequent bioinformatic analysis pipeline. ENA records this information in a data model that covers input information (sample, experimental setup, machine configuration), output machine data (sequence traces, reads and quality scores) and interpreted information (assembly, mapping, functional annotation). Data arrive at ENA from a variety of sources including submissions of raw data, assembled sequences and annotation from small-scale sequencing efforts, data provision from the major European sequencing centers and routine and comprehensive exchange with their partners in the International Nucleotide Sequence Database Collaboration (INSDC). Provision of nucleotide sequence data to ENA or its INSDC partners has become a central and mandatory step in the dissemination of research findings to the scientific community. ENA works with publishers of scientific literature and funding bodies to ensure compliance with these principles and to provide optimal submission systems and data access tools that work seamlessly with the published literature. ENA is made up of a number of distinct databases that includes the EMBL Nucleotide Sequence Database (Embl-Bank), the newly established Sequence Read Archive (SRA) and the Trace Archive. The main tool for downloading ENA data is the ENA Browser, which is available through REST URLs for easy programmatic use. All ENA data are available through the ENA Browser. Note: EMBL Nucleotide Sequence Database (EMBL-Bank) is entirely included within this resource.
Proper citation: European Nucleotide Archive (ENA) (RRID:SCR_006515) Copy
http://www.ncbi.nlm.nih.gov/projects/gv/rbc/main.fcgi?cmd=init
The dbRBC database provides an open, publicly accessible platform for DNA and clinical data related to the human Red Blood Cells (RBC). A new bioinformatics resource, dbRBC, has been installed at the National Center of Biotechnology Information (NCBI). This resource combines the well established Blood Group Antigen Gene Mutation Database (BGMUT) with tools and interlinked resources developed at the NCBI. The main task of dbRBC is to provide access to publicly available genomic, protein and structural information linked to the red blood cell antigens. The site offers a number of resources: * BGMUT Database * Alignment Viewer * SBT Tool * Probe/Primer Resource * Typing Kit Interface * Obstacle
Proper citation: NCBI dbRBC (RRID:SCR_005959) Copy
Web application to search protein databases using a translated nucleotide query. Translated BLAST services are useful when trying to find homologous proteins to a nucleotide coding region. Blastx compares translational products of the nucleotide query sequence to a protein database. Because blastx translates the query sequence in all six reading frames and provides combined significance statistics for hits to different frames, it is particularly useful when the reading frame of the query sequence is unknown or it contains errors that may lead to frame shifts or other coding errors. Thus blastx is often the first analysis performed with a newly determined nucleotide sequence and is used extensively in analyzing EST sequences. This search is more sensitive than nucleotide blast since the comparison is performed at the protein level.
Proper citation: BLASTX (RRID:SCR_001653) Copy
http://www.ncbi.nlm.nih.gov/projects/homology/maps/
This page provides quick access to the Comparative mapping functions available in the Map Viewer. Currently, comparative maps are calculated using HomoloGene orthology predictions. Once the gene pairs have been established, blocks of conserved syteny can be established using the positions of each gene object in their respective builds. Sponsors: This resource is supported by NCBI.
Proper citation: Homology Maps Page (RRID:SCR_001666) Copy
http://blast.ncbi.nlm.nih.gov/Blast.cgi
Web search tool to find regions of similarity between biological sequences. Program compares nucleotide or protein sequences to sequence databases and calculates statistical significance. Used for identifying homologous sequences.
Proper citation: NCBI BLAST (RRID:SCR_004870) Copy
http://blast.ncbi.nlm.nih.gov/Blast.cgi?PROGRAM=tblastn&PAGE_TYPE=BlastSearch&LINK_LOC=blasthome
Tool to search translated nucleotide databases using a protein query.
Proper citation: TBLASTN (RRID:SCR_011822) Copy
http://www.ncbi.nlm.nih.gov/gene
Database for genomes that have been completely sequenced, have active research community to contribute gene-specific information, or that are scheduled for intense sequence analysis. Includes nomenclature, map location, gene products and their attributes, markers, phenotypes, and links to citations, sequences, variation details, maps, expression, homologs, protein domains and external databases. All entries follow NCBI's format for data collections. Content of Entrez Gene represents result of curation and automated integration of data from NCBI's Reference Sequence project (RefSeq), from collaborating model organism databases, and from many other databases available from NCBI. Records are assigned unique, stable and tracked integers as identifiers. Content is updated as new information becomes available.
Proper citation: Entrez Gene (RRID:SCR_002473) Copy
http://www.ncbi.nlm.nih.gov/homologene
Automated system for constructing putative homology groups from complete gene sets of wide range of eukaryotic species. Databse that provides system for automatic detection of homologs, including paralogs and orthologs, among annotated genes of sequenced eukaryotic genomes. HomoloGene processing uses proteins from input organisms to compare and sequence homologs, mapping back to corresponding DNA sequences. Reports include homology and phenotype information drawn from Online Mendelian Inheritance in Man, Mouse Genome Informatics, Zebrafish Information Network, Saccharomyces Genome Database and FlyBase.
Proper citation: HomoloGene (RRID:SCR_002924) Copy
http://www.ncbi.nlm.nih.gov/protein
Databases of protein sequences and 3D structures of proteins. Collection of sequences from several sources, including translations from annotated coding regions in GenBank, RefSeq and TPA, as well as records from SwissProt, PIR, PRF, and PDB.
Proper citation: NCBI Protein Database (RRID:SCR_003257) Copy
http://www.ncbi.nlm.nih.gov/CBBresearch/Wilbur/IRET/PIE/
A web service to extract Protein-protein interaction (PPI)-relevant articles from MEDLINE that provides protein interaction information (PPI) articles for biologists, baseline system performance for bio-text mining researchers and a compact PubMed-search environment for PubMed users. It accepts PubMed input formats including All Fields, Author, Journal, MeSH Terms, Publication Date, Title, and Title/Abstract with Boolean operations (AND, OR, and NOT). However, the output is the list of articles prioritized by PPI confidence rates. Some words (mostly gene/protein names) which contributed for PPI prediction are underlined and linked to Entrez or Entrez Gene. Even though our system focuses on a PubMed search environment, it also provides a CGI access for bio-text mining researchers. Using the CGI program, a list of PubMed IDs can be obtained as a query result, thus it can be utilized as a baseline system performance. PIE the search is based on a winning approach in the BioCreative III ACT competition (BC3)1. For input queries, MEDLINE articles are first retrieved through the PubMed service. PPI scores are calculated for the retrieved articles, and the articles are re-ranked based on scores. To effectively capture PPI patterns from biomedical literature, their approach utilizes both word and syntactic features for machine learning classifiers. Dependency parsing, gene mention tagging, and term-based features are utilized along with a Huber classifier.
Proper citation: PIE the search (RRID:SCR_005296) Copy
http://www.ncbi.nlm.nih.gov/structure
Database of three-dimensional structures of macromolecules that allows the user to retrieve structures for specific molecule types as well as structures for genes and proteins of interest. Three main databases comprise Structure-The Molecular Modeling Database; Conserved Domains and Protein Classification; and the BioSystems Database. Structure also links to the PubChem databases to connect biological activity data to the macromolecular structures. Users can locate structural templates for proteins and interactively view structures and sequence data to closely examine sequence-structure relationships. * Macromolecular structures: The three-dimensional structures of biomolecules provide a wealth of information on their biological function and evolutionary relationships. The Molecular Modeling Database (MMDB), as part of the Entrez system, facilitates access to structure data by connecting them with associated literature, protein and nucleic acid sequences, chemicals, biomolecular interactions, and more. It is possible, for example, to find 3D structures for homologs of a protein of interest by following the Related Structure link in an Entrez Protein sequence record. * Conserved domains and protein classification: Conserved domains are functional units within a protein that act as building blocks in molecular evolution and recombine in various arrangements to make proteins with different functions. The Conserved Domain Database (CDD) brings together several collections of multiple sequence alignments representing conserved domains, in addition to NCBI-curated domains that use 3D-structure information explicitly to define domain boundaries and provide insights into sequence/structure/function relationships. * Small molecules and their biological activity: The PubChem project provides information on the biological activities of small molecules and is a component of NIH''''s Molecular Libraries Roadmap Initiative. PubChem includes three databases: PCSubstance, PCBioAssay, and PCCompound. The PubChem data are linked to other data types (illustrated example) in the Entrez system, making it possible, for example, to retrieve information about a compound and then Link to its biological activity data, retrieve 3D protein structures bound to the compound and interactively view their active sites, and find biosystems that include the compound as a component. * Biological Systems: A biosystem, or biological system, is a group of molecules that interact directly or indirectly, where the grouping is relevant to the characterization of living matter. The NCBI BioSystems Database provides centralized access to biological pathways from several source databases and connects the biosystem records with associated literature, molecular, and chemical data throughout the Entrez system. BioSystem records list and categorize components (illustrated example), such as the genes, proteins, and small molecules involved in a biological system. The companion FLink icon FLink tool, in turn, allows you to input a list of proteins, genes, or small molecules and retrieve a ranked list of biosystems.
Proper citation: NCBI Structure (RRID:SCR_004218) Copy
http://www.ncbi.nlm.nih.gov/unigene
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 11, 2023. Web tool for an organized view of the transcriptome. Collection of the computationally identified transcripts from the same locus. Information on protein similarities, gene expression, cDNA clones, and genomic location. System for automatically partitioning GenBank sequences into a non redundant set of gene oriented clusters.
Proper citation: UniGene (RRID:SCR_004405) Copy
https://www.ncbi.nlm.nih.gov/Web/Search/entrezfs.html
Web portal for global query cross database search and retrieval system that provides access to all databases simultaneously with a single query string and user interface. Retrieves nucleotide and protein sequence data, gene centered and genomic mapping information, 3D structures, and references. Covers databases including protein sequence data from PIR-International, PRF, Swiss-Prot, and PDB and nucleotide sequence data from GenBank that includes information from EMBL and DDBJ.
Proper citation: Entrez (RRID:SCR_016640) Copy
https://www.ncbi.nlm.nih.gov/sites/batchentrez
Software program for loading numbers of genome records. Allows the retrieval of a large number of nucleotide sequences or protein sequences, in a batch mode, by importing a file containing a list of the desired GI or accession numbers.
Proper citation: Batch Entrez (RRID:SCR_016634) Copy
http://www.ncbi.nlm.nih.gov/gtr/
Central location for voluntary submission of genetic test information by providers including the test''s purpose, methodology, validity, evidence of the test''s usefulness, and laboratory contacts and credentials. GTR aims to advance the public health and research into the genetic basis of health and disease. GTR is accepting registration of clinical tests for Mendelian disorders, complex tests and arrays, and pharmacogenetic tests. These tests may include multiple methods and may include multiple major method categories such as biochemical, cytogenetic, and molecular tests. GTR is not currently accepting registration of tests for somatic disorders, research tests or direct-to-consumer tests.
Proper citation: Genetic Testing Registry (RRID:SCR_005565) Copy
http://www.ncbi.nlm.nih.gov/biosystems/
Database that provides access to biological systems and their component genes, proteins, and small molecules, as well as literature describing those biosystems and other related data throughout Entrez. A biosystem, or biological system, is a group of molecules that interact directly or indirectly, where the grouping is relevant to the characterization of living matter. BioSystem records list and categorize components, such as the genes, proteins, and small molecules involved in a biological system. The companion FLink tool, in turn, allows you to input a list of proteins, genes, or small molecules and retrieve a ranked list of biosystems. A number of databases provide diagrams showing the components and products of biological pathways along with corresponding annotations and links to literature. This database was developed as a complementary project to (1) serve as a centralized repository of data; (2) connect the biosystem records with associated literature, molecular, and chemical data throughout the Entrez system; and (3) facilitate computation on biosystems data. The NCBI BioSystems Database currently contains records from several source databases: KEGG, BioCyc (including its Tier 1 EcoCyc and MetaCyc databases, and its Tier 2 databases), Reactome, the National Cancer Institute's Pathway Interaction Database, WikiPathways, and Gene Ontology (GO). It includes several types of records such as pathways, structural complexes, and functional sets, and is desiged to accomodate other record types, such as diseases, as data become available. Through these collaborations, the BioSystems database facilitates access to, and provides the ability to compute on, a wide range of biosystems data. If you are interested in depositing data into the BioSystems database, please contact them.
Proper citation: NCBI BioSystems Database (RRID:SCR_004690) Copy
https://www.ncbi.nlm.nih.gov/Web/Newsltr/Spring04/blastlab.html
Software tool as a program within the standalone BLAST package used to cluster either protein or nucleotide sequences. Used to make non redundant sequence sets.
Proper citation: BLASTClust (RRID:SCR_016641) Copy
https://www.ncbi.nlm.nih.gov/orffinder
Software tool to search for open reading frames (ORFs) in the DNA sequence. The program returns the range of each ORF, along with its protein translation. Used to search newly sequenced DNA for potential protein encoding segments, verify predicted protein. Limited to the subrange of the query sequence up to 50 kb long.
Proper citation: Open Reading Frame Finder (RRID:SCR_016643) Copy
http://www.ncbi.nlm.nih.gov/CCDS/
Database (anonymous FTP) resulting from a collaborative effort to identify a core set of human and mouse protein coding regions that are consistently annotated and of high quality. The long term goal is to support convergence towards a standard set of gene annotations. Collaborators are EBI, NCBI, UCSC, WTSI and the initial results are also available from the participants'''' genome browser Web sites. In addition, CCDS identifiers are indicated on the relevant NCBI RefSeq and Entrez Gene records and in Map Viewer displays of RNA (RefSeq) and Gene annotations on the reference assembly.
Proper citation: Consensus CDS (RRID:SCR_006729) Copy
http://www.ncbi.nlm.nih.gov/RefSeq/HIVInteractions/
A database of interactions between HIV-1 and human proteins published in the peer-reviewed literature. The goal is to provide a concise, yet detailed, summary of all known interactions of HIV-1 proteins with host cell proteins, other HIV-1 proteins, or proteins from disease organisms associated with HIV/AIDS. For each HIV-1 human protein interaction the following information is provided: * NCBI Reference Sequence (RefSeq) protein accession numbers. * NCBI Entrez Gene ID numbers. * Amino acids from each protein that are known to be involved in the interaction. * Brief description of the protein-protein interaction. * Keywords to support searching for interactions. * PubMed identification numbers (PMIDs) for all journal articles describing the interaction. In addition, all protein-protein interactions documented in the database are integrated into Entrez Gene records and listed in the ''HIV-1 protein interactions'' section of Entrez Gene reports. The database is also tightly linked to other databases through Entrez Gene, enabling users to search for an abundance of information related to HIV pathogenesis and replication.
Proper citation: HIV-1 Human Protein Interaction Database (RRID:SCR_006879) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.