Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://www.ncbi.nlm.nih.gov/Genbank/
NIH genetic sequence database that provides annotated collection of all publicly available DNA sequences for almost 280 000 formally described species (Jan 2014) .These sequences are obtained primarily through submissions from individual laboratories and batch submissions from large-scale sequencing projects, including whole-genome shotgun (WGS) and environmental sampling projects. Most submissions are made using web-based BankIt or standalone Sequin programs, and GenBank staff assigns accession numbers upon data receipt. It is part of International Nucleotide Sequence Database Collaboration and daily data exchange with European Nucleotide Archive (ENA) and DNA Data Bank of Japan (DDBJ) ensures worldwide coverage. GenBank is accessible through NCBI Entrez retrieval system, which integrates data from major DNA and protein sequence databases along with taxonomy, genome, mapping, protein structure and domain information, and biomedical journal literature via PubMed. BLAST provides sequence similarity searches of GenBank and other sequence databases. Complete bimonthly releases and daily updates of GenBank database are available by FTP.
Proper citation: GenBank (RRID:SCR_002760) Copy
https://www.ncbi.nlm.nih.gov/genbank/dbgss/
Database of unannotated short single-read primarily genomic sequences from GenBank including random survey sequences clone-end sequences and exon-trapped sequences. The GSS division of GenBank is similar to the EST division, with the exception that most of the sequences are genomic in origin, rather than cDNA (mRNA). It should be noted that two classes (exon trapped products and gene trapped products) may be derived via a cDNA intermediate. Care should be taken when analyzing sequences from either of these classes, as a splicing event could have occurred and the sequence represented in the record may be interrupted when compared to genomic sequence. The GSS division contains (but is not limited to) the following types of data: * random single pass read genome survey sequences. * cosmid/BAC/YAC end sequences * exon trapped genomic sequences * Alu PCR sequences * transposon-tagged sequences Although dbGSS sequences are incorporated into the GSS Division of GenBank, annotation in dbGSS is more comprehensive and includes detailed information about the contributors, experimental conditions, and genetic map locations.
Proper citation: NCBI Genome Survey Sequences Database (RRID:SCR_002146) Copy
A portal to biomedical and genomic information. NCBI creates public databases, conducts research in computational biology, develops software tools for analyzing genome data, and disseminates biomedical information for the better understanding of molecular processes affecting human health and disease.
Proper citation: NCBI (RRID:SCR_006472) Copy
http://www.ncbi.nlm.nih.gov/HTGS/
Database of high-throughput genome sequences from large-scale genome sequencing centers, including unfinished and finished sequences. It was created to accommodate a growing need to make unfinished genomic sequence data rapidly available to the scientific community in a coordinated effort among the International Nucleotide Sequence databases, DDBJ, EMBL, and GenBank. Sequences are prepared for submission by using NCBI's software tools Sequin or tbl2asn. Each center has an FTP directory into which new or updated sequence files are placed. Sequence data in this division are available for BLAST homology searches against either the htgs database or the month database, which includes all new submissions for the prior month. Unfinished HTG sequences containing contigs greater than 2 kb are assigned an accession number and deposited in the HTG division. A typical HTG record might consist of all the first-pass sequence data generated from a single cosmid, BAC, YAC, or P1 clone, which together make up more than 2 kb and contain one or more gaps. A single accession number is assigned to this collection of sequences, and each record includes a clear indication of the status (phase 1 or 2) plus a prominent warning that the sequence data are unfinished and may contain errors. The accession number does not change as sequence records are updated; only the most recent version of a HTG record remains in GenBank.
Proper citation: High Throughput Genomic Sequences Division (RRID:SCR_002150) Copy
Maintains and provides archival, retrieval and analytical resources for biological information. Central DDBJ resource consists of public, open-access nucleotide sequence databases including raw sequence reads, assembly information and functional annotation. Database content is exchanged with EBI and NCBI within the framework of the International Nucleotide Sequence Database Collaboration (INSDC). In 2011, DDBJ launched two new resources: DDBJ Omics Archive and BioProject. DOR is archival database of functional genomics data generated by microarray and highly parallel new generation sequencers. Data are exchanged between the ArrayExpress at EBI and DOR in the common MAGE-TAB format. BioProject provides organizational framework to access metadata about research projects and data from projects that are deposited into different databases.
Proper citation: DNA DataBank of Japan (DDBJ) (RRID:SCR_002359) Copy
https://www.ncbi.nlm.nih.gov/assembly
Database providing information on structure of assembled genomes, assembly names and other meta-data, statistical reports, and links to genomic sequence data. The Archive links the raw sequence information found in the Trace Archive with assembly information found in publicly available sequence repositories (GenBank/EMBL/DDBJ).
Proper citation: NCBI Assembly Archive Viewer (RRID:SCR_012917) Copy
Center with mission to conduct and support medical research and research training and to disseminate science-based information on diabetes and other endocrine and metabolic diseases. The NIDDK supports a wide range of medical research through grants to universities and other medical research institutions across the country.
Proper citation: NIDDK - National Institute of Diabetes and Digestive and Kidney Diseases (RRID:SCR_012895) Copy
International collaboration of the International Nucleotide Sequence Databases (INSD), DDBJ, ENA, and GenBank, maintained for over 18 years. Individuals submitting data to the international sequence databases should be aware of INSDC policy.
Proper citation: INSDC (RRID:SCR_011967) Copy
http://www.ncbi.nlm.nih.gov/mapview/map_search.cgi?taxid=7165
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 11, 2023. A database for the Anopheles gambiae str. PEST genome that was sequenced using a whole genome shotgun approach. The database aims to contribute to the understanding of mosquito genome structure and organization and will assist the development of malaria control strategies and improved anti-malarial drugs and vaccines. Sequences were generated and assembled into contigs for submission to GenBank.
Proper citation: Anopheles gambiae (African malaria mosquito) genome view (RRID:SCR_004402) Copy
http://www.ncbi.nlm.nih.gov/biosample
Database containing descriptions of biological source materials used in experimental assays. Sources include: GenBank, Sequence Read Archive (SRA), Coriell, ATCC. Submissions are supported by a web-based Submission Portal that guides users through a series of forms for input of rich metadata describing their samples. As the capacity and complexity of biological data sets expands, databases face new challenges in ensuring that the information is adequately organized and described. The NCBI BioSample database is being developed to help address the challenges by providing the means by which data generators can organize and describe a broad range of sample types, and link to corresponding sets of experimental data in archival databases.
Proper citation: NCBI BioSample (RRID:SCR_004854) Copy
Global registry of research data repositories from all academic disciplines that allows the easy identification of appropriate research data repositories, both for data producers and users. Information icons display principal attributes of a repository that can be used for multi-faceted searches. Repository operators can suggest their infrastructures to be listed via a simple application form. A repository is indexed when the minimum requirements are met, i.e. mode of access to the data and repository, as well as the terms of use.
Proper citation: re3data.org (RRID:SCR_006782) Copy
http://erilllab.umbc.edu/research/software/xfitom/
A fully customizable program that uses a graphical user interface to locate transcription factor-binding sites in genomic sequences. xFITOM scans DNA or RNA sequences for putative binding sites as defined by a collection of aligned known sites, a consensus sequence in IUPAC degenerate-base format, or a combination of the two.
Proper citation: xFITOM (RRID:SCR_014445) Copy
NIH initiative to support production of cDNA libraries, clones and 5'/3' sequences and to provide set of full-length (open reading frame) sequences and cDNA clones of expressed genes for Xenopus laevis and Xenopus tropicalis. Clones distribution is outsourced to for profit companies. Project concluded in September 2008. Resources generated by XGC are publicly accessible to biomedical research community. All sequences are deposited into GenBank.Corresponding clones are available through IMAGE clone distribution network. With conclusion of XGC project, GenBank records of XGC sequences will be frozen, without further updates. Since knowledge of what constitutes full-length coding region for some of genes and transcripts for which we have XGC clones will likely change in future, users planning to order XGC clones will need to monitor for these changes. Users can make use of genome browsers and gene-specific databases, such as UCSC Genome browser, NCBI's Map Viewer, and Entrez Gene, to view relevant regions of genome (browsers) or gene-related information (Entrez Gene).
Proper citation: Xenopus Gene Collection (RRID:SCR_007023) Copy
Part of zebrafish genome project. ZGC project to produce cDNA libraries, clones and sequences to provide complete set of full-length (open reading frame) sequences and cDNA clones of expressed genes for zebrafish. All ZGC sequences are deposited in GenBank and clones can be purchased from distributors of IMAGE consortium. With conclusion of ZGC project in September 2008, GenBank records of ZGC sequences will be frozen, without further updates. Since definition of what constitutes full-length coding region for some of genes and transcripts for which we have ZGC clones will likely change in future, users planning to order ZGC clones will need to monitor for these changes. Users can make use of genome browsers and gene-specific databases, such as UCSC Genome browser, NCBI's Map Viewer, and Entrez Gene, to view relevant regions of genome (browsers) or gene-related information (Entrez Gene).
Proper citation: Zebrafish Gene Collection (RRID:SCR_007054) Copy
https://www.ncbi.nlm.nih.gov/labs/virus/vssi
Community portal for viral sequence data from RefSeq, GenBank and other NCBI repositories. Integrative, value added resource designed to support retrieval, display and analysis of curated collection of virus sequences and large sequence datasets. Used to increase usability of data archived in GenBank and other NCBI repositories.
Proper citation: NCBI Virus (RRID:SCR_018253) Copy
http://heimanlab.com/cut2.html
Software tool to find restriction endonucleases. Helps restriction map nucleotide sequences. Tool with customizable interface, platform independent accessibility, interfaces to NCBI's GenBank, DNA sequence database, and NEB's REBase, and restriction enzyme database. In addition to restriction site mapping, Webcutter 2 also performs degenerate digests, including option of finding restriction sites that can be introduced into sequence by silent mutagenesis.
Proper citation: Webcutter (RRID:SCR_017638) Copy
http://www.premierbiosoft.com/protein_quantification_software/index.html
Software package as comprehensive qualitative and quantitative suite for proteomics. Used to validate and quantify proteins by combining results from popular mass spectrometry platforms and database search engines. Provides customizable interface to support any form of biological annotation. Used to compare protein quantitative results in relation to biological pathways, protein localization, protein function, or to transcript abundance. Every protein identification can be linked to any external or internal knowledge database. Custom links are provided to GenBank, UniProt, IPI, and SwissProt databases or in-house LIMS.
Proper citation: PremierBiosoft Proteo IQ Software (RRID:SCR_018072) Copy
http://linux1.softberry.com/spldb/SpliceDB.html
Database of canonical and non-canonical mammalian splice sites. The information about verified splice site sequences for canonical and non-canonical sites is presented with the supporting evidence. Weight matrices were built for the major splice groups, which can be incorporated into gene prediction programs.
Proper citation: SpliceDB (RRID:SCR_006262) Copy
http://www.sci.unisannio.it/docenti/rampone/
Data set of Homo Sapiens Exons, Introns and Splice regions extracted from GenBank Rel.123 with an aim of giving standardized material to train and to assess the prediction accuracy of computational approaches for gene identification and characterization. From the complete GenBank (Primate Sequences Division) Rel.123 (162,557 entries), entries of Human Nuclear DNA including Complete CDS and more than one Exon have been selected, and 4523 exons and 3802 introns have been extracted from these entries. Details about extracted exons and introns are reported (Locus, number, Start and End position in the entry, sequence, length, G+C content, presence of not AGCT data (nucleotide scan check)). Statistics are also reported (overall nucleotides, average G+C content, nucleotide scan check results, number of not GT starting / AG ending introns, minimum / maximum / average length, length standard deviation). 3799+3799 donor and acceptor sites, as windows of 140 nucleotides around each splice site have been extracted. After discarding sequences not including canonical GTAG junctions (65+74), including insufficient data (not enough material for a 140 nucleotide window) (686+589), including not AGCT bases (29+30), and redundant (218+226) there are 2796+ 2880 windows. Finally, there are 271,937 + 332,296 windows of false splice sites, selected by searching canonical GTAG pairs in not splicing positions. The false sites in a range of +/- 60 from a true splice site are marked as proximal.
Proper citation: HS3D - Homo Sapiens Splice Sites Dataset (RRID:SCR_002939) Copy
https://scicrunch.org/resolver/SCR_002250
THIS RESOURCE IS NO LONGER IN SERVICE. Documented Jul 19, 2024. Metadatabase manually curated that provides web accessible tools related to genomics, transcriptomics, proteomics and metabolomics. Used as informative directory for multi-omic data analysis.
Proper citation: OMICtools (RRID:SCR_002250) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.