Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://bioinf.bio.uth.gr/nat-ncs2
Web server for the detection and evolutionary classification of prokaryotic and eukaryotic nucleobase-cation symporters of the NAT/NCS2 family. Used to scan, identify and evolutionary classify NAT/NCS2 nucleobase transporter proteins.
Proper citation: NAT/NCS2 Hound (RRID:SCR_016473) Copy
https://genome.jgi.doe.gov/programs/fungi/1000fungalgenomes.jsf
Web application to provide genomic information for fungi. Includes sequenced fungal genomes, those in progress, and selected nominations. Nomination of new species for genome sequencing in the families or only one reference genome possible after providing DNA/RNA samples for their sequencing. Used to explore the diversity of fungi important for energy and the environment.
Proper citation: 1000 Fungal Genome Project (RRID:SCR_016463) Copy
https://github.com/WGS-TB/MentaLiST
Software for a MLST (multi-locus sequence typing) caller, based on a k-mer counting algorithm and written in the Julia language. Designed and implemented to handle large typing schemes.
Proper citation: MentaLiST (RRID:SCR_016469) Copy
https://github.com/SCANDAN-Team/SCANDAN-DICOM-labelling
Software tool for rules for DICOM tag based labelling. Regular expression used during the SCANDAN project to label MRI scans based on DICOM tag.
Proper citation: SCANDAN-DICOM-labelling (RRID:SCR_028365) Copy
https://github.com/DerrickWood/kraken2
Software tool as second version of Kraken taxonomic sequence classification system.
Proper citation: kraken2 (RRID:SCR_026838) Copy
https://github.com/BackofenLab/HVSeeker/tree/main
Software tool for distinguishing between bacterial and phage sequences. Consists of two separate models: one analyzing DNA sequences and the other focusing on proteins.
Proper citation: HVSeeker (RRID:SCR_026120) Copy
A database of hierarchical classification of enzymes that relates specific sequence-structure features to specific chemical capabilities. The SFLD classifies evolutionarily related enzymes according to shared chemical functions and maps these shared functions to conserved active site features. The classification is hierarchical, where broader levels encompass more distantly related proteins with fewer shared features. It thus serves as the analysis and archive site for superfamilies targeted by the Enzyme Function Initiative, and is developed by the Babbitt Laboratory in collaboration with the UCSF Resource for Biocomputing, Visualization, and Informatics. The resource also provides a collection of tools and data for investigating sequence-structure-function relationships and hypothesizing function.
Proper citation: Structure-function linkage database (RRID:SCR_001375) Copy
https://www.ddbj.nig.ac.jp/dra/index-e.html
Archive database for output data generated by next-generation sequencing machines including Roche 454 GS System, Illumina Genome Analyzer, Applied Biosystems SOLiD System, and others. DRA is a member of the International Nucleotide Sequence Database Collaboration (INSDC) and archiving the data in a close collaboration with NCBI Sequence Read Archive (SRA) and EBI Sequence Read Archive (ERA). Please submit the trace data from conventional capillary sequencers to DDBJ Trace Archive., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: DDBJ Sequence Read Archive (RRID:SCR_001370) Copy
https://services.healthtech.dtu.dk/services/NetNGlyc-1.0/
Server that predicts N-Glycosylation sites in human proteins using artificial neural networks that examine the sequence context of Asn-Xaa-Ser/Thr sequons. NetNGlyc 1.0 is also available as a stand-alone software package, with the same functionality as the service above. Ready-to-ship packages exist for the most common UNIX platforms.
Proper citation: NetNGlyc (RRID:SCR_001570) Copy
https://services.healthtech.dtu.dk/services/YinOYang-1.2/
Server that produces neural network predictions for O-beta-GlcNAc attachment sites in eukaryotic protein sequences. This server can also use NetPhos, to mark possible phosphorylated sites and hence identify Yin-Yang sites. YinOYang 1.2 is available as a stand-alone software package, with the same functionality. Ready-to-ship packages exist for the most common UNIX platforms.
Proper citation: YinOYang (RRID:SCR_001605) Copy
https://www.ebi.ac.uk/jdispatcher/msa/clustalo?stype=protein
Software package as multiple sequence alignment tool that uses seeded guide trees and HMM profile-profile techniques to generate alignments between three or more sequences. Accepts nucleic acid or protein sequences in multiple sequence formats NBRF/PIR, EMBL/UniProt, Pearson (FASTA), GDE, ALN/Clustal, GCG/MSF, RSF.
Proper citation: Clustal Omega (RRID:SCR_001591) Copy
http://gmod.org/wiki/Main_Page
A collection of open source software tools for creating and managing genome-scale biological databases. GMOD is made up databases, applications, and adaptor software that connects these components together. You can use it to create a small laboratory database of genome annotations, or a large web-accessible community database. At first GMOD just featured model organisms but now any organism with any kind of sequence associated with it is a good candidate as a subject for a GMOD database. There are GMOD databases with just protein sequence in them, with EST sequence only, those that are concerned primarily with gene expression, and even those dedicated to collections of RNA sequence. They have also heard of GMOD databases for oligonucleotides and plasmids.
Proper citation: Generic Model Organism Database Project (RRID:SCR_001731) Copy
Web application to search protein databases using a translated nucleotide query. Translated BLAST services are useful when trying to find homologous proteins to a nucleotide coding region. Blastx compares translational products of the nucleotide query sequence to a protein database. Because blastx translates the query sequence in all six reading frames and provides combined significance statistics for hits to different frames, it is particularly useful when the reading frame of the query sequence is unknown or it contains errors that may lead to frame shifts or other coding errors. Thus blastx is often the first analysis performed with a newly determined nucleotide sequence and is used extensively in analyzing EST sequences. This search is more sensitive than nucleotide blast since the comparison is performed at the protein level.
Proper citation: BLASTX (RRID:SCR_001653) Copy
http://csg.sph.umich.edu//abecasis/MACH/index.html
A Markov Chain based software tool for haplotyping, genotype imputation and disease association analysis that can resolve long haplotypes or infer missing genotypes in samples of unrelated individuals.
Proper citation: MACH 1.0 (RRID:SCR_001759) Copy
https://github.com/benedictpaten/pecan
A Java consistency based multiple sequence alignment software program.
Proper citation: Pecan (RRID:SCR_001909) Copy
http://www.aspergillus-genomes.org.uk/
A resource for viewing annotated genes arising from various Aspergillus sequencing and annotation projects, resulting from the merging of Central Aspergillus Data REpository (CADRE) and The Aspergillus Website, which took place in June 2008. The principal role of CADRE is to aid the Aspergillus research community by managing Aspergillus genome data and by providing visualization tools, ranging from relatively simple annotation displays to more complex data integration displays. In contrast, The Aspergillus Website provides a range of information to the medical community (i.e., clinicians, patients and scientists) regarding the genus Aspergillus and the diseases, such as Aspergillosis, that it can cause. CADRE has been implemented using the Ensembl v22 suite. This suite comprises: * a database schema, which has been devised for storing annotated eukaryotic genomes. The schema is implemented with the MySQL relational database management system. * several specialized programming modules for building interfaces (i.e., BioPerl and Ensembl API modules). * a series of programs (i.e., Perl CGI scripts using the API modules) for viewing genomic data within a web browser., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Aspergillus Genomes (RRID:SCR_001880) Copy
http://rgp.dna.affrc.go.jp/E/index.html
Rice Genome Research Program (RGP) is an integral part of the Japanese Ministry of Agriculture, Forestry and Fisheries (MAFF) Genome Research Project. RGP now aims to completely sequence the entire rice genome and subsequently to pursue integrated goals in functional genomics, genome informatics and applied genomics. It is jointly coordinated by the National Institute of Agrobiological Sciences (NIAS), a government research institute under MAFF and the Society for Techno-innovation of Agriculture, Forestry and Fisheries (STAFF), a semi-private research organization managed and supported by MAFF and a consortium of some twenty Japanese companies. The research is funded with yearly grants from MAFF and additional funds from the Japan Racing Association (JRA). It is now the leading member of the International Rice Genome Sequencing Project (IRGSP), a consortium of ten countries sharing the sequencing of the 12 rice chromosomes. The IRGSP adopts the clone-by-clone shotgun sequencing strategy so that each sequenced clone can be associated with a specific position on the genetic map and adheres to the policy of immediate release of the sequence data to the public domain. In December 2004, the IRGSP completed the sequencing of the rice genome. The high-quality and map-based sequence of the entire genome is now available in public databases.
Proper citation: Rice Genome Research Project (RRID:SCR_002268) Copy
http://www.structuralgenomics.org/
The Structural Genomics Project aims at determination of the 3D structure of all proteins. It also aims to reduce the cost and time required to determine three-dimensional protein structures. It supports selection, registration, and tracking of protein families and representative targets. This aim can be achieved in four steps : -Organize known protein sequences into families. -Select family representatives as targets. -Solve the 3D structure of targets by X-ray crystallography or NMR spectroscopy. -Build models for other proteins by homology to solved 3D structures. PSI has established a high-throughput structure determination pipeline focused on eukaryotic proteins. NMR spectroscopy is an integral part of this pipeline, both as a method for structure determinations and as a means for screening proteins for stable structure. Because computational approaches have estimated that many eukaryotic proteins are highly disordered, about 1 year into the project, CESG began to use an algorithm. The project has been organized into two separate phases. The first phase was dedicated to demonstrating the feasibility of high-throughput structure determination, solving unique protein structures, and preparing for a subsequent production phase. The second phase, PSI-2, has focused on implementing the high-throughput structure determination methods developed in PSI-1, as well as homology modeling and addressing bottlenecks like modeling membrane proteins. The first phase of the Protein Structure Initiative (PSI-1) saw the establishment of nine pilot centers focusing on structural genomics studies of a range of organisms, including Arabidopsis thaliana, Caenorhabditis elegans and Mycobacterium tuberculosis. During this five-year period over 1,100 protein structures were determined, over 700 of which were classified as unique due to their < 30% sequence similarity with other known protein structures. The primary goal of PSI-1 was to develop methods to streamline the structure determination process, resulted in an array of technical advances. Several methods developed during PSI-1 enhanced expression of recombinant proteins in systems like Escherichia coli, Pichia pastoris and insect cell lines. New streamlined approaches to cell cloning, expression and protein purification were also introduced, in which robotics and software platforms were integrated into the protein production pipeline to minimize required manpower, increase speed, and lower costs. The goal of the second phase of the Protein Structure Initiative (PSI-2) is to use methods introduced in PSI-1 to determine a large number of proteins and continue development in streamlining the structural genomics pipeline. Currently, the third phase of the PSI is being developed and will be called PSI: Biology. The consortia will propose work on substantial biological problems that can benefit from the determination of many protein structures Sponsors: PSI is funded by the U.S. National Institute of General Medical Sciences (NIGMS),
Proper citation: Protein Structure Initiative (RRID:SCR_002161) Copy
Collection of genome databases for vertebrates and other eukaryotic species with DNA and protein sequence search capabilities. Used to automatically annotate genome, integrate this annotation with other available biological data and make data publicly available via web. Ensembl tools include BLAST, BLAT, BioMart and the Variant Effect Predictor (VEP) for all supported species.
Proper citation: Ensembl (RRID:SCR_002344) Copy
http://www.ncbi.nlm.nih.gov/SNP/
General database of genetic variations maintained by the NCBI. Database as central repository for both single base nucleotide substitutions and short deletion and insertion polymorphisms. Distinguishes report of how to assay SNP from use of that SNP with individuals and populations. This separation simplifies some issues of data representation. However, these initial reports describing how to assay SNP will often be accompanied by SNP experiments measuring allele occurrence in individuals and populations. Community can contribute to this resource.
Proper citation: dbSNP (RRID:SCR_002338) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.