Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://code.google.com/p/lapdftext/
Software that facilitates accurate extraction of text from PDF files of research articles for use in text mining applications. It is intended for both scientists and natural language processing (NLP) engineers interested in getting access to text within specific sections of research articles. The system extracts text blocks from PDF-formatted full-text research articles and classifies them into logical units based on rules that characterize specific sections. The LA-PDFText system focuses only on the textual content of the research articles. The current version of LA-PDFText is a baseline system that extracts text using a three-stage process: * identification of blocks of contiguous text * classification of these blocks into rhetorical categories * extraction of the text from blocks grouped section-wise.
Proper citation: lapdftext (RRID:SCR_006167) Copy
http://www.broadinstitute.org/mammals/haploreg/haploreg.php
HaploReg is a tool for exploring annotations of the noncoding genome at variants on haplotype blocks, such as candidate regulatory SNPs at disease-associated loci. Using linkage disequilibrium (LD) information from the 1000 Genomes Project, linked SNPs and small indels can be visualized along with their predicted chromatin state in nine cell types, conservation across mammals, and their effect on regulatory motifs. HaploReg is designed for researchers developing mechanistic hypotheses of the impact of non-coding variants on clinical phenotypes and normal variation.
Proper citation: HaploReg (RRID:SCR_006796) Copy
http://www.mbl.org/mbl_main/atlas.html
High-resolution electronic atlases for mouse strains c57bl/6j, a/j, and dba/2j in either coronal or horizontal section. About this Atlas: The anterior-posterior coordinates are taken from an excellent print atlas of a C57BL/6J brain by K. Franklin and G. Paxinos (The Mouse Brain in Stereotaxic Coordinates, Academic Press, San Diego, 1997, ISBN Number 0-12-26607-6; Library of Congress: QL937.F72). The abbreviations we have used to label the sections conform to those in the Franklin-Paxinos atlas. A C57BL/6J mouse brain may contain as many as 75 million neurons, 23 million glial cells, 7 million endothelial cells associated with blood vessels, and 3 to 4 million miscellaneous pial, ependymal, and choroid plexus cells (see data analysis in Williams, 2000). We have not yet counted total cell number in DBA/2J mice, but the counts are probably appreciably lower.The brain and sections were all processed as described in our methods section. The enlarged images have a pixel count of 1865 x 1400 and the resolution is 4.5 microns/pixel for the processed sections.Plans: In the next several years we hope to add several additional atlases of the same sort for other strains of mice. A horizontal C57BL/6J atlas and a DBA/2J coronal atlas were completed by Tony Capra, summer 2000, and additional atlases may be made over the next several years. As describe in the MBL Procedures Section is not hard to make your own strain-specific atlas from the high resolution images in the MBL.
Proper citation: Mouse Brain Atlases (RRID:SCR_007127) Copy
http://www.cpc.unc.edu/projects/addhealth
Longitudinal study of a nationally representative sample of adolescents in grades 7-12 in the United States during the 1994-95 school year. Public data on about 21,000 people first surveyed in 1994 are available on the first phases of the study, as well as study design specifications. It also includes some parent and biomarker data. The Add Health cohort has been followed into young adulthood with four in-home interviews, the most recent in 2008, when the sample was aged 24-32. Add Health combines longitudinal survey data on respondents social, economic, psychological and physical well-being with contextual data on the family, neighborhood, community, school, friendships, peer groups, and romantic relationships, providing unique opportunities to study how social environments and behaviors in adolescence are linked to health and achievement outcomes in young adulthood. The fourth wave of interviews expanded the collection of biological data in Add Health to understand the social, behavioral, and biological linkages in health trajectories as the Add Health cohort ages through adulthood. The restricted-use contract includes four hours of free consultation with appropriate staff; after that, there''s a fee for help. Researchers can also share information through a listserv devoted to the database.
Proper citation: Add Health (National Longitudinal Study of Adolescent Health) (RRID:SCR_007434) Copy
http://posa.sanfordburnham.org/fatcat-cgi/cgi/FSN/fsn.pl
Flexible Structural Neighborhood is a database of structural neighbors of proteins as seen by FATCAT - a flexible protein structure alignment program. The server accepts either a protein (PDB ID) or a domain (SCOP ID) as a query. For the former case, the server first displays the information of chains and domains of a given protein. Afterwards, users can retrieve similar structures for a domain (if domain information is available, i.e., the protein is collected by SCOP), or for a chain otherwise. The protein structure database we collected for similar structure search includes a representative set at 90% sequence identity of SCOP domains, and of up-to-date PDB entries that are not included in the latest release of SCOP.
Proper citation: FATCAT Flexible Structural Neighborhood (RRID:SCR_007665) Copy
https://github.com/wyp1125/MCScanx
Software toolkit for detection and evolutionary analysis of gene synteny and collinearity.
Proper citation: MCScanX (RRID:SCR_022067) Copy
http://www.icn.ucl.ac.uk/motorcontrol/imaging/propatlas.htm
A probabilistic atlas of the cerebellar lobules in the space defined by the MNI152 template. The anatomical definitions are based on the fMRI atlas of an individual cerebellum by Schmahmann et al. (2000). To obtain a representative anatomical atlas, we separately masked the lobules on T1-weighted MRI scans (1mm isotropic resolution) of 20 healthy young participants (10 male, 10 female, average age 23.7 yrs). Using a different set of 23 participants, we also masked the deep cerebellar nucelei. These cerebella were then aligned using different commonly used normalization algorithms. The resultant probabilistic maps allow for the valid assignment of functional activations to specific cerebellar lobules and the nuclei, while providing a quantitative measure of the certainty of such assignments. Furthermore, maximum probability maps derived from these atlases can be used to define regions of interest (ROIs) in functional neuroimaging and neuroanatomical research. The atlas is included in the newer releases of FSL and the Anatomy toolbox. More version of the atlases for use with MRICroN are also available.
Proper citation: Probabilistic atlas of the human cerebellum (RRID:SCR_008797) Copy
Community-driven organization that develops and disseminates software for geophysics and related fields. They host codes in a wide range of disciplines in geodynamics and computational science including geodynamo, long-term tectonics, magma migration, mantle dynamics, seismology, and short-term crustal dynamics.
Proper citation: Computational Infrastructure for Geodynamics (RRID:SCR_003371) Copy
http://research.amnh.org/atol/files/
Project whose aim is to produce a robust phylogeny of all the deepest branches within a mega-diverse group, the spiders, by combining a massive amount of newly generated comparative genomic data with a substantial set of new and re-assessed data on morphology and behavior. They propose to collect a huge amount of genomic information in order to test and improve the results achieved by over 50 detailed morphological cladistic analyses conducted by more than 30 investigators during the past 15 years. The insignificant amount of genomic work to date on spiders has been uncoordinated and of little utility for broad-scale phylogenetic investigation. The advent of high-throughput DNA sequencing, however, makes it feasible to examine substantial parts of the genome across a dense sampling of spider taxa. They propose to sequence at least 50 loci (genome samples of 500-1,000 or more base pairs that can be sequenced as single pieces in both directions simultaneously) for representatives of at least 500 genera of spiders and their closest relatives (the whipscorpion orders Amblypygi, Uropygi, and Schizomida). These genera will be carefully selected by a sampling strategy designed to maximize the resolution of deep branches within spider phylogeny, and will purposefully include all the previously most-favored study organisms of ethologists, ecologists, physiologists, and developmental and molecular biologists, thus integrating and contextualizing their research. Data matrices will be produced that combine the new genomic data with a new, comprehensive survey of morphological and behavioral homologies, offering a unique index to all comparative data on one large group. New computer software, designed in large part by members of their group and using massively parallel processing to achieve supercomputing capability, makes such analyses feasible.
Proper citation: Tree of Life: Phylogeny of Spiders (RRID:SCR_003801) Copy
https://www.opensciencedatacloud.org/
Service that provides petabyte-scale cloud resources to analyze, manage, and share scientific data. It is designed to serve medium to large sized research projects by managing and operating a secure cloud computing infrastructure that can be shared across a project. This Science as a Service approach to research saves scientists and their funders valuable time and money. All of the software developed is open source and hosted on GitHub. The OSDC also has 1PB of public data in a wide variety of disciplines. The data sets can downloaded over the internet or high performance networks such as Internet2, as well as computed over directly on the OSDC.
Proper citation: Open Science Data Cloud (RRID:SCR_003523) Copy
The Dynamic Regulatory Events Miner (DREM) allows one to model, analyze, and visualize transcriptional gene regulation dynamics. The method of DREM takes as input time series gene expression data and static transcription factor-gene interaction data (e.g. ChIP-chip data), and produces as output a dynamic regulatory map. The dynamic regulatory map highlights major bifurcation events in the time series expression data and transcription factors potentially responsible for them. DREM 2.0 was released and supports a number of new features including: * new static binding data for mouse, human, D. melanogaster, A. thaliana * a new and more flexible implementation of the IOHMM supports dynamic binding data for each time point or as a mix of static/dynamic TF input * expression levels of TFs can be used to improve the models learned by DREM * the motif finder DECOD can be used in conjuction with DREM and help find DNA motifs for unannotated splits * new features for the visualization of expressed TFs, dragging boxes in the model view, and switching between representations
Proper citation: Dynamic Regulatory Events Miner (RRID:SCR_003080) Copy
A web application which provides altmetrics to help researchers measure and share the impacts of their research outputs. After making a profile, scientists can track which of their publications are most popular through number of citations, frequency of PDF downloads, etc. Information from research outputs such as journal articles, blog posts, datasets, and software contribute to a user's impact, which is viewable in their profile.
Proper citation: ImpactStory (RRID:SCR_002632) Copy
Project to create a scalable infrastructure that enables linking phenotypes across different fields of biology by the semantic similarity of their descriptions.
Proper citation: Phenoscape (RRID:SCR_003799) Copy
http://www.cs.cmu.edu/~jernst/stem/
The Short Time-series Expression Miner (STEM) is a Java program for clustering, comparing, and visualizing short time series gene expression data from microarray experiments (~8 time points or fewer). STEM allows researchers to identify significant temporal expression profiles and the genes associated with these profiles and to compare the behavior of these genes across multiple conditions. STEM is fully integrated with the Gene Ontology (GO) database supporting GO category gene enrichment analyses for sets of genes having the same temporal expression pattern. STEM also supports the ability to easily determine and visualize the behavior of genes belonging to a given GO category or user defined gene set, identifying which temporal expression profiles were enriched for these genes. (Note: While STEM is designed primarily to analyze data from short time course experiments it can be used to analyze data from any small set of experiments which can naturally be ordered sequentially including dose response experiments.) Platform: Windows compatible, Mac OS X compatible, Linux compatible, Unix compatible
Proper citation: Short Time-series Expression Miner (STEM) (RRID:SCR_005016) Copy
An open-source general packing algorithm that packs 3D objects onto surfaces, into volumes, and around volumes. It provides a general architecture to allow various packing algorithms to interoperate efficiently in the same model. autoPack can incorporate any packing solution into its modular python program architecture, but is currently optimized to provide a novel solution to the loose packing problem which places objects of discrete size into place (compared to advancing front, popcorn, or other fast tight-packing solutions that allow objects to scale to arbitrary masses.) Most popular 3D software programs now contain robust physics engines based on Bullet that can separate small collections of overlapping objects or allow volumes to be filled by pouring shapes from generators, but these approaches fails for large complex systems and result in either overlapping geometry, crashed software, or non-random gradients. Most packing algorithms are designed to position objects as efficiently as possible, but autoPack allows the user to select from random loose packing to highly organized packing methods����??even to choose both methods at the same time. autoPack positions 3D geometries into, onto, and around volumes with minimal to zero overlap. autoPack mixes several packing approaches and procedural growth algorithms. autoPack can thus place objects with forces and constraints to allow a high degree of control ranging from completely random distributions to highly ordered structures. * zero to minimal overlaps depending on the method used * accuracy vs speed parameters selected by the user * zero edge effects * complete control, from fully random to fully ordered distributions * agent-based interaction, weighting, and collision control
Proper citation: Autopack (RRID:SCR_006830) Copy
https://github.com/CBL-ORION/orion
Project to develop tools that explore single neuron function via sophisticated image analysis. ORION software bridges advanced optical imaging and compartmental modeling of neuronal function by rapidly, accurately, and robustly generating, from structural image data, a cylindrical morphology model suitable for simulating neuronal function.
Proper citation: ORION (RRID:SCR_010621) Copy
Archive of earthquake data for research in seismology and earthquake engineering in Southern California recorded or processed by the Southern California Seismic Network (SCSN). Users can access information on: * Recent earthquakes detected by the SCSN * Significant southern California earthquakes and faults * The southern California earthquake catalog, spanning from 1933 to present * Waveform and metadata files of SCSN seismic stations from 1977 to present * Data sets created by SCEC scientists to assist in ongoing and future research
Proper citation: Southern California Earthquake Data Center (RRID:SCR_000663) Copy
Web accessible database of data extracted from scientific literature, focusing on proteins that are drug-targets or candidate drug-targets and for which structural data are present in Protein Data Bank . Website supports query types including searches by chemical structure, substructure and similarity, protein sequence, ligand and protein names, affinity ranges and molecular weight . Data sets generated by BindingDB queries can be downloaded in form of annotated SDfiles for further analysis, or used as basis for virtual screening of compound database uploaded by user. Data are linked to structural data in PDB via PDB IDs and chemical and sequence searches, and to literature in PubMed via PubMed IDs .
Proper citation: BindingDB (RRID:SCR_000390) Copy
Data related to the National Critical Zone Observatory Program including in-situ environmental sensors, field instruments, remote sensing, and surface and subsurface imaging. The Program serves the international scientific community through research, infrastructure, data, and models. They focus on how components of the Critical Zone interact, shape Earth's surface, and support life. A primary goal is to develop high-resolution 4D datasets that inform our theoretical framework, constrain our conceptual and coupled systems models, and test our model-generated hypotheses. They are developing cross-CZO capabilities to easily share, integrate, analyze and preserve the wide range of multi-disciplinary data generated by CZOs.
Proper citation: Critical Zone Observatories (RRID:SCR_002199) Copy
http://www.marine-geo.org/portals/seismic/
Seismic Reflection Field Data from the academic research community. Their partner Academic Seismic Portal at UTIG offers additional seismic resources, http://www.ig.utexas.edu/sdc/
Proper citation: Academic Seismic Portal at LDEO (RRID:SCR_002194) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the NIF Resources search. From here you can search through a compilation of resources used by NIF and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that NIF has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on NIF then you can log in from here to get additional features in NIF such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into NIF you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within NIF that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.