Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
An information extracting and processing package for biological literature that can be used online or installed locally via a downloadable software package, http://www.textpresso.org/downloads.html Textpresso's two major elements are (1) access to full text, so that entire articles can be searched, and (2) introduction of categories of biological concepts and classes that relate two objects (e.g., association, regulation, etc.) or describe one (e.g., methods, etc). A search engine enables the user to search for one or a combination of these categories and/or keywords within an entire literature. The Textpresso project serves the biological and biomedical research community by providing: * Full text literature searches of model organism research and subject-specific articles at individual sites. Major elements of these search engines are (1) access to full text, so that the entire content of articles can be searched, and (2) search capabilities using categories of biological concepts and classes that relate two objects (e.g., association, regulation, etc.) or identify one (e.g., cell, gene, allele, etc). The search engines are flexible, enabling users to query the entire literature using keywords, one or more categories or a combination of keywords and categories. * Text classification and mining of biomedical literature for database curation. They help database curators to identify and extract biological entities and facts from the full text of research articles. Examples of entity identification and extraction include new allele and gene names and human disease gene orthologs; examples of fact identification and extraction include sentence retrieval for curating gene-gene regulation, Gene Ontology (GO) cellular components and GO molecular function annotations. In addition they classify papers according to curation needs. They employ a variety of methods such as hidden Markov models, support vector machines, conditional random fields and pattern matches. Our collaborators include WormBase, FlyBase, SGD, TAIR, dictyBase and the Neuroscience Information Framework. They are looking forward to collaborating with more model organism databases and projects. * Linking biological entities in PDF and online journal articles to online databases. They have established a journal article mark-up pipeline that links select content of Genetics journal articles to model organism databases such as WormBase and SGD. The entity markup pipeline links over nine classes of objects including genes, proteins, alleles, phenotypes, and anatomical terms to the appropriate page at each database. The first article published with online and PDF-embedded hyperlinks to WormBase appeared in the September 2009 issue of Genetics. As of January 2011, we have processed around 70 articles, to be continued indefinitely. Extension of this pipeline to other journals and model organism databases is planned. Textpresso is useful as a search engine for researchers as well as a curation tool. It was developed as a part of WormBase and is used extensively by C. elegans curators. Textpresso has currently been implemented for 24 different literatures, among them Neuroscience, and can readily be extended to other corpora of text.
Proper citation: Textpresso (RRID:SCR_008737) Copy
http://openconnectomeproject.org/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 9, 2023. Connectomes repository to facilitate the analysis of connectome data by providing a unified front for connectomics research. With a focus on Electron Microscopy (EM) data and various forms of Magnetic Resonance (MR) data, the project aims to make state-of-the-art neuroscience open to anybody with computer access, regardless of knowledge, training, background, etc. Open science means open to view, play, analyze, contribute, anything. Access to high resolution neuroanatomical images that can be used to explore connectomes and programmatic access to this data for human and machine annotation are provided, with a long-term goal of reconstructing the neural circuits comprising an entire brain. This project aims to bring the most state-of-the-art scientific data in the world to the hands of anybody with internet access, so collectively, we can begin to unravel connectomes. Services: * Data Hosting - Their Bruster (brain-cluster) is large enough to store nearly any modern connectome data set. Contact them to make your data available to others for any purpose, including gaining access to state-of-the-art analysis and machine vision pipelines. * Web Viewing - Collaborative Annotation Toolkit for Massive Amounts of Image Data (CATMAID) is designed to navigate, share and collaboratively annotate massive image data sets of biological specimens. The interface is inspired by Google Maps, enhanced to allow the exploration of 3D image data. View the fork of the code or go directly to view the data. * Volume Cutout Service - RESTful API that enables you to select any arbitrary volume of the 3d database (3ddb), and receive a link to download an HDF5 file (for matlab, C, C++, or C#) or a NumPy pickle (for python). Use some other programming language? Just let them know. * Annotation Database - Spatially co-registered volumetric annotations are compactly stored for efficient queries such as: find all synapses, or which neurons synapse onto this one. Create your own annotations or browse others. *Sample Downloads - In addition to being able to select arbitrary downloads from the datasets, they have also collected a few choice volumes of interest. * Volume Viewer - A web and GPU enabled stand-alone app for viewing volumes at arbitrary cutting planes and zoom levels. The code and program can be downloaded. * Machine Vision Pipeline - They are building a machine vision pipeline that pulls volumes from the 3ddb and outputs neural circuits. - a work in progress. As soon as we have a stable version, it will be released. * Mr. Cap - The Magnetic Resonance Connectome Automated Pipeline (Mr. Cap) is built on JIST/MIPAV for high-throughput estimation of connectomes from diffusion and structural imaging data. * Graph Invariant Computation - Upload your graphs or streamlines, and download some invariants. * iPad App - WholeSlide is an iPad app that accesses utilizes our open data and API to serve images on the go.
Proper citation: Open Connectome Project (RRID:SCR_004232) Copy
http://proteome.gs.washington.edu/software/bibliospec/documentation/index.html
BiblioSpec enables the identification of peptides from tandem mass spectra by searching against a database of previously identified spectra. This suite of software tools is for creating and searching MS/MS peptide spectrum libraries. BiblioSpec is available free of charge for noncommercial use through an interactive web-site at http://depts.washington.edu/ventures/UW_Technology/Express_Licenses/bibliospec.php The BiblioSpec package contains the following programs: * BlibBuild creates a library of peptide MS/MS spectra from MS2 files. * BlibFilter removes redundant spectra from a library. * BlibSearch searches a spectrum library for matches to query spectra, reporting the results in an SQT file. In addition to the primary programs, the following auxiliary programs are available: * BlibStats writes summary statistics describing a library. * BlibToMS2 writes a library in MS2 file format. * BlibUpdate adds, deletes, or annotates spectra. * BlibPpMS2 processes spectra (bins peaks, removes noise, normalizes intensity) as done in BlibSearch and prints the resulting spectra to a text file. Several reference libraries are available for download. These libraries are updated regularly and are for use under the Linux operating system. You will find libraries for * Escherichia coli * Saccharomyces cerevisiae * Caenorhabditis elegans
Proper citation: BiblioSpec (RRID:SCR_004349) Copy
http://nematode.lab.nig.ac.jp/
Expression pattern map of the 100Mb genome of the nematode Caenorhabditis elegans through EST analysis and systematic whole mount in situ hybridization. NEXTDB is the database to integrate all information from their expression pattern project and to make the data available to the scientific community. Information available in the current version is as follows: * Map: Visual expression of the relationships among the cosmids, predicted genes and the cDNA clones. * Image: In situ hybridization images that are arranged by their developmental stages. * Sequence: Tag sequences of the cDNA clones are available. * Homology: Results of BLASTX search are available. Users of the data presented on our web pages should not publish the information without our permission and appropriate acknowledgment. Methods are available for: * In situ hybridization on whole mount embryos of C.elegans * Protocols for large scale in situ hybridization on C.elegans larvae
Proper citation: NEXTDB (RRID:SCR_004480) Copy
http://cbl-gorilla.cs.technion.ac.il/
A tool for identifying and visualizing enriched GO terms in ranked lists of genes. It can be run in one of two modes: * Searching for enriched GO terms that appear densely at the top of a ranked list of genes or * Searching for enriched GO terms in a target list of genes compared to a background list of genes.
Proper citation: GOrilla: Gene Ontology Enrichment Analysis and Visualization Tool (RRID:SCR_006848) Copy
http://genetrail.bioinf.uni-sb.de/
A web-based application that analyzes gene sets for statistically significant accumulations of genes that belong to some functional category. Considered category types are: KEGG Pathways, TRANSPATH Pathways, TRANSFAC Transcription Factor, GeneOntology Categories, Genomic Localization, Protein-Protein Interactions, Coiled-coil domains, Granzyme-B clevage sites, and ELR/RGD motifs. The web server provides two statistical approaches, "Over-Representation Analysis" (ORA) comparing a reference set of genes to a test set, and "Gene Set Enrichment Analysis" (GSEA) scoring sorted lists of genes., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: GeneTrail (RRID:SCR_006250) Copy
http://www.benoslab.pitt.edu/comir/
Data analysis service that predicts whether a given mRNA is targeted by a set of miRNAs. ComiR uses miRNA expression to improve and combine multiple miRNA targets for each of the four prediction algorithms: miRanda, PITA, TargetScan and mirSVR. The composite scores of the four algorithms are then combined using a support vector machine trained on Drosophila Ago1 IP data.
Proper citation: ComiR (RRID:SCR_013023) Copy
Anatomical atlas about structural anatomy of Caenorhabditis elegans. Provides simple interface allowing user to easily navigate through every anatomical structure of worm. Contains set of images which can be sorted by different characteristics: sex, genotype, age, body portion or tissue type. Includes links to other major worm websites and databases. Application for viewing and downloading thousands of unpublished electron micrographs and associated data. These images have been generated by several labs in the C. elegans community, including the MRC, the Hall lab (Center for C. elegans Anatomy), and the Culotti and Riddle labs.
Proper citation: WormAtlas (RRID:SCR_002861) Copy
https://github.com/lucventurini/mikado/
Mikado is a lightweight Python3 pipeline whose purpose is to facilitate the identification of expressed loci from RNA-Seq data * and to select the best models in each locus.
Proper citation: Mikado (RRID:SCR_016159) Copy
Software for designing CRISPR/Cas guide RNA with reduced off target sites. Used for rational design of CRISPR/Cas target. Web server for selecting rational CRISPR/Cas targets from input sequence. Server currently incorporates genomic sequences of human, mouse, rat, marmoset, pig, chicken, frog, zebrafish, Ciona, fruit fly, silkworm, Caenorhabditis elegans, Arabidopsis, rice, Sorghum and budding yeast.
Proper citation: CRISPRdirect (RRID:SCR_018186) Copy
Freely accessible phenotype-centered database with integrated analysis and visualization tools. It combines diverse data sets from multiple species and experiment types, and allows data sharing across collaborative groups or to public users. It was conceived of as a tool for the integration of biological functions based on the molecular processes that subserved them. From these data, an empirically derived ontology may one day be inferred. Users have found the system valuable for a wide range of applications in the arena of functional genomic data integration.
Proper citation: Gene Weaver (RRID:SCR_003009) Copy
An algorithm for the identification of microRNA targets. Details are provided (3' UTR alignments with predicted sites, links to various public databases etc) regarding: # microRNA target predictions in vertebrates (Krek et al, Nature Genetics 37:495-500 (2005)) # microRNA target predictions in seven Drosophila species (Grn et al, PLoS Comp. Biol. 1:e13 (2005)) # microRNA targets in three nematode species (Lall et al, Current Biology 16, 1-12 (2006)) # human microRNA targets that are not conserved but co-expressed (i.e. the microRNA and mRNA are expressed in the same tissue) (Chen and Rajewsky, Nat Genet 38, 1452-1456 (2006)) co-expressed targets
Proper citation: PicTar (RRID:SCR_003343) Copy
http://www.ihop-net.org/UniPub/iHOP/
Information system that provides a network of concurring genes and proteins extends through the scientific literature touching on phenotypes, pathologies and gene function. It provides this network as a natural way of accessing millions of PubMed abstracts. By using genes and proteins as hyperlinks between sentences and abstracts, the information in PubMed can be converted into one navigable resource, bringing all advantages of the internet to scientific literature research. Moreover, this literature network can be superimposed on experimental interaction data (e.g., yeast-two hybrid data from Drosophila melanogaster and Caenorhabditis elegans) to make possible a simultaneous analysis of new and existing knowledge. The network contains half a million sentences and 30,000 different genes from humans, mice, D. melanogaster, C. elegans, zebrafish, Arabidopsis thaliana, yeast and Escherichia coli.
Proper citation: Information Hyperlinked Over Proteins (RRID:SCR_004829) Copy
A database of human molecular interaction networks that integrates human protein-protein and transcriptional regulatory interactions from 15 distinct resources and aims to give direct and easy access to the integrated data set and to enable users to perform network-based investigations. The database includes tools (i) to search for molecular interaction partners of query genes or proteins in the integrated dataset, (ii) to inspect the origin, evidence and functional annotation of retrieved proteins and interactions, (iii) to visualize and adjust the resulting interaction network, (iv) to filter interactions based on method of derivation, evidence and type of experiment as well as based on gene expression data or gene lists and (v) to analyze the functional composition of interaction networks.
Proper citation: Unified Human Interactome (RRID:SCR_005805) Copy
A gene and protein interactions database designed specifically for the model organism Drosophila including protein-protein, transcription factor-gene, microRNA-gene, and genetic interactions. For advanced searches and dynamic graphing capabilities the IM Browser and a DroID Cytoscape plugin are available.
Proper citation: DroID - Drosophila Interactions Database (RRID:SCR_006634) Copy
Database for conserved sequence motifs identified by genome scale motif discovery, similarity, clustering, co-occurrence and coexpression calculations. Sequence inputs include low-coverage genome sequence data and ENCODE data. The database offers information on atomic motifs, motif groups and patterns. In promoter-based cisRED databases, sequence search regions for motif discovery extend from 1.5 Kb upstream to 200b downstream of a transcription start site, net of most types of repeats and of coding exons. Many transcription factor binding sites are located in such regions. For each target gene's search region, a base set of probabilistic ab initio discovery tools is used, in parallel, to find over-represented atomic motifs. Discovery methods use comparative genomics with over 40 vertebrate input genomes. In ChIP-seq-based cisRED databases, sequence search regions for motif discovery correspond to significant peaks that represent genome-wide sites of protein-DNA binding. Because such peaks occur in a wide range of genic and intergenic locations, ChIP-seq and promoter-based databases are complementary. Currently, motif discovery for ChIP-seq data uses scan-based approaches that make more explicit use of sets of sequences known to be functional transcription factor binding sites, and that consider a wide range of levels of conservation. For the human STAT1 ChIP-seq database search regions in the target species (human) was selected +/- 300 bp around the ChIP-seq peak maximum. Repeats and coding regions were masked. Multiple sequence alignment were used to assemble orthologous input sequences from other species.
Proper citation: cisRED: cis-regulatory element (RRID:SCR_002098) Copy
http://spliceosomedb.ucsc.edu/
A database of proteins and RNAs that have been identified in various purified splicing complexes. Various names, orthologs and gene identifiers of spliceosome proteins have been cataloged to navigate the complex nomenclature of spliceosome proteins. Links to gene and protein records are also provided for the spliceosome components in other databases. To navigate spliceosome assembly dynamics, tools were created to compare the association of spliceosome proteins with complexes that form at specific stages of spliceosome assembly based on a compendium of mass spectrometry experiments that identified proteins in purified splicing complexes.
Proper citation: Spliceosome Database (RRID:SCR_002097) Copy
Cross-species microarray expression database focusing on high-throughput expression data relevant for germline development, meiosis and gametogenesis as well as the mitotic cell cycle. The database contains a unique combination of information: 1) High-throughput expression data obtained with whole-genome high-density oligonucleotide microarrays (GeneChips). 2) Sample annotation (mouse over the sample name and click on it) using the Multiomics Information Management and Annotation System (MIMAS 3.0). 3) In vivo protein-DNA binding data and protein-protein interaction data (available for selected species). 4) Genome annotation information from Ensembl version 50. 5) Orthologs are identified using data from Ensembl and OMA and linked to each other via a section in the report pages. The portal provides access to the Saccharomyces Genomics Viewer (SGV) which facilitates online interpretation of complex data from experiments with high-density oligonucleotide tiling microarrays that cover the entire yeast genome. The database displays only expression data obtained with high-density oligonucleotide microarrays (GeneChips)., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 15,2026.
Proper citation: GermOnline (RRID:SCR_002807) Copy
A database of high-quality protein-protein interactions in different organisms.
Proper citation: HINT (RRID:SCR_002762) Copy
A platform composed of three modules: the Database, the Search Engine, and rSNPs, for the computational identification of transcription factor binding sites (TFBSs) in multiple genomes, that combines TRANSFAC and JASPAR data with the search power of profile hidden Markov models (HMMs). The Database contains putative TFBSs found in the upstream sequences of genes from the human, mouse and D.melanogaster genomes. For each gene, they scanned the region from 10,000 base pairs upstream of the transcript start to 50 base pairs downstream of the coding sequence start against all their models. Therefore, the database contains putative binding sites in the gene promoter and in the initial introns and non-coding exons. Information displayed for each putative binding site includes the transcription factor name, its position (absolute on the chromosome, or relative to the gene), the score of the prediction, and the region of the gene the site belongs to. If the selected gene has homologs in any of the other two organisms, the program optionally displays the putative TFBSs in the homologs. The Search Engine allows the identification, visualization and selection of putative TFBSs occurring in the promoter or other regions of a gene from the human, mouse, D.melanogaster, C.elegans or S.cerevisiae genomes. In addition, it allows the user to upload a sequence to query and to build a model by supplying a multiple sequence alignment of binding sites for a transcription factor of interest. rSNPs MAPPER is designed to identify Single Nucleotide Polymorphisms (SNPs) that may have an effect on the presence of one or more TFBSs.
Proper citation: MAPPER - Multi-genome Analysis of Positions and Patterns of Elements of Regulation (RRID:SCR_003077) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.