Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
The database of protein-chemical structural interactions includes all existing 3D structures of complexes of proteins with low molecular weight ligands. When one considers the proteins and chemical vertices of a graph, all these interactions form a network. Biological networks are powerful tools for predicting undocumented relationships between molecules. The underlying principle is that existing interactions between molecules can be used to predict new interactions. For pairs of proteins sharing a common ligand, we use protein and chemical superimpositions combined with fast structural compatibility screens to predict whether additional compounds bound by one protein would bind the other. The current version includes data from the Protein Data Bank as of August 2011. The database is updated monthly.
Proper citation: ProtChemSI (RRID:SCR_006115) Copy
A curated repository of more than 206000 regulatory associations between transcription factors (TF) and target genes in Saccharomyces cerevisiae, based on more than 1300 bibliographic references. It also includes the description of 326 specific DNA binding sites shared among 113 characterized TFs. Further information about each Yeast gene has been extracted from the Saccharomyces Genome Database (SGD). For each gene the associated Gene Ontology (GO) terms and their hierarchy in GO was obtained from the GO consortium. Currently, YEASTRACT maintains a total of 7130 terms from GO. The nucleotide sequences of the promoter and coding regions for Yeast genes were obtained from Regulatory Sequence Analysis Tools (RSAT). All the information in YEASTRACT is updated regularly to match the latest data from SGD, GO consortium, RSA Tools and recent literature on yeast regulatory networks. YEASTRACT includes DISCOVERER, a set of tools that can be used to identify complex motifs found to be over-represented in the promoter regions of co-regulated genes. DISCOVERER is based on the MUSA algorithm. These algorithms take as input a list of genes and identify over-represented motifs, which can then be compared with transcription factor binding sites described in the YEASTRACT database.
Proper citation: Yeast Search for Transcriptional Regulators And Consensus Tracking (RRID:SCR_006076) Copy
High quality ribosomal RNA databases providing comprehensive, quality checked and regularly updated datasets of aligned small (16S/18S, SSU) and large subunit (23S/28S, LSU) ribosomal RNA (rRNA) sequences for all three domains of life (Bacteria, Archaea and Eukarya). Supplementary services include a rRNA gene aligner, online tools for probe and primer evaluation and optimized browsing, searching and downloading on the website. The extensively curated SILVA taxonomy and the new non-redundant SILVA datasets provide an ideal reference for high-throughput classification of data from next-generation sequencing approaches. Alignment tool, SINA, is available for download as well as available for use online.
Proper citation: SILVA (RRID:SCR_006423) Copy
http://prism.ccbb.ku.edu.tr/hotregion/index.php
Hot spots are energetically important residues at protein interfaces and they are not randomly distributed across the interface but rather clustered. These clustered hot spots form hot regions. Hot regions are important for the stability of protein complexes, as well as providing specificity to binding sites. HotRegion provides the hot region information of the interfaces by using predicted hot spot residues, and structural properties of these interface residues such as pair potentials of interface residues, accessible surface area (ASA) and relative ASA values of interface residues of both monomer and complex forms of proteins. Also, the 3D visualization of the interface and interactions among hot spot residues are provided. The number of interfaces in the database is 147909 and still growing.
Proper citation: HotRegion - A Database of Cooperative Hotspots (RRID:SCR_006022) Copy
ViralZone is a SIB Swiss Institute of Bioinformatics web-resource for all viral genus and families, providing general molecular and epidemiological information, along with virion and genome figures. Each virus or family page gives an easy access to UniProtKB/Swiss-Prot viral protein entries. ViralZone project is handled by the virus program of SwissProt group. Proteins popups were developed in collaboration with Prof. Christian von Mering and Andrea Franceschini, Bioinformatics Group , Institute of Molecular Life Sciences, University of Zurich, Winterthurerstrasse 190, CH-8057 Zurich, Switzerland, funded in part by the SIB Swiss Institute of bioinformatics. All pictures in ViralZone are copyright of the SIB Swiss Institute of Bioinformatics.
Proper citation: ViralZone (RRID:SCR_006563) Copy
A database of elecrophysiological properties text-mined from the biomedical literature as a function of neuron type. Specifically, NeuroElectro seeks to extract information about the electrophysiological properties (e.g. resting membrane potentials and membrane time constants) of diverse neuron types from the existing literature and place it into a centralized database. There are 252 neurons currently available, with the naming convention established in NeuroLex.
Proper citation: neuroelectro (RRID:SCR_006274) Copy
Scansite searches for motifs within proteins that are likely to be phosphorylated by specific protein kinases or bind to domains such as SH2 domains, 14-3-3 domains or PDZ domains. The Motifscanner program utilizes an entropy approach that assesses the probability of a site matching the motif using the selectivity values and sums the logs of the probability values for each amino acid in the candidate sequence. The program then indicates the percentile ranking of the candidate motif in respect to all potential motifs in proteins of a protein database. When available, percentile scores of some confirmed phosphorylation sites for the kinase of interests or confirmed binding sites of the domain of interest are provided for comparison with the scores of the candidate motifs.
Proper citation: Scansite (RRID:SCR_007026) Copy
http://www.iiserpune.ac.in/~coee/histome/
Database of human histone variants, sites of their post-translational modifications and various histone modifying enzymes. The database covers 5 types of histones, 8 types of their post-translational modifications and 13 classes of modifying enzymes. Many data fields are hyperlinked to other databases (e.g. UnprotKB/Swiss-Prot, HGNC, OMIM, Unigene etc.). Additionally, this database also provides sequences of promoter regions (-700 TSS +300) for all gene entries. These sequences were extracted from the UCSC genome browser. Sites of post-translational modifications of histones were manually searched from PubMed listed literature. Current version contains information for about ~50 histone proteins and ~150 histone modifying enzymes. HIstome is a combined effort of researchers from two institutions, Advanced Center for Treatment, Research and Education in Cancer (ACTREC), Navi Mumbai and Center of Excellence in Epigenetics (CoEE), Indian Institute of Science Education and Research (IISER), Pune.
Proper citation: HIstome: The Histone Infobase (RRID:SCR_006972) Copy
http://microkit.biocuckoo.org/
MiCroKit database is the first integrative resource to pin point most of identified components and related scientific information of midbody, centrosome and kinetochore. In this work, we have collected all proteins identified to be localized on kinetochore, centrosome, and/or midbody from two fungi (S. cerevisiae and S. pombe) and five animals, including C. elegans, D. melanogaster, X. laevis, M. musculus and H. sapiens. From the related literature of PubMed, numerous proteins have been manually curated to be localized on at least one of the sub-cellular localizations of kinetochore, centrosome and midbody. And to promise the quality of data, based on the rationale of Seeing is believing (Bloom K et al., 2005), these proteins have been unambiguously observed under fluorescent microscope as directly supportive evidences. Then an integrated and searchable database MiCroKit - Midbody, Centrosome and Kinetochore has been established. The version 1.0 of MiCroKit database was set up on Nov. 2nd, 2005, containing 1,065 unique proteins. The MiCroKit version 2.0 was released on Jun. 5th, 2006, with 1,120 entries. Currently, the MiCroKit 3.0 database was updated on July 9, 2009, containing 1,489 unique protein entries. The online service of MiCroKit 3.0 was implemented in PHP + MySQL + JavaScript. And the local packages of MiCroKit 3.0 were developed in JAVA 1.5 (J2SE). The database will be updated routinely as new microkit proteins are reported.
Proper citation: Midbody, Centrosome and Kinetochore (RRID:SCR_007052) Copy
BioCarta Pathways allows users to observe how genes interact in dynamic graphical models. Online maps available within this resource depict molecular relationships from areas of active research. In an open source approach, this community-fed forum constantly integrates emerging proteomic information from the scientific community. It also catalogs and summarizes important resources providing information for over 120,000 genes from multiple species. Find both classical pathways as well as current suggestions for new pathways.
Proper citation: BioCarta Pathways (RRID:SCR_006917) Copy
http://www.imgt.org/IMGTindex/LIGM.html
IMGT/LIGM-DB is a comprehensive database of immunoglobulin (IG) and T cell receptor (TR) nucleotide sequences from human and other vertebrate species (270). IMGT/LIGM-DB includes all germline (non-rearranged) and rearranged IG and TR genomic DNA (gDNA) and complementary DNA (cDNA) sequences published in generalist databases. IMGT/LIGM-DB allows searches from the Web interface according to biological and immunogenetic criteria through five distinct modules depending on the user interest. Users can search the catalogue by accession number, mnemonic, definition, creation date, length, or annotation level. They also have the option to search through taxonomic classification, keywords, and annotated labels. For a given entry, nine types of display are available including the IMGT flat file, the translation of the coding regions and the analysis by the IMGT/V-QUEST tool (see parent org. below). IMGT/LIGM-DB distributes expertly annotated sequences. The annotations hugely enhance the quality and the accuracy of the distributed detailed information. They include the sequence identification, the gene and allele classification, the constitutive and specific motif description, the codon and amino acid numbering, and the sequence obtaining information, according to the main concepts of IMGT-ONTOLOGY. They represent the main source of IG and TR gene and allele knowledge stored in IMGT/GENE-DB and in the IMGT reference directory., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: IMGT/LIGM-DB (RRID:SCR_006931) Copy
Database that collects all arabidopsis transcription factors (totally 1922 Loci; 2290 Gene Models) and classifies them into 64 families. It uses not only locus (gene), but also gene model (transcript, protein) and the detail information is for each gene model not for locus. It adds multiple alignment of the DNA-binding domain of each family, Neighbor-Joining phylogenetic tree of each family, the GO annotation, homolog with the Database of Rice Transcription Factors (DRTF). It also keeps old information items such as the unique cloned and sequenced information of about 1200 transcription factors, protein domains, 3D structure information with BLAST hits against PDB, predicted Nuclear Location Signals, UniGene information, as well as links to literature reference.
Proper citation: Database of Arabidopsis Transcription Factors (RRID:SCR_007101) Copy
http://www.polygenicpathways.co.uk
Database of disease genes and risk factors and of host pathogen/interactomes. Lists genes, pathways and environmental risk factors positively associated with diseases and conditions such as Alzheimer's disease, schizophrenia, multiple sclerosis, childhood obesity, anorexia nervosa, HIV-1/AIDS, and helicobacter pylori. Details of polymorphisms as well as negative/positive association data can be found via Useful links. Throughout the site are links to Entrez Gene and Pubmed.
Proper citation: Polygenic Pathways (RRID:SCR_006962) Copy
http://www.broadinstitute.org/mammals/haploreg/haploreg.php
HaploReg is a tool for exploring annotations of the noncoding genome at variants on haplotype blocks, such as candidate regulatory SNPs at disease-associated loci. Using linkage disequilibrium (LD) information from the 1000 Genomes Project, linked SNPs and small indels can be visualized along with their predicted chromatin state in nine cell types, conservation across mammals, and their effect on regulatory motifs. HaploReg is designed for researchers developing mechanistic hypotheses of the impact of non-coding variants on clinical phenotypes and normal variation.
Proper citation: HaploReg (RRID:SCR_006796) Copy
A free, simple to use web service dedicated to reconstructing and analysing phylogenetic relationships between molecular sequences. Phylogeny.fr runs and connects various bioinformatics programs to reconstruct a robust phylogenetic tree from a set of sequences.
Proper citation: Phylogeny.fr (RRID:SCR_010266) Copy
http://diana.imis.athena-innovation.gr/DianaTools/index.php?r=lncBase/index
Database that hosts elaborated information for both predicted and experimentally verified, miRNA-lncRNA interactions. The database consists of two distinct modules. The Experimental Module contains detailed information for more than 5,000 interactions, between 2,958 lncRNAs and 120 miRNAs, ranging from miRNA and lncRNA related facts to information specific to their interaction, the experimental validation methodologies and their outcomes. The Prediction Module, which is based on the latest version of DIANA-microT target prediction algorithm (DIANA-microT-CDS), contains detailed information for more than 10 million interactions, between 56,097 lncRNAs and 3,078 miRNAs, ranging from miRNA and lncRNA related details to specific information regarding their interaction sites, graphical representation of their binding and the predicted score. This module exhibits a unique feature for searching the database. Users are able to add genomic locations to their queries thus browsing every miRNA-lncRNA interaction that has at least one MRE located inside the queried locus.
Proper citation: DIANA-LncBase (RRID:SCR_010840) Copy
GDR is a curated and integrated web-based relational database. GDR contains comprehensive data of the genetically anchored peach physical map, annotated EST databases of apple, peach, almond, cherry, rose, raspberry and strawberry, Rosaceae maps and markers and all publicly available Rosaceae sequences. Annotations of ESTs include contig assembly, putative function, simple sequence repeats, ORFs, Gene Ontology and anchored position to the peach physical map where applicable. Our integrated map viewer provides graphical interface to the genetic, transcriptome and physical mapping information. We continue to add Rosaceae map data to CMap, a web-based tool that allows users to view comparisons of genetic and physical maps. ESTs, BACs and markers can be queried by various categories and the search result sites are linked to the integrated map viewer or to the WebFPC physical map sites. In addition to browsing and querying the database, users can compare their sequences with the annotated GDR sequences via a dedicated sequence similarity server running either the BLAST or FASTA algorithm, search their sequences for microsatellites using the SSR server or assemble their ESTs using the CAP3 Server.
Proper citation: Genome Database for Rosaceae (RRID:SCR_012756) Copy
http://www.cleanex.isb-sib.ch/
CleanEx is a database which provides access to public gene expression data via unique approved gene symbols and which represents heterogeneous expression data produced by different technologies in a way that facilitates joint analysis and cross-dataset comparisons. To achieve this goal, each single gene expression experiment is regularly mapped on a permanent target identifier consisting of a physical description of the targeted RNA. There is one entry per gene. To have a complete view of the transcript and its product, we also link each entry to the corresponding protein. We further provide the genomic position of the transcription start site from EPD, when available. Otherwise we give the annotated start site position in Ensembl.
Proper citation: CleanEx (RRID:SCR_012911) Copy
MBGD is a database for comparative analysis of completely sequenced microbial genomes, the number of which is now growing rapidly. The aim of MBGD is to facilitate comparative genomics from various points of view such as ortholog identification, paralog clustering, motif analysis and gene order comparison. The heart of MBGD function is to create orthologous or homologous gene cluster table. For this purpose, similarities between all genes are precomputed and stored into the database, in addition to the annotations of genes such as function categories that were assigned by the original authors and motifs that were found in the translated sequence. Using these homology data, MBGD dynamically creates orthologous gene cluster table. Users can change a set of organisms or cutoff parameters to create their own orthologous grouping. Based on this cluster table, users can further analyze multiple genomes from various points of view with the functions such as global map comparison, local map comparison, multiple sequence alignment and phylogenetic tree construction.
Proper citation: MBGD - Microbial Genome Database (RRID:SCR_012824) Copy
http://www.informatics.jax.org/
Community model organism database for laboratory mouse and authoritative source for phenotype and functional annotations of mouse genes. MGD includes complete catalog of mouse genes and genome features with integrated access to genetic, genomic and phenotypic information, all serving to further the use of the mouse as a model system for studying human biology and disease. MGD is a major component of the Mouse Genome Informatics.Contains standardized descriptions of mouse phenotypes, associations between mouse models and human genetic diseases, extensive integration of DNA and protein sequence data, normalized representation of genome and genome variant information. Data are obtained and integrated via manual curation of the biomedical literature, direct contributions from individual investigators and downloads from major informatics resource centers. MGD collaborates with the bioinformatics community on the development and use of biomedical ontologies such as the Gene Ontology (GO) and the Mammalian Phenotype (MP) Ontology.
Proper citation: Mouse Genome Database (RRID:SCR_012953) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.