Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
A database of general chemical information. The datasets are comprised of various available chemical datasets annotated with interesting properties to train and test machine-learning prediction and searching methods. Tools provided include ChemicalSearch, Virtual Chemical Space, Reaction Explorer, Datasets, and supplemental material. ChemicalSearch is a tool that allows users to find a chemical by basic criteria like molecular weight and predicted logP, or by the more abstract notion of structural similarity. Virtual Chemical Space is a tool which lets users interactively deconstruct target compounds into component precursors and reconstruct similar building-blocks into combinatorial libraries representing the virtual chemical space near the target compound. Reaction Explorer is a synthesis explorer and mechanism explorer. It provides an interactive system for learning and practicing reactions, syntheses and mechanisms in organic chemistry, with advanced support for the automatic generation of random problems, curved-arrow mechanism diagrams, and inquiry-based learning.
Proper citation: ChemDB: The UC Irvine ChemDB (RRID:SCR_007594) Copy
http://chicken.genomics.org.cn
ChickVD hosts high-quality sequence variation data, variation analysis in the context of chicken genes, cDNAs, chicken orthologs of human disease genes, genetic markers, quantitative trait loci (QTLs) etc . All data are uniquely mapped onto the RJF draft genome and graphically represented in MapView, an efficient visualization tool that allows users to browse sequence variations in the genomic and functional context. The sub-viewer TraceView assists users to view the vivid graphics of the original traces around the detected SNP. Users may query the data by the online search tool and define concrete limitations to extract records that are best suited to their research needs. For the convenience of data presentation in ChickVD, different types of sequence variations (substitutions, insertions or deletions) are all referred as ???SNPs''. ChickVD is updated constantly as more data generated and is under the continued improvement for its content and functionality
Proper citation: Chicken Variation Database (RRID:SCR_007595) Copy
http://defensins.bii.a-star.edu.sg/
The defensins knowledgebase is a manually curated database and information source devoted to the defensin family of antimicrobial peptides. The current version of the database holds a comprehensive collection of 363 defensin records each containing sequence, structure and activity information. A web-based interface provides access to the information and allows for text-based searching on the data fields. With the rapidly increasing interest in defensins, we hope that the knowledgebase will prove to be a valuable resource in the field of antimicrobial peptide research.
Proper citation: Defensins Knowledgebase (RRID:SCR_007623) Copy
A database to provide cleansed EST sequences of classified dbEST libraries. All dbEST libraries were classified according to organism, sequencing center, and eVOC ontologies (for human libraries). For each dbEST library, we provide three different EST sequences: raw, pre-cleansed, and user-cleansed. pre-cleansed ESTs are obtained from major contamination databases and cleaned of contaminated sequences. User-cleansed ESTs, however, involve the use of an automatic user-cleansing pipeline, in which sequences in a user-selected library are cleansed on-the-fly according to user-input options. CleanEST contains 62,008,259 EST sequences (24,000 libraries) with contamination information.
Proper citation: Cleansed EST Database (RRID:SCR_007587) Copy
DBTGR provides information on tunicate gene regulation, such as the location of expression, or the identified regulatory elements present in promoter sequences. The database also contains the promoters of homologous genes in multiple species to allow identification of conserved cis elements.
Proper citation: DataBase of Tunicate Gene Regulation (RRID:SCR_007620) Copy
:DDOC provides a comprehensive compilation of the published research related to the genes associated with ovarian cancer. DDOC provides details of the cell line, tissue or cell type, expression status, disease stage, tumor grade, OC type and laboratory method provided in the literature. The links to the relevant sources of data used to extract information related to genes are also included. Many aspects of the information provided in the DDOC were curated by biologists, which increases its accuracy. DDOC is freely accessible for academic and non-profit users.
Proper citation: Dragon Database for Exploration of Ovarian Cancer Genes (RRID:SCR_007621) Copy
CATH is a hierarchical classification of protein domain structures, which clusters proteins at four major levels: Class (C), Architecture (A), Topology (T) and Homologous superfamily (H). The boundaries and assignments for each protein domain are determined using a combination of automated and manual procedures which include computational techniques, empirical and statistical evidence, literature review and expert analysis Users can search CATH by ID/Sequence/text. They can also browse CATH from the top of the hierarchy, or download CATH data.
Proper citation: CATH: Protein Structure Classification (RRID:SCR_007583) Copy
http://source.rcsb.org/jfatcatserver/ceHome.jsp
CE is a databases of alignments for all polypeptide chains. A representative set of proteins is available and kept current with the PDB, a method for calculating pairwise structure alignments. CE aligns two polypeptide chains using characteristics of their local geometry as defined by vectors between C alpha positions. Matches are termed aligned fragment pairs (AFPs). Heuristics are used in defining a set of optimal paths joining AFPs with gaps as needed. The path with the best RMSD is subject to dynamic programming to achieve an optimal alignment. For specific families of proteins additional characteristics are used to weight the alignment. Complete details are described in the paper (PDF format). Databases of alignments for all polypeptide chains and a representative set of proteins is available and kept current with the PDB
Proper citation: Combinatorial Extension (CE) (RRID:SCR_007585) Copy
CASRdb is a calcium-sensing receptor locus-specific database for mutations causing familial (benign) hypocalciuric hypercalcemia, neonatal severe hyperparathyroidism, and autosomal dominant hypocalcemia. The information can be searched by mutation, genotype-phenotype, clinical data, in vitro analyses, and authors of publications describing the mutations. CASRdb is regularly updated for new mutations and it also provides a mutation submission form to ensure up-to-date information. The home page of this database provides links to different web pages that are relevant to the CASR, as well as disease clinical pages, sequence of the CASR gene exons, and position of mutations in the CASR. The CASRdb will help researchers to better understand and analyze the mutations, and aid in structure-function analyses.
Proper citation: CASRDB- Calcium Sensing Receptor Database (RRID:SCR_007581) Copy
dbPTM is a database that compiles information on protein post-translational modifications (PTM) such as the modified sites, solvent accessibility of surrounding amino acids, protein secondary and tertiary structures, protein domains, and protein variations. The version 2.0 of dbPTM integrates the experimentally validated PTM sites with referable literatures from Swiss-Prot, Phospho.ELM, O-GLYCBASE, and UbiProt. In all of the collected PTM information, about 25 types of PTM with enough experimentally validated sites are trained the profile hidden Markov models (HMMs) to detect the potential PTM sites with 100% specificity against Swiss-Prot proteins. To help users investigating more detail in each type of PTM, the substrate peptide specificity such as positional amino acid frequency, solvent accessibility and secondary structure surrounding the modified sites are also provided. Moreover, the information of orthologous protein clusters is provided to users for analyzing whether the PTM sites located in the evolutionary conserved regions or not., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: dbPTM: An informational repository of proteins and post-translational modifications (RRID:SCR_007619) Copy
http://genecards.weizmann.ac.il/genenote/
THIS RESOURCE IS NO LONGER IN SERVICE, documented June 14, 2013. GeneNote is a database of human genes and their expression profiles in healthy tissues. It is based on Weizmann Institute of Science DNA array experiments, which were performed on the Affymetrix HG-U95 set A-E. It offers: An expression profile (tissue vector) for each gene in the human genome Gene and tissue clustering based on expression profiles A full genome ranking procedure according to the gene''s tendency for tissue specificity, from tissue-specific to housekeeping genes.
Proper citation: GeneNote (RRID:SCR_007679) Copy
http://research.nhgri.nih.gov/histones/
Histone Database is a database of histones and their corresponding sequences. Sequence- and text-based searches were performed on NCBI's redundant and non-redundant (nr) peptide sequence databases. These databases are derived from GenBank, EMBL, and DDBJ translated DNA coding regions, plus protein sequences from the PDB (Protein Data Bank), SWISS-PROT, the PIR (Protein Information Resource), and the PRF (Protein Research Foundation). :Users can search by keyword, sequence fragment, category, organism, and redundancy of the set.
Proper citation: Histone Database (RRID:SCR_007711) Copy
http://www.biocheminfo.org/klotho/
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 16, 2013. A database of biochemical compound information. All files are available for download, and all entries are cataloged by accession number. Klotho is part of a larger attempt to model biological processes, beginning with biochemistry.
Proper citation: Klotho: Biochemical Compounds Declarative Database (RRID:SCR_007714) Copy
http://urgi.versailles.inra.fr/Genefarm/
GeneFarm is a database of structural and functional annotation of plant gene and protein families. The goal of the GeneFarm project is to obtain homogeneous, reliable, documented and traceable annotations for plant nuclear genes and gene products and to enter them into added-value database. The improved annotation will allow better data mining of the plant genomes (mainly Arabidopsis thaliana), and more secure planning and design of experiments. It is also a necessary step for building knowledge management tools for integrating plant genomic data, either for plant breeding or to get a broader interactive view of plant biological processes, like gene interaction networks. This re-annotation project, launched is mainly focused on gene families. A complete annotation pipeline using the most efficient prediction tools has been defined. The involved partners, each contributing with genes from his/her field of expertise, have exhaustively annotated families of homologous genes. A database named GeneFarm (Gene Families for Arabidopsis Management) gathers all these expert-curated annotations of plant gene families. Furthermore, collaboration with the Swiss Institute of Bioinformatics is underway to integrate the GeneFarm data into the protein knowledgebase Swiss-Prot.
Proper citation: GeneFarm (RRID:SCR_007674) Copy
http://caps.ncbs.res.in/gendis/home.html
Genomic Distribution of structural Superfamilies identifies and classifies evolutionary related proteins at the superfamily level in whole genome databases. GenDiS has been curated in direct correspondence with SCOP and represents 4001 highly resolved domains in 1194 structural superfamilies across protein sequence databases. Sequences showing reliable homology to entries in SCOP and PASS2 databases have been obtained from the non-redundant protein sequence database and aligned. Similar alignments of the superfamily members are provided in the genome level. GenDiS provides a platform for cross genome comparison at the superfamily level. GenDis relates proteins sequence information across all strata of taxonomy. One may navigate through the database to obtain structural homologues across different levels in taxonomic classification. The nomenclature of the various genomes and their hierarchy is in direct correspondence with the taxonomy database maintained at the NCBI. Sequence homologues for the various structural members are obtained from the non-redundant protein sequence database employing sensitive sequence search methods. Multiple approaches such as PSI-BLAST, HMMsearch of the HMMer suite and an interacting motif constrained PHI-BLAST have been employed to identify homologues in the sequence databases.
Proper citation: Genomic Distribution of structural Superfamilies (RRID:SCR_007670) Copy
http://genecards.weizmann.ac.il/geneannot/
GeneAnnot provides a revised and improved annotation of Affymetrix probe-sets from HG-U95, HG-U133 and HG-U133 Plus2.0. Probe-sets are related to GeneCards genes, by direct sequence comparison of probes to GenBank, RefSeq and Ensembl mRNA sequences, while assigning sensitivity and specificity scores to each probe-set to gene match. Where such matches are not found, probe-sets are annotated by their relation to GenBank mRNA sequences and UniGene clusters. The results are integrated with the GeneCards, GeneLoc and GeneNote databases. HG-U95, HG-U133, HG-U133
Proper citation: GeneAnnot (RRID:SCR_007673) Copy
http://www.hepseq.org/Public/Web_Front/main.php
HepSEQ is the International Repository for Hepatitis B Virus Strain Data. It is web-accessible, quality-based, molecular, clinical and epidemiological database for hepatitis B infection and provides a tool for the research community or for those involved in hepatitis B case management. This database currently has 1012 patient records and 1253 viral sequences. The quality of all submitted sequences is checked. The tools provided include: SeqMatch: search the database for matching sequences Genotyper: genotype HBV strains (based on HBV surface antigen genes) Gene Mutation: display the sequences that contain mutations in HBV coding regions Mutation Annotator: annotate sequences for mutation known to be associated with anti-viral resistance This web database development is funded by the UK Department of Health is curated and is hosted by the Health Protection Agency.
Proper citation: Hepatitis Virus B Database (RRID:SCR_007705) Copy
https://omictools.com/heg-db-tool
Genomic database that includes prediction of which genes are highly expressed in prokaryotic complete genomes under strong translational selection.
Proper citation: Highly Expressed Genes Database (HEG-DB) (RRID:SCR_007704) Copy
http://www.compbio.dundee.ac.uk/kinomer
Kinomer is a multilevel HMM library that models these protein kinase groups. It allows accurate identification of protein kinases and classification to the appropriate kinase group. Profile hidden Markov models (HMMs) are statistical descriptions of sequence conservation from multiple sequence alignments, and have been shown to outperform standard pairwise sequence comparison methods, both in terms of sensitivity and specificity. HMMs form the basis of protein family and domain description libraries such as SUPERFAMILY and Pfam.
Proper citation: Kinomer (RRID:SCR_007707) Copy
GENATLAS contains relevant information with respect to gene mapping and genetic diseases. GENATLAS compiles the information relevant to the mapping efforts of the Human Genome Project. This information is collected from more than 48,000 articles in the literature, collected in more than 870 reviews. The articles are daily analyzed by annotators to update the GENATLAS database. Only the objects with a known cytogenetic location are retained. GENATLAS repertories three kinds of objects Genes database ( more than 21.000 entries) Phenotypes database ( 4104 entries , 2000 cloned) References database linked to the two previous ( more than 48000 entries)
Proper citation: GenAtlas (RRID:SCR_007669) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.