Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
A curated repository of more than 206000 regulatory associations between transcription factors (TF) and target genes in Saccharomyces cerevisiae, based on more than 1300 bibliographic references. It also includes the description of 326 specific DNA binding sites shared among 113 characterized TFs. Further information about each Yeast gene has been extracted from the Saccharomyces Genome Database (SGD). For each gene the associated Gene Ontology (GO) terms and their hierarchy in GO was obtained from the GO consortium. Currently, YEASTRACT maintains a total of 7130 terms from GO. The nucleotide sequences of the promoter and coding regions for Yeast genes were obtained from Regulatory Sequence Analysis Tools (RSAT). All the information in YEASTRACT is updated regularly to match the latest data from SGD, GO consortium, RSA Tools and recent literature on yeast regulatory networks. YEASTRACT includes DISCOVERER, a set of tools that can be used to identify complex motifs found to be over-represented in the promoter regions of co-regulated genes. DISCOVERER is based on the MUSA algorithm. These algorithms take as input a list of genes and identify over-represented motifs, which can then be compared with transcription factor binding sites described in the YEASTRACT database.
Proper citation: Yeast Search for Transcriptional Regulators And Consensus Tracking (RRID:SCR_006076) Copy
Database providing integrated access to genome sequence, expression data and literature curation for Tuberculosis (TB) that houses genome assemblies for numerous strains of Mycobacterium tuberculosis (MTB) as well assemblies for over 20 strains related to MTB and useful for comparative analysis. TBDB stores pre- and post-publication gene-expression data from M. tuberculosis and its close relatives, including over 3000 MTB microarrays, 95 RT-PCR datasets, 2700 microarrays for human and mouse TB related experiments, and 260 arrays for Streptomyces coelicolor. (July 2010) To enable wide use of these data, TBDB provides a suite of tools for searching, browsing, analyzing, and downloading the data.
Proper citation: Tuberculosis Database (RRID:SCR_006619) Copy
ViralZone is a SIB Swiss Institute of Bioinformatics web-resource for all viral genus and families, providing general molecular and epidemiological information, along with virion and genome figures. Each virus or family page gives an easy access to UniProtKB/Swiss-Prot viral protein entries. ViralZone project is handled by the virus program of SwissProt group. Proteins popups were developed in collaboration with Prof. Christian von Mering and Andrea Franceschini, Bioinformatics Group , Institute of Molecular Life Sciences, University of Zurich, Winterthurerstrasse 190, CH-8057 Zurich, Switzerland, funded in part by the SIB Swiss Institute of bioinformatics. All pictures in ViralZone are copyright of the SIB Swiss Institute of Bioinformatics.
Proper citation: ViralZone (RRID:SCR_006563) Copy
A versatile web-server application for the analysis and visualization of array-CGH data.
Proper citation: waviCGH (RRID:SCR_006662) Copy
http://probeexplorer.cicancer.org/principal.php
Probe Explorer is an open access web-based bioinformatics application designed to show the association between microarray oligonucleotide probes and transcripts in the genomic context, but flexible enough to serve as a simplified genome and transcriptome browser. Coordinates and sequences of the genomic entities (loci, exons, transcripts), including vector graphics outputs, are provided for fifteen metazoa organisms and two yeasts. Alignment tools are used to built the associations between Affymetrix microarrays probe sequences and the transcriptomes (for human, mouse, rat and yeasts). Search by keywords is available and user searches and alignments on the genomes can also be done using any DNA or protein sequence query. Platform: Online tool
Proper citation: ProbeExplorer (RRID:SCR_007116) Copy
http://www.broadinstitute.org/annotation/tetraodon/
This database have been funded by the National Human Genome Research Institute (NHGRI) to produce shotgun sequence of the Tetraodon nigriviridis genome. The strategy involves Whole Genome Shotgun (WGS) sequencing, in which sequence from the entire genome is generated. Whole genome shotgun libraries were prepared from Tetraodon genomic DNA obtained from the laboratory of Jean Weissenbach at Genoscope. Additional sequence data of approximately 2.5X coverage of Tetraodon has also been generated by Genoscope in plasmid and BAC end reads. Broad and Genoscope intend to pool their data and generate whole genome assemblies. Tetraodon nigroviridis is a freshwater pufferfish of the order Tetraodontiformes and lives in the rivers and estuaries of Indonesia, Malaysia and India. This species is 20-30 million years distant from Fugu rubripes, a marine pufferfish from the same family. The gene repertoire of T. nigroviridis is very similar to that of other vertebrates. However, its relatively small genome of 385 Mb is eight times more compact than that of human, mostly because intergenic and intronic sequences are reduced in size compared to other vertebrate genomes. These genome characteristics along with the large evolutionary distance between bony fish and mammals make Tetraodon a compact vertebrate reference genome - a powerful tool for comparative genetics and for quick and reliable identification of human genes.
Proper citation: Tetraodon nigroviridis Database (RRID:SCR_007123) Copy
http://www.sugp.caltech.edu/SpBase/
SpBase is designed to present the results of the genome sequencing project for the purple sea urchin. The sequences and annotations emerging from this effort are organized in a database that provides the research community access to those data not normally presented through National Center for Biotechnology Information and other large databases. Additionally, the unique information on that links gene identities and sequences to the plate and well location to the library filters from the Sea Urchin genome Resource will also be presented. The software used to organize and present the sea urchin genome comes from GMOD, a collection of open source software tools for creating and managing genome-scale biological databases. That sea urchins eggs and embryos have long remained a popular research subject for cell and developmental biologists is one rationale for sequencing the genome. In addition, studies of embryonic development in the California Purple Sea Urchin, Strongylocentrotus purpuratus , have paralleled the emergence of molecular techniques ranging from the characterization of genomic repeat sequences in the 1970''s to the elucidation of gene regulatory networks in recent times. The parent of this site, SUGP, was meant to provide a focal point for the exchange of genomic information as the genome of the Purple sea urchin was being sequenced. Over these past years it has served as a repository for small sequencing projects and a source of sequence information useful for gene discovery projects. Here one could find information on macro-array libraries of cDNAs from the purple sea urchin and genomic DNA from several species. In addition, a Sequence Tag Connector (STC) collection has been assembled from 5% of the genome sequence and a very extensive repeat sequence catalog prepared. All of the sequence data that we maintained at SUGP was incorporated into the new SPBase. Of course, it is all in public sequence databases such as the National Center for Biological Information as well. Some additional sequence information is available at the Resource Center of the German Human Genome Project. With the publication of The Genome of the Sea Urchin Strongylocentrotus purpuratus by The Sea Urchin Genome Sequencing Consortium a link to the first 9941 gene annotations are now publicly available. The effort to sequence the whole purple sea urchin genome was a cooperative one that included contributions from the Sea Urchin Genome Facility here at the Center for Computational Regulatory Genomics, Beckman Institute, Caltech, and support from the Human Genome Research Institute of the National Institutes of Health. The sequencing was done at the Baylor College of Medicine, Human Genome Sequencing Center, Houston, Texas. Funding was approved based on an initiative submitted by the Sea Urchin Genome Advisory Committee.
Proper citation: SpBase - Strongylocentrotus purpuratus: the Sea Urchin Genome Database (RRID:SCR_007441) Copy
A database of human, chimpanzee, mouse, and rat proteases and protease inhibitors, as well as as the growing number of hereditary diseases caused by mutations in protease genes. Analysis of the human and mouse genomes has allowed us to annotate 581 human, 580 chimpanzee, 667 mouse, and 655 rat protease genes. Proteases are classified in five different classes according to their mechanism of catalysis. Proteases are a diverse and important group of enzymes representing >2% of the human, chimpanzee, mouse and rat genomes. This group of enzymes is implicated in numerous physiological processes. The importance of proteases is illustrated by the existence of 99 different hereditary diseases due to mutations in protease genes. Furthermore, proteases have been implicated in multiple human pathologies, including vascular diseases, rheumatoid arthritis, neurodegenerative processes, and cancer. During the last ten years, our laboratory has identified and characterized more than 60 human protease genes. Due to the importance of proteolytic enzymes in human physiology and pathology, we have recently introduced the concept of Degradome, as the complete repertoire of proteases expressed by a tissue or organism. Thanks to the recent completion of the human, chimpanzee, mouse, and rat genome sequencing projects, we were able to analyze and compare for the first time the complete protease repertoire in those mammalian organisms, as well as the complement of protease inhibitor genes. This webpage also contains the Supplementary Material of Human and mouse proteases: a comparative genomic approach Nat Rev Genet (2003) 4: 544-558, Genome sequence of the brown Norway rat yields insights into mammalian evolution Nature (2004) 428: 493-521, A genomic analysis of rat proteases and protease inhibitors Genome Res. (2004) 14: 609-622, and Comparative genomic analysis of human and chimpanzee proteases Genomics (2005) 86: 638-647.
Proper citation: Mammalian Degradome Database (RRID:SCR_007624) Copy
http://www.ncbi.nlm.nih.gov/genomes/GenomesHome.cgi?taxid=2759&hopt=html
Curated sequence data and related information on organelles from NCBI Refseq for the community to use as a standard. The animal mitochondrial records are considered reviewed; that is, they have been manually curated by the NCBI staff. Other mitochondrial and chloroplast genome records are provisional and are presented with varying levels of review compared to the primary record used to build the RefSeq. Additionally, protein clusters for the metazoan and plastid genomes proteins can be reviewed with Entrez Protein Clusters.
Proper citation: Organelle Genome Resources (RRID:SCR_007838) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on May 12,2023. Database of expression patterns of C. elegans promoter::GFP constructs. A text description of the observed pattern is provided, indicating the stage(s) and tissue(s) in which GFP is expressed. Also available for some strains are the corresponding 2D and 3D images. Investigators may browse the entire list, search by gene name, tissue, stage, and pattern. Search results may be downloaded in .csv and .txt formats. All of the strains in the expression pattern database are displayed in the browse page. The records are organized by gene; information such as locus name, genomic location (WormBase), the presence of images and videos, and the actual expression pattern are shown in a tabular format.
Proper citation: Expression Patterns for C. elegans promoter GFP fusions (RRID:SCR_001619) Copy
http://hanalyzer.sourceforge.net/
An open-source data integration system designed to assist biologists in explaining the results observed in genome-scale experiments as well as generating new hypotheses. It combines information extraction techniques, semantic data integration, and reasoning and facilitates network visualization. The Hanalyzer source code and binaries are available for download.
Proper citation: Hanalyzer (RRID:SCR_000923) Copy
Database of human genes that provides concise genomic, proteomic, transcriptomic, genetic and functional information on all known and predicted human genes. Information featured in GeneCards includes orthologies, disease relationships, mutations and SNPs, gene expression, gene function, pathways, protein-protein interactions, related drugs and compounds and direct links to cutting edge research reagents and tools such as antibodies, recombinant proteins, clones, expression assays and RNAi reagents.
Proper citation: GeneCards (RRID:SCR_002773) Copy
Model organism database that serves as central repository and web-based resource for zebrafish genetic, genomic, phenotypic and developmental data. Data represented are derived from three primary sources: curation of zebrafish publications, individual research laboratories and collaborations with bioinformatics organizations. Data formats include text, images and graphical representations.Serves as primary community database resource for laboratory use of zebrafish. Developed and supports integrated zebrafish genetic, genomic, developmental and physiological information and link this information extensively to corresponding data in other model organism and human databases.
Proper citation: Zebrafish Information Network (ZFIN) (RRID:SCR_002560) Copy
ooTFD (object-oriented Transcription Factors Database) is a successor to TFD, the original Transcription Factors Database. This database is aimed at capturing information regarding the polypeptide interactions which comprise and define the properties of transcription factors. ooTFD contains information about transcription factor binding sites, as well as composite relationships within transcription factors, which frequently occur as multisubunit proteins that form a complex interface to cellular processes outside the transcription machinery through protein-protein interactions. ooTFD contains information represented in TFD but also allows the representation of containment, composite, and interaction relationships between transcription factor polypeptides. It is designed to represent information about all transcription factors, both eukaryotic and prokaryotic, basal as well as regulatory factors, and multiprotein complexes as well as monomers.
Proper citation: object-oriented Transcription Factors Database (RRID:SCR_002435) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 23, 2016. ELISA is an online database that combines functional annotation with structure and sequence homology modeling to place proteins into sequence-structure-function neighborhoods. The atomic unit of the database is a set of sequences and structural templates that those sequences encode. A graph that is built from the structural comparison of these templates is called PDUG (protein domain universe graph). It introduces a method of functional inference through a probabilistic calculation done on an arbitrary set of PDUG nodes. Further, all PDUG structures are mapped onto all fully sequenced proteomes allowing an easy interface for evolutionary analysis and research into comparative proteomics. ELISA is the first database with applicability to evolutionary structural genomics explicitly in mind.
Proper citation: Evolutionary Lineage Inferred from Structural Analysis (RRID:SCR_002343) Copy
DoTS (Database Of Transcribed Sequences) is a human and mouse transcript index created from all publicly available transcript sequences. The input sequences are clustered and assembled to form the DoTS Consensus Transcripts that comprise the index. These transcripts are assigned stable identifiers of the form DT.123456 (and are often referred to as dots). The transcripts are in turn clustered to form putative DoTS Genes. These are assigned stable identifiers of the form DG.1234356. As of September 1, 2004, the DoTS annotation team has manually annotated 43,164 human and 78,054 mouse DoTS Transcripts (DTs), corresponding to 3,939 human and 7,752 mouse DoTS Genes (DGs). Use the manually annotated gene query to see the DoTS Transcripts that have been manually annotated. The focus of the DoTS project is integrating the various types of data (e.g., EST sequences, genomic sequence, expression data, functional annotation) in a structured manner which facilitates sophisticated queries that are otherwise not easy to perform. DoTS is built on the GUS Platform which includes a relational database that uses controlled vocabularies and ontologies to ensure that biologically meaningful queries can be posed in a uniform fashion. An easy way to start using the site is to search for DoTS Transcripts using an existing cDNA or mRNA sequence. Click on the BLAST tab at the top of the page and enter your sequence in the form provided. All the transcripts with significant sequence similarity to your query sequence will be displayed. Or use one of the provided queries to retrieve transcripts using a number of criteria. These queries are listed on the query page, which can also be reached by clicking on the tab marked query at the top of the page. Finally, the boolean query page allows these queries to be combined in a variety of ways. Sponsors: Funding provided by -NIH grant RO1-HG-01539-03 -DOE grant DE-FG02-00ER62893
Proper citation: Database of Transcribed Sequences (RRID:SCR_002334) Copy
Alternative splicing essentially increases the diversity of the transcriptome and has important implications for physiology, development and the genesis of diseases. This resource uses a different approach to investigate alternative splicing (instead of the conventional case-by case fashion) and integrates all transcripts derived from a gene into a single splicing graph. ASG is a database of splicing graphs for human genes, using transcript information from various major sources (Ensembl, RefSeq, STACK, TIGR and UniGene). Each transcript corresponds to a path in the graph, and alternative splicing is displayed by bifurcations. This representation preserves the relationships between different splicing variants and allows us to investigate systematically all possible putative transcripts. Web interface allows users to display the splicing graphs, to interactively assemble transcripts and to access their sequences as well as neighboring genomic regions. ASG also provide for each gene, an exhaustive pre-computed catalog of putative transcriptsin total more than 1.2 million sequences. It has found that ~65 of the investigated genes show evidence for alternative splicing, and in 5 of the cases, a single gene might produce over 100 transcripts.
Proper citation: Alternate splicing gallery (RRID:SCR_008129) Copy
https://wiki.cgb.indiana.edu/display/DGC/Home
The Daphnia Genomics Consortium (DGC) is an international network of investigators committed to mounting the freshwater crustacean Daphnia as a model system for ecology, evolution and the environmental sciences. Along with research activities, the DGC is: (1) coordinating efforts towards developing the Daphnia genomic toolbox, which will then be available for use by the general community; (2) facilitating collaborative cross-disciplinary investigations; (3) developing bioinformatic strategies for organizing the rapidly growing genome database; and (4) exploring emerging technologies to improve high throughput analyses of molecular and ecological samples. If we are to succeed in creating a new model system for modern life-sciences research, it will need to be a community-wide effort. Research activities of the DGC are primarily focused on creating genomic tools and information. When completed, the current projects will offer a first view of the Daphnia genome''s topography, including regions of high and low recombination, the distribution of transposable, repetitive and regulatory elements, the size and structure of genes and of their neighborhoods. This information is crucial in formulating testable hypotheses relating genetics and demographics to the evolutionary potential or constraints of natural populations. Projects aiming to compile identifiable genes with their function are also underway, together with robust methods to verify these findings. Finally, these tools are being tested, by exploring their uses in key ecological and toxicological investigations. Each project benefits from the leadership and expertise of many individuals. For further details, begin by contacting the project directors. The DGC consists of biologists from a broad spectrum of subdisciplines, including limnology, ecotoxicology, quantitative and population genetics, systematics, molecular biology and evolution, developmental biology, genomics and bioinformatics. In many regards, the rapid early success of the consortium results from its grass-roots origin promoting an international composition, under a cooperative model, with significant scientific breadth. We hold to this approach in building this network and encourage more people to participate. All the while, the DGC is structured to effectively reach specific goals. The consortium includes an advisory board (composed of experts of the various subdisciplines), whose responsibility is to act as the research community''s agent in guiding the development of Daphnia genomic resources. The advisors communicate directly to DGC members, who are either contributing genomic tools or actively seeking funds for this function. The consortium''s main body (given the widespread interest in applying genomic tools in environmental studies) are the affiliates, who make use of these tools for their research and who are soliciting support.
Proper citation: Daphnia genomics consortium (RRID:SCR_008148) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 29, 2016. The BayGenomics gene-trap resource provides researchers with access to thousands of mouse embryonic stem (ES) cell lines harboring characterized insertional mutations in both known and novel genes. The major goal of BayGenomics is to identify genes relevant to cardiovascular and pulmonary disease.
Proper citation: BayGenomics (RRID:SCR_008168) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented on August 20,2019.The COG-database has become a powerful tool in the field of comparative genomics. The construction of this data-base is based on sequence homologies of proteins from different completely sequenced genomes. Highly homologous proteins are assigned to clusters of orthologous groups. The updated collection of orthologous protein sets for prokaryotes and eukaryotes is expected to be a useful platform for functional annotation of newly sequenced genomes, including those of complex eukaryotes, and genome-wide evolutionary studies. The availability of multiple, essentially complete genome sequences of prokaryotes and eukaryotes spurred both the demand and the opportunity for the construction of an evolutionary classification of genes from these genomes. Such a classification system based on orthologous relationships between genes appears to be a natural framework for comparative genomics and should facilitate both functional annotation of genomes and large-scale evolutionary studies. Here is a major update of the previously developed system for delineation of Clusters of Orthologous Groups of proteins (COGs) from the sequenced genomes of prokaryotes and unicellular eukaryotes and the construction of clusters of predicted orthologs for 7 eukaryotic genomes, which we named KOGs after eukaryotic orthologous groups. The COG collection currently consists of 138,458 proteins, which form 4873 COGs and comprise 75% of the 185,505 (predicted) proteins encoded in 66 genomes of unicellular organisms. The eukaryotic orthologous groups (KOGs) include proteins from 7 eukaryotic genomes: three animals (the nematode Caenorhabditis elegans, the fruit fly Drosophila melanogaster and Homo sapiens), one plant, Arabidopsis thaliana, two fungi (Saccharomyces cerevisiae and Schizosaccharomyces pombe), and the intracellular microsporidian parasite Encephalitozoon cuniculi. The current KOG set consists of 4852 clusters of orthologs, which include 59,838 proteins, or approximately 54% of the analyzed eukaryotic 110,655 gene products. Compared to the coverage of the prokaryotic genomes with COGs, a considerably smaller fraction of eukaryotic genes could be included into the KOGs; addition of new eukaryotic genomes is expected to result in substantial increase in the coverage of eukaryotic genomes with KOGs. Examination of the phyletic patterns of KOGs reveals a conserved core represented in all analyzed species and consisting of approximately 20% of the KOG set. This conserved portion of the KOG set is much greater than the ubiquitous portion of the COG set (approximately 1% of the COGs). In part, this difference is probably due to the small number of included eukaryotic genomes, but it could also reflect the relative compactness of eukaryotes as a clade and the greater evolutionary stability of eukaryotic genomes.
Proper citation: Phylogenetic Clusters of Orthologous Groups Ranking (RRID:SCR_008223) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.