Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
A comprehensive collection of experimentally determined and computationally predicted CCCTC-binding factor (CTCF) binding sites (CTCFBS) from the literature. The database is designed to facilitate the studies on insulators and their roles in demarcating functional genomic domains. The CTCFBS Prediction Tool allows users to scan sequences for the single best match to CTCF position weight matrices. Currently (March 2014), the database contains almost 15 million experimentally determined CTCF binding sites across several species. CTCF binding sites were collected from published papers containing CTCF binding sites identified using ChIPSeq or similar methods, data from the ENCODE project, and a set of approximately 100 manually curated binding sites identified by low-throughput experiments. Users can browse insulator sequence features, function annotations, genomic contexts including histone methylation profiles, flanking gene expression patterns and orthologous regions in other mammalian genomes. Users can also retrieve data by text search, sequence search and genomic range search.
Proper citation: CTCFBSDB (RRID:SCR_002279) Copy
http://www.ncbi.nlm.nih.gov/HTGS/
Database of high-throughput genome sequences from large-scale genome sequencing centers, including unfinished and finished sequences. It was created to accommodate a growing need to make unfinished genomic sequence data rapidly available to the scientific community in a coordinated effort among the International Nucleotide Sequence databases, DDBJ, EMBL, and GenBank. Sequences are prepared for submission by using NCBI's software tools Sequin or tbl2asn. Each center has an FTP directory into which new or updated sequence files are placed. Sequence data in this division are available for BLAST homology searches against either the htgs database or the month database, which includes all new submissions for the prior month. Unfinished HTG sequences containing contigs greater than 2 kb are assigned an accession number and deposited in the HTG division. A typical HTG record might consist of all the first-pass sequence data generated from a single cosmid, BAC, YAC, or P1 clone, which together make up more than 2 kb and contain one or more gaps. A single accession number is assigned to this collection of sequences, and each record includes a clear indication of the status (phase 1 or 2) plus a prominent warning that the sequence data are unfinished and may contain errors. The accession number does not change as sequence records are updated; only the most recent version of a HTG record remains in GenBank.
Proper citation: High Throughput Genomic Sequences Division (RRID:SCR_002150) Copy
Portal for studies of genome structure and genetic variation, gene expression and gene function. Provides services including DNA sequencing of model and non-model genomes using both Next Generation and Sanger sequencing , Gene expression analysis using both microarrays and Next Generation Sequencing, High throughput genotyping of SNP and copy number variants, Data collection and analysis supported in-house high performance computing facilities and expertise, Extensive EST clone collections for a number of animal species, all of commercially available microarray tools from Affymetrix, Illumina, Agilent and Nimblegen, Parentage testing using microsatellites and smaller SNP panels. ARK-Genomics has developed network of researchers whom they support through each stage of their genomics research, from grant application, experimental design and technology selection, performing wet laboratory protocols, through to analysis of data often in conjunction with commercial partners.
Proper citation: ARK-Genomics: Centre for Functional Genomics (RRID:SCR_002214) Copy
http://www.genoscope.cns.fr/spip/spip.php?lang=en
French national sequencing center with the following resources: * Sequencing ** Genoscope Projects * Environmental genomics ** Microbial diversity in wastewater ** Metabolic genomics * Bioinformatics ** Atelier for comparative genomics ** Computational Systems Biology ** Servers resources *** GGB for Generic Genome Browser: graphic interface for various databases (sequence, annotation, syntenies...) for a given organism. *** MaGe for Magnifying Microbial Genomes: annotation system for microbial genomes.
Proper citation: Genoscope (RRID:SCR_002172) Copy
Maintains and provides archival, retrieval and analytical resources for biological information. Central DDBJ resource consists of public, open-access nucleotide sequence databases including raw sequence reads, assembly information and functional annotation. Database content is exchanged with EBI and NCBI within the framework of the International Nucleotide Sequence Database Collaboration (INSDC). In 2011, DDBJ launched two new resources: DDBJ Omics Archive and BioProject. DOR is archival database of functional genomics data generated by microarray and highly parallel new generation sequencers. Data are exchanged between the ArrayExpress at EBI and DOR in the common MAGE-TAB format. BioProject provides organizational framework to access metadata about research projects and data from projects that are deposited into different databases.
Proper citation: DNA DataBank of Japan (DDBJ) (RRID:SCR_002359) Copy
http://www.ncbi.nlm.nih.gov/genome
Database that organizes information on genomes including sequences, maps, chromosomes, assemblies, and annotations in six major organism groups: Archaea, Bacteria, Eukaryotes, Viruses, Viroids, and Plasmids. Genomes of over 1,200 organisms can be found in this database, representing both completely sequenced organisms and those for which sequencing is in progress. Users can browse by organism, and view genome maps and protein clusters. Links to other prokaryotic and archaeal genome projects, as well as BLAST tools and access to the rest of the NCBI online resources are available.
Proper citation: NCBI Genome (RRID:SCR_002474) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on March 17, 2022. A secure repository for storing, cataloging, and accessing cancer genome sequences, alignments, and mutation information from the Cancer Genome Atlas (TCGA) consortium and related projects. CGHub gives scientific researchers the statistical power of large cancer genome datasets to attack the molecular complexity of cancer.
Proper citation: Cancer Genomics Hub (RRID:SCR_002657) Copy
http://bioweb.ensam.inra.fr/esther
Database and tools for analysis of protein and nucleic acid sequences belonging to superfamily of alpha/beta hydrolases homologous to cholinesterases. Covers multiple species, including human, mouse caenorhabditis and drosophila., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: ESTHER (RRID:SCR_002621) Copy
http://www.nitrc.org/projects/penncnv
A free software tool for Copy Number Variation (CNV) detection from SNP genotyping arrays. Currently it can handle signal intensity data from Illumina and Affymetrix arrays. With appropriate preparation of file format, it can also handle other types of SNP arrays and oligonucleotide arrays. PennCNV implements a hidden Markov model (HMM) that integrates multiple sources of information to infer CNV calls for individual genotyped samples. It differs form segmentation-based algorithm in that it considered SNP allelic ratio distribution as well as other factors, in addition to signal intensity alone. In addition, PennCNV can optionally utilize family information to generate family-based CNV calls by several different algorithms. Furthermore, PennCNV can generate CNV calls given a specific set of candidate CNV regions, through a validation-calling algorithm.
Proper citation: PennCNV (RRID:SCR_002518) Copy
http://genomefoundation.org/index.php/Main_Page
The Genome Foundation (AKA Genome Research Foundation) is a fully government accredited and registered non-profit research foundation. GRF aims to provide genome philosophy, science, and technology. GRF is a nonprofit publisher, and research and advocacy organization to promote completely free publication of knowledge with minimum restriction. Our core objectives are to: * Provide ways to overcome unnecessary barriers to immediate availability, access, and use of research * Pursue a publishing strategy that optimizes the openness, quality, and integrity of the publication process * Develop innovative approaches to the assessment, organization, and reuse of ideas and data Genome Foundation Research * Personalized Medicine * Personal Genomics * AngioGenesis drug * Bioinformatics * RNA expression * Protein structure * Human Genome Rights Projects at Genome Foundation * The Human Genome Rights * Human Genome Rights Petition * Free Personal Genome Sequencing Project * Free Personal Genome Sequencing Petition * Tiger Genome Initiative: Amur Tiger and big cat genomes * Whale Genome Project
Proper citation: Genome Research Foundation (RRID:SCR_006056) Copy
http://bio-bigdata.hrbmu.edu.cn/diseasemeth/
Human disease methylation database. DiseaseMeth version 2.0 is focused on aberrant methylomes of human diseases. Used for understanding of DNA methylation driven human diseases.
Proper citation: DiseaseMeth (RRID:SCR_005942) Copy
http://www.nematodes.org/nembase4/
NEMBASE is a comprehensive Nematode Transcriptome Database including 63 nematode species, over 600,000 ESTs and over 250,000 proteins. Nematode parasites are of major importance in human health and agriculture, and free-living species deliver essential ecosystem services. The genomics revolution has resulted in the production of many datasets of expressed sequence tags (ESTs) from a phylogenetically wide range of nematode species, but these are not easily compared. NEMBASE4 presents a single portal into extensively functionally annotated, EST-derived transcriptomes from over 60 species of nematodes, including plant and animal parasites and free-living taxa. Using the PartiGene suite of tools, we have assembled the publicly available ESTs for each species into a high-quality set of putative transcripts. These transcripts have been translated to produce a protein sequence resource and each is annotated with functional information derived from comparison with well-studied nematode species such as Caenorhabditis elegans and other non-nematode resources. By cross-comparing the sequences within NEMBASE4, we have also generated a protein family assignment for each translation. The data are presented in an openly accessible, interactive database. An example of the utility of NEMBASE4 is that it can examine the uniqueness of the transcriptomes of major clades of parasitic nematodes, identifying lineage-restricted genes that may underpin particular parasitic phenotypes, possible viral pathogens of nematodes, and nematode-unique protein families that may be developed as drug targets.
Proper citation: NEMBASE (RRID:SCR_006070) Copy
One of eight Bioinformatics Resource Centers nationwide providing comprehensive web-based genomics resources including a relational database and web application supporting data storage, annotation, analysis, and information exchange to support scientific research directed at viruses belonging to the Arenaviridae, Bunyaviridae, Filoviridae, Flaviviridae, Paramyxoviridae, Poxviridae, and Togaviridae families. These centers serve the scientific community and conduct basic and applied research on microorganisms selected from the NIH/NIAID Category A, B, and C priority pathogens that are regarded as possible bioterrorist threats or as emerging or re-emerging infectious diseases. The VBRC provides a variety of analytical and visualization tools to aid in the understanding of the available data, including tools for genome annotation, comparative analysis, whole genome alignments, and phylogenetic analysis. Each data release contains the complete genomic sequences for all viral pathogens and related strains that are available for species in the above-named families. In addition to sequence data, the VBRC provides a curation for each virus species, resulting in a searchable, comprehensive mini-review of gene function relating genotype to biological phenotype, with special emphasis on pathogenesis.
Proper citation: VBRC (RRID:SCR_005971) Copy
FungiDB is a database for functional and evolutionary comparison of fungal genomes. FungiDB is a functional genomic resource for pan-fungal genomes that was developed in partnership with the Eukaryotic Pathogen Bioinformatic resource center (http://EuPathDB.org). FungiDB uses the same infrastructure and user interface as EuPathDB, which allows for sophisticated and integrated searches to be performed using an intuitive graphical system. The current release of FungiDB contains genome sequence and annotation from 18 species spanning several fungal classes, including the Ascomycota classes, Eurotiomycetes, Sordariomycetes, Saccharomycetes and the Basidiomycota orders, Pucciniomycetes and Tremellomycetes, and the basal "Zygomycete" lineage Mucormycotina. Additionally, FungiDB contains cell cycle microarray data, hyphal growth RNA-sequence data and yeast two hybrid interaction data. The underlying genomic sequence and annotation combined with functional data, additional data from the FungiDB standard analysis pipeline and the ability to leverage orthology provides a powerful resource for in silico experimentation.
Proper citation: FungiDB (RRID:SCR_006013) Copy
http://athina.biol.uoa.gr/bioinformatics/GENEVITO/
A JAVA-based computer application that serves as a workbench for genome-wide analysis through visual interaction. GeneViTo offers an inspectional view of genomic functional elements, concerning data stemming both from database annotation and analysis tools for an overall analysis of existing genomes. The application deals with various experimental information concerning both DNA and protein sequences (derived from public sequence databases or proprietary data sources) and meta-data obtained by various prediction algorithms, classification schemes or user-defined features. Interaction with a Graphical User Interface (GUI) allows easy extraction of genomic and proteomic data referring to the sequence itself, sequence features, or general structural and functional features. Emphasis is laid on the potential comparison between annotation and prediction data in order to offer a supplement to the provided information, especially in cases of poor annotation, or an evaluation of available predictions. Moreover, desired information can be output in high quality JPEG image files for further elaboration and scientific use. GeneViTo has already been applied to visualize the genomes of two microbial organisms: the bacterion Chlamydia trachomatis and the archaeon Methanococcus jannaschii. The application is compatible with Linux or Windows ME-2000-XP operating systems, provided that the appropriate Java Runtime Environment (Java 1.4.1) is already installed in the system.
Proper citation: GeneVito (RRID:SCR_006211) Copy
The Deciphering Developmental Disorders (DDD) study aims to find out if using new genetic technologies can help doctors understand why patients get developmental disorders. To do this we have brought together doctors in the 23 NHS Regional Genetics Services throughout the UK and scientists at the Wellcome Trust Sanger Institute, a charitably funded research institute which played a world-leading role in sequencing (reading) the human genome. The DDD study involves experts in clinical, molecular and statistical genetics, as well as ethics and social science. It has a Scientific Advisory Board consisting of scientists, doctors, a lawyer and patient representative, and has received National ethical approval in the UK. Over the next few years, we are aiming to collect DNA and clinical information from 12,000 undiagnosed children in the UK with developmental disorders and their parents. The results of the DDD study will provide a unique, online catalogue of genetic changes linked to clinical features that will enable clinicians to diagnose developmental disorders. Furthermore, the study will enable the design of more efficient and cheaper diagnostic assays for relevant genetic testing to be offered to all such patients in the UK and so transform clinical practice for children with developmental disorders. Over time, the work will also improve understanding of how genetic changes cause developmental disorders and why the severity of the disease varies in individuals. The Sanger Institute will contribute to the DDD study by performing genetic analysis of DNA samples from patients with developmental disorders, and their parents, recruited into the study through the Regional Genetics Services. Using microarray technology and the latest DNA sequencing methods, research teams will probe genetic information to identify mutations (DNA errors or rearrangements) and establish if these mutations play a role in the developmental disorders observed in patients. The DDD initiative grew out of the groundbreaking DECIPHER database, a global partnership of clinical genetics centres set up in 2004, which allows researchers and clinicians to share clinical and genomic data from patients worldwide. The DDD study aims to transform the power of DECIPHER as a diagnostic tool for use by clinicians. As well as improving patient care, the DDD team will empower researchers in the field by making the data generated securely available to other research teams around the world. By assembling a solid resource of high-quality, high-resolution and consistent genomic data, the leaders of the DDD study hope to extend the reach of DECIPHER across a broader spectrum of disorders than is currently possible.
Proper citation: Deciphering Developmental Disorders (RRID:SCR_006171) Copy
A visualization hub displaying sequencing data from the Roadmap Epigenomics project. It hosts high volume of tracks from ENCODE and Roadmap Epigenomics projects, supports multiple organisms, visualizes chromatin-interaction data (e.g. Hi-C), performs gene set view, gene plot, and many others. All delivered on the web at high performance.
Proper citation: VizHub (RRID:SCR_006209) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 17, 2013. A public resource for sharing general proteomics information including data (Tranche repository), tools, and news. Joining or creating a group/project provides tools and standards for collaboration, project management, data annotation, permissions, permanent storage, and publication.
Proper citation: Proteome Commons (RRID:SCR_006234) Copy
http://pathogenseq.lshtm.ac.uk/estmoi
A per-based software to estimate multiplicity of infection (MOI) in parasite genomic sequence data. It is primarily developed to address the limitations of current laboratory (PCR) based estimates of multiplicity using high throughput sequence data. It requires a BAM (alignment output of short reads to the reference genome), VCF (a file with information on variant calls) and FASTA (reference genome) files. # Short reads are aligned to a reference genome using BWA, BOWTIE, SMALT or other short read aligners to generate a BAM file. # Single Nucleotide Polymorphisms (SNPs) are then identified using SAMTools/BCFtools and stored in the VCF format. # The reference FASTA file is expected to be indexed using ''samtools faidx'' to generate a *.fai file. estMOI generates files containing MOI estimates for each SNP combinations (file with name *.log) and a summary for all chromosomes (file with name *.txt).
Proper citation: estMOI (RRID:SCR_006192) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on July 7, 2022. Federation of International Mouse Resources (FIMRe) is a collaborating group of Mouse Repository and Resource Centers worldwide whose collective goal is to archive and provide strains of mice as cryopreserved embryos and gametes, ES cell lines, and live breeding stock to the research community. Goals of the Federation of International Mouse Resources: * Coordinate repositories and resource centers to: ** archive valuable genetically defined mice and ES cell lines being created worldwide ** meet research demand for these genetically defined mice and ES cell lines * Establish consistent, highest quality animal health standards in all resource centers * Provide genetic verification and quality control for genetic background and mutations * Provide resource training to enhance user ability to utilize cryopreserved resources
Proper citation: Federation of International Mouse Resources (RRID:SCR_006137) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.