Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
A human full-length cDNA sequence analysis database focused on mRNA varieties caused by variations of transcription start site (TSS) and splicing. Also available is ATGpr, a program for identifying the translational initiation codons in cDNA sequences. Data are derived from several full-length cDNA studies in Japan. Human gene number was estimated to be 20-25 thousand. However, the number of human mRNA varieties was predicted to be about 100 thousand. The varieties are thought to be caused by variations of TSS and splicing. In their previous human cDNA project, about 30 thousand of FLJ human full-length sequenced cDNAs were deposited to DDBJ/GenBank/EMBL, and they obtained about 1.4 million of 5''-end sequences (5''-EST) of FLJ full-length cDNAs from about 100 kinds of cDNA libraries consist of human tissues and cells constructed by oligo-capping method. The majority of the insert cDNA sizes were over 2 kb and the full-length rate of 5''-end was 90. And our FLJ cDNAs were covered about 80 of human genes. About 22 thousand of finished grades of full-length sequenced cDNAs were obtained in this project. The sequence analysis databases is focused on mRNA variations using human genome and cDNA sequences, FLJ full-length sequenced cDNAs, 5-ESTs of FLJ full-length cDNAs and other cDNA sequences described below. After those sequences were mapped onto the human genome sequences, clustering of the cDNA sequences were done based on the mapping results.
Proper citation: FLJ Human cDNA Database (RRID:SCR_008253) Copy
The Distributed Annotation System (DAS) defines a communication protocol used to exchange annotations on genomic or protein sequences. It is motivated by the idea that such annotations should not be provided by single centralized databases, but should instead be spread over multiple sites. Data distribution, performed by DAS servers, is separated from visualization, which is done by DAS clients. The advantages of this system are that control over the data is retained by data providers, data is freed from the constraints of specific organisations and the normal issues of release cycles, API updates and data duplication are avoided. DAS is a client-server system in which a single client integrates information from multiple servers. It allows a single machine to gather up sequence annotation information from multiple distant web sites, collate the information, and display it to the user in a single view. Little coordination is needed among the various information providers. DAS is heavily used in the genome bioinformatics community. Over the last years we have also seen growing acceptance in the protein sequence and structure communities. A DAS-enabled website or application can aggregate complex and high-volume data from external providers in an efficient manner. For the biologist, this means the ability to plug in the latest data, possibly including a user''s own data. For the application developer, this means protection from data format changes and the ability to add new data with minimal development cost. Here are some examples of DAS-enabled applications or websites for end users: :- Dalliance Experimental Web/Javascript based Genome Viewer :- IGV Integrative Genome Viewer java based browser for many genomes :- Ensembl uses DAS to pull in genomic, gene and protein annotations. It also provides data via DAS. :- Gbrowse is a generic genome browser, and is both a consumer and provider of DAS. :- IGB is a desktop application for viewing genomic data. :- SPICE is an application for projecting protein annotations onto 3D structures. :- Dasty2 is a web-based viewer for protein annotations :- Jalview is a multiple alignment editor. :- PeppeR is a graphical viewer for 3D electron microscopy data. :- DASMI is an integration portal for protein interaction data. :- DASher is a Java-based viewer for protein annotations. :- EpiC presents structure-function summaries for antibody design. :- STRAP is a STRucture-based sequence Alignment Program. Hundreds of DAS servers are currently running worldwide, including those provided by the European Bioinformatics Institute, Ensembl, the Sanger Institute, UCSC, WormBase, FlyBase, TIGR, and UniProt. For a listing of all available DAS sources please visit the DasRegistry. Sponsors: The initial ideas for DAS were developed in conversations with LaDeana Hillier of the Washington University Genome Sequencing Center.
Proper citation: Distributed Annotation System (RRID:SCR_008427) Copy
The project began as a pilot study to identify inherited genetic susceptibility to prostate and breast cancer. CGEMS has developed into a robust research program involving genome-wide association studies (GWASs) for a number of cancers to identify common genetic variants that affect a person''s risk of developing cancer. In collaboration with extramural scientists, NCI''s Division of Cancer Epidemiology and Genetics (DCEG) has carried out genome-wide scans for breast, prostate, pancreatic, and lung cancers, while a GWAS of bladder cancer is currently underway. By making the data available to both intramural and extramural research scientists, as well as those in the private sector through rapid posting, NIH can leverage its resources to ensure that the dramatic advances in genomics are incorporated into rigorous population-based studies. Ultimately, findings from these studies may yield new preventive, diagnostic, and therapeutic interventions for cancer. Sponsors: This resource is supported by the U.S. National Institues Of Health.
Proper citation: CGEMS (RRID:SCR_008445) Copy
https://www.encodeproject.org/
Consortium to build comprehensive parts list of functional elements in human genome. This includes elements that act at protein and RNA levels, and regulatory elements that control cells and circumstances in which gene is active. Data from 2012-present.
Proper citation: Encode (RRID:SCR_015482) Copy
Visualization and analysis software for interactive visual exploration and mining of fiber-tracts and brain networks with their genetic determinants and functional outcomes. BECA includes an fMRI and Diseases Analysis version as well as a Genome Explorer version.
Proper citation: BECA (RRID:SCR_015846) Copy
http://www.alliancegenome.org/
Organization that aims to develop and maintain sustainable genome information resources to promote understanding of the genetic and genomic basis of human biology, health, and disease. The Alliance is composed of FlyBase, Mouse Genome Database (MGD), the Gene Ontology Consortium (GOC), Saccharomyces Genome Database (SGD), Rat Genome Database (RGD), WormBase, and the Zebrafish Information Network (ZFIN).
Proper citation: Alliance of Genome Resources (RRID:SCR_015850) Copy
https://github.com/thackl/cross-species-scaffolding
Software that generates in silico mate-pair reads from single-/paired-end reads of your organism of interest, and a closely related reference genome. It can improve draft genomes by using preferred scaffolding software with the newly created read data. Software that generates in silico mate-pair reads from single-/paired-end reads of your organism of interest, and a closely related reference genome. It can improve draft genomes by using preferred scaffolding software with the newly created read data. Super-scaffolding of draft genome assemblies with in silico mate-pair libraries derived from (closely) related references.
Proper citation: Cross-species scaffolding (RRID:SCR_015932) Copy
https://github.com/harry-thorpe/piggy
Pipeline for analyzing intergenic regions in bacteria. It is designed to be used in conjunction with Roary (https://github.com/sanger-pathogens/Roary).
Proper citation: Piggy (RRID:SCR_015941) Copy
Web application to perform automated model construction and genome annotation for large-scale metabolic networks. Platform for accessing, analyzing and manipulating genome-scale metabolic networks (GSM) as well as biochemical pathways.
Proper citation: MetaNetX (RRID:SCR_015882) Copy
http://www.vicbioinformatics.com/software.barrnap.shtml
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software to predict the location of ribosomal RNA genes in genomes. It supports bacteria, archaea, mitochondria, and eukaryotes. It takes FASTA DNA sequence as input, writes GFF3 as output, and supports multithreading., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Barrnap (RRID:SCR_015995) Copy
http://standage.github.io/AEGeAn
Software toolkit for the analysis and evaluation of genome annotations. The toolkit includes a variety of analysis programs, e.g. for comparing distinct sets of gene structure annotations (ParsEval), computation of gene loci (LocusPocus) and more.
Proper citation: Aegean (RRID:SCR_015965) Copy
https://github.com/EvolBioInf/andi
Software tool for rapidly computing and estimating evolutionary distance between closely related genomes. Because andi does not compute full alignments it scales even up to thousands of bacterial genomes.
Proper citation: andi (RRID:SCR_015971) Copy
http://plantgrn.noble.org/LegumeIP/
LegumeIP is an integrative database and bioinformatics platform for comparative genomics and transcriptomics to facilitate the study of gene function and genome evolution in legumes, and ultimately to generate molecular based breeding tools to improve quality of crop legumes. LegumeIP currently hosts large-scale genomics and transcriptomics data, including: * Genomic sequences of three model legumes, i.e. Medicago truncatula, Glycine max (soybean) and Lotus japonicus, including two reference plant species, Arabidopsis thaliana and Poplar trichocarpa, with the annotation based on UniProt TrEMBL, InterProScan, Gene Ontology and KEGG databases. LegumeIP covers a total 222,217 protein-coding gene sequences. * Large-scale gene expression data compiled from 104 array hybridizations from L. japonicas, 156 array hybridizations from M. truncatula gene atlas database, and 14 RNA-Seq-based gene expression profiles from G. max on different tissues including four common tissues: Nodule, Flower, Root and Leaf. * Systematic synteny analysis among M. truncatula, G. max, L. japonicus and A. thaliana. * Reconstruction of gene family and gene family-wide phylogenetic analysis across the five hosted species. LegumeIP features comprehensive search and visualization tools to enable the flexible query on gene annotation, gene family, synteny, relative abundance of gene expression.
Proper citation: LegumeIP (RRID:SCR_008906) Copy
http://hymenopteragenome.org/beebase/
Gene sequences and genomes of Bombus terrestris, Bombus impatiens, Apis mellifera and three of its pathogens, that are discoverable and analyzed via genome browsers, blast search, and apollo annotation tool. The genomes of two additional species, Apis dorsata and A. florea are currently under analysis and will soon be incorporated.BeeBase is an archive and will not be updated. The most up-to-date bee genome data is now available through the navigation bar on the HGD Home page.
Proper citation: BeeBase (RRID:SCR_008966) Copy
http://rgd.mcw.edu/rgdCuration/?module=portal&func=show&name=renal
An integrated resource for information on genes, QTLs and strains associated with a variety of kidney and renal system conditions such as Renal Hypertension, Polycystic Kidney Disease and Renal Insufficiency, as well as Kidney Neoplasms.
Proper citation: Renal Disease Portal (RRID:SCR_009030) Copy
http://fnih.org/work/past-programs/genetic-association-information-network-gain
The Genetic Association Information Network (GAIN) supports a series of Genome-Wide Association Studies (GWAS) designed to identify specific points of DNA variation associated with the occurrence of a particular common disease. Initially focusing on six major common diseases, GAIN focused on combining the results with clinical data to create a significant new resource for genetic researchers.
Proper citation: Genetic Association Information Network (GAIN) (RRID:SCR_013703) Copy
A SEED-quality automated service that annotates complete or nearly complete bacterial and archaeal genomes across the entire phylogenetic tree. RAST can also be used to analyze draft genomes.
Proper citation: RAST Server (RRID:SCR_014606) Copy
http://www.vicbioinformatics.com/software.prokka.shtml
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software tool for the rapid annotation of prokaryotic genomes. It produces GFF3, GBK and SQN files that are ready for editing in Sequin and ultimately submitted to Genbank/DDJB/ENA. A typical 4 Mbp genome can be fully annotated in less than 10 minutes on a quad-core computer, and scales well to 32 core SMP systems., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Prokka (RRID:SCR_014732) Copy
An easy-to-use, highly customizable genome browser you can use to visualize and explore genomic data and annotations, including RNA-Seq, ChIP-Seq, tiling array data, and more.
Proper citation: IGB (RRID:SCR_011792) Copy
https://www.genome-cloud.com/user/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on August 29, 2019. A cloud platform for next-generation sequencing analysis and storage. Services include: * g-Analysis: Automated genome analysis pipelines at your fingertips * g-Cluster: Easy-of-use and cost-effective genome research infrastructure * g-Storage: A simple way to store, share and protect data * g-Insight: Accurate analysis and interpretation of biological meaning of genome data
Proper citation: GenomeCloud (RRID:SCR_011886) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.