Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Portal to the PSORT family of computer programs for the prediction of protein localization sites in cells, as well as other datasets and resources relevant to localization prediction. The standalone versions are available for download for larger analyses.
Proper citation: Psort (RRID:SCR_007038) Copy
Comprehensive set of protein domain families automatically generated from UniProt Knowledge Database. Automated clustering of homologous domains generated from global comparison of all available protein sequences., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: ProDom (RRID:SCR_006969) Copy
http://www.genoscope.cns.fr/externe/tetraodon/
The initial objective of Genoscope was to compare the genomic sequences of this fish to that of humans to help in the annotation of human genes and to estimate their number. This strategy is based on the common genetic heritage of the vertebrates: from one species of vertebrate to another, even for those as far apart as a fish and a mammal, the same genes are present for the most part. In the case of the compact genome of Tetraodon, this common complement of genes is contained in a genome eight times smaller than that of humans. Although the length of the exons is similar in these two species, the size of the introns and the intergenic sequences is greatly reduced in this fish. Furthermore, these regions, in contrast to the exons, have diverged completely since the separation of the lineages leading to humans and Tetraodon. The Exofish method, developed at Genoscope, exploits this contrast such that the conserved regions which can be identified by comparing genomic sequences of the two species, correspond only to coding regions. Using preliminary sequencing results of the genome of Tetraodon in the year 2000, Genoscope evaluated the number of human genes at about 30,000, whereas much higher estimations were current. The progress of the annotation of the human genome has since supported the Genoscope hypothesis, with values as low as 22,000 genes and a consensus of around 25,000 genes. The sequencing of the Tetraodon genome at a depth of about 8X, carried out as a collaboration between Genoscope and the Whitehead Institute Center for Genome Research (now the Broad Institute), was finished in 2002, with the production of an assembly covering 90 of the euchromatic region of the genome of the fish. This has permitted the application of Exofish at a larger scale in comparisons with the genome of humans, but also with those of the two other vertebrates sequenced at the time (Takifugu, a fish closely related to Tetraodon, and the mouse). The conserved regions detected in this way have been integrated into the annotation procedure, along with other resources (cDNA sequences from Tetraodon and ab initio predictions). Of the 28,000 genes annotated, some families were examined in detail: selenoproteins, and Type 1 cytokines and their receptors. The comparison of the proteome of Tetraodon with those of mammals has revealed some interesting differences, such as a major diversification of some hormone systems and of the collagen molecules in the fish. A search for transposable elements in the genomic sequences of Tetraodon has also revealed a high diversity (75 types), which contrasts with their scarcity; the small size of the Tetraodon genome is due to the low abundance of these elements, of which some appear to still be active. Another factor in the compactness of the Tetraodon genome, which has been confirmed by annotation, is the reduction in intron size, which approaches a lower limit of 50-60 bp, and which preferentially affects certain genes. The availability of the sequences from the genomes of humans and mice on one hand, and Takifugu and Tetraodon on the other, provide new opportunities for the study of vertebrate evolution. We have shown that the level of neutral evolution is higher in fish than in mammals. The protein sequences of fish also diverge more quickly than those of mammals. A key mechanism in evolution is gene duplication, which we have studied by taking advantage of the anchoring of the majority of the sequences from the assembly on the chromosomes. The result of this study speaks strongly in favor of a whole genome duplication event, very early in the line of ray-finned fish (Actinopterygians). An even stronger evidence came from synteny studies between the genomes of humans and Tetraodon. Using a high-resolution synteny map, we have reconstituted the genome of the vertebrate which predates this duplication - that is, the last common ancestor to all bony vertebrates (most of the vertebrates apart from cartilaginous fish and agnaths like lamprey). This ancestral karyotype contains 12 chromosomes, and the 21 Tetraodon chromosomes derive from it by the whole genome duplication and a surprisingly small number of interchromosomal rearrangements. On the contrary, exchanges between chromosomes have been much more frequent in the lineage that leads to humans. Sponsors: The project was supported by the Consortium National de Recherche en Genomique and the National Human Genome Research Institute.
Proper citation: Tetraodon Genome Browser (RRID:SCR_007079) Copy
http://medgen.ugent.be/rtprimerdb/
Database for primer and probe sequences used in real-time PCR assays employing popular chemistries (SYBR Green I, Taqman, Hybridization Probes, Molecular Beacon) to prevent time-consuming primer design and experimental optimization, and to introduce a certain level of uniformity and standardization among different laboratories. Researchers are encouraged to submit their validated primer and probe sequence, so that other users can benefit from their expertise. The database can be queried using the official gene name or symbol, Entrez or Ensembl Gene identifier, SNP identifier, or oligonucleotide sequence. Different options make it possible to restrict a query to a particular application (Gene Expression Quantification/Detection, DNA Copy Number Quantification/Detection, SNP Detection, Mutation Analysis, Fusion Gene Quantification/Detection, Chromatin immunoprecipitation (ChIP)), organism (Human, Mouse, Rat, and others) or detection chemistry.
Proper citation: RTPrimerDB- The Real-Time PCR and Probe Database (RRID:SCR_007106) Copy
http://weizhong-lab.ucsd.edu/cd-hit/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software program for clustering biological sequences with many applications in various fields such as making non-redundant databases, finding duplicates, identifying protein families, filtering sequence errors and improving sequence assembly etc. It is very fast and can handle extremely large databases. CD-HIT helps to significantly reduce the computational and manual efforts in many sequence analysis tasks and aids in understanding the data structure and correct the bias within a dataset. The CD-HIT package has CD-HIT, CD-HIT-2D, CD-HIT-EST, CD-HIT-EST-2D, CD-HIT-454, CD-HIT-PARA, PSI-CD-HIT, CD-HIT-OTU and over a dozen scripts. * CD-HIT (CD-HIT-EST) clusters similar proteins (DNAs) into clusters that meet a user-defined similarity threshold. * CD-HIT-2D (CD-HIT-EST-2D) compares 2 datasets and identifies the sequences in db2 that are similar to db1 above a threshold. * CD-HIT-454 identifies natural and artificial duplicates from pyrosequencing reads. * CD-HIT-OTU cluster rRNA tags into OTUs The usage of other programs and scripts can be found in CD-HIT user''s guide. CD-HIT was originally developed by Dr. Weizhong Li at Dr. Adam Godzik''s Lab at the Burnham Institute (now Sanford-Burnham Medical Research Institute)., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: CD-HIT (RRID:SCR_007105) Copy
http://cagt.bu.edu/page/PRECISE_about
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on May 12,2023. Database of interactions between amino acid residues of enzyme and its ligands. Provides summary of interactions between amino acid residues of enzyme and its various ligands including substrate and transition state analogues, cofactors, inhibitors, and products., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: PRECISE (RRID:SCR_007874) Copy
A fungal rDNA internal transcribed spacer (ITS) sequence database (although additional genes and genetic markers are also welcome) to facilitate identification of environmental samples of fungal DNA. Additional important features include user annotation of INSD sequences to add metadata on, e.g., locality, habitat, soil, climate, and interacting taxa. The user can furthermore annotate INSD sequences with additional species identifications that will appear in the results of any analyses done. UNITE focuses on high-quality ITS sequences generated from fruiting bodies collected and identified by experts and deposited in public herbaria. In addition, it also holds all fungal ITS sequences in the International Nucleotide Sequence Databases (INSD: NCBI, EMBL, DDBJ). Both sets of sequences may be used in any analyses carried out. UNITE is accompanied by a project management system called PlutoF, where users can store field data, document the sequencing lab procedures, manage sequences, and make analyses. PlutoF intends to make it possible for taxonomists, ecologists, and biogeographers to use a common platform for data storage, handling, and analyses, with the intent of facilitating an integration of these disciplines. A user can have an unlimited number of projects but still make analyses across any project data available to him.
Proper citation: UNITE (RRID:SCR_006518) Copy
A comparative platform for green plant genomics. Families of orthologous and paralogous genes that represent the modern descendents of ancestral gene sets are constructed at key phylogenetic nodes. These families allow easy access to clade specific orthology / paralogy relationships as well as clade specific genes and gene expansions. As of release v9.1, Phytozome provides access to forty-one sequenced and annotated green plant genomes which have been clustered into gene families at 20 evolutionarily significant nodes. Where possible, each gene has been annotated with PFAM, KOG, KEGG, and PANTHER assignments, and publicly available annotations from RefSeq, UniProt, TAIR, JGI are hyper-linked and searchable., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Phytozome (RRID:SCR_006507) Copy
http://www.ncbi.nlm.nih.gov/projects/genome/assembly/grc/
Consortium that puts sequences into a chromosome context and provides the best possible reference assembly for human, mouse, and zebrafish via FTP. Tools to facilitate the curation of genome assemblies based on the sequence overlaps of long, high quality sequences.
Proper citation: Genome Reference Consortium (RRID:SCR_006553) Copy
Model organism database for the social amoeba Dictyostelium discoideum that provides the biomedical research community with integrated, high quality data and tools for Dictyostelium discoideum and related species. dictyBase houses the complete genome sequence, ESTs, and the entire body of literature relevant to Dictyostelium. This information is curated to provide accurate gene models and functional annotations, with the goal of fully annotating the genome to provide a ''''reference genome'''' in the Amoebozoa clade. They highlight several new features in the present update: (i) new annotations; (ii) improved interface with web 2.0 functionality; (iii) the initial steps towards a genome portal for the Amoebozoa; (iv) ortholog display; and (v) the complete integration of the Dicty Stock Center with dictyBase. The Dicty Stock Center currently holds over 1500 strains targeting over 930 different genes. There are over 100 different distinct amoebozoan species. In addition, the collection contains nearly 600 plasmids and other materials such as antibodies and cDNA libraries. The strain collection includes: * strain catalog * natural isolates * MNNG chemical mutants * tester strains for parasexual genetics * auxotroph strains * null mutants * GFP-labeled strains for cell biology * plasmid catalog The Dicty Stock Center can accept Dictyostelium strains, plasmids, and other materials relevant for research using Dictyostelium such as antibodies and cDNA or genomic libraries.
Proper citation: Dictyostelium discoideum genome database (RRID:SCR_006643) Copy
DPVweb provides a central source of information about viruses, viroids and satellites of plants, fungi and protozoa. Comprehensive taxonomic information, including brief descriptions of each family and genus, and classified lists of virus sequences are provided. The database also holds detailed, curated, information for all sequences of viruses, viroids and satellites of plants, fungi and protozoa that are complete or that contain at least one complete gene. For comparative purposes, it also contains a single representative sequence of all other fully sequenced virus species with an RNA or single-stranded DNA genome. The start and end positions of each feature (gene, non-translated region and the like) have been recorded and checked for accuracy. As far as possible, nomenclature for genes and proteins are standardized within genera and families. Sequences of features (either as DNA or amino acid sequences) can be directly downloaded from the website in FASTA format. The sequence information can also be accessed via client software for PC computers (freely downloadable from the website) that enable users to make an easy selection of sequences and features of a chosen virus for further analyses. The public sequence databases contain vast amounts of data on virus genomes but accessing and comparing the data, except for relatively small sets of related viruses can be very time consuming. The procedure is made difficult because some of the sequences on these databases are incorrectly named, poorly annotated or redundant. The NCBI Reference Sequence project (1) provides a comprehensive, integrated, non-redundant set of sequences, including genomic DNA, transcript (RNA) and protein products, for major research organisms. This now includes curated information for a single sequence of each fully sequenced virus species. While this is a welcome development, it can only deal with complete sequences. An important feature of DPV is the opportunity to access genes (and other features) of multiple sequences quickly and accurately. Thus, for example, it is easy to obtain the nucleotide or amino acid sequences of all the available accessions of the coat protein gene of a given virus species or for a group of viruses. To increase its usefulness further, DPVweb also contains a single representative sequence of all other fully sequenced virus species with an RNA or single-stranded DNA (ssDNA) genome. Sponsors: This site is supported by the Association of Applied Biologists and the Zhejiang Academy of Agricultural Sciences, Hangzhou, People''s Republic of China.
Proper citation: Descriptions of Plant Viruses (RRID:SCR_006656) Copy
https://github.com/cmayer/BaitFisher-package
Software toolkit for multispecies target DNA enrichment probe design. It consists of two programs: BaitFisher and BaitFilter, which are designed to construct hybrid enrichment baits for multiple sequence alignments or annotated features in multiple sequence alignments.
Proper citation: Baitfisher (RRID:SCR_015985) Copy
http://www.sanger.ac.uk/science/tools/seqtools
Software for multiple sequence alignment viewing, editing and phylogeny. It includes a set of user-configurable modes to color residues used to create high-quality reference alignments.
Proper citation: Belvu (RRID:SCR_015989) Copy
Simulation software for experimental evolution of microorganisms. Aevol is a digital genetics model for the study of structural variations of the genome (e.g. number of genes, synteny, proportion of coding sequences).
Proper citation: Aevol (RRID:SCR_015966) Copy
https://bioinf.eva.mpg.de/anfo/
Software for short read alignment and mapping of sequencing reads where the DNA sequence is somehow modified and/or there is more divergence between sample and reference than what fast mappers will handle.
Proper citation: Anfo (RRID:SCR_015972) Copy
http://disulfind.dsi.unifi.it/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023, Software for predicting the disulfide bonding state of cysteines and their disulfide connectivity, starting from a protein sequence alone and may be useful in other genomic annotation tasks.
Proper citation: DISULFIND (RRID:SCR_016072) Copy
https://github.com/osallou/cassiopee-c
Software to scan an input genomic sequence (dna/rna/protein). It searchs for a subsequence that has an exact match, substitutions (Hamming distance), and/or insertion/deletions with supporting alphabet ambiguity.
Proper citation: Cassiopee (RRID:SCR_016056) Copy
http://cdbfasta.sourceforge.net/
Software tool for indexing and retrieval of nucleotide sequences from FASTA (DNA and protein sequence alignment software) record databases. It has the option to compress data records.
Proper citation: Cdbfasta (RRID:SCR_016057) Copy
http://www.xavierdidelot.xtreemhost.com/clonalframe.htm
Software package for the inference of bacterial microevolution using multilocus sequence data. It is used to identify the clonal relationships between the members of a sample, while also estimating the chromosomal position of homologous recombination events that have disrupted the clonal inheritance.
Proper citation: Clonalframe (RRID:SCR_016060) Copy
https://dazzlerblog.wordpress.com
Software alignment tool to find all significant local alignments between long and noisy, up to 15% on average reads encoded in a Dazzler database. Used for DNA sequence assembly, specifically for next generation long-read sequencers such as the Pacbio RS II and Sequel sequencers.
Proper citation: Daligner (RRID:SCR_016066) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.