Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Alignment software for large-scale protein contact or protein-protein interaction prediction optimized for speed through shorter runtimes. FreeContact provides the opportunity to compute contact predictions in any environment (desktop or cloud).
Proper citation: FreeContact (RRID:SCR_016113) Copy
http://gene3d.biochem.ucl.ac.uk/Gene3D/
A large database of CATH protein domain assignments for ENSEMBL genomes and Uniprot sequences. Gene3D is a resource of form studying proteins and the component domains. Gene3D takes CATH domains from Protein Databank (PDB) structures and assigns them to the millions of protein sequences with no PDB structures using Hidden Markov models. Assigning a CATH superfamily to a region of a protein sequence gives information on the gross 3D structure of that region of the protein. CATH superfamilies have a limited set of functions and so the domain assignment provides some functional insights. Furthermore most proteins have several different domains in a specific order, so looking for proteins with a similar domain organization provides further functional insights. Strict confidence cut-offs are used to ensure the reliability of the domain assignments. Gene3D imports functional information from sources such as UNIPROT, and KEGG. They also import experimental datasets on request to help researchers integrate there data with the corpus of the literature. The website allows users to view descriptions for both single proteins and genes and large protein sets, such as superfamilies or genomes. Subsets can then be selected for detailed investigation or associated functions and interactions can be used to expand explorations to new proteins. The Gene3D web services provide programmatic access to the CATH-Gene3D annotation resources and in-house software tools. These services include Gene3DScan for identifying structural domains within protein sequences, access to pre-calculated annotations for the major sequence databases, and linked functional annotation from UniProt, GO and KEGG., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Gene3D (RRID:SCR_007672) Copy
http://www.bioinformatics.org/peakanalyzer/wiki/
A set of standalone software programs for the automated processing of any genomic loci, with an emphasis on datasets consisting of ChIP-derived signal peaks. The software is able to identify individual binding / modification sites from enrichment loci, retrieve peak region sequences for motif discovery, and integrate experimental data with different classes of annotated elements throughout the genome. PeakAnalyzer requires a peak file and a feature annotation file in BED or GTF format. Complete annotation files for the current builds of the human (HG19) and mouse (MM9) genomes are provided with the software distribution.
Proper citation: PeakAnalyzer (RRID:SCR_001194) Copy
https://services.healthtech.dtu.dk/services/NetNGlyc-1.0/
Server that predicts N-Glycosylation sites in human proteins using artificial neural networks that examine the sequence context of Asn-Xaa-Ser/Thr sequons. NetNGlyc 1.0 is also available as a stand-alone software package, with the same functionality as the service above. Ready-to-ship packages exist for the most common UNIX platforms.
Proper citation: NetNGlyc (RRID:SCR_001570) Copy
https://services.healthtech.dtu.dk/services/YinOYang-1.2/
Server that produces neural network predictions for O-beta-GlcNAc attachment sites in eukaryotic protein sequences. This server can also use NetPhos, to mark possible phosphorylated sites and hence identify Yin-Yang sites. YinOYang 1.2 is available as a stand-alone software package, with the same functionality. Ready-to-ship packages exist for the most common UNIX platforms.
Proper citation: YinOYang (RRID:SCR_001605) Copy
https://www.ebi.ac.uk/jdispatcher/msa/clustalo?stype=protein
Software package as multiple sequence alignment tool that uses seeded guide trees and HMM profile-profile techniques to generate alignments between three or more sequences. Accepts nucleic acid or protein sequences in multiple sequence formats NBRF/PIR, EMBL/UniProt, Pearson (FASTA), GDE, ALN/Clustal, GCG/MSF, RSF.
Proper citation: Clustal Omega (RRID:SCR_001591) Copy
https://github.com/benedictpaten/pecan
A Java consistency based multiple sequence alignment software program.
Proper citation: Pecan (RRID:SCR_001909) Copy
http://athina.biol.uoa.gr/bioinformatics/PRED-GPCR/
A prediction tool for GPCR Family Classification from sequence alone based on a probabilistic method that uses family-specific profile Hidden Markov Models. The PRED-GPCR system is based on a probabilistic method that uses family specific profile HMMs in order to determine to which GPCR family a query sequence belongs or resembles. The approach proposed in this method exploits the descriptive power of profile HMMs along with an exhaustive discrimination assessment method to select only highly selective and sensitive profiles, for each family. The collection of these profiles constitutes a signature library, which is scanned, for significant matches with a given query sequence. The output report for a query sequence consists of two sections: * A ranked list of the profile HMM matches, below the selected individual motif E-value cutoff, along with their corresponding family. * A ranked list of the Combined P-values, E-values as well as the number of profiles matched for each family. To cross-evaluate your results you can browse through Swiss-Prot, Trembl, Pfam and Prosite family related entries.
Proper citation: PRED-GPCR (RRID:SCR_006196) Copy
Database of Drosophila genetic and genomic information with information about stock collections and fly genetic tools. Gene Ontology (GO) terms are used to describe three attributes of wild-type gene products: their molecular function, the biological processes in which they play a role, and their subcellular location. Additionally, FlyBase accepts data submissions. FlyBase can be searched for genes, alleles, aberrations and other genetic objects, phenotypes, sequences, stocks, images and movies, controlled terms, and Drosophila researchers using the tools available from the "Tools" drop-down menu in the Navigation bar.
Proper citation: FlyBase (RRID:SCR_006549) Copy
https://github.com/uclinfectionimmunity/Decombinator
Software suite for analysis of T cell receptor repertoire data. Used for fast, efficient analysis of T cell receptor (TcR) repertoire samples, designed to be accessible to those with no previous programming experience.
Proper citation: Decombinator (RRID:SCR_006732) Copy
Database of peer-reviewed, continually updated annotation for the Pseudomonas aeruginosa PAO1 reference strain genome expanded to include all Pseudomonas species to facilitate cross-strain and cross-species genome comparisons with high quality comparative genomics. The database contains robust assessment of orthologs, a novel ortholog clustering method, and incorporates five views of the data at the sequence and annotation levels (Gbrowse, Mauve and custom views) to facilitate genome comparisons. Other features include more accurate protein subcellular localization predictions and a user-friendly, Boolean searchable log file of updates for the reference strain PAO1. The current annotation is updated using recent research literature and peer-reviewed submissions by a worldwide community of PseudoCAP (Pseudomonas aeruginosa Community Annotation Project) participating researchers. If you are interested in participating, you are invited to get involved. Many annotations, DNA sequences, Orthologs, Intergenic DNA, and Protein sequences are available for download.
Proper citation: Pseudomonas Genome Database (RRID:SCR_006590) Copy
Collection of data related to crop plant and model organism Zea mays. Used to synthesize, display, and provide access to maize genomics and genetics data, prioritizing mutant and phenotype data and tools, structural and genetic map sets, and gene models and to provide support services to the community of maize researchers. Data stored at MaizeGDB was inherited from the MaizeDB and ZmDB projects. Sequence data are from GenBank. Data are searchable by phenotype, traits, Pests, Gel Pattern, and Mutant Images.
Proper citation: MaizeGDB (RRID:SCR_006600) Copy
http://db-mml.sjtu.edu.cn/ICEberg/
ICEberg is an integrated database that provides comprehensive information about integrative and conjugative elements (ICEs) found in bacteria. ICEs are conjugative self-transmissible elements that can integrate into and excise from a host chromosome. An ICE contains three typical modules, integration and excision, conjugation, and regulation modules, that collectively promote vertical inheritance and periodic lateral gene flow. Many ICEs carry likely virulence determinants, antibiotic-resistant factors and/or genes coding for other beneficial traits. ICEberg offers a unique, highly organized, readily explorable archive of both predicted and experimentally supported ICE-relevant data. It currently contains details of 428 ICEs found in representatives of 124 bacterial species, and a collection of >400 directly related references. A broad range of similarity search, sequence alignment, genome context browser, phylogenetic and other functional analysis tools are readily accessible via ICEberg. ICEberg will facilitate efficient, multidisciplinary and innovative exploration of bacterial ICEs and be of particular interest to researchers in the broad fields of prokaryotic evolution, pathogenesis, biotechnology and metabolism. The ICEberg database will be maintained, updated and improved regularly to ensure its ongoing maximum utility to the research community.
Proper citation: ICEberg (RRID:SCR_006026) Copy
http://sourceforge.net/projects/gasic/
A method to correct read alignment results for the ambiguities imposed by similarities of genomes.
Proper citation: GASiC (RRID:SCR_006765) Copy
Integrative database of germ-line V genes from the immunoglobulin loci of human and mouse. It presents V gene sequences extracted from the EMBL nucleotide sequence database and Ensembl together with links to the respective source sequences. Based on the properties of the source sequences, V genes are classified into 3 different classes: * Class 1: genomic and rearranged evidence * Class 2: genomic evidence only * Class 3: rearranged evidence only This allows careful sequence quality validation by the user. References to other immunological databases ( KABAT, IMGT/LIGM and VBASE ) are given to provide all public annotation data for each V gene. The VBASE2 database can be accessed either by the Direct Query interface or by the DNAPLOT Query interface. The Sequences given by the user are aligned with DNAPLOT against the VBASE2 database. Direct Query allows to enter sequence IDs and names (Field 1), choose species, locus, V gene family and class (Field 2) or search for 100% sequences (Field 3). At the DNAPLOT Query, the sequences given by the user are aligned with DNAPLOT against the VBASE2 database. The DNAPLOT program offers V gene nucleotide sequence alignment referring to the IMGT V gene unique numbering. The Quick Search can be used either for Direct Query to search for sequence IDs and V gene names or for DNAPLOT Query for up to 5 sequences. The new Fab Analysis allows you to align Fab, scFab, scAb or scFv sequences with DNAPLOT against the VBASE2 database, where both heavy and light chain are analyzed.
Proper citation: VBASE2 (RRID:SCR_007082) Copy
http://mafft.cbrc.jp/alignment/server/
Software package as multiple alignment program for amino acid or nucleotide sequences. Can align up to 500 sequences or maximum file size of 1 MB. First version of MAFFT used algorithm based on progressive alignment, in which sequences were clustered with help of Fast Fourier Transform. Subsequent versions have added other algorithms and modes of operation, including options for faster alignment of large numbers of sequences, higher accuracy alignments, alignment of non-coding RNA sequences, and addition of new sequences to existing alignments.
Proper citation: MAFFT (RRID:SCR_011811) Copy
https://github.com/MikkelSchubert/adapterremoval
Software program to remove residual adapter sequences from next generation sequencing reads. Used for cleaning of next-generation sequencing reads. AdapterRemoval v2 introduces improvements in throughput, through use of single instruction, multiple data (SIMD; SSE1 and SSE2) instructions and multi-threading support; handles datasets containing reads or read-pairs with different adapters or adapter pairs; provides simultaneous demultiplexing and adapter trimming; has ability to reconstruct adapter sequences from paired-end reads for poorly documented data sets; provides native gzip and bzip2 support.
Proper citation: AdapterRemoval (RRID:SCR_011834) Copy
http://sift.bii.a-star.edu.sg/
Data analysis service to predict whether an amino acid substitution affects protein function based on sequence homology and the physical properties of amino acids. SIFT can be applied to naturally occurring nonsynonymous polymorphisms and laboratory-induced missense mutations. (entry from Genetic Analysis Software) Web service is also available.
Proper citation: SIFT (RRID:SCR_012813) Copy
http://genetics.bwh.harvard.edu/pph2/
Software tool which predicts possible impact of amino acid substitution on structure and function of human protein using straightforward physical and comparative considerations. PolyPhen-2 is new development of PolyPhen tool for annotating coding nonsynonymous SNPs.
Proper citation: PolyPhen: Polymorphism Phenotyping (RRID:SCR_013189) Copy
http://research-pub.gene.com/gmap/
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 29, 2016. A software program for mapping and aligning cDNA sequences to a genome. The program maps and aligns a single sequence with minimal startup time and memory requirements, and provides fast batch processing of large sequence sets. The program generates accurate gene structures, even in the presence of substantial polymorphisms and sequence errors, without using probabilistic splice site models. Methodology underlying the program includes a minimal sampling strategy for genomic mapping, oligomer chaining for approximate alignment, sandwich DP for splice site detection, and microexon identification with statistical significance testing.
Proper citation: GMAP (RRID:SCR_008992) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.