Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://amphoranet.pitgroup.org/
Webserver implementation of the AMPHORA2 workflow for phylogenetic analysis of metagenomic shotgun sequencing data. It is capable of assigning a probability-weighted taxonomic group for each phylogenetic marker gene found in the input metagenomic sample.
Proper citation: AmphoraNet (RRID:SCR_005009) Copy
http://www.ebi.ac.uk/biosamples/
Database that aggregates sample information for reference samples (e.g. Coriell Cell lines) and samples for which data exist in one of the EBI''''s assay databases such as ArrayExpress, the European Nucleotide Archive or PRoteomics Identificates DatabasE. It provides links to assays for specific samples, and accepts direct submissions of sample information. The goals of the BioSample Database include: # recording and linking of sample information consistently within EBI databases such as ENA, ArrayExpress and PRIDE; # minimizing data entry efforts for EBI database submitters by enabling submitting sample descriptions once and referencing them later in data submissions to assay databases and # supporting cross database queries by sample characteristics. The database includes a growing set of reference samples, such as cell lines, which are repeatedly used in experiments and can be easily referenced from any database by their accession numbers. Accession numbers for the reference samples will be exchanged with a similar database at NCBI. The samples in the database can be queried by their attributes, such as sample types, disease names or sample providers. A simple tab-delimited format facilitates submissions of sample information to the database, initially via email to biosamples (at) ebi.ac.uk. Current data sources: * European Nucleotide Archive (424,811 samples) * PRIDE (17,001 samples) * ArrayExpress (1,187,884 samples) * ENCODE cell lines (119 samples) * CORIELL cell lines (27,002 samples) * Thousand Genome (2,628 samples) * HapMap (1,417 samples) * IMSR (248,660 samples)
Proper citation: BioSample Database at EBI (RRID:SCR_004856) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 11,2023. SuperCAT hosts typing databases for the Bacillus cereus group of bacteria. The databases contain MultiLocus Sequence Typing (MLST), MultiLocus Enzyme Electrophoresis (MLEE), and Amplified Fragment Length Polymorphism (AFLP) phylogenetic data. multilocus, sequence, Bacillus cereus, bacteria, Genomics, non-vertebrate, taxonomy, identification
Proper citation: SuperCAT (RRID:SCR_004882) Copy
A clade oriented, community curated database containing genomic, genetic, phenotypic and taxonomic information for plant genomes. Genomic information is presented in a comparative format and tied to important plant model species such as Arabidopsis. SGN provides tools such as: BLAST searches, the SolCyc biochemical pathways database, a CAPS experiment designer, an intron detection tool, an advanced Alignment Analyzer, and a browser for phylogenetic trees. The SGN code and database are developed as an open source project, and is based on database schemas developed by the GMOD project and SGN-specific extensions.
Proper citation: SGN (RRID:SCR_004933) Copy
http://cgi-www.daimi.au.dk/cgi-chili/datfap/frontdoor.py
A database of transcription factors from 13 plant species, and PCR primers for around 90% of them.
Proper citation: DATFAP (RRID:SCR_005413) Copy
A publicly available database of Transposed elements (TEs) which are located within protein-coding genes of 7 organisms: human, mouse, chicken, zebrafish, fruilt fly, nematode and sea squirt. Using TranspoGene the user can learn about the many aspects of the effect these TEs have on their hosting genes, such as: exonization events (including alternative splicing-related data), insertion of TEs into introns, exons, and promoters, specific location of the TE over the gene, evolutionary divergence of the TE from its consensus sequence and involvement in diseases. TranspoGene database is quickly searchable through its website, enables many kinds of searches and is available for download. TranspoGene contains information regarding specific type and family of the TEs, genomic and mRNA location, sequence, supporting transcript accession and alignment to the TE consensus sequence. The database also contains host gene specific data: gene name, genomic location, Swiss-Prot and RefSeq accessions, diseases associated with the gene and splicing pattern. The TranspoGene and microTranspoGene databases can be used by researchers interested in the effect of TE insertion on the eukaryotic transcriptome.
Proper citation: TranspoGene (RRID:SCR_005634) Copy
http://www.gene-regulation.com/pub/databases.html#transfac
Manually curated database of eukaryotic transcription factors, their genomic binding sites and DNA binding profiles. Used to predict potential transcription factor binding sites.
Proper citation: TRANSFAC (RRID:SCR_005620) Copy
http://dynamine.ibsquare.be/submission/
An NMR based method for protein folding prediction. Users can enter a UniProt identifier, FASTA sequences, or upload a file containing FASTA sequences and results are returned., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: DynaMine (RRID:SCR_014559) Copy
http://www.cbs.dtu.dk/services/SignalP/
Web application for prediction of the presence and location of signal peptide cleavage sites in amino acid sequences from different organisms. The method incorporates a prediction of cleavage sites and a signal peptide/non-signal peptide prediction based on a combination of several artificial neural networks.
Proper citation: SignalP (RRID:SCR_015644) Copy
http://icebox.lbl.gov:8080/ApolloWebDemo/jbrowse/
WebApollo is an extensible web-based sequence annotation editor for community annotation. No software download is required and the annotations are saved to a centralized database with real-time annotation updating. (The edit server mediates annotation changes made by multiple users.) The Web based client uses JBrowse, is fast and highly interactive. WebApollo accesses many types of genomic data including access to public data from UCSC, Ensembl, and GMOD Chado databases. Source code (BSD License) * Client source code: https://github.com/berkeleybop/jbrowse * Annotation editing engine: http://code.google.com/p/apollo-web * Data model and I/O layer: http://code.google.com/p/gbol * Trellis server code: http://code.google.com/p/genomancer
Proper citation: WebApollo: A Web-Based Sequence Annotation Editor for Community Annotation (RRID:SCR_005321) Copy
http://www.ch.embnet.org/software/COILS_form.html
COILS is a program that compares a sequence to a database of known parallel two-stranded coiled-coils and derives a similarity score. By comparing this score to the distribution of scores in globular and coiled-coil proteins, the program then calculates the probability that the sequence will adopt a coiled-coil conformation.
Proper citation: COILS: Prediction of Coiled Coil Regions in Proteins (RRID:SCR_008440) Copy
Ratings or validation data are available for this resource
http://www.bioinformatics.babraham.ac.uk/projects/trim_galore/
Software tool to automate quality and adapter trimming as well as quality control, with some added functionality to remove biased methylation positions for RRBS sequence files for directional, non-directional or paired-end sequencing. Wrapper around Cutadapt and FastQC to consistently apply adapter and quality trimming to FastQ files, with extra functionality for Reduced Representation Bisulfite Sequencing data.
Proper citation: Trim Galore (RRID:SCR_011847) Copy
http://www.glycosciences.de/tools/linucs/
Service that directly converts the commonly used extended representation of complex carbohydrates into the preferred canonical description or into its inverted form. Input: A structure using the extended, non-graphic nomenclature (in ASCII writing) to describe complex carbohydrates as recommended by IUPAC. Output: A linear, unique notation. The source code (written in C), will be distributed so that software developers can easily implement their algorithm within their own application. LINUCS was chosen to fulfill to following conditions: * Input of extended, non-graphic nomenclature to describe carbohydrate structures. * Resulting linear code is closely related to notations and abbreviations recommended by IUPAC. * Number of additional rules to define the priority of the branches is low * Extended nomenclature of complex carbohydrates contains all information to define the hierarchy. * LINUCS is applicable to all types of carbohydrates (macrocyclic system are currently not implemented) . * Remaining unassigned linkage information are tolerated
Proper citation: LINUCS (RRID:SCR_001571) Copy
A curated collection of chaperonin sequence data collected from public databases or generated by a network of collaborators exploiting the cpn60 target in clinical, phylogenetic and microbial ecology studies. The database contains all available sequences for both group I and group II chaperonins. Users can search the database by Chaperonin type, group (I or II), BLAST, or other options, and can also enter and analyze FASTA sequences.
Proper citation: cpnDB: A Chaperonin Database (RRID:SCR_002263) Copy
https://github.com/OpenGene/AfterQC
Software that performs automatic filtering, trimming, error removing, and quality control for fastq data.
Proper citation: AfterQC (RRID:SCR_016390) Copy
https://github.com/dvera/albacore
Data processing basecaller for the Oxford Nanopore sequencer that identifies DNA sequences directly from raw data. It enhances accuracy of the single-read sequence data, contributing to high consensus accuracy for nanopore sequence data.
Proper citation: Albacore (RRID:SCR_015897) Copy
https://github.com/isovic/racon
Software tool as de novo genome assembly from long uncorrected reads. Used to correct raw contigs generated by rapid assembly methods which do not include consensus step. Supports data produced by Pacific Biosciences and Oxford Nanopore Technologies.
Proper citation: Racon (RRID:SCR_017642) Copy
https://github.com/TransDecoder/TransDecoder
Software tool to identify candidate coding regions within transcript sequences, such as those generated by de novo RNA-Seq transcript assembly using Trinity, or constructed based on RNA-Seq alignments to genome using Tophat and Cufflinks.Starts from FASTA or GFF file. Can scan and retain open reading frames (ORFs) for homology to known proteins by using BlastP or Pfam search and incorporate results into obtained selection. Predictions can then be visualized by using genome browser such as IGV.
Proper citation: TransDecoder (RRID:SCR_017647) Copy
https://www.sanger.ac.uk/science/tools/reapr
Software tool to identify errors in genome assemblies without need for reference sequence. Can be used in any stage of assembly pipeline to automatically break incorrect scaffolds and flag other errors in assembly for manual inspection. Reports mis-assemblies and other warnings, and produces new broken assembly based on error calls.
Proper citation: Recognition of Errors in Assemblies using Paired Reads (RRID:SCR_017625) Copy
https://github.com/lufuhao/ExonerateTransferAnnotation
Software tool as pipeline to make anntotations using cDNA and CDS sequences.
Proper citation: ExonerateTransferAnnotation (RRID:SCR_017557) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.