Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://compbio.cs.sfu.ca/software-novelseq
Software pipeline to detect novel sequence insertions using high throughput paired-end whole genome sequencing data.
Proper citation: NovelSeq (RRID:SCR_003136) Copy
http://mrcanavar.sourceforge.net/
Copy number caller that analyzes the whole-genome next-generation sequence mapping read depth to discover large segmental duplications and deletions. It also has the capability of predicting absolute copy numbers of genomic intervals.
Proper citation: mrCaNaVaR (RRID:SCR_003135) Copy
http://www.ichip.de/software/SplicingCompass.html
Software for detection of differential splicing between two different conditions using RNA-Seq data.
Proper citation: SplicingCompass (RRID:SCR_003249) Copy
http://abi.inf.uni-tuebingen.de/Services/MultiLoc2
An extensive high-performance subcellular protein localization prediction system that incorporates phylogenetic profiles and Gene Ontology terms to yield higher accuracies compared to its previous version. Moreover, it outperforms other prediction systems in two benchmarks studies. A downloadable version of MultiLoc2 for local use is also available.
Proper citation: MultiLoc (RRID:SCR_003151) Copy
Tool used to design PCR primers from DNA sequence - often in high-throughput genomics applications. It does everything from mispriming libraries to sequence quality data to the generation of internal oligos.
Proper citation: Primer3 (RRID:SCR_003139) Copy
Database enables integration of genomic and phenomic data by providing access to primary experimental data, data collection protocols and analysis tools. Data represent behavioral, morphological and physiological disease-related characteristics in naive mice and those exposed to drugs, environmental agents or other treatments. Collaborative standardized collection of measured data on laboratory mouse strains to characterize them in order to facilitate translational discoveries and to assist in selection of strains for experimental studies. Includes baseline phenotype data sets as well as studies of drug, diet, disease and aging effect., protocols, projects and publications, and SNP, variation and gene expression studies. Provides tools for online analysis. Data sets are voluntarily contributed by researchers from variety of institutions and settings, or retrieved by MPD staff from open public sources. MPD has three major types of strain-centric data sets: phenotype strain surveys, SNP and variation data, and gene expression strain surveys. MPD collects data on classical inbred strains as well as any fixed-genotype strains and derivatives that are openly acquirable by the research community. New panels include Collaborative Cross (CC) lines and Diversity Outbred (DO) populations. Phenotype data include measurements of behavior, hematology, bone mineral density, cholesterol levels, endocrine function, aging processes, addiction, neurosensory functions, and other biomedically relevant areas. Genotype data are primarily in the form of single-nucleotide polymorphisms (SNPs). MPD curates data into a common framework by standardizing mouse strain nomenclature, standardizing units (SI where feasible), evaluating data (completeness, statistical power, quality), categorizing phenotype data and linking to ontologies, conforming to internal style guides for titles, tags, and descriptions, and creating comprehensive protocol documentation including environmental parameters of the test animals. These elements are critical for experimental reproducibility.
Proper citation: Mouse Phenome Database (MPD) (RRID:SCR_003212) Copy
https://github.com/brunonevado/Pipeliner
Software for evaluating the performance of bioinformatics pipelines for Next Generation re-Sequencing.
Proper citation: Pipeliner (RRID:SCR_003171) Copy
Database to catalog experimentally determined interactions between proteins combining information from a variety of sources to create a single, consistent set of protein-protein interactions that can be downloaded in a variety of formats. The data were curated, both, manually and also automatically using computational approaches that utilize the the knowledge about the protein-protein interaction networks extracted from the most reliable, core subset of the DIP data. Because the reliability of experimental evidence varies widely, methods of quality assessment have been developed and utilized to identify the most reliable subset of the interactions. This CORE set can be used as a reference when evaluating the reliability of high-throughput protein-protein interaction data sets, for development of prediction methods, as well as in the studies of the properties of protein interaction networks. Tools are available to analyze, visualize and integrate user's own experimental data with the information about protein-protein interactions available in the DIP database. The DIP database lists protein pairs that are known to interact with each other. By interact they mean that two amino acid chains were experimentally identified to bind to each other. The database lists such pairs to aid those studying a particular protein-protein interaction but also those investigating entire regulatory and signaling pathways as well as those studying the organization and complexity of the protein interaction network at the cellular level. Registration is required to gain access to most of the DIP features. Registration is free to the members of the academic community. Trial accounts for the commercial users are also available.
Proper citation: Database of Interacting Proteins (DIP) (RRID:SCR_003167) Copy
Software R-package for running gene set analysis using various statistical methods, from different gene level statistics and a wide range of gene-set collections. The Piano package contains functions for combining the results of multiple runs of gene set analyses.
Proper citation: Piano (RRID:SCR_003200) Copy
http://cmb.molgen.mpg.de/2ndGenerationSequencing/Solas/
Software package for the statistical language R, devoted to the analysis of next generation short read data of RNA-seq transcripts. It provides predictions of alternative exons in a single condition/cell sample, predictions of differential alternative exons between two conditions/cell samples, and quantification of alternative splice forms in a single condition/cell sample.
Proper citation: Solas (RRID:SCR_003168) Copy
http://www.broadinstitute.org/cancer/software/genepattern
A powerful genomic analysis platform that provides access to hundreds of tools for gene expression analysis, proteomics, SNP analysis, flow cytometry, RNA-seq analysis, and common data processing tasks. A web-based interface provides easy access to these tools and allows the creation of multi-step analysis pipelines that enable reproducible in silico research.
Proper citation: GenePattern (RRID:SCR_003201) Copy
http://pir.georgetown.edu/pirwww/dbinfo/pirsf.shtml
A SuperFamily classification system, with rules for functional site and protein name, to facilitate the sensible propagation and standardization of protein annotation and the systematic detection of annotation errors. The PIRSF concept is being used as a guiding principle to provide comprehensive and non-overlapping clustering of UniProtKB sequences into a hierarchical order to reflect their evolutionary relationships. The PIRSF classification system is based on whole proteins rather than on the component domains; therefore, it allows annotation of generic biochemical and specific biological functions, as well as classification of proteins without well-defined domains. There are different PIRSF classification levels. The primary level is the homeomorphic family, whose members are both homologous (evolved from a common ancestor) and homeomorphic (sharing full-length sequence similarity and a common domain architecture). At a lower level are the subfamilies which are clusters representing functional specialization and/or domain architecture variation within the family. Above the homeomorphic level there may be parent superfamilies that connect distantly related families and orphan proteins based on common domains. Because proteins can belong to more than one domain superfamily, the PIRSF structure is formally a network. The FTP site provides free download for PIRSF.
Proper citation: PIRSF (RRID:SCR_003352) Copy
A tool for creating logos representing both sequence alignments and profile hidden Markov models. The interactive logos enable scrolling, zooming, and inspection of underlying values. Skylign can avoid sampling bias in sequence alignments by down-weighting redundant sequences and by combining observed counts with informed priors. It also simplifies the representation of gap parameters, and can optionally scale letter heights based on alternate calculations of the conservation of a position.
Proper citation: Skylign (RRID:SCR_001176) Copy
https://rdrr.io/bioc/yaqcaffy/
Software package for quality control of Affymetrix GeneChip expression data and reproducibility analysis of human whole genome chips with the MAQC reference datasets.
Proper citation: yaqcaffy (RRID:SCR_001295) Copy
http://julian-gehring.github.io/les/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 23,2022. Software package that estimates Loci of Enhanced Significance (LES) in tiling microarray data. These are regions of regulation such as found in differential transcription, CHiP-chip, or DNA modification analysis. The package provides a universal framework suitable for identifying differential effects in tiling microarray data sets, and is independent of the underlying statistics at the level of single probes.
Proper citation: les (RRID:SCR_001291) Copy
http://ccb.jhu.edu/software/sim4cc/
Software tool as cross species spliced alignment program.Heuristic sequence alignment tool for comparing cDNA sequence with genomic sequence containing homolog of gene in another species.
Proper citation: sim4cc (RRID:SCR_001204) Copy
http://qualimap.bioinfo.cipf.es/
Software application written in Java and R that provides both a Graphical User Inteface (GUI) and a command-line interface to facilitate the quality control of alignment sequencing data. It examines sequencing alignment data in SAM / BAM files according to the features of the mapped reads and provides an overall view of the data that helps to the detect biases in the sequencing and/or mapping of the data and eases decision-making for further analysis.
Proper citation: QualiMap (RRID:SCR_001209) Copy
http://pathology.wustl.edu/VirusHunter/
A fully automated and modular software package for mining sequence data to identify sequences of microbial origin. The pipeline was optimized for analysis of data generated by the Roche/454 next-generation sequencing platform but can be applied to longer sequences (Sanger sequencing data or assembled contigs) as well. Microbial sequences are identified on the basis of BLAST alignments and the taxonomic classification of the reference sequence(s) to which a read is aligned. Viruses are the focal point of VirusHunter as released, but it can be easily modified to generate parallel outputs for bacterial or parasitic species. To date, VirusHunter has been applied to thousands of specimens, including human, animal and environmental samples, resulting in the detection of many known and novel viruses.
Proper citation: VirusHunter (RRID:SCR_001198) Copy
http://www.biobase-international.com/product/genome-trax
Service that provides a comprehensive compilation of variant knowledge that allows you to identify pathogenic variants in human whole genome or exome sequences. It makes it easy to upload a complete genome?s worth of variations and identify the biologically relevant subset of known mutations, mutations that are novel and appear in a candidate disease genes, or mutations that are predicted to have a deleterious effect. The database includes a comprehensive collection of disease causing mutations from HGMD Professional, regulatory sites from TRANSFAC , and disease genes, drug targets and pathways from PROTEOME, as well as pharmacogenomic variants. It integrates the best public data-sets on somatic mutations, allele frequencies and clinical variants, in their most up-to-date version, for a total of more than 165 million annotations. It is possible to identify known pathogenic variants, remove harmless common variants, and obtain deleterious predictions for novel variants. With family data, it is possible to identify variants that are de novo, compound heterozygous only in the offspring. All of the results can be downloaded to Excel for further review. For core facilities and bioinformaticians, the complete underlying data is made available for download and easy integration into custom analysis pipelines. Genome Trax data is optimized to work with many other software packages, such as ANNOVARTM, CLC bio, Alamut, SimulConsult, and Cartagenia.
Proper citation: Genome Trax (RRID:SCR_001234) Copy
DNA motif discovery software adapted for ChIP-Seq data. It is an iterative algorithm that combines greedy optimization with bootstrapping and uses coverage profiles as motif positional preferences. It does not require truncation of long DNA segments and it is practical for processing up to tens of thousands of data sequences
Proper citation: ChIPMunk (RRID:SCR_001191) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.