Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://pathogenseq.lshtm.ac.uk/estmoi
A per-based software to estimate multiplicity of infection (MOI) in parasite genomic sequence data. It is primarily developed to address the limitations of current laboratory (PCR) based estimates of multiplicity using high throughput sequence data. It requires a BAM (alignment output of short reads to the reference genome), VCF (a file with information on variant calls) and FASTA (reference genome) files. # Short reads are aligned to a reference genome using BWA, BOWTIE, SMALT or other short read aligners to generate a BAM file. # Single Nucleotide Polymorphisms (SNPs) are then identified using SAMTools/BCFtools and stored in the VCF format. # The reference FASTA file is expected to be indexed using ''samtools faidx'' to generate a *.fai file. estMOI generates files containing MOI estimates for each SNP combinations (file with name *.log) and a summary for all chromosomes (file with name *.txt).
Proper citation: estMOI (RRID:SCR_006192) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on July 7, 2022. Federation of International Mouse Resources (FIMRe) is a collaborating group of Mouse Repository and Resource Centers worldwide whose collective goal is to archive and provide strains of mice as cryopreserved embryos and gametes, ES cell lines, and live breeding stock to the research community. Goals of the Federation of International Mouse Resources: * Coordinate repositories and resource centers to: ** archive valuable genetically defined mice and ES cell lines being created worldwide ** meet research demand for these genetically defined mice and ES cell lines * Establish consistent, highest quality animal health standards in all resource centers * Provide genetic verification and quality control for genetic background and mutations * Provide resource training to enhance user ability to utilize cryopreserved resources
Proper citation: Federation of International Mouse Resources (RRID:SCR_006137) Copy
http://www.nervenet.org/main/dictionary.html
A mouse-related portal of genomic databases and tables of mouse brain data. Most files are intended for you to download and use on your own personal computer. Most files are available in generic text format or as FileMaker Pro databases. The server provides data extracted and compiled from: The 2000-2001 Mouse Chromosome Committee Reports, Release 15 of the MIT microsatellite map (Oct 1997), The recombinant inbred strain database of R.W. Elliott (1997) and R. W. Williams (2001), and the Map Manager and text format chromosome maps (Apr 2001). * LXS genotype (Excel file): Updated, revised positions for 330 markers genotyped using a panel of 77 LXS strain. * MIT SNP DATABASE ONLINE: Search and sort the MIT Single Nucleotide Polymorphism (SNP) database ONLINE. These data from the MIT-Whitehead SNP release of December 1999. * INTEGRATED MIT-ROCHE SNP DATABASE in EXCEL and TEXT FORMATS (1-3 MB): Original MIT SNPs merged with the new Roche SNPs. The Excel file has been formatted to illustrate SNP haplotypes and genetic contrasts. Both files are intended for statistical analyses of SNPs and can be used to test a method outlined in a paper by Andrew Grupe, Gary Peltz, and colleagues (Science 291: 1915-1918, 2001). The Excel file includes many useful equations and formatting that will help in navigating through this large database and in testing the in silico mapping method. * Use of inbred strains for the study of individual differences in pain related phenotypes in the mouse: Elissa J. Chesler''s 2002 dissertation, discussing issues relevant to the integration of genomic and phenomic data from standard inbred strains including genetic interactions with laboratory environmental conditions and the use of various in silico inbred strain haplotype based mapping algorithms for QTL analysis. * SNP QTL MAPPER in EXCEL format (572 KB, updated January 2002 by Elissa Chesler): This Excel workbook implements the Grupe et al. mapping method and outputs correlation plots. The main spreadsheet allows you to enter your own strain data and compares them to haplotypes. Be very cautious and skeptical when using this spreadsheet and the technique. Read all of the caveates. This excel version of the method was developed by Elissa Chesler. This updated version (Jan 2002) handles missing data. * MIT SNP Database (tab-delimited text format): This file is suitable for manipulation in statistics and spreadsheet programs (752 KB, Updated June 27, 2001). Data have been formatted in a way that allows rapid acquisition of the new data from the Roche Bioscience SNP database. * MIT SNP Database (FileMaker 5 Version): This is a reformatted version of the MIT Single Nucleotide Polymorphism (SNP) database in FileMaker 5 format. You will need a copy of this application to open the file (Mac and Windows; 992 KB. Updated July 13, 2001 by RW). * Gene Mapping and Map Manager Data Sets: Genetic maps of mouse chromosomes. Now includes a 10th generation advanced intercross consisting of 500 animals genetoyped at 340 markers. Lots of older files on recombinant inbred strains. * The Portable Dictionary of the Mouse Genome, 21,039 loci, 17,912,832 bytes. Includes all 1997-98 Chromosome Committee Reports and MIT Release 15. * FullDict.FMP.sit: The Portable Dictionary of the Mouse Genome. This large FileMaker Pro 3.0/4.0 database has been compressed with StuffIt. The Dictionary of the Mouse Genome contains data from the 1997-98 chromosome committee reports and MIT Whitehead SSLP databases (Release 15). The Dictionary contains information for 21,039 loci. File size = 4846 KB. Updated March 19, 1998. * MIT Microsatellite Database ONLINE: A database of MIT microsatellite loci in the mouse. Use this FileMaker Pro database with OurPrimersDB. MITDB is a subset of the Portable Dictionary of the Mouse Genome. ONLINE. Updated July 12, 2001. * MIT Microsatellite Database: A database of MIT microsatellite loci in the mouse. Use this FileMaker Pro database with OurPrimersDB. MITDB is a subset of the Portable Dictionary of the Mouse Genome. File size = 3.0 MB. Updated March 19, 1998. * OurPrimersDB: A small database of primers. Download this database if you are using numerous MIT primers to map genes in mice. This database should be used in combination with the MITDB as one part of a relational database. File size = 149 KB. Updated March 19, 1998. * Empty copy (clone) of the Portable Dictionary in FileMaker Pro 3.0 format. Download this file and import individual chromosome text files from the table into the database. File size = 231 KB. Updated March 19, 1998. * Chromosome Text Files from the Dictionary: The table lists data on gene loci for individual chromosomes.
Proper citation: Mouse Genome Databases (RRID:SCR_007147) Copy
http://www.genoscope.cns.fr/externe/tetraodon/
The initial objective of Genoscope was to compare the genomic sequences of this fish to that of humans to help in the annotation of human genes and to estimate their number. This strategy is based on the common genetic heritage of the vertebrates: from one species of vertebrate to another, even for those as far apart as a fish and a mammal, the same genes are present for the most part. In the case of the compact genome of Tetraodon, this common complement of genes is contained in a genome eight times smaller than that of humans. Although the length of the exons is similar in these two species, the size of the introns and the intergenic sequences is greatly reduced in this fish. Furthermore, these regions, in contrast to the exons, have diverged completely since the separation of the lineages leading to humans and Tetraodon. The Exofish method, developed at Genoscope, exploits this contrast such that the conserved regions which can be identified by comparing genomic sequences of the two species, correspond only to coding regions. Using preliminary sequencing results of the genome of Tetraodon in the year 2000, Genoscope evaluated the number of human genes at about 30,000, whereas much higher estimations were current. The progress of the annotation of the human genome has since supported the Genoscope hypothesis, with values as low as 22,000 genes and a consensus of around 25,000 genes. The sequencing of the Tetraodon genome at a depth of about 8X, carried out as a collaboration between Genoscope and the Whitehead Institute Center for Genome Research (now the Broad Institute), was finished in 2002, with the production of an assembly covering 90 of the euchromatic region of the genome of the fish. This has permitted the application of Exofish at a larger scale in comparisons with the genome of humans, but also with those of the two other vertebrates sequenced at the time (Takifugu, a fish closely related to Tetraodon, and the mouse). The conserved regions detected in this way have been integrated into the annotation procedure, along with other resources (cDNA sequences from Tetraodon and ab initio predictions). Of the 28,000 genes annotated, some families were examined in detail: selenoproteins, and Type 1 cytokines and their receptors. The comparison of the proteome of Tetraodon with those of mammals has revealed some interesting differences, such as a major diversification of some hormone systems and of the collagen molecules in the fish. A search for transposable elements in the genomic sequences of Tetraodon has also revealed a high diversity (75 types), which contrasts with their scarcity; the small size of the Tetraodon genome is due to the low abundance of these elements, of which some appear to still be active. Another factor in the compactness of the Tetraodon genome, which has been confirmed by annotation, is the reduction in intron size, which approaches a lower limit of 50-60 bp, and which preferentially affects certain genes. The availability of the sequences from the genomes of humans and mice on one hand, and Takifugu and Tetraodon on the other, provide new opportunities for the study of vertebrate evolution. We have shown that the level of neutral evolution is higher in fish than in mammals. The protein sequences of fish also diverge more quickly than those of mammals. A key mechanism in evolution is gene duplication, which we have studied by taking advantage of the anchoring of the majority of the sequences from the assembly on the chromosomes. The result of this study speaks strongly in favor of a whole genome duplication event, very early in the line of ray-finned fish (Actinopterygians). An even stronger evidence came from synteny studies between the genomes of humans and Tetraodon. Using a high-resolution synteny map, we have reconstituted the genome of the vertebrate which predates this duplication - that is, the last common ancestor to all bony vertebrates (most of the vertebrates apart from cartilaginous fish and agnaths like lamprey). This ancestral karyotype contains 12 chromosomes, and the 21 Tetraodon chromosomes derive from it by the whole genome duplication and a surprisingly small number of interchromosomal rearrangements. On the contrary, exchanges between chromosomes have been much more frequent in the lineage that leads to humans. Sponsors: The project was supported by the Consortium National de Recherche en Genomique and the National Human Genome Research Institute.
Proper citation: Tetraodon Genome Browser (RRID:SCR_007079) Copy
Database containing the DNA sequence and annotation of the entire human chromosome 7, encompassing nearly 158 million nucleotides of DNA and 1917 gene structures, are presented; the most up to date collation of sequence, gene, and other annotations from all databases (eg. Celera published, NCBI, Ensembl, RIKEN, UCSC) as well as unpublished data. To generate a higher order description, additional structural features such as imprinted genes, fragile sites, and segmental duplications were integrated at the level of the DNA sequence with medical genetic data, including 440 chromosome rearrangement breakpoints associated with disease. The objective of this project is to generate a comprehensive description of human chromosome 7 to facilitate biological discovery, disease gene research and medical genetic applications. There are over 360 disease-associated genes or loci on chromosome 7. A major challenge ahead will be to represent chromosome alterations, variants, and polymorphisms and their related phenotypes (or lack thereof), in an accessible way. In addition to being a primary data source, this site serves as a weighing station for testing community ideas and information to produce highly curated data to be submitted to other databases such as NCBI, Ensembl, and UCSC. Therefore, any useful data submitted will be curated and shown in this database. All Chromosome 7 genomic clones (cosmids, BACs, YACs) listed in GBrowser and in other data tables are freely distributed.
Proper citation: Chromosome 7 Annotation Project (RRID:SCR_007134) Copy
Next generation sequencing and genotyping services provided to investigators working to discover genes that contribute to disease. On-site statistical geneticists provide insight into analysis issues as they relate to study design, data production and quality control. In addition, CIDR has a consulting agreement with the University of Washington Genetics Coordinating Center (GCC) to provide statistical and analytical support, most predominantly in the areas of GWAS data cleaning and methods development. Completed studies encompass over 175 phenotypes across 530 projects and 620,000 samples. The impact is evidenced by over 380 peer-reviewed papers published in 100 journals. Three pathways exist to access the CIDR genotyping facility: * NIH CIDR Program: The CIDR contract is funded by 14 NIH Institutes and provides genotyping and statistical genetic services to investigators approved for access through competitive peer review. An application is required for projects supported by the NIH CIDR Program. * The HTS Facility: The High Throughput Sequencing Facility, part of the Johns Hopkins Genetic Resources Core Facility, provides next generation sequencing services to internal JHU investigators and external scientists on a fee-for-service basis. * The JHU SNP Center: The SNP Center, part of the Johns Hopkins Genetic Resources Core Facility, provides genotyping to internal JHU investigators and external scientists on a fee-for-service basis. Data computation service is included to cover the statistical genetics services provided for investigators seeking to identify genes that contribute to human disease. Human Genotyping Services include SNP Genome Wide Association Studies, SNP Linkage Scans, Custom SNP Studies, Cancer Panel, MHC Panels, and Methylation Profiling. Mouse Genotyping Services include SNP Scans and Custom SNP Studies.
Proper citation: Center for Inherited Disease Research (RRID:SCR_007339) Copy
https://www.mc.vanderbilt.edu/victr/dcc/projects/acc/index.php/Main_Page
A national consortium formed to develop, disseminate, and apply approaches to research that combine DNA biorepositories with electronic medical record (EMR) systems for large-scale, high-throughput genetic research. The consortium is composed of seven member sites exploring the ability and feasibility of using EMR systems to investigate gene-disease relationships. Themes of bioinformatics, genomic medicine, privacy and community engagement are of particular relevance to eMERGE. The consortium uses data from the EMR clinical systems that represent actual health care events and focuses on ethical issues such as privacy, confidentiality, and interactions with the broader community.
Proper citation: eMERGE Network: electronic Medical Records and Genomics (RRID:SCR_007428) Copy
Resource for experimentally validated human and mouse noncoding fragments with gene enhancer activity as assessed in transgenic mice. Most of these noncoding elements were selected for testing based on their extreme conservation in other vertebrates or epigenomic evidence (ChIP-Seq) of putative enhancer marks. Central public database of experimentally validated human and mouse noncoding fragments with gene enhancer activity as assessed in transgenic mice. Users can retrieve elements near single genes of interest, search for enhancers that target reporter gene expression to particular tissue, or download entire collections of enhancers with defined tissue specificity or conservation depth.
Proper citation: VISTA Enhancer Browser (RRID:SCR_007973) Copy
Central repository for high quality frequently updated manual annotation of vertebrate finished genome sequence. Human, mouse and zebrafish are in the process of being completely annotated, whereas for other species the annotation is only of specific genomic regions of particular biological interest. The majority of the annotation is from the HAVANA group at the Welcome Trust Sanger Institute. Users can BLAST, search for specific text, export, and download data. Genomes and details of the projects for each species are available through the homepages for human mouse and zebrafish. The website is built upon code from the EnsEMBL (http://www.ensembl.org) project. Some Ensembl features are not available in Vega. From the users point of view perhaps the most significant of these is MartView. However due to their inclusion in Ensembl, Vega human and mouse data can be queried using Ensembl MartView. Vega contains annotation of the human MHC region in eight haplotypes, and the LRC region in three haplotypes. Vega also contains annotation on the Insulin Dependent Diabetes (IDD) regions on non-reference assemblies for mouse.
Proper citation: VEGA (RRID:SCR_007907) Copy
The Rfam database is a collection of RNA families, each represented by multiple sequence alignments, consensus secondary structures and covariance models (CMs). The families in Rfam break down into three broad functional classes: Non-coding RNA genes, structured cis-regulatory elements and self-splicing RNAs. Typically these functional RNAs often have a conserved secondary structure which may be better preserved than the RNA sequence. The CMs used to describe each family are a slightly more complicated relative of the profile hidden Markov models (HMMs) used by Pfam. CMs can simultaneously model RNA sequence and the structure in an elegant and accurate fashion. Rfam is also available via FTP. You can find data in Rfam in various ways... * Analyze your RNA sequence for Rfam matches * View Rfam family annotation and alignments * View Rfam clan details * Query Rfam by keywords * Fetch families or sequences by NCBI taxonomy * Enter any type of accession or ID to jump to the page for a Rfam family, sequence or genome
Proper citation: Rfam (RRID:SCR_007891) Copy
The web portal provides comprehensive local database of human genome variants with a user-friendly web page that provides a one-stop annotating and funtonal prediction service which is both convenient and up-to-date. A query can be accepted as either a dbSNP Id or a chromosomal location and our system will instantly provide all the annotation information in an interactive LD panel. The system can also simultaneously prioritize this variant based on additive effect mode by corresponding annotation information and evaluate the variant effect that is then displayed in a prioritization tree. Furthermore, cohort sequencing continuously produces lots of un-annotated variants such as rare variants or de novo variants, and our system can even fit this data by accepting genomic coordinates (hg19) to offer maximal annotations. Main Functions Over 40 up-to-date annotation items for human single nucleotide variations; Functional prediction for different types of variants; Dynamic LD panel for both HapMap and 1000 Genomes Project populations; Prioritization score and tree viewer based on variant functional model.
Proper citation: SNVrap (RRID:SCR_010512) Copy
http://www.cbil.upenn.edu/cgi-bin/tess/tess
TESS is a web tool for predicting transcription factor binding sites in DNA sequences. It can identify binding sites using site or consensus strings and positional weight matrices from the TRANSFAC, JASPAR, IMD, and our CBIL-GibbsMat database. You can use TESS to search a few of your own sequences or for user-defined CRMs genome-wide near genes throughout genomes of interest. Search for CRMs Genome-wide: TESS now has the ability to search whole genomes for user defined CRMs. Try a search in the AnGEL CRM Searches section of the navigation bar.. You can search for combinations of consensus site sequences and/or PWMs from TRANSFAC or JASPAR. Search DNA for Binding Sites: TESS also lets you search through your own sequence for TFBS. You can include your own site or consensus strings and/or weight matrices in the search. Use the Combined Search under ''Site Searches'' in the menu or use the box for a quick search. TESS assigns a TESS job number to all sequence search jobs. The job results are stored on our server for a period of time specified in the search submit form. During this time you may recall the search results using the form on this page. TESS can also email results to you as a tab-delimited file suitable for loading into a spreadsheet program. Query for Transcription Factor Info: TESS also has data browsing and querying capabilities to help you learn about the factors that were predicted to bind to your sequence. Use the Query TRANSFAC or Query Matrices links above or use the search interface provided from the home page.
Proper citation: TESS: Transcription Element Search System (RRID:SCR_010739) Copy
http://svdetect.sourceforge.net/Site/Home.html
Software application for the isolation and the type prediction of intra- and inter-chromosomal rearrangements from paired-end/mate-pair sequencing data provided by the high-throughput sequencing technologies. This tool aims to identify structural variations with both clustering and sliding-window strategies, and helping in their visualization at the genome scale. It is compatible with SOLiD and Illumina (>=1.3) reads.
Proper citation: SVDetect (RRID:SCR_010812) Copy
http://bio-bwa.sourceforge.net/
Software for aligning sequencing reads against large reference genome. Consists of three algorithms: BWA-backtrack, BWA-SW and BWA-MEM. First for sequence reads up to 100bp, and other two for longer sequences ranged from 70bp to 1Mbp.
Proper citation: BWA (RRID:SCR_010910) Copy
http://www.animalgenome.org/pig/genome/db/
Database facilitating information integration and mining within the pig and across species of all genomics / genetics research results accumulated over the years including pig gene expression, quantitative trait loci (QTL), candidate gene, and whole genome association study (WGAS) results. The key functions developed so far include pig gene pages (a centralized gene search tool), a local copy of Biomart (for customizable genome information queries), genome feature alignment tools (Pig QTLdb and Gbrowse), integrated gene expression information (ANEXDB and ESTdb), a dedicated pig genome and gene set BLAST server, and virtual comparative map database and tools (VCmap). By developing the PGD, it is our aim to collaboratively utilize existing databases and tools via networked functions, such as web services, database API, etc., to maximize the potential of all related databases through the PGD implementation.
Proper citation: Pig Genome Database (RRID:SCR_006367) Copy
A comparative platform for green plant genomics. Families of orthologous and paralogous genes that represent the modern descendents of ancestral gene sets are constructed at key phylogenetic nodes. These families allow easy access to clade specific orthology / paralogy relationships as well as clade specific genes and gene expansions. As of release v9.1, Phytozome provides access to forty-one sequenced and annotated green plant genomes which have been clustered into gene families at 20 evolutionarily significant nodes. Where possible, each gene has been annotated with PFAM, KOG, KEGG, and PANTHER assignments, and publicly available annotations from RefSeq, UniProt, TAIR, JGI are hyper-linked and searchable., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Phytozome (RRID:SCR_006507) Copy
We at NRSP-8 bioinformatics coordination program strive to serve the animal genomics research community to better use computer tools and methods, to best utilize available resources, and in working with researchers in the community, to effectively share, combine, manage, manipulate, and analyze information from genomics/genetics studies. This site is designed as an information center to serve the national animal genome research projects of cattle, chicken, pigs, sheep, horse, and aquaculture species. This is home to databases and web sites (being) built for structural, functional and application oriented studies of the animal genomics, to serve the purpose of research, education and related activities in the scientific, industrial and educational communities in the states and world wide. The challenges in bioinformatics support/research for animal genomics may involve * Effective data collection, organization and management * Rapid development of most needed bioinformatics tools and resources * Efficient use of these tools for innovative data analysis Projects: * Animal Trait Ontology (ATO) Project * Virtual Comparative Genomics * The Past, the Current, and the Potentials * Collaborative and Hosted Works
Proper citation: NAGRP Bioinformatics Coordination Program (RRID:SCR_006564) Copy
http://www.ncbi.nlm.nih.gov/projects/genome/assembly/grc/
Consortium that puts sequences into a chromosome context and provides the best possible reference assembly for human, mouse, and zebrafish via FTP. Tools to facilitate the curation of genome assemblies based on the sequence overlaps of long, high quality sequences.
Proper citation: Genome Reference Consortium (RRID:SCR_006553) Copy
http://www.geenivaramu.ee/en/tools/gwama
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software tool for meta analysis of whole genome association data.
Proper citation: GWAMA (RRID:SCR_006624) Copy
Model organism database for the social amoeba Dictyostelium discoideum that provides the biomedical research community with integrated, high quality data and tools for Dictyostelium discoideum and related species. dictyBase houses the complete genome sequence, ESTs, and the entire body of literature relevant to Dictyostelium. This information is curated to provide accurate gene models and functional annotations, with the goal of fully annotating the genome to provide a ''''reference genome'''' in the Amoebozoa clade. They highlight several new features in the present update: (i) new annotations; (ii) improved interface with web 2.0 functionality; (iii) the initial steps towards a genome portal for the Amoebozoa; (iv) ortholog display; and (v) the complete integration of the Dicty Stock Center with dictyBase. The Dicty Stock Center currently holds over 1500 strains targeting over 930 different genes. There are over 100 different distinct amoebozoan species. In addition, the collection contains nearly 600 plasmids and other materials such as antibodies and cDNA libraries. The strain collection includes: * strain catalog * natural isolates * MNNG chemical mutants * tester strains for parasexual genetics * auxotroph strains * null mutants * GFP-labeled strains for cell biology * plasmid catalog The Dicty Stock Center can accept Dictyostelium strains, plasmids, and other materials relevant for research using Dictyostelium such as antibodies and cDNA or genomic libraries.
Proper citation: Dictyostelium discoideum genome database (RRID:SCR_006643) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.