Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
https://www.ncbi.nlm.nih.gov/orffinder
Software tool to search for open reading frames (ORFs) in the DNA sequence. The program returns the range of each ORF, along with its protein translation. Used to search newly sequenced DNA for potential protein encoding segments, verify predicted protein. Limited to the subrange of the query sequence up to 50 kb long.
Proper citation: Open Reading Frame Finder (RRID:SCR_016643) Copy
https://www.ncbi.nlm.nih.gov/IEB/ToolBox/CPP_DOC/lxr/source/scripts/projects/magicblast/
Software tool for making a BLAST database from various data structures. It maps large sets of next-generation RNA or DNA sequencing runs against a whole genome or transcriptome.
Proper citation: MagicBlast (RRID:SCR_015513) Copy
http://www.ncbi.nlm.nih.gov/Structure/cblast/cblast.cgi?
The NCBI Related Structures tool allows you to find 3D structures from the Molecular Modeling Database (MMDB) that are similar in sequence to a query protein. Although the query protein may not yet have a resolved structure, the 3D shape of a similar protein sequence can shed light on the putative shape and biological function of the query protein. CBLAST is a tool that compares a query protein sequence against all protein sequences from resolved 3D structures by using protein BLAST against the PDB data set. The purpose is to find representative 3D structures for the query and/or its homologs, as available. Each record in the Entrez Protein database has been CBLAST''ed and the search results are available as Related Structures in the Links menu of Entrez Protein records. You can also enter a protein query sequence directly into the CBLAST search page in order to find its sequence-similar 3D structure records. The search results can be viewed in Cn3D (hence the name CBLAST), which displays an alignment of the query protein to the related structure''s sequence and allows you to interactively examine the sequence-structure relationship.
Proper citation: CBLAST (RRID:SCR_004711) Copy
http://www.ncbi.nlm.nih.gov/pcsubstance?db=pcsubstance
As one of three primary databases of PubChem (Pcsubstance, Pccompound, and PCBioAssay), PubChem Substance Database contains descriptions of chemical samples, from a variety of sources, and links to PubMed citations, protein 3D structures, and biological screening results that are available in PubChem BioAssay. If the contents of a chemical sample are known, the description includes links to PubChem Compound. A PubChem FTP is available and new data is accepted into the repository. Pcsubstance contains more than 81 million records (2011).
Proper citation: PubChem Substance (RRID:SCR_004742) Copy
http://www.ncbi.nlm.nih.gov/biosystems/
Database that provides access to biological systems and their component genes, proteins, and small molecules, as well as literature describing those biosystems and other related data throughout Entrez. A biosystem, or biological system, is a group of molecules that interact directly or indirectly, where the grouping is relevant to the characterization of living matter. BioSystem records list and categorize components, such as the genes, proteins, and small molecules involved in a biological system. The companion FLink tool, in turn, allows you to input a list of proteins, genes, or small molecules and retrieve a ranked list of biosystems. A number of databases provide diagrams showing the components and products of biological pathways along with corresponding annotations and links to literature. This database was developed as a complementary project to (1) serve as a centralized repository of data; (2) connect the biosystem records with associated literature, molecular, and chemical data throughout the Entrez system; and (3) facilitate computation on biosystems data. The NCBI BioSystems Database currently contains records from several source databases: KEGG, BioCyc (including its Tier 1 EcoCyc and MetaCyc databases, and its Tier 2 databases), Reactome, the National Cancer Institute's Pathway Interaction Database, WikiPathways, and Gene Ontology (GO). It includes several types of records such as pathways, structural complexes, and functional sets, and is desiged to accomodate other record types, such as diseases, as data become available. Through these collaborations, the BioSystems database facilitates access to, and provides the ability to compute on, a wide range of biosystems data. If you are interested in depositing data into the BioSystems database, please contact them.
Proper citation: NCBI BioSystems Database (RRID:SCR_004690) Copy
http://www.ncbi.nlm.nih.gov/pubmed/
Public bibliographic database that provides access to citations for biomedical literature from MEDLINE, life science journals, and online books. Citations may include links to full-text content from PubMed Central and publisher web sites. PubMed citations and abstracts include fields of biomedicine and health, covering portions of life sciences, behavioral sciences, chemical sciences, and bioengineering. Provides access to additional relevant web sites and links to other NCBI molecular biology resources. Publishers of journals can submit their citations to NCBI and then provide access to full-text of articles at journal web sites using LinkOut.
Proper citation: PubMed (RRID:SCR_004846) Copy
http://www.ncbi.nlm.nih.gov/probe
Public registry of nucleic acid reagents designed for use in a wide variety of biomedical research applications including genotyping, gene expression studies, SNP discovery, genome mapping, and gene silencing. Probe records contain information on reagent distributors, probe effectiveness, and computed sequence similarities. The database is constantly updated, with over 11,000,000 probes available. Users may deposit their data into NCBI Probe Database.
Proper citation: NCBI Probe (RRID:SCR_004816) Copy
http://www.ncbi.nlm.nih.gov/sra
Repository of raw sequencing data from next generation of sequencing platforms including including Roche 454 GS System, Illumina Genome Analyzer, Applied Biosystems SOLiD System, Helicos Heliscope, Complete Genomics, and Pacific Biosciences SMRT. In addition to raw sequence data, SRA now stores alignment information in form of read placements on reference sequence. Data submissions are welcome. Archive of high throughput sequencing data,part of international partnership of archives (INSDC) at NCBI, European Bioinformatics Institute and DNA Database of Japan. Data submitted to any of this three organizations are shared among them.
Proper citation: NCBI Sequence Read Archive (SRA) (RRID:SCR_004891) Copy
http://www.ncbi.nlm.nih.gov/popset
Database containing a set of DNA sequences that have been collected to analyse the evolutionary relatedness of a population. The population could originate from different members of the same species, or from organisms from different species. Users may submit a Popset using Sequin.
Proper citation: NCBI Popset (RRID:SCR_005049) Copy
Web application to search nucleotide databases using a nucleotide query. Algorithms: blastn, megablast, discontiguous megablast.
Proper citation: BLASTN (RRID:SCR_001598) Copy
https://ftp.ncbi.nlm.nih.gov/pub/mhc/mhc/Final%20Archive/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on August 23, 2019 Database was open, publicly accessible platform for DNA and clinical data related to human Major Histocompatibility Complex (MHC). Data from IHWG workshops were provided as well., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: dbMHC (RRID:SCR_002302) Copy
http://www.ncbi.nlm.nih.gov/HTGS/
Database of high-throughput genome sequences from large-scale genome sequencing centers, including unfinished and finished sequences. It was created to accommodate a growing need to make unfinished genomic sequence data rapidly available to the scientific community in a coordinated effort among the International Nucleotide Sequence databases, DDBJ, EMBL, and GenBank. Sequences are prepared for submission by using NCBI's software tools Sequin or tbl2asn. Each center has an FTP directory into which new or updated sequence files are placed. Sequence data in this division are available for BLAST homology searches against either the htgs database or the month database, which includes all new submissions for the prior month. Unfinished HTG sequences containing contigs greater than 2 kb are assigned an accession number and deposited in the HTG division. A typical HTG record might consist of all the first-pass sequence data generated from a single cosmid, BAC, YAC, or P1 clone, which together make up more than 2 kb and contain one or more gaps. A single accession number is assigned to this collection of sequences, and each record includes a clear indication of the status (phase 1 or 2) plus a prominent warning that the sequence data are unfinished and may contain errors. The accession number does not change as sequence records are updated; only the most recent version of a HTG record remains in GenBank.
Proper citation: High Throughput Genomic Sequences Division (RRID:SCR_002150) Copy
http://www.ncbi.nlm.nih.gov/genome
Database that organizes information on genomes including sequences, maps, chromosomes, assemblies, and annotations in six major organism groups: Archaea, Bacteria, Eukaryotes, Viruses, Viroids, and Plasmids. Genomes of over 1,200 organisms can be found in this database, representing both completely sequenced organisms and those for which sequencing is in progress. Users can browse by organism, and view genome maps and protein clusters. Links to other prokaryotic and archaeal genome projects, as well as BLAST tools and access to the rest of the NCBI online resources are available.
Proper citation: NCBI Genome (RRID:SCR_002474) Copy
https://www.ncbi.nlm.nih.gov/sutils/pasc/viridty.cgi
Web tool for analysis of pairwise identity distribution within viral families. Used for virus sequence-based classification. Data in the system are updated every day to reflect changes in virus taxonomy and additions of new virus sequences to the public database.
Proper citation: PASC (RRID:SCR_016642) Copy
https://www.ncbi.nlm.nih.gov/geo/
Functional genomics data repository supporting MIAME-compliant data submissions. Includes microarray-based experiments measuring the abundance of mRNA, genomic DNA, and protein molecules, as well as non-array-based technologies such as serial analysis of gene expression (SAGE) and mass spectrometry proteomic technology. Array- and sequence-based data are accepted. Collection of curated gene expression DataSets, as well as original Series and Platform records. The database can be searched using keywords, organism, DataSet type and authors. DataSet records contain additional resources including cluster tools and differential expression queries.
Proper citation: Gene Expression Omnibus (GEO) (RRID:SCR_005012) Copy
https://www.ncbi.nlm.nih.gov/geo/
THIS RESOURCE IS NO LONGER IN SERVICE, documented on January 19, 2022.
Proper citation: NCBI Epigenomics (RRID:SCR_006151) Copy
http://www.ncbi.nlm.nih.gov/CCDS/
Database (anonymous FTP) resulting from a collaborative effort to identify a core set of human and mouse protein coding regions that are consistently annotated and of high quality. The long term goal is to support convergence towards a standard set of gene annotations. Collaborators are EBI, NCBI, UCSC, WTSI and the initial results are also available from the participants'''' genome browser Web sites. In addition, CCDS identifiers are indicated on the relevant NCBI RefSeq and Entrez Gene records and in Map Viewer displays of RNA (RefSeq) and Gene annotations on the reference assembly.
Proper citation: Consensus CDS (RRID:SCR_006729) Copy
http://www.ncbi.nlm.nih.gov/RefSeq/HIVInteractions/
A database of interactions between HIV-1 and human proteins published in the peer-reviewed literature. The goal is to provide a concise, yet detailed, summary of all known interactions of HIV-1 proteins with host cell proteins, other HIV-1 proteins, or proteins from disease organisms associated with HIV/AIDS. For each HIV-1 human protein interaction the following information is provided: * NCBI Reference Sequence (RefSeq) protein accession numbers. * NCBI Entrez Gene ID numbers. * Amino acids from each protein that are known to be involved in the interaction. * Brief description of the protein-protein interaction. * Keywords to support searching for interactions. * PubMed identification numbers (PMIDs) for all journal articles describing the interaction. In addition, all protein-protein interactions documented in the database are integrated into Entrez Gene records and listed in the ''HIV-1 protein interactions'' section of Entrez Gene reports. The database is also tightly linked to other databases through Entrez Gene, enabling users to search for an abundance of information related to HIV pathogenesis and replication.
Proper citation: HIV-1 Human Protein Interaction Database (RRID:SCR_006879) Copy
http://www.ncbi.nlm.nih.gov/unists
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 22, 2016. Database of sequence tagged sites (STSs) derived from STS-based maps and other experiments. STSs are defined by PCR primer pairs and are associated with additional information such as genomic position, genes, and sequences. Chromosome maps are labeled by name of the originating organism, the map title, total markers, total UniSTSs and links to view maps as well as research documents available through PubMed, another NCBI database. The search functions within UniSTS allow the user to search by gene marker, chromosome, gene symbol and gene description terms to locate markers on specified genes. A representation of the UniSTS datasets is available by ftp. NOTE: All data from this resource have been moved to the Probe database, http://www.ncbi.nlm.nih.gov/probe. You can retrieve all UniSTS records by searching the probe database using the search term unists(properties). (use brackets insead of parenthesis). Additionally, legacy data remain on the NCBI FTP Site in the UniSTS Repository (ftp://ftp.ncbi.nih.gov/pub/ProbeDB/legacy_unists).
Proper citation: UniSTS (RRID:SCR_006843) Copy
http://www.ncbi.nlm.nih.gov/COG
A database for phylogenetic classification for proteins encoded in complete genomes. Clusters of Orthologous Groups of proteins (COGs) were delineated by comparing protein sequences encoded in complete genomes, representing major phylogenetic lineages. Each COG consists of individual proteins or groups of paralogs from at least 3 lineages and thus corresponds to an ancient conserved domain. Please be aware that COGs hasn't been updated in many years and will not be.
Proper citation: COG (RRID:SCR_007139) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the dkNET Resources search. From here you can search through a compilation of resources used by dkNET and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that dkNET has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on dkNET then you can log in from here to get additional features in dkNET such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into dkNET you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within dkNET that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.