Database Mining

Database mining is the computational process of extracting patterns, relationships, and useful knowledge from large, structured collections of data. In genetics, algorithms query and integrate genomic databases, then use statistical analysis, sequence comparison, and machine learning to identify associations among variants, genes, phenotypes, and biological pathways. These analyses support gene discovery, variant prioritization, disease-association studies, and interpretation of high-throughput sequencing results. By converting dispersed datasets into testable hypotheses, database mining helps researchers compare genomes, uncover genetic relationships, and guide experimental validation, while requiring careful attention to data quality, bias, and reproducibility.

Database Mining - Related Videos

Education

JoVE Science Education - Information Literacy

Database Searching

0 Views •

2026

Effective research begins with the ability to navigate academic databases efficiently. Unlike general web search engines, scholarly databases are designed to index peer-reviewed articles, conference proceedings, and other credible academic publications. Developing a structured search strategy ensures that researchers can identify relevant, high-quality sources with precision. For example, when studying how sleep patterns relate to circadian rhythms in humans, researchers should move beyond...

Research

JoVE Journal - Biology
Free Sample

The ITS2 Database

0 Views •

Cited by 52 •

2012

The ITS2 Database is a workbench for phylogenetic inference simultaneously considering sequence and secondary structure of the internal transcribed spacer 2. This includes data collection with accurate annotation, structure prediction, multiple sequence-structure alignment and fast tree calculation. In a nutshell, this workbench simplifies first phylogenetic analyses to a few clicks.

Performing Data Mining And Integrative Analysis Of Biomarker in Breast Cancer Using Multiple Publicly Accessible Databases

0 Views •

Cited by 1 •

2019

Here, we present a protocol to explore the biomarker and survival predictor of breast cancer based on the comprehensive analysis of pooled clinical datasets derived from a variety of publicly accessible databases, using the strategy of expression, correlation and survival analysis step by step.

Research

JoVE Journal - Biology
Free Sample

Mining Spatial Transcriptomics Datasets using DeepSpaceDB

0 Views •

2025

This article introduces a protocol for using DeepSpaceDB, a dynamic, interactive database for spatial transcriptomics, offering analysis workflows and examples to explore tissue organization and disease-related gene expression.

Research

JoVE Journal - Medicine
Free Sample

Generation of Comprehensive Thoracic Oncology Database - Tool for Translational Research

0 Views •

Cited by 5 •

2011

A thoracic oncology database was developed to serve as a comprehensive repository for clinical and laboratory data for the purposes of translational research. The database will serve translational cancer researchers within the Thoracic Oncology Research Program. This database is adaptable to other cancer models, as well as other human diseases.

View All Results

FAQs

Related Topics