University of Maryland Global Campus (UMGC) | Bioinformatics Capstone Project
Welcome to the official repository for The Genomic Oracle, a cascaded machine learning pipeline designed for high-precision DNA sequence classification and phenotypic prediction.
As genomic datasets grow exponentially, the need for rapid, automated sequence annotation is critical. The Genomic Oracle serves as an intelligent routing network, evaluating raw DNA sequences through a multi-stage gauntlet to classify their biological function, structural feature type, and associated phenotypic risks.
Our platform utilizes a highly specialized, branching AI architecture hosted on Hugging Face ZeroGPU infrastructure:
Integration: The pipeline concludes with an automated API routing to NCBI and Ensembl databases for real-world chromosomal coordinate mapping and validation.
We are a team of graduate researchers specializing in data science, machine learning, and bioinformatics.
Created by: Kadir Galindo, Duncan Hall, Rebecca Mellinger, George Paccione