A twelve-slide breakdown of DNABERT, the origin of the DNA language model: how it reads DNA as overlapping k-mers, why pre-training once and fine-tuning for each task changed genomics, how contiguous-span masking teaches it context, and how one model reached above 0.9 across 690 ENCODE binding-site datasets and even transferred from human to mouse.
Read the paper: DNABERT: pre-trained Bidirectional Encoder Representations from Transformers model for DNA-language in genome (Bioinformatics, 2021).