Genomics & Single-Cell · Zhou et al., Northwestern
BPE-tokenised genomic BERT for motif, splice site and regulatory classification.
1–30s
Runs on On-demand GPU
Runs on our servers with on-demand compute. A first run needs time to load the model; active capacity can be reused and scales down when idle.Run this model
About this model
DNABERT-2 replaces k-mer tokenisation with byte-pair encoding, which makes it both faster and better across the GUE benchmark suite. A solid, cheap baseline for genomic classification tasks.
Standardized I/O contract
Every model in the Hub speaks the same contract, which is what lets the Router and the agent call any of them without special-casing.
Inputs
DNA sequence
Outputs
Sequence representation.