The MultiMolecule team has confirmed that the provided model and checkpoints are producing the same intermediate representations as the original implementation.
The team releasing Borzoi did not write this model card for this model so this model card has been written by the MultiMolecule team.
Model Details
Borzoi is the successor of
Enformer
. It extends the Enformer recipe (convolution stem + Transformer trunk + binned multi-track output) to a 524,288 bp input window and 32 bp output bins, and adds a U-Net style upsampling tail so the binned positional axis matches a higher-resolution coverage prediction. A long DNA window of 524 kb is downsampled by a convolution stem and a width-growing residual convolution tower, projected to 1,536 channels by a U-Net bottleneck, processed by 8 Transformer blocks with Transformer-XL style relative positional encoding, then upsampled by two skip-connected U-Net stages with depthwise-separable convolutions, center-cropped to 6,144 bins, and projected to per-species coverage tracks with a softplus activation. The output is
binned
: it has shape
(batch_size, target_length, num_tracks)
where each bin summarizes 32 bp of sequence and
num_tracks
is the number of genomic coverage experiments for the selected species. Borzoi was trained jointly on RNA-seq, CAGE, ATAC-seq, DNase-seq, and ChIP-seq tracks. Please refer to the
Training Details
section for more information on the training process.
Variants
Borzoi releases separate human and mouse checkpoints for the corresponding species track sets.
The table reports the human checkpoint. The mouse checkpoint predicts 2,608 tracks.
FLOPs and MACs are measured on the canonical 524,288 bp Borzoi input window.
Developed by
: Johannes Linder, Divyanshi Srivastava, Han Yuan, Vikram Agarwal, David R. Kelley
Model type
: Convolutional stem followed by Transformer trunk and U-Net upsampling tail for binned multi-track RNA-seq and chromatin coverage prediction
The binned positional axis is treated as the "token" axis: each output position corresponds to one
genomic bin rather than a single nucleotide. The
species
configuration option selects the
human
(7,611 tracks) or
mouse
(2,608 tracks) species track set for the converted checkpoint.
Interface
Input length
: fixed 524,288 bp DNA window
Output binning
: 32 bp per output bin; 6,144 output bins per window (after center-cropping the U-Net upsampling tail)
Species track set
: select
human
(7,611 tracks) or
mouse
(2,608 tracks) via the
species
config option
Output
:
(batch_size, target_length, num_tracks)
Training Details
Borzoi was trained to predict bulk RNA-seq coverage together with chromatin tracks (DNase-seq, ATAC-seq, ChIP-seq) and CAGE from the human and mouse reference genomes.
Training Data
The model was trained on a large compendium of functional genomics experiments aligned to the human (hg38) and mouse (mm10) reference genomes. The genome was divided into 524 kb windows; for each window the per-32-bp coverage of every experiment served as the regression target. The training set is dominated by RNA-seq coverage (the modality Borzoi extends over Enformer); the remaining tracks include the chromatin and CAGE modalities used by Enformer.
Training Procedure
Pre-training
The model was trained to minimize a Poisson-multinomial regression loss between predicted and observed coverage, using a softplus output activation to keep the predicted coverage non-negative. Training used the Adam optimizer with a warmup schedule and global gradient-norm clipping; reverse-complement and small genomic-shift data augmentations were applied during training.
Citation
@article{linder2025predicting,
author = {Linder, Johannes and Srivastava, Divyanshi and Yuan, Han and Agarwal, Vikram and Kelley, David R.},
title = {Predicting RNA-seq coverage from DNA sequence as a unifying model of gene regulation},
journal = {Nature Genetics},
year = 2025,
volume = 57,
number = 4,
pages = {949--961},
doi = {10.1038/s41588-024-02053-6},
publisher = {Nature Publishing Group}
}
The artifacts distributed in this repository are part of the MultiMolecule project.
If MultiMolecule supports your research, please cite the MultiMolecule project as follows:
@software{chen_2024_12638419,
author = {Chen, Zhiyuan and Zhu, Sophia Y.},
title = {MultiMolecule},
doi = {10.5281/zenodo.12638419},
publisher = {Zenodo},
url = {https://doi.org/10.5281/zenodo.12638419},
year = 2024,
month = may,
day = 4
}
Contact
Please use GitHub issues of
MultiMolecule
for any questions or comments on the model card.
Please contact the authors of the
Borzoi paper
for questions or comments on the paper/model.
borzoi-human huggingface.co is an AI model on huggingface.co that provides borzoi-human's model effect (), which can be used instantly with this multimolecule borzoi-human model. huggingface.co supports a free trial of the borzoi-human model, and also provides paid use of the borzoi-human. Support call borzoi-human model through api, including Node.js, Python, http.
borzoi-human huggingface.co is an online trial and call api platform, which integrates borzoi-human's modeling effects, including api services, and provides a free online trial of borzoi-human, you can try borzoi-human online for free by clicking the link below.
multimolecule borzoi-human online free url in huggingface.co:
borzoi-human is an open source model from GitHub that offers a free installation service, and any user can find borzoi-human on GitHub to install. At the same time, huggingface.co provides the effect of borzoi-human install, users can directly use borzoi-human installed effect in huggingface.co for debugging and trial. It also supports api for free installation.