This is an intermediate version of Evo 2 7B, trained with up to a 262,144 token context length.
Evo 2 is a state-of-the-art DNA language model trained autoregressively on trillions of DNA tokens.
For instructions, details, and examples, please refer to the
GitHub and
paper.
Evo 2 40B and 7B checkpoints, trained up to 1 million sequence length, are available here:
Please refer to the
Evo 2 GitHub repository for detailed usage instructions and examples.