cs.LGAug 23, 2026

Mol-JEPA: A multimodal Joint Embedding Predictive Architecture for Molecules

Authors: Florian RottachSebastian SchieferdeckerWilliam RudmanRandall BalestrieroCarsten Eickhoff

Organizations: University of Tübingen · 2Boehringer Ingelheim · 3The University of Texas at Austin · 4Brown University

Abstract

Despite recent advances in molecular foundation models, several limitations remain, such as chemically invalid augmentations, modality collapse, and incomplete representation of biochemical environments. To address these challenges, we present \textbf{Mol-JEPA}, a scalable framework for learning molecular world models. Rather than relying on suboptimal molecular perturbations, our model uses modality masking to exploit information from molecular structures, cellular phenotypes, binding affinities, ADMET profiles, quantum chemistry simulations and other drug discovery data. Across various benchmarks, we show that the representations learned by Mol-JEPA deliver strong performance, demonstrating the value of incorporating biochemical context through latent space prediction.

Explore similar work

CardsList