MolMix: A Simple Yet Effective Baseline for Multimodal Molecular Representation Learning

Manolache, Andrei; Tantaru, Dragos; Niepert, Mathias

Computer Science > Machine Learning

arXiv:2410.07981 (cs)

[Submitted on 10 Oct 2024 (v1), last revised 24 Oct 2024 (this version, v2)]

Title:MolMix: A Simple Yet Effective Baseline for Multimodal Molecular Representation Learning

Authors:Andrei Manolache, Dragos Tantaru, Mathias Niepert

View PDF HTML (experimental)

Abstract:In this work, we propose a simple transformer-based baseline for multimodal molecular representation learning, integrating three distinct modalities: SMILES strings, 2D graph representations, and 3D conformers of molecules. A key aspect of our approach is the aggregation of 3D conformers, allowing the model to account for the fact that molecules can adopt multiple conformations-an important factor for accurate molecular representation. The tokens for each modality are extracted using modality-specific encoders: a transformer for SMILES strings, a message-passing neural network for 2D graphs, and an equivariant neural network for 3D conformers. The flexibility and modularity of this framework enable easy adaptation and replacement of these encoders, making the model highly versatile for different molecular tasks. The extracted tokens are then combined into a unified multimodal sequence, which is processed by a downstream transformer for prediction tasks. To efficiently scale our model for large multimodal datasets, we utilize Flash Attention 2 and bfloat16 precision. Despite its simplicity, our approach achieves state-of-the-art results across multiple datasets, demonstrating its effectiveness as a strong baseline for multimodal molecular representation learning.

Comments:	Machine Learning for Structural Biology Workshop, NeurIPS 2024 v2: Added optimizer references
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2410.07981 [cs.LG]
	(or arXiv:2410.07981v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2410.07981

Submission history

From: Andrei Manolache [view email]
[v1] Thu, 10 Oct 2024 14:36:58 UTC (1,465 KB)
[v2] Thu, 24 Oct 2024 08:34:50 UTC (1,466 KB)

Computer Science > Machine Learning

Title:MolMix: A Simple Yet Effective Baseline for Multimodal Molecular Representation Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:MolMix: A Simple Yet Effective Baseline for Multimodal Molecular Representation Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators