REVIEW 5 cited by
Molecule Attention Transformer
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Designing a single neural network architecture that performs competitively across a range of molecule property prediction tasks remains largely an open challenge, and its solution may unlock a widespread use of deep learning in the drug discovery industry. To move towards this goal, we propose Molecule Attention Transformer (MAT). Our key innovation is to augment the attention mechanism in Transformer using inter-atomic distances and the molecular graph structure. Experiments show that MAT performs competitively on a diverse set of molecular prediction tasks. Most importantly, with a simple self-supervised pretraining, MAT requires tuning of only a few hyperparameter values to achieve state-of-the-art performance on downstream tasks. Finally, we show that attention weights learned by MAT are interpretable from the chemical point of view.
Forward citations
Cited by 5 Pith papers
-
Predicting Therapeutic Outcome via Aligning Patient-Specific Knowledge Graph and Gene-Level Perturbation Representations
Aligning patient-specific gene-regulatory graphs with LINCS-pretrained perturbation embeddings via CLIP-style contrastive learning improves clinical drug-response prediction on TCGA and zero-shot I-SPY2.
-
From Molecules to Mixtures: Learning Representations of Olfactory Mixture Similarity using Inductive Biases
POMMix, a graph-based model with attention and cosine similarity heads, extends the Principal Odor Map to predict human perceptual similarity of odor mixtures, reporting a test correlation of 0.78 on a compiled datase...
-
Multi-Alignment Contrastive Learning for Enzyme--Reaction Retrieval
FGW-CLIP, a contrastive method that aligns enzymes and reactions while also aligning within-domain EC structure with a Gromov-Wasserstein regularizer, reports state-of-the-art retrieval on EnzymeMap and ReactZyme.
-
KEPLA: A Knowledge-Enhanced Deep Learning Framework for Accurate Protein-Ligand Binding Affinity Prediction
KEPLA jointly optimizes knowledge graph embeddings and cross attention to improve protein-ligand affinity prediction, reporting RMSE 1.202 on PDBbind core and 1.459 on CSAR-HiQ, beating all tested baselines.
-
Multi-Level Fusion Graph Neural Network for Molecule Property Prediction
MLFGNN, a GAT combined with a DyT-augmented Graph Transformer and fingerprint cross-attention, reports best regression scores on five MoleculeNet benchmarks and best classification scores on two of five.
Discussion (0). Continue with ORCID to comment.