Pith. sign in

REVIEW 1 cited by

GIMLET: A Unified Graph-Text Model for Instruction-Based Molecule Zero-Shot Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2306.13089 v3 pith:55J2E2FX submitted 2023-05-28 cs.LG cs.CLq-bio.BM

classification cs.LGcs.CLq-bio.BM
keywords tasksgimletgraphinstructionsmoleculemodelmodelszero-shot
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Molecule property prediction has gained significant attention in recent years. The main bottleneck is the label insufficiency caused by expensive lab experiments. In order to alleviate this issue and to better leverage textual knowledge for tasks, this study investigates the feasibility of employing natural language instructions to accomplish molecule-related tasks in a zero-shot setting. We discover that existing molecule-text models perform poorly in this setting due to inadequate treatment of instructions and limited capacity for graphs. To overcome these issues, we propose GIMLET, which unifies language models for both graph and text data. By adopting generalized position embedding, our model is extended to encode both graph structures and instruction text without additional graph encoding modules. GIMLET also decouples encoding of the graph from tasks instructions in the attention mechanism, enhancing the generalization of graph features across novel tasks. We construct a dataset consisting of more than two thousand molecule tasks with corresponding instructions derived from task descriptions. We pretrain GIMLET on the molecule tasks along with instructions, enabling the model to transfer effectively to a broad range of tasks. Experimental results demonstrate that GIMLET significantly outperforms molecule-text baselines in instruction-based zero-shot learning, even achieving closed results to supervised GNN models on tasks such as toxcast and muv.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Unify Graph Learning with Text: Unleashing LLM Potentials for Session Search

    cs.CV 2025-05 conditional novelty 6.0 of 10

    A session graph serialized into symbolic text, plus self-supervised graph pre-training tasks, lets an LLM outperform existing session search rankers on AOL and Tiangong-ST.

Pith tools