REVIEW 1 cited by
word2ket: Space-efficient Word Embeddings inspired by Quantum Entanglement
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Deep learning natural language processing models often use vector word embeddings, such as word2vec or GloVe, to represent words. A discrete sequence of words can be much more easily integrated with downstream neural layers if it is represented as a sequence of continuous vectors. Also, semantic relationships between words, learned from a text corpus, can be encoded in the relative configurations of the embedding vectors. However, storing and accessing embedding vectors for all words in a dictionary requires large amount of space, and may stain systems with limited GPU memory. Here, we used approaches inspired by quantum computing to propose two related methods, {\em word2ket} and {\em word2ketXS}, for storing word embedding matrix during training and inference in a highly efficient way. Our approach achieves a hundred-fold or more reduction in the space required to store the embeddings with almost no relative drop in accuracy in practical natural language processing tasks.
Forward citations
Cited by 1 Pith paper
-
Quantum-Classical Hybrid Molecular Autoencoder for Advancing Classical Decoding
A Word2Ket-based quantum autoencoder with an attention LSTM decoder reconstructs SMILES strings, reaching 84% quantum fidelity and 60% Levenshtein similarity on the QM9 training set.
Discussion (0). Continue with ORCID to comment.