REVIEW 2 cited by
Discrete Autoencoders for Sequence Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Recurrent models for sequences have been recently successful at many tasks, especially for language modeling and machine translation. Nevertheless, it remains challenging to extract good representations from these models. For instance, even though language has a clear hierarchical structure going from characters through words to sentences, it is not apparent in current language models. We propose to improve the representation in sequence models by augmenting current approaches with an autoencoder that is forced to compress the sequence through an intermediate discrete latent space. In order to propagate gradients though this discrete representation we introduce an improved semantic hashing technique. We show that this technique performs well on a newly proposed quantitative efficiency measure. We also analyze latent codes produced by the model showing how they correspond to words and phrases. Finally, we present an application of the autoencoder-augmented model to generating diverse translations.
Forward citations
Cited by 2 Pith papers
-
On Designing Diffusion Autoencoders for Efficient Generation and Representation Learning
Small binary latents conditioned via cross-attention let a diffusion autoencoder generate from a uniform Bernoulli prior with fewer steps while keeping representation quality.
-
Spatial-Temporal Expert Learning for Video-based Person Re-identification
Proposes dynamic expert selection with input-aware and spatial-temporal mechanisms plus an extendable scheme to improve fine-grained feature use in video person Re-ID.
Discussion (0). Continue with ORCID to comment.