Pith. sign in

Paper Citation Record · LEDGER

StableMask: Refining Causal Masking in Decoder-only Transformer

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2402.04779.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.04779 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T16:30:21.408915Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T17:41:03.773415Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6b6335bc-c303-4b7a-8fe7-655a47397362 · inbound

When Attention Sink Emerges in Language Models: An Empirical View cites this paper.

When Attention Sink Emerges in Language Models: An Empirical View StableMask: Refining Causal Masking in Decoder-only Transformer

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:41:03.776034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-16T17:41:03.674759Z digest=sha256:798e1bcbb764a2b93dd62f25029730305991b5e17c60a6d916099504fb951e98

Observation 33f3f5fc-2681-4d3b-9fc6-324a5e6fc640 · inbound

When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training cites this paper.

When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training StableMask: Refining Causal Masking in Decoder-only Transformer

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T16:30:21.408915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T16:30:21.408915Z digest=sha256:510aff66bc17782053fe77ca4d6be6458a3a52ab4a1f3449f0bdd4853647049c

Observation 1ffe2807-6b8b-4513-a546-39fb773b7951 · inbound

NanoVLMs: How small can we go and still make coherent Vision Language Models? cites this paper.

NanoVLMs: How small can we go and still make coherent Vision Language Models? StableMask: Refining Causal Masking in Decoder-only Transformer

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T13:35:05.176959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:35:05.176959Z digest=sha256:65825e048d51b6e9281c93ec6369345f4238b6665c756d291dc1d2f24790e923

Observation 4ea73cad-5d08-4c8d-8e82-edd91e5c2160 · inbound

ModRWKV: Transformer Multimodality in Linear Time cites this paper.

ModRWKV: Transformer Multimodality in Linear Time StableMask: Refining Causal Masking in Decoder-only Transformer

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.506938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.506938Z digest=sha256:edf0f5b02e3148a0066f65d2a10b2997eaae73ecafe3718b40b3fce399158955

Observation 2f3fbd5e-e1e4-4441-85fe-c254df22750a · inbound

Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding cites this paper.

Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding StableMask: Refining Causal Masking in Decoder-only Transformer

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T15:03:28.377330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:03:28.377330Z digest=sha256:744d91c359e163f4489688a2fbb94c8c1f821af944ecb1545467e13014b64118

Observation 274dc6ef-292a-4c05-a83e-d634165b65ee · inbound

Rethinking Causal Mask Attention for Vision-Language Inference cites this paper.

Rethinking Causal Mask Attention for Vision-Language Inference StableMask: Refining Causal Masking in Decoder-only Transformer

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:33:33.351182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:33:33.351182Z digest=sha256:adfffbe718eaf7f2e849e8be801acea5188abb6e057f3ad19bc4a33b8df9bf14