Pith. sign in

Paper Citation Record · LEDGER

Vision Transformer with Quadrangle Attention

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2303.15105.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.15105 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:41:22.021905Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T06:02:37.541296Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cd6e5d04-8c6e-4f0c-8feb-a0d5fe834643 · inbound

Survey of different Large Language Model Architectures: Trends, Benchmarks, and Challenges cites this paper.

Survey of different Large Language Model Architectures: Trends, Benchmarks, and Challenges Vision Transformer with Quadrangle Attention

Reference 256

Resolution
unresolved
no resolver link, observed 2026-08-11T22:41:22.021905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:41:22.021905Z digest=sha256:f8448332ad02e382179a6d73b172bd8d013191745598ea4ff7e02422ed92eb7b

Observation 846b657f-d763-4204-ad3c-939633a8cbf2 · inbound

LLaVA-Octopus: Unlocking Instruction-Driven Adaptive Projector Fusion for Video Understanding cites this paper.

LLaVA-Octopus: Unlocking Instruction-Driven Adaptive Projector Fusion for Video Understanding Vision Transformer with Quadrangle Attention

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-23T06:02:37.545249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-23T06:01:00.775721Z digest=sha256:2e1e6a32d985a8217ec9bd10cc466088ea7c9e0ea075521771a1ab02fb60e776

Observation 1ffd1b68-9ee3-4d03-8634-e20268023f95 · inbound

Facial Dynamics in Video: Instruction Tuning for Improved Facial Expression Perception and Contextual Awareness cites this paper.

Facial Dynamics in Video: Instruction Tuning for Improved Facial Expression Perception and Contextual Awareness Vision Transformer with Quadrangle Attention

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-10T20:33:54.289229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:33:54.289229Z digest=sha256:ea6029ac69c097c95d6315abe634b0490a113786b3a7f61757883d5e40fe9358

Observation 087b8d9f-8df3-488a-8814-c9a5f9fc6b88 · inbound

LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs cites this paper.

LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs Vision Transformer with Quadrangle Attention

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-06T22:24:35.992190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:24:35.992190Z digest=sha256:cc62e0b2cc3b3775d18db1951e8643a68aa48219e06bc94975cfa3c7358db85b

Observation fd7ae621-bd2f-40f0-8604-32794fccae95 · inbound

ReactBench: A Benchmark for Topological Reasoning in MLLMs on Chemical Reaction Diagrams cites this paper.

ReactBench: A Benchmark for Topological Reasoning in MLLMs on Chemical Reaction Diagrams Vision Transformer with Quadrangle Attention

Reference 3

Resolution
malformed identifier
arxiv_id, observed 2026-05-10T09:03:25.057325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T09:01:28.605086Z digest=sha256:2a52aa1f0810c9588214735621ad3f2145de0bbcfc1c35d51195968effb53526

Observation 50cae3d4-20d9-4b12-aaec-50482c25540f · inbound

X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment cites this paper.

X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment Vision Transformer with Quadrangle Attention

Reference 158

Resolution
unresolved
no resolver link, observed 2026-08-01T07:12:17.721095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:12:17.721095Z digest=sha256:1bd7dc0cd793b06cd1b2c7464d9da470974f5de0d30a88c7bfe16eefe030ce83