Pith. sign in

Paper Citation Record · LEDGER

OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2509.05578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.05578 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T05:06:33.852576Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:48:39.596474Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5f1683f6-57be-4be2-9d80-8e61963892de · inbound

ShelfGaussian: Shelf-Supervised Open-Vocabulary Gaussian-based 3D Scene Understanding cites this paper.

ShelfGaussian: Shelf-Supervised Open-Vocabulary Gaussian-based 3D Scene Understanding OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:28:58.098776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T03:24:32.577420Z digest=sha256:fc8e00dee1a4043d234332271edaaff71ed9f9ebbd903de5bc8c130759898dd4

Observation 9f063d58-4fb3-4dd6-8835-ef025ac71ba9 · inbound

Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes cites this paper.

Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:54:09.204266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T11:50:37.423645Z digest=sha256:ea36a673987ec2dadf3f4caf5551291980da07f054b1d327a24fad34af8a43f7

Observation 5ffcdbe2-573d-4277-9544-3875d8f2d94a · inbound

The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook cites this paper.

The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision

Reference 124

Resolution
unresolved
no resolver link, observed 2026-07-13T14:03:01.974171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:03:01.974171Z digest=sha256:a3f70052e8b9da4e4de91fd43c108f194ffa20c3e085f24b77d92bd95b7f205c

Observation 129cd057-4381-4156-be54-a34a21c1153c · inbound

Height-Guided Projection Reparameterization for Camera-LiDAR Occupancy cites this paper.

Height-Guided Projection Reparameterization for Camera-LiDAR Occupancy OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:36:06.388360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T17:24:36.149795Z digest=sha256:d42c2ba210c42acf11349e8f50566956fb6f19f926e144e8903230219bf97436

Observation 698acbad-fd1c-4fd1-8a42-201e9edaf122 · inbound

Height-Guided Projection Reparameterization for Camera-LiDAR Occupancy cites this paper.

Height-Guided Projection Reparameterization for Camera-LiDAR Occupancy OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:41:43.960049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T04:02:19.480937Z digest=sha256:cbbae3b6fe42fb204c9398818b346324a297978a356d4e2278fb9dc92a001635

Observation 207aa84a-c5bf-4bc7-96ca-eaf65bf772c0 · inbound

TPS-Drive: Task-Guided Representation Purification for VLM-based Autonomous Driving cites this paper.

TPS-Drive: Task-Guided Representation Purification for VLM-based Autonomous Driving OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:43:45.704784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T17:42:49.902997Z digest=sha256:3225c74292f4a070cce21b14427556659dd8d5f4c188367b4c6239bafb38667c

Observation 673c8019-968e-4da9-89d5-713c18b52b3e · inbound

M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks cites this paper.

M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.758810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T06:16:07.090870Z digest=sha256:3df8628f4dbfe5fe177a34eef6bafc28b975523c1e7f5a0223565024530d40ef

Observation 1a5a9899-08ed-4210-ba36-436a01495968 · inbound

Teaching Vision-Language-Action Models What to See and Where to Look cites this paper.

Teaching Vision-Language-Action Models What to See and Where to Look OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:48:39.597762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T16:42:13.520913Z digest=sha256:822459eb4d217f6aec416541ff1bb678506be431a553dc144eb3954e139df791

Observation 8d12a04d-ee7a-4419-9c93-78de1d3c9166 · inbound

GPOcc++: Unified Sparse Gaussian Occupancy Prediction with Visual Geometry Priors cites this paper.

GPOcc++: Unified Sparse Gaussian Occupancy Prediction with Visual Geometry Priors OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T05:06:33.852576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:06:33.852576Z digest=sha256:f35e0289a43c43607f4db95c8594202240c817c28b54cf5602a2ef74c7a5f652