Pith. sign in

Paper Citation Record · LEDGER

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference

As of 5 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 2 inbound Pith citation observations for arXiv:2605.02739.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.02739 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-08T17:49:40.820795Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T03:23:47.033536Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact9
  • verified fuzzy1
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0a8761d6-1317-4627-8e54-b283aba21db5 · outbound

This paper cites Revisiting Feature Prediction for Learning Visual Representations from Video.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference Revisiting Feature Prediction for Learning Visual Representations from Video

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T12:40:24.452580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:69722b9bc4071d807ecf188b3a9d289d05abed022fe968ed998e88c563e4989f

Observation f703e831-dc06-48ad-8577-37a1f8a4d6af · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-11T17:11:14.715264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:df96c8c1b49251e71e534c5e4bdd37f5468c1baee7952ddbd8f5a5e714097fde

Observation ae339173-9ef0-4f2e-acc5-2befc3beb167 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-11T17:11:14.464836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:8d77f9dc877258c684bd599e4483841787cc3f86cc10ed76dbedb164b07e0c0f

Observation b30eee6c-ec9e-47cb-8a6c-b6f96814cf84 · outbound

This paper cites SQAP-VLA: A Synergistic Quantization-Aware Pruning Framework for High-Performance Vision-Language-Action Models.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference SQAP-VLA: A Synergistic Quantization-Aware Pruning Framework for High-Performance Vision-Language-Action Models

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:11:14.246348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:71a962ecf2b8707b171e90f52add515a4ea87184640e266ae553032f2ac1b669

Observation c8e3087e-b066-408f-8efe-2d371bedc4ab · outbound

This paper cites arXiv preprint arXiv:2511.18950 (2025).

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference arXiv preprint arXiv:2511.18950 (2025)

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:11:14.827432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:3b4d4617f32eb2d0021f8136bc3083e3b1e2fcd0b9b09f4c0bfc8e8b1f7592f1

Observation ae6ce2d8-bf45-4619-9bc9-e132bbeba89c · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:11:14.495324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:1bc4efbf8dfec03bdfa65af0fa41b098fc1ce9368e6e84e40b0f071f16429647

Observation 07e8a87a-679e-4656-b5d9-029ff9a003a0 · outbound

This paper cites Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-26T03:04:58.134498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:482b0c71599f47d120289c07e52730ecc7550de2b46b260be7423a24ad4db24f

Observation 75ec6636-16a9-48fb-ab9e-8e62662b3964 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-11T17:11:14.924749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:4fe1934549d4dba7031ccbb5acc0c1d4c00c9993e825c8e1ea0c557c2b9aae0c

Observation 920aff8b-a957-4f57-9785-0153d80c0084 · outbound

This paper cites SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-26T02:03:02.651395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:53b15f7bf3c8f1873882fd25b1f25777894738b9dd5463061b86f0adff683b96

Observation 914138b2-21be-4b2a-9d18-0ab2604a01a4 · outbound

This paper cites Scaling Proprioceptive-Visual Learning with Heterogeneous Pre-trained Transformers.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference Scaling Proprioceptive-Visual Learning with Heterogeneous Pre-trained Transformers

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:11:14.816154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:6b397cd80a5bb6f882ab080dee62d56ee839b0bba376da3aab28c0e78f97d016

Observation 63df7d0a-c5e7-4a05-9075-32a9c2218a05 · outbound

This paper cites arXiv preprint arXiv:2502.02175 (2025).

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference arXiv preprint arXiv:2502.02175 (2025)

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:11:15.083137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:a4696429d3916139fb25dc27637317d0b16aeeea2fdfcb1a2f9292267d282e61

Observation 557a335d-fd68-4611-9203-0e46311f8909 · outbound

This paper cites Dyq-vla: Temporal-dynamic-aware quanti- zation for embodied vision-language-action models.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference Dyq-vla: Temporal-dynamic-aware quanti- zation for embodied vision-language-action models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:11:14.470207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:c9540b214ed020d8749d727ed136894774338367cc62886c62a537325fab802c

Observation 9b803458-10fd-4dab-b82f-00e09e2cba34 · outbound

This paper cites max-autotune.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference max-autotune

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T06:36:52.134951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:dc87f936a71f9e3bb3c0f5d35fa0607b76b2e8f64254f9846cf2c0139a0806b2

Pith citing papers

Observation 2e45585a-2c9f-4e9f-807e-b52c0d84d0be · inbound

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control cites this paper.

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T11:27:05.757435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:27:05.757435Z digest=sha256:af001ef8614813868aca88403a837216c55cb8612bddd6c578b8c26d0d286174

Observation 4730151a-369f-4c89-898a-02df7368b868 · inbound

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control cites this paper.

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T03:23:47.033536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T03:23:47.033536Z digest=sha256:640e9aa1b13832b6698a51b5315ebd22b2bce26b9b9bec7e11682db85d431260