Pith. sign in

Paper Citation Record · LEDGER

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

As of 9 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2608.05042.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05042 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:44:18.247570Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact2
  • verified fuzzy35
  • unresolved27
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation da635c3e-14ba-454c-ad3a-f0147571b3d3 · outbound

This paper cites OpenVLA: An open-source vision-language-action model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation OpenVLA: An open-source vision-language-action model,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:11.836856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:11.836856Z digest=sha256:b48140011758c0a31afde441bbec72e513fb9c848041cd6d1f18ff9382ebe255

Observation 30ef66ac-b196-4113-845a-b1c1b5b72135 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:11.920374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:11.920374Z digest=sha256:f59ee2f3b9bf374324c969c0366acb3fe4cbe2ee850f32c449c0599bf2007531

Observation 4c4286e5-ac21-4ebc-becb-0504efc4af64 · outbound

This paper cites Wall-OSS-0.5 Technical Report.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Wall-OSS-0.5 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:12.106434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:12.106434Z digest=sha256:d453240e7ff27e9b2e67f41f5f705400ac6b4fca9716819dde76a0005c4a93a8

Observation 9ae2648e-564a-4774-8ef4-35d5a4bf726b · outbound

This paper cites Vision-language foundation models as effective robot imitators,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Vision-language foundation models as effective robot imitators,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.055685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.235379Z digest=sha256:9fb8493bc1fed2119b92e5adebd0a88c613887a6b91827ebb4fff71c4a7ffa97

Observation 72f95a25-9bc3-4002-aaa5-25cae51bd56a · outbound

This paper cites RT-2: Vision-language-action models transfer web knowledge to robotic control,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RT-2: Vision-language-action models transfer web knowledge to robotic control,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.046255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.390934Z digest=sha256:c99152293d776d1f6d7224c81fe141335e28ece51e4c85e97870de3eaeec3e94

Observation 36ec393a-1107-4976-8e73-5b110a128bc7 · outbound

This paper cites Perceiver-Actor: A multi-task transformer for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Perceiver-Actor: A multi-task transformer for robotic manipulation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.036180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.523105Z digest=sha256:0b94aae2ee69ffe720c54e9f8dac945f2f60dbac827f84c736cea794f4880872

Observation 62fa746e-6e30-4d58-bd9a-fbd27cc63775 · outbound

This paper cites 3D Diffuser Actor: Policy diffusion with 3D scene representations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation 3D Diffuser Actor: Policy diffusion with 3D scene representations,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.027748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.644589Z digest=sha256:6e2152b0ba8004403dfb4d3e34cc0075c94be949cfaf0b6c197973c6b4d3e0a6

Observation 8b5524d3-a092-490e-a37f-9b0b724fa9be · outbound

This paper cites Act3D: 3D feature field transformers for multi-task robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Act3D: 3D feature field transformers for multi-task robotic manipulation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.019261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.852518Z digest=sha256:1f4c5366ba23fa88162d43b07fcc37481f9acd060a467fb4f05011d5c7f3e902

Observation c172a669-c510-4cd3-84cc-a552d268c8b2 · outbound

This paper cites RVT: Robotic view transformer for 3D object manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RVT: Robotic view transformer for 3D object manipulation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.010416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.979457Z digest=sha256:33eb4b20d4d3ba41c42163701404d417dc2de3a389ef46a4552083243ff0e64e

Observation 132b4569-ec0d-4f4b-9f54-2f83a980c37f · outbound

This paper cites RVT-2: Learning precise manipulation from few demonstrations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RVT-2: Learning precise manipulation from few demonstrations,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.001863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.078472Z digest=sha256:f9400ecd8aefc9fa51f488a561fd9768ed3f8f95f7a47499c0af3fb432758b14

Observation 111db71a-e83c-4bf4-ae07-e402e17c8359 · outbound

This paper cites 3D-VLA: A 3D vision-language-action generative world model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation 3D-VLA: A 3D vision-language-action generative world model,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.993333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.210283Z digest=sha256:f2a56b783a1053945cc05c048077dd49eac3812fb3fec0edac9ba88ade9663dd

Observation abec4001-6626-4f0a-aa60-e1d5f1aae51a · outbound

This paper cites SpatialVLA: Exploring spatial representations for visual- language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SpatialVLA: Exploring spatial representations for visual- language-action models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.984740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.350018Z digest=sha256:805d6254b993723bb869f1ecc305def7d543ed0a4b77872959179a5f98e88571

Observation e7176949-b0a9-44bd-8dc7-6660c0d0d8c7 · outbound

This paper cites RLBench: The robot learning benchmark & learning environment,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RLBench: The robot learning benchmark & learning environment,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.976335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.499499Z digest=sha256:be23bb7167b742ae453324a0a81637f56c2cd76a7ae47db0efb3795bd8b55f26

Observation edbb9f7e-c70c-4f75-96bf-10ee976757d3 · outbound

This paper cites THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:13.672028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:13.672028Z digest=sha256:797081fc5d0c572e7fde2439730ca3f28e4be02a27c52d13353822bf21d6ace3

Observation 6973dddc-827f-46de-a672-d9483518cd54 · outbound

This paper cites Towards generalizable vision- language robotic manipulation: A benchmark and LLM-guided 3D policy,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Towards generalizable vision- language robotic manipulation: A benchmark and LLM-guided 3D policy,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.966940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.754442Z digest=sha256:e570f5d97485bb1420187c9bf3d22a973ef394f10f2bd442dee664ac8d076f88

Observation fe57631e-5ac7-41f1-a37f-d84c9630ec0d · outbound

This paper cites RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:13.849049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:13.849049Z digest=sha256:574eb3c2e629cdc1cf712cbe93869028471f30ac70c87f32e6c654fe40a7bea4

Observation 10fe969f-77ba-4af8-8513-e8fb104d71cf · outbound

This paper cites SAM2Act: Integrating visual foundation model with a memory architecture for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SAM2Act: Integrating visual foundation model with a memory architecture for robotic manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.956456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.970866Z digest=sha256:d2672f39781bb863bceea5be93eac9fb349774858e573357d0956e0ea22a7d1d

Observation 496c8dc8-a445-49a6-8865-e1fe4a147bc7 · outbound

This paper cites BridgeVLA: Input-output alignment for efficient 3D ma- nipulation learning with vision-language models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation BridgeVLA: Input-output alignment for efficient 3D ma- nipulation learning with vision-language models,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.945963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:14.064287Z digest=sha256:d699e33c1a2fe9b105fe9b1dd49f647ac9c0730c7832f43c8914fd484e7c343a

Observation b4bcc8d6-2459-47ba-b90b-71f01c45a501 · outbound

This paper cites RT-1: Robotics transformer for real-world control at scale,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RT-1: Robotics transformer for real-world control at scale,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.935088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:14.135586Z digest=sha256:254fa296bdde0a96a262fac556c345f072701ef8775becb6f025cda1106fb59c

Observation a9af96c5-bcb5-4b3d-a666-3695f9e631c1 · outbound

This paper cites π 0: A vision-language-action flow model for general robot control,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation π 0: A vision-language-action flow model for general robot control,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.924717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:14.302562Z digest=sha256:1d17643d1344aa00f8059c3756c714dda2b444ab93afc9db61de028a4dfc6b01

Observation 371af7de-dfa1-40a4-806f-f70e409abec5 · outbound

This paper cites FAST: Efficient action tokenization for vision- language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation FAST: Efficient action tokenization for vision- language-action models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.914178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:14.470124Z digest=sha256:01cade7ffd141a23ec7e5ad8e1b3f2076bdc62745cf7758f410ded0976b7511f

Observation 2c2b5ade-1432-48d0-830c-4575e227a1bf · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:14.652245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:14.652245Z digest=sha256:6e3f4d9067cd2fed2a2e3c4a35f32c486fb6b15f670f87656e9c52fdd807b6b8

Observation 371df10c-c4af-4cec-bc5e-716115ee111e · outbound

This paper cites ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:14.797218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:14.797218Z digest=sha256:c098bd8d95a1bee5df5ffb616599be400a4156774e1d20b73690634a9f6d1bbb

Observation 8cf5754c-5969-4399-95a6-b3c9f23c9da4 · outbound

This paper cites GEN-0: Embodied foundation models that scale with physical interaction,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GEN-0: Embodied foundation models that scale with physical interaction,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.902689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:14.955267Z digest=sha256:d09e0707fe01937ba1076251243eec3ac1277c2d4fd4591027415cc51a89c2be

Observation 5ddf7257-1253-40f0-a59d-5e17719dbaa1 · outbound

This paper cites GEN-1: Scaling embodied foundation models to mastery,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GEN-1: Scaling embodied foundation models to mastery,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.892286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.131468Z digest=sha256:9305d5b938a2b24d7d88015ac7c73a8932339e94b05c18823daea0e5ebcedc08

Observation ec1cd722-0cfb-4adb-babe-5a00236107cd · outbound

This paper cites GENE-26.5: Advancing robotic manipulation to human level,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GENE-26.5: Advancing robotic manipulation to human level,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.881643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.233329Z digest=sha256:8883ded18a5299dd381b0f3575c68df31f1f6f77d25ef6eae58c712df75d0653

Observation 821c082a-2eb8-4751-a1ee-4bec48253783 · outbound

This paper cites ACT-2 preview: Generalizing reliability,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation ACT-2 preview: Generalizing reliability,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.871047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.384527Z digest=sha256:43d16a13fe25522266214e04043b1849a88f544ea558450e514b57821a315426

Observation e1484038-fc1c-4daa-839f-7bfc9bf6ca4f · outbound

This paper cites Open X-Embodiment: Robotic learning datasets and RT-X models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Open X-Embodiment: Robotic learning datasets and RT-X models,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.860168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.548311Z digest=sha256:8c55cac9c5f7ea420eb2e586056578be25690bd14a4f02c2e7fc3d08bad64de1

Observation cd1f516d-f3ef-42f3-9164-897156e98ac9 · outbound

This paper cites PolarNet: 3D point clouds for language-guided robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PolarNet: 3D point clouds for language-guided robotic manipulation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.848674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.653628Z digest=sha256:ee49525c415c9f7830204bfe14b857b8f294b2c5695f24e80fde724d6996fece

Observation eb140a8d-ce8b-4c83-9b44-15e09803f790 · outbound

This paper cites M2T2: Multi-task masked transformer for object-centric pick and place,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation M2T2: Multi-task masked transformer for object-centric pick and place,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.837170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.829628Z digest=sha256:9fd3d3c1bd06e6c5c2e05815b09b766302ac5c45cf2ccee0a70872461ed21752

Observation c4caa6bb-3519-40db-92e1-62c5f229d026 · outbound

This paper cites FP3: A 3D Foundation Policy for Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation FP3: A 3D Foundation Policy for Robotic Manipulation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:16.040209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:16.040209Z digest=sha256:7819af653192fede6d12bba520dc6dfe075737158e0e93cc90bad5d2507446a2

Observation 12c3079f-f9fd-4bd7-b466-94c5828442f2 · outbound

This paper cites Coarse-to-fine Q- attention: Efficient learning for visual robotic manipulation via discreti- sation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Coarse-to-fine Q- attention: Efficient learning for visual robotic manipulation via discreti- sation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.826462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:16.225713Z digest=sha256:b3b92f96f6ba0a7c5f1aef65b036eadf62e96e122e3649771fdf60a73b6d0024

Observation deeacec7-9962-42cf-ab9b-33007842f9c6 · outbound

This paper cites PointVLA: Injecting the 3D world into vision-language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PointVLA: Injecting the 3D world into vision-language-action models,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.814607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:16.363898Z digest=sha256:b7eef5dc268925b4c13b67ec851f8a3009f48a4ee8361bbcf25282961f0f3f25

Observation 518fea34-2208-4803-80f6-e2a854ba2e3a · outbound

This paper cites Lift3D policy: Lifting 2D foundation models for robust 3D robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Lift3D policy: Lifting 2D foundation models for robust 3D robotic manipulation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.803293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:16.530929Z digest=sha256:0b74107e1ff57c4a62095f7996ca5963dbdaea7f5b95168347a0ed9973cc8a64

Observation deb3e36c-995d-4e7d-9f68-269b8fdf977a · outbound

This paper cites DINOv2: Learning robust visual features without supervision,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation DINOv2: Learning robust visual features without supervision,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.791935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:16.777096Z digest=sha256:8c7406bb3f3aa02f403814d984f4c2e5f636b62bb61b91c7582bf78e84387bd0

Observation dd8f4b74-cad6-4f31-b450-c15e922ec900 · outbound

This paper cites OG- VLA: Orthographic image generation for 3D-aware vision-language action model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation OG- VLA: Orthographic image generation for 3D-aware vision-language action model,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:16.921827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:16.921827Z digest=sha256:927511121e21be32e4e749e6f142b1da122bd36cf847b347d64642014d01616d

Observation 0c222d1c-b006-4f1c-a73e-1d935006267a · outbound

This paper cites Instruction-driven history-aware policies for robotic manipulations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Instruction-driven history-aware policies for robotic manipulations,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.011327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.011327Z digest=sha256:854c50270bd20fae68fd99fed7286d1cc701aa4c697e55b8205a4aa71402df02

Observation e8352d45-348c-455d-9e25-5fba4dab6404 · outbound

This paper cites Causal World Modeling for Robot Control.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Causal World Modeling for Robot Control

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.098344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.098344Z digest=sha256:f4b379da07429ba41d809293e18cf751b9bf8582727fd8c92a2ac6cd9b512d79

Observation 19f9ba14-2da4-423f-9532-901e793b4ed8 · outbound

This paper cites World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:44:18.473850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:17.241212Z digest=sha256:881c35aa1bde237c3b9a40a155d83b45c3d95d260d82500bd8c194f912c35202

Observation f0b9dd4f-6e94-459e-9229-cbfa958934ba · outbound

This paper cites MemoryWAM: Efficient World Action Modeling with Persistent Memory.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation MemoryWAM: Efficient World Action Modeling with Persistent Memory

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.373966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.373966Z digest=sha256:29e0653de4b9685a53e34d93e66fb8b607f88be77f7d901518df19eb2b8029e4

Observation 0e8265e4-ae1c-47cd-b236-f55b4fc3419c · outbound

This paper cites TraceVLA: Visual trace prompting enhances spatial- temporal awareness for generalist robotic policies,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation TraceVLA: Visual trace prompting enhances spatial- temporal awareness for generalist robotic policies,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.773326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:17.477048Z digest=sha256:a238aac2f5ae6a6061b087b97746a6f13ee019f868c7aeec7fe1580b72e70b27

Observation f3038fbd-6cc5-4a77-a2eb-dfc5a65702df · outbound

This paper cites RoboMemory: A brain-inspired multi-memory agentic framework for interactive environmental learning in physical embodied systems,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboMemory: A brain-inspired multi-memory agentic framework for interactive environmental learning in physical embodied systems,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.604914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.604914Z digest=sha256:0e58614e63e615ac3fd6266961d99e33d0944c5409570cc65158b3b9faecd4de

Observation 066b05a7-3096-4d78-a0b3-e3387bb86b98 · outbound

This paper cites MemoryVLA: Perceptual-cognitive memory in vision- language-action models for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation MemoryVLA: Perceptual-cognitive memory in vision- language-action models for robotic manipulation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.761545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:17.711839Z digest=sha256:8dd78987bf04f545b8db2650780562927763b11f20040bfdc623527a980497d7

Observation f8a91fcd-be30-4561-8774-439ae3e83a76 · outbound

This paper cites Gated Memory Policy: In-Context Memorization and Adaptation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Gated Memory Policy: In-Context Memorization and Adaptation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.858828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.858828Z digest=sha256:1800426d0ed0da787518017cae8a0ef71e50ca9e4157f32a5ce8382779979f35

Observation b26ee186-b1d0-4ab4-8c02-0d0e5df61fff · outbound

This paper cites You only scan once: A dynamic scene reconstruction pipeline for 6-DoF robotic grasping of novel objects,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation You only scan once: A dynamic scene reconstruction pipeline for 6-DoF robotic grasping of novel objects,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.750220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:17.940226Z digest=sha256:6b362e599bd1d1d17186306d3ccc8b855afc88baba7d2463c3faa45cc6d3815a

Observation 3619bfb8-af15-4b3b-91be-2ab580a5fb28 · outbound

This paper cites Mem-World: Memory-Augmented Action-Conditioned World Models for Persistent Robot Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Mem-World: Memory-Augmented Action-Conditioned World Models for Persistent Robot Manipulation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:44:18.373945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.000401Z digest=sha256:bc2ffd4690ac43c881d57891839288fff5a1297e14b2db0698454676e4f9a280

Observation 62d4a375-70d3-4ef6-b231-9adb322ae235 · outbound

This paper cites Coarse-to-fine imitation learning: Robot manipulation from a single demonstration,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Coarse-to-fine imitation learning: Robot manipulation from a single demonstration,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.103744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.103744Z digest=sha256:3053e659471944942bbfd16bf190e64d60dd82d06b88811ebee06cc2d8c4c36b

Observation c7340192-801b-4d8f-bdc4-a939dbc5ee48 · outbound

This paper cites RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.160988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.160988Z digest=sha256:88e2b9a9c8b8170dea8c7c2c2992f8462a9eb7966f51bd93c029658e508effe4

Observation 501a8a7a-ba8b-4f0f-b121-2bf46dfbbda3 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PaliGemma: A versatile 3B VLM for transfer

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.188176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.188176Z digest=sha256:d2dd5341cd8027f4b89896c9a75b57055c3218b1305126615fe82d395ce80904

Observation 1bcbf714-17a6-4890-b530-0d816e21e272 · outbound

This paper cites Sigmoid loss for language image pre-training,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Sigmoid loss for language image pre-training,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.200108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.200108Z digest=sha256:d57b642a92e5980739370f871eeb4ab60b48ff4ff0c400af3961c0528bf5c4b5

Observation 2bea7c9e-b7be-4b08-bd06-822a2d3a4ac7 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Gemma: Open Models Based on Gemini Research and Technology

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.203663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.203663Z digest=sha256:0b74a4058bdff1646d66816a9285b7f017594a5627c678b17322f9348f95d731

Observation 3db468e6-720e-4258-be7d-34058747e6b7 · outbound

This paper cites RAFT: Recurrent all-pairs field transforms for optical flow,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RAFT: Recurrent all-pairs field transforms for optical flow,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.722936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.207095Z digest=sha256:e2109780fa760c3b096ee268aa188e5fe69f0854f39772aa41caf057bc7241dd

Observation 6b10100f-42fc-4a01-b8d3-4adf5a1fce6c · outbound

This paper cites an unresolved cited work.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.210377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.210377Z digest=sha256:0f89d029377b3d88888474a01f17995321a4bfc1863227598db7e2055774f0bf

Observation 2ffc9fc0-45bf-4d03-b590-220ec82989dc · outbound

This paper cites On the continuity of rotation representations in neural networks,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation On the continuity of rotation representations in neural networks,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.213450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.213450Z digest=sha256:6981557b691ee5b6cafb3e49b23ccde6333126b8827242bb4c5b6244340b55d7

Observation 06e3c500-0a7f-4650-bbef-286b2a71ad30 · outbound

This paper cites V-REP: A versatile and scalable robot simulation framework,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation V-REP: A versatile and scalable robot simulation framework,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.694734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.216594Z digest=sha256:7a70375273740649486994a5bc3d5aace4b506d542a90056fd23f9dffdc88e68

Observation ad897fab-5612-4722-991c-6aaf338a20a4 · outbound

This paper cites Perceiver IO: A general architecture for structured inputs & outputs,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Perceiver IO: A general architecture for structured inputs & outputs,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.682780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.220226Z digest=sha256:c443c688e51180aa97818cf5619d6e7110621d8be236f178b953e33a153aadd2

Observation ec093242-9c4c-44f6-a66a-7587bc53cf7c · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Diffusion policy: Visuomotor policy learning via action diffusion,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.670900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.223196Z digest=sha256:4ac116192869651645dd4f1ef7f2cb454e6b77e5dcac93a526dee3d6cbaf8476

Observation e2fd6278-b836-44b6-b6a4-eb6503c2a138 · outbound

This paper cites Learning fine-grained bimanual manipulation with low-cost hardware,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Learning fine-grained bimanual manipulation with low-cost hardware,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.226213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.226213Z digest=sha256:d42f6a82dd17ec292bddc2780c9a2448f21e0c0a1b96ec88fec1adbe12d846f6

Observation 78055c9f-cb59-46fe-bac4-338dfdddb5e7 · outbound

This paper cites X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.229394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.229394Z digest=sha256:deff4dce0a239ed6b8a263ef86265b958c9966d9f65d5484ff7cd9db3508c8b4

Observation 748933a2-7d86-4b24-82ee-7c8ab8cb267d · outbound

This paper cites Fast-WAM: Do World Action Models Need Test-time Future Imagination?.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.232308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.232308Z digest=sha256:a228ba1f0356f64947fc3b6663c44c79b194fb169a7ed28c5cc16db776fb72f0

Observation 5477a936-baab-46d6-8daf-b8bae834b0e5 · outbound

This paper cites R3M: A Universal Visual Representation for Robot Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation R3M: A Universal Visual Representation for Robot Manipulation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.235552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.235552Z digest=sha256:0e15adf3d5427c665b9f2c5e84f259dd651f0d4405955afc12747ea4d46d9d72

Observation 222af3a0-3646-46b8-a56f-b3f57fb18cd0 · outbound

This paper cites Masked Visual Pre-training for Motor Control.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Masked Visual Pre-training for Motor Control

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.238620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.238620Z digest=sha256:059958a3c9c27971b06e6847f240abfd82a90222675c78f96d0af3038e2e7ec4

Observation e3908d42-3017-427a-bc38-7ad0c4f92bfd · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.241421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.241421Z digest=sha256:c2518eaab34f8c80b9cecbe1730359138bddb46892029643e2b7381f3ac3c3f8

Observation 09676efc-b973-4a1e-b88d-813e430e0d34 · outbound

This paper cites SAPIEN: A simulated part-based interactive environ- ment,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SAPIEN: A simulated part-based interactive environ- ment,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.650643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.244590Z digest=sha256:fafa302da57685fe06b448c5ff8d508ca0084e415cf1eb2de7e4cabea2c73297

Observation 884d5e23-93a0-4550-b4cc-d1e6055735fa · outbound

This paper cites Point transformer V3: Simpler faster stronger,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Point transformer V3: Simpler faster stronger,

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T10:44:18.638849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.247570Z digest=sha256:5c29effa7d5848a342d0b77ac2c4c49c73e8878fb2ec93c973d5cbcd752ccb23

Pith citing papers

No inbound Pith citation observations are available.