Pith. sign in

Paper Citation Record · LEDGER

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

As of 8 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2608.05042.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05042 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:44:18.247570Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact2
  • verified fuzzy35
  • unresolved27
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation da635c3e-14ba-454c-ad3a-f0147571b3d3 · outbound

This paper cites OpenVLA: An open-source vision-language-action model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation OpenVLA: An open-source vision-language-action model,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:11.836856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:11.836856Z digest=sha256:49ec477a5b90babe9cf0b167918ede21aa9be0de256a721dd28f023b088cbcac

Observation 30ef66ac-b196-4113-845a-b1c1b5b72135 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:11.920374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:11.920374Z digest=sha256:0205ce16d895204e082769a6feb8ece54a2b780d62dcbf1c4bcf818fac313d9a

Observation 4c4286e5-ac21-4ebc-becb-0504efc4af64 · outbound

This paper cites Wall-OSS-0.5 Technical Report.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Wall-OSS-0.5 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:12.106434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:12.106434Z digest=sha256:0427e364f85832770e355d241115d06525e2ba5c63f08035f77ceb68d0800023

Observation 9ae2648e-564a-4774-8ef4-35d5a4bf726b · outbound

This paper cites Vision-language foundation models as effective robot imitators,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Vision-language foundation models as effective robot imitators,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.055685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.235379Z digest=sha256:d752694031846128510cf76e7acac19cb051d70b242f23db5ae58ab997c9b1f3

Observation 72f95a25-9bc3-4002-aaa5-25cae51bd56a · outbound

This paper cites RT-2: Vision-language-action models transfer web knowledge to robotic control,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RT-2: Vision-language-action models transfer web knowledge to robotic control,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.046255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.390934Z digest=sha256:f89ab01dad2a28118a7f02b849618320ea1b90a09e4c87247e1c353c010a1bfb

Observation 36ec393a-1107-4976-8e73-5b110a128bc7 · outbound

This paper cites Perceiver-Actor: A multi-task transformer for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Perceiver-Actor: A multi-task transformer for robotic manipulation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.036180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.523105Z digest=sha256:9d962cb759bebafb75e88816ee715cdd8facae28070526945cf30f7b24c25edb

Observation 62fa746e-6e30-4d58-bd9a-fbd27cc63775 · outbound

This paper cites 3D Diffuser Actor: Policy diffusion with 3D scene representations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation 3D Diffuser Actor: Policy diffusion with 3D scene representations,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.027748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.644589Z digest=sha256:e4d208231a6ea17878697dffaf875ba883b38cf8dd9c0864bcbabf9bcf1d7400

Observation 8b5524d3-a092-490e-a37f-9b0b724fa9be · outbound

This paper cites Act3D: 3D feature field transformers for multi-task robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Act3D: 3D feature field transformers for multi-task robotic manipulation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.019261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.852518Z digest=sha256:38c6cd094b52150c7c66cdae959e0b710797c5ae0e77848c18978c9ae76c6444

Observation c172a669-c510-4cd3-84cc-a552d268c8b2 · outbound

This paper cites RVT: Robotic view transformer for 3D object manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RVT: Robotic view transformer for 3D object manipulation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.010416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:12.979457Z digest=sha256:97f25c676f2677ad77b492bb422880dd2aa0091f423b6bb1dab000361e91140d

Observation 132b4569-ec0d-4f4b-9f54-2f83a980c37f · outbound

This paper cites RVT-2: Learning precise manipulation from few demonstrations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RVT-2: Learning precise manipulation from few demonstrations,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.001863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.078472Z digest=sha256:7991b20da6ff12d5ce1a3326768fd212903bbe29c8a9d0e1bc9c1487d4777d65

Observation 111db71a-e83c-4bf4-ae07-e402e17c8359 · outbound

This paper cites 3D-VLA: A 3D vision-language-action generative world model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation 3D-VLA: A 3D vision-language-action generative world model,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.993333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.210283Z digest=sha256:050b989aaae770cb5e4021269be74b443d7aa1002890407e501a830f9dccfe00

Observation abec4001-6626-4f0a-aa60-e1d5f1aae51a · outbound

This paper cites SpatialVLA: Exploring spatial representations for visual- language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SpatialVLA: Exploring spatial representations for visual- language-action models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.984740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.350018Z digest=sha256:a2fbdede49723bd24dbb00f8b5463773d0adf2c7685a1b2ee0511c3941188a17

Observation e7176949-b0a9-44bd-8dc7-6660c0d0d8c7 · outbound

This paper cites RLBench: The robot learning benchmark & learning environment,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RLBench: The robot learning benchmark & learning environment,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.976335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.499499Z digest=sha256:db6aa049e58645d74a9df398549d304928b6ae35dbb6c93e3e0aaf304d9b9c91

Observation edbb9f7e-c70c-4f75-96bf-10ee976757d3 · outbound

This paper cites THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:13.672028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:13.672028Z digest=sha256:ca3eeda22d6eba098423b08a9f8bc345babf9e0b4c6a77b5f2be7229f5dc4445

Observation 6973dddc-827f-46de-a672-d9483518cd54 · outbound

This paper cites Towards generalizable vision- language robotic manipulation: A benchmark and LLM-guided 3D policy,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Towards generalizable vision- language robotic manipulation: A benchmark and LLM-guided 3D policy,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.966940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.754442Z digest=sha256:c9e7d404b5e64f25dd79b8a16fcb06b8da2d9c053fc139d6b5889798cce16c05

Observation fe57631e-5ac7-41f1-a37f-d84c9630ec0d · outbound

This paper cites RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:13.849049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:13.849049Z digest=sha256:6d0ce5b39c65d0d53c648256fd36290816d685b19e05aef1be47843cfeb542f8

Observation 10fe969f-77ba-4af8-8513-e8fb104d71cf · outbound

This paper cites SAM2Act: Integrating visual foundation model with a memory architecture for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SAM2Act: Integrating visual foundation model with a memory architecture for robotic manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.956456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:13.970866Z digest=sha256:f11e066c716b748e524c32c36579f2fc85ed97309ad77aeeffff0ea9f50e1da1

Observation 496c8dc8-a445-49a6-8865-e1fe4a147bc7 · outbound

This paper cites BridgeVLA: Input-output alignment for efficient 3D ma- nipulation learning with vision-language models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation BridgeVLA: Input-output alignment for efficient 3D ma- nipulation learning with vision-language models,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.945963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:14.064287Z digest=sha256:90ca3aabc61b2eed09dc1b59bd29d85fd162c6eb7b1538fa54a53c273f3cd1bb

Observation b4bcc8d6-2459-47ba-b90b-71f01c45a501 · outbound

This paper cites RT-1: Robotics transformer for real-world control at scale,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RT-1: Robotics transformer for real-world control at scale,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.935088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:14.135586Z digest=sha256:727049531d74b1319b8d28f0975d4e91f3c42916c5e26c79a75ea75cfef0a2ac

Observation a9af96c5-bcb5-4b3d-a666-3695f9e631c1 · outbound

This paper cites π 0: A vision-language-action flow model for general robot control,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation π 0: A vision-language-action flow model for general robot control,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.924717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:14.302562Z digest=sha256:a7820e62f1d378e58cd02e70c603a7f829327050d4295bf6cfae71e675853709

Observation 371af7de-dfa1-40a4-806f-f70e409abec5 · outbound

This paper cites FAST: Efficient action tokenization for vision- language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation FAST: Efficient action tokenization for vision- language-action models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.914178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:14.470124Z digest=sha256:5eb5ab419eeb1b048ec988545a96a46f93c62c0a60be772af29a22798a1d9a78

Observation 2c2b5ade-1432-48d0-830c-4575e227a1bf · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:14.652245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:14.652245Z digest=sha256:fb5aad32564c0cfd5ee2f3735dd96538047abfa95a031dc3a1beac70de7d5dc4

Observation 371df10c-c4af-4cec-bc5e-716115ee111e · outbound

This paper cites ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:14.797218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:14.797218Z digest=sha256:5810251a6c64e44d9d6403166d0ccb46215c3d96b501610a4fdf6830c2d1e721

Observation 8cf5754c-5969-4399-95a6-b3c9f23c9da4 · outbound

This paper cites GEN-0: Embodied foundation models that scale with physical interaction,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GEN-0: Embodied foundation models that scale with physical interaction,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.902689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:14.955267Z digest=sha256:e3d121d4fe24812e1b74298fb99e7a5363b525646fd1388aa47ac4486d717007

Observation 5ddf7257-1253-40f0-a59d-5e17719dbaa1 · outbound

This paper cites GEN-1: Scaling embodied foundation models to mastery,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GEN-1: Scaling embodied foundation models to mastery,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.892286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.131468Z digest=sha256:654ff7c17f850722ecfa91bd79d02bf66313dd341af7c1b8373c0d3fc341edef

Observation ec1cd722-0cfb-4adb-babe-5a00236107cd · outbound

This paper cites GENE-26.5: Advancing robotic manipulation to human level,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GENE-26.5: Advancing robotic manipulation to human level,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.881643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.233329Z digest=sha256:a4d6c6226ff59cf3af3603cff3071e79c61bf25911f81f4aa48004b830a93d9c

Observation 821c082a-2eb8-4751-a1ee-4bec48253783 · outbound

This paper cites ACT-2 preview: Generalizing reliability,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation ACT-2 preview: Generalizing reliability,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.871047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.384527Z digest=sha256:876e27702e1f28218a36e85e1789ea452a27522b84f4f1019c66eb326d84e9db

Observation e1484038-fc1c-4daa-839f-7bfc9bf6ca4f · outbound

This paper cites Open X-Embodiment: Robotic learning datasets and RT-X models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Open X-Embodiment: Robotic learning datasets and RT-X models,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.860168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.548311Z digest=sha256:7fdeaee96aeb59888759779d1e9c3d4ce34c01fb79cc0cad498241516f3e078e

Observation cd1f516d-f3ef-42f3-9164-897156e98ac9 · outbound

This paper cites PolarNet: 3D point clouds for language-guided robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PolarNet: 3D point clouds for language-guided robotic manipulation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.848674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.653628Z digest=sha256:1d312e414f202c9a4e522a52f33bcc19f4b629784b69e8e6e9499b66f82052b8

Observation eb140a8d-ce8b-4c83-9b44-15e09803f790 · outbound

This paper cites M2T2: Multi-task masked transformer for object-centric pick and place,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation M2T2: Multi-task masked transformer for object-centric pick and place,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.837170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:15.829628Z digest=sha256:1241be9348c9b824ac0d766591a28fa4c8ecd23de79d4a984cccbb68cbf6babe

Observation c4caa6bb-3519-40db-92e1-62c5f229d026 · outbound

This paper cites FP3: A 3D Foundation Policy for Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation FP3: A 3D Foundation Policy for Robotic Manipulation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:16.040209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:16.040209Z digest=sha256:36830345b51ac3779a69e1ad62e817b50cfd30857eb66e04adce2b33d82861a2

Observation 12c3079f-f9fd-4bd7-b466-94c5828442f2 · outbound

This paper cites Coarse-to-fine Q- attention: Efficient learning for visual robotic manipulation via discreti- sation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Coarse-to-fine Q- attention: Efficient learning for visual robotic manipulation via discreti- sation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.826462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:16.225713Z digest=sha256:0aedc89c055ef7fa59ad20be471f377e387c2ae2aeb8c161431de395249e0024

Observation deeacec7-9962-42cf-ab9b-33007842f9c6 · outbound

This paper cites PointVLA: Injecting the 3D world into vision-language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PointVLA: Injecting the 3D world into vision-language-action models,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.814607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:16.363898Z digest=sha256:6e86364dfef5e01a413fcb5f8c4f631ede131a1b70aa4a3c6020ba2dceee3168

Observation 518fea34-2208-4803-80f6-e2a854ba2e3a · outbound

This paper cites Lift3D policy: Lifting 2D foundation models for robust 3D robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Lift3D policy: Lifting 2D foundation models for robust 3D robotic manipulation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.803293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:16.530929Z digest=sha256:8550304d83c7c9c1a8eee9b6e9a165220682b6c2406d39574c771e197d2a7f38

Observation deb3e36c-995d-4e7d-9f68-269b8fdf977a · outbound

This paper cites DINOv2: Learning robust visual features without supervision,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation DINOv2: Learning robust visual features without supervision,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.791935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:16.777096Z digest=sha256:0827f95f61450ea4b7f50372a7c2b3e7075cd4377b8653cc3805cc1a44334517

Observation dd8f4b74-cad6-4f31-b450-c15e922ec900 · outbound

This paper cites OG- VLA: Orthographic image generation for 3D-aware vision-language action model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation OG- VLA: Orthographic image generation for 3D-aware vision-language action model,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:16.921827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:16.921827Z digest=sha256:74a7e4feb9a1cff7021631bcac870bfab183c49fbb7182357cc5245c77bc2e99

Observation 0c222d1c-b006-4f1c-a73e-1d935006267a · outbound

This paper cites Instruction-driven history-aware policies for robotic manipulations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Instruction-driven history-aware policies for robotic manipulations,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.011327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.011327Z digest=sha256:6e0b739b7f00ea8cfad9172f2d54b4d1876881074b5fb19ce735e9dc095ffe21

Observation e8352d45-348c-455d-9e25-5fba4dab6404 · outbound

This paper cites Causal World Modeling for Robot Control.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Causal World Modeling for Robot Control

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.098344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.098344Z digest=sha256:640d9fb3a2ac40c2be71ccc0a7c50d5d223f56309bad2fd6f5b6878e25d2ce8e

Observation 19f9ba14-2da4-423f-9532-901e793b4ed8 · outbound

This paper cites World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:44:18.473850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:17.241212Z digest=sha256:d7387f75fa86c286cb0c7e56f8795b2007e933b670627450df09c555bf2a0f91

Observation f0b9dd4f-6e94-459e-9229-cbfa958934ba · outbound

This paper cites MemoryWAM: Efficient World Action Modeling with Persistent Memory.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation MemoryWAM: Efficient World Action Modeling with Persistent Memory

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.373966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.373966Z digest=sha256:3e0ed763672177527918060ce3725db738bc445359856a3c07f813050f592c49

Observation 0e8265e4-ae1c-47cd-b236-f55b4fc3419c · outbound

This paper cites TraceVLA: Visual trace prompting enhances spatial- temporal awareness for generalist robotic policies,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation TraceVLA: Visual trace prompting enhances spatial- temporal awareness for generalist robotic policies,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.773326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:17.477048Z digest=sha256:6e830f5f36e2c39b6b13578d73ab2dfe8bc17df5c868a2018dd5ce61ddb2af3a

Observation f3038fbd-6cc5-4a77-a2eb-dfc5a65702df · outbound

This paper cites RoboMemory: A brain-inspired multi-memory agentic framework for interactive environmental learning in physical embodied systems,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboMemory: A brain-inspired multi-memory agentic framework for interactive environmental learning in physical embodied systems,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.604914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.604914Z digest=sha256:eeab5d524396eb5080b46e42c36280ce0e0f67ff6585bd82f19da3532f553b75

Observation 066b05a7-3096-4d78-a0b3-e3387bb86b98 · outbound

This paper cites MemoryVLA: Perceptual-cognitive memory in vision- language-action models for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation MemoryVLA: Perceptual-cognitive memory in vision- language-action models for robotic manipulation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.761545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:17.711839Z digest=sha256:5a2032bc7ee66268f4c2b8a4ac5210a6d46b044f12ffc3ff9d112b5724faed75

Observation f8a91fcd-be30-4561-8774-439ae3e83a76 · outbound

This paper cites Gated Memory Policy: In-Context Memorization and Adaptation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Gated Memory Policy: In-Context Memorization and Adaptation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.858828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.858828Z digest=sha256:70d8af23946fe1e7c2d6868a21dfd1bf15e242fce7f1b4a5318eea1acbea6226

Observation b26ee186-b1d0-4ab4-8c02-0d0e5df61fff · outbound

This paper cites You only scan once: A dynamic scene reconstruction pipeline for 6-DoF robotic grasping of novel objects,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation You only scan once: A dynamic scene reconstruction pipeline for 6-DoF robotic grasping of novel objects,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.750220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:17.940226Z digest=sha256:0adf0c64b871f79ec1b64fe9f0f974008148aa59a1a18f4ec97847281c26cd0d

Observation 3619bfb8-af15-4b3b-91be-2ab580a5fb28 · outbound

This paper cites Mem-World: Memory-Augmented Action-Conditioned World Models for Persistent Robot Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Mem-World: Memory-Augmented Action-Conditioned World Models for Persistent Robot Manipulation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:44:18.373945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.000401Z digest=sha256:bfd94d62c67629c813486c3a104903140abd86a03530e35230690d7fee82c9d8

Observation 62d4a375-70d3-4ef6-b231-9adb322ae235 · outbound

This paper cites Coarse-to-fine imitation learning: Robot manipulation from a single demonstration,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Coarse-to-fine imitation learning: Robot manipulation from a single demonstration,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.103744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.103744Z digest=sha256:8d133253df9d4ade989dd43520e5adef15819356e44f3f39e9a03683a2118fc4

Observation c7340192-801b-4d8f-bdc4-a939dbc5ee48 · outbound

This paper cites RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.160988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.160988Z digest=sha256:254f79565f2a14188a862126bf0cb893d76e6b3a08ea0a3becb8a4d4f82c6669

Observation 501a8a7a-ba8b-4f0f-b121-2bf46dfbbda3 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PaliGemma: A versatile 3B VLM for transfer

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.188176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.188176Z digest=sha256:13deee7038952a293aa58f6b6ece6d1aa8420c46387703efca9bace9c73dcc59

Observation 1bcbf714-17a6-4890-b530-0d816e21e272 · outbound

This paper cites Sigmoid loss for language image pre-training,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Sigmoid loss for language image pre-training,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.200108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.200108Z digest=sha256:7dbdc44dbcba27a624048c7b9adbc6a8774e438263788c66686047661c704e08

Observation 2bea7c9e-b7be-4b08-bd06-822a2d3a4ac7 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Gemma: Open Models Based on Gemini Research and Technology

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.203663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.203663Z digest=sha256:72b8da78396e66edfc03080ad0171705322b0b495d5d5cf2759975e3344600cd

Observation 3db468e6-720e-4258-be7d-34058747e6b7 · outbound

This paper cites RAFT: Recurrent all-pairs field transforms for optical flow,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RAFT: Recurrent all-pairs field transforms for optical flow,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.722936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.207095Z digest=sha256:0be0a7acf50c306dcdebe553ac64c144dcd7c107bdc246ca511a05fe7e21b634

Observation 6b10100f-42fc-4a01-b8d3-4adf5a1fce6c · outbound

This paper cites an unresolved cited work.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.210377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.210377Z digest=sha256:d32c6e09c018c40180e71db69dc742262ab4b0223ef8294f8dbdc8895e4f0050

Observation 2ffc9fc0-45bf-4d03-b590-220ec82989dc · outbound

This paper cites On the continuity of rotation representations in neural networks,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation On the continuity of rotation representations in neural networks,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.213450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.213450Z digest=sha256:2a97dd614a51319ea9aa4646c9fbaa87657bdba9202e670e44ef621b5e2a8e8b

Observation 06e3c500-0a7f-4650-bbef-286b2a71ad30 · outbound

This paper cites V-REP: A versatile and scalable robot simulation framework,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation V-REP: A versatile and scalable robot simulation framework,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.694734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.216594Z digest=sha256:23032d671abc5a631acda23fe004ff009548dab8364e3a96c3e4f80ac0692fa3

Observation ad897fab-5612-4722-991c-6aaf338a20a4 · outbound

This paper cites Perceiver IO: A general architecture for structured inputs & outputs,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Perceiver IO: A general architecture for structured inputs & outputs,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.682780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.220226Z digest=sha256:190cbedd82c5d11c1616c34882588f84a2e5c27c048f495ad7b30a79f7d2d5b0

Observation ec093242-9c4c-44f6-a66a-7587bc53cf7c · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Diffusion policy: Visuomotor policy learning via action diffusion,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.670900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.223196Z digest=sha256:2418773f39c795c50068295232308f8d247f0d735f04056d707e7a728d8f02fb

Observation e2fd6278-b836-44b6-b6a4-eb6503c2a138 · outbound

This paper cites Learning fine-grained bimanual manipulation with low-cost hardware,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Learning fine-grained bimanual manipulation with low-cost hardware,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.226213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.226213Z digest=sha256:e80e7d3e7bde33ec4969fe249b46e70592686894f9939453a47f1324d27a18fe

Observation 78055c9f-cb59-46fe-bac4-338dfdddb5e7 · outbound

This paper cites X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.229394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.229394Z digest=sha256:b07707a59ed83bc05f53a106aa276798a1870af5878c84c41da2377b77bb563a

Observation 748933a2-7d86-4b24-82ee-7c8ab8cb267d · outbound

This paper cites Fast-WAM: Do World Action Models Need Test-time Future Imagination?.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.232308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.232308Z digest=sha256:e14b5d832da5989002c3e7f58bd980ab201998905817142b8152e8eee7209c6a

Observation 5477a936-baab-46d6-8daf-b8bae834b0e5 · outbound

This paper cites R3M: A Universal Visual Representation for Robot Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation R3M: A Universal Visual Representation for Robot Manipulation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.235552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.235552Z digest=sha256:53ae1520f3c0b553f4ab5fcee58250ba1e1dd622ba73e5be4836131016ee79a6

Observation 222af3a0-3646-46b8-a56f-b3f57fb18cd0 · outbound

This paper cites Masked Visual Pre-training for Motor Control.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Masked Visual Pre-training for Motor Control

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.238620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.238620Z digest=sha256:122be395445f2c6c6ebf693f125c32d3ca33df933c79f8e931572a6e82373092

Observation e3908d42-3017-427a-bc38-7ad0c4f92bfd · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.241421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.241421Z digest=sha256:3f771f3c0f258869458efbf4b02865a571ccddf487710f2a24e83aba314f4c76

Observation 09676efc-b973-4a1e-b88d-813e430e0d34 · outbound

This paper cites SAPIEN: A simulated part-based interactive environ- ment,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SAPIEN: A simulated part-based interactive environ- ment,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.650643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.244590Z digest=sha256:64d0864794bd9f7f50b00598b8ac8d9b204a65adbcb07d9a134ce55c775930ad

Observation 884d5e23-93a0-4550-b4cc-d1e6055735fa · outbound

This paper cites Point transformer V3: Simpler faster stronger,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Point transformer V3: Simpler faster stronger,

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T10:44:18.638849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:44:18.247570Z digest=sha256:4dde5a1b930dd71ea247cc1be94c1d52fbc1e7d252e6381613ada01e44163fc2

Pith citing papers

No inbound Pith citation observations are available.