Pith. sign in

Paper Citation Record · LEDGER

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models

As of 15 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 0 inbound Pith citation observations for arXiv:2608.06729.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06729 v1

Coverage vector

measured 87 of 87 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:33:40.117731Z

measured 87 of 87 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

87 of 87 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved80
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae7e2e93-53ed-489e-8bf6-92381f8214a2 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models PaliGemma: A versatile 3B VLM for transfer

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.148006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.148006Z digest=sha256:26fd5c15e05c985a0c43babdbf076624677ecb9320558a20a7fef52e8206e450

Observation 1763a70b-8f94-4838-aea3-9f4acd2bf58d · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.153666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.153666Z digest=sha256:fa7800a5c4865429115625951ebbc5fc6b83313bce27610ed7f153d840661a2b

Observation d0729d75-683b-4540-9d75-9442a5284b0f · outbound

This paper cites IEEE Robotics and Automation Letters , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models IEEE Robotics and Automation Letters , year=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.158941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.158941Z digest=sha256:c0dcd29b68971d0268209f0b55b6ae0bef5e6ee4b1b94b58eb2417611fe2f262

Observation 7c7b2dd4-08cb-4a5a-bfa4-2e82a53d2a33 · outbound

This paper cites What Matters in Building Vision-Language-Action Models for Generalist Robots.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models What Matters in Building Vision-Language-Action Models for Generalist Robots

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.183731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.183731Z digest=sha256:b282e010d64761ef5ccccc41df86efea96f340c2d9ca436a1490accd8c868a23

Observation ea6c5f2f-8764-428a-a367-1c6c15cd2cb9 · outbound

This paper cites HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.268148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.268148Z digest=sha256:f4b88463443d7ea671237ff674489a105432626b68588d777c459f385b21716a

Observation 0cfc02c2-dbc6-46fb-9324-08180790e94d · outbound

This paper cites an unresolved cited work.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.283509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.283509Z digest=sha256:bd4889b459420e97ac726c2d916cba3e9680d65750e7340e12f0c042235c61df

Observation efaa64dd-ad82-4936-bd72-302bd4c7eb17 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.289124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.289124Z digest=sha256:238ab482de45cdd594817ad6fab32a5b222c27e18b93f026c190df8a831a1e56

Observation ca3c8834-8946-4a8f-87ad-cadb67b3be5c · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.293367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.293367Z digest=sha256:78c39b54a9c537975c1664fd845a2b58f0096738af17a42598e361f2ab330b3d

Observation a159239b-43ff-4994-88b1-4d2cf859d8aa · outbound

This paper cites Conference on Robot Learning , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Conference on Robot Learning , pages=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.297762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.297762Z digest=sha256:52a9dd8a362648b2dce28e873751c90e601765c4d381e6284ab1228a7e2e20b5

Observation 91dea7aa-eeec-4f30-98dc-51c2f9f7398f · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.381582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.381582Z digest=sha256:954c83f32428d07bd48c420f80245caf2e9c13dfbc8ba0524d1e4760cae7db53

Observation f9e79d84-e7ae-474c-a7ec-c7e8ea9ee1c9 · outbound

This paper cites RT-H: Action Hierarchies Using Language.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RT-H: Action Hierarchies Using Language

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.488119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.488119Z digest=sha256:3d0d3a6f0f861792cada308fa6d8551de1847b6b7f93650a469a523da4af6359

Observation 353f80cf-62d5-4ade-8b96-1a9c810582fa · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models LLaMA: Open and Efficient Foundation Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.492031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.492031Z digest=sha256:d34bc08c56c982b97b62010438e4e9195c332371a1455cc4c205e60ee6fe2718

Observation 8cc9df09-01a3-4ae3-ac50-3f0d41b3d231 · outbound

This paper cites an unresolved cited work.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.497050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.497050Z digest=sha256:2e8d365a905929c543728858a2ecb2f8151a6c7d52b89d00885a26b998ec46ee

Observation 8a0c119c-abc1-4766-bacc-09239d225052 · outbound

This paper cites Advances in neural information processing systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in neural information processing systems , volume=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.501329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.501329Z digest=sha256:db4bffe22c5bd1d4231fa0cdf1069db987ea6a6d0d528b19b6320d3229c5c2c9

Observation 4d22abc2-52d9-4713-8225-26bfc05b6e51 · outbound

This paper cites 2023 , journal =.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models 2023 , journal =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.505225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.505225Z digest=sha256:2040b8f64ccf42860cd50bbfefedf3fce9d4681c2a6a3b72ee3da9113d03377c

Observation 99f222f3-f23f-475d-a98b-c0c742f15921 · outbound

This paper cites Advances in neural information processing systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in neural information processing systems , volume=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.619247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.619247Z digest=sha256:9067b237a5f148f00e4e3d5a5c8b3e5ad9fdecef89832f7ab0099d5937c5ad43

Observation 8408abc6-ef5b-44b8-a463-a4cd5cf473bc · outbound

This paper cites an unresolved cited work.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:33:41.423444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:38.683398Z digest=sha256:b8aac1279887ec6a4f88ef659d9bcd5fb8fa38cd38786e4aa7410e5412f5aa68

Observation 418df141-abb1-4f02-91a3-f3b5585fedc6 · outbound

This paper cites Computer Vision -.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Computer Vision -

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.413259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:38.686536Z digest=sha256:7bd6817d257d55215bbd22787db946fe306f916545085400059034eee5385254

Observation 025c6806-f067-4b91-8594-41212a13192a · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.691558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.691558Z digest=sha256:b4f9a06484cbaca8cba83183b89010ecd5135c2a3e4f852a877d272a9dc391bc

Observation 684f71ff-c72e-4af1-9211-ad0792094b98 · outbound

This paper cites GeoVLA: Empowering 3D Representations in Vision-Language-Action Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.696655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.696655Z digest=sha256:aa7b88c30df209a6ffdf8f20cd7b2aebd195ef3e868ce15a8ba00d672e4cdc6f

Observation 2c0dd2ac-bf53-4dcb-afb7-27fb5a82e3bc · outbound

This paper cites IEEE Robotics and Automation Letters , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models IEEE Robotics and Automation Letters , volume=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.700126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.700126Z digest=sha256:eaa6e7b67bdfdb1934bb6caddec270b853917d3b13fad6c88bdf61cb61756d89

Observation a2e0bae6-d0d5-4e6e-8386-f08a305fa272 · outbound

This paper cites arXiv preprint arXiv:2506.07961 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2506.07961 , year=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.794118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.794118Z digest=sha256:30fa86a3c19a31506ec27703b304b5efcb0e8f0e91e95fb90d6255d1b6dbaf00

Observation 3bed30ce-dd15-4cf7-9159-d8450e8655fc · outbound

This paper cites arXiv preprint arXiv:2506.22242 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2506.22242 , year=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.870725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.870725Z digest=sha256:d6a094b2e6ee18ca38eea3d86dcd169967f2d3cd1933b231d719e73ec24b82aa

Observation ffa41d8a-b812-453b-b475-eec3757968b8 · outbound

This paper cites arXiv preprint arXiv:2510.17439 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2510.17439 , year=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.875114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.875114Z digest=sha256:04a437f41ee58608905a2873d7f7fa1ca03d29c0440cfc4bf2f9c627d9b7c847

Observation e9beff1d-e16c-4b87-acc8-b52f80ebefbe · outbound

This paper cites 3D-VLA: A 3D Vision-Language-Action Generative World Model.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.878597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.878597Z digest=sha256:cc26a7d9eeaa0d011c8199ed60c72a0ad3f666a56180b4f269ff64c6bef2c8e1

Observation 8a05c435-c8a2-49b0-93f9-5a3ec9b93b0a · outbound

This paper cites arXiv preprint arXiv:2507.00416 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2507.00416 , year=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.882575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.882575Z digest=sha256:413a14245532b98d2f2ac71b68fc27ef42dc6470d3033c020437d399aa620075

Observation 801b711f-8a16-467e-b694-942cdcf552df · outbound

This paper cites AlphaMath Almost Zero: Process Supervision without Process.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models AlphaMath Almost Zero: Process Supervision without Process

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.971255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.971255Z digest=sha256:a64b9eaa3a6873fd8079c781021146eab209c2f2d91d3461e683f3b942a4f6cc

Observation 974d7a01-3d29-4984-aff4-c142f277c92a · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2025 , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Findings of the Association for Computational Linguistics: ACL 2025 , pages=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.975423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.975423Z digest=sha256:e865af09008bc80b4fc440dd210ae5dc361779a1b84b28c63fd4e54ddb935041

Observation 7f19ee55-8d2b-410b-bf90-83d05d1f441c · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.979347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.979347Z digest=sha256:a018c91ee9c1fa0127eaac0ba306eedd674c8f4f669a3d8441783e97d798f392

Observation 6e95a505-16e4-4cda-bfd2-331d85b41f05 · outbound

This paper cites Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.389106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:38.983031Z digest=sha256:0b53e2d655df307c7f9df6f2617e5e13036259022cf93ba23842b66ab31935c0

Observation 921a7363-0122-470c-b298-c0b43d6364cd · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.033472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.033472Z digest=sha256:d8054635ecee6da8968c66c6312871c6eeb4e42b7af70c471e468288aed4ca84

Observation 1bcaec27-b361-476f-b672-df4542556946 · outbound

This paper cites Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.091894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.091894Z digest=sha256:bd48b5a391f384ec77cee418884689dfc0f13c9be29a123d3eae0bd03bbc711a

Observation dd585a45-c7b8-4582-8fed-cfd5017f5521 · outbound

This paper cites RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.096472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.096472Z digest=sha256:da653de226136fe8619830ed5cdf8237d666fccc37cfabb083e932be2e9ded93

Observation 6fcb863b-8d34-49f3-bdb6-5348083076a1 · outbound

This paper cites Verifier-free Test-Time Sampling for Vision-Language-Action Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Verifier-free Test-Time Sampling for Vision-Language-Action Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.100502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.100502Z digest=sha256:d3dafff52c097f4bab9664e78ed3fe5da536bd425e0e5687edd8357efcdcd40a

Observation 28bfe46d-e9f3-482b-9162-6da73869dcc0 · outbound

This paper cites arXiv preprint arXiv:2510.10975 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2510.10975 , year=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.104348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.104348Z digest=sha256:4c85d6ef06c8036295560b23c6da5b46dd6479e18870fdf16f164931b9c0467c

Observation 24dccad6-a362-4f62-9ccb-e2b96c7d4114 · outbound

This paper cites arXiv preprint arXiv:2601.00675 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2601.00675 , year=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.160032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.160032Z digest=sha256:e48a083fcdb549a203d342fac27dadb2abd87e463f4dd54b830c415b14908d7a

Observation b008e761-fc45-4e5e-b572-177f675dd105 · outbound

This paper cites DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.203775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.203775Z digest=sha256:ad8b05ff31f820c9c282458f9697d7cee08a0330b9ccf563b6265952be7f1a29

Observation 0111b435-3b3b-44a4-88d2-9353815a56a2 · outbound

This paper cites Conference on Robot Learning , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Conference on Robot Learning , pages=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.209322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.209322Z digest=sha256:7d3d7afb8ae4d6823cd7132c7bdbb866698a688569b478aa811231b4bb076aac

Observation 87c4d7ed-efee-4e1d-9445-c5caa928a19d · outbound

This paper cites RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.212850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.212850Z digest=sha256:809d148237c874f6623d5e69439c2b67500248805d7798bdabb4060eb06fc033

Observation 80dd6fb4-2975-4007-b080-c3fba28930fb · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in Neural Information Processing Systems , volume=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.262864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.262864Z digest=sha256:db69ad4ebb68423df1fc7bb52454f1d5cdc72e9cdd2dfd61a27cc34aaaab8508

Observation ec6d4e48-53ec-4841-b49e-765388534aab · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.287949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.287949Z digest=sha256:dd4d2d8ecf49cb0ab02b60d21ae1ca763bc6d526a5146ff794ca383a2d39dacd

Observation bc59b65e-895c-4e78-b10c-449d389761e1 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.292896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.292896Z digest=sha256:9719bc785a3a90438b13b7dc964cbeb110e308f97d1046792ebc221b89e68a1d

Observation ffea7ef0-1caf-4aaa-a27e-7decbd33ab69 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.297696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.297696Z digest=sha256:a30fa37a25fc2c1b105b9db46a1125b924ef6cc3c3e826dafadc6f8ab316f156

Observation 23d72755-8f67-4d36-bd1b-4d2bfaa0ebd8 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.301691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.301691Z digest=sha256:8cd116f39f0c99ec9e3424a1364f1c0eff304f41d33cc32f656d7d25b8d072c4

Observation 01f7e032-03b0-4ca9-9032-83978cc94e92 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in Neural Information Processing Systems , volume=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.353990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.353990Z digest=sha256:55d79f135a0b020fa98970025e7fd316a878b838a75566d0e9f00c1d159e5cbe

Observation 6cb7f7ad-b15e-40c1-8bf0-485001dbee02 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.358576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.358576Z digest=sha256:3fbfaf23c02a057a58d1c6a33338078d804e8493c88325bbccafab2ac9b6d4a8

Observation b4ffd6d4-740b-4a24-8bbf-a2c81d120ec7 · outbound

This paper cites ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.362769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.362769Z digest=sha256:d40975c24db6ffe6676139de3dbb3097a6ad2d981bdabdf728b1d91996a0526c

Observation 62f9fa2f-2332-47c7-b234-eb9ccb9992b9 · outbound

This paper cites SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.367292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.367292Z digest=sha256:dab33218258fe2fe21a97e6ed3361ac1027284e502133a232295cff74b82e6cf

Observation 8c311ac6-cc58-4af1-b442-d92bed8e7ea4 · outbound

This paper cites Evaluating Real-World Robot Manipulation Policies in Simulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Evaluating Real-World Robot Manipulation Policies in Simulation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.371997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.371997Z digest=sha256:0994085a0a30d4757825cc3990af3f24dc5c8e4c6ba7615f76bd0fd90c951a7c

Observation 5b344481-fde6-48e3-b820-939cd989bafd · outbound

This paper cites DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.375454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.375454Z digest=sha256:5c2e80ca9223ed7b32221f8f4b55950bea7d276ba9474cb4219387aae9b9b280

Observation 686c96fe-47bf-4e27-b40e-8108637e46a7 · outbound

This paper cites International conference on machine learning , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models International conference on machine learning , pages=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.379418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.379418Z digest=sha256:7a58031f1868a8cd88d33c7ec4b4e03bdd49e717787b8b7d6140ef2aabe5d44e

Observation b6ddcaa8-62ba-48f8-aa9e-2e228d718057 · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.392414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.392414Z digest=sha256:aa67873d90c5b1676f3d954c1b4cea8b37cc625daadf1ab504bf92ab4502270f

Observation 60ca2448-e9bd-4003-ae0f-6ce16ba6df92 · outbound

This paper cites , author=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models , author=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.473885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.473885Z digest=sha256:2a20f7c16d0bc3c33ddb578812e565de83aaf3858d1029b4228263e65cbc2df9

Observation f1c1451d-d63e-46f4-b1f5-b44bd5a73354 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Octo: An Open-Source Generalist Robot Policy

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.537822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.537822Z digest=sha256:6fd6eb58cd8072c61238f58ddd7fb21397b274a60c08d1963d0042e812659cc4

Observation 0a34adf9-9825-40be-a7b7-a21df0bfa7c2 · outbound

This paper cites MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.561942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.561942Z digest=sha256:9d2177c64df7562c0ea7eec1c86aa3b21de246b0121dd564239a85a4e6a18a98

Observation 4da556c2-42cf-425a-b1a7-af0af792fa5f · outbound

This paper cites 2024 IEEE International Conference on Robotics and Automation (ICRA) , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models 2024 IEEE International Conference on Robotics and Automation (ICRA) , pages=

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.566446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.566446Z digest=sha256:6dfcfd2799a8e3c8459effb713d219fc4dadb0b0b0efe41e15dab59d88571d3e

Observation f762a73a-42df-4efb-b706-eb386f54e39d · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.311095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:39.570305Z digest=sha256:58f52d5763b8d29926b58310940b9bd4ba63842b425dae4369a640e4afc212d8

Observation abbfdb0f-96c0-4d4e-9908-ca777fba6274 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.573805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.573805Z digest=sha256:539e1664894a6fefff34abda49e3ba1e81ce9daa127b0e4fc283e3f9f2047033

Observation 3de06352-e39b-44b0-92c2-f68f8bc7ddab · outbound

This paper cites Advances in neural information processing systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in neural information processing systems , volume=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.578402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.578402Z digest=sha256:29b5847fe1f696eccd959516e663365a7d86709df97c1ba60d05ecf30b0cc8ef

Observation b334f3a3-4c84-4115-bc12-0c502d664b1e · outbound

This paper cites European conference on computer vision , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models European conference on computer vision , pages=

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.634116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.634116Z digest=sha256:61c04599fb60ef68efc286a9023417a97d68c0315f8b6e92daa161233ec3c945

Observation 9f3f74d0-5ab8-4336-87b5-c1b280691917 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in Neural Information Processing Systems , volume=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.728583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.728583Z digest=sha256:3622963cff877fec8a7beee0ff2e6217dbc2e432434a2e1675abfa267fa9342d

Observation bcad87bf-fb90-4a15-a5fc-3386623c3a3e · outbound

This paper cites 9th Annual Conference on Robot Learning , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models 9th Annual Conference on Robot Learning , year=

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.280516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:39.732138Z digest=sha256:721f6c0eeb27af89ca2575cb8334d57ec028e324ff6706afb67bfc2dd78e466c

Observation 9e5e6f1f-6f55-43a1-aec8-30ede2ab26a8 · outbound

This paper cites Depth Anything 3: Recovering the Visual Space from Any Views.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Depth Anything 3: Recovering the Visual Space from Any Views

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.735865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.735865Z digest=sha256:93a4b0bb825f4341ef7b1d1913b94eafa48a4b416dc4c3a850b530ae42e90887

Observation 0c33db0d-3f55-4324-83ce-01cc86ad7d2f · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.739135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.739135Z digest=sha256:5d6ad2149e5106ee1f8261590823bda460d3eefc34e07f0e1cc534a675c58fa6

Observation 10296197-c141-4541-a43b-bd0dfb0f63f6 · outbound

This paper cites IEEE Robotics and Automation Letters , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models IEEE Robotics and Automation Letters , volume=

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.270224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:39.744353Z digest=sha256:d08f8fb9353f106e2d9c015331d7c03efad14b7617488a670c63373fc2e1247f

Observation b3c62d0e-f7be-4c67-9049-a55541de67a0 · outbound

This paper cites arXiv preprint arXiv:2603.12942 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2603.12942 , year=

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.748139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.748139Z digest=sha256:2a16c781c48a3997d757f8cece1494770b470b31402c692f552e9b887292814e

Observation 217fca06-6500-4992-a60c-3739e205c05c · outbound

This paper cites arXiv preprint arXiv:2511.09516 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2511.09516 , year=

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.751081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.751081Z digest=sha256:991a5e73dd0eaa5abff8eb6e058a83824632dff39252a98b424169331c5d42e7

Observation e3608715-e71f-4952-9ffe-bc43eae14b7b · outbound

This paper cites Spatial Memory for Out-of-Vision Manipulation in Vision-Language-Action.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Spatial Memory for Out-of-Vision Manipulation in Vision-Language-Action

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.754892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.754892Z digest=sha256:5c35b7d7d7d481cb940c03005b14f6e58cacf05cf8adce28ff9f06105b05245a

Observation e6a6648e-c7a3-4fc3-ba67-1a971de1217d · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.871201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.871201Z digest=sha256:b976b2d76ce84ecc8cef877b4ccb0318bba73c3d9fbd9507fe630d0eed78df45

Observation da6697d6-8d2a-4f8a-acd7-b65ac646bc6c · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.959379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.959379Z digest=sha256:3d031a40a8faa9a1d6c1cd9bb2e379ac5dffbca1582b18467e0fd12bc56f7a49

Observation be4a9eeb-d7cb-430f-b815-adb4031cd6e6 · outbound

This paper cites DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.964543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.964543Z digest=sha256:f04b5d2068f919fb6c9509b6ed23f3ad165f2d258b34f0c6ea49593f344b1193

Observation b8fa5fcd-f5e0-416d-8c1c-162408133c7e · outbound

This paper cites The International Journal of Robotics Research , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models The International Journal of Robotics Research , volume=

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.968623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.968623Z digest=sha256:f55d8a76e3fb10fde636c6c6c0fc61f5ccd14c94512186e45e97447cc586ce0c

Observation 50c141b7-f62e-4f38-8685-38709cd06859 · outbound

This paper cites International Conference on Learning Representations , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models International Conference on Learning Representations , volume=

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.972759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.972759Z digest=sha256:84646e864eedb54371d51cf10232e87d357af16e14f600257799a4b76557657f

Observation 70e15055-bb60-4e35-8792-896d4f1fd471 · outbound

This paper cites arXiv preprint arXiv:2601.17885 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2601.17885 , year=

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.976530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.976530Z digest=sha256:b72894b78b2524d774bbc34959befdbe1a44b7e9540567d528213c23ed6d8dae

Observation cd0a5095-a81e-4042-b2cc-7c11d0bf3421 · outbound

This paper cites arXiv preprint arXiv:2603.03596 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2603.03596 , year=

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.980491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.980491Z digest=sha256:8b0dab01d4e6a805efe33fee6a9c10a57f9fe511dfe488ed359cd66b4dbf0125

Observation 7356e200-f8de-419a-9f72-b57d7124bfe1 · outbound

This paper cites International Conference on Machine Learning , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models International Conference on Machine Learning , pages=

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.245491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:39.984092Z digest=sha256:9de6d71edf406ee4a48b4a764db321a1c90c2d9b96d01b6c75502a751f44f188

Observation 1a06449e-2bd5-4e54-b807-016c520003fb · outbound

This paper cites 2011 10th IEEE international symposium on mixed and augmented reality , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models 2011 10th IEEE international symposium on mixed and augmented reality , pages=

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.998452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.998452Z digest=sha256:607fe83da4ca6171e366577b722a02134148651660af6ed217b060cea24c9dbe

Observation 282bb211-58b9-49b5-a321-fc7623c404df · outbound

This paper cites Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.052053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.052053Z digest=sha256:d76ee8b2bd0af38c863031ce42f82d1b5d66779c90792ebe4370f6346fe1fdfc

Observation cbcab0b7-be21-459d-a1dd-d7d2f6d4e54c · outbound

This paper cites RynnVLA-002: A Unified Vision-Language-Action and World Model.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RynnVLA-002: A Unified Vision-Language-Action and World Model

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.086994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.086994Z digest=sha256:92a7caa50e088188e24d429b551f037091af6d2ccb857fbbda7e516f0537e533

Observation ba421f0d-dcf6-4f75-bc15-eb4da3008fa7 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in Neural Information Processing Systems , volume=

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.226904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:40.091060Z digest=sha256:c520420946732699abdaec84520f84f8dcbd48dbc553b9653a32afc4622df81a

Observation d3b3487e-4528-4459-846e-406b4becffbb · outbound

This paper cites International Conference on Learning Representations , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models International Conference on Learning Representations , year=

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.094674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.094674Z digest=sha256:ccb7451f961d44332cc9b99a683eac05a78163d17bb7751f251cf08a9db4dded

Observation c44050e2-1ab4-4ddd-b3f5-5cad689ec8ce · outbound

This paper cites Classifier-Free Diffusion Guidance.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Classifier-Free Diffusion Guidance

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.098346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.098346Z digest=sha256:9b4615542506240d84bd09e89e60e0c33027217d7a1a5c62ba67051290cd766e

Observation c4c54e4c-464a-488a-93a5-16ccbfcb7628 · outbound

This paper cites IEEE Robotics and Automation Letters , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models IEEE Robotics and Automation Letters , volume=

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.102244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.102244Z digest=sha256:63e1709c5ab4f10175f5df64c8d4159327dfcbc00de95ea7e63320fda7d32041

Observation 623d2989-3534-45f8-b3ce-72f24b64da33 · outbound

This paper cites Mask World Model: Predicting What Matters for Robust Robot Policy Learning.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Mask World Model: Predicting What Matters for Robust Robot Policy Learning

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.105970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.105970Z digest=sha256:e6068e9491529c24d656e5dae3063edb7e621fc32f985a0d71b30976628f3fc1

Observation c8fe6e67-3b19-4e38-9dad-4a20fba9c767 · outbound

This paper cites Transactions on Machine Learning Research Journal , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Transactions on Machine Learning Research Journal , year=

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.109759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.109759Z digest=sha256:14bd7b5f73ba5ff1570d7ae88d8e04e11d03455476211f4f5aff93fc9924b7bf

Observation 6c3c7702-35ce-4323-8246-6f111d3673ff · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.113681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.113681Z digest=sha256:693edf271c7a8b092e8777278f058712ff2e39d066663277679d993c4e75b7a1

Observation dd029b01-e1d7-4015-8164-7d463853202e · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.117731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.117731Z digest=sha256:6a45457c79fc92e3d01137d16ca1ea66b2c5986b7fa73491a341bfbc747c14f2

Pith citing papers

No inbound Pith citation observations are available.