Pith. sign in

Paper Citation Record · LEDGER

PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2406.20083.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.20083 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:40:50.074329Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:59:58.789222Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8d98a95b-8c72-4354-9805-96035f458c02 · inbound

Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks cites this paper.

Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:51:36.314657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-16T19:51:36.137985Z digest=sha256:218cd10a0b3c53957fabfc327cb6281d8159c8e1a3b882ac35bcdcd6d4d25ec9

Observation b52d2d76-c472-4fdf-9dbe-4fc2d50b8800 · inbound

Humans Coexist, So Must Embodied Artificial Agents cites this paper.

Humans Coexist, So Must Embodied Artificial Agents PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 156

Resolution
unresolved
no resolver link, observed 2026-08-08T21:27:27.959757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:27:27.959757Z digest=sha256:abc81fa9f2c549aedb6bf848393905421bab48422f5b5e8022a16095642e53ca

Observation 96661c6e-ae07-4d10-b069-242e8ec1ad3b · inbound

SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning cites this paper.

SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-23T01:32:22.394182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-23T01:27:33.123243Z digest=sha256:2784dafe24891aeb3be3e34444a8d0f6eb83eb814b5772fd31af0f99b8eeecda

Observation 45dad172-cba8-4303-8fc8-e6aa0fbe6606 · inbound

CaRL: Learning Scalable Planning Policies with Simple Rewards cites this paper.

CaRL: Learning Scalable Planning Policies with Simple Rewards PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T10:40:50.074329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:40:50.074329Z digest=sha256:a1665a1094b6e08952c69fd32a5ee24d6d097e648db95666a5d1038c8b6bca5b

Observation 62c9138d-c652-4003-809b-0b608d034b4b · inbound

GraspMolmo: Generalizable Task-Oriented Grasping via Large-Scale Synthetic Data Generation cites this paper.

GraspMolmo: Generalizable Task-Oriented Grasping via Large-Scale Synthetic Data Generation PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T20:18:08.177749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:18:08.177749Z digest=sha256:2c334b1a4c5c6b4e3710383b5db78b1bb204c15fc0218576468e8875d65fdc3f

Observation 4db77e89-7dfd-4933-9eca-fcb3fc8ac11c · inbound

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning cites this paper.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.383944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:9ba8949d2a4718511fafaa1d00b8a5889fa870672e2267fea8316c4dd52ab9f1

Observation 973b0c07-589c-45e7-b8c3-0ebac85911ba · inbound

TrackVLA: Embodied Visual Tracking in the Wild cites this paper.

TrackVLA: Embodied Visual Tracking in the Wild PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:58:37.641043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:58:37.641043Z digest=sha256:e18e7e5d0635092ccffd5507e602a00396588578def23295f7526d7bf7ddbbe8

Observation 9094c532-8c36-46a6-90f9-d5ec889bbc71 · inbound

Spatially-Enhanced Recurrent Memory for Long-Range Mapless Navigation via End-to-End Reinforcement Learning cites this paper.

Spatially-Enhanced Recurrent Memory for Long-Range Mapless Navigation via End-to-End Reinforcement Learning PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:18.949638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:08:18.949638Z digest=sha256:08ec006849c48e58144086acac40ca80ac15569aae7ffaeeedda925c5008dc56

Observation 2609d05e-91ce-4d80-b285-5d697180c3b5 · inbound

Move to Understand a 3D Scene: Bridging Visual Grounding and Exploration for Efficient and Versatile Embodied Navigation cites this paper.

Move to Understand a 3D Scene: Bridging Visual Grounding and Exploration for Efficient and Versatile Embodied Navigation PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T20:02:33.151309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:02:33.151309Z digest=sha256:b8c19a4d26fd97d7ecfbf2ad48c0519ddc420ac0a240facfdb4c6e284f1d87e3

Observation 7f994f2f-378c-42cd-9787-ed92b06e433b · inbound

What Matters in RL-Based Methods for Object-Goal Navigation? An Empirical Study and A Unified Framework cites this paper.

What Matters in RL-Based Methods for Object-Goal Navigation? An Empirical Study and A Unified Framework PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T12:55:45.971617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:55:45.971617Z digest=sha256:370bac3a8c8217fb8a5f2561693eaaae192e8318a0c4678dcf3f0ded480a4c94

Observation 64939207-7595-4452-9653-adf359a9355e · inbound

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning cites this paper.

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T12:37:55.577906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:37:55.577906Z digest=sha256:78cb752e9a4df1b7507cd8dd34e1d6be53e0a3ea345ca7544e595740746693cf

Observation d5898c4e-c6b5-487c-ba29-b3f7e789d1c6 · inbound

Learning Category-level Last-meter Navigation from RGB Demonstrations of a Single-instance cites this paper.

Learning Category-level Last-meter Navigation from RGB Demonstrations of a Single-instance PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T16:59:18.522659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:59:18.522659Z digest=sha256:a44b7c230870c9cd499cfe2be5e66b0b17b880b9b88355aacf158d4bbf3110a8

Observation 98733f2a-028e-490d-ae83-7f4eb624c6b2 · inbound

LangMap: A Human-Verified Benchmark for Hierarchical Open-Vocabulary Goal Navigation cites this paper.

LangMap: A Human-Verified Benchmark for Hierarchical Open-Vocabulary Goal Navigation PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T05:29:00.199388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:29:00.199388Z digest=sha256:489ca4b4fc677e98e8048110c5b2b4e34c30f62ca6740e2692c29b2a0444bc6c

Observation f79e5685-a1ac-4962-95bd-42ebfe7f76da · inbound

Beyond Isolation: A Unified Benchmark for General-Purpose Navigation cites this paper.

Beyond Isolation: A Unified Benchmark for General-Purpose Navigation PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:06:27.540703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-12T04:33:53.557357Z digest=sha256:17da7a0a3e5fd18315342fd7c87972703575e99c6db15bf5dcba905b6d4d5401

Observation c703fc46-7629-4c76-9435-e048e2eeb295 · inbound

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation cites this paper.

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:11:27.557573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-12T03:36:24.941205Z digest=sha256:e93faf33c7e381934cfb2a6c1c8dedbdd284022a65d2a0007111291f8a91f40f

Observation 865eeb90-a212-487d-ab73-f762f7c1eac5 · inbound

NavOL: Navigation Policy with Online Imitation Learning cites this paper.

NavOL: Navigation Policy with Online Imitation Learning PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:52:21.897249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T05:51:53.848800Z digest=sha256:3ffa5a1eb0f2af8ed17eacdf7f7f1f80765ef32a32aae9b476858a8b25f75b85

Observation 211ea54a-848a-48b2-84ca-bb888e50d6b9 · inbound

Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration? cites this paper.

Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration? PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-06-28T17:12:24.810484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-06-28T17:08:59.111306Z digest=sha256:5aa7cb867bc055c0574d8e4996eb23bfd7e0b13c7ba81873a7740dbe8723bc7d

Observation 45b4e964-0638-44a8-acc2-a9771bf3da60 · inbound

AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning cites this paper.

AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:07:38.963945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-27T13:25:59.194721Z digest=sha256:e8fa3b794a427817600302e0981948cbfd1610b966af20b9253a78330ac6a5ee

Observation 3ad73e5a-3432-4eb9-8d15-9163b6ab93cc · inbound

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System cites this paper.

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:28:59.301191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-27T00:28:04.678371Z digest=sha256:988d7873f6275a82da3982eb64bffcf3222f63892b55fae9c21ea491f1839cc3

Observation 23419fa8-3bef-4fd7-82e1-ecb037d7a1bf · inbound

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System cites this paper.

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-30T10:54:36.766064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-30T10:47:20.911408Z digest=sha256:f58f652093ca114c31414945155f83ddb7832fbeda29983fe346c35a095711bd

Observation e394be1c-2789-4dd4-aecf-dc286a9b18a3 · inbound

SurveilNav: Collaborative Object Goal Navigation with Robot and Surveillance System cites this paper.

SurveilNav: Collaborative Object Goal Navigation with Robot and Surveillance System PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:59:58.791632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-26T00:01:02.654432Z digest=sha256:a0dcb7ff91b5fa2a57ba826e4c984bee12a8215da4ec631a379016f3ef7d8fe8

Observation e5e82b27-d246-4077-9a42-36fd39efb6be · inbound

ABot-N1: Toward a General Visual Language Navigation Foundation Model cites this paper.

ABot-N1: Toward a General Visual Language Navigation Foundation Model PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 76

Resolution
unresolved
no resolver link, observed 2026-07-14T12:10:21.115628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:10:21.115628Z digest=sha256:c4f3e140f824f9f307526f44e014cfd966783d5e664313dc7301362541f948a1

Observation f45fc376-bbdd-4e1b-bf69-de16055f0d5f · inbound

ABot-N1: Toward a General Visual Language Navigation Foundation Model cites this paper.

ABot-N1: Toward a General Visual Language Navigation Foundation Model PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-02T07:19:45.435794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:19:45.435794Z digest=sha256:91b53fc5b8e39a3bad1028e1dc1ab311d2f2c308f8b564c1d7fc4439375f311a