Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T22:38:02.945576Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2603.18444.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T22:38:02.945576Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ee13414a-61dc-4dc7-8021-5d2f7949d656 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards ObjectNav Revisited: On Evaluation of Embodied Agents Navigating to Objects
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bab5cdb9-12ab-4224-913b-44d72fe7cb78 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Bridging zero-shot object navigation and foundation models through pixel-guided navigation skill,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b88404a6-763b-4eed-8e85-21e4a1cc880e · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards V oila: Visual- observation-only imitation learning for autonomous navigation,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58a525dd-6cc7-4ab9-a1b1-e3980464365a · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Zson: Zero-shot object-goal navigation using multimodal goal embed- dings,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 714bca8c-96b0-49a3-a054-9c28dc97cfad · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Zero-shot active visual search (zavis): Intelligent object search for robotic assistants,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d31a1de-b4a2-47d7-9873-0616e56f2d04 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Zero-shot object goal visual navigation,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c8d42a3-5f05-4a76-bfee-4d1884f8e15f · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Hierarchies of planning and reinforcement learning for robot navigation,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab6858b9-e292-4437-a020-e20ccafd8cc0 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Esc: Exploration with soft commonsense constraints for zero- shot object navigation,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d577743e-939c-408b-9d4f-db2b5c29def5 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Sg-nav: Online 3d scene graph prompting for llm-based zero-shot object navigation,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdf9663e-e929-4c8d-849f-e69b0e9f0c6a · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Vlfm: Vision- language frontier maps for zero-shot semantic navigation,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c53bf7a-33c8-4c47-b896-a78f15fd179d · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards V oronav: voronoi-based zero-shot object navigation with large language model,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35611784-b131-465b-9273-15d5173dfbb2 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards TopV-Nav: Unlocking the Top-View Spatial Reasoning Potential of MLLM for Zero-shot Object Navigation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9b4a430-86be-40b7-83e2-0d1fa2aa0616 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Cows on pasture: Baselines and benchmarks for language-driven zero-shot object navigation,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55c31205-0279-4a85-98aa-d00e603c4030 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Star-searcher: A complete and efficient aerial system for autonomous target search in complex unknown environments,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a017583-2025-415a-a5cb-6a88614ab482 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Occupancy Grids: A Stochastic Spatial Representation for Active Robot Perception
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 422750fd-de7f-483e-8577-edb12fd81549 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards A frontier-based approach for autonomous exploration,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a29c07d-e87f-4229-ba10-40fa2b9788c6 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Information gain-based exploration using rao-blackwellized particle filters,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f96670a-d8de-4044-9e9d-eea88f0a39c0 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Gamap: Zero-shot object goal navigation with multi-scale geometric-affordance guidance,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a193f278-1470-4c96-bb39-37a1850a4e56 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards L3mvn: Leveraging large language models for visual target navigation,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2a3ee19-2892-410c-8e51-7d8a77e153c0 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22462d54-9508-4ef8-a93f-6e0834f090bb · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Qwen2.5-VL Technical Report
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed768df3-022a-48c8-adae-1b0ba64dc7a9 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards GPT-4 Technical Report
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9507e9ee-f4c0-4ff2-9de5-f963df4d93df · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards OpenFMNav: Towards open-set zero-shot object navigation via vision-language foundation models,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74175aed-4548-43bc-ae8e-b3906f67fe21 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Egtr: Extracting graph from transformer for scene graph generation,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5389475a-3a47-4205-a3db-5ae648be79e9 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Hi- erarchical Open-V ocabulary 3D Scene Graphs for Language-Grounded Robot Navigation,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fd13e79-6973-4696-92ce-2953a7facb63 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Expanding scene graph boundaries: Fully open-vocabulary scene graph generation via visual- concept alignment and retention,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4e1b900-4d03-457e-92b4-c0aadcf8c549 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Cognav: Cognitive process modeling for object goal navigation with llms,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e50425c1-1a04-4463-9ac3-e78fabd282b4 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Beliefmapnav: 3d voxel- based belief map for zero-shot object navigation,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfbd27a7-057d-4c53-a748-9e96f56effc5 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Object goal navigation using goal-oriented semantic exploration,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d929c26c-16d1-43c2-a77f-57a031226a53 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Habitat-Matterport 3D Dataset (HM3D): 1000 Large-scale 3D Environments for Embodied AI
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f9942e5-bbd3-4c18-8de3-a0ef9ac544b4 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Matterport3d: Learning from rgb-d data in indoor environments,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d0cc65b-40b5-424c-90e2-fb4bfedd6878 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Trihelper: Zero-shot object navigation with dynamic assistance,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf0da835-7476-4d63-811f-5304dcff8de9 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards How to not train your dragon: Training-free embodied object goal navigation with semantic frontiers,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e5c6740-0425-4705-be05-4dd17b64711b · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Fuel: Fast uav exploration using incremental frontier structure and hierarchical planning,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8a622e0-f555-4d40-a1b3-684ba0591a86 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Think holistically, act down-to-earth: A semantic navigation strategy with continuous environmental representation and multi-step forward planning,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd5a7e54-6a4a-4643-84a6-91ee7833e11f · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Open scene graphs for open-world object- goal navigation,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61156cb8-76b5-4303-9e8e-532b688f8a68 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Agent-centric relation graph for object visual navigation,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation beb2cef6-86b4-4c61-a403-8ee98f9aae9b · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Imaginenav: Prompting vision- language models as embodied navigator through scene imagination,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3dfd92a-bb50-42e0-a110-bb69bb6477b8 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards BERT: Pre- training of deep bidirectional transformers for language understanding,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05eb2658-667e-4015-8179-7f85b963f23f · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2632f36-53a3-49c2-858f-6f2d1d164a4f · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Chatnav: Leveraging llm to zero-shot semantic reasoning in object navigation,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3649dc97-8331-421d-8d70-ddaa592d0f1b · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards E²ba: Environment exploration and backtracking agent for visual language object navigation,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0ba9d8e-8ec6-4fc3-86a9-149c4d3179a5 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Strive: Structured representation integrating vlm reasoning for efficient object navigation,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffaeb69b-ceef-4fb5-b252-f5004682a512 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Visual semantic navigation using scene priors,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48e4afe0-a828-4c03-bf06-5a2a94f15a2a · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Imagine before go: Self-supervised generative map for object goal navigation,
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec82943f-9036-4602-94bf-d260b783dda4 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Hierarchical object-to-zone graph for object navigation,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f547d67-5653-4d00-b199-d77e75053ac1 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Search for or navigate to? dual adaptive thinking for object navigation,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a57d928d-7ec9-4039-91c1-f435bcaacdc0 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Layout-based causal inference for object navigation,
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 985abf22-8817-4706-bebe-607ce89eaf30 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Poni: Potential functions for objectgoal navigation with interaction-free learning,
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fcb54d5-3554-4f55-bb2c-dfef01caf4aa · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Learning object relation graph and tentative policy for visual navigation,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b25dea6e-8860-4165-98be-2444f0c34dbc · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Yolov7: Trainable bag-of-freebies sets new state-of-the-art for real-time object detectors,
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 613d9442-cc37-4e34-ba49-1f29faefd448 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Grounding dino: Marrying dino with grounded pre-training for open-set object detection,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c30f039-8c2d-42e5-a8bf-61ad55d85adb · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Faster Segment Anything: Towards Lightweight SAM for Mobile Applications
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07179b9a-ed6c-4973-bb92-8507893ebaf6 · outbound
Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards Unigoal: Towards universal zero-shot goal-oriented navigation,
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.