Pith. sign in

Paper Citation Record · LEDGER

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation

As of 10 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 0 inbound Pith citation observations for arXiv:2607.14586.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.14586 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T01:41:52.473153Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 18a3fdb3-4f74-4a0e-acdd-06fce3c33428 · outbound

This paper cites HM3D- OVON: A dataset and benchmark for open-vocabulary object goal navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation HM3D- OVON: A dataset and benchmark for open-vocabulary object goal navigation,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:48.500807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:48.500807Z digest=sha256:557a248adb0354619db2f4d55f9c0918ea0442e8e2eec39b3b14139fbf859fa0

Observation f9b56fb0-b2b5-417e-b0a8-d256ce72b834 · outbound

This paper cites GOAT-Bench: A benchmark for multi-modal lifelong navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation GOAT-Bench: A benchmark for multi-modal lifelong navigation,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:48.673352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:48.673352Z digest=sha256:f468def374eabb76e7b4d759970f56c4d4d35aba182cf86572da5def372e3a54

Observation eadc8c0c-5395-45cd-9797-00c5b107faed · outbound

This paper cites Task-oriented Sequential Grounding and Navigation in 3D Scenes.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Task-oriented Sequential Grounding and Navigation in 3D Scenes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:48.849680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:48.849680Z digest=sha256:27b90a99912f226833e6dc44e0ceb0d7f7a33bce23e00ebbb91b85414d74f896

Observation e0f17fa7-441f-42a7-a952-368e7d833415 · outbound

This paper cites Object goal navigation using goal-oriented semantic exploration,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Object goal navigation using goal-oriented semantic exploration,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:49.022785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:49.022785Z digest=sha256:3086c93a9bb3067a9e08b20cd08d35df1ec3909b24aeb42a2076b59baa991824

Observation 4c796ac1-d23e-45b0-b045-ec58a7283cb8 · outbound

This paper cites VLFM: Vision-language frontier maps for zero- shot semantic navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation VLFM: Vision-language frontier maps for zero- shot semantic navigation,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:49.144127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:49.144127Z digest=sha256:8422e9c2ee847e6a7beb8f1222a3ee7ddd3aef0ef6c6bdd424cbbe0524edb79e

Observation 8a067809-8a12-46ae-aa37-5bb25bb21965 · outbound

This paper cites Move to understand a 3D scene: Bridging visual ground- ing and exploration for efficient and versatile embodied navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Move to understand a 3D scene: Bridging visual ground- ing and exploration for efficient and versatile embodied navigation,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:49.291227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:49.291227Z digest=sha256:e955485f6774a442eca91ce46523232390f00498cdd066b83834bd37374a7efe

Observation 50a43686-fe47-4752-9f7b-62aac0e018f3 · outbound

This paper cites Unifying 3D vision-language understanding via prompt- able queries,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Unifying 3D vision-language understanding via prompt- able queries,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:49.436063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:49.436063Z digest=sha256:4a493422cbb2b968a8c5a90762b5a33377fdaf2549d83987f36ba839e3da6b72

Observation 3d527683-1a5e-4d70-b51b-fc4b5c3fe068 · outbound

This paper cites NavGPT: Explicit reasoning in vision- and-language navigation with large language models,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation NavGPT: Explicit reasoning in vision- and-language navigation with large language models,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:49.573559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:49.573559Z digest=sha256:aabf55e7b88ae01af562a26128d5e6a21c01b81abe576abf6429e530ce075019

Observation 5aa67a84-1514-45aa-b566-4eae579fd2a3 · outbound

This paper cites NaVid: Video-based VLM plans the next step for vision-and-language navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation NaVid: Video-based VLM plans the next step for vision-and-language navigation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:49.701297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:49.701297Z digest=sha256:7ba176868cf351cac274b8edb215f4d04fc9ce6fd36823a0999149fdda988267

Observation dce93dd5-08cb-4fd7-8c8e-71b966f293d7 · outbound

This paper cites L3MVN: Leveraging large language models for visual target navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation L3MVN: Leveraging large language models for visual target navigation,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:49.873639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:49.873639Z digest=sha256:f46d96f55136ac2d5a50cdb3ca8e6a9673110e13df55b3eb66941b62f2557d6e

Observation 69f1c1d8-2f6f-4429-a47c-f4ca7c899a6a · outbound

This paper cites SayNav: Grounding large language models for dynamic planning to navigation in new environments,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation SayNav: Grounding large language models for dynamic planning to navigation in new environments,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.039088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.039088Z digest=sha256:99a95d6d3d606d310074382736fd848d2d5787107ea93b28effcbda7941d4cd8

Observation 892e9a91-5e9e-48ad-8e2b-d1d3b5a24d47 · outbound

This paper cites SG-Nav: Online 3D scene graph prompting for LLM-based zero-shot object navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation SG-Nav: Online 3D scene graph prompting for LLM-based zero-shot object navigation,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.157628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.157628Z digest=sha256:347a75a8ca467b473054dab59c99ea6bbdfe2f377e9c265204543ed5eee35558

Observation 79a23742-38c4-4ea3-91ac-b217f9128ac8 · outbound

This paper cites A frontier-based approach for autonomous exploration,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation A frontier-based approach for autonomous exploration,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.256336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.256336Z digest=sha256:628dc14956a748c844a44b85bfbd6ac9f692e22e75b03c35927f9829c59953a4

Observation fc107c30-578f-49fe-ac81-c41f2db81e81 · outbound

This paper cites PONI: Potential functions for ObjectGoal navigation with interaction-free learning,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation PONI: Potential functions for ObjectGoal navigation with interaction-free learning,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.349270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.349270Z digest=sha256:ee83570c9328bbbab72acbfe27d862c2337fdb4b6e20d68538969712b1214a38

Observation da2cbd47-b561-478a-b215-033197827d8c · outbound

This paper cites PIRLNav: Pre- training with imitation and RL finetuning for ObjectNav,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation PIRLNav: Pre- training with imitation and RL finetuning for ObjectNav,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.444445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.444445Z digest=sha256:2abc08d718a663e0f4412f683c0e1d22a30181c58e2ba8fd41d23ebd6ffb8315

Observation 09e1a243-b281-4ec8-b3e7-5648140213bb · outbound

This paper cites Uni-NaVid: A video-based vision-language-action model for unifying embodied navigation tasks,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Uni-NaVid: A video-based vision-language-action model for unifying embodied navigation tasks,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.542590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.542590Z digest=sha256:5b1da6770aca41727b8ab5789b874c32f71d2ff1ab4a481033c9983718d19e15

Observation 6a3d1d78-6de2-4ea9-ab8d-aa01ba6f415d · outbound

This paper cites V oroNav: V oronoi-based zero-shot object navigation with large language model,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation V oroNav: V oronoi-based zero-shot object navigation with large language model,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.628115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.628115Z digest=sha256:0446c7f466cf25caff3a0c4f67d7d4e67c96431afe4e19052f35327c7db38cb5

Observation 802f9793-758c-4bf7-977d-bc16ec2717fb · outbound

This paper cites MapGPT: Map-guided prompting with adaptive path planning for vision-and-language navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation MapGPT: Map-guided prompting with adaptive path planning for vision-and-language navigation,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.724452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.724452Z digest=sha256:95e54d1ac402b200d61341d1a414870439d52e84de0e8603f19b95a2d269ab59

Observation 498b2f0e-96c6-4f64-a22a-dc8358fbb9a0 · outbound

This paper cites ESC: Exploration with soft commonsense constraints for zero-shot object navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation ESC: Exploration with soft commonsense constraints for zero-shot object navigation,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.790386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.790386Z digest=sha256:ef3662828dd1d04b84740c5f012f65f9f84b14c1a896a7c85137cdf7efdd0fc7

Observation 24574323-eae1-4655-80a9-1195d1452786 · outbound

This paper cites Flamingo: A visual language model for few-shot learning,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Flamingo: A visual language model for few-shot learning,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.867928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.867928Z digest=sha256:ef82995c27e26eacd14ab45f48093e4bcbc2cc75659176620287e9bdfcbf5b2e

Observation 9e2fba9b-0600-4a47-887f-d57a4a4582dd · outbound

This paper cites Visual instruction tuning,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Visual instruction tuning,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:50.920017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:50.920017Z digest=sha256:17d7ae8e6df01a08d0df7a567cbc782b063898f87c7c296c4ca79973fdf56725

Observation 6c8710cb-484d-441d-8c36-40d5c00b95ea · outbound

This paper cites BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.006655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.006655Z digest=sha256:8c571d614f6aeac4dd479843d59ee4b7ffb6a8e2848ea8c86fb663f7ec430821

Observation b29a0332-5521-4879-b866-2c8ecbf728ad · outbound

This paper cites 3D-LLM: Injecting the 3D world into large language models,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation 3D-LLM: Injecting the 3D world into large language models,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.065132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.065132Z digest=sha256:37217ce63a67015a33ede4e8d8cc2437992573d9bb86ae63b1f7def005f5b7c5

Observation 08cbdb82-3867-4f6f-b8e6-85aead536ee5 · outbound

This paper cites An embodied generalist agent in 3D world,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation An embodied generalist agent in 3D world,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.157750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.157750Z digest=sha256:19a28373158439c26921633bc4c37bdb2e7588241ebb3520cd795db84aa30ba1

Observation e34d34f8-d4b2-4fb4-af0b-4e5b6366163d · outbound

This paper cites LL3DA: Visual interactive instruction tuning for omni- 3D understanding,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation LL3DA: Visual interactive instruction tuning for omni- 3D understanding,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.255377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.255377Z digest=sha256:64fb0510fea79b68316f766b046c6fe82e1bb320d3c0402d80e5c3588c34b4b4

Observation 9487fa09-aa21-4e69-8356-ece45640f4d8 · outbound

This paper cites EmbodiedGPT: Vision-language pre-training via em- bodied chain of thought,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation EmbodiedGPT: Vision-language pre-training via em- bodied chain of thought,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.351345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.351345Z digest=sha256:697c8ddb8b394a1c15ffde296f5c9bc3c1bb650d999ff5c402d0ac6b423ae908

Observation 384b949d-474e-4ee8-aa98-4e388da0437a · outbound

This paper cites Dynam3D: Dynamic layered 3D tokens empower VLM for vision-and-language navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Dynam3D: Dynamic layered 3D tokens empower VLM for vision-and-language navigation,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.444318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.444318Z digest=sha256:c857738671370d354cc55ac3890b8eb90a854701edcb0775cd3a83cad0cc9134

Observation ac9517b3-6973-40af-8537-4165a3d72d82 · outbound

This paper cites AstraNav-Memory: Contexts compression for long memory,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation AstraNav-Memory: Contexts compression for long memory,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.535275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.535275Z digest=sha256:1b005fbee1c239f315fea44609abc673ecb52892de308c11570a8d1ad8f9108c

Observation 396570cb-60cd-49dc-b30b-6293bab5e135 · outbound

This paper cites DINOv3.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation DINOv3

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.639055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.639055Z digest=sha256:a26c204241278fa9dad642e4e89adee4f1fc7274a4dd9efd3576454ae5af123c

Observation 091fa90e-0c94-4c79-a71b-30384284b95c · outbound

This paper cites Qwen2.5-VL Technical Report.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Qwen2.5-VL Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.731349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.731349Z digest=sha256:797f40546220f941f4f06517763250e89d49e35a1b12c0ed21e77bfcecb36d03

Observation eee35148-b0bf-4a26-ba8a-f59e8a7f2822 · outbound

This paper cites LoRA: Low-rank adaptation of large language models,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation LoRA: Low-rank adaptation of large language models,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.859604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.859604Z digest=sha256:fcbd098a8bf19eabe73a6b87df5bed26b793a27dc8aee6bd123c08578bfe2486

Observation 8d1a3746-65ce-4bce-a4be-876c4a5f6a2d · outbound

This paper cites Habitat: A platform for embodied AI research,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Habitat: A platform for embodied AI research,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:51.952688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:51.952688Z digest=sha256:2dca73072ec15ac775eed002f6e6da9e9e53fe23de6323af17662d4da15c391c

Observation cc49ba36-6586-4ca4-96c2-dfda73da8a48 · outbound

This paper cites The design of stretch: A compact, lightweight mobile manipulator for indoor human environments,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation The design of stretch: A compact, lightweight mobile manipulator for indoor human environments,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:52.050418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:52.050418Z digest=sha256:f535643a334e195e481bf9034dacacfd2b0633f742c47567105af8dc355a9f33

Observation 7e48a6d5-2dde-493f-9901-72f9c75493a3 · outbound

This paper cites TANGO: Training- free embodied AI agents for open-world tasks,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation TANGO: Training- free embodied AI agents for open-world tasks,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:52.146704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:52.146704Z digest=sha256:51d765273286294f8eac162a015f48b496cf864eac13dcb4553e3d8d136f9594

Observation 6e30485e-cbee-4082-a043-3c7a8f232d83 · outbound

This paper cites MSGNav: Unleashing the power of multi-modal 3D scene graph for zero-shot embodied navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation MSGNav: Unleashing the power of multi-modal 3D scene graph for zero-shot embodied navigation,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:52.243259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:52.243259Z digest=sha256:28221d27d4a5f826e4ce248c548885a9013644f98d680a0d4372dcd03908fe13

Observation 3f91037b-3f67-4718-8b22-302b24b279ee · outbound

This paper cites Cows on pasture: Baselines and benchmarks for language-driven zero-shot object navigation,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Cows on pasture: Baselines and benchmarks for language-driven zero-shot object navigation,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:52.342874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:52.342874Z digest=sha256:cfa93a66b4a10e11c7cea8ba47f6139e0184b5424963d98238152ca25e536c0e

Observation 24499448-6823-4f3b-a2b8-66be4601a539 · outbound

This paper cites Embodied VideoAgent: Persistent memory from ego- centric videos and embodied sensors enables dynamic scene under- standing,.

SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation Embodied VideoAgent: Persistent memory from ego- centric videos and embodied sensors enables dynamic scene under- standing,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T01:41:52.473153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:41:52.473153Z digest=sha256:510bc102ea013b5f50d3e21de6e294c42d331f8b495ae7ce20d4670661eacac1

Pith citing papers

No inbound Pith citation observations are available.