Pith. sign in

Paper Citation Record · LEDGER

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization

As of 17 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 0 inbound Pith citation observations for arXiv:2507.10894.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10894 v1

Coverage vector

measured 77 of 77 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:26:43.437116Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

77 of 77 outbound references displayed

  • verified exact0
  • verified fuzzy62
  • unresolved14
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 067f3fef-cfd5-4b9b-914a-1a43d89c144b · outbound

This paper cites A survey of embodied ai: From simulators to research tasks,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization A survey of embodied ai: From simulators to research tasks,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.494005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:37.385842Z digest=sha256:024711612480a37abeb845e96bee43990e734f875f072947c8b4fa22a008f003

Observation a740e2b8-e30e-4bd1-8417-2c33ac260644 · outbound

This paper cites Behavior-1k: A benchmark for embodied ai with 1,000 everyday activities and realistic simulation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Behavior-1k: A benchmark for embodied ai with 1,000 everyday activities and realistic simulation,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.485004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:37.447109Z digest=sha256:26697f511e59b04d91dd41952db406f0814c01500b8134dccba3f82c33b5fb49

Observation c84f1d60-1240-4862-8b95-285523181373 · outbound

This paper cites Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:37.507750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:37.507750Z digest=sha256:6d35077a4b44005cd36fa3c37078a9c98b8ced532673c6e81a0f521b8b330f26

Observation d95ca566-acea-41bd-a67c-56d81f7d35c0 · outbound

This paper cites Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.476143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:37.597645Z digest=sha256:486ac160c0265bf75da11a2809354546b734cff684245839be686fbd2011212d

Observation 08615390-332c-426b-a2cc-25b031ee047f · outbound

This paper cites Room-across-room: Multilingual vision-and-language navigation with dense spatiotemporal grounding,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Room-across-room: Multilingual vision-and-language navigation with dense spatiotemporal grounding,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.466733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:37.642304Z digest=sha256:8e856990750aa8f79056159fc066740b563f7a89ad01a93bf487e7462a553c32

Observation 320f098a-a661-45fd-9ce5-b5d5716e3908 · outbound

This paper cites Talk2nav: Long-range vision-and-language navigation with dual attention and spatial memory,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Talk2nav: Long-range vision-and-language navigation with dual attention and spatial memory,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.457258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:37.732340Z digest=sha256:1765fc59993f2832760af72ce7d31b996492b625a675c98d53d87ef97516306b

Observation 3fd76607-44a9-47e8-b314-11137f882b18 · outbound

This paper cites Reverie: Remote embodied visual referring expression in real indoor environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Reverie: Remote embodied visual referring expression in real indoor environments,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.447485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:37.821615Z digest=sha256:9b798a74fbd25f5db0d249f0faef98beeb113d652e58bd35867e8a9680356e50

Observation 79b91aaf-2067-483d-9252-0e4951e56f13 · outbound

This paper cites Speaker-follower models for vision-and-language nav- igation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Speaker-follower models for vision-and-language nav- igation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.438579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:37.926692Z digest=sha256:f3b2c82d644dc037534c450550ca2379beefcee90135e97f8d6012d86c2e19d9

Observation e7b3cd13-b28a-4c39-ac02-363b48101623 · outbound

This paper cites Learning to navigate unseen environments: Back trans- lation with environmental dropout,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Learning to navigate unseen environments: Back trans- lation with environmental dropout,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.428820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.019873Z digest=sha256:c7e7214f80414b8bd23ea4354c626ab97b3b0168b3264c1e2ffe619bd188fe67

Observation 2f3b20a6-0499-4de2-ae3e-cef0c369a98f · outbound

This paper cites Visual landmark selection for generating grounded and interpretable navigation instructions,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Visual landmark selection for generating grounded and interpretable navigation instructions,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.420060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.083529Z digest=sha256:f40f28929455505702bb0eb6e012df85d924e5160d964cf2b6eaa8c41b6500de

Observation 86a5f0cd-6322-460f-b3e6-7fc111e24d61 · outbound

This paper cites Crossmap transformer: A crossmodal masked path transformer using double back-translation for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Crossmap transformer: A crossmodal masked path transformer using double back-translation for vision-and-language navigation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.410397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.155998Z digest=sha256:b4f086fba28d2ac3c312ba0c861a31e4a0205e7c20212ffbf65ca6277fea1837

Observation f84da2ac-b995-4950-8876-1ebac4b38a3f · outbound

This paper cites Improved speaker and navigator for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Improved speaker and navigator for vision-and-language navigation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.401744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.239153Z digest=sha256:91222875bbac9374395a8990b26ed10baceb60838482f358ae50a3baf32c5f53

Observation 0c5fa32d-cc31-4fe4-a8e0-f6298bde3b1b · outbound

This paper cites Res-sts: Referring expression speaker via self-training with scorer for goal-oriented vision-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Res-sts: Referring expression speaker via self-training with scorer for goal-oriented vision-language navigation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.392673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.295300Z digest=sha256:003518af84a8ebf2aefbe6a798b83581baa1b97aa47779750fd6a1a3cecbf54a

Observation a0c9bbcc-6905-443c-a24a-1dcc90d9ed0e · outbound

This paper cites Touchdown: Natural language navigation and spatial reasoning in visual street environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Touchdown: Natural language navigation and spatial reasoning in visual street environments,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.381772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.367875Z digest=sha256:ca4de8506e103d693f1241b770fa49786eb721b101fb8bddf753f455b4c9b780

Observation 6a7ef5d2-e715-423b-8567-3ba1fa884db6 · outbound

This paper cites Vision-language navigation with self-supervised auxil- iary reasoning tasks,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Vision-language navigation with self-supervised auxil- iary reasoning tasks,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.372815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.452079Z digest=sha256:795d90276b3500ec89b5c56c51df23f699ced56a67442f40e85e19ab62dca564

Observation 1eb139da-2ee3-49ec-85ea-c8e94ab8b33b · outbound

This paper cites Towards navigation by reasoning over spatial config- urations,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Towards navigation by reasoning over spatial config- urations,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.363650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.554838Z digest=sha256:66ddb1dd2d130bcaef449e5b7ba4e2fab91451749d16b887b24eb037eb935d79

Observation 34f7b84e-9f1c-4bf3-ac1d-e560385cd547 · outbound

This paper cites Language-guided navigation via cross-modal ground- ing and alternate adversarial learning,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Language-guided navigation via cross-modal ground- ing and alternate adversarial learning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.354831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.609153Z digest=sha256:6a7bf3a103019f105eb824084fa2cca327ac93f092035725ab1333a23f068374

Observation 21b5f186-52d0-4f3a-990a-1bcc4428e6e2 · outbound

This paper cites A dual semantic-aware recurrent global-adaptive net- work for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization A dual semantic-aware recurrent global-adaptive net- work for vision-and-language navigation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.345561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.716252Z digest=sha256:8b462dc80524e6586c4b6aae98724ea680f6eb5b681ec6d4c86e989f210c6128

Observation 007a28bd-c66b-4057-9f25-ab84b9db3c2a · outbound

This paper cites Vision-and-language navigation via causal learning,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Vision-and-language navigation via causal learning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.335692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.796758Z digest=sha256:8784dd3451657602acba46a2323bb85eb3dce833b2868fc634c86408c5a661b1

Observation bcacd41b-15ed-4c7e-8901-c55c2688cf0f · outbound

This paper cites Waypoint models for instruction-guided navigation in continuous environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Waypoint models for instruction-guided navigation in continuous environments,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.327313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.871862Z digest=sha256:5bf48323a788974e464e4bd0e4686ca5c527b423a67a2272b3bfce47c6e65304

Observation eec50d35-4578-4d91-bbad-ed303647c806 · outbound

This paper cites Bridging the gap between learning in discrete and continuous environments for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Bridging the gap between learning in discrete and continuous environments for vision-and-language navigation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.317948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:38.964356Z digest=sha256:a174b0d645e09e5d3c3394869049aa6150eaea3906ceddd1eb2882ce668f13e7

Observation fcf4d2f5-53ea-4e97-923c-2184a576ffa8 · outbound

This paper cites Instruction-aligned hierarchical waypoint planner for vision-and-language navigation in continuous environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Instruction-aligned hierarchical waypoint planner for vision-and-language navigation in continuous environments,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.308510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.031354Z digest=sha256:95e83fef3dd2e6beda7fa705d50d314de24df7cd2dae0fc0db61ba4cc2c189b3

Observation 8ddd9a38-05e6-472b-add2-8ea5728405b9 · outbound

This paper cites Improving vision-and-language navigation with image-text pairs from the web,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Improving vision-and-language navigation with image-text pairs from the web,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.299440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.088275Z digest=sha256:e9bbe77828b577c49c3b406300bb70c496cd611f8c9ae2c6fdbbff5fc22e24e5

Observation b6dd2d05-fc7a-4002-85c8-3741ea750657 · outbound

This paper cites Vln bert: A recurrent vision-and-language bert for navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Vln bert: A recurrent vision-and-language bert for navigation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.290467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.181868Z digest=sha256:d6f7249d2688932ca4e65f156e95d1b453420afb132e3195370136af73ce569b

Observation 8f05ce81-8683-43c7-bc49-de2b6baf70ba · outbound

This paper cites Think global, act local: Dual-scale graph transformer for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Think global, act local: Dual-scale graph transformer for vision-and-language navigation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.281913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.255109Z digest=sha256:97ea8b5ba641bd6fb112477c689595c3b6546d088e41c497e0d7008542cf16c0

Observation fcbc33e1-5ace-47b3-81b0-fc401782c2ed · outbound

This paper cites Multimodal evolutionary encoder for continuous vision- language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Multimodal evolutionary encoder for continuous vision- language navigation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.273145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.371677Z digest=sha256:35e34c209e9eb1156e9ee18d160b3568f480b1b7fa7f0b538366d6454a1aebff

Observation 7884320a-c9f4-469a-9b2e-152643a064d4 · outbound

This paper cites Lana: A language-capable navigator for instruction following and generation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Lana: A language-capable navigator for instruction following and generation,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.264496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.423682Z digest=sha256:b0e5ad68de3c8ac9c9246126b92e8f3dd6173949c458ef45b0c85d6ba590e24b

Observation 630df405-98ab-4d04-88f1-866f562a9a4d · outbound

This paper cites Pasts: Progress-aware spatio-temporal transformer speaker for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Pasts: Progress-aware spatio-temporal transformer speaker for vision-and-language navigation,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.255999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.509361Z digest=sha256:46315232e57398f234617676787a416de6443c4fab50774f555e47ef48a16f89

Observation 8a48d74d-8bc1-4388-ad64-42f79fd1f07b · outbound

This paper cites Envedit: Environment editing for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Envedit: Environment editing for vision-and-language navigation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.245998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.600011Z digest=sha256:b3530bf04b8bc0fe75b1c59921a22ea8856945b36cb6b7558e80b7c648eba582

Observation eb44402d-c6aa-4704-8cc9-b4d15dea0189 · outbound

This paper cites Less is more: Generating grounded navigation instruc- tions from landmarks,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Less is more: Generating grounded navigation instruc- tions from landmarks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.236709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.657739Z digest=sha256:2be2c7e4cf9be7a1a771aa3bda6be98b846e25ef25881e69148574e8d42290cc

Observation 81ba1bbc-87b9-4838-a7af-147d1a0540c9 · outbound

This paper cites Learning vision-and-language navigation from youtube videos,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Learning vision-and-language navigation from youtube videos,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.226926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.749105Z digest=sha256:e8951337df2b9ca341717d3fb18a55b9775db157c314de40a77ee0b7d391a049

Observation 61f4f532-c790-4ee5-9f6a-486991bddce8 · outbound

This paper cites Video captioning: A comparative review of where we are and which could be the route,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Video captioning: A comparative review of where we are and which could be the route,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.217843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.825425Z digest=sha256:adea79fbb5093a57488c89de81ee1a71bd484b32469b0f00be817db36a8b5b52

Observation 22ef3e08-ccdf-4734-8ce8-8e62679309e9 · outbound

This paper cites A review of deep learning for video captioning,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization A review of deep learning for video captioning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.209487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:39.922378Z digest=sha256:afe678ba97a19bf449f09ca99e2468c5b508a0a782656be1c64bc8770699286d

Observation d2458f56-b1f6-4903-8b55-56e5a6224c4d · outbound

This paper cites A survey of video datasets for grounded event understanding,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization A survey of video datasets for grounded event understanding,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.201062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:40.001084Z digest=sha256:5cccac71eb505d1337b114fd04e1056e60ce030b0e2e8116efe875105faa8648

Observation 77cd7376-8182-465e-9944-167429902d73 · outbound

This paper cites Dual-stream recurrent neural network for video caption- ing,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Dual-stream recurrent neural network for video caption- ing,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.191684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:40.140493Z digest=sha256:91a6e0057d0f5084b5f666a10a4003d41d9b94adb11f45cc93c3163334f3324e

Observation 8d03ee06-26ed-4750-b4cc-300ca7c23b15 · outbound

This paper cites Video captioning using global-local representation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Video captioning using global-local representation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.182805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:40.207516Z digest=sha256:fa232c68e73456754f094569c8463ae2ab18bf4f61c8d65ecc1a373571c39122

Observation 0458ea46-934a-4639-8938-dc4f54258e01 · outbound

This paper cites Evcap: Element-aware video captioning,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Evcap: Element-aware video captioning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.173422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:40.274303Z digest=sha256:66d09f8ea026926c11f6e300ae1499402345f2257078e2914eefbbee48196b07

Observation b26fa6d3-8e11-4c3e-9537-a920ae3bceae · outbound

This paper cites InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:40.342191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:40.342191Z digest=sha256:4a233c397a7d512d8b2056c23de185d5b8bfd15e6c7ef1115d96b3b5e5a5d332

Observation 6aae20aa-65e1-4b81-8d3c-1412558d63f6 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:40.424787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:40.424787Z digest=sha256:80d37885f055a327f6684c4b4d9d300a325f17e2290b2d4a7a0ab3323e27e6ff

Observation 5a0f4ef5-bca5-4426-8e87-71b62743ce9c · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Msr-vtt: A large video description dataset for bridging video and language,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.164482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:40.514568Z digest=sha256:931d39e24e4aab64401a65ba1a67b9e784e5a6ad3d802ba0aa025e35fc501c9b

Observation 0098bd94-98bd-4d4f-b0c4-c85abfd7fd19 · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Frozen in time: A joint video and image encoder for end-to-end retrieval,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.154235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:40.611316Z digest=sha256:48ffc865ae3cef8e423120b1643ca1272ce92a1b29c6edd9ceb0f403de019a06

Observation 91305db8-ed73-4a43-ae18-476114721c4e · outbound

This paper cites Beyond the nav-graph: Vision-and-language navigation in continuous environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Beyond the nav-graph: Vision-and-language navigation in continuous environments,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.144410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:40.722428Z digest=sha256:14f8214decc13960fd65653b507f7cd957bfe8a56cec7c959108cf8b0e4949bd

Observation 990bc053-7431-47a2-ba82-b380648b8fa7 · outbound

This paper cites Deep residual learning for image recognition,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Deep residual learning for image recognition,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.136022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:40.776309Z digest=sha256:d595935eee416f82a40f6d9469360b80e1004e3fb7ba8d134d911b0e2257c813

Observation 1fffbdb2-5fb0-4ea9-802f-e52533dc242c · outbound

This paper cites Object recognition from local scale-invariant features,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Object recognition from local scale-invariant features,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.127067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:40.864850Z digest=sha256:fb55838b6b8fa335b0bdd944be94e5a78fa692643b153ca7d251480d33379c95

Observation 2debacc8-7b06-45ae-b108-0611d651b798 · outbound

This paper cites Fast approximate nearest neighbors with au- tomatic algorithm configuration,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Fast approximate nearest neighbors with au- tomatic algorithm configuration,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.115716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:40.960158Z digest=sha256:86e02d52903e72d01631a724a815387c395bcb02f25ce6f31fd882d2b180a444

Observation d8d7c4b6-bd4f-43ac-a3c4-e7d6a716bff1 · outbound

This paper cites Revisiting weakly supervised pre-training of visual perception models,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Revisiting weakly supervised pre-training of visual perception models,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.105728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:41.044904Z digest=sha256:fb5bf4bffe116a7bf9ca661a190fc4a4e9f3d02c19e3aa6efa0eb4e0157e9b67

Observation 23f5e418-644f-47af-af40-1efbbb7e09b8 · outbound

This paper cites BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.096238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:41.136000Z digest=sha256:a0471c89e8ac4a436b65f0045b1047471017b846d66a13e0aa2a499e9cb1a2e8

Observation c4b53ea2-908e-4d6e-ac65-f4e5e6974249 · outbound

This paper cites End-to-end object detection with transformers,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization End-to-end object detection with transformers,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.085885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:41.246340Z digest=sha256:c1603d632b677224a5dc4bbd409b3e3cb8bda99b9fcd4fe0e979dd8f3b0dfa55

Observation 14a3a940-618e-426e-b011-b094ce65ca34 · outbound

This paper cites Visual instruction tuning,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Visual instruction tuning,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.077443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:41.321346Z digest=sha256:f61cbe211d62ae8a739b324405bf83b67c954c2bae64ea8f992b0a149df670e0

Observation 27eb50c0-19e7-4ef8-b2cd-281d96d8e02f · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Bleu: a method for automatic evaluation of machine translation,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.067482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:41.381224Z digest=sha256:6f097a2e8866862e1828f06c537c2f05183981957852a454caa9dd5aded4649f

Observation 7e4715a2-04d0-41fb-abd4-b97a82fd66a9 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Learning transferable visual models from natural language supervision,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.057996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:41.520663Z digest=sha256:2cb6df765679817ac477d6ea95ce13a23e2b4ad55488cb913598ff289e13e633

Observation fcc4dadf-e352-4213-a471-636f6c12b421 · outbound

This paper cites Cutting the gordian knot: The moving-average type–token ratio (mattr),.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Cutting the gordian knot: The moving-average type–token ratio (mattr),

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:41.645320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:41.645320Z digest=sha256:fd7727361f46d65bc9470ded8453120c2d75d8002077dc475974c799525face3

Observation 9d0a2613-97d5-4adc-93a8-0ae22ebcc604 · outbound

This paper cites Locally typical sampling,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Locally typical sampling,

Reference 53

Resolution
malformed identifier
no resolver link, observed 2026-08-06T17:26:41.755589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:41.755589Z digest=sha256:a60147b528c4abd735de00ca461fdb2f4c5dffb0b9e4c2d3bf2e632f0a59af82

Observation f0e708ec-b3a9-45e2-ae65-c66232e35414 · outbound

This paper cites Texygen: A benchmarking platform for text generation models,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Texygen: A benchmarking platform for text generation models,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.048817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:41.876239Z digest=sha256:8e9930e21ca924ccfe088cc15b82bfb90781172133fb9539533c929b8368e8a6

Observation 15627557-607d-42bc-9114-4b9486b1ce47 · outbound

This paper cites Standardizing the measurement of text diversity: A tool and a comparative analysis of scores,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Standardizing the measurement of text diversity: A tool and a comparative analysis of scores,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.017758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.017758Z digest=sha256:533707cf27a264c4ae8e7de2d46ec3053beeb20669c79732192bc31ebc06fa6c

Observation 8c2ba964-e269-4872-b6ac-2d45b22fb411 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization DINOv2: Learning Robust Visual Features without Supervision

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.152760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.152760Z digest=sha256:b317e0cea820a5f260fac6e715c07f5cc0d07639ceae022b3f0888f5dd7fe703

Observation a6530019-9259-4309-80ff-38276acef15a · outbound

This paper cites Masked autoencoders are scalable vision learners,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Masked autoencoders are scalable vision learners,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.038618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:42.420639Z digest=sha256:fc19fbbc9accbfdb014b6a79eef6545082808e458e8bb51d271bb0bb2f854f45

Observation 897009f2-c011-48d7-b139-16b23f4927cb · outbound

This paper cites Places: A 10 million image database for scene recognition,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Places: A 10 million image database for scene recognition,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.029351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:42.589056Z digest=sha256:1c6aff3b816582da0f525015880779010186b2b55e3c349a6a85c2cc2abcedb5

Observation 8f2f6ac3-43d4-4d5c-adb6-cb66239809e6 · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Llava-next: Improved reasoning, ocr, and world knowledge,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.020039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:42.755630Z digest=sha256:97ff187f253b4eb37c27f88877a7287105be5fd452004e86a03e15a26983e400

Observation c6cccb69-2e66-428c-a754-1c8b3e39b82d · outbound

This paper cites Qwen2.5-VL Technical Report.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Qwen2.5-VL Technical Report

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.870226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.870226Z digest=sha256:0148d7ffb8c1a9bd2a939334091e7ea4bef054e450fa9021b9c0a7d6883158cd

Observation 4452d2f4-3dbb-4912-beb9-97b5addd3fc4 · outbound

This paper cites GPT-4o mini: advancing cost-efficient in- telligence.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization GPT-4o mini: advancing cost-efficient in- telligence

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.011684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:42.873716Z digest=sha256:e49e726318fcae28adc8840f8b56a494607eef0e6db9af079a42c414b88356d8

Observation 4ab11380-5c49-4707-aab2-06129999fad6 · outbound

This paper cites The Llama 3 Herd of Models.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization The Llama 3 Herd of Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.877278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.877278Z digest=sha256:ca5f36b8d1282a1a8afe99931ebd86abc5c355b082a7c51bcb3c3cfba6246a71

Observation a8c8dc48-1b17-4bd8-b79c-91866c6a70e0 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Gemma 2: Improving Open Language Models at a Practical Size

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.892226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.892226Z digest=sha256:1cc4a3d01b6efad9274b2bef0937186db40b0b9c3f85ef6aa816fe9a0f5f777a

Observation b0bdb53a-ba0f-41b8-9619-687c91c69ae3 · outbound

This paper cites Qwen2.5 Technical Report.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Qwen2.5 Technical Report

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.996218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.996218Z digest=sha256:97efea54a48b521f5a9ac7864f24e29fe2295aac220807c63d34a5dceb47a05f

Observation b7bdd802-c3b5-401a-92a6-815bd1315e18 · outbound

This paper cites Openclip,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Openclip,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:43.092511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:43.092511Z digest=sha256:7c62391c95fc954ae7867bc1455836147662574f8c38be82bbb852401da08c3f

Observation fee5a68b-2a12-4e76-b3fd-6f2feac95f21 · outbound

This paper cites Torchmetrics - measuring reproducibility in pytorch,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Torchmetrics - measuring reproducibility in pytorch,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:43.191390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:43.191390Z digest=sha256:0eba5d364fdec4e225ad86e344ce29875ac5c60ccea7130b84784d23e2ab0f4d

Observation e7acea5c-bc75-4730-9fc6-307e4ca84cfd · outbound

This paper cites Coca: Contrastive captioners are image-text foundation models,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Coca: Contrastive captioners are image-text foundation models,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.001587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:43.275308Z digest=sha256:5b6bef4c6d5cd2014a9b35e585e04cc9a32051f36b67d46113b808787cb8dfd5

Observation f81e374e-5483-48e2-a264-9e627404a796 · outbound

This paper cites Llava-next: A strong zero-shot video understanding model,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Llava-next: A strong zero-shot video understanding model,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.992269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:43.314145Z digest=sha256:b6be57ce469efb2d5bf6defa58bf1d6c8076a5a04e9004d3cbb82345671b0823

Observation b1854ce4-7a63-4ad5-92b7-736cd657a4f0 · outbound

This paper cites Habitat-matterport 3d dataset (HM3d): 1000 large-scale 3d environments for embodied AI,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Habitat-matterport 3d dataset (HM3d): 1000 large-scale 3d environments for embodied AI,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.982841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:43.350446Z digest=sha256:02fb6dccbbc5e5f15b4f0e4c8cc712f13684449ec051e3b2b4d0a5334afd8748

Observation 657da7f3-f228-4c2a-9e96-5dd84ce52051 · outbound

This paper cites Scenenet rgb-d: Can 5m synthetic images beat generic imagenet pre-training on indoor segmentation?.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Scenenet rgb-d: Can 5m synthetic images beat generic imagenet pre-training on indoor segmentation?

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.973049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:43.380318Z digest=sha256:47d85cd4f1742f9814c0bf970c082de721423fac1ff0526fd40ddee78c032ef3

Observation 50046585-1b74-4c97-a40c-c8877d309220 · outbound

This paper cites GRUtopia: Dream General Robots in a City at Scale.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization GRUtopia: Dream General Robots in a City at Scale

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:43.401732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:43.401732Z digest=sha256:81772da3cb5c343b4293b6529f62104d463143ee03aa7057eed94eace805605b

Observation 0a01de21-e13d-40f3-9ec3-33838189d519 · outbound

This paper cites Sun3d: A database of big spaces reconstructed using sfm and object labels,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Sun3d: A database of big spaces reconstructed using sfm and object labels,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.927862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:43.407659Z digest=sha256:06c787c07e4b06df1d7b53501fcefd69a2207a186d044c1e64cbcf9ab1b4ddd3

Observation 505e524b-fc7c-4f03-a10f-e591e1571320 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Scannet: Richly-annotated 3d reconstructions of indoor scenes,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.849072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:43.413279Z digest=sha256:1da2f851a0133ece37927637ee5c7d768f6469ba35f036cc632e4209b76589fa

Observation a268142b-4b52-423d-a890-b6594a577936 · outbound

This paper cites A benchmark for the evaluation of rgb-d slam systems,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization A benchmark for the evaluation of rgb-d slam systems,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.798243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:43.419507Z digest=sha256:98bc36dbe36c4ba430cd8281cba19e703ccebaa2b813cad0f6ca00df4e37b867

Observation 6467cdaf-05aa-47e0-a458-28270e1a30ac · outbound

This paper cites Learning to navigate the energy landscape,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Learning to navigate the energy landscape,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.748638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:43.425274Z digest=sha256:75173e877bacb0fb77234bf871eddbcdbcd11c1a2f99b946682387a1570a8507

Observation 224e5921-e4fc-47e8-811d-6c2248a5c419 · outbound

This paper cites DIODE: A Dense Indoor and Outdoor DEpth Dataset.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization DIODE: A Dense Indoor and Outdoor DEpth Dataset

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:43.431393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:43.431393Z digest=sha256:27b8ac54c37899050e66f8a9961b767c8880ed2b22f4835546ba4572246c0590

Observation 3672e4eb-aa87-4be4-964d-ee67f5298424 · outbound

This paper cites Vision meets robotics: The kitti dataset,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Vision meets robotics: The kitti dataset,

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.696891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T17:26:43.437116Z digest=sha256:b5c913ce29df9cebd15117e3cf87c10879e3e7c65012c4a093eebc908e0d5561

Pith citing papers

No inbound Pith citation observations are available.