Pith. sign in

Paper Citation Record · LEDGER

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks

As of 15 August 2026, this Paper Citation Record lists 74 of 74 outbound references and 0 inbound Pith citation observations for arXiv:2411.17030.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17030 v1

Coverage vector

measured 74 of 74 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:41:12.121854Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

74 of 74 outbound references displayed

  • verified exact0
  • verified fuzzy55
  • unresolved17
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 87a76538-db03-4cf3-9019-3a399722a007 · outbound

This paper cites Bevbert: Multimodal map pre-training for language-guided navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Bevbert: Multimodal map pre-training for language-guided navigation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.998682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.872337Z digest=sha256:3b7a314154b21631323beef37ec0067023ecbf07dd833cb58304d66fb161e46a

Observation f1347bfa-b0b5-45a9-a855-7036e6eb5aab · outbound

This paper cites Etpnav: Evolving topo- logical planning for vision-language navigation in continuous environments.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Etpnav: Evolving topo- logical planning for vision-language navigation in continuous environments

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.986898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.876375Z digest=sha256:bc4c5b15b01b6bf2d759edaf16627ce635a5110db81101cc1fde86ff1720879b

Observation 4011b6c3-a942-4e00-9a57-48c24476a8e8 · outbound

This paper cites Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:11.879942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:11.879942Z digest=sha256:f43949312308ee7a378e4a11515a3ab37dffe49f10b71db0d8511678d70e3d24

Observation e22bb210-cc39-45d2-a090-9f82370aeffd · outbound

This paper cites Scanqa: 3d question answering for spatial scene understanding.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Scanqa: 3d question answering for spatial scene understanding

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.970597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.883947Z digest=sha256:a8ab13e257ff484f8b84a519cba00fc7610560aa0813b47a3ed3b6b85a72f085

Observation 82a71666-756a-4d72-ab5c-0c9a3acaae83 · outbound

This paper cites Matterport3d: Learning from rgb-d data in indoor environments.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Matterport3d: Learning from rgb-d data in indoor environments

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.960308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.887606Z digest=sha256:2cde6ad5aaf184294c7976bfba695984ccb06fe01867f63713454ab482b95682

Observation 231eda7a-b2a6-41d2-8a83-027011731b0c · outbound

This paper cites Object goal navigation using goal-oriented semantic exploration.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Object goal navigation using goal-oriented semantic exploration

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.950337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.891276Z digest=sha256:b3dd99d407c84c2e95b5edaef74d95ccdb571f6cbb4b5118476125588f345677

Observation b8969db9-f9b6-4a6c-9882-b91b6b68fba7 · outbound

This paper cites Scanrefer: 3d object localization in rgb-d scans using natural language.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Scanrefer: 3d object localization in rgb-d scans using natural language

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.939810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.895050Z digest=sha256:2f6504ad8f1bc384c55c749630aea875c1ab744a274a6bb37a3ca176025930ec

Observation 08e26698-f732-412f-b8db-f817edb853fb · outbound

This paper cites Weakly- supervised multi-granularity map learning for vision-and- language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Weakly- supervised multi-granularity map learning for vision-and- language navigation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.929169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.898655Z digest=sha256:58de3934ff46d5eb559b1cb9c3169fd10f8de13f00f34cd19e3a12d5a9927f0c

Observation c13ec763-50d5-4cf2-9c2b-5672480a839f · outbound

This paper cites History aware multimodal transformer for vision- and-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks History aware multimodal transformer for vision- and-language navigation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.918415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.901828Z digest=sha256:4ea6772ba6fa62ed0cd21b955fa96acc5d41bc5581a5c98a7653f793ed67ba40

Observation 2fdb889f-afea-430b-bf38-2bb66f3dd500 · outbound

This paper cites Think global, act lo- cal: Dual-scale graph transformer for vision-and-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Think global, act lo- cal: Dual-scale graph transformer for vision-and-language navigation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.906552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.905101Z digest=sha256:fafae7e671ab18525f709f1ae0201daa4a5bfe86a7c8cc6e8171553905b47dfa

Observation d395aad0-111f-4b6a-8f98-b241d1e41279 · outbound

This paper cites Grounded 3D-LLM with Referent Tokens.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Grounded 3D-LLM with Referent Tokens

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:11.908299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:11.908299Z digest=sha256:5a68545d5ca500ea5773030148033f2ad24b12d14fa316fd5d85ed395d35d490

Observation c428f878-7a90-4f20-b4e0-3ebcb61cef6c · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.895912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.912314Z digest=sha256:fa3db0eab2875275f1a3c449b3bb5e9c1a4e70108dc3b5dd4bc5c4dfad24215a

Observation 869cafd9-f504-4133-968f-6511437b6efc · outbound

This paper cites Embodied question answering.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Embodied question answering

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.885249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.915604Z digest=sha256:3b863216b079d5d57d13a453ca814c1a458b5f6d04cf547905758b70543ec19e

Observation 1885f780-6d26-43a1-98d4-9af56a37c413 · outbound

This paper cites Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:11.918948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:11.918948Z digest=sha256:1c85ca674917a5040ed0bc2ee7605ba06543ba23b09ccc49daf58a74fb4d67c8

Observation dec94811-4fb1-4952-ae31-51e3319b9022 · outbound

This paper cites Cows on pasture: Base- lines and benchmarks for language-driven zero-shot object navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Cows on pasture: Base- lines and benchmarks for language-driven zero-shot object navigation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.873615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.922474Z digest=sha256:37ba7a86e48d831881f073a1eef3928a9fbc54665150670f7b6f43d62b229c69

Observation 76e0c43a-b69d-4c27-833f-ab0f65af864f · outbound

This paper cites Cross-modal map learning for vision and language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Cross-modal map learning for vision and language navigation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.862860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.925637Z digest=sha256:51335cffdce6112d860ebbbe9737ee3b0e4b3813dda10f10d8d969e04b230491

Observation 32b708a5-5cbd-40d2-b21f-123e5ac94c43 · outbound

This paper cites Navigating to objects in the real world.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Navigating to objects in the real world

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:11.928995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:11.928995Z digest=sha256:84c543e3a3e24e72a5963e8f5eaad3b0b297aca7eccbfeb56eba81914b384d72

Observation 8ca44bc9-4fdb-402d-88a8-ff3e0ce14c12 · outbound

This paper cites Mask r-cnn.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Mask r-cnn

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:11.932150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:11.932150Z digest=sha256:e59a0058cb6e9e731296cf8caba26623297933cda26c4ac004212898f361259b

Observation 73872888-fe33-4669-bc14-3c9fecb41c28 · outbound

This paper cites Vln bert: A recurrent vision-and- language bert for navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Vln bert: A recurrent vision-and- language bert for navigation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.837099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.935910Z digest=sha256:d5725d4809bd1c1d0677ed4249d96703b7fcb068fae6bf5a1dc3067bcd7ce489

Observation d89b78d0-aa78-4caa-b59e-d6dbafd2157e · outbound

This paper cites Bridg- ing the gap between learning in discrete and continuous envi- ronments for vision-and-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Bridg- ing the gap between learning in discrete and continuous envi- ronments for vision-and-language navigation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.825192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.939141Z digest=sha256:6092543f6c9cdad5781e2ebf8c100e45f59f3a077c7ebdeb89e56c92386c24fb

Observation e9182e90-2f7b-42b1-8c19-7c45b0e82725 · outbound

This paper cites Learning naviga- tional visual representations with semantic map supervision.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Learning naviga- tional visual representations with semantic map supervision

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.813560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.942155Z digest=sha256:1f6215e9b1dad3f5481b477fdc4582f5c9c5bb0189ba62a3d9f14f5f43e48462

Observation 13d47039-932d-4663-a2f5-c1e7792a0eed · outbound

This paper cites An embodied generalist agent in 3d world.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks An embodied generalist agent in 3d world

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.802173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.945471Z digest=sha256:82e4aa721195a7c2352100eead055eab990ed935c90a6b13539712996fe0a813

Observation 4e10c0c2-b6b6-4435-a834-97ae2aa92747 · outbound

This paper cites Sceneverse: Scaling 3d vision-language learning for grounded scene un- derstanding.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Sceneverse: Scaling 3d vision-language learning for grounded scene un- derstanding

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.790379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.948651Z digest=sha256:1b683269807c25d3bb21dcb2dbd6e9c95172960401b98b63c31694c1fb655cd0

Observation 363fd063-7f4b-491e-979a-f2b99222bfa5 · outbound

This paper cites Lerf: Language embedded radiance fields.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Lerf: Language embedded radiance fields

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.779274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.951760Z digest=sha256:439c5a63d98914242282ede3ea177595d3a41daeda7d767884c441558b8ddd23

Observation 734d25ef-28dc-4fd6-95c4-bb9503c799b8 · outbound

This paper cites Segment any- thing.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Segment any- thing

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:11.955360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:11.955360Z digest=sha256:3d08c42dfdde07806db87b92b983f94d8e54c017b73a090c581371bfa1e530d3

Observation ee853e66-ad70-4388-9196-920578050de5 · outbound

This paper cites Sim-2-sim transfer for vision- and-language navigation in continuous environments.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Sim-2-sim transfer for vision- and-language navigation in continuous environments

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.762076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.958517Z digest=sha256:c9fbc34592e649e87a5f1318b0ee158ba74d550ec5b44e1e400a07b0131fe048

Observation 506da9f3-d1b5-4441-8e94-95f7f2d6f240 · outbound

This paper cites Beyond the nav-graph: Vision-and-language navigation in continuous environments.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Beyond the nav-graph: Vision-and-language navigation in continuous environments

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:11.961691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:11.961691Z digest=sha256:c4567c25ff1ec6f3c7c2640ad39500ea5c6b7b3ccc6d5cdb77a5e8ab065e115c

Observation b078446b-ad7f-45a0-a848-99aa3d927147 · outbound

This paper cites Renderable neural radiance map for visual navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Renderable neural radiance map for visual navigation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.733876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.968341Z digest=sha256:a05eabd7d84bbf5aa7c6150253df3626c96a37bdf48aded1ee8a30ed94e1c786

Observation 5c93262a-6c93-4592-a97b-2e991490be2e · outbound

This paper cites Less is more: Clipbert for video-and-language learning via sparse sampling.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Less is more: Clipbert for video-and-language learning via sparse sampling

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.723406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.971725Z digest=sha256:f4266af338a8d410a2e33b17a9c466c2a166d7604f3b460a3f06322580ec9912

Observation 9814a060-b2f8-4963-b14c-b1e8c3005b1a · outbound

This paper cites Blip- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Blip- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:11.974912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:11.974912Z digest=sha256:9968ceff85ec672e9595da2d57af70423b9bce86e93aabb05d52c7065bca696f

Observation 80dec119-ca83-433d-ae63-8da5493d4f70 · outbound

This paper cites Grounded language-image pre-training.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Grounded language-image pre-training

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.706517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.978223Z digest=sha256:68747a9ba6f6d5ae188491b6cda2b4d3ac78512ce4cdf5f853e86cfd89e938d1

Observation 96f7fa9f-6812-4373-9edf-e19d74043d10 · outbound

This paper cites Bird’s- eye-view scene graph for vision-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Bird’s- eye-view scene graph for vision-language navigation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.695814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.981613Z digest=sha256:933d25a258a219b5f30e768dad25e40ef228211bba83237da6e2cd52aeedf5d0

Observation a6a11262-f324-4279-940d-6c9f6ac64e3d · outbound

This paper cites Vision-language navigation with energy-based policy.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Vision-language navigation with energy-based policy

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.683914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.985225Z digest=sha256:b7dcfb20e44e4df0e3d54d8aa2539a5e63d3a6d7918de93d2a3dcab23e7355c6

Observation 2f963d95-0fbe-4569-9e28-cd0474b56ffb · outbound

This paper cites V olumetric envi- ronment representation for vision-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks V olumetric envi- ronment representation for vision-language navigation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.672126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.988628Z digest=sha256:a2b333e087d4215949aee65f1e06d42a93338f9f673ffc0e9c16940ac00ce13d

Observation d922b81b-96af-4e65-9622-4ba1d5d98a3b · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:11.992187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:11.992187Z digest=sha256:b516a842e3fcd828d287792d312162a66a8e7ac4c73c7a157708c40f4eb1cea9

Observation 6eaef7e8-4f0b-4f82-a952-fa7b0c0f42d4 · outbound

This paper cites Instructnav: Zero-shot system for generic instruction navigation in unexplored environment.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Instructnav: Zero-shot system for generic instruction navigation in unexplored environment

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.661182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.995710Z digest=sha256:46619b7c67bf2444c8cb403e48979c108f5a6838625586957d10527012331dfe

Observation 3af6b1bc-3dd0-4d51-85e6-a82ddbdd8796 · outbound

This paper cites Sqa3d: Situated question answering in 3d scenes.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Sqa3d: Situated question answering in 3d scenes

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.649709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.998993Z digest=sha256:dbaf792059211a61d301192a87835a6bb87dc33c3a86145ff8fc06a6b5c6996d

Observation acabc30d-ae47-41bc-822a-2efb0c3c2ae6 · outbound

This paper cites Rec- tifier nonlinearities improve neural network acoustic models.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Rec- tifier nonlinearities improve neural network acoustic models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.638752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.002258Z digest=sha256:977e0c7a144aa9f166ce176499e8c4b9eb757fd551cac85b2313ba598eec9860

Observation b219f52c-18d2-4a73-9c1c-d1b6b7334312 · outbound

This paper cites Zson: Zero-shot object-goal navigation using multimodal goal embeddings.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Zson: Zero-shot object-goal navigation using multimodal goal embeddings

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:12.005584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:12.005584Z digest=sha256:5f423ca1bdb0c93b63f3daef3b1bebf3652463258271ce7e1744d689b1bbb765

Observation 8f1cc444-c362-4453-8cf2-638a197fb633 · outbound

This paper cites Openeqa: Embodied question answering in the era of foundation models.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Openeqa: Embodied question answering in the era of foundation models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.622137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.009453Z digest=sha256:a0f7f0ae6ba32ca0bfba4094fa2c1e43192302d9d6172f2736b4d8106e8df86f

Observation 43b429da-8ac1-439c-a8d4-b5ecc2147b59 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view syn- thesis.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Nerf: Representing scenes as neural radiance fields for view syn- thesis

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:12.012861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:12.012861Z digest=sha256:3bdcac15f248df5d1e36c135e255a07dea86303299b3f482704ba49ea0f96ac6

Observation 799e14c3-c2b6-45c3-97a3-c889cc582bef · outbound

This paper cites Dinov2: Learning robust visual features without supervision.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Dinov2: Learning robust visual features without supervision

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.605027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.016089Z digest=sha256:f8256a711c93af23d2132f72881067ba281ea0988c52f54ab7c342a15e9358de

Observation 0660d706-73e8-4c6c-8ad2-570514e5ff43 · outbound

This paper cites Hop+: History-enhanced and order-aware pre-training for vision-and-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Hop+: History-enhanced and order-aware pre-training for vision-and-language navigation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.594285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.019609Z digest=sha256:ca2978596af9fc10fe1b3745b5245a9bedc984e4fe0bf881c3ebee50db0aa989

Observation 4bc9e5cb-477f-485f-8bdb-6e86c9f83573 · outbound

This paper cites Learning Generalizable Feature Fields for Mobile Manipulation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Learning Generalizable Feature Fields for Mobile Manipulation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:12.022955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:12.022955Z digest=sha256:16ea8da8ae7ef05a1299f5f6522cacf78a72c39bc702fdaa67183f156c4fc295

Observation 9029cb07-d3b0-41e2-9958-a4a5f6b4e352 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Learning transferable visual models from natural language supervi- sion

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:12.026703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:12.026703Z digest=sha256:2e647fd389bf7eb5f6778139e07d5211f1f5bc1829edd3d76238019805031ea9

Observation f94223b8-6c9a-4f83-b735-5189b4c35249 · outbound

This paper cites Habitat-matterport 3d dataset (hm3d): 1000 large-scale 3d environments for embodied ai.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Habitat-matterport 3d dataset (hm3d): 1000 large-scale 3d environments for embodied ai

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.577550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.029853Z digest=sha256:b2773acb5b4f20d19f42dad691aaab68fa462b43da4653e7af986bc8a3067276

Observation 9f13ec18-5be7-4d83-8e1e-5953cc51592b · outbound

This paper cites Poni: Potential functions for objectgoal navigation with interaction- free learning.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Poni: Potential functions for objectgoal navigation with interaction- free learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.567723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.033132Z digest=sha256:2859f1b968f4a99b8cff1c62276d322d9789ae1fbb90d0cb40cd9469bf39a63d

Observation 275dd167-0b4f-4ff0-a88c-3f515edf0292 · outbound

This paper cites Habitat: A platform for embodied ai research.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Habitat: A platform for embodied ai research

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.557450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.036557Z digest=sha256:2a54f85c8017211ffd7da539d71c12718319c1b7e5f96f0043747fdff31a3787

Observation 30efd223-28ce-48db-be1a-e388086bcf4a · outbound

This paper cites Distilled feature fields enable few-shot language-guided manipulation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Distilled feature fields enable few-shot language-guided manipulation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.547216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.040645Z digest=sha256:274a6fd1b7448d39b89c9bf3a61b0dd71f33c4911bc64427c19974190d250a1e

Observation 3a295e05-5e5f-4f1e-ae33-616a181d9ab5 · outbound

This paper cites Language- enhanced rnr-map: Querying renderable neural radiance field maps with natural language.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Language- enhanced rnr-map: Querying renderable neural radiance field maps with natural language

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.536766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.043939Z digest=sha256:9cc6d0d8b79a8d1fa596d82a9d9d52715fb263bb3371e11eec31761f9087c208

Observation 38570152-e3ba-45fe-b1f1-3103b9430c44 · outbound

This paper cites Nesf: Neural semantic fields for generalizable semantic segmentation of 3d scenes.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Nesf: Neural semantic fields for generalizable semantic segmentation of 3d scenes

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.526238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.047825Z digest=sha256:e3c014d8316c0576d8ea5d50b0dce31c85c5c40f4d7c99dedef5b8018c089484

Observation 0c292a5f-4771-4abf-b773-dc6b58d0ac3e · outbound

This paper cites Yolov7: Trainable bag-of-freebies sets new state-of- the-art for real-time object detectors.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Yolov7: Trainable bag-of-freebies sets new state-of- the-art for real-time object detectors

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.514554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.051271Z digest=sha256:9195d11cb391689d322e76920a5bc9c24e091f35e9c2bfdad1947d5c88e1f31e

Observation e754e9da-692d-4d63-8c4f-1fd9d0d90841 · outbound

This paper cites Dreamwalker: Mental planning for continuous vision- language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Dreamwalker: Mental planning for continuous vision- language navigation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.504479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.054825Z digest=sha256:0c8f4e25d480dc5a7caa5919dc9fc1f2c77b66a84274d982e8a2ccfe5b26239e

Observation 7ef24cfe-8734-4619-8497-4c8357801ad9 · outbound

This paper cites Vision-and-language naviga- tion via causal learning.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Vision-and-language naviga- tion via causal learning

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.493108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.057945Z digest=sha256:9cff0dce58d779ce4303585bbee5605f37480e5b099a8983c1472a463f1490fd

Observation c9637c06-818f-4736-9e5b-f8a8ebf87488 · outbound

This paper cites Scaling data generation in vision-and-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Scaling data generation in vision-and-language navigation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.378148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.061317Z digest=sha256:c1af8568c173b9f9c6a9c8e6be10fa864b6fdceb9976cd4668f08a36d4d668a6

Observation 797ef658-db03-4662-8867-767ab07b4058 · outbound

This paper cites Gridmm: Grid memory map for vision- and-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Gridmm: Grid memory map for vision- and-language navigation

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.367838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.064542Z digest=sha256:aa6d26e5c43f88484bc5013a73a4b9fa2f46d6cec8fcfb366af8631b0e4bceef

Observation 7f27c71e-ab7d-4931-92b5-81434c9ad481 · outbound

This paper cites Lookahead exploration with neural radiance representation for continuous vision-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Lookahead exploration with neural radiance representation for continuous vision-language navigation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.357768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.067942Z digest=sha256:39db27101d0ef2ba2410cda1b98a129e3388ba073aca64dc624042ddbd9ae5e6

Observation 6c45ba32-515f-4d69-a2b9-2e2b267c89e8 · outbound

This paper cites Sim-to-real transfer via 3d feature fields for vision-and-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Sim-to-real transfer via 3d feature fields for vision-and-language navigation

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.346864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.071256Z digest=sha256:8e19ed3c8cf17872caf8e41ee9f9f1d111c1eeedb0511bfae501d93149cb6d56

Observation 0ed69188-07b2-48da-aa16-dd27c97f58b6 · outbound

This paper cites DD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion Frames.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks DD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion Frames

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:12.074625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:12.074625Z digest=sha256:6856b7ff6520a12890915a4ece6d9f3a3a5a0ab7bb9d64e509341fe11ff936b2

Observation f260ae16-0bf0-4afc-bba2-23a88b0ccaed · outbound

This paper cites Habitat-matterport 3d semantics dataset.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Habitat-matterport 3d semantics dataset

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.336927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.078457Z digest=sha256:82e809ed01b403558d2cdeaf0f2b2d0829ae583dfa0e0e05f5eb09439cd9b88a

Observation fa71a361-cad1-48f5-91c6-85029c07c25b · outbound

This paper cites Sg-nav: Online 3d scene graph prompting for llm-based zero- shot object navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Sg-nav: Online 3d scene graph prompting for llm-based zero- shot object navigation

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.327125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.081615Z digest=sha256:6947eef7e0059283ef899ff57fb11d2548bb668b323654f5c940833fe85fea18

Observation b40685bc-3f62-49ed-a08a-561e8dd76c41 · outbound

This paper cites Vlfm: Vision-language frontier maps for zero-shot semantic navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Vlfm: Vision-language frontier maps for zero-shot semantic navigation

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.316777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.084976Z digest=sha256:6eb22bde2abae07be2d65dcb04d3351665a1c55d5e6d2899303c15ce1fc8d621

Observation a1bd8e4a-31d3-4df6-9ba7-00ac66e88785 · outbound

This paper cites Gamap: Zero-shot object goal navigation with multi-scale geometric-affordance guidance.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Gamap: Zero-shot object goal navigation with multi-scale geometric-affordance guidance

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.306045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.088244Z digest=sha256:8c56d43d26e03dd7b1e13e851cb37aacd8b852c8d96f2f831635c224a077a208

Observation 7462ac0a-36cd-40e2-8cc9-e867d1b13a85 · outbound

This paper cites Gnfactor: Multi-task real robot learning with generalizable neural feature fields.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Gnfactor: Multi-task real robot learning with generalizable neural feature fields

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.295173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.091402Z digest=sha256:e58003388c907def7f58f14a5e46c277aa4c084bc8097ec3ac955ed99af1875e

Observation 2c148112-3bef-49ce-bd06-f546961a116e · outbound

This paper cites Faster Segment Anything: Towards Lightweight SAM for Mobile Applications.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Faster Segment Anything: Towards Lightweight SAM for Mobile Applications

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:12.094566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:12.094566Z digest=sha256:f7609a0e7ee5da784d200194c13ea7cc0ee300ece3e490099b0dbc5dda397f38

Observation 4120d482-e917-4997-adff-0b203aa3310c · outbound

This paper cites VLA-3D: A Dataset for 3D Semantic Scene Understanding and Navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks VLA-3D: A Dataset for 3D Semantic Scene Understanding and Navigation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:12.098137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:12.098137Z digest=sha256:b76c336b0bc6d508ed934b5e6f95bc8b4434fde18dbbf6a2c6db74a7f7f9a2d3

Observation 8aeaf061-8751-411e-a01c-9160b7a018b3 · outbound

This paper cites Navid: Video-based vlm plans the next step for vision-and-language navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Navid: Video-based vlm plans the next step for vision-and-language navigation

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.284146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.102131Z digest=sha256:78419631b06444129f1fffe250f00d5d327be01fce0935f365ec413ff9ece267

Observation fa271cb0-fd70-420d-a31c-f992cd34dab4 · outbound

This paper cites Hierarchical object-to-zone graph for object navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Hierarchical object-to-zone graph for object navigation

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.273067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.105441Z digest=sha256:c6e64b5c89a356ead4afb6d9dc4290038c3acfb4a5cd24999c14dd96b24b43bb

Observation 3f685f3c-5c53-4d7f-821c-010b9b0811ea · outbound

This paper cites Vision-and-Language Navigation Today and Tomorrow: A Survey in the Era of Foundation Models.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Vision-and-Language Navigation Today and Tomorrow: A Survey in the Era of Foundation Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T12:41:12.108676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:41:12.108676Z digest=sha256:8070053cd9c0bd7ab9f9831614d972878c457d39f7f7a42f6368277e20f045cd

Observation 972c91b2-2fd4-46f2-9052-4f568cad6185 · outbound

This paper cites Structured3d: A large photo-realistic dataset for structured 3d modeling.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Structured3d: A large photo-realistic dataset for structured 3d modeling

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.261828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.112252Z digest=sha256:e1f965f6d60bc2bb3d56db248a4d89f56c43c9e3b076cbf43f48dc028588f499

Observation 5b35027b-bcc9-4f56-8f6d-7fd025de0259 · outbound

This paper cites Esc: Ex- ploration with soft commonsense constraints for zero-shot object navigation.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Esc: Ex- ploration with soft commonsense constraints for zero-shot object navigation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.251040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.115389Z digest=sha256:7d6ea9544de6967ab3eee89ca97ac9df387c5e4cdd36917a9c4b7b9003577d21

Observation 010eede6-6f5d-4f05-bb16-3215ce0de232 · outbound

This paper cites Target-driven visual navigation in indoor scenes using deep reinforcement learning.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Target-driven visual navigation in indoor scenes using deep reinforcement learning

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:41:12.240154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.118655Z digest=sha256:bc822ddc92954af7898f7aa03be3cf0bc08fb5a51688a3f53f5bafa822fbf1ed

Observation b7cf2336-14b3-4f69-b85b-ac37afa3c1de · outbound

This paper cites 3d-vista: Pre-trained transformer for 3d vision and text alignment.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks 3d-vista: Pre-trained transformer for 3d vision and text alignment

Reference 73

Resolution
malformed identifier
raw_fallback, observed 2026-08-12T12:41:12.228653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:12.121854Z digest=sha256:7afe93a6289109534178bc389df3950a9d18e8dd9c5510eebeb6c2db86990373

Observation dbfdaba3-38a0-45a0-bee0-57375ecc9b91 · outbound

This paper cites an unresolved cited work.

g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks Unresolved cited work

Reference 120

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T12:41:12.744699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:41:11.965153Z digest=sha256:ea2cd41dbcc6a96f0fc70f8076388f39f8c28ce3b758d90a387e2fc72ef8f19d

Pith citing papers

No inbound Pith citation observations are available.