Pith. sign in

Paper Citation Record · LEDGER

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation

As of 13 August 2026, this Paper Citation Record lists 92 of 92 outbound references and 0 inbound Pith citation observations for arXiv:2412.02402.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.02402 v2

Coverage vector

measured 92 of 92 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T23:35:20.663843Z

measured 92 of 92 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

92 of 92 outbound references displayed

  • verified exact11
  • verified fuzzy44
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 66c673e3-7849-474e-94de-fca8aa8d2b86 · outbound

This paper cites Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.231480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.231480Z digest=sha256:4e0d7c6f2d1fd84bdcdc91a69bebd02b2a6d10dded1c09d3c80c1714c7aa160a

Observation 86418cd4-de27-40e3-badd-8b3a9bff1457 · outbound

This paper cites Localizing moments in video with natural language.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Localizing moments in video with natural language

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.237165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.237165Z digest=sha256:581e9623dfce450251384023b4cbdb43f2782cb00afc87b91d34c85e69ba8664

Observation 93102e63-0f06-4130-8ba7-e079449efc61 · outbound

This paper cites Scanqa: 3d question answering for spatial scene understanding.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Scanqa: 3d question answering for spatial scene understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.242135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.242135Z digest=sha256:21dac6c81a815db3bda482e518b778aeead1dae8d6f5c0818e0221a4fdfa67c8

Observation ef71f953-536f-4a36-ad73-df001c704b25 · outbound

This paper cites WeaQA: Weak Supervision via Captions for Visual Question Answering.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation WeaQA: Weak Supervision via Captions for Visual Question Answering

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:35:21.091139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.246860Z digest=sha256:97c8b1238d3a497a70eb449d2d21df45ae8aaa6918279bc4d296f91a75aa6e78

Observation 3a6efffc-3544-44a3-a3dc-f0e247487913 · outbound

This paper cites Scanrefer: 3d object localization in rgb-d scans using natural language.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Scanrefer: 3d object localization in rgb-d scans using natural language

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.252192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.252192Z digest=sha256:0bfd23aebc48e0f2d3fe47c68f22958cfed56d4e02888d30bfe0a74972ad0475

Observation 65261456-3502-4684-a126-cbbf2bf7ae4f · outbound

This paper cites Language conditioned spatial relation reasoning for 3d object grounding.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Language conditioned spatial relation reasoning for 3d object grounding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.257074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.257074Z digest=sha256:dcb8e4b8965a87f060205eb1bfdea8fb8bc46acf3839359e94e9e6912d00a917

Observation f2705799-0cea-40c3-87ac-b5c63327b8c6 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.262126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.262126Z digest=sha256:0e6824ff2d0e991db1b3c3cfd85224e3df324b131eddaae166b749b0263d10e8

Observation c55aecd9-e080-4c3b-a533-938c503cb105 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.266674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.266674Z digest=sha256:2100e342c97bbaa8809c5ff01df25f25a54533c653a56de331b84c1203a3d79d

Observation 64fdb21e-e3e2-408e-828f-efe4e26819f7 · outbound

This paper cites Weak Supervision and Referring Attention for Temporal-Textual Association Learning.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Weak Supervision and Referring Attention for Temporal-Textual Association Learning

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:35:21.052476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.271471Z digest=sha256:837e50c6e6d3e8a83d2b918c539727edda7cac81b299aede9ebdbd2bd92e0150

Observation 0b181c94-4aff-4b1e-9827-09d26aee646c · outbound

This paper cites Video-of-thought: Step-by-step video reasoning from perception to cognition.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Video-of-thought: Step-by-step video reasoning from perception to cognition

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.276661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.276661Z digest=sha256:7161d40db9dc678b9f106371a7475c529d3c377847f8d93ac3e057bf300c35af

Observation ccacbda3-33ca-4d30-a573-22c608da1381 · outbound

This paper cites Vitron: A unified pixel-level vision llm for understanding, generating, segmenting, editing, 2024.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Vitron: A unified pixel-level vision llm for understanding, generating, segmenting, editing, 2024

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.281463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.281463Z digest=sha256:7b458496a3022cc44130b2c7976aa243d0a77861cf60f18f18a449614e021a95

Observation 5f36c49f-a69c-45f6-9919-79725e0eb90c · outbound

This paper cites Enhancing video-language representations with structural spatio-temporal alignment.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Enhancing video-language representations with structural spatio-temporal alignment

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.286155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.286155Z digest=sha256:90dd7ca65ac3114257549d1e44930880275ff3688fa8f7c08633b8a0a71e46df

Observation 5d7c917a-d0e2-4e00-97b0-1688810e1c8f · outbound

This paper cites Free-form description guided 3d visual graph network for object grounding in point cloud.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Free-form description guided 3d visual graph network for object grounding in point cloud

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.290172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.290172Z digest=sha256:176244d9c1d3f2220ec4c79b6be47c9cd2719deb51d44353b83bea7c26049e8d

Observation a072464d-e775-4f20-a086-1fc5fd323162 · outbound

This paper cites Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.294196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.294196Z digest=sha256:708c493c0d643350f9eab2763ff8b9681c43f4cb90c99c19de06fb9c03640746

Observation dd6ba9dc-c8f1-4de6-b1f3-92eb9eb59cdd · outbound

This paper cites Structured multi- modal feature embedding and alignment for image-sentence retrieval.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Structured multi- modal feature embedding and alignment for image-sentence retrieval

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.298756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.298756Z digest=sha256:f566ed5fa76b46b7c4ffe5700795c13cfb694026e146d271c342bf3665cf6053

Observation 5240e487-12a5-4aee-a80c-b9c08f1d5c58 · outbound

This paper cites 3shnet: Boosting image–sentence retrieval via visual semantic–spatial self-highlighting.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation 3shnet: Boosting image–sentence retrieval via visual semantic–spatial self-highlighting

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.302786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.302786Z digest=sha256:3470c80e247ab4b4bc6b0c78a02a332bb8fcde15a36412a4ba21ee0a6d5f6d05

Observation 8df2d7c9-d48d-4016-93a4-10e8a36c739f · outbound

This paper cites Vqa-lol: Visual question answering under the lens of logic.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Vqa-lol: Visual question answering under the lens of logic

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.948690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.307081Z digest=sha256:8b0bb5aba2fa65454efde10df90b8da45a8897d16352518b23e2ee683afc6fb7

Observation 897f0df5-a29a-4306-a367-ee33934d29fc · outbound

This paper cites Person re-identification method based on color attack and joint defence.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Person re-identification method based on color attack and joint defence

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.932779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.314186Z digest=sha256:87bb2d19c67391d81f563b70da376691d19023ad4fc1018a0e21a06066c200f2

Observation 4c7deca1-cbb8-4cee-8693-e6a5b4dfa186 · outbound

This paper cites 3d semantic segmentation with submanifold sparse convolutional networks.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation 3d semantic segmentation with submanifold sparse convolutional networks

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.319004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.319004Z digest=sha256:e1623c4688ed6ca978b52bd5122555e868e02afa4ffd8e3c08f1c51cc1c9e160

Observation dfc35995-6698-40b7-b170-4d0f3c04438d · outbound

This paper cites Segpoint: Segment any point cloud via large language model.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Segpoint: Segment any point cloud via large language model

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.906666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.323821Z digest=sha256:72fade4661d203abcdc0048476cf424e6b0a988725da73ff22e04cb8c7db793f

Observation d407bdaf-be1f-423e-a986-861936cedd2a · outbound

This paper cites 3d-llm: Injecting the 3d world into large language models.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation 3d-llm: Injecting the 3d world into large language models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.328792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.328792Z digest=sha256:0caaf224c34435cd9e74268c124f0afaad2771fa5d738b967e9b22c67a024103

Observation 9ea74b37-bbbf-49b7-9edf-769a8dc1b8c2 · outbound

This paper cites Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.333277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.333277Z digest=sha256:c303e8c6ac2d69f8801b459f2829938a4a384d61cbf4362c827efb268416b5c2

Observation 691a22b6-8e62-41fd-9020-fb9b3ad7ce03 · outbound

This paper cites Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.338471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.338471Z digest=sha256:ae4d65a9976b8570bb166cf0c915b0cef4c4af565cf7ac1f72203e83334b23d1

Observation e1e7b209-f73d-4c38-a180-d146645478ac · outbound

This paper cites Text-guided graph neural networks for referring 3d instance segmentation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Text-guided graph neural networks for referring 3d instance segmentation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.343905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.343905Z digest=sha256:cae50e59b4b5a766e345b14c1acf01608c9517e7e5ae2f33e89f2fa77bafd318

Observation 86865ac9-c6c7-45cd-b682-100b786c5381 · outbound

This paper cites Referring image segmentation via joint mask contextual embedding learning and progressive alignment network.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Referring image segmentation via joint mask contextual embedding learning and progressive alignment network

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.348573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.348573Z digest=sha256:797a70689d78e8623965b113629f7ea868f1fe5b38a68b44e194675d46480940

Observation 7ff30ec0-7027-4525-b672-9fdc7dc4c104 · outbound

This paper cites Bottom up top down detection transformers for language grounding in images and point clouds.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Bottom up top down detection transformers for language grounding in images and point clouds

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.353513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.353513Z digest=sha256:c7afa8729482986156bd416f799a7f00b467c3c0b8f6979e9fc034d4d85457ab

Observation 5e9ce923-4eec-41b9-b426-0ecfd825ae0e · outbound

This paper cites Comprehensive multi-modal interactions for referring image segmentation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Comprehensive multi-modal interactions for referring image segmentation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.857343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.358251Z digest=sha256:178eef3955333de9ddf720ffac4e8d771ea64913a23b01233f4b073027a001e2

Observation 5ecf5cb1-3442-460f-bd97-e2179ad1acd9 · outbound

This paper cites Weak Supervision helps Emergence of Word-Object Alignment and improves Vision-Language Tasks.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Weak Supervision helps Emergence of Word-Object Alignment and improves Vision-Language Tasks

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:35:20.984655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.362722Z digest=sha256:9ca92b94babd277f8b835e3d136fc8e14a735468f1736d8e865a0aad36dc4e18

Observation 5b6d9b74-b99c-47cf-8df3-40cb84a9646a · outbound

This paper cites Flexible visual grounding.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Flexible visual grounding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.367678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.367678Z digest=sha256:f442a074153a3cfac369ed3b54c256420701205921696af5c66424d919952132

Observation 23b67e82-1e64-4f4f-8748-6df33ed8ba79 · outbound

This paper cites Stratified transformer for 3d point cloud segmentation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Stratified transformer for 3d point cloud segmentation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.841054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.372326Z digest=sha256:24afe544c9d64b1fc6e8eea2fcadd52c4a78666fe7687b26f826d1b53da10af5

Observation f9134daf-ecde-4535-8558-0a223070e87b · outbound

This paper cites Mask-attention-free transformer for 3d instance segmentation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Mask-attention-free transformer for 3d instance segmentation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.823421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.377072Z digest=sha256:a18d12149975851893efd010c54bc1212b40b3a345ba5965b0490f1eb7207497

Observation a75278c8-ea1f-4ae7-ac20-58ca0c0423ab · outbound

This paper cites Large-scale point cloud semantic segmentation with superpoint graphs.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Large-scale point cloud semantic segmentation with superpoint graphs

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.806638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.381838Z digest=sha256:eb9d5f06c286692a6c3b2600fc66c27d4be4a641815563c5c593056ef69fdd96

Observation 1e2f7944-cbe1-4dc6-a99f-c6202e6f3995 · outbound

This paper cites Weakly supervised referring image segmentation with intra-chunk and inter-chunk consistency.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Weakly supervised referring image segmentation with intra-chunk and inter-chunk consistency

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.790357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.386542Z digest=sha256:e560cf4e659ce0c5fac1eca81c919a3df299f6ee842d4570c79a82ff2b976ac5

Observation 5200bb1a-b612-42a5-b091-564413e63cb1 · outbound

This paper cites Fully and weakly supervised referring expression segmentation with end-to-end learning.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Fully and weakly supervised referring expression segmentation with end-to-end learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.773055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.391780Z digest=sha256:b57e6c4c8eb46c8d7dc1d5a843c17163aafc1d60f8726152dd243dfc898bf7d9

Observation abc6104c-4c98-47d9-91da-7d4ebacc9137 · outbound

This paper cites Fine-grained semantically aligned vision-language pre-training.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Fine-grained semantically aligned vision-language pre-training

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.755948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.396223Z digest=sha256:0e24aa25e40e33833adab431477344274f500941c535f984c5c4237da472e896

Observation 2de1f385-75fe-4565-8529-1ea59c171791 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.400784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.400784Z digest=sha256:f42ec0ad82eb2b3d7d6c0595642248da0d69b6984c738ef33827e8e671e7bba1

Observation c3a78231-fd75-4047-869c-a2614e1e8b47 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.405459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.405459Z digest=sha256:019ed21a142d6695de45071899d6e3de163bb70810b3f824825e3c6fd7ab9904

Observation eff90e3a-9b6e-4419-815b-3f703f76ecce · outbound

This paper cites Toist: Task oriented instance segmentation transformer with noun-pronoun distillation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Toist: Task oriented instance segmentation transformer with noun-pronoun distillation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.717881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.410222Z digest=sha256:95189d640932eb2c1a3ce898b7b417e4128ba63cc80e985028ff1b35bd35c21f

Observation 3a6f97ec-6e78-4aa3-9678-072252225644 · outbound

This paper cites Understanding embodied reference with touch-line transformer.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Understanding embodied reference with touch-line transformer

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.700486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.414906Z digest=sha256:397384518c2c24bcd85aae5027dc23c83a1f7bbede31a94a2ecd9961033c1050

Observation 525d8460-5579-44a8-9e64-32b9638cd338 · outbound

This paper cites Transformer-empowered invariant grounding for video question answering.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Transformer-empowered invariant grounding for video question answering

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.682439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.419412Z digest=sha256:07b234c3885b37ba6dcfc86c64d3c0a5ca98b043690603de295ce58ea75d1abd

Observation a61a852c-095a-44bd-91eb-17b7c7fdc2b3 · outbound

This paper cites Laso: Language-guided affordance segmentation on 3d object.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Laso: Language-guided affordance segmentation on 3d object

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.665764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.424006Z digest=sha256:c77e3ecc1301420dcfa55536bdc9c90a4630dd02debb58f46347708d17368d05

Observation e03c661d-88ee-49b9-8d26-1c9e0b5ba7ef · outbound

This paper cites Instance segmentation in 3d scenes using semantic superpoint tree networks.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Instance segmentation in 3d scenes using semantic superpoint tree networks

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.649250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.428757Z digest=sha256:99c3d7ea5e1108b2f6dfd827ed31b61ef30ebea163e01b9d6d458cb4861b181b

Observation 401b6487-1e10-4102-b747-2d0b70773dc8 · outbound

This paper cites A Unified Framework for 3D Point Cloud Visual Grounding.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation A Unified Framework for 3D Point Cloud Visual Grounding

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:35:20.962563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.433380Z digest=sha256:2e98b2f197db7558d3bf463a94d11e4b01b24a41816be9768eab9e0461e2fc58

Observation cc8be356-efdd-4477-8666-45810174251e · outbound

This paper cites Referring image segmentation using text supervision.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Referring image segmentation using text supervision

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.632219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.438339Z digest=sha256:15dc3bea232817b9003b9cf1c27d0707af0f794936682600b40e97a94cd3c10d

Observation ad626953-6202-426f-91e9-e41633bb18d0 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.442554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.442554Z digest=sha256:4358d486a937416c521c58e5e56ee14f5e747cd1315c5ad17d491b80287c31a2

Observation be229806-f0f1-4985-ada9-86e792f6e113 · outbound

This paper cites Scaneru: Interactive 3d visual grounding based on embodied reference understanding.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Scaneru: Interactive 3d visual grounding based on embodied reference understanding

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.615114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.446933Z digest=sha256:ef216d88b52c10af50abbbbc22967443401ca31e0187f95e6b903f1ea54c5121

Observation f8f36734-64ce-4e7e-b9f4-7cc6589206c5 · outbound

This paper cites 3d-sps: Single-stage 3d visual grounding via referred point progressive selection.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation 3d-sps: Single-stage 3d visual grounding via referred point progressive selection

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.451182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.451182Z digest=sha256:2bcdf7368311bf9346077048191245bc6b5e9c56cee2fbfe3d6c8da2c4b44478

Observation d77d5778-af78-4928-b193-6f69ca26310e · outbound

This paper cites The stanford corenlp natural language processing toolkit.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation The stanford corenlp natural language processing toolkit

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.586741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.455490Z digest=sha256:f98940e02bb0bbaace31438562588c7fb2f6a8eb96a0356d6145fd215a82a594

Observation cfc0ccf2-800f-4554-ba21-1b64175489b5 · outbound

This paper cites V-net: Fully convolutional neural networks for volumetric medical image segmentation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation V-net: Fully convolutional neural networks for volumetric medical image segmentation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.459876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.459876Z digest=sha256:7d5a1e52aab7d0e5cb0bcea888637ad4a5ae6b40a1213a6f10696b5cb317c90e

Observation 5586e9d9-e49f-45ac-aefa-cb2cf100d86e · outbound

This paper cites Weakly supervised video moment retrieval from text queries.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Weakly supervised video moment retrieval from text queries

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.559697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.464165Z digest=sha256:4bc3acb2edf3ebf4beaf34149aa0a9c33bb63c8f6c380b88fc33647d2f64dc20

Observation 62bd7a5c-4078-4e29-8bda-43fc4ea8c32e · outbound

This paper cites Bridging the Gap between 2D and 3D Visual Question Answering: A Fusion Approach for 3D VQA.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Bridging the Gap between 2D and 3D Visual Question Answering: A Fusion Approach for 3D VQA

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:35:20.926325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.468888Z digest=sha256:db7d8e538f7cca0c1d69d630f4b531e30417e0468701ccfb94c58fe7a2380ff7

Observation a818af2c-e3fb-4f91-91c5-c9471079b479 · outbound

This paper cites Deep hough voting for 3d object detection in point clouds.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Deep hough voting for 3d object detection in point clouds

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.542229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.473978Z digest=sha256:9d05ad2eb63f223caa74dd527c08be5e3bc65a4a7250945b9d805e454fc41150

Observation 8387d02a-0b29-4b4a-88af-94d117592b61 · outbound

This paper cites Pointnet++: Deep hierarchical feature learning on point sets in a metric space.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Pointnet++: Deep hierarchical feature learning on point sets in a metric space

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.525560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.478744Z digest=sha256:41126f2f08e5caf8998aae0f1e06b1c7c515d8aa4cee69158282cff2bc440917

Observation 0d4b4a55-8d99-4297-a8f9-931ded1ef39a · outbound

This paper cites X-refseg3d: Enhancing referring 3d instance segmentation via structured cross-modal graph neural networks.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation X-refseg3d: Enhancing referring 3d instance segmentation via structured cross-modal graph neural networks

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.507677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.483640Z digest=sha256:bc303468786d6cb8aa5980cf8e0aff6eaadf5f148e2cf85aaf5bfe36ceade5d1

Observation 338825f4-7d69-4a2c-aa16-491c5f84b807 · outbound

This paper cites Learning transferable visual models from natural language supervision.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Learning transferable visual models from natural language supervision

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.488477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.488477Z digest=sha256:5e728e243510b4653c42d700e355e90ac6d04d326f1703e472f86d0bd1f59cac

Observation 3e890971-ef3d-4bd0-bad6-5816dabfeefc · outbound

This paper cites You only look once: Unified, real-time object detection.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation You only look once: Unified, real-time object detection

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.493295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.493295Z digest=sha256:e2917ec9663942f9cdbd4967d1e0806a8c48d1a54f223a4797d514b5208ba945

Observation cf4a6b87-c014-4444-8323-8ef14e958b4f · outbound

This paper cites Mask3D: Mask Transformer for 3D Semantic Instance Segmentation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Mask3D: Mask Transformer for 3D Semantic Instance Segmentation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.498039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.498039Z digest=sha256:75a851a04ecfb6d35612975c933114c70e4fb7f88cf904e23815d5bbfe58bfd3

Observation 444bc12b-700f-4626-8c0f-217496f9a719 · outbound

This paper cites Mpnet: Masked and permuted pre-training for language understanding.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Mpnet: Masked and permuted pre-training for language understanding

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.469640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.503343Z digest=sha256:501591f61a2e8d3af466a31490001ff7882dc7ad5c8d8ace5f55bd1544508ae3

Observation 5b8c2274-ed08-4afc-bd08-7f279f839f40 · outbound

This paper cites ReCLIP: A strong zero-shot baseline for referring expression comprehension.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation ReCLIP: A strong zero-shot baseline for referring expression comprehension

Reference 59

Resolution
verified exact
doi, observed 2026-08-11T23:35:20.731696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.508347Z digest=sha256:7b77ddb19a7d292392a856a478003443a31963d2565d2b9c473131f96fc266d7

Observation 24502735-7495-43bd-a1bc-ad5f5dce8b80 · outbound

This paper cites Superpoint transformer for 3d scene instance segmentation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Superpoint transformer for 3d scene instance segmentation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.453210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.513161Z digest=sha256:cce5881ec8cd950ea17674847501587dad23737e7d15c885015c5879871d6979

Observation 6da5567f-2150-4860-bd4c-c5646015dad6 · outbound

This paper cites Text augmented spatial aware zero-shot referring image segmentation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Text augmented spatial aware zero-shot referring image segmentation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.517871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.517871Z digest=sha256:4b63f2fcc955d53d157082d12503fd2f40d06739bae3225328cfd836762af91a

Observation 476a9e87-7043-404f-ac7e-9bb720dfc74d · outbound

This paper cites Interpretable Counting for Visual Question Answering.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Interpretable Counting for Visual Question Answering

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:35:20.887589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.522490Z digest=sha256:a7f1e6880cb65b0b3ed0b7248d89c9a9035b82f16d50f269180e455652c3d5d1

Observation 0740510a-5569-468f-935b-f2a0b1e7cb39 · outbound

This paper cites Attention is all you need.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Attention is all you need

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.527617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.527617Z digest=sha256:309684369d5a2cedc2e88eff5f05927462fc9d03a55039453ba4dc07b0297712

Observation 613c5d7b-3253-4f5b-8695-fbcf3c40f772 · outbound

This paper cites Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.532072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.532072Z digest=sha256:d69feadb9d4f14847e9244ec0a9d43319a9280dc92f008aa91a0ffbc50f21c53

Observation d7df6c1e-4204-4312-8ea8-84228356509b · outbound

This paper cites 3D-STMN: Dependency-Driven Superpoint-Text Matching Network for End-to-End 3D Referring Expression Segmentation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation 3D-STMN: Dependency-Driven Superpoint-Text Matching Network for End-to-End 3D Referring Expression Segmentation

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:35:20.851094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.536982Z digest=sha256:0be941a82624e751a01caef4e3637dfbb689b62e96660e1d68e66498d0ebadba

Observation ede7de22-f7a1-4904-9b6b-ac5363ee29b4 · outbound

This paper cites 3D-GRES: Generalized 3D Referring Expression Segmentation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation 3D-GRES: Generalized 3D Referring Expression Segmentation

Reference 66

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:35:20.829312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.542038Z digest=sha256:1df03cdee866cfd80817690e6c79e9fc5a9d2cdd49fc068942a42c51ffc79869

Observation c75f2d21-f628-43ba-a8a8-52e06e9695b2 · outbound

This paper cites Rethinking and improving relative position encoding for vision transformer.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Rethinking and improving relative position encoding for vision transformer

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.426059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.546980Z digest=sha256:bcb5707952dca601983675e7aab2c783222bc39709ff6ab20d04ca0a39500e47

Observation 5cc55593-1a95-40cc-ba82-8f53107d540f · outbound

This paper cites Towards Semantic Equivalence of Tokenization in Multimodal LLM.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Towards Semantic Equivalence of Tokenization in Multimodal LLM

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.551468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.551468Z digest=sha256:8bdf845b7de796735d2b08cf70a93165d35611a7d57eacef0610b2d5f035b932

Observation ad106fa6-16d3-4087-9cb3-533c3fe885a2 · outbound

This paper cites Next-gpt: Any-to-any multimodal llm.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Next-gpt: Any-to-any multimodal llm

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.408387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.556187Z digest=sha256:5a4a8adaba2fdbc8c401202797edc013bce7bcd611113394e8242768272f78a9

Observation 9eff1af7-21a2-49c2-8c9e-546f7c1ce603 · outbound

This paper cites Eda: Explicit text-decoupling and dense alignment for 3d visual grounding.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Eda: Explicit text-decoupling and dense alignment for 3d visual grounding

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.393168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.560544Z digest=sha256:5586f5ab94756366dc5748a506238ecf12bc8a00a63c4b17e8b2b58599a03067

Observation 059ce938-98dd-4ae4-95f1-973f9f4bfa34 · outbound

This paper cites A Unified Framework for 3D Scene Understanding.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation A Unified Framework for 3D Scene Understanding

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:35:20.790252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.565037Z digest=sha256:38a334e3ba7ad47bc9bbcb8d68cee01a2ad29ccf15d9b333bce77f22b657efac

Observation 438d85db-fb73-4b6b-9456-8d093fe80274 · outbound

This paper cites Sat: 2d semantics assisted training for 3d visual grounding.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Sat: 2d semantics assisted training for 3d visual grounding

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.378358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.569392Z digest=sha256:6a8f1e71d9ff657ca53d9f85d60b70f956d9f5e4a2e00da99634a9871021fdde

Observation 7e80e01f-b43f-476c-bce1-a5182a3aa5ab · outbound

This paper cites In- stancerefer: Cooperative holistic understanding for visual grounding on point clouds through instance multi-level contextual referring.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation In- stancerefer: Cooperative holistic understanding for visual grounding on point clouds through instance multi-level contextual referring

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.573330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.573330Z digest=sha256:e39ab30966d26a14b21c731a8c0aee66a430cb17e1e4ec69dd4da8e5add8eaa1

Observation d33a0356-a7ef-4a71-a945-043ce0987598 · outbound

This paper cites CK-transformer: Commonsense knowledge enhanced transformers for referring expression comprehension.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation CK-transformer: Commonsense knowledge enhanced transformers for referring expression comprehension

Reference 74

Resolution
verified exact
doi, observed 2026-08-11T23:35:20.703936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.577618Z digest=sha256:46edc026d0da22c5d1afea52e82deeb17cee0db4163206780bd6cba403c327bc

Observation 86a2e0f9-8981-4c1e-b7df-990ee68c6a58 · outbound

This paper cites 3dvg-transformer: Relation modeling for visual grounding on point clouds.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation 3dvg-transformer: Relation modeling for visual grounding on point clouds

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.581722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.581722Z digest=sha256:1a3d724b30dbaafea5674aae6927c46fa8aa80bdf2628a3a7b894db1bdcfed79

Observation 0fdcb3aa-f801-4fb5-91cf-7c23597f5216 · outbound

This paper cites Towards Learning a Generalist Model for Embodied Navigation.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Towards Learning a Generalist Model for Embodied Navigation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T23:35:20.585860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:35:20.585860Z digest=sha256:f3a19f378d67d8439e430f743e6115c989d10da3b5baf2fffda1a12846772314

Observation d54b8337-8471-4c91-a165-0ab361378200 · outbound

This paper cites left”, “right.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation left”, “right

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.342961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.590132Z digest=sha256:68014b27cd37f87948ec29e3ebb1e5a8f17f8c27d7b298417c37998e597c6aee

Observation 4cc595ef-4e5c-4217-a288-46daf24a8e36 · outbound

This paper cites Guidelines: • The answer NA means that the abstract and introduction do not include the claims made in the paper.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that the abstract and introduction do not include the claims made in the paper

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.326917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.595708Z digest=sha256:8c332d70ccf5234fa123c2d9656870538c61e9a73393a7d7fff50ebafffe278b

Observation de3711d3-b46c-4c83-b6bf-241876dd6d00 · outbound

This paper cites Limitations.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Limitations

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.311479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.600631Z digest=sha256:3d7b51d28505dc0dcbc78a6cf00cbed092197de983a22d35b8a29057b31a4d4c

Observation 59c0f4a0-4694-41b7-b8c4-15d84606ef57 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include theoretical results.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that the paper does not include theoretical results

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.295825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.605602Z digest=sha256:b35a975f161252df0003d7bf780b26c59e59a77b613ca5fd133f526719c2a106

Observation d6c2dc30-f5d8-488c-ae09-38aad1bbf276 · outbound

This paper cites Detailed experimental settings are provided in Sec.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Detailed experimental settings are provided in Sec

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.279306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.610880Z digest=sha256:8456b5c5f6143df06816ae43d0bc5c29f14a15ba1884c892e596f1b85db82270

Observation 9c7ff5a0-ce16-4925-b393-4d7f88552dbf · outbound

This paper cites Guidelines: • The answer NA means that paper does not include experiments requiring code.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that paper does not include experiments requiring code

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.263315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.616330Z digest=sha256:09f6bb64fb898d164f8daaa7d68e733859590a28251736697b39d9e0cbe558df

Observation d7a9b8d7-2e54-414b-942f-c7960e97ee15 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that the paper does not include experiments

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.247433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.621089Z digest=sha256:fb3bd4469f9a62599eb9fc9adb9b9ca747bfc69c5588c369ee88c3782c2bf116

Observation b486fab9-93d7-4c22-9610-d7ae331c1cc3 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that the paper does not include experiments

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.231667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.626328Z digest=sha256:802946a2a3ca1fa6a37ee8d2f69d7821e7ae2429c456fbae2487e97f8c12e130

Observation 93ee714e-964b-475e-9f53-dbaa90981f14 · outbound

This paper cites 4 and Appendix.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation 4 and Appendix

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.216461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.630884Z digest=sha256:b8760273c1d7e49a298baf4bb6b2dcf3641cc44e9017786cec92d3aada14f799

Observation ecc6d0c8-f297-42da-b618-b0ccba2d105e · outbound

This paper cites Guidelines: • The answer NA means that the authors have not reviewed the NeurIPS Code of Ethics.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that the authors have not reviewed the NeurIPS Code of Ethics

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.200530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.636245Z digest=sha256:739b7d6b22233026d62e191e72d9d3fbb0aa51d0d375659d6202cc162980e02e

Observation 91270bbf-d720-496c-b11d-6b9d2c9ed8fa · outbound

This paper cites Guidelines: • The answer NA means that there is no societal impact of the work performed.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that there is no societal impact of the work performed

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.184463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.640947Z digest=sha256:3dd9925664f2ecbf6ced3d51338eefe0d20e58c2cca493c0027c5b8258438c96

Observation 5ffc036e-e8eb-412e-a69b-100558a344ae · outbound

This paper cites Guidelines: • The answer NA means that the paper poses no such risks.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that the paper poses no such risks

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.167682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.645494Z digest=sha256:3727c910c31bd1304a600627535b11b77fce45e3c9899f872750fe121be281ad

Observation 1e1d8fec-e327-4ffe-ad79-99b3cab1b7fa · outbound

This paper cites Guidelines: • The answer NA means that the paper does not use existing assets.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that the paper does not use existing assets

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.152549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.650172Z digest=sha256:f69a994e9400bf54f8a320758642b69bd20b8c868225ddcbd73bd41b0d464886

Observation 782fdb6f-179a-49c2-9882-02215be6d888 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not release new assets.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that the paper does not release new assets

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.138405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.654903Z digest=sha256:65cdad319847c4459980d8b03ae13957640518bc63aeac12c7b41b20fb4b15d5

Observation 3245ba21-ac48-4dea-8810-0c24557492e3 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.122814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.659462Z digest=sha256:184cf7bc5b38eb651d3f703eac800a45b6b6ca80a4e8eeed89ee744bcb9e4ca3

Observation 51bc7f0f-8e6b-4b00-86cb-ec6946077897 · outbound

This paper cites Guidelines: 27 • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects.

RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation Guidelines: 27 • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:35:21.107264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:35:20.663843Z digest=sha256:cef8db2e176f06c3a266a8a1c957ad0ec7bda8c7643b8a910171529b0661cc29

Pith citing papers

No inbound Pith citation observations are available.