Pith. sign in

Paper Citation Record · LEDGER

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation

As of 20 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2506.16058.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16058 v2

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:50:23.565560Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

66 of 66 outbound references displayed

  • verified exact3
  • verified fuzzy43
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a97559ff-9004-4296-ab05-ec06c700efd3 · outbound

This paper cites Self-calibrated clip for training-free open-vocabulary segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Self-calibrated clip for training-free open-vocabulary segmentation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.250219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.250219Z digest=sha256:e6c361bbcee290fc565c5219a95eeccbdf990db1d8a4988202196e9629190fe8

Observation 8ddce7f4-5588-41a2-83f5-52828cd4f189 · outbound

This paper cites UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.255593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.255593Z digest=sha256:37c1108f8d36413f700770403a3d00a39db0472b309a0d9a6bf89d85dfaff5a8

Observation 97e86502-505d-4e63-a351-faf64d93292e · outbound

This paper cites Brostow, Jamie Shotton, Julien Fauqueur, and Roberto Cipolla.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Brostow, Jamie Shotton, Julien Fauqueur, and Roberto Cipolla

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.726791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.261190Z digest=sha256:95f9eb6fa6a3d5911fa712b02c06ce036d271f9c7b7fc3ddd7f3e3a597c4a5d8

Observation 668c0a39-61f2-4d63-b2c7-837a4f063b19 · outbound

This paper cites Zero-shot semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Zero-shot semantic segmentation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.711443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.266818Z digest=sha256:d2915b4a59247f9b1a02d62f78236dfc40e4972a492d10028218f5a7d91c26fc

Observation c50bfe99-7026-4e63-8648-f42be4d18a4a · outbound

This paper cites End-to- end object detection with transformers.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation End-to- end object detection with transformers

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.696514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.271584Z digest=sha256:ea65671f313a6330e7caf3152b85e3ce57ed1a0fffd5c12d88038c8c58b3e697

Observation e38a447f-63b6-4c93-95d4-88c76ddd1afe · outbound

This paper cites an unresolved cited work.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:50:24.681255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.276968Z digest=sha256:3c8287e76de58847d7d32cca8d83e068392a13ee6e95831cbb500baaf63f6d3d

Observation 97cb7be6-a06c-463a-8359-e2eb0820f02c · outbound

This paper cites UNITER: UNiversal Image-TExt Representation Learning.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation UNITER: UNiversal Image-TExt Representation Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.281943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.281943Z digest=sha256:133ea38f1992d754f3fbd521d76659b6976fc7cd6142919ddf9363533d7c85c9

Observation cab951f1-9f45-4213-978b-f68eea78264e · outbound

This paper cites Mask2Former for Video Instance Segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Mask2Former for Video Instance Segmentation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.286782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.286782Z digest=sha256:172dedfa3c512d30d2cc84e033b4f7112ba2b2120bec5a952230db44c7f61749

Observation c5b765f5-6d38-48ca-89e7-f31e3eecfc23 · outbound

This paper cites Schwing, Alexan- der Kirillov, and Rohit Girdhar.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Schwing, Alexan- der Kirillov, and Rohit Girdhar

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.292230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.292230Z digest=sha256:ccca4777078e01b5bd9ec77830773c0356ef3fea3cce60887238a542c545af4f

Observation d3f43add-dd73-459f-9379-39f466b93474 · outbound

This paper cites Cat-seg: Cost aggregation for open-vocabulary semantic segmenta- tion.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Cat-seg: Cost aggregation for open-vocabulary semantic segmenta- tion

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.656705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.297117Z digest=sha256:2724f98df1ed8ed0e6370eb928286580b59a3f5a08926b0f45212cf6e6a83a90

Observation 4723c5f7-8b26-4713-a00e-80203d198265 · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation The cityscapes dataset for semantic urban scene understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.301829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.301829Z digest=sha256:56b9f042e8eb36e6b5d252474219b9e0f9906d4592a8c51a2316f5a09a6760ce

Observation e4678ec7-af9e-4dda-ac7b-b7463e10667d · outbound

This paper cites De- coupling zero-shot semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation De- coupling zero-shot semantic segmentation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.630931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.307089Z digest=sha256:86318b3cc30ec22f17ead0041d6624e74a176b2be9e3059d4c9cd6b0846de9e3

Observation 9f2bc313-d6c4-40a1-8738-f963b9a62445 · outbound

This paper cites The pascal visual object classes challenge: A retrospective.IJCV, 111:98–136, 2015.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation The pascal visual object classes challenge: A retrospective.IJCV, 111:98–136, 2015

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.615216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.311413Z digest=sha256:3396331aa7b5ce8c7a0cd4f45211c3230bb51eff1d7d61add081dc693badd74d

Observation 0d1d0ab4-f453-41f8-a227-bfef21cb9f24 · outbound

This paper cites Large-scale unsu- pervised semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Large-scale unsu- pervised semantic segmentation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.600200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.316014Z digest=sha256:935bf0056a119067c8cf5c6d73f601b1ef3e6ac4644caac343cb4b522d0ab9f5

Observation edc5a565-09b0-4bc9-8784-dc05196a8f99 · outbound

This paper cites Scaling Open-Vocabulary Image Segmentation with Image-Level Labels.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Scaling Open-Vocabulary Image Segmentation with Image-Level Labels

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.320589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.320589Z digest=sha256:c0feecdcf559e4c86466814a29ee3b2f1158690aa2c70875ed2f2624c3bc9602

Observation 9e39c27a-a7ff-4470-9077-2fd5cb6a8031 · outbound

This paper cites Random walks for image segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Random walks for image segmentation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.584747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.325420Z digest=sha256:a9e5beef2d893076d6574bf84d3a882ad548c5c6f3f16df6e4eeca62bdd7a683

Observation 39bb6abf-29bb-4c4f-8712-ea9fb4931231 · outbound

This paper cites Global knowledge calibration for fast open-vocabulary segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Global knowledge calibration for fast open-vocabulary segmentation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.567916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.330380Z digest=sha256:efb064598a43573d31d3228c242eccfb6d09ebfd6ca29bc12c2bbf75993317a4

Observation 25db6e78-78b1-40f8-b904-20340d80d6c3 · outbound

This paper cites Primitive gener- ation and semantic-related alignment for universal zero-shot segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Primitive gener- ation and semantic-related alignment for universal zero-shot segmentation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.552376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.335212Z digest=sha256:8949e7e6ee55baad28140f35ede3627e88321d09fbbb19b0842fc1f42e432fec

Observation b676ed18-8618-4d5d-a61e-2ef7499ba897 · outbound

This paper cites Planning-oriented autonomous driving.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Planning-oriented autonomous driving

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.340138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.340138Z digest=sha256:fa9571ce3dbc737a56fde6a6e3a67335ce2824fc149385dc702833632759b85d

Observation 3268e0aa-7e79-441c-93b2-a1f7b2c8fd6c · outbound

This paper cites Densely Connected Parameter-Efficient Tuning for Referring Image Segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Densely Connected Parameter-Efficient Tuning for Referring Image Segmentation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.344365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.344365Z digest=sha256:6466c8f33088ac20ff77bac63b71ae17ace965039d2efe1f683ddc38c64dcbb4

Observation de7f347b-2b56-4e45-98e4-5c4184e9f8af · outbound

This paper cites Segment and Caption Anything.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Segment and Caption Anything

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:50:23.795324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.348944Z digest=sha256:1fa4d615bcd4b87031f62f5ef2ed9d17a793111a07275d3709100fa1c4df49cc

Observation cc55fbdc-facc-4b62-86d5-2b7d9f022012 · outbound

This paper cites Proxydet: Synthesizing proxy novel classes via classwise mixup for open-vocabulary object detection.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Proxydet: Synthesizing proxy novel classes via classwise mixup for open-vocabulary object detection

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.525350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.354116Z digest=sha256:b0a9515a6cdb4da831baf570ae34462c1948b35fbf6257d84805fcf2f3d281ed

Observation 98fa837c-4a11-4694-8f06-5a56b65b9df6 · outbound

This paper cites Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.511326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.359615Z digest=sha256:78cd090d22b1fb45b0dae89f28df00b897715ad459fe0f56f44abd6cc9900825

Observation d147be5f-999c-4b9c-a8d1-e16f67e5a740 · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Scaling up visual and vision-language representation learning with noisy text supervision

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.495762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.364378Z digest=sha256:27738a2b413a316424246c259f1bd0025a564940430ad5fbda49b6451b75f593

Observation 8e05eb0c-bd88-40b9-a4bf-6d017e7c101b · outbound

This paper cites Learning Mask-aware CLIP Representations for Zero-Shot Segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Learning Mask-aware CLIP Representations for Zero-Shot Segmentation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.369382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.369382Z digest=sha256:c7c105e297f0b7a3ceb047515d276ce39d34e6ac0634c9317f3cd80fdd559beb

Observation 1fc46008-179d-4f55-97b9-d16c91fbdb51 · outbound

This paper cites Collaborative vision-text rep- resentation optimizing for open-vocabulary segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Collaborative vision-text rep- resentation optimizing for open-vocabulary segmentation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.479936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.374458Z digest=sha256:2f0de90b80b7e8518a15aea65c8a2c677673fb7400fd4f1d734ea2135fb74297

Observation 498432c4-95c4-45e0-855e-1e3a1fe63fd2 · outbound

This paper cites Weinberger, Serge J.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Weinberger, Serge J

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.464721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.379399Z digest=sha256:12c7cdb760f1a96d783d420dc46103f18007c90fa8c9afb721a25438d1fc1570

Observation 0a889458-a699-4ddb-8b2b-e8528c40033f · outbound

This paper cites Unicoder-vl: A universal encoder for vision and lan- guage by cross-modal pre-training.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Unicoder-vl: A universal encoder for vision and lan- guage by cross-modal pre-training

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.448213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.384236Z digest=sha256:903e10337eae271ac89a407c64ca4d8246b0986854c34f1e764fa691828b08e0

Observation 0ebd4417-b1aa-4019-927d-55a82dc79e9f · outbound

This paper cites Ordinalclip: Learning rank prompts for language-guided ordinal regression.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Ordinalclip: Learning rank prompts for language-guided ordinal regression

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.432092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.388956Z digest=sha256:9b3c6cc43e2262e7596200996a80e1163ac6e01f01a6b73ccd4f2101aebff87b

Observation c811938a-aac7-4ac5-8e09-acf2ef80452c · outbound

This paper cites Oscar: Object-semantics aligned pre-training for vision-language tasks.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Oscar: Object-semantics aligned pre-training for vision-language tasks

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.416066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.393647Z digest=sha256:28ce30b247dc7b646d1a437d50da1687556a2fecb4f3837a28207f5f025c6949

Observation 68a2fe49-d63e-45c2-bf79-9736b651e201 · outbound

This paper cites Open-Vocabulary Semantic Segmentation with Mask-adapted CLIP.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Open-Vocabulary Semantic Segmentation with Mask-adapted CLIP

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.398271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.398271Z digest=sha256:b451c15b73a7e0eb4e8c6504c925c419dfd7c3cd9df3ed0778c28b6d4dad2d41

Observation fd7982dc-77c4-4d1e-ab7a-7ab4b057b9c1 · outbound

This paper cites Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll ´ar, and C.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll ´ar, and C

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.401125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.402939Z digest=sha256:868247c585f9d11a131ac34483a6a961814ed750674f95b3095a0355cfe6bfe5

Observation 6efe546e-7a4c-4061-8f09-b8ab3b5e83b7 · outbound

This paper cites Quality- aware and selective prior enhancement memory network for video object segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Quality- aware and selective prior enhancement memory network for video object segmentation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.385247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.407360Z digest=sha256:70acb71a295043a30ad05105ab8c541e3b10d0188cb957f02aee033aa84e67e7

Observation 66b8b1e0-6604-4c24-998d-108842b10359 · outbound

This paper cites Global spectral filter memory network for video object segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Global spectral filter memory network for video object segmentation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.370046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.411715Z digest=sha256:915f26db02a7fc1ebf1cc7d215fcce54fa8fa4779392f1f09e6f893e586faf3a

Observation 05c253fb-20d2-474e-b290-67a9a5a52fb9 · outbound

This paper cites Learning quality-aware dynamic memory for video object segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Learning quality-aware dynamic memory for video object segmentation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.355059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.416391Z digest=sha256:92f6524bf764915ce4a1772f24ce63341bf0a29365afe57e4567907c0449f46c

Observation 4568aa7d-2900-49b7-a87c-311f3614fb37 · outbound

This paper cites Universal Segmentation at Arbitrary Granularity with Language Instruction.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Universal Segmentation at Arbitrary Granularity with Language Instruction

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.420753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.420753Z digest=sha256:1e7f908a5b4d21199ae622bc11c77e6303acdfbd2f2654aad8ebc47115cbafea

Observation baedb819-8184-4044-9af5-e317f128b7de · outbound

This paper cites Open-vocabulary segmentation with semantic-assisted calibration.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Open-vocabulary segmentation with semantic-assisted calibration

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.340256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.426153Z digest=sha256:044dbbc1f6fa1314c6453e73bbe075df8abe81f6112ea70d1b27963482b7d585

Observation 0b47705e-28cb-468b-a303-fe33b669b1b1 · outbound

This paper cites Learning high-quality dynamic memory for video object segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Learning high-quality dynamic memory for video object segmentation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.324665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.430987Z digest=sha256:1659761cd5591d278f89414f867f4b779477f822d902bf8e085d26393d77da7d

Observation 4358f59c-72f4-4cdf-905f-4afcf8fea044 · outbound

This paper cites ThinkBot: Embodied Instruction Following with Thought Chain Reasoning.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation ThinkBot: Embodied Instruction Following with Thought Chain Reasoning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.435750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.435750Z digest=sha256:c51eb27c0ba10f0748b0428cd7d4b852e07c70ada0291519a009a002293266c9

Observation 1b876d93-17fc-45af-9517-fa78d3341338 · outbound

This paper cites Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.310125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.440778Z digest=sha256:5932684af85d9e5030f479f89b961059384a9a0cef472d7a0e37bec5659e1b67

Observation 08c7c79b-2c56-4764-b50f-6073c683c43e · outbound

This paper cites SOC: Semantic-Assisted Object Cluster for Referring Video Object Segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation SOC: Semantic-Assisted Object Cluster for Referring Video Object Segmentation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.445973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.445973Z digest=sha256:07ac82e6a0aac83de575f3474635904b0c42ff3c3dbbf9213a891cde394e9fb9

Observation 3ba8fd79-3fbd-4731-bda1-b25a15a2fbd1 · outbound

This paper cites CoHD: A Counting-Aware Hierarchical Decoding Framework for Generalized Referring Expression Segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation CoHD: A Counting-Aware Hierarchical Decoding Framework for Generalized Referring Expression Segmentation

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:50:23.696452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.450624Z digest=sha256:4fb6590f4ac69fb0c62c3e07ab8131c87970c5870ed5be9c05f3ea12fe327e76

Observation 881002cc-da08-4e94-9b36-86b82af250f6 · outbound

This paper cites Matrix analysis and applied linear algebra.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Matrix analysis and applied linear algebra

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.294036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.456174Z digest=sha256:c10081e96edbaaa7a42d562ad083e5c7b2118cddddab0b992c6388701b8d80e4

Observation 92649031-db64-4c2c-9ab0-735192d697d6 · outbound

This paper cites an unresolved cited work.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:50:24.279433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.460868Z digest=sha256:d0c5b5a85e6160a9ce942f757ba8cfaa353c3333e3091ac12e7ec676f1d72362

Observation 72488162-0515-4ac4-886e-9cf16112dca5 · outbound

This paper cites Siri: A simple selective retraining mechanism for transformer-based visual grounding.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Siri: A simple selective retraining mechanism for transformer-based visual grounding

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.264057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.465327Z digest=sha256:91a3c0beb7ba71d53b1b1e2d33f4b7e1b7dcab6215fe9a684c907f016b217749

Observation f6eb4e40-6224-4237-b50d-b3d76dde49b8 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Learning transferable visual models from natural language supervision

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.248952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.470029Z digest=sha256:4355959461c8fe48a4216b3b31e6e23b1aa3df62c5b46828bca11d783b254639

Observation ccf2a863-c91f-49d6-a381-782b87cdf801 · outbound

This paper cites Hierarchical Memory for Long Video QA.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Hierarchical Memory for Long Video QA

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.474555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.474555Z digest=sha256:0d5568077e38d64a2151c3655c4e1e8a6b601ab104506e22a188c5266235a84a

Observation 191b1b35-0da6-4e33-8a0d-247b3a060ea4 · outbound

This paper cites Uni-adafocus: Spatial- temporal dynamic computation for video recognition.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Uni-adafocus: Spatial- temporal dynamic computation for video recognition

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.233510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.479341Z digest=sha256:553e14aa5d3c54db522d9886187a4deba27603239ee39ad9af4b3893e20e1106

Observation 8be1fb6b-ddb6-4466-92f6-3e7cc8b144e0 · outbound

This paper cites Iterprime: Zero-shot referring image segmen- tation with iterative grad-cam refinement and primary word emphasis.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Iterprime: Zero-shot referring image segmen- tation with iterative grad-cam refinement and primary word emphasis

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.218075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.484430Z digest=sha256:eb74594db4b123cf86d5445c3bafa153351568b34f3a37e5eeda8f269e13456f

Observation cb973aa9-23da-4116-afe2-42bef671ea2b · outbound

This paper cites Sam2-love: Segment anything model 2 in language- aided audio-visual scenes.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Sam2-love: Segment anything model 2 in language- aided audio-visual scenes

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.202554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.489538Z digest=sha256:4cc934478efde1b43722290413226136dd03450cd74f481b024842ee7a55d33e

Observation 494fb9ba-fec2-4b55-8b91-96c449cbf13d · outbound

This paper cites HyperSeg: Towards Universal Visual Segmentation with Large Language Model.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation HyperSeg: Towards Universal Visual Segmentation with Large Language Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.494720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.494720Z digest=sha256:9830cf61182ee63b429ca6b4de8bef591bac1dbe6f6e73925597ef3cac446df2

Observation 0cb8fa4d-6035-476f-ae78-1a5d6d762ba8 · outbound

This paper cites InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.499385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.499385Z digest=sha256:f69052537bc257542226137adeb368d3ad6b918cceb95e36782af1544e5c6c81

Observation 6c4d208a-bab4-40b6-90d7-116debeb9aed · outbound

This paper cites A large-scale benchmark for food im- age segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation A large-scale benchmark for food im- age segmentation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.187846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.504013Z digest=sha256:d3251fa42dea0c6712000bfc63019b22eda30ff0e09dfb7ae732157c8c4191e7

Observation e17d219c-1cdf-4ab9-a8b5-4882523f970d · outbound

This paper cites Detectron2.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Detectron2

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.170493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.508334Z digest=sha256:65c3563586f9384f04a588bb39ccf6e6c3ed24011f2802ac7b7a43071cb00d3b

Observation 6f439054-5bbb-428e-959b-d363152575d9 · outbound

This paper cites Semantic projection network for zero- and few-label semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Semantic projection network for zero- and few-label semantic segmentation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.155465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.512584Z digest=sha256:0e321308d4d664c426f9db0b8778df5b447af1a2937a148fdaf883fe63da7dca

Observation 8541a39c-81cf-4695-a0d3-dce7b50db631 · outbound

This paper cites Bridging the gap: A unified video comprehension framework for moment retrieval and highlight detection.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Bridging the gap: A unified video comprehension framework for moment retrieval and highlight detection

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.140251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.516744Z digest=sha256:671ddb999d3a7b5be17558022f473eb6520334131e1e9e02499626690a1a1011

Observation a8f54fe3-8ea2-4cab-997b-ab0e6d579725 · outbound

This paper cites Sed: A simple encoder-decoder for open- vocabulary semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Sed: A simple encoder-decoder for open- vocabulary semantic segmentation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.124526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.521827Z digest=sha256:2c4c570feee297ae58a335a1aae8519351f1dcc23e899c267012e5785e703b93

Observation fe3547b4-b4dd-4f37-83c1-97cfd2ad475c · outbound

This paper cites Alvarez, and Ping Luo.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Alvarez, and Ping Luo

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.109620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.526934Z digest=sha256:118527717325e0b647d745803d8b244fda5805ca33c0e2e7f0e0353ed8378529

Observation fe505afc-f573-406b-8518-9bbc306c6719 · outbound

This paper cites Open-vocabulary panop- tic segmentation with text-to-image diffusion models.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Open-vocabulary panop- tic segmentation with text-to-image diffusion models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.094278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.531413Z digest=sha256:cd9b444418a639f0a73d7fd4c0324d51040be46372f1fbfe3a12a660f6aac0e5

Observation 17813763-b4c7-4cc1-a2ea-c59571aa80ca · outbound

This paper cites A Simple Baseline for Open-Vocabulary Semantic Segmentation with Pre-trained Vision-language Model.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation A Simple Baseline for Open-Vocabulary Semantic Segmentation with Pre-trained Vision-language Model

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:50:23.626640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.536291Z digest=sha256:5bb96f728b9e871bfb1a5b7657a3730545cd47ce1d90854d3bab9b5f11972145

Observation cd68d2cc-bf5d-4dde-bcd9-7e618daee1d8 · outbound

This paper cites Side adapter network for open-vocabulary semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Side adapter network for open-vocabulary semantic segmentation

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.077918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.541877Z digest=sha256:b5f161dc9f3667102356bb291eb397c1ace6506a1f13d65af411268406e25b00

Observation d0e369b9-683d-46a4-8293-042169a04220 · outbound

This paper cites Masq- clip for open-vocabulary universal image segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Masq- clip for open-vocabulary universal image segmentation

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.061442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.546341Z digest=sha256:2dc1e04b07194e97eaef0c6dedd1de4b81a7034942b667324b913bed44ff2a42

Observation bd2b2bcc-c07e-4916-84ae-123c09abd7b5 · outbound

This paper cites Convolutions die hard: Open-vocabulary seg- mentation with single frozen convolutional clip.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Convolutions die hard: Open-vocabulary seg- mentation with single frozen convolutional clip

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.046307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.550790Z digest=sha256:d962cdce1097f71de2df54827afe27e766a5e47b470c86dce3485821bf2c757a

Observation 0e8d8dd4-b8d1-4016-b6c3-8bb89b40456c · outbound

This paper cites Prototypical matching and open set rejection for zero-shot semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Prototypical matching and open set rejection for zero-shot semantic segmentation

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.024982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.555457Z digest=sha256:5f04729ba3ee94606d238230b54fae86e7d1f5d819810addf44f614a98b5a577

Observation fe2be44c-1e16-44a6-ad2d-f8e543fb0667 · outbound

This paper cites Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.560907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.560907Z digest=sha256:27c256a3cf2503019a9b2ccd4afa417688b877b7f40c09ef544a9bcaf150ddb2

Observation d2d361ab-ea73-4d7a-9d59-648b450a27b6 · outbound

This paper cites Scene parsing through ADE20K dataset.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Scene parsing through ADE20K dataset

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.002543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T23:50:23.565560Z digest=sha256:1e9e2b0fb945953843432711db98823cab20061523abba16f5d2421b961102ee

Pith citing papers

No inbound Pith citation observations are available.