Pith. sign in

Paper Citation Record · LEDGER

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning

As of 10 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2505.16974.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16974 v2

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:55:00.551518Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact1
  • verified fuzzy31
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation df149757-9719-42f6-bd58-513ab091fd15 · outbound

This paper cites Qwen2.5-VL Technical Report.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:55.240935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:55.240935Z digest=sha256:057d43eb3bf75cf072239f3fcf27d7f9140fdef31d7c8d054247f0e1bda285ee

Observation 4578038d-0ddb-43cf-9d97-397ef8b17960 · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Graph of thoughts: Solving elaborate problems with large language models

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:07.079856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:55.363794Z digest=sha256:6e12a2b80169c397c089b2e1f38fa5a5cb44383d39b8fb30e907de060fa4a2b5

Observation c2019963-a214-4127-b28d-8ac89486bb4b · outbound

This paper cites Zero-shot semantic segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Zero-shot semantic segmentation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:06.889923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:55.471678Z digest=sha256:a7302527ba32bbc2f91e3f7b6b0afcbb583c5c8fe71dba07ec35b05d3345a7de

Observation 40020426-045a-4894-9068-7e192cc80a6b · outbound

This paper cites Universeg: Universal medical image segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Universeg: Universal medical image segmentation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:06.752714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:55.632824Z digest=sha256:e7df9788d7e3dad7acb27d96471d8be71f77d5421a2fe46272225c739fba8766

Observation 49588e13-dea1-4ec5-ad0d-8333c24c4045 · outbound

This paper cites Coco-stuff: Thing and stuff classes in context.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Coco-stuff: Thing and stuff classes in context

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:06.536458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:55.725912Z digest=sha256:7284b89e2630eb1d81bc27fa0d8faf9a2f425bd674036ecf8810772d8d9356f8

Observation c43c090b-7aa8-44f6-92ac-1a41cc932c59 · outbound

This paper cites Open-vocabulary Panoptic Segmentation with Embedding Modulation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Open-vocabulary Panoptic Segmentation with Embedding Modulation

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:55:00.904928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:55.892210Z digest=sha256:00fef1d97d2e9902259bdd8143b431f2dd0a18f55783ebba70db44b110c06279

Observation b198457e-bddd-4661-9983-dad10af30503 · outbound

This paper cites Cat-seg: Cost aggregation for open-vocabulary semantic segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Cat-seg: Cost aggregation for open-vocabulary semantic segmentation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:06.385908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:56.030234Z digest=sha256:5e127c5e26dfa5701053111de82ebbcfba65ec1f3ac6936f96c1903f8166578d

Observation 0777d304-fb0f-4b12-97de-423a7a0ee9e0 · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning The cityscapes dataset for semantic urban scene understanding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:56.160013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:56.160013Z digest=sha256:44dda2072dcb94c1023edce2f92353767c31408ecb81ff32cc57721dc8c578dc

Observation 9cc0c596-ff2e-4e0c-995a-acd6b46f4722 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:56.327756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:56.327756Z digest=sha256:e9b25b9870768fdf9de59354cd431a038bd31186f8a57a486446fce08104146b

Observation aacaaa08-2934-415c-b5d3-e424452e28a6 · outbound

This paper cites Decoupling zero-shot semantic segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Decoupling zero-shot semantic segmentation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:06.166465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:56.447706Z digest=sha256:2fee623652199036061c7b260723524dcd11e0005ccd492fef1d3e50cbaa0404

Observation 83d48697-11f7-4868-a57c-3cf4d82be658 · outbound

This paper cites Open-Vocabulary Universal Image Segmentation with MaskCLIP.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Open-Vocabulary Universal Image Segmentation with MaskCLIP

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:56.581313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:56.581313Z digest=sha256:56493a5d96889d29473b9c401fa17089d925ca4dd3ba1512355d51510fcc6f2a

Observation 9cc51b88-a13b-40c0-a803-a88b04d905cb · outbound

This paper cites The pascal visual object classes (voc) challenge.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning The pascal visual object classes (voc) challenge

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:05.917178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:56.708615Z digest=sha256:b84d18200eb2b3054a080fbe86499efa68584d72deb70ad32972eadd608d03ee

Observation 5bd78bbe-6f87-441c-aa74-09563c718b7e · outbound

This paper cites Scaling open-vocabulary image segmentation with image-level labels.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Scaling open-vocabulary image segmentation with image-level labels

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:05.707881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:56.830080Z digest=sha256:3c1ee917b55c13873bca0e7796d3d707cac0e71acb37bf66a60b93a940676e55

Observation 164f693e-1b65-4fc7-b736-8e0dadaeb057 · outbound

This paper cites Zero-shot semantic segmentation with decoupled one-pass network.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Zero-shot semantic segmentation with decoupled one-pass network

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:05.548234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:56.932257Z digest=sha256:aba79f72a2ebb1e72d087af526a379fd174064d97a3030b19d278791f9805d25

Observation d139baad-6cbd-4e96-87c4-e70358865c18 · outbound

This paper cites Global knowledge calibration for fast open-vocabulary segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Global knowledge calibration for fast open-vocabulary segmentation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:05.221150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:57.039423Z digest=sha256:c1fbc8987f2f50667f6f78d93877adf5c570a468124a4f99a7b11b01633a4dbc

Observation 7e81d9c2-8bca-4c7f-84d3-e4114901a61a · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Scaling up visual and vision-language representation learning with noisy text supervision

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:04.972743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:57.163043Z digest=sha256:8e31d3e639b97cf5b2f4c4df8b889c8399df217071bac748de530dbdd82ffd67

Observation f5f74da9-0825-4969-988a-a1e57ad87eed · outbound

This paper cites Learning mask-aware clip representations for zero-shot segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Learning mask-aware clip representations for zero-shot segmentation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:04.750944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:57.307185Z digest=sha256:466936fb721803845377eadbe9e658f4fd238388cd6ab8e469ca5465a4997535

Observation 27036efb-9f51-4bc9-afa4-098bdee7fa1a · outbound

This paper cites Collaborative vision-text representation optimizing for open-vocabulary segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Collaborative vision-text representation optimizing for open-vocabulary segmentation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:04.621215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:57.419543Z digest=sha256:99f74dd4d065458a27df0aa8d32b6c6edd5955c58235d1e418a46dd16406daf4

Observation 574db350-6d7f-4060-8ecf-1d6ff4801ad9 · outbound

This paper cites Fineclip: Self-distilled region-based clip for better fine-grained understanding.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Fineclip: Self-distilled region-based clip for better fine-grained understanding

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:04.482616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:57.531488Z digest=sha256:4c9fc528e0b72d627123b32fe9ce2448c30d0666e2342e325c1d8f553b738d9f

Observation 8320a18e-e446-40ac-95b9-c2e7059435f7 · outbound

This paper cites Language-driven semantic segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Language-driven semantic segmentation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:04.388296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:57.647856Z digest=sha256:962b321196ef97380bf543ce0af4b0c14e5c06e1b9273f47545d5da6e848be21

Observation b4383c59-9070-4d5a-ad65-30260036bff8 · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:57.749830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:57.749830Z digest=sha256:8c62395e08d1e0b914a0ffa1b15a0d695aad11400b5772bde1bcc66ce2513900

Observation 6dc1f31a-8330-4627-89cf-e9498b80aeee · outbound

This paper cites Open-vocabulary semantic segmentation with mask-adapted clip.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Open-vocabulary semantic segmentation with mask-adapted clip

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:57.910363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:57.910363Z digest=sha256:4d3b54ba5a1680363c29ee4bc5a55df0a5104cd8448d6167656eb8eb11243947

Observation 6e3c758c-288f-4ebd-aff3-ecf00a8b8acf · outbound

This paper cites Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:58.058849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:58.058849Z digest=sha256:630286710d488837d36959e37c5ca6c45728506b20146e3eaa750bf719b6cb96

Observation a7ad0235-1ac3-4b39-ba25-af92070b4925 · outbound

This paper cites Remoteclip: A vision language foundation model for remote sensing.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Remoteclip: A vision language foundation model for remote sensing

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:58.183536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:58.183536Z digest=sha256:0e4106dadbde1670247a77fb0bbbcbb9b51c058bf03709307b75aff45123f039

Observation 612b4647-fc5a-4c37-ae34-c2c78ff30e6d · outbound

This paper cites Visual instruction tuning.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Visual instruction tuning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:58.265405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:58.265405Z digest=sha256:326cf74187b3306e49a11323adb3e04d9efa9199b9641f95822283558b83f236

Observation 3e3b7e24-d71e-4fa2-9d5c-c84289256291 · outbound

This paper cites Open-vocabulary segmentation with semantic-assisted calibration.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Open-vocabulary segmentation with semantic-assisted calibration

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:04.111055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:58.325734Z digest=sha256:9134c2b2a444b0871f8a9d788c358a664c2d9f11e2a7bc83cd74ebd992d8e880

Observation 828f60e5-2be0-455d-b510-ef587b841767 · outbound

This paper cites Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:58.395860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:58.395860Z digest=sha256:4d33b4208434cdf06c2764ca4d5fc0d5ff2a4e8c0f4730de52b2859643b6e5d7

Observation 49bfb041-e989-46cd-9089-36ccbf3c9196 · outbound

This paper cites The role of context for object detection and semantic segmentation in the wild.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning The role of context for object detection and semantic segmentation in the wild

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:04.002636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:58.464432Z digest=sha256:63e3dbe688e41e47f2d8dcb1a4286241c4b48f13ddf22a5ab0d2e824c4b1f33f

Observation 0fd97115-b573-4bf3-b7eb-eb594cca3a09 · outbound

This paper cites Open vocabulary semantic segmentation with patch aligned contrastive learning.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Open vocabulary semantic segmentation with patch aligned contrastive learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:03.819317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:58.528576Z digest=sha256:b6272668747ee1b1c4e0630fbfc01f57e0c918b9eca5d1818fcc85d73514122c

Observation 07b727f7-4407-4391-87fe-a8bfe5b40612 · outbound

This paper cites Multi-modal fusion transformer for end-to-end autonomous driving.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Multi-modal fusion transformer for end-to-end autonomous driving

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:03.640330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:58.607678Z digest=sha256:fafbb299119f7bae1379870c27cca03fd5f505b428089415f0beb4b7363186ed

Observation 65821ae5-7d69-417f-b360-c715d3232fee · outbound

This paper cites FreeSeg: Unified, Universal and Open-Vocabulary Image Segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning FreeSeg: Unified, Universal and Open-Vocabulary Image Segmentation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:58.663806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:58.663806Z digest=sha256:4362f056c1050ff03a2c177ac95a125e4732292ed41b3f5ef16c34814e619517

Observation 40fd47e4-5ccb-4d95-8057-d06f4e269c23 · outbound

This paper cites Learning transferable visual models from natural language supervision.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Learning transferable visual models from natural language supervision

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:03.433810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:58.739918Z digest=sha256:fead2671d7af9c649e5894d954b9939b0c4b739b24e5281da689150752b0e805

Observation 58b812b2-610e-4deb-9b50-9716536baee4 · outbound

This paper cites Making monolingual sentence embeddings multilingual using knowledge distillation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Making monolingual sentence embeddings multilingual using knowledge distillation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:03.276249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:58.824637Z digest=sha256:2de7b4ebdcad499e4271fec1fec357ee3e0bb10a8c605c56e80aa86f4411e913

Observation 980ea46c-73a8-4e5a-8a0a-02ee1dcdcd9d · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning LLaMA: Open and Efficient Foundation Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:58.916036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:58.916036Z digest=sha256:dc45c28729f5a73899c28d099849907e448e11255c9e07363b2856cd46cf3076

Observation 88db6878-8b79-429d-91bb-cd8e387a1eed · outbound

This paper cites Hierarchical Open-vocabulary Universal Image Segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Hierarchical Open-vocabulary Universal Image Segmentation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:59.006598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:59.006598Z digest=sha256:15caf077789be6b92bb81797603502a6aed5ceab6a4a1de615c81ba0ec1404ab

Observation d0328e41-79fb-439d-bfa9-29ee22989ef7 · outbound

This paper cites Skyscript: A large and semantically diverse vision-language dataset for remote sensing.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Skyscript: A large and semantically diverse vision-language dataset for remote sensing

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:03.090513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:59.128753Z digest=sha256:6509a3371e45aa9855ddfc562d90e7fd063d4f13f26b50fa284e62258c926dc9

Observation d05a8e1a-f177-4d60-9cec-c94438c6bcee · outbound

This paper cites an unresolved cited work.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:55:02.915036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:59.217814Z digest=sha256:6b0a0ede9ad05c44d27e354ae864115351545c9e2e4a04ea46989d72e91d7fe4

Observation 724773e4-fce9-46a7-ad91-2321f1791518 · outbound

This paper cites Towards open vocabulary learning: A survey.IEEE Transactions on Pattern Analysis and Machine Intelligence, 46(7):5092–5113, 2024.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Towards open vocabulary learning: A survey.IEEE Transactions on Pattern Analysis and Machine Intelligence, 46(7):5092–5113, 2024

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:59.291116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:59.291116Z digest=sha256:c72b351185a71e1613792d9e6f34e6fa29c4dcde62c6bc841750975195de374e

Observation aaa4a721-8054-4507-a337-1d5eaba0f8da · outbound

This paper cites Semantic projection network for zero-and few-label semantic segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Semantic projection network for zero-and few-label semantic segmentation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:02.723235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:59.370838Z digest=sha256:b1143c9a282886b7cf20d67895922566e6d7ea8bfc25f4714b5826405a0a73ea

Observation 5c91a28c-d88e-495f-b5df-05e93952e17b · outbound

This paper cites Sed: A simple encoder-decoder for open-vocabulary semantic segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Sed: A simple encoder-decoder for open-vocabulary semantic segmentation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:02.473823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:59.427761Z digest=sha256:8c2aa30bb1944cfee3420f7650602defe8d75699bb56391cbe0a125d0a48cecd

Observation aa57dabe-d4a1-4abf-9dbc-0aa92135ffcd · outbound

This paper cites FG-CLIP: Fine-Grained Visual and Textual Alignment.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning FG-CLIP: Fine-Grained Visual and Textual Alignment

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:59.478902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:59.478902Z digest=sha256:d5b98b6db5a8e340a06742adf5756cbbfb4d60676b13f306799a043ad3aa76c9

Observation c3f26acd-faf6-4078-bc6c-5d03ddd068f7 · outbound

This paper cites Groupvit: Semantic segmentation emerges from text supervision.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Groupvit: Semantic segmentation emerges from text supervision

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:02.312747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:59.483387Z digest=sha256:d01c914b4089be7a89dfe104e93e90e456596ac9fa800a561523f13e185b4b07

Observation 3bb9c853-c64c-4a00-8b97-13231b6c07a4 · outbound

This paper cites Open- vocabulary panoptic segmentation with text-to-image diffusion models.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Open- vocabulary panoptic segmentation with text-to-image diffusion models

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:02.077359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:59.487505Z digest=sha256:bdc30edb1e0c617c4f2c7cc3c6bd17b6c9cccdcf0216764a1b4ed1b5a36316d2

Observation 2f433679-d701-49d4-a2de-51dfafef1f24 · outbound

This paper cites Side adapter network for open- vocabulary semantic segmentation.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Side adapter network for open- vocabulary semantic segmentation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:01.904214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:59.513964Z digest=sha256:0104265fd8859f77a83b1a3e25185d69b7c786a2c09996cf1ea80f9871ba8f6d

Observation 7b46e5e6-873c-453c-ba2c-2b4741ec673e · outbound

This paper cites A simple baseline for open-vocabulary semantic segmentation with pre-trained vision-language model.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning A simple baseline for open-vocabulary semantic segmentation with pre-trained vision-language model

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:01.697431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:54:59.662103Z digest=sha256:a964b13581a47ccd68a354493993073240243b8e66a940a36818b3b548e01f46

Observation 4ff54c76-ce13-40da-86d2-51777c03cf60 · outbound

This paper cites Qwen2.5 Technical Report.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Qwen2.5 Technical Report

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:59.779546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:59.779546Z digest=sha256:c6a8fc5392ef73f71c4b71a9d879832ebc94f8de006b252eeb36fe0b1c3c7810

Observation ebb9b491-ccbe-4ff3-853c-9da6848c6fd7 · outbound

This paper cites Griffiths, Yuan Cao, and Karthik Narasimhan.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Griffiths, Yuan Cao, and Karthik Narasimhan

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:59.859535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:59.859535Z digest=sha256:4223fe4c402ad2207f1fc28adc35d59e330918a8502c03232bcdc22a44eef2d3

Observation b30d5edf-523b-4beb-99bc-3da5699f1760 · outbound

This paper cites Convolutions Die Hard: Open-Vocabulary Segmentation with Single Frozen Convolutional CLIP.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Convolutions Die Hard: Open-Vocabulary Segmentation with Single Frozen Convolutional CLIP

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:54:59.953787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:54:59.953787Z digest=sha256:c4829641211e233da7ffe167859bf92207f41c5819d39550ca1755e909574641

Observation 953fa962-b130-421d-9353-a104c2b3654a · outbound

This paper cites an unresolved cited work.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:55:01.547819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:55:00.053025Z digest=sha256:104c82cda03d3c76cfe91a7d853e7e58c4289012674bbcd9006575747577493e

Observation 56f356e3-2342-4a39-b49e-c1289f74e590 · outbound

This paper cites Open vocabulary scene parsing.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Open vocabulary scene parsing

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:01.417148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:55:00.211996Z digest=sha256:f77212292695dd2cb9d39a443bb05a25473ac7bb36ee464c09d9926090f651e3

Observation 77bd9a7c-9559-4bf6-bba7-f28c2265ac16 · outbound

This paper cites A foundation model for joint segmentation, detection and recognition of biomedical objects across nine modalities.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning A foundation model for joint segmentation, detection and recognition of biomedical objects across nine modalities

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:01.236717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:55:00.310535Z digest=sha256:e5ca62d97ab1fdb5edaa0c9bc7b3b3adb8d0abe034c5b8db75e2f020f9b585a7

Observation 815b547b-6a51-4578-85c1-b63601e02c2c · outbound

This paper cites Semantic understanding of scenes through the ade20k dataset.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Semantic understanding of scenes through the ade20k dataset

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:00.412108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:00.412108Z digest=sha256:cea43988f39d4eb5a4d2348e6bbac0eb5b210e35a84b0ff8aa955c6fc3ded23b

Observation c2b565aa-38f2-4291-94d4-e5c6c474e51e · outbound

This paper cites Extract free dense labels from clip.

OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning Extract free dense labels from clip

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:01.051537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:55:00.551518Z digest=sha256:16b37616daefb35b87278df16d5898b2ee5ec917ff4c4294fd0b944efb111076

Pith citing papers

No inbound Pith citation observations are available.