Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:18.789405Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 100 of 114 outbound references and 5 inbound Pith citation observations for arXiv:2505.18816.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:18.789405Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:06:22.967237Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:09:55.361578Z
100 of 114 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2535e581-f34c-49aa-a4cd-90ff023a2e52 · outbound
Reasoning Segmentation for Images and Videos: A Survey UI-Net: Interactive Artificial Neural Networks for Iterative Image Segmentation Based on a User Model
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5104d3b1-8a63-48c0-bd05-b019af82bf66 · outbound
Reasoning Segmentation for Images and Videos: A Survey ViCaS: A Dataset for Combining Holistic and Pixel-level Video Understanding using Captions with Grounded Segmentation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5628c86b-4642-451e-80ca-aa9b60782db4 · outbound
Reasoning Segmentation for Images and Videos: A Survey Burst: A benchmark for unifying object recognition, 39 segmentation and tracking in video, in: Proceedings of the IEEE/CVF winter conference on applications of computer vision, pp
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1596bf16-837f-4dc4-bc2f-653d81fd4952 · outbound
Reasoning Segmentation for Images and Videos: A Survey One Token to Seg Them All: Language Instructed Reasoning Segmentation in Videos
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98bc1725-92aa-476a-a66b-6da70b393a0b · outbound
Reasoning Segmentation for Images and Videos: A Survey Cores: Orchestrating the dance of reasoning and segmentation, in: European Conference on Computer Vision, Springer
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2799ce2-0ec9-4498-aadb-2f0dd8d18212 · outbound
Reasoning Segmentation for Images and Videos: A Survey Xmem++: Production-level video segmentation from few annotated frames, in: Pro- ceedings of the IEEE/CVF International Conference on Computer Vision, pp
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e7199cf-1e6b-407e-b00d-f1c96ac9424c · outbound
Reasoning Segmentation for Images and Videos: A Survey Coco-stuff: Thing and stuff classes in context, in: Proceedings of the IEEE conference on computer vision and pattern recognition, pp
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54b9c109-2a0d-4591-b8e9-1cb2d4b7aada · outbound
Reasoning Segmentation for Images and Videos: A Survey Coco-stuff: Thing and stuff classes in context, in: Proceedings of the IEEE conference on computer vision and pattern recognition, pp
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f732ecc8-2281-48dc-9fa4-6ee936b81096 · outbound
Reasoning Segmentation for Images and Videos: A Survey Pixel-Level Reasoning Segmentation via Multi-turn Conversations
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2672107-2874-4dd1-9b28-4eb09fc36747 · outbound
Reasoning Segmentation for Images and Videos: A Survey End-to-end object detection with transformers, in: European conference on computer vision, Springer
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c17b8da-d672-4c2d-8dcf-003f75f2aa62 · outbound
Reasoning Segmentation for Images and Videos: A Survey Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cf067c6-6e1d-49e2-b69c-d469484d8642 · outbound
Reasoning Segmentation for Images and Videos: A Survey Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0023a925-9097-4578-b971-51f60d688532 · outbound
Reasoning Segmentation for Images and Videos: A Survey Sam4mllm: Enhance multi-modal large language model for referring ex- pression segmentation, in: European Conference on Computer Vision, Springer
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4250dca0-38ea-4da1-bb92-80a94ba49762 · outbound
Reasoning Segmentation for Images and Videos: A Survey Masked-attention mask transformer for universal image segmentation, in: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ed77509-9ee5-4d18-adb0-fd733b49b0dc · outbound
Reasoning Segmentation for Images and Videos: A Survey Xmem: Long-term video object seg- mentation with an atkinson-shiffrin memory model, in: European Confer- ence on Computer Vision, Springer
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47fd545a-b396-4222-bb23-690101e7ba9a · outbound
Reasoning Segmentation for Images and Videos: A Survey Vocabulary-free image classification
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf1d2c9b-d29f-4510-bfab-118019625441 · outbound
Reasoning Segmentation for Images and Videos: A Survey Scannet: Richly-annotated 3d reconstructions of indoor scenes, in: Proceedings of the IEEE conference on computer vision and pattern recognition, pp
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e90944a5-7290-4d20-87f3-5469fca853cc · outbound
Reasoning Segmentation for Images and Videos: A Survey Tao: A large-scale benchmark for tracking any object, in: Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part V 16, Springer
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82302091-de9e-46e9-bae2-48fc6dbb0626 · outbound
Reasoning Segmentation for Images and Videos: A Survey Motion-Grounded Video Reasoning: Understanding and Perceiving Motion at Pixel Level
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4b13401-579b-4aad-ac1a-9744ffa96064 · outbound
Reasoning Segmentation for Images and Videos: A Survey Mevis: A large-scale benchmark for video segmentation with motion expressions, in: Proceed- ings of the IEEE/CVF International Conference on Computer Vision, pp
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 783e111d-b744-4e99-b275-4ac4768809cb · outbound
Reasoning Segmentation for Images and Videos: A Survey Mose: A new dataset for video object segmentation in complex scenes, in: Pro- ceedings of the IEEE/CVF international conference on computer vision, pp
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 579de55c-4d1f-4f10-b152-c97f38c742c9 · outbound
Reasoning Segmentation for Images and Videos: A Survey Panoptic Segmentation: A Review
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bcf025c8-490b-4b28-b496-4f577043241f · outbound
Reasoning Segmentation for Images and Videos: A Survey Oops! predicting unintentional action in video, in: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56a0ce34-9da4-4754-8128-2335acf07537 · outbound
Reasoning Segmentation for Images and Videos: A Survey A Survey for Foundation Models in Autonomous Driving
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abe7de4f-a2cf-4111-ac9e-f8fe5fd37ed8 · outbound
Reasoning Segmentation for Images and Videos: A Survey Part- aware panoptic segmentation, in: Proceedings of the IEEE/CVF Confer- ence on Computer Vision and Pattern Recognition, pp
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 770e0673-4881-4b20-abde-e9f044696652 · outbound
Reasoning Segmentation for Images and Videos: A Survey The Devil is in Temporal Token: High Quality Video Reasoning Segmentation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3443753-da5e-48db-b968-d39eaebd1a60 · outbound
Reasoning Segmentation for Images and Videos: A Survey The iapr tc-12 benchmark: A new evaluation resource for visual information systems, in: International workshop ontoImage, pp
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0496995-8774-403f-a4dd-cde1f9e7a271 · outbound
Reasoning Segmentation for Images and Videos: A Survey Lvis: A dataset for large vocabu- lary instance segmentation, in: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b31db6e-4004-49a0-9ed4-73dddcf0523d · outbound
Reasoning Segmentation for Images and Videos: A Survey A survey on instance segmentation: state of the art
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99f2b6af-40a8-4263-8878-38917cb669dd · outbound
Reasoning Segmentation for Images and Videos: A Survey A brief survey on semantic segmentation with deep learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2a94a13-f9d0-4335-aa48-30e7e1af3941 · outbound
Reasoning Segmentation for Images and Videos: A Survey LoRA: Low-Rank Adaptation of Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ff5aa0d-221e-4702-92d0-32ff4c5ba8a4 · outbound
Reasoning Segmentation for Images and Videos: A Survey Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3d0aa40-7f13-4279-b0aa-637e14b53d26 · outbound
Reasoning Segmentation for Images and Videos: A Survey MMR: A Large-scale Benchmark Dataset for Multi-target and Multi-granularity Reasoning Segmentation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 253e3da3-74e8-4901-9f12-5f9800cc4d08 · outbound
Reasoning Segmentation for Images and Videos: A Survey Referitgame: Referring to objects in photographs of natural scenes, in: Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP), pp
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e889dbfc-6c15-40ee-b6c1-ad95c4e89f14 · outbound
Reasoning Segmentation for Images and Videos: A Survey Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21a0058f-f297-4cf8-8f18-d9b02ba6c83c · outbound
Reasoning Segmentation for Images and Videos: A Survey Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bdf61cd-b914-4ae6-9b78-ede9ac9bb40a · outbound
Reasoning Segmentation for Images and Videos: A Survey Segment Anything
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35d9f209-ff49-4b67-aded-df2dbded454b · outbound
Reasoning Segmentation for Images and Videos: A Survey Visual genome: Connecting language and vision using crowdsourced dense image annota- tions
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07e0de33-8137-44ee-ba4c-eb01420f4378 · outbound
Reasoning Segmentation for Images and Videos: A Survey The open images dataset v4: Unified image classification, object detection, and visual relationship detection at scale
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4177e72-78d3-4531-aaf8-79acb0d3d0d0 · outbound
Reasoning Segmentation for Images and Videos: A Survey Lisa: Reasoning segmentation via large language model, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31641866-bf98-47f1-9203-f901a609f30f · outbound
Reasoning Segmentation for Images and Videos: A Survey Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language mod- els, in: International conference on machine learning, PMLR
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8b427f1-981a-4b88-9931-54c2b2ae607d · outbound
Reasoning Segmentation for Images and Videos: A Survey SegEarth-R1: Geospatial Pixel Reasoning via Large Language Model
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b28ed6f8-e70a-43ca-aa56-40198daea904 · outbound
Reasoning Segmentation for Images and Videos: A Survey Robust referring video object segmentation with cyclic structural consensus, in: Proceed- ings of the IEEE/CVF International Conference on Computer Vision, pp
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebc5c99f-5a2a-488f-b500-da0ab41c0c1b · outbound
Reasoning Segmentation for Images and Videos: A Survey Llama-vid: An image is worth 2 tokens in large language models, in: European Conference on Computer Vision, Springer
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b393a59-01d6-491b-9e1f-1e25d2744064 · outbound
Reasoning Segmentation for Images and Videos: A Survey Microsoft coco: Common objects in context, in: Computer Vision–ECCV 2014: 13th European Confer- ence, Zurich, Switzerland, September 6-12, 2014, Proceedings, Part V 13, Springer
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9b0163e8-9b0b-48eb-ba6a-16065bb2e529 · outbound
Reasoning Segmentation for Images and Videos: A Survey Gres: Generalized referring expression segmentation, in: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 09a281c4-b093-437f-84cf-ce514a7497c9 · outbound
Reasoning Segmentation for Images and Videos: A Survey Visual instruction tuning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 17fd4db2-02f2-476e-a215-7c8d40dae819 · outbound
Reasoning Segmentation for Images and Videos: A Survey Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af282039-22b4-42fe-8e33-9d88f65ff8df · outbound
Reasoning Segmentation for Images and Videos: A Survey Ground abstract structure concepts of scaffolding systems for automatic compliance check- ing based on reasoning segmentation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6c3dddf-c8cf-49f0-8e4d-304427d957a7 · outbound
Reasoning Segmentation for Images and Videos: A Survey Video anomaly detection and explanation via large language models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc0c66bc-dbca-49f9-9a4f-6239e505fd84 · outbound
Reasoning Segmentation for Images and Videos: A Survey Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41bdf037-fc82-43e5-b1cd-8f4d9b98bd02 · outbound
Reasoning Segmentation for Images and Videos: A Survey Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4552696c-2c40-4fcf-8ca9-645ee71e25db · outbound
Reasoning Segmentation for Images and Videos: A Survey Large-scale video panoptic segmentation in the wild: A benchmark, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pat- tern Recognition, pp
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7785866-13af-4b67-a815-1f83c89251f6 · outbound
Reasoning Segmentation for Images and Videos: A Survey Image segmentation using deep learning: A survey
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0d0e917-465d-46e7-9e8d-6def35c33a95 · outbound
Reasoning Segmentation for Images and Videos: A Survey GPT-4 Technical Report
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1d4d6a45-bf66-432c-8b22-6c4f1b2ab48f · outbound
Reasoning Segmentation for Images and Videos: A Survey GPT-4V(ision) System Card
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 53191d63-744b-4e11-9acd-8cc39efae377 · outbound
Reasoning Segmentation for Images and Videos: A Survey DINOv2: Learning Robust Visual Features without Supervision
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e9a4395-0340-4ad8-92db-1bdbc82aa975 · outbound
Reasoning Segmentation for Images and Videos: A Survey Flickr30k entities: Collecting region-to-phrase correspondences for richer image-to-sentence models, in: Proceedings of the IEEE international conference on computer vision, pp
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 264110e8-d488-4fd7-ba05-e7207fa041de · outbound
Reasoning Segmentation for Images and Videos: A Survey The 2017 DAVIS Challenge on Video Object Segmentation
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 184032c2-86a5-4be1-8391-b84572ca2e99 · outbound
Reasoning Segmentation for Images and Videos: A Survey Occluded video instance segmentation: A benchmark
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 316d3aa7-825b-4253-8944-3f4469407fdb · outbound
Reasoning Segmentation for Images and Videos: A Survey Reasoning to attend: Try to understand how< seg> token works
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a67f527-dab8-4bf5-bee0-62f351497b25 · outbound
Reasoning Segmentation for Images and Videos: A Survey Learning trans- ferable visual models from natural language supervision, in: International conference on machine learning, PMLR
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e994d17d-0097-423f-9f4c-bbda7b8edf20 · outbound
Reasoning Segmentation for Images and Videos: A Survey Paco: Parts and attributes of common objects, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f5f67165-0da6-4ff5-9b25-611289af9cd7 · outbound
Reasoning Segmentation for Images and Videos: A Survey Glamm: Pixelground- inglargemultimodalmodel, in: ProceedingsoftheIEEE/CVFConference on Computer Vision and Pattern Recognition, pp
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5e00767-6e8e-4f2e-bc96-94c6afd01c47 · outbound
Reasoning Segmentation for Images and Videos: A Survey SAM 2: Segment Anything in Images and Videos
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13e78c09-8f72-4f36-a0a7-99b403efb154 · outbound
Reasoning Segmentation for Images and Videos: A Survey Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 018d04b0-a9d8-4d8e-9aa5-5213379dccff · outbound
Reasoning Segmentation for Images and Videos: A Survey Pixellm: Pixel reasoning with large multimodal model, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b7865683-1aa7-4cf3-8385-19d6aa57d54a · outbound
Reasoning Segmentation for Images and Videos: A Survey Object Hallucination in Image Captioning
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f488d60a-fcf2-408d-908e-3ea4f8cb5ed5 · outbound
Reasoning Segmentation for Images and Videos: A Survey Unresolved cited work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8f6a796f-b9c3-43eb-afa1-2f3b46f58aae · outbound
Reasoning Segmentation for Images and Videos: A Survey Position: Foundation Models Need Digital Twin Representations
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6154626e-f598-42ac-b4cf-248e91f05f2b · outbound
Reasoning Segmentation for Images and Videos: A Survey Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ab7aff1f-a7d5-44b9-82d4-6c5c1e4bc53a · outbound
Reasoning Segmentation for Images and Videos: A Survey RVTBench: A Benchmark for Visual Reasoning Tasks
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6260a082-1e7a-49cf-8066-b91dd9288ce4 · outbound
Reasoning Segmentation for Images and Videos: A Survey Operating Room Workflow Analysis via Reasoning Segmentation over Digital Twins
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92574bd9-3902-4306-b588-39bf8b41ed14 · outbound
Reasoning Segmentation for Images and Videos: A Survey Online Reasoning Video Segmentation with Just-in-Time Digital Twins
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d201cc74-11a5-4ff9-b09c-545f0a9bf9ca · outbound
Reasoning Segmentation for Images and Videos: A Survey Unresolved cited work
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e75943ca-b73f-4740-b246-86923e47b367 · outbound
Reasoning Segmentation for Images and Videos: A Survey EVA-CLIP: Improved Training Techniques for CLIP at Scale
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed7bddb8-2d12-4292-9d58-85e2c9f5c624 · outbound
Reasoning Segmentation for Images and Videos: A Survey Growcut: Interactive multi-label nd image segmentation by cellular automata, in: proc
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 02c0e07d-8a62-4754-b277-3d52456ef1a1 · outbound
Reasoning Segmentation for Images and Videos: A Survey Reducing the annotation effort for video object segmentation datasets, in: Proceed- ings of the IEEE/CVF Winter Conference on Applications of Computer Vision, pp
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7b617589-b580-4d20-a999-c99295fe8c3a · outbound
Reasoning Segmentation for Images and Videos: A Survey Prima: Multi-image vision-language models for reasoning segmentation
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5194841-dd28-44a5-bad6-18c571d4eda9 · outbound
Reasoning Segmentation for Images and Videos: A Survey Ov-vis: Open-vocabulary video instance segmentation
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4f1441ef-eeb7-4181-9c0a-80a193c00a37 · outbound
Reasoning Segmentation for Images and Videos: A Survey Towards open-vocabulary video instance segmentation, in: proceedings of the IEEE/CVF international conference on computer vision, pp
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 77313945-d7f1-4d7f-8ddb-150254cd2b55 · outbound
Reasoning Segmentation for Images and Videos: A Survey Llm-seg: Bridging image segmentation and large language model reasoning, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 135e0f7e-ac33-4b4b-822b-7803a27d02ac · outbound
Reasoning Segmentation for Images and Videos: A Survey Unidentified video objects: A benchmark for dense, open-world segmentation, in: Proceed- ings of the IEEE/CVF international conference on computer vision, pp
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 69049e09-7e92-4305-a151-604abd22817f · outbound
Reasoning Segmentation for Images and Videos: A Survey SegLLM: Multi-round Reasoning Segmentation
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c58eabf6-7f01-4e6e-92c4-d9cbc0d0ad04 · outbound
Reasoning Segmentation for Images and Videos: A Survey LaSagnA: Language-based Segmentation Assistant for Complex Queries
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84d1e245-ea60-4124-8039-1caec776250d · outbound
Reasoning Segmentation for Images and Videos: A Survey InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87f917ea-32ca-4d10-a3e9-72157c73ce39 · outbound
Reasoning Segmentation for Images and Videos: A Survey Chain-of-thought prompting elicits reasoning in large language models
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4eeaab8-288b-41c1-a4a6-1bde2a30ff2b · outbound
Reasoning Segmentation for Images and Videos: A Survey Ov- parts: Towards open-vocabulary part segmentation
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d94a26f9-d9b2-495b-a66b-3aec75517bdd · outbound
Reasoning Segmentation for Images and Videos: A Survey Phrasecut: Language- based image segmentation in the wild, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 482a4945-58ae-4189-81d3-31cdb93c7b47 · outbound
Reasoning Segmentation for Images and Videos: A Survey See say and segment: Teaching lmms to overcome false premises, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1682cf05-b69a-4cd6-a225-5c8f435a83c4 · outbound
Reasoning Segmentation for Images and Videos: A Survey Gsva: Generalized segmentation via multimodal large language models, in: Pro- ceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2aeeff76-74e6-4d7e-a14a-eace992817c2 · outbound
Reasoning Segmentation for Images and Videos: A Survey Youtube-vos: Sequence-to-sequence video object segmentation, in: Proceedings of the European conference on computer vision (ECCV), pp
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8374649-c868-43fb-a8d0-65a273149fd8 · outbound
Reasoning Segmentation for Images and Videos: A Survey Visa: Reasoning video object segmentation via large language models, in: European Conference on Computer Vision, Springer
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a40796a9-b8ff-4ef7-8d6d-11208ac46186 · outbound
Reasoning Segmentation for Images and Videos: A Survey Panop- tic scene graph generation, in: European Conference on Computer Vision, Springer
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f8c2825-a571-46e7-ab1b-7d7dde838826 · outbound
Reasoning Segmentation for Images and Videos: A Survey Video instance segmentation, in: Pro- ceedings of the IEEE/CVF international conference on computer vision, pp
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5eadc24a-072c-40ed-a1b0-bfaa947003b7 · outbound
Reasoning Segmentation for Images and Videos: A Survey Depth Anything V2
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d846f45-b74d-4917-bef0-fe69edc5c59d · outbound
Reasoning Segmentation for Images and Videos: A Survey An improved baseline for reasoning segmentation with large language model
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 61a7ef6b-27f3-4b50-8a02-cc88f72dbcf5 · outbound
Reasoning Segmentation for Images and Videos: A Survey Empowering Segmentation Ability to Multi-modal Large Language Models
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation add5e5f7-086d-41d7-be11-5af856289345 · outbound
Reasoning Segmentation for Images and Videos: A Survey Follow the rules: reasoning for video anomaly detection with large language models, in: European Conference on Computer Vision, Springer
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8ad9b85-c90d-4b59-b68f-c088e2d7806d · outbound
Reasoning Segmentation for Images and Videos: A Survey Lavt: Language-aware vision transformer for referring image segmenta- tion, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6507a37d-5882-44da-844d-3acb78c39a8e · inbound
Temporally-Constrained Video Reasoning Segmentation and Automated Benchmark Construction Reasoning Segmentation for Images and Videos: A Survey
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dba12a3-d13f-4f62-8e1e-4ff91d86e7a5 · inbound
GTPBD-MM: A Global Terraced Parcel and Boundary Dataset with Multi-Modality Reasoning Segmentation for Images and Videos: A Survey
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99dbdaa5-f54f-43cb-989b-8cf35a603dd1 · inbound
An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentation Reasoning Segmentation for Images and Videos: A Survey
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation faaf51be-153a-4a61-9ba3-1c7b12e6a4a4 · inbound
From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models Reasoning Segmentation for Images and Videos: A Survey
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b5da47c9-260f-41ab-9dba-72e58d18e44a · inbound
DGSeg: Dynamic Gating of Semantic-Spatial Guided Predictions for Reasoning Segmentation Reasoning Segmentation for Images and Videos: A Survey
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.