Pith. sign in

Paper Citation Record · LEDGER

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception

As of 10 August 2026, this Paper Citation Record lists 100 of 116 outbound references and 4 inbound Pith citation observations for arXiv:2508.11256.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.11256 v1

Coverage vector

measured 100 of 116 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:10:19.724032Z

measured 104 of 104 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T13:42:25.796548Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T13:33:28.299770Z

Reference resolution

100 of 116 outbound references displayed

  • verified exact0
  • verified fuzzy64
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b6e465b7-0a4d-4f1e-99af-92c722ea4481 · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Faster r-cnn: Towards real-time object detection with region proposal networks,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:09.840942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:09.840942Z digest=sha256:9973400ff6c7bcf9ba932f0ac7c43799117a56f21b159301d0bb27e922db4838

Observation b13cb74d-faf2-41ad-aa96-8b19e342e8fa · outbound

This paper cites DAB-DETR: Dynamic anchor boxes are better queries for DETR,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception DAB-DETR: Dynamic anchor boxes are better queries for DETR,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:09.985717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:09.985717Z digest=sha256:f8022062d6632b3e268b008a6b1488fa025cb6cd1b7f5f85b91446ec8eff047a

Observation a81e8b4b-1baf-4134-a92b-8061e6e25e0a · outbound

This paper cites U-net: Convolutional net- works for biomedical image segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception U-net: Convolutional net- works for biomedical image segmentation,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:10.172898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:10.172898Z digest=sha256:69b18cc531178e1da83b09fa357f672ebbf6ec6bef0964ab30e49283ecbff95b

Observation 676ab988-432f-499b-bcb3-7877e37249b7 · outbound

This paper cites Masked-attention mask transformer for universal image segmenta- tion,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Masked-attention mask transformer for universal image segmenta- tion,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:10.358448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:10.358448Z digest=sha256:f70da9cb338665031cf0f4fc95f3a530347d8663c6fba3c4916d79849bfba438

Observation 26dd1416-6cde-424a-8587-73b8f536edc4 · outbound

This paper cites Mask dino: Towards a unified transformer-based framework for object detection and segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Mask dino: Towards a unified transformer-based framework for object detection and segmentation,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:10.521983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:10.521983Z digest=sha256:27567092c3c38f02c85d2280d17d8d8a69df423cad9047f6d4a5c96b4c5cd832

Observation 3720bf71-4fe1-469c-8eca-f51954994a68 · outbound

This paper cites Enhanced training of query-based object detection via selective query recollection,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Enhanced training of query-based object detection via selective query recollection,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:10.665833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:10.665833Z digest=sha256:705be1ced2842b480e39a11988f70f2ef057540425fe24635d8725371ef693b8

Observation fee31577-53ec-4b70-972d-1b8c35b02531 · outbound

This paper cites Deformable detr: Deformable transformers for end-to-end object detection,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Deformable detr: Deformable transformers for end-to-end object detection,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:10.784765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:10.784765Z digest=sha256:42d74951c19abd17360dd3e8899256b028fd25d5d2ede0ac1b8c9f3f162d2bf6

Observation a55b8869-a91b-4a9c-9e25-458a2551f58c · outbound

This paper cites Open-vocabulary object detection using captions,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Open-vocabulary object detection using captions,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:10.884468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:10.884468Z digest=sha256:47e24b89220826614978f81235457ed5a18dd1040535b05be7ee06bb74f0b135

Observation 41ae203e-0c3f-479d-86a3-f58951966f6f · outbound

This paper cites Aligning bag of regions for open-vocabulary object detection,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Aligning bag of regions for open-vocabulary object detection,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:10.958133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:10.958133Z digest=sha256:1802b61bbb9096973f2ed0a2f4304a96f1ab0838a27f4e0316cdbd62d3fad83c

Observation 80667852-540e-44d3-9631-282ed239b717 · outbound

This paper cites Cora: Adapting clip for open- vocabulary detection with region prompting and anchor pre-matching,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Cora: Adapting clip for open- vocabulary detection with region prompting and anchor pre-matching,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.041423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.041423Z digest=sha256:3efb9eb1a56d42f0e09fdd8e7605d24e34267ed3bd4d4d0d38239088d434314e

Observation 4b505aa2-f179-4821-b956-5f601838d1c6 · outbound

This paper cites Cat- seg: Cost aggregation for open-vocabulary semantic segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Cat- seg: Cost aggregation for open-vocabulary semantic segmentation,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.095095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.095095Z digest=sha256:c0bd71d84ad1febb1c649aad417790c4d6b400dffd0bf6e73c77a2f46d170044

Observation 005594f2-0264-46a5-acf0-9f44c9ce3e10 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Learning transferable visual models from natural language supervi- sion,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.167225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.167225Z digest=sha256:66207fac8a03b45bc31a5800b6ddd1fbe4ad608b62057a077539c86814352566

Observation 254a089b-43b8-4ebb-bb04-7832a204a03b · outbound

This paper cites Scaling language- image pre-training via masking,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Scaling language- image pre-training via masking,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.249205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.249205Z digest=sha256:1538a131c2d399a9afe14a8748c0c704855921e0affd4740cc4e28e50a52144e

Observation 51e3ac40-5b5c-4e61-b2e7-35fc905e8009 · outbound

This paper cites Clim: Contrastive language-image mosaic for region representation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Clim: Contrastive language-image mosaic for region representation,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.367823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.367823Z digest=sha256:2208b696ef38e34a0758ede5cdaab40c9626017f44829fa97651ef23f32cd025

Observation 3cbf97b9-dfb9-46f1-b7aa-f2fa527463e9 · outbound

This paper cites CLIPSelf: Vision transformer distills itself for open-vocabulary dense prediction,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception CLIPSelf: Vision transformer distills itself for open-vocabulary dense prediction,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.471990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.471990Z digest=sha256:ad0bf165bff21e3fc228be67988277c59f7fab7936bcd87184d03dd3211e40e3

Observation 4ffefd41-6b38-4499-b361-7086adddb965 · outbound

This paper cites Regionclip: Region-based language- image pretraining,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Regionclip: Region-based language- image pretraining,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.550247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.550247Z digest=sha256:4bd21400d132ac6c3b731dba150543762ac38eda5de0a98329d30f5bf201bbff

Observation d66d7f3c-8cb1-4afd-8f2a-3fd23eae852f · outbound

This paper cites Ov-dquo: Open-vocabulary detr with denoising text query train- ing and open-world unknown objects supervision,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Ov-dquo: Open-vocabulary detr with denoising text query train- ing and open-world unknown objects supervision,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.632230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.632230Z digest=sha256:0fd8aa6f59b5ccce4f9fec431fc0a66298cbd57de8dbad56b16e446228231a90

Observation 92b6f947-9229-436a-9b7a-8db1109c8f04 · outbound

This paper cites Open-vocabulary object de- tection via vision and language knowledge distillation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Open-vocabulary object de- tection via vision and language knowledge distillation,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.712835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.712835Z digest=sha256:ba4d5192332b41ee3cbf5ec0151cfea4517e7ca988a5202a6aac9f5147e88365

Observation b511101e-e005-4623-89cc-81a6ce653cf9 · outbound

This paper cites F-vlm: Open-vocabulary object detection upon frozen vision and language models,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception F-vlm: Open-vocabulary object detection upon frozen vision and language models,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.778514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.778514Z digest=sha256:967da8a812402e36506c87aba918abfe69440ac2a542a40b7cb8393ff5e1df10

Observation 59667dd5-5997-409f-ad3a-25b3c4c980f5 · outbound

This paper cites Open-vocabulary semantic segmentation with mask-adapted clip,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Open-vocabulary semantic segmentation with mask-adapted clip,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:11.880929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:11.880929Z digest=sha256:d94c9598d71ded873b9cea5e1d4863406990ba37008bc644975d85f92569ab6b

Observation c4ab2810-05bc-487f-9d45-591ebcfe8151 · outbound

This paper cites High-resolution image synthesis with latent diffusion models,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception High-resolution image synthesis with latent diffusion models,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.005807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.005807Z digest=sha256:19d79090f3052b1d19157003ee6ec68b8d30fb83ac6dd126a863c788b554b0a8

Observation 3cd92037-2c8a-4627-b85f-0d5216336030 · outbound

This paper cites A survey on open-vocabulary detection and seg- mentation: Past, present, and future,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception A survey on open-vocabulary detection and seg- mentation: Past, present, and future,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.093572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.093572Z digest=sha256:fa929120f6551032c13e62cd48b7b7ab33303b66cc908185ab79bcf42d9c7a93

Observation 6ebec5cc-a8ed-426f-8684-035a8e05981c · outbound

This paper cites Towards open vocabulary learning: A survey,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Towards open vocabulary learning: A survey,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.169894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.169894Z digest=sha256:4647c0917fe980ca9adb359be21d5b98c832f08221d8f1c499668f0ddfa1c79b

Observation b9605788-1ab7-4438-beb8-d969d2a32bbe · outbound

This paper cites Object-aware distillation pyramid for open-vocabulary object detection,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Object-aware distillation pyramid for open-vocabulary object detection,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.210528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.210528Z digest=sha256:96c1c02cddbe3e82c7c5301e3b15af88b3adec4e196473a1b48a1754425501e2

Observation 23979c0f-7801-4bfa-aca1-778effe26441 · outbound

This paper cites Groupvit: Semantic segmentation emerges from text supervision,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Groupvit: Semantic segmentation emerges from text supervision,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.287880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.287880Z digest=sha256:54c374bc78e000cf5dabb950cd0bacc254b7fbad537c2c81debe5886e915e527

Observation 8a0721e4-2a9c-4870-9444-f119a472b4eb · outbound

This paper cites Scaling open-vocabulary image segmentation with image-level labels,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Scaling open-vocabulary image segmentation with image-level labels,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.348029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.348029Z digest=sha256:d115af48a408162ef77f8f39b70bd1a8c71cc3c00549ecdc29bf8b39538ba7ca

Observation b36b47c6-cb06-4053-a69a-b1310328eeb5 · outbound

This paper cites Taming self-training for open-vocabulary object detection,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Taming self-training for open-vocabulary object detection,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.474425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.474425Z digest=sha256:705486be3f98d68eb674c224709bc997011dcffb11cc10161d70ef03b95e5a09

Observation 54c81a43-db65-43fc-9ea9-fe6458e595fe · outbound

This paper cites Detecting twenty-thousand classes using image-level supervision,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Detecting twenty-thousand classes using image-level supervision,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.594935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.594935Z digest=sha256:812ab129410385684d8ef7b5233b486ca5f45fa866ddacb3197c332f30f25eaa

Observation 87e73afb-085b-43e8-8cb6-617f97ab90ce · outbound

This paper cites Open-vocabulary detr with conditional matching,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Open-vocabulary detr with conditional matching,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.694061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.694061Z digest=sha256:5883e03ded5a7b67bfd12bf7008a7f4724f938d8031e8f07afbe55d58a9bbf8a

Observation 7933283e-d56e-48c3-8910-4478e6c73780 · outbound

This paper cites Global knowledge calibration for fast open-vocabulary segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Global knowledge calibration for fast open-vocabulary segmentation,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.791515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.791515Z digest=sha256:85bab7ae49720f462c52317e13ac2f7ecda348e510ed73595618d12fe8fd53b1

Observation 70b6f036-7c3c-46b1-9a72-94c8db72d67a · outbound

This paper cites Densegrounding: Improving dense language- vision semantics for ego-centric 3d visual grounding,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Densegrounding: Improving dense language- vision semantics for ego-centric 3d visual grounding,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:12.869705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:12.869705Z digest=sha256:939021b99dda23f59d7095a6a224829c6ad42bb977f6c8c85210da24c55cd0b9

Observation 65466bf6-0e39-45c0-bee0-9607668b5824 · outbound

This paper cites Detect anything 3d in the wild,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Detect anything 3d in the wild,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:13.005986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:13.005986Z digest=sha256:eecff95fcfb37b65b29dc2a606726431be7e2c7575840e95563f3875af125bce

Observation 9e3e68d5-8023-4b95-8956-d49f03946f03 · outbound

This paper cites Sam3d: Segment anything in 3d scenes,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Sam3d: Segment anything in 3d scenes,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:13.097939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:13.097939Z digest=sha256:159eeb03777dc3c72c63a887c1052e92d91c04173b43e9ecbe4cadf6b8fd31af

Observation a294c18a-4d55-4924-82c5-b2ba4e5da431 · outbound

This paper cites Ovir-3d: Open-vocabulary 3d instance retrieval without training on 3d data,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Ovir-3d: Open-vocabulary 3d instance retrieval without training on 3d data,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:13.236236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:13.236236Z digest=sha256:8fa5aebcb7a47cd8d1c732aa740ab4ae5be91d149ebe1f92bcaa2ccfa492059c

Observation f63594e5-8b8c-45dd-82cd-bcd277462492 · outbound

This paper cites Open- vocabulary object 6d pose estimation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Open- vocabulary object 6d pose estimation,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:13.309695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:13.309695Z digest=sha256:36b66c14c329791cd13cd2ca0723e19aaca62001b4eceac6866969b04abf8d7c

Observation c5725e7e-534b-46ba-b95f-7ec2ef16536b · outbound

This paper cites Open3dis: Open-vocabulary 3d instance segmentation with 2d mask guidance,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Open3dis: Open-vocabulary 3d instance segmentation with 2d mask guidance,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:13.348117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:13.348117Z digest=sha256:63459f2bd7ee1791a5013f6bdb72dccc1b04ce2872d19284d4c23523ba68c2a4

Observation 97ff88ea-e009-408f-9049-a35ac8abe379 · outbound

This paper cites Openmask3d: Open-vocabulary 3d instance segmenta- tion,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Openmask3d: Open-vocabulary 3d instance segmenta- tion,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.104216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:13.444699Z digest=sha256:26de3b372e6d084d6112f0067dec5704cd579335520009d79481e6d58fc182a6

Observation a6081716-0575-49b1-b84c-03f10a955575 · outbound

This paper cites Clip-vis: Adapting clip for open-vocabulary video instance segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Clip-vis: Adapting clip for open-vocabulary video instance segmentation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.095328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:13.562789Z digest=sha256:8900e0f58e98adb817af9358b7a6cd6b28c381c599ae97d5e6e4b269d80ea578

Observation ccb798fd-ca9f-45e8-9765-4fb34d4485ae · outbound

This paper cites Semantic and sequential alignment for referring video object segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Semantic and sequential alignment for referring video object segmentation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.086407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:13.698246Z digest=sha256:add85bbb71c0451402c36c6f96a28edea075fe5ced86c55c920237f2e1a2f5f3

Observation a6de5595-2620-4bed-8115-fd83c51a5ce7 · outbound

This paper cites Unified embedding align- ment for open-vocabulary video instance segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Unified embedding align- ment for open-vocabulary video instance segmentation,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.076767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:13.776707Z digest=sha256:2f9a20e2908ade5e57b8de9e57688c993be68e4abd933ca2b2d26587b7e42fa5

Observation 58b3e3ad-cc6a-4666-a1a0-406302272a52 · outbound

This paper cites Sigmoid loss for language image pre-training,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Sigmoid loss for language image pre-training,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.066746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:13.860179Z digest=sha256:ae5e96c81ae4ed944e7e42c6db1d78d632ae005878e8e56c022b4edde698af4a

Observation a556ba8d-5184-490a-9d16-65668d7f8f84 · outbound

This paper cites Learning mask-aware clip representations for zero-shot segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Learning mask-aware clip representations for zero-shot segmentation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.057529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:13.940164Z digest=sha256:280908fd5439500014ed34fb7c67c620cb126da7238b8191cc6793b1a7a1c3b4

Observation cadb0cd1-c63c-48b8-be3f-afa148e45b02 · outbound

This paper cites Language-driven semantic segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Language-driven semantic segmentation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.048518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:14.042602Z digest=sha256:79e747c31bad7236d48c3d9a397075c4e10856c095deaf5c77c276bb4560ff6b

Observation 65b5c39a-8931-4ff1-bd2b-2c2ff2964766 · outbound

This paper cites Open vocabulary semantic segmentation with patch aligned contrastive learning,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Open vocabulary semantic segmentation with patch aligned contrastive learning,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.039381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:14.204613Z digest=sha256:31f5d83e7cce9340413133a475a31b54edfa17a799ab9ea4aed4d1ead15312b3

Observation b25a119c-0a21-4bb7-8a81-19d80f4db41c · outbound

This paper cites Sam-clip: Merging vision foundation models towards semantic and spatial under- standing,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Sam-clip: Merging vision foundation models towards semantic and spatial under- standing,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.028511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:14.294410Z digest=sha256:403a8bc6d644bf57565c687cc133f181be049d8e85caa20ca573a603c7b4cb38

Observation c51eff8f-92a3-4ad1-a9d2-80ea322a2df9 · outbound

This paper cites Open- vocabulary sam: Segment and recognize twenty-thousand classes in- teractively,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Open- vocabulary sam: Segment and recognize twenty-thousand classes in- teractively,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.018831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:14.362385Z digest=sha256:67cc65de44ecbbc83dadc0af5a84d41b8695f028ee33e13761dde7f6265f1ead

Observation 61478947-f6cf-4a1a-a6fe-e8d95eb959e3 · outbound

This paper cites Frozenseg: Harmo- nizing frozen foundation models for open-vocabulary segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Frozenseg: Harmo- nizing frozen foundation models for open-vocabulary segmentation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:24.009620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:14.437271Z digest=sha256:ab03dc8f33cf0b4b293e270b5fa4ce9612f6e10fcc22754667a6ddafd3d64572

Observation aee0f0a6-dac6-4a7c-94a9-c223bfc651c2 · outbound

This paper cites Segment anything,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Segment anything,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.998121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:14.544984Z digest=sha256:532ac2ed84eb764b99b2c68920f93e03e65709d317fc34da81398a7c7a294b67

Observation f0d793ae-1c9d-4c32-b3eb-7c4d95387681 · outbound

This paper cites Rep- resentation alignment for generation: Training diffusion transformers is easier than you think,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Rep- resentation alignment for generation: Training diffusion transformers is easier than you think,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.987948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:14.679168Z digest=sha256:70c051ae9e93739f45ce50cb7f8c8b0fefc873fe42883be628819be42f8dd37f

Observation 986b009c-be4d-49bf-b945-6923fd7e86fc · outbound

This paper cites A convnet for the 2020s,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception A convnet for the 2020s,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.976472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:14.803588Z digest=sha256:eb466a5c77c805167be07851bd9e459b50bc6e3b549d3fc02d961af99cc5403b

Observation 6619ccf9-cac6-4c3b-a27e-5dcf43a1f0f7 · outbound

This paper cites Deep residual learning for image recognition,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Deep residual learning for image recognition,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.965878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:14.906924Z digest=sha256:0478c74655dbe2fa181ca69f311f3a310b045359c843b13c27268b1991e3b23b

Observation 9dac5372-388a-45fd-83ee-fafa1ae2fb26 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.956105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:14.996849Z digest=sha256:774a11e0883ff5fa1d39bf0370a9c3af225f1a40cf6e678fca901048e50f49e3

Observation cbe289d1-a07c-4257-b58e-5fff0805e3be · outbound

This paper cites Attention is all you need,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Attention is all you need,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.945286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:15.076869Z digest=sha256:545d1613216ea9c0fbd37faa19ba4e95a4830d88d0035661501eabf6cb19cb3a

Observation f708fa24-1712-4b72-a4a0-53afdca8449b · outbound

This paper cites Emerging properties in self-supervised vision transformers,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Emerging properties in self-supervised vision transformers,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.936097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:15.183668Z digest=sha256:dbdf27b84b1d362f8552ffa94a39d42a1b178e1c0e1ba76ed27ff8a7c1462f23

Observation 92be49f7-c9f9-4161-83cc-5667cf040cd3 · outbound

This paper cites Dinov2: Learning robust visual features without supervision,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Dinov2: Learning robust visual features without supervision,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.924042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:15.290363Z digest=sha256:6966ad945b24d7249a34e0fab1453c7da80a1ead21e5d6dfbdb844b53404b277

Observation 25aedd9a-1bc6-4701-8e1b-018f207f79f8 · outbound

This paper cites Sam 2: Segment anything in images and videos,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Sam 2: Segment anything in images and videos,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.913865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:15.436149Z digest=sha256:56b886d1d2d0a558e54f98e5f75bc1820e6092b2c399669ab3cdcec372a4e98f

Observation be12a41f-a993-416a-a2ec-6610f7d176a2 · outbound

This paper cites Sclip: Rethinking self-attention for dense vision-language inference,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Sclip: Rethinking self-attention for dense vision-language inference,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.902701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:15.507929Z digest=sha256:17bd953178e4fc9c3d650491afb9ab3967987f984b1f6274c81cd0c942698e7c

Observation bd6458fc-e7a9-4dc9-bd71-68a633858bc7 · outbound

This paper cites Explore the potential of clip for training-free open vocabulary semantic segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Explore the potential of clip for training-free open vocabulary semantic segmentation,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.891619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:15.616014Z digest=sha256:a40a25d2d6e2af36e5a7a75df498204ce1c9080502bca173e7e499eb987982d6

Observation 5dd4cad7-b9f6-48e7-9757-c62281e5afcd · outbound

This paper cites Clearclip: Decomposing clip representations for dense vision-language inference,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Clearclip: Decomposing clip representations for dense vision-language inference,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.881675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:15.780654Z digest=sha256:0320c31679f1146396821c1d2a427ebee34e5cb79aab1499595f8f81fcc91632

Observation 4b4c6637-59b2-4460-949b-50c2aaf538e0 · outbound

This paper cites Clip-dinoiser: Teaching clip a few dino tricks,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Clip-dinoiser: Teaching clip a few dino tricks,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.871784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:15.922960Z digest=sha256:db95293f2bbba5a5444e13cdd8ec0aa0d2431dcb287fd134b00a286d84558fcb

Observation 41efa2dc-4218-4e97-864a-f07ff8cdb94b · outbound

This paper cites Diffusion model is secretly a training-free open vocabulary semantic segmenter,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Diffusion model is secretly a training-free open vocabulary semantic segmenter,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.862851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:16.113871Z digest=sha256:cea1fe31cd5b13dbbe2798c85ad4aa52482333418b50ec23830f1b4e1b25fd61

Observation 2ba01bef-113a-4131-ba3e-c5c23fb7c13a · outbound

This paper cites Dynamic prompt learning: Addressing cross-attention leakage for text-based image editing,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Dynamic prompt learning: Addressing cross-attention leakage for text-based image editing,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.852800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:16.260625Z digest=sha256:754f1398281fab7a8350484f4ab7af2ffab69166dc3f3b1cba0e9f6e4aaf6860

Observation 600b1258-c5b5-49a7-8d24-4f7667c938e0 · outbound

This paper cites Cliper: Hierarchically improving spatial representation of clip for open-vocabulary semantic segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Cliper: Hierarchically improving spatial representation of clip for open-vocabulary semantic segmentation,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.843352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:16.380258Z digest=sha256:685c74ebf083cc11a50d38f5e90c3665a82438b2c77c0d63ca6179036b75272f

Observation 2f936345-1d95-44fb-b4dd-c5bcd21ee3b3 · outbound

This paper cites Silc: Improving vision language pretraining with self-distillation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Silc: Improving vision language pretraining with self-distillation,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.834223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:16.477605Z digest=sha256:9d3c8bc5167d4b6eede0ddabe721fcda5fd8cce1a3ef045c25f3028ffaec5f3a

Observation 79b85762-b0b9-45ce-8510-9b50eb619c28 · outbound

This paper cites Exploring open-vocabulary semantic segmentation from clip vision encoder distillation only,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Exploring open-vocabulary semantic segmentation from clip vision encoder distillation only,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.824247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:16.681775Z digest=sha256:bf32283ddb105a81bdd62244c59ada32ec0263b48c9de9348a5f501da0ad0a75

Observation d5193b3c-ebda-45e2-bbfb-feada88a460f · outbound

This paper cites Mask r-cnn,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Mask r-cnn,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.814331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:16.782295Z digest=sha256:d51148c41352ee30e7ef827945ce06f1b56833d40ded104741f1a2dca1c22ce0

Observation 205d5565-0b63-4285-ad22-aede1d0d3094 · outbound

This paper cites Relational knowledge distilla- tion,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Relational knowledge distilla- tion,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.804002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:16.829689Z digest=sha256:e33d8bc005cfc8043dd5234fa80225d90b297bd24d748cee05cc3d6675742fd4

Observation 5277c74d-472f-4a68-bc85-9d836843269a · outbound

This paper cites Microsoft coco: Common objects in context,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Microsoft coco: Common objects in context,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.796031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:16.894190Z digest=sha256:6ff4c200e1e7c03195cc3b86430dfa8d237dc0dd14d758630098f4ae180d9e14

Observation 972f6476-b4b4-464e-84df-73550f4a0523 · outbound

This paper cites Decoupled weight decay regularization,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Decoupled weight decay regularization,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.785855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:16.986846Z digest=sha256:267f98ba22c5daf437b0174a58bbf239d4c03873de55508fc7db2e6ade999375

Observation c5f7be36-a448-4687-b029-a22148928726 · outbound

This paper cites Eva-clip: Improved training techniques for clip at scale,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Eva-clip: Improved training techniques for clip at scale,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.775193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:17.061374Z digest=sha256:6b4b8462a0db3c374bf022e9b00e938805c676c99d5bec2929da2eff905ffb0a

Observation 49b7510c-7d13-4880-83a3-33a7159b7b58 · outbound

This paper cites Vision transform- ers need registers,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Vision transform- ers need registers,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.765935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:17.191747Z digest=sha256:337a4bf088ef23938f3c38971bd5bd08543f7f5d35afa50306ea121dd952e72a

Observation ac1612bb-55dc-4acf-b6c9-47460689030c · outbound

This paper cites Language-grounded indoor 3d semantic segmentation in the wild,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Language-grounded indoor 3d semantic segmentation in the wild,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.755803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:17.276128Z digest=sha256:5faf136bb21e5aedac8b06a234ad3ec918b566bc3f0fa9b410ec5203b4c61f23

Observation 5275d7da-9d4d-4006-9db1-1ea7e7fc4c0a · outbound

This paper cites Isbnet: a 3d point cloud instance segmentation network with instance-aware sampling and box-aware dynamic convolution,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Isbnet: a 3d point cloud instance segmentation network with instance-aware sampling and box-aware dynamic convolution,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.746015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:17.352021Z digest=sha256:c183c9142b69f8823c7db8b4290db82bde983ff76741a6a287c169367d4d21a4

Observation 987cc53e-740e-42e2-923e-af661b8ca986 · outbound

This paper cites Mask3d: Mask transformer for 3d semantic instance segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Mask3d: Mask transformer for 3d semantic instance segmentation,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.736986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:17.411695Z digest=sha256:6070ea5c9608d3c4972f4c3e75d575d23e027106c341d4628bb3def2cfa19151

Observation 9e692ed7-6df9-4e5c-81c2-946d6f8602a6 · outbound

This paper cites Openscene: 3d scene understanding with open vocabularies,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Openscene: 3d scene understanding with open vocabularies,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.729040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:17.473786Z digest=sha256:e41d10b095bbb707d67ed881a6beeb4b95d6b233d4b8cbe7176f305c40e9f9f8

Observation b2983b14-f286-460e-be1d-1c5dd73729b2 · outbound

This paper cites A density-based algorithm for discovering clusters in large spatial databases with noise,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception A density-based algorithm for discovering clusters in large spatial databases with noise,

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.720545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:17.540970Z digest=sha256:a4609de1f161f9d3be3e5a096fdba555217630e8ecb4c00baea8bb75c8d96d75

Observation b13a1cb6-7574-49f7-a680-3c51c19c1489 · outbound

This paper cites Openins3d: Snap and lookup for 3d open-vocabulary instance seg- mentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Openins3d: Snap and lookup for 3d open-vocabulary instance seg- mentation,

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.711917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:17.680891Z digest=sha256:d461580244c766896187f4817bf73594f91460e82172107828dff6d93784c123

Observation 7de92333-3e91-4dfd-b235-963bf0d22236 · outbound

This paper cites In defense of on- line models for video instance segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception In defense of on- line models for video instance segmentation,

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.696296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:17.852582Z digest=sha256:4d646b9c35ee0879e202cd0793b724d2488a193fc9adbb01e8f456e736b87c7e

Observation 7efe7155-1222-4a61-9e89-ad21db25f6c7 · outbound

This paper cites Simple online and realtime tracking,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Simple online and realtime tracking,

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.688453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:17.953660Z digest=sha256:56125f58238e9675457ccdd0ddc76e2d369fdf177872fe0e8a5cbfb5bab746b8

Observation f8650d5b-ca4f-4e2c-9d15-bb844a8fbff5 · outbound

This paper cites Opening up open world tracking,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Opening up open world tracking,

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.678693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.020767Z digest=sha256:efa63b01ae1f8eb7fa895b52e6dd6006c0a4d5852a2bd747326fbc28a912e966

Observation 1bda6104-ec21-4a83-bebc-6bacaeb08a6f · outbound

This paper cites Xmem: Long-term video object segmentation with an atkinson-shiffrin memory model,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Xmem: Long-term video object segmentation with an atkinson-shiffrin memory model,

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.671067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.097949Z digest=sha256:5e2ba162c1db036a08e69ccc3d354409b11ebeb3e5963eb9aafb3995f633605c

Observation ad7dbb41-7e45-494e-8aa1-47daaebdad62 · outbound

This paper cites Towards open-vocabulary video instance segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Towards open-vocabulary video instance segmentation,

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.662817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.158794Z digest=sha256:c5565fa281fe27dc2cc6d36a0ff9ac7189a8b14910c481b8cca78f396e934313

Observation 5e01609f-edcc-41c0-8839-39878b1d24b9 · outbound

This paper cites Video instance segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Video instance segmentation,

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.704463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.232954Z digest=sha256:edf8e6dc244b0944b91e2e4fde6312893f38a402cf5c2dc76ee3d13c3554227d

Observation 3c4543e9-4a42-49c0-99cb-81610213c364 · outbound

This paper cites Lvis: A dataset for large vocabulary instance segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Lvis: A dataset for large vocabulary instance segmentation,

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.654980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.300865Z digest=sha256:be8d4fecea7984d143b6e7656be8faeb914ad1499918dfcfdd83bccdb21d35d5

Observation 33b1d745-9ae6-4fab-91c6-5b7a5caa95db · outbound

This paper cites Occluded video instance segmentation: A benchmark,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Occluded video instance segmentation: A benchmark,

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.647349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.362458Z digest=sha256:80c9bc74284fc7b66eb49e00b034dc7846bd06f71bffa05252ec9bffeadd6dcc

Observation 897b2fb6-e632-4d5f-aed8-d1bf679cb521 · outbound

This paper cites Burst: A benchmark for unifying object recognition, segmentation and tracking in video,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Burst: A benchmark for unifying object recognition, segmentation and tracking in video,

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.638704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.428705Z digest=sha256:4ae6ff466e624244ea042cdfa7e15925dd169a25288a0f2c39a86a760b6da41e

Observation 6f2340bc-7af9-4092-9e04-cac822a5bb02 · outbound

This paper cites Fs6d: Few-shot 6d pose estimation of novel objects,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Fs6d: Few-shot 6d pose estimation of novel objects,

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.629572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.508453Z digest=sha256:aa3ec83482d3dd7c428af6bb278b1b33f0b4a6b6f71183a82df2fb9c360316a5

Observation 11d3065d-c56c-43d4-86b7-dc4ee7ebef54 · outbound

This paper cites Semantically-enriched 3d models for common-sense knowledge,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Semantically-enriched 3d models for common-sense knowledge,

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.621276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.606981Z digest=sha256:b5b06b24f9627f944e4839e4f75effc327b20a48a7dcc41abae458afaa7b2374

Observation 400933f0-9f5a-40b3-9f94-ab77701c2ef2 · outbound

This paper cites Normalized object coordinate space for category-level 6d object pose and size estimation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Normalized object coordinate space for category-level 6d object pose and size estimation,

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.612037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.686622Z digest=sha256:7ecb9ffe6ecd586513f6556d949a3001759c43eaf343c414b5298cc239bb6417

Observation aba4668e-f6a4-421e-9de1-5324e94dbd0f · outbound

This paper cites Bop: Benchmark for 6d object pose estimation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Bop: Benchmark for 6d object pose estimation,

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.604203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.797052Z digest=sha256:f24b7fe264f834458eb23e22042bb81ff60c522a06b79014022fb2605b11afbf

Observation 15346383-ffe8-4f89-990c-41b4c8cb622e · outbound

This paper cites Bop challenge 2020 on 6d object localization,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Bop challenge 2020 on 6d object localization,

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.596699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.894916Z digest=sha256:e950d14ead2b3c75f3c293c3828f733dcda112d1dbbe820d7b9ff8901329515c

Observation 2c6b7203-98e8-47a2-815b-c7245133b94a · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted win- dows,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Swin transformer: Hierarchical vision transformer using shifted win- dows,

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.589444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:18.985082Z digest=sha256:49c18c42f890fab02cb3b88f5df7d4db36cff9f91692ba4f67661ae3341750ec

Observation 9fbd2366-3de6-43a1-b8b5-861ae8405fb9 · outbound

This paper cites End-to-end object detection with transformers,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception End-to-end object detection with transformers,

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.579542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:19.045927Z digest=sha256:3858235aa11634a864d3c004dd10dc09b29080f59f31362cfabeb27a31576234

Observation 40a25b6f-db23-4027-955a-9d5abd0cdeea · outbound

This paper cites Region-aware pretraining for open-vocabulary object detection with vision transformers,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Region-aware pretraining for open-vocabulary object detection with vision transformers,

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.570341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:19.131544Z digest=sha256:87fc972ccb29b1a8dce4ff22202814b0a72a790fb14bac6c65c0336d976307f9

Observation 2fba74c1-a24c-448e-83bc-cfa1b4c1fc96 · outbound

This paper cites Contrastive feature masking open-vocabulary vision trans- former,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Contrastive feature masking open-vocabulary vision trans- former,

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.562591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:19.218112Z digest=sha256:980ca577f7e831ea644444771c81b3dda4e12edad2db496aab6c50e0ff1477ec

Observation 941618b2-67e7-4535-9552-166819674ef0 · outbound

This paper cites A simple baseline for open-vocabulary semantic segmentation with pre-trained vision-language model,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception A simple baseline for open-vocabulary semantic segmentation with pre-trained vision-language model,

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.547138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:19.393386Z digest=sha256:60a743883dfddfd21ce292936f7876f21a0d6179979cdddae7ba9797320f3f26

Observation b58ed32b-0a2a-4836-886a-5f7a700fb054 · outbound

This paper cites Side adapter network for open-vocabulary semantic segmentation,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Side adapter network for open-vocabulary semantic segmentation,

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.538678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:19.456741Z digest=sha256:85399a9f94e3bf0a0d38fc70362d1c8ae3f679d976ee8ecf5b6688ffe094d7b5

Observation b08179dd-ad6b-48fa-bdd1-4767272c1b8f · outbound

This paper cites Open-vocabulary panoptic segmentation with text-to-image diffusion models,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Open-vocabulary panoptic segmentation with text-to-image diffusion models,

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.528893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:19.565524Z digest=sha256:47761c7aad18ed83c6a2a65c1ba9e24c3a18a403fbf2dea8f0ece2a0a0bc2c78

Observation d6babb57-d5f7-4925-be5f-4a8e1d18117f · outbound

This paper cites Convolutions die hard: Open-vocabulary segmentation with single frozen convolutional clip,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Convolutions die hard: Open-vocabulary segmentation with single frozen convolutional clip,

Reference 101

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.520963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:19.641618Z digest=sha256:4b8529a172cebefdd361d4ff939368a8930f24eea8b6f33c3000a7ad76f68008

Observation 7482feaa-5a78-4153-a5c2-3bf695e44b61 · outbound

This paper cites Extract free dense labels from clip,.

Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception Extract free dense labels from clip,

Reference 102

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:10:23.511695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:10:19.724032Z digest=sha256:3e5fcb96ee9d16638b255d09cefa584884c073476c8a47d52f92dadbc85d5cd2

Pith citing papers

Observation f308d82d-f00e-4cfe-8560-93a223e5e313 · inbound

Beyond Binary Contrast: Modeling Continuous Skeleton Action Spaces with Transitional Anchors cites this paper.

Beyond Binary Contrast: Modeling Continuous Skeleton Action Spaces with Transitional Anchors Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:51:03.571495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T04:32:04.547197Z digest=sha256:880b10d682c6ae609b7b457b071888267c0d2ea6c852756974c12d59aa4772ab

Observation d2be249b-5a9f-4509-b7a5-0bb8fe90acfc · inbound

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech cites this paper.

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception

Reference 105

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:19:03.217532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T21:14:58.814362Z digest=sha256:b714a29dddadd53b3bb57aa091d26ca2d5ef0b4c6e6f9e809d8e11d4bfa88790

Observation 0d581b57-d70b-4243-b350-5a8155b7ea8f · inbound

Dynamic-dLLM: Dynamic Cache-Budget and Adaptive Parallel Decoding for Training-Free Acceleration of Diffusion LLM cites this paper.

Dynamic-dLLM: Dynamic Cache-Budget and Adaptive Parallel Decoding for Training-Free Acceleration of Diffusion LLM Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.301224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T13:27:45.796650Z digest=sha256:315e90aff9919aafbdc01e0fad8cd669424e5211ef621986d804035bddd93d92

Observation 2648fb87-7a24-4a22-bfd8-9cbc79320386 · inbound

DC-Leap: Training-Free Acceleration of dLLMs via Draft-Guided Contiguous Leaping Decoding cites this paper.

DC-Leap: Training-Free Acceleration of dLLMs via Draft-Guided Contiguous Leaping Decoding Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T13:42:25.796548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:42:25.796548Z digest=sha256:1559867be257ae7a730ad235660e9104b0e40470457532ad398b06a6742eed22