Pith. sign in

Paper Citation Record · LEDGER

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation

As of 13 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 0 inbound Pith citation observations for arXiv:2607.02922.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.02922 v1

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T06:06:47.233814Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

86 of 86 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved86
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d4f6b78e-f922-41b3-abab-2b27e0411ec9 · outbound

This paper cites GPT-4 Technical Report.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:3608233b43031c3ce75ccff747db751bfa983183f020dedb2109faa22c1a3ab2

Observation fd14e178-fc05-4119-afee-5954dfb61084 · outbound

This paper cites arXiv preprint arXiv:1412.69801412(6) (2014).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation arXiv preprint arXiv:1412.69801412(6) (2014)

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:bf894838af94bba86c58140fe06a1c8ac905650cdc8fa49ab15f8bcfb84830df

Observation a731ca61-a436-431d-9f5e-44dc3a6c75bb · outbound

This paper cites Machine Intelligence Research (2026).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research (2026)

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:19d8916db7e1caaf2e2a80c0716826924f3838f10af1fffa797834b7315bc331

Observation b43e701c-f674-426e-b1be-b193255ae447 · outbound

This paper cites Machine Intelligence Research (2026).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research (2026)

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:97084c2ab01cbb3dfb5777006fff2d06b007d42bbdb49f51ab14856c74b9c988

Observation a985ebc9-90f4-4304-a321-f9f016748f5a · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:0dd7c354d6a7cbc3fad6a4742e3a2d4bf068f33f31f7e6eaa6a7fcce508e480d

Observation e9344be0-e3fc-440a-bf3a-297b14c07ac1 · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:6ecf1e02dfae64789d61d58865493cf5ff68694b832afc3ca94f0d43ecb3857a

Observation 43e53513-0d5a-4527-b976-2a3b4b884d8c · outbound

This paper cites Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:a7e6b34f1e53bc00a3d3f844ae71ec290452ede00f6e0800d7948f7318e8fbc7

Observation 286409b6-3575-4587-9819-987bf083cbab · outbound

This paper cites Token Merging: Your ViT But Faster.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Token Merging: Your ViT But Faster

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:6a5c1d4fdd1d46ead0c6c4fc125ce7d6361d60d152be3552fdd9413bdf8bd567

Observation 49536b96-5fe3-4f53-82ea-8fa6a26f86de · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:2fe2762b6f9f47f0d022f37287cd96fbfcb0c833514beebf9f504bb0ca167de2

Observation 60d65879-347f-4703-ae19-0bdd6bbc6dad · outbound

This paper cites Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:07fdb33a5f67d0fae4dff6af8623265cdb2f9d35a44f368df6abbd67952ef03b

Observation 2273dcbc-1b9f-4414-873f-eea92e22b690 · outbound

This paper cites In: ECCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ECCV

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:5a2d7c6091ddec139465954dd3496456e91b1dfd81f29abb240894651c156fe6

Observation e5a4a78c-c2e6-4cd1-b2a9-ad8f673a58f8 · outbound

This paper cites LongVILA: Scaling Long-Context Visual Language Models for Long Videos.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation LongVILA: Scaling Long-Context Visual Language Models for Long Videos

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:77de2b757214b1accd45c1406ecc447113552ce9bc651e1dedc633811683f133

Observation ee914331-38a4-4e45-bce5-deeb1d576497 · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:84a80aa44ed916b81ac3bdb03fc46243ed8e59cd118ec271506ff8e5bf0d66ba

Observation 2428d069-2089-40df-9b83-285c6c5ecf9e · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:a027b93d9099291e9fda602f4eaa3f75c066a6659c4195ad82e1a02f0be49989

Observation 7c353ec4-af22-4fbb-a93c-8a486d8cfdcc · outbound

This paper cites Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:194f695ffdd2111eaf7079cd76b44cd569b6be310e2b2705e0750523316bd660

Observation d83652dd-aaa0-4204-acaf-6e0c470f5855 · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:af5e5ca56808774bd7f72528fcf954e6d0261c40a7f08d8b4f72f5fc3656b412

Observation 86b087dd-01bb-4cd4-a2dd-e66afb5d88c4 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:cb28d4d042aa805d1993dc3ed1646a6298eda5043393ab21ec58c7284d36785d

Observation f6cd77ae-5788-4da2-ae8e-044d04b1e68c · outbound

This paper cites arXiv e-prints pp.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation arXiv e-prints pp

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:1e11efb5fc99809a4017a856a7406e62d74c6fa9fb3873c335f7fc0229afc73d

Observation 58980e28-8142-4b70-ada0-6d47a6e9d725 · outbound

This paper cites Machine Intelligence Research (2026).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research (2026)

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:0ec7b7ec713a7c8a59aecc9496f6c4d174d146559051e3d949488a025fe052e3

Observation 839d4958-a5a6-42d6-90d3-5e302b800bc3 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:2d6bda523fe1370c80ea3e571f88e3285920e8bd7758392b89c1a6588cd0cbcc

Observation 28d79346-3a6e-4693-ba8c-432a47844da1 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:2ccfe4bc26c214c965d8b5351386f6f7eb532512b6370510c5e3e0207cd94724

Observation 6d6050d8-1e45-4a09-b003-b4fc68496778 · outbound

This paper cites In: First Conference on Language Modeling (2024).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: First Conference on Language Modeling (2024)

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:6af0963847d627174a168366eb85c1b511c83be258305c51ff27b8782b48a198

Observation cf3a40f1-d2d0-42db-a3b5-6523e2ea6d9c · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:33ba06a21ff1bf0198cd27fb57f4551e4370f8d947610743a2ed21ea0430bd55

Observation 13095465-ae21-4a06-a09f-14e954dfa0d5 · outbound

This paper cites Efficiently Modeling Long Sequences with Structured State Spaces.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Efficiently Modeling Long Sequences with Structured State Spaces

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:c200614b5f786b15cd89386461e8db507cb12b342ae21b8dd6652c4e8de5e027

Observation 39dc3970-31d6-4ab3-8225-8dac56440536 · outbound

This paper cites In: ICLR (2022).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICLR (2022)

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:769941db57b5b874865095c0b0f352f056941e540200600f8fa67cf780c0f108

Observation 8be4e71c-fc84-401f-abc4-486ced89d171 · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:29bd3f8c3fba1d9c80f7322319d54559739e672ec347c77f543bca6478d52ef8

Observation 7f46380b-9f22-48d3-85d4-b3e7202a9ec6 · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:3f6c9e45cd986c6cb9f44ea69a38a9b0011e6b87bed38a73066caeca284bbce8

Observation 91864828-bab9-45e9-9bf7-a21edaed84d4 · outbound

This paper cites Machine Intelligence Research (2026) STAC 17.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research (2026) STAC 17

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:110bf51f82067a551b7432d58aaa2bfb2c60b37a463e00326941ea4ae488c4d8

Observation 0b37b6cf-bf96-4d8b-b8d6-9dacca63d7c1 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:f8aa1dde8518eaeb1da755b403034f6d3ce663e569b432a847c96b79b6a7990c

Observation 5f7d59a9-f7df-4e15-b2c1-8a903d976df7 · outbound

This paper cites In: ICLR (2022).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICLR (2022)

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:73b77e4b80f6b4793cbb1efaec03161252c9997025a4d3b2eb910b75abc18ddf

Observation 66d601ba-6c07-44d3-af1e-3ff8c8562a84 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:be38773ace8525a27fd25fe39750afd544b84ff39c39633c69af4c2a66ca3854

Observation 5434a24e-294f-4761-a24c-9514f4e5b6b0 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:bf56e6e028fb8ac969153830a39a7d7d1b5f2eeb7c31de84f9a2a92fc5ca41e5

Observation f8435b2d-3802-4302-800f-b4156d23f27b · outbound

This paper cites an unresolved cited work.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:c8005a1400ba781dacaa04cf8354af601de5f6164fb008cbb450dce186855137

Observation d95f45dd-7816-4adb-838a-ccf224d39e04 · outbound

This paper cites Mixtral of Experts.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Mixtral of Experts

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:b659ffe301ffc32368d78199b6f063bd87aba3f735d19bbd217a0843d7a809d5

Observation cb8fd7a6-1ec7-4bec-a92a-cc179f284e2d · outbound

This paper cites arXiv preprint arXiv:2503.04130 (2025).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation arXiv preprint arXiv:2503.04130 (2025)

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:c4dae724546c285607ef75a1b1c465b9b881a90875c0a8e58f236eaec210eac5

Observation b6c65628-7cc4-42a5-80d4-65265b56df39 · outbound

This paper cites Journal of Basic Engineering82(1), 35–45 (1960).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Journal of Basic Engineering82(1), 35–45 (1960)

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:1200d3be60843c4d885665b50d46d0a2b3c146e642798bde9d2620ce32ea79c6

Observation 461e5042-f45e-45f8-8ec3-bbba9c00e7e7 · outbound

This paper cites Computational Visual Media11(3), 655–667 (2025).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Computational Visual Media11(3), 655–667 (2025)

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:9dc2902c811d4a33cf05a67ce8d85a7885e54494916a3b1ade3c9dbc951bd6bd

Observation 07d234a6-566e-4f88-aab3-4927bb33ded9 · outbound

This paper cites In: Asian conference on computer vision.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: Asian conference on computer vision

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:bcda6b67ed6d7edbce1e87836f73a47dd84b3796ddcf3a96f15515fa3fa85cef

Observation ad9c4785-e53e-4274-88b2-db8d0ad6a5df · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:6aae96914d63d0ef75a89d8e16cdd098b50dc602e637b700e24b3a3c64f48b52

Observation 50c01312-37ba-4f83-9556-43bc100d8206 · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:ed85fa3348798840c3887d2542fd2f4152e2a81c89f387eab0f2f19453025ba8

Observation 44799c7f-5fff-4b2c-ae49-e3e226033daa · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:4e3db3d8a2f5a70a299279400054d69df5f817c390a7a8452d230645a76a9d76

Observation 1a136f9a-79bc-4341-a9d8-7663fd96a47c · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation LLaVA-OneVision: Easy Visual Task Transfer

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:37add7ce7e8c4b809ad7513a37034c81c135ee9a4c780df423d760c506821233

Observation 3be3c91e-b7b6-44b9-98fe-723c075550cb · outbound

This paper cites In: ECCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ECCV

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:84cbc17e37288581113a7ae3fb85e04339d77ee01af55eae3bd9494d9e247f51

Observation 52ffdd98-43f9-447a-9a6b-11525769d7f8 · outbound

This paper cites an unresolved cited work.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:3b1538e3c65cdb5c9ccd55d771287d06032e7c066ae09e6912213284e6343302

Observation 0aa1c8bd-29c0-4fc9-9186-2e06fef0cd9d · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:c9b9534448e7a68d2e2db1f64c21e846da71baf7d1c6d96b1c5c9e8287ce48d9

Observation fca6b943-2e8b-4f21-8027-042bedadfe68 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:a4fe7b682ce10c1f29af10a77ef0417b69f1237389c0a9e314f39af7c2ba881b

Observation 19bdcc5a-64fa-4aae-a66c-1a1086274fe5 · outbound

This paper cites an unresolved cited work.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:0555cf9186d576076c7023f2fcd3a82df0e688f6aec6156d4a4af09d2c643f52

Observation e73b5dfc-90ce-4da9-bcaf-296b92770c4d · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:9253af8fb2907e9442c72af893193e780fbf9a046f24fcede1d02b76af5ddd8b

Observation fdd9ddd9-0a28-4f01-9f4d-50e2fbfc5cdf · outbound

This paper cites NeurIPS37, 103031–103063 (2024).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation NeurIPS37, 103031–103063 (2024)

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:c591030573dd739ea9e78e6e637675758c51513928442af1112afad75c040f38

Observation d14f4efd-bb8a-4503-ac09-5e79097118ee · outbound

This paper cites Machine Intelligence Research21(4), 670–683 (2024).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research21(4), 670–683 (2024)

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:6f4b52d67667d0675d9b91454e58c6386141e7cdd46254c686f89b1f94f40bdf

Observation 7b2ce9c1-ab38-4dd4-ba6f-2c5e01d42977 · outbound

This paper cites arXiv e-prints pp.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation arXiv e-prints pp

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:a557d6d5cc5bdb30fac2c313095560d10b230b551065cc6868f0efefa1b24f72

Observation fe4f295d-3fd1-4abb-9653-153300840e30 · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:92b90e3d4145f9db366252d29d80d21a9351ed97c09cb7dd1a9b3fca02991a8f

Observation a216cdec-57c7-42b8-b8f9-1d092d8cb6a7 · outbound

This paper cites In: Proceedings of the IEEE/CVF Interna- tional Conference on Computer Vision.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: Proceedings of the IEEE/CVF Interna- tional Conference on Computer Vision

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:befa1c1641085cfb7d53f3dc790120f641d90a45e85c989bbf4a943314a6b637

Observation a02ea75c-7acd-46fc-861e-5f796ca732c7 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:d206aee01d81451f3d8712e058ddb8a88049ba886e164598dc882f5b6e67078a

Observation 3ce13bc8-a305-4cf6-b2b4-3c4e22d205cd · outbound

This paper cites Computational Visual Media12(1), 71–84 (2026).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Computational Visual Media12(1), 71–84 (2026)

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:3bd43c86e7a58ad5bc391a757e06fa01b151ae1a181093b9bda84fcf00fdf188

Observation e57e284a-daf4-49c6-8eb7-e0d902de0f7a · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 56

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:f5df9c11ce75d20e90f1e02669c1251e550c9dc7022c3130b135d5ae46c25e46

Observation ce403898-ed40-42f8-a75f-c14eccd175b4 · outbound

This paper cites In: ECCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ECCV

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:2be2cd234d1b0d19c5ebd5583fa417821a08a6bf5bc22ee7fbd17f83ad8cf6ff

Observation 99e6b471-2162-4ab1-a085-5dd6ecea0bf9 · outbound

This paper cites In: AAAI.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: AAAI

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:4b0cfcaaa8f8a6f59cc0bcde72f784af061136e2e197e413c3c0c8c3e2947dcc

Observation e5a10d60-7622-49ec-be24-312751ef4129 · outbound

This paper cites The 2017 DAVIS Challenge on Video Object Segmentation.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation The 2017 DAVIS Challenge on Video Object Segmentation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:617ead669c73fd2b500e87bcc94117e3882befc285b79a9b0c62b88a6824b931

Observation 75917728-ec17-4d9f-a731-f863bc70975d · outbound

This paper cites an unresolved cited work.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:43d4ccd41b2b304241dc3de593fcc616723a84d0f17aaf8bfc8e43da37af9182

Observation 6575c041-ef40-4a1d-a2bb-b7e44731f92f · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 61

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:bc52e12539aeca80b91a63b850f8433bc0aea8870b0ca6995068cf8b2793dfd9

Observation 622741c0-e100-44a9-9d93-a76f3c5d7186 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:e140c164e527b89f0cd4bd4efcb2268ffd10cd73de322f8b28c01f3b426eac94

Observation 20f518b8-8b32-4cc0-9ed3-94ebeef0aa29 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation SAM 2: Segment Anything in Images and Videos

Reference 63

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:b7226f8ae0067ae6e30008133024ad772404859e9a1bc370979fb343d7397578

Observation 702d9478-8f9f-4295-8b22-b14605d92368 · outbound

This paper cites In: ECCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ECCV

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:ec98341e2287b97563aa43784dfe3b161b7a1b14f84c1216abb45f9f147e2977

Observation 54e0cb18-43af-407f-b79f-dc31dac84aca · outbound

This paper cites TempMe: Video Temporal Token Merging for Efficient Text-Video Retrieval.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation TempMe: Video Temporal Token Merging for Efficient Text-Video Retrieval

Reference 65

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:771830192b1a4a2357b701cf3991efa8b9f99acd8f2f1337a02a377e01bb16eb

Observation 6b9a8a68-1aca-4181-a66f-ff7e7e52d9f7 · outbound

This paper cites LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:05ea09bd21dca63a86c2fb37847c8fa659352d249ce433b5ba08e1f3ce38c78c

Observation b695f4fe-31ef-4978-bec3-337a0f1036c8 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:13a483f634442b792085b2a485b2f10ef27d37e92c5259d8f1a08ed0338b929b

Observation c2dab446-ef64-436f-a05f-c75209fd2426 · outbound

This paper cites arXiv preprint arXiv:2508.04369 (2025).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation arXiv preprint arXiv:2508.04369 (2025)

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:5a72b092a3dc2d1ff875133c604192658d5c2dc18407dc71f70426b92284a84c

Observation dac7d91b-fe0b-4f15-b4be-493fd14a03ee · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:35f79a0322b28080a9b91bd4678a076972e47f5328e78cb12afcd1b82d8aa7a0

Observation 7b85d7d3-ee20-421a-84cf-f9c4bade943e · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Gemini: A Family of Highly Capable Multimodal Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:6ad587219279ce420eb8be989c84f77708bee73f82a756d19496566b73b5b534

Observation 1a80582f-508d-4f15-9231-5a4d7d6deb9c · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation LLaMA: Open and Efficient Foundation Language Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:45a7165b38c1f3f5e25202d9552ef053d6f0846eabca6d21f5a04e3a45d05c37

Observation cbb240cc-5e93-47cc-9801-319ebd9e1b52 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:fb8fcc10f7c7ef6cf944f4bbf72997b367858b5af28afd2c22e93b8b37a9d34f

Observation 7189ef69-2297-4e59-a673-1eb616517dc2 · outbound

This paper cites In: European Conference on Computer Vision.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: European Conference on Computer Vision

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:3df1f5f025ba4bbdfaab55d747b13643669b157d10c58386155853be7e9be6b7

Observation 73e56558-2548-4938-968d-0b83975b64fc · outbound

This paper cites Machine Intelligence Research (2026).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research (2026)

Reference 74

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:28f8f2184573cd28955880262a5223d779c55da3b30c5cc2918b766987c2ee47

Observation c4f5f634-8189-4261-afb8-f4f95487a011 · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 75

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:9d5bd2078593cf2b555d9e9658cb7f2396799f484b5d157d853ba80e7e3c4e32

Observation b8a6e547-2ca3-4297-a687-b5c140103a6d · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 76

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:e5468e8a0090a7aa8197cd3cdf384cbc0c2a64ac6cc92487fa63a058ff1351e3

Observation ba25135d-c025-4de1-b363-f149987afac8 · outbound

This paper cites YouTube-VOS: A Large-Scale Video Object Segmentation Benchmark.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation YouTube-VOS: A Large-Scale Video Object Segmentation Benchmark

Reference 77

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:ad5bac8951dceb334811fd0adad19e1f196e4541fcd3e8ce1ea71445c225d35b

Observation b10c7c05-e02b-4bea-8777-3fdf15331f2c · outbound

This paper cites In: ECCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ECCV

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:2aae0d26aeae17660dea272adb8848e0079bb48ef0669e599e8807621b871879

Observation 7ed5a5f4-7b5f-4753-8be1-b4a8a65b32ec · outbound

This paper cites Vivim: a Video Vision Mamba for Medical Video Segmentation.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Vivim: a Video Vision Mamba for Medical Video Segmentation

Reference 79

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:b3ab5d51e99ac13a4209775678036f952a0c7296856e006ece6b2b7520373ead

Observation 290f63e8-2ef4-48ad-aabc-29890a19ab04 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 80

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:ebea3a34139244a44d539ff7c5473f32629b8bf461d609628f7ffd357f919e25

Observation f05d300c-997b-450d-ab5b-ea4fbd2a9409 · outbound

This paper cites Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

Reference 81

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:f2cdfe6636140984ffee67a5d6e5a7f9113ef661db80f40886044fae12e1dc6d

Observation 33889bac-850f-4aeb-bee4-9b6d824e4b2e · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 82

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:e18d11f9d137f50b62aa64ef30f46789db8942b6b6279649b5e9e4b4b235e3f1

Observation 8989a9ea-f1ed-4eb1-8c16-26734c29cd64 · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 83

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:2cb7633a24f08221efc1b42529080fba5b91122e5a484bc0f4eca580aa73fe0a

Observation e369217a-d568-4ef6-a4b4-d63108ace0bb · outbound

This paper cites In: AAAI.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: AAAI

Reference 84

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:51b32d65cf4282b4d40653baded3a7ca79e5c5b206825bdcd926194a7c8ba75c

Observation 8314b72e-9a63-4657-acf5-cebd372b1211 · outbound

This paper cites Tracking with Human-Intent Reasoning.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Tracking with Human-Intent Reasoning

Reference 85

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:6ada3e0099010cf8091f2941d27cddb3e3e2645ed932b7f8b7f034b87d7acfa5

Observation 796b549c-8f3b-487a-87dc-907d78ce2ac4 · outbound

This paper cites an unresolved cited work.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Unresolved cited work

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:ac8bcec72e43646217ea1c9621f317bcdb1c1d57efcf56c1c64d88f7a6ab00d7

Pith citing papers

No inbound Pith citation observations are available.