Pith. sign in

Paper Citation Record · LEDGER

Expanding Event Modality Applications through a Robust CLIP-Based Encoder

As of 13 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 1 inbound Pith citation observation for arXiv:2412.03093.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.03093 v2

Coverage vector

measured 69 of 69 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:50:59.205085Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:11:07.359765Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-12T21:11:07.571282Z

Reference resolution

69 of 69 outbound references displayed

  • verified exact1
  • verified fuzzy46
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5a3803d5-f844-40b9-b0e2-2d1dd701cf12 · outbound

This paper cites Unconstrained Open Vocabulary Image Classification: Zero-Shot Transfer from Text to Image via CLIP Inversion.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Unconstrained Open Vocabulary Image Classification: Zero-Shot Transfer from Text to Image via CLIP Inversion

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-11T22:50:59.554755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.885676Z digest=sha256:62133b6f397099c0943c56180ab68f02dfd4bff6469d9429fef80853302df360

Observation 21aaf855-2519-4707-9cb8-c6a27c331f11 · outbound

This paper cites Learning to prompt clip for monocular depth estimation: Exploring the limits of human language.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Learning to prompt clip for monocular depth estimation: Exploring the limits of human language

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.433416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.891175Z digest=sha256:3b8e8e704d349d1b130ea627d674fe1839384345947f5006f50998e567df69f4

Observation eb684b8d-ab9e-4dc2-9641-c8fa36a46e46 · outbound

This paper cites Graph-based object classifica- tion for neuromorphic vision sensing.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Graph-based object classifica- tion for neuromorphic vision sensing

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.417437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.895537Z digest=sha256:bcc7f126171ac22d01d9e412f841b2fb967e44b45a9dea5ec94714657df191ce

Observation 99e2be03-d3e9-4f23-9226-c565ca6f72d1 · outbound

This paper cites A differentiable recurrent surface for asynchronous event-based data.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder A differentiable recurrent surface for asynchronous event-based data

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:58.900169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:58.900169Z digest=sha256:68d5230ad461a940a219d8a1fd53c0ea81868693689fb7e71dbc3d5196458f05

Observation 26243ba6-08ed-428d-a34b-bdb8d6887a49 · outbound

This paper cites Recent Event Camera Innovations: A Survey.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Recent Event Camera Innovations: A Survey

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:58.904660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:58.904660Z digest=sha256:a6809c8cb81029503fda7092c05194a15a1ee6782d760beea38774d1d197c9d2

Observation aee9cbf6-abff-4961-bce8-536885420f46 · outbound

This paper cites Generic attention- model explainability for interpreting bi-modal and encoder- decoder transformers.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Generic attention- model explainability for interpreting bi-modal and encoder- decoder transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:58.909239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:58.909239Z digest=sha256:cd432848c890b4577efbfdabfd066f36cbe9ccaa3baec7aa9d5a12f20ef9a9d0

Observation 480333b1-dff8-459d-bfc2-3aa1abd5ae4a · outbound

This paper cites Image-based clip-guided essence transfer.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Image-based clip-guided essence transfer

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.382619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.913988Z digest=sha256:e1ffcd6572364e49a107dbc374b7b3db0a7fa272c6399df2896e9bb4d8d13064

Observation aa6fdf6f-8c32-4c59-b456-ed63bda860e1 · outbound

This paper cites Understanding Transferable Representation Learning and Zero-shot Transfer in CLIP.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Understanding Transferable Representation Learning and Zero-shot Transfer in CLIP

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:58.918866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:58.918866Z digest=sha256:dce03305aea2d0f85416c5ee45eb39689cfbb6be6dae648e8e4adf73c47c78e1

Observation 67ed22b5-5dd1-457f-979a-7a765117fb13 · outbound

This paper cites Transfer clip for gen- eralizable image denoising.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Transfer clip for gen- eralizable image denoising

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.367097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.924531Z digest=sha256:719f6ebd28476917b5606c6b73a05b451abba98ceadd73163b0285998f023ba8

Observation ae8588e5-f47e-4675-a62b-d32245793f08 · outbound

This paper cites Label-free event-based object recognition via joint learning with image reconstruction from events.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Label-free event-based object recognition via joint learning with image reconstruction from events

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.351572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.929215Z digest=sha256:bcab4990cd5d63bd8998e4a278ec0fd6cb70c93a8530572a001117d9129da8f5

Observation 821f10f3-40db-4eb5-9a7f-7e2762fc03f1 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Imagenet: A large-scale hierarchical image database

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:58.933715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:58.933715Z digest=sha256:2519b048eef80d9510d6f9b0395fc8148d1042072d24be12207886c6fbbd7d40

Observation 9ec5e25b-5e50-4012-8c8b-bc700b095cc4 · outbound

This paper cites Event-based vision: A survey.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Event-based vision: A survey

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:58.938395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:58.938395Z digest=sha256:9de2d088eb9e59ce0cdba1ae17e9571fec824772ac3218fbd22a8dd469e13800

Observation 6525283f-46b6-4912-a813-9c16659e70fb · outbound

This paper cites Clip-adapter: Better vision-language models with feature adapters.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Clip-adapter: Better vision-language models with feature adapters

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:58.942907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:58.942907Z digest=sha256:ea070dde405b2076bc7c60cf633641863a84e89b342f57b06d6cbeebc6d5dc8b

Observation b75a3667-2d90-42c7-a3e8-77ce2902b382 · outbound

This paper cites End-to-end learning of repre- sentations for asynchronous event-based data.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder End-to-end learning of repre- sentations for asynchronous event-based data

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.306212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.947666Z digest=sha256:7bab0f85bd5bed1875a6c4c151868c50baf990d6e3a81a77388c5c4a973fcd7a

Observation dff4b428-4070-449a-b297-5a12a913d13f · outbound

This paper cites Imagebind: One embedding space to bind them all.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Imagebind: One embedding space to bind them all

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:58.952187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:58.952187Z digest=sha256:e58f62517f36e4245901f93b832c9c6c2f2868e47d42dbe01fa7cc50df580f38

Observation 18402977-93d2-46b5-ab04-9e28f98c9812 · outbound

This paper cites Cyclip: Cyclic contrastive language-image pretraining.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Cyclip: Cyclic contrastive language-image pretraining

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.281357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.956599Z digest=sha256:9ee08842dc03609a7e175adf45bf22e58a0e5424d764f11462f0601a7cc234ad

Observation f8cc1f31-9f7d-4acd-99a8-3a3b7e68e90a · outbound

This paper cites EventDrop: data augmentation for event-based learning.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder EventDrop: data augmentation for event-based learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:58.961315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:58.961315Z digest=sha256:8a6cfc8e130c762158069599dce38f728914aa37572fda324ecffbedf1105447

Observation d05b3de5-a5e3-44e6-a307-84faf0566dbb · outbound

This paper cites Open-vocabulary Object Detection via Vision and Language Knowledge Distillation.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Open-vocabulary Object Detection via Vision and Language Knowledge Distillation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:58.966310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:58.966310Z digest=sha256:52d1dcf89d34bab4bd6b754071307883f57563679397a8c8635f02bea26c2d2b

Observation f10f3fc0-618c-4b23-b88e-a0ae84a32d77 · outbound

This paper cites Cliptrans: trans- ferring visual knowledge with pre-trained models for multi- modal machine translation.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Cliptrans: trans- ferring visual knowledge with pre-trained models for multi- modal machine translation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.267130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.971400Z digest=sha256:1bb7a27cbbc876e950676e3fc48cac75e269f63cea695f21ed8804c1496b616b

Observation cbb52e21-8882-4bb8-978e-77107628dd1d · outbound

This paper cites Audioclip: Extending clip to image, text and au- dio.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Audioclip: Extending clip to image, text and au- dio

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.251795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.976575Z digest=sha256:dd1fb46ff9fbe94599190dff5c554fbbc04aa3d536423d73c4e87d37de409d41

Observation 4ab15f38-f9fd-4267-b5fe-dbdcc8dfd211 · outbound

This paper cites Lidarclip or: How i learned to talk to point clouds.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Lidarclip or: How i learned to talk to point clouds

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.236514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.981508Z digest=sha256:9804f9810d7a0a0d3caf26648728d785713043efc7738171ff24c0bab8a78134

Observation e51c5ef6-3c2b-47c2-a15e-27c33a327757 · outbound

This paper cites Learning monocular dense depth from events.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Learning monocular dense depth from events

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.221553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.985914Z digest=sha256:ffe63463387df0ff2a829c5ea389109cd47dc3ab88ba9522421a775de3feb84a

Observation 21a750da-7840-47ff-a6dd-24be073fac6c · outbound

This paper cites Clip2point: Transfer clip to point cloud classifica- tion with image-depth pre-training.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Clip2point: Transfer clip to point cloud classifica- tion with image-depth pre-training

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.206923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.990327Z digest=sha256:b235440c055799aa2fd8174cc8c4aea92651af035f7ea9811550f61f9e44307c

Observation cb4d03dd-41ab-410b-afb9-ee46c30d5b43 · outbound

This paper cites Transferring pre-trained multimodal rep- resentations with cross-modal similarity matching.Advances in Neural Information Processing Systems, 35:30826–30839,.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Transferring pre-trained multimodal rep- resentations with cross-modal similarity matching.Advances in Neural Information Processing Systems, 35:30826–30839,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.192561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.994950Z digest=sha256:670f1795e63ee011794b85c93ef0f3ea36a6c6109aeb457becf520b73a922123

Observation 38763687-7684-490b-b514-2fa62c9a0fc6 · outbound

This paper cites N-imagenet: Towards robust, fine-grained object recognition with event cameras.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder N-imagenet: Towards robust, fine-grained object recognition with event cameras

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.177439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:58.999832Z digest=sha256:fe7daec1f3c7bfc824e0a808446eaca9c4622810ed747566583e474d69ba59f1

Observation aacfd93b-eb1c-46fd-895b-00de0a73496d · outbound

This paper cites Masked event modeling: Self-supervised pretraining for event cameras.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Masked event modeling: Self-supervised pretraining for event cameras

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.162450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.004538Z digest=sha256:2c9bad9c49bf3ee4daa9225e716b0a095b31383cd89485a72c5135b8831bcd5b

Observation b8e4a749-3b78-4b57-8dbb-540a101cc57d · outbound

This paper cites Graph-based asyn- chronous event processing for rapid object recognition.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Graph-based asyn- chronous event processing for rapid object recognition

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.146500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.009309Z digest=sha256:c441429eafa14700053671278d231ac7af4d9eadf6f92e07b11f4726961a419e

Observation 4b9817b1-5673-4b39-b5ee-65e15572118c · outbound

This paper cites A 128×128 120 db 15 µs latency asynchronous temporal con- trast vision sensor.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder A 128×128 120 db 15 µs latency asynchronous temporal con- trast vision sensor

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.130610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.014054Z digest=sha256:87bf7096c3907f81e49188a9592548c67f2da2b49d2c8d02bd19245b3f9da3ea

Observation ce31b631-7d20-47b8-9215-dff4762ae36b · outbound

This paper cites Fast classification and action recognition with event-based imaging.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Fast classification and action recognition with event-based imaging

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.115128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.019107Z digest=sha256:5f3e211962be13249722d31133f37ad605e83b1e2ba49aa57114c9a4dfca0ecd

Observation d76037cb-227c-4add-b0ea-d28d86e55093 · outbound

This paper cites Revisiting temporal modeling for clip-based image-to-video knowledge transferring.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Revisiting temporal modeling for clip-based image-to-video knowledge transferring

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.098636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.023635Z digest=sha256:d170b120636552e84d3f2c2798fc0c71bcec732bcfef5327c5b30ab17732c975

Observation c31ee535-63f5-44d5-b71d-4fc4b8911437 · outbound

This paper cites A revisit of sparse coding based anomaly detection in stacked rnn framework.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder A revisit of sparse coding based anomaly detection in stacked rnn framework

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.083593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.028113Z digest=sha256:9fb4a7f22b1a7e185716b4c9b892401774747b8dee6b151627ec0ea42008a84b

Observation 610a87d8-6882-4d0a-a88c-51ebcfba46d2 · outbound

This paper cites Event-based asynchronous sparse con- volutional networks.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Event-based asynchronous sparse con- volutional networks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.067564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.032607Z digest=sha256:09f90a2b77c22a144828d1a56ac147684be568181edd369b975096a61750a74f

Observation b9d7b78a-05da-45da-9e93-43dbc95d8ce5 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Representation Learning with Contrastive Predictive Coding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.036732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.036732Z digest=sha256:5d546ad31150f9a4c16cd2c4d9eddbb1f8577129dd6363953c1ed15dae6762dd

Observation e903d5c0-4ae6-4e1e-ba89-ce6e30a1c2ac · outbound

This paper cites Converting static image datasets to spiking neuromorphic datasets using saccades.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Converting static image datasets to spiking neuromorphic datasets using saccades

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.049682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.041271Z digest=sha256:e8ce5079075489c6f6a41abf3296f9a88517b9d68fe6433c04394b96d243841e

Observation ba2d2925-aacc-4ad5-9998-6987f24a9ff5 · outbound

This paper cites St-adapter: Parameter-efficient image-to-video transfer learning.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder St-adapter: Parameter-efficient image-to-video transfer learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.030027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.045553Z digest=sha256:fb5a34c824af7eb409c163260a1e3c0d2730ea88db46f09775a412dd73cd3e57

Observation a45bbe5c-8f3f-44b8-9287-82c7e912bc46 · outbound

This paper cites Back to event basics: Self-supervised learning of image reconstruc- tion for event cameras via photometric constancy.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Back to event basics: Self-supervised learning of image reconstruc- tion for event cameras via photometric constancy

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:51:00.003968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.049815Z digest=sha256:ef9ad9045a0bb11b6179b7d1cd65d6a083b3272b3980adfc0dbd00a7e271feb3

Observation 1c8d35f9-b3af-40f9-8798-e83fa468ea0a · outbound

This paper cites Ecodepth: Effective conditioning of diffusion models for monocular depth estimation.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Ecodepth: Effective conditioning of diffusion models for monocular depth estimation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.986866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.054048Z digest=sha256:a3ffc585cfe966b2c4c2641b15b230b1002660418d9cd345c80286709dd218f0

Observation 1d08db6c-b962-442c-a151-871459664118 · outbound

This paper cites Esc: Dataset for environmental sound classi- fication.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Esc: Dataset for environmental sound classi- fication

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.971263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.058301Z digest=sha256:1e71ce407ac86d195df745c6add28040e2a4ce92009a4c4f4c8d1534a771c042

Observation ecbe0db7-8e86-475d-83e9-a32fb68fb399 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Learning transferable visual models from natural language supervi- sion

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.062483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.062483Z digest=sha256:9280e6a61a89590cb953650d2b04c2ec687dfb66296cfbf438d15ea9988c645a

Observation 0be796d2-49a0-49be-8103-34442c6f9ae2 · outbound

This paper cites Events-to-video: Bringing modern computer vision to event cameras.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Events-to-video: Bringing modern computer vision to event cameras

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.945811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.066835Z digest=sha256:8562d03f96d7e00cb1adcede863e23ef0364de578ab4d0ec80999461e3744fb4

Observation 8143f752-74fa-4110-ac9b-5f1c3f1e5646 · outbound

This paper cites Aegnn: Asynchronous event-based graph neural networks.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Aegnn: Asynchronous event-based graph neural networks

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.930028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.071339Z digest=sha256:c672fdded490e014844cba358270dbb3e4f5486d050bfcb446bb1fa8a78aa54b

Observation c90db542-dcfe-4075-af68-8cd57408210d · outbound

This paper cites Towards understanding the modality gap in clip.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Towards understanding the modality gap in clip

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.911004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.075846Z digest=sha256:cf8588481d9cb835b1533d0b9b815941aacbeac58fe85abfd7399cbadd112876

Observation a8980ce5-4d15-4c1b-b3fe-8b58a2602538 · outbound

This paper cites Speechclip: Integrating speech with pre-trained vision and language model.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Speechclip: Integrating speech with pre-trained vision and language model

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.894007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.080601Z digest=sha256:759a92c30873ba33b690d62a075e3a6ee410dcb5b5f5e6d66de27a52c9e1bfd8

Observation 11182a22-d430-4fda-a36d-2a4509b13efd · outbound

This paper cites Hats: Histograms of aver- aged time surfaces for robust event-based object classifica- tion.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Hats: Histograms of aver- aged time surfaces for robust event-based object classifica- tion

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.877001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.085198Z digest=sha256:7e4d8e3cf557cce96a09edb98030d0cccade8bd0e7aab4e65c04db9e01f2120f

Observation b33dcb96-d29c-4b9a-868e-fbe91c00ce65 · outbound

This paper cites Real-world anomaly detection in surveillance videos.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Real-world anomaly detection in surveillance videos

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.859060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.090125Z digest=sha256:5fff30656e60565f78ddddb43d322e40ceabdc1c87e666124647c49d3f1d9c5c

Observation 263d39e2-db51-4d79-b7da-c8e824d18a69 · outbound

This paper cites Weakly-supervised video anomaly detection with robust temporal feature magni- tude learning.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Weakly-supervised video anomaly detection with robust temporal feature magni- tude learning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.841057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.094730Z digest=sha256:0ee63a0343259fe02e09d56c2cb03ef929eb5f0f8f0bcbaba2657dd37a665956

Observation 7983bc49-3237-428b-beeb-dc132f7648a4 · outbound

This paper cites Clipn for zero-shot ood detection: Teaching clip to say no.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Clipn for zero-shot ood detection: Teaching clip to say no

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.099593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.099593Z digest=sha256:923cf2965b14c01a9a98b09f3c5f06ead968939fff9ce30db8abec6326c57094

Observation 9a4916f5-f44c-4cd2-b026-16ebf663c7d5 · outbound

This paper cites Ev-gait: Event- based robust gait recognition using dynamic vision sensors.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Ev-gait: Event- based robust gait recognition using dynamic vision sensors

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.809896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.104100Z digest=sha256:2ee63f290c98cf974de3b25a5328f00bd2ce1a4ba9da3de8f97253da96813c54

Observation 222294a2-f625-4ea9-9c6c-353d409e862c · outbound

This paper cites Transferring clip’s knowledge into zero-shot point cloud semantic seg- mentation.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Transferring clip’s knowledge into zero-shot point cloud semantic seg- mentation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.793298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.108998Z digest=sha256:45a9523c874c834c0b0b653af249e95a4ac9b2bc7660feae6a318cad9665d2ac

Observation a38b14db-b490-42a5-b068-6f7016b2f919 · outbound

This paper cites Exploiting spatial sparsity for event cameras with visual transformers.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Exploiting spatial sparsity for event cameras with visual transformers

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.772319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.114001Z digest=sha256:b0063013b14ce91641b05c71bb501cc4f45c225e6ec2100e704f94bf913d352d

Observation b9dd9e0b-0e7c-4526-aced-5c5066f0178f · outbound

This paper cites Improving zero-shot generalization for clip with synthesized prompts.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Improving zero-shot generalization for clip with synthesized prompts

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.755069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.118718Z digest=sha256:d8679bf7256928c64ca609c75df7dfdfe5cbfe51ffbb689e3cf5cf8fea0dd11b

Observation 18b32784-3ca4-45f3-bca4-73dc4cd275e3 · outbound

This paper cites Wav2clip: Learning robust audio repre- sentations from clip.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Wav2clip: Learning robust audio repre- sentations from clip

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.736793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.123455Z digest=sha256:12423cd45595418d59fe0ea3cfb28c5c43f8289885dd6388b15fc9789026a70a

Observation 4d0a44aa-cd69-436a-932c-a1b353d559ce · outbound

This paper cites EventCLIP: Adapting CLIP for Event-based Object Recognition.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder EventCLIP: Adapting CLIP for Event-based Object Recognition

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.128397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.128397Z digest=sha256:29ed42ba9f9895774bc4c149da58cd00009b32a5123776d83c89b0f36e836bc2

Observation 89907a92-55ef-41c4-b7d4-18babd47e028 · outbound

This paper cites Event camera data pre-training.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Event camera data pre-training

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.709095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.133543Z digest=sha256:421e9f01b11c4364c6ce4de1bf8482b81e17838481f081544f353093525427b0

Observation 2c8cb96e-955c-4098-a4ff-bd3a07fd7c3c · outbound

This paper cites Event-guided low- light video semantic segmentation.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Event-guided low- light video semantic segmentation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.138237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.138237Z digest=sha256:fdb86464df89693578fca4f080fc58001feb6c209253fc2ddd68c64e962b0d7f

Observation 219accb7-f14a-47e9-b123-0bd5e55f2d29 · outbound

This paper cites Clip2: Contrastive language- image-point pretraining from real-world point cloud data.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Clip2: Contrastive language- image-point pretraining from real-world point cloud data

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.694032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.143005Z digest=sha256:118bbaff3346f56ddb4f6ee7423390647005c54d74c6d521ec6fd16c94a02015

Observation 6543e224-f664-4552-bb42-c0293b2afbe2 · outbound

This paper cites Vision-language models for vision tasks: A survey.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Vision-language models for vision tasks: A survey

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.147647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.147647Z digest=sha256:884376afb24a0b8cc9d83cd95534c1697c488ff26331795a4d3842dadd6e4b8c

Observation 62302795-f6d8-4aa6-985e-2b3f2f64be63 · outbound

This paper cites Pointclip: Point cloud understanding by clip.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Pointclip: Point cloud understanding by clip

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.152995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.152995Z digest=sha256:2cbc225634383a30cb260ccc63f961349c3854f9d84853b4f1706d5bbbdea83a

Observation 709d7b41-c539-405b-be9b-6cd80f411f32 · outbound

This paper cites Can language understand depth? In Proceedings of the 30th ACM International Conference on Multimedia, pages 6868–6874,.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Can language understand depth? In Proceedings of the 30th ACM International Conference on Multimedia, pages 6868–6874,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.660915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.157632Z digest=sha256:d8bb7b2383eac5d2634ea9578970b892d4bc3a30fb7b53df4d551ab558a815d1

Observation be37f930-1e1e-4df5-b071-e24d318eb53f · outbound

This paper cites Tip- adapter: Training-free adaption of clip for few-shot classi- fication.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Tip- adapter: Training-free adaption of clip for few-shot classi- fication

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.645427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.163001Z digest=sha256:5630db319520951cb359ed772a7acf0455828d70971ddb3181a714fcc7b2f46a

Observation 4d696e3a-5504-4bf1-bf75-2c953456a5a9 · outbound

This paper cites Eventdance: Unsupervised source- free cross-modal adaptation for event-based object recogni- tion.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Eventdance: Unsupervised source- free cross-modal adaptation for event-based object recogni- tion

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.168079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.168079Z digest=sha256:17caa634f9b688ded91b86bd0a5bdb5401f6bd9e02efb9509dc6b5be6d362245

Observation 1729ad24-a7d8-45bb-865d-b702ea4f1737 · outbound

This paper cites Deep Learning for Event-based Vision: A Comprehensive Survey and Benchmarks.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Deep Learning for Event-based Vision: A Comprehensive Survey and Benchmarks

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.172895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.172895Z digest=sha256:98d5322d4de7d2c8992f5b5ef793b6ec687c46485c3289fe32b1f01bf9c6dec6

Observation 5faf4cd9-bc2a-48e3-97e1-14d0023a07a4 · outbound

This paper cites Preventing zero-shot transfer degradation in continual learning of vision-language mod- els.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Preventing zero-shot transfer degradation in continual learning of vision-language mod- els

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.619395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.177859Z digest=sha256:6d256b81d6d00b1d42067cc21421ca4ba2c25203ade4fbe9eb96eebdf5055a3d

Observation 3c1a4a4b-bc96-4dff-b5a0-0b2f641e11b1 · outbound

This paper cites Regionclip: Region- based language-image pretraining.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Regionclip: Region- based language-image pretraining

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.603685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.182321Z digest=sha256:756d5481751a87b02a2d8bf9fc394754009dacdb8bbfc3a18fa66335064512b6

Observation 0215a477-74e3-46bc-9a79-ef3bd1cb3c63 · outbound

This paper cites EventBind: Learning a Unified Representation to Bind Them All for Event-based Open-world Understanding.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder EventBind: Learning a Unified Representation to Bind Them All for Event-based Open-world Understanding

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.186930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.186930Z digest=sha256:e874b7e5a489682de595f21f4a8abfbb0f748746dec1e3e8822ba7e0c42558e0

Observation 5f7b8734-13ba-4d03-8bc9-23fc078b2407 · outbound

This paper cites Eventbind: Learning a unified representation to bind them all for event-based open-world understanding.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Eventbind: Learning a unified representation to bind them all for event-based open-world understanding

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.587898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.191727Z digest=sha256:16eff17ddf390fec15b0d46c36232ea496f3155a2909ac189ad00ec434fcff7c

Observation 6af4ced4-d257-4474-aab1-5bc29f273ffb · outbound

This paper cites Anomalyclip: Object-agnostic prompt learn- ing for zero-shot anomaly detection.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder Anomalyclip: Object-agnostic prompt learn- ing for zero-shot anomaly detection

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.196331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.196331Z digest=sha256:55f0b9bf44f42c36aa4dfbb9c62aff18ee2f6a0d27f68df8cfe306916d368155

Observation 02bf9f7d-2fd7-4689-bd2e-37722b975cfa · outbound

This paper cites A brief introduction to weakly supervised learning.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder A brief introduction to weakly supervised learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:50:59.571875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:50:59.200685Z digest=sha256:03cff9a899e17012c76e81a5e4187013b898427e00fcd0096d7129e3521bbe78

Observation 6f641583-3407-4a50-abe9-08ecb4bcf23a · outbound

This paper cites LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment.

Expanding Event Modality Applications through a Robust CLIP-Based Encoder LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T22:50:59.205085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:50:59.205085Z digest=sha256:8fb381865765e6e1416a6354c82e2180f8b596fb615a45c384f918dc31d948ca

Pith citing papers

Observation ed8b3608-7fe0-4662-8ff6-435354978716 · inbound

Continuous GNN-based Anomaly Detection on Edge using Efficient Adaptive Knowledge Graph Learning cites this paper.

Continuous GNN-based Anomaly Detection on Edge using Efficient Adaptive Knowledge Graph Learning Expanding Event Modality Applications through a Robust CLIP-Based Encoder

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-12T21:11:07.578297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T21:11:07.359765Z digest=sha256:cec06ed8ea4b26ac1c24c7389c34ebd6fc9883780f9ea9c1a34670a8cbea607d