Pith. sign in

Paper Citation Record · LEDGER

Group Relative Augmentation for Data Efficient Action Detection

As of 7 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2507.21353.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21353 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:55:41.858385Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact2
  • verified fuzzy27
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9e6b3f22-96c8-4fdf-bd98-99d6283d1359 · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval.

Group Relative Augmentation for Data Efficient Action Detection Frozen in time: A joint video and image encoder for end-to-end retrieval

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:48.486773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:36.785009Z digest=sha256:10242d9551d77f75623a8bd7308d89ad8ce0caf88a45c39dad632bb2bd31a683

Observation 349e228c-2515-4f54-bef2-3dc0a1259e28 · outbound

This paper cites Exploiting vlm localizability and semantics for open vocabulary action detection.

Group Relative Augmentation for Data Efficient Action Detection Exploiting vlm localizability and semantics for open vocabulary action detection

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:48.249842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:36.938943Z digest=sha256:040936ae4f7717fc1d8cf34f0dde5de0bea0054c635817d8826833c09b0d33fe

Observation a4e19a9e-a116-40f2-8a43-dda1673bbca1 · outbound

This paper cites Frozen feature augmentation for few-shot image classification.

Group Relative Augmentation for Data Efficient Action Detection Frozen feature augmentation for few-shot image classification

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:48.040466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:37.050985Z digest=sha256:808a68b60ad7c0acf82113fe255fe5910bc75cabe7561e68668928f0d8bd4138

Observation 1315f334-b352-4329-9662-6fdb5718d5f4 · outbound

This paper cites Visualgpt: Data- efficient adaptation of pretrained language models for image captioning.

Group Relative Augmentation for Data Efficient Action Detection Visualgpt: Data- efficient adaptation of pretrained language models for image captioning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:47.779631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:37.167650Z digest=sha256:59c8755bae81d003c132dc5a2b5f315fcc05dd878a1709487a042bb88f9e553d

Observation fddb37e3-7c7c-4e32-9b25-f5f1054d910e · outbound

This paper cites Adversarial Feature Augmentation and Normalization for Visual Recognition.

Group Relative Augmentation for Data Efficient Action Detection Adversarial Feature Augmentation and Normalization for Visual Recognition

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:55:42.537495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:37.290334Z digest=sha256:c4f88c6789139b8ec5a4db7617fd4657a88a7aa61d4a75a12ed9c564b0fb7479

Observation 254383df-acec-4f27-a082-f36dbfe6d5b6 · outbound

This paper cites PerceptionLM: Open-Access Data and Models for Detailed Visual Understanding.

Group Relative Augmentation for Data Efficient Action Detection PerceptionLM: Open-Access Data and Models for Detailed Visual Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:37.435979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:37.435979Z digest=sha256:496f07786e5f34bd0c36948db70c081db94f1fb75e538f40cc4108bfe0784fff

Observation 10f9b7f0-37fe-406c-ae7c-6055b09f2a18 · outbound

This paper cites Autoaugment: Learning augmentation strategies from data.

Group Relative Augmentation for Data Efficient Action Detection Autoaugment: Learning augmentation strategies from data

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:47.546084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:37.606762Z digest=sha256:1703de5b6f4bd3bb0d6fea91040c26774206503fd1e73547d8e14ac129952f28

Observation a1ef18b4-e09c-4fbb-b061-84da64288753 · outbound

This paper cites Clip-adapter: Better vision-language models with fea- ture adapters.

Group Relative Augmentation for Data Efficient Action Detection Clip-adapter: Better vision-language models with fea- ture adapters

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:47.336907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:37.752750Z digest=sha256:3edcd8568c6076b713903493cca7d285e75cc88323c965d6d04fd198ecf8ecaf

Observation 94005041-74c4-4cf2-bfba-70c183cdab04 · outbound

This paper cites Ava: A video dataset of spatio-temporally localized atomic visual actions.

Group Relative Augmentation for Data Efficient Action Detection Ava: A video dataset of spatio-temporally localized atomic visual actions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:47.133805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:37.911344Z digest=sha256:2f3e3b4f6a7323c4e6f138d468f1551782d1abc123f2730d9cff04786ef8bff5

Observation f85fcdf5-329b-4cfb-a473-79972efaf6f6 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Group Relative Augmentation for Data Efficient Action Detection Lora: Low-rank adaptation of large language models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:38.054743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:38.054743Z digest=sha256:f6693adf107809ec859a8b63ff88048b683a4af26bb2563c97a88b4fbb881e50

Observation db1f0c87-1b3e-4238-9dd2-754a670352ac · outbound

This paper cites Interaction-aware prompting for zero-shot spatio-temporal action detection.

Group Relative Augmentation for Data Efficient Action Detection Interaction-aware prompting for zero-shot spatio-temporal action detection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:46.917614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:38.161397Z digest=sha256:4c137729b19ffab8db989f635970d46887efb04dadcd1bd10697a1da664cb424

Observation 4fd9715a-7dda-45ce-8a63-e2764a9c28a3 · outbound

This paper cites Interaction-aware prompting for zero-shot spatio-temporal action detection.

Group Relative Augmentation for Data Efficient Action Detection Interaction-aware prompting for zero-shot spatio-temporal action detection

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:46.678765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:38.274684Z digest=sha256:414ef2412bc31207bb057e90742c998b6b0bb04a788468edf46919a02cc39669

Observation 6a36ac45-cfb3-456a-ac98-2a8c50b20fbd · outbound

This paper cites Spatio-temporal context prompting for zero-shot action detection.

Group Relative Augmentation for Data Efficient Action Detection Spatio-temporal context prompting for zero-shot action detection

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:46.474804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:38.390618Z digest=sha256:bcb6f74034f69d4528969b368ff82826ab0b168f0f367e8db616d83f7f02cdb2

Observation 0b32c2b8-056c-47cf-ac57-d1f5024f7549 · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision.

Group Relative Augmentation for Data Efficient Action Detection Scaling up visual and vision-language representation learning with noisy text supervision

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:38.531397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:38.531397Z digest=sha256:141cb47fdcac691519ea195180bd36cd8783c6a2bcb87fe901d30439921233c1

Observation 5bbc9bb6-f8f0-4f41-bc00-37147d0d2d25 · outbound

This paper cites Region-aware pretraining for open- vocabulary object detection with vision transformers.

Group Relative Augmentation for Data Efficient Action Detection Region-aware pretraining for open- vocabulary object detection with vision transformers

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:46.257650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:38.641284Z digest=sha256:301a4b00675ed04c93a1685cf2df501e2747e51c69027fb0e538c89ac8905ba5

Observation 69262b86-ce9e-44f5-8156-17915e2b7ea9 · outbound

This paper cites On feature normalization and data augmentation.

Group Relative Augmentation for Data Efficient Action Detection On feature normalization and data augmentation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:46.032460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:38.754542Z digest=sha256:0ce2dbd00abcbe4f07b7cb800d3853616c5b49fc92b02bd25fae5b201d8d293d

Observation 27a371b8-107f-4dbc-b820-05748e87c9ce · outbound

This paper cites Blip: Bootstrapping language- image pre-training for unified vision-language understanding and generation.

Group Relative Augmentation for Data Efficient Action Detection Blip: Bootstrapping language- image pre-training for unified vision-language understanding and generation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:45.806299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:38.881211Z digest=sha256:92280601c20a7a786abcda79e5ba720014e13cd329a66ca79d64bc9346aa19f6

Observation 96be9944-1086-4730-a0c1-ae2ec5354e4f · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Group Relative Augmentation for Data Efficient Action Detection Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:38.992416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:38.992416Z digest=sha256:c7aee172a4a96d16ef616276211598d979358e2767e64319bbab58e07e47f2f1

Observation 0f55c61e-cd19-4f48-a533-c1f51e3fd8ab · outbound

This paper cites Learning Object-Language Alignments for Open-Vocabulary Object Detection.

Group Relative Augmentation for Data Efficient Action Detection Learning Object-Language Alignments for Open-Vocabulary Object Detection

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:39.145776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:39.145776Z digest=sha256:11a273158af37c4e6831ea96a2d17a91ce08af4c78079f4849e4a67b868cf4f6

Observation af6a9cf9-bbff-42cd-8e17-e1b110e90cce · outbound

This paper cites UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation.

Group Relative Augmentation for Data Efficient Action Detection UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:39.270857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:39.270857Z digest=sha256:82639825a077938482e98375d157f9537519202223b250f1f14b855ad9c304d2

Observation c71580c9-2a89-4a71-bd0b-32a618059631 · outbound

This paper cites Moma: Multi-object multi-actor activity parsing.

Group Relative Augmentation for Data Efficient Action Detection Moma: Multi-object multi-actor activity parsing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:45.534123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:39.376277Z digest=sha256:f0cee7fdc006045e400e1af9c292d0f34c77b78a75199b2137fa1011fbab2871

Observation f39fccbd-82bb-4898-96f7-bd52d3dc0bbf · outbound

This paper cites Film: Visual reasoning with a general conditioning layer.

Group Relative Augmentation for Data Efficient Action Detection Film: Visual reasoning with a general conditioning layer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:39.478935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:39.478935Z digest=sha256:3325ca27cfc50df22c6c87d793953b1c47f11c3bbedb8fd0110975e84e85cabc

Observation 81659a12-2f1c-4cb7-98ec-2f0e983214af · outbound

This paper cites Learning transferable visual models from natural language supervision.

Group Relative Augmentation for Data Efficient Action Detection Learning transferable visual models from natural language supervision

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:45.265455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:39.630606Z digest=sha256:574c01088195e70744d111cca1b86265e87070f92cd547c809f13ad1f291b464

Observation 228b6ebe-8c05-44ea-9c66-b84aab72939c · outbound

This paper cites Videomae: Masked autoen- coders are data-efficient learners for self-supervised video pre-training.

Group Relative Augmentation for Data Efficient Action Detection Videomae: Masked autoen- coders are data-efficient learners for self-supervised video pre-training

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:45.081638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:39.727853Z digest=sha256:d776936c9aec8440584294b7e05a460c1358465fe33da0eb7415e4e4482745c0

Observation c5aa8eb1-7c74-400e-815e-eb4541b7dc87 · outbound

This paper cites Internvideo2: Scaling foundation models for multimodal video understanding.

Group Relative Augmentation for Data Efficient Action Detection Internvideo2: Scaling foundation models for multimodal video understanding

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:44.837052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:39.848204Z digest=sha256:ca75cd106983f282cf3f6192c39261f1a95a688f2799fbaaf85a2ea51dc259ed

Observation 742852c3-85d0-42b1-a66e-8bb309ec9bde · outbound

This paper cites Feature adaptation with clip for few-shot classification.

Group Relative Augmentation for Data Efficient Action Detection Feature adaptation with clip for few-shot classification

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:44.655749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:39.955529Z digest=sha256:64dd11ef73017448d74e53aad9582ff499f3932a3362b9acdda505af2228e193

Observation f6f3f7f4-705c-4f09-99ec-ec3291c07080 · outbound

This paper cites Cora: Adapting clip for open- vocabulary detection with region prompting and anchor pre-matching.

Group Relative Augmentation for Data Efficient Action Detection Cora: Adapting clip for open- vocabulary detection with region prompting and anchor pre-matching

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:44.410536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:40.070428Z digest=sha256:98f92b86d462ba649fb6af37fc9175e6f88d29fef87f92f265b7d95481730959

Observation 90e63cc7-bf95-4ea5-aa75-fa1289e5b836 · outbound

This paper cites VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding.

Group Relative Augmentation for Data Efficient Action Detection VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:40.226378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:40.226378Z digest=sha256:a3cc5e2a252db0d4eefae2f7c043b2f1b3b69d8586cc820d5958b29437963497

Observation be5f889f-9d9e-4c31-bc0c-d8ab0347161f · outbound

This paper cites VideoCoCa: Video-Text Modeling with Zero-Shot Transfer from Contrastive Captioners.

Group Relative Augmentation for Data Efficient Action Detection VideoCoCa: Video-Text Modeling with Zero-Shot Transfer from Contrastive Captioners

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:40.391207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:40.391207Z digest=sha256:579b6f52472449618ec8afcaedb1ec71dce384e3c962cdcc51ac22770646f083

Observation a435b824-26a2-43b8-ab32-aa11018d5207 · outbound

This paper cites Vid2seq: Large-scale pretraining of a visual language model for dense video captioning.

Group Relative Augmentation for Data Efficient Action Detection Vid2seq: Large-scale pretraining of a visual language model for dense video captioning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:44.192366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:40.486806Z digest=sha256:d97705fa1d2a21d99546089c6bf0f0dc7b158a0c08d399c8976b4ea49ef0ab68

Observation c0d2707b-60cc-43ad-ac3a-837849da7257 · outbound

This paper cites Image Data Augmentation for Deep Learning: A Survey.

Group Relative Augmentation for Data Efficient Action Detection Image Data Augmentation for Deep Learning: A Survey

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:40.626249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:40.626249Z digest=sha256:36d263dfcce7f6304d27f3ebef2c529a207260537bdfe35895f0257ea410ece2

Observation 1dfee674-368b-4c65-a70e-9e3d971d1b73 · outbound

This paper cites Textmania: Enriching visual feature by text-driven manifold augmentation.

Group Relative Augmentation for Data Efficient Action Detection Textmania: Enriching visual feature by text-driven manifold augmentation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:44.008802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:40.747236Z digest=sha256:10c6bf1f8ba3db07aa92c8f23401861246f82d1dd91ff4755f9b513980f1f073

Observation 029b2d6d-5e9d-45b5-aa3e-3c75799d9622 · outbound

This paper cites CoCa: Contrastive Captioners are Image-Text Foundation Models.

Group Relative Augmentation for Data Efficient Action Detection CoCa: Contrastive Captioners are Image-Text Foundation Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:40.853210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:40.853210Z digest=sha256:7946cf90930b6d96b2597626d6972d55c931e45fecdbb4f39e5b1286f284ea3f

Observation f9b956d0-0445-4635-96b6-4c7a73d489f2 · outbound

This paper cites Cutmix: Regularization strategy to train strong classifiers with local- izable features.

Group Relative Augmentation for Data Efficient Action Detection Cutmix: Regularization strategy to train strong classifiers with local- izable features

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:43.803420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:40.973725Z digest=sha256:102f2c3c195c365294b2dc67a0d25e6d3ec6a8a9dc8aec7adb78a214fc8e7c49

Observation 67fffa54-4971-474c-8cf5-517ee3aec1ce · outbound

This paper cites Fasa: Feature augmentation and sampling adaptation for long-tailed instance segmentation.

Group Relative Augmentation for Data Efficient Action Detection Fasa: Feature augmentation and sampling adaptation for long-tailed instance segmentation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:43.561454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:41.106384Z digest=sha256:ef644e96afd9bb91489c4498e9e9253a3172400a673f99f75d349bda061c8c18

Observation 3778ec90-35a5-4105-af86-eea23753776b · outbound

This paper cites Open- vocabulary object detection using captions.

Group Relative Augmentation for Data Efficient Action Detection Open- vocabulary object detection using captions

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:43.357916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:41.202202Z digest=sha256:344b2ad05d19a709d86c55cae87d989cdcc9200c34b8753633b512b36520acea

Observation f06c7f30-fcb4-4a29-84aa-7b381c0b3191 · outbound

This paper cites Sigmoid loss for language image pre-training.

Group Relative Augmentation for Data Efficient Action Detection Sigmoid loss for language image pre-training

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:43.153849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:41.316739Z digest=sha256:09b2936c327adc1bc96e1591cc0deb4e39282991d53e4508440f32110263e5b5

Observation 894229cf-580e-40d6-bde8-c47c77ae6cb1 · outbound

This paper cites Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling.

Group Relative Augmentation for Data Efficient Action Detection Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:41.440158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:41.440158Z digest=sha256:c434eef30bc0e3ce146edd70a7415adefe23d51ce1a060275f310478ada6ea04

Observation 5fead464-266c-48d7-bad9-279b38e5a680 · outbound

This paper cites Don't Judge by the Look: Towards Motion Coherent Video Representation.

Group Relative Augmentation for Data Efficient Action Detection Don't Judge by the Look: Towards Motion Coherent Video Representation

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:55:42.090764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:41.609971Z digest=sha256:ef0255c3f3380755297325fda3f055fd8689aa601f50446fb203176184165526

Observation ee2b1069-a638-46dd-8642-57333de3b69e · outbound

This paper cites Regionclip: Region- based language-image pretraining.

Group Relative Augmentation for Data Efficient Action Detection Regionclip: Region- based language-image pretraining

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:42.975241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:41.733634Z digest=sha256:81122ff737a86e048172f1a5290e5819cd25a773a26ee677c55362e48f58b880

Observation 15a2ed48-d5a7-425b-9651-0d15ca654883 · outbound

This paper cites Not all features matter: Enhancing few-shot clip with adaptive prior refine- ment.

Group Relative Augmentation for Data Efficient Action Detection Not all features matter: Enhancing few-shot clip with adaptive prior refine- ment

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:42.726408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T12:55:41.858385Z digest=sha256:c2b8d992a8b76c05c401ff076bb15db3ccd98638f8893d99268e4ac14a227bce

Pith citing papers

No inbound Pith citation observations are available.