Pith. sign in

Paper Citation Record · LEDGER

Group Relative Augmentation for Data Efficient Action Detection

As of 17 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2507.21353.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21353 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:55:41.858385Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact2
  • verified fuzzy27
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9e6b3f22-96c8-4fdf-bd98-99d6283d1359 · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval.

Group Relative Augmentation for Data Efficient Action Detection Frozen in time: A joint video and image encoder for end-to-end retrieval

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:48.486773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:36.785009Z digest=sha256:aef4a4f09eb43df519264c8cd954ea9a01a51e7c9389eec70840ee34be191375

Observation 349e228c-2515-4f54-bef2-3dc0a1259e28 · outbound

This paper cites Exploiting vlm localizability and semantics for open vocabulary action detection.

Group Relative Augmentation for Data Efficient Action Detection Exploiting vlm localizability and semantics for open vocabulary action detection

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:48.249842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:36.938943Z digest=sha256:07a82cfb65ab2f131f0d3a7718c2930b46f4e232981d3671f034aa3b107d7d5f

Observation a4e19a9e-a116-40f2-8a43-dda1673bbca1 · outbound

This paper cites Frozen feature augmentation for few-shot image classification.

Group Relative Augmentation for Data Efficient Action Detection Frozen feature augmentation for few-shot image classification

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:48.040466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:37.050985Z digest=sha256:db287efdee215a9cf66276e7b1c96358b1a954e903a175370d4b1bbe5a6df455

Observation 1315f334-b352-4329-9662-6fdb5718d5f4 · outbound

This paper cites Visualgpt: Data- efficient adaptation of pretrained language models for image captioning.

Group Relative Augmentation for Data Efficient Action Detection Visualgpt: Data- efficient adaptation of pretrained language models for image captioning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:47.779631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:37.167650Z digest=sha256:bbcd0d7d8ae942345dade67b188408cdf268424cb280658289108f810e0f33a9

Observation fddb37e3-7c7c-4e32-9b25-f5f1054d910e · outbound

This paper cites Adversarial Feature Augmentation and Normalization for Visual Recognition.

Group Relative Augmentation for Data Efficient Action Detection Adversarial Feature Augmentation and Normalization for Visual Recognition

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:55:42.537495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:37.290334Z digest=sha256:0971aaee1e223a97140e31f8954dce75f560dcfb3d1aa3bc9c7f80bcd4fbfba0

Observation 254383df-acec-4f27-a082-f36dbfe6d5b6 · outbound

This paper cites PerceptionLM: Open-Access Data and Models for Detailed Visual Understanding.

Group Relative Augmentation for Data Efficient Action Detection PerceptionLM: Open-Access Data and Models for Detailed Visual Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:37.435979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:37.435979Z digest=sha256:33bde0c5d21c47feffebb1f57017b98a02246a2dc5e911403eb4640163c4a028

Observation 10f9b7f0-37fe-406c-ae7c-6055b09f2a18 · outbound

This paper cites Autoaugment: Learning augmentation strategies from data.

Group Relative Augmentation for Data Efficient Action Detection Autoaugment: Learning augmentation strategies from data

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:47.546084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:37.606762Z digest=sha256:fbbdb01a9c72e0d3a99e9ed6a585ffc1828af2a6d0d13f43731fd68e78daced7

Observation a1ef18b4-e09c-4fbb-b061-84da64288753 · outbound

This paper cites Clip-adapter: Better vision-language models with fea- ture adapters.

Group Relative Augmentation for Data Efficient Action Detection Clip-adapter: Better vision-language models with fea- ture adapters

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:47.336907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:37.752750Z digest=sha256:14b4e97afda5d999303b01395f31ce029e13e30994e50ec49ba8cc342b429ab6

Observation 94005041-74c4-4cf2-bfba-70c183cdab04 · outbound

This paper cites Ava: A video dataset of spatio-temporally localized atomic visual actions.

Group Relative Augmentation for Data Efficient Action Detection Ava: A video dataset of spatio-temporally localized atomic visual actions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:47.133805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:37.911344Z digest=sha256:5d3e1f14037cb8f2188c19f0c7f43d530529a2633ac7ec325a51702508cb6f48

Observation f85fcdf5-329b-4cfb-a473-79972efaf6f6 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Group Relative Augmentation for Data Efficient Action Detection Lora: Low-rank adaptation of large language models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:38.054743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:38.054743Z digest=sha256:0fddecaec39c3e1bada455566fe7b73073db418782a68e5d3c96ef2d3fee609f

Observation db1f0c87-1b3e-4238-9dd2-754a670352ac · outbound

This paper cites Interaction-aware prompting for zero-shot spatio-temporal action detection.

Group Relative Augmentation for Data Efficient Action Detection Interaction-aware prompting for zero-shot spatio-temporal action detection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:46.917614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:38.161397Z digest=sha256:e20807ed9297e94b3991fe06ebf8b567229d498e36b327f5a9b1d2604233ff7e

Observation 4fd9715a-7dda-45ce-8a63-e2764a9c28a3 · outbound

This paper cites Interaction-aware prompting for zero-shot spatio-temporal action detection.

Group Relative Augmentation for Data Efficient Action Detection Interaction-aware prompting for zero-shot spatio-temporal action detection

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:46.678765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:38.274684Z digest=sha256:a4e637c816295d295712e19509360bc8f164027f9a784b10aa4dbb76147f14af

Observation 6a36ac45-cfb3-456a-ac98-2a8c50b20fbd · outbound

This paper cites Spatio-temporal context prompting for zero-shot action detection.

Group Relative Augmentation for Data Efficient Action Detection Spatio-temporal context prompting for zero-shot action detection

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:46.474804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:38.390618Z digest=sha256:a058524743068791875dcede79df91b78d4cbd55a211ccd96ac83acc3f19bcaa

Observation 0b32c2b8-056c-47cf-ac57-d1f5024f7549 · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision.

Group Relative Augmentation for Data Efficient Action Detection Scaling up visual and vision-language representation learning with noisy text supervision

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:38.531397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:38.531397Z digest=sha256:393e660becf9fa5200476a334781b2a9e375a6e4edf863d67a581a652d02bd08

Observation 5bbc9bb6-f8f0-4f41-bc00-37147d0d2d25 · outbound

This paper cites Region-aware pretraining for open- vocabulary object detection with vision transformers.

Group Relative Augmentation for Data Efficient Action Detection Region-aware pretraining for open- vocabulary object detection with vision transformers

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:46.257650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:38.641284Z digest=sha256:f05d12ba8de79273aaa85e22ca72b15b70ddc2c469425b95118962ee8df9c171

Observation 69262b86-ce9e-44f5-8156-17915e2b7ea9 · outbound

This paper cites On feature normalization and data augmentation.

Group Relative Augmentation for Data Efficient Action Detection On feature normalization and data augmentation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:46.032460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:38.754542Z digest=sha256:ae42095d3b6de17ffe30ab10d510c397289b0e6ddcf2c2062edaefbbf9764e45

Observation 27a371b8-107f-4dbc-b820-05748e87c9ce · outbound

This paper cites Blip: Bootstrapping language- image pre-training for unified vision-language understanding and generation.

Group Relative Augmentation for Data Efficient Action Detection Blip: Bootstrapping language- image pre-training for unified vision-language understanding and generation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:45.806299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:38.881211Z digest=sha256:37908932e9aae27d9038726807262409641b577b07d03d4a199981add48a9d64

Observation 96be9944-1086-4730-a0c1-ae2ec5354e4f · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Group Relative Augmentation for Data Efficient Action Detection Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:38.992416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:38.992416Z digest=sha256:5da21071a687f475ed12f11fc78b5e0813e4ea48f03e906c8274e8347c251057

Observation 0f55c61e-cd19-4f48-a533-c1f51e3fd8ab · outbound

This paper cites Learning Object-Language Alignments for Open-Vocabulary Object Detection.

Group Relative Augmentation for Data Efficient Action Detection Learning Object-Language Alignments for Open-Vocabulary Object Detection

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:39.145776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:39.145776Z digest=sha256:64b2ff3dbe7f97cc2fc8900320fe07a45550d7abc28300fe47637bac9ffe44f5

Observation af6a9cf9-bbff-42cd-8e17-e1b110e90cce · outbound

This paper cites UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation.

Group Relative Augmentation for Data Efficient Action Detection UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:39.270857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:39.270857Z digest=sha256:1b1e1d01878739509a55f1c323c3e96fad1ea60f9b0410800659e8d32e4a5bc1

Observation c71580c9-2a89-4a71-bd0b-32a618059631 · outbound

This paper cites Moma: Multi-object multi-actor activity parsing.

Group Relative Augmentation for Data Efficient Action Detection Moma: Multi-object multi-actor activity parsing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:45.534123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:39.376277Z digest=sha256:bc637869174a6aff1b3c9a2cc1aafab79a8b32923ff96793d25f0fdd2c4791b3

Observation f39fccbd-82bb-4898-96f7-bd52d3dc0bbf · outbound

This paper cites Film: Visual reasoning with a general conditioning layer.

Group Relative Augmentation for Data Efficient Action Detection Film: Visual reasoning with a general conditioning layer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:39.478935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:39.478935Z digest=sha256:99978b7498963afcb3ca3034304e94cf2b79c5db17348315e23c6e89e1c1225b

Observation 81659a12-2f1c-4cb7-98ec-2f0e983214af · outbound

This paper cites Learning transferable visual models from natural language supervision.

Group Relative Augmentation for Data Efficient Action Detection Learning transferable visual models from natural language supervision

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:45.265455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:39.630606Z digest=sha256:64336bd270f1087fc6f0f5eba722d677d0f793a85471f0e2211eeb08382bf33b

Observation 228b6ebe-8c05-44ea-9c66-b84aab72939c · outbound

This paper cites Videomae: Masked autoen- coders are data-efficient learners for self-supervised video pre-training.

Group Relative Augmentation for Data Efficient Action Detection Videomae: Masked autoen- coders are data-efficient learners for self-supervised video pre-training

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:45.081638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:39.727853Z digest=sha256:fc1a0753a609e8e241781c74ffbe389a4bb99f751dd355951e937e6ac53010ff

Observation c5aa8eb1-7c74-400e-815e-eb4541b7dc87 · outbound

This paper cites Internvideo2: Scaling foundation models for multimodal video understanding.

Group Relative Augmentation for Data Efficient Action Detection Internvideo2: Scaling foundation models for multimodal video understanding

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:44.837052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:39.848204Z digest=sha256:d8c242c4b376fcb4f57e25edad9641126bf570f2cacb47c68c96363bc7abd6ac

Observation 742852c3-85d0-42b1-a66e-8bb309ec9bde · outbound

This paper cites Feature adaptation with clip for few-shot classification.

Group Relative Augmentation for Data Efficient Action Detection Feature adaptation with clip for few-shot classification

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:44.655749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:39.955529Z digest=sha256:aaa470a13663d72bafa5262d65d7c0e7136605dfbb0c414046d42909e7faf1de

Observation f6f3f7f4-705c-4f09-99ec-ec3291c07080 · outbound

This paper cites Cora: Adapting clip for open- vocabulary detection with region prompting and anchor pre-matching.

Group Relative Augmentation for Data Efficient Action Detection Cora: Adapting clip for open- vocabulary detection with region prompting and anchor pre-matching

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:44.410536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:40.070428Z digest=sha256:a94a316fcccf687c7a041d9b94f683de367187ce03d025513afe090dcabb72ab

Observation 90e63cc7-bf95-4ea5-aa75-fa1289e5b836 · outbound

This paper cites VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding.

Group Relative Augmentation for Data Efficient Action Detection VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:40.226378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:40.226378Z digest=sha256:42c3e985f266556e15a908cee09809f260263f4ddb29d67970e2f00a82da5ab5

Observation be5f889f-9d9e-4c31-bc0c-d8ab0347161f · outbound

This paper cites VideoCoCa: Video-Text Modeling with Zero-Shot Transfer from Contrastive Captioners.

Group Relative Augmentation for Data Efficient Action Detection VideoCoCa: Video-Text Modeling with Zero-Shot Transfer from Contrastive Captioners

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:40.391207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:40.391207Z digest=sha256:83e5808aaf395117e5a6e46535dd42d303a0973e39f2ee79aae480da6804caf1

Observation a435b824-26a2-43b8-ab32-aa11018d5207 · outbound

This paper cites Vid2seq: Large-scale pretraining of a visual language model for dense video captioning.

Group Relative Augmentation for Data Efficient Action Detection Vid2seq: Large-scale pretraining of a visual language model for dense video captioning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:44.192366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:40.486806Z digest=sha256:66083d0e73ca923b533a8a4b215895efb87a4d364231658d89c3017564992956

Observation c0d2707b-60cc-43ad-ac3a-837849da7257 · outbound

This paper cites Image Data Augmentation for Deep Learning: A Survey.

Group Relative Augmentation for Data Efficient Action Detection Image Data Augmentation for Deep Learning: A Survey

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:40.626249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:40.626249Z digest=sha256:235b88ed5bc00848853a17c5a5b4c177dda3194a8c06d923ab67ea954af92864

Observation 1dfee674-368b-4c65-a70e-9e3d971d1b73 · outbound

This paper cites Textmania: Enriching visual feature by text-driven manifold augmentation.

Group Relative Augmentation for Data Efficient Action Detection Textmania: Enriching visual feature by text-driven manifold augmentation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:44.008802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:40.747236Z digest=sha256:1c56a2500248cbd51dbaf389a840affeb447808ba7f3d5275e47a5f328ad1cc1

Observation 029b2d6d-5e9d-45b5-aa3e-3c75799d9622 · outbound

This paper cites CoCa: Contrastive Captioners are Image-Text Foundation Models.

Group Relative Augmentation for Data Efficient Action Detection CoCa: Contrastive Captioners are Image-Text Foundation Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:40.853210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:40.853210Z digest=sha256:6cdd8855bf6a2988218aec3e13d7f18643f4ea01b448d6f06a9e874b71432a4a

Observation f9b956d0-0445-4635-96b6-4c7a73d489f2 · outbound

This paper cites Cutmix: Regularization strategy to train strong classifiers with local- izable features.

Group Relative Augmentation for Data Efficient Action Detection Cutmix: Regularization strategy to train strong classifiers with local- izable features

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:43.803420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:40.973725Z digest=sha256:515bc4a0c3be08ec57587e0db7b2061eed01ddd453c33f73167d9ffeae1a8ef4

Observation 67fffa54-4971-474c-8cf5-517ee3aec1ce · outbound

This paper cites Fasa: Feature augmentation and sampling adaptation for long-tailed instance segmentation.

Group Relative Augmentation for Data Efficient Action Detection Fasa: Feature augmentation and sampling adaptation for long-tailed instance segmentation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:43.561454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:41.106384Z digest=sha256:91ce65ada5976d1a2310408498dbeb8304873440d360b7891e2cb3b15746103a

Observation 3778ec90-35a5-4105-af86-eea23753776b · outbound

This paper cites Open- vocabulary object detection using captions.

Group Relative Augmentation for Data Efficient Action Detection Open- vocabulary object detection using captions

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:43.357916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:41.202202Z digest=sha256:d71b8150364de70e5fa33374f891583765d80bf3a0135c7bd5a149697aeea5dc

Observation f06c7f30-fcb4-4a29-84aa-7b381c0b3191 · outbound

This paper cites Sigmoid loss for language image pre-training.

Group Relative Augmentation for Data Efficient Action Detection Sigmoid loss for language image pre-training

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:43.153849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:41.316739Z digest=sha256:72378f872d4be6d3ae5e70048c58027ebc1f084a38a07c545a53bbc09e2d5183

Observation 894229cf-580e-40d6-bde8-c47c77ae6cb1 · outbound

This paper cites Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling.

Group Relative Augmentation for Data Efficient Action Detection Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:41.440158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:41.440158Z digest=sha256:2429477bf4758ec42314431eb821db606ecee152c342baf0b014a0b3afdb0251

Observation 5fead464-266c-48d7-bad9-279b38e5a680 · outbound

This paper cites Don't Judge by the Look: Towards Motion Coherent Video Representation.

Group Relative Augmentation for Data Efficient Action Detection Don't Judge by the Look: Towards Motion Coherent Video Representation

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:55:42.090764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:41.609971Z digest=sha256:684d56e4a26e28c07ec67700f0dc4e2775f29b0391ccaaae713e2874c8204b13

Observation ee2b1069-a638-46dd-8642-57333de3b69e · outbound

This paper cites Regionclip: Region- based language-image pretraining.

Group Relative Augmentation for Data Efficient Action Detection Regionclip: Region- based language-image pretraining

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:42.975241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:41.733634Z digest=sha256:6ad9c09ba84133cba5159d3a93b227b3377b72751f0cba2ac9733d0c85e8d6a8

Observation 15a2ed48-d5a7-425b-9651-0d15ca654883 · outbound

This paper cites Not all features matter: Enhancing few-shot clip with adaptive prior refine- ment.

Group Relative Augmentation for Data Efficient Action Detection Not all features matter: Enhancing few-shot clip with adaptive prior refine- ment

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:55:42.726408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T12:55:41.858385Z digest=sha256:c9b9322a76a6eb30a9633aa1b9b47f9bf9a0b4783fe8ec89a80a2efabb0b2431

Pith citing papers

No inbound Pith citation observations are available.