Pith. sign in

Paper Citation Record · LEDGER

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings

As of 20 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:1908.03477.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.03477 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T14:15:57.272572Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy30
  • unresolved9
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 09347ffc-8b47-4a24-89f0-7dcf42b6192e · outbound

This paper cites an unresolved cited work.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:15:58.209242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.015401Z digest=sha256:66d6dd91dce522ae1b82fa870b48c6fa9af53f49f199809d0d9b955b5616905c

Observation 70e071cc-fe06-4f05-8010-266a3834f327 · outbound

This paper cites Re-ID done right: towards good practices for person re-identification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Re-ID done right: towards good practices for person re-identification

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.023091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.023091Z digest=sha256:57797f48a1f93ac89977f0d0cd5d8e7a72bc4cb051577b1c8e1a0d411d9a3372

Observation a7526b53-9e57-488c-8b35-236a261a4bca · outbound

This paper cites NetVLAD: CNN architecture for weakly supervised place recognition.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings NetVLAD: CNN architecture for weakly supervised place recognition

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.187494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.029900Z digest=sha256:9ea107a1dfbed5722427fca473002d26a5dc1f6b20cca1673e16bb13f243981a

Observation eb52af3d-3229-4db1-b3f3-9c81d452bc61 · outbound

This paper cites An empirical study and analysis of generalized zero- shot learning for object recognition in the wild.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings An empirical study and analysis of generalized zero- shot learning for object recognition in the wild

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.160419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.036174Z digest=sha256:8d722f79563e2fcf2e21ed9a0be8a790fe219d1cf685ba3eb2a8c45b6315f711

Observation 40f1f8b8-8e51-435a-8b4e-ce75457b613a · outbound

This paper cites Beyond triplet loss: a deep quadruplet network for person re-identification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Beyond triplet loss: a deep quadruplet network for person re-identification

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.139875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.043008Z digest=sha256:d9a7fc6d3eba1d874595c5150bc2f5ddb2f45dfe3cf8b6bf9118334195dff1a5

Observation 0b862c3b-5658-46a5-81f0-01806def26c8 · outbound

This paper cites Scaling egocentric vision: The epic-kitchens dataset.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Scaling egocentric vision: The epic-kitchens dataset

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.120942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.055743Z digest=sha256:05721119c32f7678796326bbc0b2e308b47b4a4fd56100b9ee1fb510af1c555a

Observation c4095747-5b8f-4326-9d83-56a92a40de35 · outbound

This paper cites Predict- ing visual features from text for image and video caption re- trieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Predict- ing visual features from text for image and video caption re- trieval

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.103558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.061105Z digest=sha256:ac8f60cc0aaa7228293eee9727e1d46569741dfff9d56966aa224a7d84210669

Observation 86de2abf-eb5e-4198-8626-0b4c4a2c1b66 · outbound

This paper cites Dual Encoding for Zero-Example Video Retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Dual Encoding for Zero-Example Video Retrieval

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:15:57.473760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.065902Z digest=sha256:2b0acd62ebf03a1bdfaf508275558eda9c166939bae00996728ba57466b31030

Observation 91628ae4-1ca5-4ff3-a5a9-f518899f9046 · outbound

This paper cites Improving image-sentence embeddings using large weakly annotated photo collections.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Improving image-sentence embeddings using large weakly annotated photo collections

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.086177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.071958Z digest=sha256:59c931de18f2dab257728d2e644caabb76d45918d728e6bde64c6804cfc4f1ba

Observation 24c39c6a-2d93-41d4-ab83-5847f64df17e · outbound

This paper cites Deep image retrieval: Learning global representations for image search.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Deep image retrieval: Learning global representations for image search

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.065837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.077021Z digest=sha256:91d4e22434dcd1a9baaf1f33cf768775fa46b39993cd35539f0d958dac98d0df

Observation 437a4b4c-65a0-4334-91a8-3b7703c905ea · outbound

This paper cites Beyond instance-level im- age retrieval: Leveraging captions to learn a global visual representation for semantic retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Beyond instance-level im- age retrieval: Leveraging captions to learn a global visual representation for semantic retrieval

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.046935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.082890Z digest=sha256:d39b79b962342888105a2d126b4bc46463fb73b6d7419d0c9e468cc8632ff351

Observation 8356357a-0ab1-4ed0-a846-89dbbf0fb2a7 · outbound

This paper cites Something Something.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Something Something

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.028366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.088103Z digest=sha256:6d93bdfc59c952a6ef095435ac68f6da0396f78775ec8e4f34231f22b33b84d2

Observation 48d26ae9-a8df-4a56-a7a8-807789c60ed5 · outbound

This paper cites Ava: A video dataset of spatio-temporally localized atomic visual actions.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Ava: A video dataset of spatio-temporally localized atomic visual actions

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.009196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.093921Z digest=sha256:639d5d3dd99450321cb8d7e25bc89c9fcafe1f0c5b87b1264f11ba1dc9561f9b

Observation ce2de66e-540f-4bbb-87b6-0d4b12c9a203 · outbound

This paper cites Youtube2text: Recognizing and describing arbitrary activities using semantic hierarchies and zero-shot recognition.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Youtube2text: Recognizing and describing arbitrary activities using semantic hierarchies and zero-shot recognition

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.985015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.099943Z digest=sha256:deb03c9be7777fe517df3f35a4f9548a3c2e91eab56160aca7296350ebc6f47a

Observation b1399349-65c3-4448-af8d-9686b1960d0c · outbound

This paper cites an unresolved cited work.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:15:57.966058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.106075Z digest=sha256:b42fd4c58e2e525efc0b3ff5287c58ad7d386aac0bf04cf37e82fad36c74cf17

Observation 47c7d464-2874-4da8-b38a-bda26323ec06 · outbound

This paper cites In Defense of the Triplet Loss for Person Re-Identification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings In Defense of the Triplet Loss for Person Re-Identification

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.111220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.111220Z digest=sha256:0e530c2c7d4f4e0103de0c62f36286c20f8f479e38ab5e72a1ce56979325ed1c

Observation cce0550f-0af1-45af-bb9b-8b78364b9faa · outbound

This paper cites Deep metric learning using triplet network.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Deep metric learning using triplet network

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.945593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.116903Z digest=sha256:26cb95d8f986fcfede4bf9bed71ec32d1adf03a5a816df6305298deda3678872

Observation 46f38826-3731-4783-a1ae-56bcc6ed91d4 · outbound

This paper cites Unifying visual-semantic embeddings with multimodal neu- ral language models.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unifying visual-semantic embeddings with multimodal neu- ral language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.927132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.122863Z digest=sha256:c6e5c75b3090a4c2c8e3abe3c8f3dc7cabc3778b3838e1cc687ed41e2f184900

Observation 6d57a355-c19c-4736-9c72-0db32e089aa0 · outbound

This paper cites Learning two-branch neural networks for image-text match- ing tasks.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning two-branch neural networks for image-text match- ing tasks

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.908471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.127947Z digest=sha256:cdd27401647b2adae5839e836430033b9629465ace5eb9c1b36a16850c3962fa

Observation 165f5939-6426-4545-b699-ce487380717f · outbound

This paper cites On the effectiveness of task granularity for transfer learning.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings On the effectiveness of task granularity for transfer learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.133392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.133392Z digest=sha256:3f7534bc19711c63f0c2bbc089eb39f4aeec4eb32981feefa18b3c2a1198b939

Observation 611c895f-318a-438d-82f6-b8eef42c219e · outbound

This paper cites Learnable pooling with Context Gating for video classification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learnable pooling with Context Gating for video classification

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.139487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.139487Z digest=sha256:81990904badc29c3d12245e4438c56a116a165f2ff1b1fac74da850c5c2709ea

Observation ad39cbdd-ded8-4ef3-a79b-62d059664eaf · outbound

This paper cites Learning a Text-Video Embedding from Incomplete and Heterogeneous Data.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning a Text-Video Embedding from Incomplete and Heterogeneous Data

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.144801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.144801Z digest=sha256:3463cdaa1b349ba17c2992eae603dc0c3f516144d935ed463ecad8aaa94447b6

Observation aac8fecd-fc60-415d-9bb5-d576ca70a05c · outbound

This paper cites HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.150310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.150310Z digest=sha256:8a7796a7bf5e996407cb1ad1bb70ccd2a436387564838ad7744d373f4b525d41

Observation 8e5e29eb-be5f-4f25-9112-4a3b710ce7a2 · outbound

This paper cites Learning joint embedding with multimodal cues for cross-modal video-text retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning joint embedding with multimodal cues for cross-modal video-text retrieval

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.878011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.157125Z digest=sha256:a0f30ae7f0587e9bca71c7e8668b4ea684b425fea5ed2be516216cc92dd98969

Observation bcaa3663-4b73-45ec-8b97-3e3a89210ea3 · outbound

This paper cites Learning joint representations of videos and sentences with web image search.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning joint representations of videos and sentences with web image search

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.859538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.162683Z digest=sha256:66dfb3a62be93db580d41644543bb7182126215c502e03f3a7d585d8df14ea21

Observation 23fb74cc-65c8-4b6a-88d3-d2106a05bdee · outbound

This paper cites Enhancing video summarization via vision-language embed- ding.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Enhancing video summarization via vision-language embed- ding

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.839909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.168392Z digest=sha256:536ceff84b36999ede423c11b9b3039200b3e505d63c0437aff9c85f024b41c7

Observation 9e4c1524-2906-440f-998e-d771ed3b57aa · outbound

This paper cites CNN image retrieval learns from BoW: Unsupervised fine-tuning with hard examples.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings CNN image retrieval learns from BoW: Unsupervised fine-tuning with hard examples

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.818744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.174858Z digest=sha256:cbc30e2cfbbe7c2d747de16ddc19f43b060220f04dfdfb84c080e72e079dc196

Observation 2aa1b1a3-a3bd-4aa1-b917-fb7cae7cb855 · outbound

This paper cites Recognizing fine-grained and composite ac- tivities using hand-centric features and script data.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Recognizing fine-grained and composite ac- tivities using hand-centric features and script data

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.797187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.181674Z digest=sha256:c4e0124c369f883ebbb343db900c83a1afc27ffc58f6ac5678336fe0c46c1cf6

Observation b72f6f25-cffc-4b26-8a4f-4af865fa33aa · outbound

This paper cites Facenet: A unified embedding for face recognition and clus- tering.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Facenet: A unified embedding for face recognition and clus- tering

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.773385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.190379Z digest=sha256:a408711685094a12327ea2c603cf824e34a39f96a030f342b761b0f5ae7f54f3

Observation 9967b099-b0d6-4ea4-8209-3ee0cbe3c6e2 · outbound

This paper cites Higher-order Network for Action Recognition.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Higher-order Network for Action Recognition

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:15:57.347997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.195289Z digest=sha256:4bf42a608924ea82ba1f0ea791cd2de7fbf561b47a445f88f7aec640c58e7996

Observation a2eab173-ae32-4ee7-b040-75ac41588ab3 · outbound

This paper cites Sigurdsson, G ¨ul Varol, Xiaolong Wang, Ali Farhadi, Ivan Laptev, and Abhinav Gupta.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Sigurdsson, G ¨ul Varol, Xiaolong Wang, Ali Farhadi, Ivan Laptev, and Abhinav Gupta

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.745079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.201647Z digest=sha256:e9b1a1f02e8d882922501c80c1a11b091772ad5c2e13611c75beacae51a8c746

Observation 888645c1-88be-4105-b163-bea0dd21e166 · outbound

This paper cites Improved deep metric learning with multi- class n-pair loss objective.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Improved deep metric learning with multi- class n-pair loss objective

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.726059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.207509Z digest=sha256:643d85a5c58c40bc8672b43055252d755e37d0ccfe54f0baf43dc910f3bd2d69

Observation ffd6b78b-0223-4de0-8d70-c4c8d8e10659 · outbound

This paper cites Cross modal embeddings for video and audio retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Cross modal embeddings for video and audio retrieval

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.705322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.213105Z digest=sha256:f3e8ec5b3fa21b55bba13966fbc5a789e2e8ca95b05a4a57ba0fd59349a529c9

Observation 8555b7f9-6dfc-4b21-8d2a-0122fb6050c0 · outbound

This paper cites Learning Language-Visual Embedding for Movie Understanding with Natural-Language.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning Language-Visual Embedding for Movie Understanding with Natural-Language

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.218372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.218372Z digest=sha256:aa4584b271d651b9febe71277eb6bb9a30e658818e412c8a27c9b05ee3ed268b

Observation df1b3d13-f745-42f3-b12f-0d53586ee99e · outbound

This paper cites Learn- ing fine-grained image similarity with deep ranking.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learn- ing fine-grained image similarity with deep ranking

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.683785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.224034Z digest=sha256:3c1c823731baf216c7810c7258f0349e91dbba05ff318668367ca811ea861a97

Observation 939fc6d5-a252-4f04-9e2f-fbf3a1fbdf57 · outbound

This paper cites Learning deep structure-preserving image-text embeddings.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning deep structure-preserving image-text embeddings

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.653864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.230640Z digest=sha256:3d4408b65cf62841ccc1e875f8e24d1bbe74b594241261613f8ac1c9151e125f

Observation ebca65a9-0e2d-4061-8e1c-0fc8167dc29a · outbound

This paper cites Temporal segment networks: Towards good practices for deep action recogni- tion.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Temporal segment networks: Towards good practices for deep action recogni- tion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.634381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.237032Z digest=sha256:33db1ef98b2bd5104ba3723971449673745de2c6d5a180dc945c78b90f2dd03d

Observation 1bb59219-bbbf-433e-aa3d-aacdf9ac1e7b · outbound

This paper cites Learning visual actions using multiple verb-only labels.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning visual actions using multiple verb-only labels

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.611464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.243318Z digest=sha256:66ed26ebee7cf5740a3d1d23bbd029cdb3c36f0bdd54ac3906317a20d17a1a9f

Observation e00a62a1-00b5-42b8-8970-f9002b97fd6a · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Msr-vtt: A large video description dataset for bridging video and language

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.591146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.249711Z digest=sha256:104bccb54b76f685af22a10b15b6cd424c0f939b513c495fcbbea5d867f0e84f

Observation 51108899-bda2-4313-823d-5bb99511e90c · outbound

This paper cites Jointly modeling deep video and compositional text to bridge vision and language in a unified framework.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Jointly modeling deep video and compositional text to bridge vision and language in a unified framework

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.556663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.255506Z digest=sha256:c8ba2681e6f472a36c1c8cb0f90082da2cef89c95a4a00103e40ddccb719b169

Observation ecc408dc-a83e-4a8f-8726-c74941c7f6b9 · outbound

This paper cites A joint se- quence fusion model for video question answering and re- trieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings A joint se- quence fusion model for video question answering and re- trieval

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.533791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.262176Z digest=sha256:7cde6ad60f49c5cb3faf9e651e14e6e22c5c2ff5847e4eab90561685b6c2acdd

Observation 78072827-6038-4ace-9a0b-2f39734e026c · outbound

This paper cites Zero-shot learning via semantic similarity embedding.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Zero-shot learning via semantic similarity embedding

Reference 42

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T14:15:57.515796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-14T14:15:57.272572Z digest=sha256:cb1f7e71da85412e1f020f22e3d89fce28834056739d2c48775892ff0b7e24e1

Pith citing papers

No inbound Pith citation observations are available.