Pith. sign in

Paper Citation Record · LEDGER

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings

As of 20 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:1908.03477.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.03477 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T14:15:57.272572Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy30
  • unresolved9
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 09347ffc-8b47-4a24-89f0-7dcf42b6192e · outbound

This paper cites an unresolved cited work.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:15:58.209242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.015401Z digest=sha256:a9835544214e900db555f28cf7a7721b24203d58ced79638360ad675e102d961

Observation 70e071cc-fe06-4f05-8010-266a3834f327 · outbound

This paper cites Re-ID done right: towards good practices for person re-identification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Re-ID done right: towards good practices for person re-identification

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.023091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.023091Z digest=sha256:57797f48a1f93ac89977f0d0cd5d8e7a72bc4cb051577b1c8e1a0d411d9a3372

Observation a7526b53-9e57-488c-8b35-236a261a4bca · outbound

This paper cites NetVLAD: CNN architecture for weakly supervised place recognition.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings NetVLAD: CNN architecture for weakly supervised place recognition

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.187494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.029900Z digest=sha256:6f73112e65144b212738e3787642f9a77a5bab0b4ab1574302b1e06edec4de5b

Observation eb52af3d-3229-4db1-b3f3-9c81d452bc61 · outbound

This paper cites An empirical study and analysis of generalized zero- shot learning for object recognition in the wild.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings An empirical study and analysis of generalized zero- shot learning for object recognition in the wild

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.160419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.036174Z digest=sha256:aea387a53615374840fff4e0416c4a46fa68b20343cd218370d43f1978af245c

Observation 40f1f8b8-8e51-435a-8b4e-ce75457b613a · outbound

This paper cites Beyond triplet loss: a deep quadruplet network for person re-identification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Beyond triplet loss: a deep quadruplet network for person re-identification

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.139875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.043008Z digest=sha256:259dfa3aa3923812987cb4b9ae7d0776f4c22521e643e8b21a96e48ac4e7e2ad

Observation 0b862c3b-5658-46a5-81f0-01806def26c8 · outbound

This paper cites Scaling egocentric vision: The epic-kitchens dataset.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Scaling egocentric vision: The epic-kitchens dataset

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.120942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.055743Z digest=sha256:1749e145a4f984c57c2f53dc6df632a660feea44f57c8046be65eaa57f60fbbe

Observation c4095747-5b8f-4326-9d83-56a92a40de35 · outbound

This paper cites Predict- ing visual features from text for image and video caption re- trieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Predict- ing visual features from text for image and video caption re- trieval

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.103558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.061105Z digest=sha256:7a3e543bf65de5b3bfa6eac333bcf5170c3f610257552b30791636a2517ed8b9

Observation 86de2abf-eb5e-4198-8626-0b4c4a2c1b66 · outbound

This paper cites Dual Encoding for Zero-Example Video Retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Dual Encoding for Zero-Example Video Retrieval

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:15:57.473760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.065902Z digest=sha256:cb18388618d910fb739e7d6f0bba45ed10e17a3f62d2f2e2ca2a447adc9ffa8f

Observation 91628ae4-1ca5-4ff3-a5a9-f518899f9046 · outbound

This paper cites Improving image-sentence embeddings using large weakly annotated photo collections.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Improving image-sentence embeddings using large weakly annotated photo collections

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.086177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.071958Z digest=sha256:e9ffe43e3660d8df6e7528c0cc849090ba52912678a234b1618973c52a46cf88

Observation 24c39c6a-2d93-41d4-ab83-5847f64df17e · outbound

This paper cites Deep image retrieval: Learning global representations for image search.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Deep image retrieval: Learning global representations for image search

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.065837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.077021Z digest=sha256:8583ba5287121df1dc5fbc0d580feba83547ed4c7fd0236c3fe11ab73e3221ac

Observation 437a4b4c-65a0-4334-91a8-3b7703c905ea · outbound

This paper cites Beyond instance-level im- age retrieval: Leveraging captions to learn a global visual representation for semantic retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Beyond instance-level im- age retrieval: Leveraging captions to learn a global visual representation for semantic retrieval

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.046935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.082890Z digest=sha256:e53b46628d731c0c82ab5558457ae2e65f9309b47711a378acc45d2e9d7e2b77

Observation 8356357a-0ab1-4ed0-a846-89dbbf0fb2a7 · outbound

This paper cites Something Something.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Something Something

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.028366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.088103Z digest=sha256:41fb59c290cf42df8244f581df8c9dbe65a5c88d0d0a6e487d806fbc997d3511

Observation 48d26ae9-a8df-4a56-a7a8-807789c60ed5 · outbound

This paper cites Ava: A video dataset of spatio-temporally localized atomic visual actions.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Ava: A video dataset of spatio-temporally localized atomic visual actions

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.009196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.093921Z digest=sha256:09166c80a5e0153042388a6fe258cb8476bafa2c9f034b29a39e7779933e9875

Observation ce2de66e-540f-4bbb-87b6-0d4b12c9a203 · outbound

This paper cites Youtube2text: Recognizing and describing arbitrary activities using semantic hierarchies and zero-shot recognition.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Youtube2text: Recognizing and describing arbitrary activities using semantic hierarchies and zero-shot recognition

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.985015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.099943Z digest=sha256:4dc520d7103aea4964fb427cda3de2fbd43be9bf8cd39c48186488e669de5425

Observation b1399349-65c3-4448-af8d-9686b1960d0c · outbound

This paper cites an unresolved cited work.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:15:57.966058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.106075Z digest=sha256:71a49777fb496372db5b7641de5635046d53548e97f9b97ad28f232ab384ab33

Observation 47c7d464-2874-4da8-b38a-bda26323ec06 · outbound

This paper cites In Defense of the Triplet Loss for Person Re-Identification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings In Defense of the Triplet Loss for Person Re-Identification

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.111220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.111220Z digest=sha256:0e530c2c7d4f4e0103de0c62f36286c20f8f479e38ab5e72a1ce56979325ed1c

Observation cce0550f-0af1-45af-bb9b-8b78364b9faa · outbound

This paper cites Deep metric learning using triplet network.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Deep metric learning using triplet network

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.945593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.116903Z digest=sha256:3767e9773a0c09b36887730dfe562824637f61b16b1d24b212cbe4da45feba3b

Observation 46f38826-3731-4783-a1ae-56bcc6ed91d4 · outbound

This paper cites Unifying visual-semantic embeddings with multimodal neu- ral language models.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unifying visual-semantic embeddings with multimodal neu- ral language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.927132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.122863Z digest=sha256:f9ecbc070481cf69ac3e5843e6fa8f73e4933dd8aae9c0d2226fa13072937d80

Observation 6d57a355-c19c-4736-9c72-0db32e089aa0 · outbound

This paper cites Learning two-branch neural networks for image-text match- ing tasks.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning two-branch neural networks for image-text match- ing tasks

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.908471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.127947Z digest=sha256:6452a23391d5c7b8db9304645e3b7f7b0635b6dbcc42ce9cdf5db659a3182a53

Observation 165f5939-6426-4545-b699-ce487380717f · outbound

This paper cites On the effectiveness of task granularity for transfer learning.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings On the effectiveness of task granularity for transfer learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.133392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.133392Z digest=sha256:3f7534bc19711c63f0c2bbc089eb39f4aeec4eb32981feefa18b3c2a1198b939

Observation 611c895f-318a-438d-82f6-b8eef42c219e · outbound

This paper cites Learnable pooling with Context Gating for video classification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learnable pooling with Context Gating for video classification

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.139487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.139487Z digest=sha256:81990904badc29c3d12245e4438c56a116a165f2ff1b1fac74da850c5c2709ea

Observation ad39cbdd-ded8-4ef3-a79b-62d059664eaf · outbound

This paper cites Learning a Text-Video Embedding from Incomplete and Heterogeneous Data.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning a Text-Video Embedding from Incomplete and Heterogeneous Data

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.144801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.144801Z digest=sha256:3463cdaa1b349ba17c2992eae603dc0c3f516144d935ed463ecad8aaa94447b6

Observation aac8fecd-fc60-415d-9bb5-d576ca70a05c · outbound

This paper cites HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.150310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.150310Z digest=sha256:8a7796a7bf5e996407cb1ad1bb70ccd2a436387564838ad7744d373f4b525d41

Observation 8e5e29eb-be5f-4f25-9112-4a3b710ce7a2 · outbound

This paper cites Learning joint embedding with multimodal cues for cross-modal video-text retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning joint embedding with multimodal cues for cross-modal video-text retrieval

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.878011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.157125Z digest=sha256:881753758f241a22d75d7a72eded0754e33cf0cb8f6a93ed43f6b5cc290cbcdc

Observation bcaa3663-4b73-45ec-8b97-3e3a89210ea3 · outbound

This paper cites Learning joint representations of videos and sentences with web image search.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning joint representations of videos and sentences with web image search

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.859538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.162683Z digest=sha256:49822d389eb04e63cb6c36831d483cc70ddc2cbb2bfa2ff01d8a8f2466b1ea28

Observation 23fb74cc-65c8-4b6a-88d3-d2106a05bdee · outbound

This paper cites Enhancing video summarization via vision-language embed- ding.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Enhancing video summarization via vision-language embed- ding

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.839909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.168392Z digest=sha256:ad2e6dd00d4c7362e51cf9c58daca14c6f4c2166c5d9d166468aea170fc6e7ec

Observation 9e4c1524-2906-440f-998e-d771ed3b57aa · outbound

This paper cites CNN image retrieval learns from BoW: Unsupervised fine-tuning with hard examples.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings CNN image retrieval learns from BoW: Unsupervised fine-tuning with hard examples

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.818744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.174858Z digest=sha256:ab24f0dd7fb4ffe6d69108de793c1ed1e8bdee256b4a90d22e8d5f8f9bce4cfd

Observation 2aa1b1a3-a3bd-4aa1-b917-fb7cae7cb855 · outbound

This paper cites Recognizing fine-grained and composite ac- tivities using hand-centric features and script data.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Recognizing fine-grained and composite ac- tivities using hand-centric features and script data

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.797187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.181674Z digest=sha256:43dd2696ae9b5105511086a97c682a7732e2e6471a4d3773dc544e967e372fac

Observation b72f6f25-cffc-4b26-8a4f-4af865fa33aa · outbound

This paper cites Facenet: A unified embedding for face recognition and clus- tering.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Facenet: A unified embedding for face recognition and clus- tering

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.773385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.190379Z digest=sha256:3f0864856354b594cfb7dbeadf8b6b0ff7fef825927925d437ead063852ec4d0

Observation 9967b099-b0d6-4ea4-8209-3ee0cbe3c6e2 · outbound

This paper cites Higher-order Network for Action Recognition.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Higher-order Network for Action Recognition

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:15:57.347997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.195289Z digest=sha256:7b57243cadbe82ff5c9e992c3010ad0f6ddfba5d6b043f3a0a038e3d89e181f4

Observation a2eab173-ae32-4ee7-b040-75ac41588ab3 · outbound

This paper cites Sigurdsson, G ¨ul Varol, Xiaolong Wang, Ali Farhadi, Ivan Laptev, and Abhinav Gupta.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Sigurdsson, G ¨ul Varol, Xiaolong Wang, Ali Farhadi, Ivan Laptev, and Abhinav Gupta

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.745079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.201647Z digest=sha256:11b4c32e412b3a3d9480dcc54c38c7103da6cfc752d78c918ce3605ce2a79d05

Observation 888645c1-88be-4105-b163-bea0dd21e166 · outbound

This paper cites Improved deep metric learning with multi- class n-pair loss objective.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Improved deep metric learning with multi- class n-pair loss objective

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.726059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.207509Z digest=sha256:df8d0cb0f93f560d67c7846a1e63c629f0451049af2ea635dcc7674f39787d27

Observation ffd6b78b-0223-4de0-8d70-c4c8d8e10659 · outbound

This paper cites Cross modal embeddings for video and audio retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Cross modal embeddings for video and audio retrieval

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.705322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.213105Z digest=sha256:01f508d03ce466ff6b93b359481cf9ddf778eb42e96fabba51c260e1bc484298

Observation 8555b7f9-6dfc-4b21-8d2a-0122fb6050c0 · outbound

This paper cites Learning Language-Visual Embedding for Movie Understanding with Natural-Language.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning Language-Visual Embedding for Movie Understanding with Natural-Language

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.218372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.218372Z digest=sha256:aa4584b271d651b9febe71277eb6bb9a30e658818e412c8a27c9b05ee3ed268b

Observation df1b3d13-f745-42f3-b12f-0d53586ee99e · outbound

This paper cites Learn- ing fine-grained image similarity with deep ranking.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learn- ing fine-grained image similarity with deep ranking

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.683785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.224034Z digest=sha256:8e624d0082c8dfea217d0fcf515654cc00dc798dd3c1cb4d0a32a5df7c6b7ad0

Observation 939fc6d5-a252-4f04-9e2f-fbf3a1fbdf57 · outbound

This paper cites Learning deep structure-preserving image-text embeddings.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning deep structure-preserving image-text embeddings

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.653864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.230640Z digest=sha256:2e82afe90ec1a84f513665bbfb482e45c94a39082152c59d9ca68b91b4f5d769

Observation ebca65a9-0e2d-4061-8e1c-0fc8167dc29a · outbound

This paper cites Temporal segment networks: Towards good practices for deep action recogni- tion.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Temporal segment networks: Towards good practices for deep action recogni- tion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.634381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.237032Z digest=sha256:88aeb84ed3c7fcf6aeac5f52ccf516e1764773f5d694bafd4602d29a006407ef

Observation 1bb59219-bbbf-433e-aa3d-aacdf9ac1e7b · outbound

This paper cites Learning visual actions using multiple verb-only labels.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning visual actions using multiple verb-only labels

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.611464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.243318Z digest=sha256:9070859d9c9d0d4ee7c6e21e369b1a52ba7cc5f459bfcdcf458b69cb6e57436c

Observation e00a62a1-00b5-42b8-8970-f9002b97fd6a · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Msr-vtt: A large video description dataset for bridging video and language

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.591146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.249711Z digest=sha256:98dbc2624072e6fadf2a7791a7177913202f7b3724a9827c6fe2b879217e957e

Observation 51108899-bda2-4313-823d-5bb99511e90c · outbound

This paper cites Jointly modeling deep video and compositional text to bridge vision and language in a unified framework.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Jointly modeling deep video and compositional text to bridge vision and language in a unified framework

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.556663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.255506Z digest=sha256:39514b12b95656b43bd1b59e43370bb6d79ed8d68fdaa2c1be1231887da6265d

Observation ecc408dc-a83e-4a8f-8726-c74941c7f6b9 · outbound

This paper cites A joint se- quence fusion model for video question answering and re- trieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings A joint se- quence fusion model for video question answering and re- trieval

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.533791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.262176Z digest=sha256:faf6495e5ee8fce9adfd15a2cb18970edfb9d6982fa571c8e9d12a8fbccf9c59

Observation 78072827-6038-4ace-9a0b-2f39734e026c · outbound

This paper cites Zero-shot learning via semantic similarity embedding.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Zero-shot learning via semantic similarity embedding

Reference 42

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T14:15:57.515796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T14:15:57.272572Z digest=sha256:28caf08527735ae45de17bd9d7a6eae74cfc2b7eca5e90fc2a2ac5dd3f08429e

Pith citing papers

No inbound Pith citation observations are available.