Pith. sign in

Paper Citation Record · LEDGER

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living

As of 10 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2502.03459.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03459 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:44:58.910397Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a09c5ffe-d7c5-4a20-a2c9-ce6ba00e2cae · outbound

This paper cites , " * write output.state after.block = add.period write newline.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.739680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.739680Z digest=sha256:3b36e83355550a8222513a5b23ec825b89f67db07f56756633efbc2acc89fbf6

Observation 4f6807ed-6db9-4ec3-98c4-febcecead612 · outbound

This paper cites write newline.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.743392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.743392Z digest=sha256:cff006216683d4136d07e4711b51ca2cdb89e3ac3b2529b0a9a35e1ba815c542

Observation cc067b34-46b2-4595-bee6-e9419a29f7f9 · outbound

This paper cites GPT-4 Technical Report.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.747112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.747112Z digest=sha256:b47b468761a0b287ab79cfdcad347bf2abe3306f132be43b60c203f1e5fbf23b

Observation f80afc5e-07e6-468e-982f-0eec120e20ca · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.384190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.750548Z digest=sha256:dc18631524e2ddfc6a28a0c6f2f8dc9c723112fbd9cac93efc3dca325494c14b

Observation 45365206-6f49-4d6a-b69d-8c516c560bfb · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.375481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.753636Z digest=sha256:7954fb1d5a295e96a349a81720e8072d8b7706de6f19bd84b8803f9626e85ff9

Observation f7a4e3ea-b264-435d-b50f-aba1b6f51a87 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.367040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.757401Z digest=sha256:0670c0440786d4c847a9ff88a8affa98b21f7137701c2c372a814ad42fe8c0f5

Observation b1657822-4607-476a-9942-8927d45b8dff · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.358290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.760473Z digest=sha256:2515ce5dd7536f26b2f819734314a09acbe7f8829b9af57eebf76dc01a637649

Observation 74124eed-04fb-4b42-aace-8b2990a08336 · outbound

This paper cites E.; Stoica, I.; and Xing, E.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living E.; Stoica, I.; and Xing, E

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.763620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.763620Z digest=sha256:3a01f9fd6829cc3028a9cb7b2a0e1edfd19941a14b9e10bedbf69d460f2a153b

Observation a46e8080-61a8-43a2-a237-a1b6a75d6267 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.344668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.766847Z digest=sha256:6afbaebde37c4ba16323af0c162e4448758d1e9574cb92d593fcb118f320552e

Observation 84ed8ecd-1680-455f-8062-2924e9e12524 · outbound

This paper cites Vision Transformers Need Registers.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Vision Transformers Need Registers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.770111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.770111Z digest=sha256:90f58a755dfa1b99bc95a9ec1ab6abafaedde999ee7ea724b042df24cfafd066

Observation 9768034e-e4d3-4f33-86a9-e67abcf99ef3 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.335786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.774535Z digest=sha256:f891078dae47cfd6d9d0909699efd279700b3c3cadb85dcbfceb78e346090048

Observation 6c824462-e257-48e1-9d20-41056c32535c · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.327639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.777364Z digest=sha256:ac14c7de3b688979fa8668dac29442f3766d9fd7652cdd2992ae080136e65ecd

Observation 4862d20d-3eb7-45d6-a138-0dc844aa3e0f · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.319320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.780311Z digest=sha256:981b7dd4c1767e674d35ecff55d1d16a65665aae820d82de589aaa83d3992504

Observation bc61fd4b-297a-4c8a-b3d0-f1e4cc34aef3 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.310476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.783021Z digest=sha256:f8277bed6253e5d327df140f57fabedae7919e0a46864058121d6a7c2ebeb8ee

Observation 1b8533ee-f3aa-4fc6-a219-9df01b1f972b · outbound

This paper cites K.; Sun, Y.; Patel, P.; and Black, M.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living K.; Sun, Y.; Patel, P.; and Black, M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:44:59.301511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.785815Z digest=sha256:ca99e8137b7ddeb3f5bbd54e68287aae777cba7b0e13329b556d53d0fac5de7b

Observation 4801c7b1-12b4-448e-a577-15719e8fa9f9 · outbound

This paper cites V.; Joulin, A.; and Misra, I.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living V.; Joulin, A.; and Misra, I

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.788512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.788512Z digest=sha256:0e26f932d826b14a664fcec493ea74cfb0bbe5fd7839a3cf8aaeb5f79c428df6

Observation 25f595f6-74c1-43a4-b478-16ed7a7bf540 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.287126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.791341Z digest=sha256:cdc38785582dd3f0d864c170f0f2dd36ea83a9c1b393731b40e445e54e08a674

Observation 8fcb828b-0ccd-480b-b82b-3bbfb925e32e · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.278740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.794349Z digest=sha256:a21f41b2cd536da94b9ee7fd7f52e410e20c215a948c2ace3fe9370c8412b3a6

Observation 04f3e484-de05-46b8-91b7-3b8f7a01426c · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.270413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.797434Z digest=sha256:c2bea8df60ee8447ccedc2cb264191ad46fcb70c1bbe8568cfa0a970dc4e98ad

Observation 2fb2dafd-00bc-47e2-977d-ad1b43a34929 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Distilling the Knowledge in a Neural Network

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.800197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.800197Z digest=sha256:958c32df3677977d132352e5142d4b7da953d62822cd7a080d85319d04551e3b

Observation 15bbbb5d-93ec-47fb-964e-e2320581e371 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.261860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.803531Z digest=sha256:1df3467a7b2d6c1b978885d385ec4f1f2143d313ef1610458b1d8c714eedc2cb

Observation c58284bd-b098-4379-af3f-343171e0e46b · outbound

This paper cites The Kinetics Human Action Video Dataset.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living The Kinetics Human Action Video Dataset

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.806472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.806472Z digest=sha256:7a049e243752030bdf83f08c1a20b4203bc236dfafd56d5be2d6b67c555eb33a

Observation 5f68b827-35f3-43fd-920a-3bd5c78e0da8 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.252627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.809749Z digest=sha256:86cb50f532fd351ca4836d062a77782a67cc3764d8e2abb3a784dbf4820b1b48

Observation f0185f63-076d-4d4e-bb3c-4941aa59afee · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.812526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.812526Z digest=sha256:3110f428d1bca72e1a7b19aa733e20959e28ed7c0dd219bbe233d30576108526

Observation 03564d78-d7cf-46c8-9c14-5f20059224c2 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.243761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.816041Z digest=sha256:f7385ce6201d6936159f96a7265ea654efa2544e576be6745f5520bdb5609f34

Observation 18821cef-6e58-4eb4-af69-e4f3a8c919e3 · outbound

This paper cites Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.818749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.818749Z digest=sha256:e0a0f4471e5c99d61d9047cda64bdfb1f7d206dd7e2ef935c74db9b810264748

Observation 67c4e852-cfd2-4f61-99a9-f7bc149e8cb3 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.234701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.821991Z digest=sha256:692df4e517f423997af901fb1173e366099e2497726630497f8441d17afc47a9

Observation 7c18f423-58a6-4cc6-bb56-b31f4dae4b8e · outbound

This paper cites The Llama 3 Herd of Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living The Llama 3 Herd of Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.824885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.824885Z digest=sha256:8e4d9206d46d3c30f038e7ce4754f3f8a11d6514bc91a3d0c18b779a36cc9bc9

Observation ff28a171-010d-40ad-8c36-95aef13f6509 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.225734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.827928Z digest=sha256:7dd158a562a02cfba694935bd0ee4b3e0af3939bbfa6480402c67da71700972c

Observation 749eccfb-893d-40db-ab5d-dcbd09789a76 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.217434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.830716Z digest=sha256:444ea972385cbf7ab593a5c9380105911ef1b12fdf4ce0a9b00a46461b0c1fc8

Observation b78052dd-89a8-4c07-bc01-c7ccf6886ff5 · outbound

This paper cites Multimodal Open-Vocabulary Video Classification via Pre-Trained Vision and Language Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Multimodal Open-Vocabulary Video Classification via Pre-Trained Vision and Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.833389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.833389Z digest=sha256:a6ca046743bd6b469c65b6faffe843f78882f7b345c58aa558a4ce17c296986f

Observation 49db148f-90d3-4744-b915-d2662f195c39 · outbound

This paper cites W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; Krueger, G.; and Sutskever, I.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; Krueger, G.; and Sutskever, I

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.836457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.836457Z digest=sha256:3c8866d6d4af2728b042827966333ec7e8061e312fab28828a745dc5df98fa93

Observation 0525f685-af0a-46c7-b05d-aeb541a94456 · outbound

This paper cites U.; Maaz, M.; Khan, S.; and Khan, F.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living U.; Maaz, M.; Khan, S.; and Khan, F

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:44:59.203460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.839261Z digest=sha256:7372bfeb53aabce0f2d5f66778c307fb0f9d2737978071e987be29cdd300bf5a

Observation a3eec8cb-faec-43b6-bc86-fae6cf883a5c · outbound

This paper cites LLAVIDAL: A Large LAnguage VIsion Model for Daily Activities of Living.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living LLAVIDAL: A Large LAnguage VIsion Model for Daily Activities of Living

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-08-09T04:44:59.026175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.841940Z digest=sha256:b0cb839eeb4e76198cef989b59083e57ac8accd8449b5bf192dc74e07cd10975

Observation 55c85646-dcd8-473f-b317-6461ed194c75 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.195267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.845095Z digest=sha256:118b72634bc42a0deb672758c37464c5828975096f1b2e4f588e1b6816b82fbc

Observation dd19305e-74af-4d64-8c16-5ffdbf45c023 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.187474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.847930Z digest=sha256:6aae7610aec507441bbc5d01b016e6062f89c9a632aeae1012e0722149c110c8

Observation 7d778381-38af-498e-91b2-57f95639bc67 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.178768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.850749Z digest=sha256:fde9233ea762a3cc302246b801767745afbdd1f41cac17556a781323603f7a92

Observation f28fa74d-1357-4fd5-ba01-e31703e02d8b · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.170988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.853607Z digest=sha256:8ad0ac33d317cb763e574fd3ec5b929479ef31d8b44bda6a689e09f08e089d79

Observation c94a6557-7f97-4894-9c2a-ed49b206b94a · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.162729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.856293Z digest=sha256:8d37fb7a1db33bea8185ed5d28d5fcea41bc5607bcf87460f83fa42818b05634

Observation d428651c-b0cb-4bde-8702-dc59826bd4d6 · outbound

This paper cites A.; Varol, G.; Wang, X.; Farhadi, A.; Laptev, I.; and Gupta, A.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living A.; Varol, G.; Wang, X.; Farhadi, A.; Laptev, I.; and Gupta, A

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:44:59.154375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.858982Z digest=sha256:f32f4e536754b2952ab64ba422da07b06bd7374439bc947b2689b5c87c0859c4

Observation 3c6ceb8f-dc0a-4012-bc1f-c1ed69ce8b63 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.145730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.861661Z digest=sha256:29d0327d44f3f543a785f500c0acbb963ebfcffee9ccaacb396fbe1216dbc844

Observation 809ffe0b-5ff6-47f5-86df-a1297608fc55 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.864413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.864413Z digest=sha256:402c0243a2e7142d94e57b138c20e356815e7c69761ac57e453faca588afd293

Observation 9b476259-d7b6-4e9e-afec-275cefff6106 · outbound

This paper cites PandaGPT: One Model To Instruction-Follow Them All.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living PandaGPT: One Model To Instruction-Follow Them All

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.867443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.867443Z digest=sha256:6798189649e4579ff954ae53b51e12d1769173608bb574b95299f0b3113f8163

Observation 7fbbbcde-d7f9-43f9-8396-ae4a38f4a4a1 · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.870821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.870821Z digest=sha256:fdf4c6e1ce127c8fbc371e06786f6b32ff1433f33497d6fd525f6ce1b33b8bb5

Observation 48467c18-75a4-4263-b5f4-cb0eb9e2c9e4 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living LLaMA: Open and Efficient Foundation Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.873900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.873900Z digest=sha256:56eecd6d8b69eb83fed6dde372439fc9535f32880d539a5f112dadd418c4f6f5

Observation 717a7f45-29f6-4f8d-a1c0-9a0d749f5d34 · outbound

This paper cites ActionCLIP: A New Paradigm for Video Action Recognition.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living ActionCLIP: A New Paradigm for Video Action Recognition

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.877082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.877082Z digest=sha256:e8eed75ed4671ed1e53bc1a1d65b65fa27191c6fdfe48a16f2e54f1b844fa8d3

Observation 9f8917dd-0a2a-4a17-889c-e93499148733 · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living CogVLM: Visual Expert for Pretrained Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.880265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.880265Z digest=sha256:6fa8b58543b7d9d9b2e6d86a2ce1cb670378a977ebd793e116d54aa88fbf40a2

Observation ea4c7457-1e85-4795-9aef-25e4d221c7f4 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.136172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.883755Z digest=sha256:e03233724fab2aa4e33d51f751bcd0131379fa696a7b21e95df2f1db80ded4bf

Observation c67b4078-98c3-43da-9dad-e60ba0f55d4f · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.127192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.886845Z digest=sha256:6a1c6255ea618565f4deccd42f8e93525b5fd4f8ac6cdc9b3691101db950561c

Observation c24f294b-10e9-4497-82aa-61c797d398aa · outbound

This paper cites CrossGLG: LLM Guides One-shot Skeleton-based 3D Action Recognition in a Cross-level Manner.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living CrossGLG: LLM Guides One-shot Skeleton-based 3D Action Recognition in a Cross-level Manner

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.889585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.889585Z digest=sha256:4f5d83c41778ed9f94b937d66f15b2fa2dff9c45b1007af5c02b58cfade5d9bd

Observation 5767af60-c35c-4b60-a239-ae281cc10ef7 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.117989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.892781Z digest=sha256:64831f8f4db50a947e781cd8585d348b5057b3745ef90f167e7629234803d1ac

Observation dafa6a34-9a96-4a51-b004-cbb2b41f7904 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.895795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.895795Z digest=sha256:33d9d52407ea988706b40c03139beacd3bad4f028fde49e8e7fc472190116353

Observation fe6622ee-4080-42c8-9e88-a4f32c5ec3cb · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.898707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.898707Z digest=sha256:d7e56e8e751bbdf9f43cb374f5f2601954e2dfdba9d71b9d7048bb3559521d5a

Observation 9602cfe7-07a7-48d8-836c-4fcd743cae2c · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.108218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.901971Z digest=sha256:4127a21a2dbf63f71a2b16f50da8f3dc09f8f48ce25bf2506ba887a4e14efdda

Observation 91f593e9-4257-4750-9abc-3b818c95ddbc · outbound

This paper cites Hypergraph Transformer for Skeleton-based Action Recognition.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Hypergraph Transformer for Skeleton-based Action Recognition

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.904658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.904658Z digest=sha256:42e2411b39b7aa701ef214a5f094d3b384cff64078428690875feb9a08c7ffe6

Observation 8cae0090-2e98-432c-8121-24a2515bc481 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.099260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.907612Z digest=sha256:df04500edbe69761bf7bdf5db54db73631fd442f3d9c44892ee7b41e539944bb

Observation 5d6eaf64-e9e8-4b4f-931f-72eaeb08b46f · outbound

This paper cites LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.910397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.910397Z digest=sha256:8bfa27e632d0557833c5c10d424ca7cfdcb619851fda1099e487da5b43951d68

Pith citing papers

No inbound Pith citation observations are available.