Pith. sign in

Paper Citation Record · LEDGER

LMM-Regularized CLIP Embeddings for Image Classification

As of 16 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2412.11663.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11663 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:45:44.879975Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0ccaf34c-13f4-4856-83cb-9026ace0d370 · outbound

This paper cites Learning transferable visual models from natural language supervision.

LMM-Regularized CLIP Embeddings for Image Classification Learning transferable visual models from natural language supervision

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.817992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.817992Z digest=sha256:2129a8cfb2dad2340444eddb09ea915918257ffa0ee7f57e70f85fec0fffdbbe

Observation bff2d027-2e6c-4b3b-8206-82be550f3672 · outbound

This paper cites GestureDiffuCLIP: Gesture Diffusion Model with CLIP Latents.

LMM-Regularized CLIP Embeddings for Image Classification GestureDiffuCLIP: Gesture Diffusion Model with CLIP Latents

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.822200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.822200Z digest=sha256:75968b7a5859c19f0be05ffc05680d5c4ee5b53baf7160e3a2d02fc7dd6f0f44

Observation e3017430-f5b6-449f-9971-2adb852e1ad2 · outbound

This paper cites Delving into clip latent space for video anomaly recognition.

LMM-Regularized CLIP Embeddings for Image Classification Delving into clip latent space for video anomaly recognition

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.069767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T14:45:44.826196Z digest=sha256:db219b7e3c7208b71ae1719e19598364db9dd3095adfb1b0da88918b58199a2f

Observation ad9fdcca-b4c0-493d-8815-ad472a137480 · outbound

This paper cites A Survey of Large Language Models.

LMM-Regularized CLIP Embeddings for Image Classification A Survey of Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.829990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.829990Z digest=sha256:48a998e85ea283701b1bce7b891a0de7db066e32c47b4cb7c9d75cbc62b9d1af

Observation 80e931cc-c0db-493e-9f9d-2e8744e85ee5 · outbound

This paper cites A Survey on Multimodal Large Language Models.

LMM-Regularized CLIP Embeddings for Image Classification A Survey on Multimodal Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.834033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.834033Z digest=sha256:b94058cf09567260a7aa40d38d4eaacf48e7a414a956a0c403da751d9d23ec1d

Observation 51e91cb0-b2c8-41fb-85bd-152915dd2aba · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

LMM-Regularized CLIP Embeddings for Image Classification MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.838160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.838160Z digest=sha256:b35dadd4d9d484358cf237d5f67b4342f4752816aa3a0b2148cc7db1b09615cf

Observation 7bf1e90a-37f0-4fc5-aaa3-3e90e3a7a762 · outbound

This paper cites Graph embedded convolutional neural networks in human crowd detection for drone flight safety.

LMM-Regularized CLIP Embeddings for Image Classification Graph embedded convolutional neural networks in human crowd detection for drone flight safety

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.059429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T14:45:44.842426Z digest=sha256:19429b65cc03602a3ab99211726ae949505770a5564d6048b36e009196f3184b

Observation c6f337d1-6a24-4637-85cc-1d5291fac1d6 · outbound

This paper cites Learning to prompt for vision-language models.

LMM-Regularized CLIP Embeddings for Image Classification Learning to prompt for vision-language models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.845777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.845777Z digest=sha256:704afeeaf9ad1caff0eee04fbd22e944faa2635eb62815aa9006b812d87bc702

Observation f02c98b0-69ba-48dc-9003-aed864977b57 · outbound

This paper cites Conditional prompt learning for vision-language models.

LMM-Regularized CLIP Embeddings for Image Classification Conditional prompt learning for vision-language models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.042173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T14:45:44.849154Z digest=sha256:040c4e2d61de000073ca93169cdbfa6f0a705b5756d447a6475e645f09188c00

Observation e8c474e3-522b-422d-87c2-a096e4333eca · outbound

This paper cites Language in a bottle: Language model guided concept bottlenecks for interpretable image classification.

LMM-Regularized CLIP Embeddings for Image Classification Language in a bottle: Language model guided concept bottlenecks for interpretable image classification

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.030547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T14:45:44.852508Z digest=sha256:f469fec87b9d0c8cd6d284de4da77a4a443c1de4904b11152d91b2a31a1405d6

Observation 0140eeeb-c6d1-4c3a-a6d3-7ad68c9eaf54 · outbound

This paper cites Language models are few-shot learners.

LMM-Regularized CLIP Embeddings for Image Classification Language models are few-shot learners

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.855619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.855619Z digest=sha256:4a0b543ed8b09522b0381b255a19744f2ca42fd1e75bfa74fb4660a4cd7be745

Observation 98ff77d4-3c2f-48f5-a56e-80b242af9b1e · outbound

This paper cites En- hancing clip with gpt-4: Harnessing visual descriptions as prompts.

LMM-Regularized CLIP Embeddings for Image Classification En- hancing clip with gpt-4: Harnessing visual descriptions as prompts

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.012438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T14:45:44.859052Z digest=sha256:649f9a40a079ebda8f6c078c835ad30bd41d8395768770c43c5066f2f811a2e1

Observation 4f2ecd95-2697-4bb9-9e48-2285ca6f157c · outbound

This paper cites GPT-4 Technical Report.

LMM-Regularized CLIP Embeddings for Image Classification GPT-4 Technical Report

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.862809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.862809Z digest=sha256:9a5485157d7e1b132f04a87e8a6a684e16dcf6d6b711a4ab16125070391f7517

Observation b685c802-8fae-4abc-8f85-0037d437b2a5 · outbound

This paper cites Exploiting lmm-based knowledge for image classification tasks.

LMM-Regularized CLIP Embeddings for Image Classification Exploiting lmm-based knowledge for image classification tasks

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.001520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T14:45:44.866432Z digest=sha256:3065fea3bdd9e66aa43eeefd47d7693930e58ed2bc70200fc985341f2e5d573b

Observation 684881ee-5e3d-472e-af5c-cd6cb9315d5e · outbound

This paper cites Disturbing image detection using lmm-elicited emotion embeddings.

LMM-Regularized CLIP Embeddings for Image Classification Disturbing image detection using lmm-elicited emotion embeddings

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:44.990963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T14:45:44.869730Z digest=sha256:571bdb9fb4f038472dbe56be93ba8f4307f926d038ad74bd5d73e1c8e0c67e49

Observation 07e99017-8748-4185-9ee3-97db8e30c678 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

LMM-Regularized CLIP Embeddings for Image Classification UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.872958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.872958Z digest=sha256:fa66ad85a659866dbe87a6a8f8bbdd0eea2ef131cc48c3c7aaa87578d68269ad

Observation 16b33324-d2fa-4497-9fe9-38f7cc0f2802 · outbound

This paper cites an unresolved cited work.

LMM-Regularized CLIP Embeddings for Image Classification Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:45:44.980271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T14:45:44.876510Z digest=sha256:0cd69df790e25c81b32c3554d44655650cdf85b93276be396efa47b8354845a6

Observation fc70647e-bf9e-4efa-9771-47db5b465892 · outbound

This paper cites Learning from failure: De-biasing classifier from biased classifier.

LMM-Regularized CLIP Embeddings for Image Classification Learning from failure: De-biasing classifier from biased classifier

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:44.968822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T14:45:44.879975Z digest=sha256:9dcc164b3d3b5837528e792e6cd7fd1c74e9d409acf5d200848532a16620dbea

Pith citing papers

No inbound Pith citation observations are available.