Pith. sign in

Paper Citation Record · LEDGER

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models

As of 19 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2412.06292.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06292 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:53:22.619176Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T14:49:39.643676Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T14:49:39.758012Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved13
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c2de93c0-e140-413a-9279-2a2bee636f38 · outbound

This paper cites Zero-shot 3d shape correspon- dence.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Zero-shot 3d shape correspon- dence

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.379194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.379194Z digest=sha256:27ecebee859172efdc04f8d97ef12b5751683cac593bc85901b33b6a245349c6

Observation 4dff6c20-b5f1-463f-8ac6-c950b1c21202 · outbound

This paper cites Satr: Zero-shot semantic segmentation of 3d shapes.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Satr: Zero-shot semantic segmentation of 3d shapes

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.521990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.385098Z digest=sha256:6db5309396e6e31c6aabc5e79740d253125837c9c92123023554518331acddaa

Observation 582e0e54-aceb-4f25-9f97-0410510bf224 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Flamingo: a visual language model for few-shot learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.389449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.389449Z digest=sha256:33c43a873f118a768a5034acf4a1eb17250497964de831fc3bde1c15f6eaef18

Observation 00ab1aeb-bb33-4a83-98a1-05180b557ac2 · outbound

This paper cites Claude 3.5 sonnet model card addendum.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Claude 3.5 sonnet model card addendum

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.489572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.394568Z digest=sha256:799669422f3ab7acdfc63afdefc92524058f6c5946087eeb734b1ade4a318405

Observation a087627a-9cb8-40bb-862a-825fbf1e4aed · outbound

This paper cites Ncp: Neural cor- respondence prior for effective unsupervised shape match- ing.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ncp: Neural cor- respondence prior for effective unsupervised shape match- ing

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.468354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.400519Z digest=sha256:5ad4a72a686af5988b9aca59eb0891da8c53e4ea19ad4860ae821af1b40f3c27

Observation f7dd7139-ff39-43db-ac0d-2391eb2563cc · outbound

This paper cites Language Models are Few-Shot Learners.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Language Models are Few-Shot Learners

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.405338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.405338Z digest=sha256:63bcc1907887390d14009dbef04b0034122f293265ff336a776237b1497af258

Observation e4a6b8df-d928-4025-8a99-7c7a08556dc7 · outbound

This paper cites Understanding the Limits of Vision Language Models Through the Lens of the Binding Problem.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Understanding the Limits of Vision Language Models Through the Lens of the Binding Problem

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.410059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.410059Z digest=sha256:bd91a528164125c9d7d9efea5b45d5451912edc6e1e400374492c52a93ce8ccb

Observation df6fb406-9c59-4005-8af0-f8f2fb23d155 · outbound

This paper cites Unsuper- vised learning of intrinsic structural representation points.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Unsuper- vised learning of intrinsic structural representation points

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.448792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.414870Z digest=sha256:9eabb8f7fd006dcabd5b00c8a2940e5df2ac3eb2d2c48eeb35b957d127423df7

Observation 53c55fa8-823e-4c88-96d4-b78f4bd50586 · outbound

This paper cites Schelling points on 3d surface meshes.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Schelling points on 3d surface meshes

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.429886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.419429Z digest=sha256:7e6dffdce5810c0f2bf09b6848448cb5d6d5ae3a7cb6120b1a34051bb0418646

Observation 470d239a-0fdf-4916-a21b-88394179106c · outbound

This paper cites 3dmv: Joint 3d-multi- view prediction for 3d semantic scene segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3dmv: Joint 3d-multi- view prediction for 3d semantic scene segmentation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.410607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.423666Z digest=sha256:9bf5729551f772dbfa53923828457033a3fcb405611abb427e7d2f89e8f9509d

Observation f77e4a32-2e12-4a46-aae2-878e19db2235 · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.427754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.427754Z digest=sha256:db6c58599dba2bb61cbd5ad881cd7951ee5c7d50c62ca1e86b5c65283f2d7fcb

Observation b39572e5-31b4-49c3-ab91-3ab8e891485f · outbound

This paper cites Unsupervised learning of category-specific symmetric 3d keypoints from point sets.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Unsupervised learning of category-specific symmetric 3d keypoints from point sets

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.382461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.432903Z digest=sha256:d1a1ade19d66cc2d723f16ae677add967fc9ea351f168403ff762c6c386555c2

Observation a1009d42-3e04-49a0-88bd-73d605595d87 · outbound

This paper cites Mvtn: Multi-view transformation network for 3d shape recognition.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Mvtn: Multi-view transformation network for 3d shape recognition

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.365304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.437148Z digest=sha256:0bce7a0a9939b8c37b9e84bccd96d6b619f87c4f52963f0a2045d81643e2cf3d

Observation 1871914a-392b-44be-989b-a3599c4d210b · outbound

This paper cites V oint cloud: Multi-view point cloud representation for 3d understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models V oint cloud: Multi-view point cloud representation for 3d understanding

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.349447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.441286Z digest=sha256:f6390df48f4116b65e2a2498e6d8e4936b379f135817e3571c7305de1f97be35

Observation 84a880fa-5ed3-415a-bb82-e681933076af · outbound

This paper cites Mvtn: Learning multi-view transforma- tions for 3d understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Mvtn: Learning multi-view transforma- tions for 3d understanding

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.331028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.445167Z digest=sha256:8111724f6acac85cadb22f3b1f9d67a8f2e39b183adef51276a72f9e66cdd810

Observation bc0da9e4-1860-4f29-9481-693403af7c7b · outbound

This paper cites Unsupervised keypoints from pretrained diffusion models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Unsupervised keypoints from pretrained diffusion models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.316667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.449699Z digest=sha256:62e6bd923dcd9f4f80824ddfea13a8d607facdfe42406d36224f07728b8ac860

Observation 596c8d96-1212-4984-93d2-282a39e49667 · outbound

This paper cites 3d-llm: In- jecting the 3d world into large language models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d-llm: In- jecting the 3d world into large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.454016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.454016Z digest=sha256:f3a21a1b49b9f81e49c38dc9e3f64bd89c680c0dbf099e9792636da84efffd19

Observation bd089bf1-b03f-462c-8563-3d35bf8b5767 · outbound

This paper cites 3d-sis: 3d se- mantic instance segmentation of rgb-d scans.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d-sis: 3d se- mantic instance segmentation of rgb-d scans

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.292681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.458692Z digest=sha256:cd549a708e3d2bcce9c6981767127dc846f2650d31bbd570c583932841074425

Observation 454c4372-5c82-4eb0-82ff-789fa2034217 · outbound

This paper cites Segment3d: Learning fine-grained class-agnostic 3d segmentation without manual labels.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Segment3d: Learning fine-grained class-agnostic 3d segmentation without manual labels

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.278453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.462899Z digest=sha256:ea267d26117f4fc7b78a1f01613550ebad1413add034af627bafdd32d626fc03

Observation 7d10388c-56b0-4ca7-9d4b-a6f7e0b2404c · outbound

This paper cites Keypointdeformer: 9 Unsupervised 3d keypoint discovery for shape control.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Keypointdeformer: 9 Unsupervised 3d keypoint discovery for shape control

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.260658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.467348Z digest=sha256:126c8ed3f81554e0fb4f78c16c400d9877efff8140f4019fbf2d5c55e65889f3

Observation 9e92acb9-cf9f-4a0e-b476-61628feddb11 · outbound

This paper cites Multi-view pointnet for 3d scene understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Multi-view pointnet for 3d scene understanding

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.245437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.472270Z digest=sha256:d9abc71002dad0bd10989f7918446ffd0a28d990ceba65f498397bae0aa4070c

Observation d2165763-6c21-4ef5-a648-823767707ba2 · outbound

This paper cites 3d shape segmentation with projective convolutional networks.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d shape segmentation with projective convolutional networks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.222504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.476533Z digest=sha256:ffeabbf3a33a21a19da5657369523b4c3ec718b114be513f7955e00618022085

Observation b4bb8479-64b2-47de-849b-dd1fa8933abe · outbound

This paper cites Berg, Wan-Yen Lo, Piotr Doll ´ar, and Ross Girshick.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Berg, Wan-Yen Lo, Piotr Doll ´ar, and Ross Girshick

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.480826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.480826Z digest=sha256:e60e898d566e0f9602c0eb8a1fbfe149d76913c82559f1f7f6392f7f45127a57

Observation 0bdbfc53-5879-42ec-9077-6889205c388f · outbound

This paper cites Virtual multi-view fusion for 3d semantic segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Virtual multi-view fusion for 3d semantic segmentation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.188206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.485866Z digest=sha256:f2831f2d14f00362e106154c9c68594f5e287cfc73a30917005d8d6eff20e34d

Observation e2c303d6-0c1c-48af-bc3b-6faf051f9de1 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.490024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.490024Z digest=sha256:6b6f32210ac8f8dd41430d524f79a1727238fe182b84a05ac8dd33e088253638

Observation eac79e8c-9230-404e-9faa-24db22f008a2 · outbound

This paper cites Visual instruction tuning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Visual instruction tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.494848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.494848Z digest=sha256:dcebfd42242d86d6ea8cec7f347cb27f80f7ea1b6bb827457a889f3fc67c5a65

Observation 72cfa958-02dc-4787-a8b9-329f22a9cae1 · outbound

This paper cites Improved baselines with visual instruction tuning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Improved baselines with visual instruction tuning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.146801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.499396Z digest=sha256:47f2cefbf994995cd9a015bcb608b0ff3bbd50ecabf1ac93e3c4b8803527cf10

Observation fcd044aa-136b-48e1-b8d8-b2c514ab1d65 · outbound

This paper cites 3d-to-2d distillation for indoor scene parsing.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d-to-2d distillation for indoor scene parsing

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.131318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.503575Z digest=sha256:5d730398840ec1d78d678134ef10c874282b407c5dd216ea20c4ece26d391706

Observation 71a2aa23-2dfa-40ad-a580-96362a4f41e8 · outbound

This paper cites Learning to segment 3d point clouds in 2d image space.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Learning to segment 3d point clouds in 2d image space

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.116382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.508556Z digest=sha256:d630829465bfa43e78fa92be3d982ae5ffb222a776e8834937a9ea0d7f84056d

Observation 378073f9-a7d3-461a-b7ef-d29ea10de51b · outbound

This paper cites Egoloc: Revisiting 3d object localiza- tion from egocentric videos with visual queries.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Egoloc: Revisiting 3d object localiza- tion from egocentric videos with visual queries

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.093838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.512221Z digest=sha256:54b110937ac18386e8604e2b267688c8d1f465c4b97234229373c9655d98adfb

Observation cc9a7285-2371-41ab-9431-02124fdbfb9b · outbound

This paper cites Tracknerf: Bundle adjusting nerf from sparse and noisy views via feature tracks, 2024.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Tracknerf: Bundle adjusting nerf from sparse and noisy views via feature tracks, 2024

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.073993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.516273Z digest=sha256:a762b209fb3028c5cbb943e9eb2277d3a984225cdc62c8f1a1e2f814a7c27d9d

Observation afb41c8c-3f9f-4502-8ed9-0bd41c70655a · outbound

This paper cites Gpt-4 technical report, 2023.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Gpt-4 technical report, 2023

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.048343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.520516Z digest=sha256:735f13bc2fecb71b5a55c3b681e7958e482267b13ec12c693ae438690fe92196

Observation 455dfd8b-5d4b-4918-ab37-1a55510ad9ae · outbound

This paper cites Synthesize diagnose and optimize: Towards fine- grained vision-language understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Synthesize diagnose and optimize: Towards fine- grained vision-language understanding

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.031427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.525276Z digest=sha256:4659f00b6d77f4919345f6578036fbefb63e1583e1404b93b5b99bf47d3337d8

Observation 46b9cf11-d76a-4e8f-9c9a-e24ae4926aa2 · outbound

This paper cites Shapellm: Universal 3d object understanding for embodied interaction.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Shapellm: Universal 3d object understanding for embodied interaction

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.014512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.529670Z digest=sha256:ec31ba36b7c706ee0c45d8ec5d43867d8796cde90603f9adfa9f1c433a5aa468

Observation f1ac552a-836b-475c-b835-1e4d326928ce · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Learn- ing transferable visual models from natural language super- vision

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.998627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.534012Z digest=sha256:3a91931639e45be83f81b93365c05411c4f518d2098e54aeb31cc4d6d716125b

Observation 39761a1e-ab19-4615-912c-6fc91ee58bac · outbound

This paper cites Vision language models are blind: Failing to translate detailed visual features into words.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Vision language models are blind: Failing to translate detailed visual features into words

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.538799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.538799Z digest=sha256:f4f97aaf5249a80814b43d001b966a4fa9ecdcfb8d2910540b9cd87f9bdaa6c4

Observation 31eb0969-dfc1-488c-8d08-2a5b9fb900fe · outbound

This paper cites The Strategy of Conflict: with a new Preface by the Author.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models The Strategy of Conflict: with a new Preface by the Author

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.979666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.543453Z digest=sha256:628b61419bff59c494362974d25eab2936bdd18f56ff569bd0e8ebcbd771b606

Observation 925db8d7-fa42-4053-8bc0-8ba92a4a68d8 · outbound

This paper cites Mask3D: Mask Trans- former for 3D Semantic Instance Segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Mask3D: Mask Trans- former for 3D Semantic Instance Segmentation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.548106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.548106Z digest=sha256:f2887b1a1388c9b975ba0c0db45faed98e4bad0c7769d2bd77bc7311e3088fea

Observation 5b34724c-af20-4f5f-824e-be6c83aea78c · outbound

This paper cites Skele- ton merger: an unsupervised aligned keypoint detector.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Skele- ton merger: an unsupervised aligned keypoint detector

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.949277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.553206Z digest=sha256:e2f350c56b9be7836ad5298de48b1bcb5056e6922b7587d6c59fbba4bc0a3ce2

Observation ed0f51ef-4cab-4922-b4fa-84dc6f63c8d0 · outbound

This paper cites What does clip know about a red circle? vi- sual prompt engineering for vlms.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models What does clip know about a red circle? vi- sual prompt engineering for vlms

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.935249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.557535Z digest=sha256:4da859c26b59db583ed6cd6efa2cb5a6366f56b9242999b8fed4cf7438e34b23

Observation 294a82b1-c7bf-42fa-8116-ce496b9ff8a2 · outbound

This paper cites Discovery of latent 3d key- points via end-to-end geometric reasoning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Discovery of latent 3d key- points via end-to-end geometric reasoning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.920888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.561859Z digest=sha256:7e5b23157a4280daca255773a9a390890b717d5484f4887f7e31e9a54cd96e0d

Observation 353481c6-fc2a-4a3d-854d-4c931b2e2dec · outbound

This paper cites Ldls: 3- d object segmentation through label diffusion from 2-d im- ages.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ldls: 3- d object segmentation through label diffusion from 2-d im- ages

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.905419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.566259Z digest=sha256:7a9f0d089852e21bb13f7f44550ca24826794ee0c3989d874e8c10a75739f61f

Observation 4dc85b4c-cbf5-42e4-bb6c-dfdffee05be7 · outbound

This paper cites Learning 3d keypoint descriptors for non-rigid shape matching.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Learning 3d keypoint descriptors for non-rigid shape matching

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.889417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.570113Z digest=sha256:73c13834089269bdd73ab7f9909dfba30dea28c09d4dd81d61e5981784f0a436

Observation 6b129e28-9bf1-4a00-ae2f-7e21ddee4aa1 · outbound

This paper cites Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.871745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.574293Z digest=sha256:8adac6f59c403eda20a9b85d9eb9003a0f3dae7ecbe330a82e940630a88ccc37

Observation 88e02f07-2665-424c-a728-65d545fac1b2 · outbound

This paper cites Back to 3d: Few-shot 3d keypoint detection with back-projected 2d features.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Back to 3d: Few-shot 3d keypoint detection with back-projected 2d features

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.856075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.578146Z digest=sha256:b61fe960be659bf6bf64ae42be5d16d49b32e8aaf870ee3d0008231d897a100f

Observation b8f9746f-4395-4082-abf0-1c4842dc3551 · outbound

This paper cites Clip-dinoiser: Teaching clip a few dino tricks for open- 10 vocabulary semantic segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Clip-dinoiser: Teaching clip a few dino tricks for open- 10 vocabulary semantic segmentation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.840684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.582597Z digest=sha256:d94fef875864667120c309c1de34293c509c123a51edb5739ba4c01be05e5155

Observation d80eac84-56fd-48b9-9b57-fb748e5681a0 · outbound

This paper cites Pointllm: Empowering large language models to understand point clouds.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Pointllm: Empowering large language models to understand point clouds

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.824142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.586781Z digest=sha256:05664b7310276779e6871138709d840c3bd9b376ede9b391b4ca651699e0aec1

Observation 203fd48a-fece-48b8-b6a6-746e9186fad3 · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.809708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.591199Z digest=sha256:1781ae9c06eed79e98e7ffd23e6953a03365cb3706a563fb68f0c2b058a868ee

Observation aef7633b-abfa-4629-9c13-1e13f93132ff · outbound

This paper cites 3dfeat-net: Weakly su- pervised local 3d features for point cloud registration.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3dfeat-net: Weakly su- pervised local 3d features for point cloud registration

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.795136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.595690Z digest=sha256:e3eafb056ac9f8fe7236a8222c6f35409bdf7081496f02e391dd310734a352d0

Observation 8b26c0c1-3579-445e-97dd-21a712c5e4c7 · outbound

This paper cites KeypointNet: A Large-scale 3D Keypoint Dataset Aggregated from Numerous Human Annotations.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models KeypointNet: A Large-scale 3D Keypoint Dataset Aggregated from Numerous Human Annotations

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.600534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.600534Z digest=sha256:2e599d91f173f1ac0a34def294e2385d8ed0a2b4614f96e7986080d76edbd3cf

Observation 276695df-040c-467d-9e47-b5b8d2137c6d · outbound

This paper cites Ukpgan: A general self-supervised keypoint detector.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ukpgan: A general self-supervised keypoint detector

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.780211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.605173Z digest=sha256:1e0e0b4e5190285509f48f0247fff32e88e3d4298af94b610d15ee93a63dfb3d

Observation 6f564cf9-386f-447d-8be6-636e7c1d5f14 · outbound

This paper cites Good at captioning bad at counting: Benchmarking gpt-4v on earth observation data.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Good at captioning bad at counting: Benchmarking gpt-4v on earth observation data

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.765899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T19:53:22.610243Z digest=sha256:872b4c896d47b54eeaba77207426782ff9984e2110e5fed7e11300acd7177c36

Observation 902ee351-1ecc-4d3f-bc1e-01008d542817 · outbound

This paper cites Uni3D: Exploring Unified 3D Representation at Scale.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Uni3D: Exploring Unified 3D Representation at Scale

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.614137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.614137Z digest=sha256:bb1bcd9ff0e6432a2999c107a1cdb489ca51922aeb9f50aca37d4ce07da6748b

Observation 9b53a1ec-0163-47b0-b616-b3d452b8b755 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 54

Resolution
malformed identifier
no resolver link, observed 2026-08-11T19:53:22.619176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.619176Z digest=sha256:e3f19106d390a78c1ce5cc099f4444fb7cdd423ab2967d229aef7f6def09ff32

Pith citing papers

Observation 7f9daa15-1422-4e9f-ba24-35f50659d139 · inbound

Estimating 2D Keypoints of Surgical Tools Using Vision-Language Models with Low-Rank Adaptation cites this paper.

Estimating 2D Keypoints of Surgical Tools Using Vision-Language Models with Low-Rank Adaptation ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T14:49:39.762699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T14:49:39.643676Z digest=sha256:925d4d9ac08275ae2117af7ec41c103efb1ee3f0e90277f4f3c162f194b73aa4