Pith. sign in

Paper Citation Record · LEDGER

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models

As of 19 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2412.06292.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06292 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:53:22.619176Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T14:49:39.643676Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T14:49:39.758012Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved13
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c2de93c0-e140-413a-9279-2a2bee636f38 · outbound

This paper cites Zero-shot 3d shape correspon- dence.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Zero-shot 3d shape correspon- dence

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.379194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.379194Z digest=sha256:27ecebee859172efdc04f8d97ef12b5751683cac593bc85901b33b6a245349c6

Observation 4dff6c20-b5f1-463f-8ac6-c950b1c21202 · outbound

This paper cites Satr: Zero-shot semantic segmentation of 3d shapes.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Satr: Zero-shot semantic segmentation of 3d shapes

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.521990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.385098Z digest=sha256:77f93f2ac6343f947a1c5a98d521c8f5efd24f5d6fcfb41be1ff578ab1af480f

Observation 582e0e54-aceb-4f25-9f97-0410510bf224 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Flamingo: a visual language model for few-shot learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.389449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.389449Z digest=sha256:33c43a873f118a768a5034acf4a1eb17250497964de831fc3bde1c15f6eaef18

Observation 00ab1aeb-bb33-4a83-98a1-05180b557ac2 · outbound

This paper cites Claude 3.5 sonnet model card addendum.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Claude 3.5 sonnet model card addendum

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.489572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.394568Z digest=sha256:2a6891a77c69a1d9aa6d958ace0ac7d7f1485c5234ce1c4137978c7e34a8de95

Observation a087627a-9cb8-40bb-862a-825fbf1e4aed · outbound

This paper cites Ncp: Neural cor- respondence prior for effective unsupervised shape match- ing.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ncp: Neural cor- respondence prior for effective unsupervised shape match- ing

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.468354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.400519Z digest=sha256:1b002b5df2345a65924c554f828aec990da00c522deaed2b07006966c4b21297

Observation f7dd7139-ff39-43db-ac0d-2391eb2563cc · outbound

This paper cites Language Models are Few-Shot Learners.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Language Models are Few-Shot Learners

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.405338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.405338Z digest=sha256:63bcc1907887390d14009dbef04b0034122f293265ff336a776237b1497af258

Observation e4a6b8df-d928-4025-8a99-7c7a08556dc7 · outbound

This paper cites Understanding the Limits of Vision Language Models Through the Lens of the Binding Problem.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Understanding the Limits of Vision Language Models Through the Lens of the Binding Problem

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.410059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.410059Z digest=sha256:bd91a528164125c9d7d9efea5b45d5451912edc6e1e400374492c52a93ce8ccb

Observation df6fb406-9c59-4005-8af0-f8f2fb23d155 · outbound

This paper cites Unsuper- vised learning of intrinsic structural representation points.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Unsuper- vised learning of intrinsic structural representation points

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.448792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.414870Z digest=sha256:064c3d09c6af9fa13be38cf6961e3dd26e52aee508ec34989cf793c767a13a37

Observation 53c55fa8-823e-4c88-96d4-b78f4bd50586 · outbound

This paper cites Schelling points on 3d surface meshes.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Schelling points on 3d surface meshes

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.429886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.419429Z digest=sha256:833a932ad73cfda61b2bc7d1747662e1aaf61e60cbccf78ebcb7a3392f95987a

Observation 470d239a-0fdf-4916-a21b-88394179106c · outbound

This paper cites 3dmv: Joint 3d-multi- view prediction for 3d semantic scene segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3dmv: Joint 3d-multi- view prediction for 3d semantic scene segmentation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.410607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.423666Z digest=sha256:58bfec2c82f59cbab7cd7daa4326c7f2a71523d948a4a5933bc3720c9ee1eff8

Observation f77e4a32-2e12-4a46-aae2-878e19db2235 · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.427754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.427754Z digest=sha256:db6c58599dba2bb61cbd5ad881cd7951ee5c7d50c62ca1e86b5c65283f2d7fcb

Observation b39572e5-31b4-49c3-ab91-3ab8e891485f · outbound

This paper cites Unsupervised learning of category-specific symmetric 3d keypoints from point sets.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Unsupervised learning of category-specific symmetric 3d keypoints from point sets

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.382461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.432903Z digest=sha256:d1d5854eace4efd105dc0a77ac87cba65a125ed9aa4098c8aee8b1b6f65f04dd

Observation a1009d42-3e04-49a0-88bd-73d605595d87 · outbound

This paper cites Mvtn: Multi-view transformation network for 3d shape recognition.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Mvtn: Multi-view transformation network for 3d shape recognition

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.365304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.437148Z digest=sha256:82b8a06ca1b0689386e81e41199505888b060e5b0e18067c2a2ae1fd782ce561

Observation 1871914a-392b-44be-989b-a3599c4d210b · outbound

This paper cites V oint cloud: Multi-view point cloud representation for 3d understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models V oint cloud: Multi-view point cloud representation for 3d understanding

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.349447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.441286Z digest=sha256:b7f845956b6b4f7a560af9f88543fc6fff1aa50c999c4e28cece888fbab86039

Observation 84a880fa-5ed3-415a-bb82-e681933076af · outbound

This paper cites Mvtn: Learning multi-view transforma- tions for 3d understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Mvtn: Learning multi-view transforma- tions for 3d understanding

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.331028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.445167Z digest=sha256:8e24483bba376275d1c4a500f3285094b865f84ae06bca5b3d89f3ce3b3bf7ba

Observation bc0da9e4-1860-4f29-9481-693403af7c7b · outbound

This paper cites Unsupervised keypoints from pretrained diffusion models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Unsupervised keypoints from pretrained diffusion models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.316667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.449699Z digest=sha256:727ab29dddaa2d1654f41f05b08b82b08a2765871e4fc747a1802734ff23c604

Observation 596c8d96-1212-4984-93d2-282a39e49667 · outbound

This paper cites 3d-llm: In- jecting the 3d world into large language models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d-llm: In- jecting the 3d world into large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.454016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.454016Z digest=sha256:f3a21a1b49b9f81e49c38dc9e3f64bd89c680c0dbf099e9792636da84efffd19

Observation bd089bf1-b03f-462c-8563-3d35bf8b5767 · outbound

This paper cites 3d-sis: 3d se- mantic instance segmentation of rgb-d scans.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d-sis: 3d se- mantic instance segmentation of rgb-d scans

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.292681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.458692Z digest=sha256:09b2591b4cd73e0c3e29abc4d4402f4b1014bbb55282396449c0f7350440d973

Observation 454c4372-5c82-4eb0-82ff-789fa2034217 · outbound

This paper cites Segment3d: Learning fine-grained class-agnostic 3d segmentation without manual labels.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Segment3d: Learning fine-grained class-agnostic 3d segmentation without manual labels

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.278453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.462899Z digest=sha256:b16b9711dd1c428dd1a5a2834c7eac7488807a9262d1fa75f0d67af89042e307

Observation 7d10388c-56b0-4ca7-9d4b-a6f7e0b2404c · outbound

This paper cites Keypointdeformer: 9 Unsupervised 3d keypoint discovery for shape control.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Keypointdeformer: 9 Unsupervised 3d keypoint discovery for shape control

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.260658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.467348Z digest=sha256:def14c59698d26956c69babe85ebed79e33cb05035cfa5bdfb6c3c7968475f1c

Observation 9e92acb9-cf9f-4a0e-b476-61628feddb11 · outbound

This paper cites Multi-view pointnet for 3d scene understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Multi-view pointnet for 3d scene understanding

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.245437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.472270Z digest=sha256:1e77380181487a783f8efd02b899ad3b9c4d4ec0628f8aa427bd17bd7b765621

Observation d2165763-6c21-4ef5-a648-823767707ba2 · outbound

This paper cites 3d shape segmentation with projective convolutional networks.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d shape segmentation with projective convolutional networks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.222504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.476533Z digest=sha256:5813b9e134f84cdeb66d51085095df2d3063f2d0b054e632475d9d636307b4c2

Observation b4bb8479-64b2-47de-849b-dd1fa8933abe · outbound

This paper cites Berg, Wan-Yen Lo, Piotr Doll ´ar, and Ross Girshick.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Berg, Wan-Yen Lo, Piotr Doll ´ar, and Ross Girshick

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.480826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.480826Z digest=sha256:e60e898d566e0f9602c0eb8a1fbfe149d76913c82559f1f7f6392f7f45127a57

Observation 0bdbfc53-5879-42ec-9077-6889205c388f · outbound

This paper cites Virtual multi-view fusion for 3d semantic segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Virtual multi-view fusion for 3d semantic segmentation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.188206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.485866Z digest=sha256:5580c9a0ab27995217bc0778367eb1ca39e8f57f0033a7d236f39804dcec8073

Observation e2c303d6-0c1c-48af-bc3b-6faf051f9de1 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.490024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.490024Z digest=sha256:6b6f32210ac8f8dd41430d524f79a1727238fe182b84a05ac8dd33e088253638

Observation eac79e8c-9230-404e-9faa-24db22f008a2 · outbound

This paper cites Visual instruction tuning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Visual instruction tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.494848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.494848Z digest=sha256:dcebfd42242d86d6ea8cec7f347cb27f80f7ea1b6bb827457a889f3fc67c5a65

Observation 72cfa958-02dc-4787-a8b9-329f22a9cae1 · outbound

This paper cites Improved baselines with visual instruction tuning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Improved baselines with visual instruction tuning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.146801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.499396Z digest=sha256:0086c32333af79739227c66c50508eb9c22d5b178e8a45c4ae86375c01bc7b91

Observation fcd044aa-136b-48e1-b8d8-b2c514ab1d65 · outbound

This paper cites 3d-to-2d distillation for indoor scene parsing.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d-to-2d distillation for indoor scene parsing

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.131318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.503575Z digest=sha256:da46a457a3ddf761f8318992d050500ddd8d5a034a878fefdf309216e8da37d8

Observation 71a2aa23-2dfa-40ad-a580-96362a4f41e8 · outbound

This paper cites Learning to segment 3d point clouds in 2d image space.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Learning to segment 3d point clouds in 2d image space

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.116382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.508556Z digest=sha256:94da5d048f588ab4f0fa881f848eafb1ef1277a59a50c21d1d7c7758c6cb88f6

Observation 378073f9-a7d3-461a-b7ef-d29ea10de51b · outbound

This paper cites Egoloc: Revisiting 3d object localiza- tion from egocentric videos with visual queries.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Egoloc: Revisiting 3d object localiza- tion from egocentric videos with visual queries

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.093838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.512221Z digest=sha256:de8da094bc5934a931c203dbb92c7284dd72747083e56502ec1ffff551433ee3

Observation cc9a7285-2371-41ab-9431-02124fdbfb9b · outbound

This paper cites Tracknerf: Bundle adjusting nerf from sparse and noisy views via feature tracks, 2024.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Tracknerf: Bundle adjusting nerf from sparse and noisy views via feature tracks, 2024

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.073993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.516273Z digest=sha256:78e2d67f503fee580e88f32cdec914acfdd12053d473e416546e093220fa397a

Observation afb41c8c-3f9f-4502-8ed9-0bd41c70655a · outbound

This paper cites Gpt-4 technical report, 2023.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Gpt-4 technical report, 2023

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.048343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.520516Z digest=sha256:080f4c3e6088da51a4ffdbb5e5dfb6ee83e63e52ef3dc3142cf256abeb5eb7f9

Observation 455dfd8b-5d4b-4918-ab37-1a55510ad9ae · outbound

This paper cites Synthesize diagnose and optimize: Towards fine- grained vision-language understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Synthesize diagnose and optimize: Towards fine- grained vision-language understanding

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.031427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.525276Z digest=sha256:84148013dee6b22cd10631649c62c8abed403e62dc2b39b8dd83fb2519ea251c

Observation 46b9cf11-d76a-4e8f-9c9a-e24ae4926aa2 · outbound

This paper cites Shapellm: Universal 3d object understanding for embodied interaction.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Shapellm: Universal 3d object understanding for embodied interaction

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.014512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.529670Z digest=sha256:004470dace8cd0bffb6e32aa9bc46f8a98e211838a0197a110fbf95bfc882e9c

Observation f1ac552a-836b-475c-b835-1e4d326928ce · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Learn- ing transferable visual models from natural language super- vision

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.998627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.534012Z digest=sha256:d844c0261a29836d52095fd276732e2b084dd847207dad2133308a6c7ac13b22

Observation 39761a1e-ab19-4615-912c-6fc91ee58bac · outbound

This paper cites Vision language models are blind: Failing to translate detailed visual features into words.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Vision language models are blind: Failing to translate detailed visual features into words

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.538799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.538799Z digest=sha256:f4f97aaf5249a80814b43d001b966a4fa9ecdcfb8d2910540b9cd87f9bdaa6c4

Observation 31eb0969-dfc1-488c-8d08-2a5b9fb900fe · outbound

This paper cites The Strategy of Conflict: with a new Preface by the Author.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models The Strategy of Conflict: with a new Preface by the Author

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.979666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.543453Z digest=sha256:78c5017f041bc16385a6601ab93760e580cf0a790f68a4d3baeb4ee01ee0cf1b

Observation 925db8d7-fa42-4053-8bc0-8ba92a4a68d8 · outbound

This paper cites Mask3D: Mask Trans- former for 3D Semantic Instance Segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Mask3D: Mask Trans- former for 3D Semantic Instance Segmentation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.548106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.548106Z digest=sha256:f2887b1a1388c9b975ba0c0db45faed98e4bad0c7769d2bd77bc7311e3088fea

Observation 5b34724c-af20-4f5f-824e-be6c83aea78c · outbound

This paper cites Skele- ton merger: an unsupervised aligned keypoint detector.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Skele- ton merger: an unsupervised aligned keypoint detector

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.949277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.553206Z digest=sha256:0eac1a53c660f70c5ec1fe3ab065503c8cd294c44f5dacd532ef0f73db03c83e

Observation ed0f51ef-4cab-4922-b4fa-84dc6f63c8d0 · outbound

This paper cites What does clip know about a red circle? vi- sual prompt engineering for vlms.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models What does clip know about a red circle? vi- sual prompt engineering for vlms

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.935249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.557535Z digest=sha256:291f4c6859bbf5a321d4852b6075a2fab3a5a341364d8cfa0799435b7f99cae5

Observation 294a82b1-c7bf-42fa-8116-ce496b9ff8a2 · outbound

This paper cites Discovery of latent 3d key- points via end-to-end geometric reasoning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Discovery of latent 3d key- points via end-to-end geometric reasoning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.920888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.561859Z digest=sha256:bec9ce4c50e58e7811b10fe5b24e6658352a41184387fa656c327867c4ff3391

Observation 353481c6-fc2a-4a3d-854d-4c931b2e2dec · outbound

This paper cites Ldls: 3- d object segmentation through label diffusion from 2-d im- ages.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ldls: 3- d object segmentation through label diffusion from 2-d im- ages

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.905419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.566259Z digest=sha256:e2c9d0758ef2665949db339ee09c9a0bd4063341c7e37cec47f5a6e900ea83f0

Observation 4dc85b4c-cbf5-42e4-bb6c-dfdffee05be7 · outbound

This paper cites Learning 3d keypoint descriptors for non-rigid shape matching.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Learning 3d keypoint descriptors for non-rigid shape matching

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.889417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.570113Z digest=sha256:fb74c607e1c2d559cd849f3d9ee9d1c7322d4a3c267ba2cf5d007fb5ab2aaa45

Observation 6b129e28-9bf1-4a00-ae2f-7e21ddee4aa1 · outbound

This paper cites Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.871745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.574293Z digest=sha256:2cd1ea7823c32e7d895f9a0b7f4b5c8001a2e8e3970aba6ffc5252aced47cc7c

Observation 88e02f07-2665-424c-a728-65d545fac1b2 · outbound

This paper cites Back to 3d: Few-shot 3d keypoint detection with back-projected 2d features.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Back to 3d: Few-shot 3d keypoint detection with back-projected 2d features

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.856075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.578146Z digest=sha256:31f7d413562932c09d8ffcf1c0fcf990562621bc95627cd4c3c877b3e2227bd1

Observation b8f9746f-4395-4082-abf0-1c4842dc3551 · outbound

This paper cites Clip-dinoiser: Teaching clip a few dino tricks for open- 10 vocabulary semantic segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Clip-dinoiser: Teaching clip a few dino tricks for open- 10 vocabulary semantic segmentation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.840684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.582597Z digest=sha256:c89d8bc965526f7153c67584be78b9acd82c6e4e79b30d04698b2f32fb3b4fa7

Observation d80eac84-56fd-48b9-9b57-fb748e5681a0 · outbound

This paper cites Pointllm: Empowering large language models to understand point clouds.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Pointllm: Empowering large language models to understand point clouds

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.824142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.586781Z digest=sha256:064dab17313126daad08b0cc2ab73983cd2e8e889f282d9ab7f88660c632c86d

Observation 203fd48a-fece-48b8-b6a6-746e9186fad3 · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.809708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.591199Z digest=sha256:e1d72db29e803ccef949afac83bbbdbcb6a4662d0517988791b8e94f2f4af4e1

Observation aef7633b-abfa-4629-9c13-1e13f93132ff · outbound

This paper cites 3dfeat-net: Weakly su- pervised local 3d features for point cloud registration.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3dfeat-net: Weakly su- pervised local 3d features for point cloud registration

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.795136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.595690Z digest=sha256:a760ba4ba70fe27870d271ab84507546e4574443e4f16aa0795850ac4b951d1b

Observation 8b26c0c1-3579-445e-97dd-21a712c5e4c7 · outbound

This paper cites KeypointNet: A Large-scale 3D Keypoint Dataset Aggregated from Numerous Human Annotations.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models KeypointNet: A Large-scale 3D Keypoint Dataset Aggregated from Numerous Human Annotations

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.600534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.600534Z digest=sha256:2e599d91f173f1ac0a34def294e2385d8ed0a2b4614f96e7986080d76edbd3cf

Observation 276695df-040c-467d-9e47-b5b8d2137c6d · outbound

This paper cites Ukpgan: A general self-supervised keypoint detector.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ukpgan: A general self-supervised keypoint detector

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.780211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.605173Z digest=sha256:d7c9d9b4f66216e62806e6d9213d9a8c45a04d5eee50680ae0140a2ed3ab0ebf

Observation 6f564cf9-386f-447d-8be6-636e7c1d5f14 · outbound

This paper cites Good at captioning bad at counting: Benchmarking gpt-4v on earth observation data.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Good at captioning bad at counting: Benchmarking gpt-4v on earth observation data

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.765899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T19:53:22.610243Z digest=sha256:f775f7bb41b04ce1ba75cf13e07355b80ae3398b66c2df07c1f8bb2c9e9259fb

Observation 902ee351-1ecc-4d3f-bc1e-01008d542817 · outbound

This paper cites Uni3D: Exploring Unified 3D Representation at Scale.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Uni3D: Exploring Unified 3D Representation at Scale

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.614137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.614137Z digest=sha256:bb1bcd9ff0e6432a2999c107a1cdb489ca51922aeb9f50aca37d4ce07da6748b

Observation 9b53a1ec-0163-47b0-b616-b3d452b8b755 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 54

Resolution
malformed identifier
no resolver link, observed 2026-08-11T19:53:22.619176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.619176Z digest=sha256:e3f19106d390a78c1ce5cc099f4444fb7cdd423ab2967d229aef7f6def09ff32

Pith citing papers

Observation 7f9daa15-1422-4e9f-ba24-35f50659d139 · inbound

Estimating 2D Keypoints of Surgical Tools Using Vision-Language Models with Low-Rank Adaptation cites this paper.

Estimating 2D Keypoints of Surgical Tools Using Vision-Language Models with Low-Rank Adaptation ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T14:49:39.762699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T14:49:39.643676Z digest=sha256:b74eb4bf17fb17b0a14cd10f28bf9d5003929afcf5ff14ff6e66f0f53d94383d