Pith. sign in

Paper Citation Record · LEDGER

PointLLM: Empowering Large Language Models to Understand Point Clouds

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2308.16911.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.16911 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:13:33.303592Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 58b9a34b-0eaf-450d-acd2-6515ad04d066 · inbound

A Survey on Multimodal Large Language Models cites this paper.

A Survey on Multimodal Large Language Models PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 145

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T02:56:41.954357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T02:56:41.658658Z digest=sha256:2b316114771d7ba03ef013341a5b8cbef705c3635c923e5f1ff8220f696f3275

Observation b303b711-1fca-418d-8524-c287e939785d · inbound

SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models cites this paper.

SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T03:03:26.779545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-17T03:03:26.723464Z digest=sha256:499a2ca34549749e0a89df271bb36f7a68a899886ea8278196acc63742351651

Observation 19f05398-b6b3-49d5-b118-89ea80e7978f · inbound

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? cites this paper.

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:29:30.146288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-17T01:29:30.032408Z digest=sha256:e7082336aac78725af8994ce7ecc135b77489a6b2c14973d7599da1217a56a64

Observation 84121b69-b692-4d17-b069-7a39544b3a85 · inbound

LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models cites this paper.

LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:01:54.002228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-11T06:01:53.730356Z digest=sha256:c45cbc8d5494dfe4617d6ffeb4ec8e62ae35f1d4ae329a0575e45936fe0865fd

Observation 079f1ba2-31be-41fb-8c42-89323504386d · inbound

sMoRe: Enhancing Object Manipulation and Organization in Mixed Reality Spaces with LLMs and Generative AI cites this paper.

sMoRe: Enhancing Object Manipulation and Organization in Mixed Reality Spaces with LLMs and Generative AI PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T18:13:33.303592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:13:33.303592Z digest=sha256:e360d3f209cfa52ea2bcec1ed1f2ecf4b53170c567f9317b1d48143ab7133242

Observation e9cb2987-b74a-4121-a0b9-134b25bc07f9 · inbound

PerLA: Perceptive 3D Language Assistant cites this paper.

PerLA: Perceptive 3D Language Assistant PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T05:52:39.158355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:52:39.158355Z digest=sha256:61e4da071926d93b9029af5e7e8933b7a0f1864b222bfba63114ce3d3f6cbb56

Observation 6e6ffa06-16de-4985-b0fd-5779beaf94e6 · inbound

LSceneLLM: Enhancing Large 3D Scene Understanding Using Adaptive Visual Preferences cites this paper.

LSceneLLM: Enhancing Large 3D Scene Understanding Using Adaptive Visual Preferences PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T04:35:37.009769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:35:37.009769Z digest=sha256:e852841226a828ce8b31e334c290d0c779b8deeab72f4fe2a12d602545fcc1f1

Observation 41a64d97-2b2f-4ee0-a5ce-84dfd9281f06 · inbound

SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model cites this paper.

SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T04:22:04.042646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:22:04.042646Z digest=sha256:405b4bd2fc8f5eec99955e698a388f3225bb0a01882212619d93f839e5a77949

Observation 456deeb1-e6b4-43a2-a432-afa5f61dfdc2 · inbound

AIpparel: A Multimodal Foundation Model for Digital Garments cites this paper.

AIpparel: A Multimodal Foundation Model for Digital Garments PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T22:02:03.998697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:02:03.998697Z digest=sha256:9ae8d7ac4742346885f40ee2046cbc687f0ec68dc89d053dd85dc805dd823055

Observation 87ca801d-b520-4f0f-88bc-98fafc8c6841 · inbound

MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models cites this paper.

MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T19:31:43.992134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T19:31:43.992134Z digest=sha256:d206cbeb1614c69dad7b17c31a4ddb779e03aa5feb326e47083e3dcb6cfb3790

Observation c56ffbaf-ca7a-48d0-99c4-ddd88aa970cc · inbound

ObjVariantEnsemble: Advancing Point Cloud LLM Evaluation in Challenging Scenes with Subtly Distinguished Objects cites this paper.

ObjVariantEnsemble: Advancing Point Cloud LLM Evaluation in Challenging Scenes with Subtly Distinguished Objects PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T11:54:12.249927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:54:12.249927Z digest=sha256:5c82e858d33b7598df0a4e4fd10f656bf2b8d3ba730d14883861178173fdc985

Observation e00940a1-fa89-420c-abb2-07f3d0ab5619 · inbound

3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer cites this paper.

3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T22:38:21.594099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:38:21.594099Z digest=sha256:82a1e41a6dc01bba49e0bfb4a34203ca755b2dfd2237f5a54e9012dc9e7e7d6d

Observation 0e08e82a-1f50-4ed4-9adc-c01f18a53f3b · inbound

CL3DOR: Contrastive Learning for 3D Large Multimodal Models via Odds Ratio on High-Resolution Point Clouds cites this paper.

CL3DOR: Contrastive Learning for 3D Large Multimodal Models via Odds Ratio on High-Resolution Point Clouds PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T21:48:27.526667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:48:27.526667Z digest=sha256:c2316125a10ac8ee7e67c0c92a3a6f0d76782abc540fd5e2e254182c5a075873

Observation eb0cf1c9-3fa3-4b9b-9c30-52b079489d09 · inbound

3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding cites this paper.

3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T20:39:12.665081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:39:12.665081Z digest=sha256:12c6e037241e309c1cfea44018cebce016af358ab77a96df358982db5303c8a9

Observation a1f4ca3b-ebdc-441c-a061-efbf5d027bd3 · inbound

Revisiting 3D LLM Benchmarks: Are We Really Testing 3D Capabilities? cites this paper.

Revisiting 3D LLM Benchmarks: Are We Really Testing 3D Capabilities? PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:02.124890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T04:54:02.124890Z digest=sha256:0a846c4aa3ba5a7b7921353d4e41c846176de12e0b91b3a590384582875f6673

Observation 31f97f64-9b56-4d71-aa35-d99cc0f5a528 · inbound

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision cites this paper.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.866129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.866129Z digest=sha256:ecc4c9a2b2ee521b20d41d4db6da451d71194baf2bebb0902029e9f8414556ce

Observation 344c3894-eb95-42dc-9bf5-a910bba323e6 · inbound

MINT-CoT: Enabling Interleaved Visual Tokens in Mathematical Chain-of-Thought Reasoning cites this paper.

MINT-CoT: Enabling Interleaved Visual Tokens in Mathematical Chain-of-Thought Reasoning PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:53.592815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:53.592815Z digest=sha256:69f4c95dec0add0c216e5ad4fa58e7f964d85e1793fe4e7c3b9b5176e52459ad

Observation 8e24a87e-fa6e-47ec-80aa-350393141259 · inbound

Enhancing Spatial Reasoning in Multimodal Large Language Models through Reasoning-based Segmentation cites this paper.

Enhancing Spatial Reasoning in Multimodal Large Language Models through Reasoning-based Segmentation PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:53:02.513697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:53:02.513697Z digest=sha256:f9d829ef8c6d9b5f85687017ec742c7dd81359c7d8c2d9b58ab40b3f9bae6823

Observation 6a496748-550e-4b94-9a5a-8e115ba8c9a7 · inbound

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation cites this paper.

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:17.158713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:17.158713Z digest=sha256:18fc0dd60eea9e379cdf9d042ad15c172e08445a23d12b2f08d4ed43ab8d5ab2

Observation 44692333-e098-417a-a654-422bfac6020c · inbound

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models cites this paper.

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:50:14.728483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-21T14:48:21.787919Z digest=sha256:cb666245d648836dc72d91274318fe73cafe478712fc8ae64cf821c7040eb51a

Observation fb7f232f-1d86-428d-a350-aee2585f70f7 · inbound

Chat-Scene++: Exploiting Context-Rich Object Identification for 3D LLM cites this paper.

Chat-Scene++: Exploiting Context-Rich Object Identification for 3D LLM PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:32:59.393660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-14T21:31:00.764078Z digest=sha256:e9bb5f9f72243db1a7f6c5392666d933c56bd3232ec1cd9b488c5454286cb3b1

Observation 84ec85aa-a649-40e1-9ef7-8dde3a05a5b0 · inbound

Affordance Agent Harness: Verification-Gated Skill Orchestration cites this paper.

Affordance Agent Harness: Verification-Gated Skill Orchestration PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:06:34.630394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-09T18:40:53.380512Z digest=sha256:68afb75473f0c28bfe0057ebcde03095ccc95fe7b7de4a2f854374777ef37bdf

Observation 7f8de01e-6136-4854-b565-d4b5e9b64ae4 · inbound

Affordance Agent Harness: Verification-Gated Skill Orchestration cites this paper.

Affordance Agent Harness: Verification-Gated Skill Orchestration PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:10:55.785639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-11T01:55:07.248106Z digest=sha256:1773633bb72dc97717fb7d76da4686493632492cbadab473aff36b1c240cbca3

Observation 468c9436-4a99-4780-a720-9b29f161dc2b · inbound

From 3D Perception to Safety Reasoning: A Graph-Based Framework for Real-Time Underground Mine Monitoring cites this paper.

From 3D Perception to Safety Reasoning: A Graph-Based Framework for Real-Time Underground Mine Monitoring PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:46:27.779918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T10:44:00.642366Z digest=sha256:aca23fae9ecbcf6341ace8f9bde4d9e0460bc6b8853433f06aa9205a55c49a41

Observation c5a4d80e-b56e-4371-8523-ef12c672fff3 · inbound

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models cites this paper.

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:26:45.808225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T06:58:39.117228Z digest=sha256:f015e3e812dda0f008d1ed6f6d5138bddebe1a4c9aad71bbfc5b67d3ff36ed50

Observation c563d40b-a124-4e43-a2e1-12f1e116d70e · inbound

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning cites this paper.

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:18:57.806629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-27T01:23:40.564561Z digest=sha256:9439f63aece0f813bd5bf91776916f1955f397c1bb299da663d023891f6b8fdf

Observation ef0f21d6-cfd6-4951-8373-88c9de408106 · inbound

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation cites this paper.

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 227

Resolution
unresolved
no resolver link, observed 2026-08-02T06:23:49.818915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:23:49.818915Z digest=sha256:b8f1350e4710e1f1b6aeb68ba0ea04cd395e65e776055aac6edbd8b733f03ddf

Observation 8d5cbfb9-d64d-4d45-8f92-dad97c174035 · inbound

CR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene Retrieval cites this paper.

CR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene Retrieval PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T13:28:41.335244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:28:41.335244Z digest=sha256:c3ebc29de1e458b9213f11d5f536eefcaaa9439c96dcba1f7cf2da428f00eff6