Pith. sign in

Paper Citation Record · LEDGER

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval

As of 8 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 0 inbound Pith citation observations for arXiv:2507.21489.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21489 v1

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:47:08.442715Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

73 of 73 outbound references displayed

  • verified exact0
  • verified fuzzy46
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e4aac53a-8eba-4245-841e-e1b4050517d1 · outbound

This paper cites GPT-4 Technical Report.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:02.365253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:02.365253Z digest=sha256:546446f4bf8a023bb0f8d1ef320ab59652d2c7bd29db12ddd9987e897c67b941

Observation b8ea0d4c-9b69-4cf2-862a-5c3cb8f73ddb · outbound

This paper cites Learning representations and generative models for 3d point clouds.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Learning representations and generative models for 3d point clouds

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.859434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:02.433911Z digest=sha256:e84389a3a90448eb78c9dd8b93108705f4bfd7e145e0a37560097ee8f3c8f755

Observation 999751b7-9af4-4ca8-a8cc-b92716dae7db · outbound

This paper cites Flamingo: a visual language model for few-shot learning.NeurIPS, 35: 23716–23736, 2022.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Flamingo: a visual language model for few-shot learning.NeurIPS, 35: 23716–23736, 2022

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:02.604846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:02.604846Z digest=sha256:b2d44c5b28af266211a30b0b82c479d8d601116a2f7e3731469e6a3fcad3c137

Observation a43544f7-af1f-4cbc-8234-efa0de1b77c9 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:02.780795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:02.780795Z digest=sha256:4ac775fdf7c06584968b6f236eb8dae8e08e1df5075aa24eea2be4d3fbc8679a

Observation 4fcf1764-3975-4614-a34d-31129cb75b39 · outbound

This paper cites Qwen2.5-VL Technical Report.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Qwen2.5-VL Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:02.884100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:02.884100Z digest=sha256:b6cef42e50927316508ea6481060b18e38e2a7c17fbba7582e611db7163d2fda

Observation ee5ee636-8bd2-4873-aacc-a58928464e15 · outbound

This paper cites Shape matching and object recognition using shape contexts.IEEE TPAMI, 24(4):509–522, 2002.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Shape matching and object recognition using shape contexts.IEEE TPAMI, 24(4):509–522, 2002

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.846394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.016796Z digest=sha256:48e0e8d7a23164a6ed8f70ec746b9be26fbe0deba643e552ff6e29f483b3bb1a

Observation 58286fd9-015a-4f4e-984a-91c3fb966a67 · outbound

This paper cites On visual similarity based 3d model retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval On visual similarity based 3d model retrieval

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.838131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.094828Z digest=sha256:12de57cc9ab865ec6f85645fae2aab72982dfc16ed9083aa2816ad6d13226118

Observation 5c7843a9-6416-421f-b833-a417b1c5031d · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:03.150233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:03.150233Z digest=sha256:6416b49e2ec762634bc184f27770232e05ada0ff77d16aae3c5da43859aea664

Observation 5eb4c32d-f262-4321-8e27-7fbdae07ad64 · outbound

This paper cites Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.830563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.233035Z digest=sha256:bdc6bd395f015a9b63816e63de6f52adbf952e4454a8946a8810a4b15b9d169f

Observation fe416e7f-c5ba-4648-9501-e4e7818801cb · outbound

This paper cites Pra-net: Point relation-aware network for 3d point cloud analysis.IEEE TIP, 30:4436–4448, 2021.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Pra-net: Point relation-aware network for 3d point cloud analysis.IEEE TIP, 30:4436–4448, 2021

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.822525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.286593Z digest=sha256:ef9dd8aa83f31b99464d41077f05d2a1dabbd04a2c061bbddb1d14fbaf7441de

Observation cae22099-756b-4a79-ae48-34d4b763b966 · outbound

This paper cites Abo: Dataset and benchmarks for real-world 3d object un- derstanding.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Abo: Dataset and benchmarks for real-world 3d object un- derstanding

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.814815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.332545Z digest=sha256:4e9199b36323c4ac9765a7aca992ca1bbfaf0e8ab672a387a5af29b7a80675fc

Observation 8c13730c-450c-4356-a4b1-c9c1ad63dfa5 · outbound

This paper cites Siamese cnn-bilstm architecture for 3d shape representation learning.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Siamese cnn-bilstm architecture for 3d shape representation learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:03.405200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:03.405200Z digest=sha256:1059434004ab3d4e5103bad83a8a2a9c3f84d632767a1c54be7fcd517db46747

Observation f48d7503-23da-495c-9c5c-1ca5fc3aaaf0 · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Objaverse: A universe of annotated 3d objects

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.802001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.490927Z digest=sha256:df8b87d94895d7ab5dfcbcdb278595ba6c42ca7bc6b25e188fcf3829dd67f2e7

Observation fc43a5cb-36cb-4c0f-adff-13eabd631bfa · outbound

This paper cites Equivariant multi-view networks.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Equivariant multi-view networks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:03.553089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:03.553089Z digest=sha256:c00049c109edbcc52aa38f5d57a293d3282712a3cc3a9621a0f12dad3d517adb

Observation 2495eeda-fb6a-4edf-9f30-bbecd515182c · outbound

This paper cites Gvcnn: Group-view convolutional neural networks for 3d shape recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Gvcnn: Group-view convolutional neural networks for 3d shape recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:03.614094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:03.614094Z digest=sha256:61a4d0f01b92557a58f2fdb081ff983ba9fc08330bef110de6e8eae3fef9ea77

Observation 1a816f62-9280-4509-8bfc-ac887bb3fe15 · outbound

This paper cites Meshnet: Mesh neural network for 3d shape rep- resentation.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Meshnet: Mesh neural network for 3d shape rep- resentation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.783681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.664107Z digest=sha256:886740d610983dcd867055f65ff7791cf7cf14a7af1d01b4ba368136e5b2fcee

Observation f4b2d03c-7c35-4a48-9b51-6edd2ecf2adc · outbound

This paper cites Shrec’22 track: Open-set 3d object retrieval.Computers & Graphics, 107:231–240, 2022.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Shrec’22 track: Open-set 3d object retrieval.Computers & Graphics, 107:231–240, 2022

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.775924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.721045Z digest=sha256:df6b5144c9e2007f568dba138649023d32d4089cdb95d372166f2795581d5aba

Observation 4d4f070d-1df2-4fa2-867c-d24670e22292 · outbound

This paper cites Hypergraph-based multi-modal represen- tation for open-set 3d object retrieval.IEEE TPAMI, 2023.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Hypergraph-based multi-modal represen- tation for open-set 3d object retrieval.IEEE TPAMI, 2023

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.768338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.779910Z digest=sha256:2869e5b10e90e31002c5e966c29fb8d97a3947a0252b28a45c3d72dc41657773

Observation bcff649d-f534-48c4-a44d-4bbf168d559c · outbound

This paper cites 3d object retrieval based on similarity calculation in 3d computer aided design systems.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval 3d object retrieval based on similarity calculation in 3d computer aided design systems

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.760873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.851713Z digest=sha256:dfd14303d4b9dc5d999bef7db4dd150bc9b1552df7d67c546c05c6c2c3b84be2

Observation 10d8b2a7-003e-42e0-a19a-46879c8ada4a · outbound

This paper cites Clip-adapter: Better vision-language models with feature adapters.IJCV, 132(2):581–595, 2024.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Clip-adapter: Better vision-language models with feature adapters.IJCV, 132(2):581–595, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.752571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.918093Z digest=sha256:8e55dfbe060e990fa3f51ef6abd98bd0757c497e023d054783ce3b457d1c0bb1

Observation 76847e5b-f7f0-4c99-b861-0e6491b9ae51 · outbound

This paper cites Deep learning for 3d point clouds: A survey.IEEE TPAMI, 43(12):4338–4364, 2020.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Deep learning for 3d point clouds: A survey.IEEE TPAMI, 43(12):4338–4364, 2020

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.744371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:03.981913Z digest=sha256:55e93f0cd8d657ebb7bd750bc5aeafc6bdff7b3584cb5a8cb5554e0bdba87a22

Observation a962f3c3-cf0f-4410-8616-8d0bc53dfd3d · outbound

This paper cites 3d2seqviews: Aggregating sequential views for 3d global feature learning by cnn with hierarchical attention ag- gregation.IEEE TIP, 28(8):3986–3999, 2019.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval 3d2seqviews: Aggregating sequential views for 3d global feature learning by cnn with hierarchical attention ag- gregation.IEEE TIP, 28(8):3986–3999, 2019

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:04.041414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:04.041414Z digest=sha256:080568f67fa9f5a534d5878bb606b396e8776cbff6e5665b471a10cb102a6668

Observation 88effd1b-bd57-4880-80f8-ad08d88e8368 · outbound

This paper cites Triplet-center loss for multi-view 3d object retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Triplet-center loss for multi-view 3d object retrieval

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.731351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:04.169624Z digest=sha256:992a36272e0c94e84813e7f43be61cb1e844af321d58e0d964848b61a67c231c

Observation 3dfcf5d2-0e93-49e3-bbaa-ab326a267836 · outbound

This paper cites View n-gram network for 3d object retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval View n-gram network for 3d object retrieval

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:04.239337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:04.239337Z digest=sha256:53787629d3642b25de459487a4a08a810c5249d342221be691b14f3139566725

Observation 815893c3-e0a6-4b80-abef-14fc8b3f754c · outbound

This paper cites Latformer: Locality-aware point- view fusion transformer for 3d shape recognition.PR, 151: 110413, 2024.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Latformer: Locality-aware point- view fusion transformer for 3d shape recognition.PR, 151: 110413, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.717957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:04.300096Z digest=sha256:535d297fbd01b6da13bbbafb66bfee75d3868aa66d7c2e961f1a08a10e41d5e5

Observation 37b4c765-2053-4fb8-807c-91540a2c986a · outbound

This paper cites Clip goes 3d: Leveraging prompt tuning for language grounded 3d recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Clip goes 3d: Leveraging prompt tuning for language grounded 3d recognition

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.709871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:04.368902Z digest=sha256:350d3b868b6feff75d02e3ecec2ccc972f2ac9ee85888e3a10f4aa2ec6d876ef

Observation 7386090c-3e0a-4375-9fd0-c1e07f7a3495 · outbound

This paper cites LoRA: Low-rank adaptation of large language models.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval LoRA: Low-rank adaptation of large language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.701250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:04.423232Z digest=sha256:aeaca5d031fb1d232eb7e4f2a395c22a623906157442f991dcd4a8db4ce0e7a0

Observation 117d2cdc-be8b-4cf3-a377-2cbd543781f4 · outbound

This paper cites Scal- able deep multimodal learning for cross-modal retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Scal- able deep multimodal learning for cross-modal retrieval

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.693541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:04.470427Z digest=sha256:eff242def2617d17cda2159a37c3409f24b8896294d21f53defef4184e9d8bf0

Observation 4daccc3d-7e4b-4ebc-9ace-3d51d8f055bd · outbound

This paper cites Clip2point: Transfer clip to point cloud classifica- tion with image-depth pre-training.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Clip2point: Transfer clip to point cloud classifica- tion with image-depth pre-training

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.685113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:04.527676Z digest=sha256:b4cf8432b325c3fe0b97fcd03b0319a519e9aaca61fdc5588a3671f955515bf6

Observation 01dbc503-4c9e-4fd6-b786-753e0d282054 · outbound

This paper cites Developing an engineering shape benchmark for cad models.Computer-Aided Design, 38(9):939–953, 2006.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Developing an engineering shape benchmark for cad models.Computer-Aided Design, 38(9):939–953, 2006

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.676422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:04.572553Z digest=sha256:8178b9e21c62424718322dde2d964cd9235bba98a6e4804d5850c4f4b406c504

Observation 3cb8bc34-6b35-4991-9fec-1da2e7051ed5 · outbound

This paper cites Cross-modal center loss for 3d cross-modal retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Cross-modal center loss for 3d cross-modal retrieval

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.529357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:04.655578Z digest=sha256:fdece3eacbd584c28ea111e47c60f71a377870ef59d06a60fdf4b49a1fd8a57e

Observation 1a92f109-709f-4ad2-a78c-e501fd6c3e7b · outbound

This paper cites Rotationnet: Joint object categorization and pose estimation using multiviews from unsupervised viewpoints.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Rotationnet: Joint object categorization and pose estimation using multiviews from unsupervised viewpoints

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:04.728566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:04.728566Z digest=sha256:8cdf4a9a7a628707644a87b5b99493c19ef7f9353aad000210d7dd32537af2c8

Observation 44a49962-a314-47a7-b401-c029592c416c · outbound

This paper cites Rotation invariant spherical harmonic repre- sentation of 3 d shape descriptors.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Rotation invariant spherical harmonic repre- sentation of 3 d shape descriptors

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.259037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:04.817478Z digest=sha256:daa46c3bec7dd597206293c05674bcce5da2add3289afcc85a7ea4589cf8ccd2

Observation 9243c747-567d-40b2-b973-5ffb1feeeb1c · outbound

This paper cites Re- search challenges for digital archives of 3d cultural her- itage models.Journal on Computing and Cultural Heritage (JOCCH), 2(3):1–17, 2010.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Re- search challenges for digital archives of 3d cultural her- itage models.Journal on Computing and Cultural Heritage (JOCCH), 2(3):1–17, 2010

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:13.992856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:04.879461Z digest=sha256:55cf40be51f4d2e0127e42e494c1d82312996fc8bc826ec99716f9bc4dd6a287

Observation 4ab4e967-f98d-4b2d-a1d5-b3c9b8692492 · outbound

This paper cites Angular triplet- center loss for multi-view 3d shape retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Angular triplet- center loss for multi-view 3d shape retrieval

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:04.974880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:04.974880Z digest=sha256:d2be787d092fad6a07c219d9103db500ed59cf81dfacb09064b6f0b56d456be4

Observation bc8403e9-a325-4f6f-8ae6-c364dcd4624e · outbound

This paper cites Meshmae: Masked autoencoders for 3d mesh data analysis.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Meshmae: Masked autoencoders for 3d mesh data analysis

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:13.736307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:05.049444Z digest=sha256:7ede8033e0282ff40fcd709866306d0ead1898706e0d02fa5933bea07c3c5b02

Observation 5e48812c-ecfb-49ac-b19d-496b2f16b128 · outbound

This paper cites Improved baselines with visual instruction tuning.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Improved baselines with visual instruction tuning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:13.482566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:05.138939Z digest=sha256:f3e573054a4687c422203095bd7def5e91566932c9f16a29c14cc295d72d141e

Observation 264362b9-fa81-4531-b6c5-7806627be3d0 · outbound

This paper cites Visual instruction tuning.NeurIPS, 36, 2024.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Visual instruction tuning.NeurIPS, 36, 2024

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:13.133129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:05.175939Z digest=sha256:e6ea1f1fb24b14f7f638a7c3ed7cc1c90928cb8afc4c2ff3ce7ef37995834329

Observation 8a090f86-d450-4e86-b979-d7a3d59422da · outbound

This paper cites Openshape: Scaling up 3d shape representation towards open-world understanding.NeurIPS, 36, 2023.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Openshape: Scaling up 3d shape representation towards open-world understanding.NeurIPS, 36, 2023

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:12.863259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:05.277656Z digest=sha256:0091fcaa5dc77efe0495bb02179420bbb1edeeb1347d1bb313cbafc3dad2b9e6

Observation c71387ad-3397-4a1b-96cc-a86cebe9f79f · outbound

This paper cites Point2sequence: Learning the shape representa- tion of 3d point clouds with an attention-based sequence to sequence network.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Point2sequence: Learning the shape representa- tion of 3d point clouds with an attention-based sequence to sequence network

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:05.339949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:05.339949Z digest=sha256:09aa7d7264a9177362c603847a9d3388c6818f8cbeed18b3dc74c72b4fc032c0

Observation 387e2f23-ae6f-4852-91fb-4a6812ffd73a · outbound

This paper cites Relation-shape convolutional neural network for point cloud analysis.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Relation-shape convolutional neural network for point cloud analysis

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:05.426858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:05.426858Z digest=sha256:a47a1861ea2544db3745f78aa276e4de984cd25ce0bb22366de54e1770a8b65f

Observation 4e2d60e5-80c3-4e40-92a4-26d27bc026b4 · outbound

This paper cites An Empirical Study of Scaling Instruct-Tuned Large Multimodal Models.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval An Empirical Study of Scaling Instruct-Tuned Large Multimodal Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:05.523805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:05.523805Z digest=sha256:e9ddfa7345489076671b703c191b93fdcdb6449914d59b78e6c777c879727411

Observation 0315681a-189f-43f7-a31b-854b87843fc2 · outbound

This paper cites 3d learning objects for augmented/virtual reality educational ecosystems.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval 3d learning objects for augmented/virtual reality educational ecosystems

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:12.619734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:05.621772Z digest=sha256:06349719fa5cff553887c9b878cdc76cf3d1d2c88933b6cccfb512010487f91a

Observation bb366d37-50fe-4ac3-8559-cc8366f9941c · outbound

This paper cites V oxnet: A 3d con- volutional neural network for real-time object recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval V oxnet: A 3d con- volutional neural network for real-time object recognition

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:12.433140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:05.685811Z digest=sha256:7bd2d177286e3bfaf400bf1344096dac820a5eb2347e548dc428a0b6d3c2f424

Observation 16e478c5-a186-41a6-9a3c-e69da20444fb · outbound

This paper cites Mmjn: Multi-modal joint networks for 3d shape recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Mmjn: Multi-modal joint networks for 3d shape recognition

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:12.273646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:05.754575Z digest=sha256:60b608b31f6e270b98890161ba7b9e67b0d0164dd4294a0c62b4129920e0caea

Observation 30ad683e-6621-4b01-bd3f-c78dfdf80f9f · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Representation Learning with Contrastive Predictive Coding

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:05.854039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:05.854039Z digest=sha256:8654a35a4e19c3c051c9aff359aa2d45c5455b77398d26898babd84a1e45056f

Observation 4fcbc6c2-0bbd-44ee-bec3-5f48dbc128ec · outbound

This paper cites Kosmos-2: Grounding Multimodal Large Language Models to the World.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Kosmos-2: Grounding Multimodal Large Language Models to the World

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:05.946843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:05.946843Z digest=sha256:958da0b55e10e3bd2fbbda2e5bd557e4f62fc1d29d49d27b47241428d19e6787

Observation c7c831b5-0ae5-464f-9eaa-86b921d388a5 · outbound

This paper cites Pointnet: Deep learning on point sets for 3d classification and segmentation.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Pointnet: Deep learning on point sets for 3d classification and segmentation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.016588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.016588Z digest=sha256:0b01428cb3faa2977a3392473e550aa6cd7deffdd8fd2d6f14250fff241e7627

Observation 8029e1ed-a18e-4386-ac80-e70f2aabe21a · outbound

This paper cites Pointnet++: Deep hierarchical feature learning on point sets in a metric space.NeurIPS, 30, 2017.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Pointnet++: Deep hierarchical feature learning on point sets in a metric space.NeurIPS, 30, 2017

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.111380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.111380Z digest=sha256:f60ab5e2727395d7d505713fe97ba0a5fb5140c65ac860dbaf936c3db5557f86

Observation d8739b54-3665-4a47-b467-36176d8b2325 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Learn- ing transferable visual models from natural language super- vision

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:12.064361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:06.175609Z digest=sha256:cac966a5e6f9b9c9e26e03f9cf4526dc5910f2ebad3eb2e9c343f104bfd8ecbc

Observation dc8afc3f-b0ee-46a0-8976-33c0606fed60 · outbound

This paper cites Clip for all things zero-shot sketch-based image retrieval, fine- grained or not.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Clip for all things zero-shot sketch-based image retrieval, fine- grained or not

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.906256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:06.305599Z digest=sha256:daf03253ce8411ca2261c3c097c6f90e0c984adf66192ae9d6ec2e906c23fdd5

Observation bda63105-a27b-41c7-a280-a5f7059b94f3 · outbound

This paper cites Deep- voxels: Learning persistent 3d feature embeddings.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Deep- voxels: Learning persistent 3d feature embeddings

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.406724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.406724Z digest=sha256:5e6888d128c237dbcaf814d187a665302e623e6aafb84a69f313957b6dce4904

Observation 093cc193-e569-423b-bba7-a872a33dc293 · outbound

This paper cites MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.495224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.495224Z digest=sha256:0097188122f9342bb08ec17ddb08bf8e38d43e3f7965221055070f4c5e902d09

Observation 09ecf3ad-d5ec-4928-bba1-82ec6115d2f7 · outbound

This paper cites Multi-view convolutional neural networks for 3d shape recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Multi-view convolutional neural networks for 3d shape recognition

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.726653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:06.598342Z digest=sha256:e8699e48c5377e29372cf1fa414f6dcac9d8fd618dcb7926bc45f2d5f5356e23

Observation e7c38b4e-a674-400a-9ecb-4b353a3fbb1a · outbound

This paper cites A survey of content based 3d shape retrieval methods.Proceedings Shape Modeling Applications, 2004., pages 145–156, 2004.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval A survey of content based 3d shape retrieval methods.Proceedings Shape Modeling Applications, 2004., pages 145–156, 2004

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.555829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:06.714822Z digest=sha256:0297e5fed5252ce149551d6698140e981275bb92de71ff73be4571f96a7d239c

Observation ff79b94f-e89e-434f-93ce-072747635032 · outbound

This paper cites O-cnn: Octree-based convolutional neural networks for 3d shape analysis.ACM TOG, 36(4):1–11,.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval O-cnn: Octree-based convolutional neural networks for 3d shape analysis.ACM TOG, 36(4):1–11,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.823291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.823291Z digest=sha256:fed3199ff612917254a23d92c7e099ab24c744c57dc3801e273d1f8931d4e2ae

Observation a403d6d1-6d31-47a6-a48e-6014de054a3b · outbound

This paper cites The All-Seeing Project: Towards Panoptic Visual Recognition and Understanding of the Open World.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval The All-Seeing Project: Towards Panoptic Visual Recognition and Understanding of the Open World

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.899224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.899224Z digest=sha256:574941b14a7459d3e0ad6b8230c71f408292101199a2f4c605ade835d0fbdbe0

Observation e314d4b6-7a1d-4cc5-a012-1192015977bc · outbound

This paper cites Visionllm: Large language model is also an open- ended decoder for vision-centric tasks.NeurIPS, 36, 2024.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Visionllm: Large language model is also an open- ended decoder for vision-centric tasks.NeurIPS, 36, 2024

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.375033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:06.982283Z digest=sha256:9c8f9b9cb040806c05e788e48633504a5a1caf04ea46cb9b1e1b5b468b67d0b6

Observation 1380f257-583f-429f-9251-464e7b45e0d7 · outbound

This paper cites Dynamic graph cnn for learning on point clouds.ACM TOG, 38(5): 1–12, 2019.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Dynamic graph cnn for learning on point clouds.ACM TOG, 38(5): 1–12, 2019

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:07.080674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:07.080674Z digest=sha256:68de3bf96bbd87a22fc56e7b4445fde0793dac2abc0f4fb959abc67896296a3d

Observation c94a750d-7e34-4468-93b5-fbe661deca4c · outbound

This paper cites Teda: Boosting vision-lanuage models for zero-shot 3d object retrieval via testing-time distribution alignment.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Teda: Boosting vision-lanuage models for zero-shot 3d object retrieval via testing-time distribution alignment

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.173636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:07.146620Z digest=sha256:f6d05fff40811e5b6cbcd1c5cfa8cb6ec657e4638fac966c360d92fd097c56d7

Observation d7f9ed69-099a-468d-876a-b5a4a4f2abc6 · outbound

This paper cites View-gcn: View-based graph convolutional network for 3d shape analysis.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval View-gcn: View-based graph convolutional network for 3d shape analysis

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:07.221969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:07.221969Z digest=sha256:7fd4d90780cfa26f8566306cfcaa97f57d160f445f47d21aad64e2876b2d13b8

Observation 85db8e0f-ccce-4c71-b8e5-487058e71b2e · outbound

This paper cites Multi- modal semantic autoencoder for cross-modal retrieval.Neu- rocomputing, 331:165–175, 2019.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Multi- modal semantic autoencoder for cross-modal retrieval.Neu- rocomputing, 331:165–175, 2019

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.027826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:07.328959Z digest=sha256:1b533369275ebc2398ae149d36d52db85c86c914b157751387419b5a10af1cbe

Observation 275c8697-f4eb-44d6-a860-a909dc8abef6 · outbound

This paper cites 3d shapenets: A deep representation for volumetric shapes.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval 3d shapenets: A deep representation for volumetric shapes

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:10.885400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:07.421895Z digest=sha256:48d7c5d7186a98b21127cf6b896e8b493191f822fceef2ea382b2f5b6039923a

Observation fca0dbe6-d017-462c-a767-61c1cd11760a · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:10.685990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:07.509907Z digest=sha256:ed8d33a730498c5ab3c7970bee87fcc8997efd19b969e4c0cfd30d442ee43ffe

Observation ae7e5cdc-a3e6-4b3b-bd74-2ecaf4757c92 · outbound

This paper cites Ulip-2: Towards scal- able multimodal pre-training for 3d understanding.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Ulip-2: Towards scal- able multimodal pre-training for 3d understanding

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:10.425943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:07.617149Z digest=sha256:828919a177641af079606eb2eac3e9a66cfb6c97b298f53a98cfdc75abfb021e

Observation 89ddbaee-7950-45b1-af27-331e577b8107 · outbound

This paper cites Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:07.732480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:07.732480Z digest=sha256:02c39cb92600f9a2b0d69f3b8ed1dd3e230e9d9ee763b25d7c9d5491b63cea65

Observation e8df5b5d-f60d-4d5e-8225-a4105bb3babe · outbound

This paper cites Pointclip: Point cloud understanding by clip.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Pointclip: Point cloud understanding by clip

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:10.117082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:07.848450Z digest=sha256:ab59b8e275d49bca2eb5287f41b1bdfe40587c91a99e46c3411ac3cc39788360

Observation 05b9e1ca-ef95-4da3-b4f4-2028089a2c1e · outbound

This paper cites Pointweb: Enhancing local neighborhood features for point cloud processing.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Pointweb: Enhancing local neighborhood features for point cloud processing

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:07.936148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:07.936148Z digest=sha256:bb6924adeff465b346766d3d005392009aa63a1b2af18e9317553278e9ee4ced

Observation f10a92bd-5966-4dc6-8d42-9d3c9d6401a6 · outbound

This paper cites Multi-channel weight-sharing autoencoder based on cascade multi-head attention for multimodal emo- tion recognition.IEEE TMM, 2022.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Multi-channel weight-sharing autoencoder based on cascade multi-head attention for multimodal emo- tion recognition.IEEE TMM, 2022

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:09.781355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:08.051113Z digest=sha256:2819cdc3b40c383dd27281fd0233729b77d08d84ae1c486420582ed1754175ff

Observation e481f9f8-bba0-40ff-a853-2e15e11a0493 · outbound

This paper cites Learn- ing placeholders for open-set recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Learn- ing placeholders for open-set recognition

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:09.475429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:08.150545Z digest=sha256:661d4b65f75ef2fe0c99fa885adf11d14cd65e3ea0bf1d5fb6b517f9c9b92291

Observation 28f81f8b-4d9b-48af-8aed-88c93bcbcdff · outbound

This paper cites Uni3d: Exploring unified 3d representation at scale.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Uni3d: Exploring unified 3d representation at scale

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:09.202898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:08.268396Z digest=sha256:7fe014a3885faef14e726953f5038ca59e38c89299d72f678714531be22e45b8

Observation 45e08547-3a65-4d11-af5e-f1a239455cb7 · outbound

This paper cites Conditional prompt learning for vision-language models.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Conditional prompt learning for vision-language models

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:08.939053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:08.364823Z digest=sha256:b68a9709386fe591cfb24a4d2fcc5cd2cba44c0ef946d6d9275ce84e1c09481c

Observation 99ffab05-b885-4d1f-8e5d-d0f8446993a1 · outbound

This paper cites a synthetic 3D model view of [cls] with different an- gles.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval a synthetic 3D model view of [cls] with different an- gles

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:08.708015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:47:08.442715Z digest=sha256:6b325b802b8bc1485c7faeacaac9da0561a07e4978a76d46abace59d2086de20

Pith citing papers

No inbound Pith citation observations are available.