Pith. sign in

Paper Citation Record · LEDGER

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval

As of 17 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 0 inbound Pith citation observations for arXiv:2507.21489.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21489 v1

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:47:08.442715Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

73 of 73 outbound references displayed

  • verified exact0
  • verified fuzzy46
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e4aac53a-8eba-4245-841e-e1b4050517d1 · outbound

This paper cites GPT-4 Technical Report.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:02.365253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:02.365253Z digest=sha256:3dc37538a41442f586b3a0b983931255c997fe41e9d5bb01ee0f1eb4228e97aa

Observation b8ea0d4c-9b69-4cf2-862a-5c3cb8f73ddb · outbound

This paper cites Learning representations and generative models for 3d point clouds.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Learning representations and generative models for 3d point clouds

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.859434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:02.433911Z digest=sha256:5636885fdc682834a3f7ca58fc3a9629447dd9a415e562dc0dde78f197794ae7

Observation 999751b7-9af4-4ca8-a8cc-b92716dae7db · outbound

This paper cites Flamingo: a visual language model for few-shot learning.NeurIPS, 35: 23716–23736, 2022.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Flamingo: a visual language model for few-shot learning.NeurIPS, 35: 23716–23736, 2022

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:02.604846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:02.604846Z digest=sha256:18f497d99609b721b7f315c93d5d04835aa91811b25dfa7b6afa7ecc60309fe5

Observation a43544f7-af1f-4cbc-8234-efa0de1b77c9 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:02.780795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:02.780795Z digest=sha256:cdf6d8a50c9731b7f35ee3b41fdd9a9de45985b16470d6a29bec08bd79d201aa

Observation 4fcf1764-3975-4614-a34d-31129cb75b39 · outbound

This paper cites Qwen2.5-VL Technical Report.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Qwen2.5-VL Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:02.884100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:02.884100Z digest=sha256:a02a6d2e3decaede19f307881ff53844ccd095df55a71458637e098f56918281

Observation ee5ee636-8bd2-4873-aacc-a58928464e15 · outbound

This paper cites Shape matching and object recognition using shape contexts.IEEE TPAMI, 24(4):509–522, 2002.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Shape matching and object recognition using shape contexts.IEEE TPAMI, 24(4):509–522, 2002

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.846394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.016796Z digest=sha256:cb56770f792bdf880503654f91d76601e2d203fc9575a915ff9d8e6314c8558c

Observation 58286fd9-015a-4f4e-984a-91c3fb966a67 · outbound

This paper cites On visual similarity based 3d model retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval On visual similarity based 3d model retrieval

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.838131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.094828Z digest=sha256:c911f783a28022f65ff3cc219cc7e647b3c018d64755aaa95b7b749374dab6a8

Observation 5c7843a9-6416-421f-b833-a417b1c5031d · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:03.150233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:03.150233Z digest=sha256:2d78f62f73b6f5d63f14cff909a641ed477d248d367f2ed02b48f72bbf1abff5

Observation 5eb4c32d-f262-4321-8e27-7fbdae07ad64 · outbound

This paper cites Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.830563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.233035Z digest=sha256:4e91eaa8a4dddfb3f93ffa9f90385450b87299a5fc8ba9802c87568f4e16611a

Observation fe416e7f-c5ba-4648-9501-e4e7818801cb · outbound

This paper cites Pra-net: Point relation-aware network for 3d point cloud analysis.IEEE TIP, 30:4436–4448, 2021.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Pra-net: Point relation-aware network for 3d point cloud analysis.IEEE TIP, 30:4436–4448, 2021

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.822525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.286593Z digest=sha256:838d8472952f52e40f65fd2fc1b4068ac70117ecae43dd863845ae90659918dc

Observation cae22099-756b-4a79-ae48-34d4b763b966 · outbound

This paper cites Abo: Dataset and benchmarks for real-world 3d object un- derstanding.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Abo: Dataset and benchmarks for real-world 3d object un- derstanding

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.814815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.332545Z digest=sha256:de164cbc0789b17b42b58ee60814c36424c48b3d483d5e97f2e41e28eda15d03

Observation 8c13730c-450c-4356-a4b1-c9c1ad63dfa5 · outbound

This paper cites Siamese cnn-bilstm architecture for 3d shape representation learning.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Siamese cnn-bilstm architecture for 3d shape representation learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:03.405200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:03.405200Z digest=sha256:58ebdbb01ba520974c978221e62a457c3e408f0dc038a146cdac0384987befce

Observation f48d7503-23da-495c-9c5c-1ca5fc3aaaf0 · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Objaverse: A universe of annotated 3d objects

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.802001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.490927Z digest=sha256:6c056fc7dd54cc9406435c1d02adaf3fac6adcf3566154807a3adf548340b0f5

Observation fc43a5cb-36cb-4c0f-adff-13eabd631bfa · outbound

This paper cites Equivariant multi-view networks.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Equivariant multi-view networks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:03.553089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:03.553089Z digest=sha256:e68cac06da3e8417cb8d1795214bfc113cd28910e361b89f85d05346d797f562

Observation 2495eeda-fb6a-4edf-9f30-bbecd515182c · outbound

This paper cites Gvcnn: Group-view convolutional neural networks for 3d shape recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Gvcnn: Group-view convolutional neural networks for 3d shape recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:03.614094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:03.614094Z digest=sha256:a5b8ddbd538c56b3e3e602dd2f0c4605a1f64c01d9a7ee837a190a65b8619293

Observation 1a816f62-9280-4509-8bfc-ac887bb3fe15 · outbound

This paper cites Meshnet: Mesh neural network for 3d shape rep- resentation.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Meshnet: Mesh neural network for 3d shape rep- resentation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.783681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.664107Z digest=sha256:eee3a9617f3fb514d8cbd80d604d6d9a8cbc8e2bc6c2f32255475e22ccd9a4fc

Observation f4b2d03c-7c35-4a48-9b51-6edd2ecf2adc · outbound

This paper cites Shrec’22 track: Open-set 3d object retrieval.Computers & Graphics, 107:231–240, 2022.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Shrec’22 track: Open-set 3d object retrieval.Computers & Graphics, 107:231–240, 2022

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.775924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.721045Z digest=sha256:450ad31b41f8c535eb27b94d18e9fa1d9bfe14b16a64234b385b46ec07fbfe9b

Observation 4d4f070d-1df2-4fa2-867c-d24670e22292 · outbound

This paper cites Hypergraph-based multi-modal represen- tation for open-set 3d object retrieval.IEEE TPAMI, 2023.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Hypergraph-based multi-modal represen- tation for open-set 3d object retrieval.IEEE TPAMI, 2023

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.768338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.779910Z digest=sha256:55e518b77a100d06f2996a14387f64b663a146bb95bfde636963c02448bc5f13

Observation bcff649d-f534-48c4-a44d-4bbf168d559c · outbound

This paper cites 3d object retrieval based on similarity calculation in 3d computer aided design systems.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval 3d object retrieval based on similarity calculation in 3d computer aided design systems

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.760873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.851713Z digest=sha256:89d2b67e62e0adf8d9da938025e05f12e61fbdfe22eff9c48bcca1157aab89f7

Observation 10d8b2a7-003e-42e0-a19a-46879c8ada4a · outbound

This paper cites Clip-adapter: Better vision-language models with feature adapters.IJCV, 132(2):581–595, 2024.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Clip-adapter: Better vision-language models with feature adapters.IJCV, 132(2):581–595, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.752571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.918093Z digest=sha256:f05294d629d53da57d58a6f83b12cd8b82b60e4a50e4c561e806473cf3b70fe2

Observation 76847e5b-f7f0-4c99-b861-0e6491b9ae51 · outbound

This paper cites Deep learning for 3d point clouds: A survey.IEEE TPAMI, 43(12):4338–4364, 2020.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Deep learning for 3d point clouds: A survey.IEEE TPAMI, 43(12):4338–4364, 2020

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.744371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:03.981913Z digest=sha256:b0176872681e2c8b663480a749b5312fbe62c3105c60300f66243952ebc640ad

Observation a962f3c3-cf0f-4410-8616-8d0bc53dfd3d · outbound

This paper cites 3d2seqviews: Aggregating sequential views for 3d global feature learning by cnn with hierarchical attention ag- gregation.IEEE TIP, 28(8):3986–3999, 2019.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval 3d2seqviews: Aggregating sequential views for 3d global feature learning by cnn with hierarchical attention ag- gregation.IEEE TIP, 28(8):3986–3999, 2019

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:04.041414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:04.041414Z digest=sha256:f97b36a77f77a2db5f20ff930e9517ff394d402b9fff278883192ba1ce0b406b

Observation 88effd1b-bd57-4880-80f8-ad08d88e8368 · outbound

This paper cites Triplet-center loss for multi-view 3d object retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Triplet-center loss for multi-view 3d object retrieval

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.731351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:04.169624Z digest=sha256:66b2283149bb7fb088dee45d85a2bc29e42ae48afe176bda6820862a4301aab9

Observation 3dfcf5d2-0e93-49e3-bbaa-ab326a267836 · outbound

This paper cites View n-gram network for 3d object retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval View n-gram network for 3d object retrieval

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:04.239337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:04.239337Z digest=sha256:0a60ff016bdbd639f7b63d0dba5dabc3ac8e5ebacfc548d47d41a742ef67a101

Observation 815893c3-e0a6-4b80-abef-14fc8b3f754c · outbound

This paper cites Latformer: Locality-aware point- view fusion transformer for 3d shape recognition.PR, 151: 110413, 2024.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Latformer: Locality-aware point- view fusion transformer for 3d shape recognition.PR, 151: 110413, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.717957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:04.300096Z digest=sha256:4070c14a0a8e6cfab4adfb455c7f01da595465be4f4859bd04d78819b55e8204

Observation 37b4c765-2053-4fb8-807c-91540a2c986a · outbound

This paper cites Clip goes 3d: Leveraging prompt tuning for language grounded 3d recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Clip goes 3d: Leveraging prompt tuning for language grounded 3d recognition

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.709871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:04.368902Z digest=sha256:526637479d1f8af5aaef363ecef2681f40f0e321daf05735fc0dcb8839b0560b

Observation 7386090c-3e0a-4375-9fd0-c1e07f7a3495 · outbound

This paper cites LoRA: Low-rank adaptation of large language models.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval LoRA: Low-rank adaptation of large language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.701250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:04.423232Z digest=sha256:74521322a6bde3408f52022c6ce2f43fd376ef1ab099c1e3482131e2d3b68282

Observation 117d2cdc-be8b-4cf3-a377-2cbd543781f4 · outbound

This paper cites Scal- able deep multimodal learning for cross-modal retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Scal- able deep multimodal learning for cross-modal retrieval

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.693541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:04.470427Z digest=sha256:24131f6fd244193e01a6ca6b3c9af228cfe2ac284f8199fc3266f828462a0c2b

Observation 4daccc3d-7e4b-4ebc-9ace-3d51d8f055bd · outbound

This paper cites Clip2point: Transfer clip to point cloud classifica- tion with image-depth pre-training.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Clip2point: Transfer clip to point cloud classifica- tion with image-depth pre-training

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.685113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:04.527676Z digest=sha256:e1cb404db2cfe66cd2b9e7451f4ebc993daaf66f3c92d8a664fba179e5038113

Observation 01dbc503-4c9e-4fd6-b786-753e0d282054 · outbound

This paper cites Developing an engineering shape benchmark for cad models.Computer-Aided Design, 38(9):939–953, 2006.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Developing an engineering shape benchmark for cad models.Computer-Aided Design, 38(9):939–953, 2006

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.676422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:04.572553Z digest=sha256:854f5867d3195133e8ac4cf4fdf8a4b469fd963535c7aaf782684603bbfe83f4

Observation 3cb8bc34-6b35-4991-9fec-1da2e7051ed5 · outbound

This paper cites Cross-modal center loss for 3d cross-modal retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Cross-modal center loss for 3d cross-modal retrieval

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.529357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:04.655578Z digest=sha256:4dc66957550ec6f73bd4152a822884afa51a3e525aeff4cb0bf59a42bdd410ba

Observation 1a92f109-709f-4ad2-a78c-e501fd6c3e7b · outbound

This paper cites Rotationnet: Joint object categorization and pose estimation using multiviews from unsupervised viewpoints.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Rotationnet: Joint object categorization and pose estimation using multiviews from unsupervised viewpoints

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:04.728566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:04.728566Z digest=sha256:711962f8c5eeab44528ce24c001e5183737a2ccb0a71dc2cf87ea5ca0f8e562b

Observation 44a49962-a314-47a7-b401-c029592c416c · outbound

This paper cites Rotation invariant spherical harmonic repre- sentation of 3 d shape descriptors.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Rotation invariant spherical harmonic repre- sentation of 3 d shape descriptors

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:14.259037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:04.817478Z digest=sha256:2cfd63a781f2e3a7fcd680c582be1d882ea22d9009816a91a0e933e6104d86f7

Observation 9243c747-567d-40b2-b973-5ffb1feeeb1c · outbound

This paper cites Re- search challenges for digital archives of 3d cultural her- itage models.Journal on Computing and Cultural Heritage (JOCCH), 2(3):1–17, 2010.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Re- search challenges for digital archives of 3d cultural her- itage models.Journal on Computing and Cultural Heritage (JOCCH), 2(3):1–17, 2010

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:13.992856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:04.879461Z digest=sha256:86ed3176d6228288f345cd928a2befb9e49ff71438e16e6aff9deddaf7e63ff8

Observation 4ab4e967-f98d-4b2d-a1d5-b3c9b8692492 · outbound

This paper cites Angular triplet- center loss for multi-view 3d shape retrieval.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Angular triplet- center loss for multi-view 3d shape retrieval

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:04.974880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:04.974880Z digest=sha256:7a0ad87adf7cf35b563a483f0e8352db7fb8746e60755909fd963cb481076128

Observation bc8403e9-a325-4f6f-8ae6-c364dcd4624e · outbound

This paper cites Meshmae: Masked autoencoders for 3d mesh data analysis.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Meshmae: Masked autoencoders for 3d mesh data analysis

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:13.736307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:05.049444Z digest=sha256:0f1333925a6ef2762e3fbdbf33b19ea975b4ff8436f5244bfa370f983da965da

Observation 5e48812c-ecfb-49ac-b19d-496b2f16b128 · outbound

This paper cites Improved baselines with visual instruction tuning.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Improved baselines with visual instruction tuning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:13.482566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:05.138939Z digest=sha256:356d88954932394099d97211de2a7bab533e39b81b965c8a9fb3f20b564338be

Observation 264362b9-fa81-4531-b6c5-7806627be3d0 · outbound

This paper cites Visual instruction tuning.NeurIPS, 36, 2024.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Visual instruction tuning.NeurIPS, 36, 2024

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:13.133129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:05.175939Z digest=sha256:f6fe74d522c9ead1d57e0506eae4085646834f4a652e707e908c2eb3b5300be4

Observation 8a090f86-d450-4e86-b979-d7a3d59422da · outbound

This paper cites Openshape: Scaling up 3d shape representation towards open-world understanding.NeurIPS, 36, 2023.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Openshape: Scaling up 3d shape representation towards open-world understanding.NeurIPS, 36, 2023

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:12.863259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:05.277656Z digest=sha256:d5690feff420f17a3f569a040f13146d70475b9ddb75c2bcd6c8afdfcfd49fe0

Observation c71387ad-3397-4a1b-96cc-a86cebe9f79f · outbound

This paper cites Point2sequence: Learning the shape representa- tion of 3d point clouds with an attention-based sequence to sequence network.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Point2sequence: Learning the shape representa- tion of 3d point clouds with an attention-based sequence to sequence network

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:05.339949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:05.339949Z digest=sha256:9a1d47690299f2ca5fc2343b3411363f002fa15443fbc88f3e1096ff1cce75b1

Observation 387e2f23-ae6f-4852-91fb-4a6812ffd73a · outbound

This paper cites Relation-shape convolutional neural network for point cloud analysis.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Relation-shape convolutional neural network for point cloud analysis

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:05.426858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:05.426858Z digest=sha256:d3c8a43f209c17661b4a1e083d211c105ca1368406484ec270bb44ba1c286ef5

Observation 4e2d60e5-80c3-4e40-92a4-26d27bc026b4 · outbound

This paper cites An Empirical Study of Scaling Instruct-Tuned Large Multimodal Models.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval An Empirical Study of Scaling Instruct-Tuned Large Multimodal Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:05.523805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:05.523805Z digest=sha256:5070574ff313ef51ea28bd46ca5048d2f820da47f120a7e7eabccf39ad4e898e

Observation 0315681a-189f-43f7-a31b-854b87843fc2 · outbound

This paper cites 3d learning objects for augmented/virtual reality educational ecosystems.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval 3d learning objects for augmented/virtual reality educational ecosystems

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:12.619734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:05.621772Z digest=sha256:9ba79a6a18ed66f8bca153543461b2003af39711a70d825742b4dbb3679cb75c

Observation bb366d37-50fe-4ac3-8559-cc8366f9941c · outbound

This paper cites V oxnet: A 3d con- volutional neural network for real-time object recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval V oxnet: A 3d con- volutional neural network for real-time object recognition

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:12.433140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:05.685811Z digest=sha256:a25770c27f2dc2ddd3a953ee5f762759e0b7e7ae6822c0d2b0b9fa6dbd40b1f9

Observation 16e478c5-a186-41a6-9a3c-e69da20444fb · outbound

This paper cites Mmjn: Multi-modal joint networks for 3d shape recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Mmjn: Multi-modal joint networks for 3d shape recognition

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:12.273646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:05.754575Z digest=sha256:dc0429e9797b827ebaba1a68f51f496913bb399123a3a773987873c0fd661bce

Observation 30ad683e-6621-4b01-bd3f-c78dfdf80f9f · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Representation Learning with Contrastive Predictive Coding

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:05.854039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:05.854039Z digest=sha256:930e509dc382b52239f495cd89918ae33bc502429950cc8c57e7d1994d3452f1

Observation 4fcbc6c2-0bbd-44ee-bec3-5f48dbc128ec · outbound

This paper cites Kosmos-2: Grounding Multimodal Large Language Models to the World.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Kosmos-2: Grounding Multimodal Large Language Models to the World

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:05.946843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:05.946843Z digest=sha256:821e67bdabbdbf2e3f2263fad282816dbe96dad65c0421a416672c4d4712c6cc

Observation c7c831b5-0ae5-464f-9eaa-86b921d388a5 · outbound

This paper cites Pointnet: Deep learning on point sets for 3d classification and segmentation.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Pointnet: Deep learning on point sets for 3d classification and segmentation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.016588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.016588Z digest=sha256:785646731a0233df157f43e6268ea4642ab27d4f4f5f8f1d7b75d346dc8a3fa0

Observation 8029e1ed-a18e-4386-ac80-e70f2aabe21a · outbound

This paper cites Pointnet++: Deep hierarchical feature learning on point sets in a metric space.NeurIPS, 30, 2017.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Pointnet++: Deep hierarchical feature learning on point sets in a metric space.NeurIPS, 30, 2017

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.111380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.111380Z digest=sha256:1d70555f08da4324ef89602629c4e3d9a236a4ed2f6ff83178c2f4019dc4bd4f

Observation d8739b54-3665-4a47-b467-36176d8b2325 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Learn- ing transferable visual models from natural language super- vision

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:12.064361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:06.175609Z digest=sha256:dc108123eed2041f8a5523e8d990dd883c8bdc282c241431a83d8d8b6f096580

Observation dc8afc3f-b0ee-46a0-8976-33c0606fed60 · outbound

This paper cites Clip for all things zero-shot sketch-based image retrieval, fine- grained or not.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Clip for all things zero-shot sketch-based image retrieval, fine- grained or not

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.906256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:06.305599Z digest=sha256:f4b9216869d8e905a8a28f0c87322f7880795e1fc98afd214fb01a1b5bd3da65

Observation bda63105-a27b-41c7-a280-a5f7059b94f3 · outbound

This paper cites Deep- voxels: Learning persistent 3d feature embeddings.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Deep- voxels: Learning persistent 3d feature embeddings

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.406724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.406724Z digest=sha256:ce5242f304b100bb7ad805069aff7fa2739ff1939ec3a843e6906bdb70a677b8

Observation 093cc193-e569-423b-bba7-a872a33dc293 · outbound

This paper cites MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.495224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.495224Z digest=sha256:aafc1f91848eacff7f8d7c66be72324e273f1a292fcf9555da3e15b687b5e1a9

Observation 09ecf3ad-d5ec-4928-bba1-82ec6115d2f7 · outbound

This paper cites Multi-view convolutional neural networks for 3d shape recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Multi-view convolutional neural networks for 3d shape recognition

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.726653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:06.598342Z digest=sha256:b4029cd39a51567f3062447469c1af7d01448977ebdb69874bf0ac9dba2f0b95

Observation e7c38b4e-a674-400a-9ecb-4b353a3fbb1a · outbound

This paper cites A survey of content based 3d shape retrieval methods.Proceedings Shape Modeling Applications, 2004., pages 145–156, 2004.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval A survey of content based 3d shape retrieval methods.Proceedings Shape Modeling Applications, 2004., pages 145–156, 2004

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.555829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:06.714822Z digest=sha256:18af06449d8bb4df688eb1569a88cbf90eba66685fcf2318c614a409c3ee95ae

Observation ff79b94f-e89e-434f-93ce-072747635032 · outbound

This paper cites O-cnn: Octree-based convolutional neural networks for 3d shape analysis.ACM TOG, 36(4):1–11,.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval O-cnn: Octree-based convolutional neural networks for 3d shape analysis.ACM TOG, 36(4):1–11,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.823291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.823291Z digest=sha256:7ce6e981caaffb0f5319da73f44b91ce5369fe55993e6041505259c95e1d1605

Observation a403d6d1-6d31-47a6-a48e-6014de054a3b · outbound

This paper cites The All-Seeing Project: Towards Panoptic Visual Recognition and Understanding of the Open World.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval The All-Seeing Project: Towards Panoptic Visual Recognition and Understanding of the Open World

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:06.899224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:06.899224Z digest=sha256:2b69821afdda3b125fd6b92ccf2a328d2a164999e9b1cb827d0efe91117f8719

Observation e314d4b6-7a1d-4cc5-a012-1192015977bc · outbound

This paper cites Visionllm: Large language model is also an open- ended decoder for vision-centric tasks.NeurIPS, 36, 2024.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Visionllm: Large language model is also an open- ended decoder for vision-centric tasks.NeurIPS, 36, 2024

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.375033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:06.982283Z digest=sha256:2d6572c23c55617c4b97b0109b1a127f8bb4519740240ad4ce6516522121c526

Observation 1380f257-583f-429f-9251-464e7b45e0d7 · outbound

This paper cites Dynamic graph cnn for learning on point clouds.ACM TOG, 38(5): 1–12, 2019.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Dynamic graph cnn for learning on point clouds.ACM TOG, 38(5): 1–12, 2019

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:07.080674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:07.080674Z digest=sha256:bd3c8a4efeb1ed8293e4780c21f526eb9976dd3cc2ce9d03b49825d71d2d34e6

Observation c94a750d-7e34-4468-93b5-fbe661deca4c · outbound

This paper cites Teda: Boosting vision-lanuage models for zero-shot 3d object retrieval via testing-time distribution alignment.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Teda: Boosting vision-lanuage models for zero-shot 3d object retrieval via testing-time distribution alignment

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.173636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:07.146620Z digest=sha256:f73ee364b7545195e002a2784ebef5e32539d3aafe45b7d31cf8ba8ad0464e86

Observation d7f9ed69-099a-468d-876a-b5a4a4f2abc6 · outbound

This paper cites View-gcn: View-based graph convolutional network for 3d shape analysis.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval View-gcn: View-based graph convolutional network for 3d shape analysis

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:07.221969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:07.221969Z digest=sha256:702a4ffd048ec29a6fe0a9b7677a1914de04bdbbf3281725f8f0c3ab4d300e25

Observation 85db8e0f-ccce-4c71-b8e5-487058e71b2e · outbound

This paper cites Multi- modal semantic autoencoder for cross-modal retrieval.Neu- rocomputing, 331:165–175, 2019.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Multi- modal semantic autoencoder for cross-modal retrieval.Neu- rocomputing, 331:165–175, 2019

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:11.027826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:07.328959Z digest=sha256:8994697bea55fa005209feac5d7e8c52c54a328e2e4fe50e798efbde3c1c1212

Observation 275c8697-f4eb-44d6-a860-a909dc8abef6 · outbound

This paper cites 3d shapenets: A deep representation for volumetric shapes.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval 3d shapenets: A deep representation for volumetric shapes

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:10.885400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:07.421895Z digest=sha256:212b33894074ee1ea478fdaac1961eb6f81ca0e7b07b57cc3553f2512f642853

Observation fca0dbe6-d017-462c-a767-61c1cd11760a · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:10.685990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:07.509907Z digest=sha256:f7cf8fc2a5a4e3d9b7ede30dbc96477a6f9c66416878aab40c9633a404bd3954

Observation ae7e5cdc-a3e6-4b3b-bd74-2ecaf4757c92 · outbound

This paper cites Ulip-2: Towards scal- able multimodal pre-training for 3d understanding.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Ulip-2: Towards scal- able multimodal pre-training for 3d understanding

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:10.425943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:07.617149Z digest=sha256:e8f640b8616fa81c011c4cec406db4b91d96ad6636c772ba9f15958d9a368f85

Observation 89ddbaee-7950-45b1-af27-331e577b8107 · outbound

This paper cites Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:07.732480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:07.732480Z digest=sha256:f114a581a9afa059cce9767c2d811761ee6d2a59bd1e1d30432c6cf6fabc3c4b

Observation e8df5b5d-f60d-4d5e-8225-a4105bb3babe · outbound

This paper cites Pointclip: Point cloud understanding by clip.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Pointclip: Point cloud understanding by clip

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:10.117082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:07.848450Z digest=sha256:9c9910026227cd22767256457096c096a8f20e47bfdbb84483729859dfd61a29

Observation 05b9e1ca-ef95-4da3-b4f4-2028089a2c1e · outbound

This paper cites Pointweb: Enhancing local neighborhood features for point cloud processing.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Pointweb: Enhancing local neighborhood features for point cloud processing

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T12:47:07.936148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:47:07.936148Z digest=sha256:5927b91eacd4cfa8352d7aad423b61ea1a80d81bd8275a730fc8ba20d098c18e

Observation f10a92bd-5966-4dc6-8d42-9d3c9d6401a6 · outbound

This paper cites Multi-channel weight-sharing autoencoder based on cascade multi-head attention for multimodal emo- tion recognition.IEEE TMM, 2022.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Multi-channel weight-sharing autoencoder based on cascade multi-head attention for multimodal emo- tion recognition.IEEE TMM, 2022

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:09.781355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:08.051113Z digest=sha256:da0eac42476d833e14025720b7b319e2e0b337802acd5a89a286b09651a9e8ff

Observation e481f9f8-bba0-40ff-a853-2e15e11a0493 · outbound

This paper cites Learn- ing placeholders for open-set recognition.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Learn- ing placeholders for open-set recognition

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:09.475429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:08.150545Z digest=sha256:9eb3d0ca1bfe7a43953a535a42f1c178a635484b7f0def297a4466831c3bf44c

Observation 28f81f8b-4d9b-48af-8aed-88c93bcbcdff · outbound

This paper cites Uni3d: Exploring unified 3d representation at scale.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Uni3d: Exploring unified 3d representation at scale

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:09.202898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:08.268396Z digest=sha256:de773be368a0c3cb8cb258a20d9f6dee826cb4b0174c380f263ba41047d42ab2

Observation 45e08547-3a65-4d11-af5e-f1a239455cb7 · outbound

This paper cites Conditional prompt learning for vision-language models.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval Conditional prompt learning for vision-language models

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:08.939053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:08.364823Z digest=sha256:9b491eb10592d4c4142dc6404e9b983fab62fa6053ab726a78c960eec18ced9b

Observation 99ffab05-b885-4d1f-8e5d-d0f8446993a1 · outbound

This paper cites a synthetic 3D model view of [cls] with different an- gles.

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval a synthetic 3D model view of [cls] with different an- gles

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:47:08.708015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T12:47:08.442715Z digest=sha256:3623056ebfdf019a3204c713c5df768b5956f7cc1407baa8696f1742e22754cf

Pith citing papers

No inbound Pith citation observations are available.