Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:39:12.809687Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 3 inbound Pith citation observations for arXiv:2501.07819.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:39:12.809687Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:46:14.473013Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-17T22:45:24.456612Z
68 of 68 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 86537c9f-e7a5-4eb3-9b3e-f488dcd4dcec · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Chatgpt: Optimizing language models for dialogue,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9c87d6d7-194d-4d0b-bc4c-e48f9b4be5f9 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding LLaMA: Open and Efficient Foundation Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a7c1dc8-ba53-4303-b5b0-eb004f24549e · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding The Llama 3 Herd of Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10c6979b-6176-466d-acbd-8629b503412b · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Gemini: A Family of Highly Capable Multimodal Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00713a15-c5e2-4d3d-83c9-7a4288d67de6 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4ce9ec64-8924-40b1-b6f9-e35cca63b96b · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25316d5b-6cd3-4a81-9267-ad2b14393f95 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Improved Baselines with Visual Instruction Tuning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54232287-b629-49b0-9e68-f757043a9a7e · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding X-InstructBLIP: A Framework for aligning X-Modal instruction-aware representations to LLMs and Emergent Cross-modal Reasoning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 882b9a20-4a37-4b81-b3f2-b8c0f70adf50 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Scanqa: 3d question answering for spatial scene understanding,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a9548686-ba0b-4b81-b9d0-bb565b686606 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding SQA3D: Situated Question Answering in 3D Scenes
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0db83eb1-3ce6-41d2-bf9c-925ac1f187f3 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 911bfe9a-57be-4a5c-8eda-440d9954f8d6 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Deep hough voting for 3d object detection in point clouds,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a4988914-e146-436c-984a-5b43be4ccc57 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding 3d-llm: Injecting the 3d world into large language models,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee3aac66-f8e7-40e5-8e73-9779bf3879af · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Segment anything,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 75a1cecc-f363-4e4d-871c-366da6dbc5d8 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Masked-attention mask transformer for universal image segmentation,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation abb9dd29-7bf7-41cf-accf-bb0bdf76d39d · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Training language models to follow instructions with human feedback,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e4a6011-888d-492b-9274-6208796adb90 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding An end-to-end transformer model for 3d object detection,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 67d4dd9f-7124-46bb-9445-7cb3bbb22906 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 40b8c5a7-a891-41b9-993e-4fe940c5c8e7 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30ec9318-0d55-4ac6-94ed-7f53fb0d6833 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Scannet: Richly-annotated 3d reconstructions of indoor scenes,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76cad131-e458-4a6f-a361-83fd307fa6a0 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Rio: 3d object instance re-localization in changing indoor environments,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cc31d721-d77e-4517-a99e-e737e9b2a136 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Exploring the limits of transfer learning with a unified text-to-text transformer,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0ae896f7-fc7f-4c93-91e3-2cbb38a8dd87 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Improving language understanding by generative pre-training,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da4c0956-d103-42e3-9ca8-86c7b227b352 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Attention is all you need,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5dbffb5-a6f5-4d57-ad0f-4e15ee6b49e7 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Language Models are Few-Shot Learners
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb4d0a92-6165-4aeb-9219-f20695ae9012 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Palm: Scaling IEEE TRANSACTIONS ON MULTIMEDIA 12 language modeling with pathways,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1fd794bd-8bd1-4963-8aed-041f62183640 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Baichuan 2: Open Large-scale Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f536e8c-4164-41ed-8283-950e810934fa · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Flamingo: a visual language model for few-shot learning,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caaab393-4c06-4497-b3f3-6061c181e667 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Ferret: Refer and Ground Anything Anywhere at Any Granularity
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c7cb602-e538-4525-b269-1e726bc37eeb · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding OneLLM: One Framework to Align All Modalities with Language
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2b8a48b-ad0d-4ca3-bfdb-530e479272af · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Scanrefer: 3d object localization in rgb-d scans using natural language,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 911dd5a4-4baf-46cb-b6c8-95473bfd0bbf · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Scan2cap: Context-aware dense captioning in rgb-d scans,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c5b17490-9bf3-4fdb-a87a-868bbcd0b204 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Proposalcontrast: Unsupervised pre-training for lidar-based 3d object detection,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6b07d920-cd08-4cb2-a785-23756e104843 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Graph neural network and spatiotemporal transformer attention for 3d video object detection from point clouds,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8a1fc6ef-ad7c-4e90-85ed-171f1f181541 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Is-fusion: Instance-scene collaborative fusion for multimodal 3d object detection,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5d0485a7-4762-46db-bb92-6ca5cd33ec80 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Lsk3dnet: Towards effective and efficient 3d perception with large sparse kernels,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8944b8aa-cb8d-408c-a235-f165081895d1 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Shape2Scene: 3D Scene Representation Learning Through Pre-training on Shape Data
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe59aff1-0f92-4177-867f-3237c083a435 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Clustering based point cloud representation learning for 3d analysis,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cf5e707a-e2e3-4ed6-8816-3c514a1cde6a · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Weakly supervised 3d object detection from lidar point cloud,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation eb0cf1c9-3fa3-4b9b-9c30-52b079489d09 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding PointLLM: Empowering Large Language Models to Understand Point Clouds
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9690402-a39d-4b5e-b218-a55365a7d2d2 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding ShapeLLM: Universal 3D Object Understanding for Embodied Interaction
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 199f1a3b-46a0-4550-9002-399de16f956c · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7146b3f7-ccfe-4720-bd8b-5d861a27bb73 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Openshape: Scaling up 3d shape representation towards open-world understanding,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 11682bad-c7ed-437e-a279-7bb85f95b4d8 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Contrast with reconstruct: Contrastive 3d representation learning guided by generative pretraining,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 84f41382-658e-4b23-9bbc-0f1133c29f1c · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43aef20e-db89-4dff-8531-be3e8189dbf2 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding BAMBOO: A Comprehensive Benchmark for Evaluating Long Text Modeling Capacities of Large Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2961d74e-47ac-4eb1-aefa-c5cf3c4b02df · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e072ceeb-fb24-4ee5-8742-ac729be88d55 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Visual instruction tuning,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 51a66ac0-39ab-465c-86b7-5954b17470f3 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 489413b1-2a61-445c-aeb1-fbb29118e12b · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Language models are unsupervised multitask learners,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27ae5591-3cc4-40d7-8fc5-4dfe36296ef3 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Learning transferable visual models from natural language supervision,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b03ea6e5-2791-4900-a1b2-af686b0e5f11 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding 4d spatio-temporal convnets: Minkowski convolutional neural networks,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb37209c-9864-44f0-96fd-5607582103c0 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Point transformer,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9dbd8e82-1bb1-4f20-b1c0-e73f73c3c873 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding slam: Dense slam meets automatic differentiation,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5287a31b-8269-4532-9396-18bd72f82627 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92a66cc1-1d39-49e0-aa03-98ddec7f179b · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Context-aware alignment and mutual masking for 3d-language pre-training,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a7b91f80-1c30-452a-b690-2b3b0e639472 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Ll3da: Visual interactive instruction tuning for omni-3d understanding reasoning and planning,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 01d1941d-f9f0-4048-aa9b-c0a58889375c · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Cider: Consensus- based image description evaluation,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2888d537-fda7-40f5-806d-4849acac4193 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Bleu: a method for automatic evaluation of machine translation,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5c93929f-3281-45a8-b720-fba181e57d35 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Meteor: An automatic metric for mt evaluation with improved correlation with human judgments,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b33b90a1-ac3f-4f2d-9fc1-8d37aa5a8a1e · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Rouge: A package for automatic evaluation of summaries,
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98e26045-219e-4bfe-be49-e6b22eecc441 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Mask3d: Mask transformer for 3d semantic instance segmentation,
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 812c78b1-b1b7-4895-a80b-d5fd3c69dd3a · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Judging llm-as-a-judge with mt-bench and chatbot arena,
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 836d19e6-d68b-4282-8b6c-9606a30d2992 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding 3d-vista: Pre-trained transformer for 3d vision and text alignment,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9b6f0acb-dab6-4861-9b44-8b2625642794 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Scaling Instruction-Finetuned Language Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e46bb015-5a37-4b73-94d5-2350e39a3335 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Decoupled Weight Decay Regularization
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d894135c-e36f-4193-8e0e-be7883d1cb40 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding V otenet: A deep learning label fusion method for multi-atlas segmentation,
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 87d23797-1061-4368-a0bb-6ba93a11f832 · outbound
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding Deep modular co-attention networks for visual question answering,
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 58d23528-1edd-4483-a4f0-a17c2c013d05 · inbound
Uni3D-MoE: Scalable Multimodal 3D Scene Understanding via Mixture of Experts 3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e4e8dc4-e183-471f-a647-449bd00eab2b · inbound
PySeizure: A single machine learning classifier framework to detect seizures in diverse datasets 3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc07f7e3-e66e-40f2-bf53-9bebee77e807 · inbound
Abstract 3D Perception for Spatial Intelligence in Vision-Language Models 3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.