Pith. sign in

Paper Citation Record · LEDGER

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models

As of 9 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2509.03837.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03837 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:42:48.527090Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7661e347-db4e-4686-9caa-9f1ff590eb2b · outbound

This paper cites Artificial General Intelligence (AGI)- Native Wireless Systems: A Journey Beyond 6G,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Artificial General Intelligence (AGI)- Native Wireless Systems: A Journey Beyond 6G,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:50.339169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:42:46.682072Z digest=sha256:8d724eaa16ddfbee569b0968da149dd2d86c88442e1ffeb65bc8ef432b3ea41d

Observation 1eb62e7e-2fe7-4694-a9d1-4df9397e1ca0 · outbound

This paper cites Joint Sensing, Communication, and AI: A Trifecta for Resilient THz User Experiences,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Joint Sensing, Communication, and AI: A Trifecta for Resilient THz User Experiences,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:50.176014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:42:46.741361Z digest=sha256:fe60add34f7e05e4db40d5c8bdb9c896721ef4277e3461b6895de6e64e4e58e7

Observation 3955825f-f184-4f05-a21c-3c29af1e81e0 · outbound

This paper cites Sensing-Assisted High Reliable Communication: A Transformer- Based Beamforming Approach,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Sensing-Assisted High Reliable Communication: A Transformer- Based Beamforming Approach,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:50.027458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:42:46.826433Z digest=sha256:8632a0316edcc1aab55aec6146e5a4b3a1b8e20393b6f029f8911582c3ef6c42

Observation 7b419d69-ff08-4621-83d0-3bedcc8fadfd · outbound

This paper cites Multimodal Transformers for Wireless Communications: A Case Study in Beam Prediction.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Multimodal Transformers for Wireless Communications: A Case Study in Beam Prediction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:46.916383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:46.916383Z digest=sha256:4f8aaa53916685cf8ec06f6bb6784ab469fa42336205ae123dfbbf4e4123d3a6

Observation 5667bcab-5266-4e07-855e-e6878ba03c14 · outbound

This paper cites Vision-Aided 6G Wireless Communications: Blockage Prediction and Proactive Handoff,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Vision-Aided 6G Wireless Communications: Blockage Prediction and Proactive Handoff,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.886913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:42:46.979068Z digest=sha256:b53460de35bd2faf2ed862d977a5490b2be42c64138d8e99a4961d84c87b9d22

Observation 697b5786-72d8-47c7-9598-6df083e26354 · outbound

This paper cites Passive Radar at the Roadside Unit to Configure Millimeter Wave Vehicle-to-Infrastructure Links,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Passive Radar at the Roadside Unit to Configure Millimeter Wave Vehicle-to-Infrastructure Links,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.746056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:42:47.075053Z digest=sha256:c3a15bfba9df849da661081516ad1fd22edb22bc061b0a4be20a546396269fa2

Observation e0a2efc5-5f75-4781-9b54-9e3039174fb5 · outbound

This paper cites BeamLLM: Vision-Empowered mmWave Beam Prediction with Large Language Models.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models BeamLLM: Vision-Empowered mmWave Beam Prediction with Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.149512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.149512Z digest=sha256:096c6324c94e742b8599b543082bcebcc19642f30d767baa3dab40907ca79a3f

Observation 8fbda038-5192-4732-bea9-46dde18623d6 · outbound

This paper cites LLM4CP: Adapting Large Language Models for Channel Prediction,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models LLM4CP: Adapting Large Language Models for Channel Prediction,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.581498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:42:47.254853Z digest=sha256:e5fad6144ce81ee108ce08daf886d475562b0e8f18bcefd4e3d0c7831c8e9581

Observation 0c398674-2cee-4ea3-98f3-984b53cfc4fb · outbound

This paper cites Port-LLM: A Port Prediction Method for Fluid Antenna based on Large Language Models.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Port-LLM: A Port Prediction Method for Fluid Antenna based on Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.360252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.360252Z digest=sha256:0282d9d4d7e6f4587aadbf834be6a0521369d735b25c35af9b4d88abb6468110

Observation c860dd26-4d5f-4bfb-88ba-c6d787b66330 · outbound

This paper cites Visual Instruction Tuning.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Visual Instruction Tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.458653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.458653Z digest=sha256:1b82bfbced266c5fbfe2016355dee0c798b7318839f8518f2a798a838031478d

Observation b89035b5-0201-4c0c-9ff5-2e2c1b0fb6d1 · outbound

This paper cites Large Language Models Empower Multimodal Integrated Sensing and Communication,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Large Language Models Empower Multimodal Integrated Sensing and Communication,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.442427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:42:47.533042Z digest=sha256:91471d988a9b98e5b51460f2bafa1ec7a8ff78ba3c082251f2c06b0857840544

Observation a70ff868-81ef-4b14-b30e-b0358a998655 · outbound

This paper cites Spatial-RAG: Spatial Retrieval Augmented Generation for Real-World Geospatial Reasoning Questions.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Spatial-RAG: Spatial Retrieval Augmented Generation for Real-World Geospatial Reasoning Questions

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.639839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.639839Z digest=sha256:263e0c6e79ede5e6f98fd4fbacc59954917bffa2edbe705b06392d5e31f65d0f

Observation db04601f-b74a-4f9f-bf04-711e93fad118 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.749665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.749665Z digest=sha256:9c7758d98a1d526afe66732e9d19e5eec550aefbbb6cd4d380e76d6583522855

Observation a5998c3e-1889-4f78-ab29-279feb60abf6 · outbound

This paper cites BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird’s-Eye View Representation,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird’s-Eye View Representation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.281872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:42:47.879642Z digest=sha256:e0ceca47268bf15b24f95aac702608c20c8deded31aac60022bc29abeb003fc9

Observation 0b72d4ed-8679-4e5f-a2b2-d9c4406ad130 · outbound

This paper cites BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.959381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.959381Z digest=sha256:ab874639a18192b55c6c925f60bebb00c325f79bd0721699cda1a2127c916f1b

Observation de258daa-0b44-4e0d-95ac-a9fccce2c92d · outbound

This paper cites CARLA: An Open Urban Driving Simulator.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models CARLA: An Open Urban Driving Simulator

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:48.041403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:48.041403Z digest=sha256:61f8b73ae01f0e7e902c41450095e4858a0c7d3ae6954e3a7d39fb0dfbb23a87

Observation 35bc6856-6a6c-435f-8c9c-baacca2566c0 · outbound

This paper cites an unresolved cited work.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:48.123049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:48.123049Z digest=sha256:e9c91a11d29dc3f4c924f43471392677a09c8947b54d018393edaa252e8f01c6

Observation f5ce4959-c9c0-45fb-92f9-845783b3d391 · outbound

This paper cites Llama 3.2: Revolutionizing edge AI and vision with open, customiz- able models.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Llama 3.2: Revolutionizing edge AI and vision with open, customiz- able models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:49.131525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:42:48.203558Z digest=sha256:e32ce74eb5a9e20b7d84d9a3089392abdc351de0ee7efaa91a3709d4b4a5b67f

Observation 202cb994-76c8-49af-ba6d-910b44e7c37a · outbound

This paper cites Decoupled Weight Decay Regularization.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Decoupled Weight Decay Regularization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:48.288180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:48.288180Z digest=sha256:f0b91c98715f651f18b62ccf849f131297e29a52b9855a95c00162d66390421f

Observation acc90f7d-326a-4382-b3af-da66c481179d · outbound

This paper cites Long Short-Term Memory,.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Long Short-Term Memory,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:42:48.986723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T10:42:48.375998Z digest=sha256:b126e8f9fa243a62cba87f15ea15fdbae52e13643f91d1b68bb1eba3ae394c8d

Observation 64e6b651-f7b1-4f91-907c-3fabaf230c5f · outbound

This paper cites Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:48.461598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:48.461598Z digest=sha256:9590880215291ad15e8c1445f142c7f8a95ae8ac31ac83b353c11408b102006b

Observation 33df7c63-133b-45d6-bb14-8a8156e47940 · outbound

This paper cites Attention Is All You Need.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models Attention Is All You Need

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:48.527090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:48.527090Z digest=sha256:44634e00e95870e575aaa1d25ef7791a70c2c04d8590b93623fd92ec788d9605

Pith citing papers

No inbound Pith citation observations are available.