Pith. sign in

Paper Citation Record · LEDGER

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency

As of 7 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2507.07938.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07938 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:33:26.603654Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved13
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 20539d96-2825-4042-a2bb-3fb6e3d6b48d · outbound

This paper cites A survey of autonomous driving: Common practices and emerging technolo- gies,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency A survey of autonomous driving: Common practices and emerging technolo- gies,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:24.489572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:24.489572Z digest=sha256:76884b6e9896b8573418cda27a0ed8f9282748882178e86e90145f232492a0df

Observation 8520d665-ff66-47ee-8dbd-4e72f7cb95e7 · outbound

This paper cites ClusT3: Information Invariant Test-Time Training.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency ClusT3: Information Invariant Test-Time Training

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:24.607371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:24.607371Z digest=sha256:c99bad613c2342d60b260266426ad8327619b345e220347547d684ed11e0831c

Observation f82fddac-1ecb-415b-9d97-1fe8922bc95b · outbound

This paper cites Planning-oriented autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Planning-oriented autonomous driving,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:24.769558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:24.769558Z digest=sha256:1452a9848f973c8f1d0f54e1d03a07354b7e85b4e9888fc91fc954f78adc7735

Observation 6c117f6c-590a-4b39-85ec-4e8a6c0c984d · outbound

This paper cites Deep multi-modal object detection and semantic segmentation for au- tonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Deep multi-modal object detection and semantic segmentation for au- tonomous driving,

Reference 4

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T18:33:27.877104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:33:24.896397Z digest=sha256:f32b89edf384a0adec28d8ab39ab50abf97fa7fa01615c42bbd3a3bb9ce26087

Observation 7e43c368-c3a3-400b-b3cd-23fdfd13c3fc · outbound

This paper cites Ex- plainable artificial intelligence (XAI),.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Ex- plainable artificial intelligence (XAI),

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.053146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.053146Z digest=sha256:89d0544779d3b5feb6d162d15c55b4fe4b2ce66714de21dfd4ce62fa1356785f

Observation 2a9fe5ed-4484-4588-930e-de14a4d03178 · outbound

This paper cites Why did the AI make that decision?,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Why did the AI make that decision?,

Reference 6

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T18:33:27.596507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:33:25.148590Z digest=sha256:71c6f9a691b68d6cb60ff7d37386a428624d4bfec22059e8b50fba91ec28679e

Observation 71590303-b49f-48f7-9bed-69c9d9e00b4d · outbound

This paper cites Ex- plainable artificial intelligence (XAI),.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Ex- plainable artificial intelligence (XAI),

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.251827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.251827Z digest=sha256:9cb5277c7b51a30e439aaabe5e5bbe6515940cd702a968f04a0f82e405e12fe9

Observation 45f9e0cd-7f03-410d-b6de-9df6a1259202 · outbound

This paper cites Interpretable autonomous driving: A sur- vey of recent advances,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Interpretable autonomous driving: A sur- vey of recent advances,

Reference 8

Resolution
malformed identifier
no resolver link, observed 2026-08-06T18:33:25.417228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.417228Z digest=sha256:87221ede9bb9b7e144e35797ef36eb3311392c18a1ccd60fee2c29977ddafc73

Observation 63ae09ef-8a4d-4c8b-af86-25bb78cb87e8 · outbound

This paper cites Attention- based multimodal framework for au- tonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Attention- based multimodal framework for au- tonomous driving,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.485181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.485181Z digest=sha256:a0331b83166960a62bfffb0afd50f3f6853999cb5489eeee157ae54a0bf0801c

Observation bbbe183c-1677-42ea-b03b-7709cd130195 · outbound

This paper cites DeepDriving: Learning affordance for direct perception in autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency DeepDriving: Learning affordance for direct perception in autonomous driving,

Reference 10

Resolution
malformed identifier
no resolver link, observed 2026-08-06T18:33:25.565184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.565184Z digest=sha256:101a8fa6780d4a534fb036e37b3650af53f8c316ff2b6dc0e837fae39084d3ea

Observation fdfd001c-ed76-4d3c-8e29-360eeecfa416 · outbound

This paper cites CARLA: An Open Urban Driving Simulator.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency CARLA: An Open Urban Driving Simulator

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.619982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.619982Z digest=sha256:a450bca2b1bd101e877e671c6e2f271b3b5d1a0eafa16048c8c39852c496e2c6

Observation 13d5428e-6ba5-44f3-9677-ad9af44217c8 · outbound

This paper cites VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.677404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.677404Z digest=sha256:6fea8c91464695127d6d05bdefe483485b5da8dfd807f5d398b35f962d1f6028

Observation 62b49ade-ec11-4b84-97cf-ba2a095ad09f · outbound

This paper cites BERT: Pre-training of deep bidi- rectional transformers for language under- standing,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency BERT: Pre-training of deep bidi- rectional transformers for language under- standing,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.751970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.751970Z digest=sha256:1a5b10a83007a3cf0df435f655381e7aacf2103429af80671cf519de95b87818

Observation 07b0436b-7af7-4f2d-b463-bff89e3a65bc · outbound

This paper cites nuScenes: A mul- timodal dataset for autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency nuScenes: A mul- timodal dataset for autonomous driving,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.938197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.938197Z digest=sha256:6a462150fe465b3b385cda728dbde2a31bfd78f592965bd6baa44523f6c732f2

Observation d66212bc-28ca-4069-88b0-d919a9919170 · outbound

This paper cites BART: Denoising sequence- to-sequence pre-training for natural lan- guage generation, translation, and com- prehension,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency BART: Denoising sequence- to-sequence pre-training for natural lan- guage generation, translation, and com- prehension,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:26.045042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:26.045042Z digest=sha256:b0e6498446400fbb22f4ad06d913be5512793c62802f11ffd9e4de5a6f54f296

Observation 5f648b4d-208e-469f-bbe9-b5e6a4e582f2 · outbound

This paper cites Language-augmented Bird’s-eye View Maps for autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Language-augmented Bird’s-eye View Maps for autonomous driving,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:26.182936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:26.182936Z digest=sha256:cf315b28eb82bf65fbf5a02ff4cce23e88d30ac7651f0817a9fb833a1fceac9f

Observation 059c3613-b99e-4c49-a31e-bcee2249e905 · outbound

This paper cites Vi- sual question answering and natural lan- guage explanations for autonomous driv- ing,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Vi- sual question answering and natural lan- guage explanations for autonomous driv- ing,

Reference 18

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T18:33:27.076617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:33:26.293828Z digest=sha256:09feb0212c4dd339d18f8e55a9cc71bc1b4ca65d11945c46e0845dcc29d6c5ff

Observation 0ad41744-5391-44a0-8385-d63e9f736a87 · outbound

This paper cites Antagonising explanation and revealing bias directly through sequencing and multimodal inference.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Antagonising explanation and revealing bias directly through sequencing and multimodal inference

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:33:26.829568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:33:26.341130Z digest=sha256:30917ecb39bdb187c87fb0259d0558b2f00a3fc699b181adf0eb3ab2f422e1e3

Observation b7c750e8-da9a-4a73-90ee-405e352b7312 · outbound

This paper cites an unresolved cited work.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:33:28.597555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:33:26.410084Z digest=sha256:f5c29ba25335a175ab1efdf397caa1c50e44f983f8194d0fdf810327d420f375

Observation 6f7ac32a-f619-4442-8762-8e0c5f5a6ff2 · outbound

This paper cites ROUGE: A package for automatic evaluation of summaries,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency ROUGE: A package for automatic evaluation of summaries,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:33:28.484960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:33:26.474538Z digest=sha256:12cb513758fa2379b0088ba26672fc92d471ff5631da5860dc08801233c759bd

Observation 1a763c5f-6605-4540-8378-1459416a7ed9 · outbound

This paper cites METEOR: An automatic metric for MT evaluation with im- proved correlation with human judgments,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency METEOR: An automatic metric for MT evaluation with im- proved correlation with human judgments,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:33:28.231684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:33:26.603654Z digest=sha256:050b9f2ba19aac8895f718ea5ce64428fcfdf1515b13bdf42681452897bbbec0

Pith citing papers

No inbound Pith citation observations are available.