Pith. sign in

Paper Citation Record · LEDGER

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency

As of 10 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2507.07938.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07938 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:33:26.603654Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved13
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 20539d96-2825-4042-a2bb-3fb6e3d6b48d · outbound

This paper cites A survey of autonomous driving: Common practices and emerging technolo- gies,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency A survey of autonomous driving: Common practices and emerging technolo- gies,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:24.489572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:24.489572Z digest=sha256:405254763f8f857ec239753f264f3c00ee6e54bc6d284e2bd636f15500da9f71

Observation 8520d665-ff66-47ee-8dbd-4e72f7cb95e7 · outbound

This paper cites ClusT3: Information Invariant Test-Time Training.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency ClusT3: Information Invariant Test-Time Training

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:24.607371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:24.607371Z digest=sha256:a37a7e995b6275b2155bfeddd6562e0118f3ef83365be58a806ec4de12bae9e9

Observation f82fddac-1ecb-415b-9d97-1fe8922bc95b · outbound

This paper cites Planning-oriented autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Planning-oriented autonomous driving,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:24.769558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:24.769558Z digest=sha256:eca209304ecec3821ab017156a23d831b8fb37df651cdced21103766dc26e243

Observation 6c117f6c-590a-4b39-85ec-4e8a6c0c984d · outbound

This paper cites Deep multi-modal object detection and semantic segmentation for au- tonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Deep multi-modal object detection and semantic segmentation for au- tonomous driving,

Reference 4

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T18:33:27.877104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:33:24.896397Z digest=sha256:f3fc040d4532f3d233729a8717c847e06ec58ac167b23b7a527b3799740c2e66

Observation 7e43c368-c3a3-400b-b3cd-23fdfd13c3fc · outbound

This paper cites Ex- plainable artificial intelligence (XAI),.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Ex- plainable artificial intelligence (XAI),

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.053146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.053146Z digest=sha256:ca758604b4e3fbbcd02f65b596804b87611e53e4356765f158a052e16cb0d483

Observation 2a9fe5ed-4484-4588-930e-de14a4d03178 · outbound

This paper cites Why did the AI make that decision?,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Why did the AI make that decision?,

Reference 6

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T18:33:27.596507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:33:25.148590Z digest=sha256:16500a5fdd10be5c6e6ec395504cfa4229c3767632246e4e1b5a146f02c4cd71

Observation 71590303-b49f-48f7-9bed-69c9d9e00b4d · outbound

This paper cites Ex- plainable artificial intelligence (XAI),.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Ex- plainable artificial intelligence (XAI),

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.251827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.251827Z digest=sha256:eb5a2ffadf317c2881e2cc69669e349f5b04a11aa278cb20e4281e5478aa1604

Observation 45f9e0cd-7f03-410d-b6de-9df6a1259202 · outbound

This paper cites Interpretable autonomous driving: A sur- vey of recent advances,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Interpretable autonomous driving: A sur- vey of recent advances,

Reference 8

Resolution
malformed identifier
no resolver link, observed 2026-08-06T18:33:25.417228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.417228Z digest=sha256:d8abc1154161d9d840e97d0dad96461e681d28450fd4b9a54771f5b8fbedac73

Observation 63ae09ef-8a4d-4c8b-af86-25bb78cb87e8 · outbound

This paper cites Attention- based multimodal framework for au- tonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Attention- based multimodal framework for au- tonomous driving,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.485181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.485181Z digest=sha256:1e8cca03cf0b838621567e8dce80d377e926e7b118c43479baddd6428a7d3c0a

Observation bbbe183c-1677-42ea-b03b-7709cd130195 · outbound

This paper cites DeepDriving: Learning affordance for direct perception in autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency DeepDriving: Learning affordance for direct perception in autonomous driving,

Reference 10

Resolution
malformed identifier
no resolver link, observed 2026-08-06T18:33:25.565184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.565184Z digest=sha256:8a2a6f45748038f6bb69f9ee0a393d87c19d4f8485e3d94b716aa65442dbeff5

Observation fdfd001c-ed76-4d3c-8e29-360eeecfa416 · outbound

This paper cites CARLA: An Open Urban Driving Simulator.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency CARLA: An Open Urban Driving Simulator

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.619982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.619982Z digest=sha256:1bd62592ab8bfdc9af6119ae15892ef6167c623c9c1c330ff70a7cc6ef1b7f14

Observation 13d5428e-6ba5-44f3-9677-ad9af44217c8 · outbound

This paper cites VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.677404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.677404Z digest=sha256:f031db21f8af3eef5a2633448624fa708ac4d97018484f8dad691525d11e9981

Observation 62b49ade-ec11-4b84-97cf-ba2a095ad09f · outbound

This paper cites BERT: Pre-training of deep bidi- rectional transformers for language under- standing,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency BERT: Pre-training of deep bidi- rectional transformers for language under- standing,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.751970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.751970Z digest=sha256:bd82c429d2df16678b414d09320ea4f9b6e39271c7606a6075fce6232918f120

Observation 07b0436b-7af7-4f2d-b463-bff89e3a65bc · outbound

This paper cites nuScenes: A mul- timodal dataset for autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency nuScenes: A mul- timodal dataset for autonomous driving,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.938197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.938197Z digest=sha256:f6f4f3e7a76d79d048b28f7fe70243ecf4f58660425a13ab7230009980aa6e44

Observation d66212bc-28ca-4069-88b0-d919a9919170 · outbound

This paper cites BART: Denoising sequence- to-sequence pre-training for natural lan- guage generation, translation, and com- prehension,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency BART: Denoising sequence- to-sequence pre-training for natural lan- guage generation, translation, and com- prehension,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:26.045042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:26.045042Z digest=sha256:8405a2e42c5773afb751fde4a2fd47c288d8af1474657c98492dd4e49062be48

Observation 5f648b4d-208e-469f-bbe9-b5e6a4e582f2 · outbound

This paper cites Language-augmented Bird’s-eye View Maps for autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Language-augmented Bird’s-eye View Maps for autonomous driving,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:26.182936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:26.182936Z digest=sha256:2bbf60bce694aff50018ebc2970d15c8f0b5dd408d7cc26885b3a6e1aa60ba58

Observation 059c3613-b99e-4c49-a31e-bcee2249e905 · outbound

This paper cites Vi- sual question answering and natural lan- guage explanations for autonomous driv- ing,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Vi- sual question answering and natural lan- guage explanations for autonomous driv- ing,

Reference 18

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T18:33:27.076617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:33:26.293828Z digest=sha256:7a9bf060c051f585aa5210d9bc1677e908c5628ba17d58bb558ec070bcc2bfc4

Observation 0ad41744-5391-44a0-8385-d63e9f736a87 · outbound

This paper cites Antagonising explanation and revealing bias directly through sequencing and multimodal inference.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Antagonising explanation and revealing bias directly through sequencing and multimodal inference

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:33:26.829568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:33:26.341130Z digest=sha256:521d644c67bf9222877c790e69119d94709cb7f65621173ee6022c461d5fe421

Observation b7c750e8-da9a-4a73-90ee-405e352b7312 · outbound

This paper cites an unresolved cited work.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:33:28.597555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:33:26.410084Z digest=sha256:1306b1cee0dba57fb4b0f6e09185da0929dd1af5355f07b58d3479a296acf27d

Observation 6f7ac32a-f619-4442-8762-8e0c5f5a6ff2 · outbound

This paper cites ROUGE: A package for automatic evaluation of summaries,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency ROUGE: A package for automatic evaluation of summaries,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:33:28.484960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:33:26.474538Z digest=sha256:97b70c0f4fa6ec8e31fb1c699b19bc45209075a17b0189c0b5286e91d9c6f270

Observation 1a763c5f-6605-4540-8378-1459416a7ed9 · outbound

This paper cites METEOR: An automatic metric for MT evaluation with im- proved correlation with human judgments,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency METEOR: An automatic metric for MT evaluation with im- proved correlation with human judgments,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:33:28.231684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T18:33:26.603654Z digest=sha256:1a4b8738bf5dd67cc15f88c2cafa991e2015831ceca0c7254c441169f2342774

Pith citing papers

No inbound Pith citation observations are available.