REVIEW 4 cited by
Real-time Driver Monitoring Systems on Edge AI Device
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
As road accident cases are increasing due to the inattention of the driver, automated driver monitoring systems (DMS) have gained an increase in acceptance. In this report, we present a real-time DMS system that runs on a hardware-accelerator-based edge device. The system consists of an InfraRed camera to record the driver footage and an edge device to process the data. To successfully port the deep learning models to run on the edge device taking full advantage of the hardware accelerators, model surgery was performed. The final DMS system achieves 63 frames per second (FPS) on the TI-TDA4VM edge device.
Forward citations
Cited by 4 Pith papers
-
HeteroMosaic: Exposing and Exploiting Heterogeneous Execution Opportunities for Energy-Efficient Edge LLM Inference
HeteroMosaic uses micro-batching and trace-guided co-optimization to split edge LLM prefill across iGPU and NPU, achieving up to 1.73-2.05x speedups and 45.3% energy reduction on AMD Ryzen AI.
-
TileFuse: A Fused Mixed-Precision Kernel Library for Efficient Quantized LLM Inference on AMD NPUs
TileFuse introduces fused kernels and data layouts for W4A16/W8A16 on AMD XDNA2 NPUs, reporting up to 2.0x lower LLM prefilling latency and 64.6% lower energy versus baselines.
-
HeteroMosaic: Exposing and Exploiting Heterogeneous Execution Opportunities for Energy-Efficient Edge LLM Inference
HeteroMosaic co-schedules edge LLM inference across iGPU and NPU via roofline analysis and micro-batches, claiming up to ~2× speedup and ~45% energy reduction on AMD Ryzen AI SoCs.
-
TileFuse: A Fused Mixed-Precision Kernel Library for Efficient Quantized LLM Inference on AMD NPUs
TileFuse introduces a fused kernel library enabling AWQ W4A16/W8A16 quantized LLM inference on AMD NPUs, reporting up to 2.0x lower prefilling latency and 64.6% lower energy on Ryzen AI laptops.
Discussion (0). Sign in to comment.