REVIEW 4 cited by
Benchmarking Deep Learning Models on NVIDIA Jetson Nano for Real-Time Systems: An Empirical Investigation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The proliferation of complex deep learning (DL) models has revolutionized various applications, including computer vision-based solutions, prompting their integration into real-time systems. However, the resource-intensive nature of these models poses challenges for deployment on low-computational power and low-memory devices, like embedded and edge devices. This work empirically investigates the optimization of such complex DL models to analyze their functionality on an embedded device, particularly on the NVIDIA Jetson Nano. It evaluates the effectiveness of the optimized models in terms of their inference speed for image classification and video action detection. The experimental results reveal that, on average, optimized models exhibit a 16.11% speed improvement over their non-optimized counterparts. This not only emphasizes the critical need to consider hardware constraints and environmental sustainability in model development and deployment but also underscores the pivotal role of model optimization in enabling the widespread deployment of AI-assisted technologies on resource-constrained computational systems. It also serves as proof that prioritizing hardware-specific model optimization leads to efficient and scalable solutions that substantially decrease energy consumption and carbon footprint.
Forward citations
Cited by 4 Pith papers
-
TS-MAMP: A Remanufactured Agricultural Robot Powered by Second-Life EV Components and NMS-Free On-Device Weed Detection
A remanufactured agricultural robot using retired EV powertrains and an edge-deployed NMS-free YOLOv10n detector achieves below-$450 drivetrain/chassis cost and 80.87% mAP@0.5 on the new Wanxi Crop-Weed dataset.
-
RobotxR1: Enabling Embodied Robotic Intelligence on Large Language Models through Closed-Loop Reinforcement Learning
Combining supervised fine-tuning with closed-loop RL lets small Qwen LLMs tune an MPC controller, and the 3B model scores 63.3% versus 58.5% for GPT-4o on the paper's custom control adaptability metric.
-
Large Language Models on Small Resource-Constrained Systems: Performance Characterization, Analysis and Trade-offs
A measurement study characterizing LLM inference latency, power, memory, and energy on Jetson Orin devices across model sizes, power modes, and quantization, with a public testing utility.
-
FEDEXCHANGE: Bridging the Domain Gap in Federated Object Detection for Free
FEDEXCHANGE improves cross-domain federated object detection by server-side clustering and exchanging client decoder models, achieving higher mAP in some domains at no extra local compute.
Discussion (0). Continue with ORCID to comment.