Pith. sign in

REVIEW 2 cited by

An Exploratory Study of AI System Risk Assessment from the Lens of Data Distribution and Uncertainty

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2212.06828 v1 pith:VOEEPT4P submitted 2022-12-13 cs.LG cs.AIcs.SE

classification cs.LGcs.AIcs.SE
keywords levelsystemassessmentexploratoryrisksystemsapplicationsbeen
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deep learning (DL) has become a driving force and has been widely adopted in many domains and applications with competitive performance. In practice, to solve the nontrivial and complicated tasks in real-world applications, DL is often not used standalone, but instead contributes as a piece of gadget of a larger complex AI system. Although there comes a fast increasing trend to study the quality issues of deep neural networks (DNNs) at the model level, few studies have been performed to investigate the quality of DNNs at both the unit level and the potential impacts on the system level. More importantly, it also lacks systematic investigation on how to perform the risk assessment for AI systems from unit level to system level. To bridge this gap, this paper initiates an early exploratory study of AI system risk assessment from both the data distribution and uncertainty angles to address these issues. We propose a general framework with an exploratory study for analyzing AI systems. After large-scale (700+ experimental configurations and 5000+ GPU hours) experiments and in-depth investigations, we reached a few key interesting findings that highlight the practical need and opportunities for more in-depth investigations into AI systems.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. MoDitector: Module-Directed Testing for Autonomous Driving Systems

    cs.SE 2025-02 conditional novelty 6.0 of 10

    MoDitector generates collision scenarios that are caused by errors in a user-specified ADS module, reporting 55.3, 75.3, 71.7, and 14.3 module-induced critical scenarios for perception, prediction, planning, and contr...

  2. The "I Don't Know" Filter: Enhancing Agentic Reliability in Function Calling

    cs.SE 2026-07 conditional novelty 5.5 of 10

    A multi-sample uncertainty classifier filter improves agent reliability (IDKS) by abstaining on uncertain function calls across open SLMs and BFCL-style benchmarks.

Pith tools