Pith. sign in

REVIEW 5 cited by

On the Impossible Safety of Large AI Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2209.15259 v2 pith:FQSZTPOQ submitted 2022-09-30 cs.LG cs.AIcs.CR

classification cs.LGcs.AIcs.CR
keywords largemodelslaimslearningmachinesecurityaccuracyaccurate
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Large AI Models (LAIMs), of which large language models are the most prominent recent example, showcase some impressive performance. However they have been empirically found to pose serious security issues. This paper systematizes our knowledge about the fundamental impossibility of building arbitrarily accurate and secure machine learning models. More precisely, we identify key challenging features of many of today's machine learning settings. Namely, high accuracy seems to require memorizing large training datasets, which are often user-generated and highly heterogeneous, with both sensitive information and fake users. We then survey statistical lower bounds that, we argue, constitute a compelling case against the possibility of designing high-accuracy LAIMs with strong security guarantees.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. More Edits, More Stable: Understanding the Lifelong Normalization in Sequential Model Editing

    cs.LG 2026-05 unverdicted novelty 7.0 of 10

    Online value-gradient normalization in lifelong LLM editing produces bounded, asymptotically orthogonal parameter updates; an explicit warm-up and full whitening (StableEdit) strengthen this effect and improve long-ho...

  2. A Case for Specialisation in Non-Human Entities

    cs.CY 2025-02 conditional novelty 6.0 of 10

    A position paper making the case that specialised, well-specified AI systems are more robust, secure, and governable than general-purpose AGI systems, and that hard-to-specify tasks need specified governance.

  3. Reality Check: A New Evaluation Ecosystem Is Necessary to Understand AI's Real World Effects

    cs.CY 2025-05 conditional novelty 4.0 of 10

    A position paper argues that understanding AI's second-order effects requires moving from static benchmarks to an ecosystem of field testing, red teaming, and contextual evaluation.

  4. `Do as I say not as I do': A Semi-Automated Approach for Jailbreak Prompt Attack against Multimodal LLMs

    cs.CR 2025-02 reject novelty 4.0 of 10

    A voice jailbreak that buries a forbidden question between benign prompts reportedly succeeds against Gemini 67 to 93 percent of the time, but the metric comes from the target model judging itself and is not reliable.

  5. AI Governance through Markets

    econ.GN 2025-01 conditional novelty 4.0 of 10

    Market governance mechanisms, supported by standardized AI disclosures, can create financial incentives for responsible AI development, according to this policy paper.

Pith tools