Pith. sign in

REVIEW 2 cited by

Exploiting Code Symmetries for Learning Program Semantics

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2308.03312 v9 pith:4ODVX2VT submitted 2023-08-07 cs.LG cs.CRcs.PL

classification cs.LGcs.CRcs.PL
keywords codeprogramsymmetriesgroupsemanticsanalysisllmsmodels
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

This paper tackles the challenge of teaching code semantics to Large Language Models (LLMs) for program analysis by incorporating code symmetries into the model architecture. We introduce a group-theoretic framework that defines code symmetries as semantics-preserving transformations, where forming a code symmetry group enables precise and efficient reasoning of code semantics. Our solution, SymC, develops a novel variant of self-attention that is provably equivariant to code symmetries from the permutation group defined over the program dependence graph. SymC obtains superior performance on five program analysis tasks, outperforming state-of-the-art code models without any pre-training. Our results suggest that code LLMs that encode the code structural prior via the code symmetry group generalize better and faster.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Metamorphic Testing of Deep Code Models: A Systematic Literature Review

    cs.SE 2025-07 conditional novelty 5.0 of 10

    A systematic review of 45 papers shows metamorphic testing of code models relies mostly on identifier renaming and dead code insertion, targets encoder-only models like CodeBERT, and under-covers generative tasks, new...

  2. Large Language Models in Cybersecurity: Applications, Vulnerabilities, and Defense Techniques

    cs.CR 2025-07 conditional novelty 4.0 of 10

    A survey that maps LLM applications, vulnerabilities, and defenses across eight cybersecurity domains, but with significant citation and rigor problems.

Pith tools