Pith. sign in

Paper Citation Record · LEDGER

MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2311.07689.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.07689 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T17:02:47.226707Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0384d0bf-208d-4fad-b987-297f8324a428 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.631750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:1fa18585053d4075566ed382bbb232d73e45d63160c0e9f661614e21fcc5bc5c

Observation accb2930-5513-4b33-85bf-6e8578db0435 · inbound

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions cites this paper.

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 237

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:55:50.429288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T21:54:26.670284Z digest=sha256:004f13307c5c38750edbc73238f612c9b11b44897472ac30b9a1d82d71c3ffde

Observation 4f1f5435-9257-4719-b770-3521efb562b6 · inbound

Jailbreaking to Jailbreak cites this paper.

Jailbreaking to Jailbreak MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T17:02:47.226707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:02:47.226707Z digest=sha256:c15c345acd3fb9f77df1d7d8062b1cac65d6e6a2b116a3d17502dd2461bf3009

Observation 438fc101-1d13-405f-bbe4-fda065063267 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 194

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:27.191626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:27.191626Z digest=sha256:c0715de6ab02d216d5948d494687b6c92320b1f6e02f306e9147c272c95078f2

Observation 308cfed0-9064-49df-be80-4ccaa8235d0b · inbound

Addressing Bias in LLMs: Strategies and Application to Fair AI-based Recruitment cites this paper.

Addressing Bias in LLMs: Strategies and Application to Fair AI-based Recruitment MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T01:09:01.473343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:09:01.473343Z digest=sha256:077ed874bb5e9fb79b479454615e9c71cad6dfa359f663de7f0af6979bb47a45

Observation 1d795c97-49bb-4593-8c82-ed9875ae1870 · inbound

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety cites this paper.

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:00.457694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:00.457694Z digest=sha256:9af239aaba29da14416d01ff36041a65d7692234d5b57aa8b3ad47ce35eaa530

Observation 7cecb0e4-914f-4a43-a672-5e06b0012366 · inbound

MGC: A Compiler Framework Exploiting Compositional Blindness in Aligned LLMs for Malware Generation cites this paper.

MGC: A Compiler Framework Exploiting Compositional Blindness in Aligned LLMs for Malware Generation MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:45:53.270366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:45:53.270366Z digest=sha256:d383d1c4cbcf57aa3fe63868294dbc8063456b130b5cb7365e5708a9d14b2867

Observation f3f047c8-b8c2-4675-afe1-33facb0f8ab6 · inbound

SEALGuard: Safeguarding the Multilingual Conversations in Southeast Asian Languages for LLM Software Systems cites this paper.

SEALGuard: Safeguarding the Multilingual Conversations in Southeast Asian Languages for LLM Software Systems MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:42.583296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:28:42.583296Z digest=sha256:23146b1ae62c0d6b5a53344c31e833ad5791c828055eb41f64ecdee23f5de439

Observation 794e5e78-c718-48ad-a1c2-d83518a3af0d · inbound

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers cites this paper.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.312168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:50.312168Z digest=sha256:53ceb23f91db7fcccfeb3adcff1a264ba0a6884dc3827fdcac0c65f15b422a33

Observation eb6e87cc-a43e-4207-88c8-8f34e930529b · inbound

Agentic Web: Weaving the Next Web with AI Agents cites this paper.

Agentic Web: Weaving the Next Web with AI Agents MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:33.631306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T13:05:33.631306Z digest=sha256:e2fecbba718c85e83e4da02966e0bbc79e2572c97a0cfa94d9ee2407fe2a1e8c

Observation 48e07090-f92e-460f-b050-f2a98aa1cd0d · inbound

RedCoder: Automated Multi-Turn Red Teaming for Code LLMs cites this paper.

RedCoder: Automated Multi-Turn Red Teaming for Code LLMs MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:44.699105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:00:44.699105Z digest=sha256:2c590b1eb3cb484b63b232011cc23d505ae11f2d7e05951d98ab58bb727b71ff

Observation 7b4a7287-a7aa-4929-b01a-c995f4b45c8b · inbound

Paladin: Defending LLM-enabled Phishing Emails with a New Trigger-Tag Paradigm cites this paper.

Paladin: Defending LLM-enabled Phishing Emails with a New Trigger-Tag Paradigm MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T22:33:21.814669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:33:21.814669Z digest=sha256:989714fa2009eef05b0848ffdd27f2e90a40e1f49f76702b80bb4f32ce18050e

Observation deaa8e17-b61c-41e6-a21e-e8e8b54a9ce0 · inbound

When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models cites this paper.

When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:31:11.613656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T09:29:14.842228Z digest=sha256:a93c9dc7ab3ae13dbad0b8da0b16cea56bf073724a84df8e6e425e4d9dd2e940

Observation c65ad069-e119-4b26-8c17-bc5fd43554ae · inbound

Reasoning Structure Matters for Safety Alignment of Reasoning Models cites this paper.

Reasoning Structure Matters for Safety Alignment of Reasoning Models MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:56:04.094447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T02:36:57.093584Z digest=sha256:d198d04e6c856cab52156bc5091ec9cbef43b85b71f826517845d8db288fcf88

Observation a1816671-9d2b-4b1f-9103-cf72de8f98e5 · inbound

Adaptive Instruction Composition for Automated LLM Red-Teaming cites this paper.

Adaptive Instruction Composition for Automated LLM Red-Teaming MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:06:03.758073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T23:34:36.280092Z digest=sha256:6bb357897e934a407090f4e135a9189785bae8f3a8968087466249fa32a70c90

Observation b9b35e52-538f-496d-ae95-ae79ffdf4ba9 · inbound

Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models cites this paper.

Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:41:31.607404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T21:12:39.195573Z digest=sha256:fdcc11414dbde96411769278d102fce29eb1a925a613640d19cb3a1cf1351726

Observation 9b07d03e-0cd4-4ea3-92ff-e1e8d009ff59 · inbound

A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts cites this paper.

A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:45:39.712159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T18:11:29.066362Z digest=sha256:1e5f76e6acdef4a0891495a14d323a46f9ef57051de25a514d96c6766ea6b454

Observation 124d926e-fc74-4822-aaef-26572f79a101 · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:50:11.171050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-25T20:58:53.119386Z digest=sha256:910bcb2b417054d9467dad90f8b6ff2fda68e44dff7fe797cd06dd51900af2a6

Observation f76f131f-7dd2-4942-8d8e-14168e587f8a · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T10:16:36.929637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:16:36.929637Z digest=sha256:977746e838719c9dce5da04e02242d4184601b38109e222d92b2e7eab8a564d0

Observation a0666eaf-1d3e-4a28-a74c-712c4e336e3c · inbound

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks cites this paper.

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:49:52.898247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T05:51:01.446139Z digest=sha256:2c52723feb20524849aa94ed14f74f2bdb381172445bf3b56a46e5e98280b07d

Observation 64677a4b-47eb-40e2-af4b-f333ce154e4c · inbound

GPT-Red: Automated Red Teaming via Self-Play at Scale cites this paper.

GPT-Red: Automated Red Teaming via Self-Play at Scale MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T01:12:44.140186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:12:44.140186Z digest=sha256:ee7e88ba3a918e0c33efb8b0af778b9d9b8a16c61082116fbb6d45f84688edd3

Observation edb960e9-8d02-478f-ab19-7cf4cadd7b88 · inbound

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation cites this paper.

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T00:29:56.778717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:29:56.778717Z digest=sha256:b69ea68787b61776616919770196ced877e7f97a93084c31486d0af530e4fb7f