Pith. sign in

Paper Citation Record · LEDGER

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code

As of 17 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 1 inbound Pith citation observation for arXiv:2509.07006.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.07006 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:26:53.128079Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-11T02:25:54.160882Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T03:40:54.567829Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact12
  • verified fuzzy10
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fa249d92-4588-4500-9d48-7da9d67f5c93 · outbound

This paper cites an unresolved cited work.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Unresolved cited work

Reference 1

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.402666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.848031Z digest=sha256:f81772149d7fb265e03d1009b66b1f016e560bb31f706551a0fe67b2a263a795

Observation ee3b24b1-85f2-4f71-bce5-fd0b8c4bb81d · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Constitutional AI: Harmlessness from AI Feedback

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.854837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.854837Z digest=sha256:97056fcfedd8001420bba996b9e4f1755640c2e949d77baada7f2fc1f207ace6

Observation 5dd75a43-3384-41bf-93bc-c88949164e74 · outbound

This paper cites FeedbackLogs: Recording and Incorporating Stakeholder Feedback into Machine Learning Pipelines.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code FeedbackLogs: Recording and Incorporating Stakeholder Feedback into Machine Learning Pipelines

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.861299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.861299Z digest=sha256:65a01eb036a392b8abea6f64dbdb3bfafc12a411ce442d1cd8bc071fe65ccba3

Observation f4710fcb-8f86-43d5-9753-9eb79054e862 · outbound

This paper cites Modelling moral reasoning and ethical responsibility with logic programming.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Modelling moral reasoning and ethical responsibility with logic programming

Reference 4

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.379609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.866555Z digest=sha256:f6b1e2c3f3d9dac0d109a733ca7056612a0e5406a46406096672f9c488d79bd6

Observation 9b074c3b-e88a-44b2-91a6-923c6b9fd60f · outbound

This paper cites When Should Algorithms Resign? A Proposal for AI Governance.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code When Should Algorithms Resign? A Proposal for AI Governance

Reference 5

Resolution
metadata mismatch
raw_fallback, observed 2026-08-15T16:26:54.288168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.873492Z digest=sha256:8e5385f34e7d49c1cfb11557b445189cada22de495b4dc8647445ba05ac64374

Observation 4634952b-d99c-49e9-8d58-be547e70cea8 · outbound

This paper cites an unresolved cited work.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.879008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.879008Z digest=sha256:b228e227ee29c891c40931b66cf3aeef1bbbc6671d3187054c0b78e815c88a55

Observation 27ebfc6d-0a57-422b-afbf-31ccba477b2b · outbound

This paper cites Vera Liao, Prasanna Sattigeri, Riccardo Fogliato, Gabrielle Melnikov, Ranganath Krishnan, Jason Stanley, Omesh Tickoo, Lior Nachman, Adrian Cheng, and Kush R.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Vera Liao, Prasanna Sattigeri, Riccardo Fogliato, Gabrielle Melnikov, Ranganath Krishnan, Jason Stanley, Omesh Tickoo, Lior Nachman, Adrian Cheng, and Kush R

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.886800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.886800Z digest=sha256:a4e4974d5d1c6b63f9135f7eb49d74990adbcb21dd3f74c1e6c74db2e436d2ca

Observation 50a5a9e7-0dad-4833-a37d-1a24d3103977 · outbound

This paper cites Superintelligence: Paths, Dangers, Strategies.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Superintelligence: Paths, Dangers, Strategies

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:26:54.574130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.898956Z digest=sha256:0d78c8a613f563ff3de58c63b8f836872fade06b05acf069f0bb15ff5cfb2af5

Observation 76001f45-0b05-4cfa-b021-6319504e43cb · outbound

This paper cites Harms from Increasingly Agentic Algorithmic Systems.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Harms from Increasingly Agentic Algorithmic Systems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.904380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.904380Z digest=sha256:3166e301402e8425a39fd3d9e25950c104ab5e74cb8382bf25422e9a19f8e309

Observation 2a3d0433-460d-48c7-a47f-6c05cfdda440 · outbound

This paper cites Confucian Ethics and AI: Towards Harmonious Human-Machine Interaction.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Confucian Ethics and AI: Towards Harmonious Human-Machine Interaction

Reference 10

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.360584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.909399Z digest=sha256:195175f19c362465d1ee0fe0357f04a090fcb0155475e045f878181be94e60b3

Observation 3bf52532-aed5-4752-bc30-a077d466382c · outbound

This paper cites Deep reinforcement learning from human preferences.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Deep reinforcement learning from human preferences

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.914337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.914337Z digest=sha256:991ae6332b47fdba223718601ef855cdd38d0e893021d175bb7d51614078a9ce

Observation a5f7ced4-a4d3-4e3d-8c11-48759187c401 · outbound

This paper cites Achieving EU AI Act Compliance by Integrating Governance as Code (GaC) and Machine Learning Operations (MLOps).

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Achieving EU AI Act Compliance by Integrating Governance as Code (GaC) and Machine Learning Operations (MLOps)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:26:54.558537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.919504Z digest=sha256:afa45e2bf50406b5f1fa3cb33d22bbdf66d0eeddc3a2fc0c3673a4d5a86bc67b

Observation 53599dfe-1688-4d51-a0ba-da9f713a2a55 · outbound

This paper cites Collins, Ilia Sucholutsky, Umang Bhatt, Adrian Weller, Thomas L.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Collins, Ilia Sucholutsky, Umang Bhatt, Adrian Weller, Thomas L

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.924465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.924465Z digest=sha256:283bd9d33c803393614b23c80778e3794631898319cea6c489db1be7f2bc22dc

Observation da7549c6-f490-45e1-97b7-79b27da8778f · outbound

This paper cites Process Reinforcement through Implicit Rewards.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Process Reinforcement through Implicit Rewards

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.929255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.929255Z digest=sha256:74272bea7fc05c8ba734cc3c96202e0268f54f9ff0124d2e0723344e099e78de

Observation 65cc20c8-df71-4118-a21b-3cfdb5aeeff8 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.934071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.934071Z digest=sha256:80cc3784a4bdb2082a04b16737c9e0c3908c4bd9d61120db10cb0a193e286ca6

Observation 3d8c81c6-47db-424d-86f4-7710bdbcbe2b · outbound

This paper cites Dennis, Michael Fisher, Marija Slavkovik, and Matt P.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Dennis, Michael Fisher, Marija Slavkovik, and Matt P

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.938613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.938613Z digest=sha256:a22dedef97b1dc3759426ee132e687e69b4a5c0222dcd44f2247bc8bf004d4de

Observation 8d01316c-08d5-4570-8588-65f276f0ff1c · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.943079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.943079Z digest=sha256:825c116053a852cb7dec87c5b2dcbc14b73a92a751ab244b9d6d1a19a5345295

Observation 81014b0b-6e4e-45c7-8b0e-1772d12fe360 · outbound

This paper cites Ubuntu and Artificial Intelligence: Towards an African Ethical Framework.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Ubuntu and Artificial Intelligence: Towards an African Ethical Framework

Reference 18

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.315826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.947833Z digest=sha256:b81572e3bd413485205660ba6eaea126402f7af2ffdeab957fa4e7973924f036

Observation 8841a5c9-780d-42b7-8cbf-7897e6162dc7 · outbound

This paper cites Buddhist Ethics and AI: Compassion-Based Approaches to Artificial Intelligence.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Buddhist Ethics and AI: Compassion-Based Approaches to Artificial Intelligence

Reference 19

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.297447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.952640Z digest=sha256:b104ab9176eb68b256bb778961cb0f2aea3303e9000fdc76a302ca9430624f03

Observation 85d26e96-f27d-4929-a946-9dcdd74c21c2 · outbound

This paper cites Buen Vivir: Today's Tomorrow.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Buen Vivir: Today's Tomorrow

Reference 20

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.279919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.957988Z digest=sha256:83f789cf9dbf3dbe9b45a10da0346953b14053d85953bffb672b0440b9f7e093

Observation b9a7e6de-b239-475b-8458-3a85bdabf65b · outbound

This paper cites Introduction to AI Safety, Ethics, and Society.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Introduction to AI Safety, Ethics, and Society

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:26:54.542396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.963236Z digest=sha256:8bc69ab465023ddf318dc4a9c63d10d6f16ce1258103951e056eb574a3d90700

Observation 099095ce-d0ee-4835-ab59-06ea1161a122 · outbound

This paper cites Towards interactive evaluations for interaction harms in human-AI systems.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Towards interactive evaluations for interaction harms in human-AI systems

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.967518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.967518Z digest=sha256:bf4d9141cd635af34af9f74fc50016ca4f453fa45d42e32454379e7fe29af880

Observation 1a1e38cd-4044-4975-8536-8cc3de103757 · outbound

This paper cites Gordon, Caglar Gulcehre, Dongyeop Kang, Maarten Sap, Amy Zhang, and He He.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Gordon, Caglar Gulcehre, Dongyeop Kang, Maarten Sap, Amy Zhang, and He He

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:26:54.525122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.973342Z digest=sha256:723bec6294182519015deffa0461a697bd257cb033259e01e6882c0d735519fd

Observation 58dca08c-85ea-4e18-abab-d3727fca59e7 · outbound

This paper cites High spin axion insulator.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code High spin axion insulator

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.978011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.978011Z digest=sha256:73a736a553e6f1379714f8984c564d67aaf6806d3e8df71303fc4c30fbc4e505

Observation 38e519cd-c847-4145-9cf5-935b972e2d39 · outbound

This paper cites The Road to Artificial SuperIntelligence: A Comprehensive Survey of Superalignment.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code The Road to Artificial SuperIntelligence: A Comprehensive Survey of Superalignment

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:52.982927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:52.982927Z digest=sha256:c2f69a4ac16eca3be2702cf3a23b3ded91d4a327e02357568bb4b15d1bad62a5

Observation 4171fc9f-ad2e-4430-992e-a19bb54d166e · outbound

This paper cites Confucian Values in AI Development: A Framework for Ethical Technology.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Confucian Values in AI Development: A Framework for Ethical Technology

Reference 26

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.260399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.989197Z digest=sha256:986055d2c2a4935634123cc07a7d97a3ae602cc131570e697913de31900edc85

Observation 4abdc95f-7c89-4e08-ad0a-8c235a3b69fd · outbound

This paper cites Topological eigenvalues braiding and quantum state transfer near a third-order exceptional point.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Topological eigenvalues braiding and quantum state transfer near a third-order exceptional point

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T16:26:53.879887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:52.994667Z digest=sha256:8bf833c88c3464e9baca46ff25a363621e570f0dc30d8a6e186c5d3972362e03

Observation 46604d68-be3d-4041-9bc1-a805f513d267 · outbound

This paper cites Monte Carlo Tree Search for Comprehensive Exploration in LLM-Based Automatic Heuristic Design.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Monte Carlo Tree Search for Comprehensive Exploration in LLM-Based Automatic Heuristic Design

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.000305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.000305Z digest=sha256:0bb65cdb48dffe223bb220d69f5b6c411ae801c33671d385667ba1d9746cab37

Observation 28c0d855-d4d7-4bfe-8fdd-1f0195ece22c · outbound

This paper cites From Rationality to Relationality: Ubuntu as an Ethical and Human Rights Framework for Artificial Intelligence Governance.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code From Rationality to Relationality: Ubuntu as an Ethical and Human Rights Framework for Artificial Intelligence Governance

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:26:54.508883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.006468Z digest=sha256:53bc6dc279a230501554cdc387c0050848871abece6e709bb6cc6e452ffaac03

Observation a1bfe3e2-dd21-4ac2-b700-66b33ea006b1 · outbound

This paper cites Decolonial AI: Decolonial Theory as Sociotechnical Foresight in Artificial Intelligence.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Decolonial AI: Decolonial Theory as Sociotechnical Foresight in Artificial Intelligence

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.011677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.011677Z digest=sha256:11266f476a3fe610e5391706564ac1b834e9bc9b66397aa6a81aa8ecb4c8156c

Observation 70fb6c68-b68f-47fc-983d-8a83dfe813fc · outbound

This paper cites The Ubuntu Way: Ensuring Ethical AI Integration in Health Research.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code The Ubuntu Way: Ensuring Ethical AI Integration in Health Research

Reference 31

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.233222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.017589Z digest=sha256:f78b5de89ac7aa0ec3f3ed59507773cd1bc597a64fd294bd2cf88923a8f2b42b

Observation 94a99a66-1493-4e33-9421-1d4bb37eb2e0 · outbound

This paper cites Cognitive imperialism in artificial intelligence: counteracting bias with indigenous epistemologies.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Cognitive imperialism in artificial intelligence: counteracting bias with indigenous epistemologies

Reference 32

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.215912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.023430Z digest=sha256:b6f676eaa539e05e9621f266851cc34ec7d56e051a988a79a36d72d8a88b2caf

Observation c63e8aef-0dcc-4da3-961c-a44abc494cb2 · outbound

This paper cites Open Policy Agent Documentation.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Open Policy Agent Documentation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:26:54.487672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.028318Z digest=sha256:45ac5315fbcd19825fadaf4e03dde31b8b9946fcfcde4d701714581ba778f045

Observation b076bd39-16dc-44db-bdb6-ef0622c5dca8 · outbound

This paper cites Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke E.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke E

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:26:54.467615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.034880Z digest=sha256:966a110f1fc18503fdb89c0500da3f206f61b059327e01c30d115ddc9709dc34

Observation 36c42f60-7e48-42f9-aa52-f2607889018c · outbound

This paper cites GOPAL: Governance Open Policy Agent Library for AI System Evaluations.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code GOPAL: Governance Open Policy Agent Library for AI System Evaluations

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:26:54.448176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.041556Z digest=sha256:ddffb8d2a0f9cbc891f673ca3bdb841003076e535d502f3ca67cc4983c8763bf

Observation ec199ffe-87f7-48e4-8e40-91968b938a93 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:26:54.433327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.047214Z digest=sha256:a816f3546747a9d25cf58af83439e0534b50c8438c069e10a5704fb28b75777e

Observation 9ce9153c-f6de-4247-aa07-34229d5929b4 · outbound

This paper cites an unresolved cited work.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:26:54.416667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.052578Z digest=sha256:3874ae5f2f9a977fe615d928ff67bbab8c8133a8c14f340e0ee5fdec5b6bde13

Observation 911f9df6-488b-4741-9d87-ff27681510b0 · outbound

This paper cites Re-imagining Algorithmic Fairness in India and Beyond.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Re-imagining Algorithmic Fairness in India and Beyond

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.058557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.058557Z digest=sha256:c9a04a9f0674386f41711ed0069a614c322f15b05834454db9178d8bdbc332e6

Observation b385bd80-f7b8-4fa4-b8b0-6fac6d0bc5d7 · outbound

This paper cites Proximal Policy Optimization Algorithms.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Proximal Policy Optimization Algorithms

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.065616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.065616Z digest=sha256:2ebdd6fa471b6eeea9e818460662c4f9bd59e9953d8b6c950bf2a55ffcfb96a3

Observation 16c3ad40-dfa6-426e-9308-7e77f3e7b1b3 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.070700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.070700Z digest=sha256:2fc7d1741fb13d26f8ab885445db65446a84a8aced3a1cfb079d53b1df3d1903

Observation 5c26cbf3-529f-47ab-b7f6-9b5d8f13278e · outbound

This paper cites Susiddha AI Project: Dharmic Frameworks for AI Development.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Susiddha AI Project: Dharmic Frameworks for AI Development

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:26:54.399547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.076236Z digest=sha256:c73acc0534300a0d357af9e08624c1ed824b16d6ced6f97b79e48bc8ab5141e2

Observation 841b2c01-0a83-46dc-b905-ae5f810e6df9 · outbound

This paper cites Varshney.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Varshney

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.080891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.080891Z digest=sha256:a032bd6a1f28cca24a040562bf9f64c190399aaf63abfd6a91831f334735e06e

Observation 0a6ddb81-80ca-40fd-8781-4fe238aa8792 · outbound

This paper cites On the Fairness of Causal Algorithmic Recourse.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code On the Fairness of Causal Algorithmic Recourse

Reference 43

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.198823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.087658Z digest=sha256:d25979272b779e55d71f2801402dc221fc7ea950f8b03856d46911fb87236285

Observation 611d6fb4-3e7b-439f-927d-a8a3be294c7f · outbound

This paper cites Machine Ethics: Creating an Ethical Intelligent Agent.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Machine Ethics: Creating an Ethical Intelligent Agent

Reference 44

Resolution
verified exact
doi, observed 2026-08-15T16:26:53.181323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.094667Z digest=sha256:877764fde08e00f4582222ef80aa1a7f663fe4290055475d2b7dc30cca5a9cb5

Observation 5ef44ee1-9582-49a2-993b-acf1684f6fa5 · outbound

This paper cites SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.100627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.100627Z digest=sha256:cc3a7bc6aa21662ffe71b314fc100f1b3ed3dcef88ebba78dfac206007d4ec66

Observation 6011e29e-71e3-440b-ae73-55888d3890eb · outbound

This paper cites UC-MOA: Utility-Conditioned Multi-Objective Alignment for Distributional Pareto-Optimality.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code UC-MOA: Utility-Conditioned Multi-Objective Alignment for Distributional Pareto-Optimality

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-15T16:26:53.630010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T16:26:53.106287Z digest=sha256:6f6728f9c3d4c8474628da9e8dbe5021496e8e3882dd15f0871859c6e0b0dd82

Observation e31dfc4c-6f6e-4ced-9795-9e5d86379c3f · outbound

This paper cites Chemotactic motility-induced phase separation.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Chemotactic motility-induced phase separation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.111340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.111340Z digest=sha256:523c31029266458f850d94738e263730cf9b644341191cc4d951d7e2d58d7a6d

Observation aeb34150-c008-4bce-a39f-c5c9f64e5118 · outbound

This paper cites Adaptive Group Policy Optimization: Towards Stable Training and Enhanced Performance.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Adaptive Group Policy Optimization: Towards Stable Training and Enhanced Performance

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.117357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.117357Z digest=sha256:7bf65c6d25a68861715b07e81b623fe137494b51a37eecd5aae1f184ff620a43

Observation 6e052142-40ff-4270-a482-cb3d6643cbe3 · outbound

This paper cites Debate Helps Weak-to-Strong Generalization.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Debate Helps Weak-to-Strong Generalization

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.123046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.123046Z digest=sha256:b99ca73c3eed80e23a9bcc022a1f07b8d688e5f7cb20947d748b09c7289355e0

Observation b6f0452f-7794-49f0-bcf7-1e02412030ed · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code Understanding R1-Zero-Like Training: A Critical Perspective

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:53.128079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:53.128079Z digest=sha256:dc2e0aec8a110d63c48611f0965547f6233b85163eda78dac2640b30c67d133e

Pith citing papers

Observation 36f0ab0e-b25a-4e32-bc3b-6acc883ef9f3 · inbound

An Automated Framework for Cybersecurity Policy Compliance Assessment Against Security Control Standards cites this paper.

An Automated Framework for Cybersecurity Policy Compliance Assessment Against Security Control Standards ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:40:54.569469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-11T02:25:54.160882Z digest=sha256:baa20c9fa2de8a87fa7ecdb4d75174819e46b988290578076890f5cab6899d18