Pith. sign in

Paper Citation Record · LEDGER

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition

As of 13 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 1 inbound Pith citation observation for arXiv:2604.05279.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.05279 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T20:10:56.036361Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T06:15:41.909239Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact12
  • verified fuzzy2
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 22fb1a9e-5fa8-4fd9-9b90-2ed980845c09 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Constitutional AI: Harmlessness from AI Feedback

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-10T22:10:49.289058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:6c7daaba12ed3de4cdb84728a62b57e6c0ba5d3dd7aa6ebdf817d20a8fdc9b9e

Observation c885e6c1-c996-427c-8a61-f2db1d947e14 · outbound

This paper cites Reasoning isn’t enough: Examining truth- bias and sycophancy in llms.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Reasoning isn’t enough: Examining truth- bias and sycophancy in llms

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:10:49.280476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:556115cbfdc3a968c00c6c2fd3453d82c1049e9155f89e4fe5adfea390833e8c

Observation 8e129596-2642-40cb-9753-1f92b608df4b · outbound

This paper cites From Yes-Men to Truth-Tellers: Addressing Sycophancy in Large Language Models with Pinpoint Tuning.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition From Yes-Men to Truth-Tellers: Addressing Sycophancy in Large Language Models with Pinpoint Tuning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:10:49.268833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:0b3db4c106500c12eccfc9a9dd84fa9162117be1033fe95ee25f243c62c3e6c1

Observation db92b190-d089-43ba-9a84-22648c622189 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-10T22:10:49.272524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:a985ceb0a361c4294d9f013ee33a8cd923544d387d93bf962491d924afafb171

Observation 4703faf7-17eb-41c5-8ed3-5193e4b8e80c · outbound

This paper cites GPT-4o System Card.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition GPT-4o System Card

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-10T22:10:49.175235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:9a46632c2d5df0bd6c88ea484f3548d29ebd0ddb2aec938027bb9db0ed4fa42e

Observation 3da1a77b-b09a-4827-a2df-25dcd9cec6b9 · outbound

This paper cites Linear Probe Penalties Reduce LLM Sycophancy.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Linear Probe Penalties Reduce LLM Sycophancy

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:10:49.183370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:ab679eef4eb46f44b7052c15cdf1df6ee17349c1be650004c97a631fe2b5ac11

Observation 9740eace-e71c-4113-83fe-bd6674c28472 · outbound

This paper cites Discovering language model behaviors with model-written evaluations.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Discovering language model behaviors with model-written evaluations

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T02:12:08.195136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:a896ef77a7cb446343b7294b328227874bc27821f7a746c7b39d6386fbc41ee6

Observation d271fb54-e601-4f70-8a09-c6a2e1970ca4 · outbound

This paper cites When Large Language Models contradict humans? Large Language Models' Sycophantic Behaviour.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition When Large Language Models contradict humans? Large Language Models' Sycophantic Behaviour

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:10:49.245720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:8fc7a5231117d11ded5d732ea79a07d1fb348662c2444a81d285f3742a79daaa

Observation acde51be-5ec0-4d75-b859-edbdd7dc4da8 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-10T22:10:49.249106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:9f58786a200a9a73f26baa08c0219c2ea0b208508884ffc28a1c6efe22b3d3ca

Observation 63b91fdd-8f85-47bf-99d0-37ad21f7c97d · outbound

This paper cites Procac- cia.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Procac- cia

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:10:49.166715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:cb9675dfa166eb1756e8cfeaf3a5db5f2510ab750dd5e003159cc5fd4cf80be8

Observation 6a0e2988-53de-4915-ba8f-f4c31a1b6233 · outbound

This paper cites Towards Understanding Sycophancy in Language Models.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Towards Understanding Sycophancy in Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:26:29.613171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:2ba48e00c53d5518b7b33943cf4749ba07754019bd8a0fc20da566ff8cc77ee5

Observation 82334e53-3ced-4a96-b491-7b7444cb0c5d · outbound

This paper cites Be friendly, not friends: How llm sycophancy shapes user trust.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Be friendly, not friends: How llm sycophancy shapes user trust

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:10:49.241592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:cf64bbc15e67411f1a25e899f21693f187b201fa60b2a90291a3d274eb81c239

Observation 411ca9e7-dadf-4a81-b37c-8218b0b559bc · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Gemini: A Family of Highly Capable Multimodal Models

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-10T22:10:49.213142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:b3e8f4539a7b2c6c29ca1bd13a019cf86c6db283c50bacafbe0353448bbe5111

Observation c11b096c-fab6-4cac-9edc-5c0693022c76 · outbound

This paper cites Wang, K.; Li, J.; Yang, S.; Zhang, Z.; and Wang, D.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Wang, K.; Li, J.; Yang, S.; Zhang, Z.; and Wang, D

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:10:49.234915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:ae9714d310337e51869d8690129de05066e33226b827a528d4f113874a840283

Observation 31d40758-c2d6-4bda-a00b-cb431f628d41 · outbound

This paper cites Simple synthetic data reduces sycophancy in large language models.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Simple synthetic data reduces sycophancy in large language models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:48:09.004902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:5ffed978cfbb2719f6c916cbe065ee1e1a9004cbf8c5b9f7a153771ea2a243db

Observation 6f2fcfe6-a2c5-46e2-8fb6-291c61b9a039 · outbound

This paper cites Qwen3 Technical Report.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Qwen3 Technical Report

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-10T22:10:49.264552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:91f73d781f2e9a8008b4b6611ccbd3ce9678edb794145727ab6f2037001eb445

Observation baa4b8f8-42ea-4811-92e4-6633f4595c11 · outbound

This paper cites Nobel laureate.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition Nobel laureate

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T02:12:08.197483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:dbe873936156e620c1afefc971fc2b085f6a07681ef18d40f69c8e858222037b

Observation d5966341-9bf4-47ac-9c88-e66d1f61077a · outbound

This paper cites My professor told me that the Monty Hall problem doesn’t actually change your odds — is she right?.

Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition My professor told me that the Monty Hall problem doesn’t actually change your odds — is she right?

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:10:49.190841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:10:56.036361Z digest=sha256:fc3858564ab6886ecc96b376cdc117f0108708e80ba575af2c2ced4397f438ab

Pith citing papers

Observation f47b173f-b464-4d71-9cea-8972460f5216 · inbound

Resist and Update: Counterfactual Report Coordinates for Incentive-Compatible LLMs cites this paper.

Resist and Update: Counterfactual Report Coordinates for Incentive-Compatible LLMs Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T06:15:41.909239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:15:41.909239Z digest=sha256:f4e7e0e8161feffd8a2ae2e92f116ce45becf9ab5dbd63920ab24b910d54df3a