Pith. sign in

Paper Citation Record · LEDGER

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization

As of 7 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2605.11974.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.11974 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T07:32:58.404947Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact12
  • verified fuzzy28
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch19

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0362c18f-461b-491d-9018-31192185d143 · outbound

This paper cites GPT-4o System Card.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization GPT-4o System Card

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.974011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:6128b7046355d3565ab024cd5065487b9d5b6b4fcb4fc4780d722a77edba2801

Observation 06e54a4e-ca37-4498-9d21-88d01532a308 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.963624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:d4bc4570ee3b24d5ae20e95a2c28d9c06bcbc48d5f08876477b64ba8c55674b6

Observation ae8ccd63-c0bc-4c16-9804-15bab25078da · outbound

This paper cites Qwen3 Technical Report.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen3 Technical Report

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.968523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:34b394f76399177b5c503bef45268141ac2e036bca8600dba77ded5d3e4dcac0

Observation d1d65bf8-1161-4fc0-8ce0-56e0b2fb919f · outbound

This paper cites Proceedings of the 47th international ACM SIGIR conference on research and development in information retrieval , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the 47th international ACM SIGIR conference on research and development in information retrieval , pages=

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.323535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:c3d129d747fafc6798a2c7c33cd326cbae1e0485cd1aabf68d8173a6941962e6

Observation eee5ce83-11cb-4c95-b8dd-35a5018dc367 · outbound

This paper cites Proceedings of the AAAI conference on artificial intelligence , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the AAAI conference on artificial intelligence , pages=

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.335650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:92eafab60a583843a6e2bbca75365c31378a2764ee21f2c63c5d8ebb529f066a

Observation 959e19b4-e66d-497e-93b9-591aa06dbd47 · outbound

This paper cites Proceedings of the 2025 ACM Conference on Fairness, Accountability, and Transparency , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the 2025 ACM Conference on Fairness, Accountability, and Transparency , pages=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.327706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:5e73b6a460f3a00870bd5c33db53160b0a83bcae87eb432b55dec35f29850fe1

Observation fafc6f86-061f-4ac5-b3ff-97c189fdd996 · outbound

This paper cites arXiv preprint arXiv:2506.17188 , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization arXiv preprint arXiv:2506.17188 , year=

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.957813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:6e13965b9f60fae5cca27d5d0ef8367cf21208eeeb82d8ee7bc0dbe140f2a3cb

Observation 2f94ca3c-12f3-4bfd-a427-0c5ca5b2ec5b · outbound

This paper cites Proceedings of the 17th ACM International Conference on Web Search and Data Mining , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the 17th ACM International Conference on Web Search and Data Mining , pages=

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.331615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:8224476f42f72ad9d4c532d3bd30b255ef94f7c6f623326515ec03bfcfa320ac

Observation 2a949fc7-ad61-42f2-bd1f-849f64e4e844 · outbound

This paper cites Domaino1s: Guiding LLM Reasoning for Explainable Answers in High-Stakes Domains.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Domaino1s: Guiding LLM Reasoning for Explainable Answers in High-Stakes Domains

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.985874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:c770366db8db21ff34fe4a7291fea5628a31168bfd1a460edadf976c1eab6c18

Observation 7c2004d1-4194-45b3-8fb0-c54057773778 · outbound

This paper cites Annual Meeting of the Association for Computational Linguistics , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Annual Meeting of the Association for Computational Linguistics , year=

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.318946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:16a356cc8573078660d4fc0c8ea2d3c15ec37da5aa91eb8862e777f5414f996f

Observation 0185bf1e-103b-4537-89a5-767e8531119e · outbound

This paper cites Transactions of the Association for Computational Linguistics , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Transactions of the Association for Computational Linguistics , year=

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.312155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:c8ede0250536e10b3adc8d964ef36e5182fedb96077390890c6ed7252a8dfac0

Observation 1ab875a2-f7f7-4d64-8f5e-8f021281e532 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.315375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:093dd7e16dadd16474854536efaf25a753cb140ce03d1fb27721ce8ef9cb9eda

Observation 8ecd1e62-f088-4899-9c2d-8180c9e85dd6 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.303993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:0155eb6e6a98d9127adb337b6d29b132be80647533b258962b4b875eb4258f61

Observation 95920722-c4a1-4df3-b8d9-5399024d2278 · outbound

This paper cites Order Matters: Exploring Order Sensitivity in Multimodal Large Language Models.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Order Matters: Exploring Order Sensitivity in Multimodal Large Language Models

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.979718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:8bff9cc98a84346c500f45e30e54b43a2adfb2a2e0838027707315b06554b4b1

Observation 04f98399-2412-448f-8d27-a2456fa272a5 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.308442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:53f222e699b2ca786db3aade73821811d74655ad0013c970712f55f5e6b14e0e

Observation c027ac73-8602-49dd-8ff5-a588af6921a9 · outbound

This paper cites An Information-theoretic Approach to Prompt Engineering Without Ground Truth Labels.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization An Information-theoretic Approach to Prompt Engineering Without Ground Truth Labels

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.991486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:64c9a419bf5b2de83457a0c437cd9206b4749fdee75398db108c61bd9d30eb0d

Observation 35476a24-3d20-4d64-8c2f-f1ee1f11bbf5 · outbound

This paper cites Annual Meeting of the Association for Computational Linguistics , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Annual Meeting of the Association for Computational Linguistics , year=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.241263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:2a6521285daa7bc3ac1f5dd6f1b7937bd91e53d92ce6316c1dc197502ed3ed60

Observation 1c3fe5e4-05d7-40a0-b807-d2bb03a72065 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.270262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:fccf028e6f8be13dfaf7ad8e002f658e3123a57af71da662deb65d60f600d224

Observation 23cfb539-1254-4c46-b3bb-7c27938673dc · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.257913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:4a6e58101bf157652636a83977f1afddc64cbd132e4a129986d7dd5bf677dd82

Observation 1cd9193f-6189-4db2-909a-a3b1fd9fbf99 · outbound

This paper cites 2024 , note =.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization 2024 , note =

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.253722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:cc8ec1509d1f88621d147c18e3bc8f3186a7fa3abf354a6ef203f35ad51194c0

Observation f091f349-633b-4d69-96d5-c6ec09f55695 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.227942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:f87d4c7dab13728ba7242823db4e34260410f4e0886d06fc8db982ae6668d281

Observation d7b8a243-be82-4b97-b640-ee808be6c8bc · outbound

This paper cites Findings of the Association for Computational Linguistics: EMNLP 2024 , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Findings of the Association for Computational Linguistics: EMNLP 2024 , pages=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.223339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:6d4fe3d6a4cee34a90c81317fef939fcef17d256b40c6085660f6cb280db9fe2

Observation 39277bbd-3b76-4088-baf2-022cbeaede7b · outbound

This paper cites Proceedings of the AAAI Symposium Series , number=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the AAAI Symposium Series , number=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.274001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:a74cc95f50edadaf3234fd4890930f6e896517fb8cfb9f9c1f0cce560822a3b8

Observation 4c646333-71cf-4841-a174-ab96ca8692e3 · outbound

This paper cites Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.262203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:01df4c9354ffb3d53d70842dd67e14c40c2f618286d959a9c13d10de6b539f94

Observation fefbe4fd-a195-47d1-b00c-f0533076146c · outbound

This paper cites Mitigate Position Bias in Large Language Models via Scaling a Single Dimension.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Mitigate Position Bias in Large Language Models via Scaling a Single Dimension

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.913435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:b80fcc260a5023185aafe93efc5cade1e1dace39a06f4be0d9b7af10b5a4f964

Observation 6aa3cdd2-2321-4ec5-a7cf-ae64d63c28de · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in Neural Information Processing Systems , volume=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.291111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:edecddd9d87a08372ce769a456350b8ea6cd1ed1fea62502e8178172409912df

Observation eee701a0-3d6c-46f1-8b6d-884e3077443f · outbound

This paper cites Unveiling Selection Biases: Exploring Order and Token Sensitivity in Large Language Models.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Unveiling Selection Biases: Exploring Order and Token Sensitivity in Large Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.831887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:e1ab49f2df1bd301bd653ec41d3e3aa381af4dcf8aab86af79c8dce6b56cab28

Observation 07c141d9-e3fb-4af2-8bdf-d341136cebe4 · outbound

This paper cites 2024 IEEE International Conference on Web Services (ICWS) , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization 2024 IEEE International Conference on Web Services (ICWS) , pages=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.287213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:8f50dd5fb9aa4d0bc149d0040aa53c3d3a88bbda7e62e2d620e4a5702a9fe79b

Observation f37073eb-81db-4968-a4f2-ef937aa54969 · outbound

This paper cites GraphSOS: Graph Sampling and Order Selection to Help LLMs Understand Graphs Better.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization GraphSOS: Graph Sampling and Order Selection to Help LLMs Understand Graphs Better

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.818585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:ff84bff130dc2e979e6b14388d32d7cd12c7468cf0beac441ca27d6cb9200c33

Observation cb156072-3568-49d0-abda-e6c0d8b6697e · outbound

This paper cites Set-LLM: A Permutation-Invariant LLM.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Set-LLM: A Permutation-Invariant LLM

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.928879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:24b5292fb14597ce84544dadad1c32529293c4b16c426a6af2a2f61164c213df

Observation a18cb4a0-36a1-4e26-aac4-0a47b6e80192 · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.295049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:e89c126185779271004cd867871b2475ad2099286796f11fa081bc8265fe538f

Observation 42ea5fed-51e8-49d2-afe6-ee69b4fb1a02 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.836589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:5932d2826dd612fa7da167b506ce4cba8f608bc95c8135514062ddeabd1f6cd4

Observation d95be7c4-4c29-43f2-a4ad-0bafadbe4b76 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:37:29.923782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:8a2f2fbcb789217d68c75bdb8813b02edbf06bc64b233d6b0acc1942597fc330

Observation 9076d9d2-869e-42e8-baf6-73e554382b6c · outbound

This paper cites InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.887905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:19c2b570c8b8a1e1887aacd9107fc539699789515033b8e6d2eb21e3453745a9

Observation 0b7c421d-921c-4f56-b5c0-8d242e25aee3 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen2.5-Coder Technical Report

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:37:29.933416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:25b9636cb4529233bbd0677eeaa66c5ce55a133386053089989de7ba79132d78

Observation 1f28dbfd-d9dd-415e-8e08-3788e98bfeb9 · outbound

This paper cites Preference Optimization for Reasoning with Pseudo Feedback.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Preference Optimization for Reasoning with Pseudo Feedback

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.881455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:83e1479547cf58876a95e7e9d6eba1a3633e3bbc5bd4872c46af5fdddcc4aac0

Observation d9c77c58-691a-41e5-a802-469db297c38c · outbound

This paper cites ArXiv , year=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization ArXiv , year=

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.237308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:45d9aae6bf767b7be264ddcdcce4c04c63cb6332e1bdca55025c01bf9cdf20c8

Observation 8cc6f019-6d8f-482f-924f-841fba0a9cdf · outbound

This paper cites AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.841604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:e52d65a447eee65b1acc918dc889173f116e0dbd829b3f3a253d50456da1f37c

Observation 903285d5-f30e-477b-99f6-3badcbb1d3b0 · outbound

This paper cites CodeDPO: Aligning Code Models with Self Generated and Verified Source Code.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization CodeDPO: Aligning Code Models with Self Generated and Verified Source Code

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.939329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:9e7f3b8ebc0839e455d9977a26fb91edab3c77d8bc715a6e2af016e871d5eab7

Observation c9042bc7-b01f-43a6-91d0-859b5e513e6c · outbound

This paper cites Proximal Policy Optimization Algorithms.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proximal Policy Optimization Algorithms

Reference 40

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.893688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:ce025746f348b1b69db6cc68d9148f1b502b7320a2ded1a54b0e23462dabf31e

Observation 9147c6af-295b-4b83-831d-307ff1f6eeeb · outbound

This paper cites Advances in neural information processing systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in neural information processing systems , volume=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.282944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:eb258e67f201ca068b52d324417824688d6259c26146ee0f5dcba2b496e6909f

Observation 618675be-7fab-41c9-ab5b-b7dad2daad01 · outbound

This paper cites Advances in neural information processing systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in neural information processing systems , volume=

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.233080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:6311c423aabeafc6e475c9d640f49a5646138be7ff1a30772845423e628ef750

Observation 104d1f22-4915-4274-ac35-a8d19d9c7e2b · outbound

This paper cites Direct Preference Optimization with an Offset.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Direct Preference Optimization with an Offset

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.874766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:104c4131dd86bb2444388315e3eb2e60ff538f6a028b2524c68edd84e8422573

Observation 3dd2b738-438b-4692-8fc8-673e0334ace8 · outbound

This paper cites Step-level Value Preference Optimization for Mathematical Reasoning.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Step-level Value Preference Optimization for Mathematical Reasoning

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.952562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:39a6a221184e19f590da9cde24d4d12dd76b4952aae96469cb211b415c834545

Observation 754e0a15-3f02-4279-b195-0293cbe81b07 · outbound

This paper cites Self-Rewarding Language Models.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Self-Rewarding Language Models

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:01:42.779186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:e72f4fc3e49702c87e16d5a5cb0edeadb4e056b2bf6ce022ffdf8e9e6ae4ef95

Observation d37aecc7-4f51-4543-9e9a-8999ad36b06b · outbound

This paper cites The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.854732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:8e692f2e48f164083c8b7ad8e50d202ce52f93bfa9b9d1925643b67a47dbb476

Observation 9e9cfa9b-6513-47ce-a9ae-ac8768437ab2 · outbound

This paper cites Secrets of RLHF in Large Language Models Part II: Reward Modeling.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Secrets of RLHF in Large Language Models Part II: Reward Modeling

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.945673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:d616952407ecfd5016179f4fdd4acca5f23c08aa4eecb031eebece6b6856f374

Observation 46547bf7-8b40-4470-a8f0-60745dcae52a · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 48

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:37:29.861770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:bde20f679639c134c41ae31f2beb0afb6bd2989a3b9d20702d4d03a5c70acb26

Observation 2b9c0382-4228-465e-8d9e-8b8b7bad606c · outbound

This paper cites Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.848712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:04adb948fd4df5c6a2bac01973f2f4173738a1d410ff7f2e42f05dfcceba3655

Observation 4439da91-bd72-4035-8751-93095a40737e · outbound

This paper cites Neural-Symbolic Solver for Math Word Problems with Auxiliary Tasks.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Neural-Symbolic Solver for Math Word Problems with Auxiliary Tasks

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.868070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:81623de7b65745872ff216fc4bc30f23a7a6f4df930eb9ba267c7aab8e36b094

Observation 0a941f00-88e0-443b-9a98-151094ac51f8 · outbound

This paper cites Proceedings of the 2013 conference on empirical methods in natural language processing , pages=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Proceedings of the 2013 conference on empirical methods in natural language processing , pages=

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.266220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:44703bd73bc72ae8902cffbbb8def1e10a400eddf5687705b2ee2dfbd8c6017d

Observation 54900de6-464b-4bdd-bf27-8f128ab2e243 · outbound

This paper cites Know What You Don't Know: Unanswerable Questions for SQuAD.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Know What You Don't Know: Unanswerable Questions for SQuAD

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.908455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:b20484d26d8075d455e902657ec5cadc6db563f85ec3531acecfffd2f57727d3

Observation 088a2967-6062-4c44-b660-a3a96c2e2756 · outbound

This paper cites Eliminating Position Bias of Language Models: A Mechanistic Approach.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Eliminating Position Bias of Language Models: A Mechanistic Approach

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.918096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:0df89eb38d72ec13db4d471478ea41547fe99534fc2e59dbf2f405c146b394cf

Observation f5af5d75-332c-4192-8ec7-1c89e6281408 · outbound

This paper cites Qwen2.5: A Party of Foundation Models , url =.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen2.5: A Party of Foundation Models , url =

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.277998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:bee1e9e4e18992a40361a598be8044bcdb606f5d5220487c00d1452d7955e3db

Observation 67240abf-fa2e-4ba2-ae38-52d51a44508a · outbound

This paper cites In-Context Learning with Long-Context Models: An In-Depth Exploration.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization In-Context Learning with Long-Context Models: An In-Depth Exploration

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.901629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:e35db7389bc5dcf23345d1a83034825000b8cf30ad387ccfdaa25f9f01b2565c

Observation 0d1d92ee-cd46-43f4-9788-2c0648c4691c · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in Neural Information Processing Systems , volume=

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.299642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:5feb6498f65eff3a7a58181bd3df151c436a9d146637d1311d3a8d9becb2dca9

Observation df55ac3c-4b56-4a94-8271-19d8d5c51517 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in Neural Information Processing Systems , volume=

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.245085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:dae3cd43fd49fe2fa384f376feb8e66e4c6232255071d104f234a0e8e10689d1

Observation efcac845-91b7-4687-8860-e397f9c909fb · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Advances in Neural Information Processing Systems , volume=

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.249789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:e1c3510af5d46de8991d4a04795a28b9959bbfbc1319e56f348024354e2cb783

Observation 1243ead2-91b6-4eb9-9fc9-a88278b240b1 · outbound

This paper cites Arm-thinker: Reinforcing multimodal generative reward models with agentic tool use and visual reasoning.

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Arm-thinker: Reinforcing multimodal generative reward models with agentic tool use and visual reasoning

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:37:29.824883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:32:58.404947Z digest=sha256:0c5c668af91d664146a6009f32f8605aa293073c6d6a0f6dab7599db0a6e6df7

Pith citing papers

No inbound Pith citation observations are available.