Pith. sign in

Paper Citation Record · LEDGER

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models

As of 14 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2506.02726.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02726 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:21:43.077882Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact3
  • verified fuzzy16
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 35378862-893b-4597-b95b-05288e3f5877 · outbound

This paper cites Large language models: A survey of their development, capabilities, and applications.Knowledge and Information Systems, pages 1–56, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Large language models: A survey of their development, capabilities, and applications.Knowledge and Information Systems, pages 1–56, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:47.731516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:38.818556Z digest=sha256:ec2e2d0d7452e71431e56e283e6bceac6a9af5ef55848a03fd9a55ef9ebc5748

Observation d7bbc86f-e931-45c9-8953-57ef7582e0ec · outbound

This paper cites Nlp for social good: A survey of challenges, opportunities, and responsible deployment.arXiv preprint arXiv:2505.22327, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Nlp for social good: A survey of challenges, opportunities, and responsible deployment.arXiv preprint arXiv:2505.22327, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:38.860198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:38.860198Z digest=sha256:281ae217d0966463c0ac172e5e1762bd12d19a798bb248eb6e5f5c3883b21eee

Observation bb19e456-97f4-405b-aae1-7bae18f537a2 · outbound

This paper cites Current applications and challenges in large language models for patient care: a systematic review.Communications Medicine, 5(1):26, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Current applications and challenges in large language models for patient care: a systematic review.Communications Medicine, 5(1):26, 2025

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:47.478778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:38.949163Z digest=sha256:7a06d8c9d8a089e32b84e5cd47ae94429a1c03d0a595fcb1a08cc211108a494f

Observation 57b394cb-018e-46af-be82-05bfc3318205 · outbound

This paper cites A Survey on Large Language Models with some Insights on their Capabilities and Limitations.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A Survey on Large Language Models with some Insights on their Capabilities and Limitations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.055557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.055557Z digest=sha256:a4f76c1bd8b9ee28ec02f58e024355a06debb34bea087a6ae5c4fdb56fd9d35a

Observation 44ab70e2-1c85-4c56-a568-e9ce0a935689 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.Advances in Neural Information Processing Systems, 36:53728–53741, 2023.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Direct preference optimization: Your language model is secretly a reward model.Advances in Neural Information Processing Systems, 36:53728–53741, 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.094280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.094280Z digest=sha256:e628a0202760622400c2d0ae7cc0fbb42aacaeea5b3334c024f51de56910348e

Observation 74c35b06-33bc-4769-a41d-ee2412b33a4e · outbound

This paper cites Reinforcement Learning Enhanced LLMs: A Survey.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Reinforcement Learning Enhanced LLMs: A Survey

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.153862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.153862Z digest=sha256:cbab2ac0db3ffe8f19bf92f5f06708ee908e5b2acf25ca7d26a858e3bed87929

Observation a381681a-455a-4ea2-9fb1-fabb69e82004 · outbound

This paper cites LLMs for Explainable AI: A Comprehensive Survey.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models LLMs for Explainable AI: A Comprehensive Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.242160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.242160Z digest=sha256:ec7c85a4cda85f49b5de73df0a9053926c82dd68d1ccc7f0b332df27032260ac

Observation ba7746db-7ab2-49c0-9c0d-d2fe3e7f91eb · outbound

This paper cites Explainability for large language models: A survey.ACM Transactions on Intelligent Systems and Technology, 15(2):1–38, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Explainability for large language models: A survey.ACM Transactions on Intelligent Systems and Technology, 15(2):1–38, 2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:47.331458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:39.392138Z digest=sha256:5cbc56530e26039bbead4626d780765081068d0b20095e00e1bc71c66a79aad1

Observation ab4988a7-bbac-49a4-8d43-ae72784b60c7 · outbound

This paper cites Understand what llm needs: Dual preference alignment for retrieval-augmented generation.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Understand what llm needs: Dual preference alignment for retrieval-augmented generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.447414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.447414Z digest=sha256:dee184f482b0b2a126bda52d0c13a9f64188423d330bea3dd8cb0b66102ecb21

Observation 4f2b9230-9f4f-41fa-b65c-682f64177050 · outbound

This paper cites Chain of preference optimization: Improving chain-of-thought reasoning in llms.Advances in Neural Information Processing Systems, 37:333–356, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Chain of preference optimization: Improving chain-of-thought reasoning in llms.Advances in Neural Information Processing Systems, 37:333–356, 2024

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.545398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.545398Z digest=sha256:693db8bd9ca212fe5261672123a1aaf42dc78b8ae4143434de1272f18e6bc7d6

Observation 5faf9fb3-54ce-4de8-99ab-631bf7637757 · outbound

This paper cites Context-DPO: Aligning Language Models for Context-Faithfulness.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Context-DPO: Aligning Language Models for Context-Faithfulness

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.613828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.613828Z digest=sha256:a4922c9f93083e497c933d6d18da6426d3d7835e0ba16dc1273fe333d037423e

Observation bb9969f1-74a8-4879-9709-705b98e488d6 · outbound

This paper cites PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:21:44.178173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:39.749604Z digest=sha256:28f333e1b94c7c0e546b1e512a098e40379a0222305da39fd13e89bfc2b78a0a

Observation 26010888-1271-453c-9889-bb799a52a1a3 · outbound

This paper cites Knowpo: Knowledge-aware preference optimization for controllable knowledge selection in retrieval- augmented language models.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Knowpo: Knowledge-aware preference optimization for controllable knowledge selection in retrieval- augmented language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:47.090260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:39.821944Z digest=sha256:84ea38dd19cfc7999a134069da63b4ba815b852c5536af957eebcacb9ad71cdb

Observation 9d28fa35-cf80-4ed3-9bc1-053bd929396d · outbound

This paper cites Iterative reasoning preference optimization.Advances in Neural Information Processing Systems, 37:116617–116637, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Iterative reasoning preference optimization.Advances in Neural Information Processing Systems, 37:116617–116637, 2024

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:39.897385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:39.897385Z digest=sha256:d32da1924f5aacf971a5e14c09e2cb41da48d6af68c846df9f149b58523329d5

Observation 903b4620-bbfa-4740-b989-c9aa6a4f3c2d · outbound

This paper cites Self-Training with Direct Preference Optimization Improves Chain-of-Thought Reasoning.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Self-Training with Direct Preference Optimization Improves Chain-of-Thought Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.015790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.015790Z digest=sha256:624f4be4aa5511eb73eec618482f8ccbb076a28f843e8c9a0a2ebdb3973009b5

Observation a6cab298-93e4-432d-a3cd-31d582c665e6 · outbound

This paper cites Solving Maxwell's Equations.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Solving Maxwell's Equations

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T11:21:44.000302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:40.090748Z digest=sha256:829c55b756b240408a33045b29f947dfe69c6b9d8fe6f8838609a3434a7927fe

Observation a9a06361-e0ee-4ac3-abe2-d72c8f9d7519 · outbound

This paper cites A Survey on Knowledge-Oriented Retrieval-Augmented Generation.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A Survey on Knowledge-Oriented Retrieval-Augmented Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.158290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.158290Z digest=sha256:279aa50aa8e038b5711d8b356355a2bcbdcd9d081e69660172a455adf62ae580

Observation 23df82cc-59c2-4a25-bebe-450afeb4def6 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.284051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.284051Z digest=sha256:217769ade59538bbc981d1a4e5ce1cd54d8bc97c523e73f1c915591b14e0f82d

Observation fa508398-e9cd-4894-bbe0-055d2e765f73 · outbound

This paper cites A Survey on Post-training of Large Language Models.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A Survey on Post-training of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.383980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.383980Z digest=sha256:1a2d30f9a15e41a33ebbad0f56304448272b91d99c580c0303af30eb1b3ecbd3

Observation fdf48b54-8869-47c6-bf9a-03f922f803a9 · outbound

This paper cites A Survey on Personalized Alignment -- The Missing Piece for Large Language Models in Real-World Applications.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A Survey on Personalized Alignment -- The Missing Piece for Large Language Models in Real-World Applications

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.479121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.479121Z digest=sha256:7eafe063fc4b2c9ea3bd1ddf70ef47e149a5c4436402cf868d542f558e006bf7

Observation 9b0ec985-fb2e-429e-bbeb-c23793eb36c8 · outbound

This paper cites MM-RLHF: The Next Step Forward in Multimodal LLM Alignment.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models MM-RLHF: The Next Step Forward in Multimodal LLM Alignment

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.615790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.615790Z digest=sha256:c81429beee452ba7ef6e203b4aa6e28ca671ca0fe276c2929e1512461a649c72

Observation 0204eac1-f9af-4dc2-881c-4b43e76c53fe · outbound

This paper cites Asynchronous RLHF: Faster and More Efficient Off-Policy RL for Language Models.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Asynchronous RLHF: Faster and More Efficient Off-Policy RL for Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.684605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.684605Z digest=sha256:c897d36fdb7cbd026401adad8e2caa376978cb17a250a16c39e7d79a01462fe3

Observation 0434a918-f900-4b85-946d-c03d54180b4b · outbound

This paper cites RLHS: Mitigating Misalignment in RLHF with Hindsight Simulation.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models RLHS: Mitigating Misalignment in RLHF with Hindsight Simulation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.801007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.801007Z digest=sha256:831c929634021600f927a88b14c2fd0ab444c769efbbd01726301be89ba02fec

Observation 82e0992e-7473-4faf-b2b3-07506c9ccda4 · outbound

This paper cites REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:40.927940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:40.927940Z digest=sha256:cb05e2763616895435051a94ae6d4c8f0802d175fcfc12c5d04f6f95763182c7

Observation 12271706-1c52-45c2-9b33-393f1be2538e · outbound

This paper cites Exploratory Preference Optimization: Harnessing Implicit Q*-Approximation for Sample-Efficient RLHF.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Exploratory Preference Optimization: Harnessing Implicit Q*-Approximation for Sample-Efficient RLHF

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.016055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.016055Z digest=sha256:8e73ac9396b9c10f1e10dce9534e5a9fd56ee276ab6f6d964533630ea5426804

Observation fad7e606-d01f-4894-84a3-b0e44b075c91 · outbound

This paper cites an unresolved cited work.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:21:46.854469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:41.058198Z digest=sha256:7e0ab77198d2d6aa4cb2479cbc5c1cf482728733beece9b97dc12a964e0ddd76

Observation bbe81107-5f30-4627-ad40-7fa6c65256cc · outbound

This paper cites Safer-Instruct: Aligning Language Models with Automated Preference Data.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Safer-Instruct: Aligning Language Models with Automated Preference Data

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.130792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.130792Z digest=sha256:2e64b2b55f1880e8a3444f76ebbadde0132053a8000942be99c3532618c71f60

Observation 710738de-a91c-4664-8a43-71f3c1dc08e5 · outbound

This paper cites Self-Boosting Large Language Models with Synthetic Preference Data.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Self-Boosting Large Language Models with Synthetic Preference Data

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.210733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.210733Z digest=sha256:eb3f4a6a3f8757bb64f31dc1eba02c8ee753622247c5a4618fff37ecd8074032

Observation a28d5641-84b1-4e89-a0e7-5e53d514116e · outbound

This paper cites A comprehensive review of large language models: issues and solutions in learning environments.Discover Sustainability, 6(1):27, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A comprehensive review of large language models: issues and solutions in learning environments.Discover Sustainability, 6(1):27, 2025

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:46.524500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:41.269690Z digest=sha256:4a0f00d8f0730b6a609171ea14e3651e6433a068aa600e076621f55a20db7826

Observation 148f7fa5-3df2-48f8-93f9-b611bc3a5f75 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Constitutional AI: Harmlessness from AI Feedback

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.322841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.322841Z digest=sha256:02b157e7e62fa9f81a3f55d8ea68f7a95f817feae231c0d9b8e7bb79a30f9ef0

Observation cdf321d2-69d2-4864-8466-81151cfdede6 · outbound

This paper cites Zhang and J.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Zhang and J

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:46.337621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:41.393033Z digest=sha256:65c23c83d7078930cc2c95be0cdf879c3c28d373e324468d5755cfa89f9e6dc6

Observation 3535564f-345b-4253-9118-ffd667784e46 · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.484804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.484804Z digest=sha256:b48fc86d18e08c75e167ec01092a31253f74bb10bc1a6327959360aeb799bc3d

Observation f5a0a599-c7ff-48f5-93b9-bc6aa98de06c · outbound

This paper cites an unresolved cited work.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:21:46.013978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:41.550399Z digest=sha256:db03ece7ef9014f38dede34e499e5e79137d3daecfbc372e826f55a06d7b37e4

Observation 4a09417b-8f90-4f45-8b86-de503559f55b · outbound

This paper cites Wang and Y.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Wang and Y

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.870818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:41.601069Z digest=sha256:71f12f1a8f3112406ebf5f4154b60a2a9b63dceece54323bd0bd18c294e02a02

Observation a92a8db8-8f5d-4ba3-9e81-588b0fa43f84 · outbound

This paper cites Feng and L.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Feng and L

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.725340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:41.653470Z digest=sha256:349e13f383f5464407d335232045f20b4bb708a657b575595f1801b841ca8678

Observation aedea759-c1d4-422f-bc74-d0649ca28f24 · outbound

This paper cites Retrieval-Augmented Generation with Graphs (GraphRAG).

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Retrieval-Augmented Generation with Graphs (GraphRAG)

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.722040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.722040Z digest=sha256:b422518dffbeb8c82f0189151101ad14289aa65d322653fc044467cec50be370

Observation a8369b8b-1fca-444f-92c3-4f022cc3f304 · outbound

This paper cites Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.804315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.804315Z digest=sha256:a6dd38f8cb185509e98cfc8c72ce289af0e8821082d203ad27903445fab160f4

Observation a465f5ba-829c-4790-8c3e-56fc59df58af · outbound

This paper cites Enhancing chain of thought prompting in large language models via reasoning patterns.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Enhancing chain of thought prompting in large language models via reasoning patterns

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.550033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:41.874814Z digest=sha256:5f261475d9376cecb549b3472d26fec294a178d676461b816e6866cfae1518be

Observation 09bcebd1-1f60-478e-90b7-79b246d1c87f · outbound

This paper cites Transformers Provably Solve Parity Efficiently with Chain of Thought.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Transformers Provably Solve Parity Efficiently with Chain of Thought

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.931755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.931755Z digest=sha256:55046964cbd78d945b322b078bcac774b4777c022477553229b1138972d8fb81

Observation 43ec6d05-cebe-4146-a919-e602bba05855 · outbound

This paper cites Chain of Draft: Thinking Faster by Writing Less.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Chain of Draft: Thinking Faster by Writing Less

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:41.999082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:41.999082Z digest=sha256:bd7ed6fc331cdf25e2117514405d0d23163249cb4da6e73b30e1f305bf391a51

Observation 4f1fe121-7f15-46fa-906c-5d85df9b3985 · outbound

This paper cites Tool learning with large language models: A survey.Frontiers of Computer Science, 19(8):198343, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Tool learning with large language models: A survey.Frontiers of Computer Science, 19(8):198343, 2025

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.046502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.046502Z digest=sha256:8f64e1f0dfa85d555d507edbdb61019a238b4971b570a60f7c297ae71671096b

Observation b3030e9e-68e9-480a-9bf1-70470f0c9d15 · outbound

This paper cites Making Large Language Models Better Reasoners with Alignment.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Making Large Language Models Better Reasoners with Alignment

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.079243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.079243Z digest=sha256:f0586e7664c0e3e8283a8e6ccfa0c1c99d323bd0b6d75b159006524bd70561ac

Observation e9913d1b-478a-4d3d-b1c3-a4700efdfbb5 · outbound

This paper cites PORT: Preference Optimization on Reasoning Traces.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models PORT: Preference Optimization on Reasoning Traces

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:21:43.674287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:42.154602Z digest=sha256:b796f705e14b358067727e0de10f4ce406d6218c4b9b0394840381ad630ef466

Observation e709ee68-ea43-4f99-907c-ab3a2df57a1d · outbound

This paper cites Beyond Chain-of-Thought: A Survey of Chain-of-X Paradigms for LLMs.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Beyond Chain-of-Thought: A Survey of Chain-of-X Paradigms for LLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.204519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.204519Z digest=sha256:102a7dab09657998ec042f0488bd79486353fae0c927281d869d798ab2cbff47

Observation 22ae8ed5-37c1-4c8a-95fe-75f8c167d131 · outbound

This paper cites Preference tree optimization: Enhancing goal- oriented dialogue with look-ahead simulations.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Preference tree optimization: Enhancing goal- oriented dialogue with look-ahead simulations

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.372913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:42.271188Z digest=sha256:fe4ce61009d44f5d7beb5ccb1c3c701651d932b415fc71eb48fca025f92c1519

Observation f98649cf-13d2-4964-9fa3-e4f9334cfae4 · outbound

This paper cites Large language models in traditional chinese medicine: A scoping review.Journal of Evidence-Based Medicine, 18(1):e12658, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Large language models in traditional chinese medicine: A scoping review.Journal of Evidence-Based Medicine, 18(1):e12658, 2025

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.203522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:42.326058Z digest=sha256:d3507326e17405d0c95ee4a59af1774cbacc561613e18c14a3efaff02c23aa04

Observation 06d4c3d7-50f9-45b8-aae0-cfddaa4ce002 · outbound

This paper cites Tcmchat: A generative large language model for traditional chinese medicine.Pharmacological Research, 210:107530, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Tcmchat: A generative large language model for traditional chinese medicine.Pharmacological Research, 210:107530, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:45.046715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:42.411645Z digest=sha256:68b3ec8e5ab115a74d476be7712bd84a3f00f2c84385a30c66004519ebecd889

Observation e99d14b3-39d3-4263-8b86-4882bb6c85ab · outbound

This paper cites Biancang: A traditional chinese medicine large language model.arXiv preprint arXiv:2411.11027, 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Biancang: A traditional chinese medicine large language model.arXiv preprint arXiv:2411.11027, 2024

Reference 48

Resolution
verified exact
raw_fallback, observed 2026-08-07T11:21:43.446236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:42.480389Z digest=sha256:321cae43ce4ccd580a2be8e1beca8543580fc33015d2b91ccdb8e88db330ee58

Observation 17d4fd0d-80d0-4afe-b56d-8bcd6eb95937 · outbound

This paper cites Qibo: A Large Language Model for Traditional Chinese Medicine.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Qibo: A Large Language Model for Traditional Chinese Medicine

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.548438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.548438Z digest=sha256:4caa87153c97f5306b62c928d2cdeca5f92a9f0cc287639afefdde7f965ea9f2

Observation 2ea30630-77fc-4016-a820-90eff300a36a · outbound

This paper cites Ai-powered lawyering: Ai reasoning models, retrieval augmented generation, and the future of legal practice.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Ai-powered lawyering: Ai reasoning models, retrieval augmented generation, and the future of legal practice

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:44.913465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:42.619817Z digest=sha256:f4a7c6df0c1c88b03c9f4e0e6ed0516c292e334e82329a2e24e457fd629413a9

Observation 645540cd-f7b2-4593-be27-28e96ffafad7 · outbound

This paper cites A comprehensive review on financial explainable ai.Artificial Intelligence Review, 58(6):1–49, 2025.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models A comprehensive review on financial explainable ai.Artificial Intelligence Review, 58(6):1–49, 2025

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:44.785971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:42.699536Z digest=sha256:c1e0f4ee8c05bc1d712d5e8af1bbc1f4dce14688b1a1147adeee75526943ecef

Observation 0595f1a3-41d5-4292-913c-e2cba6ba271f · outbound

This paper cites Findings of the association for computational linguistics: Eacl 2024.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Findings of the association for computational linguistics: Eacl 2024

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:44.608922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:42.746394Z digest=sha256:55f0b191a47574a145e6b48c9265070d4f21f2670292e08646175c011be43177

Observation 673dff5e-42c1-4845-97dc-2fade4c69f74 · outbound

This paper cites ALFA: Aligning LLMs to Ask Good Questions A Case Study in Clinical Reasoning.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models ALFA: Aligning LLMs to Ask Good Questions A Case Study in Clinical Reasoning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.818394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.818394Z digest=sha256:24254a7ff4a3e4a774374119c5a937036ae431a3919121a76f156f2894a46200

Observation 6dfbdf6e-383e-4290-b84e-a2609ada25ee · outbound

This paper cites Shennong-tcm: A traditional chinese medicine large language model.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Shennong-tcm: A traditional chinese medicine large language model

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:21:44.433484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T11:21:42.914094Z digest=sha256:9e4bb2b41648025892b59b93ed5ee92a7d3aabff0c6668222bf1f48d5f7ab01f

Observation aada8a85-a2e7-412f-bb92-483f6a257300 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Rouge: A package for automatic evaluation of summaries

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:42.986811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:42.986811Z digest=sha256:eb3d70957481304fa209c13f1bf3227cfad5ef8689844dbb95b1bc130e9f9369

Observation 5f72547f-2c79-4385-95ec-873c5c2c23f2 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

RACE-Align: Retrieval-Augmented and Chain-of-Thought Enhanced Preference Alignment for Large Language Models Bleu: a method for automatic evaluation of machine translation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T11:21:43.077882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:21:43.077882Z digest=sha256:b70bc8c53b7bced50bc093b98be70eabbbc33f6cf4b52862c332568dcd4c2774

Pith citing papers

No inbound Pith citation observations are available.