Pith. sign in

Paper Citation Record · LEDGER

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs

As of 9 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2506.00439.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00439 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:08:19.440424Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T08:25:13.253635Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T08:26:06.323797Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 443c67d2-da7f-4b78-a737-937b63d2874d · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:24.142456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:16.163758Z digest=sha256:0455833eb7d6ac557db91fcf37e7ad6b804337dcba8c48d9f5c76bcd0737d0a0

Observation 743385a2-5f84-4c8f-ba05-d098014d9804 · outbound

This paper cites Program Synthesis with Large Language Models.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Program Synthesis with Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:16.319253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:16.319253Z digest=sha256:70fbd51e519f106ca3a4d022371ca47ffa67f5d00eb4ed2e5b5e3a1520df0713

Observation 281d8725-41de-4498-806f-be8fac304914 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:23.879798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:16.406842Z digest=sha256:bbe4fbe87e674ddaffa20f3499ab0595cbf956409c49b031139ee39cf6a79c0c

Observation cab1e0bb-992c-4034-aaa2-0e08f449b658 · outbound

This paper cites OpenAI Gym.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs OpenAI Gym

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:16.525814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:16.525814Z digest=sha256:b5c01d46b61179f62f0756da4329859eeec3ebb7833e978d3dd2afbac2a0bb64

Observation 8a3c8ca9-7d47-4f21-adaf-2e88858f76bc · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:23.575586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:16.641835Z digest=sha256:a5bf4f7791f652e71294921ae488d01f22c1890eff9038cbb1eb7b6fffe4e99c

Observation dd889b3b-e7ce-40e7-890e-fee60414a3c0 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:23.241206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:16.720571Z digest=sha256:c84889770a563553fa8e0a3ad8ac705bdb3cdc2859a7a0b384c27dc3b73e2e02

Observation 2f9b2ed3-5e33-46fa-8a83-29808666e5ab · outbound

This paper cites Harnessing Multiple Large Language Models: A Survey on LLM Ensemble.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Harnessing Multiple Large Language Models: A Survey on LLM Ensemble

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:16.821916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:16.821916Z digest=sha256:9c45b33eab0cdd6ccf9748f3667e736341b385a848d9f63814089347f983bc21

Observation 577bc709-8155-4355-b95c-ba73864162d5 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:16.916932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:16.916932Z digest=sha256:ad58209d0c5330ed3826a37f0bb07c1c033a9f557a8b33aa780b067f786a4f72

Observation 9d0253c8-d3a5-47cf-90af-982cfce4ec8c · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Training Verifiers to Solve Math Word Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:16.999482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:16.999482Z digest=sha256:2a08cdc139421660b8bafb6d7c040e22199235dac66aaab018adaaffe764d0e3

Observation 8f6a594f-8d58-4ad1-9ae8-af3813541e59 · outbound

This paper cites Prompt-to-Leaderboard.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Prompt-to-Leaderboard

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:17.093651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:17.093651Z digest=sha256:efb042ee87400ebd3c7a9b8da5290a5fe065164e5034752d9e7ab7c057353b8f

Observation e459faa5-ae31-4e1a-9fd5-fa93c30876bf · outbound

This paper cites The Llama 3 Herd of Models.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:17.220995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:17.220995Z digest=sha256:94a75a6e13e2b44089d8a2856df1b378fd6616fcb4e223ae211d4b608f5e440e

Observation 1e0ad496-7944-4c0a-ac25-64718d2fb422 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:22.917576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:17.322107Z digest=sha256:fa5e49588805612775b8fde18fe3cc06cb496bd04aab373b60e3e88e32c57fb1

Observation 50d201fa-3d63-4e7d-87fb-8b73fb959a95 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:17.409173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:17.409173Z digest=sha256:d2d100a83a4c5e18ce39ff4b0df341f3525556f50efcabc9f348b417bcf2d791

Observation 524b54e9-348a-40cc-9619-ca20f41be9ba · outbound

This paper cites DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:17.485544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:17.485544Z digest=sha256:cdfbaf504308b28a2800cc29b2eaa6563cafec8b2de081318de02a03e6eacd9f

Observation bf7b077a-0ca3-4c46-91ce-8f975dbafdec · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:17.549611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:17.549611Z digest=sha256:db5de8b055d1094fe209c56408aebc8ff06ffd9e56091d03036e4e5cd5704d6e

Observation 31c77da4-634b-4055-80f2-deafd89460b6 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:22.665362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:17.663506Z digest=sha256:c063c3243b9d77cfb3c66c751b9ce60ef51d3e3f8b262548a56dee5c8ab443cc

Observation cb692c2d-2a3e-431c-97db-f4ab245b0497 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:22.473983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:17.720530Z digest=sha256:ea63ad6348401a093e88b3a675f974daadb1751ac6db799e3b70a8fa4e4dc854

Observation a549d815-80e1-47a2-8cfc-62901ee7f6a0 · outbound

This paper cites Mistral 7B.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Mistral 7B

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:17.802993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:17.802993Z digest=sha256:b054ae1d77c8678a288e114a01912c310aaf13c722be87f941180c4465f28d05

Observation dfa2395e-4830-4da5-8911-707f033af9fc · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:17.877934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:17.877934Z digest=sha256:e853140cfe00220153d8a8539c9e75e3402bda34c2de41ca0798382f3cbf482f

Observation 7a7ad712-af92-4caa-a821-be5be6d20315 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:17.939571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:17.939571Z digest=sha256:ae4d7a45aeeeb2b4fc30b539ee963ec9ab426fba5b13f04e66c8c3324055c286

Observation 6fc9b9ae-13a2-446f-9bca-db2351b04f2f · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:22.271746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:17.999572Z digest=sha256:dbaf5465c1580817d927349c8cbce01d8cffd53883e732bc9f0a02590ef43177

Observation 6f0e9384-2555-42b8-a69a-25320879844f · outbound

This paper cites Merge, Ensemble, and Cooperate! A Survey on Collaborative Strategies in the Era of Large Language Models.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Merge, Ensemble, and Cooperate! A Survey on Collaborative Strategies in the Era of Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:18.059847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:18.059847Z digest=sha256:6ae2b1d3fbf2b28968a8c1484fdbd2213c1837ceba3befbb214d8d06d78e42d7

Observation a8a2d01d-fd47-46d7-99b3-fc45d1e23abd · outbound

This paper cites SpecFuse: Ensembling Large Language Models via Next-Segment Prediction.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs SpecFuse: Ensembling Large Language Models via Next-Segment Prediction

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:18.114378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:18.114378Z digest=sha256:d3d757b757e0783cbfb4427154bb46f29aa903674f8b15fd5feb2ab527ae6142

Observation ed161b3c-92e6-4fe2-ae4f-99bbb463b4d9 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:22.050993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:18.172915Z digest=sha256:00a573da551081305287e0457ce970adce751703bd83cab9724ef37a4560f126

Observation 5e7e5fc5-6168-4ea1-acb0-2d94ddcbaa6b · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:21.793963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:18.244824Z digest=sha256:a0b55895f314a2f9fdbd2bd86a160f7467ea4c435757a81cb1c153593a0b38fb

Observation 2eebad97-9627-4187-98b5-7c9b48580100 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:18.331915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:18.331915Z digest=sha256:275570560ba91ec4d788d7c418e9fc1153f43deaf329c0daa5a81176d8fe402e

Observation eb090367-d82b-4c77-957c-c86c3591e512 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:21.551510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:18.404393Z digest=sha256:e2ba809adfd7b8c1d2e39f72744efe31e96973a8434ea6584ec5df38ca3a86fe

Observation cbd4f4bf-81fd-4b76-a2af-3649a78b6a8d · outbound

This paper cites Qwen2.5 Technical Report.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Qwen2.5 Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:18.493980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:18.493980Z digest=sha256:45d3bf7ce0310b7b26787f87f46d25e55bf4869c8339756f0149f7a8f27bcb05

Observation 7d98a62b-3cf9-424c-8783-8e14e5f84463 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:21.250204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:18.568496Z digest=sha256:25f60b91505db4dce1957382591e0355a87b5e5415153eed81c043d754a8efe9

Observation e278df90-644b-41e4-94cf-7e99d6ca1d9f · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:21.073019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:18.634061Z digest=sha256:9b9ff3cd2ead43ccd118151e0852d7d9415437d2af3ede3b64130eb6d895b922

Observation 1bc7b8fa-3005-47fa-96d9-cee705901e62 · outbound

This paper cites Proximal Policy Optimization Algorithms.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Proximal Policy Optimization Algorithms

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:18.705282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:18.705282Z digest=sha256:f7ae064c938ab22ddeca6492338180d9bd89a3863b4adad8ae63e34fb53c037c

Observation 12e159d0-0887-48d6-988b-1da76d1aba3b · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:18.763436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:18.763436Z digest=sha256:13c5bfe03465c5cff79a1998ba6e3bdcb6b110b9744d94c3c887d9940d87a902

Observation 9a8a8d3d-b74d-492e-90dd-d1ddb70e2626 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:20.872157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:18.820375Z digest=sha256:e5301b96fb924337225ff62b21e6213e846720707b1e051a5c5c088389429a02

Observation 603b89a0-84da-4573-a571-26d546695c0c · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:20.676124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:18.882584Z digest=sha256:f405cc7f80c8d13afb8ddf2410f20be6f0eee34f200c9cfb624c2fac587f01d6

Observation 632c940d-8642-40ea-8c58-88d8c2a31e12 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:20.467587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:18.950176Z digest=sha256:68a825f64a52768ab3823f8d5ec4e90909fe6e1e5c0d5517e94cd7512e368de0

Observation dc798efc-e3c8-48fc-8cae-8400349bfae1 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:20.291745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:19.026175Z digest=sha256:ffdd8a784e29a8cfa700705bb5c90f0dd517c515f8e47f97d5fd124d64860a98

Observation b6505634-88a6-4c6e-9742-ce478955637d · outbound

This paper cites Qwen2 Technical Report.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Qwen2 Technical Report

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:19.046979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:19.046979Z digest=sha256:9077a44426e681f1487361032753e3685a5aa0f3dc3bcb5361d360a763438a3d

Observation 7a3d534d-4aac-4423-bff9-389f98ca96ab · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:20.102910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:19.049442Z digest=sha256:fc07b0867a3f882e9225c8ca2078393dd02487b843d219d6671b9ed6f6f520bd

Observation 2b568b36-20c5-42d0-964f-4b7eedaef9bd · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:08:19.910986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:19.051944Z digest=sha256:fa34ecf4b3056e50e7763ed1f8fa0803414cf615c52e7158a07f1a6135f22bcc

Observation a3538b41-65a8-40cd-9408-f3867b7e9582 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:19.084919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:19.084919Z digest=sha256:3d3732625ddd7468cdd5340fdcfca5245035b58c4b4ad4343b5251a30474481a

Observation 20d790aa-3441-4b2c-ac7a-88d3688270e0 · outbound

This paper cites an unresolved cited work.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs Unresolved cited work

Reference 41

Resolution
verified exact
doi, observed 2026-08-07T12:08:19.655242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T12:08:19.177130Z digest=sha256:b49d127dd2ed63d2cb03110bbb61d4a56da3651bd7d159f8a064096b30684508

Observation 85ee61f5-4280-4ede-bbbe-f0fdf462b4c1 · outbound

This paper cites CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:19.253109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:19.253109Z digest=sha256:a455f813d95045d00612b44a6fcf42d85a3b8cde6ae7ab715f4c5d7842908b26

Observation dd5d54dc-7458-4027-80d2-c5efc4b3ad05 · outbound

This paper cites online" 'onlinestring :=.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs online" 'onlinestring :=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:19.340618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:19.340618Z digest=sha256:e23da4c42b59be71278e4b32e6e50165ce181d2137e3fae76c2791d66335728c

Observation 718a4390-f83f-422b-926b-efedb15b9e03 · outbound

This paper cites write newline.

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs write newline

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:19.440424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:19.440424Z digest=sha256:78d58ea9780fa9e1348814feb712cf58556f484afff81fb555d86ed9f94ee5df

Pith citing papers

Observation 6205a938-1b5e-4261-aab3-3c3c281b56c0 · inbound

RV-HATE: Reinforced Multi-Module Voting for Implicit Hate Speech Detection cites this paper.

RV-HATE: Reinforced Multi-Module Voting for Implicit Hate Speech Detection RLAE: Reinforcement Learning-Assisted Ensemble for LLMs

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:26:06.325773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T08:25:13.253635Z digest=sha256:281059876b86d1f45b72e545fcc9520df05b9fbae25104f6051738a59b17231d