Pith. sign in

Paper Citation Record · LEDGER

Language Model Cascades: Token-level uncertainty and beyond

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2404.10136.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.10136 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T19:37:15.904648Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T11:18:03.484596Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation daeb02fd-472a-4355-ad9d-ed733221a94d · inbound

Estimating LLM Uncertainty with Evidence cites this paper.

Estimating LLM Uncertainty with Evidence Language Model Cascades: Token-level uncertainty and beyond

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T19:37:15.904648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:37:15.904648Z digest=sha256:aba491fa57db879ce0804447d9a532c1fec9b4b4241c93439f248e556b5ec078

Observation 251bd8a1-2454-4603-9e1c-7b738f151477 · inbound

CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing cites this paper.

CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing Language Model Cascades: Token-level uncertainty and beyond

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T13:58:45.318248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T13:58:45.318248Z digest=sha256:bf147591c74d09c604bdd5fd495e47957f190659a368456b9d0ed786618ef7af

Observation 97d15c5e-8a15-45ac-8972-75ed37be2e8f · inbound

Division-of-Thoughts: Harnessing Hybrid Language Model Synergy for Efficient On-Device Agents cites this paper.

Division-of-Thoughts: Harnessing Hybrid Language Model Synergy for Efficient On-Device Agents Language Model Cascades: Token-level uncertainty and beyond

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T01:02:19.275090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T01:02:19.275090Z digest=sha256:ecd7ecc668fd94970a91c56b1bbf57d098f11c15535b71a91fa68e1122a1c6d4

Observation f6fce31a-2f5b-4d32-8dc5-c421b2fb6a1a · inbound

When One LLM Drools, Multi-LLM Collaboration Rules cites this paper.

When One LLM Drools, Multi-LLM Collaboration Rules Language Model Cascades: Token-level uncertainty and beyond

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T22:33:09.287843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:33:09.287843Z digest=sha256:04c0c6ef9724347122771b2ca7ab38cf7d74f8a4500652cc09a03683f8b2027c

Observation 5d2b0f53-21fc-48b0-a285-77fdc41ebc06 · inbound

Harnessing Multiple Large Language Models: A Survey on LLM Ensemble cites this paper.

Harnessing Multiple Large Language Models: A Survey on LLM Ensemble Language Model Cascades: Token-level uncertainty and beyond

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-23T02:25:19.676602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T02:22:28.649071Z digest=sha256:d10c02ef32784592e6db02abf0b8dffaed4acb3760ea3da84c158d45aa85af39

Observation 3bd4f856-0ba6-4dc6-8b33-a1758255070b · inbound

Maximizing Confidence Alone Improves Reasoning cites this paper.

Maximizing Confidence Alone Improves Reasoning Language Model Cascades: Token-level uncertainty and beyond

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:07:47.311678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:07:47.311678Z digest=sha256:8d068aa1c2f84ed793801e8108981950ebf05635bfeef430cb250e8bf806944a

Observation b407d3de-efad-4227-b031-c7652a78dd28 · inbound

AdaptiveLLM: A Framework for Selecting Optimal Cost-Efficient LLM for Code-Generation Based on CoT Length cites this paper.

AdaptiveLLM: A Framework for Selecting Optimal Cost-Efficient LLM for Code-Generation Based on CoT Length Language Model Cascades: Token-level uncertainty and beyond

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:30:43.623890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:30:43.623890Z digest=sha256:5f4b0503907ae6861cc9b7de2f9b3c3255d6bffd5ff6e4537afb589ac3385aaa

Observation 76a1baff-bdc8-4954-9157-54deaae169d3 · inbound

Direct Reasoning Optimization: Token-Level Reasoning Reflectivity Meets Rubric Gates for Unverifiable Tasks cites this paper.

Direct Reasoning Optimization: Token-Level Reasoning Reflectivity Meets Rubric Gates for Unverifiable Tasks Language Model Cascades: Token-level uncertainty and beyond

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:52:14.114295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T09:48:56.990745Z digest=sha256:b76c862fa499c052dfa927756be7f39987a4d00934dbbef3ca56d12e4d4302c8

Observation 9d6a1a0b-4fcd-4c19-ad82-0233580174fd · inbound

BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute cites this paper.

BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute Language Model Cascades: Token-level uncertainty and beyond

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:05:27.249295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:05:27.249295Z digest=sha256:be1ffb886c94d2030e557f05354fbc8c1b55d1adec7559fd1331343335ede31a

Observation 484d343e-59a4-4b33-83a1-625a2f519888 · inbound

Estimating the Black-box LLM Uncertainty with Distribution-Aligned Adversarial Distillation cites this paper.

Estimating the Black-box LLM Uncertainty with Distribution-Aligned Adversarial Distillation Language Model Cascades: Token-level uncertainty and beyond

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:46:15.108597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T11:01:18.405626Z digest=sha256:28d5c69d98445805daac9b2d683fbdc37ef4b552a8f50a5151e34d1140891f7a

Observation 84e50ab5-d774-4bfa-9e4f-2e812ea503f3 · inbound

When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack cites this paper.

When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack Language Model Cascades: Token-level uncertainty and beyond

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:02:53.740797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T23:58:53.524203Z digest=sha256:571a3d63b2e022dc59c1a302cc97478852217c1fc3181056f4c465b72c59f869

Observation ee436165-4d1e-4593-9ee1-ca1979644024 · inbound

Online Pandora's Box for Contextual LLM Cascading cites this paper.

Online Pandora's Box for Contextual LLM Cascading Language Model Cascades: Token-level uncertainty and beyond

Reference 115

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.284419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T21:59:43.429931Z digest=sha256:b95d8aa728eb85af342b002ab007efafd3cebd5c118fb34212b9578b010540e7

Observation 72c55d02-3357-4915-9c84-41917d6838d1 · inbound

VIA-SD: Verification via Intra-Model Routing for Speculative Decoding cites this paper.

VIA-SD: Verification via Intra-Model Routing for Speculative Decoding Language Model Cascades: Token-level uncertainty and beyond

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:18:03.485943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T09:38:43.443489Z digest=sha256:64edebea98db37d0f08b8df71bfb924db727f3eb9bc696e77d197161ebb84de7

Observation d53083ab-1603-4de3-93a9-6edf5e8c59c1 · inbound

Constraint-Anchored Reasoning Traces cites this paper.

Constraint-Anchored Reasoning Traces Language Model Cascades: Token-level uncertainty and beyond

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T20:14:10.518412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:14:10.518412Z digest=sha256:fbb808c66a354af18b8e38b138a0cd57a5865e324a5534f763e5f4f6bff0bf88

Observation 2d4a6058-9b69-41bd-bebf-b3e2ec233127 · inbound

HACO: Hedged Agent Computing for Reliable LLM Systems cites this paper.

HACO: Hedged Agent Computing for Reliable LLM Systems Language Model Cascades: Token-level uncertainty and beyond

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T13:10:58.424492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:10:58.424492Z digest=sha256:6bbf6d069343e791b9153d560469be8b77ad84ff6112644383ef5c08cc5964a9

Observation 2af96ace-88b1-461d-80cf-5d1def268016 · inbound

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents cites this paper.

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents Language Model Cascades: Token-level uncertainty and beyond

Reference 127

Resolution
unresolved
no resolver link, observed 2026-08-07T13:36:03.312250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:36:03.312250Z digest=sha256:6b021aba969d82ebb925c971e0b1ddc5f1e02ba8329b91ff25b881db0caf687e