Pith. sign in

Paper Citation Record · LEDGER

Language Model Cascades: Token-level uncertainty and beyond

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2404.10136.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.10136 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T19:37:15.904648Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T11:18:03.484596Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation daeb02fd-472a-4355-ad9d-ed733221a94d · inbound

Estimating LLM Uncertainty with Evidence cites this paper.

Estimating LLM Uncertainty with Evidence Language Model Cascades: Token-level uncertainty and beyond

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T19:37:15.904648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:37:15.904648Z digest=sha256:5a426ab969839f2f52052afc978a874568677a7cefa8666d40bb5185be52d5f9

Observation 251bd8a1-2454-4603-9e1c-7b738f151477 · inbound

CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing cites this paper.

CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing Language Model Cascades: Token-level uncertainty and beyond

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T13:58:45.318248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T13:58:45.318248Z digest=sha256:c9ec576e1119b70f8f8b037ad30563a88c8792b9060ed6cf20284d1047feb81b

Observation 97d15c5e-8a15-45ac-8972-75ed37be2e8f · inbound

Division-of-Thoughts: Harnessing Hybrid Language Model Synergy for Efficient On-Device Agents cites this paper.

Division-of-Thoughts: Harnessing Hybrid Language Model Synergy for Efficient On-Device Agents Language Model Cascades: Token-level uncertainty and beyond

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T01:02:19.275090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T01:02:19.275090Z digest=sha256:b4c62e86fac0689aba72045829174ccc2a024da9e9962d54e49eb57bd61e4977

Observation f6fce31a-2f5b-4d32-8dc5-c421b2fb6a1a · inbound

When One LLM Drools, Multi-LLM Collaboration Rules cites this paper.

When One LLM Drools, Multi-LLM Collaboration Rules Language Model Cascades: Token-level uncertainty and beyond

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T22:33:09.287843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:33:09.287843Z digest=sha256:1bd1301832a91a1b791da02261c12b1b2802c3006fbdc72d7b664fd5d4b17c72

Observation 5d2b0f53-21fc-48b0-a285-77fdc41ebc06 · inbound

Harnessing Multiple Large Language Models: A Survey on LLM Ensemble cites this paper.

Harnessing Multiple Large Language Models: A Survey on LLM Ensemble Language Model Cascades: Token-level uncertainty and beyond

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-23T02:25:19.676602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T02:22:28.649071Z digest=sha256:ba87a2a475e6810ec6936e3ee358ea55d5850644f85b9138bf3147be6b8b92e9

Observation 3bd4f856-0ba6-4dc6-8b33-a1758255070b · inbound

Maximizing Confidence Alone Improves Reasoning cites this paper.

Maximizing Confidence Alone Improves Reasoning Language Model Cascades: Token-level uncertainty and beyond

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:07:47.311678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:07:47.311678Z digest=sha256:8d068aa1c2f84ed793801e8108981950ebf05635bfeef430cb250e8bf806944a

Observation b407d3de-efad-4227-b031-c7652a78dd28 · inbound

AdaptiveLLM: A Framework for Selecting Optimal Cost-Efficient LLM for Code-Generation Based on CoT Length cites this paper.

AdaptiveLLM: A Framework for Selecting Optimal Cost-Efficient LLM for Code-Generation Based on CoT Length Language Model Cascades: Token-level uncertainty and beyond

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:30:43.623890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:30:43.623890Z digest=sha256:5f4b0503907ae6861cc9b7de2f9b3c3255d6bffd5ff6e4537afb589ac3385aaa

Observation 76a1baff-bdc8-4954-9157-54deaae169d3 · inbound

Direct Reasoning Optimization: Token-Level Reasoning Reflectivity Meets Rubric Gates for Unverifiable Tasks cites this paper.

Direct Reasoning Optimization: Token-Level Reasoning Reflectivity Meets Rubric Gates for Unverifiable Tasks Language Model Cascades: Token-level uncertainty and beyond

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:52:14.114295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T09:48:56.990745Z digest=sha256:4ddafbcf93cc60bbc4f236319fd386b460448f513bca30809159753f66f1eacf

Observation 9d6a1a0b-4fcd-4c19-ad82-0233580174fd · inbound

BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute cites this paper.

BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute Language Model Cascades: Token-level uncertainty and beyond

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:05:27.249295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:05:27.249295Z digest=sha256:be1ffb886c94d2030e557f05354fbc8c1b55d1adec7559fd1331343335ede31a

Observation 484d343e-59a4-4b33-83a1-625a2f519888 · inbound

Estimating the Black-box LLM Uncertainty with Distribution-Aligned Adversarial Distillation cites this paper.

Estimating the Black-box LLM Uncertainty with Distribution-Aligned Adversarial Distillation Language Model Cascades: Token-level uncertainty and beyond

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:46:15.108597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T11:01:18.405626Z digest=sha256:5333a8c4fafcc422ff912f83aabd3f1fd5bb64087101922c6f9d5554a3c3e368

Observation 84e50ab5-d774-4bfa-9e4f-2e812ea503f3 · inbound

When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack cites this paper.

When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack Language Model Cascades: Token-level uncertainty and beyond

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:02:53.740797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T23:58:53.524203Z digest=sha256:c55b9d20221f0f22a57e1f2b704669989e9bb6babfb4407300dd8e9ef6c4d188

Observation ee436165-4d1e-4593-9ee1-ca1979644024 · inbound

Online Pandora's Box for Contextual LLM Cascading cites this paper.

Online Pandora's Box for Contextual LLM Cascading Language Model Cascades: Token-level uncertainty and beyond

Reference 115

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.284419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T21:59:43.429931Z digest=sha256:304516b713883d6a88d181c2ef20073dcb5a6ed2c61b88c5616badbdcf4af3d8

Observation 72c55d02-3357-4915-9c84-41917d6838d1 · inbound

VIA-SD: Verification via Intra-Model Routing for Speculative Decoding cites this paper.

VIA-SD: Verification via Intra-Model Routing for Speculative Decoding Language Model Cascades: Token-level uncertainty and beyond

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:18:03.485943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T09:38:43.443489Z digest=sha256:2c73564a00edd0b85243bb876a5c530b6f80a6d5aa86efd7b3c36442c0f15e1a

Observation d53083ab-1603-4de3-93a9-6edf5e8c59c1 · inbound

Constraint-Anchored Reasoning Traces cites this paper.

Constraint-Anchored Reasoning Traces Language Model Cascades: Token-level uncertainty and beyond

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T20:14:10.518412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:14:10.518412Z digest=sha256:693226f0879035541b6615d8c99544e42386a75048809a849398804143574971

Observation 2d4a6058-9b69-41bd-bebf-b3e2ec233127 · inbound

HACO: Hedged Agent Computing for Reliable LLM Systems cites this paper.

HACO: Hedged Agent Computing for Reliable LLM Systems Language Model Cascades: Token-level uncertainty and beyond

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T13:10:58.424492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:10:58.424492Z digest=sha256:6bbf6d069343e791b9153d560469be8b77ad84ff6112644383ef5c08cc5964a9

Observation 2af96ace-88b1-461d-80cf-5d1def268016 · inbound

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents cites this paper.

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents Language Model Cascades: Token-level uncertainty and beyond

Reference 127

Resolution
unresolved
no resolver link, observed 2026-08-07T13:36:03.312250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:36:03.312250Z digest=sha256:85aa82bab7f606cbcd7b6e544f6cf524e22796ce0e8cc63615aedc3f59a892d5