Pith. sign in

Paper Citation Record · LEDGER

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference

As of 11 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 3 inbound Pith citation observations for arXiv:2501.02336.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.02336 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:18:51.833410Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:51:52.225072Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T10:51:52.148427Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact2
  • verified fuzzy1
  • unresolved34
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fda33ad9-0c0b-4d8c-ba67-65ab0890c9c1 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.001164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.001164Z digest=sha256:cdbb8634032797f7caac3aa279d358a70e4706cda7e84235a3ea51ab3efe6811

Observation 9b5e98f5-25c4-40da-91fe-12a3358c4161 · outbound

This paper cites write newline.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.047303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.047303Z digest=sha256:8c6ba911ac7d44c673e490f232209c0f1b5dc3b7b4792235b61095cc64e3ad89

Observation 5a03ccba-efd8-46fe-9ae4-cc36b50e62dc · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Yi: Open Foundation Models by 01.AI

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.053783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.053783Z digest=sha256:ab18115656019d9c60772fc3af80a3b352900bab6a617e1d5b1af0fd039d0996

Observation 1712b992-0c93-4d6d-9a86-c828002be37e · outbound

This paper cites Fast and Robust Early-Exiting Framework for Autoregressive Language Models with Synchronized Parallel Decoding.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Fast and Robust Early-Exiting Framework for Autoregressive Language Models with Synchronized Parallel Decoding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.060790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.060790Z digest=sha256:7f773c1f8bdc9322cc726cd2993a68ee27e22dcc013e98bda2d4fa7079e759d7

Observation b74bbc94-8f70-454c-a92a-8c876860bc73 · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.107322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.107322Z digest=sha256:51acbe9398d2b613b0eaea21ae79c1679406d6e94626f85b7e9d84e32988874e

Observation 568f83cd-d702-4997-a299-67ab365bc37e · outbound

This paper cites CodePlan: Repository-level Coding using LLMs and Planning.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference CodePlan: Repository-level Coding using LLMs and Planning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.112306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.112306Z digest=sha256:a73c1d3a61993b5c8eb331833c24f9ace75872184ca63f1356ada56304bba308

Observation 02806810-2f50-4b7e-835c-100a4ce0d71f · outbound

This paper cites EE-LLM: Large-Scale Training and Inference of Early-Exit Large Language Models with 3D Parallelism.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference EE-LLM: Large-Scale Training and Inference of Early-Exit Large Language Models with 3D Parallelism

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.124883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.124883Z digest=sha256:418e595de2bae798bb5d47af6be454099203760d9878a9da81b094fb78a1e71d

Observation af31485e-0ef6-4a8c-a61f-2093a3400e0c · outbound

This paper cites DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.167526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.167526Z digest=sha256:f2bdc9765f6cccd5d03e65b797f831bb43d4e9a49a19e7b709651e834a847986

Observation 2724502a-2f5d-4205-aa19-03fae0b273b9 · outbound

This paper cites SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.199241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.199241Z digest=sha256:7c781573192feef262e23830c074c43c5d7b80c930c3247aee3a76cdbd238b83

Observation cafb6c6c-2984-4419-8d33-dcdb35188236 · outbound

This paper cites Multi-News: a Large-Scale Multi-Document Summarization Dataset and Abstractive Hierarchical Model.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Multi-News: a Large-Scale Multi-Document Summarization Dataset and Abstractive Hierarchical Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.208387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.208387Z digest=sha256:c79f04ff5e1503cc28977fa5d17db86e568711e3c1441b5cc8db6b905bd65d40

Observation a87b23ba-a0c6-436c-9241-be8e2fec7b7c · outbound

This paper cites Not All Layers of LLMs Are Necessary During Inference.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Not All Layers of LLMs Are Necessary During Inference

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.213603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.213603Z digest=sha256:727eb498d028f3327213ec24331273498cbe6a8cb670d869ade7c25f8b600d64

Observation 1e410ddf-31ce-4697-a539-4964f08141cf · outbound

This paper cites Efficient Attentions for Long Document Summarization.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Efficient Attentions for Long Document Summarization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.258739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.258739Z digest=sha256:d94c2d1636910804c3f8e90d507df984b49ccd6019d5ecd96545f90af527193e

Observation 0e4d8f32-ea77-4a25-b01e-acf5c62356ad · outbound

This paper cites FFN-SkipLLM: A Hidden Gem for Autoregressive Decoding with Adaptive Feed Forward Skipping.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference FFN-SkipLLM: A Hidden Gem for Autoregressive Decoding with Adaptive Feed Forward Skipping

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.266499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.266499Z digest=sha256:d1d966f3480aee5119c87082d956b8e95f308a948e86b570d1b9609cae97e3f3

Observation 1fbe6fb6-d8c0-41bd-b5d2-ce0472f62f3e · outbound

This paper cites MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.272353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.272353Z digest=sha256:e588775b16ab06cc6bfc052cd40920cb4e44dcf1415f1aa50d0d62e6d0746e82

Observation 18ee931f-be0d-4f59-82cb-8d680a427853 · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.277675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.277675Z digest=sha256:f47edbd78e38e65298f36c53e6eea2aec4812f7b39f85066e690a708be8a5e60

Observation 85168557-8a11-4578-9a65-5a0b9a130cab · outbound

This paper cites TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.300728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.300728Z digest=sha256:348ea76b30048e74d00c8a30aa3f43ac54fc04e9b2ead555b3f74e2f69cef9c6

Observation 1324aa69-e22a-4ee8-85f8-f86f419877d2 · outbound

This paper cites an unresolved cited work.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:18:52.529670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T22:18:51.319807Z digest=sha256:0d0585f10ddbbe18b0e34251374d457303a6efc961b97dd9b921cf56feb21eec

Observation 1aa1e3ed-0a73-495f-a190-c17a574ae58d · outbound

This paper cites an unresolved cited work.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:18:52.514355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T22:18:51.325219Z digest=sha256:e2eb887c34d5221e69662b99d0fba75d2e47fb6ae8212761f359d10cbbcab7e9

Observation 840c9b4a-9ec1-4582-b927-882ab27fbc7f · outbound

This paper cites SnapKV: LLM Knows What You are Looking for Before Generation.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference SnapKV: LLM Knows What You are Looking for Before Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.331905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.331905Z digest=sha256:5e7c44a9a4615f4b123e2972e8d45756628fa76fe1840969dd5d32c5e45818a5

Observation dc25739e-3f13-4584-b622-584211f3430a · outbound

This paper cites World Model on Million-Length Video And Language With Blockwise RingAttention.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference World Model on Million-Length Video And Language With Blockwise RingAttention

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.350983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.350983Z digest=sha256:6bf70a9b66945023cb508dcb58d41fac6fde9a9e3ffd508fb9977337d6ecda80

Observation dffbf877-f025-48af-b992-15f5fbcf3e39 · outbound

This paper cites RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.369710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.369710Z digest=sha256:c084f136caee5a7cf5e45e0527597b3a7fdb0bd47caebcea14e8ccfdf6ea8afc

Observation 27c41345-81d6-4a20-81ac-b2844c98a424 · outbound

This paper cites Scaling Laws of RoPE-based Extrapolation.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Scaling Laws of RoPE-based Extrapolation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.381911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.381911Z digest=sha256:3d81032382abafc9e35be52e9cf9faec34ec5973f169fed865f2122cafba616b

Observation b6ae7ff4-8ff5-45ac-9468-cac4efb6cb83 · outbound

This paper cites Accelerating Inference in Large Language Models with a Unified Layer Skipping Strategy.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Accelerating Inference in Large Language Models with a Unified Layer Skipping Strategy

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.387701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.387701Z digest=sha256:041fd9c13f659d2088dab8d4be973202d9b5cbbdebbdf13704cc122289d96876

Observation 1eccc7a6-d42b-4e0a-b9d8-14b81af2ad4d · outbound

This paper cites an unresolved cited work.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:18:52.498814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T22:18:51.393615Z digest=sha256:aa939cbd3f37e4b2ee0b333f8889f8c2c94f2d2c71397c92a5f0174870c1c1ed

Observation 8f13a4cf-bef0-4225-ab68-08f71dfcbbe8 · outbound

This paper cites Generative Agents: Interactive Simulacra of Human Behavior.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Generative Agents: Interactive Simulacra of Human Behavior

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.398389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.398389Z digest=sha256:019690a0289f552da0b1a50d847d5faa71c3b95f9d4e669185c5b509de212fbb

Observation b879b134-b948-45df-8966-5c73e36c2626 · outbound

This paper cites YaRN: Efficient Context Window Extension of Large Language Models.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference YaRN: Efficient Context Window Extension of Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.480717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.480717Z digest=sha256:d0ca5561cb235eaf15b77afaf3f95517f0c241f47cdc19df9e175aaf2912e417

Observation b04a5975-4563-4606-94c3-091b36a05d56 · outbound

This paper cites an unresolved cited work.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:18:52.481445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T22:18:51.517257Z digest=sha256:e709c92e8b52f6bfde5ebcd584eb0e14cec28b9f3ff215f567f8cd02744bf4f6

Observation c6977aa0-a079-4793-9bb4-747a53402e01 · outbound

This paper cites Preble: Efficient Distributed Prompt Scheduling for LLM Serving.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.523731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.523731Z digest=sha256:578b88def885cd3a01a85ebda82cc8a0cc94920bf0880c856f7a9377b26894ea

Observation 38c55796-9ca6-4cc0-b47c-9737db6d0b4d · outbound

This paper cites an unresolved cited work.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Unresolved cited work

Reference 30

Resolution
malformed identifier
no resolver link, observed 2026-08-10T22:18:51.531437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.531437Z digest=sha256:6db44f0533568384f6dcf37bdc92b720687e1fde7e2df47c810fc8dc9c52930d

Observation d3dd4578-2a71-43cf-981d-5b825dde4a18 · outbound

This paper cites Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.672133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.672133Z digest=sha256:0d1e2034ce850f50e590dd9c13e67752a8cf2137ba781be85c590c892b2c5cb3

Observation 619b7794-c4b7-429e-8429-4387c32ad76d · outbound

This paper cites an unresolved cited work.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.679041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.679041Z digest=sha256:330c52a8ef548752dcb62ebd3eab7d5d6454dc3ad47caa991cedf419fa279b55

Observation f9a20ed0-b739-4e41-b30b-10236ba5cc2b · outbound

This paper cites an unresolved cited work.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:18:52.453471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T22:18:51.684393Z digest=sha256:0461e748bb775ce34c083cd91775870b0ba6944c0112690756f81c6b82707497

Observation e1e2cbd2-4286-4f28-ad37-8150a4d23bf4 · outbound

This paper cites X.; Wei, Z.; and Wen, J.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference X.; Wei, Z.; and Wen, J

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:18:52.437445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T22:18:51.689647Z digest=sha256:2645c26aa586c6800b40f8b61a468d1faeea111639aacbf99feba8abd209eee9

Observation 708cfd60-e01d-43f5-8c89-d8051f00413e · outbound

This paper cites VCSUM: A Versatile Chinese Meeting Summarization Dataset.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference VCSUM: A Versatile Chinese Meeting Summarization Dataset

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:18:51.996847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T22:18:51.697592Z digest=sha256:428144d164903e3cf5255636ef470dd11ca3dbce504a01af5fa1d722f1f8b070

Observation 362854ae-c1c3-494a-a680-f32ea8b83ee6 · outbound

This paper cites InfLLM: Training-Free Long-Context Extrapolation for LLMs with an Efficient Context Memory.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference InfLLM: Training-Free Long-Context Extrapolation for LLMs with an Efficient Context Memory

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.761837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.761837Z digest=sha256:d0a7c3fdcc9e492d3eb122242a53594227cd7b5aac9cfacf7c267d2a063421e3

Observation 208be0d1-d760-45dc-8df0-f5d66bfe54a6 · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.822998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.822998Z digest=sha256:25818ae3fcd58239e1c259839624a8c129d86e663cea723f689d457adb6bf192

Observation 890d759c-0000-4867-a1cc-4ea51db766ef · outbound

This paper cites an unresolved cited work.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:18:52.421038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T22:18:51.828162Z digest=sha256:309b29a4e67ab512927f5b15966ed821a469ad3d5f779472ff55ef3eaff73c6d

Observation de74219c-4565-421a-8cba-e24940d5915e · outbound

This paper cites Hierarchical Skip Decoding for Efficient Autoregressive Text Generation.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Hierarchical Skip Decoding for Efficient Autoregressive Text Generation

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:18:51.934033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T22:18:51.833410Z digest=sha256:40f581cd8fb1b56235b968206776e0dbe70965611ff1cc16ad022c26cd98ccf4

Pith citing papers

Observation 18ed080e-0432-4a98-8c6b-f03e02882084 · inbound

DASH: Input-Aware Dynamic Layer Skipping for Efficient LLM Inference with Markov Decision Policies cites this paper.

DASH: Input-Aware Dynamic Layer Skipping for Efficient LLM Inference with Markov Decision Policies AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:52.225072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:52.225072Z digest=sha256:d360a4665f4ce1e04563a5db4c663db1895c095ad40a14c3f092edd2af8fc46e

Observation 5d9032ef-73f0-4afd-9cf3-2e37da695058 · inbound

SkipGPT: Dynamic Layer Pruning Reinvented with Token Awareness and Module Decoupling cites this paper.

SkipGPT: Dynamic Layer Pruning Reinvented with Token Awareness and Module Decoupling AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:51:52.153846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T10:51:51.504854Z digest=sha256:06c6ea6623398bf50369b81e1c45d67e76039b9c276d4ebc57188e162d138f7d

Observation 3e695c57-6c7f-4887-8a3e-0044a15c9bc9 · inbound

QTALE: Quantization-Robust Token-Adaptive Layer Execution for LLMs cites this paper.

QTALE: Quantization-Robust Token-Adaptive Layer Execution for LLMs AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T01:10:00.166572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:10:00.166572Z digest=sha256:6e881cbc1b409bcaf5fb5b2dbf478728492f8d74c7545d687bcb6d37dd5c36e8