Pith. sign in

Paper Citation Record · LEDGER

Efficient Large Language Models: A Survey

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 46 inbound Pith citation observations for arXiv:2312.03863.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.03863 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 46 of 46 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T15:26:18.959612Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

24
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f881ece7-0004-402a-a558-b6a8cdfae25a · inbound

A Survey on the Memory Mechanism of Large Language Model based Agents cites this paper.

A Survey on the Memory Mechanism of Large Language Model based Agents Efficient Large Language Models: A Survey

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:21:39.755184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T07:21:39.440092Z digest=sha256:136a7efff96be1e364bda44159d9075fecd3e6e8f5fefacd802ea087b3ce3b03

Observation c67e25e3-312e-4b6c-be31-feb3e71905fe · inbound

A Survey on Efficient Inference for Large Language Models cites this paper.

A Survey on Efficient Inference for Large Language Models Efficient Large Language Models: A Survey

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:39:33.686744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T02:39:33.007894Z digest=sha256:ac50f28f947902be194b9f990b1c02c1a6269e5a485f5839d06a8427ad7cd628

Observation 7c916fd0-fe42-40e1-a36c-67926ae38a7d · inbound

From Cool Demos to Production-Ready FMware: Core Challenges and a Technology Roadmap cites this paper.

From Cool Demos to Production-Ready FMware: Core Challenges and a Technology Roadmap Efficient Large Language Models: A Survey

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:08:20.777408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T19:07:21.016824Z digest=sha256:b8335cf6302a6a60f8839bd0f70db5ea5cc9dc6d9012b967398414d69b219d3e

Observation b7d85cb7-4d73-4e51-bdfd-5863dc4c2aa4 · inbound

Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models cites this paper.

Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models Efficient Large Language Models: A Survey

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-09T15:26:18.959612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T15:26:18.959612Z digest=sha256:5c3ae0cbb0a05b8563bcb681ea89c877a4bf4288749d2291f6cff624d3b35cd0

Observation dc37a957-0548-46d9-ab98-bf9802b373e6 · inbound

Unveiling Simplicities of Attention: Adaptive Long-Context Head Identification cites this paper.

Unveiling Simplicities of Attention: Adaptive Long-Context Head Identification Efficient Large Language Models: A Survey

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-08T13:46:11.428921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:46:11.428921Z digest=sha256:abe05f919c411330d382794a4cf9df79950aaad371fa3126da4424b58109330c

Observation f94a8adf-9117-452c-8120-defff72b0e5f · inbound

DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving cites this paper.

DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving Efficient Large Language Models: A Survey

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:31:40.679447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T14:30:47.787654Z digest=sha256:9ff9c0b2199bd4a1e41a3d6c7b5a999b0870723d65b015d6106b0ec2cd76b383

Observation 3d0d45d0-175c-4d80-830d-6149eb3feeca · inbound

ALPS: Attention Localization and Pruning Strategy for Efficient Alignment of Large Language Models cites this paper.

ALPS: Attention Localization and Pruning Strategy for Efficient Alignment of Large Language Models Efficient Large Language Models: A Survey

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:32.512737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:31:32.512737Z digest=sha256:a388d5aa27f99e7af1391ebaaac32a19cb813c5093d9583cc62ff41810e6f4c9

Observation d6b78270-85ac-48e4-8724-6d02e4619f28 · inbound

SkipGPT: Dynamic Layer Pruning Reinvented with Token Awareness and Module Decoupling cites this paper.

SkipGPT: Dynamic Layer Pruning Reinvented with Token Awareness and Module Decoupling Efficient Large Language Models: A Survey

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:51.633694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:51:51.633694Z digest=sha256:21675fdd7e0f0c32fdcc95c6b7f295f65a31650e27e954491db25637333d43b9

Observation 0e60339d-5a4a-41cb-b92a-c1acc6c5b01e · inbound

Direct Behavior Optimization: Unlocking the Potential of Lightweight LLMs cites this paper.

Direct Behavior Optimization: Unlocking the Potential of Lightweight LLMs Efficient Large Language Models: A Survey

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:19:35.095186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:19:35.095186Z digest=sha256:3e18e4db115f8cc79c9bf4761a12674051d8afbde30e056b49670d280f5f59c8

Observation adb9d9d7-c967-4860-a65b-ae619bb15522 · inbound

NeurIPS 2025 E2LM Competition : Early Training Evaluation of Language Models cites this paper.

NeurIPS 2025 E2LM Competition : Early Training Evaluation of Language Models Efficient Large Language Models: A Survey

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:34:24.096831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:34:24.096831Z digest=sha256:7faaf04f90751bc207c74eedf4decd332441f7d0b58ef0b4072f172e806535d4

Observation 17a2411b-9c09-491b-81e1-c191ed6b30d3 · inbound

Enhancing Reasoning Capabilities of Small Language Models with Blueprints and Prompt Template Search cites this paper.

Enhancing Reasoning Capabilities of Small Language Models with Blueprints and Prompt Template Search Efficient Large Language Models: A Survey

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:10:39.570052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:10:39.570052Z digest=sha256:6c47314fc6824511aa1c0e1074f0134cad1215462ff4e651c26f950363490adc

Observation b0b4398d-94be-4292-9c31-18049f43db74 · inbound

Semantic Scheduling for LLM Inference cites this paper.

Semantic Scheduling for LLM Inference Efficient Large Language Models: A Survey

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T01:09:21.374055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:09:21.374055Z digest=sha256:6b5b567cb19d1e44b119423030a84c9f54a4cef78d76ef9feb8f26a63adc3fc2

Observation 9fedacd3-5301-4453-b94f-2a3702bf1170 · inbound

EAT: QoS-Aware Edge-Collaborative AIGC Task Scheduling via Attention-Guided Diffusion Reinforcement Learning cites this paper.

EAT: QoS-Aware Edge-Collaborative AIGC Task Scheduling via Attention-Guided Diffusion Reinforcement Learning Efficient Large Language Models: A Survey

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:41.990874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:41.990874Z digest=sha256:1a708589bacf210e977500e958d152123e0aa1b77735ef978a548bcf143cd8ff

Observation 957296a5-1acc-4ae2-932a-265bc358080a · inbound

TASE: Token Awareness and Structured Evaluation for Multilingual Language Models cites this paper.

TASE: Token Awareness and Structured Evaluation for Multilingual Language Models Efficient Large Language Models: A Survey

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T23:21:25.003665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:21:25.003665Z digest=sha256:0ad86ad9e1656c7b85c4b884962cff3505b501aa801d746247eb471d2a8fdfaa

Observation 944abb55-8cab-4d86-a9a0-aec39d7abd7e · inbound

Decoding Memories: An Efficient Pipeline for Self-Consistency Hallucination Detection cites this paper.

Decoding Memories: An Efficient Pipeline for Self-Consistency Hallucination Detection Efficient Large Language Models: A Survey

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T14:33:20.422919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:33:20.422919Z digest=sha256:4a7637003f6c231628182e06e617ce0e369dd65714ac2d4a34f3088a9690429f

Observation f98ecfeb-3a36-4d30-8512-c444a8fa88ca · inbound

DaMoC: Efficiently Selecting the Optimal Large Language Model for Fine-tuning Domain Tasks Based on Data and Model Compression cites this paper.

DaMoC: Efficiently Selecting the Optimal Large Language Model for Fine-tuning Domain Tasks Based on Data and Model Compression Efficient Large Language Models: A Survey

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T12:52:01.691397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:52:01.691397Z digest=sha256:36bedb8b6e57ca14aa8f470d9b37060f7899c0fa14ed89146be188c0bda5e68b

Observation 07f81e60-1143-420b-a45b-ddb6c846cfe2 · inbound

NeuronMLP: Efficient LLM Inference via Singular Value Decomposition Compression and Tiling on AWS Trainium cites this paper.

NeuronMLP: Efficient LLM Inference via Singular Value Decomposition Compression and Tiling on AWS Trainium Efficient Large Language Models: A Survey

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:52:21.884286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T02:51:19.275111Z digest=sha256:63903b488705dca1fcb1d4d95ce8c2ab739d13df68c7195090328077f93d7c4a

Observation 6a2f46ae-f827-4651-bef0-7063865becd7 · inbound

Toward Efficient Agents: Memory, Tool learning, and Planning cites this paper.

Toward Efficient Agents: Memory, Tool learning, and Planning Efficient Large Language Models: A Survey

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-03T09:21:42.553598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:21:42.553598Z digest=sha256:506d502a091c78f544e8550e86b6a236d7423081924aa4702a255f14b415eff9

Observation 937bf638-f0a6-42ee-acb6-fdacb1b4cc1d · inbound

On the Limits of Layer Pruning for Generative Reasoning in Large Language Models cites this paper.

On the Limits of Layer Pruning for Generative Reasoning in Large Language Models Efficient Large Language Models: A Survey

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:40:46.217656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T08:40:24.822863Z digest=sha256:90f3423ea2b9f419194d5d1ba6e704bdd33736d3d3ec612f59aca73fdd8fa487

Observation 3b288f5e-17fc-4319-afdf-832d40031302 · inbound

Triplet Feature Fusion for Equipment Anomaly Prediction : An Open-Source Methodology Using Small Foundation Models cites this paper.

Triplet Feature Fusion for Equipment Anomaly Prediction : An Open-Source Methodology Using Small Foundation Models Efficient Large Language Models: A Survey

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:56:40.989807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T21:52:27.888915Z digest=sha256:ee93d10260d28b04bed407657ddb4ae9dc68bd894d42b51b7ee1359351b3ba44

Observation 546db393-9dfb-4383-89c5-c9a5750d74af · inbound

CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation cites this paper.

CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation Efficient Large Language Models: A Survey

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:23:22.886951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:20:36.975106Z digest=sha256:85069b414d16c93bb64f3cad6b8536f7de61d24053be3df2fefddf2138b9ed2b

Observation da83b55e-56eb-4dc7-abad-db8ccbd05f43 · inbound

Unified Deployment-Aware Evaluation of Open Reasoning Language Models cites this paper.

Unified Deployment-Aware Evaluation of Open Reasoning Language Models Efficient Large Language Models: A Survey

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:01:01.270847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:49:52.594972Z digest=sha256:d8e2e967fdfd25734ab424f3ae29a3e96952eac86bcee349b9ae1ff852121f2f

Observation 82689292-8553-48f8-bda0-2e92b5a88ca2 · inbound

Unified Deployment-Aware Evaluation of Open Reasoning Language Models cites this paper.

Unified Deployment-Aware Evaluation of Open Reasoning Language Models Efficient Large Language Models: A Survey

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:44:05.726912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T09:42:33.144129Z digest=sha256:3c99a8b10c8690cbc4bf5a65bbf7aa89f21c93c9c7105455a387e97c6c2700aa

Observation 0b92a5b2-d53e-4cbb-90aa-981378b81a51 · inbound

Cloud-native and Distributed Systems for Efficient and Scalable Large Language Models -- A Research Agenda cites this paper.

Cloud-native and Distributed Systems for Efficient and Scalable Large Language Models -- A Research Agenda Efficient Large Language Models: A Survey

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:31:30.836748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T06:27:23.580445Z digest=sha256:1fe86749b0784fa0aad4c5c5b6a4d121de90ba21ddb641ea6f243f6703feb1f6

Observation 3494371a-1bef-4b0d-86cf-5492daf91652 · inbound

Compress Then Adapt? No, Do It Together via Task-aware Union of Subspaces cites this paper.

Compress Then Adapt? No, Do It Together via Task-aware Union of Subspaces Efficient Large Language Models: A Survey

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:45:40.935533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T18:10:41.778201Z digest=sha256:0b2d8771a42f636f0cc257a528006002355529364a57088b7e812d9589759224

Observation 780515ae-7217-4263-b25d-88c6ac02bdf5 · inbound

OSAQ: Outlier Self-Absorption for Accurate Low-bit LLM Quantization cites this paper.

OSAQ: Outlier Self-Absorption for Accurate Low-bit LLM Quantization Efficient Large Language Models: A Survey

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:30:44.243755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T18:23:14.935801Z digest=sha256:e990126217ed276a56d566157e48f284250d871ec9e88c4ae2b353af1d87a07e

Observation e4086864-4fc8-4b4f-9fe9-fa9cefeb88a2 · inbound

OSAQ: Outlier Self-Absorption for Accurate Low-bit LLM Quantization cites this paper.

OSAQ: Outlier Self-Absorption for Accurate Low-bit LLM Quantization Efficient Large Language Models: A Survey

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:01:18.166123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T02:59:00.997742Z digest=sha256:76280ffb89a515ce7d359b5b61c16dba9394123764f1a65c2321f9d4d6a57ddd

Observation 77eccf07-7179-4649-bec5-97c8f97a3884 · inbound

Continuous Latent Diffusion Language Model cites this paper.

Continuous Latent Diffusion Language Model Efficient Large Language Models: A Survey

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:11:10.544287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T10:04:09.646578Z digest=sha256:f4cbeee6fa298d26db683b5cd9692d42849c19962d18fc59eb8cde265e00610e

Observation 1f3178b3-5611-41c9-9c58-43cdd5293b41 · inbound

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents cites this paper.

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents Efficient Large Language Models: A Survey

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:46:27.575343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T01:57:22.554881Z digest=sha256:b5f4a61aff907478d0ffa47bfa6c6cdd7a11cb5a2ae269899d48f13731c6febc

Observation 433bdf12-0e1b-4f04-b5f3-93caea3d0838 · inbound

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents cites this paper.

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents Efficient Large Language Models: A Survey

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:45:45.467485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T22:52:54.684185Z digest=sha256:a5754dfd273957ab7e47a12ecf8a33119f06bde0df69f7e391b3718a8ce49534

Observation b641f171-32e8-480c-af79-a60632b11f34 · inbound

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents cites this paper.

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents Efficient Large Language Models: A Survey

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T05:17:31.244809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:17:31.244809Z digest=sha256:ca4424e1a7b8a471bc793619332606fa28b0030f97051712fa6e4fd2fa6c4efa

Observation 51e8d4f7-ff39-4b93-a88a-86c1aafbf229 · inbound

When is Warmstarting Effective for Scaling Language Models? cites this paper.

When is Warmstarting Effective for Scaling Language Models? Efficient Large Language Models: A Survey

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:02:53.104922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T20:00:47.126736Z digest=sha256:975ba766297a7018f89acf69bb47fa80f9c340f8e613469f0a7db339ed0bd10b

Observation 749b13f2-b228-4628-82b8-580fce174359 · inbound

From Text to Voice: A Reproducible and Verifiable Framework for Evaluating Tool Calling LLM Agents cites this paper.

From Text to Voice: A Reproducible and Verifiable Framework for Evaluating Tool Calling LLM Agents Efficient Large Language Models: A Survey

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:49:53.649954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T08:45:56.550821Z digest=sha256:ae604bc38dbc2a1fd4ff9beeee86ebeda4117a5a69d1255e81f94feff6fdb993

Observation b1b11fb2-fafe-485f-b572-1976f07f659b · inbound

Latent Action Reparameterization for Efficient Agent Inference cites this paper.

Latent Action Reparameterization for Efficient Agent Inference Efficient Large Language Models: A Survey

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:48:12.863156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T10:45:20.306945Z digest=sha256:48e1d07fc6bc84a154cd5c66092ea6bc2eb09fbb4aa7151ae7687c06ddf9d9c9

Observation 06f91f01-caa1-4730-9b11-03040e86f958 · inbound

EmbGen: Teaching with Reassembled Corpora cites this paper.

EmbGen: Teaching with Reassembled Corpora Efficient Large Language Models: A Survey

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:38:05.506945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T06:34:39.739666Z digest=sha256:0bf78232f2334b8ef33609e364d9f6911519277069d35857cc615ab5107c2549

Observation cd89a295-a623-4263-a8e0-fa0fd2fc7cda · inbound

GEMQ: Global Expert-Level Mixed-Precision Quantization for MoE LLMs cites this paper.

GEMQ: Global Expert-Level Mixed-Precision Quantization for MoE LLMs Efficient Large Language Models: A Survey

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:36:39.850168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T05:33:06.719954Z digest=sha256:20608cd638f14bfe74ce629d504f785af8c7eed785089986536faa2a625ad63d

Observation 5897d9cf-d54e-4d8d-82d6-ba32d7bd5766 · inbound

Evaluation of ML Resource Utilization Requires Model Life Cycle Assessment cites this paper.

Evaluation of ML Resource Utilization Requires Model Life Cycle Assessment Efficient Large Language Models: A Survey

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:06:14.588198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T17:27:19.467192Z digest=sha256:2649db317186e46143f574a8986a9cb7b52047b86a15ef595632f36ea442a448

Observation bf4af32d-1f4e-4508-8baf-1ea91200213a · inbound

Dynamic Linear Attention cites this paper.

Dynamic Linear Attention Efficient Large Language Models: A Survey

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:27:39.795371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T13:18:18.331031Z digest=sha256:4af4bbbb577f614471016b02fbeb05d5631d1367c9a9266983e2a93eebd4ddc2

Observation 8b7b1267-ff4b-4dbc-a978-b1bd9d5c8332 · inbound

Recency/Frequency Adaptive KV Caching for Large Language Model Serving cites this paper.

Recency/Frequency Adaptive KV Caching for Large Language Model Serving Efficient Large Language Models: A Survey

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:29:39.052734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T13:24:44.218871Z digest=sha256:43fbf31760fffef4f116c55378ff86ea550683dc8f6a2ef12d2c5426eed51131

Observation 9901a9a7-6210-4689-863a-a02f9fad655a · inbound

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks cites this paper.

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks Efficient Large Language Models: A Survey

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:49:52.913147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T05:51:01.446139Z digest=sha256:cd6ab0d17e7281b25f8a4a03f218198aa55fad9b33fc8f88b929e28d359544ed

Observation 1f5c11cf-8d01-4d77-9088-2b6d03f17395 · inbound

BlockPilot: Instance-Adaptive Policy Learning for Diffusion-based Speculative Decoding cites this paper.

BlockPilot: Instance-Adaptive Policy Learning for Diffusion-based Speculative Decoding Efficient Large Language Models: A Survey

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:55:41.384304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T05:59:15.199717Z digest=sha256:3c5c29d2ecaabdd9ace4ff0efc8fbdad95041ba5144ff82f855d52da689a9b2b

Observation 229976d7-c436-4347-81a0-fa0ab7f87b43 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents Efficient Large Language Models: A Survey

Reference 117

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T01:36:44.270544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:68e0e86fd1292f6dbdff3b8341bec1f3ebfb8b674a42dc439938d9143e5145e9

Observation b70e6bcf-a08e-4646-8684-be985470a96c · inbound

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy cites this paper.

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy Efficient Large Language Models: A Survey

Reference 245

Resolution
unresolved
no resolver link, observed 2026-07-14T06:30:16.612345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:30:16.612345Z digest=sha256:cc5c36c6263831ede4bbae1c8cb03f2f6217287c598b0540dba185e181c918e6

Observation 399498e4-c7cc-4f3e-bdcb-93978b1ec766 · inbound

Token Reduction Is Not Cost Reduction cites this paper.

Token Reduction Is Not Cost Reduction Efficient Large Language Models: A Survey

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T06:45:14.741063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:45:14.741063Z digest=sha256:3e1a77d787efd2bb1c8cc59a5220342762923611dc3575f8c849e2a70c7ece59

Observation db817337-0139-4590-ad76-a8058ee2b9cf · inbound

Token Reduction Is Not Cost Reduction cites this paper.

Token Reduction Is Not Cost Reduction Efficient Large Language Models: A Survey

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T04:23:31.693317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:23:31.693317Z digest=sha256:7b06426562a65c6e87f78c24431f882b59ede4c8f4f446f0c74122ff08757fb7

Observation 7f446d23-3496-4481-ab49-d225a2486905 · inbound

CHS-SQL: A Text-to-SQL approach based on Confidence-Guided Heuristic Search Schema Linking process cites this paper.

CHS-SQL: A Text-to-SQL approach based on Confidence-Guided Heuristic Search Schema Linking process Efficient Large Language Models: A Survey

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T11:00:12.321867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:00:12.321867Z digest=sha256:f891d3ef749e983317f3c5fd7b472daa108ee2ee25b7eacabe394a379a89ab8b