Pith. sign in

Paper Citation Record · LEDGER

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric

As of 9 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 3 inbound Pith citation observations for arXiv:2508.11269.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.11269 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:05:57.621012Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T01:55:09.658053Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-11T01:57:51.086130Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact8
  • verified fuzzy8
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch6

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9bae77a1-71c0-483c-a593-0c037772ef94 · outbound

This paper cites Vaswani, N.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Vaswani, N

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.962138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.001668Z digest=sha256:489f86bf66b2452890d22f4fb528eb773e112dcfe996fff5ee5a91f969cb0bd8

Observation 96f77f29-410f-4f25-9bb8-a92ca2e2fcf5 · outbound

This paper cites Brown, B.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Brown, B

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.814023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.062872Z digest=sha256:c04ee40337fc7b4ceaa2edd211e77360de8a28d153327c8a3d08783aaa6d39a0

Observation 8dccf6c3-f44c-49ce-b933-6412e6fed1cd · outbound

This paper cites Ouyang, J.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Ouyang, J

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.679901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.170129Z digest=sha256:28879af8f3a1304185447721e7bfac7ac57d04ec41af2f8c9a078bbc09deb215

Observation 8fef53c5-e5d5-4b98-9de7-582cb720414b · outbound

This paper cites GLM: General Language Model Pretraining with Autoregressive Blank Infilling.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric GLM: General Language Model Pretraining with Autoregressive Blank Infilling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:55.270317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:55.270317Z digest=sha256:68c011294188aa389f436ee74affe5cfd2446be81dbdd6d5c6d885d2b08c5f16

Observation 5d61cd7a-20f7-4efc-848a-8769c1d9bbf3 · outbound

This paper cites GLM-130B: An Open Bilingual Pre-trained Model.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric GLM-130B: An Open Bilingual Pre-trained Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:55.316579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:55.316579Z digest=sha256:47ccc8e28f62d3c5204a0e67ddfba51140be5fea4d184bf8e59936b61ff67a8b

Observation 158519bc-5207-48cf-8ba1-422161b784da · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric LLaMA: Open and Efficient Foundation Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:55.382267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:55.382267Z digest=sha256:424063b2aaf13fc38d579b0de66deeb13e86081065f627985416313e669bcdf6

Observation 1f208a59-b529-4f93-9737-2d38836ba360 · outbound

This paper cites Shahriar, K.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Shahriar, K

Reference 7

Resolution
verified exact
doi, observed 2026-08-05T20:05:58.465668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.446911Z digest=sha256:886700ed84856e53d4483b532be5b4200ad6408a7f38c1653b874cb7816d59b6

Observation a279ddc0-739d-4855-aa19-5c7c67ef6a17 · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 8

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:06:00.584585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.504516Z digest=sha256:c84ed5b7a375a89d186c0b3b3dede4ee8a2b05e54b25c15ff3afb43b26bf1a2e

Observation c496601d-97d8-4903-b08b-21d81420e44d · outbound

This paper cites Zhang, F.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Zhang, F

Reference 9

Resolution
verified exact
doi, observed 2026-08-05T20:05:58.233180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.592127Z digest=sha256:e40a3272c0056c246c3e931e112996b149466f15455bb35df644eba4c6e2727d

Observation ca6de3fc-395e-4fb4-88b3-b30e3576e016 · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 10

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:06:00.384999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.641971Z digest=sha256:f9d35436bad74639d25fcdc49ad8504eeef1f5bacbe5cfd75b189ee7c639f2df

Observation 05560984-3a0d-411a-8ddf-2f7ee5ef093a · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 11

Resolution
verified exact
raw_fallback, observed 2026-08-05T20:06:00.171684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.724786Z digest=sha256:8492902e8ecd6dae10e38d3e8205988ab7634d06c346ccdb55e2ed91a8ba6393

Observation 63c32996-0d7a-4a66-b9be-48bcc1c7ab1c · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 12

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:05:59.884792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.786076Z digest=sha256:b3bba520ec0572d1e5abff9b500c21aaa5dcbb49e1f1a35f3fcf8151926b9c9d

Observation 4641ba15-0b3d-4877-a593-1a2cb85897fb · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 13

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:05:59.579192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.865354Z digest=sha256:7f2df46513e456cd7bf4839028e0b6a5369869e6af75f76bb67e004ad88ae055

Observation 467b45b3-22a2-4cbf-8c32-67f0d2e1e0fc · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 14

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:05:59.359007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.932000Z digest=sha256:cf67a48eb1727ef0f7025e375e795eb43b73f7961021019bc21f81611e5f386a

Observation 250d5fb4-a44b-49c3-b4f7-e454699144f7 · outbound

This paper cites Varghese, N.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Varghese, N

Reference 15

Resolution
verified exact
doi, observed 2026-08-05T20:05:58.117010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:55.985556Z digest=sha256:5c4230f1af056be6506a69e5985ba5757550c496b4f1faa822a62b15f69dc375

Observation 6175d6d0-fd09-4f91-b986-0ccff5a9f19c · outbound

This paper cites Bianco, R.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Bianco, R

Reference 16

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:05:59.122227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:56.031251Z digest=sha256:89c472e3d4fea71a591614063ae330c3d9be5b932663dc82d2b2bcf3ec6047d9

Observation 0bd623d0-3e81-4e66-b604-3a679a725a75 · outbound

This paper cites Coleman, D.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Coleman, D

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.550910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:56.114028Z digest=sha256:a181a989f2cac727c15fbc76857ab1ec516698dcc12d4cf2cdde24fe523a7e5f

Observation 318fe081-99b2-49ea-95e9-37f8c7e19890 · outbound

This paper cites Kuzmin, M.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Kuzmin, M

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.430693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:56.185910Z digest=sha256:213ac7e21fda99a742196137dda2637701b60e8baae92fd14307d304380cb2e8

Observation 0fcb3489-a28a-4da2-84f5-e73f3e652083 · outbound

This paper cites FP8 Formats for Deep Learning.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric FP8 Formats for Deep Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.251641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.251641Z digest=sha256:26cfa13050af269cd944d174bba054575fd341761cc60cf9ab439c839f1ee8a2

Observation 5de54629-625a-4e03-bbf4-08768e3860ab · outbound

This paper cites Rodriguez, E.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Rodriguez, E

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.307042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:56.339619Z digest=sha256:116d495b23988fd121adb038402be1672e24593349432000c10a996a8f907ff7

Observation 0a90f51a-5dea-4e9c-9b51-09e640e2fd34 · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:06:01.145418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:56.411430Z digest=sha256:3ebd26df0a7ca635f59582ae0edcd1ee1c5a9c64c852cc517374dfb65b4af054

Observation a84879eb-6ed3-43c0-88cd-f65798ae0201 · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 22

Resolution
verified exact
doi, observed 2026-08-05T20:05:57.958969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:56.470438Z digest=sha256:43c540beec154b7200985a25dc594cfe65acffc3e717a8b4d244228aa087d610

Observation 38c6f5dd-736c-457b-b665-cebc0f133e81 · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.559532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.559532Z digest=sha256:6db8dabd9e6a83cc086fda3268e216135edb92c680e85e7372aa6943f635564c

Observation 7ef26c64-5753-4e4c-8c44-322cb9e87bcd · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 24

Resolution
verified exact
raw_fallback, observed 2026-08-05T20:05:58.831268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:56.625867Z digest=sha256:dbc3d09e5a09c73e61b8f1c1c7a1e829cdac6519cc9a72d598c1fa6c9335a9ac

Observation 50393aab-ef14-4e4e-8ded-1339af24ddff · outbound

This paper cites Benchmarking LLM powered Chatbots: Methods and Metrics.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.713568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.713568Z digest=sha256:30d84d8b86485505faab86a4b2cc17cfe7ceb9c9d4a8b5f865d0b6a3ecb15ba4

Observation 2b6d70b8-8968-4298-8a25-a3a8aaa67cfe · outbound

This paper cites LLMRec: Benchmarking Large Language Models on Recommendation Task.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric LLMRec: Benchmarking Large Language Models on Recommendation Task

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.773541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.773541Z digest=sha256:6f6f8f30257eb6f107663589ce566c2ac199c776facdd767758891921b475514

Observation 77854e3e-7cf4-4d08-91ed-216b311768a7 · outbound

This paper cites Benchmarking Large Language Models for News Summarization.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Benchmarking Large Language Models for News Summarization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.847389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.847389Z digest=sha256:ef297f0041e7e12ae3a8044d53611c5fbdb96992380f78cd209baebbb8a1127a

Observation a69ab616-d340-4ff6-8fdd-169e10513b8a · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 28

Resolution
verified exact
doi, observed 2026-08-05T20:05:57.855990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:56.934150Z digest=sha256:ad0f077a2a8086e57f9c1cf2a856c2dd4178542dc8115b20f2fcaf1d41bc0e7f

Observation d7737b23-1a9c-49b8-a376-a73641473bbd · outbound

This paper cites Nugteren, CLBlast: A tuned OpenCL BLAS library, in: Proceedings of the International Workshop on OpenCL, IWOCL ’18, ACM, Oxford, UK, 2018.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Nugteren, CLBlast: A tuned OpenCL BLAS library, in: Proceedings of the International Workshop on OpenCL, IWOCL ’18, ACM, Oxford, UK, 2018

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.997246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.997246Z digest=sha256:c1147b14ff08868f139b9b20d9686f1cef9c311773e6dbf4f6a018a7e46331d8

Observation 1e6ca89c-a1ba-433f-af5e-ec660fd72a9d · outbound

This paper cites Gebraad, A.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Gebraad, A

Reference 30

Resolution
verified exact
doi, observed 2026-08-05T20:05:57.726370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:57.065743Z digest=sha256:7192048f82523c944732da8decc23da4ac9724636f37b1a56c60768367265665

Observation bbb99788-89e8-4c14-871c-d812d8e0286e · outbound

This paper cites LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:57.150331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:57.150331Z digest=sha256:b8dce263954ef73cbfdf93d69f516dd55b23572b54d059fb5fb6ac6c842a092a

Observation 5ce29606-fc79-4210-b1bb-06a683abf5fc · outbound

This paper cites GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:57.229767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:57.229767Z digest=sha256:acc46d528e426a1e38e77459a968d5809a9898669827f5ddb7abcf52720e4f4f

Observation 236f72f6-978c-4337-b9ca-9a91c9c7f68c · outbound

This paper cites LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:57.318565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:57.318565Z digest=sha256:9d11a3c3d0cf447912b7b673c75548b069f31f5a55138cc8bf87d740bcdb70f4

Observation 4fa0a500-24c5-49f2-82bf-1b3cef5443ed · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:06:01.031012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:57.382306Z digest=sha256:44084bb8d5eddf8b2cc42bf2f72413b2d94f6b12620a656b08ac5b1344fb9b3d

Observation 98f79e35-1f90-4ce3-bf3d-8f70e23be6f2 · outbound

This paper cites Banner, Y.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Banner, Y

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:00.895185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:57.493137Z digest=sha256:47990837724f310f53d262600a57c9b995f18797fee941eaa46532083edc3567

Observation 14aecba7-3b88-4f69-ba96-79c87f35e852 · outbound

This paper cites Efficient LLM Inference on CPUs.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Efficient LLM Inference on CPUs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:57.550624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:57.550624Z digest=sha256:6038c0fcef8f0d1b8436a1b40365f187259a4a3e0e85f70efe473c788ecbd024

Observation 66c69ea8-ba93-4365-80ff-de6b573f562f · outbound

This paper cites Huyen, Evaluation metrics for language modeling, The Gradient (2019).

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Huyen, Evaluation metrics for language modeling, The Gradient (2019)

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:00.765994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:05:57.621012Z digest=sha256:fd6743c4b1a7fe1cc2adf493727f250749761e344d37aea0eb596795ff291ac3

Pith citing papers

Observation 05089264-03e2-453c-85ae-e6d8114f2fbe · inbound

CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration cites this paper.

CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:07:13.076124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T04:06:24.409303Z digest=sha256:ca41d9d1fc4fba95613a6e457604ee560e79aaf3bc0856a90a50711e3d25a391

Observation 13b88533-4d67-4fff-8396-c5c91f6487bd · inbound

Think Before You Grid-Search: Floor-First Triage for LLM Serving cites this paper.

Think Before You Grid-Search: Floor-First Triage for LLM Serving Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-08T22:45:40.079768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-08T22:38:12.637901Z digest=sha256:521230f77fffe493c6d85c2184994a94d386dbf430f53ceb151d90230cdfa8dc

Observation 6d2d85db-7cfb-4463-b785-3896ab79837e · inbound

Think Before You Grid-Search: Floor-First Triage for LLM Serving cites this paper.

Think Before You Grid-Search: Floor-First Triage for LLM Serving Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-11T01:57:51.114012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T01:55:09.658053Z digest=sha256:3cb9a657497bb415a8f3c69f87a3d5ad25b8f6ad3e373bfcab93350c8e522340