Pith. sign in

Paper Citation Record · LEDGER

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric

As of 14 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 3 inbound Pith citation observations for arXiv:2508.11269.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.11269 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:05:57.621012Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T01:55:09.658053Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-11T01:57:51.086130Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact8
  • verified fuzzy8
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch6

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9bae77a1-71c0-483c-a593-0c037772ef94 · outbound

This paper cites Vaswani, N.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Vaswani, N

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.962138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.001668Z digest=sha256:ec2ca8cb9a71f72a393339db53a025d931f648acaef625c047ff21184b90efe0

Observation 96f77f29-410f-4f25-9bb8-a92ca2e2fcf5 · outbound

This paper cites Brown, B.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Brown, B

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.814023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.062872Z digest=sha256:e0e19f3ea732fdbc3682a428238007c4f0c7be8fc1805bd7ed0644756b9982c3

Observation 8dccf6c3-f44c-49ce-b933-6412e6fed1cd · outbound

This paper cites Ouyang, J.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Ouyang, J

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.679901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.170129Z digest=sha256:7a4022bfaf29a79f911b8f9942f342cb64be73d4c8689c6ce710b9b1d4f3c207

Observation 8fef53c5-e5d5-4b98-9de7-582cb720414b · outbound

This paper cites GLM: General Language Model Pretraining with Autoregressive Blank Infilling.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric GLM: General Language Model Pretraining with Autoregressive Blank Infilling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:55.270317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:55.270317Z digest=sha256:5dee9a592cc02a48de6c42bc8795824c071e86ed34f84e51f977343a3fd5d637

Observation 5d61cd7a-20f7-4efc-848a-8769c1d9bbf3 · outbound

This paper cites GLM-130B: An Open Bilingual Pre-trained Model.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric GLM-130B: An Open Bilingual Pre-trained Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:55.316579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:55.316579Z digest=sha256:47ccc8e28f62d3c5204a0e67ddfba51140be5fea4d184bf8e59936b61ff67a8b

Observation 158519bc-5207-48cf-8ba1-422161b784da · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric LLaMA: Open and Efficient Foundation Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:55.382267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:55.382267Z digest=sha256:424063b2aaf13fc38d579b0de66deeb13e86081065f627985416313e669bcdf6

Observation 1f208a59-b529-4f93-9737-2d38836ba360 · outbound

This paper cites Shahriar, K.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Shahriar, K

Reference 7

Resolution
verified exact
doi, observed 2026-08-05T20:05:58.465668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.446911Z digest=sha256:c0c7c76604d5b6b0b9e4a2c0c607cc14f14100d0d6561cae3b8e4764fa9388c3

Observation a279ddc0-739d-4855-aa19-5c7c67ef6a17 · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 8

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:06:00.584585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.504516Z digest=sha256:be42ba4f589d4f32b6ff709e6b80a1c6c4f6a045363ae5f027f99da661ed0fb3

Observation c496601d-97d8-4903-b08b-21d81420e44d · outbound

This paper cites Zhang, F.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Zhang, F

Reference 9

Resolution
verified exact
doi, observed 2026-08-05T20:05:58.233180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.592127Z digest=sha256:145d27310b25ebc541326dbad1a5c040f15611a409e7661388c648ffefc2b9b9

Observation ca6de3fc-395e-4fb4-88b3-b30e3576e016 · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 10

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:06:00.384999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.641971Z digest=sha256:f77ed1af911479bb302d8c7688854a90280f0e831370390910c52a33953fe717

Observation 05560984-3a0d-411a-8ddf-2f7ee5ef093a · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 11

Resolution
verified exact
raw_fallback, observed 2026-08-05T20:06:00.171684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.724786Z digest=sha256:813f1d875d69d8e587ab5ac0f447fb51b4f9ca868075b2d014f2f416e8ffc73e

Observation 63c32996-0d7a-4a66-b9be-48bcc1c7ab1c · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 12

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:05:59.884792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.786076Z digest=sha256:d71aea1a8a3f0f20670a02523356653e70b5af4fa7b7be3cdef4d0653afbf221

Observation 4641ba15-0b3d-4877-a593-1a2cb85897fb · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 13

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:05:59.579192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.865354Z digest=sha256:cf41a9d799914f8d9e6a512886995d96db9291179d6f4434cc2b293fde6d24d5

Observation 467b45b3-22a2-4cbf-8c32-67f0d2e1e0fc · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 14

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:05:59.359007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.932000Z digest=sha256:7809627df6bc4338b90e78bc0730b5b57132449c2406801635bbda57df84152d

Observation 250d5fb4-a44b-49c3-b4f7-e454699144f7 · outbound

This paper cites Varghese, N.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Varghese, N

Reference 15

Resolution
verified exact
doi, observed 2026-08-05T20:05:58.117010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:55.985556Z digest=sha256:17807a109d0959ceae6101cea5e0bbbea7416c51786267d98d513ac6c0669f0d

Observation 6175d6d0-fd09-4f91-b986-0ccff5a9f19c · outbound

This paper cites Bianco, R.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Bianco, R

Reference 16

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T20:05:59.122227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:56.031251Z digest=sha256:5361e02b84bdad51c7a77a7e16eb306b0f026330dfa9d122575e2577901e20e9

Observation 0bd623d0-3e81-4e66-b604-3a679a725a75 · outbound

This paper cites Coleman, D.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Coleman, D

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.550910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:56.114028Z digest=sha256:7cfa3393c4b0584493afcd2a8b063dfc3a3fd7298368051f60794357bc9a9dd7

Observation 318fe081-99b2-49ea-95e9-37f8c7e19890 · outbound

This paper cites Kuzmin, M.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Kuzmin, M

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.430693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:56.185910Z digest=sha256:9caa11625f7ca7146745e95df35d968f49fc48947a0972e94d941d0b3aef0123

Observation 0fcb3489-a28a-4da2-84f5-e73f3e652083 · outbound

This paper cites FP8 Formats for Deep Learning.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric FP8 Formats for Deep Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.251641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.251641Z digest=sha256:3147dd62700528a1e399082d6966d9cebaa53a64a5b79d9602a74384032be665

Observation 5de54629-625a-4e03-bbf4-08768e3860ab · outbound

This paper cites Rodriguez, E.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Rodriguez, E

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:01.307042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:56.339619Z digest=sha256:d14ca7bcc13b99cdb2c20c68b307d80b800842f1d632d4d6fb94ce5a4a9b71f1

Observation 0a90f51a-5dea-4e9c-9b51-09e640e2fd34 · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:06:01.145418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:56.411430Z digest=sha256:3eec5d955e1adb4c0f74f0237276c90c8a2a773fe073f119c024c4932a1509c7

Observation a84879eb-6ed3-43c0-88cd-f65798ae0201 · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 22

Resolution
verified exact
doi, observed 2026-08-05T20:05:57.958969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:56.470438Z digest=sha256:00fb875cbae4f0984192b7708e6448a515b246a54a458e71563349445b2869d6

Observation 38c6f5dd-736c-457b-b665-cebc0f133e81 · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.559532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.559532Z digest=sha256:6db8dabd9e6a83cc086fda3268e216135edb92c680e85e7372aa6943f635564c

Observation 7ef26c64-5753-4e4c-8c44-322cb9e87bcd · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 24

Resolution
verified exact
raw_fallback, observed 2026-08-05T20:05:58.831268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:56.625867Z digest=sha256:4906119c1e74e2ab616aef7b154114d4ead03c71db268cb5bec55cfeb4f632d3

Observation 50393aab-ef14-4e4e-8ded-1339af24ddff · outbound

This paper cites Benchmarking LLM powered Chatbots: Methods and Metrics.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.713568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.713568Z digest=sha256:5509ef01f59845da9607a4ec94c2b40f5239a2a58a3c398bcf4a640bbd4bcfe0

Observation 2b6d70b8-8968-4298-8a25-a3a8aaa67cfe · outbound

This paper cites LLMRec: Benchmarking Large Language Models on Recommendation Task.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric LLMRec: Benchmarking Large Language Models on Recommendation Task

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.773541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.773541Z digest=sha256:cf427a1a2cdc0b72e665e8c01af3e03de2e3420b97c64ccd79392e1e807e4f50

Observation 77854e3e-7cf4-4d08-91ed-216b311768a7 · outbound

This paper cites Benchmarking Large Language Models for News Summarization.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Benchmarking Large Language Models for News Summarization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.847389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.847389Z digest=sha256:ba777c21d0260be815c05889f5007f1a99eea4e8493b59860e46699fb8322298

Observation a69ab616-d340-4ff6-8fdd-169e10513b8a · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 28

Resolution
verified exact
doi, observed 2026-08-05T20:05:57.855990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:56.934150Z digest=sha256:3cc5b9d71246ccf259decd06832835a15807de4899d1fb3e321e21a9432fc84a

Observation d7737b23-1a9c-49b8-a376-a73641473bbd · outbound

This paper cites Nugteren, CLBlast: A tuned OpenCL BLAS library, in: Proceedings of the International Workshop on OpenCL, IWOCL ’18, ACM, Oxford, UK, 2018.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Nugteren, CLBlast: A tuned OpenCL BLAS library, in: Proceedings of the International Workshop on OpenCL, IWOCL ’18, ACM, Oxford, UK, 2018

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.997246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.997246Z digest=sha256:c1147b14ff08868f139b9b20d9686f1cef9c311773e6dbf4f6a018a7e46331d8

Observation 1e6ca89c-a1ba-433f-af5e-ec660fd72a9d · outbound

This paper cites Gebraad, A.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Gebraad, A

Reference 30

Resolution
verified exact
doi, observed 2026-08-05T20:05:57.726370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:57.065743Z digest=sha256:0375feb72ccb2b3cfb3a624c5d0371f8919504136dbce584e4662f0839d4ce13

Observation bbb99788-89e8-4c14-871c-d812d8e0286e · outbound

This paper cites LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:57.150331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:57.150331Z digest=sha256:40dab99abde2c7ec4d5dd822e4ebd1c181ec522a64ba0775cecfe9848a493b3a

Observation 5ce29606-fc79-4210-b1bb-06a683abf5fc · outbound

This paper cites GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:57.229767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:57.229767Z digest=sha256:f3b90be221a696d3a0c6c246a146b025140de9d8eb6ece46ab479373b6bbf463

Observation 236f72f6-978c-4337-b9ca-9a91c9c7f68c · outbound

This paper cites LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:57.318565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:57.318565Z digest=sha256:9d11a3c3d0cf447912b7b673c75548b069f31f5a55138cc8bf87d740bcdb70f4

Observation 4fa0a500-24c5-49f2-82bf-1b3cef5443ed · outbound

This paper cites an unresolved cited work.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:06:01.031012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:57.382306Z digest=sha256:e2d4df5fb29af20efbe6fb86e19c58901d0d05a18f652d74131c6713b3a9206d

Observation 98f79e35-1f90-4ce3-bf3d-8f70e23be6f2 · outbound

This paper cites Banner, Y.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Banner, Y

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:00.895185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:57.493137Z digest=sha256:d31853b51639a1700d88304ea3444a05029377af5e00dfc4d468eb88aeadad27

Observation 14aecba7-3b88-4f69-ba96-79c87f35e852 · outbound

This paper cites Efficient LLM Inference on CPUs.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Efficient LLM Inference on CPUs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:57.550624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:57.550624Z digest=sha256:40d0a2533475aa80fc5c74c9000469beb8fe91b72004061009c8017eb71d954a

Observation 66c69ea8-ba93-4365-80ff-de6b573f562f · outbound

This paper cites Huyen, Evaluation metrics for language modeling, The Gradient (2019).

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Huyen, Evaluation metrics for language modeling, The Gradient (2019)

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:06:00.765994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:05:57.621012Z digest=sha256:b0db557b6418bcea9b5dfbf3df14c7a31c2a8d590f3847e0326def94f235b31d

Pith citing papers

Observation 05089264-03e2-453c-85ae-e6d8114f2fbe · inbound

CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration cites this paper.

CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:07:13.076124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T04:06:24.409303Z digest=sha256:1ff98efe3d24a63d172d1260d1c6b40425312dbe20722f3ab8484113c0ef42b8

Observation 13b88533-4d67-4fff-8396-c5c91f6487bd · inbound

Think Before You Grid-Search: Floor-First Triage for LLM Serving cites this paper.

Think Before You Grid-Search: Floor-First Triage for LLM Serving Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-08T22:45:40.079768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-08T22:38:12.637901Z digest=sha256:28ac76658373defb4c448bbadfbf76dbe13141087ccd21a61ba57d586f9823de

Observation 6d2d85db-7cfb-4463-b785-3896ab79837e · inbound

Think Before You Grid-Search: Floor-First Triage for LLM Serving cites this paper.

Think Before You Grid-Search: Floor-First Triage for LLM Serving Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-11T01:57:51.114012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-11T01:55:09.658053Z digest=sha256:1e29fd59a76c516c4a4493d05070ada9a1aa451e7050a07128bbf9175562dbf0