Pith. sign in

Paper Citation Record · LEDGER

Benchmarking LLM powered Chatbots: Methods and Metrics

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2308.04624.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.04624 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:31:49.439924Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T18:00:00.857339Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5664f3c8-1074-4e5e-acfa-73bdd4e5d990 · inbound

A Survey on Large Language Model based Autonomous Agents cites this paper.

A Survey on Large Language Model based Autonomous Agents Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 174

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:03:01.086336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T04:03:00.340349Z digest=sha256:acf347fb742929c040398e26c531c9d4e8c6f9837c7c075ff83d40a230d64622

Observation a7931328-d07c-46a9-9942-539836dc4cbf · inbound

Political-LLM: Large Language Models in Political Science cites this paper.

Political-LLM: Large Language Models in Political Science Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 219

Resolution
unresolved
no resolver link, observed 2026-08-11T19:52:04.741322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:52:04.741322Z digest=sha256:282846aec71184f062653033752a0cf05e6fda8b2789e490f5e80ccf6e0b1f57

Observation d12eee62-f31e-4fdd-97a7-d92ef964210b · inbound

Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle cites this paper.

Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:55:17.084581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:55:17.084581Z digest=sha256:069c0f85315e2d7f8c595e8f8ffdc67d0b697f739b5f77811f1eb09390e60dbe

Observation c84ef91f-04a1-432b-8215-a76615f1af0e · inbound

Random-Set Large Language Models cites this paper.

Random-Set Large Language Models Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:49.439924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:31:49.439924Z digest=sha256:2d782982dd735cc2ec0356ba4088bd3dec0631e311c3d7344df6ddea8ac37706

Observation 50393aab-ef14-4e4e-8ded-1339af24ddff · inbound

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric cites this paper.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.713568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.713568Z digest=sha256:8ced02abb1639a01bddb1f4004dd1d352ce8bd6b2e373581b36bbb70bf4c14b1

Observation 3daa7b15-70cc-40f7-8d32-26d0aebf6f61 · inbound

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks cites this paper.

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:33:21.285901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T14:32:56.579146Z digest=sha256:82e6558d005f2f530a4c7a189e323fead4a8c9509a4420a93a11bf74883c877f

Observation b954c345-9089-41a9-b539-89645cd2f615 · inbound

GuidaPA: Privacy-Preserving Chatbot for Public Administration via Federated Learning cites this paper.

GuidaPA: Privacy-Preserving Chatbot for Public Administration via Federated Learning Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:26:14.271175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T17:04:39.452280Z digest=sha256:7d862dbc39ff5e40c53f9bd4585b473920bbf8118229b4332b4adf77798c9078

Observation 1352548e-497e-4e23-a3b8-24500c6c272e · inbound

Accuracy and Satisfaction in Multi-Turn LLM Dialogues for NFR Assessment cites this paper.

Accuracy and Satisfaction in Multi-Turn LLM Dialogues for NFR Assessment Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T18:00:00.858874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-25T23:12:20.065631Z digest=sha256:909f6c630cb99168fd2e15c59d24a0839236da4611b271cee9ade65bfdacfa9a