Pith. sign in

Paper Citation Record · LEDGER

Benchmarking LLM powered Chatbots: Methods and Metrics

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2308.04624.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.04624 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:31:49.439924Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T18:00:00.857339Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5664f3c8-1074-4e5e-acfa-73bdd4e5d990 · inbound

A Survey on Large Language Model based Autonomous Agents cites this paper.

A Survey on Large Language Model based Autonomous Agents Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 174

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:03:01.086336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T04:03:00.340349Z digest=sha256:045058e160e1af141ad6ae005f8ec638ff1358ed3c878ec9f0e0ce2ad0188a96

Observation a7931328-d07c-46a9-9942-539836dc4cbf · inbound

Political-LLM: Large Language Models in Political Science cites this paper.

Political-LLM: Large Language Models in Political Science Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 219

Resolution
unresolved
no resolver link, observed 2026-08-11T19:52:04.741322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:52:04.741322Z digest=sha256:7205b7288e2bb8dffca68d5982bf1ad495a8a1a35289123acec8aa447cd6d361

Observation d12eee62-f31e-4fdd-97a7-d92ef964210b · inbound

Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle cites this paper.

Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:55:17.084581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:55:17.084581Z digest=sha256:069c0f85315e2d7f8c595e8f8ffdc67d0b697f739b5f77811f1eb09390e60dbe

Observation c84ef91f-04a1-432b-8215-a76615f1af0e · inbound

Random-Set Large Language Models cites this paper.

Random-Set Large Language Models Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:49.439924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:31:49.439924Z digest=sha256:2d782982dd735cc2ec0356ba4088bd3dec0631e311c3d7344df6ddea8ac37706

Observation 50393aab-ef14-4e4e-8ded-1339af24ddff · inbound

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric cites this paper.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.713568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.713568Z digest=sha256:8ced02abb1639a01bddb1f4004dd1d352ce8bd6b2e373581b36bbb70bf4c14b1

Observation 3daa7b15-70cc-40f7-8d32-26d0aebf6f61 · inbound

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks cites this paper.

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:33:21.285901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T14:32:56.579146Z digest=sha256:08a6ab1f59ce52ecd67525464c43e5925353fd4cf2e8d73e5f6f0a5dad76ecfa

Observation b954c345-9089-41a9-b539-89645cd2f615 · inbound

GuidaPA: Privacy-Preserving Chatbot for Public Administration via Federated Learning cites this paper.

GuidaPA: Privacy-Preserving Chatbot for Public Administration via Federated Learning Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:26:14.271175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T17:04:39.452280Z digest=sha256:0f7a51485883dddf14c24ea58950fe485ded7af6c934d098715e508ffc4126ac

Observation 1352548e-497e-4e23-a3b8-24500c6c272e · inbound

Accuracy and Satisfaction in Multi-Turn LLM Dialogues for NFR Assessment cites this paper.

Accuracy and Satisfaction in Multi-Turn LLM Dialogues for NFR Assessment Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T18:00:00.858874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-25T23:12:20.065631Z digest=sha256:84991c6c28a72760e813f8fdee6d394ddaffb9902075cec674eb3652e4690231