Pith. sign in

Paper Citation Record · LEDGER

Benchmarking LLM powered Chatbots: Methods and Metrics

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2308.04624.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.04624 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:05:56.713568Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T18:00:00.857339Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5664f3c8-1074-4e5e-acfa-73bdd4e5d990 · inbound

A Survey on Large Language Model based Autonomous Agents cites this paper.

A Survey on Large Language Model based Autonomous Agents Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 174

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:03:01.086336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T04:03:00.340349Z digest=sha256:4c2fcc7927fbd5f2c6b5f45bf16ab95c9fc3adeaaa8fd010f56e2ea6991584d2

Observation 50393aab-ef14-4e4e-8ded-1339af24ddff · inbound

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric cites this paper.

Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:56.713568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:56.713568Z digest=sha256:30d84d8b86485505faab86a4b2cc17cfe7ceb9c9d4a8b5f865d0b6a3ecb15ba4

Observation 3daa7b15-70cc-40f7-8d32-26d0aebf6f61 · inbound

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks cites this paper.

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:33:21.285901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T14:32:56.579146Z digest=sha256:0ccf67b3722b27df40760a4db0597f2c578b451fbe690466edd24119ed6856f4

Observation b954c345-9089-41a9-b539-89645cd2f615 · inbound

GuidaPA: Privacy-Preserving Chatbot for Public Administration via Federated Learning cites this paper.

GuidaPA: Privacy-Preserving Chatbot for Public Administration via Federated Learning Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:26:14.271175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T17:04:39.452280Z digest=sha256:2e885a4bdcae729b0690e9f7044a7bd46379000202e6cedbfdbaee600d9701cf

Observation 1352548e-497e-4e23-a3b8-24500c6c272e · inbound

Accuracy and Satisfaction in Multi-Turn LLM Dialogues for NFR Assessment cites this paper.

Accuracy and Satisfaction in Multi-Turn LLM Dialogues for NFR Assessment Benchmarking LLM powered Chatbots: Methods and Metrics

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T18:00:00.858874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-25T23:12:20.065631Z digest=sha256:e9050bd8a5bb211227324817d11c10c7f15a2cdf33f640e4de1cb4fe4ef0a409