Pith. sign in

Paper Citation Record · LEDGER

IntellAgent: A Multi-Agent Framework for Evaluating Conversational AI Systems

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2501.11067.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.11067 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:44:21.636050Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T10:58:14.019840Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 86576c80-50e8-48f5-b106-d7eb39242f96 · inbound

$\tau^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment cites this paper.

$\tau^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment IntellAgent: A Multi-Agent Framework for Evaluating Conversational AI Systems

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:52:17.378438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T07:52:17.174347Z digest=sha256:6477e6c035c20012044c9cc1660ec4bc7d6092f0e0f0cdaad613345c7dd0748f

Observation a7a1f55d-5608-463f-ba76-40eaeb9259b8 · inbound

Evaluation and Benchmarking of LLM Agents: A Survey cites this paper.

Evaluation and Benchmarking of LLM Agents: A Survey IntellAgent: A Multi-Agent Framework for Evaluating Conversational AI Systems

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T12:44:21.636050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:44:21.636050Z digest=sha256:8802f681055fff3d8d415cc12f43fe09d9480d10f48da3fdf9ee729376efe07e

Observation 988949c2-74aa-4222-9d9f-bdbecac438b3 · inbound

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory cites this paper.

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory IntellAgent: A Multi-Agent Framework for Evaluating Conversational AI Systems

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:15.534657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-14T23:13:15.016486Z digest=sha256:e9077699ed110764469b67bc4bf05a905a50e5548d991372dd8944a8ed00265a

Observation 2c5c510e-7a6e-4cbd-98c4-77944c47aee0 · inbound

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments cites this paper.

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments IntellAgent: A Multi-Agent Framework for Evaluating Conversational AI Systems

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:49.458088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T19:07:46.077831Z digest=sha256:928c91c3930de70f3b2309059b69284e0c0c79c0bf18befd6e2ce2cb72f714a4

Observation 83caedcb-f771-4e56-98fc-588ed22b31f6 · inbound

MANTRA: Synthesizing SMT-Validated Compliance Benchmarks for Tool-Using LLM Agents cites this paper.

MANTRA: Synthesizing SMT-Validated Compliance Benchmarks for Tool-Using LLM Agents IntellAgent: A Multi-Agent Framework for Evaluating Conversational AI Systems

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:06:12.706281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T10:18:08.444296Z digest=sha256:705547f5ec1be722e3d2ae84cf3359e970331aa4cb10ff1025f358221a79a9df

Observation 09cbdbfa-b8e1-48e5-93ca-fd187250d11e · inbound

Interactive Evaluation Requires a Design Science cites this paper.

Interactive Evaluation Requires a Design Science IntellAgent: A Multi-Agent Framework for Evaluating Conversational AI Systems

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:14.021687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T10:55:08.135630Z digest=sha256:b368b61d15597954e374a821a4e1ab3848cc659dbd485cbe8dae5d9042b066cf