Pith. sign in

REVIEW 1 cited by

Intention-aware policy graphs: answering what, how, and why in opaque agents

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2409.19038 v1 pith:U2T7WOOC submitted 2024-09-27 cs.AI cs.LGcs.MAcs.RO

classification cs.AIcs.LGcs.MAcs.RO
keywords agentbehaviourmodelagentsemergentexplainingincreasingmeasurements
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Agents are a special kind of AI-based software in that they interact in complex environments and have increased potential for emergent behaviour. Explaining such emergent behaviour is key to deploying trustworthy AI, but the increasing complexity and opaque nature of many agent implementations makes this hard. In this work, we propose a Probabilistic Graphical Model along with a pipeline for designing such model -- by which the behaviour of an agent can be deliberated about -- and for computing a robust numerical value for the intentions the agent has at any moment. We contribute measurements that evaluate the interpretability and reliability of explanations provided, and enables explainability questions such as `what do you want to do now?' (e.g. deliver soup) `how do you plan to do it?' (e.g. returning a plan that considers its skills and the world), and `why would you take this action at this state?' (e.g. explaining how that furthers or hinders its own goals). This model can be constructed by taking partial observations of the agent's actions and world states, and we provide an iterative workflow for increasing the proposed measurements through better design and/or pointing out irrational agent behaviour.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Explaining Autonomous Vehicles with Intention-aware Policy Graphs

    cs.AI 2025-05 conditional novelty 5.0 of 10

    Intention-aware Policy Graphs, applied to 830 nuScenes driving scenes, attribute desires and intentions to an autonomous vehicle and produce interpretable global and local teleological explanations, with a reported 75...

Pith tools