Pith. sign in

Paper Citation Record · LEDGER

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems

As of 14 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2411.16305.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16305 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:21:29.811392Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact6
  • verified fuzzy1
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9a1ab78b-4450-499a-bbcd-b08c5a09dfb0 · outbound

This paper cites online" 'onlinestring :=.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.713062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.713062Z digest=sha256:c84a1fbfb03e33d4fd3c64b05c63ff7b9636cbca93aefe257f3a73ed944b0556

Observation ee639f2c-5ea7-46a5-8173-3a5cea167f1c · outbound

This paper cites write newline.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.716933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.716933Z digest=sha256:aa42259f39f8a1a43bc1e0a6e10fd884fec6119388018061d1698e029f144f65

Observation edd2f15f-eeac-4f55-bd68-54ae0e9ce912 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.720707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.720707Z digest=sha256:0b7066b5ef03733bc73c81a6d034c76d062822e2c0100d888b71a55679f6e34e

Observation 3a0f2923-9c29-473c-860e-1a99341eafe8 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.138011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.723836Z digest=sha256:36cbae635192998d811c1ee6f1e08a4980bd9b3d2bad30fd3a4841e1ae20d1d8

Observation 370f2bb2-8153-44bb-915d-369bc7023e33 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.727730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.727730Z digest=sha256:a550347b08bc23beb8a1d1f8e4eda56f6bed6b0f8c76fca23e1104405afbfb54

Observation 85fb7f01-5dfa-4baa-b86f-2a76c5804ca6 · outbound

This paper cites Fantastic Rewards and How to Tame Them: A Case Study on Reward Learning for Task-oriented Dialogue Systems.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Fantastic Rewards and How to Tame Them: A Case Study on Reward Learning for Task-oriented Dialogue Systems

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-12T13:21:30.042810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.731432Z digest=sha256:b8bbd0de2ccf1fee31c5f169a046be9bc3d86ac842af6235c08b42f7ced11c42

Observation d474e5a5-656e-4339-8f74-0770b2699a28 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.734744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.734744Z digest=sha256:0e56259e421e06a0872836b31e30224ced4c5d4bcefdd4da2b87d7bb6ba2a769

Observation 70d0da9a-847c-4e6d-b030-2a0f79bf87fe · outbound

This paper cites Direct Language Model Alignment from Online AI Feedback.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Direct Language Model Alignment from Online AI Feedback

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.737705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.737705Z digest=sha256:e8f7d748eb9aa7c704838ec8d9b3cb83a8a3ef3c4ab673f36544c08b3fa4637b

Observation e25c4f12-dee5-40eb-b64e-c7a50b3108a4 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.119508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.741474Z digest=sha256:a6f46f5c58de6d5ef9afbf499d0226f73f6f33758015fe9e529137cb5c4b788f

Observation b49ed11c-28d4-47cb-8783-72fad357df85 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.110775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.744386Z digest=sha256:d1e37ff3ef3346b6c0b2e8f16640d5800d61c86b70ca443ceacfdf08d2578ccb

Observation f0cadb74-aeed-48e5-90f7-b8e7a72f75e5 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.747563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.747563Z digest=sha256:89b290384822e611c0fba009ae5ec480206a815494c3474bad8940728f08203c

Observation 6a7b6c4a-9986-4f80-8d94-cc473a568bd2 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.750542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.750542Z digest=sha256:2367e58137dafee6eced7492e1c8beea08a715256f435a8f8fdf8a880a2b71e2

Observation cc4ccdb4-716a-41bc-b090-938c3513ae55 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 13

Resolution
verified exact
doi, observed 2026-08-12T13:21:29.908535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.753658Z digest=sha256:364350e9530d8bd7ad746c38876c4045c7836f46b0bd28073af43e40c11e4d6d

Observation 0ff90f6f-dbd6-4c11-b62f-72f28cfb0af3 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.756846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.756846Z digest=sha256:f588b8084c5ef44afbbfcb7fba2fcc351d6d7227e964544c53e483c9c8564bf8

Observation 4787ab63-517d-42cd-a98b-94170976ca8f · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 15

Resolution
verified exact
doi, observed 2026-08-12T13:21:29.893734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.760761Z digest=sha256:ea2c54a297a0fc6c3feb529b0de00436d67def7dcc71b17dc3b4b0b0276bd482

Observation f8bb8ee9-5596-4b00-a7ca-02a87ce3efc4 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.763702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.763702Z digest=sha256:27a18e9f9187445c0b6bb39aa6c08fa8608e24d57cd45b3608a2ff31306c21b7

Observation 2b4c96ad-641f-41e5-8ee7-329d9940e8ee · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.096285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.767018Z digest=sha256:4ca521f2a4ffebe8807b5334b76a870a7785df76a418269a93a3604ce7cc6011

Observation 3c5fc5d2-1b52-4598-ab79-2501be555e4c · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.087226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.770322Z digest=sha256:4d3b5593c23f3dceac71ad556c947fc1dd47dec0ae05775f1f0c32e11bc054c8

Observation 14f000d8-c9f2-418c-9441-e58a0610b521 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.773351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.773351Z digest=sha256:b85e709095ae73112a1caa93de51ae0255137dfaefc9a15b020bdcc4f9bf28bb

Observation 18689bd7-5d58-47bb-a92e-76ee5f5ad2b9 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.078402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.777098Z digest=sha256:05607f852a6830c54d7803d1f89b3f7f670d6cc103f4cd3a46bb31617013594c

Observation 147f44a1-524b-4f2e-a26b-1151440ccada · outbound

This paper cites Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul Christiano.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul Christiano

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:21:30.070122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.780448Z digest=sha256:8315127c06002f6980e5a3f3896de8384a927d27d3b882027854b656870d0489

Observation b64e833c-860b-46d4-acff-a6ac3725ec08 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.783528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.783528Z digest=sha256:d050ce7e8ac792757b9c046a218f1e3446c3bb86dda322eee0c1a59513254cf0

Observation cb389fd4-7381-4313-aac2-5ae22a9c550b · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.786820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.786820Z digest=sha256:fdaa4e4cd35c641faf94c74f463ed7a91f468aa2a8d0463d6a0e06154f3e5432

Observation 58e00e15-7f05-4765-ac24-b6510ddb3bb5 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 24

Resolution
verified exact
doi, observed 2026-08-12T13:21:29.867782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.790007Z digest=sha256:16823812f0c65fb434b0776a2aa0fccb1150343edcebab390ed1de848057faed

Observation 581f2955-01b1-4b5a-9aef-201958dbdffe · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 25

Resolution
verified exact
doi, observed 2026-08-12T13:21:29.857513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.792931Z digest=sha256:ad5b682070624a6df4e96d2b7e3a2b43487d8599e8341868e1eb67f468ff821c

Observation a1d66b31-f1f8-4beb-8e96-8bc4f7c167a0 · outbound

This paper cites Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.795760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.795760Z digest=sha256:e8723a722fd957868377f17a51e0c016c5f107e14be7e5a3c6aaff3a0e915c17

Observation d3e6ef04-a348-436d-927e-e1944adc3002 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 27

Resolution
verified exact
doi, observed 2026-08-12T13:21:29.847443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.799004Z digest=sha256:9c0facd394e6a91a1ecf1d880550414962020750c373b42147cedce6ddfe1491

Observation 6a00af7b-fc8f-449f-8ac2-b7bd86da7101 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.060366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.801980Z digest=sha256:75ae0ee09a922a2dc07054f388182c7b7ded9a6c8f8b640630acf88189fc48a5

Observation 73bae19c-4601-4a25-835b-ebb84839b14b · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.051751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.805328Z digest=sha256:755e071dd0685b670fb619c8f3a59e9b5e176363029e770fa8ad7cbd156dc6c2

Observation 88ed6aa7-b314-4e67-a241-a07720b594bf · outbound

This paper cites Description-Driven Task-Oriented Dialog Modeling.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Description-Driven Task-Oriented Dialog Modeling

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.808227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.808227Z digest=sha256:52e556d72cb9a80fae86b885f6d23900aaf53176d0f15031f5dd6e7f20533c87

Observation 5ba1ee13-5092-4b67-9b48-d6da995e0258 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.811392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.811392Z digest=sha256:98deb43def7c8c4509478ec2c22c8ae219d21e8d5d5ef5e7ebac4f5b29a26510

Pith citing papers

No inbound Pith citation observations are available.