Pith. sign in

Paper Citation Record · LEDGER

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems

As of 13 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2411.16305.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16305 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:21:29.811392Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact6
  • verified fuzzy1
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9a1ab78b-4450-499a-bbcd-b08c5a09dfb0 · outbound

This paper cites online" 'onlinestring :=.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.713062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.713062Z digest=sha256:6d8b8c78c1b8e8c317ae5b7c90e9fec515bcfa2d23371d7f12c30731616d6d60

Observation ee639f2c-5ea7-46a5-8173-3a5cea167f1c · outbound

This paper cites write newline.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.716933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.716933Z digest=sha256:e331c4296eafbd1023d9d4b6f01b940cefc7ceafacf880b7286b9dc05a42f6bf

Observation edd2f15f-eeac-4f55-bd68-54ae0e9ce912 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.720707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.720707Z digest=sha256:4665f2b0e77ba045135064bc564b7b16f73d07617c9d9fb89978fe6e92b9d841

Observation 3a0f2923-9c29-473c-860e-1a99341eafe8 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.138011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.723836Z digest=sha256:cc7025b488c82058f263fbf196915200f9165e4463a9e5784834888928f7e523

Observation 370f2bb2-8153-44bb-915d-369bc7023e33 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.727730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.727730Z digest=sha256:4e264b211f0e413bcf6b939bd1e20d39646f417b85373138e7150d37bd7669cc

Observation 85fb7f01-5dfa-4baa-b86f-2a76c5804ca6 · outbound

This paper cites Fantastic Rewards and How to Tame Them: A Case Study on Reward Learning for Task-oriented Dialogue Systems.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Fantastic Rewards and How to Tame Them: A Case Study on Reward Learning for Task-oriented Dialogue Systems

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-12T13:21:30.042810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.731432Z digest=sha256:fdf621c18b207128be256fc7273bc7b38a127be9dffbbcc35631f6b7f8b504c2

Observation d474e5a5-656e-4339-8f74-0770b2699a28 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.734744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.734744Z digest=sha256:1f62972b023c92ae96af1949e4d3340e3d649bdbd9e5ba19a22d47eb3b969dfd

Observation 70d0da9a-847c-4e6d-b030-2a0f79bf87fe · outbound

This paper cites Direct Language Model Alignment from Online AI Feedback.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Direct Language Model Alignment from Online AI Feedback

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.737705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.737705Z digest=sha256:b190dd5a23a89b1f0564613502b1594a4a98973c5bd8bc58b32e856c4b05fa90

Observation e25c4f12-dee5-40eb-b64e-c7a50b3108a4 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.119508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.741474Z digest=sha256:0b9a0623ffcff72315fa018c866e383215f78b0b337f7b62b1d5c82bf6f33c30

Observation b49ed11c-28d4-47cb-8783-72fad357df85 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.110775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.744386Z digest=sha256:4d1f11a75531b10a03f149109265db0ad6b628e9fba2433346f5c3e60b36a09d

Observation f0cadb74-aeed-48e5-90f7-b8e7a72f75e5 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.747563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.747563Z digest=sha256:2f82704507e8245e9899fbdd90f9bdd7cc4380cfc75811b4f2845b5007ee6800

Observation 6a7b6c4a-9986-4f80-8d94-cc473a568bd2 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.750542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.750542Z digest=sha256:79bd420ecb3a69ddb3813dbb709d463d4d90554507e26259c5e02447e1fd8a75

Observation cc4ccdb4-716a-41bc-b090-938c3513ae55 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 13

Resolution
verified exact
doi, observed 2026-08-12T13:21:29.908535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.753658Z digest=sha256:95f449a1435717c7679de99355973fc41b0cad8209e15b43489587123653ddf2

Observation 0ff90f6f-dbd6-4c11-b62f-72f28cfb0af3 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.756846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.756846Z digest=sha256:f1fd2437511cd93e9c0c36f75464e33b1f0738e671376851f516c9d5ac74cd71

Observation 4787ab63-517d-42cd-a98b-94170976ca8f · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 15

Resolution
verified exact
doi, observed 2026-08-12T13:21:29.893734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.760761Z digest=sha256:20484b2f82f1330b5f95ed8cb153fe450a5943043da9b1f1403d74db27078b06

Observation f8bb8ee9-5596-4b00-a7ca-02a87ce3efc4 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.763702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.763702Z digest=sha256:088bed98651d6fdc11c6cf6a5e214c86d3f0fd273326b77ca0c5011ce89416ab

Observation 2b4c96ad-641f-41e5-8ee7-329d9940e8ee · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.096285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.767018Z digest=sha256:425eca10dc440ec487a3c4fc34eea3c0e16c3dc7bb8cf52fc59ed75c5b013b19

Observation 3c5fc5d2-1b52-4598-ab79-2501be555e4c · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.087226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.770322Z digest=sha256:8db4c7f1c3910067e4491472b35d2cfce8cabc6f5a75a3bdef150a586ffc2af4

Observation 14f000d8-c9f2-418c-9441-e58a0610b521 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.773351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.773351Z digest=sha256:3fe28e7fb20245a0339b67e966bf98da60b420edfb9696b890e3029b1d7aa9b6

Observation 18689bd7-5d58-47bb-a92e-76ee5f5ad2b9 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.078402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.777098Z digest=sha256:bf57354f6d4ca76b69cca2d326568200254a4c109272c3d52735b6424b7e765c

Observation 147f44a1-524b-4f2e-a26b-1151440ccada · outbound

This paper cites Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul Christiano.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul Christiano

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:21:30.070122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.780448Z digest=sha256:da53b4bdd643d9dd75226e47e2f50fd70ffc6368d4334caad6aa316a4f9069f5

Observation b64e833c-860b-46d4-acff-a6ac3725ec08 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.783528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.783528Z digest=sha256:2637992025d0c4498aef44d74023ca14daf6ccf4e2644abb3e67e777c9918093

Observation cb389fd4-7381-4313-aac2-5ae22a9c550b · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.786820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.786820Z digest=sha256:c116edb98acd88b7f4236f067cec3fa288bcf261d0657cdf561de4c566d7b491

Observation 58e00e15-7f05-4765-ac24-b6510ddb3bb5 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 24

Resolution
verified exact
doi, observed 2026-08-12T13:21:29.867782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.790007Z digest=sha256:1f44bbadc874893ddabecb2da4013d57653bd4535be3efbf3775d4decd2d0822

Observation 581f2955-01b1-4b5a-9aef-201958dbdffe · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 25

Resolution
verified exact
doi, observed 2026-08-12T13:21:29.857513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.792931Z digest=sha256:621ac592a48d14f15c3fbd7857cd3d7b54528fb5e900765496650f102eca3974

Observation a1d66b31-f1f8-4beb-8e96-8bc4f7c167a0 · outbound

This paper cites Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.795760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.795760Z digest=sha256:4251431731528af8e5849977f36d45ef80e00efbb31f285c9be4d32c228c968d

Observation d3e6ef04-a348-436d-927e-e1944adc3002 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 27

Resolution
verified exact
doi, observed 2026-08-12T13:21:29.847443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.799004Z digest=sha256:3def99df899918bf153eab9105af6ba8b95f6b737bb07955a7c01940a67f3985

Observation 6a00af7b-fc8f-449f-8ac2-b7bd86da7101 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.060366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.801980Z digest=sha256:82e510a71e39058f1e33b4827aa91b4c19e49af625c57d7dbd735921a6d02293

Observation 73bae19c-4601-4a25-835b-ebb84839b14b · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:21:30.051751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T13:21:29.805328Z digest=sha256:1c04e93d914a74686094d9557b97271e50d3d9e054e4b7ea423d0fec5ec20adc

Observation 88ed6aa7-b314-4e67-a241-a07720b594bf · outbound

This paper cites Description-Driven Task-Oriented Dialog Modeling.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Description-Driven Task-Oriented Dialog Modeling

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.808227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.808227Z digest=sha256:3c04b652e6157d732c5483d3c0665beaaa30b1c925c8adb55194bbb016c86f47

Observation 5ba1ee13-5092-4b67-9b48-d6da995e0258 · outbound

This paper cites an unresolved cited work.

Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:29.811392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:21:29.811392Z digest=sha256:a9cac81b84bebddb21844cf85b06f2383aa902b02d90ee6d0fbdadc5e4a18771

Pith citing papers

No inbound Pith citation observations are available.