Pith. sign in

Paper Citation Record · LEDGER

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents

As of 21 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2608.06735.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06735 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:42:05.373272Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9152c2b9-3b2b-458e-92e4-7b6245e2b556 · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.715055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.275537Z digest=sha256:df68baae0e793dc16dc68a8d4ff012464a0a6a95ffa44977af5023bd3b12d36e

Observation f5d32cb0-2825-40e4-b40f-41a8b1a669fa · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.700118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.280364Z digest=sha256:c3bc9ccefd8bf28ba16e97717e41874a80fcbe62b0b807113beb70708d429f49

Observation ffc82e8f-c860-4ae5-bfa2-ee5be51354f4 · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.686731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.286754Z digest=sha256:d88be817762ac2ccc0f3c9e195764dc701c689fe2c434671551e24452467e626

Observation 848057f3-c15b-4700-b066-73dcd34bfd14 · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.674962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.290749Z digest=sha256:f2425e99e60113e57f9eeb5c453d639bbeb3ab112027741e1cd4168c5b718dd0

Observation abe397af-c06d-4860-9490-b99e5d27cca4 · outbound

This paper cites Conversation policy.Use natural, concise, spoken language suitable for a phone call.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Conversation policy.Use natural, concise, spoken language suitable for a phone call

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:42:05.656135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.295187Z digest=sha256:e322b04a12dbbe99f4017ec6aa50a4b609c1251ee086011638134b39a1f5ca79

Observation 511a2d5f-6ba4-4e33-90a9-5305b045cdc1 · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.591747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.324146Z digest=sha256:e9754ffd101d85fa97cf9eada03f62b9b9f36f15202c27b68bb7dccf979a5865

Observation 42deba47-34cc-4dfc-a512-89c5ab3de2b6 · outbound

This paper cites The numbers again denote the units received by your side.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents The numbers again denote the units received by your side

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:42:05.479487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.357085Z digest=sha256:c740ba61aa2dd90711eb3ad5199b71b455987f2827e9c03ad2304f083fd346e2

Observation 57e2d3b5-9f1d-484a-90d4-b915dc64326b · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.464398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.363538Z digest=sha256:a36132b4b150ad258d4057c5b250ff8dc00232ff9224bbb003e9700e1a18ccb5

Observation d6252ad3-9758-4e97-9bb4-d571f7ac2cd2 · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.642733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.302088Z digest=sha256:25fdb8660c516cf42d881b57eb0a381f38775486c2b3c676a886f7f2056aeed4

Observation 9e2f75ca-31fd-461a-9888-1487efc594cf · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.630436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.307124Z digest=sha256:2f9ea95f00c922f4b222bb0c15f56fc1a2488a3f93efed4259f822589113df5a

Observation 45e784fa-6814-49bc-ba22-b08286bc591a · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.618388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.315981Z digest=sha256:812b06a3c73670d5859f1f079e5c137cfe15e04440f3e220597d861f654628a2

Observation 0542f156-00db-45df-8fcc-24f5747c3c59 · outbound

This paper cites 5.ConfirmAdded,IgnoreRequest, andRejectRequestare available only after the customer has verbally agreed to add WeChat.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents 5.ConfirmAdded,IgnoreRequest, andRejectRequestare available only after the customer has verbally agreed to add WeChat

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:42:05.605061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.319782Z digest=sha256:b289d4de82ed58d060482c909616a1b5db07eccc40f75e92eea7c2a8aa0e47c0

Observation 6ed4281b-5d16-4bc6-bfb6-83d777f1bacf · outbound

This paper cites The other participant’s private values are never visible.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents The other participant’s private values are never visible

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:42:05.573671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.328322Z digest=sha256:45d2038534704cb9d2cfd37ef424c7a3222b62b33f75cc925d9e2e373eb72feb

Observation 26f7e311-159b-4010-b7d6-ad0d4b169bcb · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.556899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.332769Z digest=sha256:2ccc1cec2b73485613570c96e3fdb87672d6a81a874b7d32304d2c2e141d3b69

Observation 7582fd19-d2ea-4305-9e4a-2b61382ba7e6 · outbound

This paper cites You may ask a question, make an offer, accept, or compromise.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents You may ask a question, make an offer, accept, or compromise

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:42:05.542793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.337931Z digest=sha256:a61f4acab2c8e8e90661ad0c7bb3f4d7678cd7e7ad55c6549c519d1985abdd57

Observation fdfbc270-b2c4-486a-a0da-e412a0f7bad3 · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.525433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.342394Z digest=sha256:fb0ca37d8cda8c27cbc9f0669802a25a5eea33b50c12211b559e897bdde02bdb

Observation 03669721-d273-4282-8a05-2f8ee2e2f207 · outbound

This paper cites The numbers specify the units received by your side; the counterpart receives the remaining units.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents The numbers specify the units received by your side; the counterpart receives the remaining units

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:42:05.509130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.346471Z digest=sha256:0a71991026bdb145668699bcdeeea442835e49331a58ce8f96dd64411eebf07b

Observation 41f2ded0-5555-41f3-953f-4d0391d1834e · outbound

This paper cites A complete valid dialogue contains at least one<try>allocation and two matching<selection>confirmations, one from each participant.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents A complete valid dialogue contains at least one<try>allocation and two matching<selection>confirmations, one from each participant

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:42:05.493839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.351806Z digest=sha256:9accc10df7f8a56b30c06f77749c1ac25ffc295440f2d6c1fe645833ef8ad5fd

Observation f633128c-439d-48e5-a64c-3e730c8748b8 · outbound

This paper cites an unresolved cited work.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:42:05.450141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.367801Z digest=sha256:70bd7cf2059c8bb70d87a7017cc5ad7ced381d72150871887bfa138d38d66294

Observation 9fa43245-8c08-4c22-a1d7-5c972b62aeea · outbound

This paper cites If it proposes a concrete allocation, include the required<try>allocation.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents If it proposes a concrete allocation, include the required<try>allocation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:42:05.437937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.373272Z digest=sha256:b14acec56babe6f9e110e1911c205a1005882cc14ef74b69668b03fd2e70f10b

Observation a8f233ed-cd6c-40bd-b95f-093db46b5d8a · outbound

This paper cites Wan, Z.; Li, Y.; Wen, X.; Song, Y.; Wang, H.; Yang, L.; Schmidt,M.;Wang,J.;Zhang,W.;Hu,S.;andWen,Y.2025.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents Wan, Z.; Li, Y.; Wen, X.; Song, Y.; Wang, H.; Yang, L.; Schmidt,M.;Wang,J.;Zhang,W.;Hu,S.;andWen,Y.2025

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:42:05.755766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.260985Z digest=sha256:0ac3ce875bb81f9ac99ee6b97ac3ee1300153761bacff05b58a264eb3ba66215

Observation 948c72f5-c49e-4377-b317-a01ccdaa08e2 · outbound

This paper cites get to the point.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents get to the point

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:42:05.731482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T21:42:05.266864Z digest=sha256:9b7902456eecb40ed670853107449584899cfd9acc1fa1e2d447ea9d919b663e

Observation caa4df75-ef76-42cc-8fd1-a90ed92355b4 · outbound

This paper cites OpenAI o1 System Card.

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents OpenAI o1 System Card

Reference 5001

Resolution
unresolved
no resolver link, observed 2026-08-10T21:42:05.253792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:42:05.253792Z digest=sha256:54be7337c267efff3eb43088e2a420fa1c16699d894d4cd13a557f4e588de135

Pith citing papers

No inbound Pith citation observations are available.