Pith. sign in

Paper Citation Record · LEDGER

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards

As of 7 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2507.17147.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.17147 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:02:27.120609Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4e08cdc8-3d48-4f94-b62b-dc6259fe8d12 · outbound

This paper cites "Let Your Characters Tell Their Story": A Dataset for Character-Centric Narrative Understanding.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards "Let Your Characters Tell Their Story": A Dataset for Character-Centric Narrative Understanding

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.847892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.847892Z digest=sha256:fe16115c5af949644439dac6b702c495edf7c54b34d600411bcf057fe8d9d596

Observation 6baeb213-9b97-482c-b2d0-cea8eb952187 · outbound

This paper cites Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.853811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.853811Z digest=sha256:875caca1bca87d37fc1aeca54a951feee73557ac395065361298479833981514

Observation 8404f121-5560-4e99-bd44-75e8dbaf0d3f · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.859380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.859380Z digest=sha256:2a66be8a08fe5bde2204ea4c383f3403279f6f6941b0da89a85bd4305dbe8009

Observation eee4d1e4-2fc4-4139-bb67-4553658a309a · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:02:28.315684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:02:26.864407Z digest=sha256:fd3a737c59d0dd27f82ce040fb9a80476e437cadf1a97631315a037592a317b4

Observation 6a359f50-1ac9-471b-9300-3a38ec70366e · outbound

This paper cites SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.869430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.869430Z digest=sha256:8b5066ec118140eef052a141806143d58451fef1ba3484e0b744bc66ccf759df

Observation d1ae54d3-fa04-443b-b95c-98e4bde4b0b7 · outbound

This paper cites The Oscars of AI Theater: A Survey on Role-Playing with Language Models.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards The Oscars of AI Theater: A Survey on Role-Playing with Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.874359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.874359Z digest=sha256:6acc285226eeb16bb164939f3621dd749040132780543a0105e66d71d42d119f

Observation de71f6c6-1874-44a7-8913-4a7fa5d75eda · outbound

This paper cites Large Language Models Meet Harry Potter: A Bilingual Dataset for Aligning Dialogue Agents with Characters.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Large Language Models Meet Harry Potter: A Bilingual Dataset for Aligning Dialogue Agents with Characters

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.881130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.881130Z digest=sha256:4cb5fa4653e19fb536a19866678694688a1c28039ca0f2c5b1ab93e9f67ae8f6

Observation 34fc88e1-6798-46c4-a002-f76722a43669 · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.886510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.886510Z digest=sha256:102a645f42335d679e8ae8ef2894c4b2e4b1b66b88cce806b07098356ff9958d

Observation 8779d4cb-ae39-4f93-ab54-d6954c64d3b0 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.891631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.891631Z digest=sha256:a68bc33c46e235f76f806cb7aeec4eecab299597a1a759693699bee2c725cf60

Observation e2fdbe4c-4f7a-4356-8895-fc782284d778 · outbound

This paper cites ReTool: Reinforcement Learning for Strategic Tool Use in LLMs.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards ReTool: Reinforcement Learning for Strategic Tool Use in LLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.897593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.897593Z digest=sha256:958ebed144bb48ce1606d4d3202bc91922938aad05f9b6ef3323023e746f3f93

Observation 27b42656-2331-4b91-97f9-4384be482d25 · outbound

This paper cites Reasoning Does Not Necessarily Improve Role-Playing Ability.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Reasoning Does Not Necessarily Improve Role-Playing Ability

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.902549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.902549Z digest=sha256:719215d19cbf0cb6d89812304953bac1e5b5e413078709eb9061a010e37cf489

Observation f37ea5c4-dded-4731-a649-4f4bc3453f5f · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.907285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.907285Z digest=sha256:d583f7868f6ac1fe2e7fa3fbbc8b0552eae1d7cf947a78eda50d2d5c191c15a8

Observation d83cbc5b-2a6c-4bea-a3f7-91699ac33105 · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.912224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.912224Z digest=sha256:6f3f5f0e86530a81303f4ddc68b8b26748cd6b6632be1aca083e43ea2e840975

Observation d03403b8-d6dc-437d-bc49-73343c21339e · outbound

This paper cites Enhancing Persona Consistency for LLMs' Role-Playing using Persona-Aware Contrastive Learning.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Enhancing Persona Consistency for LLMs' Role-Playing using Persona-Aware Contrastive Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.917233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.917233Z digest=sha256:9dc7dc48783d860151c57bf58fb058c5e9df2e2e3647ebc0bfc5d0f2a1bf31e3

Observation 257be1cb-b242-431b-9311-391f1800d631 · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.922458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.922458Z digest=sha256:218b805b5918afb8f957498b0c27e85c3252930cf92f2d2e482dc8b9c3d73939

Observation a658fafb-be38-4542-b42f-8a19164d6021 · outbound

This paper cites CodeTool: Enhancing Programmatic Tool Invocation of LLMs via Process Supervision.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards CodeTool: Enhancing Programmatic Tool Invocation of LLMs via Process Supervision

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:02:27.845222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:02:26.927340Z digest=sha256:c735ac64af9d8c0351f68a098411449f7087d0d51bc60923856c3b59cadb257e

Observation 31143d57-8403-4307-9567-3d65736b6bca · outbound

This paper cites S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.932292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.932292Z digest=sha256:78be6183f2a70fe0a7ba0e53e51a955498d7aebcbb25c2a14a609f1ebd1448ff

Observation d52518ef-05f8-4ece-af1d-d3a08849b68f · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.945889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.945889Z digest=sha256:b7a5e1a325833e3de8090b84f9f5296c78f954d3d6d1a92a20f97827326a375c

Observation 8a0aa8bb-b4c2-4c7a-bece-7701bdd1429f · outbound

This paper cites OpenAI o1 System Card.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards OpenAI o1 System Card

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.951446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.951446Z digest=sha256:33e0c458df33e90df7d83e92d0475e223461f32061f6f990d0d343b977750607

Observation be5bbd53-e46e-476f-a306-c6903f25c1c8 · outbound

This paper cites Generative Agents: Interactive Simulacra of Human Behavior.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Generative Agents: Interactive Simulacra of Human Behavior

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.957432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.957432Z digest=sha256:931457e5c1ba84cb9432499fc07518fd2e00a68ee889757316aa1c2e3457d5fd

Observation b79f26c2-cf4b-49dd-854a-eea4dfdfb9ab · outbound

This paper cites ToolRL: Reward is All Tool Learning Needs.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards ToolRL: Reward is All Tool Learning Needs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.964109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.964109Z digest=sha256:4b237a6d99a4054c123c3986dfcd0151a9af7ff614fc5f64be559bf112918dae

Observation e44a1e8b-e94a-4210-a901-687b2bd9ca24 · outbound

This paper cites Sequence Level Training with Recurrent Neural Networks.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Sequence Level Training with Recurrent Neural Networks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.969880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.969880Z digest=sha256:e15cfcb8202df488eb10914ddca9d08e81f66d204cf66b791ff52261ca14d649

Observation 3a714bbd-ba6b-41b8-ba6d-ce8cc4df7a99 · outbound

This paper cites Towards Empathetic Open-domain Conversation Models: a New Benchmark and Dataset.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Towards Empathetic Open-domain Conversation Models: a New Benchmark and Dataset

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.977002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.977002Z digest=sha256:b70ff38def00001004060dc000ceabe2af6da9a439da3edbf47996e4dc5e397e

Observation 82dad115-11ff-4a67-b5ab-5acc990cc05b · outbound

This paper cites Hybrid Group Relative Policy Optimization: A Multi-Sample Approach to Enhancing Policy Optimization.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Hybrid Group Relative Policy Optimization: A Multi-Sample Approach to Enhancing Policy Optimization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.981947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.981947Z digest=sha256:47433007d95b702644a818f0ccf69b04e125b65e36d790a1dab32ed2bb9dd592

Observation 16d1ca7c-6a60-4fed-a258-b9511d9a9a83 · outbound

This paper cites Character-LLM: A Trainable Agent for Role-Playing.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Character-LLM: A Trainable Agent for Role-Playing

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.986704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.986704Z digest=sha256:0ff480b7abc3f278e69e41341e9ae555d25bf108e2e6dccb236c6829f43e2b29

Observation c912c350-cb68-413a-8165-e4826f483f2d · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.991162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.991162Z digest=sha256:0e23cca2e256824df9587ecad6021c5eb9c433e5f8643209d28ce8449cbccdc6

Observation f8c35e5a-56a0-4f74-8a97-ca5fc8a3c988 · outbound

This paper cites Reflexion: Language Agents with Verbal Reinforcement Learning.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Reflexion: Language Agents with Verbal Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.997108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.997108Z digest=sha256:c5efb3c179f1dea0c127e9c6cf25e885f2f2c7538a726ac526d248b5250ffb25

Observation d9188b5c-73c5-433e-ad33-23f967423673 · outbound

This paper cites LLMs are Also Effective Embedding Models: An In-depth Overview.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards LLMs are Also Effective Embedding Models: An In-depth Overview

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.002338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.002338Z digest=sha256:301eed1f6c562c932426d615e6ae3d4209de8bbbd8143698384ceffd10693e5d

Observation 780e735e-0711-4724-87e5-ee6ca8b66a2c · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:02:28.238415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:02:27.007114Z digest=sha256:6103979df94905d1722876218fc24e84bc302a40604f4172fe99fdc4415c962a

Observation 1b9f1258-f911-435d-bc9d-b463dafdd5a4 · outbound

This paper cites LaMDA: Language Models for Dialog Applications.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards LaMDA: Language Models for Dialog Applications

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.012282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.012282Z digest=sha256:051b1a26bb3f01866468682cf5d1f4ab506528463027bbef1e9cef218a2a802d

Observation 074a3ce3-3475-4ce4-aa06-237d8f1ca4fa · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:02:28.219625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:02:27.018200Z digest=sha256:ef6d4b9c3dabcc367807c134e7fe62334d4ba966a8497f64861e00c150d5dfb1

Observation 77f8b97f-9572-4759-8943-c46baed89086 · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.025256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.025256Z digest=sha256:9f4fa7a811bbfc5526e2018b79d8582eb41cfbb3754bcf661f3086b9731a4aca

Observation ba4f8c21-3cfb-443c-9d7e-eadd3004018c · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:02:28.197057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:02:27.029872Z digest=sha256:5aed1537d09cbec7d400d8d5598b5fc51da55f56edd37e7c4699dc27b7bfc81b

Observation 37128dca-b97d-4e47-885c-456b1aca554e · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.034810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.034810Z digest=sha256:1101440247b9e5bc506dc59808e48f4fa158d34a2747f390d02d17616c506ded

Observation f8c65467-306b-4c3d-b1ba-4d81bb74ca5a · outbound

This paper cites InCharacter: Evaluating Personality Fidelity in Role-Playing Agents through Psychological Interviews.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards InCharacter: Evaluating Personality Fidelity in Role-Playing Agents through Psychological Interviews

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.041691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.041691Z digest=sha256:1bff7c7f04bc4b41bc810035c166b1b5f9ec9de01bdb3f38e4ee9e429a326e61

Observation 9c47933e-6571-489f-9442-701dc090511f · outbound

This paper cites RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.047188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.047188Z digest=sha256:3a7303cce654f92984d0950d071193b3cfd4b91dd116474a7160e4b8672980ba

Observation 4cc375d0-a90f-443d-914e-9ca25978bceb · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.052414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.052414Z digest=sha256:239fc6ea9cbeea135be75a14a0db34e77a1a4f222e2f2416569883817705fc03

Observation 45ce1ce7-8878-4f08-8668-00c68c09cf87 · outbound

This paper cites an unresolved cited work.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:02:28.178487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:02:27.057998Z digest=sha256:f10a222b1b1aab547d1987cc9c811a9ed1f3ca1a6c54ddda2616cc4dd31f050d

Observation 5a833151-3a8c-444a-bd4c-2785aed6747d · outbound

This paper cites Towards Enhanced Immersion and Agency for LLM-based Interactive Drama.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Towards Enhanced Immersion and Agency for LLM-based Interactive Drama

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:02:27.357082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:02:27.064815Z digest=sha256:ee8eaf691295e0ad574497d19ec16d09123741fb2093e978e9e25db58214b0b2

Observation eb81e13e-6023-4429-91fb-db29af084e8a · outbound

This paper cites Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.070162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.070162Z digest=sha256:390d37a348c920acbe2f6b8b1f288187e38fde1267b7d00e1bc180ced9271451

Observation 13c7b42e-b213-4504-a242-79b5da48fc55 · outbound

This paper cites Guess What I am Thinking: A Benchmark for Inner Thought Reasoning of Role-Playing Language Agents.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Guess What I am Thinking: A Benchmark for Inner Thought Reasoning of Role-Playing Language Agents

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.075822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.075822Z digest=sha256:f69d3048970b352e8bcbec86fd6fc216c87e0243957b330c631669f4c6c0f6d6

Observation c0defd52-837d-4df0-bb27-fd32f605d5ff · outbound

This paper cites Character is Destiny: Can Role-Playing Language Agents Make Persona-Driven Decisions?.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Character is Destiny: Can Role-Playing Language Agents Make Persona-Driven Decisions?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.082286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.082286Z digest=sha256:8cb39b8447991dc48d176275f58e3388a4d46e05e1d079e4302059097db767d6

Observation df9319ef-3d34-419a-9d45-e224e83e0dbc · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.091819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.091819Z digest=sha256:91144b87e11d0b03a7a30fe20114c01856b7eb93255209e7ba9ac0a5087dce23

Observation 6712f72c-af34-4fb2-b409-a4feb87d338e · outbound

This paper cites Few-Shot Character Understanding in Movies as an Assessment to Meta-Learning of Theory-of-Mind.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Few-Shot Character Understanding in Movies as an Assessment to Meta-Learning of Theory-of-Mind

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.097864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.097864Z digest=sha256:837f593960c93d86a9588017d1b0ee5bf871ee93446ff242e3a6ad1026831974

Observation e5b23e0c-4134-49fd-acbd-d751861cbbbe · outbound

This paper cites Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:02:27.228027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:02:27.103368Z digest=sha256:64f2b52e3e563f4e5a0797e687c671a7ee11cd53a354799f0d1694788afd58dd

Observation cd211747-996a-4f65-94cd-623efe742dce · outbound

This paper cites CharacterGLM: Customizing Chinese Conversational AI Characters with Large Language Models.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards CharacterGLM: Customizing Chinese Conversational AI Characters with Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.109104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.109104Z digest=sha256:9b8ffaa9b472d6772556d99cbaa157be99fb419fb2ce602975f58d2ea6ac2e24

Observation 43d9a8bf-47c6-4f54-b457-e72406ae6de8 · outbound

This paper cites online" 'onlinestring :=.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards online" 'onlinestring :=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.114741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.114741Z digest=sha256:b42bc3026a2ca60cd3d2b52204bdf6675f0693a24c78af9bc74b3922521f196f

Observation 9c3c8003-d20e-40c4-9cd0-5422385fcd77 · outbound

This paper cites write newline.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards write newline

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:27.120609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:27.120609Z digest=sha256:f1a5cf5b6d74d4e8957a79a6090e859bb31c184a0e576f8c91f3454a718e9759

Pith citing papers

No inbound Pith citation observations are available.