Pith. sign in

Paper Citation Record · LEDGER

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models?

As of 7 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 1 inbound Pith citation observation for arXiv:2506.13065.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13065 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:43:08.528481Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T15:31:25.079191Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T15:33:25.520674Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy19
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 35508956-b02a-4755-8f58-3dc3c487d173 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:12.930373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:06.990283Z digest=sha256:3b5f1ef1731eb97e313f8990eab2e637e6c873ed9223a93540c83fd5d0a52a3b

Observation 3c17f80e-9116-49fd-a4f7-4b8307fc33f9 · outbound

This paper cites A Sequence-to-Sequence Model for User Simulation in Spoken Dialogue Systems.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? A Sequence-to-Sequence Model for User Simulation in Spoken Dialogue Systems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.413338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.413338Z digest=sha256:f6d5ea1418e0afb51ddb959115190750d2bef503e3222db1643861f753bdf0b2

Observation e1dfb8d4-a5d3-4b33-8f76-5f98ad7c42d0 · outbound

This paper cites Emergent autonomous scientific research capabilities of large language models.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Emergent autonomous scientific research capabilities of large language models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.473711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.473711Z digest=sha256:e01d2589dffa480f63448520aa013aabd3d90abf32712cf1286f53e02b1c319d

Observation f5353aec-fa41-471f-b8fa-55b6bc307c1f · outbound

This paper cites The distractors must be related to certain parts of the information in the question.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The distractors must be related to certain parts of the information in the question

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:11.685098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.590113Z digest=sha256:5ebe99ddcc300e36a55b37d0a335f915eb5d33eb80bdbcc6178fa9a109c0fb1d

Observation 20551643-e8ac-48af-b96b-677e5edc69af · outbound

This paper cites This is necessary to ensure each question is challenging.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? This is necessary to ensure each question is challenging

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:11.421325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.642886Z digest=sha256:c54950719bd4ff5bc71c5ca0b539f988850e9e80e74762d5693651fe61641599

Observation 1a539ba5-65b4-49da-bbb9-ce6a994edf43 · outbound

This paper cites If they do not, suggest adding the relevant distracting information or modifying the options.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? If they do not, suggest adding the relevant distracting information or modifying the options

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:10.499118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.930913Z digest=sha256:3bdad7149e6d6ec4ceb3b0883b9cab7dbe0c052290b4980a325c4a257f96fd12

Observation ff39faf5-77e3-4a69-892b-c93f7ddf8d9d · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:10.321988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.968105Z digest=sha256:56b3c0db0ac6f8ff323920c4c7a6a0ab84465644b04c56b6ce43fe3e88437edd

Observation 3420088b-b267-4d3a-9d72-8a86c395c3ce · outbound

This paper cites Please provide specific modification suggestions for the question set and give your feedback to the question author in a reasonable tone.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Please provide specific modification suggestions for the question set and give your feedback to the question author in a reasonable tone

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:10.128969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:08.055680Z digest=sha256:09a82b5a7adc112a0d4fd4cfaaacac1fa49f7504b4a63ce070c3dff1917f78bd

Observation db7c632b-8ec6-4419-b7cb-b1f5fecf5a77 · outbound

This paper cites LiveBench: A Challenging, Contamination-Limited LLM Benchmark.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? LiveBench: A Challenging, Contamination-Limited LLM Benchmark

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.828900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.828900Z digest=sha256:ae5f52acfc724d2d35f52d0e9d0850238187c9dec01a8bb24371228624b703fb

Observation 9edd8e6d-ef77-473d-974a-340009677af8 · outbound

This paper cites SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.885426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.885426Z digest=sha256:df1b8729302733084506df06656b50726397a075391f9dcbcba58f192494261f

Observation 087c0df3-dd1b-4d1a-a7db-86ede95fe0fd · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:12.939017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:06.934518Z digest=sha256:d6ba1b80ccf82e5f5cbfd81c09496163ff170e1172054fa212d48ca72c5f46d5

Observation 79f74805-d35e-4822-a887-d59f223b79ea · outbound

This paper cites Question: {Question_Content} Options: {Options} CoT Prompt for Evaluation The following is a {Question_Type}.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Question: {Question_Content} Options: {Options} CoT Prompt for Evaluation The following is a {Question_Type}

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.913951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.106368Z digest=sha256:1a89136807039a088800bf5c30f8aaf942e888ad7ab3f0fe748e9e96c397b843

Observation aa1e3507-a675-4c8f-80eb-74bc52971eb4 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:12.783342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.136703Z digest=sha256:7c5452f637b2b160613c01d9a1ed21aeb7cb5a295e9e760ffbeda83983a91826

Observation 95eb9b0c-ec08-46a3-953c-f56210fcb7e1 · outbound

This paper cites A, B, C, D, E, F.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? A, B, C, D, E, F

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.921343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.196600Z digest=sha256:2d6442b643c0989eeb72999607445645782e4bdd4e6ba0c9a4e89e0fc29305fc

Observation c710a47d-e5b7-4dbd-8c20-ed454e2711f1 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:12.722206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.249369Z digest=sha256:372fb9dfab7b79e963ac1b0f0c51c894561f238e24657b63c2036b81d8f18b53

Observation 6cd9dd59-f3c1-4ea8-8cd4-21fe05c3a290 · outbound

This paper cites The question should not contain any direct description related to the predicted motivation.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The question should not contain any direct description related to the predicted motivation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.556860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.285303Z digest=sha256:b3b618fd7a1cd49be0d637815d748062288055be80e87693b47a67c35bb7107e

Observation 770135ea-b476-473d-9e32-c73d2db2ee5f · outbound

This paper cites The question should not contain any direct description related to the predicted behavior.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The question should not contain any direct description related to the predicted behavior

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.439760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.340325Z digest=sha256:892794ec05420e9910291040f2253b4641b632c3e392bca0b0a70e44efd7831c

Observation f5c20ab1-dffd-4c2d-8346-44a45e97c11b · outbound

This paper cites The question should only include the complex scenario and the character’s profile.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The question should only include the complex scenario and the character’s profile

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.298433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.393854Z digest=sha256:f032f11367b79c659e8693f428b094c8a9d2dc707025ae7116e645d570723aa3

Observation 8ae918ab-fd12-4ce2-87cb-054e2557ac0c · outbound

This paper cites Please rewrite this scenario by cor- recting any logical inconsistencies, and add relevant details to make the scenario, profile, motivation, and behavior more vivid and complex.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Please rewrite this scenario by cor- recting any logical inconsistencies, and add relevant details to make the scenario, profile, motivation, and behavior more vivid and complex

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.194449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.446648Z digest=sha256:ab532ea20f33b4a1c5f08137ee305c01f199e7a2bdc98c84eeb971c6d4c16a4d

Observation 8d36fe18-465a-4713-81cc-9dd2da8644ef · outbound

This paper cites However, ensure that the motivation and behavior are only related to real human needs, not to any POIs or products in the text.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? However, ensure that the motivation and behavior are only related to real human needs, not to any POIs or products in the text

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.009779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.498704Z digest=sha256:544e5e1d22044648a740ae906cab107498f89541b73450b9495146bac83a4f9b

Observation e82ebf1b-7ea1-4f8d-a9e0-dec0bc1035e7 · outbound

This paper cites Therefore, please ensure that each question has enough rich and complex scenario and profile information to support correct reasoning.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Therefore, please ensure that each question has enough rich and complex scenario and profile information to support correct reasoning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:11.843382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.535893Z digest=sha256:8964634f6bfbc711af95459dc7888c8955f45bd6507344a310bedb00bc88aaa1

Observation cec63912-de7a-4c5f-b789-d14f939bbb56 · outbound

This paper cites The motivation reasoning question should include additional behavioral information about the character.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The motivation reasoning question should include additional behavioral information about the character

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:11.284235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.672766Z digest=sha256:825f501a2354b1d38a8d9e4b471a493bcb5faf91ffe6fb7d3e4c0f9fdb69795e

Observation 69957a79-e6f8-4ad5-91d8-628f379fe3c0 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:11.126370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.721601Z digest=sha256:04384be96d1bad355734e43a65ab919a1a78a266b827661b30bb748d9ff597e9

Observation 338d7ffd-2e14-41fc-a359-b53a0bfc8773 · outbound

This paper cites If not, suggest modifications to the scenario or character profile to make the information clearer or more comprehensive.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? If not, suggest modifications to the scenario or character profile to make the information clearer or more comprehensive

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:11.005281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.779326Z digest=sha256:17728eb57e5f40860cb82824b6b9b96a01fe171457fcc578a220e7d3b6ed8f82

Observation b5de10a9-74f6-4f29-bde3-08ced9ab848e · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:10.837908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.830422Z digest=sha256:2598062259b01306fe78024d5c8c458f38f867f20b992bf3a3c4efd72505d7a9

Observation 28d6b908-c370-4144-8929-7d2eb874c42c · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:10.685253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:07.882397Z digest=sha256:18402e25de4c6a0cb4b1a2b4cc5532c3f2e0b28cd71614f19ca5a16f8036e880

Observation efa51cdc-7d40-45da-b9dd-67536a8bfa85 · outbound

This paper cites The question should not include any description related to the predicted motivation.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The question should not include any description related to the predicted motivation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:09.972730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:08.123433Z digest=sha256:891fa0074e3bfef962ba5f09c93880defed2f8505319d970b4bdf0740f4d417b

Observation 315c78b7-171a-40e0-8c17-0823a9400072 · outbound

This paper cites The question should not include any description related to the predicted behavior.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The question should not include any description related to the predicted behavior

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:09.766268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:08.173303Z digest=sha256:b5e8b440158cb8cd63444e2e5e8be04181d90f2fd7048fb74de2094ccf74a9c5

Observation ef08e0bb-8316-430e-9133-a3402584a0b6 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:09.592265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:08.253355Z digest=sha256:ad764abc5aaf17c8274268f65a8e1c05e651d7cf5bc5c55421c5948c26f1f2cd

Observation 14a6bdec-1881-4797-a14f-10c4c2a7777f · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:09.356741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:08.328734Z digest=sha256:b8ca40a1b4fe049fc639a0499904640701b342dd85fa61f433ca5c903a267fc1

Observation fc6598fb-7e8f-4487-b224-c7e20c52c968 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:09.107701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:08.405819Z digest=sha256:416368cbf7fb074a3e439cef299ac975380d6bcb64ff7c8c04d7cc7d762be29b

Observation f69e1631-63be-471c-8d31-56af8534c153 · outbound

This paper cites Respondents should only reason based on the question provided, without seeing any other information.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Respondents should only reason based on the question provided, without seeing any other information

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:08.960846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:08.452146Z digest=sha256:b2b71b9bc576983671dead2122ba91ef72d916cc1251bf22782bc4420c16458e

Observation ab9cadef-d8bf-4422-8d6f-9030e73f35f5 · outbound

This paper cites self-promotion.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? self-promotion

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:08.797069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:08.528481Z digest=sha256:3f691959fbc1036575ef4dafa96db49a684b29c24fccb895f28c060e0af73382

Observation 78b84e43-6e4c-4d62-9ab0-8e01b86a96ed · outbound

This paper cites Clever Hans or Neural Theory of Mind? Stress Testing Social Reasoning in Large Language Models.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Clever Hans or Neural Theory of Mind? Stress Testing Social Reasoning in Large Language Models

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.727480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.727480Z digest=sha256:24789bc9e2d035ff29171364aa6afe1c3742db9820ba0b99f5973935b6c28805

Observation 55be7344-65e9-466b-a782-a9139d670165 · outbound

This paper cites Neural User Simulation for Corpus-based Policy Optimisation for Spoken Dialogue Systems.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Neural User Simulation for Corpus-based Policy Optimisation for Spoken Dialogue Systems

Reference 2010

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:43:08.701502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:06.553376Z digest=sha256:56152b9ecd445d59ecd0452f94fe8639e2a9aabbea5d04f305c2d581dd6e9523

Observation 958c50c7-34da-4976-b1f5-a9465258e5a3 · outbound

This paper cites To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.779316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.779316Z digest=sha256:31bcfb0a62718d0a316297a24755b5d02754c4d72cc437412d24d4c40cc85820

Observation a88f81ec-669c-4dc9-b645-24232aca3cb7 · outbound

This paper cites Towards Social AI: A Survey on Understanding Social Interactions.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Towards Social AI: A Survey on Understanding Social Interactions

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.603376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.603376Z digest=sha256:3cbb8d44fea2a6392b2cd75b0c9b13af12ff9f22a7045b13648609c11f2b941e

Observation cfb77f19-2672-4a00-926a-64ccf72b072b · outbound

This paper cites InInternational Conference on Machine Learning, pages 337–371.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? InInternational Conference on Machine Learning, pages 337–371

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.949017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:43:06.349846Z digest=sha256:c9044051cc6312374ed6bc1c81015aef25bb312551a12a9a45729058adb704cc

Observation 7bfdf243-313a-4357-8b98-54601d3885ee · outbound

This paper cites EmoBench: Evaluating the Emotional Intelligence of Large Language Models.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? EmoBench: Evaluating the Emotional Intelligence of Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.686448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.686448Z digest=sha256:bfc226bcf4504d2ea41caa055456e4bcc4f3a61659c65d6faf8bafeda45bbeac

Pith citing papers

Observation 3eb5105e-3f69-4b94-8008-38b16aec4baa · inbound

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench cites this paper.

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models?

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:33:25.522229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T15:31:25.079191Z digest=sha256:2ec2c0361236c7b0d2cbfc6b67654cd924434defc6ce3f73d5a5a734d259b10c