Pith. sign in

Paper Citation Record · LEDGER

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs

As of 9 August 2026, this Paper Citation Record lists 100 of 105 outbound references and 0 inbound Pith citation observations for arXiv:2506.00577.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00577 v1

Coverage vector

measured 100 of 105 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:06:31.530153Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 105 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved78
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f694234f-5dec-43c7-befb-205ba2169f5f · outbound

This paper cites Coop- eration, competition, and maliciousness: LLM-stakeholders interactive negotiation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Coop- eration, competition, and maliciousness: LLM-stakeholders interactive negotiation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:21.951181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:21.951181Z digest=sha256:ca9d0af3fb36ee5922e15b1f0afc035f1fab18099a44e16fb6470d9db4ec23af

Observation 443cce06-abcb-4c99-a6d0-97d8c785eb34 · outbound

This paper cites Playing repeated games with large language models.Nature Human Behaviour, pages 1–11, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Playing repeated games with large language models.Nature Human Behaviour, pages 1–11, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.052417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.052417Z digest=sha256:b04dbfca7024ad2db5a20ce20d203d7198ff2fc2028a1c8909b1e49d10391272

Observation e2931e38-8a20-4bb6-b2bd-bff93c00e524 · outbound

This paper cites Mechanistic interpretability for AI safety - a review.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Mechanistic interpretability for AI safety - a review

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.136800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.136800Z digest=sha256:b225b16e2f1a375692663e9a1ec03a083ad1d81ebefeb495294899be74d75a2b

Observation 210504bd-cb33-47f0-bff2-3653adb85416 · outbound

This paper cites Token Merging: Your ViT But Faster.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Token Merging: Your ViT But Faster

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.218936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.218936Z digest=sha256:4c65559b49c5cfb5e375fc0c3638b49276cc4e4653666a4d60e3ab04aac09c35

Observation 43b7d5e5-a7f6-4964-b224-0ebb8b06d1c8 · outbound

This paper cites Sparks of artificial general intelligence: Early experiments with gpt-4, 2023.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Sparks of artificial general intelligence: Early experiments with gpt-4, 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.273793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.273793Z digest=sha256:b7bb87efa8acfcef069509ae0ad993481b08d5987de749a0ecfc97eb2c81f05c

Observation f0ce41f8-077f-4755-a7a4-ecc180ad2fcf · outbound

This paper cites Cambridge University Press, 2006.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Cambridge University Press, 2006

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.352074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.352074Z digest=sha256:32c62260f968bc2131fe07f88be940a12711cc3ce9c1f223559f4b11cea9d458

Observation dac041fb-c4cd-4cf8-9ea0-13761f3a6c74 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.421425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.421425Z digest=sha256:cf8f21d168e63d5e22f6bfd794f0105429600d92936ac535241997e1a298c327

Observation 569cab53-abb2-4a2a-95e1-70a634fb5e9b · outbound

This paper cites The computational limits of state-space models and mamba via the lens of circuit complexity.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs The computational limits of state-space models and mamba via the lens of circuit complexity

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.491870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.491870Z digest=sha256:a886918a4c72069f5edeaa003f2d4237d2fb8b63e5d0c6684725e5dcc758ede8

Observation 23407aeb-deec-4e61-8efd-6ab0c3408842 · outbound

This paper cites Universal Approximation of Visual Autoregressive Transformers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Universal Approximation of Visual Autoregressive Transformers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.588160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.588160Z digest=sha256:4a06aa80204ac97bcc7d0b01aa3493c6f3d5ee4393f3a01907192f87929c75c3

Observation 38411d10-08ef-4cab-b791-44b800942f0d · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.681246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.681246Z digest=sha256:673af2267bf89d1bea3832fea47b2d270b3a5d2af90c83518c14dff9b228f9ba

Observation b289405f-e8b5-4c51-a684-04a12f98d677 · outbound

This paper cites Gamebench: Evaluating strategic reasoning abilities of LLM agents.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gamebench: Evaluating strategic reasoning abilities of LLM agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.751511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.751511Z digest=sha256:7c61fa6d505819148b537b92fdbeef12460196d6551c7274e5d984ddf52d06c7

Observation 18f51814-d2e8-435f-a617-5a56d97353a1 · outbound

This paper cites Learning to Estimate Shapley Values with Vision Transformers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Learning to Estimate Shapley Values with Vision Transformers

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.865547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.865547Z digest=sha256:9bca13e188c1276ad00f770b2b0d0b203cecb052d3ec9258eb597c91149aa31a

Observation 58d951ef-d171-4c18-ac61-b4b26545e6e6 · outbound

This paper cites Unsloth, 2023.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Unsloth, 2023

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.955760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.955760Z digest=sha256:4565c686f8108e3d0fef75a98fa5517eb160bb79525c643e1e7f093e57ec86a7

Observation 729722a4-164e-4911-9b9c-1e0f97919eb5 · outbound

This paper cites Evaluating language model agency through negotiations.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Evaluating language model agency through negotiations

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.018165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.018165Z digest=sha256:2c9878804837b0beb88e80df1cef2da70d8609c1a2cf8c1174cf11863d2186f4

Observation 057366c2-69bc-4bc3-9328-bf277a615388 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.094647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.094647Z digest=sha256:5ebb4a000acc2252817f211d64ec47a1d6f9235e7047ecab1ecd37bf0abb7979

Observation 935647c9-1fc1-4056-9afa-59307dc08514 · outbound

This paper cites A survey on the optimization of large language model-based agents.arXiv preprint arXiv:2503.12434, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs A survey on the optimization of large language model-based agents.arXiv preprint arXiv:2503.12434, 2025

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.184862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.184862Z digest=sha256:59629f06ea475c2d1d99d64bb75fbd565f70bd6fab858fac95e9067cf73b8a00

Observation 78462c23-68fd-4afa-90ea-9abe3336a673 · outbound

This paper cites GTBench: Uncovering the strategic reasoning capabilities of LLMs via game-theoretic evaluations.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs GTBench: Uncovering the strategic reasoning capabilities of LLMs via game-theoretic evaluations

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.263043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.263043Z digest=sha256:a102566dfe2f7fb3a12fbd50fe312c4e743cb33b0d0abc050fe0fe970fd620e5

Observation d170b8d1-0e69-49af-b9ce-1c1e129cb80b · outbound

This paper cites Can large language models serve as rational players in game theory? a systematic analysis.Proceedings of the AAAI Conference on Artificial Intelligence, 38(16):17960–17967, Mar.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Can large language models serve as rational players in game theory? a systematic analysis.Proceedings of the AAAI Conference on Artificial Intelligence, 38(16):17960–17967, Mar

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.351443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.351443Z digest=sha256:a867d5cdb5c244b1d7f6a791ac58ceb0faf679ac9c434b34f4c8ba0059decd0e

Observation 08d21ab5-7251-497d-8290-82d8e73e003b · outbound

This paper cites How far are we from agi: Are llms all we need?Transactions on Machine Learning Research, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs How far are we from agi: Are llms all we need?Transactions on Machine Learning Research, 2024

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.407980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.407980Z digest=sha256:d8e4d2125095585a4c4828abb25942ab1eb77bf5da9e0deb0b2c626a4f3f1bf3

Observation 2f0c5408-b06e-4ac8-80b9-f23d6c96e868 · outbound

This paper cites Dataset with 200 million 3-by-3 strategic games for comparing perfectly transparent equilibria with nash equilibria, 2020-10-07.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Dataset with 200 million 3-by-3 strategic games for comparing perfectly transparent equilibria with nash equilibria, 2020-10-07

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.502734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.502734Z digest=sha256:3328b796a3dca697093bb9c414161cc159743cb5b503dde046db23f67be4ffae

Observation 2a44a10b-a7ca-42d9-bd40-fc6966575dc7 · outbound

This paper cites Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.589925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.589925Z digest=sha256:0254ed8993c6ba68f815d051a6c88e4bd2782c05f3d9eea00a171aeccead9c1d

Observation 359e17ef-f1be-4c9f-9182-3869723f85f5 · outbound

This paper cites The llama 3 herd of models.arXiv e-prints, pages arXiv–2407, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs The llama 3 herd of models.arXiv e-prints, pages arXiv–2407, 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.661807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.661807Z digest=sha256:1acf9026b28433950a570f0c6ef0155e4b3d3f56c7c187f87a414cc5c3469186

Observation d23c612c-b3be-4b30-95c7-0ce5d8316d40 · outbound

This paper cites Econnli: Evaluating large language models on economics reasoning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Econnli: Evaluating large language models on economics reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.754861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.754861Z digest=sha256:e55b9d32adf5328d6d2a70eaf3efec46a177ff38f9e83ef127ed0acfce0a0dd4

Observation be29577f-849e-4dae-9470-5c184001a066 · outbound

This paper cites To- wards lossless dataset distillation via difficulty-aligned trajectory matching.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs To- wards lossless dataset distillation via difficulty-aligned trajectory matching

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.847212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.847212Z digest=sha256:0ad1d870d728c1c0f56b0372e1ce620cceab19058bf904fba603ca4d8d6f2040

Observation b76e24e1-81c4-41bd-a7a2-3ff195e5c8ca · outbound

This paper cites A multi-llm-agent-based framework for economic and public policy analysis.arXiv preprint arXiv:2502.16879, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs A multi-llm-agent-based framework for economic and public policy analysis.arXiv preprint arXiv:2502.16879, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.967091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.967091Z digest=sha256:cf9887829654038e07d1fc59520de703315db04c1a4b8a79ee4da034f86b49a4

Observation 6b5ffdc1-e762-4f38-8c78-caeab0cec311 · outbound

This paper cites Measuring mathematical problem solving with the MATH dataset.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Measuring mathematical problem solving with the MATH dataset

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.064982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.064982Z digest=sha256:1dbb0835e74939eb8ccfe1820a917fd47a19f03cc01ece6e8bc3969d186e05cc

Observation 62d4e287-69ae-44f6-9a94-d2d28fc662be · outbound

This paper cites Training Compute-Optimal Large Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Training Compute-Optimal Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.136956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.136956Z digest=sha256:3c936034db638722a862840a6964ac37f6e4a792b602cb82cf08067bd3a90d7e

Observation d7427233-0f6d-49d8-aee7-92997177b82b · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1 (2):3, 2022.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Lora: Low-rank adaptation of large language models.ICLR, 1 (2):3, 2022

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.214408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.214408Z digest=sha256:d5a187f0d27842a149310e06200972c03f4a1c1e23f2bf876785b2ba52313164

Observation 832c3562-b10c-46b6-bd30-aa450877c049 · outbound

This paper cites Game-theoretic LLM: Agent Workflow for Negotiation Games.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Game-theoretic LLM: Agent Workflow for Negotiation Games

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.323823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.323823Z digest=sha256:2b28b51cad30ab448aba15cf4a44dcadc5b68f358fb821a296b68013f4a9948c

Observation 5f413b8b-5f10-4bf3-a094-4be155a82a71 · outbound

This paper cites Scaling Laws for Neural Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Scaling Laws for Neural Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.445884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.445884Z digest=sha256:6e235583b200e1dacc083b96c8718f38f34c85d5e48ea1cd8a2d984dc8e1d63e

Observation 88d3205a-c42b-49d8-9a7a-7580d0b0664c · outbound

This paper cites Large language models are zero-shot reasoners.Advances in neural information processing systems, 35:22199–22213, 2022.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Large language models are zero-shot reasoners.Advances in neural information processing systems, 35:22199–22213, 2022

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.549727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.549727Z digest=sha256:2c2bc8f3d617618e4cad022879e7e3528ceaedad28876595948b3689aeb56d90

Observation 4c7422eb-915f-4c85-9266-3449b88e1a37 · outbound

This paper cites Scaling Laws for Precision.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Scaling Laws for Precision

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.610518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.610518Z digest=sha256:fb59e6a3e7628c13b3eeac7f3aafd1df9911b601356ef806e01af4042b1ecde0

Observation dd143728-d871-4e76-acc2-87fdb62feda1 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gonzalez, Hao Zhang, and Ion Stoica

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.686532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.686532Z digest=sha256:b44622be38451cae90e0fcc9dc7c7b348ec6c3791ea85594e5a26a29c6171874

Observation 87cab2af-d06a-4866-bc01-8119fe4218ef · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.761893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.761893Z digest=sha256:ba9786db2731405a98936cb520cc518e62d3809cd074375fca16fbae9123cb2a

Observation 6046f0a2-ce3f-42a3-9ec0-a638152b5b75 · outbound

This paper cites Investigating Tax Evasion Emergence Using Dual Large Language Model and Deep Reinforcement Learning Powered Agent-based Simulation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Investigating Tax Evasion Emergence Using Dual Large Language Model and Deep Reinforcement Learning Powered Agent-based Simulation

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:06:32.997747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:24.846435Z digest=sha256:e050571f2a128aad42fa4666c41fd4dcc47c418668449f0d251045e45a74e3a0

Observation 53e5b513-16c3-4520-a731-5a8ad4c56db7 · outbound

This paper cites CAMEL: Communicative agents for "mind" exploration of large language model society.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs CAMEL: Communicative agents for "mind" exploration of large language model society

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.941592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.941592Z digest=sha256:29773dd936c059ddb1ca09f0067d59bd4c9b2b15dc29f63c7179f5d757ce4972

Observation 9b0e3d55-1701-44f7-9195-7a06fea9b1a5 · outbound

This paper cites EconAgent: Large language model-empowered agents for simulating macroeconomic activities.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs EconAgent: Large language model-empowered agents for simulating macroeconomic activities

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.995146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.995146Z digest=sha256:17876fc4bd52573177e9d96402aa72f32119829142d1f30532b378fb828e99d2

Observation 622b3240-6a49-4b5d-ba52-4a2c48cb0687 · outbound

This paper cites LIMR: Less is More for RL Scaling.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs LIMR: Less is More for RL Scaling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.160677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.160677Z digest=sha256:799bc8dd06de1d984068f105984270d293832315f44d4731b7b972c18871bba5

Observation c9e7634b-437a-4278-82b9-d9489b0af0f2 · outbound

This paper cites Beyond linear approximations: A novel pruning approach for attention matrix.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Beyond linear approximations: A novel pruning approach for attention matrix

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.227804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.227804Z digest=sha256:eae6d30a85d621c62a00d5379a632df585f95a1470ee6bc14251c5f29c44bfbc

Observation 8f8e26c4-5f19-4551-a456-634f196ae0a8 · outbound

This paper cites Looped relu mlps may be all you need as programmable computers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Looped relu mlps may be all you need as programmable computers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.334573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.334573Z digest=sha256:17d4bed163596438fcf0ec022c8bdf8c08e3d19d56c288e10c64c2c2713b9459

Observation e61494a9-9d23-4fd6-a885-542d46411b31 · outbound

This paper cites MARFT: Multi-Agent Reinforcement Fine-Tuning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs MARFT: Multi-Agent Reinforcement Fine-Tuning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.439085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.439085Z digest=sha256:6fe9d76c86cd0be538c0f2fd9aabcd4be88d0c1f79b8b16f6a92b302bab4b51a

Observation e3dfada5-9506-4340-81e2-ead41ad9199a · outbound

This paper cites Let’s verify step by step.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Let’s verify step by step

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.554524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.554524Z digest=sha256:3a43a6f6ff2ccd99181f03e34523b529243d8544c91b33373bf47e9adac15f56

Observation 82868fab-5092-4b42-a382-667d5d7ee9d9 · outbound

This paper cites Awq: Activation-aware weight quanti- zation for on-device llm compression and acceleration.Proceedings of Machine Learning and Systems, 6:87–100, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Awq: Activation-aware weight quanti- zation for on-device llm compression and acceleration.Proceedings of Machine Learning and Systems, 6:87–100, 2024

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.665528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.665528Z digest=sha256:5b4d4f13a56cabef54e4e1165b12153fcb3e8f75dab1d782259ee41f828f4963

Observation 6b97be4f-8fef-4016-847b-682b060c615b · outbound

This paper cites Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.772527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.772527Z digest=sha256:a489ff6730f38d622f98df5bc158f5499bd5136060900a7b09956fceaaa22ef4

Observation ff4f1b8e-554e-4206-b230-010114d1bf38 · outbound

This paper cites Shifting ai efficiency from model-centric to data-centric compression.arXiv preprint arXiv:2505.19147, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Shifting ai efficiency from model-centric to data-centric compression.arXiv preprint arXiv:2505.19147, 2025

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.913312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.913312Z digest=sha256:bd3c414216be271c52159ab93c28b07d7d96ed00a155a6a5e72acebccca2d397

Observation 51d64124-5b14-4299-95a6-a225a23de0af · outbound

This paper cites Fin-r1: A large language model for financial reasoning through reinforcement learning.arXiv preprint arXiv:2503.16252, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Fin-r1: A large language model for financial reasoning through reinforcement learning.arXiv preprint arXiv:2503.16252, 2025

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.001796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.001796Z digest=sha256:fe9030bc59bf1a2c91d125874e9ee52f0716bf2fcc9bcc43c80be776bbbb66ac

Observation cd71bb0a-2636-4789-b4ce-25546c1b8c53 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Understanding R1-Zero-Like Training: A Critical Perspective

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.113033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.113033Z digest=sha256:0e322da696310743de669e40133f12a96b3c8c26ec4fd42a1e0269be143eb591

Observation 369c1bfb-1127-45e2-81a2-06fafbcf2e37 · outbound

This paper cites Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.263283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.263283Z digest=sha256:f6cca506f4212ba2cdd4a34fa5b34416dc803817d14494278bf4ec4067c67856

Observation b3cd4cae-b09e-4f27-af57-b966764d8f8c · outbound

This paper cites The llama 4 herd: The beginning of a new era of natively multimodal ai innovation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs The llama 4 herd: The beginning of a new era of natively multimodal ai innovation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.368480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.368480Z digest=sha256:a07d30077edc23ef1b038ef7b9b8a503bffa47abd87a7fb7b77b563ee55cb5f1

Observation 5c810c38-5696-4ffa-b85f-d853269205ff · outbound

This paper cites Sql-r1: Training natural language to sql reasoning model by reinforcement learning.arXiv preprint arXiv:2504.08600, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Sql-r1: Training natural language to sql reasoning model by reinforcement learning.arXiv preprint arXiv:2504.08600, 2025

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.523847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.523847Z digest=sha256:674a3e3119a53ab0daff97e01f37c7a9ff6410870e149482e8cbf361e9c63972

Observation badcbb50-571f-427e-a078-27bfc2ac3e72 · outbound

This paper cites American invitational mathematics examination - aime.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs American invitational mathematics examination - aime

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.930734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:26.649702Z digest=sha256:6b7b0572a76b9e3b67b46376f0d69523de7bf4a903bfa8e8e63107ff360a221f

Observation b69bc38b-9a00-4c10-bef3-72f236b54b3f · outbound

This paper cites Tractable multi-agent reinforcement learning through behavioral economics.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Tractable multi-agent reinforcement learning through behavioral economics

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.753318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:26.747299Z digest=sha256:b7d99e0d1eb413ca67d2e9b8d31c4c5c5229f1ab9724b2dac8406690ee7c6187

Observation 5add7113-ec19-4883-a36c-48466dd1d472 · outbound

This paper cites s1: Simple test-time scaling.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs s1: Simple test-time scaling

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.861404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.861404Z digest=sha256:781ee610c4d1d6b06e5c1c259d42ec568f67b296a01d956d1de1bb271222ee27

Observation 12090523-9901-471d-b2e4-122ba93475f2 · outbound

This paper cites Introducing chatgpt.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Introducing chatgpt

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.575161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:26.961132Z digest=sha256:fe01c6747f732e0cac5334effb6b2e07a8081cabb57b2155ad260b57d6b8c17f

Observation b845319f-0298-4e2f-beec-612482c53615 · outbound

This paper cites Hello gpt-4o.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Hello gpt-4o

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.395650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:27.048192Z digest=sha256:75443cf2405f9da07061a8bcd50d34da7095ae8b654089e29d89d3057b94456e

Observation a6f3e788-7ece-44ee-bf4a-a250b80ba323 · outbound

This paper cites OpenAI o1 System Card.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs OpenAI o1 System Card

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.119384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.119384Z digest=sha256:1c61bae700ac5395e94a5129b65921474be2fcb2c62d0376ab94f86efea833a3

Observation 5b328b6b-b6c3-41d1-bd03-f630763a32d9 · outbound

This paper cites an unresolved cited work.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:06:37.275463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:27.216457Z digest=sha256:821eb7d2ce7d1323032e417d9c746cf0f5991134e16b5cbd3ac9d205dbc927af

Observation 465d0b79-109e-445b-8d40-b22776430b9d · outbound

This paper cites O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.316374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.316374Z digest=sha256:607fe1b9549f169ab4fc3476b8217f11a5fb30ed2368b6e7abdc6ba037cfc72d

Observation 7d6f6829-cb33-4c11-a102-3e1783d8a453 · outbound

This paper cites Corrupted by reasoning: Reasoning language models become free-riders in public goods games.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Corrupted by reasoning: Reasoning language models become free-riders in public goods games

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.994085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:27.402952Z digest=sha256:ae1bd9909bab14559c972367cb3f204e7f30d7f6ff2b0c89ede652d71589f5e1

Observation 484538ba-e022-4c90-a7d6-50b35e7c37a2 · outbound

This paper cites Fino1: On the Transferability of Reasoning-Enhanced LLMs and Reinforcement Learning to Finance.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Fino1: On the Transferability of Reasoning-Enhanced LLMs and Reinforcement Learning to Finance

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.524566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.524566Z digest=sha256:a3febf3c36005d24f6549204ee1ae59b5dce09783170efddd84db310dfc325bf

Observation e4e1efd2-5964-4ac2-91e2-50cca814d6f0 · outbound

This paper cites Econlogicqa: A question-answering benchmark for evaluating large language models in economic sequential reasoning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Econlogicqa: A question-answering benchmark for evaluating large language models in economic sequential reasoning

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.714530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:27.593164Z digest=sha256:867fdd54a68abfa042ca2c22b8e4b3c4ce56e0c0a950a448e36b48893ce177db

Observation de166244-9af6-4647-87a8-6bd42a7ff3ff · outbound

This paper cites STEER: Assessing the economic rationality of large language models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs STEER: Assessing the economic rationality of large language models

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.404265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:27.640402Z digest=sha256:8f361d8f120099572317edac761a83aa622d432f748f2e05b597da8dabcdcb7d

Observation 86407cb3-b286-4ac6-b4e8-cbc85dafa5e5 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.756368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.756368Z digest=sha256:301df5d2df0a88297371120646edeaf64ef32892fdf9201d64cdaaa39fcfa0ff

Observation 58dbb699-28ff-4408-94ca-3944ef3011be · outbound

This paper cites Glee: A unified framework and benchmark for language-based economic environments.arXiv preprint arXiv:2410.05254, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Glee: A unified framework and benchmark for language-based economic environments.arXiv preprint arXiv:2410.05254, 2024

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.868455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.868455Z digest=sha256:85ebf7742ad50fcf3d20e8f7490c316fc34f9aecdb971b091156203e17bfe597

Observation b11444a5-f7a3-402c-a49a-e8ad87ef4d27 · outbound

This paper cites FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.949635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.949635Z digest=sha256:3d2e380239fdc6edc56b73c112b4ed8abfa8a8afb802757412fefc1d6c0bc90e

Observation fbf8b06b-ad6f-47be-9caf-c8607bd63db2 · outbound

This paper cites Lazydit: Lazy learning for the acceleration of diffusion transformers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Lazydit: Lazy learning for the acceleration of diffusion transformers

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.032139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.032139Z digest=sha256:2858529068f8ff82000924e221cdd7d0c5704afe770e15758729b1dd45e145aa

Observation 8e13b2ec-89c0-4941-b761-31ad46a413ef · outbound

This paper cites Numerical pruning for efficient autoregressive models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Numerical pruning for efficient autoregressive models

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.199738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:28.116146Z digest=sha256:9380813d3e1e6c36bcd6653df8a1b4436b8f4eabaa5740cfb69ba4444d8ab0bb

Observation 7c4569fe-9af5-4cdc-a238-d1ce01f1af07 · outbound

This paper cites Blumberg, Stephen Marcus McAleer, Yaodong Yang, and Jun Wang.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Blumberg, Stephen Marcus McAleer, Yaodong Yang, and Jun Wang

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.016775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:28.194139Z digest=sha256:889e5ca09534ca062b0f0b724f230e80f0dc31b57722e688be4d31386371cf35

Observation b96aebd1-5915-490a-b807-7b216794f814 · outbound

This paper cites Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.437391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.437391Z digest=sha256:ebda914ab3ab51f23df530aa8d215a397f222087d00e851fc1ff7eaf77bd9f7f

Observation 9afceeff-96e3-4937-aa25-901a8575507d · outbound

This paper cites Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.538769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.538769Z digest=sha256:87d2b55022f21d3120b55f72a805be05d9a73338ac9b43cdaa9e7a8e9d2e698f

Observation 7f4c9919-748c-4c67-af21-c92ff3cd9c53 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gemma 2: Improving Open Language Models at a Practical Size

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.633246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.633246Z digest=sha256:68ab23b915f74b466eb3d4ac39b50fafd87d0326ac63a1934381441c68b93a56

Observation 83ee0e5d-377c-4b73-8c39-b220d1f59c15 · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, March 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Qwq-32b: Embracing the power of reinforcement learning, March 2025

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.738712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.738712Z digest=sha256:4ea2764e3480ee249b0cae9cce6571462d2ab427d5a22e27acba408d88f31ed9

Observation 33df0520-313a-4297-bebc-fb18006bf2db · outbound

This paper cites Competing large language models in multi-agent gaming environments.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Competing large language models in multi-agent gaming environments

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.649550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:28.848667Z digest=sha256:98d19b45c033aa0df15c482e9b153afbec9e56c952bc6c0f49083a5d1c175701

Observation 539d6bf5-cd33-42d2-b2e6-e4e8c770ddd6 · outbound

This paper cites Attention is all you need.Advances in neural information processing systems, 30, 2017.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Attention is all you need.Advances in neural information processing systems, 30, 2017

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.940776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.940776Z digest=sha256:17c176cccb0b2b6e18cb38c59df0fd45af432fb73c4cf2a16cd7c8b3dab0a488

Observation f7f1bdb2-acbc-49bc-99bd-0ce49608aba9 · outbound

This paper cites Trl: Transformer reinforcement learning.https://github.com/huggingface/trl, 2020.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Trl: Transformer reinforcement learning.https://github.com/huggingface/trl, 2020

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.462498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:29.053198Z digest=sha256:a55276c75ab7c0bbe3ef9b6e2a71953f9d664615cfb1ae5d75398f451ee802e4

Observation 5872265f-aeb3-4ad1-b395-d9969c4b32b8 · outbound

This paper cites Drupi: Dataset reduction using privileged information.arXiv preprint arXiv:2410.01611, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Drupi: Dataset reduction using privileged information.arXiv preprint arXiv:2410.01611, 2024

Reference 76

Resolution
verified exact
raw_fallback, observed 2026-08-07T12:06:32.294519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:29.177484Z digest=sha256:f6493cd7de81321b5054ebaecf4cad255399a9611d87f6a8237cb2cd9c04fac8

Observation d632b52e-18a2-4fb0-93d8-16b0c4d40b47 · outbound

This paper cites Data whisperer: Efficient data selection for task-specific llm fine-tuning via few-shot in-context learning.Annual Meeting of the Association for Computational Linguistics, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Data whisperer: Efficient data selection for task-specific llm fine-tuning via few-shot in-context learning.Annual Meeting of the Association for Computational Linguistics, 2025

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.265495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:29.287822Z digest=sha256:11f0ea32bbe33fb7b83643c0c63343be3aa2a609dba10cafc883e66e05e93056

Observation 79df8869-3b74-4f49-b450-ddf1cc753295 · outbound

This paper cites Gnothi seauton: Empowering faithful self-interpretability in black-box transformers.International Conference on Learning Representations, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gnothi seauton: Empowering faithful self-interpretability in black-box transformers.International Conference on Learning Representations, 2025

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.079828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:29.377015Z digest=sha256:9a33ac5ab3a896b4e03d376d95f6da0d1f922fa7d329d5ac55840125b0f25645

Observation d23a7c22-2193-4bf3-88b3-25e112173a45 · outbound

This paper cites Not all samples should be utilized equally: Towards understanding and improving dataset distillation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Not all samples should be utilized equally: Towards understanding and improving dataset distillation

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.898641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:29.479013Z digest=sha256:8a3b5332d4c441a43cfefec1e4df0d7e5b11538514f01fbebbdff3102d6fe63b

Observation 34d93961-4df5-44c8-acac-48d347e969f4 · outbound

This paper cites Dataset distillation with neural characteristic function: A minmax perspective.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Dataset distillation with neural characteristic function: A minmax perspective

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.731618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:29.555815Z digest=sha256:78079d2d655a9c7fee0b21906ae9aa6a388452cbe9f55b0fbd9b674bad245c84

Observation 67e8d35c-14d8-4c3f-be32-2a71abd0edbd · outbound

This paper cites Dataset Distillation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Dataset Distillation

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:29.647628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:29.647628Z digest=sha256:2592354d018aa2bf733b8828e686cc3a29137fb05149d0fce23a96b1693edb22

Observation d87a9b8e-8592-4dae-b3b3-deed3808e6e4 · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:29.723448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:29.723448Z digest=sha256:cd7ffb98880eca4505741e9048cbc6ace4c755ebeb454f53bc00a346a0d54bcf

Observation 4cad0334-7bca-40f2-960e-9e1905b9cb76 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Chain-of-thought prompting elicits reasoning in large language models

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.503202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:29.817142Z digest=sha256:6b6a80f24a6afd35d1b9c493b9ef7d7518ef9fa66a0d43b66df5851161a96ee9

Observation c2d08aef-ed6d-4ed8-82b2-bff548805eed · outbound

This paper cites Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:29.931517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:29.931517Z digest=sha256:8fcfd35f8a9a7feefbc9d91914aefdfe06f915a05db46fe70924e211035be5a3

Observation 97693e2d-168c-4a5c-bb10-113493f57727 · outbound

This paper cites AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.017600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.017600Z digest=sha256:34a2763a155edfd156e6d232debcc04ab8b9bae007dc25a908af579200abec37

Observation 3208fa00-3ea2-4407-a70c-a64010242d19 · outbound

This paper cites LESS: Selecting Influential Data for Targeted Instruction Tuning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.108486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.108486Z digest=sha256:ae18e5b2fc4ec3323aaf246bd7c91cd8da26e3b597fe7177297d0d19875fd129

Observation 3531be30-ba7f-40c4-938e-4e62d108b0b0 · outbound

This paper cites Rethinking Data Selection at Scale: Random Selection is Almost All You Need.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Rethinking Data Selection at Scale: Random Selection is Almost All You Need

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.202322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.202322Z digest=sha256:858e6489bf58e736ce06e58fabec0cf7da5019c9742f38a04839b3298f06caab

Observation 0a9f2820-51ba-4ed6-9437-8a187c990f2a · outbound

This paper cites TradingAgents: Multi-Agents LLM Financial Trading Framework.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs TradingAgents: Multi-Agents LLM Financial Trading Framework

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.294760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.294760Z digest=sha256:da02f193ed44019a1b7eb721a25dbf5f6014db593528c0dba00bcf459ab7fb8f

Observation 9ea7fc11-a432-4451-997c-5360e06c048b · outbound

This paper cites Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.415442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.415442Z digest=sha256:023468a082f9e0e30b15896ed1f69e878a3128356c976ef620c8bbb303eee081

Observation 3656115e-bb98-4084-90fd-29c1b6d0d3f7 · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.507784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.507784Z digest=sha256:e2bc2ad67521c7db75270e37abbb7493c144d3762db26f6700d598609ac86316

Observation 50517e4f-ff1a-4ec5-a543-8a84daaee2d2 · outbound

This paper cites Rethinking dataset pruning from a generalization perspective.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Rethinking dataset pruning from a generalization perspective

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.321086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:30.648635Z digest=sha256:5baad10e8793a4521260519ee9cd50089508bcd0c9c0ad6b758e3dba1efa5f57

Observation 99b30453-2e05-42e1-96a8-30f71ecb7eda · outbound

This paper cites Karlsson.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Karlsson

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.157744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:30.760323Z digest=sha256:503b6cafb41c6076ce9bddbf5bf687624faff75da58209941cb7bbba2f5f021a

Observation 47912837-3515-4385-8c12-b4a6a3066139 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.829964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.829964Z digest=sha256:e7896681e56023a71f36edbb6a3ae9c14f0d75b0e2259de04a1fbb66fc70fe63

Observation 4ca12435-349e-44b7-a3ca-57e4072a1358 · outbound

This paper cites Qwen3 Technical Report.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Qwen3 Technical Report

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.893164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.893164Z digest=sha256:deaffd7253d3f5e1f02775ec4ec6ee41da30d05ac8917cef8815586bc81c7699

Observation af35aa0e-2447-4432-a27c-63e7c01ad41a · outbound

This paper cites LIMO: Less is More for Reasoning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs LIMO: Less is More for Reasoning

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.014267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.014267Z digest=sha256:e3b16f8178c86df00b535cb5eca8863a2f3cc7fd82dc2fa0708eeaeab3a1e196

Observation 54bf7264-cdcf-438b-b43d-4d4ca27322d4 · outbound

This paper cites an unresolved cited work.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Unresolved cited work

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.090651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.090651Z digest=sha256:9868a3586047c200faa62d8dd5855e3c50e2fc7a8afa8cbba17ec52d2338171b

Observation 141f96ca-1c14-4896-8ee8-d6c13be14e3f · outbound

This paper cites Synergistic multi-agent framework with trajectory learning for knowledge-intensive tasks.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Synergistic multi-agent framework with trajectory learning for knowledge-intensive tasks

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:33.991524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:31.204430Z digest=sha256:9e90c33ddb0c2fbdc57a454203bd52e4fe51a60540bc424139a70c590ec1bda5

Observation dbde382f-1e63-4088-84d9-1214d603fa35 · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.295933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.295933Z digest=sha256:0ea99c118ed0c6b488f9e8768360c90ec901c1657ac7e42d3cf3ceb1ecf7ee3f

Observation 7bdd97bf-f6ef-49f8-ae05-56eb1f7b797f · outbound

This paper cites Multi-agent reinforcement learning: A selective overview of theories and algorithms.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Multi-agent reinforcement learning: A selective overview of theories and algorithms

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:33.842562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:06:31.387736Z digest=sha256:0255dd942655dd131666265c11fb8d991ef9ba1a0fba5eeff85d73db0624a48e

Observation a652ec66-449a-4c3f-b792-6bd0b635ab19 · outbound

This paper cites Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.530153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.530153Z digest=sha256:787b6af967a76bfa63d33dc17e3a073aae7f006fb7057cd6fa4b1ed806cc560c

Pith citing papers

No inbound Pith citation observations are available.