Pith. sign in

Paper Citation Record · LEDGER

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs

As of 7 August 2026, this Paper Citation Record lists 100 of 105 outbound references and 0 inbound Pith citation observations for arXiv:2506.00577.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00577 v1

Coverage vector

measured 100 of 105 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:06:31.530153Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 105 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved78
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f694234f-5dec-43c7-befb-205ba2169f5f · outbound

This paper cites Coop- eration, competition, and maliciousness: LLM-stakeholders interactive negotiation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Coop- eration, competition, and maliciousness: LLM-stakeholders interactive negotiation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:21.951181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:21.951181Z digest=sha256:ca9d0af3fb36ee5922e15b1f0afc035f1fab18099a44e16fb6470d9db4ec23af

Observation 443cce06-abcb-4c99-a6d0-97d8c785eb34 · outbound

This paper cites Playing repeated games with large language models.Nature Human Behaviour, pages 1–11, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Playing repeated games with large language models.Nature Human Behaviour, pages 1–11, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.052417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.052417Z digest=sha256:b04dbfca7024ad2db5a20ce20d203d7198ff2fc2028a1c8909b1e49d10391272

Observation e2931e38-8a20-4bb6-b2bd-bff93c00e524 · outbound

This paper cites Mechanistic interpretability for AI safety - a review.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Mechanistic interpretability for AI safety - a review

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.136800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.136800Z digest=sha256:b225b16e2f1a375692663e9a1ec03a083ad1d81ebefeb495294899be74d75a2b

Observation 210504bd-cb33-47f0-bff2-3653adb85416 · outbound

This paper cites Token Merging: Your ViT But Faster.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Token Merging: Your ViT But Faster

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.218936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.218936Z digest=sha256:4c65559b49c5cfb5e375fc0c3638b49276cc4e4653666a4d60e3ab04aac09c35

Observation 43b7d5e5-a7f6-4964-b224-0ebb8b06d1c8 · outbound

This paper cites Sparks of artificial general intelligence: Early experiments with gpt-4, 2023.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Sparks of artificial general intelligence: Early experiments with gpt-4, 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.273793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.273793Z digest=sha256:b7bb87efa8acfcef069509ae0ad993481b08d5987de749a0ecfc97eb2c81f05c

Observation f0ce41f8-077f-4755-a7a4-ecc180ad2fcf · outbound

This paper cites Cambridge University Press, 2006.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Cambridge University Press, 2006

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.352074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.352074Z digest=sha256:32c62260f968bc2131fe07f88be940a12711cc3ce9c1f223559f4b11cea9d458

Observation dac041fb-c4cd-4cf8-9ea0-13761f3a6c74 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.421425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.421425Z digest=sha256:2a3604f2eb714068b791b29ed9fa4c0551b1a18c4ec68171875f34f1c250f312

Observation 569cab53-abb2-4a2a-95e1-70a634fb5e9b · outbound

This paper cites The computational limits of state-space models and mamba via the lens of circuit complexity.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs The computational limits of state-space models and mamba via the lens of circuit complexity

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.491870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.491870Z digest=sha256:a886918a4c72069f5edeaa003f2d4237d2fb8b63e5d0c6684725e5dcc758ede8

Observation 23407aeb-deec-4e61-8efd-6ab0c3408842 · outbound

This paper cites Universal Approximation of Visual Autoregressive Transformers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Universal Approximation of Visual Autoregressive Transformers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.588160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.588160Z digest=sha256:72cf1b0623a84610979d373169d2a44b52e1b35f78c4eededf07827c50f8c687

Observation 38411d10-08ef-4cab-b791-44b800942f0d · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.681246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.681246Z digest=sha256:090c255f5fb17d841544ecea8984cb812099f464ad854860d883f8e8b21c9cf8

Observation b289405f-e8b5-4c51-a684-04a12f98d677 · outbound

This paper cites Gamebench: Evaluating strategic reasoning abilities of LLM agents.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gamebench: Evaluating strategic reasoning abilities of LLM agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.751511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.751511Z digest=sha256:7c61fa6d505819148b537b92fdbeef12460196d6551c7274e5d984ddf52d06c7

Observation 18f51814-d2e8-435f-a617-5a56d97353a1 · outbound

This paper cites Learning to Estimate Shapley Values with Vision Transformers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Learning to Estimate Shapley Values with Vision Transformers

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.865547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.865547Z digest=sha256:5b694400cf4ad07a892433279cd4e5e50f460e7fce00ceb3bbd86693f3bc6a6b

Observation 58d951ef-d171-4c18-ac61-b4b26545e6e6 · outbound

This paper cites Unsloth, 2023.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Unsloth, 2023

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.955760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.955760Z digest=sha256:4565c686f8108e3d0fef75a98fa5517eb160bb79525c643e1e7f093e57ec86a7

Observation 729722a4-164e-4911-9b9c-1e0f97919eb5 · outbound

This paper cites Evaluating language model agency through negotiations.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Evaluating language model agency through negotiations

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.018165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.018165Z digest=sha256:2c9878804837b0beb88e80df1cef2da70d8609c1a2cf8c1174cf11863d2186f4

Observation 057366c2-69bc-4bc3-9328-bf277a615388 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.094647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.094647Z digest=sha256:5ebb4a000acc2252817f211d64ec47a1d6f9235e7047ecab1ecd37bf0abb7979

Observation 935647c9-1fc1-4056-9afa-59307dc08514 · outbound

This paper cites A survey on the optimization of large language model-based agents.arXiv preprint arXiv:2503.12434, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs A survey on the optimization of large language model-based agents.arXiv preprint arXiv:2503.12434, 2025

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.184862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.184862Z digest=sha256:59629f06ea475c2d1d99d64bb75fbd565f70bd6fab858fac95e9067cf73b8a00

Observation 78462c23-68fd-4afa-90ea-9abe3336a673 · outbound

This paper cites GTBench: Uncovering the strategic reasoning capabilities of LLMs via game-theoretic evaluations.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs GTBench: Uncovering the strategic reasoning capabilities of LLMs via game-theoretic evaluations

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.263043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.263043Z digest=sha256:a102566dfe2f7fb3a12fbd50fe312c4e743cb33b0d0abc050fe0fe970fd620e5

Observation d170b8d1-0e69-49af-b9ce-1c1e129cb80b · outbound

This paper cites Can large language models serve as rational players in game theory? a systematic analysis.Proceedings of the AAAI Conference on Artificial Intelligence, 38(16):17960–17967, Mar.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Can large language models serve as rational players in game theory? a systematic analysis.Proceedings of the AAAI Conference on Artificial Intelligence, 38(16):17960–17967, Mar

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.351443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.351443Z digest=sha256:a867d5cdb5c244b1d7f6a791ac58ceb0faf679ac9c434b34f4c8ba0059decd0e

Observation 08d21ab5-7251-497d-8290-82d8e73e003b · outbound

This paper cites How far are we from agi: Are llms all we need?Transactions on Machine Learning Research, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs How far are we from agi: Are llms all we need?Transactions on Machine Learning Research, 2024

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.407980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.407980Z digest=sha256:d8e4d2125095585a4c4828abb25942ab1eb77bf5da9e0deb0b2c626a4f3f1bf3

Observation 2f0c5408-b06e-4ac8-80b9-f23d6c96e868 · outbound

This paper cites Dataset with 200 million 3-by-3 strategic games for comparing perfectly transparent equilibria with nash equilibria, 2020-10-07.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Dataset with 200 million 3-by-3 strategic games for comparing perfectly transparent equilibria with nash equilibria, 2020-10-07

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.502734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.502734Z digest=sha256:3328b796a3dca697093bb9c414161cc159743cb5b503dde046db23f67be4ffae

Observation 2a44a10b-a7ca-42d9-bd40-fc6966575dc7 · outbound

This paper cites Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.589925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.589925Z digest=sha256:cacbac53bd7b5ff54faa8914e94d6a0d0c7dbfdde35e872f5c14ff8f68c473a0

Observation 359e17ef-f1be-4c9f-9182-3869723f85f5 · outbound

This paper cites The llama 3 herd of models.arXiv e-prints, pages arXiv–2407, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs The llama 3 herd of models.arXiv e-prints, pages arXiv–2407, 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.661807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.661807Z digest=sha256:1acf9026b28433950a570f0c6ef0155e4b3d3f56c7c187f87a414cc5c3469186

Observation d23c612c-b3be-4b30-95c7-0ce5d8316d40 · outbound

This paper cites Econnli: Evaluating large language models on economics reasoning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Econnli: Evaluating large language models on economics reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.754861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.754861Z digest=sha256:e55b9d32adf5328d6d2a70eaf3efec46a177ff38f9e83ef127ed0acfce0a0dd4

Observation be29577f-849e-4dae-9470-5c184001a066 · outbound

This paper cites To- wards lossless dataset distillation via difficulty-aligned trajectory matching.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs To- wards lossless dataset distillation via difficulty-aligned trajectory matching

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.847212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.847212Z digest=sha256:0ad1d870d728c1c0f56b0372e1ce620cceab19058bf904fba603ca4d8d6f2040

Observation b76e24e1-81c4-41bd-a7a2-3ff195e5c8ca · outbound

This paper cites A multi-llm-agent-based framework for economic and public policy analysis.arXiv preprint arXiv:2502.16879, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs A multi-llm-agent-based framework for economic and public policy analysis.arXiv preprint arXiv:2502.16879, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.967091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.967091Z digest=sha256:cf9887829654038e07d1fc59520de703315db04c1a4b8a79ee4da034f86b49a4

Observation 6b5ffdc1-e762-4f38-8c78-caeab0cec311 · outbound

This paper cites Measuring mathematical problem solving with the MATH dataset.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Measuring mathematical problem solving with the MATH dataset

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.064982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.064982Z digest=sha256:1dbb0835e74939eb8ccfe1820a917fd47a19f03cc01ece6e8bc3969d186e05cc

Observation 62d4e287-69ae-44f6-9a94-d2d28fc662be · outbound

This paper cites Training Compute-Optimal Large Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Training Compute-Optimal Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.136956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.136956Z digest=sha256:8d130b1c1fe668904bb8bb2c8245b4fc56a9567c823dea257e5d3a6aa9b6c13d

Observation d7427233-0f6d-49d8-aee7-92997177b82b · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1 (2):3, 2022.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Lora: Low-rank adaptation of large language models.ICLR, 1 (2):3, 2022

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.214408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.214408Z digest=sha256:d5a187f0d27842a149310e06200972c03f4a1c1e23f2bf876785b2ba52313164

Observation 832c3562-b10c-46b6-bd30-aa450877c049 · outbound

This paper cites Game-theoretic LLM: Agent Workflow for Negotiation Games.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Game-theoretic LLM: Agent Workflow for Negotiation Games

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.323823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.323823Z digest=sha256:2b28b51cad30ab448aba15cf4a44dcadc5b68f358fb821a296b68013f4a9948c

Observation 5f413b8b-5f10-4bf3-a094-4be155a82a71 · outbound

This paper cites Scaling Laws for Neural Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Scaling Laws for Neural Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.445884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.445884Z digest=sha256:6e235583b200e1dacc083b96c8718f38f34c85d5e48ea1cd8a2d984dc8e1d63e

Observation 88d3205a-c42b-49d8-9a7a-7580d0b0664c · outbound

This paper cites Large language models are zero-shot reasoners.Advances in neural information processing systems, 35:22199–22213, 2022.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Large language models are zero-shot reasoners.Advances in neural information processing systems, 35:22199–22213, 2022

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.549727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.549727Z digest=sha256:2c2bc8f3d617618e4cad022879e7e3528ceaedad28876595948b3689aeb56d90

Observation 4c7422eb-915f-4c85-9266-3449b88e1a37 · outbound

This paper cites Scaling Laws for Precision.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Scaling Laws for Precision

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.610518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.610518Z digest=sha256:fb59e6a3e7628c13b3eeac7f3aafd1df9911b601356ef806e01af4042b1ecde0

Observation dd143728-d871-4e76-acc2-87fdb62feda1 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gonzalez, Hao Zhang, and Ion Stoica

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.686532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.686532Z digest=sha256:b44622be38451cae90e0fcc9dc7c7b348ec6c3791ea85594e5a26a29c6171874

Observation 87cab2af-d06a-4866-bc01-8119fe4218ef · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.761893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.761893Z digest=sha256:a678f27c72acd8b8bfdca81ede2bee2b7b96d4b9fc5c580ecfd063950202a624

Observation 6046f0a2-ce3f-42a3-9ec0-a638152b5b75 · outbound

This paper cites Investigating Tax Evasion Emergence Using Dual Large Language Model and Deep Reinforcement Learning Powered Agent-based Simulation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Investigating Tax Evasion Emergence Using Dual Large Language Model and Deep Reinforcement Learning Powered Agent-based Simulation

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:06:32.997747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:24.846435Z digest=sha256:dc7023ed5168e15a61e75136a8b41ad5222151985410cf35b076bae77a727969

Observation 53e5b513-16c3-4520-a731-5a8ad4c56db7 · outbound

This paper cites CAMEL: Communicative agents for "mind" exploration of large language model society.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs CAMEL: Communicative agents for "mind" exploration of large language model society

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.941592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.941592Z digest=sha256:29773dd936c059ddb1ca09f0067d59bd4c9b2b15dc29f63c7179f5d757ce4972

Observation 9b0e3d55-1701-44f7-9195-7a06fea9b1a5 · outbound

This paper cites EconAgent: Large language model-empowered agents for simulating macroeconomic activities.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs EconAgent: Large language model-empowered agents for simulating macroeconomic activities

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.995146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.995146Z digest=sha256:17876fc4bd52573177e9d96402aa72f32119829142d1f30532b378fb828e99d2

Observation 622b3240-6a49-4b5d-ba52-4a2c48cb0687 · outbound

This paper cites LIMR: Less is More for RL Scaling.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs LIMR: Less is More for RL Scaling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.160677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.160677Z digest=sha256:3a7482283521277b1bed96173efa5e74e171942f132737c1f211616bf2dc3aef

Observation c9e7634b-437a-4278-82b9-d9489b0af0f2 · outbound

This paper cites Beyond linear approximations: A novel pruning approach for attention matrix.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Beyond linear approximations: A novel pruning approach for attention matrix

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.227804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.227804Z digest=sha256:eae6d30a85d621c62a00d5379a632df585f95a1470ee6bc14251c5f29c44bfbc

Observation 8f8e26c4-5f19-4551-a456-634f196ae0a8 · outbound

This paper cites Looped relu mlps may be all you need as programmable computers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Looped relu mlps may be all you need as programmable computers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.334573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.334573Z digest=sha256:17d4bed163596438fcf0ec022c8bdf8c08e3d19d56c288e10c64c2c2713b9459

Observation e61494a9-9d23-4fd6-a885-542d46411b31 · outbound

This paper cites MARFT: Multi-Agent Reinforcement Fine-Tuning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs MARFT: Multi-Agent Reinforcement Fine-Tuning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.439085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.439085Z digest=sha256:6fe9d76c86cd0be538c0f2fd9aabcd4be88d0c1f79b8b16f6a92b302bab4b51a

Observation e3dfada5-9506-4340-81e2-ead41ad9199a · outbound

This paper cites Let’s verify step by step.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Let’s verify step by step

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.554524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.554524Z digest=sha256:3a43a6f6ff2ccd99181f03e34523b529243d8544c91b33373bf47e9adac15f56

Observation 82868fab-5092-4b42-a382-667d5d7ee9d9 · outbound

This paper cites Awq: Activation-aware weight quanti- zation for on-device llm compression and acceleration.Proceedings of Machine Learning and Systems, 6:87–100, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Awq: Activation-aware weight quanti- zation for on-device llm compression and acceleration.Proceedings of Machine Learning and Systems, 6:87–100, 2024

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.665528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.665528Z digest=sha256:5b4d4f13a56cabef54e4e1165b12153fcb3e8f75dab1d782259ee41f828f4963

Observation 6b97be4f-8fef-4016-847b-682b060c615b · outbound

This paper cites Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.772527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.772527Z digest=sha256:a2bf5102021bbd8091f50acbaa1ee52a41636fd8f44b7479c85745d882a7ddc9

Observation ff4f1b8e-554e-4206-b230-010114d1bf38 · outbound

This paper cites Shifting ai efficiency from model-centric to data-centric compression.arXiv preprint arXiv:2505.19147, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Shifting ai efficiency from model-centric to data-centric compression.arXiv preprint arXiv:2505.19147, 2025

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.913312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.913312Z digest=sha256:bd3c414216be271c52159ab93c28b07d7d96ed00a155a6a5e72acebccca2d397

Observation 51d64124-5b14-4299-95a6-a225a23de0af · outbound

This paper cites Fin-r1: A large language model for financial reasoning through reinforcement learning.arXiv preprint arXiv:2503.16252, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Fin-r1: A large language model for financial reasoning through reinforcement learning.arXiv preprint arXiv:2503.16252, 2025

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.001796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.001796Z digest=sha256:fe9030bc59bf1a2c91d125874e9ee52f0716bf2fcc9bcc43c80be776bbbb66ac

Observation cd71bb0a-2636-4789-b4ce-25546c1b8c53 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Understanding R1-Zero-Like Training: A Critical Perspective

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.113033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.113033Z digest=sha256:0e322da696310743de669e40133f12a96b3c8c26ec4fd42a1e0269be143eb591

Observation 369c1bfb-1127-45e2-81a2-06fafbcf2e37 · outbound

This paper cites Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.263283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.263283Z digest=sha256:f6cca506f4212ba2cdd4a34fa5b34416dc803817d14494278bf4ec4067c67856

Observation b3cd4cae-b09e-4f27-af57-b966764d8f8c · outbound

This paper cites The llama 4 herd: The beginning of a new era of natively multimodal ai innovation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs The llama 4 herd: The beginning of a new era of natively multimodal ai innovation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.368480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.368480Z digest=sha256:a07d30077edc23ef1b038ef7b9b8a503bffa47abd87a7fb7b77b563ee55cb5f1

Observation 5c810c38-5696-4ffa-b85f-d853269205ff · outbound

This paper cites Sql-r1: Training natural language to sql reasoning model by reinforcement learning.arXiv preprint arXiv:2504.08600, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Sql-r1: Training natural language to sql reasoning model by reinforcement learning.arXiv preprint arXiv:2504.08600, 2025

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.523847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.523847Z digest=sha256:674a3e3119a53ab0daff97e01f37c7a9ff6410870e149482e8cbf361e9c63972

Observation badcbb50-571f-427e-a078-27bfc2ac3e72 · outbound

This paper cites American invitational mathematics examination - aime.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs American invitational mathematics examination - aime

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.930734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:26.649702Z digest=sha256:2aa101e7e00906ac1fe55912d7a1fc5ba045360b8b9d4ab1b757b63dfc1d5b9f

Observation b69bc38b-9a00-4c10-bef3-72f236b54b3f · outbound

This paper cites Tractable multi-agent reinforcement learning through behavioral economics.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Tractable multi-agent reinforcement learning through behavioral economics

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.753318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:26.747299Z digest=sha256:a8ee6bbc20eeac919bae4c6e10f8af5155e080a586e958d0b13315f604d13a13

Observation 5add7113-ec19-4883-a36c-48466dd1d472 · outbound

This paper cites s1: Simple test-time scaling.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs s1: Simple test-time scaling

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.861404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.861404Z digest=sha256:781ee610c4d1d6b06e5c1c259d42ec568f67b296a01d956d1de1bb271222ee27

Observation 12090523-9901-471d-b2e4-122ba93475f2 · outbound

This paper cites Introducing chatgpt.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Introducing chatgpt

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.575161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:26.961132Z digest=sha256:6e36aa75e179aecd5fd331355f04e57bbf319de47b7826e7dfa09ee741371664

Observation b845319f-0298-4e2f-beec-612482c53615 · outbound

This paper cites Hello gpt-4o.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Hello gpt-4o

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.395650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:27.048192Z digest=sha256:83f88c6e6a4521e4e38f0e4961ed6abac717ce7bf5b9155d2a57412e932bd69b

Observation a6f3e788-7ece-44ee-bf4a-a250b80ba323 · outbound

This paper cites OpenAI o1 System Card.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs OpenAI o1 System Card

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.119384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.119384Z digest=sha256:1c61bae700ac5395e94a5129b65921474be2fcb2c62d0376ab94f86efea833a3

Observation 5b328b6b-b6c3-41d1-bd03-f630763a32d9 · outbound

This paper cites an unresolved cited work.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:06:37.275463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:27.216457Z digest=sha256:fff59ecc8102ea836febdeea2fb3bd335c1822708af29d6ce0db6922602d889a

Observation 465d0b79-109e-445b-8d40-b22776430b9d · outbound

This paper cites O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.316374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.316374Z digest=sha256:607fe1b9549f169ab4fc3476b8217f11a5fb30ed2368b6e7abdc6ba037cfc72d

Observation 7d6f6829-cb33-4c11-a102-3e1783d8a453 · outbound

This paper cites Corrupted by reasoning: Reasoning language models become free-riders in public goods games.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Corrupted by reasoning: Reasoning language models become free-riders in public goods games

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.994085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:27.402952Z digest=sha256:5b8ae5588dd277cd0f82dd33d846d71423dddfb29eb37500135d48169defee16

Observation 484538ba-e022-4c90-a7d6-50b35e7c37a2 · outbound

This paper cites Fino1: On the Transferability of Reasoning-Enhanced LLMs and Reinforcement Learning to Finance.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Fino1: On the Transferability of Reasoning-Enhanced LLMs and Reinforcement Learning to Finance

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.524566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.524566Z digest=sha256:f924930a0f5f8fceed5f69660e2464ccf3b1de046c9a79c4f1d28d40f6a0a079

Observation e4e1efd2-5964-4ac2-91e2-50cca814d6f0 · outbound

This paper cites Econlogicqa: A question-answering benchmark for evaluating large language models in economic sequential reasoning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Econlogicqa: A question-answering benchmark for evaluating large language models in economic sequential reasoning

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.714530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:27.593164Z digest=sha256:9e26dfbdf567571f21e95e881b3cf2b3d6703dc3044588c56b876f99cee6b9d1

Observation de166244-9af6-4647-87a8-6bd42a7ff3ff · outbound

This paper cites STEER: Assessing the economic rationality of large language models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs STEER: Assessing the economic rationality of large language models

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.404265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:27.640402Z digest=sha256:8d8467cd8265f350fe21017c11273647aac5972e27dc9257ce2da701d8d15cf0

Observation 86407cb3-b286-4ac6-b4e8-cbc85dafa5e5 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.756368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.756368Z digest=sha256:301df5d2df0a88297371120646edeaf64ef32892fdf9201d64cdaaa39fcfa0ff

Observation 58dbb699-28ff-4408-94ca-3944ef3011be · outbound

This paper cites Glee: A unified framework and benchmark for language-based economic environments.arXiv preprint arXiv:2410.05254, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Glee: A unified framework and benchmark for language-based economic environments.arXiv preprint arXiv:2410.05254, 2024

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.868455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.868455Z digest=sha256:85ebf7742ad50fcf3d20e8f7490c316fc34f9aecdb971b091156203e17bfe597

Observation b11444a5-f7a3-402c-a49a-e8ad87ef4d27 · outbound

This paper cites FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.949635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.949635Z digest=sha256:3d2e380239fdc6edc56b73c112b4ed8abfa8a8afb802757412fefc1d6c0bc90e

Observation fbf8b06b-ad6f-47be-9caf-c8607bd63db2 · outbound

This paper cites Lazydit: Lazy learning for the acceleration of diffusion transformers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Lazydit: Lazy learning for the acceleration of diffusion transformers

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.032139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.032139Z digest=sha256:2858529068f8ff82000924e221cdd7d0c5704afe770e15758729b1dd45e145aa

Observation 8e13b2ec-89c0-4941-b761-31ad46a413ef · outbound

This paper cites Numerical pruning for efficient autoregressive models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Numerical pruning for efficient autoregressive models

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.199738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:28.116146Z digest=sha256:66ca3aa600c9d0772bdc7ddc1717488dc728ad498b92bd20fd811096688bf060

Observation 7c4569fe-9af5-4cdc-a238-d1ce01f1af07 · outbound

This paper cites Blumberg, Stephen Marcus McAleer, Yaodong Yang, and Jun Wang.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Blumberg, Stephen Marcus McAleer, Yaodong Yang, and Jun Wang

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.016775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:28.194139Z digest=sha256:f796059fcc8dae4d6fb2839277d7a19790c097b28a13660eed2444d7793d14cf

Observation b96aebd1-5915-490a-b807-7b216794f814 · outbound

This paper cites Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.437391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.437391Z digest=sha256:ebda914ab3ab51f23df530aa8d215a397f222087d00e851fc1ff7eaf77bd9f7f

Observation 9afceeff-96e3-4937-aa25-901a8575507d · outbound

This paper cites Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.538769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.538769Z digest=sha256:8ec97c620281ef7e1ccf8bcfbc1ba92fa797e9585a4d6504043d93075e74b3cd

Observation 7f4c9919-748c-4c67-af21-c92ff3cd9c53 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gemma 2: Improving Open Language Models at a Practical Size

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.633246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.633246Z digest=sha256:68ab23b915f74b466eb3d4ac39b50fafd87d0326ac63a1934381441c68b93a56

Observation 83ee0e5d-377c-4b73-8c39-b220d1f59c15 · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, March 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Qwq-32b: Embracing the power of reinforcement learning, March 2025

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.738712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.738712Z digest=sha256:4ea2764e3480ee249b0cae9cce6571462d2ab427d5a22e27acba408d88f31ed9

Observation 33df0520-313a-4297-bebc-fb18006bf2db · outbound

This paper cites Competing large language models in multi-agent gaming environments.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Competing large language models in multi-agent gaming environments

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.649550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:28.848667Z digest=sha256:b109d595102e1fd96dcef63e250c4e25a0ad961b06a3af802d2b224304f14db4

Observation 539d6bf5-cd33-42d2-b2e6-e4e8c770ddd6 · outbound

This paper cites Attention is all you need.Advances in neural information processing systems, 30, 2017.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Attention is all you need.Advances in neural information processing systems, 30, 2017

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.940776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.940776Z digest=sha256:17c176cccb0b2b6e18cb38c59df0fd45af432fb73c4cf2a16cd7c8b3dab0a488

Observation f7f1bdb2-acbc-49bc-99bd-0ce49608aba9 · outbound

This paper cites Trl: Transformer reinforcement learning.https://github.com/huggingface/trl, 2020.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Trl: Transformer reinforcement learning.https://github.com/huggingface/trl, 2020

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.462498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.053198Z digest=sha256:e6c61773896fd3882b37ab8a4321a7b8ac1e53a7aa6243e2d92a4a6786497396

Observation 5872265f-aeb3-4ad1-b395-d9969c4b32b8 · outbound

This paper cites Drupi: Dataset reduction using privileged information.arXiv preprint arXiv:2410.01611, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Drupi: Dataset reduction using privileged information.arXiv preprint arXiv:2410.01611, 2024

Reference 76

Resolution
verified exact
raw_fallback, observed 2026-08-07T12:06:32.294519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.177484Z digest=sha256:ea19382b68e7c7bee28eae8b9a60619825a1e9fa87c550ecade885c3d524a877

Observation d632b52e-18a2-4fb0-93d8-16b0c4d40b47 · outbound

This paper cites Data whisperer: Efficient data selection for task-specific llm fine-tuning via few-shot in-context learning.Annual Meeting of the Association for Computational Linguistics, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Data whisperer: Efficient data selection for task-specific llm fine-tuning via few-shot in-context learning.Annual Meeting of the Association for Computational Linguistics, 2025

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.265495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.287822Z digest=sha256:3cec1da5d9918a368a8c12c175b95982e7b4f4821646da95904e653aafa6587b

Observation 79df8869-3b74-4f49-b450-ddf1cc753295 · outbound

This paper cites Gnothi seauton: Empowering faithful self-interpretability in black-box transformers.International Conference on Learning Representations, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gnothi seauton: Empowering faithful self-interpretability in black-box transformers.International Conference on Learning Representations, 2025

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.079828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.377015Z digest=sha256:bb28e01c12edeb61de5884d2b8ee2991ec3aac2b5e129d0785e190833bbb0970

Observation d23a7c22-2193-4bf3-88b3-25e112173a45 · outbound

This paper cites Not all samples should be utilized equally: Towards understanding and improving dataset distillation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Not all samples should be utilized equally: Towards understanding and improving dataset distillation

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.898641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.479013Z digest=sha256:0bdf3bbdc4a07bec6157aa174ff47249779dce07184855c36e9cd347d68c4b8e

Observation 34d93961-4df5-44c8-acac-48d347e969f4 · outbound

This paper cites Dataset distillation with neural characteristic function: A minmax perspective.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Dataset distillation with neural characteristic function: A minmax perspective

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.731618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.555815Z digest=sha256:d0171775eaa6c9d14b3161dccff809d5d4b12c294857b6c6c7247d7b7b6aa185

Observation 67e8d35c-14d8-4c3f-be32-2a71abd0edbd · outbound

This paper cites Dataset Distillation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Dataset Distillation

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:29.647628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:29.647628Z digest=sha256:2592354d018aa2bf733b8828e686cc3a29137fb05149d0fce23a96b1693edb22

Observation d87a9b8e-8592-4dae-b3b3-deed3808e6e4 · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:29.723448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:29.723448Z digest=sha256:cd7ffb98880eca4505741e9048cbc6ace4c755ebeb454f53bc00a346a0d54bcf

Observation 4cad0334-7bca-40f2-960e-9e1905b9cb76 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Chain-of-thought prompting elicits reasoning in large language models

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.503202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.817142Z digest=sha256:12ce1d63b800d2b3239c3d1bf64b83a041fa7bccbe8ddbc8d94bf4fa3671cb43

Observation c2d08aef-ed6d-4ed8-82b2-bff548805eed · outbound

This paper cites Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:29.931517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:29.931517Z digest=sha256:f47f0af58146c1e37f4debc0227b70bc2e3df029230912bcfbec766a23f46484

Observation 97693e2d-168c-4a5c-bb10-113493f57727 · outbound

This paper cites AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.017600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.017600Z digest=sha256:34a2763a155edfd156e6d232debcc04ab8b9bae007dc25a908af579200abec37

Observation 3208fa00-3ea2-4407-a70c-a64010242d19 · outbound

This paper cites LESS: Selecting Influential Data for Targeted Instruction Tuning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.108486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.108486Z digest=sha256:ae18e5b2fc4ec3323aaf246bd7c91cd8da26e3b597fe7177297d0d19875fd129

Observation 3531be30-ba7f-40c4-938e-4e62d108b0b0 · outbound

This paper cites Rethinking Data Selection at Scale: Random Selection is Almost All You Need.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Rethinking Data Selection at Scale: Random Selection is Almost All You Need

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.202322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.202322Z digest=sha256:bc05c978d0f99582417c00730fdde72d9ef932d4f989e0b6dd5c192fa756353e

Observation 0a9f2820-51ba-4ed6-9437-8a187c990f2a · outbound

This paper cites TradingAgents: Multi-Agents LLM Financial Trading Framework.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs TradingAgents: Multi-Agents LLM Financial Trading Framework

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.294760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.294760Z digest=sha256:b5dcacab0636c5126eebda796deb092b6d6f3c615ae400ee01bd6bf3981b7772

Observation 9ea7fc11-a432-4451-997c-5360e06c048b · outbound

This paper cites Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.415442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.415442Z digest=sha256:023468a082f9e0e30b15896ed1f69e878a3128356c976ef620c8bbb303eee081

Observation 3656115e-bb98-4084-90fd-29c1b6d0d3f7 · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.507784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.507784Z digest=sha256:e2bc2ad67521c7db75270e37abbb7493c144d3762db26f6700d598609ac86316

Observation 50517e4f-ff1a-4ec5-a543-8a84daaee2d2 · outbound

This paper cites Rethinking dataset pruning from a generalization perspective.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Rethinking dataset pruning from a generalization perspective

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.321086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:30.648635Z digest=sha256:4981db8f38dbba002e109084a362c012a79b86b664dae66f77b3d89c0cc59378

Observation 99b30453-2e05-42e1-96a8-30f71ecb7eda · outbound

This paper cites Karlsson.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Karlsson

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.157744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:30.760323Z digest=sha256:11c0cf11733146c73aa44f4abe95be706017a00e8484bf99c3888bdafb821bc7

Observation 47912837-3515-4385-8c12-b4a6a3066139 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.829964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.829964Z digest=sha256:e7896681e56023a71f36edbb6a3ae9c14f0d75b0e2259de04a1fbb66fc70fe63

Observation 4ca12435-349e-44b7-a3ca-57e4072a1358 · outbound

This paper cites Qwen3 Technical Report.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Qwen3 Technical Report

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.893164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.893164Z digest=sha256:deaffd7253d3f5e1f02775ec4ec6ee41da30d05ac8917cef8815586bc81c7699

Observation af35aa0e-2447-4432-a27c-63e7c01ad41a · outbound

This paper cites LIMO: Less is More for Reasoning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs LIMO: Less is More for Reasoning

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.014267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.014267Z digest=sha256:75b6f4c937e3b9cceaa3d4baa692f074a311a8154a5b144760adee95141d0c09

Observation 54bf7264-cdcf-438b-b43d-4d4ca27322d4 · outbound

This paper cites an unresolved cited work.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Unresolved cited work

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.090651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.090651Z digest=sha256:9868a3586047c200faa62d8dd5855e3c50e2fc7a8afa8cbba17ec52d2338171b

Observation 141f96ca-1c14-4896-8ee8-d6c13be14e3f · outbound

This paper cites Synergistic multi-agent framework with trajectory learning for knowledge-intensive tasks.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Synergistic multi-agent framework with trajectory learning for knowledge-intensive tasks

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:33.991524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:31.204430Z digest=sha256:e2478b6d1dfe1feff0394f33e4b93f763aabd98303dda655d9c8f4c953309cbb

Observation dbde382f-1e63-4088-84d9-1214d603fa35 · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.295933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.295933Z digest=sha256:0ea99c118ed0c6b488f9e8768360c90ec901c1657ac7e42d3cf3ceb1ecf7ee3f

Observation 7bdd97bf-f6ef-49f8-ae05-56eb1f7b797f · outbound

This paper cites Multi-agent reinforcement learning: A selective overview of theories and algorithms.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Multi-agent reinforcement learning: A selective overview of theories and algorithms

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:33.842562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:31.387736Z digest=sha256:0e54f7792bd25f1e137725aadcefd9744f295efbe08eb38320499f508cd11a6d

Observation a652ec66-449a-4c3f-b792-6bd0b635ab19 · outbound

This paper cites Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.530153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.530153Z digest=sha256:787b6af967a76bfa63d33dc17e3a073aae7f006fb7057cd6fa4b1ed806cc560c

Pith citing papers

No inbound Pith citation observations are available.