Pith. sign in

Paper Citation Record · LEDGER

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs

As of 7 August 2026, this Paper Citation Record lists 100 of 105 outbound references and 0 inbound Pith citation observations for arXiv:2506.00577.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00577 v1

Coverage vector

measured 100 of 105 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:06:31.530153Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 105 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved78
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f694234f-5dec-43c7-befb-205ba2169f5f · outbound

This paper cites Coop- eration, competition, and maliciousness: LLM-stakeholders interactive negotiation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Coop- eration, competition, and maliciousness: LLM-stakeholders interactive negotiation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:21.951181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:21.951181Z digest=sha256:9ac626c5bfc5709ce4d5e1f30dc00911ad75c456d5dc751f7691d54b17c58f4e

Observation 443cce06-abcb-4c99-a6d0-97d8c785eb34 · outbound

This paper cites Playing repeated games with large language models.Nature Human Behaviour, pages 1–11, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Playing repeated games with large language models.Nature Human Behaviour, pages 1–11, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.052417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.052417Z digest=sha256:b3622fd6b399a5f382944f721b99c8b5dd53047496fb65cb5c47af48d94b527f

Observation e2931e38-8a20-4bb6-b2bd-bff93c00e524 · outbound

This paper cites Mechanistic interpretability for AI safety - a review.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Mechanistic interpretability for AI safety - a review

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.136800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.136800Z digest=sha256:9b205d1195cf2bb9eb0deb73745f57f88875a23fd83ef952fb0f49035ce5dbbb

Observation 210504bd-cb33-47f0-bff2-3653adb85416 · outbound

This paper cites Token Merging: Your ViT But Faster.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Token Merging: Your ViT But Faster

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.218936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.218936Z digest=sha256:ce3dc45d87209b83c5f0b25921049f7ad31b8bfccbc187e8c0a08a5fa4d97e1d

Observation 43b7d5e5-a7f6-4964-b224-0ebb8b06d1c8 · outbound

This paper cites Sparks of artificial general intelligence: Early experiments with gpt-4, 2023.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Sparks of artificial general intelligence: Early experiments with gpt-4, 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.273793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.273793Z digest=sha256:fbbc92e33d362c6178ed701e0040c6ddbb2524aaec8138f983b03381bc5cb139

Observation f0ce41f8-077f-4755-a7a4-ecc180ad2fcf · outbound

This paper cites Cambridge University Press, 2006.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Cambridge University Press, 2006

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.352074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.352074Z digest=sha256:72042d5fd64240a66682f3acbdcc95d9c66fde1b4d5b61c7913b18022a81d877

Observation dac041fb-c4cd-4cf8-9ea0-13761f3a6c74 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.421425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.421425Z digest=sha256:e5f15f884af9a4325ede18e02c36a0aba3685bf33c19a0ea6d3bb698c0072b76

Observation 569cab53-abb2-4a2a-95e1-70a634fb5e9b · outbound

This paper cites The computational limits of state-space models and mamba via the lens of circuit complexity.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs The computational limits of state-space models and mamba via the lens of circuit complexity

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.491870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.491870Z digest=sha256:b625b48f1440a2b8399363c2fdbdb20488e91b643ac23028eb2c924379f2b37c

Observation 23407aeb-deec-4e61-8efd-6ab0c3408842 · outbound

This paper cites Universal Approximation of Visual Autoregressive Transformers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Universal Approximation of Visual Autoregressive Transformers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.588160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.588160Z digest=sha256:2aaf41e1da9eb41864155cac7e11c77738a4b73353e7eecbc11678b13d27fbfd

Observation 38411d10-08ef-4cab-b791-44b800942f0d · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.681246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.681246Z digest=sha256:a5a5a28b65dc39e7388a3d0616e82031e19a8ff56715825cccd8be7581f52a87

Observation b289405f-e8b5-4c51-a684-04a12f98d677 · outbound

This paper cites Gamebench: Evaluating strategic reasoning abilities of LLM agents.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gamebench: Evaluating strategic reasoning abilities of LLM agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.751511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.751511Z digest=sha256:bc44154ade0fba708aae7f43c5542bc26b552a456f8e8b3cb64fbcc286daa4af

Observation 18f51814-d2e8-435f-a617-5a56d97353a1 · outbound

This paper cites Learning to Estimate Shapley Values with Vision Transformers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Learning to Estimate Shapley Values with Vision Transformers

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.865547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.865547Z digest=sha256:bb48aac63163411e3ac0c5af315eae04b567d0974098cd953476a3ff8069fa79

Observation 58d951ef-d171-4c18-ac61-b4b26545e6e6 · outbound

This paper cites Unsloth, 2023.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Unsloth, 2023

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:22.955760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:22.955760Z digest=sha256:6e2b60481a910596f07d77853a2f40fb4b30e1e1ba24868691caa6017e1ea5a9

Observation 729722a4-164e-4911-9b9c-1e0f97919eb5 · outbound

This paper cites Evaluating language model agency through negotiations.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Evaluating language model agency through negotiations

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.018165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.018165Z digest=sha256:98e45aae2b180f05dbb880c702f0afb332585dfe70b9655e3cf3fb56da2e2dea

Observation 057366c2-69bc-4bc3-9328-bf277a615388 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.094647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.094647Z digest=sha256:192aae38d689b998135af9bf8ef22cb98abd4b514efc65c1c0d5080228bef700

Observation 935647c9-1fc1-4056-9afa-59307dc08514 · outbound

This paper cites A survey on the optimization of large language model-based agents.arXiv preprint arXiv:2503.12434, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs A survey on the optimization of large language model-based agents.arXiv preprint arXiv:2503.12434, 2025

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.184862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.184862Z digest=sha256:42646253ef482dd3591e0eb279a533868d4f191b6e4d7ab2bfad7553220be11a

Observation 78462c23-68fd-4afa-90ea-9abe3336a673 · outbound

This paper cites GTBench: Uncovering the strategic reasoning capabilities of LLMs via game-theoretic evaluations.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs GTBench: Uncovering the strategic reasoning capabilities of LLMs via game-theoretic evaluations

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.263043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.263043Z digest=sha256:529e0e4fa3da0dd45290a0151d7d3d85d8c4bab82aee627bf19b693bf813d63a

Observation d170b8d1-0e69-49af-b9ce-1c1e129cb80b · outbound

This paper cites Can large language models serve as rational players in game theory? a systematic analysis.Proceedings of the AAAI Conference on Artificial Intelligence, 38(16):17960–17967, Mar.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Can large language models serve as rational players in game theory? a systematic analysis.Proceedings of the AAAI Conference on Artificial Intelligence, 38(16):17960–17967, Mar

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.351443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.351443Z digest=sha256:d84b674e3adf893ee3d87ba51d2e2cf886f41c9975e571da6e5d7d06bbef9336

Observation 08d21ab5-7251-497d-8290-82d8e73e003b · outbound

This paper cites How far are we from agi: Are llms all we need?Transactions on Machine Learning Research, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs How far are we from agi: Are llms all we need?Transactions on Machine Learning Research, 2024

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.407980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.407980Z digest=sha256:7e4d0ae30bacface38ba80d2a4d211af4b00568d19574b3bacfcbc2ecb2f7b28

Observation 2f0c5408-b06e-4ac8-80b9-f23d6c96e868 · outbound

This paper cites Dataset with 200 million 3-by-3 strategic games for comparing perfectly transparent equilibria with nash equilibria, 2020-10-07.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Dataset with 200 million 3-by-3 strategic games for comparing perfectly transparent equilibria with nash equilibria, 2020-10-07

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.502734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.502734Z digest=sha256:b00c00179e9b0460eedd043c81e57bffc685020cc984fc0f6dacf2e2ef926c96

Observation 2a44a10b-a7ca-42d9-bd40-fc6966575dc7 · outbound

This paper cites Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.589925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.589925Z digest=sha256:75d240cc5e6c2792706dd74f86f7efd780fa876b79e980eef5d263137ba3106c

Observation 359e17ef-f1be-4c9f-9182-3869723f85f5 · outbound

This paper cites The llama 3 herd of models.arXiv e-prints, pages arXiv–2407, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs The llama 3 herd of models.arXiv e-prints, pages arXiv–2407, 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.661807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.661807Z digest=sha256:8f6e9b922f567dbe4e507039e40bf381189ba92f6136794b1da5a1a584fa2083

Observation d23c612c-b3be-4b30-95c7-0ce5d8316d40 · outbound

This paper cites Econnli: Evaluating large language models on economics reasoning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Econnli: Evaluating large language models on economics reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.754861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.754861Z digest=sha256:20ec6588a48d7f2fc16edd0a48147eb2c57491d680f03fb7fbd9bdfb1977f1a5

Observation be29577f-849e-4dae-9470-5c184001a066 · outbound

This paper cites To- wards lossless dataset distillation via difficulty-aligned trajectory matching.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs To- wards lossless dataset distillation via difficulty-aligned trajectory matching

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.847212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.847212Z digest=sha256:2b1661e12df33825adc079ec80b1a81a68a6389687966e2bddab8446838dc4fe

Observation b76e24e1-81c4-41bd-a7a2-3ff195e5c8ca · outbound

This paper cites A multi-llm-agent-based framework for economic and public policy analysis.arXiv preprint arXiv:2502.16879, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs A multi-llm-agent-based framework for economic and public policy analysis.arXiv preprint arXiv:2502.16879, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.967091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.967091Z digest=sha256:dcbd46597d9f6c108e80a7273bf51badb171a2d31a30475362fcf4ea40fd4a6d

Observation 6b5ffdc1-e762-4f38-8c78-caeab0cec311 · outbound

This paper cites Measuring mathematical problem solving with the MATH dataset.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Measuring mathematical problem solving with the MATH dataset

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.064982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.064982Z digest=sha256:9b0560c3672cbd4c1d24d1bd444a7b30db2c0e619f74c23b2f0c01099ce1166f

Observation 62d4e287-69ae-44f6-9a94-d2d28fc662be · outbound

This paper cites Training Compute-Optimal Large Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Training Compute-Optimal Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.136956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.136956Z digest=sha256:0e3e43f9d2449be1500eaa48942a6c9fb191d4fb17ade5c712fb01b39f5736a5

Observation d7427233-0f6d-49d8-aee7-92997177b82b · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1 (2):3, 2022.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Lora: Low-rank adaptation of large language models.ICLR, 1 (2):3, 2022

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.214408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.214408Z digest=sha256:87a18a14e4849a39694d7bac801331dbec28efc64919fbb368d0072b9fcd323a

Observation 832c3562-b10c-46b6-bd30-aa450877c049 · outbound

This paper cites Game-theoretic LLM: Agent Workflow for Negotiation Games.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Game-theoretic LLM: Agent Workflow for Negotiation Games

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.323823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.323823Z digest=sha256:733870e84a5b2a4d7ef4b65e125203340deaaa5f60bcaed04ecd533d20a2aff9

Observation 5f413b8b-5f10-4bf3-a094-4be155a82a71 · outbound

This paper cites Scaling Laws for Neural Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Scaling Laws for Neural Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.445884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.445884Z digest=sha256:52a719a1a5dfbf1e7f4bb1c0a18697e9e44c43059b85856c7fe36d1408f2fa6c

Observation 88d3205a-c42b-49d8-9a7a-7580d0b0664c · outbound

This paper cites Large language models are zero-shot reasoners.Advances in neural information processing systems, 35:22199–22213, 2022.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Large language models are zero-shot reasoners.Advances in neural information processing systems, 35:22199–22213, 2022

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.549727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.549727Z digest=sha256:3ac7bea00926b747ec005f7b8df6c02dab8d0a0e3fc9c3479520b0428cdaaf62

Observation 4c7422eb-915f-4c85-9266-3449b88e1a37 · outbound

This paper cites Scaling Laws for Precision.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Scaling Laws for Precision

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.610518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.610518Z digest=sha256:dc837fa6db64ca1124df0d68769155c5cd0b31fcdcce3bac4c094bd0e3c515c5

Observation dd143728-d871-4e76-acc2-87fdb62feda1 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gonzalez, Hao Zhang, and Ion Stoica

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.686532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.686532Z digest=sha256:6074b7fa69514d4ce54dd51f9d7c6f111a1f2e442a1559f917878355003887ea

Observation 87cab2af-d06a-4866-bc01-8119fe4218ef · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.761893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.761893Z digest=sha256:1fb5a3cf9db7b7464674687ef17aaff4f471a6c517f6484b0051a5c9d8322a06

Observation 6046f0a2-ce3f-42a3-9ec0-a638152b5b75 · outbound

This paper cites Investigating Tax Evasion Emergence Using Dual Large Language Model and Deep Reinforcement Learning Powered Agent-based Simulation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Investigating Tax Evasion Emergence Using Dual Large Language Model and Deep Reinforcement Learning Powered Agent-based Simulation

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:06:32.997747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:24.846435Z digest=sha256:9e6c2dde3276708a1aec7be4c79680ba7d692f0fc7000016463ff3ace858d2ab

Observation 53e5b513-16c3-4520-a731-5a8ad4c56db7 · outbound

This paper cites CAMEL: Communicative agents for "mind" exploration of large language model society.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs CAMEL: Communicative agents for "mind" exploration of large language model society

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.941592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.941592Z digest=sha256:42e3366457f39a9d999f5df70d918f911bc85c687c6a83ae74247086ac79e1b5

Observation 9b0e3d55-1701-44f7-9195-7a06fea9b1a5 · outbound

This paper cites EconAgent: Large language model-empowered agents for simulating macroeconomic activities.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs EconAgent: Large language model-empowered agents for simulating macroeconomic activities

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:24.995146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:24.995146Z digest=sha256:b81d80c908bfbe76272491a15bb4f3a5991fcec9451c19c30898ba076b0246f2

Observation 622b3240-6a49-4b5d-ba52-4a2c48cb0687 · outbound

This paper cites LIMR: Less is More for RL Scaling.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs LIMR: Less is More for RL Scaling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.160677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.160677Z digest=sha256:2c5e8230c4e4c7b94ca301d3f8b807167ae63dfc706c01759e4e9e2bd7e1f8d9

Observation c9e7634b-437a-4278-82b9-d9489b0af0f2 · outbound

This paper cites Beyond linear approximations: A novel pruning approach for attention matrix.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Beyond linear approximations: A novel pruning approach for attention matrix

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.227804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.227804Z digest=sha256:141c9936ab7f2fe2286d9031512e0609f398035f62ddc87c608b98443d7ea094

Observation 8f8e26c4-5f19-4551-a456-634f196ae0a8 · outbound

This paper cites Looped relu mlps may be all you need as programmable computers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Looped relu mlps may be all you need as programmable computers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.334573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.334573Z digest=sha256:988bb8895808af962fc2d16c5de9f22456752de5f392620b7e2168cb594d4f32

Observation e61494a9-9d23-4fd6-a885-542d46411b31 · outbound

This paper cites MARFT: Multi-Agent Reinforcement Fine-Tuning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs MARFT: Multi-Agent Reinforcement Fine-Tuning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.439085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.439085Z digest=sha256:ab1f933fcaa66545dedc328cc659187e9f30dc2cf5abaf123734b572934f4e30

Observation e3dfada5-9506-4340-81e2-ead41ad9199a · outbound

This paper cites Let’s verify step by step.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Let’s verify step by step

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.554524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.554524Z digest=sha256:7ba53d3b613832fd5804f34537631c475abd00a4530d5c6cf7d04e22065f014c

Observation 82868fab-5092-4b42-a382-667d5d7ee9d9 · outbound

This paper cites Awq: Activation-aware weight quanti- zation for on-device llm compression and acceleration.Proceedings of Machine Learning and Systems, 6:87–100, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Awq: Activation-aware weight quanti- zation for on-device llm compression and acceleration.Proceedings of Machine Learning and Systems, 6:87–100, 2024

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.665528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.665528Z digest=sha256:43cd6559732b3e8ed49b8f34189e9b00ce3aec6dc4bc043cbc7c5496152d9f60

Observation 6b97be4f-8fef-4016-847b-682b060c615b · outbound

This paper cites Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.772527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.772527Z digest=sha256:25eb135a8d43535e8bb8ada9a4f3a4794673454149c68f48375e7d832bf19b4c

Observation ff4f1b8e-554e-4206-b230-010114d1bf38 · outbound

This paper cites Shifting ai efficiency from model-centric to data-centric compression.arXiv preprint arXiv:2505.19147, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Shifting ai efficiency from model-centric to data-centric compression.arXiv preprint arXiv:2505.19147, 2025

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:25.913312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:25.913312Z digest=sha256:b6a923b5f0dfaafcef1769e046740263c37b72cf031fe92b56d015b14f17298e

Observation 51d64124-5b14-4299-95a6-a225a23de0af · outbound

This paper cites Fin-r1: A large language model for financial reasoning through reinforcement learning.arXiv preprint arXiv:2503.16252, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Fin-r1: A large language model for financial reasoning through reinforcement learning.arXiv preprint arXiv:2503.16252, 2025

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.001796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.001796Z digest=sha256:13daf15627df21b271f415671e297405accb0d9b9b41dd71fd38396f324c503f

Observation cd71bb0a-2636-4789-b4ce-25546c1b8c53 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Understanding R1-Zero-Like Training: A Critical Perspective

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.113033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.113033Z digest=sha256:6a0a9b3872bfcc7204d2e0c03111645d88b268fee1b68cd7267068eb86d34d7f

Observation 369c1bfb-1127-45e2-81a2-06fafbcf2e37 · outbound

This paper cites Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.263283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.263283Z digest=sha256:d52994c62d8b1ea57ddd1ed05aff193e71be870170ff024ca9c347433c9d5b5c

Observation b3cd4cae-b09e-4f27-af57-b966764d8f8c · outbound

This paper cites The llama 4 herd: The beginning of a new era of natively multimodal ai innovation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs The llama 4 herd: The beginning of a new era of natively multimodal ai innovation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.368480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.368480Z digest=sha256:981c8b407fbc30ad7f187f77884891880743ee84148e659ce2276edded16bb82

Observation 5c810c38-5696-4ffa-b85f-d853269205ff · outbound

This paper cites Sql-r1: Training natural language to sql reasoning model by reinforcement learning.arXiv preprint arXiv:2504.08600, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Sql-r1: Training natural language to sql reasoning model by reinforcement learning.arXiv preprint arXiv:2504.08600, 2025

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.523847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.523847Z digest=sha256:4ebc18f7bd99e1eed25565cc400bd24283d74a26dfd5f5330f3cc105f5674779

Observation badcbb50-571f-427e-a078-27bfc2ac3e72 · outbound

This paper cites American invitational mathematics examination - aime.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs American invitational mathematics examination - aime

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.930734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:26.649702Z digest=sha256:335c7b756d6996a601e26b2f46d1eea564ec55db2a8a348111d0ade944a95a05

Observation b69bc38b-9a00-4c10-bef3-72f236b54b3f · outbound

This paper cites Tractable multi-agent reinforcement learning through behavioral economics.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Tractable multi-agent reinforcement learning through behavioral economics

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.753318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:26.747299Z digest=sha256:f768b877b779d48f2a9bff01397463776324b886435a0e022e74c6b64931310c

Observation 5add7113-ec19-4883-a36c-48466dd1d472 · outbound

This paper cites s1: Simple test-time scaling.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs s1: Simple test-time scaling

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:26.861404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:26.861404Z digest=sha256:634322ef2430e1cdbee4706ecb75dfc7e759214994390d1cd24e54404f553838

Observation 12090523-9901-471d-b2e4-122ba93475f2 · outbound

This paper cites Introducing chatgpt.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Introducing chatgpt

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.575161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:26.961132Z digest=sha256:ad75e3d08c099c0eb51eb09ee5c42dab93279eb043f35fb71519776869dfde77

Observation b845319f-0298-4e2f-beec-612482c53615 · outbound

This paper cites Hello gpt-4o.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Hello gpt-4o

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:37.395650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:27.048192Z digest=sha256:715bce8dfbd0ffb0af5bb3cf600bd7104c21e793a26ac5e76e67bdd86d9764c3

Observation a6f3e788-7ece-44ee-bf4a-a250b80ba323 · outbound

This paper cites OpenAI o1 System Card.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs OpenAI o1 System Card

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.119384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.119384Z digest=sha256:3cfaa28318884f7d42a282febfec718caf5e1d79967f433e8f55536b6dde4f33

Observation 5b328b6b-b6c3-41d1-bd03-f630763a32d9 · outbound

This paper cites an unresolved cited work.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:06:37.275463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:27.216457Z digest=sha256:8a29ce21819350bae3beb692825d3305dbcb3f9224208b1c8ca1c60a5407fe94

Observation 465d0b79-109e-445b-8d40-b22776430b9d · outbound

This paper cites O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.316374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.316374Z digest=sha256:6f3a248188c953040f096a8731fa9f348aab44af29ef98c0c6cbaeffca0a644e

Observation 7d6f6829-cb33-4c11-a102-3e1783d8a453 · outbound

This paper cites Corrupted by reasoning: Reasoning language models become free-riders in public goods games.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Corrupted by reasoning: Reasoning language models become free-riders in public goods games

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.994085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:27.402952Z digest=sha256:2f1999fdcaa58921d8c5ca4b0363f3856d2f26d9e2c6a72b989ed7daf8a37f13

Observation 484538ba-e022-4c90-a7d6-50b35e7c37a2 · outbound

This paper cites Fino1: On the Transferability of Reasoning-Enhanced LLMs and Reinforcement Learning to Finance.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Fino1: On the Transferability of Reasoning-Enhanced LLMs and Reinforcement Learning to Finance

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.524566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.524566Z digest=sha256:042c1b90c92d7460b1e0433fcdaefdcec83ca1d71faf8126a5a74397adaed91b

Observation e4e1efd2-5964-4ac2-91e2-50cca814d6f0 · outbound

This paper cites Econlogicqa: A question-answering benchmark for evaluating large language models in economic sequential reasoning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Econlogicqa: A question-answering benchmark for evaluating large language models in economic sequential reasoning

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.714530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:27.593164Z digest=sha256:0c962a253f9d548708e59e6e5ecdac0db624ccb9a6e5b2c667d5e7cee82efd04

Observation de166244-9af6-4647-87a8-6bd42a7ff3ff · outbound

This paper cites STEER: Assessing the economic rationality of large language models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs STEER: Assessing the economic rationality of large language models

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.404265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:27.640402Z digest=sha256:0f83c6510cc63cfc998de069516ee4f8dfdce7eac2d2d18c4161bb6c081e3636

Observation 86407cb3-b286-4ac6-b4e8-cbc85dafa5e5 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.756368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.756368Z digest=sha256:ff5fbf457036c456476cde8481123099f2ba5423a464c06630d1b1bab64a756a

Observation 58dbb699-28ff-4408-94ca-3944ef3011be · outbound

This paper cites Glee: A unified framework and benchmark for language-based economic environments.arXiv preprint arXiv:2410.05254, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Glee: A unified framework and benchmark for language-based economic environments.arXiv preprint arXiv:2410.05254, 2024

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.868455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.868455Z digest=sha256:ce40e64a3bff70c29fb4c671c0a8696d278c4b6b30f79bc13fbd46e89a453098

Observation b11444a5-f7a3-402c-a49a-e8ad87ef4d27 · outbound

This paper cites FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.949635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.949635Z digest=sha256:f767ca3dba3b86129da4b7aa9f41c9a47dc9eabc7c4e6274f798c46fd02db1cc

Observation fbf8b06b-ad6f-47be-9caf-c8607bd63db2 · outbound

This paper cites Lazydit: Lazy learning for the acceleration of diffusion transformers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Lazydit: Lazy learning for the acceleration of diffusion transformers

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.032139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.032139Z digest=sha256:1a031e8c9a6c7ca8593267a27fd37cd104b1d50aeca56a71ae0aeb492fee7dde

Observation 8e13b2ec-89c0-4941-b761-31ad46a413ef · outbound

This paper cites Numerical pruning for efficient autoregressive models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Numerical pruning for efficient autoregressive models

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.199738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:28.116146Z digest=sha256:315671a602db382afc8d634ebbb0e7770f83e69ccea0a8263fa9cc86d581afd1

Observation 7c4569fe-9af5-4cdc-a238-d1ce01f1af07 · outbound

This paper cites Blumberg, Stephen Marcus McAleer, Yaodong Yang, and Jun Wang.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Blumberg, Stephen Marcus McAleer, Yaodong Yang, and Jun Wang

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:36.016775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:28.194139Z digest=sha256:1389a7f60e12f7fc39b826e591528567850d2f650801d903819df3c55bd53a78

Observation b96aebd1-5915-490a-b807-7b216794f814 · outbound

This paper cites Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.437391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.437391Z digest=sha256:38f9840b4a2d9aa023a86692e6faa5600b7fdf5f861c7db9db5c0af6d0e2f9d7

Observation 9afceeff-96e3-4937-aa25-901a8575507d · outbound

This paper cites Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.538769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.538769Z digest=sha256:5709878cb20a2b2f2dbcdd282377aacfa9dba6a045f0bd659879afd47e7cc84d

Observation 7f4c9919-748c-4c67-af21-c92ff3cd9c53 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gemma 2: Improving Open Language Models at a Practical Size

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.633246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.633246Z digest=sha256:2a7c68cf6fa9646b4c3cc3c4299ac21953124517a76ff92c297d9cdcc1454dd3

Observation 83ee0e5d-377c-4b73-8c39-b220d1f59c15 · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, March 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Qwq-32b: Embracing the power of reinforcement learning, March 2025

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.738712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.738712Z digest=sha256:2054e72072d893fa2a66d4fa917b75b65b09541f76a39a3a7859243a8984308f

Observation 33df0520-313a-4297-bebc-fb18006bf2db · outbound

This paper cites Competing large language models in multi-agent gaming environments.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Competing large language models in multi-agent gaming environments

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.649550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:28.848667Z digest=sha256:ee5ae7607b3b8483b363fce8280f95a64cae7399c9d8e6d2bca8f1f863e7e1c3

Observation 539d6bf5-cd33-42d2-b2e6-e4e8c770ddd6 · outbound

This paper cites Attention is all you need.Advances in neural information processing systems, 30, 2017.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Attention is all you need.Advances in neural information processing systems, 30, 2017

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:28.940776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:28.940776Z digest=sha256:f49c04d2a78d55bd5e73c180814655b5aff4fb8df2a9abccb105609336264651

Observation f7f1bdb2-acbc-49bc-99bd-0ce49608aba9 · outbound

This paper cites Trl: Transformer reinforcement learning.https://github.com/huggingface/trl, 2020.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Trl: Transformer reinforcement learning.https://github.com/huggingface/trl, 2020

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.462498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.053198Z digest=sha256:025ea18aaa78bd74743f993ba47f1067067c63e2e1f5d7e7ed1f1fb33b800630

Observation 5872265f-aeb3-4ad1-b395-d9969c4b32b8 · outbound

This paper cites Drupi: Dataset reduction using privileged information.arXiv preprint arXiv:2410.01611, 2024.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Drupi: Dataset reduction using privileged information.arXiv preprint arXiv:2410.01611, 2024

Reference 76

Resolution
verified exact
raw_fallback, observed 2026-08-07T12:06:32.294519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.177484Z digest=sha256:7f7862742ea465f54b53e73a62c0ef9e6f698ba6b0bd33bb7ffa384a78ef89b8

Observation d632b52e-18a2-4fb0-93d8-16b0c4d40b47 · outbound

This paper cites Data whisperer: Efficient data selection for task-specific llm fine-tuning via few-shot in-context learning.Annual Meeting of the Association for Computational Linguistics, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Data whisperer: Efficient data selection for task-specific llm fine-tuning via few-shot in-context learning.Annual Meeting of the Association for Computational Linguistics, 2025

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.265495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.287822Z digest=sha256:4f7772c0d58d27bb7bd1c04e4d8455fe8d67be132641daf4043f1930923ba42e

Observation 79df8869-3b74-4f49-b450-ddf1cc753295 · outbound

This paper cites Gnothi seauton: Empowering faithful self-interpretability in black-box transformers.International Conference on Learning Representations, 2025.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Gnothi seauton: Empowering faithful self-interpretability in black-box transformers.International Conference on Learning Representations, 2025

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:35.079828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.377015Z digest=sha256:ba33e96fa664f1e2cf4906ec562cd508d2d820aaa8ebad84b6c07d40a7abe630

Observation d23a7c22-2193-4bf3-88b3-25e112173a45 · outbound

This paper cites Not all samples should be utilized equally: Towards understanding and improving dataset distillation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Not all samples should be utilized equally: Towards understanding and improving dataset distillation

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.898641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.479013Z digest=sha256:8b746bca6ef5548dce8b7ad95dfd8528864c5fc4095b3f31f54be5385312e033

Observation 34d93961-4df5-44c8-acac-48d347e969f4 · outbound

This paper cites Dataset distillation with neural characteristic function: A minmax perspective.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Dataset distillation with neural characteristic function: A minmax perspective

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.731618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.555815Z digest=sha256:044b6eb93228b522496ab342b801f6f2289ed6e5aa555aaa7667989096018b62

Observation 67e8d35c-14d8-4c3f-be32-2a71abd0edbd · outbound

This paper cites Dataset Distillation.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Dataset Distillation

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:29.647628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:29.647628Z digest=sha256:1f68723761acb0df42fd727fa7c6a4dd997c974a2ca8368da497325bbf9e7c95

Observation d87a9b8e-8592-4dae-b3b3-deed3808e6e4 · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:29.723448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:29.723448Z digest=sha256:c48ad120d731656586ffbca11ee795324f442df2e8e95bba1cf6101eea93e5fb

Observation 4cad0334-7bca-40f2-960e-9e1905b9cb76 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Chain-of-thought prompting elicits reasoning in large language models

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.503202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:29.817142Z digest=sha256:2fe2fd35c5cda133cb7d5b79569414da02c7f3c9c5e3e7c99545b7bc65c28772

Observation c2d08aef-ed6d-4ed8-82b2-bff548805eed · outbound

This paper cites Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:29.931517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:29.931517Z digest=sha256:52a6c3a57d15fb7be0d49216d625493480daec3ac2ee7313940c8b8bb5d68001

Observation 97693e2d-168c-4a5c-bb10-113493f57727 · outbound

This paper cites AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.017600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.017600Z digest=sha256:2c851685d0c29d3e0937b151e18d5281bed013dbee51ca1dd4e7bb587308033f

Observation 3208fa00-3ea2-4407-a70c-a64010242d19 · outbound

This paper cites LESS: Selecting Influential Data for Targeted Instruction Tuning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.108486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.108486Z digest=sha256:5f9a4f92865b2f188fc94e7160d3d9dcd4d1661db03e359b83428403d22018b5

Observation 3531be30-ba7f-40c4-938e-4e62d108b0b0 · outbound

This paper cites Rethinking Data Selection at Scale: Random Selection is Almost All You Need.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Rethinking Data Selection at Scale: Random Selection is Almost All You Need

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.202322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.202322Z digest=sha256:b7aee807445018c1bdf10d397fb9e97f51b0de6941bc76acc470cbe742a1530f

Observation 0a9f2820-51ba-4ed6-9437-8a187c990f2a · outbound

This paper cites TradingAgents: Multi-Agents LLM Financial Trading Framework.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs TradingAgents: Multi-Agents LLM Financial Trading Framework

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.294760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.294760Z digest=sha256:a4ca3dbf75cb96937bee2274746a651ece1ddbf048642611642edac6eeab9ba7

Observation 9ea7fc11-a432-4451-997c-5360e06c048b · outbound

This paper cites Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.415442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.415442Z digest=sha256:f35761dbeb5d197e7e8a6c68806797f0dfdcde038771a624d654b7add0701146

Observation 3656115e-bb98-4084-90fd-29c1b6d0d3f7 · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.507784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.507784Z digest=sha256:8fc0b857d6a07be370c0139f43359dacc8d7ea6c61dc2ab6b629cd450b84e690

Observation 50517e4f-ff1a-4ec5-a543-8a84daaee2d2 · outbound

This paper cites Rethinking dataset pruning from a generalization perspective.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Rethinking dataset pruning from a generalization perspective

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.321086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:30.648635Z digest=sha256:e0dbb04558f1bf69eb365d31d1d93f33e1de4205af79c3cc4b52d6863956ac43

Observation 99b30453-2e05-42e1-96a8-30f71ecb7eda · outbound

This paper cites Karlsson.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Karlsson

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:34.157744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:30.760323Z digest=sha256:a166f8b4f2ccbf20339f097df5f36bd822dfe1620b414f5b83464427f2e31fda

Observation 47912837-3515-4385-8c12-b4a6a3066139 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.829964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.829964Z digest=sha256:8bf78e9c54054b2e63344f99c9b035e0400ae8fab9f1d00b957d2415b6f7db6c

Observation 4ca12435-349e-44b7-a3ca-57e4072a1358 · outbound

This paper cites Qwen3 Technical Report.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Qwen3 Technical Report

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:30.893164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:30.893164Z digest=sha256:5e8f90414f1f50d8662e3d7684dbc0de1b08433dff93115acf9a5c2285e1546e

Observation af35aa0e-2447-4432-a27c-63e7c01ad41a · outbound

This paper cites LIMO: Less is More for Reasoning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs LIMO: Less is More for Reasoning

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.014267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.014267Z digest=sha256:766eac2d1ba5914fef98942cbc9dfcdc705b4c48d738615f6c597ed2a8d12158

Observation 54bf7264-cdcf-438b-b43d-4d4ca27322d4 · outbound

This paper cites an unresolved cited work.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Unresolved cited work

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.090651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.090651Z digest=sha256:d4d791c38267d699e51f084a98ba9a5e41f22009df354d6ecbfa897afb1b8092

Observation 141f96ca-1c14-4896-8ee8-d6c13be14e3f · outbound

This paper cites Synergistic multi-agent framework with trajectory learning for knowledge-intensive tasks.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Synergistic multi-agent framework with trajectory learning for knowledge-intensive tasks

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:33.991524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:31.204430Z digest=sha256:a1b98719d361979a8efda3574a33345c919f2a0b646b89c3daa66aae639d5266

Observation dbde382f-1e63-4088-84d9-1214d603fa35 · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.295933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.295933Z digest=sha256:2888f0c0627266540f4c6c16812756939bfeea7f27c828712322d5fc75fd9a8f

Observation 7bdd97bf-f6ef-49f8-ae05-56eb1f7b797f · outbound

This paper cites Multi-agent reinforcement learning: A selective overview of theories and algorithms.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Multi-agent reinforcement learning: A selective overview of theories and algorithms

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:06:33.842562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:06:31.387736Z digest=sha256:b8742138d79b5020cd60bebf96e0db11ec92749b9218a2cc115551598cb83600

Observation a652ec66-449a-4c3f-b792-6bd0b635ab19 · outbound

This paper cites Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning.

Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:31.530153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:31.530153Z digest=sha256:378b5b4bf43c3f2d7df7c3b6aef81849772c5212d7fd8295a9319198343ed563

Pith citing papers

No inbound Pith citation observations are available.