Pith. sign in

REVIEW 1 cited by

Combining Reinforcement Learning and Inverse Reinforcement Learning for Asset Allocation Recommendations

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2201.01874 v1 pith:PQXKKBJQ submitted 2022-01-06 cs.LG cs.AIq-fin.CPq-fin.PM

classification cs.LGcs.AIq-fin.CPq-fin.PM
keywords fundlearningmanagersreinforcementallocationassetfunctionimprove
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

We suggest a simple practical method to combine the human and artificial intelligence to both learn best investment practices of fund managers, and provide recommendations to improve them. Our approach is based on a combination of Inverse Reinforcement Learning (IRL) and RL. First, the IRL component learns the intent of fund managers as suggested by their trading history, and recovers their implied reward function. At the second step, this reward function is used by a direct RL algorithm to optimize asset allocation decisions. We show that our method is able to improve over the performance of individual fund managers.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Regret-Optimized Portfolio Enhancement through Deep Reinforcement Learning and Future Looking Rewards

    q-fin.PM 2025-02 conditional novelty 5.0 of 10

    A PPO agent with a hindsight regret reward, block-bootstrapped synthetic data, and a transaction-cost curriculum rebalances a 60/40 portfolio and beats it on return in three out-of-sample periods.

Pith tools