Pith. sign in

REVIEW 1 cited by

Applying Opponent Modeling for Automatic Bidding in Online Repeated Auctions

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2212.02723 v3 pith:UP5KFARL submitted 2022-12-06 cs.GT

classification cs.GT
keywords biddersalgorithmbiddingstrategiesstrategyauctionautomaticlearn
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Online auction scenarios, such as bidding searches on advertising platforms, often require bidders to participate repeatedly in auctions for identical or similar items. Most previous studies have only considered the process by which the seller learns the prior-dependent optimal mechanism in a repeated auction. However, in this paper, we define a multiagent reinforcement learning environment in which strategic bidders and the seller learn their strategies simultaneously and design an automatic bidding algorithm that updates the strategy of bidders through online interactions. We propose Bid Net to replace the linear shading function as a representation of the strategic bidders' strategy, which effectively improves the utility of strategy learned by bidders. We apply and revise the opponent modeling methods to design the PG (pseudo-gradient) algorithm, which allows bidders to learn optimal bidding strategies with predictions of the other agents' strategy transition. We prove that when a bidder uses the PG algorithm, it can learn the best response to static opponents. When all bidders adopt the PG algorithm, the system will converge to the equilibrium of the game induced by the auction. In experiments with diverse environmental settings and varying opponent strategies, the PG algorithm maximizes the utility of bidders. We hope that this article will inspire research on automatic bidding strategies for strategic bidders.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. HOBA: Hierarchical On-Policy Bidding Agents for Adaptive Online Advertising

    cs.AI 2026-06 reject novelty 6.0 of 10

    A three-tier bidding system (LLM hyperparameters, SARSA expert selection, fixed expert pool) reports +3.6% target-cost and +8.1% conversion gains in production A/B tests.

Pith tools