Pith. sign in

REVIEW 1 cited by

A Bag of Tricks for Scaling CPU-based Deep FFMs to more than 300m Predictions per Second

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2407.10115 v1 pith:374ZKCXO submitted 2024-07-14 cs.LG cs.AIcs.IR

classification cs.LGcs.AIcs.IR
keywords deepffmsclick-throughcpu-onlydetailin-housemodelprediction
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Field-aware Factorization Machines (FFMs) have emerged as a powerful model for click-through rate prediction, particularly excelling in capturing complex feature interactions. In this work, we present an in-depth analysis of our in-house, Rust-based Deep FFM implementation, and detail its deployment on a CPU-only, multi-data-center scale. We overview key optimizations devised for both training and inference, demonstrated by previously unpublished benchmark results in efficient model search and online training. Further, we detail an in-house weight quantization that resulted in more than an order of magnitude reduction in bandwidth footprint related to weight transfers across data-centres. We disclose the engine and associated techniques under an open-source license to contribute to the broader machine learning community. This paper showcases one of the first successful CPU-only deployments of Deep FFMs at such scale, marking a significant stride in practical, low-footprint click-through rate prediction methodologies.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. DCN^2: Interplay of Implicit Collision Weights and Explicit Cross Layers for Large-Scale Recommendation

    cs.IR 2025-06 conditional novelty 5.0 of 10

    DCN^2 augments DCNv2 with collision-weighted lookups, a dense-only cross layer, and an FFM-like similarity layer, and reports improved offline and online recommendation performance.

Pith tools