Pith. sign in

REVIEW 1 cited by

SAGE: Structured Attribute Value Generation for Billion-Scale Product Catalogs

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2309.05920 v1 pith:VD74QYWR submitted 2023-09-12 cs.IR cs.AIcs.CL

classification cs.IRcs.AIcs.CL
keywords attributesagevaluescatalogspredictingproducttaskacross
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We introduce SAGE; a Generative LLM for inferring attribute values for products across world-wide e-Commerce catalogs. We introduce a novel formulation of the attribute-value prediction problem as a Seq2Seq summarization task, across languages, product types and target attributes. Our novel modeling approach lifts the restriction of predicting attribute values within a pre-specified set of choices, as well as, the requirement that the sought attribute values need to be explicitly mentioned in the text. SAGE can infer attribute values even when such values are mentioned implicitly using periphrastic language, or not-at-all-as is the case for common-sense defaults. Additionally, SAGE is capable of predicting whether an attribute is inapplicable for the product at hand, or non-obtainable from the available information. SAGE is the first method able to tackle all aspects of the attribute-value-prediction task as they arise in practical settings in e-Commerce catalogs. A comprehensive set of experiments demonstrates the effectiveness of the proposed approach, as well as, its superiority against state-of-the-art competing alternatives. Moreover, our experiments highlight SAGE's ability to tackle the task of predicting attribute values in zero-shot setting; thereby, opening up opportunities for significantly reducing the overall number of labeled examples required for training.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. CatalogAgent: A Supervisor-mediated Self-Learning System Enabling Context Engineering for GenAI Models

    cs.AI 2026-07 conditional novelty 6.0 of 10

    A supervisor AI mediates generator/evaluator disagreements on product attributes and feeds summarized lessons back into worker prompts, improving accuracy by up to about 15% on selected attributes.

Pith tools