Pith. sign in

REVIEW 5 cited by

Using Large Language Models to Generate, Validate, and Apply User Intent Taxonomies

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2309.13063 v3 pith:VR2UNBOD submitted 2023-09-14 cs.IR cs.AIcs.CL

classification cs.IRcs.AIcs.CL
keywords userdataintentintentssearchapplygeneratelarge
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Log data can reveal valuable information about how users interact with Web search services, what they want, and how satisfied they are. However, analyzing user intents in log data is not easy, especially for emerging forms of Web search such as AI-driven chat. To understand user intents from log data, we need a way to label them with meaningful categories that capture their diversity and dynamics. Existing methods rely on manual or machine-learned labeling, which are either expensive or inflexible for large and dynamic datasets. We propose a novel solution using large language models (LLMs), which can generate rich and relevant concepts, descriptions, and examples for user intents. However, using LLMs to generate a user intent taxonomy and apply it for log analysis can be problematic for two main reasons: (1) such a taxonomy is not externally validated; and (2) there may be an undesirable feedback loop. To address this, we propose a new methodology with human experts and assessors to verify the quality of the LLM-generated taxonomy. We also present an end-to-end pipeline that uses an LLM with human-in-the-loop to produce, refine, and apply labels for user intent analysis in log data. We demonstrate its effectiveness by uncovering new insights into user intents from search and chat logs from the Microsoft Bing commercial search engine. The proposed work's novelty stems from the method for generating purpose-driven user intent taxonomies with strong validation. This method not only helps remove methodological and practical bottlenecks from intent-focused research, but also provides a new framework for generating, validating, and applying other kinds of taxonomies in a scalable and adaptable way with reasonable human effort.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 11 citations worldwide. Full citation record

  1. AI is the Strategy: From Agentic AI to Autonomous Business Models onto Strategy in the Age of AI

    cs.CY 2025-06 conditional novelty 6.0 of 10

    The paper defines Autonomous Business Models as business models where agentic AI executes value creation, delivery, and capture, and argues this changes strategy, competition, and leadership.

  2. Reading between the Lines: Can LLMs Identify Cross-Cultural Communication Gaps?

    cs.CL 2025-02 conditional novelty 6.0 of 10

    A user study of 57 preselected Goodreads reviews finds culture-specific comprehension gaps in most texts, while GPT-4o identifies the relevant spans with 0.49 precision and 0.65 recall across India, Mexico, and the USA.

  3. Automatic Labelling with Open-source LLMs using Dynamic Label Schema Integration

    cs.CL 2025-01 conditional novelty 6.0 of 10

    RAC ranks label descriptions by semantic similarity, then runs iterative binary LLM checks one label at a time, improving zero-shot labelling accuracy and enabling a precision-coverage trade-off.

  4. Large Action Models: From Inception to Implementation

    cs.AI 2024-12 conditional novelty 6.0 of 10

    A four-phase training pipeline converts a 7B language model into a Windows GUI action model that reaches 81.2% offline and 71.0% online task success on the authors' Word test set, beating text-only GPT-4o.

  5. Synergizing LLMs and Knowledge Graphs: A Novel Approach to Software Repository-Related Question Answering

    cs.SE 2024-12 conditional novelty 4.0 of 10

    An LLM plus a repository knowledge graph answers software repository questions with 84% accuracy when few-shot chain-of-thought prompting is added, outperforming an intent-based bot and web-search GPT-4o.

Pith tools