Khaled Boughanmi, Kamel Jedidi, Nour Jedidi
arXiv 18 Oct 2025 · Statistics — Machine Learning
arXiv:2510.16551 · PDF · DOI · OpenAlex · Extracted main text
This research proposes a systematic, large language model (LLM) approach for extracting product and service attributes, features, and associated sentiments from customer reviews. Grounded in marketing theory, the framework distinguishes perceptual attributes from actionable features, producing interpretable and managerially actionable insights. We apply the methodology to 20,000 Yelp reviews of Starbucks stores and evaluate eight prompt variants on a random subset of reviews. Model performance is assessed through agreement with human annotations and predictive validity for customer ratings. Results show high consistency between LLMs and human coders and strong predictive validity, confirming the reliability of the approach. Human coders required a median of six minutes per review, whereas the LLM processed each in two seconds, delivering comparable insights at a scale unattainable through manual coding. Managerially, the analysis identifies attributes and features that most strongly influence customer satisfaction and their associated sentiments, enabling firms to pinpoint "joy points," address "pain points," and design targeted interventions. We demonstrate how structured review data can power an actionable marketing dashboard that tracks sentiment over time and across stores, benchmarks performance, and highlights high-leverage features for improvement. Simulations indicate that enhancing sentiment for key service features could yield 1-2% average revenue gains per store.
appendix boundary found by appendix_command · 71% of the source is main text. Read the extracted text to check this.
The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.
| Reference | Intensity | Mentions | Sections | Main text | |
|---|---|---|---|---|---|
| 1 | Büschken, Joachim and Greg M Allenby (2020) Improving text analysis using sentence conjunctions and punctuation | 0.928 | 4 | 3 | 100% |
| 2 | Chakraborty, Ishita, Minkyung Kim, and K Sudhir (2022) Attribute sentiment scoring with online text reviews: Accounting for language structure and missing attributes | 0.843 | 3 | 3 | 100% |
| 3 | Wei, Jason, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed… (2022) Chain-of-Thought Prompting Elicits Reasoning in Large Language Models | 0.737 | 3 | 2 | 100% |
| 4 | Bojic, Ljubisa, Olga Zagovora, Asta Zelenkauskaite, Vuk Vukovic, Mil… (2025) Evaluating Large Language Models Against Human Annotators in Latent Content Analysis: Sentiment, Political Leaning, Emotional In… | 0.644 | 2 | 2 | 100% |
| 5 | Brown, Tom B., Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D.… (2020) Language models are few-shot learners | 0.644 | 2 | 2 | 100% |
| 6 | Gutman, Jonathan (1982) A means-end chain model based on consumer categorization processes | 0.644 | 2 | 2 | 100% |
| 7 | Luca, Michael (2016) Reviews, reputation, and revenue: The case of Yelp. com | 0.644 | 2 | 2 | 100% |
| 8 | Schoenmueller, Verena, Oded Netzer, and Florian Stahl (2020) The polarity of online reviews: Prevalence, drivers and implications | 0.644 | 2 | 2 | 100% |
| 9 | Yelp Inc (2025) Yelp Open Dataset, (2024) | 0.644 | 2 | 2 | 100% |
| 10 | Berger, Jonah, Ashlee Humphreys, Stephan Ludwig, Wendy W Moe, Oded N… (2019) Uniting the Tribes: Using Text for Marketing Insight | 0.511 | 2 | 1 | 100% |
Showing the top 10 of 58 scored citations.