Search arXiv⌕ Search

arXiv · 2009.03258

Personalized Review Ranking for Improving Shopper's Decision Making: A Term Frequency based Approach

Abstract

User-generated reviews serve as crucial references in shopper's decision-making process. Moreover, they improve product sales and validate the reputation of the website as a whole. Thus, it becomes important to design reviews ranking methods that help shoppers make informed decisions quickly. However, reviews ranking has its unique challenges. First, there is no relevance labels for reviews. A relevant review for shopper A might not be relevant to shopper B. Second, since shoppers cannot click on reviews, we have no ways of getting relevance feedback. Eventually, reviews ranking suffers from the lack of ground truth due to the variability in the standard of relevance for different users. In this paper, we aim to address the challenges of helping users to find information they might be interested in from the sea of customer reviews. Using the Amazon Customer Reviews Dataset collected and organized by UCSD, we first constructed user profiles based on user's personal web trails, recent shopping history and previous reviews, incorporated user profiles into our ranking algorithm, and assigned higher ranks to reviews that address individual shopper's concerns to the largest extent. Also, we leveraged user profiles to recommend products based on reviews texts. We evaluated our model based on both empirical evaluations and numerical evaluations of review scores. The results from both evaluation methods reveal a significant increase in the quality of top reviews as well as user satisfaction for over 1000 products. Our reviews based recommendation system also suggests that there's a large chance of user viewing and liking the product we recommend. Our work shows the basic steps of developing a ranking method that learns from a particular end-user's preferences.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Akhil Sai Peddireddy. 2020-09-07. Personalized Review Ranking for Improving Shopper's Decision Making: A Term Frequency based Approach. https://arxiv.org/abs/2009.03258

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Component Benchmark: Hierarchical Model Profiling for Large-scale Recommendation Systems

Large-scale recommendation models pose distinct, under-explored profiling challenges. Most recommendation model architectures are structurally heterogeneous, intermixing memory-bandwidth-bound operations, small compute-bound dense layers, dynamic shapes from jagged categorical features, and low-arithmetic-intensity operations. Recommendation models evolve rapidly as modeling engineers experiment with compositions, often written without visibility into hardware execution characteristics. Standard profiling tools offer either end-to-end throughput or operator-level traces, but cannot attribute performance to the submodules that practitioners reason about. We present Component Benchmark (CB), a profiling system that independently characterizes each submodule performance in a hierarchical manner, providing a tree-structured, interactive visualization that brings performance clarity to ML practitioners. At its core, CB provides a simple yet extensible, submodule-based benchmarking framework with a plugin architecture that enables hierarchical performance analysis. These large-scale recommendation models are TB-scale, run on thousands of GPUs and ingest 100B examples per day. We demonstrate CB's effectiveness on common open sourced models and discuss how CB has been leveraged to accelerate modern recommendation model performance analysis and optimization.

cs.IR↗

Recommendation World Models for Future-State Control

Sequential recommendation optimizes which items to rank, while each displayed slate also shapes subsequent feedback and user state. We study how a trained ranker can support decisions about these future consequences. We introduce UA-TWM, a utility-anchored world-model interface that constructs nearby slate actions, estimates their target-relevant consequences, and selects an alternative subject to utility constraints. The reference slate serves as a fallback when no alternative qualifies. A logged-replay instantiation combines utility and target-gain estimates with calibrated failure-risk prediction; a closed-loop instantiation uses one-step state-action prediction and updates its decisions after observed feedback. We evaluate transfer across twelve sequential backbones on MovieLens-25M and KuaiRand-Pure, and repeated target-directed interaction in KuaiSim. Attaching the interface improves Recall@20, NDCG@20, and future-state alignment for every matched logged backbone. Selection ablations reveal the utility and risk costs of aggressive target pursuit, while closed-loop diagnostics isolate the contribution of action-conditioned prediction. Local consequence modeling thus enables target-aware selection around a trained sequential ranker.

cs.IR↗

RecToolBench: Benchmarking Recommendation-Specific Tool Orchestration under Fuzzy User Intent

Recent advances in agentic recommender systems are shifting recommender systems from passive filtering engines to instruction-following agents that use external tools to resolve user intent. However, existing benchmarks often assume explicit user intent, simplified tool environments, or isolated function calls, leaving realistic tool orchestration for recommendation underexplored. To bridge this gap, we propose RecToolBench, a Model Context Protocol (MCP)-based benchmark for evaluating tool-using recommender agents under fuzzy user instructions. RecToolBench contains more than 1,200 executable tasks across three recommendation domains, 13 MCP servers, and 32 tools, spanning single-tool calls, parallel tool calls, sequential tool chains, and hybrid tool orchestration. We construct RecToolBench with a scalable synthesize--fuzzify--judge pipeline that generates executable fuzzy recommendation tasks, and evaluates agent trajectories using rule-based execution checks and rubric-based LLM evaluation. Experiments on representative LLMs show that syntactically valid tool calls do not guarantee successful recommendations. Models struggle with semantic parameter grounding, multi-step evidence integration, and grounded final recommendations, especially as orchestration complexity increases. Our results identify tool orchestration under fuzzy user intent as a major bottleneck for agentic recommender systems. Our data and code are available at https://github.com/ShawnChenn/RecToolBench.

cs.IR↗