Query reformulation bridges user intent and retrieval in e-commerce search, yet production systems optimize rewrite quality and retrieval effectiveness separately, leaving the two stages structurally misaligned. Path-based architectures unify them end-to-end but were designed for personalization, where relevance is not an explicit constraint—search additionally requires the rewrite to remain faithful to the user’s stated query intent. Transplanted directly, these models learn a shortcut we term the generic-word dominance effect: they favor generic rewrites that score well on paths but drift from query intent. To address this, we propose SPEAR (Selection-aware Personalized End-to-end Adaptive Rewriting and Retrieval), which integrates three components that each target one failure mode: (1) a dual-embedding backbone with auxiliary loss and gradient isolation that shields recall-side semantics from being eroded by CTR-driven ranking signals; (2) a multiplicative gating aggregator that lets a rewrite score high only when both its confidence and item relevance are strong, eliminating the generic-word shortcut; (3) a Dynamic Rewrite Selector that jointly generates request-specific rewrite weights and user-query-conditioned scale and bias terms, allowing both rewrite preference and relevance calibration to adapt to each request. Offline evaluation on 100K held-out industrial search sessions shows that the proposed framework improves rewrite semantic similarity@10 by +18.2% and click recall@10 by +99.5% over the production baseline. In online A/B testing, SPEAR achieves +0.259% in query-view CTR and +0.733% in average reading depth, confirming that improved rewrite selection translates into stronger retrieval and deeper user engagement. The proposed SPEAR system has been fully deployed in Dewu’s community search platform since 2025. Our code is available at https://github.com/mallocagi1-cell/spear.
This paper presents a scalable system for discovery-augmented search that leverages intent-conditioned recall expansion, and addresses the cost-quality tradeoff of generative retrieval through a two-stage hybrid architecture.
Ji Xin, Xiao Xiao, Ishan Bhatt et al.· arXiv.org· 0 citations
Generative Retrieval (GR) is promising for e-commerce search, yet existing methods struggle to maintain query-intent consistency throughout the training pipeline. First, semantic ID (SID) construction based on static product information limits the ability of SIDs to encode product-intent associations. Second, although...
Jiayi Tuo, He-Han Li, Dong-Jun Fu et al.· 0 citations
E-commerce platforms increasingly display clickable query suggestions alongside items in the user feed, enabling users to refine or expand their intent without manually reformulating queries. Existing approaches either mine suggestions from historical logs -- limited to past behavior and blind to long-tail, personalize...
Shu-Wei Yuan, Mingyu Ding, Lu-Xin Liu et al.· 0 citations
This work proposes UniR, a decoder-only Transformer that unifies Generative and Multi-Objective ranking within a single heterogeneous sequence comprising user context, SID trajectory, and item features, validating the practicality of unified model in large-scale recommendation systems.
Efficient vector similarity search is critical for Retrieval-Augmented Generation (RAG) systems and other real-time AI applications. However, most existing methods are optimized for isolated queries and fail to leverage the continuity and correlation inherent in real-world query streams, such as those in multi-turn dia...
Zhuanglin Zheng, Yuxiang Zeng, Yunzhen Chi et al.· Proceedings of the 32nd ACM...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.