ARXIV:2601.18225 · AGENTS · SUBMITTED 02 APR · 02:30 UTC · FRESHNESS STALE

VerifiedSource: PDF linkedPartialPaperPack: 3 of 4 citation fields filledMissingMissing fields: authorsPartialProof: unverified proof status

ShopSimulator: Evaluating and Exploring RL-Driven LLM Agent for Shopping Assistants

arXiv

An RL-driven LLM shopping assistant that improves product search and personalization in e-commerce.

Blocked on Code›Score7.0Evidence unverified

Opportunity summary

Pain An RL-driven LLM shopping assistant that improves product search and personalization in e-commerce.

Evidence 0 refs | 0 sources | 17% coverage

Blocker Evidence unverified

Open Build Read PDF Signal Canvas Track

PROBLEM

An RL-driven LLM shopping assistant that improves product search and personalization in e-commerce. To perform thorough, user-tailored product searches, agents should interpret personal preferences, engage in multi-turn dialogues, and ultimately retrieve and discriminate among…

METHOD

Full abstract

Large language model (LLM)-based agents are increasingly deployed in e-commerce shopping. To perform thorough, user-tailored product searches, agents should interpret personal preferences, engage in multi-turn dialogues, and ultimately retrieve and discriminate among highly similar products. However, existing research has yet to provide a unified simulation environment that consistently captures all of these aspects, and always focuses solely on evaluation benchmarks without training support. In this paper, we introduce ShopSimulator, a large-scale and challenging Chinese shopping environment. Leveraging ShopSimulator, we evaluate LLMs across diverse scenarios, finding that even the best-performing models achieve less than 40% full-success rate. Error analysis reveals that agents struggle with deep search and product selection in long trajectories, fail to balance the use of personalization cues, and to effectively engage with users. Further training exploration provides practical guidance for overcoming these weaknesses, with the combination of supervised fine-tuning (SFT) and reinforcement learning (RL) yielding significant performance improvements. Code and data will be released at https://github.com/ShopAgent-Team/ShopSimulator.

RESULT

ScienceToStartup currently rates this 7.0/10 on the public viability pass. However, existing research has yet to provide a unified simulation environment that consistently captures all of these aspects, and always focuses solely on evaluation…

WHY NOW

Agents moved forward this cycle; last verified April 2026. Public score 7.0/10.

Continue into Read for claims, analysis, references, and neighboring papers.

Opportunity summary

Score7.0

PainAn RL-driven LLM shopping assistant that improves product search and personalization in e-commerce.

Evidence0 refs | 0 sources | 17% coverage

Blockermissing authors

Analysis summary

An RL-driven LLM shopping assistant that improves product search and personalization in e-commerce.

VerifiedSource: PDF linkedPartialPaperPack: 3 of 4 citation fields filledMissingMissing fields: authorsPartialProof: unverified proof status

References(15)

Reference metadata pending (afe48a826867fec5a06c940f62fc92591edd45e7)

Reference metadata pending (e05795cdb5b0c11e72016c79122777c5b3a7ed48)

Reference metadata pending (7444c7eceba6bbedddbe278c97ce1b75ae2a05de)

Reference metadata pending (05735041d30f9c0fc05e2b2a183a7543caab6965)

Reference metadata pending (8f1f2c9c04143df7e7b2a096eae2c9655a3d0978)

Reference metadata pending (a9c8f5afe6a47a59209617a0b8c324d4dc0d19db)

Reference metadata pending (7956e2c2e829658db221b18232c8e83403adc1a7)

Reference metadata pending (a3683fa4943f1bea24f769584e30c062d15305e2)

Reference metadata pending (6cb372f508d25322d5c5568c324c132af4e4be32)

Reference metadata pending (79d11dd9c97e8727eed7ea31e31fb561e42c9523)

Reference metadata pending (bdd52171bdd1de6b5c22f5e345ee0b1efff48085)

Reference metadata pending (64e802ea8e9dbe247c31fb06184c04dbf9e55e4e)

Reference metadata pending (b486982fa7c68a8a08df1111ba9607119419c488)

Reference metadata pending (0d1c76d45afa012ded7ab741194baf142117c495)

Reference metadata pending (23525374cfd3af714f3ffb7a203b1ef3253333fe)

{ "contract_version": "paper-r2", "paper_id": "d3992e2e-c85c-4c2d-9b1e-01457dd2075f", "arxiv_id": "2601.18225", "canonical_route": "/paper/shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants", "active_tab": "synced from current hash by the drawer client", "selected_artifact": "shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants", "endpoints": { "paper_pack": "/api/v1/paper/shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants/paper-pack", "build_passport": "/api/v1/paper/shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants/build-passport", "mcp_resource": "sciencetostartup://surfaces/paper-workspace" } }

{ "surface": "paper", "mode": "paper", "query": "ShopSimulator: Evaluating and Exploring RL-Driven LLM Agent for Shopping Assistants", "normalized_query": "2601.18225", "route": "/paper/shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants", "paper_ref": "shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants", "topic_slug": null, "benchmark_ref": null, "dataset_ref": null }

{ "@context": "https://schema.org", "@graph": [ { "@type": "WebPage", "@id": "https://sciencetostartup.com/paper/shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants#webpage", "url": "https://sciencetostartup.com/paper/shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants", "name": "ShopSimulator: Evaluating and Exploring RL-Driven LLM Agent for Shopping Assistants", "description": "An RL-driven LLM shopping assistant that improves product search and personalization in e-commerce.", "isPartOf": { "@id": "https://sciencetostartup.com/#website" } }, { "@type": "ScholarlyArticle", "@id": "https://sciencetostartup.com/paper/shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants#scholarlyArticle", "headline": "ShopSimulator: Evaluating and Exploring RL-Driven LLM Agent for Shopping Assistants", "description": "An RL-driven LLM shopping assistant that improves product search and personalization in e-commerce.", "url": "https://sciencetostartup.com/paper/shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants", "sameAs": "https://arxiv.org/abs/2601.18225", "identifier": { "@type": "PropertyValue", "propertyID": "arXiv", "value": "2601.18225" }, "isAccessibleForFree": true, "isPartOf": { "@id": "https://sciencetostartup.com/#website" }, "datePublished": "2026-01-26T07:24:28.000Z", "citation": [ { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "afe48a826867fec5a06c940f62fc92591edd45e7" }, "url": "https://www.semanticscholar.org/paper/afe48a826867fec5a06c940f62fc92591edd45e7" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "e05795cdb5b0c11e72016c79122777c5b3a7ed48" }, "url": "https://www.semanticscholar.org/paper/e05795cdb5b0c11e72016c79122777c5b3a7ed48" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "7444c7eceba6bbedddbe278c97ce1b75ae2a05de" }, "url": "https://www.semanticscholar.org/paper/7444c7eceba6bbedddbe278c97ce1b75ae2a05de" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "05735041d30f9c0fc05e2b2a183a7543caab6965" }, "url": "https://www.semanticscholar.org/paper/05735041d30f9c0fc05e2b2a183a7543caab6965" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "8f1f2c9c04143df7e7b2a096eae2c9655a3d0978" }, "url": "https://www.semanticscholar.org/paper/8f1f2c9c04143df7e7b2a096eae2c9655a3d0978" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "a9c8f5afe6a47a59209617a0b8c324d4dc0d19db" }, "url": "https://www.semanticscholar.org/paper/a9c8f5afe6a47a59209617a0b8c324d4dc0d19db" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "7956e2c2e829658db221b18232c8e83403adc1a7" }, "url": "https://www.semanticscholar.org/paper/7956e2c2e829658db221b18232c8e83403adc1a7" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "a3683fa4943f1bea24f769584e30c062d15305e2" }, "url": "https://www.semanticscholar.org/paper/a3683fa4943f1bea24f769584e30c062d15305e2" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "6cb372f508d25322d5c5568c324c132af4e4be32" }, "url": "https://www.semanticscholar.org/paper/6cb372f508d25322d5c5568c324c132af4e4be32" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "79d11dd9c97e8727eed7ea31e31fb561e42c9523" }, "url": "https://www.semanticscholar.org/paper/79d11dd9c97e8727eed7ea31e31fb561e42c9523" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "bdd52171bdd1de6b5c22f5e345ee0b1efff48085" }, "url": "https://www.semanticscholar.org/paper/bdd52171bdd1de6b5c22f5e345ee0b1efff48085" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "64e802ea8e9dbe247c31fb06184c04dbf9e55e4e" }, "url": "https://www.semanticscholar.org/paper/64e802ea8e9dbe247c31fb06184c04dbf9e55e4e" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "b486982fa7c68a8a08df1111ba9607119419c488" }, "url": "https://www.semanticscholar.org/paper/b486982fa7c68a8a08df1111ba9607119419c488" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "0d1c76d45afa012ded7ab741194baf142117c495" }, "url": "https://www.semanticscholar.org/paper/0d1c76d45afa012ded7ab741194baf142117c495" }, { "@type": "ScholarlyArticle", "identifier": { "@type": "PropertyValue", "propertyID": "SemanticScholar", "value": "23525374cfd3af714f3ffb7a203b1ef3253333fe" }, "url": "https://www.semanticscholar.org/paper/23525374cfd3af714f3ffb7a203b1ef3253333fe" } ], "additionalProperty": [ { "@type": "PropertyValue", "propertyID": "viabilityScore", "value": 7 }, { "@type": "PropertyValue", "propertyID": "researchDomain", "value": "Agents" } ] }, { "@type": "BreadcrumbList", "itemListElement": [ { "@type": "ListItem", "position": 1, "name": "Home", "item": "https://sciencetostartup.com" }, { "@type": "ListItem", "position": 2, "name": "Agents", "item": "https://sciencetostartup.com/topics" }, { "@type": "ListItem", "position": 3, "name": "ShopSimulator: Evaluating and Exploring RL-Driven LLM Agent ", "item": "https://sciencetostartup.com/paper/shopsimulator-evaluating-and-exploring-rl-driven-llm-agent-for-shopping-assistants" } ] } ] }