An Onsite AI Assistant Does Four Jobs. Most Teams Buy It for One
Every vendor in this category sells conversion. The teams getting the most out of an onsite assistant are the ones using it as a demand asset, a support channel and a listening post at the same time.
Viren Inaniyan · September 23, 2026 · AI Agent Storefronts

An onsite AI shopping assistant is usually bought as a conversion tool and then quietly does three other jobs that nobody budgeted for. The gap between those two facts is where most of the disappointment in this category comes from, and most of the upside.
We have been in a run of evaluation conversations with mid-size retailers this year. The pattern is consistent: the brief arrives as "improve website conversion", and within twenty minutes the same buyer is asking about order tracking, about whether ChatGPT can see their catalog, and about what the assistant is hearing that their analytics is not. Those are four jobs, not one.
Job one: conversion, the one on the invoice
This is the job every vendor sells and the only one they publish numbers for. Rep AI publishes "10-30% Lift in CVR". Envive publishes a "4X AVg Conversion Lift" for visitors who engage with its storefront. Alhena publishes a 20% AOV increase at Victoria Beckham. All captured from their own sites on 23 September 2026.
Read those carefully before you budget against them. Only one denominator in that set is stated, and we took the whole category's numbers apart in the conversion-claims audit. The short version: engaged-visitor conversion and sitewide conversion are different claims, and a lift quoted without the denominator tells you almost nothing about what will happen to your business.
Conversion is a real job. It is just the one where the published evidence is weakest.
Job two: support, before and after the sale
Buyers raise this unprompted, and it is usually the point where the conversation gets specific. Pre-sale is sizing, shipping windows, compatibility, returns policy. Post-sale is order tracking and refund status, which means the assistant needs read access to order state, not just the catalog.
The category is converging here from both directions. Gorgias came from support and now publishes that "1 in 7 (14%) Shopping Assistant conversations ends in an attributed order". The sales-first vendors all ship deflection metrics. The practical consequence for a buyer: ask which direction the vendor came from, because the roadmap and the pricing model still point that way.
Post-sale scope is the question to press on. An assistant that answers "where is my order" needs a live order lookup and an identity check, and that is a different integration from a catalog feed. Several products in this category stop at pre-sale and do not say so on the pricing page.
Job three: the catalog work is shared with off-site agents
This is the one almost nobody sells, and it is the one with the longest shelf life.
To answer "something warm and waterproof under £150 that does not look like hiking gear", an assistant needs attributes: material, waterproof rating, fit, colour family, price, availability at the variant level. Most catalogs do not carry them. The enrichment work to make an onsite assistant useful is substantial and it is one-time, with ongoing sync as products are added.
Here is the part that changes the business case. That is the same data an external agent reads. When a shopper asks ChatGPT, Gemini, Perplexity or Alexa for Shopping what to buy, products whose attributes cannot satisfy a constraint drop out of the answer before price is considered. The enrichment you do for the widget in the corner is the enrichment that decides whether you appear in an answer you will never see.
The surfaces are separate. The work is not. Any evaluation that scores an onsite assistant purely on its own conversion is undercounting it, because a second surface is being paid for at the same time. We set out what the external side actually reads in which product data agents use.
Job four: the demand you cannot currently hear
Site search records queries a shopper already knew how to phrase in your vocabulary. Everything else is invisible.
An assistant records the rest: occasion, recipient, relationship, budget framing, compatibility, constraint. A gifting retailer hears "something for my sister who just moved house, under 3,000". Nothing in a search log looks like that, and nothing in a category tree answers it.
Three things fall out of that log that no other system produces:
Gaps in the catalog. Repeated requests with no good match are a buying-plan input.
Gaps in the attributes. Questions the assistant could not answer are exactly the fields missing from your product data, which loops straight back to job three.
The vocabulary shoppers actually use, which is the input to category naming, filter design and the copy on your product pages.
We would go further: for a catalog under a few thousand SKUs, this log is often worth more in the first quarter than the conversion delta, because it changes decisions rather than moving a rate. That is a claim about where attention goes, not a measured number, and we would rather say it plainly than dress it up as a statistic.
What this means for how you evaluate
Score all four. A vendor that is excellent at conversion and has no post-sale scope, no attribute export and no query log is one product being sold as four.
Concretely, in the demo:
- Ask for the denominator behind every conversion figure, and the attribution window.
- Ask whether post-sale order and refund lookups are in scope, or a later tier.
- Ask whether the enriched attributes are exportable and who holds the semantic index. If you cannot get the enrichment back out, you are renting the work that also decides your visibility in external agents.
- Ask what the query log exposes, and whether unanswered questions are reported as a first-class metric rather than buried.
The four-job frame also settles the sequencing question. If qualified traffic is arriving and not converting, start at job one. If shoppers are never arriving because assistants are naming somebody else, no onsite tool reaches them - that gap is measured by Citation Rank, and House of Zelena is what closing it looked like over six months.
The agentic readiness scan checks whether your catalog can answer a constraint question at all, which is the precondition for jobs one, three and four.
Sources
- Rep AI, Envive and Alhena homepages, captured 23 September 2026.
- Gorgias, "AI shopping assistants are now officially influencing sales", gorgias.com/research, captured 23 September 2026.
- Tru Commerce citation corpus, 26,629 citations, 250 prompts, 5 engines, US/UK/India, captured 16 September 2026.
- Buyer-question patterns are drawn from Tru Commerce evaluation calls in September 2026, reported in aggregate and unattributed.
FAQ
Four things, though it is almost always sold as one. It converts traffic that already arrived, it answers pre-sale and post-sale support questions, it produces the enriched catalog that off-site agents like ChatGPT and Gemini read, and it captures demand signals your site search never records.
Both, and the split matters when you compare vendors. Support-first products lead with deflection rate and CSAT. Sales-first products lead with conversion and basket size. Read the metric a vendor puts first and you learn which product you are being sold.
It does not directly, but the work does. An assistant needs complete, structured, variant-level product attributes to answer a constraint question, and that is the same data an external agent needs before it will recommend you. The enrichment is shared even though the surfaces are separate.
Questions phrased the way shoppers think rather than the way your taxonomy is built - occasion, recipient, compatibility, constraint. Site search only records queries a shopper already knew how to phrase. The assistant records the rest, and that log is a merchandising input nothing else gives you.
Whichever one your loss is in. If qualified traffic arrives and leaves, measure conversion. If support volume is the problem, measure resolution. If you are not being named by AI assistants at all, no onsite tool reaches that shopper and the catalog work comes first.
Continue reading
September 23, 2026
Every Published AI Shopping Assistant Conversion Claim, Audited
Rep AI publishes 10-30% CVR lift. Envive publishes 4x. Alhena publishes 20% AOV. Gorgias publishes 14% of conversations. All four are probably true and none of them answer the question you are asking.
September 23, 2026
AI Shopping Assistant vs Chatbot vs Site Search vs Agent: The Four Are Not the Same
Four things get sold under one name. A chatbot deflects tickets, an assistant sells, site search retrieves, and an external agent decides whether a shopper reaches your site at all.
September 23, 2026
What It Actually Takes to Put an AI Shopping Assistant on Your Store
A script tag on a clean Shopify catalog is a two-day job. A custom stack with location-bound inventory and seventy attributes per SKU is a two-to-three week job. Here is what separates them.