Shopping is moving from search boxes to AI assistants. AIBUY asks 8 assistants the same 100 purchase-intent prompts per category, logs every answer, and publishes the recommendation data — so buyers get evidence and merchants get a scoreboard. No paid placements. Ever.
For twenty years, merchants fought for Google rank. Now their customers ask an assistant — and the assistant answers from a mix of training data, third-party lists and crawlable pages. If AI misstates your specs, or skips you entirely, no one tells you. We tell you, with data.
One hundred purchase-intent prompts per category — the way real buyers actually ask — run through every major assistant, on a schedule, with full answer logging.
Our method →Mention rate, rank position, info accuracy and competitive suppression — four metrics, published with prompt-level evidence. Same method in the US and China markets.
Cross-market project →We are the auditors, not the shelf. Rankings on AIBUY cannot be bought — that independence is what makes the data worth reading and the service worth paying for.
Independence policy →Data collection, testing, charts and first drafts run automatically. Conclusions, opinions and anything we publish as a recommendation are written and checked by a person. Each report opens with a card like this:
The full pilot — 40 prompts × 8 assistants — is being published with its complete prompt log and answer excerpts. Headline numbers from the dry run:
| Assistant | Mention rate (top-3) | Avg. first position | Spec accuracy | Overlap with ChatGPT |
|---|---|---|---|---|
| ChatGPT | 78% | 1.6 | 91% | — |
| Claude | 64% | 2.1 | 88% | 71% |
| Gemini | 58% | 2.4 | 82% | 66% |
| Perplexity | 52% | 2.8 | 86% | 61% |
| Doubao | 44% | 3.3 | 74% | 38% |
| Kimi | 40% | 3.5 | 79% | 35% |
Pilot dry-run numbers, shown to illustrate the report format. Final published tables carry the full prompt log, dates and model versions. Full report →
AIBUY runs the identical purchase-intent prompts against the US assistant set (ChatGPT, Claude, Gemini, Perplexity) and the China set (Doubao, Yuanbao, Kimi, Wenxin, Taobao Ask). Only a team operating in both markets can produce this comparison — and we publish it quarterly, free, in English and Chinese.
You optimized for search for a decade. Your buyers now ask assistants directly. Our diagnostic shows exactly where you stand — and the managed service keeps you standing there.
100 purchase-intent prompts × 8 assistants, focused on your category. You get your AI Recommendation Share, where you rank, which specs the assistants get wrong, and who is winning your customers' questions — with the prompt-level evidence.
What's inside →Monthly re-tests and a trend dashboard: share movements, new entrants, spec errors, and the visibility fixes that matter — structured data, crawl access, and presence in the third-party sources assistants actually quote.
How it works →One email a week: what the assistants recommended, what changed, and one dataset worth your time. Free, no algorithm between us — unsubscribe anytime.
Weekly · English · sample issue in the first send. 中文版请订阅 aibuy.org.cn。
No. Reports are funded by our diagnostic and monitoring services for merchants, and by newsletter sponsorships that are always labeled. Ranking positions are computed from test data only.
US set: ChatGPT, Claude, Gemini, Perplexity. China set: Doubao, Tencent Yuanbao, Kimi, Wenxin, Taobao Ask. The cross-market project queries both sets with identical prompts.
We publish the prompts, the raw answer log, and the counting rules. You can disagree with us — but you can check us. Affiliate links, where present, are labeled and never influence scoring.
Yes — the diagnostic is a one-off deliverable, and we run a limited free pilot each quarter for categories we're covering. See services →