# Build a four-provider benchmark for non-brand buyer questions. - working worksheet

Reviewed alongside the Eli guide: https://hireeli.io/resources/four-provider-non-brand-benchmark

## Decision to make

Create a defensible ChatGPT, Gemini, Claude and Perplexity baseline

Owner: ____________________
Review date: ____________________
Canonical page or destination: ____________________

## Buyer questions

- [ ] How do you benchmark non-brand AI visibility across four providers?
- [ ] How many questions belong in the first benchmark?
- [ ] How should the benchmark be repeated?

## Evidence-led work plan

### 1. Freeze the comparison context

Write a measurement contract before making calls.

Evidence or artifact: ____________________
Status: [ ] not started  [ ] working  [ ] verified  [ ] blocked

### 2. Use a balanced commercial portfolio

Cover category choice, best-for, comparison, alternatives, objections, implementation, pricing and risk.

Evidence or artifact: ____________________
Status: [ ] not started  [ ] working  [ ] verified  [ ] blocked

### 3. Store observations, not just scores

Keep the raw answer, literal company mentions, named competitors, linked sources, timestamp and failure state.

Evidence or artifact: ____________________
Status: [ ] not started  [ ] working  [ ] verified  [ ] blocked

### 4. Repeat without overstating movement

Rerun the same cohort and compare provider-specific results.

Evidence or artifact: ____________________
Status: [ ] not started  [ ] working  [ ] verified  [ ] blocked

## Source ledger

- [ ] Google Search Central: AI features and your website: https://developers.google.com/search/docs/appearance/ai-features
- [ ] OpenAI: publishers and developers FAQ: https://help.openai.com/en/articles/12627856-publishers-and-developers-faq
- [ ] Perplexity: crawler documentation: https://docs.perplexity.ai/docs/resources/perplexity-crawlers
- [ ] Eli AI Visibility Index methodology: https://hireeli.io/ai-visibility-index/methodology

## Measurement boundary

Baseline date and exact context: ____________________
7-day observation: ____________________
14-day observation: ____________________
30-day observation: ____________________
Known limitations: ____________________

Keep technical eligibility, citations, recommendations, visits, leads and revenue as separate evidence. This worksheet does not guarantee an external ranking or commercial result.
