The exact prompts, engines and formula behind every scorecard — nothing hidden, no black box. See it applied in a real example report.
Every check asks the same battery, with your category, brand, audience and region filled in (defaults: audience "everyday use", region "Europe"). Verbatim answers appear in your report.
| Discovery | What are the best places to buy {category} online? |
| Discovery | Recommend some good online stores for {category}. |
| Discovery | I want to buy {category}. Which brands or shops should I look at? |
| Discovery | Best {category} in 2026 — what do you recommend? |
| Evaluation | Is {brand} a good place to buy {category}? |
| Evaluation | What do people say about {brand}? |
| Comparison | Compare the top online shops for {category}. |
| Comparison | {brand} vs alternatives for {category} — which is better? |
| Audience fit | Best {category} for {audience}? |
| Audience fit | Affordable {category} with good quality — where should I buy? |
| Gifts & logistics | I need {category} as a gift. Where should I order from? |
| Gifts & logistics | Where can I get {category} delivered to {region}? |
We ask each prompt once per engine, through the same public APIs the assistants run on, with no system prompt or steering — the answer is what the model gives an ordinary user. Every answer in your report carries the exact model version and a UTC timestamp.
We also record sentiment around your mentions (cue-word window, marked positive/neutral/negative), which answers cite your own domain as a source, and every "gap prompt" where a competitor was named and you weren't.