The verdict
How many of 150 would buy, book, sign up, or stay unsure — plus your peer percentile in the selected lane.
Intent distribution · mean score · benchmark rankSynthetic audience testing · built on peer-reviewed methodology
Paste your URL. A fixed panel of 150 realistic buyers from your industry reads your page and reacts in plain words. You see how many would buy — or book — how you rank against 1,100+ scored pages, and exactly what stops the rest — plus a heatmap of where attention actually lands and a UX review of the friction standing in their way.
Report in ~5 minutes · $49 one-time · auto-refund if we can't produce it · see a sample report →
What's a synthetic panel? 150 AI-simulated buyers, each modeled on how real customers in your market think and decide — a method validated against 9,300 real survey answers. Not real people, and we never pretend otherwise.
We read only the public page at the URL you paste — no account, no tracking pixels on your site, deleted on request. How we handle data →
“How likely are you to subscribe to this?”
“The free tier sounds generous, but I can't tell what the paid plan actually adds. I'd try it and probably never upgrade.”
Built for your industry
A coaching site is judged by coaching prospects — ROI skeptics, comparison shoppers, people ready to invest. A SaaS page is judged by SaaS buyers. Same method, your market.
Marketing, design, dev. See who would book an intro call — and what stops the rest.
Private practice. See who would book a consultation, in plain words.
Landing pages, app listings, launches. Test variants and competitors before traffic.
Three ways to run it: Diagnostic $49 · Variant ranking $79 — rank 2–4 versions, the method's strongest mode · Competitor scan $99. See pricing ↓
In every report
The panel's answer is only useful if you can see what caused it. Every report ties purchase intent to objections, attention, UX friction, and the next changes worth testing.
Open a full sample report →How many of 150 would buy, book, sign up, or stay unsure — plus your peer percentile in the selected lane.
Intent distribution · mean score · benchmark rankThe recurring objections behind the score, clustered by theme and backed by plain-language buyer quotes.
Top objections · affected segments · verbatim reactionsWhat each buyer type notices, what they miss, and which parts of the page fail to carry enough attention.
Attention heatmap · fold math · reading orderA severity-ranked UX review of the page's primary task, accessibility basics, and the places visitors get stuck.
UX inspection · task walkthrough · fix-first listConcrete copy and layout changes tied back to the objections they address, so the report turns into a work queue.
Copy fixes · proof gaps · next-test ideasHow it works
The same four stages a research agency runs over a month — compressed into minutes, at a price a small business survives.
Screenshot + copy extraction of your website, landing page, or app listing. The panel reacts to exactly what visitors see — headline, images, pricing, all of it.
We suggest the industry lane your page should be benchmarked against. Nothing runs until you confirm the fixed panel and baseline.
Each persona reacts in free text — no forced ratings. Research shows direct numeric ratings from AI are unrealistic; natural reactions are where the signal lives.¹
Semantic Similarity Rating maps every reaction onto a purchase-intent distribution, benchmarked against 1,100+ scored pages — plus the top objections, quoted.
The science
150 Strangers implements Semantic Similarity Rating (SSR) — a technique developed by researchers at PyMC Labs and Colgate-Palmolive and tested against 57 real consumer surveys.
Maier, Aslak, Fiaschi, Rismal, Fletcher, Luhmann, Dow, Pappas & Wiecki (2025). arXiv:2510.08338 · open-source reference implementation on GitHub
Straight answers
A research tool you can't trust is worthless. So here is exactly where the method is strong — and where we'll refuse to oversell it.
Pricing
One-time payments. No subscription. If the scrape fails or the report can't be produced, it's auto-refunded.
Questions
Naively, yes — ask a model for a 1–5 rating and you get useless, middle-clustered answers. That's exactly what the underlying research demonstrated, and why we don't do it. SSR elicits natural-language reactions and maps them to ratings by semantic meaning. Validated against 9,300 real respondents across 57 surveys, it recovered ~90% of the reliability a repeated human panel achieves. Synthetic panels aren't a replacement for talking to customers — they're a way to arrive at those conversations with a sharper page.
It's for you — service businesses are now the majority of pages we score. Your page is judged by a fixed panel built around how your buyers decide (booking a call or consultation, not installing an app), and your percentile compares you against scored sites in your own industry. Start from your industry page — coaches, agencies, therapists — and the whole flow speaks your language.
We read your page and suggest an industry lane, then stop and show it to you. You can switch lanes before a single persona is generated. The scored panel is fixed for that lane, so your percentile compares against peer pages judged by the same buyer types.
Because the method can't honestly deliver one, and we'd rather be trusted than impressive. The validated strength of SSR is relative measurement — rankings, percentiles against a benchmark corpus, and the qualitative why. Synthetic intent distributions are systematically wider and slightly more critical than human ones, which makes them great discriminators and bad absolute forecasters.
The method works because models have absorbed enormous amounts of real customer conversation about most consumer domains. If your domain has thin coverage — deep tech, novel B2B categories — synthetic reactions get less trustworthy. We detect this and put a low-confidence banner on the report rather than hiding it. If a report is flagged and you don't find it useful, ask for a refund.
We screenshot and extract copy from the URL you give us, run the analysis, and store your report so you can revisit it. We don't train models on your pages, don't resell your data, and only ever test publicly accessible URLs you submit.
Constantly. This page's structure, copy, and even its price were chosen by running variants through our own panel — our two-price-point test came back identical to three decimal places across two independent 150-persona runs. We publish self-tests as we run them: see the 731-site coaching teardown in Guides. A method that can't survive its own scrutiny doesn't deserve yours.
Five minutes from now
Or you could keep guessing, spend on traffic, and find out from a flat conversion chart three weeks from now.
$49 · auto-refund if we can't produce your report