# Message test — https://koalatative.com/

After reading your page, only 2 of 15 personas could name a reason to pick you over a similar option.

- **Page tested:** https://koalatative.com/
- **Audience tested against:** Ecom, saas, or b2b companies in doing at least 10 million per year. The buyer would typically be a senior leader in marketing, growth, UX, or product. They either have done no AB testing at all or have tried it and haven't seen good results from it. They don't really trust the data. Their website isn't converting as well as it should and the stuff the internal team tries isn't working. They are paying 5 or 6 figures per month on ad spend and only converting a tiny fraction of the traffic.
- **Personas:** 15 simulated
- **Report:** https://grader.wynter.com/r/koalatative-website-optimization-content-servi-_d_Bo6c

> These answers are generated by AI, scored on Wynter's B2B Message
> Layers framework using behaviorally-diverse simulated personas. The
> methodology is real and the critique is directional. What a simulated
> persona cannot have is a live budget, a renewal coming up, or a boss
> asking about this quarter.

---

## 01 · The scores

Every persona answered all four questions. These are four independent
proportions of the same panel, not stages of a funnel.

| Layer | Question | Cleared the bar | Strength | Of those who passed |
| --- | --- | --- | --- | --- |
| 1. Clarity | Do they understand what you do? | 15/15 | 87% | 6 without hesitation, 9 with reservations |
| 2. Relevance | Can they tell what it solves, and who it's for? | 15/15 | 78% | all with reservations |
| 3. Value | Do they actually want it? | 7/15 | 48% | all with reservations |
| 4. Differentiation | Is there a reason to pick you over the alternatives? | 2/15 | 30% | all with reservations |

**Brand alignment** (a side metric, not one of the four layers) — 4/15, 37% strength (all with reservations). Does the page read like the company you actually are?

**Fix first: Value.** Earliest failing layer, walking the sequence in order — not simply the lowest score.

---

## 02 · What to change, layer by layer

Ordered worst-first. Specific edits, not a restatement of the score.

### Differentiation

**Rewrite the six subheads to carry the claim alone.**

"Alignment with business objectives" and "Balancing rigor with speed" could sit on any agency site. Make each subhead state the specific practice — "We work in your stack, not ours" — so a scanner gets the argument without body copy.

*effort medium · impact high · tested against Headings stand alone*

**Move "No 90-day discovery period" into the hero.**

The anti-agency promises are buried six items down a "What's it's like working with us" list. Lead with no discovery period and working in your existing stack — that is the reason to pick this over any CRO shop.

*effort low · impact high · tested against Front-load the meaning*

**Say what work you refuse to do.**

"No slide deck theater" is the sharpest line and it stops at presentations. Extend the anti-pattern list — no vanity tests, no retainer for research nobody reads — since that specificity is what readers repeat back.

*effort low · impact medium · tested against Give a reason to choose you*

### Value

**Add a named client result beside the revenue-gains claim.**

"deliver revenue gains in weeks, not months" sits alone with no number, client or before/after. Put one named engagement with a lift figure and timeframe directly under it.

*effort medium · impact high · tested against Proof next to the claim*

**Replace "High-ROI" and "world-class" with measured outcomes.**

Both phrases are self-awarded grades. Swap in what a client actually got: tests shipped per month, win rate, revenue per visitor movement.

*effort low · impact high · tested against Specifics beat superlatives*

**State what the first 30 days produce.**

"start delivering outcomes right away" never says what arrives or when. Name the deliverables of month one — tests live, roadmap, tracking fixed — so buyers can price the engagement.

*effort medium · impact medium · tested against Concrete over abstract*

### Clarity

**Label Katsed as a side project, not a service.**

The episode title "we're building AB test analysis software" makes readers ask whether they are buying an agency or a tool. One line separating the podcast content from the service offering settles it.

*effort low · impact medium · tested against Plain language*

### Relevance

**Separate the two service tiers visually and label them.**

"done for you" and "done with you" run together in one block, forcing a re-read. Give each its own heading and a line on who picks it.

*effort low · impact medium · tested against Headings stand alone*

### Brand alignment (side metric)

**Cut "boutique" from the opening line.**

"we're a boutique CRO agency" pre-disqualifies the page for anyone with a budget approval chain. Keep founder access as the benefit — direct senior attention — without the word that signals small.

*effort low · impact high*

**Name the buyer's role and company size upfront.**

Nothing says who this is for. Add a line naming the reader — ecommerce growth lead, VP of Digital at a company running X tests a month — so the right buyer self-selects instead of guessing.

*effort low · impact high · tested against Name the audience*

**Raise the register of "This is business, not life or death."**

The casual aside undercuts the statistical-rigor claim it sits inside and reads as practitioner banter to a budget holder. State the tradeoff plainly: which confidence thresholds you use and why.

*effort low · impact medium · tested against Plain language*

---

## 03 · What is working

### The two-tier service structure is understood on first read

Four points said the problem statement and the two service modes — outsource or build internal capability — are stated clearly upfront. One respondent noted the tiers are visually mashed together and required re-reading.

> "done for you High-ROI conversion rate optimization" for those who want it fully outsourced, and "done with you Accelerate internal experimentation capability" for teams building their own program in-house. That's not buried, it's right up top in two clear tracks.
> 
> — Head of Growth, B2B Software, 201-500

> the run-on formatting of "done for you High-ROI conversion rate optimization" and "done with you Accelerate internal experimentation capability" mashed together without clear visual separation made me have to re-read to figure out those were two distinct service tiers
> 
> — Director of Product, SaaS, 501-1000

### The 'no discovery period' and anti-agency lines are the copy that lands

Five points named the no-discovery-period promise, stack flexibility and specific anti-agency anti-pattern messaging as credible differentiation from generic competitors. Respondents could repeat these claims back unprompted.

> The "no 90-day discovery period, strategy set before we even get rolling" line and "we'll adapt to your stack" are the two things that would actually tip me toward them over a generic competitor
> 
> — Director of Marketing, SaaS, 1001-5000

> "no 90-day discovery period" and "no slide deck theater" lines would actually pull me toward them versus a competitor — that's a specific, credible dig at how agencies like this usually waste the first quarter
> 
> — VP of Marketing, E-commerce, 5000+

> Tone-wise it's fine for someone like me — direct, no jargon-soup, "no slide deck theater" and "no 90-day discovery period" are the kind of operator-to-operator lines that land — but it still reads generic rather than written for an e-commerce VP of Product specifically
> 
> — VP of Product, E-commerce, 201-500

> The "no 90-day discovery period" and "no slide deck theater" lines actually stand out to me — they're specific complaints I've had about past agency engagements
> 
> — VP of Product, E-commerce, 1001-5000

> The one thing that'd actually tilt me toward them versus a generic agency is the "No 90-day discovery period" line — "The strategy will be set before we even get rolling so we can start delivering outcomes right away"
> 
> — Senior Marketing Manager, B2B Software, 5000+

---

## 04 · What the personas said

### The Katsed software product is unexplained and confuses what the company sells

Four points flagged the secondary software mention as unclear, leaving respondents unsure whether this is an agency or a vendor. One read building an in-house AB test tool as a funding-priority concern.

> There's a mention of "Katsed," some A/B test analysis software they're building on the side, but that's clearly secondary to the core service. So this isn't a "product" in the SaaS sense at all — it's a services shop
> 
> — Senior Marketing Manager, B2B Software, 501-1000

> the fact that they're mid-build on their own tool ("we're building AB test analysis software") makes me wonder if I'm paying for a service or subsidizing their product roadmap
> 
> — Director of Product, SaaS, 501-1000

> the Katsed mention buried in the podcast blurb — 'we're building AB test analysis software' — that made me pause, because it's dropped in almost as an aside rather than positioned as a product
> 
> — Senior Marketing Manager, B2B Software, 5000+

> That single unexplained noun dropped into an otherwise clear services page is what muddies it.
> 
> — Director of Marketing, SaaS, 201-500

### The page never names who it is for at what size

Six points said the target buyer profile is left implicit, with no calibration for organization size and no named persona. One respondent found the audience clear from the language about teams and internal programs.

> "inconsistent or unreliable A/B test results" isn't addressed head-on anywhere; I'd need a case study or methodology detail
> 
> — Senior Marketing Manager, B2B Software, 501-1000

> the problem (low conversion, no experimentation muscle) and the two service modes are stated, not inferred. The intended reader isn't named explicitly by title or company size, but it's obviously a marketing/growth leader
> 
> — Senior Marketing Manager, B2B Software, 501-1000

> "done for you High-ROI conversion rate optimization" for those who want it fully outsourced, and "done with you Accelerate internal experimentation capability" for teams building their own program in-house. That's not buried, it's right up top in two clear tracks.
> 
> — Head of Growth, B2B Software, 201-500

> I'd need a named buyer persona with my actual scale problem — something like "for SaaS companies doing €5-50M ARR who've plateaued on conversion rate" — plus a number showing they've worked at that scale
> 
> — Director of Product, SaaS, 501-1000

> So problem (losing conversions/revenue you're leaving on the table) and buyer segment (companies with or without an internal CRO team) are both stated, not inferred. What's missing is which size/type of company this is calibrated for — a 1001-5000 person org doesn't know if I'm the target
> 
> — Head of Growth, B2B Software, 1001-5000

> the phrasing "your team" and "internal experimentation program" makes it obvious this is aimed at a business with an existing marketing/product function, not a solo founder
> 
> — Director of Marketing, SaaS, 1001-5000

### Not a single respondent found proof for any results claim

Fourteen points across value, clarity and differentiation cite the total absence of case studies, client logos, before/after numbers or named clients. Several said they would not advance past a first call without them.

> there's no baseline, no case study, no client logo, no "we lifted X client's conversion by Y%" to anchor that claim, so I can't tell if "weeks" means a 2% lift or a 20% one
> 
> — Senior Marketing Manager, B2B Software, 501-1000

> If a competitor on my shortlist has even one named client with a real percentage lift and timeframe, they win by default — this page gives me tone and process promises, not evidence.
> 
> — Senior Marketing Manager, B2B Software, 501-1000

> there's zero proof on this page: no case study, no client name, no actual before/after conversion lift number. It's all promise, no evidence.
> 
> — Director of Marketing, SaaS, 1001-5000

> it's asserted, not proven; there's no case study, no client logo, no "we took X site from Y% to Z% lift" anywhere on this page
> 
> — VP of Marketing, E-commerce, 5000+

> there's nothing on the page telling me who's actually used them, what results they got, or what "revenue gains in weeks not months" actually means in dollars
> 
> — Head of Growth, B2B Software, 201-500

> there's zero proof on this page: no case study, no client name, no actual lift percentage, just a claim of "statistically valid A/B tests."
> 
> — Director of Product, SaaS, 501-1000

> I just don't yet have proof behind "weeks not months" or what their testing methodology actually is
> 
> — VP of Product, E-commerce, 1001-5000

> "statistically valid A/B tests" and "weeks, not months" are still just claims on a page; I'd want to see one or two case studies with actual before/after revenue numbers
> 
> — VP of Product, E-commerce, 1001-5000

> there's no case study, no logo, no actual lift number on the page, so I have nothing to check the claim against
> 
> — Senior Marketing Manager, B2B Software, 5000+

> "no 90-day discovery period" and "no slide deck theater" lines would actually pull me toward them versus a competitor — that's a specific, credible dig at how agencies like this usually waste the first quarter
> 
> — VP of Marketing, E-commerce, 5000+

> The "no 90-day discovery period" and "no slide deck theater" lines actually stand out to me — they're specific complaints I've had about past agency engagements
> 
> — VP of Product, E-commerce, 1001-5000

> "inconsistent or unreliable A/B test results" isn't addressed head-on anywhere; I'd need a case study or methodology detail
> 
> — Senior Marketing Manager, B2B Software, 501-1000

### Boutique and founder-led signals read as disqualifying for enterprise buyers

Six points said the copy carries no enterprise scale signals and raises doubt the playbook scales to larger organizations. Respondents read the brand as a small, mid-market, founder-led shop.

> "boutique" and "founders-led" signals small/scrappy, which makes me wonder if they can handle our volume of tests and stakeholders
> 
> — VP of Marketing, E-commerce, 5000+

> I don't see any signal they've handled a 500+ employee org before, so I'd want to ask directly whether their playbook scales to a company our size or if it's built around leaner teams
> 
> — VP of Marketing, E-commerce, 501-1000

> Tone-wise it's fine for someone like me — direct, no jargon-soup, "no slide deck theater" and "no 90-day discovery period" are the kind of operator-to-operator lines that land — but it still reads generic rather than written for an e-commerce VP of Product specifically
> 
> — VP of Product, E-commerce, 201-500

> I'd need a line that names the buyer and their scale directly — something like 'for enterprise marketing teams running dozens of experiments across multiple markets' or a stat on traffic/revenue size they typically work with
> 
> — Senior Marketing Manager, B2B Software, 5000+

> the blog/podcast content skews more practitioner-hobbyist (Claude.ai tutorials, cohort analysis guides) than enterprise-buyer, so it feels more aimed at a director of growth at a mid-size company than at someone signing off a large budget
> 
> — VP of Marketing, E-commerce, 5000+

> Small — a handful of people, founder-led, probably under 10 years old, riding on the two founders' personal reputations rather than a built brand
> 
> — VP of Marketing, E-commerce, 5000+

### The casual practitioner tone will not survive a CFO conversation

Three points said the tone reads as peer-to-peer or practitioner-hobbyist rather than aimed at budget decision-makers, mixing practitioner community voice with vendor positioning.

> The tone is casual and a bit cute ("Koalas in the Wild," "a safe place where we talk about interesting things") which reads more like they're writing for peers in the CRO/analytics world than for a director trying to justify ad spend to a CFO.
> 
> — Director of Marketing, SaaS, 1001-5000

> the blog/podcast content skews more practitioner-hobbyist (Claude.ai tutorials, cohort analysis guides) than enterprise-buyer, so it feels more aimed at a director of growth at a mid-size company than at someone signing off a large budget
> 
> — VP of Marketing, E-commerce, 5000+

---

## 05 · The hardest read

An adversarial pass over the findings. Every claim below was checked
against the panel's own answers; unsupported ones were dropped.

- **The differentiation that lands is unverifiable, so the page's strongest copy is also its most fragile** *(high)*
  Five points repeat the no-discovery-period and anti-agency promises unprompted, but fourteen points across value, clarity and differentiation cite zero case studies, logos or before/after numbers. Memorable claims with no proof read as bravado.
- **The page cannot survive a first call, let alone a procurement cycle** *(high)*
  Nine of fifteen respondents found no proof for any results claim and several said they would not advance past a first call; three more said the tone is pitched at peers rather than budget holders.
- **The page actively disqualifies itself from enterprise deals it appears to want** *(high)*
  Six points read the brand as a small, founder-led mid-market shop with no scale signals, while six more say no buyer size is ever named. The page lets enterprise buyers self-eliminate by default.
- **Understanding the service structure does not tell anyone whether to buy it** *(medium)*
  Four points grasp the two service modes on first read, yet six say the target buyer is implicit with no size calibration. Clear mechanics attached to an unnamed audience produce recall without qualification.
- **The company's own category is unresolved on the page** *(medium)*
  Four points flagged the Katsed software mention as leaving them unsure whether this is an agency or a vendor, and one read the in-house AB test tool as a funding-priority worry. The secondary product undermines the primary pitch.
- **Voice and buyer are misaligned in the same paragraph** *(medium)*
  Three points describe a practitioner-community or hobbyist register mixed with vendor positioning, while six read boutique founder-led signals. The page speaks to operators while asking decision-makers for budget.

---

## 06 · Who answered

| # | Role | Industry | Company size |
| --- | --- | --- | --- |
| 1 | Senior Marketing Manager | B2B Software | 501-1000 |
| 2 | Director of Marketing | SaaS | 1001-5000 |
| 3 | VP of Marketing | E-commerce | 5000+ |
| 4 | Head of Growth | B2B Software | 201-500 |
| 5 | Director of Product | SaaS | 501-1000 |
| 6 | VP of Product | E-commerce | 1001-5000 |
| 7 | Senior Marketing Manager | B2B Software | 5000+ |
| 8 | Director of Marketing | SaaS | 201-500 |
| 9 | VP of Marketing | E-commerce | 501-1000 |
| 10 | Head of Growth | B2B Software | 1001-5000 |
| 11 | Director of Product | SaaS | 5000+ |
| 12 | VP of Product | E-commerce | 201-500 |
| 13 | Senior Marketing Manager | B2B Software | 501-1000 |
| 14 | Director of Marketing | SaaS | 1001-5000 |
| 15 | VP of Marketing | E-commerce | 5000+ |

---

## 07 · Before you act on this

The methodology is real, and the critique is directional. What a
simulated persona cannot have is a live budget, a renewal coming up, or
a boss asking about this quarter. **Validate anything you're betting on
with real ICPs who are actually in-market.** Being wrong is more
expensive than you think. Finding out is cheaper than you'd guess.

Wynter runs message testing with verified B2B professionals — trusted
by HubSpot, RingCentral, Shopify, Cognism, Paddle, Veeam, Rippling and
Miro. <https://wynter.com>

This report is kept for 60 days from 2026-08-25, then deleted along with the personas and their answers.

