Daycare website A/B testing is the practice of showing two versions of the same page to real visitors and counting which one produces more tour requests.

It is the method underneath daycare website conversion, the work of turning the families already finding you into booked tours, and it is how you know a change helped instead of just feeling better.

A/B testing for child care websites runs into one problem the e-commerce case studies never mention: a single-site center can get a few hundred visits in a month, while a valid test can need thousands per version.

So the honest version has three questions: how big a test needs to be, how to run one without fooling yourself, and what to do when no test is worth running.

How many visitors a daycare A/B test needs

Three inputs decide it: your baseline rate, the smallest lift worth detecting, and your significance threshold.

On the threshold, Optimizely's sample-size calculator calls 95% an accepted standard for statistical significance, though you can set your own.

On the sample, Evan Miller's rule of thumb for classic fixed-sample tests is n = 16·σ²/δ² per variation, where σ² is p(1−p) and p is your baseline conversion rate.

Visitors needed per variation, by rule of thumb

Proving 5% became 6%about 7,600
Proving 10% became 12%about 3,600
Illustrative rates, not benchmarks. Arithmetic from Evan Miller's sample-size rule of thumb for classic fixed-sample A/B tests.

At an illustrative 5% tour-request rate, detecting a lift to 6% works out to about 7,600 visitors per variation, roughly 15,000 in total.

A bigger swing, 10% to 12%, takes about 3,600 per variation.

Both rates are illustrations; no benchmark for a center's tour-request rate is published, so your own baseline goes into the formula.

Many single-site centers see fewer visitors in a month than the first example needs, which is the problem this page works around.

How to run one test without fooling yourself

Pick one page and one change

The page with the most to lose, usually the tour request form or your ads' landing page, and one change: fields, headline, or photos.

Pick one primary metric

Tour requests or completed bookings, counted the same way on both versions; a test judged by form fills alone can win on paper while tours stay flat.

Fix the sample size or end date first

Decide before launch how many visitors each version needs, or the stop date, and write it down.

Let it run, then look once

Call it when the preset sample is reached and your threshold is met, not when a Tuesday looks good.

The discipline in the last two steps is what broken tests miss: in Evan Miller's worked example (2010), checking after every visitor and stopping at the first result past a 5% significance threshold produced a false positive 26.1% of the time, and peeking ten times turned a reported 1% into about 5%.

Some modern tools use sequential statistics that allow monitoring, but fix-the-sample-first costs nothing in any tool.

Judging matters too: LineLeader, the child care CRM vendor, reports in its Q1 2026 benchmark summary that an inquiry's conversion probability drops significantly after 30 days and is minimal after 60 without structured re-engagement, waitlists excluded, so a form-fill win that never becomes a held tour is not a win.

That is why I judge every test by tours booked and children enrolled, measured with read access to your CRM or waitlist numbers, never by form fills alone.

What is worth testing on a center website

Start where the ask happens: the daycare tour request form has the field-by-field design, so the question here is what to test, not how to build.

On structure, Baymard Institute's conclusion across 10+ years of checkout research is that the total number of form fields affects usability more than the number of steps.

That is checkout research, not child care research, and no percentage lift is published for either change, which is why it is a test and not a copy job.

Step-by-step forms have a real case here though: NN/g's guidance on wizard-style forms (Raluca Budiu, 2017) is that they suit occasional processes and can branch, like asking the child's age before showing that room's tour times, at the cost of extra clicks.

Beyond the form, the shortlist that earns a test: the first screen a visitor sees, how the tuition page answers cost, the online booking flow, trust signals, photos you hold written parent permission to use, and where the click-to-call button sits.

Split testing daycare forms one bet at a time is what keeps a small center's testing calendar realistic.

The rules a test variant can trip

  • Never draft a parent quote for a B version: 16 CFR 465.2 bans creating reviews or testimonials that misrepresent the reviewer or the experience, and the FTC names AI-generated ones as an example.
  • A parent's Google review quoted on your homepage or in an ad becomes a testimonial, per the FTC's own Q&A, and publishing testimonials on your own website is disseminating them, so a center can be liable for the fake or false ones it chose.
  • 16 CFR 465.5 requires clear and conspicuous disclosure when an officer or manager writes a review of their own business, and officers include owners.
  • Under 16 CFR 255.1, quotation marks present a quote as the parent's exact words, and an edit may not change its meaning.
  • Under 16 CFR 255.2(b), a testimonial about one child's outcome reads as what families generally get, so disclose what is generally expected when it is not; "my son was reading by age 3" is my illustration, not the FTC's.

The rule, 16 CFR Part 465, has been in effect since October 21, 2024 and lets courts impose civil penalties for knowing violations; the Guides at 16 CFR Part 255, last revised in 2023, are the FTC's interpretation, where inconsistent practices may result in corrective action.

If a variant tests an enrollment offer or registration-fee waiver on the Google Business Profile, Google's guidelines for representing your business require the promotion to link to its terms, and the business must honor what it posts (as of October 2026).

Confirm how these rules apply to your variants with your state licensing agency or a lawyer.

When traffic is too low to test

Some months no controlled test is worth running, and there are three better ways to spend the time.

Fix what is already broken without a test: a form that fails on phones or a missing tuition answer loses families regardless, and the damage shows in your daycare website conversion rate measured before you touch anything.

Make one bigger change at a time and judge it against the last three months, knowing before-and-after comparisons are not controlled tests because the season and traffic mix move underneath you.

Or feed a test with traffic: paid clicks concentrated on a daycare landing page reach a real sample in weeks instead of quarters, the one lever that shrinks the calendar.

CRO testing a preschool website at this pace is slower than the case studies promise, and still beats redesigning on a hunch.

One controlled test a month on a center's website and landing pages, judged by tours booked and children enrolled, is the work More Booked Enrollments does, a center running managed Meta ads with me also gets one landing-page A/B test a month, and current prices are on the pricing page.

Frequently asked questions

How much traffic does a daycare website need for an A/B test?

It depends on your baseline and the size of the change: by Evan Miller's rule of thumb, proving an illustrative 5% tour-request rate improved to 6% takes about 7,600 visitors per variation, while a jump from 10% to 12% needs about 3,600. Most single-site centers close a gap that big by running the test for months or by pointing paid traffic at a landing page.

What should I test first on a daycare website?

Start with the page that carries the ask, usually the tour request form, and test the change you would bet on before seeing any data. Small traffic can only detect big swings, and field count is where the usability research points first.

How long should an A/B test run before I call it?

Until the sample size or end date you fixed before launch is reached, which can be weeks or months on a center's traffic. With classic fixed-sample tests, Evan Miller's worked example shows that stopping at the first significant-looking reading produced a false positive 26.1% of the time at a 5% threshold.

Can I test two versions of a parent testimonial against each other?

Yes, but the FTC's rules travel with the test: a parent's Google review quoted on your site becomes a testimonial, quotation marks present the words as exact, and an edit cannot change the review's meaning. A testimonial about one child's outcome also reads as a typical result, so disclose what families should generally expect.

What if my test ends with no clear winner?

A null result is a real result: the change did not move tour requests enough to detect, so keep the version that costs less to run and put the next test on a bigger swing. Rerunning the same test and hoping the sample breaks differently is not a method.