Define practical customer tasks.
Tasks match the business and use plain language. Typical examples include finding current hours, confirming a service area, requesting a quote, or booking an appointment.
Readiness Check methodology
Each task runs three times against the live website. A separate AI judge reviews the evidence before any result becomes a finding.
Meta’s September 8 launch of Muse provides a current example: a personal AI agent with its own browser that can navigate websites, fill forms, book appointments, and—with the user’s approval—make purchases. That makes task completion by agents a live website concern rather than a theoretical one.
ViaLayer tests the broader condition this creates: whether an outside agent can understand a business and complete a defined customer path on its public website.
Muse is context, not a ViaLayer integration or test target. Results describe the systems and tasks named in each report.
The check uses the same public paths available to a customer. No private access, prepared demo, or cooperation from the site is used during a run.
Tasks match the business and use plain language. Typical examples include finding current hours, confirming a service area, requesting a quote, or booking an appointment.
An AI agent receives a real browser session and attempts each task through the public website. The run records steps, site responses, and screenshots.
Every task starts again twice. Three independent runs distinguish a repeatable site problem from ordinary variation in one agent run.
A separate AI judge reviews the transcript and evidence for every run. The result does not rely on the acting agent's self-reported success.
The report states the outcome, consistency, supporting evidence, and confidence classification. Mixed results remain mixed.
Every finding carries an explicit class and the exact result across the three independent runs.
The same material outcome occurs across the runs and is supported by a site response, transcript, or screenshot. The exact count appears beside it.
The result varies across runs or depends on interpretation. The report states the mixed outcomes and why confidence is limited.
The available runs do not support a reliable conclusion. We explain what prevented classification and do not present a confirmed problem.
Every outcome can be traced to an attempted task and its evidence.
A Readiness Check tests defined tasks at a specific point in time. Different systems and customer prompts can behave differently.
Agent behavior varies. The check repeats each task and reports mixed outcomes rather than converting them into a cleaner story.
Websites, directories, forms, and agent behavior change. Monitoring is required when ongoing stability matters.
No administrative access is needed for the initial Readiness Check. Any later access is separately scoped, limited, and revocable.
Testing stops before payment, contract acceptance, account changes, destructive actions, or other irreversible steps unless a separate written procedure explicitly authorizes them.
If a meaningful test needs a form or booking submission, the identifier, allowed details, expected handling, and cleanup responsibility are agreed in advance.
A payment, legal, authentication, sensitive-data, or security boundary is reported rather than crossed.
Review a wholly fictional example, or ask ViaLayer to test one practical task on your live website.