What an AI search agency should deliver in the first 90 days
A phase-by-phase table of what a B2B SaaS should receive from an AI search agency in the first 90 days, built on LoudFace's eight-stage method, with the reading that proves each stage worked.
In the first 90 days a B2B SaaS should receive, in this order, starting in week one, a per-engine baseline of how often AI answers name you, proof the engines' crawlers can reach your pages, one fixed description of your company on every surface, pages built as units an engine can lift, original material a competitor cannot restate, the first third-party pages that name you, and a report that reads every visibility number against booked demos and signups. The pages ship inside month one. Early citations can land inside days. Dominant share of answer on a competitive prompt cluster takes months. An agency that promises the second on the timeline of the first is selling the wrong number.
On this page
- The first 90 days, phase by phase
- Why the order matters more than the dates
- What the first 30 days should look like on your side
- Weeks 1 to 8: pages built for extraction
- Months 2 to 3: other people's pages
- What the report should say at day 90
- What it costs
- How we know the sequence holds
- Frequently Asked Questions
On r/AskMarketing, a buyer weighing an AEO agency wrote: "If I were evaluating an agency, I'd ask for a really practical breakdown of the first 90 days. What are they auditing?"

The first 90 days, phase by phase
The table is LoudFace's own eight-stage method laid across a calendar quarter, the same length as the three-month minimum on a fixed-scope engagement. The stages are fixed and the order is fixed. Most of them open in week one and run in parallel, which is why the first pages ship in the first week rather than the second month. The windows are ranges, because the engines re-select their sources week to week and nobody honest attaches a date to a citation.
| Phase | When | What ships | How you verify it |
|---|---|---|---|
| 1. Baseline, per engine | week 1 | Share of answers naming your brand, citations of your URLs, average position when cited, and sentiment, on every tracked prompt, on ChatGPT, Perplexity and Google AI Overviews separately | You hold a dated number per engine before any page changes. Every later claim is compared against it |
| 2. Access | week 1 | Access rules in robots.txt that admit the search crawlers (OAI-SearchBot, PerplexityBot and Google's own crawler), pages indexed and snippet-eligible, text present in the HTML before scripts run | Your server logs show which crawler fetched which page and when. Where your hosting hides the logs, the agency says so |
| 3. Entity | week 1 | One sentence that describes the company, one category name, the named verticals, and the proof that travels with them, placed on the homepage, the service pages and every roster entry that names you | Your team asks ChatGPT what your company does. The answer should use your sentence, not its own flattest guess |
| 4. Artifact | from week 1 | The first calibration articles, drafted and reviewed with you inside week one, then the pages built for the buyer prompts you chose, each leading with a unit an engine lifts: a ranked list with verdicts, a comparison table with real figures, or a short answer at the top. Volume climbs to 20 or more articles a month from month two | The page appears in the sources of an AI answer. Retrieval is the first reading; being named is the second, and they are counted separately |
| 5. Original material | from week 1 | Your own measurements, your experts' corrections written into a knowledge base you approve, and claims tied to persisted primary sources | Every number on a shipped page traces to a source the agency can show you |
| 6. Corroboration | months 2 to 3 | The first third-party pages that name you, starting with the lists the engines already retrieve for your prompts | The engines' own source lists for your prompts begin to include pages that are not yours |
| 7. Placement | months 2 to 3 | 8 to 12 placements a month on pages where the engines and your buyers already look, never a link blast | Each placement is on a page you could have found in an AI answer's sources yourself |
| 8. Reporting through to revenue | weekly | A dashboard that refreshes every morning and a written report every Friday, carrying readings per engine joined to your Search Console, AI-referred visits by first touch, signups and booked demos, and revenue from your CRM, with a weekly showcase and a call every second week | A visibility line that moved while the pipeline stayed flat is reported as a failure, in those words |
Why the order matters more than the dates
Three different things happen inside an AI answer, and they fail separately. Retrieval comes first, when the engine fetches a page. Citation is the engine listing that URL as a source, which it often does not do. Naming your brand in the text a buyer reads is a third event again, and the rarest of the three. We measured that chain on ourselves across 120 recent answers, 40 per engine. ChatGPT retrieved a loudface.co page in 24 of its 40 answers and named LoudFace in 7 of them. Over 30 days, mentions of LoudFace across the three engines rose from 968 to 1,905 while citations of our URLs moved from 4,075 to 4,183, close to flat. Being read is not being recommended.
Every stage in the table exists to move one link of that chain. The baseline tells you where the chain breaks today. Access and entity work make retrieval possible and give the engines a name they can repeat. Build the page as a unit an engine can lift and a retrieval turns into a citation. Being named is the stubborn one: a brand that is only ever named from its own pages has a ceiling, which is what corroboration and placement are there to raise. And reporting closes the loop on the only reading that pays the invoice.
The order is about dependency rather than delay. You measure before you claim anything moved, and you fix the name before you ask an engine to repeat it, which is why the baseline and the entity sentence land in week one and not after the first pages. Neither one holds a page back. The structural work runs through the whole first quarter, underneath the shipping, and a site that never gets it gives the engines little reason to select anything on it, however well any single page is written. That is the part agencies leave out, because it produces nothing to screenshot for a monthly report. So the pitch becomes "fast AEO": listicles aimed at prompts the client cannot credibly win yet. The client churns at month six when nothing has moved.
Retrieved is not named: 40 ChatGPT answers on LoudFace's tracked prompts
Of 40 ChatGPT answers sampled on LoudFace's tracked prompts, 24 retrieved a loudface.co page and 7 named LoudFace. Being read is not being recommended.
Source: LoudFace methodology page, 120 answers sampled across three engines
What the first 30 days should look like on your side
Packed. Month one is the busiest month of the engagement, on our side and on yours. Kickoff runs within 48 hours of signature. Inside week one the shared Slack channel opens, access is collected, the technical fixes ship (canonical tags, sitemap, H1s, schema), baseline tracking goes live per engine and per prompt, the entity sentence is fixed, and the first calibration articles are drafted and reviewed with you. You see shipped progress within five days, and week one does not end in silence. The cadence starts on day one too: a dashboard that refreshes every morning, a written report every Friday, Slack through the week, a weekly showcase, and a call every second week. Our own pricing page says it in one line: "No lengthy onboarding, no bloated statements of work. We learn the gap, scope the first initiatives, and start shipping on a weekly cadence."
On your side the packed part is three jobs. You hand over access: Search Console, your hosting or CDN logs where they exist, the CMS, the analytics property. You answer one uncomfortable question: which prompts do your buyers actually type, in their words, not your product's. That prompt set is the measurement frame for the whole engagement, and it is worth a week of argument. And you read the first calibration articles while they are still drafts, because your corrections are what the later pieces are built from.
The baseline document lands in that same week, beside the shipped fixes rather than ahead of them. It carries a number per engine per prompt. On a B2B SaaS site that has never been measured this way, expect the number to be near zero on the buyer prompts that matter and high on the prompt that is the company's own name. That gap is normal. It is also the first thing to be honest about in the kickoff, because it sets the shape of the next 60 days.
A crawler report comes with it, in the same week, while the fixes are already going live. We read raw server logs where your hosting gives us access, rather than trusting a tracker's estimate, because the log is the only direct observation of which AI crawler fetched which page. Probability-based trackers query the engines from outside and estimate what they are citing. A log records what actually arrived, with a timestamp. If your hosting hides them, the agency should say so and the crawler picture comes from the tracked readings alone. An agency that never mentions logs is measuring you from the outside only.
The entity work in the same window is unglamorous and it is where a B2B SaaS most often loses the name. If your category label reads one way on the homepage, another on the service pages, and your verticals drift with it, the engine invents its own summary. Its invention is usually the flattest thing it can say about you. One sentence, one category, the named verticals, the proof that travels with them, on every surface. The entity problem is the same one that makes engines confuse two companies with similar names. That is a copy job and a discipline job, and it runs in the same week as the first articles, so those pages ship with the name already fixed.
Weeks 1 to 8: pages built for extraction
The first pages do not wait for the foundation. Calibration articles are drafted and reviewed with you in week one, and the volume climbs to 20 or more articles a month from month two. What moves across these eight weeks is how much of the corpus the pages cover.
Format decides citation. In our own 90-day study of 128,515 citations in the B2B SaaS growth-agency category, listicles carried 53.17% of every citation, more than every other page type combined. Our most-cited page is a listicle. It carried 819 citations in the 30 days to 1 September 2026.
The mechanism is simple once you have read enough AI answers. An engine lifts a pre-formatted unit: a ranked list that names brands with a one-line verdict, a comparison table with real figures, or a short answer at the top of a page. A page that buries the same content in prose gets fetched and skipped. So every page an agency builds for a buyer prompt in this phase should lead with the unit that prompt wants, in the first screen, and you should be able to see the unit in the draft before it ships.
This is also where the fast part of the timeline lives. The industry line that AI citations take six to twelve months is wrong as a default. A well-structured page on a brand with even modest authority can be cited by Google AI Overviews and Perplexity within a day of publishing, on a prompt that has no entrenched winners. The slow part is not the first citation. The slow part is climbing to a dominant share on a competitive prompt cluster, which is a months-long job and the one you are actually paying for.
Read a proposal against that split. An agency describing initial citation gains within four to eight weeks and compounding results over three to six months is describing the middle speed. One describing a fixed day-by-day ship list, prompt set by day 7, baseline by day 14, first batch by day 30, is describing its own project plan, which is fine, as long as it does not present the plan's dates as the engines' dates. The engines decide, and they change their selections week to week. Anyone guaranteeing a placement is either not measuring or not telling you.
What sits inside the page decides whether your brand survives the answer, and original material starts with the very first article. Your own measurements. Your experts' corrections, captured into a knowledge base you approve entry by entry, so the next piece starts from what your people know rather than from what the model remembers. Claims tied to persisted primary sources. A competitor can restate a public statistic. They cannot restate yours.
Months 2 to 3: other people's pages
From month two the work moves off your own domain, and this is the stage most agencies never reach, because it cannot be done from a content calendar. Our own data makes the case against us here. Of the 1,000 most-cited pages the three engines used in our category over 30 days, 974 are somebody else's, and 8 of those mention LoudFace. In our own sample of 120 recent AI answers, 40 on each engine, LoudFace was never named unless one of our own pages was in the sources. A brand that is only ever named from its own pages has a ceiling. Raising that ceiling is what months two and three are for.
So the off-domain work is corroboration and placement, running at 8 to 12 placements a month through months two and three. Corroboration means the first third-party pages that name you, starting with the lists the engines already retrieve for your prompts. Those lists are visible: open an AI answer on your buyer prompt, read its sources, and you have the target list. Placement means selective placements where the engines and your buyers already look. Never a link blast, never a directory sweep. Each placement should be a page you could have found in an AI answer's sources yourself.
This phase is also where an engine-by-engine read starts to matter. Each engine trusts a different corpus. Google AI Overviews leans on YouTube, LinkedIn and Reddit. ChatGPT pulls listicles, arXiv and Reddit. Perplexity, in our category, takes 81% of its citations from listicles. A plan that treats "AI search" as one surface will move one engine and stall on the others, and the blended number will hide which one stalled.
What the report should say at day 90
Not "visibility is up." A day-90 report from a serious agency reads every engine signal against the commercial events on your side.
- Per-engine share of answers, citations, position when cited and sentiment, on the same prompt set as the day-1 baseline, on the same engines, with the baseline printed beside it.
- Search demand from your own Search Console, clicks and impressions. Read them next to what the engines do before they retrieve anything: ChatGPT fans a prompt out into narrower sub-queries on 47% of its answers and Perplexity on 2%, measured across 3,718 AI answers to our own tracked buyer prompts between 26 July and 25 August 2026. Those fan-out queries are what the engine searches, not what your buyer typed.
- AI-referred visits by first touch, labelled as a floor, because the reading depends on a referrer being present and not every AI visit carries one.
- The signups, booked demos and any other lead capture your site collects, joined to first touch where the data exists and marked unknown where it does not.
- Revenue, joined to your CRM.
A program that moves the visibility numbers and leaves the pipeline flat is a program we call failing, and we say that in the report rather than leading with the chart that went up. If a report you receive at day 90 leads with the chart that went up, ask for the pipeline line beneath it.
What it costs
LoudFace runs this as a retainer from $5k a month. The figure depends on tier, scope and complexity, and it is scoped on an intro call. A fixed scope runs inside the same retainer with a three-month minimum, the same team, cadence and Scoreboard pointed at the defined deliverable until it ships. A published 90-day program elsewhere in the category quotes roughly $5,000 to $15,000 a month depending on prompt-set scope and content volume. The floor is comparable. The difference is in what the money buys in the first 30 days: the fixes, the baseline and the first pages together, or a listicle written before anyone knew where the chain broke.
How we know the sequence holds
LoudFace is a full-stack organic growth agency for B2B SaaS: one program across SEO, AEO/GEO, content and Webflow, built for the answer engine rather than classic SEO silos. I would not sell a sequence we had not run on ourselves, so we ran it on LoudFace first. LoudFace went from 0.18% of AI answers in its own category to 10% in 90 days, with the readings published as we went, including the ones that made us look weak. In the 30 days to 2 September 2026, we are named in 12.95% of AI answers on our tracked prompt set, and our average position when cited is 2.8, across a tracked panel of 50 brands, published on the methodology page. The readings carry on in our own AEO case study. The retrieve, cite, name chain is measured on our own domain in 30-day windows, floors labelled.
That is also the test to apply to any agency you are weighing against our evaluation scorecard: ask for their own domain's per-engine numbers over time, and ask whether the weak ones are published next to the strong ones. An agency that will not run its method on itself is asking you to be the experiment.
Frequently asked questions
Answers to the questions readers ask most about this topic.
How long before an AI search agency shows results?
There are three speeds, and a proposal should name which one it means. First citation pickup can happen within a day of publishing a well-structured page on a prompt without entrenched winners, on Google AI Overviews and Perplexity especially. A consistent slot in the cited-source set, surviving the engines' re-evaluations, takes weeks. Dominant share of answer on a competitive prompt cluster, with branded search lift and pipeline behind it, takes months. Technical fixes and early wins usually move inside the first 30 to 60 days; content authority compounds over 90 to 180 days.
What should I receive in the first 30 days?
Kickoff inside 48 hours of signature, then a packed week one. You hand over access and agree the prompt set that becomes the measurement frame. You get a dated per-engine baseline on your buyer prompts, the technical fixes live (canonical tags, sitemap, H1s, schema), one fixed company description and category label placed on every surface that names you, and the first calibration articles drafted and reviewed with you. A weekly showcase and a written report every Friday start in the same week.
How do I verify an agency's claims at day 90?
Ask for the same reading you got at day one, and verify it the way you would before hiring: the same prompts, the same engines, the same metric labels, with the baseline printed beside the new number. Then ask for the pipeline line. Visibility is the share of tracked prompts where your brand appears; position is where you sit inside the answer when you do; share of voice is the stricter, lower figure. Refuse a blended number across engines, because the blend hides which engine is losing. Refuse a number with no window or prompt count attached.
Does an AI search engagement replace SEO?
No, and an agency that sells them as separate programs leaves half your pipeline uncovered. Google AI Overviews requires a page to be indexed and snippet-eligible before it can appear as a supporting link, so the crawl and indexing work of SEO is the access stage of AI search. We sell the two as one program for that reason. What changes is the artifact on the page and the reading you hold the work to: share of answer alongside clicks, on each engine separately.
What does a good B2B SaaS AEO engagement include?
The eight stages in the table, in that order, run inside one retainer by one team: baseline per engine, crawler access, brand entity, liftable artifacts, original material, third-party corroboration, selective placement, and per-engine reporting joined to revenue. Serious agencies in the category describe the same shape in their own words, AEO integrated across a full organic program of SEO, content, PR and links rather than bolted on. What separates them is whether anything ships in the first month, and whether the report at the end reads visibility against booked demos.
What if my category has entrenched citation winners already?
Then the artifact stage starts with prompts your buyers ask where the winners are thin, usually the vertical and situation prompts rather than the head term, and the corroboration stage starts sooner. An agency that opens with the head term against three entrenched lists is choosing the prompt it can write about over the prompt you can win. Ask to see the per-prompt baseline and the competitor set behind each prompt before any page is drafted. That decision is made in phase one, and it is the one that most determines what day 90 looks like.


