Getting Your Firm Cited by AI,
Not Just Ranked by Google
A growing share of people with a legal question now ask an AI assistant before they open a search engine. The assistant answers, cites two or three sources, and the conversation frequently ends there. AEO is the work of being one of those sources.
Definition
What AEO Is
Search engine optimization competes for a position in a list of links. The user sees ten results and chooses.
Answer engine optimization competes to be the source an AI system quotes when it answers a question directly. The user sees one answer and a handful of citations. There is no page two.
The two disciplines overlap heavily — clean structure, fast delivery, genuine depth and clear authorship help with both — but they diverge in what they reward. SEO rewards a page that is comprehensively about a topic. AEO rewards a passage that cleanly answers a specific question in a form a model can lift and attribute.
That distinction drives most of what follows.
Why Legal
Why This Matters for Law Firms Specifically
Legal questions are close to the ideal case for AI-assisted answering. They’re high-stakes enough that people research before acting, complex enough that a direct explanation genuinely helps, and common enough that the same questions recur constantly. Do I need probate if there’s a will? How long do I have to file? What does a retainer actually cover?
Someone asking those questions is early in a process that often ends with hiring a lawyer. Historically they found you through a search result. Increasingly they get an answer, and the firms named in that answer get the consideration.
The firms currently being cited are rarely the largest. They’re the ones who published a clear answer to that specific question in a form the model could use.
Evidence
What Appears to Matter
Here we separate what’s confirmed from what’s inferred, because much of this field is inference.
Confirmed: crawler access
This is the one thing that is binary and verifiable. If your robots.txt blocks an AI crawler, that system cannot read your site, and no amount of content quality compensates.
Many sites block these by default — some hosting platforms and security plugins add the rules automatically, and plenty of site owners blocked them deliberately during the 2023–24 period when AI training was the dominant concern. Those decisions are frequently still in place and forgotten.
The agents worth checking include GPTBot, OAI-SearchBot and ChatGPT-User (OpenAI), ClaudeBot (Anthropic), PerplexityBot (Perplexity), and Google-Extended, which governs Google’s AI products separately from ordinary Googlebot access. Agent names change; this list is worth re-checking every six months.
There’s a real decision embedded here, and it isn’t only technical. Allowing these crawlers means your content may be used in training as well as retrieval. Most firms conclude that visibility is worth it. Some don’t. It should be a deliberate choice rather than a default nobody examined.
Confirmed: structured data is machine-readable
Schema markup — FAQPage, Article, Organization, Person, Service — is an explicit, unambiguous statement of what a page contains and who published it. Search engines have used it for years to generate rich results.
What’s confirmed is that it makes content machine-parseable. What’s inferred is how much weight AI systems give it in citation selection. The reasoning is straightforward — a system choosing sources benefits from unambiguous signals about authorship and content type — but no provider publishes a weighting.
We implement it because the cost is low, the SEO benefit is independently real, and the AEO logic is sound.
Inferred: answer-first structure
Content organized as clear questions with complete, self-contained answers appears to be cited more readily than the same information embedded in flowing prose.
The mechanism is intuitive. A system assembling an answer needs a passage that stands alone. A sentence that begins “Most firms need eight to ten” can be lifted and attributed. A paragraph that begins “This is a question we hear often, and the answer depends on several factors” cannot — the actual answer arrives three sentences later, entangled with setup.
Every FAQ answer we write leads with the answer. It reads slightly blunt to a human. It’s substantially more extractable.
Inferred: specificity and verifiability
Concrete, dated, sourced claims appear to be preferred over general ones. “A 98.5% collection rate at this firm, 2026 year to date” is checkable. “Industry-leading collections” is not.
This aligns with what model providers say publicly about wanting to cite reliable sources, and it’s consistent with what we observe. It’s also just better writing.
Inferred: entity clarity and demonstrated expertise
Who wrote this, what are their credentials, and does the site demonstrate genuine knowledge of the subject? For legal content specifically — where Google has long applied elevated scrutiny under E-E-A-T — attributed authorship with real credentials appears to matter more than in other verticals.
Person schema on attorney bios, named authorship on articles, and visible credentials are all cheap to implement and defensible on their own merits.
The Work
What We Actually Do
- Audit crawler access — confirm which AI agents can reach your site, and make the allow-or-block decision explicit rather than inherited.
- Implement the full schema stack — Organization, LocalBusiness, Person, Service, FAQPage, BreadcrumbList, Article — validated against Google’s Rich Results Test.
- Structure content for extraction — question-shaped headings, answer-first paragraphs, self-contained passages that make sense lifted out of context.
- Build genuine topical coverage — a firm with one page on a subject has one chance to be the source. A firm with fifteen has fifteen. This is where the content program does the work. Website & Content Program →
- Attribute everything — named authors, real credentials, Person schema, dated claims.
- Write specifically — concrete numbers with sources and dates, not adjectives.
Boundaries
What We Don’t Promise
We don’t guarantee that any AI system will cite your firm, for any query, at any point. Nobody controls citation selection, the mechanisms are undisclosed, and they change without warning. Any vendor promising AI visibility outcomes is promising something they cannot deliver.
What we commit to is the work: crawler access confirmed, schema implemented and validated, content structured for extraction, and publishing volume sufficient to build real coverage. That’s the part anyone can actually be held to.
Common Questions