Do small categories get fewer named businesses than large ones?

A look at whether category size predicts how many distinct businesses an AI engine lists in its answer, across four size segments spanning from the smallest to the largest quarter of the sample.

When an AI engine answers a question about a small, niche market category, does it name fewer distinct businesses than it would for a huge, well-known category?Measured 2026-04-27
The number

by_segment.1 smallest quarter.classes.names 6 or more.pct

A reasonable guess, before looking at any data, is that AI engines behave like a thin encyclopedia for small categories and a thick one for big categories. A huge category, something like coffee shops or accounting software, has thousands of real businesses competing for attention, a long documented history, and presumably a deep well of named entities for a language model to draw on. A small category, a niche tool or a regional service type, has fewer real businesses to draw on in the first place. So it would be unsurprising if engines named fewer distinct businesses when answering about small categories, simply because there is less to name.

This study tested that guess directly. It took 193 groups covering 752,124 rows of answer data (none excluded), split every category into one of four size segments by row count, from a smallest quarter of 88,655 rows up to a largest quarter of 261,745 rows, and classed every answer set as naming six or more distinct businesses or naming five or fewer. If category size drove naming behavior, the smallest quarter should show a noticeably lower share of six-or-more answers than the largest quarter.

It does not. In the smallest quarter, 93.2% of answer sets named six or more distinct businesses, against 6.8% that named five or fewer, out of 88,655 rows in that segment. That is barely different from the pattern in the whole sample, where 684,307 of 752,124 answer sets (91%) named six or more businesses. Small categories are not thin on names. Whatever mechanism produces a rich, multi-name answer does not appear to depend on how many real competitors exist in the category.

Why would this be true? The likely explanation is that an AI engine is not enumerating a market from a private census of real businesses, it is generating a plausible-sounding list from whatever patterns are strongest in its training data and retrieval context for that kind of question. A prompt asking "what are some good options for X" pulls a similar list-shaped response whether X is a huge category or a small one, because the shape of the answer is driven by the shape of the question and the model's general habit of listing several options, not by an internal inventory of how many businesses actually exist. A model does not know, in any strong sense, how big the category is before it starts answering. It infers a plausible list size from the pattern of similar prompts it has seen, and that pattern is fairly stable across category sizes.

An alternative explanation worth naming: maybe the size segments in this study are not tracking real-world category size well, so the smallest quarter here is not actually made of niche markets. That is worth checking, but it would take a different kind of study, one that validates segment size against an independent measure of real business counts. What this study can say is narrower and still useful: within the row-count based segmentation used here, the smallest and largest quarters produced almost identical naming behavior, 93.2% versus 90%.

For a company in a small or niche category, the practical implication is that being small is not a reason to expect thin, generic answers from AI engines. The engine is not holding back names because your category is small. If your business is not showing up in a six-or-more-name answer, category size is not the excuse. Something else, likely how the business is described, linked, or reviewed online, is more likely responsible than the size of the market it competes in.

Free strategy session

Want to know how AI answers describe you?

We run the same measurement on your category. Fifteen minutes with founder Omar Jenblat, your own numbers, no deck.

Omar Jenblat, Founder & CEO of BusySeed
Omar JenblatFounder & CEO, BusySeed
  • Your category measured the same way
  • Your own numbers, not a sample deck
  • Fifteen minutes, no obligation

First, who are we meeting?

Three fields, then pick your time. We read up on you before the call so we open with something useful.

No sales sequence. If you never pick a time, we leave it there.