| Publisher | Answers citing it | Answer penetration |
|---|
| healthline.com | 172/272 | 63.2% |
| forbes.com | 169/272 | 62.1% |
| fortune.com | 119/272 | 43.8% |
| pmc.ncbi.nlm.nih.gov | 82/272 | 30.1% |
| consumerlab.com | 81/272 | 29.8% |
| seed.com | 59/272 | 21.7% |
| medicalnewstoday.com | 59/272 | 21.7% |
| everydayhealth.com | 54/272 | 19.9% |
| health.usnews.com | 49/272 | 18% |
| garagegymreviews.com | 49/272 | 18% |
| nih.gov | 46/272 | 16.9% |
| target.com | 42/272 | 15.4% |
Healthline was cited in 63.2% of answers with retrieval evidence, followed by Forbes at 62.1% and Fortune at 43.8%. PMC, a National Library of Medicine archive domain, appeared in 30.1%, while ConsumerLab appeared in 29.8%. Retail domains also entered the evidence set: Target appeared in 15.4% and Amazon in 13.2% of retrieval-supported answers. The source layer is therefore led by health publishers, comparison content, research sources, and retailer pages rather than brand-owned sites alone.
Who is winning and who is at risk?
Winning on measured visibility
Culturelle is in the leading group because it ranked first with a 56.4% appearance rate and appeared across all four engines. Seed is in the leading group because it ranked second at 50.5%, appeared across all four engines, and had a 21.7% owned-domain citation rate. Garden of Life and Align are also leading contenders on measured appearance, at 44.0% and 42.9%, though their intervals overlap with Culturelle's interval within this sample.
At risk from uneven coverage
Renew Life is at risk from uneven engine coverage because Claude named it in 0% of 72 answers, despite a 70.2% rate in ChatGPT's 57 answers. Ritual is at risk from a similar concentration pattern: 43.1% in Perplexity and 33.3% in Gemini, but 3.5% in ChatGPT. Visbiome and Klaire Labs are at risk from a Claude absence, while Physician's Choice had a 37.5% Perplexity rate but low or zero appearance elsewhere. These are visibility risks in this measured set, not assessments of product performance.
What structural patterns does the tally show?
The top of the leaderboard is concentrated among four brands. Culturelle, Seed, Garden of Life, and Align each appeared in more than 40.0% of answers returned, while Florastor ranked fifth at 30.8%. Culturelle's 50.5% to 62.2% interval overlaps Seed's 44.7% to 56.4% interval and reaches the upper boundary of Garden of Life's 38.2% to 49.9% interval. It also overlaps Align's 37.1% to 48.8% interval within the sample margin.
Engine convergence is strongest for the brands with 100.0% cross-engine consistency, but rate levels still vary widely by engine. The source tally is more concentrated outside brand domains: Healthline, Forbes, Fortune, PMC, and ConsumerLab all had higher answer penetration than the highest brand-owned domain, Seed at 21.7%. The brief named Physician's Choice, and the tally measured two name variants: Physician's Choice at 3.7% and Physician's Choice with a curly apostrophe at 11.0%. That naming split is a measurement issue worth resolving in future tracking.
What changed during the measurement window?
The data does not measure more than one time period, so it cannot identify change during the measurement window. It does show a cross-engine split within the collected answers: several brands were frequent in ChatGPT or Perplexity and absent from Claude.
What should probiotic brands measure next?
- Audit name normalization for Physician's Choice. The tally records a 3.7% appearance rate for one spelling and 11.0% for another. Re-measure after defining a single approved brand and product-name dictionary.
- Build pages that match strain-specific and digestive-health questions with clear product identity, formula details, and evidence references. Re-measure appearance by demand cluster to test whether the pages increase brand mentions in those prompt groups.
- Strengthen cited pages on owned sites for brands with high appearance but lower owned-domain citation rates. Culturelle appeared in 56.4% of returned answers but its domain was cited in 8.8% of retrieval-supported answers. Re-measure both rates separately.
- Diagnose engine-specific absences. Renew Life had a 70.2% ChatGPT rate and a 0% Claude rate. Compare the same prompt clusters in both systems after page updates.
- Investigate reliance on one engine. Physician's Choice appeared in 37.5% of Perplexity answers but at 1.8%, 2.8%, and 0% in ChatGPT, Gemini, and Claude. Re-measure cross-engine consistency after publication changes.
- Publish and maintain evidence pages that can compete with health-publisher and comparison content. Healthline appeared in 63.2% of retrieval-supported answers and Forbes in 62.1%. Re-measure owned-domain citation rate against this source layer.
What does this mean for probiotic brands?
Interpretation: visibility in these probiotic and gut health prompts depends on more than a brand being named. The strongest measured positions combine broad engine coverage with early mention placement and a cited owned site. Seed came closest to that combination among the leading consumer brands, while Culturelle had the highest overall appearance rate but a lower owned-domain citation rate.
Interpretation: a brand that appears frequently in one system can still have fragile coverage. Brands with zero appearance in Claude or low rates outside a single engine need to test whether their product pages and supporting evidence are discoverable across different retrieval systems. Publisher citations remain a separate competitive field.
How was this measured?
This flagship study measured 24 stratified prompts three times per engine across Perplexity, ChatGPT, Gemini, and Claude. Appearance rates use 273 answers returned. Citation and source rates use 272 answers with retrieval evidence. The study recorded 285 attempted answers, 13 answers without retrieval evidence, and no failed answers. Results describe these prompts on the measurement date and do not measure clinical outcomes, certification status, consumer demand, or product quality.
Study method
How this study was measured
- How was the sample set? The frame was fixed first: 24 prompts were allocated across the demand clusters before any questions were asked.
- Were engines asked identical questions? Yes: Perplexity (sonar-pro), ChatGPT (OpenAI), Gemini (Google), Claude (Anthropic), 3 trial(s) each.
- 273 answers returned of 285 attempted, 272 with retrieval evidence.
- How were product lines handled? They were assigned to parent brands under a registry fixed before scoring.
- How is uncertainty handled? Appearance rate uses a 95% Wilson interval; citation metrics are counted only for answers with evidence.
Appearance and citation rates cannot be compared as if they use the same denominator. Brand-name normalization also matters, as the separate Physician's Choice entries show. The automatic appendix contains the auditable study detail.