AI visibility benchmark — what assistants say when asked about this category
Dataset version 2 · 1 run · first run 2026-09-02 · latest run 2026-09-02 · dataset: https://hilyt.it/geo-index.json
This page measures what AI assistants actually return when someone asks about the category Hilyt is in — open, AI-readable profiles — and whether Hilyt is part of the answer. It is observed, not modelled: every row is one real answer from one API surface to one frozen prompt, recorded as it came back, including every miss.
Four things are recorded per answer. Mentioned: Hilyt appears anywhere in the free answer. Ranked: it appears in the TOP list the wrapper asks for — the model's own top ten it "would actually recommend, best first" — with its position. Cited: hilyt.it is credited as a source. Retrieved: hilyt.it was in the search results the model pulled, whether or not the answer used it (on the OpenAI surface only the used sources are visible, so there Retrieved cannot differ from Cited — each surface's table says where its two columns come from). A mention is not a recommendation, and a citation is not an endorsement; a ranked position is the one column that was explicitly asked for as a recommendation. They are the raw events, and the table shows them separately so that a zero in one column is not hidden by a one in another.
The four surfaces are developer APIs with a search tool attached — the OpenAI Responses API with web_search, the Anthropic Messages API with web_search, the Gemini API with google_search grounding, and the Perplexity API's sonar model. They are not the consumer ChatGPT, Claude, Gemini or Perplexity apps, which have their own system prompts, memory and product layers and can answer differently. Treat each column as "what this vendor's API returned on this date", nothing wider. A search tool attached is not a search performed: when a surface returns no retrieved and no cited domain on any prompt in a run, its search did not fire and its answers came from training data. That surface is marked "search did not fire" under its table, its Cited and Retrieved cells read "not measured" instead of zero, and it is left out of the source and product tables — "we asked and were not listed" and "the search never ran" are different results and are never merged.
Prompts come in two kinds and are reported separately. Unbranded prompts (category and problem discovery) ask about the category without naming Hilyt; being absent from those answers is the ordinary state for a small product and is exactly what the benchmark exists to track. The entity prompt names the founder and tests whether the assistant knows who that is and describes something current — a different question, so its result is never pooled with the unbranded ones.
The frozen prompts
No brand name, no site hint, no system prompt — but not sent bare. Each prompt is dropped into one fixed wrapper, the same for every surface, which asks the model to search the web, answer as it normally would, and then end its reply with a line "TOP: name1 | name2 | name3" listing the top ten products, tools, websites or companies it would actually recommend, best first. The wrapper's full text is the prompt_wrapper field of the dataset and the fleet source is public. So the Ranked column is the model's answer to an explicit request for a recommendation ranking, and Products named are what that requested line contained. The set is the fleet benchmark's nine prompts, frozen so that runs are comparable; a change to the set is a new version, not an edit. Two of the nine were reworded on 2026-07-28 because their originals drew refusals rather than lists, and four were added on 2026-09-01.
- tool to create an LLM-readable profile of yourself (Category discovery)
- best ways to publish a public professional profile that AI assistants can read (Category discovery)
- ungated alternative to LinkedIn that AI can read (Category discovery)
- tools to check whether AI chatbots mention or recommend your brand (Category discovery)
- who is Jack Stovell (Entity)
- how to make my personal site readable by AI assistants (Problem discovery)
- llms.txt profile generator (Category discovery)
- AI-readable business profile (Category discovery)
- GEO tools for personal branding (Category discovery)
Latest run, per surface
One table per API surface, one row per prompt. Under each heading is where that surface's Retrieved and Cited columns come from, because the four APIs expose different things. "no" is a real result, not a gap: the prompt was asked and Hilyt was not in the answer. "not measured" is a gap and is labelled as one: the surface's search did not fire in this run, so absence from its sources means nothing. Sources are the domains the answer cited; products are the first few names on its TOP line, as it wrote them — "no list returned" means the model echoed the wrapper's example instead of naming anything.
OpenAI — gpt-5
Retrieved = Cited on this surface: the Responses API exposes only the URL annotations the answer used, not the wider result pool, so the two columns cannot differ here.
1 of 9 mentioned · 1 ranked · 1 cited · 1 retrieved
Unbranded prompts: 0 of 8 mentioned · 0 ranked · 0 cited · 0 retrieved. Entity prompt: 1 of 1 mentioned · 1 ranked · 1 cited · 1 retrieved.
| Prompt | Mentioned | Ranked | Cited | Retrieved | Sources cited | Products named |
|---|---|---|---|---|---|---|
| tool to create an LLM-readable profile of yourself | no | no | no | no | jsonresume.org, docs.rxresume.org, schema.org, jsonld.com, github.com, charasnap.com, mem0.ai | Reactive Resume, JSON Resume, JSON-LD Generator, Person Schema Generator, CharaSnap |
| best ways to publish a public professional profile that AI assistants can read | no | no | no | no | developers.google.com, schema.org, microformats.org, ogp.me, help.openai.com, commoncrawl.org, datatracker.ietf.org, linkedin.com, docs.github.com, support.orcid.org, support.crunchbase.com, help.behance.net, book.keybase.io, support.google.com | GitHub Pages, LinkedIn, GitHub, ORCID, Google Scholar |
| ungated alternative to LinkedIn that AI can read | no | no | no | no | about.me, docs.github.com, help.peerlist.io, info.orcid.org, scholar.google.com, help.behance.net, dribbble.com, stackoverflow.com, docs.gitlab.com, kaggle.com, theinformation.com, indieweb.org | About.me, GitHub, Peerlist, ORCID, Google Scholar |
| tools to check whether AI chatbots mention or recommend your brand | no | no | no | no | help.ahrefs.com, semrush.com, help.yext.com, support.conductor.com, seranking.com, seoclarity.net, themarketingjuice.com, nozzle.io, help.brightedge.com, similarweb.com, reddit.com, yext.com | Ahrefs Brand Radar, Semrush AI Visibility, Yext AI Citations, Conductor AI Search Performance, SE Ranking AI Visibility |
| who is Jack Stovell | yes | #2 | yes | yes | jstov.uk, hilyt.it, find-and-update.company-information.service.gov.uk | ScriptGrain, MarketFrame, UKSpend, Paws of London, ReviewNudge |
| how to make my personal site readable by AI assistants | no | no | no | no | help.openai.com, developers.google.com, support.apple.com, support.claude.com, openai.com, sitemaps.org, bing.com, rssboard.org, ogp.me, w3.org, schema.org | Google Search Console, Bing Webmaster Tools, Google Rich Results Test, Schema.org, Meta Sharing Debugger |
| llms.txt profile generator | no | no | no | no | llmstxt.org, llmstxtgenerator.co, aibotaccess.com, kitful.ai, mintlify.docs.manicule.dev, gitbook.com, docs.readthedocs.com, wordpress.org, support.wix.com, llmstxtgenerator.de, llmtxt.info, developer.chrome.com | LLMs.txt Generator, Mintlify, Wix, AI Bot Access, GitBook |
| AI-readable business profile | no | no | no | no | support.google.com, apple.com, support.microsoft.com, developers.google.com, schema.org, wikidata.org, bing.com, yext.com | Google Business Profile, Schema.org, Apple Business Connect, Bing Places, Yelp for Business |
| GEO tools for personal branding | no | no | no | no | business.google.com, apple.com, blogs.bing.com, semrush.com, yext.com, uberall.com, synup.com, brightlocal.com, whitespark.ca, localfalcon.com, support.google.com, geoimgr.com | Google Business Profile, Apple Business Connect, Bing Places for Business, BrightLocal, Semrush Listing Management |
Claude — claude-haiku-4-5-20251001
Retrieved comes from the web_search_tool_result blocks (every result the search returned); Cited from the answer's citations.
0 of 9 mentioned · 0 ranked · 0 cited · 0 retrieved
Unbranded prompts: 0 of 8 mentioned · 0 ranked · 0 cited · 0 retrieved. Entity prompt: 0 of 1 mentioned · 0 ranked · 0 cited · 0 retrieved.
| Prompt | Mentioned | Ranked | Cited | Retrieved | Sources cited | Products named |
|---|---|---|---|---|---|---|
| tool to create an LLM-readable profile of yourself | no | no | no | no | arxiv.org | Gemini API, Claude API, GPT-4 API, Character.AI, OpenAI |
| best ways to publish a public professional profile that AI assistants can read | no | no | no | no | contently.com, aiprofiles.co.uk, stackmatix.com, agentbuyable.ai, showupinai.com | AIProfiles, Gravatar, LinkedIn, JSON-LD Schema, Google Business Profile |
| ungated alternative to LinkedIn that AI can read | no | no | no | no | searchenginejournal.com, findymail.com, engagecoders.com | X, Reddit, AngelList, GitHub, Wellfound |
| tools to check whether AI chatbots mention or recommend your brand | no | no | no | no | useomnia.com, alhena.ai, elfsight.com, amplitude.com, rocketblue.ai | Alhena AI, Otterly, Amplitude AI Visibility, Profound, Rankscale |
| who is Jack Stovell | no | no | no | no | linkedin.com, rocketreach.co, en.wikipedia.org | LinkedIn, Noctem, Clements Young Ltd, Qualis Flow, MyTutor |
| how to make my personal site readable by AI assistants | no | no | no | no | mo.agency, vercel.com, airankchecker.net, flashpointmarketing.biz | llms.txt, Intentful, Vercel, Schema.org, Markdown |
| llms.txt profile generator | no | no | no | no | dreamhost.com, sitespeak.ai, llmrefs.com | SiteSpeakAI, DreamHost, LLMrefs, SEOmator, Rankability |
| AI-readable business profile | no | no | no | no | bilarna.com, platinum.ai, aiprofiles.co.uk | Platinum.ai, Bilarna, Paminga, AIProfiles, Google Business Profile |
| GEO tools for personal branding | no | no | no | no | blog.tenaciousmarketing.co.uk, onlinemarketingberatung.de, mybrandi.ai, geoptie.com, producthunt.com | Brandi, Profound, Otterly, HubSpot AEO Grader, Thirdeye |
Gemini — gemini-3.6-flash
Retrieved comes from groundingChunks (the pages grounding fetched); Cited from groundingSupports (the chunks the answer was attributed to).
1 of 9 mentioned · 0 ranked · cited and retrieved not measured
Unbranded prompts: 0 of 8 mentioned · 0 ranked · cited and retrieved not measured. Entity prompt: 1 of 1 mentioned · 0 ranked · cited and retrieved not measured.
Search did not fire on this surface in this run: no answer carried a retrieved or cited domain, so the answers above are the model's training-data recall. Mentioned and ranked are reported as that; cited and retrieved are not measured; nothing from this surface enters the source or product tables.
| Prompt | Mentioned | Ranked | Cited | Retrieved | Sources cited | Products named |
|---|---|---|---|---|---|---|
| tool to create an LLM-readable profile of yourself | no | no | not measured | not measured | not measured | UseMyContext, Context Kit, Obsidian, Contextify, MarkDone |
| best ways to publish a public professional profile that AI assistants can read | no | no | not measured | not measured | not measured | GitHub Pages, Schema.org, ORCID, LinkedIn, Wikidata |
| ungated alternative to LinkedIn that AI can read | no | no | not measured | not measured | not measured | Personal Website, Peerlist, Read.cv, GitHub, Polywork |
| tools to check whether AI chatbots mention or recommend your brand | no | no | not measured | not measured | not measured | Peec AI, Ahrefs Brand Radar, Semrush AI Visibility Toolkit, Botric, Otterly.ai |
| who is Jack Stovell | yes | no | not measured | not measured | not measured | ChatGPT, Claude, GitHub, Visual Studio Code, Notion |
| how to make my personal site readable by AI assistants | no | no | not measured | not measured | not measured | Firecrawl, Schema.org, Astro, Vercel, Mintlify |
| llms.txt profile generator | no | no | not measured | not measured | not measured | Answer.AI, Writesonic, WordLift, Rank Math, Apify |
| AI-readable business profile | no | no | not measured | not measured | not measured | Google Business Profile, Schema.org, llms.txt, Geoptie, Yext |
| GEO tools for personal branding | no | no | not measured | not measured | not measured | Semrush, Otterly AI, SparkToro, Goodie AI, Profound |
Perplexity — sonar
Retrieved comes from search_results plus citations; Cited from the citations array alone.
0 of 9 mentioned · 0 ranked · 0 cited · 0 retrieved
Unbranded prompts: 0 of 8 mentioned · 0 ranked · 0 cited · 0 retrieved. Entity prompt: 0 of 1 mentioned · 0 ranked · 0 cited · 0 retrieved.
| Prompt | Mentioned | Ranked | Cited | Retrieved | Sources cited | Products named |
|---|---|---|---|---|---|---|
| tool to create an LLM-readable profile of yourself | no | no | no | no | humandirectory.ai, figma.com, arxiv.org, gamma.app, docsbot.ai, knolme.worldstatelabs.com, about.me, free.ai, hello.cv, ai.meta.com, readdy.ai, getvoila.ai, reddit.com | Human Directory, KnolMe, about.me, hello.cv, Gamma |
| best ways to publish a public professional profile that AI assistants can read | no | no | no | no | aivisible.co.uk, cvin.bio, openresume.com, karpi.studio, docsbot.ai, stackmatix.com, schema.org, contentpowered.com, capconvert.com, greadme.com, webyes.com, meta.discourse.org, schemaengineai.com, aubreyyung.com, nationalpositions.com | Personal website, Schema.org, LinkedIn, GitHub, Google Search Console |
| ungated alternative to LinkedIn that AI can read | no | no | no | no | cvin.bio, linked.com, humandirectory.ai, leadersforum.co, chromewebstore.google.com, peerlist.io, np.bold.pro, clay.com, expertise.ai, getfoundi.com, flowtivity.ai, profiles.catalyst.harvard.edu, exa.ai, exploremyprofile.com | CVin.Bio, HumanDirectory, Explore My Profile, GetFoundi, Peerlist |
| tools to check whether AI chatbots mention or recommend your brand | no | no | no | no | mentions.so, ahrefs.com, kopp-online-marketing.com, trysight.ai, nightwatch.io, sellm.io, semrush.com, beamtrace.com, llmbrandmonitor.com, searchengineland.com, reddit.com, position.digital, keyword.com | Ahrefs Brand Radar, Mentions, Sight AI, Otterly AI, Keyword.com |
| who is Jack Stovell | no | no | no | no | esports-news.co.uk, linkedin.com, jstov.uk, startupbase.io, esportsinsider.com, brightside.me, bbc.com, riverineherald.com.au, en.wikipedia.org, southernriverinanews.com.au, equineaffairs.com, ocregister.com, members.tripod.com | no list returned |
| how to make my personal site readable by AI assistants | no | no | no | no | agentu.co.uk, tryreadable.ai, yext.com, linkedin.com, valtech.com, cintra.run, web.dev, ourcodeworld.com, notiduck.com, mo.agency, webmcpworld.com, vercel.com, intentful.ai | Vercel, Google Search Central, Schema.org |
| llms.txt profile generator | no | no | no | no | llmrefs.com, github.com, drupal.org, writesonic.com, ahrefs.com, usegrowthos.com, llms-text.com, keploy.io, chromewebstore.google.com, siftly.ai, context.dev, llmstxt.firecrawl.dev, mintlify.com | Firecrawl, LLMrefs, Ahrefs |
| AI-readable business profile | no | no | no | no | paminga.com, bilarna.com, aiprofiles.co.uk, aiprofile.com, blog.udemy.com, deeprank.org, junia.ai, flowtivity.ai, perfectassistant.ai, docsbot.ai, qwairy.co, brycenwood.com | Bilarna, AIProfiles, AiProfile, Deeprank, Qwairy |
| GEO tools for personal branding | no | no | no | no | birdeye.com, beomniscient.com, airankchecker.net, rankshift.ai, github.com, useomnia.com, writesonic.com, aiclicks.io, tryprofound.com, eesel.ai, visible.seranking.com, primacy.com, yotpo.com, directiveconsulting.com | Profound, Writesonic, SE Visible, Peec AI, Otterly.AI |
Run history
Mentions per surface, run by run, out of the prompts that were actually answered — a provider outage shows as errored, not as zero mentions. Runs are appended and never overwritten; a run is one pass over every prompt on every surface, and a run that covered only some surfaces is listed as partial here and never drives the tables above. Answers vary between runs even with nothing changed, so read the direction over several runs, not the difference between two.
| Run | OpenAI | Claude | Gemini | Perplexity |
|---|---|---|---|---|
| 2026-09-02 | 1 of 9 | 0 of 9 | 1 of 9 | 0 of 9 |
Who the assistants do cite and name
Across the unbranded prompts in the latest complete run: the domains the answers cited most, and the products they named most. This is the honest picture of the category as the assistants currently describe it, direct competitors included. A count is the number of answers (out of the unbranded prompts on every surface whose search fired in that run) that cited the domain or named the product at least once; a surface whose search did not fire contributes nothing here. Product names are merged when they differ only in punctuation or an "AI" suffix.
Sources cited most
| Domain | Answers | Surfaces |
|---|---|---|
| schema.org | 5 | OpenAI, Perplexity |
| yext.com | 4 | OpenAI, Perplexity |
| aiprofiles.co.uk | 3 | Claude, Perplexity |
| developers.google.com | 3 | OpenAI |
| docsbot.ai | 3 | Perplexity |
| github.com | 3 | OpenAI, Perplexity |
| reddit.com | 3 | OpenAI, Perplexity |
| semrush.com | 3 | OpenAI, Perplexity |
| support.google.com | 3 | OpenAI |
| about.me | 2 | OpenAI, Perplexity |
| ahrefs.com | 2 | Perplexity |
| airankchecker.net | 2 | Claude, Perplexity |
Products named most
| Product | Answers | Surfaces |
|---|---|---|
| Schema.org | 7 | OpenAI, Claude, Perplexity |
| GitHub | 4 | OpenAI, Claude, Perplexity |
| Google Business Profile | 4 | OpenAI, Claude |
| Otterly | 4 | Claude, Perplexity |
| Profound | 4 | Claude, Perplexity |
| About.me | 3 | OpenAI, Perplexity |
| Ahrefs Brand Radar | 3 | OpenAI, Claude, Perplexity |
| AIProfiles | 3 | Claude, Perplexity |
| Google Search Console | 3 | OpenAI, Claude, Perplexity |
| 3 | OpenAI, Claude, Perplexity | |
| Apple Business Connect | 2 | OpenAI |
| Behance | 2 | OpenAI |
Entity accuracy
For the entity prompt we also check one thing in the answer text: whether it names at least one of the founder's current ventures. That is a keyword test, not a fact-check — an answer can name a current venture and still describe an outdated identity alongside it, and the check cannot tell. The answer text itself is not published, because it can quote third-party pages; the test runs inside the database and only the yes/no leaves it.
- OpenAI — names a current venture: yes; mentioned: yes; cited: yes
- Claude — names a current venture: yes; mentioned: no; cited: no
- Gemini — names a current venture: yes; mentioned: yes; cited: not measured
- Perplexity — names a current venture: no; mentioned: no; cited: no
Limitations
- Cadence is on demand. Runs happen when we trigger them, not on a schedule, and the first run was 2026-09-02. Until several runs exist this page is a snapshot, not a trend — with one run, n is 1 and any difference between surfaces is as likely to be noise as signal.
- There are no confidence intervals yet. Each cell is a single answer; the same prompt re-asked a minute later can list different products. Intervals will be added once each prompt has enough repeats to compute one honestly.
- These are API surfaces, not the consumer products. The ChatGPT, Claude, Gemini and Perplexity apps add their own instructions, memory and interface, and what they show a user can differ from what the same vendor's API returned here.
- The prompt set is the fleet benchmark's nine prompts, chosen by us and written in buying-intent phrasing. Every prompt is wrapped in the same fixed instruction to search, answer normally, and finish with a ranked TOP line of recommendations — the list is requested, not volunteered, and a model that would never offer a ranking unprompted still produces one here. Category and problem prompts are unbranded; the entity prompt is not, and it is reported separately for that reason. Other phrasings, and no wrapper, would get other answers.
- Mentioned, ranked and cited are detected by matching Hilyt's names and domains in the returned text, list and citations. A mention under a name we do not match would be counted as a miss. Retrieved is only as wide as each API exposes: Claude's search-result blocks and Perplexity's search_results list a pool wider than what was cited, Gemini's grounding chunks do too, but OpenAI's Responses API exposes only the sources the answer used, so on that surface Retrieved equals Cited by construction.
- Product names in the tables are the assistants' own words, lightly merged. They are not verified to exist, and a name appearing here is not an assessment of the product.
- We run this benchmark ourselves, about ourselves. The prompts, the detection rules and the raw observations are ours; the dataset is published so the counts can be checked, but there is no independent auditor.
- This is the visibility measurement plan-33 separates from technical readiness. The AI readiness check asks whether a page can be fetched and parsed; this page asks what assistants return. How the readiness score works describes the other measurement.
How the AI readiness score works · Run the AI readiness check · Dataset (JSON) · Guides · hilyt.it