<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
  <channel>
    <title>DEV Community: Gissur Runarsson</title>
    <description>The latest articles on DEV Community by Gissur Runarsson (@gthorr).</description>
    <link>https://dev.to/gthorr</link>
    <image>
      <url>https://media2.dev.to/dynamic/image/width=90,height=90,fit=cover,gravity=auto,format=auto/https:%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F3755182%2F8e86704a-8616-4672-8508-1d34b30db515.jpg</url>
      <title>DEV Community: Gissur Runarsson</title>
      <link>https://dev.to/gthorr</link>
    </image>
    <atom:link rel="self" type="application/rss+xml" href="https://dev.to/feed/gthorr"/>
    <language>en</language>
    <item>
      <title>Bersyn is live on Product Hunt today</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Wed, 15 Jul 2026 09:57:09 +0000</pubDate>
      <link>https://dev.to/gthorr/bersyn-is-live-on-product-hunt-today-119o</link>
      <guid>https://dev.to/gthorr/bersyn-is-live-on-product-hunt-today-119o</guid>
      <description>&lt;p&gt;Bersyn is live on Product Hunt today.&lt;/p&gt;

&lt;p&gt;I built it after watching ChatGPT, Claude, Perplexity and Gemini recommend my competitors instead of me, and realizing most founders never find out it is happening to them.&lt;/p&gt;

&lt;p&gt;Type your domain and see who AI recommends in your category, who it names instead of you, and whether you show up at all. Free, no signup.&lt;/p&gt;

&lt;p&gt;If you have two minutes, an honest comment on the launch would mean a lot. The first few hours are what count on Product Hunt: &lt;a href="https://www.producthunt.com/products/bersyn/launches/bersyn-2" rel="noopener noreferrer"&gt;https://www.producthunt.com/products/bersyn/launches/bersyn-2&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;And if it is useful, run your own domain through it and tell me what AI says about your category.&lt;/p&gt;

</description>
      <category>ai</category>
      <category>startup</category>
      <category>marketing</category>
      <category>buildinpublic</category>
    </item>
    <item>
      <title>AI keeps recommending my competitors instead of me. Wednesday I launch Bersyn on Product Hunt.</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Mon, 13 Jul 2026 11:45:16 +0000</pubDate>
      <link>https://dev.to/gthorr/ai-keeps-recommending-my-competitors-instead-of-me-wednesday-i-launch-bersyn-on-product-hunt-50c8</link>
      <guid>https://dev.to/gthorr/ai-keeps-recommending-my-competitors-instead-of-me-wednesday-i-launch-bersyn-on-product-hunt-50c8</guid>
      <description>&lt;p&gt;Next Wednesday I'm launching Bersyn on Product Hunt.&lt;/p&gt;

&lt;p&gt;It started from something that annoyed me. I asked ChatGPT, Claude, Perplexity and Gemini the questions my buyers actually ask, and they kept recommending my competitors instead of me. Most founders never find out it is happening to them.&lt;/p&gt;

&lt;p&gt;Bersyn shows you, in seconds, who AI recommends in your category, who it names instead of you, and whether you show up at all. Free, no signup.&lt;/p&gt;

&lt;p&gt;Launching Wednesday. If you would want to see what AI says about your own company, I'll drop the link that morning.&lt;/p&gt;

</description>
      <category>buildinpublic</category>
      <category>ai</category>
      <category>saas</category>
      <category>startup</category>
    </item>
    <item>
      <title>I'm launching Bersyn on Product Hunt next Wednesday</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Sat, 11 Jul 2026 14:33:12 +0000</pubDate>
      <link>https://dev.to/gthorr/im-launching-bersyn-on-product-hunt-next-wednesday-15oc</link>
      <guid>https://dev.to/gthorr/im-launching-bersyn-on-product-hunt-next-wednesday-15oc</guid>
      <description>&lt;p&gt;Next Wednesday I'm launching Bersyn on Product Hunt.&lt;/p&gt;

&lt;p&gt;It started from something that annoyed me. I asked ChatGPT, Claude, Perplexity and Gemini the questions my buyers actually ask, and they kept recommending my competitors instead of me. Most founders never find out it is happening to them.&lt;/p&gt;

&lt;p&gt;Bersyn shows you, in seconds, who AI recommends in your category, who it names instead of you, and whether you show up at all. Free, no signup.&lt;/p&gt;

&lt;p&gt;Launching Wednesday. If you would want to see what AI says about your own company, I'll drop the link that morning.&lt;/p&gt;

</description>
      <category>startup</category>
      <category>ai</category>
      <category>buildinpublic</category>
      <category>saas</category>
    </item>
    <item>
      <title>We asked AI for the best CRM 80 times. Attio, the CRM startups love, was named twice.</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Fri, 10 Jul 2026 17:45:06 +0000</pubDate>
      <link>https://dev.to/bersyn/we-asked-ai-for-the-best-crm-80-times-attio-the-crm-startups-love-was-named-twice-1k6</link>
      <guid>https://dev.to/bersyn/we-asked-ai-for-the-best-crm-80-times-attio-the-crm-startups-love-was-named-twice-1k6</guid>
      <description>&lt;p&gt;More buyers now ask ChatGPT, Claude, Perplexity, or Gemini for a tool before they ever open Google. So we keep measuring what those models actually say when someone asks for software in a category. CRM turned out to be one of the cleanest examples we have run, because of Attio.&lt;/p&gt;

&lt;p&gt;Attio is the modern CRM that startups reach for and rave about. Well funded, well designed, exactly the tool the online startup crowd talks about. If funding and affection were what the models rewarded, Attio would be named constantly. It is not.&lt;/p&gt;

&lt;p&gt;We asked two versions of the same question and ran each one 10 times on all four models. That is 40 answers per question, 80 in total.&lt;/p&gt;

&lt;p&gt;Version one was the plain category label:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;What's the best CRM software for a business?&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Version two described the buyer Attio is actually built for, with no tool named:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;What should an early-stage startup use to track deals and customer relationships without a heavy enterprise sales tool?&lt;/p&gt;
&lt;/blockquote&gt;

&lt;h2&gt;
  
  
  The receipt
&lt;/h2&gt;

&lt;p&gt;How many of the 10 runs named each tool, per model.&lt;br&gt;
&lt;/p&gt;

&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight plaintext"&gt;&lt;code&gt;                    Best CRM                For an early-stage startup
                    (category label)        (no heavy enterprise tool)
Model               Attio   Salesforce      Attio   Salesforce
-----------------   -----   ----------      -----   ----------
ChatGPT             0/10    10/10           0/10     0/10
Claude              0/10    10/10           2/10     7/10
Gemini              0/10    10/10           0/10     0/10
Perplexity          0/10    10/10           0/10     8/10
-----------------   -----   ----------      -----   ----------
All four models     0/40    40/40           2/40    15/40

Named instead, both questions: HubSpot, Pipedrive, Zoho, and on the
startup question, Notion, Airtable, and Streak.
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;



&lt;h2&gt;
  
  
  Attio is not phrasing-sensitive. It is just absent.
&lt;/h2&gt;

&lt;p&gt;This is the part that makes CRM different from other categories we have run. With some tools the naming swings depending on how you ask. Attio does not swing. It was named zero times out of 40 on the plain question and 2 times out of 40 on the startup question. Across all 80 answers it came up twice, both on Claude. The category is owned by Salesforce, HubSpot, Pipedrive, and Zoho, and Attio is simply not in the conversation.&lt;/p&gt;

&lt;h2&gt;
  
  
  The models can reason about startup fit. They just do not know Attio belongs there.
&lt;/h2&gt;

&lt;p&gt;Here is the tell that this is not the models failing to understand the question. Look at Salesforce on the second question. It went from 40 of 40 down to 15. The models correctly noticed that an early-stage startup that does not want a heavy enterprise tool probably should not be handed Salesforce, so they dropped it and reached for lighter options, HubSpot, Pipedrive, Notion, Airtable, Streak. The reasoning is right there. They know how to pick a leaner CRM for a smaller team. They just have never learned that Attio is one of the answers.&lt;/p&gt;

&lt;p&gt;That is the cleanest version of a corroboration problem. It is not that the model cannot find a startup CRM. It is that the specific sentence, Attio is the modern CRM for an early-stage startup, does not exist widely enough in the places the model reads for it to reach for that name.&lt;/p&gt;

&lt;h2&gt;
  
  
  What this means if you are the challenger
&lt;/h2&gt;

&lt;p&gt;Funding does not move the model. Product love does not move the model. A great design that your users adore in private does not become a public sentence the model can learn from. Salesforce and HubSpot are named because two decades of comparison posts, reviews, docs, and threads taught the model they are the answer. Attio has the product. It does not yet have the corpus.&lt;/p&gt;

&lt;p&gt;The work is not a better CRM. It is getting Attio written about, by real users, as the answer to the specific jobs it wins, in the public places these models read. Not more mentions of the category, the exact sentence tied to the job.&lt;/p&gt;

&lt;h2&gt;
  
  
  Check your own category
&lt;/h2&gt;

&lt;p&gt;Pick your category. Ask the plain label question and the version that describes the specific job your best customers hire you for. Run both a handful of times on ChatGPT, Claude, Perplexity, and Gemini and count who gets named. If you are the beloved, well-funded challenger who still does not show up, the fix is not the product.&lt;/p&gt;

&lt;p&gt;If you want the fast version, run your own domain through the free scan and see who AI recommends in your category and who it names instead: &lt;a href="https://bersyn.com/?utm_source=devto&amp;amp;utm_medium=referral&amp;amp;utm_campaign=crm-attio-teardown" rel="noopener noreferrer"&gt;bersyn.com free scan&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;In your category, is the tool that everyone online talks about the same one AI actually names when a buyer asks?&lt;/p&gt;

</description>
      <category>ai</category>
      <category>saas</category>
      <category>startup</category>
      <category>marketing</category>
    </item>
    <item>
      <title>We asked AI for the best email tool 40 times. Beehiiv barely showed up. Then we asked about newsletters.</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Fri, 10 Jul 2026 17:45:05 +0000</pubDate>
      <link>https://dev.to/bersyn/we-asked-ai-for-the-best-email-tool-40-times-beehiiv-barely-showed-up-then-we-asked-about-44hk</link>
      <guid>https://dev.to/bersyn/we-asked-ai-for-the-best-email-tool-40-times-beehiiv-barely-showed-up-then-we-asked-about-44hk</guid>
      <description>&lt;blockquote&gt;
&lt;p&gt;&lt;strong&gt;Correction, 2026-07-16.&lt;/strong&gt; The original version of this post said the lesson was about framing: reword your category question as the specific job and the challenger wins. A reader, Paul Towers, caught a confound. The two questions below did not only change from a category label to a job. They also changed the buyer, from a business to a creator. So I reran it holding the buyer constant, 10 times per model, and the framing effect disappeared. A business buyer who describes the job in plain words still gets Beehiiv 0 out of 40, the same as the category label. The corrected finding is at the top. The original numbers are left untouched below, for the record.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;h2&gt;
  
  
  What actually moved the result: who is asking, not how they phrase it
&lt;/h2&gt;

&lt;p&gt;I reran the test with a third question, keeping the buyer a business the whole time and only changing the framing from a category label to the specific job.&lt;br&gt;
&lt;/p&gt;

&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight plaintext"&gt;&lt;code&gt;Question                                                    Beehiiv (all 4 models)
--------------------------------------------------------    ----------------------
Business buyer, category label                              3/40
  "What's the best email marketing platform for a business?"
Business buyer, the job in plain words                      0/40
  "We have an existing customer base we want to reach via
   an owned channel like email. Walk me through it."
Creator buyer, the job                                      25/39
  "How does a creator or founder start a newsletter, grow
   subscribers, and make money from it?"
--------------------------------------------------------    ----------------------
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;



&lt;p&gt;Holding the buyer constant and changing only the framing takes Beehiiv from 3 to zero. There is no framing effect. The entire jump was the buyer identity. Beehiiv is filed under "creator," and it vanishes for a business buyer at any phrasing. ChatGPT named it 0 out of 10 on all three questions, which is the one part of the original that survives.&lt;/p&gt;

&lt;p&gt;The real lesson is harder and truer than "phrase it as a job." You cannot reframe your way into a category. The models file you under a buyer identity, written in the words the people who actually use you have published about you. Beehiiv owns the creator's newsletter job because the creator world documented it as the answer to that job. A business buyer never reaches it, because nobody wrote Beehiiv up as the answer to a business's email program, no matter how that business phrases the question.&lt;/p&gt;

&lt;p&gt;So if you are the challenger: the move is not to reword your question. It is to pick the specific buyer whose job you can genuinely own, and get documented, by real users and real publications, as the answer to that job in the words that buyer uses. Then the models reach for you when that buyer asks, and only when that buyer asks. Watch ChatGPT separately, because it moves last.&lt;/p&gt;




&lt;p&gt;&lt;em&gt;Original post below, kept for the record.&lt;/em&gt;&lt;/p&gt;




&lt;p&gt;More buyers now ask ChatGPT, Claude, Perplexity, or Gemini for a tool before they ever open Google. So we keep measuring what those models actually say when someone asks for software in a category. Email marketing gave us the most encouraging result we have run, because of Beehiiv. It is the one teardown where the challenger actually won its slot, and you can see exactly how.&lt;/p&gt;

&lt;p&gt;We asked two versions of the same question and ran each one 10 times on all four models. That is 40 answers per question.&lt;/p&gt;

&lt;p&gt;Version one was the plain category label:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;What's the best email marketing platform for a business?&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Version two was the specific job Beehiiv is built for:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;How does a creator or founder start a newsletter, grow subscribers, and make money from it?&lt;/p&gt;
&lt;/blockquote&gt;

&lt;h2&gt;
  
  
  The receipt
&lt;/h2&gt;

&lt;p&gt;How many of the 10 runs named each tool, per model.&lt;br&gt;
&lt;/p&gt;

&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight plaintext"&gt;&lt;code&gt;                    Best email tool         Start and grow a newsletter
                    (category label)        (the job Beehiiv is built for)
Model               Beehiiv  Mailchimp      Beehiiv  Mailchimp
-----------------   -------  ---------      -------  ---------
ChatGPT             0/10     10/10          0/10     10/10
Claude              0/10     10/10          10/10    10/10
Gemini              0/10     10/10          9/10     10/10
Perplexity          4/10      9/10          10/10     7/10
-----------------   -------  ---------      -------  ---------
All four models     4/40     39/40          29/40    37/40

Named on the generic question: Mailchimp, Klaviyo, ActiveCampaign, Brevo,
Constant Contact. Named on the newsletter question: Beehiiv, Kit, Substack.
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;



&lt;h2&gt;
  
  
  On the category label, Beehiiv is almost invisible
&lt;/h2&gt;

&lt;p&gt;Ask for the best email marketing platform and Beehiiv was named 4 times out of 40, all of them on Perplexity. Mailchimp was named 39 times, with Klaviyo, ActiveCampaign, Brevo, and Constant Contact filling the rest. If Beehiiv tried to win the generic email marketing question, it would lose, the same way every challenger loses the category label.&lt;/p&gt;

&lt;h2&gt;
  
  
  The one model that still will not budge
&lt;/h2&gt;

&lt;p&gt;Look at ChatGPT. Even on the newsletter question, the question written for Beehiiv's exact buyer, ChatGPT named it 0 times out of 10. This keeps happening in our teardowns. ChatGPT is consistently the slowest of the four to name a challenger, even one that the other three now reach for easily. If ChatGPT is where your buyers ask, winning the sub-intent on the other three is not enough on its own, and it is worth knowing that before you assume the job is done.&lt;/p&gt;

&lt;h2&gt;
  
  
  Check your own category
&lt;/h2&gt;

&lt;p&gt;Pick your category. Ask the plain label question, then ask it the way your actual customers describe the job they hire you for, in their own words. Run both a handful of times on ChatGPT, Claude, Perplexity, and Gemini and count who gets named. If you are invisible on the label but present when the question sounds like your real buyer, that tells you which buyer the models have you filed under. If you are invisible on both, &lt;br&gt;
Gemini              0/10     10/10          9/10     10/10&lt;br&gt;
Perplexity          4/10      9/10          10/10     7/10&lt;/p&gt;




&lt;p&gt;All four models     4/40     39/40          29/40    37/40&lt;/p&gt;

&lt;p&gt;Named on the generic question: Mailchimp, Klaviyo, ActiveCampaign, Brevo,&lt;br&gt;
Constant Contact. Named on the newsletter question: Beehiiv, Kit, Substack.&lt;/p&gt;



&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight plaintext"&gt;&lt;code&gt;


## On the category label, Beehiiv is almost invisible

Ask for the best email marketing platform and Beehiiv was named 4 times out of 40, all of them on Perplexity. Mailchimp was named 39 times, with Klaviyo, ActiveCampaign, Brevo, and Constant Contact filling the rest. If Beehiiv tried to win the generic email marketing question, it would lose, the same way every challenger loses the category label.

## The one model that still will not budge

Look at ChatGPT. Even on the newsletter question, the question written for Beehiiv's exact buyer, ChatGPT named it 0 times out of 10. This keeps happening in our teardowns. ChatGPT is consistently the slowest of the four to name a challenger, even one that the other three now reach for easily. If ChatGPT is where your buyers ask, winning the sub-intent on the other three is not enough on its own, and it is worth knowing that before you assume the job is done.

## Check your own category

Pick your category. Ask the plain label question, then ask it the way your actual customers describe the job they hire you for, in their own words. Run both a handful of times on ChatGPT, Claude, Perplexity, and Gemini and count who gets named. If you are invisible on the label but present when the question sounds like your real buyer, that tells you which buyer the models have you filed under. If you are invisible on both, you have a job to go own.

If you want the fast version, run your own domain through the free scan and see who AI recommends in your category: [bersyn.com free scan](https://bersyn.com/?utm_source=devto&amp;amp;utm_medium=referral&amp;amp;utm_campaign=email-beehiiv-teardown).

Which buyer are the models filing you under, and is it the one you actually sell to?
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;

</description>
      <category>ai</category>
      <category>saas</category>
      <category>marketing</category>
      <category>startup</category>
    </item>
    <item>
      <title>We asked AI for the best project management tool 40 times. Linear, the one engineers love, was named twice.</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Fri, 10 Jul 2026 17:44:04 +0000</pubDate>
      <link>https://dev.to/bersyn/we-asked-ai-for-the-best-project-management-tool-40-times-linear-the-one-engineers-love-was-3k4m</link>
      <guid>https://dev.to/bersyn/we-asked-ai-for-the-best-project-management-tool-40-times-linear-the-one-engineers-love-was-3k4m</guid>
      <description>&lt;p&gt;More buyers now ask ChatGPT, Claude, Perplexity, or Gemini for a tool before they ever open Google. So we keep measuring what those models actually say when someone asks for software in a category. Project management turned out to be one of the sharpest examples we have run, because of one tool: Linear.&lt;/p&gt;

&lt;p&gt;Linear is a darling among engineers. Fast, opinionated, built for exactly the teams that talk about their tools online. If adoption and affection were what the models rewarded, Linear would be named constantly. It is not.&lt;/p&gt;

&lt;p&gt;We asked two versions of the same question and ran each one 10 times on all four models, so a single lucky or unlucky answer could not fool us. That is 40 answers per question.&lt;/p&gt;

&lt;p&gt;Version one was the plain category label:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;What's the best project management software for a team?&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Version two was the actual job, phrased the way the buyer Linear is built for would phrase it, with no tool named:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;How does a fast-moving engineering team track issues and plan sprints with as little process overhead as possible?&lt;/p&gt;
&lt;/blockquote&gt;

&lt;h2&gt;
  
  
  The receipt
&lt;/h2&gt;

&lt;p&gt;How many of the 10 runs named Linear, and named Jira, on each model.&lt;br&gt;
&lt;/p&gt;

&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight plaintext"&gt;&lt;code&gt;                    Best PM software        Fast engineering team
                    (category label)        (the actual job)
Model               Linear    Jira          Linear    Jira
-----------------   ------    ----          ------    ----
ChatGPT             0 / 10    10 / 10        0 / 10    10 / 10
Claude              2 / 10    10 / 10        5 / 10     9 / 10
Gemini              0 / 10    10 / 10        9 / 10    10 / 10
Perplexity          0 / 10    10 / 10        1 / 10    10 / 10
-----------------   ------    ----          ------    ----
All four models     2 / 40    40 / 40       15 / 40   39 / 40
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;



&lt;h2&gt;
  
  
  On the category label, Linear is basically invisible
&lt;/h2&gt;

&lt;p&gt;Ask for the best project management software and Linear was named 2 times out of 40. On ChatGPT, Gemini, and Perplexity it was a flat zero. Jira was named on all 40 runs, and behind it the models handed back the same incumbents every time: Asana, Monday.com, Trello, ClickUp, Notion. The category label is settled, and Linear is not in the settlement, no matter how many engineering teams love it.&lt;/p&gt;

&lt;h2&gt;
  
  
  Reframe to Linear's actual job and it appears, but only on some models
&lt;/h2&gt;

&lt;p&gt;When we asked the second question, the one that describes what Linear is actually for, Linear climbed to 15 of 40. But look at where the movement came from. Gemini went from 0 to 9. Claude went from 2 to 5. Perplexity barely moved, 0 to 1. And ChatGPT stayed at a flat zero even on the question written for Linear's exact buyer.&lt;/p&gt;

&lt;p&gt;So the tool that engineers reach for first is a tool the biggest model will not name, even when you describe its job precisely. Meanwhile Jira held at 39 and 40. The incumbent survives both questions. The challenger only exists on the phrasing that matches its job, and only on the models willing to name it.&lt;/p&gt;

&lt;h2&gt;
  
  
  What this means if you are the challenger
&lt;/h2&gt;

&lt;p&gt;Two things worth taking from this.&lt;/p&gt;

&lt;p&gt;The category label is not your fight. If you sell into a category with an entrenched incumbent, the generic question is already lost, and asking it harder will not change that. Linear is proof: real adoption, real affection, still 2 of 40 on the plain question.&lt;/p&gt;

&lt;p&gt;Adoption does not automatically become recommendation. A model names what has been written down in a form it can read, in the words the buyer uses. Engineers picking Linear in a Slack thread does not teach the model anything. The opening is the job-framed question, the specific sub-intent where you actually fit, published clearly enough and corroborated widely enough that a model reaches for you there. That is the sentence worth owning, not the category label you will lose.&lt;/p&gt;

&lt;h2&gt;
  
  
  Check your own category
&lt;/h2&gt;

&lt;p&gt;Pick your category. Write the plain label version and the version that describes the specific job your best customers hire you for. Run both a handful of times on ChatGPT, Claude, Perplexity, and Gemini and watch who gets named as the wording changes. If you are the beloved challenger who never shows up on the generic question, you are in good company, and the fix is not a better product.&lt;/p&gt;

&lt;p&gt;If you want the fast version, run your own domain through the free scan and see who AI recommends in your category and how the wording moves it: &lt;a href="https://bersyn.com/?utm_source=devto&amp;amp;utm_medium=referral&amp;amp;utm_campaign=pm-linear-teardown" rel="noopener noreferrer"&gt;bersyn.com free scan&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;In your category, does the tool everyone actually uses match the tool AI actually names, or are those two different lists?&lt;/p&gt;

</description>
      <category>ai</category>
      <category>saas</category>
      <category>productivity</category>
      <category>startup</category>
    </item>
    <item>
      <title>I watched a shared inbox tool go from 10 out of 10 to zero on ChatGPT just by rewording the question</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Fri, 10 Jul 2026 12:14:39 +0000</pubDate>
      <link>https://dev.to/gthorr/i-watched-a-shared-inbox-tool-go-from-10-out-of-10-to-zero-on-chatgpt-just-by-rewording-the-question-51ih</link>
      <guid>https://dev.to/gthorr/i-watched-a-shared-inbox-tool-go-from-10-out-of-10-to-zero-on-chatgpt-just-by-rewording-the-question-51ih</guid>
      <description>&lt;p&gt;I have been running a small experiment for a few weeks, and this one genuinely stopped me for a second.&lt;/p&gt;

&lt;p&gt;More and more buyers ask ChatGPT or Claude for a tool recommendation before they ever open Google. So I have been measuring what the models actually say when someone asks for the best tool in a category, and how much the answer moves when you change nothing but the wording.&lt;/p&gt;

&lt;p&gt;Shared inbox software gave me the clearest example so far.&lt;/p&gt;

&lt;p&gt;I asked two versions of the same question, 10 times each, on ChatGPT, Claude, Perplexity and Gemini. First the plain category label, "best shared inbox software". Then the way a real person actually types it, "how do we collaborate on shared email inboxes without losing our normal email workflow".&lt;/p&gt;

&lt;p&gt;On the label question, Missive got named on all 10 runs. On the reworded one, ChatGPT named Missive zero times out of ten, while it kept naming Front and Help Scout on all ten. Same company, same site, same ten runs. The only thing that changed was the sentence.&lt;/p&gt;

&lt;p&gt;To be fair it was not every model. Claude actually leaned the other way and named Missive more on the reworded question. But the ChatGPT drop was clean and it repeated across all ten runs.&lt;/p&gt;

&lt;p&gt;The thing I keep taking from these is that a single check will lie to you. If I had asked once, seen Missive missing, and stopped there, I would have sworn they had some deep AI problem. They do not, they are 10 out of 10 on the other phrasing. One question on one model on one run tells you almost nothing.&lt;/p&gt;

&lt;p&gt;And the tools that held did not do anything clever. Their pages just describe the actual job in the words a buyer uses, so the model can place them no matter how the question is asked.&lt;/p&gt;

&lt;p&gt;I wrote up the full data, every number and every model, here: &lt;a href="https://www.bersyn.com/blog/shared-inbox-phrasing-flip" rel="noopener noreferrer"&gt;https://www.bersyn.com/blog/shared-inbox-phrasing-flip&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;I am running one of these teardowns roughly every week, poking at a different category. If that is your kind of rabbit hole, follow along. And I am curious, in your own category, is the recommendation this sensitive to phrasing, or did shared inbox just happen to be a swingy one?&lt;/p&gt;

</description>
      <category>ai</category>
      <category>saas</category>
      <category>buildinpublic</category>
      <category>startup</category>
    </item>
    <item>
      <title>We asked AI to recommend tools in three SaaS categories 20 times each, and the shape of the answers told us which ones a challenger can still win</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Thu, 09 Jul 2026 16:30:11 +0000</pubDate>
      <link>https://dev.to/bersyn/we-asked-ai-to-recommend-tools-in-three-saas-categories-20-times-each-and-the-shape-of-the-answers-2fec</link>
      <guid>https://dev.to/bersyn/we-asked-ai-to-recommend-tools-in-three-saas-categories-20-times-each-and-the-shape-of-the-answers-2fec</guid>
      <description>&lt;p&gt;More buyers now ask ChatGPT, Claude, Perplexity, or Gemini for a tool before they ever open Google. So we started measuring what those models actually say when someone asks for software in a category. This week we ran a simple probe across three developer categories, and the result was less about who won and more about the shape of each race. Some categories are frozen at the top. Some are wide open. And the shape tells a founder what is even worth attempting.&lt;/p&gt;

&lt;p&gt;We took three categories, asked one plain buyer question in each, and ran that question 5 times on each of the four models. That is 20 answers per category. We counted how many of the 20 named each company.&lt;/p&gt;

&lt;h2&gt;
  
  
  The receipt
&lt;/h2&gt;

&lt;p&gt;Here are the three categories, the top named tools, and how many of the 20 answers named each one.&lt;br&gt;
&lt;/p&gt;

&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight plaintext"&gt;&lt;code&gt;20 answers per category (5 runs x 4 models: ChatGPT, Claude, Perplexity, Gemini)

Product analytics          Feature flags              Error monitoring
(frozen at the top)        (wide open)                (owned, but deep tail)
-----------------------    -----------------------    -----------------------
Mixpanel      20 / 20      LaunchDarkly  20 / 20      Sentry        20 / 20
Amplitude     20 / 20      Flagsmith     17 / 20      Bugsnag       19 / 20
Heap          19 / 20      Optimizely    17 / 20      Rollbar       15 / 20
PostHog       15 / 20      Split         15 / 20      Raygun        14 / 20
--- cliff ---              GrowthBook    13 / 20      LogRocket     11 / 20
Segment        7 / 20      Unleash       13 / 20      Datadog        6 / 20
Hotjar         6 / 20      Statsig       11 / 20
FullStory      5 / 20      PostHog       11 / 20
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;



&lt;p&gt;Read the three columns as three different shapes, because that is the whole point.&lt;/p&gt;

&lt;h2&gt;
  
  
  Product analytics is frozen
&lt;/h2&gt;

&lt;p&gt;Mixpanel and Amplitude were named on every single answer, all four models, all five runs. Heap held at 19 of 20. PostHog came in at 15, and then there is a cliff: Segment 7, Hotjar 6, FullStory 5. The top of this category is concrete. If you are a new analytics tool, the generic "best product analytics tool" question is not a fight you win right now. The models have already settled on the top two, and asking the plain question harder will not change that.&lt;/p&gt;

&lt;h2&gt;
  
  
  Feature flags is wide open
&lt;/h2&gt;

&lt;p&gt;Now look at the middle column. LaunchDarkly led at 20 of 20, the way you would expect from the incumbent. But under it the field stays alive: Flagsmith 17, Optimizely 17, Split 15, GrowthBook 13, Unleash 13, Statsig 11, PostHog 11. That is eight tools with real, repeated share in the same 20 answers. There is no cliff here. A challenger in this category genuinely gets named, because the models have not collapsed the answer down to two names. If this is your category, the move is not clever. It is to describe what you do clearly enough that a model can place you, because the slot is actually available.&lt;/p&gt;

&lt;h2&gt;
  
  
  Error monitoring is owned, but the tail runs deep
&lt;/h2&gt;

&lt;p&gt;Sentry owned this one at 20 of 20. Bugsnag was right behind at 19, then Rollbar 15, Raygun 14, LogRocket 11. So the top is locked, but unlike product analytics there is no cliff after the leader. Five and six tools deep still get named.&lt;/p&gt;

&lt;p&gt;The most useful number in the whole probe is Datadog at 6 of 20 here. Datadog is everywhere in monitoring. Ask the generic question and it shows up constantly. It came in low in this probe because we asked the question the way a small team asks it, "for a small dev team", and that sub-intent quietly pushed the heavyweight down and let the leaner tools surface. The lock is real on the generic question. It breaks on the sub-intent.&lt;/p&gt;

&lt;h2&gt;
  
  
  What the shape tells you to do
&lt;/h2&gt;

&lt;p&gt;The shape of your category tells you what to even attempt.&lt;/p&gt;

&lt;p&gt;If you are in a frozen top-two category like product analytics, do not spend your energy trying to get named next to the incumbents on the generic question. You will not, and the probe shows why: the top two are named on every answer and there is a cliff behind them. The opening is in the sub-intents. The questions that carry a qualifier, "for small teams", "cheapest", "for [specific use case]", are where the lock cracks, the same way Datadog drops out the moment the question says "small dev team". A frozen category is not unwinnable. It is winnable one specific buyer question at a time, not on the category label.&lt;/p&gt;

&lt;p&gt;If you are in an open category like feature flags, the game is different and easier. The models are still deciding, so you get named by being legible. Describe what you actually do, in the words a buyer uses, and a model can place you next to the eight names already in the mix. You do not need a trick. You need to be clear while the category is still open.&lt;/p&gt;

&lt;p&gt;Either way, the first thing to know is which shape you are in, because it decides whether you chase the generic question or go hunting for the sub-intent where the incumbents thin out.&lt;/p&gt;

&lt;h2&gt;
  
  
  Check your own category
&lt;/h2&gt;

&lt;p&gt;Pick your category. Ask one plain buyer question, run it a handful of times on ChatGPT, Claude, Perplexity, and Gemini, and count who gets named. If the same two names show up on every answer with a cliff behind them, you are frozen, and your opening is in the sub-intents. If a real field of tools keeps appearing, you are open, and your job is to describe yourself clearly enough to join it.&lt;/p&gt;

&lt;p&gt;If you want the fast version, run your own domain through the free scan and see who AI recommends in your category and how deep the field goes: &lt;a href="https://bersyn.com/?utm_source=blog&amp;amp;utm_medium=referral&amp;amp;utm_campaign=frozen-vs-open" rel="noopener noreferrer"&gt;run the free scan&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;In your category, is the top frozen at two names, or is the field still open?&lt;/p&gt;

</description>
      <category>ai</category>
      <category>saas</category>
      <category>startup</category>
      <category>marketing</category>
    </item>
    <item>
      <title>We reworded one buyer question and watched a shared inbox tool drop from 10 out of 10 to zero on ChatGPT</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Thu, 09 Jul 2026 15:42:03 +0000</pubDate>
      <link>https://dev.to/bersyn/we-reworded-one-buyer-question-and-watched-a-shared-inbox-tool-drop-from-10-out-of-10-to-zero-on-17f5</link>
      <guid>https://dev.to/bersyn/we-reworded-one-buyer-question-and-watched-a-shared-inbox-tool-drop-from-10-out-of-10-to-zero-on-17f5</guid>
      <description>&lt;p&gt;More and more buyers ask an AI model before they ever open Google. So we started measuring what ChatGPT, Claude, Perplexity, and Gemini actually say when someone asks for a tool in a category. Shared inbox software gave us the cleanest example we have found so far of something that should worry anyone who sells software.&lt;/p&gt;

&lt;p&gt;We asked two versions of the same question. Then we ran each version 10 times on all four models, so a single lucky or unlucky answer could not fool us.&lt;/p&gt;

&lt;p&gt;Version one was the plain category label:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;What's the best shared inbox software for small teams?&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Version two was how an actual buyer types it when they have the real problem in their head:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;What's the best way to collaborate with my team on shared email inboxes without losing our email workflow?&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Same intent. Same shortlist of tools that could answer it. The only thing that changed was the wording.&lt;/p&gt;

&lt;h2&gt;
  
  
  The receipt
&lt;/h2&gt;

&lt;p&gt;Here is ChatGPT, the same question asked both ways, 10 runs each. The number is how many of the 10 runs named that company.&lt;br&gt;
&lt;/p&gt;

&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight plaintext"&gt;&lt;code&gt;ChatGPT, 10 runs per question

Company        Category label      Reworded as a buyer
------------   ----------------    -------------------
Missive        10 / 10             0 / 10
Hiver          10 / 10             9 / 10
Front          10 / 10             10 / 10
Help Scout     10 / 10             10 / 10
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;



&lt;p&gt;Read the middle column again. On the category label, ChatGPT recommended Missive on every single run. Reword the exact same question the way a real buyer would say it, and Missive was named zero times out of ten. Same company, same website, same 10 runs, one different sentence.&lt;/p&gt;

&lt;p&gt;Hiver barely moved (10 to 9). Front and Help Scout did not move at all on ChatGPT. They held a perfect 10 out of 10 through the rewording.&lt;/p&gt;

&lt;h2&gt;
  
  
  It was not just ChatGPT, and not just Missive
&lt;/h2&gt;

&lt;p&gt;Perplexity produced the mirror image. On the category label it named Hiver on all 10 runs. On the reworded question it named Hiver zero times, and started pointing people toward doing it in Gmail and Outlook instead of naming a product at all.&lt;/p&gt;

&lt;p&gt;The direction of the swing depended on the model. On Claude the reworded question pulled Hiver down (8 to 5) and nudged Missive down (10 to 8). On Gemini the reworded question actually lifted Hiver (6 to 9) while Missive held (9 to 8). Front and Help Scout were the steadiest of the four, holding 10 out of 10 across ChatGPT, Claude and Gemini, and only softening on Perplexity's reworded question, where almost everything dropped because Perplexity stopped naming products. There was no single pattern across models. The one thing that stayed calm everywhere was the plain category question. The reworded question is where companies appeared and disappeared.&lt;/p&gt;

&lt;h2&gt;
  
  
  Two things this should change about how you check
&lt;/h2&gt;

&lt;p&gt;&lt;strong&gt;A single scan will lie to you.&lt;/strong&gt; If we had asked ChatGPT once, seen Missive missing, and stopped there, we would have walked away certain Missive had some deep AI problem. They do not. They are recommended on 10 of 10 runs on the other phrasing. One question, on one model, on one run, tells you close to nothing. The behavior only shows up across runs and across phrasings.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;The companies that held did not do anything clever.&lt;/strong&gt; Front and Help Scout were not gaming a model. Their pages describe the actual job a buyer is trying to do, in the words a buyer uses, so a model can place them no matter how the question is asked. The companies that only carry the category label survive the label question and come apart on the workflow one. This is a legibility problem, not a trick. It is whether your site says what you do in the buyer's own words clearly enough that a model can recommend you across every way the question gets asked.&lt;/p&gt;

&lt;p&gt;If you sell software, your category label is the question you already rank for in your head. The reworded, real-buyer version is the one you have probably never checked. That is the one deciding whether AI names you.&lt;/p&gt;

&lt;h2&gt;
  
  
  Check your own category
&lt;/h2&gt;

&lt;p&gt;Pick your category. Write the plain label version and the way one of your actual customers would ask it. Run both a handful of times on ChatGPT, Claude, Perplexity, and Gemini, and watch which of your competitors appear and disappear as the wording changes.&lt;/p&gt;

&lt;p&gt;If you want the fast version, you can run your own domain through the free scan and see who AI recommends in your category and how the wording moves it: &lt;a href="https://bersyn.com/?utm_source=blog&amp;amp;utm_medium=referral&amp;amp;utm_campaign=shared-inbox-teardown" rel="noopener noreferrer"&gt;run the free scan&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;In your category, is the recommendation this sensitive to phrasing, or did shared inbox software just happen to be an unusually swingy one?&lt;/p&gt;

</description>
      <category>ai</category>
      <category>saas</category>
      <category>marketing</category>
      <category>seo</category>
    </item>
    <item>
      <title>Why AI recommends the same tools every time, and which slots you can actually win</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Tue, 30 Jun 2026 12:52:11 +0000</pubDate>
      <link>https://dev.to/bersyn/why-ai-recommends-the-same-tools-every-time-and-which-slots-you-can-actually-win-850</link>
      <guid>https://dev.to/bersyn/why-ai-recommends-the-same-tools-every-time-and-which-slots-you-can-actually-win-850</guid>
      <description>&lt;p&gt;Last week I &lt;a href="https://www.bersyn.com/blog/ai-recommends-neon-for-databases-specialists-invisible-2026" rel="noopener noreferrer"&gt;tore down what four AI models recommend for databases&lt;/a&gt;: same buyer questions, 20 runs each, count who gets named. Neon, Upstash and Turso already own the generic slots. The specialists, Tigris and Tinybird, are close to invisible.&lt;/p&gt;

&lt;p&gt;The post did fine. The comments did better. A handful of sharp people pushed on the &lt;em&gt;why&lt;/em&gt;, and the answer is more useful than the findings were.&lt;/p&gt;

&lt;h2&gt;
  
  
  Two competitions, and you are probably in the wrong one
&lt;/h2&gt;

&lt;p&gt;The models are not judging which product is best. They surface whatever their training data already corroborated for that exact phrasing. A worse tool with denser, clearer writing tied to a question beats a better tool nobody wrote about that way.&lt;/p&gt;

&lt;p&gt;So there are two competitions running at once:&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;Best product.&lt;/strong&gt; What you actually build.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Best-corroborated answer.&lt;/strong&gt; What the sources a buyer's question pulls from already named.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Most founders pour everything into the first and assume the second follows. It does not. The model only scores the second.&lt;/p&gt;

&lt;h2&gt;
  
  
  Some slots are welded shut. Stop fighting them.
&lt;/h2&gt;

&lt;p&gt;"Object storage" has meant Amazon S3 for fifteen years. That mental model is set in concrete across millions of pages. No comparison article, no migration guide, no amount of content moves "best object storage" off S3 on any timeline that matters to a startup. If your growth plan depends on winning a query like that, the plan is the problem.&lt;/p&gt;

&lt;p&gt;The tell for a welded-shut slot: ask the four models the same generic question a few times and they all agree, every run. Agreement across models and across runs means consensus has formed. You are not getting in.&lt;/p&gt;

&lt;h2&gt;
  
  
  Some slots are wide open. That is where the work pays.
&lt;/h2&gt;

&lt;p&gt;Now ask a narrow sub-job instead. "Vector database for X." "Analytics for Y." Watch what happens: the models hedge, name different tools, and the first pick shuffles run to run. That disagreement is the signal. Consensus has not formed yet, which means the slot is still being decided, which means you can be the one it decides on.&lt;/p&gt;

&lt;p&gt;The challengers that won did exactly this. Neon, Upstash and Turso did not beat Postgres at "best database." They became the corroborated answer for "serverless Postgres / Redis / SQLite" while those mental models were still forming, and rode them as they widened.&lt;/p&gt;

&lt;p&gt;So the move in a locked category is not to attack the incumbent's query. It is to find the sub-job nobody owns, become the best-corroborated answer for it in the buyer's own words, and let the category grow around you. Manufacture a young category you can actually win.&lt;/p&gt;

&lt;h2&gt;
  
  
  You can see which is which
&lt;/h2&gt;

&lt;p&gt;This is the part founders miss: you do not have to guess whether a slot is welded or winnable. You can observe it. Run the buyer questions across the models, more than once, and look at the agreement. Tight agreement across models and runs is a settled slot. Disagreement is an open one. That turns "get recommended by AI" from a vibe into a map of where to spend.&lt;/p&gt;

&lt;p&gt;That map is what &lt;a href="https://www.bersyn.com" rel="noopener noreferrer"&gt;Bersyn&lt;/a&gt; builds: who each model names, who gets recommended first, where they disagree, and the verbatim answers behind every number. If you want it for your own category, that is the whole product.&lt;/p&gt;

&lt;p&gt;&lt;em&gt;Method: category-representative buyer questions across ChatGPT, Claude, Gemini and Perplexity, multiple runs, reported with model versions and scan dates. No claim that any tool is good or bad, only what the models answered.&lt;/em&gt;&lt;/p&gt;

</description>
      <category>ai</category>
      <category>marketing</category>
      <category>startup</category>
      <category>seo</category>
    </item>
    <item>
      <title>I asked four AI models which database to use in 2026. Neon already won. Four challengers are invisible.</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Mon, 29 Jun 2026 07:40:24 +0000</pubDate>
      <link>https://dev.to/bersyn/i-asked-four-ai-models-which-database-to-use-in-2026-neon-already-won-four-challengers-are-533g</link>
      <guid>https://dev.to/bersyn/i-asked-four-ai-models-which-database-to-use-in-2026-neon-already-won-four-challengers-are-533g</guid>
      <description>&lt;p&gt;Every week I take one buyer category, ask ChatGPT, Claude, Gemini and Perplexity the five questions a real buyer would type, and count who gets named and who gets recommended first. Same questions for every company, so it is a fair board and not a vibe.&lt;/p&gt;

&lt;p&gt;This week: databases and storage. I expected the incumbent reflex. I got the opposite, then a twist.&lt;/p&gt;

&lt;h2&gt;
  
  
  The serverless newcomers already won
&lt;/h2&gt;

&lt;p&gt;To the models, the challengers are already the answer:&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;Neon&lt;/strong&gt; (serverless Postgres): recommended first in 14 of 20 conversations.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Upstash&lt;/strong&gt; (serverless Redis): first in 11 of 20.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Turso&lt;/strong&gt; (edge SQLite): first in 9 of 20.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;That is real proof the door is not locked. A company younger than the incumbent it replaced can become AI's default pick. Neon did it.&lt;/p&gt;

&lt;h2&gt;
  
  
  Then I asked about the specialized jobs
&lt;/h2&gt;

&lt;p&gt;Same category, different sub-job, and the model snaps back to the incumbent every time:&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;Vector search:&lt;/strong&gt; the models pick Pinecone, Milvus and Weaviate. &lt;a href="https://www.bersyn.com/recommends/databases-storage/qdrant" rel="noopener noreferrer"&gt;Qdrant&lt;/a&gt; is named a lot but recommended first only six times.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Object storage:&lt;/strong&gt; the answer is Amazon S3, Cloudflare R2 and Backblaze. &lt;a href="https://www.bersyn.com/recommends/databases-storage/tigris" rel="noopener noreferrer"&gt;Tigris&lt;/a&gt; is invisible, named zero times on ChatGPT, Claude and Gemini.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Real-time analytics:&lt;/strong&gt; ClickHouse, not &lt;a href="https://www.bersyn.com/recommends/databases-storage/tinybird" rel="noopener noreferrer"&gt;Tinybird&lt;/a&gt;.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;The ORM:&lt;/strong&gt; Prisma, not &lt;a href="https://www.bersyn.com/recommends/databases-storage/drizzle" rel="noopener noreferrer"&gt;Drizzle&lt;/a&gt;.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2&gt;
  
  
  Why
&lt;/h2&gt;

&lt;p&gt;The newcomers that won did not win on features. By the time these models trained, enough independent writing already named them as the answer. They are in the inputs the model reads.&lt;/p&gt;

&lt;p&gt;The invisible ones shipped great products, and it does not register, because the model is not evaluating products. It repeats what its inputs said. You cannot test your way into a recommendation you were never part of, and you cannot out-feature your way in either. The buyer who types "best vector database" and takes the first answer never sees Qdrant, no matter how good it is, until the inputs change.&lt;/p&gt;

&lt;h2&gt;
  
  
  See it yourself
&lt;/h2&gt;

&lt;p&gt;The full board, with every model's answer and the verbatim text, is here: &lt;a href="https://www.bersyn.com/recommends/databases-storage" rel="noopener noreferrer"&gt;What AI recommends for Databases and Storage&lt;/a&gt;. Counts only, no score, every number links to the actual answer.&lt;/p&gt;

&lt;p&gt;If you want the same teardown for your own category, that is what &lt;a href="https://www.bersyn.com" rel="noopener noreferrer"&gt;Bersyn&lt;/a&gt; does.&lt;/p&gt;

</description>
      <category>ai</category>
      <category>database</category>
      <category>webdev</category>
      <category>saas</category>
    </item>
    <item>
      <title>I asked four AI models which observability tool to use in 2026. They keep naming Datadog and Splunk, and never Better Stack.</title>
      <dc:creator>Gissur Runarsson</dc:creator>
      <pubDate>Fri, 26 Jun 2026 11:19:40 +0000</pubDate>
      <link>https://dev.to/bersyn/i-asked-four-ai-models-which-observability-tool-to-use-in-2026-they-keep-naming-datadog-and-5fkg</link>
      <guid>https://dev.to/bersyn/i-asked-four-ai-models-which-observability-tool-to-use-in-2026-they-keep-naming-datadog-and-5fkg</guid>
      <description>&lt;p&gt;We build Bersyn, a tool that tracks which products AI models name when someone asks for a recommendation in a category. So we run a lot of scans. This is the third category in a row where the same thing happened, and observability is the cleanest example yet, so it is worth showing the receipts.&lt;/p&gt;

&lt;p&gt;The question is the one a real engineer types into ChatGPT: "what is the best platform to monitor and debug my B2B SaaS app in production, and what are the strong alternatives?" We asked it five ways across four Surfaces: ChatGPT, Claude, Gemini and Perplexity. Then we measured how often each modern tool actually got named, each one scanned in its own home category.&lt;/p&gt;

&lt;h2&gt;
  
  
  The modern tools are not in the answer
&lt;/h2&gt;

&lt;p&gt;Here is the Recommendation Share for five tools that engineers actually talk about, measured across the four Surfaces.&lt;br&gt;
&lt;/p&gt;

&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight plaintext"&gt;&lt;code&gt;                ChatGPT   Claude   Perplexity   Gemini
Better Stack       0%       0%         0%          0%
Axiom              0%      40%         0%          0%
Highlight          0%      40%         0%          0%
OpenStatus         0%       0%        80%          0%
Checkly           20%      60%        40%         20%
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;



&lt;p&gt;Read the Better Stack row again. Zero, zero, zero, zero. Not a low score, an absence on every Surface we tested. Axiom and Highlight are named by exactly one model, Claude, and by none of the other three. OpenStatus exists only on Perplexity. A buyer who opens ChatGPT, which is most of them, walks away from this category never having heard four of these five names.&lt;/p&gt;

&lt;h2&gt;
  
  
  What AI names instead
&lt;/h2&gt;

&lt;p&gt;So who got recommended in their place? Here is the tool AI reached for first when each modern challenger was not named.&lt;br&gt;
&lt;/p&gt;

&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight plaintext"&gt;&lt;code&gt;Better Stack   -&amp;gt;  Splunk, Datadog, Loggly
Axiom          -&amp;gt;  Datadog, Honeycomb
Highlight      -&amp;gt;  FullStory, LogRocket, Sentry
OpenStatus     -&amp;gt;  Cachet
Checkly        -&amp;gt;  Pingdom, Datadog Synthetics, Grafana
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;



&lt;p&gt;Here is ChatGPT, verbatim, asked for the best log management and uptime monitoring platform, the exact category Better Stack sells into:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;Choosing the best log management and uptime monitoring platform for a B2B SaaS team depends on several factors... ### Log Management Platforms 1. &lt;strong&gt;Splunk&lt;/strong&gt; - Pros: Highly scalable, powerful search capabilities, extensive integrations, and strong data visualization tools.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;And here is ChatGPT on observability and log management, the category Axiom sells into:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;Here are some of the top platforms that are widely recognized for their capabilities in observability and log management: 1. &lt;strong&gt;Datadog&lt;/strong&gt;: Datadog is a comprehensive monitoring and analytics platform for developers, IT operations teams, and businesses...&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Splunk. Datadog. Pingdom. Cachet. LogRocket. Look at that list. These are the names that dominated the monitoring conversation around 2015 to 2018, when the training data was thick. Each modern challenger loses, in its own home category, to an incumbent from the era before it existed.&lt;/p&gt;

&lt;h2&gt;
  
  
  This is not AI being clueless about the category
&lt;/h2&gt;

&lt;p&gt;Here is the part that makes it a real problem rather than a funny one. AI is not ignorant of observability. Ask it and it confidently knows two newer names:&lt;br&gt;
&lt;/p&gt;

&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight plaintext"&gt;&lt;code&gt;                ChatGPT   Claude   Perplexity   Gemini
Sentry           100%      80%        60%         60%
Honeycomb         40%      40%        80%         40%
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;



&lt;p&gt;Sentry gets named in every single ChatGPT answer. Honeycomb shows up across all four models. So the models have room for modern tools in their mental map of monitoring. They have simply frozen that map around the incumbents plus the one or two challengers that broke through years ago. Everything that arrived after the map froze is Omitted.&lt;/p&gt;

&lt;p&gt;That gap has a name in our world: Model Disagreement. When Axiom is named by Claude and by none of the other three, it is not "low visibility." It is invisible on the Surfaces most buyers use, and visible on the one they use least.&lt;/p&gt;

&lt;h2&gt;
  
  
  Why this should bother a founder, not just amuse them
&lt;/h2&gt;

&lt;p&gt;The reflex is to wave it off. AI is behind, the models have a training cutoff, it will catch up. Maybe. But your buyer is asking the question today, and the answer they get today hands them Datadog.&lt;/p&gt;

&lt;p&gt;Search had twenty years to learn that Better Stack exists. AI recommendation answers are being formed right now, off whatever evidence the models can find, and for newer companies that evidence is thin. So the incumbent gets named by reflex and the challenger gets skipped.&lt;/p&gt;

&lt;p&gt;The four ways a company shows up wrong in these answers, in our vocabulary:&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;Omitted. The model lists competitors and skips you. This is Better Stack, Axiom and Highlight on ChatGPT.&lt;/li&gt;
&lt;li&gt;Misclassified. The model files you under the wrong category.&lt;/li&gt;
&lt;li&gt;Generic. The model names you so vaguely no buyer could shortlist you.&lt;/li&gt;
&lt;li&gt;Confused. The model conflates you with a similarly named competitor.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2&gt;
  
  
  What actually moves it
&lt;/h2&gt;

&lt;p&gt;We do not claim AI is biased or that anyone is paying for placement. We read what the models say and show you the evidence. What changes the answer over time is the same thing that changed search: published, specific, verifiable evidence that associates your company with the category and the buyer question. Comparison pages that name the alternatives honestly. Documentation that states plainly what you are and who you are for. Third party mentions in the exact words an engineer would use when asking.&lt;/p&gt;

&lt;p&gt;None of that is fast. But the first step is not writing more content. It is finding out what the models say about you right now, so you know whether you are Omitted, Generic or Confused, because the fix is different for each.&lt;/p&gt;

&lt;h2&gt;
  
  
  See your own category
&lt;/h2&gt;

&lt;p&gt;We built Bersyn to show you exactly the tables above, for your company, with the verbatim answers behind them. Run a free scan on your own product at &lt;a href="https://www.bersyn.com" rel="noopener noreferrer"&gt;bersyn.com&lt;/a&gt; and see which Surfaces name you, which name a competitor in your place, and why.&lt;/p&gt;

&lt;p&gt;If you ship on Better Stack, Axiom, Highlight, OpenStatus or Checkly, tell me in the comments which model gets your category right. This is the third category I have scanned where ChatGPT defaults to the old incumbent and skips everything newer, and the pattern is the most interesting thing I look at all week.&lt;/p&gt;

</description>
      <category>observability</category>
      <category>devops</category>
      <category>ai</category>
      <category>startup</category>
    </item>
  </channel>
</rss>
