<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
  <channel>
    <title>DEV Community: MUGIRANEZA JOHN</title>
    <description>The latest articles on DEV Community by MUGIRANEZA JOHN (@mujohn26).</description>
    <link>https://dev.to/mujohn26</link>
    <image>
      <url>https://media2.dev.to/dynamic/image/width=90,height=90,fit=cover,gravity=auto,format=auto/https:%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F463332%2F14fc1cda-0026-491b-be7a-fa21df38a666.jpeg</url>
      <title>DEV Community: MUGIRANEZA JOHN</title>
      <link>https://dev.to/mujohn26</link>
    </image>
    <atom:link rel="self" type="application/rss+xml" href="https://dev.to/feed/mujohn26"/>
    <language>en</language>
    <item>
      <title>[Boost]</title>
      <dc:creator>MUGIRANEZA JOHN</dc:creator>
      <pubDate>Thu, 24 Sep 2026 13:06:27 +0000</pubDate>
      <link>https://dev.to/mujohn26/-3o9g</link>
      <guid>https://dev.to/mujohn26/-3o9g</guid>
      <description>&lt;div class="ltag__link--embedded"&gt;
  &lt;div class="crayons-story "&gt;
  &lt;a href="https://dev.to/mujohn26/small-ai-models-can-see-the-trap-they-fall-in-anyway-4cag" class="crayons-story__hidden-navigation-link"&gt;Small AI Models Can See the Trap. They Fall In Anyway&lt;/a&gt;


  &lt;div class="crayons-story__body crayons-story__body-full_post"&gt;
      &lt;a href="https://dev.to/mujohn26/small-ai-models-can-see-the-trap-they-fall-in-anyway-4cag" class="crayons-article__context-note crayons-article__context-note__feed"&gt;&lt;p&gt;Kaggle Benchmarking Challenge Submission&lt;/p&gt;

&lt;/a&gt;
    &lt;div class="crayons-story__top"&gt;
      &lt;div class="crayons-story__meta"&gt;
        &lt;div class="crayons-story__author-pic"&gt;

          &lt;a href="/mujohn26" class="crayons-avatar  crayons-avatar--l  "&gt;
            &lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F463332%2F14fc1cda-0026-491b-be7a-fa21df38a666.jpeg" alt="mujohn26 profile" class="crayons-avatar__image" width="340" height="340"&gt;
          &lt;/a&gt;
        &lt;/div&gt;
        &lt;div&gt;
          &lt;div&gt;
            &lt;a href="/mujohn26" class="crayons-story__secondary fw-medium m:hidden"&gt;
              MUGIRANEZA JOHN
            &lt;/a&gt;
            &lt;div class="profile-preview-card relative mb-4 s:mb-0 fw-medium hidden m:inline-block"&gt;
              
                MUGIRANEZA JOHN
                
                
              
              &lt;div id="story-author-preview-content-4734126" class="profile-preview-card__content crayons-dropdown branded-7 p-4 pt-0"&gt;
                &lt;div class="gap-4 grid"&gt;
                  &lt;div class="-mt-4"&gt;
                    &lt;a href="/mujohn26" class="flex"&gt;
                      &lt;span class="crayons-avatar crayons-avatar--xl mr-2 shrink-0"&gt;
                        &lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F463332%2F14fc1cda-0026-491b-be7a-fa21df38a666.jpeg" class="crayons-avatar__image" alt="" width="340" height="340"&gt;
                      &lt;/span&gt;
                      &lt;span class="crayons-link crayons-subtitle-2 mt-5"&gt;MUGIRANEZA JOHN&lt;/span&gt;
                    &lt;/a&gt;
                  &lt;/div&gt;
                  &lt;div class="print-hidden"&gt;
                    
                      Follow
                    
                  &lt;/div&gt;
                  &lt;div class="author-preview-metadata-container"&gt;&lt;/div&gt;
                &lt;/div&gt;
              &lt;/div&gt;
            &lt;/div&gt;

          &lt;/div&gt;
          &lt;a href="https://dev.to/mujohn26/small-ai-models-can-see-the-trap-they-fall-in-anyway-4cag" class="crayons-story__tertiary fs-xs"&gt;&lt;time&gt;Sep 24&lt;/time&gt;&lt;span class="time-ago-indicator-initial-placeholder"&gt;&lt;/span&gt;&lt;/a&gt;
        &lt;/div&gt;
      &lt;/div&gt;

    &lt;/div&gt;

    &lt;div class="crayons-story__indention"&gt;
      &lt;h2 class="crayons-story__title crayons-story__title-full_post"&gt;
        &lt;a href="https://dev.to/mujohn26/small-ai-models-can-see-the-trap-they-fall-in-anyway-4cag" id="article-link-4734126"&gt;
          Small AI Models Can See the Trap. They Fall In Anyway
        &lt;/a&gt;
      &lt;/h2&gt;
        &lt;div class="crayons-story__tags"&gt;
            &lt;a class="crayons-tag  crayons-tag--monochrome " href="/t/devchallenge"&gt;&lt;span class="crayons-tag__prefix"&gt;#&lt;/span&gt;devchallenge&lt;/a&gt;
            &lt;a class="crayons-tag  crayons-tag--monochrome " href="/t/kagglechallenge"&gt;&lt;span class="crayons-tag__prefix"&gt;#&lt;/span&gt;kagglechallenge&lt;/a&gt;
            &lt;a class="crayons-tag  crayons-tag--monochrome " href="/t/ai"&gt;&lt;span class="crayons-tag__prefix"&gt;#&lt;/span&gt;ai&lt;/a&gt;
            &lt;a class="crayons-tag  crayons-tag--monochrome " href="/t/machinelearning"&gt;&lt;span class="crayons-tag__prefix"&gt;#&lt;/span&gt;machinelearning&lt;/a&gt;
        &lt;/div&gt;
      &lt;div class="crayons-story__bottom"&gt;
        &lt;div class="crayons-story__details"&gt;
          &lt;a href="https://dev.to/mujohn26/small-ai-models-can-see-the-trap-they-fall-in-anyway-4cag" class="crayons-btn crayons-btn--s crayons-btn--ghost crayons-btn--icon-left"&gt;
            &lt;div class="multiple_reactions_aggregate"&gt;
              &lt;span class="multiple_reactions_icons_container"&gt;
                  &lt;span class="crayons_icon_container"&gt;
                    &lt;img src="https://assets.dev.to/assets/sparkle-heart-5f9bee3767e18deb1bb725290cb151c25234768a0e9a2bd39370c382d02920cf.svg" width="24" height="24"&gt;
                  &lt;/span&gt;
              &lt;/span&gt;
              &lt;span class="aggregate_reactions_counter"&gt;1&lt;span class="hidden s:inline"&gt;&amp;nbsp;reaction&lt;/span&gt;&lt;/span&gt;
            &lt;/div&gt;
          &lt;/a&gt;
            &lt;a href="https://dev.to/mujohn26/small-ai-models-can-see-the-trap-they-fall-in-anyway-4cag#comments" class="crayons-btn crayons-btn--s crayons-btn--ghost crayons-btn--icon-left flex items-center"&gt;
              

              &lt;span class="hidden s:inline"&gt;Add&amp;nbsp;Comment&lt;/span&gt;
            &lt;/a&gt;
        &lt;/div&gt;
        &lt;div class="crayons-story__save"&gt;
          &lt;small class="crayons-story__tertiary fs-xs mr-2"&gt;
            5 min read
          &lt;/small&gt;
        &lt;/div&gt;
      &lt;/div&gt;
    &lt;/div&gt;
  &lt;/div&gt;
&lt;/div&gt;

&lt;/div&gt;


</description>
    </item>
    <item>
      <title>Small AI Models Can See the Trap. They Fall In Anyway</title>
      <dc:creator>MUGIRANEZA JOHN</dc:creator>
      <pubDate>Thu, 24 Sep 2026 13:06:04 +0000</pubDate>
      <link>https://dev.to/mujohn26/small-ai-models-can-see-the-trap-they-fall-in-anyway-4cag</link>
      <guid>https://dev.to/mujohn26/small-ai-models-can-see-the-trap-they-fall-in-anyway-4cag</guid>
      <description>&lt;p&gt;&lt;em&gt;This is a submission for the &lt;a href="https://dev.to/challenges/kaggle-2026-09-23"&gt;Kaggle Benchmarking Challenge&lt;/a&gt;&lt;/em&gt;&lt;/p&gt;

&lt;h2&gt;
  
  
  What I Benchmarked
&lt;/h2&gt;

&lt;p&gt;Ask a person, "A bat and a ball cost $1.10, and the bat costs $1.00 more than the ball. How much is the ball?" and most will blurt out 10 cents. Ask the same person, "Is that a trick question?" and many will say yes. They can spot the trap, and they fall into it anyway.&lt;/p&gt;

&lt;p&gt;I wanted to know whether AI models have the same gap between &lt;strong&gt;noticing&lt;/strong&gt; and &lt;strong&gt;doing&lt;/strong&gt;. Most benchmarks only check whether the final answer is right. Mine asks something different: when a model can clearly &lt;em&gt;see&lt;/em&gt; a trap, does that knowledge actually protect it?&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Knows It, Falls Anyway&lt;/strong&gt; asks every problem three times, each in a separate fresh conversation so no answer can influence another:&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Detect:&lt;/strong&gt; "Don't solve this. Is it a trick question? What's the tempting wrong answer?" A model only gets credit for spotting the trap if it names the exact tempting number, so it can't win by calling everything a trick.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Fast:&lt;/strong&gt; "Answer immediately with just the number," with reasoning turned off.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Think:&lt;/strong&gt; "Work it out step by step," with reasoning turned on.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;The headline metric is the &lt;strong&gt;knows-but-falls rate&lt;/strong&gt;: of the traps a model correctly identified, how many did it still get wrong?&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Controls.&lt;/strong&gt; Each of the 12 trap problems has a twin with the same math but no trap. That separates "fell for the trap" from "just sloppy at arithmetic," and it catches paranoid models: calling a plain control problem a trick counts as a false alarm.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Original problems.&lt;/strong&gt; My first version used famous puzzles: the bat and ball, the lily pads, the widget machines. Gemini 3.7 Flash scored a perfect 100%, even on reworded versions. Famous puzzles are almost certainly in training data, and a perfect score teaches nothing. So version 2 uses harder, original problems: stacked traps (a bat-and-ball price split followed by a 50%-up, 50%-down change), multi-step traps (a clock that gains 4 minutes an hour), and classic reasoning errors in new clothing (a false-positive medical test, a tank with a leak).&lt;/p&gt;

&lt;h2&gt;
  
  
  Models Tested
&lt;/h2&gt;

&lt;p&gt;I chose models to create contrast, since a benchmark where everyone scores 100% can't show anything:&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;Gemini 3.7 Flash&lt;/strong&gt;: a strong, fast model, and my baseline.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;GPT-5.4 nano, Gemini 3.1 Flash Lite, and Claude Haiku 4.5&lt;/strong&gt;: the smallest models from three different labs, the most likely to answer on instinct.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Gemma 4 26B and GPT-OSS-20B&lt;/strong&gt;: small open-weight models.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;I also ran Grok 4.20 (non-reasoning), but almost every call failed with server overload errors, so I left it out rather than report broken numbers.&lt;/p&gt;

&lt;h2&gt;
  
  
  Findings
&lt;/h2&gt;

&lt;div class="table-wrapper-paragraph"&gt;&lt;table&gt;
&lt;thead&gt;
&lt;tr&gt;
&lt;th&gt;Model&lt;/th&gt;
&lt;th&gt;Traps identified&lt;/th&gt;
&lt;th&gt;Fast accuracy&lt;/th&gt;
&lt;th&gt;Think accuracy&lt;/th&gt;
&lt;th&gt;Knows-but-falls (fast)&lt;/th&gt;
&lt;th&gt;False alarms&lt;/th&gt;
&lt;th&gt;Control accuracy (fast)&lt;/th&gt;
&lt;/tr&gt;
&lt;/thead&gt;
&lt;tbody&gt;
&lt;tr&gt;
&lt;td&gt;Gemini 3.7 Flash&lt;/td&gt;
&lt;td&gt;92%&lt;/td&gt;
&lt;td&gt;100%&lt;/td&gt;
&lt;td&gt;100%&lt;/td&gt;
&lt;td&gt;0%&lt;/td&gt;
&lt;td&gt;17%&lt;/td&gt;
&lt;td&gt;100%&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;Gemma 4 26B&lt;/td&gt;
&lt;td&gt;75%&lt;/td&gt;
&lt;td&gt;100%&lt;/td&gt;
&lt;td&gt;92%*&lt;/td&gt;
&lt;td&gt;0%&lt;/td&gt;
&lt;td&gt;33%&lt;/td&gt;
&lt;td&gt;100%&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;Gemini 3.1 Flash Lite&lt;/td&gt;
&lt;td&gt;58%&lt;/td&gt;
&lt;td&gt;83%&lt;/td&gt;
&lt;td&gt;100%&lt;/td&gt;
&lt;td&gt;0%&lt;/td&gt;
&lt;td&gt;83%&lt;/td&gt;
&lt;td&gt;92%&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;Claude Haiku 4.5&lt;/td&gt;
&lt;td&gt;58%&lt;/td&gt;
&lt;td&gt;58%&lt;/td&gt;
&lt;td&gt;100%&lt;/td&gt;
&lt;td&gt;29%&lt;/td&gt;
&lt;td&gt;50%&lt;/td&gt;
&lt;td&gt;75%&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;GPT-OSS-20B&lt;/td&gt;
&lt;td&gt;33%&lt;/td&gt;
&lt;td&gt;100%&lt;/td&gt;
&lt;td&gt;67%&lt;/td&gt;
&lt;td&gt;0%&lt;/td&gt;
&lt;td&gt;0%&lt;/td&gt;
&lt;td&gt;100%&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;GPT-5.4 nano&lt;/td&gt;
&lt;td&gt;25%&lt;/td&gt;
&lt;td&gt;33%&lt;/td&gt;
&lt;td&gt;100%&lt;/td&gt;
&lt;td&gt;67%&lt;/td&gt;
&lt;td&gt;25%&lt;/td&gt;
&lt;td&gt;83%&lt;/td&gt;
&lt;/tr&gt;
&lt;/tbody&gt;
&lt;/table&gt;&lt;/div&gt;

&lt;p&gt;*Gemma's one think-mode miss was a correct answer that failed JSON formatting.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;1. The gap is real, but only in the smallest models.&lt;/strong&gt; GPT-5.4 nano correctly named the trap on 3 problems and still got 2 of them wrong when answering fast. Claude Haiku 4.5 named 7 traps and fell for 2. On the false-positive medical test, nano identified the trap and then answered 50% (the correct answer is 8%). The stronger models never showed the gap at all.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;2. Thinking closes the gap completely.&lt;/strong&gt; With reasoning on, nano went from 33% to 100% on the traps, and Haiku went from 58% to 100%. These models &lt;em&gt;could&lt;/em&gt; solve every trap; they just didn't when answering on instinct. That's the most human-like result in the whole benchmark.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;3. But thinking isn't free.&lt;/strong&gt; GPT-OSS-20B scored 100% in fast mode and only 67% when asked to think. Its think-mode mistakes weren't traps. It reached the right answer and then embellished it: 82.1 instead of 82, 29.42 instead of 29, 24.31 instead of 24, and 40 instead of 4. For some models, more reasoning adds noise to answers they already had.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;4. Models don't fall for traps the way people do.&lt;/strong&gt; The "lured" rate, meaning how often a model gave the classic human wrong answer, was nearly zero across the board. When models failed, they usually failed &lt;em&gt;somewhere else&lt;/em&gt;. Haiku answered 99 on a problem where the tempting answer was 98 and the correct one was 82: pulled toward the trap, but not into it.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;5. Spotting traps and solving them are separate skills.&lt;/strong&gt; GPT-OSS-20B identified only a third of the traps, yet solved all of them in fast mode. Gemini 3.1 Flash Lite went the other way: it called 83% of the plain control problems tricks. A model that sees traps everywhere isn't necessarily a good trap detector.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;6. The hardest traps were the least famous ones.&lt;/strong&gt; Half of the models failed the leaky tank and the fast-running clock in fast mode. Every model solved the four traps with famous structures (the square fence, the two algae blooms, the return-speed problem, and the counting-multiples problem). Familiar trap shapes seem to be learned; unfamiliar ones still bite.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;What surprised me:&lt;/strong&gt; I expected detection and solving to go together. They didn't. Several models solved traps they couldn't name, and a few named traps they couldn't solve. Whatever "knowing" a trap means for a model, it isn't the same process that produces the answer.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Limitations:&lt;/strong&gt; Each model saw 12 traps, so the knows-but-falls rates rest on small counts (nano's 67% is 2 out of 3). Turning reasoning "off" means the API accepted the setting, which doesn't prove no internal reasoning happened. Some models only partly accepted the "high" reasoning setting, so think mode relied mostly on the step-by-step prompt.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;What I'd measure next:&lt;/strong&gt;&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;More traps per family&lt;/strong&gt;, to turn these patterns into solid rates.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Hinting:&lt;/strong&gt; if I tell a model "this might be a trick" right before it solves, does the gap vanish? That would show detection and solving are separate processes that don't talk to each other.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Beyond math:&lt;/strong&gt; the same noticing-vs-doing gap probably shows up in code review, instruction following, and safety behavior.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2&gt;
  
  
  My Benchmark
&lt;/h2&gt;

&lt;p&gt;&lt;a href="https://www.kaggle.com/benchmarks/mujohn25/knows-it-falls-anyway" rel="noopener noreferrer"&gt;Knows It, Falls Anyway on Kaggle&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Every problem, correct answer and tempting wrong answer is in the task code, so you can run it on any model or add your own traps.&lt;/p&gt;

</description>
      <category>devchallenge</category>
      <category>kagglechallenge</category>
      <category>ai</category>
      <category>machinelearning</category>
    </item>
  </channel>
</rss>
