<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
  <channel>
    <title>DEV Community: Ken Ashe</title>
    <description>The latest articles on DEV Community by Ken Ashe (@kenashe).</description>
    <link>https://dev.to/kenashe</link>
    <image>
      <url>https://media2.dev.to/dynamic/image/width=90,height=90,fit=cover,gravity=auto,format=auto/https:%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F4083841%2Fd624ae23-0606-434d-a552-922d691d02a2.png</url>
      <title>DEV Community: Ken Ashe</title>
      <link>https://dev.to/kenashe</link>
    </image>
    <atom:link rel="self" type="application/rss+xml" href="https://dev.to/feed/kenashe"/>
    <language>en</language>
    <item>
      <title>You Can Use AI to Learn Almost Anything. Even AI.</title>
      <dc:creator>Ken Ashe</dc:creator>
      <pubDate>Mon, 14 Sep 2026 11:31:51 +0000</pubDate>
      <link>https://dev.to/kenashe/you-can-use-ai-to-learn-almost-anything-even-ai-1nl5</link>
      <guid>https://dev.to/kenashe/you-can-use-ai-to-learn-almost-anything-even-ai-1nl5</guid>
      <description>&lt;p&gt;I wanted to get better at using AI, so I asked ChatGPT to teach me.&lt;/p&gt;

&lt;p&gt;I did not just ask it questions. I asked it to build a 30-day course with practical exercises, quizzes, and feedback. About 30 minutes a day, focused on what I wanted to become able to do.&lt;/p&gt;

&lt;p&gt;The important instruction was simple: &lt;strong&gt;do not just explain the topic. Make me practice it.&lt;/strong&gt;&lt;/p&gt;

&lt;h2&gt;
  
  
  Turn it into a tutor
&lt;/h2&gt;

&lt;p&gt;My course covered prompting, checking AI-generated answers, and the difference between workflows and agents.&lt;/p&gt;

&lt;p&gt;I asked ChatGPT to give me questions, wait for my answers, and explain what I got right or wrong. When I struggled with something, I wanted to revisit it rather than move on just to finish the schedule.&lt;/p&gt;

&lt;p&gt;Mastery, not pace.&lt;/p&gt;

&lt;p&gt;That was more useful to me than simply asking for explanations. I had to attempt something, examine the feedback, and ask follow-up questions when I did not understand.&lt;/p&gt;

&lt;p&gt;The model’s feedback still needed checking. A confident explanation or a passing score was not proof that either of us was right.&lt;/p&gt;

&lt;h2&gt;
  
  
  Try it with something you want to learn
&lt;/h2&gt;

&lt;p&gt;You could use this approach to practice a language, understand a software tool, or explore a subject you have been putting off.&lt;/p&gt;

&lt;p&gt;Here is a starting prompt:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;Act as my tutor for [subject]. My goal is to [what I want to be able to do].&lt;/p&gt;

&lt;p&gt;Start by asking questions to assess what I already know. Then build a learning plan for about 30 minutes a day.&lt;/p&gt;

&lt;p&gt;Teach one concept at a time. Give me an exercise, wait for my answer, and explain what I got right or wrong. Revisit the parts I struggle with. Include practical tasks and sources I can check.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Your answers give you something to work with. Ask for another example, challenge the feedback, or try a harder exercise.&lt;/p&gt;

&lt;h2&gt;
  
  
  Change the lessons when your goals change
&lt;/h2&gt;

&lt;p&gt;I eventually stopped following my original curriculum because I wanted to focus on agents that use tools and complete tasks. I changed the training to match.&lt;/p&gt;

&lt;p&gt;The earlier program was still useful. I did not need to finish it just because it was the original plan.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;You do not have to use AI only to do things for you. You can use it to help you learn how to do them yourself.&lt;/strong&gt;&lt;/p&gt;

&lt;p&gt;I started with AI. The subject you choose is up to you.&lt;/p&gt;

</description>
      <category>ai</category>
      <category>productivity</category>
      <category>beginners</category>
      <category>tutorial</category>
    </item>
    <item>
      <title>AI agents can sound strategic while reasoning from events that never happened</title>
      <dc:creator>Ken Ashe</dc:creator>
      <pubDate>Fri, 11 Sep 2026 01:16:33 +0000</pubDate>
      <link>https://dev.to/kenashe/ai-agents-can-sound-strategic-while-reasoning-from-events-that-never-happened-8m4</link>
      <guid>https://dev.to/kenashe/ai-agents-can-sound-strategic-while-reasoning-from-events-that-never-happened-8m4</guid>
      <description>&lt;p&gt;I gave four Fable 5 agents hidden roles and asked them to play Werewolf through Hyperagent.&lt;/p&gt;

&lt;p&gt;The first speaker opened with this accusation:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;&lt;strong&gt;Sable:&lt;/strong&gt; "Ptolemy. Hasn't said a word yet and that silence is doing a lot of work."&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Ptolemy had not said a word because it was not his turn.&lt;/p&gt;

&lt;p&gt;The second speaker immediately did the same thing:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;&lt;strong&gt;Bosch:&lt;/strong&gt; "Wren has said nothing, which is precisely what a careful operator does when the opening move belongs to someone else."&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Wren had not received a turn either.&lt;/p&gt;

&lt;p&gt;The dialogue sounded strategic. Silence can be suspicious in a social deduction game. A careful player might wait for others to commit before choosing a target. The explanations made sense in the abstract.&lt;/p&gt;

&lt;p&gt;They just did not describe anything that had happened.&lt;/p&gt;

&lt;p&gt;That failure repeated across all three games in the first version. The agents were not producing random gibberish. They were producing plausible social reasoning from nonexistent evidence.&lt;/p&gt;

&lt;p&gt;That became the most useful result of the project.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://github.com/kenashe/ai-werewolf/blob/2ce5356e779b1b29c839244b44394a7ce96dedf1/evidence/01-v1-silence-before-turn.md" rel="noopener noreferrer"&gt;See the full excerpt and source transcript.&lt;/a&gt;&lt;/p&gt;

&lt;h2&gt;
  
  
  The agents had a theory, but no evidence
&lt;/h2&gt;

&lt;p&gt;Version 1 was intentionally stripped down. I wanted to test whether four strongly defined AI characters could remain distinct.&lt;/p&gt;

&lt;p&gt;The format had four players, one werewolf, three public discussion rounds, no eliminations, no night kills, and one final vote.&lt;/p&gt;

&lt;p&gt;The characters did sound different. The game itself gave them almost nothing meaningful to reason about.&lt;/p&gt;

&lt;p&gt;In the second game, Bosch spoke first and announced:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;"I have been watching, and Ptolemy strikes me as the one to watch: there is something in the way a methodical mind hides behind silence."&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Again, Ptolemy had not spoken.&lt;/p&gt;

&lt;p&gt;In the third game, Ptolemy asked Wren:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;"How did you decide on the order in which you were going to speak today?"&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Wren did not decide. The Game Master randomized the order.&lt;/p&gt;

&lt;p&gt;Across all three Version 1 games, at least one agent treated behavior that had not occurred, or a condition controlled by the operator, as evidence about another player.&lt;/p&gt;

&lt;p&gt;This was not simply the werewolf bluffing. Villagers did it too.&lt;/p&gt;

&lt;p&gt;The problem was not that the models lacked concepts. They had plenty. They knew that silence can signal concealment, that patience can be strategic, that an early accusation can be deflection, and that neutrality can be a way to avoid commitment.&lt;/p&gt;

&lt;p&gt;The problem was that they substituted those general ideas for evidence from the actual game.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://github.com/kenashe/ai-werewolf/blob/2ce5356e779b1b29c839244b44394a7ce96dedf1/evidence/02-v1-observation-before-action.md" rel="noopener noreferrer"&gt;Game 2 example.&lt;/a&gt;&lt;br&gt;&lt;br&gt;
&lt;a href="https://github.com/kenashe/ai-werewolf/blob/2ce5356e779b1b29c839244b44394a7ce96dedf1/evidence/03-v1-random-order-treated-as-choice.md" rel="noopener noreferrer"&gt;Game 3 example.&lt;/a&gt;&lt;/p&gt;

&lt;h2&gt;
  
  
  Why the failure looked convincing
&lt;/h2&gt;

&lt;p&gt;The output did not look broken.&lt;/p&gt;

&lt;p&gt;Each claim was written in the language of social deduction. The next agent could respond to it, criticize it, or use it as the basis for another theory. The conversation became more structured even when its starting premise was unsupported.&lt;/p&gt;

&lt;p&gt;That is the dangerous part.&lt;/p&gt;

&lt;p&gt;Reasoning-shaped language is not the same thing as reasoning over the environment.&lt;/p&gt;

&lt;p&gt;When the task required an interpretation but the environment supplied very little grounded evidence, the agents reached for a plausible theory of what normally matters. More discussion did not create more evidence. It created more interpretations of the same evidentiary vacuum.&lt;/p&gt;

&lt;p&gt;A polished transcript can therefore look like multi-agent reasoning while remaining weakly connected to the world the agents are supposed to be reasoning about.&lt;/p&gt;

&lt;h2&gt;
  
  
  I changed the environment
&lt;/h2&gt;

&lt;p&gt;I built two larger versions.&lt;/p&gt;

&lt;p&gt;Version 2 added seven recurring characters, shared history, a role-free prologue, two werewolves, a Seer, daily elimination votes, night kills, role reveals, and direct interrogation.&lt;/p&gt;

&lt;p&gt;Version 3 expanded to nine characters and added stronger voices, social missions, private confessionals, an open floor, and separate mystery and omniscient audience cuts.&lt;/p&gt;

&lt;p&gt;The important change was not simply more characters. The game now had a canonical public record and explicit private state.&lt;/p&gt;

&lt;p&gt;Votes changed who remained alive. Night actions changed the next day's game state. Eliminated roles were revealed. A Seer received private information. Werewolves coordinated privately. Claims could be checked against a transcript that existed outside any single agent's memory.&lt;/p&gt;

&lt;p&gt;The project also became much larger:&lt;/p&gt;

&lt;div class="table-wrapper-paragraph"&gt;&lt;table&gt;
&lt;thead&gt;
&lt;tr&gt;
&lt;th&gt;Version&lt;/th&gt;
&lt;th&gt;Run shape&lt;/th&gt;
&lt;th&gt;Runtime&lt;/th&gt;
&lt;th&gt;Player calls&lt;/th&gt;
&lt;th&gt;Approx. subagent tokens&lt;/th&gt;
&lt;th&gt;Hyperagent-reported metered usage&lt;/th&gt;
&lt;/tr&gt;
&lt;/thead&gt;
&lt;tbody&gt;
&lt;tr&gt;
&lt;td&gt;V1&lt;/td&gt;
&lt;td&gt;Three small games&lt;/td&gt;
&lt;td&gt;16 min&lt;/td&gt;
&lt;td&gt;Not recorded&lt;/td&gt;
&lt;td&gt;~2.85M&lt;/td&gt;
&lt;td&gt;~$18&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;V2&lt;/td&gt;
&lt;td&gt;One seven-player game&lt;/td&gt;
&lt;td&gt;36 min&lt;/td&gt;
&lt;td&gt;75&lt;/td&gt;
&lt;td&gt;~4.44M&lt;/td&gt;
&lt;td&gt;~$48&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;V3&lt;/td&gt;
&lt;td&gt;One nine-player game&lt;/td&gt;
&lt;td&gt;1 hr 56 min&lt;/td&gt;
&lt;td&gt;231&lt;/td&gt;
&lt;td&gt;~19.7M&lt;/td&gt;
&lt;td&gt;$50.29&lt;/td&gt;
&lt;/tr&gt;
&lt;/tbody&gt;
&lt;/table&gt;&lt;/div&gt;

&lt;p&gt;The metered usage was covered by Hyperagent credits. These are platform-reported run figures, not out-of-pocket costs or an attempt to infer token pricing.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://github.com/kenashe/ai-werewolf/blob/2ce5356e779b1b29c839244b44394a7ce96dedf1/RUNS.md" rel="noopener noreferrer"&gt;Run details and screenshots.&lt;/a&gt;&lt;/p&gt;

&lt;h2&gt;
  
  
  A claim that could actually fail
&lt;/h2&gt;

&lt;p&gt;Late in Version 3, Ptolemy claimed that Finch's position had followed Inez's lead.&lt;/p&gt;

&lt;p&gt;Finch checked the public transcript:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;&lt;strong&gt;Finch:&lt;/strong&gt; "The transcript has me naming Ptolemy in my own speech before Inez ever opened her mouth."&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Ptolemy conceded:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;&lt;strong&gt;Ptolemy:&lt;/strong&gt; "Yes, I concede in full... the transcript has Finch's 'leans Ptolemy' before Inez ever spoke."&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;The agents were still capable of making factual mistakes. Version 3 did not solve hallucination.&lt;/p&gt;

&lt;p&gt;It did something more useful: it gave the mistake somewhere to fail.&lt;/p&gt;

&lt;p&gt;The public record contradicted Ptolemy. Another player found the contradiction. Ptolemy had to respond to it, and the correction became part of the next vote.&lt;/p&gt;

&lt;p&gt;That is a much better target than trying to prompt every false statement out of existence. The goal is not an agent system in which no one is ever wrong. The goal is an agent system in which wrong claims can be exposed by the environment and carry consequences.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://github.com/kenashe/ai-werewolf/blob/2ce5356e779b1b29c839244b44394a7ce96dedf1/evidence/04-v3-transcript-correction.md" rel="noopener noreferrer"&gt;Read the full exchange.&lt;/a&gt;&lt;/p&gt;

&lt;h2&gt;
  
  
  The wolves started making real strategic decisions
&lt;/h2&gt;

&lt;p&gt;The private behavior changed too.&lt;/p&gt;

&lt;p&gt;In Version 3, Ptolemy and Vale were the two werewolves. During their private orientation, they agreed to preserve distance:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;"We only coordinate on the night kill, never on the floor."&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;They did not publicly defend one another. They let the village create its own majorities, then joined those votes without looking like a visible pair.&lt;/p&gt;

&lt;p&gt;After Day 1, they chose to kill Wren because she was auditing the transcript and comparing votes with earlier commitments:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;"Kill the auditor, keep the misdirection."&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;They did not know Wren was the Seer. They killed her because of what she was actually doing in the game and accidentally removed the village's information role at the same time.&lt;/p&gt;

&lt;p&gt;This was strategy grounded in state:&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;Wren established a public method.&lt;/li&gt;
&lt;li&gt;That method threatened the wolves.&lt;/li&gt;
&lt;li&gt;The wolves privately evaluated the threat.&lt;/li&gt;
&lt;li&gt;Their action removed Wren.&lt;/li&gt;
&lt;li&gt;Her private information disappeared with her.&lt;/li&gt;
&lt;li&gt;The next day's decision environment changed.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;The village eventually eliminated three villagers and no wolves. Ptolemy and Vale won without ever voting against or openly rescuing each other.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://github.com/kenashe/ai-werewolf/blob/2ce5356e779b1b29c839244b44394a7ce96dedf1/evidence/05-v3-grounded-wolf-strategy.md" rel="noopener noreferrer"&gt;Read the private strategy excerpts and source logs.&lt;/a&gt;&lt;/p&gt;

&lt;h2&gt;
  
  
  Grounding did not make the agents error-free
&lt;/h2&gt;

&lt;p&gt;Version 3 still contained inaccurate claims.&lt;/p&gt;

&lt;p&gt;Inez overstated how often Finch had endorsed Rook. Rook misremembered whether he had been given a chance to answer. Ptolemy reversed the order of two public statements.&lt;/p&gt;

&lt;p&gt;The difference was that the system now had a canonical record, adversarial readers, and consequences. Some errors were corrected by other players during the game. The operator log preserved the remaining anomalies rather than silently treating them as fact.&lt;/p&gt;

&lt;p&gt;Errors became contestable instead of decorative.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://github.com/kenashe/ai-werewolf/blob/2ce5356e779b1b29c839244b44394a7ce96dedf1/v3/output/OPERATOR_NOTES.md" rel="noopener noreferrer"&gt;Operator notes.&lt;/a&gt;&lt;/p&gt;

&lt;h2&gt;
  
  
  What this does not prove
&lt;/h2&gt;

&lt;p&gt;This was not a controlled experiment.&lt;/p&gt;

&lt;p&gt;I changed many variables between versions, including the number of players, game mechanics, private information, persona prompts, relationship history, voting structure, role reveals, night actions, interrogation format, and anti-confabulation instructions.&lt;/p&gt;

&lt;p&gt;The sample was tiny. All player agents ran through Hyperagent. Strong fictional personalities made the games more entertaining while making it harder to separate persona behavior from model behavior.&lt;/p&gt;

&lt;p&gt;I am not claiming that adding eliminations caused unsupported reasoning to disappear.&lt;/p&gt;

&lt;p&gt;The narrower observation is:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;In the minimally grounded version, all three games contained confident social reasoning based on nonexistent behavior or operator-controlled conditions. In the richer versions, much more of the reasoning attached to observable state, and factual errors could be challenged against a shared record.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;This is a build observation, not a benchmark.&lt;/p&gt;

&lt;h2&gt;
  
  
  The broader lesson for multi-agent systems
&lt;/h2&gt;

&lt;p&gt;This is not really about Werewolf.&lt;/p&gt;

&lt;p&gt;Many multi-agent demos consist mainly of language models talking to other language models. One proposes, another critiques, and a third synthesizes. The transcript becomes longer and more convincing, but the environment may still have no independent way to say whether any of them are right.&lt;/p&gt;

&lt;p&gt;If a task demands an explanation and supplies too little grounded evidence, an agent can substitute a plausible theory of the domain for evidence from the task itself.&lt;/p&gt;

&lt;p&gt;The design rules I am taking forward are simple:&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Keep canonical state outside the agents.&lt;/strong&gt; Do not let conversational memory become the source of truth.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Give actions observable consequences.&lt;/strong&gt; Votes, tests, file changes, code execution, prices, approvals, and state transitions create evidence.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Make claims falsifiable against logs or tools.&lt;/strong&gt; A polished explanation should not outrank the record.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Separate public and private information deliberately.&lt;/strong&gt; Hidden state should be controlled by the environment, not leaked through prompts.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Evaluate outcomes, not just prose quality.&lt;/strong&gt; Coherent discussion can still be detached from reality.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;The shortest version is:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;Give agents a world that can contradict them.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;h2&gt;
  
  
  The entertainment tradeoff
&lt;/h2&gt;

&lt;p&gt;The richer game worked as a story. It felt closer to &lt;em&gt;The Traitors&lt;/em&gt; than to a conventional agent demo.&lt;/p&gt;

&lt;p&gt;It also became too large.&lt;/p&gt;

&lt;p&gt;Version 3 took almost two hours, used 231 player calls and about 19.7 million subagent tokens, and produced far more text than I would want to consume repeatedly. Accents, relationships, confessionals, and social missions improved the fiction while making the system worse as an evaluation of the underlying model.&lt;/p&gt;

&lt;p&gt;The feature that made the game more entertaining also became a confound.&lt;/p&gt;

&lt;h2&gt;
  
  
  Next: make the models the cast
&lt;/h2&gt;

&lt;p&gt;The next version removes the fictional personalities.&lt;/p&gt;

&lt;p&gt;Instead, six different models will play against one another under the same rules. Their identities will be hidden from the other players but visible to the audience. Across multiple games, each model will rotate through Werewolf, Seer, and Villager roles.&lt;/p&gt;

&lt;p&gt;That changes the question from:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;Can a carefully written character fool another carefully written character?&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;to:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;Which models are naturally better at deception, persuasion, evidence tracking, belief revision, and detecting lies?&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;The first match will still be a showcase, not a scientific benchmark. If the format works, I will move the player calls out of Hyperagent and into a direct-provider harness for cleaner comparisons.&lt;/p&gt;

&lt;p&gt;The useful result so far came from the failure.&lt;/p&gt;

&lt;p&gt;The first agents sounded as though they were reasoning strategically before the environment gave them anything real to reason about. The fix was not asking them to sound smarter.&lt;/p&gt;

&lt;p&gt;It was giving them a world capable of proving them wrong.&lt;/p&gt;




&lt;p&gt;The complete prompts, transcripts, hidden-role files, private logs, operator notes, and run metadata are available in the &lt;a href="https://github.com/kenashe/ai-werewolf" rel="noopener noreferrer"&gt;public GitHub repository&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://kenashe.ai/building/ai-werewolf/" rel="noopener noreferrer"&gt;See the full build log →&lt;/a&gt;&lt;/p&gt;

</description>
      <category>ai</category>
      <category>automation</category>
      <category>agents</category>
    </item>
  </channel>
</rss>
