<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
  <channel>
    <title>DEV Community: connerlambden</title>
    <description>The latest articles on DEV Community by connerlambden (@connerlambden).</description>
    <link>https://dev.to/connerlambden</link>
    <image>
      <url>https://media2.dev.to/dynamic/image/width=90,height=90,fit=cover,gravity=auto,format=auto/https:%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F3776793%2F2b4714ed-0e1b-407a-a636-4236e644b524.jpeg</url>
      <title>DEV Community: connerlambden</title>
      <link>https://dev.to/connerlambden</link>
    </image>
    <atom:link rel="self" type="application/rss+xml" href="https://dev.to/feed/connerlambden"/>
    <language>en</language>
    <item>
      <title>The GMAT redesign: three sections, one stubborn question</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Mon, 24 Aug 2026 20:31:52 +0000</pubDate>
      <link>https://dev.to/connerlambden/the-gmat-redesign-three-sections-one-stubborn-question-52hh</link>
      <guid>https://dev.to/connerlambden/the-gmat-redesign-three-sections-one-stubborn-question-52hh</guid>
      <description>&lt;p&gt;The GMAT does not look like it did in 2023. The familiar 200-800 scale is gone, replaced by 205-805. The Analytical Writing section now stands alone and unscored. Data Sufficiency, the section everyone dreaded, is folded into the same scored pool as Problem Solving. The exam is about an hour shorter.&lt;/p&gt;

&lt;p&gt;And the two questions people actually care about are the same two they asked ten years ago. Is the score coachable? And what does the number really predict?&lt;/p&gt;

&lt;h2&gt;
  
  
  What changed
&lt;/h2&gt;

&lt;p&gt;The current GMAT sits on a single scored section: Quantitative Reasoning, Verbal Reasoning, and Data Insights. Three sections, scored 60-90 each, combined into a total of 205-805.&lt;/p&gt;

&lt;p&gt;Data Insights is the section with no direct ancestor. It combines data sufficiency, table and graph analysis, and reasoning from numbers. It is the most reasoning-heavy thing the GMAT has ever put on a scored section, because it rewards speed at reading a table and deciding what is sufficient, not speed at arithmetic.&lt;/p&gt;

&lt;h2&gt;
  
  
  What the score predicts
&lt;/h2&gt;

&lt;p&gt;This is the part the ads do not want to linger on: the GMAT is a serious predictor. Study after study finds higher GMAT scores correlate with stronger performance in the first year of an MBA program, and MBA programs lean on the score in admissions because that relationship is real.&lt;/p&gt;

&lt;p&gt;That is the thing to hold onto when someone tells you prep is a scam. The score measures something. It is not measuring your worth; it is measuring a specific blend of reasoning speed, composure under time, and judgment about what you can conclude from data. All three are learnable, and all three improve with deliberate practice.&lt;/p&gt;

&lt;h2&gt;
  
  
  What prep actually buys you
&lt;/h2&gt;

&lt;p&gt;Here is where most marketing overshoots.&lt;/p&gt;

&lt;p&gt;The dramatic score jumps in the ads are real but cherry-picked. The people in them are usually retakers who had a bad first sit, or people who switched to the new content late. Controlled gains from real practice are solid, modest, and they cluster in the same narrow window: 20 to 50 points for most people who work steadily. That is a real edge. It moves the needle for admissions odds at most schools, and it is not the transformation the billboards sell.&lt;/p&gt;

&lt;p&gt;The faster, more durable win is exactly the one prep marketing has the least incentive to advertise: practice the reasoning, not the content. Data Insights rewards reading a table and deciding with sparse evidence. Verbal rewards telling an assumption from a fact. Both are trainable judgments, and both respond to volume with feedback far more than to another lecture.&lt;/p&gt;

&lt;h2&gt;
  
  
  The trap
&lt;/h2&gt;

&lt;p&gt;The most expensive mistake a 2026 GMAT candidate can make is studying the old test. Huge amounts of online material are written for the pre-2024 exam, and most of it is now wrong about the format. If your study plan still has a dedicated "Data Sufficiency" block because a tutor you like said so, you are spending time on a configuration the test does not run.&lt;/p&gt;

&lt;p&gt;The sharpest prep move is to ground everything in the current exam, use only material written for the new format, and point your practice at Data Insights and Verbal — the two sections easiest to improve and toughest to fake.&lt;/p&gt;

&lt;h2&gt;
  
  
  What actually helps
&lt;/h2&gt;

&lt;p&gt;Keep it short, because it is short.&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;Practice timed, every session.&lt;/strong&gt; The GMAT is a speed test wearing a reasoning disguise.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Review the miss harder than the hit.&lt;/strong&gt; The score-feedback loop is the whole game.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Drill Data Insights daily.&lt;/strong&gt; It is new, it is reasoning-heavy, and it is the easiest place to pick up points nobody is defending.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2&gt;
  
  
  The through-line
&lt;/h2&gt;

&lt;p&gt;The redesign did not make the GMAT more coachable. It made the test more honest: three sections, less content to hoard, more reasoning to train. That is good news for a reasoning practice that does not promise to make you smarter, only sharper at the specific judgment these sections actually measure.&lt;/p&gt;

&lt;p&gt;If you want a free daily reasoning warm-up tuned to the exact judgment the GMAT and GRE verbal sections reward, the &lt;a href="https://intelligencemax.ai/test-prep/gmat" rel="noopener noreferrer"&gt;GMAT guide on IntelligenceMax&lt;/a&gt; digs into what the redesign means and where the coaching hype overshoots.&lt;/p&gt;




&lt;p&gt;Disclosure: I founded IntelligenceMax, an adaptive reasoning practice. If any of this reads soft, say so.&lt;/p&gt;

</description>
      <category>discuss</category>
      <category>productivity</category>
    </item>
    <item>
      <title>Is the LSAT coachable? What the 2024 redesign settled</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Mon, 24 Aug 2026 19:14:05 +0000</pubDate>
      <link>https://dev.to/connerlambden/is-the-lsat-coachable-what-the-2024-redesign-settled-2n37</link>
      <guid>https://dev.to/connerlambden/is-the-lsat-coachable-what-the-2024-redesign-settled-2n37</guid>
      <description>&lt;p&gt;In August 2024 the LSAT dropped Logic Games. Studiers lost a whole section of the test overnight, their courses lost a quarter of their material, and the marketing copy went quiet for about a week before pretending it had always been this way.&lt;/p&gt;

&lt;p&gt;The change is worth sitting with, because it quietly settled a claim people have argued about for years: is the LSAT a trainable skill, or a raw-aptitude measure you either have or don't?&lt;/p&gt;

&lt;h2&gt;
  
  
  What actually changed
&lt;/h2&gt;

&lt;p&gt;The LSAT is now two Logical Reasoning sections, one Reading Comprehension section, and one unscored experimental section. Still scored 120 to 180. The writing sample is separate and unscored.&lt;/p&gt;

&lt;p&gt;That one structural edit moves a lot of weight. Logical Reasoning is now two-thirds of the scored test. The section everyone treats as "LR is where the score lives" is now most of the exam by definition.&lt;/p&gt;

&lt;h2&gt;
  
  
  Coachable, but not in the way the ads mean
&lt;/h2&gt;

&lt;p&gt;Here is the honest position, and it is a more interesting one than either side of the old argument gives you.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;The LSAT is the most coachable of the big admissions tests.&lt;/strong&gt; Not because a magic method exists, but because what it tests is unusually specific. Assumption, flaw, strengthen, weaken, inference: these are learnable moves, not reflexes you are born with. Get good at naming the job of a question stem on sight and you buy back real speed. That is real.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;But "coachable" does not mean "15 points in a month."&lt;/strong&gt; The dramatic scored jumps you see advertised are selected samples, shown because they are dramatic. Controlled gains from real practice are solid, but modest and clustered: they show up as a handful of scaled points and, more durably, in how you read under time. Call it a real edge, not a transformation.&lt;/p&gt;

&lt;h2&gt;
  
  
  The trap that costs you time
&lt;/h2&gt;

&lt;p&gt;The most expensive mistake a 2026 LSAT studier can make is studying the old test. Huge amounts of reputable material are still geared toward Logic Games that no longer exists on it. If your schedule still has a Logic Games block, you are rehearsing a test that stopped being real in 2024.&lt;/p&gt;

&lt;p&gt;The sharpest reprioritization you can do is to point almost all of your reasoning practice at the two Logical Reasoning sections, because that is where two-thirds of your score now comes from.&lt;/p&gt;

&lt;h2&gt;
  
  
  What actually helps
&lt;/h2&gt;

&lt;p&gt;This part is not complicated, and it is worth saying plainly:&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;Know your question stems cold.&lt;/strong&gt; Learn to recognize the job before reading the stimulus, so you read the argument with a purpose instead of hoping.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Read comparatively.&lt;/strong&gt; The comparative passage set rewards tracking who argues what. Practice on real dense prose, not warmed-over summaries.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Drill timed, review the miss.&lt;/strong&gt; The feedback loop is the highest-yield activity there is. Get one wrong and understand exactly why.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Retake with a plan.&lt;/strong&gt; A focused repeat beats passive repetition.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2&gt;
  
  
  The through-line
&lt;/h2&gt;

&lt;p&gt;The design of the new LSAT matches what the evidence has said all along: you get better at the reasoning you actually practice, in the form you practice it. You do not get smarter in general; you get sharper at this specific, high-value judgment.&lt;/p&gt;

&lt;p&gt;That is the whole honest frame. And it is why, if you want a daily reasoning warm-up that chases the exact distinctions the LSAT cares about, there is one with no score promise: &lt;a href="https://intelligencemax.ai/test-prep/lsat" rel="noopener noreferrer"&gt;IntelligenceMax's LSAT guide&lt;/a&gt;, which breaks down the current test format, what prep buys you, and where the hype overshoots.&lt;/p&gt;




&lt;p&gt;Disclosure: I founded IntelligenceMax, an adaptive reasoning practice. If any of this reads soft, say so.&lt;/p&gt;

</description>
      <category>discuss</category>
      <category>productivity</category>
    </item>
    <item>
      <title>Does learning formal logic make you better at reasoning? What the evidence actually says</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Mon, 24 Aug 2026 17:31:56 +0000</pubDate>
      <link>https://dev.to/connerlambden/does-learning-formal-logic-make-you-better-at-reasoning-what-the-evidence-actually-says-3p3i</link>
      <guid>https://dev.to/connerlambden/does-learning-formal-logic-make-you-better-at-reasoning-what-the-evidence-actually-says-3p3i</guid>
      <description>&lt;p&gt;Will studying logic make you better at reasoning? is one of those questions that sounds naive until you look at how long people have argued about it. In 1917, Edward Thorndike tried to settle it with his doctrine of identical elements: training transfers only to tasks that share parts with what you practiced. A century of replication attempts later, the answer has not gotten more flattering for the optimists.&lt;/p&gt;

&lt;p&gt;Here is the honest state of the evidence, written for a developer or self-learner who wants to reason better and distrusts both brain-training hype and nothing-transfers cynicism.&lt;/p&gt;

&lt;h2&gt;
  
  
  What does not hold: far transfer to general intelligence
&lt;/h2&gt;

&lt;p&gt;Meta-analyses of cognitive training find consistent near-transfer gains. Broad far-transfer to general intelligence is weak or null once you control for placebo and reporting bias:&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;Simons et al. (2016), Do Brain-Training Programs Work?, found limited evidence of far transfer beyond the trained task.&lt;/li&gt;
&lt;li&gt;Melby-Lervaag, Redick and Hulme (2016) found near-zero transfer to general ability for working memory training.&lt;/li&gt;
&lt;li&gt;Sala and Gobet (2017) found practice effects confined to the trained task across many training types.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Translation: studying formal logic can make you better at formal logic, written argument evaluation, and certain math. There is no credible lane from a month of truth tables to getting broadly smarter.&lt;/p&gt;

&lt;h2&gt;
  
  
  What does hold: near transfer and honest practice
&lt;/h2&gt;

&lt;p&gt;The transfer literature is not a wall. It is a corridor:&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;Skills transfer along shared structures, not shared surfaces. Train the exact inference you want and you get better at it. This is why LSAT drilling works for LSAT logical reasoning, and algebra drills help a math exam.&lt;/li&gt;
&lt;li&gt;Practice with feedback beats passive exposure. You learn to evaluate arguments by evaluating hundreds of them and seeing why you were wrong, not by reading textbook definitions of validity.&lt;/li&gt;
&lt;li&gt;Knowledge compounds. The best reasoners know a great deal about the thing they are reasoning about, plus they have practiced a small set of moves: premises, support, warrants, and a sharp eye for the two choices in a dilemma.&lt;/li&gt;
&lt;/ol&gt;

&lt;h2&gt;
  
  
  The practical upshot for a builder
&lt;/h2&gt;

&lt;p&gt;This evidence is why IntelligenceMax exists. It is a reasoning practice that does not claim to raise your IQ, and says so in the product. It drills the narrowest useful judgment of all: what is actually supported, with what strength, by this premise? delivered adaptively, scored honestly, with feedback on every answer.&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;Questions are generated live so the gym never goes stale, and you always see fresh items at your current boundary.&lt;/li&gt;
&lt;li&gt;The score is an item-response-style estimate of current ability on that task, with an honest uncertainty band, not an IQ certificate.&lt;/li&gt;
&lt;li&gt;The design assumption is near transfer: train the judgment you want, in the form you want it in.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2&gt;
  
  
  The short answer to the title question
&lt;/h2&gt;

&lt;p&gt;Yes, but only in ways that survive scrutiny:&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;It improves reasoning on tasks that share structure with practice.&lt;/li&gt;
&lt;li&gt;It does not reliably raise a score on a reasoning measure far from the practice.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;Anyone selling you claim two without claim one is either describing older, better-trodden practice (law, math, engineering, code) or overselling your future.&lt;/p&gt;

&lt;p&gt;See how we score reasoning and where we draw the claim boundaries on the science page at intelligencemax.ai. And if this reads soft anywhere, say so. I am trying to train exactly the judgment this piece exercises.&lt;/p&gt;

&lt;p&gt;If you want to see how one product translates this into an adaptive practice with honest scoring and no IQ claims: &lt;a href="https://intelligencemax.ai/science" rel="noopener noreferrer"&gt;intelligencemax.ai/science&lt;/a&gt;&lt;/p&gt;

</description>
      <category>discuss</category>
      <category>learning</category>
      <category>webdev</category>
    </item>
    <item>
      <title>Honest edtech claims: what reasoning practice can and cannot do</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Mon, 24 Aug 2026 15:22:21 +0000</pubDate>
      <link>https://dev.to/connerlambden/honest-edtech-claims-what-reasoning-practice-can-and-cannot-do-39i0</link>
      <guid>https://dev.to/connerlambden/honest-edtech-claims-what-reasoning-practice-can-and-cannot-do-39i0</guid>
      <description>&lt;p&gt;Edtech fails in two directions. It overclaims, selling a few games as a way to boost intelligence and reshape a life. Or it underclaims, staying so vague that it says nothing at all. Both are dishonest, and both erode trust in the whole category.&lt;/p&gt;

&lt;h2&gt;
  
  
  The transfer evidence is clear
&lt;/h2&gt;

&lt;p&gt;The training-transfer literature is settled enough to build on. Melby-Lervag and Hulme (2013) meta-analyzed working-memory training and found reliable gains on trained tasks but little transfer to general cognition. Sala and Gobet (2017) reviewed reasoning training and reached the same conclusion. A 2022 Translational Psychiatry trial gave middle-aged adults eight weeks of adaptive n-back training and found strong practice effects with no near or far transfer, at the cognitive level or in brain structure.&lt;/p&gt;

&lt;p&gt;The pattern across all of them: you improve at what you train, and sometimes at things closely related. You do not measure broadly. That is not a product failure. It is what the evidence actually shows.&lt;/p&gt;

&lt;h2&gt;
  
  
  What honest edtech looks like
&lt;/h2&gt;

&lt;p&gt;An honest reasoning product claims a narrow edge. It trains a specific skill, adapts to your level, and scores you the way a psychometric instrument does. It does not claim to raise your IQ, fix your career, or make you a genius.&lt;/p&gt;

&lt;p&gt;At &lt;a href="https://intelligencemax.ai" rel="noopener noreferrer"&gt;IntelligenceMax&lt;/a&gt; I built practice scored with item response theory. The model estimates a latent ability behind your answers and calibrates difficulty to it. That is not magic. It is the same math behind standardized tests. What it buys is honesty: we can score practice fairly and measure the right thing, without promising to transform your mind.&lt;/p&gt;

&lt;h2&gt;
  
  
  Honest claims are the strategy
&lt;/h2&gt;

&lt;p&gt;People do not trust brain training. The instant a product promises a twenty-point IQ jump, everyone discounts everything else it says. Trust survives when the claims do. We prefer narrow ones.&lt;/p&gt;

&lt;p&gt;If reasoning practice has a future, it is the narrow kind: getting better at telling what is supported from what is merely plausible, reading closely, spotting hidden assumptions. That is useful on its own. It is also the only claim the research supports.&lt;/p&gt;

</description>
      <category>cognition</category>
      <category>psychometrics</category>
    </item>
    <item>
      <title>Why we score reasoning practice with Item Response Theory (and what it does not claim)</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Sat, 22 Aug 2026 16:04:33 +0000</pubDate>
      <link>https://dev.to/connerlambden/why-we-score-reasoning-practice-with-item-response-theory-and-what-it-does-not-claim-2ann</link>
      <guid>https://dev.to/connerlambden/why-we-score-reasoning-practice-with-item-response-theory-and-what-it-does-not-claim-2ann</guid>
      <description>&lt;p&gt;When I started building &lt;a href="https://intelligencemax.ai" rel="noopener noreferrer"&gt;IntelligenceMax&lt;/a&gt;, I faced a question that every cognitive training product has to answer: how do you measure whether someone is getting better at reasoning?&lt;/p&gt;

&lt;p&gt;The obvious answer is "track their accuracy." The obvious answer is wrong.&lt;/p&gt;

&lt;h2&gt;
  
  
  The problem with raw accuracy
&lt;/h2&gt;

&lt;p&gt;If a learner gets 8 out of 10 questions right on day 1 and 9 out of 10 on day 10, did they improve? Maybe. But you cannot know without knowing the difficulty of those questions. Easy questions inflate accuracy. Hard questions depress it. Raw accuracy confounds ability with the difficulty of the specific questions you happened to ask.&lt;/p&gt;

&lt;p&gt;This is why standardized tests do not report raw accuracy. They report scaled scores estimated from the difficulty of the items answered. The SAT does not tell you "you got 47/58." It tells you "your score is 720," where 720 is an estimate of your ability that accounts for which questions you got right and wrong.&lt;/p&gt;

&lt;h2&gt;
  
  
  What IRT actually does
&lt;/h2&gt;

&lt;p&gt;Item Response Theory (IRT) models the probability that a person with a given ability level answers a given item correctly. The simplest model, the 1PL or Rasch model, uses one parameter per item: difficulty. The probability of a correct answer is a logistic function of the difference between the person's ability and the item's difficulty.&lt;/p&gt;

&lt;p&gt;$$P(\text{correct}) = \frac{1}{1 + e^{-(\theta - b)}}$$&lt;/p&gt;

&lt;p&gt;Where θ is the person's ability and b is the item's difficulty. Higher ability relative to difficulty means higher probability of a correct response.&lt;/p&gt;

&lt;p&gt;The 2PL model adds a discrimination parameter. The 3PL model adds a guessing parameter. For our use case (LLM-generated MCQs where guessing is not the primary concern), we use a 2PL model.&lt;/p&gt;

&lt;p&gt;The benefit: ability estimates are comparable across different sets of questions. If you answer 7 hard questions correctly and I answer 9 easy questions correctly, IRT can tell you your ability is higher than mine, even though my raw accuracy is higher.&lt;/p&gt;

&lt;h2&gt;
  
  
  Adaptive difficulty
&lt;/h2&gt;

&lt;p&gt;IRT also enables adaptive testing. Once you have an ability estimate, you can select the next question to maximize information about that person's ability. For a 2PL model, the most informative question for a person at ability θ is one where the difficulty b is close to θ.&lt;/p&gt;

&lt;p&gt;This is what CAT (Computerized Adaptive Testing) does. The GRE and GMAT use it. We use a simplified version: generate questions at a target difficulty based on the current ability estimate, and adjust after each response.&lt;/p&gt;

&lt;p&gt;The benefit for the learner: every question is at the edge of their ability. Not too easy (boring), not too hard (frustrating). This is the zone of proximal development, operationalized psychometrically.&lt;/p&gt;

&lt;h2&gt;
  
  
  What IRT does not do
&lt;/h2&gt;

&lt;p&gt;Here is where I want to be careful, because the cognitive training industry has a history of overclaiming.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;IRT does not measure intelligence.&lt;/strong&gt; It measures performance on a specific set of items. If those items are reasoning MCQs, IRT estimates your ability at answering reasoning MCQs. It does not estimate your general intelligence (g). Ability on a specific reasoning task and general intelligence are correlated but not identical.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;IRT does not prove transfer.&lt;/strong&gt; Getting better at our reasoning MCQs, even with adaptive difficulty and accurate scoring, does not mean you have improved at other reasoning tasks. The evidence for far-transfer from cognitive training to general cognition is weak (Melby-Lervåg &amp;amp; Hulme, 2013; Sala &amp;amp; Gobet, 2017). We do not claim our platform improves general cognitive ability. We claim it provides adaptive practice at a specific skill, scored with a method that accounts for item difficulty.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;IRT does not make claims about IQ.&lt;/strong&gt; IQ is a specific construct measured by specific standardized tests. Our ability estimate is not an IQ score. It is a performance estimate on our specific item pool. Conflating the two is what brain training companies do, and it is why the field has a credibility problem.&lt;/p&gt;

&lt;h2&gt;
  
  
  Why we chose IRT anyway
&lt;/h2&gt;

&lt;p&gt;Because it is the honest way to measure what we are doing. We are providing practice at a specific reasoning skill (distinguishing supported arguments from plausible ones). To know whether that practice is working, we need to measure ability changes over time. Raw accuracy cannot do that because question difficulty varies. IRT can, because it separates ability from difficulty.&lt;/p&gt;

&lt;p&gt;The alternative is to not measure at all, or to report raw accuracy and pretend it means something. Neither is honest.&lt;/p&gt;

&lt;h2&gt;
  
  
  The engineering trade-off
&lt;/h2&gt;

&lt;p&gt;IRT requires item parameter estimation. You need enough responses per item to estimate difficulty and discrimination. For a platform generating fresh questions with LLMs, this is a challenge: each question is new, so there are no historical parameters.&lt;/p&gt;

&lt;p&gt;Our approach: we generate questions at a target difficulty (determined by prompt engineering), then calibrate item parameters from response data as it accumulates. Early on, we rely on prompt-specified difficulty. Over time, the data corrects.&lt;/p&gt;

&lt;p&gt;This is a known trade-off in adaptive testing with generated items. The alternative is to use a static item bank with pre-calibrated parameters, but that sacrifices freshness and variety. We chose variety over parameter precision, and we are transparent about that.&lt;/p&gt;

&lt;h2&gt;
  
  
  The bottom line
&lt;/h2&gt;

&lt;p&gt;IRT-based scoring is a tool for honest measurement of a specific skill. It does not make IntelligenceMax a brain training app, an IQ test, or a general intelligence improver. It makes it a platform that practices one reasoning skill, measures performance on that skill accounting for item difficulty, and adapts difficulty to the learner.&lt;/p&gt;

&lt;p&gt;That is a narrow claim. I think narrow claims are the honest ones.&lt;/p&gt;

</description>
      <category>cognition</category>
      <category>psychometrics</category>
      <category>education</category>
    </item>
    <item>
      <title>The LSAT-to-1L reasoning carryover is real, but narrower than people think</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Sat, 22 Aug 2026 02:51:41 +0000</pubDate>
      <link>https://dev.to/connerlambden/the-lsat-to-1l-reasoning-carryover-is-real-but-narrower-than-people-think-1ii4</link>
      <guid>https://dev.to/connerlambden/the-lsat-to-1l-reasoning-carryover-is-real-but-narrower-than-people-think-1ii4</guid>
      <description>&lt;p&gt;I built an adaptive reasoning practice app, so I spend a lot of time thinking about what transfers from one reasoning-heavy task to another. The LSAT-to-1L question is the one I get most often, and it's the one where the popular answer is most wrong.&lt;/p&gt;

&lt;p&gt;The popular claim: LSAT prep makes you better at law school because law is just logic.&lt;/p&gt;

&lt;p&gt;The honest claim: LSAT Logical Reasoning trains a specific, narrow skill (spotting the distinction between an answer that's supported and one that's merely plausible), and that skill does carry into 1L exam writing, but only as one component among several. It does not replace learning doctrine, and it does not make law school easy.&lt;/p&gt;

&lt;h2&gt;
  
  
  What actually transfers
&lt;/h2&gt;

&lt;p&gt;The LSAT LR section is built around a core move: read a short argument, identify its gap, and then pick the answer that either exploits or patches that gap. The wrong answers are designed to be tempting, which means they're plausible-sounding but unsupported. The right answer is the one that does the work the question asks.&lt;/p&gt;

&lt;p&gt;That distinction, supported vs. plausible, is the same move that separates good 1L exam answers from average ones. On a law school exam, the professor gives you a fact pattern and asks you to identify the claims, defenses, and counterarguments. The students who do well are the ones who can look at a set of facts and figure out which rule is actually triggered, not which rule looks like it might be. They can distinguish a case where negligence applies from one where it looks like it applies but doesn't.&lt;/p&gt;

&lt;p&gt;That's the transfer. It's real. It's the reason LSAT score correlates with 1L GPA more than undergraduate GPA does. But it's one component.&lt;/p&gt;

&lt;h2&gt;
  
  
  What doesn't transfer
&lt;/h2&gt;

&lt;p&gt;The LSAT doesn't teach you doctrine. It doesn't teach you the elements of negligence, the rule against perpetuities, or the Erie doctrine. A 175 LSAT scorer who hasn't learned the material will still fail the exam. The reasoning skill is the scaffold, but the doctrine is the building.&lt;/p&gt;

&lt;p&gt;The LSAT also doesn't teach you the exam-writing format law professors reward. IRAC isn't hard, but it's specific, and it's not intuitive if you've never seen it. A brilliant reasoner who writes a formless exam answer will score lower than a mediocre reasoner who structures it well. The LSAT trains reasoning, not legal writing.&lt;/p&gt;

&lt;p&gt;And the LSAT doesn't teach you to read cases the way law school requires. Reading a case for the holding is a different skill from reading an LR stimulus for the flaw. Both are close reading, but the objects are different.&lt;/p&gt;

&lt;h2&gt;
  
  
  The narrow version of the claim
&lt;/h2&gt;

&lt;p&gt;The honest version of the LSAT-to-1L carryover is this: if you trained the distinction-finding muscle for the LSAT, you'll find it easier to do the distinction-finding part of 1L exams. That's it. It's a head start on one skill, not a substitute for the rest of law school.&lt;/p&gt;

&lt;p&gt;If you're a 0L and you're asking whether LSAT prep was wasted if you don't go to law school: no. The same distinction (supported vs. plausible) shows up in the GRE critical reasoning questions, in the GMAT Critical Reasoning section, and in the MCAT CARS section. It's a general reasoning skill, not a law-specific one. That's why standardized tests across fields all include some version of it.&lt;/p&gt;

&lt;h2&gt;
  
  
  The practical takeaway
&lt;/h2&gt;

&lt;p&gt;If you're heading to law school and you already did LSAT prep, the thing to do is not to rest on it. The reasoning head start is real, but it decays if you don't use it, and it's only one part of the exam. Spend your 1L energy on doctrine and exam structure. The reasoning will be there when you need it.&lt;/p&gt;

&lt;p&gt;If you're a 0L who hasn't taken the LSAT yet, the highest-value thing you can drill is the LR section, because the distinction-finding skill compounds across the whole test and then carries forward. Reading comp matters, but it rewards volume. LR rewards learning the structural move, and that move transfers.&lt;/p&gt;




&lt;p&gt;&lt;em&gt;I'm Conner. I build &lt;a href="https://intelligencemax.ai" rel="noopener noreferrer"&gt;IntelligenceMax&lt;/a&gt;, an adaptive reasoning practice app that generates LSAT-style and GRE-style questions from frontier LLMs and scores them with item response theory. It's not a test-prep replacement and it won't raise your IQ; it's a reasoning gym for the distinction-finding muscle I'm talking about above.&lt;/em&gt;&lt;/p&gt;

</description>
    </item>
    <item>
      <title>How to improve your IQ (without lying to yourself)</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Mon, 20 Jul 2026 16:54:42 +0000</pubDate>
      <link>https://dev.to/connerlambden/how-to-improve-your-iq-without-lying-to-yourself-49da</link>
      <guid>https://dev.to/connerlambden/how-to-improve-your-iq-without-lying-to-yourself-49da</guid>
      <description>&lt;p&gt;Search demand for "how to improve your IQ" and "how to improve your intelligence" usually wants one of three different things, and sellers often blur them.&lt;/p&gt;

&lt;h2&gt;
  
  
  Three claims people mix up
&lt;/h2&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Test score movement.&lt;/strong&gt; Practice effects, item familiarity, sleep, and anxiety can move an IQ score without proving a durable jump in general intelligence (g).&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Competence growth.&lt;/strong&gt; Crystallized skill and domain expertise keep rising for years with hard study. That is real improvement even when Raven-like fluid scores barely budge.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Far transfer from brain games to broad Gf.&lt;/strong&gt; Near transfer to trained tasks is common. Large lasting Gf gains are not the default finding once active controls enter the picture.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;If an article promises a free upgrade to general intelligence from puzzles alone, shrink the claim until it matches the evidence.&lt;/p&gt;

&lt;h2&gt;
  
  
  A boring adult plan that survives contact with research
&lt;/h2&gt;

&lt;ul&gt;
&lt;li&gt;Sleep and aerobic exercise first.&lt;/li&gt;
&lt;li&gt;Deliberate practice in a hard domain (math, language, code, music).&lt;/li&gt;
&lt;li&gt;Optional adaptive reasoning practice if you enjoy it.&lt;/li&gt;
&lt;li&gt;Treat any practice estimate as training feedback, not a clinical IQ diagnosis.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2&gt;
  
  
  Sources I maintain
&lt;/h2&gt;

&lt;ul&gt;
&lt;li&gt;Claim-size map: &lt;a href="https://intelligencemax.ai/guide" rel="noopener noreferrer"&gt;intelligencemax.ai/guide&lt;/a&gt;
&lt;/li&gt;
&lt;li&gt;Evidence pack: &lt;a href="https://osf.io/kja9b/" rel="noopener noreferrer"&gt;osf.io/kja9b&lt;/a&gt;
&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Disclosure: I build IntelligenceMax, an adaptive reasoning gym. Practice software is not a promise that apps raise trait IQ.&lt;/p&gt;

</description>
      <category>psychology</category>
      <category>learning</category>
      <category>evidence</category>
    </item>
    <item>
      <title>Lumosity, Elevate, Peak: same aisle, different drills</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Thu, 16 Jul 2026 21:06:32 +0000</pubDate>
      <link>https://dev.to/connerlambden/lumosity-elevate-peak-same-aisle-different-drills-2e04</link>
      <guid>https://dev.to/connerlambden/lumosity-elevate-peak-same-aisle-different-drills-2e04</guid>
      <description>&lt;p&gt;Search demand lumps Lumosity, Elevate, Peak, BrainHQ, and NeuroNation into one bucket: "brain training apps." Product design does not.&lt;/p&gt;

&lt;h2&gt;
  
  
  Three different product jobs
&lt;/h2&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Language / productivity drills&lt;/strong&gt; (Elevate-shaped): reading, writing, listening, math fluency. Progress charts track those drills.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Cognitive mini-games&lt;/strong&gt; (Lumosity / Peak-shaped): memory, attention, flexibility, speed under time pressure.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Reasoning practice&lt;/strong&gt; (what I ship): adaptive multiple-choice distinctions at your edge, with an honest practice score that is &lt;em&gt;not&lt;/em&gt; a clinical IQ claim.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;If you evaluate all three with the same KPI ("raises IQ"), you will misread the literature and mis-buy the subscription.&lt;/p&gt;

&lt;h2&gt;
  
  
  What the evidence usually supports
&lt;/h2&gt;

&lt;p&gt;Independent reviews of commercial brain training typically find:&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;Near transfer:&lt;/strong&gt; you get better at the trained tasks and close cousins.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Far transfer:&lt;/strong&gt; weak under careful controls for everyday reasoning / general intelligence.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Useful anchors: Simons et al. 2016; Melby-Lervåg, Redick &amp;amp; Hulme 2016 on working-memory training. Pick the app whose drills match the skill you actually want to practice.&lt;/p&gt;

&lt;p&gt;I keep a public claim-size map here: &lt;a href="https://intelligencemax.ai/guide" rel="noopener noreferrer"&gt;intelligencemax.ai/guide&lt;/a&gt; (OSF: &lt;a href="https://osf.io/kja9b/" rel="noopener noreferrer"&gt;osf.io/kja9b&lt;/a&gt;).&lt;/p&gt;

&lt;h2&gt;
  
  
  Why this matters for builders
&lt;/h2&gt;

&lt;p&gt;If you are shipping adaptive difficulty, do not market "adaptive" as "clinically validated IQ gains." Adaptive means the item bank tracks an ability estimate for &lt;em&gt;your&lt;/em&gt; task family. That is useful. It is still not an IQ test.&lt;/p&gt;

&lt;p&gt;Disclosure: I build &lt;a href="https://apps.apple.com/us/app/intelligencemax/id6786447264" rel="noopener noreferrer"&gt;IntelligenceMax&lt;/a&gt;, an iOS reasoning gym. Practice estimate ≠ clinical IQ.&lt;/p&gt;

&lt;h2&gt;
  
  
  Quick chooser
&lt;/h2&gt;

&lt;div class="table-wrapper-paragraph"&gt;&lt;table&gt;
&lt;thead&gt;
&lt;tr&gt;
&lt;th&gt;Want…&lt;/th&gt;
&lt;th&gt;Prefer…&lt;/th&gt;
&lt;/tr&gt;
&lt;/thead&gt;
&lt;tbody&gt;
&lt;tr&gt;
&lt;td&gt;Grammar / reading fluency games&lt;/td&gt;
&lt;td&gt;Elevate-like&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;Short attention/memory mini-games&lt;/td&gt;
&lt;td&gt;Lumosity/Peak-like&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;Hard MCQ reasoning under adaptive difficulty&lt;/td&gt;
&lt;td&gt;Reasoning gym&lt;/td&gt;
&lt;/tr&gt;
&lt;/tbody&gt;
&lt;/table&gt;&lt;/div&gt;

&lt;p&gt;Same aisle. Different drills. Claim size is the product.&lt;/p&gt;

</description>
      <category>analysis</category>
      <category>product</category>
      <category>productivity</category>
      <category>reviews</category>
    </item>
    <item>
      <title>What "adaptive" should mean in a reasoning practice app</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Thu, 16 Jul 2026 20:58:45 +0000</pubDate>
      <link>https://dev.to/connerlambden/what-adaptive-should-mean-in-a-reasoning-practice-app-3eko</link>
      <guid>https://dev.to/connerlambden/what-adaptive-should-mean-in-a-reasoning-practice-app-3eko</guid>
      <description>&lt;p&gt;Most apps that say adaptive just mean a streak counter and a harder level after three greens.&lt;/p&gt;

&lt;p&gt;I shipped IntelligenceMax on the App Store with a narrower definition.&lt;/p&gt;

&lt;h2&gt;
  
  
  Adaptive as serving, not as a vibe
&lt;/h2&gt;

&lt;p&gt;After each answer, the app updates a practice estimate and uses that to choose later item difficulty. Explanations appear after submit, not before. That is the loop: miss → see why → get something near that edge again.&lt;/p&gt;

&lt;h2&gt;
  
  
  What the score is not
&lt;/h2&gt;

&lt;p&gt;The estimate uses an IRT-shaped update and can be shown on an IQ-looking scale for readability. It is still an in-app practice metric. It is not a normed clinical IQ battery, and shipping to the App Store does not change that.&lt;/p&gt;

&lt;h2&gt;
  
  
  Why builders should care about the claim size
&lt;/h2&gt;

&lt;p&gt;App Store categories reward big cognitive verbs. Independent reviews of commercial brain training are usually clearer about near transfer than about far transfer. If your product only proves "people get better at your items," say that. Do not let the marketing department invent adult &lt;em&gt;g&lt;/em&gt;.&lt;/p&gt;

&lt;h2&gt;
  
  
  Try / read
&lt;/h2&gt;

&lt;ul&gt;
&lt;li&gt;App Store: &lt;a href="https://apps.apple.com/us/app/intelligencemax/id6786447264" rel="noopener noreferrer"&gt;https://apps.apple.com/us/app/intelligencemax/id6786447264&lt;/a&gt;
&lt;/li&gt;
&lt;li&gt;Methods: &lt;a href="https://intelligencemax.ai/science" rel="noopener noreferrer"&gt;https://intelligencemax.ai/science&lt;/a&gt;
&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;I build IntelligenceMax. Feedback on explanation quality beats applause for IQ theater.&lt;/p&gt;

</description>
      <category>product</category>
    </item>
    <item>
      <title>I graded my own ML option forecasts. Here's the Brier score.</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Tue, 14 Jul 2026 22:24:55 +0000</pubDate>
      <link>https://dev.to/connerlambden/i-graded-my-own-ml-option-forecasts-heres-the-brier-score-20fp</link>
      <guid>https://dev.to/connerlambden/i-graded-my-own-ml-option-forecasts-heres-the-brier-score-20fp</guid>
      <description>&lt;p&gt;In May I promised — out loud, on the internet, where people can screenshot you — that I'd grade Helium's published &lt;code&gt;prob_itm&lt;/code&gt; forecasts when the June 2026 AAPL contracts expired.&lt;/p&gt;

&lt;p&gt;June 26 came. The contracts died. So I graded them.&lt;/p&gt;

&lt;p&gt;This is not a victory lap. Mean Brier loss on &lt;strong&gt;n=2&lt;/strong&gt; is &lt;strong&gt;0.3846&lt;/strong&gt;. A coin flip that always says 0.5 scores 0.25 when the world resolves to 0 or 1. We did worse than a coin flip. That is information.&lt;/p&gt;

&lt;h2&gt;
  
  
  What we froze in May
&lt;/h2&gt;

&lt;div class="table-wrapper-paragraph"&gt;&lt;table&gt;
&lt;thead&gt;
&lt;tr&gt;
&lt;th&gt;Contract&lt;/th&gt;
&lt;th&gt;Helium &lt;code&gt;prob_itm&lt;/code&gt;
&lt;/th&gt;
&lt;th&gt;Market-ish implied&lt;/th&gt;
&lt;/tr&gt;
&lt;/thead&gt;
&lt;tbody&gt;
&lt;tr&gt;
&lt;td&gt;AAPL $310C 2026-06-26&lt;/td&gt;
&lt;td&gt;0.42&lt;/td&gt;
&lt;td&gt;~0.50&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;AAPL $295P 2026-06-26&lt;/td&gt;
&lt;td&gt;0.23&lt;/td&gt;
&lt;td&gt;~0.24&lt;/td&gt;
&lt;/tr&gt;
&lt;/tbody&gt;
&lt;/table&gt;&lt;/div&gt;

&lt;h2&gt;
  
  
  What the stock actually did
&lt;/h2&gt;

&lt;p&gt;&lt;strong&gt;AAPL NASDAQ close on 2026-06-26: $283.78&lt;/strong&gt;&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;$310 call → OTM → realized ITM = 0 → Brier = (0.42 − 0)² = &lt;strong&gt;0.1764&lt;/strong&gt;
&lt;/li&gt;
&lt;li&gt;$295 put → ITM → realized ITM = 1 → Brier = (0.23 − 1)² = &lt;strong&gt;0.5929&lt;/strong&gt;
&lt;/li&gt;
&lt;/ul&gt;

&lt;h2&gt;
  
  
  Honest read (short)
&lt;/h2&gt;

&lt;ol&gt;
&lt;li&gt;The call forecast was closer to reality than the market's ~0.50 shrug — but "closer" is not "good." Any non-zero probability on an OTM finish still costs you.&lt;/li&gt;
&lt;li&gt;The put is the bruise. Probabilities looked aligned pre-expiry; then the underlying fell hard enough that both market and model had understated the ITM risk. Helium's number was far from 1, so the Brier loss is ugly.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;n=2&lt;/strong&gt;. This is a promise kept, not a calibration paper. If you want to treat it as science, wait for more expiries on the &lt;a href="https://connerlambden.github.io/helium-news-explorer/calibration.html" rel="noopener noreferrer"&gt;honesty board&lt;/a&gt;.&lt;/li&gt;
&lt;/ol&gt;

&lt;h2&gt;
  
  
  Reproduce
&lt;/h2&gt;



&lt;div class="highlight js-code-highlight"&gt;
&lt;pre class="highlight shell"&gt;&lt;code&gt;git clone https://github.com/connerlambden/helium-mcp-cookbook
&lt;span class="nb"&gt;cd &lt;/span&gt;helium-mcp-cookbook
python calibration/grade_june_expiry.py
&lt;/code&gt;&lt;/pre&gt;

&lt;/div&gt;



&lt;p&gt;Frozen inputs live in &lt;a href="https://github.com/connerlambden/helium-mcp-cookbook/blob/main/calibration/june_2026_forecasts.json" rel="noopener noreferrer"&gt;&lt;code&gt;calibration/june_2026_forecasts.json&lt;/code&gt;&lt;/a&gt;.&lt;/p&gt;

&lt;h2&gt;
  
  
  Why bother publishing a bad score
&lt;/h2&gt;

&lt;p&gt;Because the alternative is the usual ML options blog post: a backtest, a vibe, a screenshot of a green P&amp;amp;L. Publishing &lt;code&gt;prob_itm&lt;/code&gt; with a fixed expiry is a tiny discipline. Grading it after the fact is the second half of that discipline.&lt;/p&gt;

&lt;p&gt;If you log your own forecasts with &lt;a href="https://github.com/connerlambden/helium-mcp-cookbook/blob/main/recipes/02_options_calibration_tracker.py" rel="noopener noreferrer"&gt;recipe 02&lt;/a&gt;, you can run the same scorer when your contracts expire. I'll keep adding rows.&lt;/p&gt;

&lt;p&gt;—&lt;/p&gt;

&lt;p&gt;Also adjacent, if you're here for news rather than options: we published a free &lt;a href="https://huggingface.co/datasets/HeliumTrades/helium-news-bias-corpus" rel="noopener noreferrer"&gt;212×37 news-outlet framing corpus&lt;/a&gt; with an &lt;a href="https://connerlambden.github.io/helium-news-explorer/" rel="noopener noreferrer"&gt;explorer&lt;/a&gt;. Different animal. Same "put numbers where people can grade them" instinct.&lt;/p&gt;

</description>
      <category>python</category>
      <category>machinelearning</category>
      <category>statistics</category>
    </item>
    <item>
      <title>An IRT-shaped practice score is not an IQ test</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Tue, 14 Jul 2026 22:10:27 +0000</pubDate>
      <link>https://dev.to/connerlambden/an-irt-shaped-practice-score-is-not-an-iq-test-42c8</link>
      <guid>https://dev.to/connerlambden/an-irt-shaped-practice-score-is-not-an-iq-test-42c8</guid>
      <description>&lt;p&gt;IntelligenceMax updates a learner estimate after every answer and displays it on the familiar IQ scale. The arithmetic resembles item response theory. That resemblance is useful, but it is easy to ask the formula to carry more meaning than it can bear.&lt;/p&gt;

&lt;p&gt;A formula can be internally consistent while its inputs remain uncertain. Here, the distinction begins with the item parameters.&lt;/p&gt;

&lt;h2&gt;
  
  
  The update rule
&lt;/h2&gt;

&lt;p&gt;For ability θ, item difficulty b, and discrimination a, the model assumes a logistic P(correct). The first estimate has a normal prior centred at 100 with SD 15. After an answer, the implementation takes one capped local update using score and Fisher information. Per-answer change is capped at 1.5 points.&lt;/p&gt;

&lt;p&gt;The calculation is the easy part. The inputs are where the uncertainty lives.&lt;/p&gt;

&lt;h2&gt;
  
  
  Model-assigned is not empirically calibrated
&lt;/h2&gt;

&lt;p&gt;The AI that generates each question also assigns difficulty and discrimination. Those values are checked for shape and range, but they are not fitted from a norming sample. Calling the update IRT-style describes the logistic form. It does not establish that 115 means the same thing as a score of 115 on a validated instrument.&lt;/p&gt;

&lt;h2&gt;
  
  
  What this number can support
&lt;/h2&gt;

&lt;p&gt;Within those limits, the estimate can support an adaptive practice loop. It cannot establish a clinical Full Scale IQ, a permanent change in general intelligence, transfer to school or work, or treatment of a cognitive condition.&lt;/p&gt;

&lt;p&gt;Reviews of commercial brain training find the strongest evidence on trained tasks and close relatives. Broad gains are less convincing against active controls (Simons et al., 2016; Melby-Lervåg et al., 2016).&lt;/p&gt;

&lt;p&gt;Until empirical calibration exists:&lt;/p&gt;

&lt;blockquote&gt;
&lt;p&gt;This is a logistic, IRT-shaped practice estimate that uses model-assigned item parameters. It is not a normed IQ score.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Public notes: &lt;a href="https://intelligencemax.ai/science" rel="noopener noreferrer"&gt;https://intelligencemax.ai/science&lt;/a&gt;&lt;br&gt;
Evidence map: &lt;a href="https://intelligencemax.ai/guide" rel="noopener noreferrer"&gt;https://intelligencemax.ai/guide&lt;/a&gt;&lt;br&gt;
OSF: &lt;a href="https://osf.io/kja9b/" rel="noopener noreferrer"&gt;https://osf.io/kja9b/&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Disclosure: I built IntelligenceMax. The technical question I most want readers to challenge is whether useful empirical calibration is possible for generated, mostly one-off items without quietly turning the system into a fixed bank.&lt;/p&gt;

</description>
      <category>algorithms</category>
      <category>datascience</category>
      <category>learning</category>
    </item>
    <item>
      <title>How to improve your IQ without lying to yourself</title>
      <dc:creator>connerlambden</dc:creator>
      <pubDate>Tue, 14 Jul 2026 21:02:57 +0000</pubDate>
      <link>https://dev.to/connerlambden/how-to-improve-your-iq-without-lying-to-yourself-15fh</link>
      <guid>https://dev.to/connerlambden/how-to-improve-your-iq-without-lying-to-yourself-15fh</guid>
      <description>&lt;p&gt;People type “how to improve IQ” into Google the way thirsty people type “how to hydrate instantly.” Fair. Also: the internet will sell you a miracle before it sells you sleep.&lt;/p&gt;

&lt;h2&gt;
  
  
  What IQ is (and is not)
&lt;/h2&gt;

&lt;p&gt;IQ is a test score. Useful. Incomplete. Noisy day to day. It is not a moral ranking, a destiny, or a personality.&lt;/p&gt;

&lt;p&gt;If your plan is “raise my number by Friday,” you are shopping for fog. Scores move with sleep, anxiety, practice effects, and which test you took.&lt;/p&gt;

&lt;h2&gt;
  
  
  What actually tends to help
&lt;/h2&gt;

&lt;p&gt;Schooling and hard domain practice still look better than most commercial brain-training promises. Exercise and sleep help cognition more reliably than another puzzle leaderboard. Reading hard books and taking feedback that can embarrass you remains undefeated.&lt;/p&gt;

&lt;h2&gt;
  
  
  Near transfer vs far transfer
&lt;/h2&gt;

&lt;p&gt;Near transfer is real: practice a skill, get better at that skill and nearby ones.&lt;/p&gt;

&lt;p&gt;Far transfer — train a game, raise general intelligence forever — is where ads get loud and meta-analyses get quiet. Simons et al. (2016) on commercial brain training. Melby-Lervåg and colleagues on working-memory training. Treat “this app raised my IQ 15 points” with the same skepticism you would bring to a late-night fitness infomercial.&lt;/p&gt;

&lt;h2&gt;
  
  
  A plan that is kinder than chasing points
&lt;/h2&gt;

&lt;ol&gt;
&lt;li&gt;Pick one hard skill you will actually use.&lt;/li&gt;
&lt;li&gt;Force weekly output that can be graded or rejected.&lt;/li&gt;
&lt;li&gt;Sleep like it is part of the training.&lt;/li&gt;
&lt;li&gt;Revisit IQ only if you need it for something concrete. Do not let it narrate your identity.&lt;/li&gt;
&lt;/ol&gt;

&lt;h2&gt;
  
  
  Disclosure
&lt;/h2&gt;

&lt;p&gt;I build &lt;a href="https://intelligencemax.ai" rel="noopener noreferrer"&gt;IntelligenceMax&lt;/a&gt;: adaptive reasoning practice. We do not claim it raises IQ. We care about transfer honesty.&lt;/p&gt;

&lt;p&gt;Longer map: &lt;a href="https://intelligencemax.ai/guide" rel="noopener noreferrer"&gt;intelligencemax.ai/guide&lt;/a&gt;&lt;br&gt;&lt;br&gt;
OSF notes: &lt;a href="https://osf.io/kja9b/" rel="noopener noreferrer"&gt;osf.io/kja9b&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Canonical essay: &lt;a href="https://intelmax.substack.com/p/how-to-improve-your-iq-without-lying" rel="noopener noreferrer"&gt;How to improve your IQ (without lying to yourself)&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Getting sharper is allowed. Lying to yourself about what transferred is optional, and expensive.&lt;/p&gt;

</description>
      <category>learning</category>
      <category>productivity</category>
      <category>science</category>
      <category>watercooler</category>
    </item>
  </channel>
</rss>
