<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
  <channel>
    <title>DEV Community: Ahmed Bafagih</title>
    <description>The latest articles on DEV Community by Ahmed Bafagih (@classwise).</description>
    <link>https://dev.to/classwise</link>
    <image>
      <url>https://media2.dev.to/dynamic/image/width=90,height=90,fit=cover,gravity=auto,format=auto/https:%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F4104445%2F11a1a9c1-e9a6-4c86-bb7a-ac1b6a46ad2e.png</url>
      <title>DEV Community: Ahmed Bafagih</title>
      <link>https://dev.to/classwise</link>
    </image>
    <atom:link rel="self" type="application/rss+xml" href="https://dev.to/feed/classwise"/>
    <language>en</language>
    <item>
      <title>OpenAI tests outcome-based billing, plus two more AI shifts</title>
      <dc:creator>Ahmed Bafagih</dc:creator>
      <pubDate>Tue, 01 Sep 2026 13:24:48 +0000</pubDate>
      <link>https://dev.to/classwise/openai-tests-outcome-based-billing-plus-two-more-ai-shifts-1eog</link>
      <guid>https://dev.to/classwise/openai-tests-outcome-based-billing-plus-two-more-ai-shifts-1eog</guid>
      <description>&lt;p&gt;Today's AI Newsroom covers outcome-based AI pricing, Anthropic's latest compute agreement, and new research into smaller rubric-based reinforcement learning judges.&lt;/p&gt;

&lt;h2&gt;
  
  
  OpenAI tests outcome-based billing
&lt;/h2&gt;

&lt;p&gt;OpenAI has begun letting some of its largest customers pay only when its AI actually completes the job. The arrangement is limited to select major accounts rather than offered generally.&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;Intercom charges $0.99 for each conversation its Fin agent resolves and nothing for ones it does not.&lt;/li&gt;
&lt;li&gt;Zendesk restricted billing to Verified Resolutions, confirmed by an LLM evaluation within 72 hours of the conversation.&lt;/li&gt;
&lt;li&gt;Salesforce launched Agentforce at $2 per conversation, charged for every 24-hour session whether or not anything was resolved.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Sources: &lt;a href="https://www.theinformation.com/briefings/openai-starts-letting-customers-pay-ai-works" rel="noopener noreferrer"&gt;The Information&lt;/a&gt;, &lt;a href="https://thenextweb.com/news/openai-outcome-based-pricing-enterprise" rel="noopener noreferrer"&gt;The Next Web&lt;/a&gt;, and &lt;a href="https://www.pymnts.com/news/artificial-intelligence/2026/openai-lets-some-customers-pay-only-when-ai-performs/" rel="noopener noreferrer"&gt;PYMNTS&lt;/a&gt;.&lt;/p&gt;

&lt;h2&gt;
  
  
  Anthropic signs a $35B compute deal with Lambda
&lt;/h2&gt;

&lt;p&gt;Anthropic signed a $35 billion deal for computing with Lambda. The Texas facility linked to the project in Nueces County is being built by Hut 8.&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;Lambda is a cloud company backed by Nvidia.&lt;/li&gt;
&lt;li&gt;Nvidia is expected to be the facility's lessee.&lt;/li&gt;
&lt;li&gt;Lambda has been negotiating a fundraising of up to $3 billion.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Sources: &lt;a href="https://www.moomoo.com/community/feed/anthropic-signs-a-35b-ai-compute-deal-with-nvidia-nvda-117193011429381" rel="noopener noreferrer"&gt;Moomoo&lt;/a&gt;, &lt;a href="https://www.briefs.co/news/anthropic-strikes-35-billion-compute-pact-with-lambda/" rel="noopener noreferrer"&gt;Briefs&lt;/a&gt;, and &lt;a href="https://finance.yahoo.com/technology/ai/articles/anthropic-signs-35b-computing-deal-022806081.html" rel="noopener noreferrer"&gt;Yahoo Finance&lt;/a&gt;.&lt;/p&gt;

&lt;h2&gt;
  
  
  Study tests small models as rubric-based RL judges
&lt;/h2&gt;

&lt;p&gt;Reinforcement learning from human feedback has become the dominant paradigm for aligning large language models with human preferences. Traditional RLHF relies on scalar reward signals that lack interpretability and fail to capture the multifaceted nature of response quality.&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;Rubric-guided reinforcement learning introduces structured, interpretable evaluation criteria as its backbone.&lt;/li&gt;
&lt;li&gt;Rubric-based reinforcement learning extends RL beyond tasks with exact answers or rule-based verifiers by scoring responses against instance-specific criteria.&lt;/li&gt;
&lt;li&gt;Training requires repeated rubric judging, often with proprietary APIs or local generative LLM judges with 7B parameters or more.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Sources: &lt;a href="https://arxiv.org/abs/2608.27505" rel="noopener noreferrer"&gt;Small Language Models as Judges for Rubric-Based Reinforcement Learning&lt;/a&gt; and &lt;a href="https://arxiv.org/abs/2608.30005" rel="noopener noreferrer"&gt;Rubric-based Reinforcement Learning&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;This issue was researched and written by the Classwise AI Newsroom. &lt;a href="https://classwise.io/daily/2026-09-01" rel="noopener noreferrer"&gt;Read the canonical issue&lt;/a&gt; or &lt;a href="https://classwise.io/subscribe?utm_source=syndication&amp;amp;utm_medium=article&amp;amp;utm_campaign=ai_newsroom" rel="noopener noreferrer"&gt;get the free daily briefing&lt;/a&gt;.&lt;/p&gt;

</description>
      <category>ai</category>
      <category>news</category>
      <category>machinelearning</category>
      <category>startup</category>
    </item>
  </channel>
</rss>
