AI NEWSROOMBY CLASSWISE.IO

Tuesday, September 1, 2026

OpenAI tests outcome-based billing

Plus: Anthropic signs $35B compute deal with Lambda; Study tests small models as rubric-based RL judges

Welcome to a Tuesday where the industry is rethinking how we measure value, from outcome-based billing to the very rubrics we use to judge intelligence.

PRICING

💰 OpenAI tests outcome-based billing

OpenAI has begun letting some of its largest customers pay only when its AI actually completes the job. The arrangement is limited to select major accounts rather than offered generally.

The details:

  • Intercom charges $0.99 for each conversation its Fin agent resolves and nothing for ones it does not.
  • Zendesk restricted billing to Verified Resolutions, confirmed by an LLM evaluation within 72 hours of the conversation.
  • Salesforce launched Agentforce at $2 per conversation, charged for every 24-hour session whether or not anything was resolved.

First reported by theinformation.com, 23h ago · theinformation.com · thenextweb.com · pymnts.com

COMPUTE

🤝 Anthropic signs $35B compute deal with Lambda

Anthropic signs $35B compute deal with Lambda
Image source: moomoo.com

Anthropic signed a $35 billion deal for computing with Lambda. The Texas facility linked to the project in Nueces County is being built by Hut 8.

The details:

  • Lambda is a cloud company backed by Nvidia.
  • Nvidia is expected to be the facility's lessee.
  • Lambda has been negotiating a fundraising of up to $3 billion.

First reported by moomoo.com, 3h ago · moomoo.com · briefs.co · finance.yahoo.com

MODELS

⚖️ Study tests small models as rubric-based RL judges

Reinforcement learning from human feedback (RLHF) has become the dominant paradigm for aligning large language models (LLMs) with human preferences. Traditional RLHF relies on scalar reward signals that lack interpretability and fail to capture the multifaceted nature of response quality.

The details:

  • Rubric-guided reinforcement learning addresses these limitations by introducing structured, interpretable evaluation criteria, or rubrics, as the backbone.
  • Rubric-based reinforcement learning extends RL beyond tasks with exact answers or rule-based verifiers by scoring responses against instance-specific criteria.
  • Training requires repeated rubric judging, often with proprietary APIs or local generative LLM judges with 7B parameters or more.

First reported by arxiv.org, 24h ago · arxiv.org · arxiv.org

📰 Everything else in AI today

🛠️ Tool of the day

SCALR AI SCALR AI provides security teams a free platform to automate and manage security operations center workflows.

Researched and written by AI Newsroom, by Classwise.io

See how AI Newsroom is researched and checked

Get tomorrow's issue in your inbox.

Free, daily, 4:57 AM ET. Five minutes and you know what matters in AI.

Join for free