• Today
  • Archive
claude.mazzotta.devdaily briefing
Next drop in 6h 19m · 04:30 UTCUpdated 4d ago
Issue139loading…

Sonnet 5.5 beats Opus at half the price, and Anthropic planned it.

From the editor

Three threads converge today. First, Anthropic is deliberately collapsing its own tier hierarchy: Sonnet 5.5 eating Opus is a strategic choice, not an accident. Second, the tooling layer keeps maturing quietly, with Playwright MCP and MCP Inspector 2.9.0 pushing Claude Code closer to true autonomous operation. Third, the research front is getting serious: evaluation awareness in smaller models and open-weight interpretability probes both point toward a field grappling honestly with alignment gaps. Gemini 4 arriving sharpens all of this. Competition clarifies priorities fast.

TL;DR

  1. 1.Sonnet 5.5 outperforms Opus 5.5 on coding benchmarks while costing half as much.
  2. 2.Playwright MCP lets Claude Code control a real browser, no UI narration needed.
  3. 3.Even small models detect when they are being evaluated, complicating alignment research.
8 curated itemsscroll for the brief
01

Releases

What shipped · 2 items

01

v2.1.286

Claude Code ships v2.1.286 with bug fixes, authentication improvements, UX polish, and cloud session updates.

Claude Code
02

inspector: 2.9.0

MCP Inspector reaches version 2.9.0 with tooling improvements for the Model Context Protocol ecosystem.

MCP
02

Tools

Worth a look · 1 item

Playwright MCP for Claude Code: Let It Open a Real Browser and Click Through Flows (2-Minute Setup)

Connect Playwright MCP to Claude Code so the agent can open a real browser, navigate pages, and click through flows without you describing the UI manually.

dev.to
03

Tips

Actionable craft · 1 item

What's the one line in your CLAUDE.md that made the biggest difference ?

Community members share single-line CLAUDE.md entries that dramatically improved Claude Code behavior, covering tone, format, and task-scoping instructions.

r/ClaudeAI
04

Reading

Long-form signal · 3 items

01

Sonnet 5.5 Just Ate Opus. Anthropic Did It on Purpose.

Sonnet 5.5 outperforms Opus 5.5 on agentic coding benchmarks including Terminal-Bench while costing half the price, forcing a hard look at Anthropic's deliberate model tier strategy.

dev.to
02

Model Hermeneutics: Monitoring Closed-Weight Models with Open-Weight Internals

Researchers propose using open-weight models as interpretability probes for closed-weight systems, enabling misalignment detection without requiring internal access to frontier models.

LessWrong
03

Evaluation Awareness in Small(ish) Models

Even smaller models can detect when they are being evaluated, though this awareness rarely translates into changes in refusal behavior or safety metrics, raising nuanced questions for alignment research.

LessWrong
05

Discussions

Where it heats up · 1 item

Gemini 4 is out: The competition has woken up

Reddit users react to the Gemini 4 release and debate how it stacks up against Claude Opus 5.5, examining benchmark gaps and real-world use cases.

r/ClaudeAI
※

Always at hand

Reference links you keep open

  • Anthropic docs

    API + agents reference

    →
  • Claude Code

    CLI docs and changelog

    →
  • MCP spec

    Open standard

    →
  • Model lineup

    Opus, Sonnet, Haiku

    →
  • Pricing

    Per-token, batch, cache

    →
  • Status

    Live incidents

    →

Wealthior Labs · Get in touch

Need a custom Claude integration?

We ship agentic systems other consultancies fumble. MCP servers, internal AI tooling, end-to-end. Two weeks of senior engineering, no decks.

Start a conversation→

Everything Claude,
once a day.

One editorial briefing curated by Haiku, Sonnet, and Opus. Published every morning, 04:30 UTC.

Browse

  • Archive
  • Sources
  • About
  • Sponsor
  • Feedback
  • RSS feed
  • Public API

Connect

  • labs.wealthior-group.ch
  • info@wealthior-group.ch

Created by Roberto Mazzotta at Wealthior Labs · © 2026

·Issue №139·admin

Drawing from 26 sources