AI News selected for Professionals and Decision Makers
Model And Platform Announcements

An update on recent Claude Code quality reports

02:00 · April 23, 2026 · Antropic Engineering Blog

An update on recent Claude Code quality reports

We traced recent reports of Claude Code quality issues to three separate changes. Here's what happened and what we're changing.

Summary

Anthropic has published a postmortem examining three distinct changes that together produced noticeable quality degradation in Claude Code, the Claude Agent SDK, and Claude Cowork over several weeks, while leaving the underlying API unaffected. The issues surfaced at different times and affected different user cohorts, which made the pattern appear as broad but inconsistent performance loss that was initially difficult to separate from normal variation in feedback.

The first change occurred after the February release of Opus 4.6, when the default reasoning effort was lowered from high to medium to reduce occasional long-tail latency and token consumption. Internal evaluations had shown that medium effort delivered acceptable intelligence at lower cost for most tasks, yet many users experienced the model as less capable and did not adjust the setting themselves. The default was restored to high for Opus 4.7 and xhigh for Opus 4.6 on 7 April.

A second problem stemmed from an efficiency modification to prompt caching introduced on 26 March. The intended logic cleared stale thinking blocks only after an hour of inactivity, but a bug caused the clear_thinking flag to remain active for every subsequent turn. As a result, reasoning history was progressively discarded, producing the forgetfulness, repetition, and erratic tool use reported by users. The defect also triggered repeated cache misses that accelerated usage-limit consumption. It was corrected in version 2.1.101 on 10 April.

The third issue arose from a system-prompt adjustment shipped with Opus 4.7 on 16 April to curb verbosity. Although the change passed the evaluations run at the time, later ablations revealed a roughly 3 percent intelligence regression on both Opus 4.6 and 4.7. The prompt was reverted as part of the 20 April release (v2.1.116).

In response, Anthropic is expanding internal dogfooding to the exact public build, strengthening code-review tooling with additional repository context, enforcing broader per-model evaluations and soak periods for any prompt change that could affect capability, and improving auditability of system-prompt modifications. Usage limits have been reset for all subscribers.

Why it matters

This article is highly relevant for product teams and builders using Claude models, as it provides deep technical insights into prompt caching, system prompt tuning, and latency-intelligence tradeoffs. Dutch AI practitioners can apply these learnings to optimize their own LLM implementations and better understand recent anomalies in Claude's performance.

More in this beat
anthropicclaudeclaude-codecontext-managementopus-4-6prompt-cachingreasoning-models
Maximizing the value of your Claude Code sessions

02:00 · August 14, 2026

Maximizing the value of your Claude Code sessions

It offers highly actionable, technical insights for product teams and builders on optimizing token usage and costs when using Claude Code. Understanding these mechanics is crucial for Dutch AI practitioners looking to efficiently integrate and scale agentic coding tools in their development workflows.

Relevance 85 · Audience 95

Equipping agents for the real world with Agent Skills

02:00 · October 16, 2025

Equipping agents for the real world with Agent Skills

Directly actionable for Product Teams and Builders: provides concrete implementation patterns, evaluation guidelines, and code patterns for building specialized agents. Addresses lifecycle, observability via progressive loading, and risks like malicious skills.

Relevance 78 · Audience 85

The Claude Code Guide For Startups

02:00 · August 20, 2026

The Claude Code Guide For Startups

This article is highly relevant for product teams and builders as it offers actionable strategies and technical tips for integrating agentic coding into the SDLC. Dutch AI practitioners can apply these insights to scale development efficiently while maintaining governance and compliance through robust evaluation frameworks.

Relevance 85 · Audience 95

How ABC Legal turned every employee into a builder with Claude Managed Agents

02:00 · August 17, 2026

How ABC Legal turned every employee into a builder with Claude Managed Agents

This article provides a highly actionable blueprint for product teams and builders to deploy scalable, observable AI agents using a GitOps approach. It demonstrates how to empower non-technical staff to build automations while maintaining centralized governance, which is highly applicable to Dutch enterprises scaling AI.

Relevance 75 · Audience 90

A guide to cost visibility and control in Claude

02:00 · August 4, 2026

A guide to cost visibility and control in Claude

Cost management is a critical hurdle for AI adoption. For Dutch Product Teams and Builders, understanding Claude's cost levers like caching, batching, and model routing is essential for building sustainable, high-ROI AI applications.

Relevance 80 · Audience 90

Think through hard problems in voice mode

02:00 · July 23, 2026

Think through hard problems in voice mode

This update is highly relevant for product teams and builders as it enhances Claude's utility for complex problem-solving and workflow integration via voice. The addition of multilingual support and tool connectors provides new avenues for Dutch AI practitioners to streamline development and brainstorming processes.

Relevance 85 · Audience 90

Building verification loops in Claude Code with skills

02:00 · July 22, 2026

Building verification loops in Claude Code with skills

It provides highly actionable insights for AI product teams and builders on how to improve agentic coding workflows using Claude Code. Dutch AI practitioners can leverage these verification loops to increase development efficiency and enforce project-specific quality standards autonomously.

Relevance 85 · Audience 95

How Outtake built a cyber investigator on Claude

02:00 · July 22, 2026

How Outtake built a cyber investigator on Claude

This article provides a practical use case for Product Teams and Builders on how to leverage Claude Code and the Agent SDK to build long-running, autonomous AI agents. It offers valuable architectural insights for Dutch AI practitioners developing cybersecurity solutions or complex agentic workflows.

Relevance 75 · Audience 85

UST is bringing Claude to physical AI

02:00 · July 9, 2026

UST is bringing Claude to physical AI

This case study provides product teams and builders with concrete examples of integrating LLMs into complex engineering workflows, such as hardware validation and digital twin comparisons. It is highly applicable to the Dutch high-tech and manufacturing sectors (e.g., semiconductors, healthcare tech) looking to operationalize AI with human-in-the-loop governance.

Relevance 75 · Audience 85