Ad
Skip to content

Matthias Bastian

Matthias is the co-founder and publisher of THE DECODER, exploring how AI is fundamentally changing the relationship between humans and computers.

Investor pressure forces Nvidia to shrink its OpenAI bet just as Anthropic's numbers defy bubble warnings

Nvidia has cut its guarantee for OpenAI’s planned data center in Ohio nearly in half, from $250 billion to just under $120 billion, after investors pushed back on the risk. Meanwhile, Anthropic is complicating the AI bubble debate with revenue that jumped from $4.7 billion to $11.5 billion in a single quarter.

Anthropic announces watermark detection API that will let third parties detect Claude's AI texts

Anthropic will soon offer a watermark detection API that lets third parties check whether text was written by Claude. The technology builds on Google’s SynthID method and tweaks the randomness during word selection without affecting text quality, Anthropic says. The approach has limits with fact-heavy text, code, and heavy rewriting.

Read full article about: Alibaba's Qwen team releases Qwen 3.8 models with open weights under the Apache 2.0 license

Alibaba's AI team Qwen has released the open model weights for Qwen3.8. The core model, Qwen3.8-27B, is a multimodal dense model with 27 billion parameters that, according to Qwen, outperforms the larger Qwen3.7-Plus in coding and office tasks. The team also touts improved agent capabilities, with the model planning more independently and completing tasks more reliably. It natively handles up to 262,000 tokens of context and can scale to one million using the YaRN method. Beyond text, it processes images and videos, including diagrams, documents, and multi-hour video. A flexible thinking mode is on by default but can be toggled per query.

Image: QwenThe weights ship under the Apache 2.0 license. Qwen has also released weights for the much larger Qwen3.8-2.4T-A95B, built to operate at the Max level. Both models are available on Hugging Face and ModelScope. A hosted version with one million tokens of context will soon be available through Qwen Cloud, Alibaba's AI service.

OpenAI's Computer History turns your clicks and keystrokes into a searchable ChatGPT memory timeline

OpenAI’s Computer History records clicks, keystrokes, and app switches on Mac and turns them into a searchable timeline for ChatGPT and Codex. The data is stored locally as unencrypted Markdown files. OpenAI says it’s not used for AI training, but memories that feed into chats may still end up as training data.

Claude Code now runs daily maintenance on Anthropic's software with a 46 percent merge rate

Anthropic is testing whether Claude Code can handle daily maintenance of the company’s own apps, from crash fuzzing to dead-code removal. In a few weeks, the AI created 388 pull requests, and 46 percent were merged after human review. Claude Code inventor Boris Cherny sees this as “early signs of life that this might be possible.”

Read full article about: Zhipu AI releases GLM-5.3, claims it's the strongest open-weights coding model

Chinese AI startup Zhipu AI has released GLM-5.3. The model shares the same base as its predecessor, GLM-5.2, and all gains come from extended post-training alone. Zhipu says GLM-5.3 is the most powerful open-weights coding model, with the biggest jumps in agent-based tasks.

Image: Z.ai

One area where top Chinese models like Kimi or Qwen still lag behind US frontier models is cybersecurity. Zhipu trained GLM-5.3 with data and environments built to find software vulnerabilities. According to Z.ai, the model "began to reason across multiple stages of exploitation, forming coherent plans for complete exploitation chains." Working with security teams in China, the company says it found 2,436 vulnerabilities across 269 projects, some up to 40 years old. The flaws are documented in a public registry.

Image: Z.ai

GLM-5.3 is available now through the GLM Coding Plan and works with coding agents like ZCode, Claude Code, or OpenCode. The model weights are set to go open source in two weeks, once security reviews wrap up.

Read full article about: Suno Studio 2.0's new chat feature lets you talk to your DAW like it's a bandmate

With Studio 2.0, Suno turns its AI music platform into a chat-driven DAW (Digital Audio Workstation) for Premier subscribers. A new beta chat feature lets users talk to Studio like a bandmate, creating instruments, vocals, or custom plugins through plain text. The update also adds MIDI import, recording, and editing, stem separation, automation curves, and unrestricted multitrack export at 32-bit/48 kHz. Plugin creation doesn't burn any credits yet, though Suno says a credit system may come later.

The unlimited export feature stands out against the backdrop of Suno's recent crackdown on AI music spam. CEO Mikey Shulman had just rolled out new download limits for lower tiers to curb mass distribution of AI-generated songs on streaming platforms. Premier subscribers have always had unrestricted exports, a sign that Suno sees paying professionals as a lower spam risk.

The company has also drawn fire for using copyrighted material to train its music models. A German court recently ruled that the service violates German copyright law and that the training doesn't qualify under the U.S. "fair use" doctrine.

Update: A previous version of this article incorrectly stated that unlimited exports were a new feature for Premier subscribers. Premier subscribers have always had unrestricted exports.

Comment Source: Suno
Read full article about: Gemini 3.7 Flash lands with coding gains and undercuts its three-week-old predecessor's price by 50%

Google has yet another Flash model. Just three weeks after Gemini 3.6 Flash, Google is already shipping Gemini 3.7 Flash. The company calls it its most capable workhorse model yet for coding and AI agents. Google credits "awesome algorithmic improvements" for the rapid jump in "intelligence".

Code quality sees the biggest gains over its predecessor. On the FrontierCode benchmark, the model scores 43.6 percent, up from 34.4 percent. On DeepSWE, it hits 65.3 percent versus 49.0 percent. According to Google's own measurements, that puts it ahead of both Claude Sonnet 5 and GPT-5.6 Terra. Google's benchmarks also show gains in web development, document comprehension, and business process automation.

Benchmark results as measured by Google. Gemini Flash 3.7 aims to set a new standard for price-to-performance. | Image: Google

The model is available through the API, AI Studio, and Antigravity. Launch pricing sits at $0.75 per million input tokens and $3.75 per million output tokens, 50 percent cheaper than 3.6 Flash at launch. Both models now share the same price point. That pricing holds through the end of the year; the models likely won't.