Skip to content

Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence

Nvidia’s Nemotron 3.5 Lightning is an open-weights model with just 3.6 billion active parameters that matches OpenAI’s gpt-oss-120b on the Intelligence Index despite being four times smaller. At nearly 670 tokens per second, it’s also the fastest model in the comparison, showing Nvidia is betting on efficiency over raw size.

Read full article about: OpenAI introduces $125 Premium Seats for ChatGPT Business as agentic AI burns through more tokens

OpenAI is raising prices for ChatGPT Business users. The company is rolling out so-called Premium Seats, which cost $125 per user per month, or $100 with annual billing. The existing standard seats stay at $25 per month ($20 with annual billing).

Premium users get five times the usage capacity of standard users and aren't bound by the five-hour limit. Usage resets weekly. Both seat types can be mixed within the same workspace, letting admins assign the right tier to each team member. OpenAI says the change reflects how teams are tackling more complex tasks with ChatGPT and need more capacity. Put plainly, agent-based AI burns through far more tokens.

The price hike isn't a surprise. The current flat-rate plans from the major labs were always likely to be loss leaders, with increases coming eventually. Then there's the opposite approach, which Microsoft has taken with Copilot, where the company has swapped out some pricey OpenAI and Anthropic models for cheaper in-house alternatives.

Read full article about: Anthropic signs $9.1 billion data center deal with Bitcoin miner Riot Platforms

Anthropic locks in $9.1 billion data center deal with Bitcoin miner Riot Platforms. That's according to Bloomberg, citing people familiar with the matter. Riot disclosed the contract a day earlier alongside its quarterly earnings but only described the tenant as a "leading frontier AI lab."

The deal covers 191 megawatts at Riot's Rockdale site in Texas, enough power for roughly 143,000 homes, according to Bloomberg. Riot will build a data center to the tenant's specs, providing the building, power connections, cooling, and operations. Anthropic has to bring its own servers and AI chips. The lease runs 20 years, with two extension options that could push the total value to $16.1 billion. The first 96 megawatts are set to go live in December 2027, with the rest following in June 2028.

The deal adds to a growing list of Anthropic infrastructure commitments. The company is paying SpaceX an estimated $1.25 billion per month through May 2029 for the Colossus 1 data center and plans to deploy two gigawatts of AMD GPUs. Amazon is investing up to $25 billion and building up to five gigawatts of Trainium capacity with Anthropic. On top of that, gigawatts of TPU capacity from Google and Broadcom are coming online starting in 2027, and a six-year, $10 billion contract with Volta Infra rounds out the portfolio.

Nvidia guarantees its own chips' value to unlock $500 billion in AI infrastructure financing

Nvidia is teaming up with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR to mobilize over $500 billion for AI infrastructure. To win over investors, the chipmaker is guaranteeing up to 25 percent of the residual value of its own hardware. The Bank of England is already warning of systemic risks if the AI sector takes a hit.

Old OCR text cripples language model training, and FineBooks wants to fix that at scale

The FineBooks project from Hugging Face and EleutherAI tested 14 open-source OCR models on more than 2,000 historical book pages. The top model, dots.mocr, hits 97.6 percent character accuracy at under two dollars per thousand pages. That’s good enough for AI training data, but not yet for scholarly transcriptions, the team says.

Meta returns to open models with Zuckerberg's plan to out-copy China and sell compute by auction

Meta has released Muse Glimmer, the first open model from its new Superintelligence Labs. It’s a 30B agent model that runs on consumer hardware once the weights are compressed, needing less than 20 GB of memory. In an accompanying essay, Mark Zuckerberg mounts an aggressive defense of distilling other labs’ models and calls for fewer restrictions on US labs, a direct counterpunch at OpenAI and Anthropic. An open-weight version of Muse Spark 1.2 should follow soon, according to the Wall Street Journal.